🏷️ DevOps
60 articles with this tag · all tags
LLM
Hugging Face passes 3 million models: abundance becomes a curation problem
2026-08-18
Altro
Qwen3.8-27B on an RTX PRO 6000: eight hours and $650 in API costs avoided
2026-08-18
Hardware
Hydra update brings VRAM and power limit controls to RTX 50 GPUs up to +3000 MHz
2026-08-18
LLM
Self-improving AI hits a wall: agents fail open-ended research
2026-08-18
Market
Microsoft rebrands its products 158 times: a registry that exposes hidden operational costs
2026-08-18
LLM
HarmProfile: frontier LLM risk is a distribution, not a failure
2026-08-18
LLM
FPO Speeds Up LLM Fine-Tuning Without Cross-Layer Backpropagation
2026-08-18
LLM
Medical LLMs: Partial Confidence Calibration and Errors in Ambiguous Cases
2026-08-18
Market
CDW raises RTX Pro 6000 price to $19,999: list update or leak?
2026-08-18
Frameworks
Rust on GPUs: memory safety beyond CUDA and HIP
2026-08-17
Frameworks
llama.cpp v0.1.0 marks the move to semantic versioning
2026-08-17
Altro
AMD Works on a New ROCm Backend for Virtualized GPU Compute in QEMU
2026-08-17
Frameworks
KTransformers 0.7 Expands AVX-512 Support to Benefit AMD EPYC Servers
2026-08-17
Hardware
GPU prices rising: PC Partner warns of budget card shortages
2026-08-17
Hardware
Linux 7.3 redeems ARM64 with NVIDIA Olympus workarounds after AI patch chaos
2026-08-17
LLM
BCMT: Blockwise Causal Memory Reduces the Weight of Global Attention
2026-08-17
LLM
Self-Explainable Latent Reasoning: One Model for Efficiency and Interpretability
2026-08-17
LLM
Coding benchmarks don't prove general capability: optimization needs diverse evaluation
2026-08-17
LLM
Depth-aware expert masking: MoE late layers absorb pruning better than early ones
2026-08-17
Altro
Linux 7.2 stable: faster I/O and refreshed AMD/Intel drivers
2026-08-16
Frameworks
Why the AI world keeps thanking Georgi Gerganov and llama.cpp
2026-08-16
Hardware
Google reportedly turns to AMD for next-generation TPU design
2026-08-16
LLM
Qwen3.8-27B beats Qwen3.6-27B in autonomous iteration on a BASIC ray tracer
2026-08-16
Altro
Qwen3.8-27B runs locally and one-shots a Super Mario clone
2026-08-15
Altro
RustConn 0.20 Polishes a GTK4 Connection Manager, Quietly Helping On-Prem Operations
2026-08-15
Frameworks
KDE Plasma 6.8 adds fine-grained control over mouse and touchpad speed
2026-08-15
LLM
Debian developers vote on LLM use in the project: governance and trust at stake
2026-08-15
LLM
Qwen 3.8 35BA3B appears in a commit: a signal before the launch
2026-08-15
LLM
Qwen 3.8 27B Release Day: Local Formats and the Deployment Shift
2026-08-15
Frameworks
Lemonade 11.6 Brings Muse-Glimmer 30B and Experimental ROCm Image Generation to Local AI
2026-08-14
LLM
LLM self-reflection: action routing beats diagnostic questions and taxonomies
2026-08-14
Hardware
Doom inside an LLM: a 34 GB checkpoint and token-based rendering
2026-08-13
LLM
Writer targets token cost containment with new LLM based on GLM-5.2
2026-08-13
LLM
DeepSeek V4 Pro 0813 on Hugging Face: A Name Is Not Enough
2026-08-13
LLM
Qwen opens official countdown for Qwen3.8-27B on Hugging Face
2026-08-13
Hardware
Nvidia doubles RTX PRO 6000 Blackwell price to $16,000, raising on-prem AI costs
2026-08-13
LLM
Retrofitting Recurrent Depth into Pretrained LLMs: Faster Latent Reasoning with Sharp Limits
2026-08-13
Altro
AI Detectors Are Failing Academic Integrity by Penalizing Transparent Use
2026-08-13
Altro
Distribird brings Bayesian calibration to local, open-weight LLMs
2026-08-13
Frameworks
Governing Conflicting LLMs: The Control Layer That Prevents Conversational Collapse
2026-08-13
Hardware
Comma.ai launches Chestnut dock with AMD GPU and open-source firmware
2026-08-13
Altro
Supply-chain attack on LiteLLM exposes terabytes of credentials
2026-08-12
Altro
AI safety concerns grow, at Ai4 Hinton, Li and Ng discuss the value of openness
2026-08-12
Hardware
Linux Unlocks Hybrid Graphics on 2018–2019 MacBook Pros: A Boost for Local Inference, Too
2026-08-12
Altro
Decoding the hidden reasoning of Claude and GPT: what it changes
2026-08-12
Hardware
LACT 0.10 Brings NVIDIA Overclocking and Blackwell Hotspot Sensing for Linux GPU Enthusiasts
2026-08-12
Frameworks
Intel LLM-Scaler Now Supports Muse Glimmer, Simplifying Local Inference on Arc (Pro) B GPUs
2026-08-12
Hardware
AMD: AI Agents Will Push the CPU-GPU Ratio Toward 1:1
2026-08-12
Altro
EU Mandates Watermarking for Local Models Too — What It Means for Open Source AI
2026-08-12
LLM
Robust for conflict, weak for morality: LLM pipelines tested on French headlines
2026-08-12
Frameworks
LLM Agents Factory: An Agent Factory That Cuts Inference Costs
2026-08-12
LLM
How topology reveals the inner evolution of Transformers
2026-08-12
Altro
Autonomous steering for vertical farms: LLM slashes energy and time by up to 68%
2026-08-12
LLM
Nemotron-3.5 Lightning: NVIDIA’s bet on efficiency for local inference
2026-08-11
Frameworks
FastFlowLM 1.0: AMD brings NPU AI under the ROCm umbrella
2026-08-11
LLM
Luth-2: French small language models beat 3x larger competitors, redefining local AI
2026-08-11
LLM
LLMs and Waste Management: WuYuEval Reveals the Limits of Generalist AI
2026-08-11
Frameworks
Self-adaptive fuzzing exposes the hallucination cracks in multimodal LLMs
2026-08-11
LLM
Training a 1B LLM from scratch for under $200: the frontier of accessible self-hosting
2026-08-11
Hardware
Linux: Open-Source NVIDIA “Nova” Driver Gains Functionality with Rust
2026-08-11