🏷️ DevOps

60 articles with this tag · all tags

Hugging Face passes 3 million models: abundance becomes a curation problem
LLM Hugging Face passes 3 million models: abundance becomes a curation problem 2026-08-18
Qwen3.8-27B on an RTX PRO 6000: eight hours and $650 in API costs avoided
Altro Qwen3.8-27B on an RTX PRO 6000: eight hours and $650 in API costs avoided 2026-08-18
Hydra update brings VRAM and power limit controls to RTX 50 GPUs up to +3000 MHz
Hardware Hydra update brings VRAM and power limit controls to RTX 50 GPUs up to +3000 MHz 2026-08-18
Self-improving AI hits a wall: agents fail open-ended research
LLM Self-improving AI hits a wall: agents fail open-ended research 2026-08-18
Microsoft rebrands its products 158 times: a registry that exposes hidden operational costs
Market Microsoft rebrands its products 158 times: a registry that exposes hidden operational costs 2026-08-18
HarmProfile: frontier LLM risk is a distribution, not a failure
LLM HarmProfile: frontier LLM risk is a distribution, not a failure 2026-08-18
FPO Speeds Up LLM Fine-Tuning Without Cross-Layer Backpropagation
LLM FPO Speeds Up LLM Fine-Tuning Without Cross-Layer Backpropagation 2026-08-18
Medical LLMs: Partial Confidence Calibration and Errors in Ambiguous Cases
LLM Medical LLMs: Partial Confidence Calibration and Errors in Ambiguous Cases 2026-08-18
CDW raises RTX Pro 6000 price to $19,999: list update or leak?
Market CDW raises RTX Pro 6000 price to $19,999: list update or leak? 2026-08-18
Rust on GPUs: memory safety beyond CUDA and HIP
Frameworks Rust on GPUs: memory safety beyond CUDA and HIP 2026-08-17
llama.cpp v0.1.0 marks the move to semantic versioning
Frameworks llama.cpp v0.1.0 marks the move to semantic versioning 2026-08-17
AMD Works on a New ROCm Backend for Virtualized GPU Compute in QEMU
Altro AMD Works on a New ROCm Backend for Virtualized GPU Compute in QEMU 2026-08-17
KTransformers 0.7 Expands AVX-512 Support to Benefit AMD EPYC Servers
Frameworks KTransformers 0.7 Expands AVX-512 Support to Benefit AMD EPYC Servers 2026-08-17
GPU prices rising: PC Partner warns of budget card shortages
Hardware GPU prices rising: PC Partner warns of budget card shortages 2026-08-17
Linux 7.3 redeems ARM64 with NVIDIA Olympus workarounds after AI patch chaos
Hardware Linux 7.3 redeems ARM64 with NVIDIA Olympus workarounds after AI patch chaos 2026-08-17
BCMT: Blockwise Causal Memory Reduces the Weight of Global Attention
LLM BCMT: Blockwise Causal Memory Reduces the Weight of Global Attention 2026-08-17
Self-Explainable Latent Reasoning: One Model for Efficiency and Interpretability
LLM Self-Explainable Latent Reasoning: One Model for Efficiency and Interpretability 2026-08-17
Coding benchmarks don't prove general capability: optimization needs diverse evaluation
LLM Coding benchmarks don't prove general capability: optimization needs diverse evaluation 2026-08-17
Depth-aware expert masking: MoE late layers absorb pruning better than early ones
LLM Depth-aware expert masking: MoE late layers absorb pruning better than early ones 2026-08-17
Linux 7.2 stable: faster I/O and refreshed AMD/Intel drivers
Altro Linux 7.2 stable: faster I/O and refreshed AMD/Intel drivers 2026-08-16
Why the AI world keeps thanking Georgi Gerganov and llama.cpp
Frameworks Why the AI world keeps thanking Georgi Gerganov and llama.cpp 2026-08-16
Google reportedly turns to AMD for next-generation TPU design
Hardware Google reportedly turns to AMD for next-generation TPU design 2026-08-16
Qwen3.8-27B beats Qwen3.6-27B in autonomous iteration on a BASIC ray tracer
LLM Qwen3.8-27B beats Qwen3.6-27B in autonomous iteration on a BASIC ray tracer 2026-08-16
Qwen3.8-27B runs locally and one-shots a Super Mario clone
Altro Qwen3.8-27B runs locally and one-shots a Super Mario clone 2026-08-15
RustConn 0.20 Polishes a GTK4 Connection Manager, Quietly Helping On-Prem Operations
Altro RustConn 0.20 Polishes a GTK4 Connection Manager, Quietly Helping On-Prem Operations 2026-08-15
KDE Plasma 6.8 adds fine-grained control over mouse and touchpad speed
Frameworks KDE Plasma 6.8 adds fine-grained control over mouse and touchpad speed 2026-08-15
Debian developers vote on LLM use in the project: governance and trust at stake
LLM Debian developers vote on LLM use in the project: governance and trust at stake 2026-08-15
Qwen 3.8 35BA3B appears in a commit: a signal before the launch
LLM Qwen 3.8 35BA3B appears in a commit: a signal before the launch 2026-08-15
Qwen 3.8 27B Release Day: Local Formats and the Deployment Shift
LLM Qwen 3.8 27B Release Day: Local Formats and the Deployment Shift 2026-08-15
Lemonade 11.6 Brings Muse-Glimmer 30B and Experimental ROCm Image Generation to Local AI
Frameworks Lemonade 11.6 Brings Muse-Glimmer 30B and Experimental ROCm Image Generation to Local AI 2026-08-14
LLM self-reflection: action routing beats diagnostic questions and taxonomies
LLM LLM self-reflection: action routing beats diagnostic questions and taxonomies 2026-08-14
Doom inside an LLM: a 34 GB checkpoint and token-based rendering
Hardware Doom inside an LLM: a 34 GB checkpoint and token-based rendering 2026-08-13
Writer targets token cost containment with new LLM based on GLM-5.2
LLM Writer targets token cost containment with new LLM based on GLM-5.2 2026-08-13
DeepSeek V4 Pro 0813 on Hugging Face: A Name Is Not Enough
LLM DeepSeek V4 Pro 0813 on Hugging Face: A Name Is Not Enough 2026-08-13
Qwen opens official countdown for Qwen3.8-27B on Hugging Face
LLM Qwen opens official countdown for Qwen3.8-27B on Hugging Face 2026-08-13
Nvidia doubles RTX PRO 6000 Blackwell price to $16,000, raising on-prem AI costs
Hardware Nvidia doubles RTX PRO 6000 Blackwell price to $16,000, raising on-prem AI costs 2026-08-13
Retrofitting Recurrent Depth into Pretrained LLMs: Faster Latent Reasoning with Sharp Limits
LLM Retrofitting Recurrent Depth into Pretrained LLMs: Faster Latent Reasoning with Sharp Limits 2026-08-13
AI Detectors Are Failing Academic Integrity by Penalizing Transparent Use
Altro AI Detectors Are Failing Academic Integrity by Penalizing Transparent Use 2026-08-13
Distribird brings Bayesian calibration to local, open-weight LLMs
Altro Distribird brings Bayesian calibration to local, open-weight LLMs 2026-08-13
Governing Conflicting LLMs: The Control Layer That Prevents Conversational Collapse
Frameworks Governing Conflicting LLMs: The Control Layer That Prevents Conversational Collapse 2026-08-13
Comma.ai launches Chestnut dock with AMD GPU and open-source firmware
Hardware Comma.ai launches Chestnut dock with AMD GPU and open-source firmware 2026-08-13
Supply-chain attack on LiteLLM exposes terabytes of credentials
Altro Supply-chain attack on LiteLLM exposes terabytes of credentials 2026-08-12
AI safety concerns grow, at Ai4 Hinton, Li and Ng discuss the value of openness
Altro AI safety concerns grow, at Ai4 Hinton, Li and Ng discuss the value of openness 2026-08-12
Linux Unlocks Hybrid Graphics on 2018–2019 MacBook Pros: A Boost for Local Inference, Too
Hardware Linux Unlocks Hybrid Graphics on 2018–2019 MacBook Pros: A Boost for Local Inference, Too 2026-08-12
Decoding the hidden reasoning of Claude and GPT: what it changes
Altro Decoding the hidden reasoning of Claude and GPT: what it changes 2026-08-12
LACT 0.10 Brings NVIDIA Overclocking and Blackwell Hotspot Sensing for Linux GPU Enthusiasts
Hardware LACT 0.10 Brings NVIDIA Overclocking and Blackwell Hotspot Sensing for Linux GPU Enthusiasts 2026-08-12
Intel LLM-Scaler Now Supports Muse Glimmer, Simplifying Local Inference on Arc (Pro) B GPUs
Frameworks Intel LLM-Scaler Now Supports Muse Glimmer, Simplifying Local Inference on Arc (Pro) B GPUs 2026-08-12
AMD: AI Agents Will Push the CPU-GPU Ratio Toward 1:1
Hardware AMD: AI Agents Will Push the CPU-GPU Ratio Toward 1:1 2026-08-12
EU Mandates Watermarking for Local Models Too — What It Means for Open Source AI
Altro EU Mandates Watermarking for Local Models Too — What It Means for Open Source AI 2026-08-12
Robust for conflict, weak for morality: LLM pipelines tested on French headlines
LLM Robust for conflict, weak for morality: LLM pipelines tested on French headlines 2026-08-12
LLM Agents Factory: An Agent Factory That Cuts Inference Costs
Frameworks LLM Agents Factory: An Agent Factory That Cuts Inference Costs 2026-08-12
How topology reveals the inner evolution of Transformers
LLM How topology reveals the inner evolution of Transformers 2026-08-12
Autonomous steering for vertical farms: LLM slashes energy and time by up to 68%
Altro Autonomous steering for vertical farms: LLM slashes energy and time by up to 68% 2026-08-12
Nemotron-3.5 Lightning: NVIDIA’s bet on efficiency for local inference
LLM Nemotron-3.5 Lightning: NVIDIA’s bet on efficiency for local inference 2026-08-11
FastFlowLM 1.0: AMD brings NPU AI under the ROCm umbrella
Frameworks FastFlowLM 1.0: AMD brings NPU AI under the ROCm umbrella 2026-08-11
Luth-2: French small language models beat 3x larger competitors, redefining local AI
LLM Luth-2: French small language models beat 3x larger competitors, redefining local AI 2026-08-11
LLMs and Waste Management: WuYuEval Reveals the Limits of Generalist AI
LLM LLMs and Waste Management: WuYuEval Reveals the Limits of Generalist AI 2026-08-11
Self-adaptive fuzzing exposes the hallucination cracks in multimodal LLMs
Frameworks Self-adaptive fuzzing exposes the hallucination cracks in multimodal LLMs 2026-08-11
Training a 1B LLM from scratch for under $200: the frontier of accessible self-hosting
LLM Training a 1B LLM from scratch for under $200: the frontier of accessible self-hosting 2026-08-11
Linux: Open-Source NVIDIA “Nova” Driver Gains Functionality with Rust
Hardware Linux: Open-Source NVIDIA “Nova” Driver Gains Functionality with Rust 2026-08-11