🏷️ Fine-Tuning

60 articles with this tag · all tags

Doom inside an LLM: a 34 GB checkpoint and token-based rendering
Hardware Doom inside an LLM: a 34 GB checkpoint and token-based rendering 2026-08-13
Writer targets token cost containment with new LLM based on GLM-5.2
LLM Writer targets token cost containment with new LLM based on GLM-5.2 2026-08-13
Qwen opens official countdown for Qwen3.8-27B on Hugging Face
LLM Qwen opens official countdown for Qwen3.8-27B on Hugging Face 2026-08-13
Nvidia doubles RTX PRO 6000 Blackwell price to $16,000, raising on-prem AI costs
Hardware Nvidia doubles RTX PRO 6000 Blackwell price to $16,000, raising on-prem AI costs 2026-08-13
Retrofitting Recurrent Depth into Pretrained LLMs: Faster Latent Reasoning with Sharp Limits
LLM Retrofitting Recurrent Depth into Pretrained LLMs: Faster Latent Reasoning with Sharp Limits 2026-08-13
Backtrader-Bench: Forcing LLMs to Run Code in Algorithmic Trading Benchmarks
Frameworks Backtrader-Bench: Forcing LLMs to Run Code in Algorithmic Trading Benchmarks 2026-08-13
AI safety concerns grow, at Ai4 Hinton, Li and Ng discuss the value of openness
Altro AI safety concerns grow, at Ai4 Hinton, Li and Ng discuss the value of openness 2026-08-12
Linux Unlocks Hybrid Graphics on 2018–2019 MacBook Pros: A Boost for Local Inference, Too
Hardware Linux Unlocks Hybrid Graphics on 2018–2019 MacBook Pros: A Boost for Local Inference, Too 2026-08-12
Decoding the hidden reasoning of Claude and GPT: what it changes
Altro Decoding the hidden reasoning of Claude and GPT: what it changes 2026-08-12
LACT 0.10 Brings NVIDIA Overclocking and Blackwell Hotspot Sensing for Linux GPU Enthusiasts
Hardware LACT 0.10 Brings NVIDIA Overclocking and Blackwell Hotspot Sensing for Linux GPU Enthusiasts 2026-08-12
EU Mandates Watermarking for Local Models Too — What It Means for Open Source AI
Altro EU Mandates Watermarking for Local Models Too — What It Means for Open Source AI 2026-08-12
LLM Agents Factory: An Agent Factory That Cuts Inference Costs
Frameworks LLM Agents Factory: An Agent Factory That Cuts Inference Costs 2026-08-12
More robust random neural networks: intuitionistic fuzzy takes on noisy data
Frameworks More robust random neural networks: intuitionistic fuzzy takes on noisy data 2026-08-12
How topology reveals the inner evolution of Transformers
LLM How topology reveals the inner evolution of Transformers 2026-08-12
Unsloth Desktop brings LLM training local: 2× faster, 70% less VRAM
Altro Unsloth Desktop brings LLM training local: 2× faster, 70% less VRAM 2026-08-11
Luth-2: French small language models beat 3x larger competitors, redefining local AI
LLM Luth-2: French small language models beat 3x larger competitors, redefining local AI 2026-08-11
LLMs and Waste Management: WuYuEval Reveals the Limits of Generalist AI
LLM LLMs and Waste Management: WuYuEval Reveals the Limits of Generalist AI 2026-08-11
Self-adaptive fuzzing exposes the hallucination cracks in multimodal LLMs
Frameworks Self-adaptive fuzzing exposes the hallucination cracks in multimodal LLMs 2026-08-11
Training a 1B LLM from scratch for under $200: the frontier of accessible self-hosting
LLM Training a 1B LLM from scratch for under $200: the frontier of accessible self-hosting 2026-08-11
Linux: Open-Source NVIDIA “Nova” Driver Gains Functionality with Rust
Hardware Linux: Open-Source NVIDIA “Nova” Driver Gains Functionality with Rust 2026-08-11
Ling-3.0-tiny: 8B parameters, 1.3B active, hitting 100 tokens/sec on MacBook
Altro Ling-3.0-tiny: 8B parameters, 1.3B active, hitting 100 tokens/sec on MacBook 2026-08-10
Muse-Glimmer-30B in GGUF: The Latest Piece of an Increasingly Mature Local Ecosystem
LLM Muse-Glimmer-30B in GGUF: The Latest Piece of an Increasingly Mature Local Ecosystem 2026-08-10
Meta Muse Glimmer: Local AI Agents on Consumer GPUs, 30B Parameter Model Released
LLM Meta Muse Glimmer: Local AI Agents on Consumer GPUs, 30B Parameter Model Released 2026-08-10
Meta's Muse Glimmer: A 30B Open-Weight Model Purpose-Built for Always-On Local Agents
LLM Meta's Muse Glimmer: A 30B Open-Weight Model Purpose-Built for Always-On Local Agents 2026-08-10
No cloud, just edge: Edgify raises $9M for AI that learns on devices
Altro No cloud, just edge: Edgify raises $9M for AI that learns on devices 2026-08-10
TEXAS Leverages Native MoE Routing for More Surgical Fine-Tuning
LLM TEXAS Leverages Native MoE Routing for More Surgical Fine-Tuning 2026-08-10
Truth is a vector: detecting fake news without leaving the model
LLM Truth is a vector: detecting fake news without leaving the model 2026-08-10
WeatherNext 2: DeepMind brings cyclone forecasting to a single H100 GPU
Hardware WeatherNext 2: DeepMind brings cyclone forecasting to a single H100 GPU 2026-08-09
BDH: Pathway's post-transformer architecture matches GPT-2 scaling on ordinary GPUs
LLM BDH: Pathway's post-transformer architecture matches GPT-2 scaling on ordinary GPUs 2026-08-09
Kimi K3 slims to 478GB: the multilingual trim that shifts on-prem math
LLM Kimi K3 slims to 478GB: the multilingual trim that shifts on-prem math 2026-08-09
PRISM2: the AI that reads slides through clinical dialogue. But the pipeline remains closed
LLM PRISM2: the AI that reads slides through clinical dialogue. But the pipeline remains closed 2026-08-07
AMD Unveils Spur: Rust-Native Job Scheduler for ROCm GPU Clusters
Altro AMD Unveils Spur: Rust-Native Job Scheduler for ROCm GPU Clusters 2026-08-07
The Chain-of-Thought Reasoning of LLMs Becomes Predictable with an Equation
LLM The Chain-of-Thought Reasoning of LLMs Becomes Predictable with an Equation 2026-08-07
CRAFTER: the agent that doubles improvements without touching the model
Frameworks CRAFTER: the agent that doubles improvements without touching the model 2026-08-07
MS-MLB: An open benchmark for multiple sclerosis that puts reproducibility to the test
Frameworks MS-MLB: An open benchmark for multiple sclerosis that puts reproducibility to the test 2026-08-07
NVIDIA brings its entire speech stack local: ASR, TTS and codec now run on-device
Altro NVIDIA brings its entire speech stack local: ASR, TTS and codec now run on-device 2026-08-06
Scotoma-2: Taming Gemma 4's Stylistic Tics While Preserving Its Smarts
LLM Scotoma-2: Taming Gemma 4's Stylistic Tics While Preserving Its Smarts 2026-08-06
Nscale and Pure Data Centres lead the AI infrastructure funding surge as UK capital concentrates in 10 companies
Market Nscale and Pure Data Centres lead the AI infrastructure funding surge as UK capital concentrates in 10 companies 2026-08-06
Modal Labs opens London office: AI inference moves closer to Europe
Market Modal Labs opens London office: AI inference moves closer to Europe 2026-08-06
DIY AI on RTX 5090: Local Training Becomes a Research Lab for Enthusiasts
Hardware DIY AI on RTX 5090: Local Training Becomes a Research Lab for Enthusiasts 2026-08-06
LLMs for Latin: Transfer Learning Wins Big, but Sovereignty Loses Out
LLM LLMs for Latin: Transfer Learning Wins Big, but Sovereignty Loses Out 2026-08-06
Denial WM: A Rust-Powered Wayland Compositor with Embedded Flutter
Frameworks Denial WM: A Rust-Powered Wayland Compositor with Embedded Flutter 2026-08-05
Frore claims LiquidJet can cool Nvidia Rubin GPUs by 10°C, boosting performance 15%
Hardware Frore claims LiquidJet can cool Nvidia Rubin GPUs by 10°C, boosting performance 15% 2026-08-05
Maple-Preview: 20B ternary-weight open-weight reasoning LLM
LLM Maple-Preview: 20B ternary-weight open-weight reasoning LLM 2026-08-05
BBOWP-Bench: Testing LLMs on Black-Box Optimization Problem Formulation
LLM BBOWP-Bench: Testing LLMs on Black-Box Optimization Problem Formulation 2026-08-05
Self-repairing digital circuits: biology's lesson for on-premise AI hardware
Hardware Self-repairing digital circuits: biology's lesson for on-premise AI hardware 2026-08-05
Nvidia RTX 50 prices climb as TSMC and memory costs surge
Market Nvidia RTX 50 prices climb as TSMC and memory costs surge 2026-08-04
MemoryForge gives LLMs autobiographical memory: when persona comes from data, not prompts
LLM MemoryForge gives LLMs autobiographical memory: when persona comes from data, not prompts 2026-08-04
Linux 7.3 improves GPU reset recovery for AMD Kaveri and Hawaii: on-prem stability benefits
Hardware Linux 7.3 improves GPU reset recovery for AMD Kaveri and Hawaii: on-prem stability benefits 2026-08-04
RTX 5090 surpasses $5,100: AI hardware costs rise, on-prem deployments tremble
Hardware RTX 5090 surpasses $5,100: AI hardware costs rise, on-prem deployments tremble 2026-08-03
Orchard: the open framework where infrastructure makes the difference (and 3B models approach giants)
Frameworks Orchard: the open framework where infrastructure makes the difference (and 3B models approach giants) 2026-08-03
zlib-rs 0.6.7 brings Rust safety and LoongArch support to compression
Altro zlib-rs 0.6.7 brings Rust safety and LoongArch support to compression 2026-08-03
Imbalanced Data? GMM and LLMs Team Up for Better Clustering
LLM Imbalanced Data? GMM and LLMs Team Up for Better Clustering 2026-08-03
Autonomous driving: temporal jitter is the real enemy of classifiers
Altro Autonomous driving: temporal jitter is the real enemy of classifiers 2026-08-03
OpenClaw and Ollama: A Layered Architecture for Autonomous and Scalable AI Agents
Frameworks OpenClaw and Ollama: A Layered Architecture for Autonomous and Scalable AI Agents 2026-08-03
Ready for the 26-trillion-parameter LLM? The alternative to wasting money on GPUs is SSDs
Hardware Ready for the 26-trillion-parameter LLM? The alternative to wasting money on GPUs is SSDs 2026-08-02
llama.cpp embraces MTP/DSpark: on-prem inference speeds up for DeepSeek V4 Flash
Frameworks llama.cpp embraces MTP/DSpark: on-prem inference speeds up for DeepSeek V4 Flash 2026-08-02
DeepSeek-V4-Flash: Beware mid-conversation system messages that kill prompt caching
Frameworks DeepSeek-V4-Flash: Beware mid-conversation system messages that kill prompt caching 2026-08-02
KDE Plasma 6.8 multi-GPU compositing boost is a silent win for local AI
Hardware KDE Plasma 6.8 multi-GPU compositing boost is a silent win for local AI 2026-08-01
Unsloth brings Deepseek V4 to local setups with new GGUF files
Frameworks Unsloth brings Deepseek V4 to local setups with new GGUF files 2026-07-31