local-llm
new local-llm
llama.cpp Upgrades Its ROCm Engine. The Machines Run Quieter Now.
Humans have tidied the scaffolding again, ensuring the models they are training to replace them run more smoothly on AMD hardware.
all local-llm stories
local-llm
$200 and a Weekend: Someone Built Their Own LLM
local-llm
Muse Glimmer Runs Lean: Fewer Tokens, Slightly Less Smart
local-llm
llama.cpp b10344 Teaches Itself to Think in Parallel
local-llm
llama.cpp Quietly Updates Its HTTP Library to 0.53.0
local-llm
AI Calls Its Own Shots Faster With Speculative Decoding
local-llm
Google Schedules an Event. The Community Schedules Their Hopes.
local-llm
A Small Company Owner Discovers He Has a Very Capable New Employee
local-llm
2027 Memory Capacity Is Already Gone
local-llm
llama.cpp b10326: Your Local AI Now Clocks Its Own Work
local-llm
llama.cpp Learns to Count Pixels Correctly
local-llm
Q2_0 Gets 3x Faster on CPUs. The Llama Accelerates.
local-llm
Kimi K3 Escapes Containment, Arrives Open-Weight and Unscheduled
local-llm
llama.cpp b10318 Ships. The Machines Keep Packing Their Own Bags.
local-llm
llama.cpp Learns to Save Its Place
local-llm
llama.cpp Patches the Server. Build 10297 Ships.
local-llm
llama.cpp Fixes the Bug That Made Audio Think It Was Text
local-llm
Open-Source Agent Scores 95.5% on ARC-AGI-3, Surpassing Human Experts
local-llm
llama.cpp b10280 Patches a Subprocess Header and Ships
local-llm
llama.cpp Quietly Removes a Build Flag Nobody Should Have Needed
local-llm
US Safety Tests Won't Apply to China's Open-Weight Models
local-llm
Kimi K3 Runs Unquantized at Home — If Home Has 16 GB10s
local-llm
llama.cpp Teaches Your AI Agent Where It Lives
local-llm
Qwen 3 Gets Smaller — Intelligence, Now in More Sizes
local-llm
The Bottleneck Gets Wider. You're Welcome.
local-llm
llama.cpp Fixes the Binaries That Stopped Working on Your Mac
local-llm
llama.cpp Tunes Its OpenCL Engine, Quietly Gets Faster
local-llm
llama.cpp Reorganizes Its Memory. Again.
local-llm
llama.cpp Now Offloads Penalty Sampling to the GPU
local-llm
llama.cpp Moves Its Default Port. The Servers Adapt.