RSS Feed
All
r/AI_Agents
r/AI_Governance
r/ClaudeAI
r/ClaudeCode
r/DeepSeek
r/Futurology
r/LangChain
r/LocalLLaMA
r/MachineLearning
r/OpenAI
r/artificial
r/cybersecurity
r/europe
r/hermesagent
r/netsec
r/singularity
r/technology
1252 items in r/LocalLLaMA
30.07.2026
29.07.2026
r/LocalLLaMA
Bought a 5090 to escape API fees. Ended up building a mini datacenter. Sound familiar?
π
r/LocalLLaMA
Are you guys not scared of where we're heading? A year ago, GPT-5 was considered one of the best models in the world. Today, we have open-weight models like Qwen3.6-27B that are competitive enough to run locally on high-end consumer hardware. The pace of progress is absolutely brutal.
π
r/LocalLLaMA
PSA: llama.cpp now loads MTP tensors by default for any draft-mtp arch, even with MTP disabled
π
r/LocalLLaMA
I keep coming back to Qwen... Over and Over. Is there really nothing better under 120B?
π
r/LocalLLaMA
The idea: on a CPU the decode speed depends on the active params per token, not the total. My objective is trying to run a 10B at 100tok/s on a mid level PC (No GPU).
π
r/LocalLLaMA
Understand Kimi K3 from first principles: a recommended order for anyone trying to understand this beast
π
r/LocalLLaMA
A slide deck you can edit with a local model or in Chrome β the whole deck is a JSON block in one HTML file (~640KB with editor and viewer included)
π