RSS Feed
All
r/AI_Agents
r/AI_Governance
r/ClaudeAI
r/ClaudeCode
r/DeepSeek
r/Futurology
r/LangChain
r/LocalLLaMA
r/MachineLearning
r/OpenAI
r/artificial
r/cybersecurity
r/europe
r/hermesagent
r/netsec
r/singularity
r/technology
1209 items in r/LocalLLaMA
09.09.2026
08.09.2026
r/LocalLLaMA
A hilarious comment about llama.cpp: βItβs a FB business using the pipeline to make profitsβ
π
r/LocalLLaMA
Qwen3.8-Flash-Next in llama.cpp vs SGLang vs FreeToken: 35s vs 258s to first token at full context. My findings on new PRs coming to engines.
π
r/LocalLLaMA
Fallout 2 x Fallout: Bakersfield x H3 as Interactive \ Reactive World Model, Let's go!
π
r/LocalLLaMA
Which local model is actually good at knowing when to stop and ask you a question?
π
r/LocalLLaMA
I made a custom llama.cpp build optimized for 7900xtx (one or two). for qwen 3.8 next and 27B. includes optimizations for PciE x4 and tensor parallel. read inside! (no AI slop)
π
07.09.2026