RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

1218 items in r/LocalLLaMA

20.08.2026
r/LocalLLaMA The boring way to run Deepseek V4 Flash-0731 130-150 tks - 16x5060ti 16GB over 2 PLX88096 switches πŸ”—
r/LocalLLaMA Aurora-80K releases! A modern tiny language model. πŸ”—
r/LocalLLaMA Tencent begins testing its new flagship model Hunyuan Hy4 πŸ”—
r/LocalLLaMA I just built a mini Kimi-K3 from Scratch under 250$. Already beats GPT-2 (124M)! πŸ”—
r/LocalLLaMA G9v3-39A5B on artificialanalysis looks good. Has anyone tested it? πŸ”—
r/LocalLLaMA New benchmark just dropped! πŸ”—
r/LocalLLaMA TinySearch v0.6.1 - still a lightweight web research tool for local LLMs, now with bring-your-own-browser support πŸ”—
r/LocalLLaMA The "local frontier" is now smarter than Sonnet 4.5 πŸ”—
r/LocalLLaMA Qwen 3.8 27B KV f16 vs q8_0 are not equivalents πŸ”—
r/LocalLLaMA Spider-man: Brand New Day, does Peter self host his AI? (Spoilers) πŸ”—
r/LocalLLaMA Qwen3.8-27B took a serious hit to *knowledge* vs 3.6 πŸ”—
r/LocalLLaMA Qwen3.8-27b has the highest level of "agency" I've ever seen in a local model πŸ”—
19.08.2026
r/LocalLLaMA Qwen3.8-23B-Mini-Me: A Depth-Pruned Qwen3.8-27B (to ~22.7BB) πŸ”—
r/LocalLLaMA I pushed Qwen3.8-27B limits again... Dflash2 - 134 tps on a RTX 3090 πŸ”—
r/LocalLLaMA DFlash2 speeds Qwen 3.8 27B up to 4 times πŸ”—
r/LocalLLaMA Introducing Qwen3.8-27B Dynamic v3 Unsloth GGUFs πŸ”—
r/LocalLLaMA AntLing’ve open-sourced 6 Base Model checkpoints for Ling-3.0-tiny & Ling-3.0-flash, covering pre-trained, mid-trained, and WSM-merged stages. πŸ”—
r/LocalLLaMA NVFP4 on VOLTA! Despite being built for Blackwell, I made four 2017 V100s run Qwen 3.8 NVFP4 natively and match my $6000 RTX 5090. πŸ”—
r/LocalLLaMA Ornith-1.5 (397B [DeepSWE 56], 35B-A3B, 9B) πŸ”—
r/LocalLLaMA We have Q3.8 35B at home: 3x new Ornith 1.5 released πŸ”—
r/LocalLLaMA updated unsloth/Qwen3.8-27B-GGUF Β· Hugging Face πŸ”—
r/LocalLLaMA Qwen 3.8 35B-A3B vagueposting πŸ”—
r/LocalLLaMA Finally found a really solid suno-like minimax music UI!! πŸ”—
r/LocalLLaMA Stop Anthropomorphisizing Intermediate Tokens: Qwen3.8 doesn't "overthink" πŸ”—
r/LocalLLaMA Am I doing something wrong? Qwen 3.8 27B seems useless for agentic coding πŸ”—