RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

1216 items in r/LocalLLaMA

22.08.2026
r/LocalLLaMA GLM and I created a llama.cpp fork optimized for AMD GFX906 (Mi50, Mi60, Radeon VII, GCN HIP) - Machine Learning, LLMs, & AI ๐Ÿ”—
r/LocalLLaMA Single RTX 5090: Qwen3.8-27B NVFP4 at a real 262K context in vLLM โ€” 77 tok/s short-context, 64.7 tok/s at 128K ๐Ÿ”—
r/LocalLLaMA I forked Ninfer 3090 and converted it to run on the CMP170HX - doubled my Qwen3.6-35B from llama.cpp ๐Ÿ”—
r/LocalLLaMA Fixed the MTP head on Ornith1.5 35B A3B. +3% TPS -33% wall clock ๐Ÿ”—
r/LocalLLaMA How to remove trendy speech from llms? ๐Ÿ”—
r/LocalLLaMA Your own GGUF ๐Ÿ”—
r/LocalLLaMA This is a great sub, regardless of what complaints people have about it. ๐Ÿ”—
r/LocalLLaMA Think you're going to get cheap DDR5 RAM? Think again, even if prices fall, scalper bots now outnumber shoppers 10 to 1 and will keep prices high ๐Ÿ”—
r/LocalLLaMA Artificial Analysis "Intelligence": A meaningless benchmark ๐Ÿ”—
r/LocalLLaMA Freetokens project is impressive ๐Ÿ”—
r/LocalLLaMA Llama.cpp version 0.2.0 is out! ๐Ÿ”—
r/LocalLLaMA Bro wtf, Qwen Lab cooked with Qwen 3.8 27B, it's so fucking good ๐Ÿ”—
r/LocalLLaMA This is why I run locally. ๐Ÿ”—
r/LocalLLaMA 16 GB VRAM purgatory discussion thread ๐Ÿ”—
r/LocalLLaMA Qwen 3.8 27b - PI AGENT vs OPENCODE - another smaple ๐Ÿ”—
21.08.2026
r/LocalLLaMA I tried to do agenic coding with Qwen 3.8 27B 3bit quant on a macbook air m2 24gb. It took 63 hours, but amazingly, the flight simulator worked. ๐Ÿ”—
r/LocalLLaMA Qwen3.8-27B different thinking levels ๐Ÿ”—
r/LocalLLaMA Qwen 3.8 Low and Medium are goated ๐Ÿ”—
r/LocalLLaMA Qwen3.8-27B Q6 is a beast at agentic coding ๐Ÿ”—
r/LocalLLaMA BEAST!! Unlocked a CMP 170HX: 6.3 โ†’ 193 TFLOPS tensor, pp512 599 โ†’ 3468, and half the community's advice about this card is wrong. ๐Ÿ”—
r/LocalLLaMA FireRedAudio & FireRedTTS3 by FireRedTeam - Huggingface ๐Ÿ”—
r/LocalLLaMA I benchmarked Ox Alpha on SWE-bench Verified Mini (50 tasks): 96% resolved. Now Iโ€™m skeptical of myself. ๐Ÿ”—
r/LocalLLaMA NVIDIA AVO got 100% on ARC-AGI-3. It completed all 183 levels across all 25 public environments, figuring out what to do with no instructions, explicit rules, or stated goals. ๐Ÿ”—
r/LocalLLaMA DeepSeek Harness v0.1.1 released ๐Ÿ”—
r/LocalLLaMA Qwen 3.8 27b is strong even at Q3_xxs ๐Ÿ”—