RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

1251 items in r/LocalLLaMA

09.08.2026
r/LocalLLaMA The Gemma team will host a special event on August 20 ๐Ÿ”—
r/LocalLLaMA Speculative decoding in a tools call ๐Ÿ”—
r/LocalLLaMA Open Model: Google Weather Next 2 ๐Ÿ”—
r/LocalLLaMA Pathway's BDH(post-transformer arch) matches GPT2 scaling from 10M to 1B params trained from scratch. runs on Normal GPUs ๐Ÿ”—
r/LocalLLaMA Two flags took the official Ling-3.0-flash INT4 from 20.8 to 38.7 tok/s on one DGX Spark ๐Ÿ”—
r/LocalLLaMA Lophius: A workbench for language model research, from the creator of Heretic ๐Ÿ”—
r/LocalLLaMA DeepSeek v4 Flash 0731 locally on CPU ๐Ÿ”—
r/LocalLLaMA Underestimated budget solution: radeon 780m iGPU ๐Ÿ”—
r/LocalLLaMA Tencent announce WorldClaw ๐Ÿ”—
r/LocalLLaMA What a deal. Thanks newegg ๐Ÿ”—
r/LocalLLaMA AMD llama.cpp: reducing MTP buffer overhead gave me 64K โ†’ 149K context for Qwen 27B ๐Ÿ”—
r/LocalLLaMA 300b on 32gb MoE-streaming findings + optimisations ๐Ÿ”—
r/LocalLLaMA DeepSeek V4 Flash 0731 hits 82.7% on Terminal-Bench 2.1 in an independent public-harness run (445 trials) ๐Ÿ”—
r/LocalLLaMA Best Embedding + Reranking Model ๐Ÿ”—
r/LocalLLaMA Updated benchmark: Deepseek V4 Flash on SlopCodeBench (local) ๐Ÿ”—
r/LocalLLaMA LFM 2.6B is a lot of fun. ๐Ÿ”—
r/LocalLLaMA ds4 flash 0731 UD-IQ2_M wrote a custom metal kernal for kimi k2 IQ1_0 in about 50 minutes ๐Ÿ”—
r/LocalLLaMA RTX 5090 96GB spotted on Alibaba? ๐Ÿ”—
r/LocalLLaMA No wonder Qwen and Gemma are so different ๐Ÿ”—
08.08.2026
r/LocalLLaMA Kimi K3 (Unsloth) IQ2-XXS from 711GB down to 478GB!!! Only Multi-language was removed to trim the size ๐Ÿ”—
r/LocalLLaMA Is Microsoft-Phi dead? ๐Ÿ”—
r/LocalLLaMA enabling PCI-E p2p for consumer Nvidia cards will yield you more than you think ๐Ÿ”—
r/LocalLLaMA Showoff Saturday: Local 4x 6000 Pro (multi-year progression) ๐Ÿ”—
r/LocalLLaMA 2027 Memory Capacity Is Reportedly Sold Out ๐Ÿ”—
02.08.2026
r/LocalLLaMA DeepSeek-V4-Flash-0731: surpasses Fable-5, Sol & Kimi-K3 on Chess Benchmark ๐Ÿ”—