RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

1205 items in r/LocalLLaMA

15.09.2026
r/LocalLLaMA ByteShape Qwen 3.8 27B: To KL Diverge or Not to KL Diverge, Part 2: Metric Boogaloo ๐Ÿ”—
r/LocalLLaMA Occamy-1.0 by Accio Lab ๐Ÿ”—
r/LocalLLaMA It's Tuesday already with no new model drops ๐Ÿ”—
r/LocalLLaMA Closed source AI is more dangerous than open source AI. ๐Ÿ”—
r/LocalLLaMA DeepSeek V4.1F Q4 on M3 Ultra with native DSpark MTP (40tps / 800tps) ๐Ÿ”—
r/LocalLLaMA Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NEO-CODER-MAX-MTP-GGUF ๐Ÿ”—
r/LocalLLaMA CrofAI "cheapest inference provider in the world" gets exposed as an OpenRouter wrapper, routing requests to smaller, cheaper models at up to 20x markup. CrofAI responds to Wire Fraud allegations by denying everything, then backtracking, then 3 hours later wiping their entire online presence ๐Ÿ”—
r/LocalLLaMA Voodoo Dynamic Quant - Now MIT Licensed ๐Ÿ”—
r/LocalLLaMA If you have a 3090, or other 30xx for local LLMs, I have something for you ๐Ÿ”—
14.09.2026
r/LocalLLaMA I think Muse Glimmer is slept on ๐Ÿ”—
r/LocalLLaMA DeepSeek engineer relections on RSI - burying my talent to yesterday ๐Ÿ”—
r/LocalLLaMA Running Qwen3.8-Flash-Next locally on a 12GB VRAM card ๐Ÿ”—
r/LocalLLaMA Animated transition from AA Intelligence Index v4.1 to v4.3 ๐Ÿ”—
r/LocalLLaMA Nvidia's RTX 5090 vanishes from online retail in the US โ€” third-party sellers now demand as much as $9,500 for Nvidia's fastest GPU ๐Ÿ”—
r/LocalLLaMA NVIDIA Unveils RTX PRO 5500 "Blackwell" Workstation GPU with 84 GB GDDR7 Memory ๐Ÿ”—
r/LocalLLaMA [Bi-Weekly Megathread] Project Showcase ๐Ÿ”—
r/LocalLLaMA Our new battle flag! ๐Ÿ”—
r/LocalLLaMA China Wants to Build a BRICS-Wide AI Cloud. That Might Be a Much Bigger Deal Than the Open Models ๐Ÿ”—
r/LocalLLaMA For the GPU poor. K2 Horizon 7B ranks between qwen 3.6 27B and qwen 3.6 35BA3b on the Artificial Analysis Intelligence Index. ๐Ÿ”—
r/LocalLLaMA UkisAI Swift-Qwen3.8-27B / -58.3% thinking, x1.95 speed while keeping the accuracy of xhigh ๐Ÿ”—
r/LocalLLaMA Xi promotes open source AI zone among BRICS countries ๐Ÿ”—
r/LocalLLaMA K2 Horizon lineup is out on AA, and once again AA plots are misleading. ๐Ÿ”—
r/LocalLLaMA llama: add Maple 20B-A1B ternary MoE architecture (CPU) by AlexGabbia ยท Pull Request #27000 ยท ggml-org/llama.cpp ๐Ÿ”—
r/LocalLLaMA What actually makes you trust a local coding agent enough to leave it running unattended? ๐Ÿ”—
r/LocalLLaMA The new k2 horizon models seem like an absolute beast ๐Ÿ”—