RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

1216 items in r/LocalLLaMA

23.08.2026
r/LocalLLaMA Any upcoming models to be excited about? πŸ”—
r/LocalLLaMA you can now use MTP in GLM-Air πŸ”—
r/LocalLLaMA Qwen 3.8 27b helped me with something unique that Opus 4 couldn't - Firmware + Software preservation and emulation on an early 2000's ARM based POS system πŸ”—
r/LocalLLaMA We quantized Qwen 3.8 27B and compared the quants on an RTX 6000 πŸ”—
r/LocalLLaMA Nvidia Customers Notified About AI-Related Price Hikes Above 15% πŸ”—
r/LocalLLaMA New qwen3.8:27b on a 39k line C to single-file HTML / three.js port πŸ”—
r/LocalLLaMA 1/100 β†’ 44/100: fine-tuning a 450M VLM on 50K browser screenshots πŸ”—
r/LocalLLaMA Qwen 3.8 27B for actual local programming πŸ”—
r/LocalLLaMA Qwen3.5-9B Triple-Loop πŸ”—
r/LocalLLaMA GMKtec is going to launch new hardware with Ryzen AI Max+ PRO 495 at IFA Berlin 2026 πŸ”—
r/LocalLLaMA Don't want to be this guy, but I need Qwen 3.8 35B A3B πŸ”—
r/LocalLLaMA I hosted Kimi K3 (2.8T parameters) using 8 B300s. 92 tok/s, $190 per million tokens πŸ”—
r/LocalLLaMA i finally switched from windows to linux and got a 30-50% boost in speed. πŸ”—
r/LocalLLaMA DeepSeek Harness is Insanely Good πŸ”—
r/LocalLLaMA Nvidia Poolside deal to compete with Chinese Open Weights πŸ”—
r/LocalLLaMA Qwen 3.8 27B is a game changer. πŸ”—
r/LocalLLaMA β€œThe All Spark” Cluster: Upgrading from 16 - 36 DGX Sparks πŸ”—
r/LocalLLaMA # Qwen3.8-27B β€” One Week Later: The r/LocalLLaMA + r/LocalLLM Verdict πŸ”—
r/LocalLLaMA I fine tuned Gemma 4 12B for a 2.7x improvement on tool calling because I can't fit anything else comfortably into my 16 GBs of Vram πŸ”—
r/LocalLLaMA Closed AI has been real quiet since Qwen 3.8 27B dropped. πŸ”—
r/LocalLLaMA Has anyone actually made 64k feel like 300k+ with recursive local agents? πŸ”—
r/LocalLLaMA Tested in Coding: Q8_K_XL Qwen3.8 27B vs BF16 Qwen3.6 27B πŸ”—
r/LocalLLaMA Best harness for long autonomous tasks πŸ”—
22.08.2026
r/LocalLLaMA I benchmark DFlash 2 (PR build) in llama.cpp on Qwen 3.8 27B against all speculative methods for 3 days. 2.26x on 100 real coding prompts, 4.68x with one n-gram drafter on top. Up to 8x on specific cases. πŸ”—
r/LocalLLaMA New 100B Liquid AI model coming soon πŸ”—