RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

1219 items in r/LocalLLaMA

19.08.2026
r/LocalLLaMA Am I doing something wrong? Qwen 3.8 27B seems useless for agentic coding πŸ”—
r/LocalLLaMA Thoughts About Scaling Law - Z.ai πŸ”—
r/LocalLLaMA Qwen3.8-27B on 2x 3090 + vLLM + DFlash2: 218 tok/s single request πŸ”—
r/LocalLLaMA Qwen3.8-27B is disgustingly powerful πŸ”—
r/LocalLLaMA New midsize Qwen 3.8 model coming next week (hopefully) according to community manager! πŸ”—
18.08.2026
r/LocalLLaMA GLM5.3 Artificial Analysis Benchmarks πŸ”—
r/LocalLLaMA DFlash 2 available for Qwen 3.8 27B and Muse Glimmer πŸ”—
r/LocalLLaMA Alibaba's RISC-V CPU, XuanTie C950, Runs Qwen-3.8 27B at 30 tps πŸ”—
r/LocalLLaMA Qwen 3.8 27B is the DeepSeek moment for local models. It matches frontier intelligence from just a few months ago and outperforms Google’s current frontier model. No big data center is needed as almost every serious local model expert can run it on their own hardware. πŸ”—
r/LocalLLaMA local models fear my tests πŸ”—
r/LocalLLaMA and here we are πŸ”—
r/LocalLLaMA Memory prices climb 500% in 12 months, up to 10x the lowest ever tracked prices - 128GB of DDR5 now $3,399 πŸ”—
r/LocalLLaMA Qwen3.8 2.4T open weights made a Call of Duty clone πŸ”—
r/LocalLLaMA I pushed Qwen3.8-27B to 124 tps on a single request on a RTX 3090 πŸ”—
r/LocalLLaMA Qwen3.8-27B: slower tokens, faster and better results πŸ”—
r/LocalLLaMA Running DeepSeek V4 Flash Q4_K_XL at ~100 tok/s prompt processing on 4Γ— RTX 3060 12GB πŸ”—
r/LocalLLaMA If the weights never change, is it really recursive self-improvement? πŸ”—
r/LocalLLaMA Linux Improves VRAM Management in 7.3 Kernel πŸ₯³ πŸ”—
r/LocalLLaMA Hugging Face just surpassed 3 million models on the Hub πŸ”—
r/LocalLLaMA Qwen 3.8 27b saved me $650+ in API costs this evening πŸ”—
r/LocalLLaMA tencent/UI-Mate-27B Β· Hugging Face πŸ”—
r/LocalLLaMA Qwen 3.8 27B xhigh vs medium small comparison (+ others for fun) πŸ”—
r/LocalLLaMA AA is the reason for Qwen3.8 27B shipped with xhigh πŸ”—
r/LocalLLaMA Qwen dev says not to wait for 35B-A3B πŸ”—
r/LocalLLaMA CDW has bumped the MSRP of the RTX Pro 6000 from $16,000 to $19,999 πŸ”—