RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

1249 items in r/LocalLLaMA

12.08.2026
r/LocalLLaMA Benched a 124B on one DGX Spark for a week and published all of it β€” 38.7 tok/s on the fastest path he found, 2.4x DeepSeek V4 Flash on the same box πŸ”—
r/LocalLLaMA DeepSeek V4-Pro-0813 Benchmarks πŸ”—
r/LocalLLaMA DeepSeek-V4-Pro-0813 is UP! πŸ”—
r/LocalLLaMA Qwen3.8-Max πŸ”—
r/LocalLLaMA Qwen3.8-2.4T-A95B Released πŸ”—
r/LocalLLaMA Qwen/Qwen3.8-2.4T-A95B Β· Released! πŸ”—
r/LocalLLaMA LiquidAI/LFM2.5-VL-3B Β· Hugging Face πŸ”—
r/LocalLLaMA Exact Qwen 3.8 27b release date and time πŸ”—
r/LocalLLaMA NVIDIA's Fastest Blackwell GPU, the 96 GB RTX PRO 6000, Now Costs $16,000, Almost Double Its Original Price πŸ”—
r/LocalLLaMA All your reasoning are belong to us πŸ”—
r/LocalLLaMA Hidden Reasoning from Claude and GPT are Decoded, and it is interesting πŸ”—
r/LocalLLaMA According to AMD, Arm, and Microsoft, agentic AI could push CPU-to-GPU ratios from 1:4 to even1:1 πŸ”—
r/LocalLLaMA It's Qwenesday my dudes πŸ”—
r/LocalLLaMA RAG for regular users? πŸ”—
r/LocalLLaMA It's the final countdown, baby! Qwen is out in just over 7 hours! πŸ”—
r/LocalLLaMA FYI: Muse Glimmer Chat Template Got Updated Recently πŸ”—
r/LocalLLaMA RTX 6000 PRO price raised to $16,000 USD on the Nvidia website πŸ”—
r/LocalLLaMA New Muse-Glimmer-30B SoTA Quants - hopefully a new lineup :) πŸ”—
r/LocalLLaMA Watching labs post benchmarks comparing their models to 3.6 27b while knowing what’s about to go down on Qwednesday πŸ”—
r/LocalLLaMA Anthropic, OpenAI, Google, Meta, Microsoft, and Mistral all signed the EU Code of Practice on Transparency of AI-Generated Content πŸ”—
11.08.2026
r/LocalLLaMA We quantized DeepSeek V4 0731 and benchmarked it against popular quants on 8Γ— RTX 5090 πŸ”—
r/LocalLLaMA 366 t/s Qwen3.6 27B NVFP4 on v100s πŸ”—
r/LocalLLaMA Local Benchmark : Muse Glimmer 30B vs Qwen 3.6 27B vs Gemma4 31B (and many other models and finetunes) πŸ”—
r/LocalLLaMA I will be parting with my 4x Spark Cluster. πŸ”—
r/LocalLLaMA All the more reason not to use Closed Models ... Claude now officially "marks" AI-generated content ... steganographically, apparently ... and there are false positives already πŸ”—