RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

1250 items in r/LocalLLaMA

10.08.2026
r/LocalLLaMA inclusionAI/Ling-3.0-tiny Β· 8B A1.3B MoEΒ· Hugging Face πŸ”—
r/LocalLLaMA DiffusionGemma Technical Report πŸ”—
r/LocalLLaMA DeepSeek V4 Flash 0731 is the β€˜killer app’ that is going to sell A LOT of DGX Sparks πŸ”—
r/LocalLLaMA Early signs that Muse-Glimmer-30B might quantize *very* well? Share your experiences. πŸ”—
r/LocalLLaMA Best Local LLMs - August 2026 πŸ”—
r/LocalLLaMA Muse Glimmer ACTUALLY fits on a single RTX 3090 πŸ”—
r/LocalLLaMA Motif-Technologies/Motif-3 official realese πŸ”—
r/LocalLLaMA Glimmer seems pretty censored? πŸ”—
r/LocalLLaMA Mark Zuckerberg on releases πŸ”—
r/LocalLLaMA model: Muse Glimmer Support by pcuenca Β· Pull Request #26841 Β· ggml-org/llama.cpp πŸ”—
r/LocalLLaMA 1M context with 17 GB model in 24 GB VRAM: "for the first time I was able to load a context of almost 1M tokens and extract 7 needles from various parts of the text" πŸ”—
r/LocalLLaMA unsloth/Muse-Glimmer-30B-GGUF Β· Hugging Face πŸ”—
r/LocalLLaMA Comparing how Cline, Kilo, and Qwen Code handle long-task context/state (and why context loops keep happening) πŸ”—
r/LocalLLaMA Introducing Muse Glimmer: an open-weight model optimized for always-on local agent workflows πŸ”—
r/LocalLLaMA Meta open sources new on-device model Muse Glimmer & Muse spark 1.2 also coming soon! πŸ”—
r/LocalLLaMA meta-models/Muse-Glimmer-30B πŸ”—
r/LocalLLaMA Why Speculative Decoding went mature in 2026? πŸ”—
r/LocalLLaMA Running Qwen 3.5 35B A3B-Q8_0 gguf on a cheap radeon 7600 at 18 token/s πŸ”—
r/LocalLLaMA omlab/VLX-Seek-1.5-10B Β· Hugging Face πŸ”—
r/LocalLLaMA So... did we give up on the rule against AI posts? πŸ”—
r/LocalLLaMA ByteDance vows to avoid AI distillation, develop new model its own way πŸ”—
r/LocalLLaMA KPMG Says Nearly Half Of Executives Pulled Back AI Agents Over Cost πŸ”—
09.08.2026
r/LocalLLaMA [NEW MODEL] SupraElegans-500K πŸ”—
r/LocalLLaMA KLQ: Training-free measured rotation quantization. Beats all training-free rotation-based quantization methods on W4A4KV4-bits. Llama 3.2 1B KLQ-quantized beats SpinQuant and gets close to ReSpinQuant without GPTQ/LDLQ rounding. πŸ”—
r/LocalLLaMA The Gemma team will host a special event on August 20 πŸ”—