RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

1249 items in r/LocalLLaMA

11.08.2026
r/LocalLLaMA Ling-3.0-flash quant ladder on one DGX Spark: the whole thing sits in a 32 to 40 tok/s band ๐Ÿ”—
r/LocalLLaMA nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16 ๐Ÿ”—
r/LocalLLaMA Encrypted reasoning from ClosedAI et al 100% recoverable ๐Ÿ”—
r/LocalLLaMA I built a weird, low-power llama.cpp server using an Intel N100 + RTX 5060Ti ๐Ÿ”—
r/LocalLLaMA Introducing Unsloth Desktop app ๐Ÿ”—
r/LocalLLaMA nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16 ยท Hugging Face ๐Ÿ”—
r/LocalLLaMA The small open weight models are scarier in AI development ๐Ÿ”—
r/LocalLLaMA We even got a fgn manifesto!! Meta is on a run! ๐Ÿ”—
r/LocalLLaMA Tested in Coding: BF16 Muse Glimmer vs BF16 Qwen3.6 27B ๐Ÿ”—
r/LocalLLaMA A friendly reminder that tomorrow is Qwednesday. ๐Ÿ”—
r/LocalLLaMA Luth-2: New State-of-the-Art French Small Language Models ๐Ÿ”—
r/LocalLLaMA I ran Muse Glimmer @ 1M context - All tests passed. ๐Ÿ”—
r/LocalLLaMA Qwen 3.8-27b coming this week ๐Ÿ”—
r/LocalLLaMA Nvidia reportedly testing lower memory configs of Rubin Ultra as memory shortage bites back โ€” designs tested include as little as 192 GB and step back to HBM4 ๐Ÿ”—
r/LocalLLaMA I gave DeepSeek V4 Flash basic vision by training a 40M connector on 100K examples ๐Ÿ”—
r/LocalLLaMA 1 Day in and I feel okay saying Muse-Glimmer-30B finally beats 3.6-27B for the size in some use-cases ๐Ÿ”—
r/LocalLLaMA Observations on Muse-Glimmer reasoning traces being noticeably different from qwen / gemma models and questions for you guys ๐Ÿ”—
10.08.2026
r/LocalLLaMA I trained a 1B-parameter LLM from scratch on 20B tokens for about $200 ๐Ÿ”—
r/LocalLLaMA Muse glimmer benchmark ๐Ÿ”—
r/LocalLLaMA Muse Spark 1.2 Open Source before Llama 4 Behemoth!!? ๐Ÿ”—
r/LocalLLaMA I made a web-design benchmark for local models (Muse Glimmer 30B vs Qwen 3.6 27b vs Deepseek V4 Flash 0731) ๐Ÿ”—
r/LocalLLaMA Please Share Your Experience About Muse Glimmer ๐Ÿ”—
r/LocalLLaMA I compared GGUF quants of Qwen3.6 27B to NVFP4, AWQ, AutoRound, and FP8 ๐Ÿ”—
r/LocalLLaMA Meta wants to create a personal super intelligence for free that runs on your Maschine. Meanwhile at anthropic: How can we get our (fake) image back of the "Humanity" Lab? ReleaseFreeOpenSourceModels Now let's just double the price on all our users, and tax our plan with 50% Fable weekly usages ๐Ÿ”—
r/LocalLLaMA inclusionAI/Ling-3.0-tiny ยท 8B A1.3B MoEยท Hugging Face ๐Ÿ”—