RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

1210 items in r/LocalLLaMA

03.09.2026
r/LocalLLaMA How does meta spark 1.3 match claude fable 5 on benchmarks? Did anyone try it in agentic coding? ๐Ÿ”—
r/LocalLLaMA local AI can't be disabled ๐Ÿ”—
r/LocalLLaMA Introducing K2 Horizon: Frontier Performance, Radically Open ๐Ÿ”—
r/LocalLLaMA IFM/K2-Horizon-MoVA-36B-A4B-GGUF ยท Hugging Face ๐Ÿ”—
r/LocalLLaMA Frontier models sabotaging local AI implementations? ๐Ÿ”—
r/LocalLLaMA Could the shortage be getting better? ๐Ÿ”—
r/LocalLLaMA It's official! Nvidia to acquire Hugging Face for 12.9 billion dollars. ๐Ÿ”—
r/LocalLLaMA We built an open-source, model-neutral agent harness and compared it with claude managed agents - for the same model, got same accuracy, upto 75% lower cost ๐Ÿ”—
r/LocalLLaMA Qwen3.6 35b Q2_XXS: Being GPU poor in 2026 is not so bad ๐Ÿ”—
r/LocalLLaMA Qwen-3.8-Next-Flash Ngram Hot-Swappable Knowledge Injector for llama.cpp ๐Ÿ”—
r/LocalLLaMA How to handle naughty model ๐Ÿ”—
r/LocalLLaMA model: add NVIDIA Nemotron-3-Puzzle-75B-A9B (NemotronHPuzzle) support by YanissAmz ยท Pull Request #25444 ยท ggml-org/llama.cpp ๐Ÿ”—
r/LocalLLaMA Can a 4B local model actually feel like an AI assistant? ๐Ÿ”—
r/LocalLLaMA My RULE of Thumb of choosing a models ๐Ÿ”—
r/LocalLLaMA KV cache might be a bigger problem for local models than parameter count ๐Ÿ”—
r/LocalLLaMA Qwen3.8-Flash-Next on 2x3090 + DDR4: 17 โ†’ 25-29 t/s decode with the expert cache PR ๐Ÿ”—
r/LocalLLaMA GLM5.3 Flash over DSV4 Flash? ๐Ÿ”—
r/LocalLLaMA Microsoft VibeVoice-ASR-Streaming Released ๐Ÿ”—
r/LocalLLaMA DeepSeek-V4-Flash vs. GLM-5.3-Flash on 2ร— DGX Spark ๐Ÿ”—
02.09.2026
r/LocalLLaMA Qwen3.8-flash-next sees corruption everywhere ๐Ÿ”—
r/LocalLLaMA Perplexity open-sourced their Mac inference server for Qwen 3.6 ๐Ÿ”—
r/LocalLLaMA Muse Spark open weights coming soon ๐Ÿ”—
r/LocalLLaMA Confirmed bolting Q8 NGram into IQ4 Qwen no speed degradation ๐Ÿ”—
r/LocalLLaMA GLM 5.3 Flash makes a black hole Minecraft mod running locally on 4x RTX PRO 6000 WS ๐Ÿ”—
r/LocalLLaMA Vision support merged for DeepSeek-V4-Flash-Vision-Exp ๐Ÿ”—