RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

167 items in r/LocalLLaMA

17.09.2026
r/LocalLLaMA Cactus Needle 3: A Sliceable 8-29MB Automation Foundation Model That Matches DeepSeek v4 Flash ๐Ÿ”—
r/LocalLLaMA Thank you :) Swift Qwen 3.8 27B now has 100k+ downloads, is #1 finetune and #9 model on HuggingFace Trending ๐Ÿ”—
r/LocalLLaMA Qwen announces Qwen3.8-Omni-Flash ๐Ÿ”—
r/LocalLLaMA IFM/K2-Horizon-7B-Uno ยท Hugging Face - 5200tps with no quality loss ๐Ÿ”—
r/LocalLLaMA AMD Plans 10% Price Hike Across GPUs, Chipsets, and Possibly CPUs ๐Ÿ”—
r/LocalLLaMA shots fired at dario from glm ๐Ÿ”—
r/LocalLLaMA Fujitsu formally announces their MONAKA 144 core CPU ๐Ÿ”—
r/LocalLLaMA 153 tok/s on 1x AMD Radeon R9700 running Qwen3.8 27b NVFP4, 470 tok/s @ 8 conc requests, Prefill @ 3,619 tok/s ๐Ÿ”—
r/LocalLLaMA Does anyone use uncensored models purely for coding? ๐Ÿ”—
r/LocalLLaMA Keeping vLLM's Prefix Cache Warm Between Agent Turns ๐Ÿ”—
r/LocalLLaMA First M5 Ultra benchmarks ๐Ÿ”—
r/LocalLLaMA Update : Small model + Engram ๐Ÿ”—
r/LocalLLaMA China's Huawei says AI chip demand outstrips supply as it steps up Nvidia challenge ๐Ÿ”—
r/LocalLLaMA Intel releases OpenVINO 2026.4 ๐Ÿ”—
r/LocalLLaMA XingChen-AGI/Xing4.0-29B-A4B MoE ๐Ÿ”—
r/LocalLLaMA I literally built the Jev architecture one year back and completely open-sourced it with model, dataset and paper ๐Ÿ”—
r/LocalLLaMA I literally built the Jev architecture one year back and completely open-sourced it with model, dataset and paper ๐Ÿ”—