RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

641 items in r/LocalLLaMA

17.09.2026
r/LocalLLaMA IFM/K2-Horizon-7B-Uno ยท Hugging Face - 5200tps with no quality loss ๐Ÿ”—
r/LocalLLaMA AMD Plans 10% Price Hike Across GPUs, Chipsets, and Possibly CPUs ๐Ÿ”—
r/LocalLLaMA shots fired at dario from glm ๐Ÿ”—
r/LocalLLaMA Fujitsu formally announces their MONAKA 144 core CPU ๐Ÿ”—
r/LocalLLaMA 153 tok/s on 1x AMD Radeon R9700 running Qwen3.8 27b NVFP4, 470 tok/s @ 8 conc requests, Prefill @ 3,619 tok/s ๐Ÿ”—
r/LocalLLaMA Does anyone use uncensored models purely for coding? ๐Ÿ”—
r/LocalLLaMA Keeping vLLM's Prefix Cache Warm Between Agent Turns ๐Ÿ”—
r/LocalLLaMA First M5 Ultra benchmarks ๐Ÿ”—
r/LocalLLaMA Update : Small model + Engram ๐Ÿ”—
r/LocalLLaMA China's Huawei says AI chip demand outstrips supply as it steps up Nvidia challenge ๐Ÿ”—
r/LocalLLaMA Intel releases OpenVINO 2026.4 ๐Ÿ”—
r/LocalLLaMA XingChen-AGI/Xing4.0-29B-A4B MoE ๐Ÿ”—
r/LocalLLaMA I literally built the Jev architecture one year back and completely open-sourced it with model, dataset and paper ๐Ÿ”—
r/LocalLLaMA I literally built the Jev architecture one year back and completely open-sourced it with model, dataset and paper ๐Ÿ”—
16.09.2026
r/LocalLLaMA Upgraded my local setup with 2 rtx pros and it's amazing. ๐Ÿ”—
r/LocalLLaMA Xiaomi is livestreaming the MIMO-V2.6-PRO/FLASH post-training! ๐Ÿ”—
r/LocalLLaMA Openjev ๐Ÿ”—
r/LocalLLaMA Qwen 3.8 27B Running for 63 hours on a RTX 3090 to solve the Riemann hypothesis ๐Ÿ”—
r/LocalLLaMA Xiaomi MiMo 2.6 Live Training Dashboard ๐Ÿ”—
r/LocalLLaMA i left gpu poor range ๐Ÿ”—
r/LocalLLaMA Frontier LLM development simplified for politicians: ๐Ÿ”—
r/LocalLLaMA Qwen3.8 Flash on 12GB VRAM - 15 tokens/s ๐Ÿ”—
r/LocalLLaMA China's open-weight AI models are now just 4 months behind frontier US offerings, Mozilla report claims โ€” models still lag in some benchmarks but are drastically cheaper to use ๐Ÿ”—
r/LocalLLaMA What's the current best LLM uncensoring method? ๐Ÿ”—
r/LocalLLaMA Qwen3.8 Max (0902) scores 45 on the Artificial Analysis Intelligence Index, up 5 points in a month and back on top of China's leaderboard, nosing out GLM-5.3 (44.9) and Kimi K3 (43.8) ๐Ÿ”—