RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

1211 items in r/LocalLLaMA

02.09.2026
r/LocalLLaMA Vision support merged for DeepSeek-V4-Flash-Vision-Exp ๐Ÿ”—
r/LocalLLaMA Do we forget about another Qwen model for a while now ? ๐Ÿ”—
r/LocalLLaMA H3-World: Turning Language Understanding into World Control ๐Ÿ”—
r/LocalLLaMA Qwen will be the king? ๐Ÿ”—
r/LocalLLaMA GLM 5.3 flash is annoying ๐Ÿ”—
r/LocalLLaMA LocalLLaMA is unironically one of the best places to go to get up to date AI news. ๐Ÿ”—
r/LocalLLaMA Running 104GB Qwen3.8-Flash-Next on 48GB Mac at ~12 tok/s ๐Ÿ”—
r/LocalLLaMA Qwen3.8-Max-0902 released ๐Ÿ”—
r/LocalLLaMA 2/5 of my CMP 170HX have died after 2 weeks and the 3rd came with defective tensor cores. Current prices DO NOT justify the risk you are taking ๐Ÿ”—
r/LocalLLaMA Everyone is t/s maxing.. 3.8.. but after a week of using it for work I'm tempted to switch back to 3.6 ๐Ÿ”—
r/LocalLLaMA Sam Altman trying to sabotage open source compute ๐Ÿ”—
01.09.2026
r/LocalLLaMA Really stunned by the Singularity comment section ๐Ÿ”—
r/LocalLLaMA Kaitchup posted Qwen3.8 27B Benchmarks for quants from Q4 to Q1 ๐Ÿ”—
r/LocalLLaMA Fingers crossed for a 122b or really anything above 31b.๐Ÿคž ๐Ÿ”—
r/LocalLLaMA Keeping up with model launches ๐Ÿ”—
r/LocalLLaMA Is it silly to get a 64GB Strix Halo (Framework Desktop) ~$2000? ๐Ÿ”—
r/LocalLLaMA Round-Robin with llama-server? ๐Ÿ”—
r/LocalLLaMA Help me set up local AI for my 85 year old aunt who is blind. ๐Ÿ”—
r/LocalLLaMA Intel hints it may get back into memory business ๐Ÿ”—
r/LocalLLaMA Deceptive model quantization from AtomicChat? ๐Ÿ”—
r/LocalLLaMA New Model: Spark-X2.5-4B, Spark-X2.5-1.7B ๐Ÿ”—
r/LocalLLaMA I pushed Qwen3.8-27B to 2.000 prefill per second and 132 decode per second on A RTX 3090. ๐Ÿ”—
r/LocalLLaMA Qwen 3.8 27b (Q4KM) oneshot a Super Mario clone ๐Ÿ”—
r/LocalLLaMA New Gemma models on arena ai ๐Ÿ”—
r/LocalLLaMA ExLlamav3 Recent Updates : CPU offload, GLM-5.3-FLASH, Qwen3.8-Flash, SC Quants ++ ๐Ÿ”—