RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

1207 items in r/LocalLLaMA

10.09.2026
r/LocalLLaMA deepseek-ai/DeepSeek-V4.1-Flash ยท Hugging Face ๐Ÿ”—
r/LocalLLaMA So relevant ๐Ÿ”—
r/LocalLLaMA Qwen3.8-Flash-Next on 2x3090 + DDR4, part 4: 2.2-2.5x faster prefill by kicking the expert cache off the GPU while the prompt runs ๐Ÿ”—
r/LocalLLaMA Local LLM / Qwen 3.8 win ๐Ÿ”—
09.09.2026
r/LocalLLaMA Running qwen 3.8 27B iq3 xxs on RTX 3060. ๐Ÿ”—
r/LocalLLaMA Apple A20 Pro debuts with 7-core GPU, 32-core Neural Engine and 50% more memory bandwidth (~115 GB/s) ๐Ÿ”—
r/LocalLLaMA Surveillance plagiarism by OpenAI ๐Ÿ”—
r/LocalLLaMA Qwen3.8-27B has the best coding ceiling you can run at home on consumer hardware, it ships with reasoning_effort defaulting to xhigh - I measured what that costs ๐Ÿ”—
r/LocalLLaMA Don't let FOMO win if you're interested in local llm from a hobby/learning aspect ๐Ÿ”—
r/LocalLLaMA Server rebuild to custom loop. 2x RTX Titans 24gb, 1x 22gb 2080ti | T: 70GB VRAM. ๐Ÿ”—
r/LocalLLaMA I trained an audio model that can generate infinite one-shots for music production and turn text prompts into fully playable synths. I'm not only releasing the model but I've also released a video on exactly how I did it (and the inferencing pipeline to let others make text based synths.) ๐Ÿ”—
r/LocalLLaMA Best Open source TTS right now for narration? ๐Ÿ”—
r/LocalLLaMA Mention if a "new model" is a finetune ๐Ÿ”—
r/LocalLLaMA SOTA ImageGen Locally NVIDIA Cosmos3(64B) INT4 quants CUDA/MLX ๐Ÿ”—
r/LocalLLaMA Qwen3.8-27B-Uncensored-Genesis-V1-GGUF ๐Ÿ”—
r/LocalLLaMA 1-bit 27B in the browser: 25โ€“30 tok/s on a 6 GB RTX 3060 Laptop (WebGPU, no install) ๐Ÿ”—
r/LocalLLaMA Why the hell is LM Studio making LM Studio so difficult to download? ๐Ÿ”—
r/LocalLLaMA GLM 5.3 Flash Q4 @ 60tps / 550tps on M3 Ultra ๐Ÿ”—
r/LocalLLaMA Now this is a serious local machine ๐Ÿ”—
r/LocalLLaMA DeepSeek-V4-Flash-Vision-Exp (285B MoE) on 10-12x RTX 3090 โ€” spec decoding, vision ๐Ÿ”—
r/LocalLLaMA Deepseek Has Soft Retired Deepseek V4 Pro ๐Ÿ”—
r/LocalLLaMA OpenAI and the Navier-Stokes controversy ๐Ÿ”—
r/LocalLLaMA new Nex model ๐Ÿ”—
r/LocalLLaMA MiMo-X-Pro-Preview and MiMo-X-Flash-Preview - New Mimo Model found in Mimo Desktop preview announcement ๐Ÿ”—
r/LocalLLaMA Is there a dummies guide for setting up qwen 27B with dflash2 and n-gram? ๐Ÿ”—