RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

1214 items in r/LocalLLaMA

26.08.2026
r/LocalLLaMA Thomson Reuters releases Thomson-1.0-Small. A law and tax focused model πŸ”—
r/LocalLLaMA Fully quantized NVFP4 Qwen3.8-27B with QUASAR QAD πŸ”—
25.08.2026
r/LocalLLaMA It's here! πŸ”—
r/LocalLLaMA Granite Speech 5.0 Turbo CTC: Extremely Fast and Accurate Transcription πŸ”—
r/LocalLLaMA Qwen3.8-Flash-Next. This architecture could be surprisingly local-friendly once the weights drop. πŸ‘€ πŸ”—
r/LocalLLaMA Mac Studio M5 Max Cost Analysis πŸ”—
r/LocalLLaMA ibm-granite/granite-4.2-30b Β· Hugging Face πŸ”—
r/LocalLLaMA Intel Arc Pro B60 Dual 48G spotted πŸ”—
r/LocalLLaMA me to the model I spent all weekend fine-tuning πŸ”—
r/LocalLLaMA It’s over folks. πŸ”—
r/LocalLLaMA Apple releases M5 ultra at 1.2TB/s bandwith πŸ”—
r/LocalLLaMA Apple introduces new Mac Studio with M5 Max and M5 Ultra - up to 512GB of unified memory πŸ”—
r/LocalLLaMA Quantization-Aware Healing: a compressed, 4-bit model that outperforms its full-precision original πŸ”—
r/LocalLLaMA Qwen 3.8 Flash Next day 0 support from unsloth πŸ”—
r/LocalLLaMA New: Llama.cpp adaptive speculation for faster inference πŸ”—
r/LocalLLaMA Qwen3.8 flash next πŸ”—
r/LocalLLaMA Qwen3.8-Flash-Next tomorrow πŸ”—
r/LocalLLaMA Glm 5.3 flash? πŸ”—
r/LocalLLaMA tencent/WeMM-Embedding 9B/4B/2B πŸ”—
r/LocalLLaMA How many of you are using Kiwix/offline Wikipedia for training? πŸ”—
r/LocalLLaMA Today I merged the first feature branch written entirely by my 4060Ti 16GB! πŸ”—
24.08.2026
r/LocalLLaMA I just tried DeepSeek Harness and it escaped from its workspace folder πŸ”—
r/LocalLLaMA Copilot you say? πŸ”—
r/LocalLLaMA JetBrains local AI (using Qwen3.6 27B) πŸ”—
r/LocalLLaMA Please join r/LowEndLocalAI, a community for running local LLMs on low spec hardware πŸ”—