RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

1252 items in r/LocalLLaMA

30.07.2026
r/LocalLLaMA Software Engineers: Do you honestly get anything useful out of LLMs? πŸ”—
r/LocalLLaMA Turbo-fieldfare: Open-source engine running Gemma 4 26B in 2 GB RAM on Apple Silicon πŸ”—
r/LocalLLaMA Think of the children, another excuse for them to go after open source AI πŸ”—
r/LocalLLaMA GLM 5.2 with vision on Hugging Face πŸ”—
r/LocalLLaMA Benchmarked: MindControl for Llama.cpp πŸ”—
r/LocalLLaMA Anyone tested the IQ1_M 342GB Pruned Kimi K3? Is it usable? πŸ”—
29.07.2026
r/LocalLLaMA Bought a 5090 to escape API fees. Ended up building a mini datacenter. Sound familiar? πŸ”—
r/LocalLLaMA Are you guys not scared of where we're heading? A year ago, GPT-5 was considered one of the best models in the world. Today, we have open-weight models like Qwen3.6-27B that are competitive enough to run locally on high-end consumer hardware. The pace of progress is absolutely brutal. πŸ”—
r/LocalLLaMA The open-weights carousel never stops. πŸ”—
r/LocalLLaMA Kimi K3 for local use (1.56TB β†’ 594GB) compressed and released by Unsloth πŸ”—
r/LocalLLaMA PSA: llama.cpp now loads MTP tensors by default for any draft-mtp arch, even with MTP disabled πŸ”—
r/LocalLLaMA Ilintar's Official Guide To Model Selection πŸ”—
r/LocalLLaMA Everyone posts day-one impressions. What's still in your stack a month later? πŸ”—
r/LocalLLaMA First Kimi K3 results on home lab ~ 4t/s πŸ”—
r/LocalLLaMA I keep coming back to Qwen... Over and Over. Is there really nothing better under 120B? πŸ”—
r/LocalLLaMA "Uncensored" LLMs are measurably more optimistic than their base models πŸ”—
r/LocalLLaMA The idea: on a CPU the decode speed depends on the active params per token, not the total. My objective is trying to run a 10B at 100tok/s on a mid level PC (No GPU). πŸ”—
r/LocalLLaMA Understand Kimi K3 from first principles: a recommended order for anyone trying to understand this beast πŸ”—
r/LocalLLaMA A slide deck you can edit with a local model or in Chrome β€” the whole deck is a JSON block in one HTML file (~640KB with editor and viewer included) πŸ”—
r/LocalLLaMA Leaks : Z.ai GLM5.5 aiming Fable-5 on August πŸ”—
r/LocalLLaMA Microsoft did it .... again! (404 for their Mage-Flow models on HF) πŸ”—
r/LocalLLaMA I built a GBNF grammar compiler that makes 8B models reliably call tools - here's how it works (deep dive) πŸ”—
r/LocalLLaMA Anyone tried the Q1 Kimi K3 yet? (555GB) πŸ”—
r/LocalLLaMA A.X-K2 released πŸ”—
r/LocalLLaMA I tried running a 1.56TB MoE model on a 6GB RTX 4050 Laptop, Here’s the result πŸ”—