RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

1205 items in r/LocalLLaMA

16.09.2026
r/LocalLLaMA China's open-weight AI models are now just 4 months behind frontier US offerings, Mozilla report claims β€” models still lag in some benchmarks but are drastically cheaper to use πŸ”—
r/LocalLLaMA What's the current best LLM uncensoring method? πŸ”—
r/LocalLLaMA Qwen3.8 Max (0902) scores 45 on the Artificial Analysis Intelligence Index, up 5 points in a month and back on top of China's leaderboard, nosing out GLM-5.3 (44.9) and Kimi K3 (43.8) πŸ”—
r/LocalLLaMA Qwen3.5 4B + grabbing logits is almost "Jev"? Or even just Qwen Reranker? πŸ”—
r/LocalLLaMA Open source will push the new frontier, here's why πŸ”—
r/LocalLLaMA Apple May Return to Server Market With Nvidia Technology πŸ”—
r/LocalLLaMA LocalJev? πŸ”—
r/LocalLLaMA You can offload most of Qwen3.8-Flash-Next's KV cache to RAM with little decode slowdown πŸ”—
r/LocalLLaMA [Release] SOTA GGUFs for Qwen3.8-Flash-Next: GSQ-RCO Providing Near Baseline Performance πŸ”—
r/LocalLLaMA Mozilla Report: China-U.S. AI Model Capability Gap Narrows to 4.4 Months πŸ”—
r/LocalLLaMA qwen4exp: add hc ops by am17an Β· Pull Request #28901 Β· ggml-org/llama.cpp πŸ”—
r/LocalLLaMA Hey, Meta. Where's those Muse Spark weights? πŸ”—
r/LocalLLaMA What's the best open weight model for Blender? That's comparable to Astra πŸ”—
r/LocalLLaMA mlabonne/LFM2.5-230M-Chess Β· Hugging Face πŸ”—
r/LocalLLaMA Qwen3.8 27b Game Dev Part 2 πŸ”—
r/LocalLLaMA LACT PR to let NVIDIA gpus go lower than stock VBIOS limit (so below 400W for 5090, or below 250W for 6000 PRO MaxQ) πŸ”—
r/LocalLLaMA Open Source Appreciation Post πŸ”—
r/LocalLLaMA What's the next local model you are excited about? πŸ”—
15.09.2026
r/LocalLLaMA Don’t buy a $9K RTX 5090.... instead. πŸ”—
r/LocalLLaMA censorship has begun on HuggingFace πŸ”—
r/LocalLLaMA A very unexpected analogy πŸ”—
r/LocalLLaMA Apple Foundation Models: local AI natively on MacOS 27 πŸ”—
r/LocalLLaMA Cut Qwen3.8-27B Reasoning Tokens by 40% -- 3.8 'ThinkingCap' benchmarked! πŸ”—
r/LocalLLaMA Koboldcpp v1.121 released πŸ”—
r/LocalLLaMA Got it unopened off Craigslist for $4k. Excited to start hosting my own models! πŸ”—