RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

162 items in r/LocalLLaMA

24.09.2026
r/LocalLLaMA UkisAI Swift Series / 27B, Flash Next and Bonsai 2 + GSQ-RCO / -63.4% thinking, x1.95 speed with xhigh accuracy πŸ”—
r/LocalLLaMA ThinkingCap 3.8-27B vs. Swift 3.8-27B vs. Qwen 3.8-27B Benchmarks πŸ”—
r/LocalLLaMA PSA: llama.cpp -cram should be increased for agentic workflows (default is 8192) πŸ”—
r/LocalLLaMA loaded 1b local ai for driving assistant integrated with ADAs. πŸ”—
r/LocalLLaMA Muse wen? πŸ”—
r/LocalLLaMA list of entities hacked by openai grows by one πŸ”—
r/LocalLLaMA R9V Update: Created and adopted KVA projections based on Deepseek V4.1 Flash + HySparse2/MiMo-V3 for Qwen3.8 Flash Next. This is a game changer for models that don't natively implement it. 1.45-1.85x speedup in prefill to 3k+ at a small deficit to perplexity. [2x R9700, 128GB DDR5] πŸ”—
r/LocalLLaMA Qwen-3.8-27B is good enough that I stopped using API πŸ”—
r/LocalLLaMA Lesson learned. Don't blindly trust repos and make sure everything is stable for a long running (multi weeks) benchmark. πŸ”—
r/LocalLLaMA Folks, have you purchased the Mac M5 Ultra with 256GB yet? We need serious benchmarks, because we only get YouTube clowns influencers results πŸ”—
r/LocalLLaMA JEV almost dead: CLM vs JEV πŸ”—
r/LocalLLaMA My foray into local ai. Two BC-250 ex mining apus running Qwen3.6-35B-A3B Q4_K_M at 60 tok/s with 64k context πŸ”—
r/LocalLLaMA Contrastive Language Models πŸ”—
r/LocalLLaMA what in the fck is an ngram πŸ”—
r/LocalLLaMA What TPS is too slow for you? πŸ”—
r/LocalLLaMA The Pope’s AI Guy Is Worried About β€˜Cartel’ Behavior Among Big Labs πŸ”—
23.09.2026
r/LocalLLaMA Qwen 3.8 Flash Next q4_k_m, 130k context, q8 cache on 16GB VRAM ann 64GB RAM, 15-20 t/s on 4080 πŸ”—
r/LocalLLaMA Using uncensored models makes working less of a headache πŸ”—
r/LocalLLaMA Qwen FN vs 27B --- Think I'm saturated. πŸ”—
r/LocalLLaMA Please Google, for the love of God. πŸ”—
r/LocalLLaMA Introducing Support for Local AI Models in the Antigravity SDK πŸ”—
r/LocalLLaMA Jev isn't new tech. Its marketing targets people who think AI started with LLMs. πŸ”—
r/LocalLLaMA apple/LensVLM-9B Β· Hugging Face πŸ”—
r/LocalLLaMA BFL releases FLUX 3 Action: a 7B robot model πŸ”—
r/LocalLLaMA MiMo-V2.6 (both Pro and Flash) is a benchmaxxed scam πŸ”—