RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

1221 items in r/LocalLLaMA

18.08.2026
r/LocalLLaMA Qwen dev says not to wait for 35B-A3B πŸ”—
r/LocalLLaMA CDW has bumped the MSRP of the RTX Pro 6000 from $16,000 to $19,999 πŸ”—
17.08.2026
r/LocalLLaMA Waiting for Qwen 3.8 35B A3B πŸ”—
r/LocalLLaMA Weirdly, no one talks about Temperature setting for the Qwen3.8 27b πŸ”—
r/LocalLLaMA "Opus 4.8 thinks too much", "Muse Glimmer sits between Gemma and Qwen, that's boring", "Gemma 4 is too lazy" πŸ”—
r/LocalLLaMA Qwen 3.8 35bA3b wen? πŸ”—
r/LocalLLaMA llama.cpp adaptive MTP PR#27210 πŸ”—
r/LocalLLaMA Qwen3.8 27B > Opus 5 Medium on Artificial Analysis Agentic Index πŸ”—
r/LocalLLaMA Qwen3.8 27B = GPT-5.6 Luna compressed into 27B πŸ”—
r/LocalLLaMA Artificial Analysis' Qwen3.8-27B benchmarks put it neck and neck with DeepSeek V4 and GPT-5.6 Luna Max πŸ”—
r/LocalLLaMA Ling 3.0 Tiny is the strongest, fastest and greatest model on my low end PC! πŸ”—
r/LocalLLaMA llama.cpp version v0.1.0 has been released πŸ”—
r/LocalLLaMA After pushing 1M+ tokens through Qwen 3.8 27B, here is my optimal llama.cpp config for 16GB VRAM (73k Context, Agentic Coding) πŸ”—
r/LocalLLaMA 100$ worth of gpu runs qwen 3.8 27b at 7.39 t/s πŸ”—
r/LocalLLaMA LLM's can't "jump" - a paper by Deepmind showing LLMs can't generate novel explanatory hypotheses πŸ”—
r/LocalLLaMA Unpopular opinion : Qwen 3.8 27b is not an overthinker πŸ”—
r/LocalLLaMA Petition to add a rule for people to add their DAMN quant levels to their posts πŸ”—
r/LocalLLaMA Ling 3.0 support merged into llama.cpp πŸ”—
r/LocalLLaMA Qwen3.8-27B Q8_0 on Strix Halo is seriously impressive πŸ”—
r/LocalLLaMA Long Review: Qwen 3.8 27B is VERY good at tapping into it's real-world knowledge. It's "overthinking" brings it to Sonnet level performance with the potential for Opus level results. πŸ”—
r/LocalLLaMA Stripe will reportedly acquire AI gateway startup OpenRouter for $7B+ πŸ”—
r/LocalLLaMA How many tokens/second output are you getting with Qwen3.8-27B? πŸ”—
r/LocalLLaMA …and I’m not afraid of losing my social credits. πŸ”—
16.08.2026
r/LocalLLaMA Qwen 3.8 27b vs 3.6 27b - how good is with a Turtle library. πŸ”—
r/LocalLLaMA Dario Amodei defends his policy proposals, warns open weights won't decentralize power, endorses pre-launch vetting, says real accomplishments will earn trust πŸ”—