RSS Feed
All
r/AI_Agents
r/AI_Governance
r/ClaudeAI
r/ClaudeCode
r/DeepSeek
r/Futurology
r/LangChain
r/LocalLLaMA
r/MachineLearning
r/OpenAI
r/artificial
r/cybersecurity
r/europe
r/hermesagent
r/netsec
r/singularity
r/technology
1216 items in r/LocalLLaMA
21.08.2026
r/LocalLLaMA
Qwen3.8-27B at 262K context on a Strix Halo + RTX 3090 Ti: 9.5 -> 153 tok/s, and it beats a dual-3090 vLLM box on HumanEval
๐
20.08.2026
r/LocalLLaMA
NVIDIA dropped an NVIDIA-hosted CUDA MCP for AI-assisted CUDA operations, such as searching official, up-to-date documentation, writing optimized GPU code, and analyzing performance data
๐
r/LocalLLaMA
Any speculation on whether or not Google will announce a new Gemma model at the Gemma SF Celebration tonight?
๐
r/LocalLLaMA
AQuA's "self-improvement" updates research state, not the agent LM. What should a local port freeze?
๐
r/LocalLLaMA
QwenMix-3.7: Kept seeing posts about Qwen3.8 and 3.6 sharing the same structure.. so I had Qwen3.8 combine them.
๐
r/LocalLLaMA
[MASSIVE TINY RELEASE] - Supra2-Medium-Base - a tiny 25M parameters model competing heavily with our previous 50M model!
๐