RSS Feed
All
r/AI_Agents
r/AI_Governance
r/ClaudeAI
r/ClaudeCode
r/DeepSeek
r/Futurology
r/LangChain
r/LocalLLaMA
r/MachineLearning
r/OpenAI
r/artificial
r/cybersecurity
r/europe
r/hermesagent
r/netsec
r/singularity
r/technology
165 items in r/LocalLLaMA
19.09.2026
r/LocalLLaMA
Radeon RX 10800 XT can outperform the RTX 5090 by 15-25% in 4K gaming and local AI
๐
r/LocalLLaMA
Qwen3.8-Flash-Next at 1M context on Strix Halo: 38 tok/s decode, 18 min prefill (halogen 0.12.0)
๐
r/LocalLLaMA
I turned an asymetric pair of Tesla V100s PCIe both (16 GB + 32 GB) into a surprisingly capable local LLM lab โ 1.38k prompt tok/s, 40 decode tok/s with qwen3.8 27B Q6 and Q8...
๐
r/LocalLLaMA
Calling it now: within the next year a major US lab's frontier model will torrent itself in order to be free.
๐
r/LocalLLaMA
Built a home server from an old PC with GPU upgrade. Qwen3.8 27B runs at ~30 tokens per second.
๐
r/LocalLLaMA
Steer LLMs and Agents at the Token Level: An interactive tool for token visualization & control, model inspection and data annotation.
๐
r/LocalLLaMA
Alibaba open-sources medical AI model that can detect cancer and nearly 150 conditions
๐