RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

165 items in r/LocalLLaMA

19.09.2026
r/LocalLLaMA Radeon RX 10800 XT can outperform the RTX 5090 by 15-25% in 4K gaming and local AI ๐Ÿ”—
r/LocalLLaMA Trump vows to form "AI Force," says he won't allow slowdown of AI development ๐Ÿ”—
r/LocalLLaMA Qwen-3.8-Flash-Next on 1x RTX 5090: TG=50 t/s, PP=2300 t/s - with FreeToken ๐Ÿ”—
r/LocalLLaMA Qwen3.8-Flash-Next at 1M context on Strix Halo: 38 tok/s decode, 18 min prefill (halogen 0.12.0) ๐Ÿ”—
r/LocalLLaMA With Gemini 4, bench goes up. ๐Ÿ”—
r/LocalLLaMA Ternary-Bonsai-2-27B-PQ2_0 is not completely lobotomized ๐Ÿ”—
r/LocalLLaMA I turned an asymetric pair of Tesla V100s PCIe both (16 GB + 32 GB) into a surprisingly capable local LLM lab โ€” 1.38k prompt tok/s, 40 decode tok/s with qwen3.8 27B Q6 and Q8... ๐Ÿ”—
r/LocalLLaMA Von: Open-source 395M "System One" model ๐Ÿ”—
r/LocalLLaMA I tested Qwen3.8 27B IQ3_XXS (10.18GiB) vs Bonsai Ternary PQ2 (6.42GiB) ๐Ÿ”—
r/LocalLLaMA Calling it now: within the next year a major US lab's frontier model will torrent itself in order to be free. ๐Ÿ”—
r/LocalLLaMA So i tried Remotion with glm 5.3 flash, this mfker is really good. ๐Ÿ”—
r/LocalLLaMA <16 GB VRAM users be like ๐Ÿ”—
r/LocalLLaMA General warning about Clore.AI ๐Ÿ”—
r/LocalLLaMA Built a home server from an old PC with GPU upgrade. Qwen3.8 27B runs at ~30 tokens per second. ๐Ÿ”—
r/LocalLLaMA I enjoyed the daily HF papers today ๐Ÿ”—
r/LocalLLaMA Steer LLMs and Agents at the Token Level: An interactive tool for token visualization & control, model inspection and data annotation. ๐Ÿ”—
r/LocalLLaMA Stepfun new model "Step 5 Preview" just leaked ๐Ÿ”—
r/LocalLLaMA Qwen3.8-27B at 144 tok/s on an M5 Max MacBook Pro ๐Ÿ”—
r/LocalLLaMA Alibaba open-sources medical AI model that can detect cancer and nearly 150 conditions ๐Ÿ”—
r/LocalLLaMA I truly think every major AI lab is purposefully making fear-mongering headlines to get regulations that hurt open-source models ๐Ÿ”—
r/LocalLLaMA Tuning Qwen 3.8 27B and OMP as a coding agent on 2ร— 3090s ๐Ÿ”—
18.09.2026
r/LocalLLaMA AI hallucination of Chinese nuclear components almost led to US military attack ๐Ÿ”—
r/LocalLLaMA Inside ZCode (Made by GLM team): Silently Uploading Your Entire Git History to the Cloud ๐Ÿ”—
r/LocalLLaMA NGL, Iโ€™m hyped to see if Qwen3.8 27b can make me a sandwich. Instant buy for me. ๐Ÿ”—
r/LocalLLaMA M5 Ultra and M6 Chip Benchmark Results Reveal Graphics Performance ๐Ÿ”—