RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

1205 items in r/LocalLLaMA

14.09.2026
r/LocalLLaMA Local On Laptop - I am due an upgrade and need your help! πŸ”—
r/LocalLLaMA Made a Windows app that translates in real-time as you type β€” seeking model recommendations & feedback πŸ”—
r/LocalLLaMA What are the current best retail GPUs for max VRAM at a reasonable price? πŸ”—
r/LocalLLaMA Running Qwen 3.8 next on 16vram+32ram - A useful/fun post for the gpu poors πŸ”—
r/LocalLLaMA RTX PRO 5500 Blackwell (84GB) released πŸ”—
r/LocalLLaMA Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored-NVFP4 - most accurate tool calling πŸ”—
r/LocalLLaMA 5090 Stock is Almost Gone πŸ”—
r/LocalLLaMA Another Qwen3.8-27b Appreciation Post πŸ”—
r/LocalLLaMA R9V Update: now ~100 tok/s in TG on Qwen3.8 Flash Next IQ4_XS on x2 R9700 + 128GB RAM. Fixed crashes with n-gram SSD streaming, improved diagnostics, plus pinned images. Q4_K_XL now supported, 50 tok/s TG. πŸ”—
r/LocalLLaMA Micron's memory wall chart. Compute up ~3x every two years, HBM bandwidth under 2x πŸ”—
r/LocalLLaMA Right to Intelligence. Protect your right to run local AI. πŸ”—
r/LocalLLaMA Decided to build a game, and test the ceiling of Qwen3.8 27b πŸ”—
r/LocalLLaMA DeepSeek V4.1 Flash beats Astra on AA's new benchmark πŸ”—
r/LocalLLaMA best local model for Japanese translation right now? πŸ”—
13.09.2026
r/LocalLLaMA Trump downplays the need to check AI development and says he doesn't want to cede edge to China -- "I think you have a lot of negative forces that are...bringing up things that won’t happen...whoever wins with AI wins" πŸ”—
r/LocalLLaMA Plot twist πŸ”—
r/LocalLLaMA Aurora1.0-150M Releases! πŸ”—
r/LocalLLaMA Talk me out of buying a 3rd Spark πŸ”—
r/LocalLLaMA 3k$ 128GB VRAM + 256GB RAM DDR4 Server πŸ”—
r/LocalLLaMA Dear 24G owners, try VLLM you might be able to run Qwen3.8 27B INT4, 144K FP8 KV on RTX 3090 with better speed. (TLDR VLLM AOT) πŸ”—
r/LocalLLaMA Trump says not to slow down AI development πŸ”—
r/LocalLLaMA Migration from Claude Code to a private local harness. Questions. πŸ”—
r/LocalLLaMA Hoping for Optimized Smarter Upcoming Models .... Like DeepSeek-V4.1-Flash( KVCache + Engram) in Small/Medium/Big sizes πŸ”—
r/LocalLLaMA Qwen3.8 flash next - untrained svg generation πŸ”—
r/LocalLLaMA Build around cmp 170hx for qwen flash next πŸ”—