RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

1253 items in r/LocalLLaMA

29.07.2026
r/LocalLLaMA I tried running a 1.56TB MoE model on a 6GB RTX 4050 Laptop, Here’s the result πŸ”—
r/LocalLLaMA Nvidia is expected to raise GeForce RTX GPU prices again by up to 30% πŸ”—
28.07.2026
r/LocalLLaMA Zuck's opinion: The AI Future Is for Everyone πŸ”—
r/LocalLLaMA SK Hynix stock fell some 40% in the last 30 days, finally cheap RAM and GPUs again? πŸ”—
r/LocalLLaMA I got Kimi-k3 running..... πŸ”—
r/LocalLLaMA Unsloth has begun dropping Kimi K3 GGUFs. The MXFP4 (it's 1.5 TB) and mmproj are already there. πŸ”—
r/LocalLLaMA Now, this: 1,100 current/former frontier-AI employees sign a petition calling for US gov't to step in for "pacing" frontier development πŸ”—
r/LocalLLaMA Should we be calling Elon a liar? πŸ”—
r/LocalLLaMA microsoft/Mage-VL Β· Hugging Face - An Efficient Codec-Native Streaming Multimodal Foundation Model πŸ”—
r/LocalLLaMA White-hat hacking IS the defense to black-hat hacking. The techniques are the same. How does Dario expect companies to do it if their models refuse? πŸ”—
r/LocalLLaMA Appreciation for Gemma 4 26b A4b πŸ”—
r/LocalLLaMA A 5B-active model doesn't know much, and I've stopped counting that as a flaw πŸ”—
r/LocalLLaMA SWE-rebench Multilingual Update (Go, Java, Python, Rust, TS). Evaluated: GLM-5.2, DeepSeek-V4 Pro, Qwen3.6-27B and others πŸ”—
r/LocalLLaMA What would it take for the frontier labs to open the weights of their old, deprecated proprietary models? πŸ”—
r/LocalLLaMA Gemini Distillation Service πŸ”—
r/LocalLLaMA DeepSeek V4 Flash, up to 32 tok/s on AMD Ryzen AI MAX+ 395 πŸ”—
r/LocalLLaMA CohereLabs/North-Mini-Code-1.0-eagle Β· Hugging Face πŸ”—
r/LocalLLaMA China Al open weight model will burst the US Al bubble market soon πŸ”—
r/LocalLLaMA spec: add DSpark speculative decoding by wjinxu Β· Pull Request #25173 Β· ggml-org/llama.cpp πŸ”—
r/LocalLLaMA Sorry, but did Dario just say that closed-weights, in-secret models are worse than open-weights ones? πŸ”—
r/LocalLLaMA Medical model: Reasoning-Medical-27B (Qwen3.6-27B finetune) πŸ”—
r/LocalLLaMA This came to mind about the current situation. No local models mean no middle-men who re-sell re-branded models, which means lower hardware demand. πŸ”—
r/LocalLLaMA First evidence of a pending qwen3.7 open weights release. Qwen3.7-flash is on open router. They referred to Qwen3.6-35b-a3b as Qwen3.6 flash so this is likely a small MoE. The prices are substantially cheaper than 3.6 flash with a native 1M context window. πŸ”—
27.07.2026
r/LocalLLaMA A user has managed to run Kimi K3 on 80xRTX 5090, via 25GbE Ethernet. πŸ”—
r/LocalLLaMA Anthropic is calling for a ban on open-weights models by proposing mandatory requirements they will probably never be able to meet πŸ”—