RSS Feed
All
r/AI_Agents
r/AI_Governance
r/ClaudeAI
r/ClaudeCode
r/DeepSeek
r/Futurology
r/LangChain
r/LocalLLaMA
r/MachineLearning
r/OpenAI
r/artificial
r/cybersecurity
r/europe
r/hermesagent
r/netsec
r/singularity
r/technology
1216 items in r/LocalLLaMA
22.08.2026
r/LocalLLaMA
GLM and I created a llama.cpp fork optimized for AMD GFX906 (Mi50, Mi60, Radeon VII, GCN HIP) - Machine Learning, LLMs, & AI
๐
r/LocalLLaMA
Single RTX 5090: Qwen3.8-27B NVFP4 at a real 262K context in vLLM โ 77 tok/s short-context, 64.7 tok/s at 128K
๐
r/LocalLLaMA
I forked Ninfer 3090 and converted it to run on the CMP170HX - doubled my Qwen3.6-35B from llama.cpp
๐
21.08.2026
r/LocalLLaMA
I tried to do agenic coding with Qwen 3.8 27B 3bit quant on a macbook air m2 24gb. It took 63 hours, but amazingly, the flight simulator worked.
๐
r/LocalLLaMA
BEAST!! Unlocked a CMP 170HX: 6.3 โ 193 TFLOPS tensor, pp512 599 โ 3468, and half the community's advice about this card is wrong.
๐
r/LocalLLaMA
I benchmarked Ox Alpha on SWE-bench Verified Mini (50 tasks): 96% resolved. Now Iโm skeptical of myself.
๐