RSS Feed
All
r/AI_Agents
r/AI_Governance
r/ClaudeAI
r/ClaudeCode
r/DeepSeek
r/Futurology
r/LangChain
r/LocalLLaMA
r/MachineLearning
r/OpenAI
r/artificial
r/cybersecurity
r/europe
r/hermesagent
r/netsec
r/singularity
r/technology
1205 items in r/LocalLLaMA
15.09.2026
r/LocalLLaMA
ByteShape Qwen 3.8 27B: To KL Diverge or Not to KL Diverge, Part 2: Metric Boogaloo
๐
r/LocalLLaMA
Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NEO-CODER-MAX-MTP-GGUF
๐
r/LocalLLaMA
CrofAI "cheapest inference provider in the world" gets exposed as an OpenRouter wrapper, routing requests to smaller, cheaper models at up to 20x markup. CrofAI responds to Wire Fraud allegations by denying everything, then backtracking, then 3 hours later wiping their entire online presence
๐
14.09.2026
r/LocalLLaMA
Nvidia's RTX 5090 vanishes from online retail in the US โ third-party sellers now demand as much as $9,500 for Nvidia's fastest GPU
๐
r/LocalLLaMA
China Wants to Build a BRICS-Wide AI Cloud. That Might Be a Much Bigger Deal Than the Open Models
๐
r/LocalLLaMA
For the GPU poor. K2 Horizon 7B ranks between qwen 3.6 27B and qwen 3.6 35BA3b on the Artificial Analysis Intelligence Index.
๐
r/LocalLLaMA
UkisAI Swift-Qwen3.8-27B / -58.3% thinking, x1.95 speed while keeping the accuracy of xhigh
๐
r/LocalLLaMA
llama: add Maple 20B-A1B ternary MoE architecture (CPU) by AlexGabbia ยท Pull Request #27000 ยท ggml-org/llama.cpp
๐