RSS Feed
All
r/AI_Agents
r/AI_Governance
r/ClaudeAI
r/ClaudeCode
r/DeepSeek
r/Futurology
r/LangChain
r/LocalLLaMA
r/MachineLearning
r/OpenAI
r/artificial
r/cybersecurity
r/europe
r/hermesagent
r/netsec
r/singularity
r/technology
1211 items in r/LocalLLaMA
01.09.2026
31.08.2026
r/LocalLLaMA
AVX2: Speed up large batch size prompt processing of IQ models by bartowski1182 ยท Pull Request #27402 ยท ggml-org/llama.cpp
๐
r/LocalLLaMA
GLM 5.3 and GLM 5.3 Flash ran locally on RTX PRO 6000 WS and built a penthouse using BlenderMCP
๐
r/LocalLLaMA
SlopTV: an infinite livestream of AI slop generated from youtube chat comments, Minimax H3 on 2x5090
๐
r/LocalLLaMA
The Chrono Trigger plot challenge - Crono awakens in his modest bedroom of 2095...
๐
r/LocalLLaMA
How bad do you think models like Qwen3.8-27B or GLM-5.3-Flash would be with H-Neurons disabled?
๐
r/LocalLLaMA
CUDA: extend MOE fusion to specdec, earlier MOE glu fusion and topk-router fusion were restricted to 1 token by ynankani ยท Pull Request #27621 ยท ggml-org/llama.cpp
๐
30.08.2026