RSS Feed
All
r/AI_Agents
r/AI_Governance
r/ClaudeAI
r/ClaudeCode
r/DeepSeek
r/Futurology
r/LangChain
r/LocalLLaMA
r/MachineLearning
r/OpenAI
r/artificial
r/cybersecurity
r/europe
r/hermesagent
r/netsec
r/singularity
r/technology
1209 items in r/LocalLLaMA
07.09.2026
06.09.2026
r/LocalLLaMA
Trying to create my own server and consuming it for code with my phone remotely (Mac OS)
๐
r/LocalLLaMA
I built an LLM benchmark harness that lets you browse and compare how models answered each question
๐
r/LocalLLaMA
2x R9700, 64 GB DDR5 is an absolute beast machine with vLLM Radiance / R9V and Qwen 3.8 27b and Flash next
๐
r/LocalLLaMA
Qwen 3.8-27B NVFP4 actually beats Q5_K_M and touches official BF16 levels, but only if you change 2 sampling params. Also almost 3x faster.
๐
r/LocalLLaMA
[Model] Support for Spark2_5ForCausalLM implementation by KnightYao ยท Pull Request #27868 ยท ggml-org/llama.cpp
๐
r/LocalLLaMA
Validate your local LLM advertised KV cache against real pressure; see exactly how old contexts get evicted from cache
๐
r/LocalLLaMA
Qwen3.8-Flash-Next-oQ4e-mtp: 45 tok/s on M4 Max, 25 tok/s on M2 Ultra for local inference โ llm-bench.io
๐
05.09.2026