RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

1212 items in r/LocalLLaMA

30.08.2026
r/LocalLLaMA Qwen 3.8 27B - Fantastic German capabilities ๐Ÿ”—
r/LocalLLaMA Are there any interesting architectural innovations that we seem to be on the verge of for LLM models or AI models that might be a big deal? (Excluding maybe N-gram, since everyone is already well aware of that one) ๐Ÿ”—
r/LocalLLaMA Demo of local document extraction (52 pages) using Arctic Embed and Bonsai 8B on an Iphone 16 (KernelAI app) ๐Ÿ”—
r/LocalLLaMA Will apple still release devices with mobile HbM in 2027 ? ๐Ÿ”—
r/LocalLLaMA NVIDIAยฎ DGX Stationโ„ข Delivering Data-Center-Class Performance from the Desktop ๐Ÿ”—
r/LocalLLaMA Whatever happened to OpenClaw and its derivatives? ๐Ÿ”—
r/LocalLLaMA Here my pretty good qwen3.8 27B setup, hope it helps ๐Ÿ”—
r/LocalLLaMA Qwen 3.8 Flash Next locally on simple mobile phone at 3.5 tok/s ๐Ÿ”—
r/LocalLLaMA Unpopular opinion Qwen 3.8 is hard to understand ๐Ÿ”—
r/LocalLLaMA Qwen3.8-Flash-Next turns 4xR9700 into a local AI powerhouse! 120 t/s TG and 12k t/s PP single request with optimized vLLM ๐Ÿ”—
r/LocalLLaMA Oh so that's where my PCIe lanes went... ๐Ÿ”—
r/LocalLLaMA Uncensored Multi-Model Releases, LongCat-Flash-Lite-Sparse with MTPs and LSAs, Qwen3.8-27B with MTPs, Qwen3.5-122B-A10B with MTPs, Qwen3-Coder-Next and Laguna-S2.1 with Vision, All Available in GGUF Format! Bonus: Links to my llama.cpp Fork for LongCat-Flash-Lite Support and J-Wash Enhanced Fork! ๐Ÿ”—
r/LocalLLaMA Don't Sleep on EXL3 Quants ๐Ÿ”—
r/LocalLLaMA my sincere condolences to all information security professionals, the following few years before the world war will be tough ๐Ÿ”—
r/LocalLLaMA Hit the jackpot girls and boys! ๐Ÿ”—
r/LocalLLaMA Me these days ๐Ÿ”—
r/LocalLLaMA I fine-tuned a 0.8B local model for dictation cleanup. It matched a hosted frontier model on this narrow task ๐Ÿ”—
r/LocalLLaMA Experience report - Qwen 3.8 Flash Next on memory rich, GPU poor setup ๐Ÿ”—
r/LocalLLaMA When you say, because I can. Limits of X870e ๐Ÿ”—
r/LocalLLaMA Some people said the Minecraft clone I fully vibecoded with Qwen3.8-27B Q4 is not that impressive because Minecraft is in the training data, so I had the model add 4 things that are probably not. ๐Ÿ”—
r/LocalLLaMA Koboldcpp v1.120 released ๐Ÿ”—
r/LocalLLaMA Qwen3.8-Flash-Next NVFP4 Day-3 support for 4xV100 ๐Ÿ”—
r/LocalLLaMA Qwen3.8-Flash-Next optimised for Macs ๐Ÿ”—
r/LocalLLaMA It's official! 192GB Framework ๐Ÿ”—
r/LocalLLaMA an unscientific qwen 3.8 flash next and glm 5.3 flash comparison ๐Ÿ”—