RSS Feed

Day Week Month Year
All r/AI_Agents r/AI_Governance r/ClaudeAI r/ClaudeCode r/DeepSeek r/Futurology r/LangChain r/LocalLLaMA r/MachineLearning r/OpenAI r/artificial r/cybersecurity r/europe r/hermesagent r/netsec r/singularity r/technology

165 items in r/LocalLLaMA

18.09.2026
r/LocalLLaMA Is HF starting to move against abliterated models? πŸ”—
r/LocalLLaMA RTX 5090 Bonsai 2 27B vs Gemma 4 12B vs Qwen 3.5 9B Japanese voxel pagoda πŸ”—
r/LocalLLaMA Prism-ML Bonsai 2 Joins Our Qwen3.8 Quantization Comparison πŸ”—
r/LocalLLaMA inclusionAI/Realtime-Venus Β· Hugging Face πŸ”—
r/LocalLLaMA MiniMax Code goes open source πŸ”—
r/LocalLLaMA China’s mysterious AI company Naive AI is valued at over $1.4 billion and could release its first open-source llm model as early as this month. πŸ”—
r/LocalLLaMA 768gb vram for less than the price of one RTX 6000 πŸ”—
r/LocalLLaMA I just got access to 1.9 TB of ram… πŸ”—
r/LocalLLaMA Question: UkisAI Swift Ternary Bonsai 2 27B? πŸ”—
r/LocalLLaMA bonsai's document reveal how much cherry picked their headlines are πŸ”—
r/LocalLLaMA US government website used AI search tool (Qwen) from China that FBI said copied Anthropic πŸ”—
r/LocalLLaMA Bonsai 1.7B can solve simple physics problems on a 12 watt intel N97 at about 9.1 t/s πŸ”—
r/LocalLLaMA ZCode was allegedly caught uploading workspace/.git records to the cloud. πŸ”—
r/LocalLLaMA JEV architecture πŸ”—
r/LocalLLaMA Qwen 3.8 Next Flash appreciation post πŸ”—
r/LocalLLaMA Made the horizontal open-source model for Jev with RLCD, and it surpasses all the Jev benchmarks. HF space, benchmark, model, repo πŸ”—
r/LocalLLaMA Ternary Bonsai is a headless chicken πŸ”—
r/LocalLLaMA I just realized Courage used local LLMs to solve his problems before any of us ever did. πŸ”—
r/LocalLLaMA still doesn’t get what Jev is…..is it just a more generalised BERT? πŸ”—
17.09.2026
r/LocalLLaMA 600tok/s single request on qwen3.6 35ba3b with Ninfer on an RTX Pro 6000. Anybody remember that Comcast ad "stupid fast"? πŸ”—
r/LocalLLaMA Installing 6 GPUs in a standard case rather than using an open-frame chassis. πŸ”—
r/LocalLLaMA Ternary Bonsai 2 27B πŸ”—
r/LocalLLaMA Ternary Bonsai 2 (27B) just released on Hugging Face. At <6GB in size, it can even run locally in-browser on WebGPU. πŸ”—
r/LocalLLaMA Cactus Needle 3: A Sliceable 8-29MB Automation Foundation Model That Matches DeepSeek v4 Flash πŸ”—
r/LocalLLaMA Thank you :) Swift Qwen 3.8 27B now has 100k+ downloads, is #1 finetune and #9 model on HuggingFace Trending πŸ”—