r/LocalLLaMA
500 articles archived · Visit source ↗ · RSS
-
-
-
-
-
-
r/LocalLLaMA community 9d ago
Intern S2 Mobius
A Qwen3.5-35B derived model with an interesting architectural difference that results in larger throughput and less token consumption (allegedly): https://huggingface.co/internlm/Intern-S2-Mobius   submitted by   /u/Miserable-Dare5090 [link]   [comments]
7 -
-
r/LocalLLaMA community 9d ago
China’s Open-Weight Models Will Be Spared US Safety Tests
  submitted by   /u/fallingdowndizzyvr [link]   [comments]
36 -
-
r/LocalLLaMA community 9d ago
Maple-Preview: 20B-A1B ternary-weight reasoning open-weight LLM
  submitted by   /u/cafedude [link]   [comments]
30 -
-
-
-
-
r/LocalLLaMA community 9d ago
Introducing Shieldstral. | Mistral AI
  submitted by   /u/tengo_harambe [link]   [comments]
28 -
-
-
r/LocalLLaMA community 9d ago
LFM2.5-2.6B is out
Released today, with emphasis on agentic capabilities. I really like their models for simple, high volume tasks ("summarize these gazillion documents") and their 8b-a1b was my go-to for certain tasks so I'm excited to see how this one performs. There's not enough love for tiny…
18 -
-
-
-
-
-
-
r/LocalLLaMA community 10d ago
More Qwen 3.8 sizes coming
  submitted by   /u/appakaradi [link]   [comments]
14 -
-
r/LocalLLaMA community 10d ago
nvidia/NVIDIA-NemotronLabs-VoiceChat-11B · Hugging Face (full duplex)
  submitted by   /u/adefa [link]   [comments]
9 -
r/LocalLLaMA community 10d ago
Only 3 days ago...
  submitted by   /u/Fun_Librarian_7699 [link]   [comments]
22 -