r/LocalLLaMA
500 articles archived · Visit source ↗ · RSS
-
-
-
-
-
-
-
-
r/LocalLLaMA community 11d ago
Quantization hurts knowledge nonlinearly - Qwen3.6 27B case study
  submitted by   /u/pmigdal [link]   [comments]
35 -
r/LocalLLaMA community 11d ago
70-class VRAM stagnation
been thinking about how the desktop 70-class has sat at 12GB for two generations now, 4070, 4070 super, 5070, all 12GB. the 1070 gave you 8GB back in 2016 and it felt generous for the price. ten years later the jump is... 4GB. and the thing is these chips arent even weak. the…
32 -
-
-
-
-
r/LocalLLaMA community 11d ago
[RELEASE] SupraBrain-50M-v0.1
Hey there! So today we're releasing SupraBrain-50M, a hybrid language model that combines Gated DeltaNet linear recurrence with Sliding-Window Attention and Surprise-Gated update mechanisms to deliver very strong performance. Here are the benchmarks:…
19 -
-
-
r/LocalLLaMA community 11d ago
Seedance 2.5 Vs Minimax H3 (Open Weight). Excellent Output Comparison!
  submitted by   /u/Hannibalj2ca [link]   [comments]
23 -
r/LocalLLaMA community 11d ago
Open weight has made to frontier
Am looking forward to this! Open weight has come near frontier for 5x less the cost per/M tokens on task completion Open weight ranking Kimi k3 Qwen 3.8 GLM 5.2 Deepseek v4 flash 07/31   submitted by   /u/Specialized-Trap404 [link]   [comments]
4 -
-
r/LocalLLaMA community 11d ago
Qwen3.8-27B announced alongside Qwen3.8-Max
https://preview.redd.it/gy0tgokdl2hh1.png?width=540&format=png&auto=webp&s=7db9e034613a915cb33d378b99ad72c31c7cc18f source: https://x.com/Alibaba_Qwen/status/2084100707423289643   submitted by   /u/TKGaming_11 [link]   [comments]
29 -
r/LocalLLaMA community 11d ago
Qwen 3.8 is live now.
Update: And yes, Qwen3.8-27B is coming too. Next week! 2.4-trillion-parameter MoE flagship delivering a comprehensive leap in coding and professional work. Autonomously codes and delivers complete projects spanning 10+ days. Open weights coming soon! It is live at…
19 -
-
r/LocalLLaMA community 11d ago
Anyone Used MiniMAx H3 yet? Open Weights are out today!
I am curious if anyone have used it. I would love to feed it key frames and test if it can create in-between frames between my keys. Anyone have tried it, any thoughts?   submitted by   /u/Hannibalj2ca [link]   [comments]
14 -
-
-
-
-
-
-
-
-
-
r/LocalLLaMA community 12d ago
Has Qwen 3.8 has dropped yet? Day 90...
  submitted by   /u/mailto_devnull [link]   [comments]
32 -
r/LocalLLaMA community 12d ago
llama.cpp just added MTP / DSpark support for DeepSeek V4 Flash
  submitted by   /u/rmhubbert [link]   [comments]
21 -
r/LocalLLaMA community 12d ago
Vacuum 16T
https://huggingface.co/tsfrm/vacuum-16t A 16.5-trillion-parameter model that contains nothing. This model is just a ████ you to the labs and companies who say that "haha I have the biggest model out there!". We the people with shitty laptops want to get a record. And I now have…
4 -
r/LocalLLaMA community 12d ago
Xberg v1 is out
Hi all, I'm happy to announce that Xberg v1 is out. Xberg is the successor to Kreuzberg, equivalent to what would have been Kreuzberg v5. It's a content intelligence framework that handles a very wide range of inputs: documents (currently 101 formats), code and data formats…
16