AA is the reason for Qwen3.8 27B shipped with xhigh
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
I know why Qwen3.8 27B shipped with xhigh reasoning as default, it's to do its best in benchmarks. Models from top labs often get benchmarked at multiple reasoning levels, but that same treatment doesn't apply to other labs. Open models are lucky to even be benchmarked at all. (See Laguna S 2.1)
So it makes total sense that Qwen team decided to ship with a default that show the model at its maximum capabilities, assuming Artificial Analysis would benchmark at the default.
And before anyone accuses Qwen team, I don't think it's benchmaxxing. That is an actual toggle that you can use if you have high bandwidth (or tolerence), and variable reasoning is pretty standard across the board. Totally reasonable to default to your best if you think you have one shot.
[link] [comments]
More from r/LocalLLaMA
-
Expert expansion with llama.cpp
Sep 6
-
Maybe a noob question, but why aren't there safetensor quants of inclusionAI/Ling-3.0-flash-Fin?
Sep 6
-
2x R9700, 64 GB DDR5 is an absolute beast machine with vLLM Radiance / R9V and Qwen 3.8 27b and Flash next
Sep 6
-
Planning to get a cheap-ish GPU. Would appreciate some advice.
Sep 6
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.