r/LocalLLaMA · · 1 min read

Best choice of model 40B+ Parameters

Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.

currently using Qwen3.6 35B as my main assistant model + coding agent

but I think sometimes it misses basical general knowledge things, and it is more like executioner that assistant.

That's why I though should I go with bigger models, But I don't want to lose speed

I am on Strix Halo

Having 30-40 t/s roughly on 131k context

Thinking to switch on Qwen3.5 122B

Any Suggestions?

submitted by /u/FeiX7
[link] [comments]

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from r/LocalLLaMA