Best choice of model 40B+ Parameters
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
currently using Qwen3.6 35B as my main assistant model + coding agent
but I think sometimes it misses basical general knowledge things, and it is more like executioner that assistant.
That's why I though should I go with bigger models, But I don't want to lose speed
I am on Strix Halo
Having 30-40 t/s roughly on 131k context
Thinking to switch on Qwen3.5 122B
Any Suggestions?
[link] [comments]
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.