r/LocalLLaMA · · 1 min read

Qwen 3.8 27B — MTP or DFlash?

Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.

Do we.know whether the 27B model will ship with a DFlash or MTP head? It's super exciting, but since 35B-A3B is my daily driver, 27B will crawl — still excited for it though!

I think 3.6 27B with MTP was about 8 tok/s for me (32GB unified memory, 780M)

submitted by /u/mailto_devnull
[link] [comments]

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from r/LocalLLaMA