r/LocalLLaMA · · 1 min read

Alright, We got Qwen3.8-27B. Now it's community's turn to make it more better & faster

Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.

Facing any issues?

Chat Template is fine?

Looping issue?

Too much reasoning thing?

How's MTP with this one?

Any other issues faced by Qwen3.6-27B & Qwen3.5-27B during release time?

If I missed any other items, please mention in your comments.

AND

  1. Share comparison with Qwen3.6-27B. On Memory & t/s stats
  2. How much memory takes for this model if you use full 256K context + unquantized KVCache + MTP? For Q4 & above quants. Particularly Q8 please, want to know it's possible to hold this in 32GB VRAM. Also share t/s stats.
  3. How good is this model on Creative writing? Better than Qwen3.6-27B?
submitted by /u/pmttyji
[link] [comments]

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from r/LocalLLaMA