r/LocalLLaMA · · 1 min read

Qwen3.5 122B-A10B · ROCmFP4 iMatrix

Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.

Qwen3.5 122B-A10B · ROCmFP4 iMatrix

Hola Strix and AMD stacker frendios. Read the Lineage and Credits, this uses charlie12345/ROCmFPX, won't work on native llama.cpp yet.

122B total · 10B active · 60.70 GiB · 28.50 tok/s MTP-off · BF16 KLD 0.041366 · Decode 28.505

Decode speed + 36.89% faster

Size - 13.47gb smaller

submitted by /u/RedParaglider
[link] [comments]

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from r/LocalLLaMA