r/LocalLLaMA · · 1 min read

Run Qwen3.8+Flash-Next and tiny models on Apple Silicon up to 3x faster

Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.

Maybe you'll like it? I hope I get to use my self-promotion credit a tiny little bit here after being in the community so long haha. I was the top of MLX.fast for a while and remain the winner on chips below M5. If you have capacity to contribute further enhancements I'd love that <3

https://github.com/struffl/ishizuki

submitted by /u/VagabondTruffle
[link] [comments]

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from r/LocalLLaMA