qwen4exp fixes in llama.cpp
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
| if you are on Qwen Flash Next make sure to update your build often merged already (by ServeurpersoCom) https://github.com/ggml-org/llama.cpp/pull/27978 https://github.com/ggml-org/llama.cpp/pull/28011 https://github.com/ggml-org/llama.cpp/pull/28023 https://github.com/ggml-org/llama.cpp/pull/28123 merged (by 0cc4m) https://github.com/ggml-org/llama.cpp/pull/28032 in progress (by danielhanchen) MERGED NOW https://github.com/ggml-org/llama.cpp/pull/27941 MTP in progress https://github.com/ggml-org/llama.cpp/pull/27836 more in progress for example [link] [comments] |
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.