unsloth/Qwen3.8-Flash-Next-GGUF is being updated
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
Looks like Unsloth is updating https://huggingface.co/unsloth/Qwen3.8-Flash-Next-GGUF to work with llama.cpp, so hopefully this will resolve the issue of having two different GGUF versions of Qwen Flash Next.
[link] [comments]
More from r/LocalLLaMA
-
Microsoft confirms OpenAI has been using Looped Transformers in the GPT-6 series
Oct 6
-
Tencent releases Octop, a self-hosted AI assistant
Oct 6
-
Mistral CEO says new AI model beats Chinese ones in some areas
Oct 6
-
We’re using GLM-5.3 Flash instead of frontier models on a massive production codebase
Oct 6
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.