IFM/K2-Horizon-MoVA-36B-A4B-GGUF · Hugging Face
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
| more sizes (probably still uploading): https://huggingface.co/IFM/K2-Horizon-32B-GGUF https://huggingface.co/IFM/K2-Horizon-7B-GGUF https://huggingface.co/IFM/K2-Horizon-3.7B-GGUF https://huggingface.co/IFM/K2-Horizon-0.9B-GGUF from IFM: K2-Horizon-MoVA-36B-A4B is the sparse member of the K2-Horizon family: a Mixture-of-Experts model with Mixture-of-Values attention (MoVA) that stores 36B parameters and runs 4B per token. We have released the final checkpoint; intermediate checkpoints, along with the data and the training code, will be released. K2-Horizon-MoVA-36B-A4B Highlights
collection: https://huggingface.co/collections/IFM/k2-horizon [link] [comments] |
More from r/LocalLLaMA
-
GGUFs in transformers natively!
Sep 23
-
Pirate Face - pirate bay for LLMs
Sep 23
-
DeepSeek and Moonshot AI face Beijing's probe over potential data leaks to Anthropic
Sep 23
-
Nathan Lambert's written Congressional testimony on the state of open models - Chinese open-weight downloads now 2x America's, >80% of OpenRouter open-model usage
Sep 23
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.