IFM/K2-Horizon-MoVA-36B-A4B-GGUF · Hugging Face
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
| more sizes (probably still uploading): https://huggingface.co/IFM/K2-Horizon-32B-GGUF https://huggingface.co/IFM/K2-Horizon-7B-GGUF https://huggingface.co/IFM/K2-Horizon-3.7B-GGUF https://huggingface.co/IFM/K2-Horizon-0.9B-GGUF from IFM: K2-Horizon-MoVA-36B-A4B is the sparse member of the K2-Horizon family: a Mixture-of-Experts model with Mixture-of-Values attention (MoVA) that stores 36B parameters and runs 4B per token. We have released the final checkpoint; intermediate checkpoints, along with the data and the training code, will be released. K2-Horizon-MoVA-36B-A4B Highlights
collection: https://huggingface.co/collections/IFM/k2-horizon [link] [comments] |
More from r/LocalLLaMA
-
Apple A20 Pro debuts with 7-core GPU, 32-core Neural Engine and 50% more memory bandwidth (~115 GB/s)
Sep 9
-
Surveillance plagiarism by OpenAI
Sep 9
-
Don't let FOMO win if you're interested in local llm from a hobby/learning aspect
Sep 9
-
Server rebuild to custom loop. 2x RTX Titans 24gb, 1x 22gb 2080ti | T: 70GB VRAM.
Sep 9
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.