CohereLabs/North-Micro-Vision-Instruct · Hugging Face
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
| North Micro Vision Instruct is a 2.4B-parameter open-weight vision-language model with native-resolution image support, released under the Apache 2.0 license. It is designed as a compact foundation for prototyping, task-specific fine-tuning, and specialized multimodal applications. Highlights
Model Details
The language backbone supports a 128K-token context window, but the validated operating range for multimodal prompts is up to 8K tokens. Longer multimodal contexts may rely on extrapolation and have not been benchmarked. Intended UseNorth Micro Vision Instruct is intended for research and development use cases such as:
Limitations
[link] [comments] |
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.