CohereLabs has released North-Micro-Vision-Instruct, a 2.4 billion parameter open-weight vision-language model. This model supports native-resolution image processing and is licensed under Apache 2.0, making it suitable for prototyping and specialized multimodal applications. It offers broad image understanding capabilities, including VQA, captioning, and OCR, with support for multiple languages and images. AI
IMPACT Provides a compact, customizable vision-language model for researchers and developers.
RANK_REASON Release of an open-weight model from a research lab. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →