TaichuAI has released ZDTaichu5.0-9B, an open-source multimodal model with approximately 9 billion parameters. This model is capable of processing images and videos at any resolution and has demonstrated strong performance on spatial and agent-based tasks. It achieved scores such as 62.5 on ViewSpatial, 47.2 on MMSI, 78.3 on MindCube-tiny, 87.7 on TAU2, and 71.4 on Claw-Eval. AI
IMPACT Provides a new open-source option for multimodal AI research and development, particularly for tasks involving spatial understanding and agentic capabilities.
RANK_REASON Release of a new open-source multimodal model with reported benchmark scores. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →