MiniMax AI has released the open weights for its H3 model, which supports multimodal inputs including text, images, and video, and can generate video with audio. The model is now available with day-0 support on platforms like vLLM-Omni and lmsysorg's SGLang Diffusion. This release aims to make the model easier to serve and build upon, running on NVIDIA and AMD hardware, and is positioned as a competitive option in the rapidly evolving open-source model landscape. AI
IMPACT Accelerates open-source multimodal model development and deployment, offering a competitive alternative for video generation tasks.
RANK_REASON Open weights release of a multimodal model from a frontier lab (MiniMax AI). [lever_c_demoted from frontier_release: ic=2 ai=1.0]
- AMD
- H3 Healthcare Three Hop Index
- lmsysorg
- MiniMax AI
- NVIDIA AI
- RTX 6000
- Seedance 2.0
- SGLang Diffusion
- Hugging Face
- OpenAI
- vLLM
- vLLM-Omni
AI-generated summary · Google Gemini · from 2 sources. How we write summaries →