PulseAugur
EN
LIVE 03:42:27

MiniMax H3: Open-weight multimodal video model released

MiniMax has announced H3, a new open-weight multimodal video generation model capable of producing up to 15-second videos with stereo sound at 2K resolution. The model processes unified context across text, images, video, and audio, and is designed for commercial content creation in areas like advertising, branding, and gaming. MiniMax plans to release the model weights soon, aiming to foster the open-source community and improve hardware compatibility, while also highlighting its cost-effectiveness compared to existing closed-source models. AI

IMPACT This release democratizes advanced video generation capabilities, potentially accelerating innovation in content creation and AI tooling.

RANK_REASON Frontier-lab model release with system card [lever_c_demoted from frontier_release: ic=1 ai=1.0]

Read on r/StableDiffusion →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

MiniMax H3: Open-weight multimodal video model released

COVERAGE [1]

  1. r/StableDiffusion TIER_2 English(EN) · /u/Hoodfu ·

    MiniMax H3: Open-weight multimodel video model

    <!-- SC_OFF --><div class="md"><p>Just saw this posted by Fal.ai and then by Hailuo themselves, the next video model will be open weight released! Here's the blurb and link to to the feature post:</p> <p>Today, we're launching MiniMax H3, a general-purpose multimodal generation m…