The weights for the comfy MiniMax-H3 model have been released, offering two distinct variants for video generation. The H3-Base-FL2VA model supports text-to-video generation with options for using one or two input images to guide the output. The H3-Base-Ref2VA model provides an omni-reference mode, capable of processing multiple images, video clips, and audio clips as input to generate video content. AI
IMPACT Provides new tools for video generation with flexible input options.
RANK_REASON Release of weights for a specific AI model. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →