PulseAugur
EN
LIVE 06:59:57

MiniMax H3 text-to-video model optimized for consumer hardware

A new text-to-video model called MiniMax H3 has been developed, capable of generating 1080p video in approximately 25 seconds. This model is designed to run on consumer hardware, with minimum requirements including a 3060 GPU with 12GB VRAM and 32GB system RAM for 480p video generation. The development team has focused on optimizing the model for efficiency within the ComfyUI framework, and open weights are expected to be released soon. AI

IMPACT This model's optimization for consumer hardware could lower the barrier to entry for AI-powered video generation.

RANK_REASON The cluster describes a new model release with open weights expected, fitting the research category. [lever_c_demoted from research: ic=1 ai=1.0]

Read on r/StableDiffusion →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

MiniMax H3 text-to-video model optimized for consumer hardware

COVERAGE [1]

  1. r/StableDiffusion TIER_2 English(EN) · /u/comfyanonymous ·

    Minimax H3, 1080p 25 seconds, text to video in native ComfyUI (open weights coming soon)

    <table> <tr><td> <a href="https://www.reddit.com/r/StableDiffusion/comments/1vd9o0r/minimax_h3_1080p_25_seconds_text_to_video_in/"> <img alt="Minimax H3, 1080p 25 seconds, text to video in native ComfyUI (open weights coming soon)" src="https://external-preview.redd.it/a3J4YWZ3a3…