PulseAugur
EN
LIVE 20:05:05

MiniMax releases open-weights music model capable of 5-minute song generation

MiniMax has released MiniMax-Music3, an open-weights model capable of generating complete five-minute songs from text inputs. The model accepts lyrics with section tags and a detailed music description, producing 32 kHz, 16-bit stereo WAV audio. Its architecture combines a Hybrid-LM, consisting of an 8B Global LLM and a 0.6B Local LLM, with a continuous synthesis stack utilizing flow matching and a Flow-VAE. MiniMax has made the weights, inference code, and serving paths publicly available, allowing for commercial use under specific conditions. AI

IMPACT Enables creators and businesses to generate full songs from text, potentially impacting music production, game development, and advertising.

RANK_REASON Open-weights model release from a frontier lab (MiniMax). [lever_c_demoted from frontier_release: ic=1 ai=1.0]

Read on MarkTechPost →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

MiniMax releases open-weights music model capable of 5-minute song generation

COVERAGE [1]

  1. MarkTechPost TIER_1 English(EN) · Asif Razzaq ·

    MiniMax Releases MiniMax-Music3: An Open-Weights Music Model Generating Complete Five-Minute Songs From Lyrics and a Structured Caption

    <p>MiniMax released MiniMax-Music3, an open-weights text-to-music model. Given lyrics with section tags and a structured caption, it generates a complete song of up to five minutes in a single pass, as 32 kHz, 16-bit stereo WAV. Here is the architecture, the three serving paths, …