MiniMax Music 3, a new open-weight model capable of generating complete songs up to five minutes long, has been released. This model utilizes a hybrid approach, combining an 8B Global LLM initialized from Qwen3-8B for long-range structure and a 0.6B Local LLM for fine-grained acoustic detail. It can be conditioned on lyrics and detailed music descriptions, producing 32 kHz, 16-bit stereo WAV audio. The model is available for local use via SGLang-Omni and diffusers, with previews and demos accessible through Hugging Face and community projects like audio.cpp. AI
IMPACT Enables creation of longer, more coherent musical pieces with fine-grained control, potentially impacting music production and AI-generated content.
RANK_REASON Model release from MiniMaxAI, a recognized AI lab.
Read on Mastodon — fosstodon.org →
- diffusers
- Flow Matching VAE
- MiniMax-Music3
- Qwen3-8B-LLM
- SGLang-Omni
- Flow Matching
- Flow-VAE
- MiniMaxAI/MiniMax-Music3
- MiniMax Music 3
- Qwen3-8B
- ComfyUI
- Hugging Face
- StableDiffusion
- audio.cpp
- MiniMax H3
AI-generated summary · Google Gemini · from 8 sources. How we write summaries →