Researchers have introduced WanSong, a novel diffusion-based model designed for generating long-form, commercial-grade songs. This model directly produces high-fidelity, multilingual music up to five minutes in length, outputting both vocals and background music in a single pass. WanSong's diffusion framework also facilitates faster inference through step-distillation and allows for efficient fine-tuning for downstream editing tasks. AI
IMPACT Introduces a new approach to AI music generation, potentially enabling more efficient and controllable creation of longer, higher-fidelity songs.
RANK_REASON The item describes a technical report and release of a new model for music generation. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →