Tencent Hunyuan and Fudan University have released Prism, a new AI model capable of generating native 2K video and audio simultaneously. The model utilizes a novel dynamic sparse attention mechanism, reportedly enabling 2.5x faster training with improved quality. Prism is released under an MIT license, with training and fine-tuning code available, though higher resolutions require multiple high-end GPUs. AI
IMPACT Sets a new standard for integrated audio-visual AI generation with an open-source license, potentially accelerating research and development in multimedia AI.
RANK_REASON Frontier-lab model release with system card. [lever_c_demoted from frontier_release: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →