A new LoRA (Low-Rank Adaptation) model, lightx2v/Minimax-h3-Turbo-SLA, has been released, offering significant improvements in inference speed for the MiniMax-H3 model. This distilled model utilizes Sparse-Linear Attention (SLA) with an 85% sparsity ratio, achieving approximately 2.5x faster inference on an NVIDIA RTX 5090 while maintaining visual quality. The model supports 768p FL2V generation and is compatible with LightX2V and ComfyUI workflows, requiring the original MiniMax-H3 model for operation. AI
IMPACT Accelerates inference for video generation models, potentially enabling more efficient workflows and real-time applications.
RANK_REASON Release of a new LoRA model with performance improvements and technical details.
Read on Hugging Face Trending Models →
- ComfyUI
- Jiang, Kaiwen
- LightX2V
- lightx2v/Minimax-h3-Turbo-SLA
- MiniMaxAI/MiniMax-H3
- MiniMax-H3 Turbo-SLA
- NVIDIA RTX 5090
- Sparse–Linear Attention
- Stoica, Ion
- Wang, Haoxu
- Wang, Ziteng
- Xi, Haocheng
- Zhao, Min
- Zheng, Kaiwen
- Zhu, Hongzhou
- LoRA
AI-generated summary · Google Gemini · from 2 sources. How we write summaries →