Cursor Research has open-sourced Mixture-of-Kittens (MoK), a specialized training kernel designed for Mixture-of-Experts (MoE) models. This megakernel fuses MoE communication and computation into a single deterministic process, reportedly achieving up to 2.37x higher throughput compared to existing baselines. MoK is optimized for high-end hardware like NVIDIA GB200 NVL72 racks, making it suitable for large-scale AI model development and training. AI
IMPACT Accelerates training for large MoE models on high-end hardware, potentially speeding up development cycles for advanced AI.
RANK_REASON Cursor Research open-sourced a specialized training kernel (Mixture-of-Kittens) for MoE models, which is a software tool rather than a new frontier model release.
- Cursor
- Mixture-of-Kittens
- Composer
- CUDA
- Cursor Research
- DeepSeek-V3
- GB300 NVL72
- Kimi 2.5
- NVIDIA Blackwell
- PyTorch
- Apache-2.0
- Composer 2.5
- Composer models
- CUDA toolkit 13.0+
- DeepSeek-V4-Pro
- GLM-5.2
- Mixture-of-Experts (MoE)
- Mixture-of-Kittens (MoK)
- NVIDIA Blackwell SM100
- NVIDIA Blackwell SM103
- NVIDIA GB200 NVL72
- NVL72
- Python 3.12+
- PyTorch 2.10+
- Qwen3.5-397B-A17B
AI-generated summary · Google Gemini · from 5 sources. How we write summaries →