PulseAugur
实时 19:36:58
English(EN) Cursor Open-Sources Mixture-of-Kittens (MoK): A Deterministic MoE Training Megakernel for GB300 NVL72 Racks

Cursor 开源 MoK 巨型内核以加速 MoE 模型训练

Cursor Research 已开源 Mixture-of-Kittens (MoK),这是一种专门用于优化 Mixture-of-Experts (MoE) 模型的训练内核。该巨型内核将通信和计算步骤融合到单一的确定性过程中,据报道与现有基线相比,吞吐量提高了 2.37 倍。MoK 专为高端硬件设计,特别需要 NVIDIA Blackwell GB200 NVL72 或 GB300 NVL72 机架,使其适用于大规模 AI 模型开发和云基础设施提供商。 AI

影响 使拥有高端 GPU 基础设施的组织能够更快地训练大型 MoE 模型。

排序理由 Cursor Research 开源了一个专门用于 MoE 模型的训练内核 (MoK),这是一个软件工具,而不是新的前沿模型发布。

在 MarkTechPost 阅读 →

AI 生成摘要 · Google Gemini · 来自 2 个来源。 我们如何撰写摘要 →

Cursor 开源 MoK 巨型内核以加速 MoE 模型训练

报道来源 [2]

  1. MarkTechPost TIER_1 English(EN) · Asif Razzaq ·

    Cursor Open-Sources Mixture-of-Kittens (MoK): A Deterministic MoE Training Megakernel for GB300 NVL72 Racks

    <p>Cursor Research has open-sourced Mixture-of-Kittens (MoK), the MoE training megakernel behind its Composer models. MoK fuses all mixture-of-experts communication and computation into a single deterministic kernel, and runs up to 2.37x faster than the strongest public baseline …

  2. r/LocalLLaMA TIER_1 English(EN) · /u/CapnHat ·

    Cursor releases their Mixture-of-Kittens megakernel for training MoE models - Claims to nearly double TFLOP/s

    <!-- SC_OFF --><div class="md"><p>Link: <a href="https://cursor.com/blog/mixture-of-kittens">https://cursor.com/blog/mixture-of-kittens</a></p> <p>GitHub: <a href="https://github.com/cursor/mixture-of-kittens">https://github.com/cursor/mixture-of-kittens</a></p> <p>Seems like a n…