PulseAugur
EN
LIVE 02:08:16
ENTITY CuTe-DSL

CuTe-DSL

PulseAugur coverage of CuTe-DSL — every cluster mentioning CuTe-DSL across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
3
3 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
0 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

1 day(s) with sentiment data

RECENT · PAGE 1/1 · 5 TOTAL
  1. SIGNIFICANT · CL_246698 ·

    Together optimizes inference on Nvidia Rubin GPU with new features · 4 sources tracked

    Together has announced advancements in their inference and OSS capabilities, leveraging new features from Nvidia's Rubin GPU. The company has updated its b200 gemms to incorporate Rubin's wider MMA steps, increased tens…

  2. TOOL · CL_247090 ·

    Together AI optimizes ThunderKittens for NVIDIA Vera Rubin Blackwell GPUs

    Together AI has gained access to NVIDIA's Vera Rubin NVL72 platform, which is based on the Blackwell architecture. Their team has updated their ThunderKittens software to leverage new features of the Vera Rubin chip, sp…

  3. RESEARCH · CL_104070 ·

    GB200 NVL72 serving costs slashed 2.5x via software upgrades

    Software optimizations for the GB200 NVL72 have drastically reduced serving costs by 2.5 times in under 70 days. These improvements, particularly the rewriting of the NVFP4 MoE kernel using CuTe-DSL and leveraging the N…

  4. TOOL · CL_86322 ·

    Modal optimizes FlashAttention-4 for faster LLM inference

    Modal has enhanced the FlashAttention-4 kernel to improve inference speed for large language models, particularly for decode-heavy workloads. Their contributions focused on adjusting parallelism strategies, such as shif…

  5. RESEARCH · CL_18472 ·

    NVIDIA open-sources cuDNN kernels after 12 years, including MoE and sparse attention

    NVIDIA has open-sourced parts of its cuDNN library, a significant move after 12 years of it being closed-source. This release includes over 20 Mixture-of-Experts (MoE) kernels and NSA sparse attention kernels. The codeb…