PulseAugur
EN
LIVE 18:56:36
ENTITY Kyūtai

Kyūtai

PulseAugur coverage of Kyūtai — every cluster mentioning Kyūtai across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
2
8 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
2 over 90d
TIER MIX · 90D
TOPICS
RELATIONSHIPS
RECENT · PAGE 1/1 · 8 TOTAL
  1. TOOL · CL_136789 ·

    Kyutai releases open-weight MuScriptor for multi-instrument music transcription

    Kyutai, a French AI lab, has introduced MuScriptor, an open-weight transformer model designed for transcribing multi-instrument music into MIDI format. The model is available in three sizes, with the largest having 1.4 …

  2. SIGNIFICANT · CL_134539 ·

    AI voice startup Gradium raises $100M seed, backed by Nvidia

    Paris-based AI voice startup Gradium has secured $100 million in a seed funding round, with Nvidia participating as a new investor. This funding will support the company's expansion into the Bay Area to compete for tale…

  3. TOOL · CL_144505 ·

    MuScriptor: Open-weight model transcribes multi-instrument music to MIDI

    Researchers have introduced MuScriptor, an open-weight model designed for multi-instrument music transcription. Unlike previous models that struggled with complex mixes or single instruments, MuScriptor can transcribe n…

  4. RESEARCH · CL_129929 ·

    MIRA AI model trained on Rocket League data released with demo

    A new AI model called MIRA has been released, designed for multiplayer interactive world modeling and trained on data from the game Rocket League. Developed through a collaboration involving General Intuition, Kyutai, a…

  5. TOOL · CL_127886 ·

    CPU TTS benchmark: Pocket TTS shows flat RTF, UTMOS struggles with naturalness

    A benchmark comparing four CPU-based text-to-speech (TTS) models—Kokoro, Supertonic, Inflect-Nano, and Pocket TTS—reveals distinct performance characteristics. Pocket TTS, utilizing a streaming language model architectu…

  6. TOOL · CL_127764 ·

    Kyutai's Pocket TTS offers CPU-based voice cloning from 5s audio

    Kyutai has released Pocket TTS, a ~100M parameter streaming language model that generates audio tokens autoregressively. This model is notable for its ability to perform zero-shot voice cloning from just 5 seconds of au…

  7. TOOL · CL_146936 ·

    New open-weight MuScriptor model transcribes multi-instrument music

    A new open-weight model named MuScriptor has been released for automatic music transcription, capable of converting audio recordings with multiple instruments into detailed note sequences. Developed through a collaborat…

  8. RESEARCH · CL_13577 ·

    Sakana AI's KAME architecture injects LLM knowledge into speech AI without latency

    Sakana AI has developed KAME, a novel tandem architecture for speech-to-speech AI that aims to combine the speed of direct systems with the knowledge depth of LLM-based approaches. KAME operates with two asynchronous co…