PulseAugur
EN
LIVE 21:03:25
ENTITY Kyūtai

Kyūtai

PulseAugur coverage of Kyūtai — every cluster mentioning Kyūtai across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
1
8 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
1 over 90d
TIER MIX · 90D
TOPICS
RELATIONSHIPS
SENTIMENT · 30D

1 day(s) with sentiment data

RECENT · PAGE 1/1 · 9 TOTAL
  1. TOOL · CL_243085 ·

    Gradium launches AI voice generator from text prompts

    Gradium, a voice AI company, has launched Voice Design, a new tool that generates synthetic voices from text descriptions. Unlike traditional voice cloning, Voice Design does not require reference audio or speaker conse…

  2. TOOL · CL_136789 ·

    Kyutai releases open-weight MuScriptor for multi-instrument music transcription

    Kyutai, a French AI lab, has introduced MuScriptor, an open-weight transformer model designed for transcribing multi-instrument music into MIDI format. The model is available in three sizes, with the largest having 1.4 …

  3. SIGNIFICANT · CL_134539 ·

    AI voice startup Gradium raises $100M seed, backed by Nvidia

    Paris-based AI voice startup Gradium has secured $100 million in a seed funding round, with Nvidia participating as a new investor. This funding will support the company's expansion into the Bay Area to compete for tale…

  4. TOOL · CL_144505 ·

    MuScriptor: Open-weight model transcribes multi-instrument music to MIDI

    Researchers have introduced MuScriptor, an open-weight model designed for multi-instrument music transcription. Unlike previous models that struggled with complex mixes or single instruments, MuScriptor can transcribe n…

  5. RESEARCH · CL_129929 ·

    MIRA AI model trained on Rocket League data released with demo

    A new AI model called MIRA has been released, designed for multiplayer interactive world modeling and trained on data from the game Rocket League. Developed through a collaboration involving General Intuition, Kyutai, a…

  6. TOOL · CL_127886 ·

    CPU TTS benchmark: Pocket TTS shows flat RTF, UTMOS struggles with naturalness

    A benchmark comparing four CPU-based text-to-speech (TTS) models—Kokoro, Supertonic, Inflect-Nano, and Pocket TTS—reveals distinct performance characteristics. Pocket TTS, utilizing a streaming language model architectu…

  7. TOOL · CL_127764 ·

    Kyutai's Pocket TTS offers CPU-based voice cloning from 5s audio

    Kyutai has released Pocket TTS, a ~100M parameter streaming language model that generates audio tokens autoregressively. This model is notable for its ability to perform zero-shot voice cloning from just 5 seconds of au…

  8. TOOL · CL_146936 ·

    New open-weight MuScriptor model transcribes multi-instrument music

    A new open-weight model named MuScriptor has been released for automatic music transcription, capable of converting audio recordings with multiple instruments into detailed note sequences. Developed through a collaborat…

  9. RESEARCH · CL_13577 ·

    Sakana AI's KAME architecture injects LLM knowledge into speech AI without latency

    Sakana AI has developed KAME, a novel tandem architecture for speech-to-speech AI that aims to combine the speed of direct systems with the knowledge depth of LLM-based approaches. KAME operates with two asynchronous co…