Pocket TTS
PulseAugur coverage of Pocket TTS — every cluster mentioning Pocket TTS across labs, papers, and developer communities, ranked by signal.
2 day(s) with sentiment data
-
CPU Voice Cloning Benchmark: Pocket TTS, Kokoro, Audio8, XTTS v2 Compared
A practical benchmark evaluated four voice cloning models (Pocket TTS, Kokoro, Audio88 & Yassin, and XTTS v2) on CPU performance. The evaluation focused on speaker similarity, naturalness, intelligibility, and latency, …
-
Llama.cpp adds on-device text-to-speech with Pocket-TTS integration
The Llama.cpp project has integrated Pocket-TTS, enabling on-device text-to-speech generation without requiring cloud services. This integration allows for local voice synthesis directly from repositories that already s…
-
CPU TTS benchmark: Pocket TTS shows flat RTF, UTMOS struggles with naturalness
A benchmark comparing four CPU-based text-to-speech (TTS) models—Kokoro, Supertonic, Inflect-Nano, and Pocket TTS—reveals distinct performance characteristics. Pocket TTS, utilizing a streaming language model architectu…
-
Kyutai's Pocket TTS offers CPU-based voice cloning from 5s audio
Kyutai has released Pocket TTS, a ~100M parameter streaming language model that generates audio tokens autoregressively. This model is notable for its ability to perform zero-shot voice cloning from just 5 seconds of au…
-
User seeks help implementing Calm TTS paper, facing voice cloning issues
A user is seeking assistance with implementing the Calm text-to-speech model described in a research paper. They have encountered difficulties in replicating the model's performance, experiencing issues with generating …