PulseAugur
EN
LIVE 21:06:05

NeuTTS-2E: Open-source on-device TTS model with 7 controllable emotions released

NeuTTS-2E is a new open-source, on-device text-to-speech model that allows users to control the emotion of the generated speech. The model, with 125 million parameters, supports seven distinct emotions and four built-in voices, enabling users to select specific emotional deliveries while maintaining the original voice. This release aims to address challenges in emotional speech data and disentangling emotion from text semantics, offering a privacy-preserving solution for local hardware. AI

IMPACT Enables more expressive and privacy-preserving speech generation on local hardware.

RANK_REASON Release of an open-source model with novel features. [lever_c_demoted from research: ic=1 ai=1.0]

Read on r/LocalLLaMA →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

NeuTTS-2E: Open-source on-device TTS model with 7 controllable emotions released

COVERAGE [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/TeamNeuphonic ·

    We built NeuTTS-2E, an open-source on-device TTS model with 7 controllable emotions

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1v3h4ni/we_built_neutts2e_an_opensource_ondevice_tts/"> <img alt="We built NeuTTS-2E, an open-source on-device TTS model with 7 controllable emotions" src="https://external-preview.redd.it/a2dtbmdqdTlmc2VoMSjC…