Researchers have developed Tacit-TTS, a novel text-to-speech system designed for efficient and transcript-free voice cloning. This system, derived from IndexTTS2, replaces traditional autoregressive decoding with a masked non-autoregressive generation approach. Tacit-TTS achieves over 10x faster speech generation compared to its predecessor for longer utterances and supports cross-lingual and non-lexical references, demonstrating its capability with diverse audio inputs. AI
IMPACT This research could lead to more efficient and versatile voice cloning tools, potentially impacting content creation and accessibility.
RANK_REASON The cluster contains a research paper detailing a new model and its technical approach. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →