Researchers have developed AudioNoisePrints, a novel method for watermarking audio generated by text-to-speech (TTS) models. This technique leverages the spatial correlation between initial noise inputs and the resulting audio in flow matching and diffusion models, enabling watermarking without retraining the TTS model or compromising audio quality. The method has demonstrated superior performance compared to existing baselines like AudioSeal, particularly under aggressive augmentation conditions, and shows promise for broader application across various TTS and vocoder models. AI
IMPACT This research offers a new technique for securing AI-generated audio content, potentially impacting content authentication and intellectual property protection in TTS applications.
RANK_REASON The cluster describes a new research paper detailing a novel method for audio watermarking. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →