audio large language models
PulseAugur coverage of audio large language models — every cluster mentioning audio large language models across labs, papers, and developer communities, ranked by signal.
2 day(s) with sentiment data
-
New benchmark probes TTS evaluators on linguistic dimensions
Researchers have developed a new benchmark to evaluate automated Text-to-Speech (TTS) systems, moving beyond simple naturalness metrics. The benchmark deconstructs speech quality into 10 distinct linguistic dimensions, …
-
New GRM framework enhances stealth of audio LLM jailbreak attacks
Researchers have developed a new framework called GRM (Gradient-Ratio Masking) to improve the stealthiness of jailbreak attacks on Audio Large Language Models (ALLMs). This method selectively applies perturbations to sp…
-
Audio LLMs unify speech editing detection and localization
Researchers have developed a new framework to unify speech editing detection and content localization using Audio Large Language Models. This approach addresses limitations in existing methods, particularly for deletion…
-
New framework boosts Audio LLM robustness against noise
Researchers have developed EchoDistill, a novel self-distillation framework designed to enhance the robustness of Audio Large Language Models (ALLMs) against real-world noise. This method aligns noisy student models wit…
-
SpeakerLLM advances audio AI with speaker-specific understanding
Researchers have developed SpeakerLLM, a novel audio large language model framework designed to enhance speaker understanding and verification in AI systems. This framework integrates speaker profiling, recording condit…