Kimi-Audio
PulseAugur coverage of Kimi-Audio — every cluster mentioning Kimi-Audio across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
New research tackles deepfake detection with cross-lingual and multimodal approaches · 4 sources tracked
Researchers are developing advanced methods for detecting audio and video deepfakes, focusing on improving generalization across different languages and unseen attack types. One approach, Language Orthogonalization, aim…
-
AI models' emotion neurons identified and validated across languages
Researchers have conducted the first neuron-level interpretability studies on large audio-language models (LALMs) to understand how they encode emotion across different languages. The studies identified "Multilingual Em…
-
New IAAN method boosts LALM acoustic perception by targeting encoder neurons
Researchers have developed a novel method called IAAN (Identifying and Amplifying Acoustic Neurons) to enhance the acoustic perception capabilities of large audio-language models (LALMs) without requiring retraining. Th…
-
FlexiSLM introduces dynamic frame rates for spoken language models
Researchers have developed FlexiSLM, a novel spoken language model that dynamically adjusts its frame rate for speech input and output. Unlike existing models that use fixed frame rates, FlexiSLM can adapt to the varyin…