Kimi-Audio
PulseAugur coverage of Kimi-Audio — every cluster mentioning Kimi-Audio across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
AI models' emotion neurons identified and validated across languages
Researchers have conducted the first neuron-level interpretability studies on large audio-language models (LALMs) to understand how they encode emotion across different languages. The studies identified "Multilingual Em…
-
New IAAN method boosts LALM acoustic perception by targeting encoder neurons
Researchers have developed a novel method called IAAN (Identifying and Amplifying Acoustic Neurons) to enhance the acoustic perception capabilities of large audio-language models (LALMs) without requiring retraining. Th…
-
FlexiSLM introduces dynamic frame rates for spoken language models
Researchers have developed FlexiSLM, a novel spoken language model that dynamically adjusts its frame rate for speech input and output. Unlike existing models that use fixed frame rates, FlexiSLM can adapt to the varyin…