AudioSet
PulseAugur coverage of AudioSet — every cluster mentioning AudioSet across labs, papers, and developer communities, ranked by signal.
2 day(s) with sentiment data
-
AV-JEPA model advances audio-visual self-supervised learning
Researchers have introduced AV-JEPA, a new self-supervised learning model that extends LeJEPA to handle both audio and visual data. This model utilizes an early-fusion Vision Transformer and modality dropout for masking…
-
Researchers Compare Token Representations Against CNNs for Bird Vocalization Detection
Researchers from DS@GT ARC explored token representations against supervised CNN backbones for the BirdCLEF+ 2026 challenge, which focuses on detecting animal vocalizations in soundscapes. They developed a baseline mode…
-
kandinskylab releases KVAE-Audio, a high-fidelity audio autoencoder
KVAE-Audio, a new continuous, full-band audio autoencoder, has been released by kandinskylab. This model effectively compresses raw audio waveforms into compact latents and reconstructs them with high fidelity across sp…
-
New probing method boosts audio SSL model evaluation
Researchers have developed a new method called binarized prototypical probes for evaluating audio self-supervised learning models. This technique addresses the information bottleneck caused by global pooling in existing…
-
New unsupervised method segments multilingual laughter in audio
Researchers have developed a new unsupervised method for segmenting acoustic laughter across multiple languages. This approach treats laughter detection as an anomaly detection problem on audio sequences, utilizing an I…