From the 1 of 4 linked papers with an AI index.
4 papers · 1 filter
Probing Spatial Structure in Pretrained Audio Representations
Chuyang Chen, Sivan Ding, Adrian S. Roman +1
The paper introduces SARL, a benchmark for evaluating how pretrained audio models encode spatial information such as source direction and room acoustics, and analyzes the strengths…
Latent Multi-view Learning for Robust Environmental Sound Representations
Sivan Ding, Julia Wilkins, Magdalena Fuentes +1
Self-supervised learning (SSL) approaches, such as contrastive and generative methods, have advanced environmental sound representation learning using unlabeled data. However, how…
Balancing Information Preservation and Disentanglement in Self-Supervised Music Representation Learning
Julia Wilkins, Sivan Ding, Magdalena Fuentes +1
Recent advances in self-supervised learning (SSL) methods offer a range of strategies for capturing useful representations from music audio without the need for labeled data. While…
Self-Supervised Multi-View Learning for Disentangled Music Audio Representations
Julia Wilkins, Sivan Ding, Magdalena Fuentes +1
Self-supervised learning (SSL) offers a powerful way to learn robust, generalizable representations without labeled data. In music, where labeled data is scarce, existing SSL metho…