From the 1 of 3 linked papers with an AI index.
3 papers
cs.SD2026
Probing Spatial Structure in Pretrained Audio Representations
Chuyang Chen, Sivan Ding, Adrian S. Roman +1
The paper introduces SARL, a benchmark for evaluating how pretrained audio models encode spatial information such as source direction and room acoustics, and analyzes the strengths…
cs.SD2025
Latent Multi-view Learning for Robust Environmental Sound Representations
Sivan Ding, Julia Wilkins, Magdalena Fuentes +1
Self-supervised learning (SSL) approaches, such as contrastive and generative methods, have advanced environmental sound representation learning using unlabeled data. However, how…
cs.SD2025
Balancing Information Preservation and Disentanglement in Self-Supervised Music Representation Learning
Julia Wilkins, Sivan Ding, Magdalena Fuentes +1
Recent advances in self-supervised learning (SSL) methods offer a range of strategies for capturing useful representations from music audio without the need for labeled data. While…