2 citations · 3 across the 8 of their papers we have counts for
4 papers · 1 filter
Unify Variables in Neural Scaling Laws for General Audio Representations via Embedding Effective Rank
Xuyao Deng, Yanjie Sun, Yong Dou +1
Scaling laws have profoundly shaped our understanding of model performance in computer vision and natural language processing, yet their application to general audio representation…
AudioSet-R: A Refined AudioSet with Multi-Stage LLM Label Reannotation
Yulin Sun, Qisheng Xu, Yi Su +4
AudioSet is a widely used benchmark in the audio research community and has significantly advanced various audio-related tasks. However, persistent issues with label accuracy and c…
AudioCIL: A Python Toolbox for Audio Class-Incremental Learning with Multiple Scenes
Qisheng Xu, Yulin Sun, Yi Su +7
Deep learning, with its robust aotomatic feature extraction capabilities, has demonstrated significant success in audio signal processing. Typically, these methods rely on static,…
Contrastive Learning-based Chaining-Cluster for Multilingual Voice-Face Association
Wuyang Chen, Yanjie Sun, Kele Xu +1
The innate correlation between a person's face and voice has recently emerged as a compelling area of study, especially within the context of multilingual environments. This paper…