1 citations · 1 across the 4 of their papers we have counts for
1 paper · 1 filter
Haoyi Duan, Yan Xia, Mingze Zhou +3
In recent years, the deployment of large-scale pre-trained models in audio-visual downstream tasks has yielded remarkable outcomes. However, these models, primarily trained on sing…