1 paper
Bixing Wu, Yuhong Zhao, Zongli Ye +3
Audio-visual joint representation learning under Cross-Modal Generalization (CMG) aims to transfer knowledge from a labeled source modality to an unlabeled target modality through…