2 papers
cs.SD2022
Unsupervised Voice-Face Representation Learning by Cross-Modal Prototype Contrast
Boqing Zhu, Kele Xu, Changjian Wang +4
We present an approach to learn voice-face representations from the talking face videos, without any identity labels. Previous works employ cross-modal instance discrimination task…
cs.SD2018
Learning Environmental Sounds with Multi-scale Convolutional Neural Network
Boqing Zhu, Changjian Wang, Feng Liu +3
Deep learning has dramatically improved the performance of sounds recognition. However, learning acoustic models directly from the raw waveform is still challenging. Current wavefo…