1 paper · 1 filter
Boqing Zhu, Kele Xu, Changjian Wang +4
We present an approach to learn voice-face representations from the talking face videos, without any identity labels. Previous works employ cross-modal instance discrimination task…