6 citations · 7 across the 5 of their papers we have counts for
1 paper · 1 filter
Boqing Zhu, Kele Xu, Changjian Wang +4
We present an approach to learn voice-face representations from the talking face videos, without any identity labels. Previous works employ cross-modal instance discrimination task…