3 papers
eess.AS2022
Preserving background sound in noise-robust voice conversion via multi-task learning
Jixun Yao, Yi Lei, Qing Wang +6
Background sound is an informative form of art that is helpful in providing a more immersive experience in real-application voice conversion (VC) scenarios. However, prior research…
cs.CV2019
Unknown Identity Rejection Loss: Utilizing Unlabeled Data for Face Recognition
Haiming Yu, Yin Fan, Keyu Chen +4
Face recognition has advanced considerably with the availability of large-scale labeled datasets. However, how to further improve the performance with the easily accessible unlabel…
cs.CV2018
iQIYI-VID: A Large Dataset for Multi-modal Person Identification
Yuanliu Liu, Bo Peng, Peipei Shi +12
Person identification in the wild is very challenging due to great variation in poses, face quality, clothes, makeup and so on. Traditional research, such as face recognition, pers…