2 citations · 2 across the 3 of their papers we have counts for
4 papers
Cross-Modal Perceptionist: Can Face Geometry be Gleaned from Voices?
Cho-Ying Wu, Chin-Cheng Hsu, Ulrich Neumann
This work digs into a root question in human perception: can face geometry be gleaned from one's voices? Previous works that study this question only adopt developments in image sy…
Voice2Mesh: Cross-Modal 3D Face Model Generation from Voices
Cho-Ying Wu, Ke Xu, Chin-Cheng Hsu +1
This work focuses on the analysis that whether 3D face models can be learned from only the speech inputs of speakers. Previous works for cross-modal face synthesis study image gene…
Adversarial defense for deep speaker recognition using hybrid adversarial training
Monisankha Pal, Arindam Jati, Raghuveer Peri +3
Deep neural network based speaker recognition systems can easily be deceived by an adversary using minuscule imperceptible perturbations to the input speech samples. These adversar…
Adversarial Attack and Defense Strategies for Deep Speaker Recognition Systems
Arindam Jati, Chin-Cheng Hsu, Monisankha Pal +3
Robust speaker recognition, including in the presence of malicious attacks, is becoming increasingly important and essential, especially due to the proliferation of several smart s…