most citedAdversarial Attacks on Spoofing Countermeasures of automatic speaker verification

16 citations · 82 across the 11 of their papers we have counts for

collaborators

13 papers

eess.AS20203 cited

Speaker Independent and Multilingual/Mixlingual Speech-Driven Talking Head Generation Using Phonetic Posteriorgrams

Huirong Huang, Zhiyong Wu, Shiyin Kang +9

Generating 3D speech-driven talking head has received more and more attention in recent years. Recent approaches mainly have following limitations: 1) most speaker-independent meth…

eess.AS202012 cited

Investigating Robustness of Adversarial Samples Detection for Automatic Speaker Verification

Xu Li, Na Li, Jinghua Zhong +5

Recently adversarial attacks on automatic speaker verification (ASV) systems attracted widespread attention as they pose severe threats to ASV systems. However, methods to defend a…

eess.AS20202 cited

Transferring Source Style in Non-Parallel Voice Conversion

Songxiang Liu, Yuewen Cao, Shiyin Kang +5

Voice conversion (VC) techniques aim to modify speaker identity of an utterance while preserving the underlying linguistic information. Most VC approaches ignore modeling of the sp…

eess.AS20206 cited

Bayesian x-vector: Bayesian Neural Network based x-vector System for Speaker Verification

Xu Li, Jinghua Zhong, Jianwei Yu +4

Speaker verification systems usually suffer from the mismatch problem between training and evaluation data, such as speaker population mismatch, the channel and environment variati…

eess.AS202011 cited

Multi-Target Emotional Voice Conversion With Neural Vocoders

Songxiang Liu, Yuewen Cao, Helen Meng

Emotional voice conversion (EVC) is one way to generate expressive synthetic speech. Previous approaches mainly focused on modeling one-to-one mapping, i.e., conversion from one em…

eess.AS20209 cited

Emotional Voice Conversion With Cycle-consistent Adversarial Network

Songxiang Liu, Yuewen Cao, Helen Meng

Emotional Voice Conversion, or emotional VC, is a technique of converting speech from one emotion state into another one, keeping the basic linguistic information and speaker ident…