most citedPseudo-Siamese Network based Timbre-reserved Black-box Adversarial Attack in Speaker Identification

1 citations · 1 across the 5 of their papers we have counts for

collaborators

5 papers

cs.SD2024

Takin: A Cohort of Superior Quality Zero-shot Speech Generation Models

Sijing Chen, Yuan Feng, Laipeng He +22

With the advent of the big data and large language model era, zero-shot personalized rapid customization has emerged as a significant trend. In this report, we introduce Takin Audi…

cs.SD2023

SALT: Distinguishable Speaker Anonymization Through Latent Space Transformation

Yuanjun Lv, Jixun Yao, Peikun Chen +3

Speaker anonymization aims to conceal a speaker's identity without degrading speech quality and intelligibility. Most speaker anonymization systems disentangle the speaker represen…

cs.SD2023

Timbre-reserved Adversarial Attack in Speaker Identification

Qing Wang, Jixun Yao, Li Zhang +2

As a type of biometric identification, a speaker identification (SID) system is confronted with various kinds of attacks. The spoofing attacks typically imitate the timbre of the t…

eess.AS2023

DualVC: Dual-mode Voice Conversion using Intra-model Knowledge Distillation and Hybrid Predictive Coding

Ziqian Ning, Yuepeng Jiang, Pengcheng Zhu +4

Voice conversion is an increasingly popular technology, and the growing number of real-time applications requires models with streaming conversion capabilities. Unlike typical (non…

cs.SD20231 cited

Pseudo-Siamese Network based Timbre-reserved Black-box Adversarial Attack in Speaker Identification

Qing Wang, Jixun Yao, Ziqian Wang +2

In this study, we propose a timbre-reserved adversarial attack approach for speaker identification (SID) to not only exploit the weakness of the SID model but also preserve the tim…