2 papers
eess.AS2026
Experience-Calibrated Contrastive Decoding for Mitigating Hallucinations in LM-Based Text-to-Speech
Chenlin Liu, Minghui Fang, Zhonghao Bi +3
Language model-based text-to-speech (LM-based TTS) remains vulnerable to speech hallucinations that deviate from the target text. Existing mitigation mainly relies on architectural…
eess.AS2025
Attacking Voice Anonymization Systems with Augmented Feature and Speaker Identity Difference
Yanzhe Zhang, Zhonghao Bi, Feiyang Xiao +3
This study focuses on the First VoicePrivacy Attacker Challenge within the ICASSP 2025 Signal Processing Grand Challenge, which aims to develop speaker verification systems capable…