4 papers
The Voiceprint Fallacy: Why Voices Are Not Unique Biometric Imprints
Tianle Yang, Cuiling Zhang, Chengzhe Sun +2
In recent years, the term voiceprint has regained attention, particularly in technological applications and policy-making contexts, often carrying the assumption that a person's vo…
Acoustic and perceptual differences between standard and accented speech and their voice clones
Tianle Yang, Chengzhe Sun, Phil Rose +1
Voice cloning is often evaluated in terms of overall quality, but less is known about accent preservation and its perceptual consequences. We compare standard and heavily accented…
Assessing the Ability of Neural TTS Systems to Model Consonant-Induced F0 Perturbation
Tianle Yang, Chengzhe Sun, Phil Rose +2
This study proposes a segmental-level prosodic probing framework to evaluate neural TTS models' ability to reproduce consonant-induced f0 perturbation, a fine-grained segmental-pro…
Forensic deepfake audio detection using segmental speech features
Tianle Yang, Chengzhe Sun, Siwei Lyu +1
This study explores the potential of using acoustic features of segmental speech sounds to detect deepfake audio. These features are highly interpretable because of their close rel…