activity
20242026
collaborators

9 papers

cs.AI2026

A survey of AI-generated voices and their detection

Chengzhe Sun, Tianle Yang, Siwei Lyu

The ability of artificial intelligence (AI) models to generate highly realistic human voices has advanced rapidly. These technologies power accessibility tools, virtual assistants…

eess.AS2026

The Voiceprint Fallacy: Why Voices Are Not Unique Biometric Imprints

Tianle Yang, Cuiling Zhang, Chengzhe Sun +2

In recent years, the term voiceprint has regained attention, particularly in technological applications and policy-making contexts, often carrying the assumption that a person's vo…

cs.SD2026

Acoustic and perceptual differences between standard and accented speech and their voice clones

Tianle Yang, Chengzhe Sun, Phil Rose +1

Voice cloning is often evaluated in terms of overall quality, but less is known about accent preservation and its perceptual consequences. We compare standard and heavily accented…

cs.CL2026

Assessing the Ability of Neural TTS Systems to Model Consonant-Induced F0 Perturbation

Tianle Yang, Chengzhe Sun, Phil Rose +2

This study proposes a segmental-level prosodic probing framework to evaluate neural TTS models' ability to reproduce consonant-induced f0 perturbation, a fine-grained segmental-pro…

cs.SD2025

Forensic deepfake audio detection using segmental speech features

Tianle Yang, Chengzhe Sun, Siwei Lyu +1

This study explores the potential of using acoustic features of segmental speech sounds to detect deepfake audio. These features are highly interpretable because of their close rel…

cs.CR2025

DeepfakeBench-MM: A Comprehensive Benchmark for Multimodal Deepfake Detection

Kangran Zhao, Yupeng Chen, Xiaoyu Zhang +8

The misuse of advanced generative AI models has resulted in the widespread proliferation of falsified data, particularly forged human-centric audiovisual content, which poses subst…