From the 1 of 8 linked papers with an AI index.
8 papers
Investigating Codec-Internal Latent Audio Watermarking for Neural Codec Robustness
Zi Hu, Houmin Sun, Linxi Li +4
Neural audio codecs are challenging transformations for audio watermarking because they re-encode, quantize, and resynthesize speech. This paper investigates continuous latent-spac…
Making Separation-First Multi-Stream Audio Watermarking Feasible via Joint Training
Houmin Sun, Zi Hu, Linxi Li +4
The paper introduces a joint training method that combines audio watermarking with source separation, allowing distinct watermarks to be embedded in individual stems and reliably r…
PC-Mix: Partial-Component Audio Spoofing Detection under Mixed Speech and Environmental Sound Conditions
Zhenshan Zhang, Xueping Zhang, Linxi Li +2
Recent studies on partial audio spoofing mainly focus on studio-recorded speech with temporal localization of spoofed segments. However, these studies often overlook realistic cond…
Overview of ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge
Xueping Zhang, Han Yin, Yang Xiao +4
The Environment-Aware Speech and Sound Deepfake Detection Challenge (ESDD2), held in conjunction with ICME 2026, evaluated systems for five component-level audio spoofing detection…
MultiAPI Spoof: A Multi-API Dataset and Local-Attention Network for Speech Anti-spoofing Detection
Xueping Zhang, Zhenshan Zhang, Yechen Wang +3
Existing speech anti-spoofing benchmarks rely on a narrow set of public models, creating a substantial gap from real-world scenarios in which commercial systems employ diverse, oft…
ESDD2: Environment-Aware Speech and Sound Deepfake Detection Challenge Evaluation Plan
Xueping Zhang, Han Yin, Yang Xiao +4
Audio recorded in real-world environments often contains a mixture of foreground speech and background environmental sounds. With rapid advances in text-to-speech, voice conversion…