6 papers
Audio Pirates: Black-box Audio Watermark Removal via Diffusion Priors
Lingfeng Yao, Xincong Zhong, Chenpei Huang +6
With the rise of AI-generated audio, watermarking has become widely used for detecting misuse and protecting intellectual property. However, adversaries may try to remove these wat…
LongCat-AudioDiT: High-Fidelity Diffusion Text-to-Speech in the Waveform Latent Space
Detai Xin, Shujie Hu, Chengzuo Yang +4
We present LongCat-AudioDiT, a novel, non-autoregressive diffusion-based text-to-speech (TTS) model that achieves state-of-the-art (SOTA) performance. Unlike previous methods that…
EchoMark: Perceptual Acoustic Environment Transfer with Watermark-Embedded Room Impulse Response
Chenpei Huang, Lingfeng Yao, Kyu In Lee +3
Acoustic Environment Matching (AEM) is the task of transferring clean audio into a target acoustic environment, enabling engaging applications such as audio dubbing and auditory im…
Yours or Mine? Overwriting Attacks Against Neural Audio Watermarking
Lingfeng Yao, Chenpei Huang, Shengyao Wang +6
As generative audio models are rapidly evolving, AI-generated audios increasingly raise concerns about copyright infringement and misinformation spread. Audio watermarking, as a pr…
Who's Wearing? Ear Canal Biometric Key Extraction for User Authentication on Wireless Earbuds
Chenpei Huang, Lingfeng Yao, Hui Zhong +5
Ear canal scanning/sensing (ECS) has emerged as a novel biometric authentication method for mobile devices paired with wireless earbuds. Existing studies have demonstrated the uniq…
SpeechVerifier: Robust Acoustic Fingerprint against Tampering Attacks via Watermarking
Lingfeng Yao, Chenpei Huang, Shengyao Wang +5
With the surge of social media, maliciously tampered public speeches, especially those from influential figures, have seriously affected social stability and public trust. Existing…