19 citations · 57 across the 46 of their papers we have counts for
46 papers
IDEAW: Robust Neural Audio Watermarking with Invertible Dual-Embedding
Pengcheng Li, Xulong Zhang, Jing Xiao +1
The audio watermarking technique embeds messages into audio and accurately extracts messages from the watermarked audio. Traditional methods develop algorithms based on expert expe…
QLSC: A Query Latent Semantic Calibrator for Robust Extractive Question Answering
Sheng Ouyang, Jianzong Wang, Yong Zhang +5
Extractive Question Answering (EQA) in Machine Reading Comprehension (MRC) often faces the challenge of dealing with semantically identical but format-variant inputs. Our work intr…
EfficientASR: Speech Recognition Network Compression via Attention Redundancy and Chunk-Level FFN Optimization
Jianzong Wang, Ziqi Liang, Xulong Zhang +2
In recent years, Transformer networks have shown remarkable performance in speech recognition tasks. However, their deployment poses challenges due to high computational and storag…
EAD-VC: Enhancing Speech Auto-Disentanglement for Voice Conversion with IFUB Estimator and Joint Text-Guided Consistent Learning
Ziqi Liang, Jianzong Wang, Xulong Zhang +3
Using unsupervised learning to disentangle speech into content, rhythm, pitch, and timbre for voice conversion has become a hot research topic. Existing works generally take into a…
CONTUNER: Singing Voice Beautifying with Pitch and Expressiveness Condition
Jianzong Wang, Pengcheng Li, Xulong Zhang +2
Singing voice beautifying is a novel task that has application value in people's daily life, aiming to correct the pitch of the singing voice and improve the expressiveness without…
Medical Speech Symptoms Classification via Disentangled Representation
Jianzong Wang, Pengcheng Li, Xulong Zhang +2
Intent is defined for understanding spoken language in existing works. Both textual features and acoustic features involved in medical speech contain intent, which is important for…