3 papers
cs.CV2026
Omni-Supervised Motion Editing: Balancing Change and Invariance through Positive-Negative Learning
Zhenwu Shi, Jingyu Gong, Peiwei Wang +7
Text-based human motion editing aims to modify existing motion sequences according to natural language instructions while maintaining the consistency of the original motion. Existi…
cs.SD2024
Vector Quantized Diffusion Model Based Speech Bandwidth Extension
Yuan Fang, Jinglin Bai, Jiajie Wang +1
Recent advancements in neural audio codec (NAC) unlock new potential in audio signal processing. Studies have increasingly explored leveraging the latent features of NAC for variou…
cs.SD2024
A Two-Stage Band-Split Mamba-2 Network For Music Separation
Jinglin Bai, Yuan Fang, Jiajie Wang +1
Music source separation (MSS) aims to separate mixed music into its distinct tracks, such as vocals, bass, drums, and more. MSS is considered to be a challenging audio separation t…