Showing cs.SDShow all
2 papers · 1 filter
cs.SD2025
EmoDubber: Towards High Quality and Emotion Controllable Movie Dubbing
Gaoxiang Cong, Jiadong Pan, Liang Li +5
Given a piece of text, a video clip, and a reference audio, the movie dubbing task aims to generate speech that aligns with the video while cloning the desired voice. The existing…
cs.SD2024
Generating High-quality Symbolic Music Using Fine-grained Discriminators
Zhedong Zhang, Liang Li, Jiehua Zhang +5
Existing symbolic music generation methods usually utilize discriminator to improve the quality of generated music via global perception of music. However, considering the complexi…