1 citations · 1 across the 2 of their papers we have counts for
1 paper · 1 filter
Yuan Zhao, Zhenqi Jia, Rui Liu +3
Automatic Video Dubbing (AVD) aims to take the given script and generate speech that aligns with lip motion and prosody expressiveness. Current AVD models mainly utilize visual inf…