1 citations · 1 across the 3 of their papers we have counts for
1 paper · 1 filter
Haomin Zhang, Chang Liu, Junjie Zheng +3
Currently, high-quality, synchronized audio is synthesized using various multi-modal joint learning frameworks, leveraging video and optional text inputs. In the video-to-audio ben…