3 citations · 5 across the 13 of their papers we have counts for
Showing 2024Show all
3 papers · 1 filter
cs.SD2024★ 1 cited
EmoDubber: Towards High Quality and Emotion Controllable Movie Dubbing
Gaoxiang Cong, Jiadong Pan, Liang Li +5
Given a piece of text, a video clip, and a reference audio, the movie dubbing task aims to generate speech that aligns with the video while cloning the desired voice. The existing…
cs.CV2024
RETTA: Retrieval-Enhanced Test-Time Adaptation for Zero-Shot Video Captioning
Yunchuan Ma, Laiyun Qing, Guorong Li +4
Despite the significant progress of fully-supervised video captioning, zero-shot methods remain much less explored. In this paper, we propose a novel zero-shot video captioning fra…
cs.CL2024★ 1 cited
StyleDubber: Towards Multi-Scale Style Learning for Movie Dubbing
Gaoxiang Cong, Yuankai Qi, Liang Li +6
Given a script, the challenge in Movie Dubbing (Visual Voice Cloning, V2C) is to generate speech that aligns well with the video in both time and emotion, based on the tone of a re…