2 papers
cs.CV2024
Let Me Finish My Sentence: Video Temporal Grounding with Holistic Text Understanding
Jongbhin Woo, Hyeonggon Ryu, Youngjoon Jang +2
Video Temporal Grounding (VTG) aims to identify visual frames in a video clip that match text queries. Recent studies in VTG employ cross-attention to correlate visual frames and t…
cs.CV2023
That's What I Said: Fully-Controllable Talking Face Generation
Youngjoon Jang, Kyeongha Rho, Jong-Bin Woo +5
The goal of this paper is to synthesise talking faces with controllable facial motions. To achieve this goal, we propose two key ideas. The first is to establish a canonical space…