Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
OmniResponse: Online Multimodal Conversational Response Generation in Dyadic Interactions
Cheng Luo, Jianghui Wang, Bing Li +2
In this paper, we introduce Online Multimodal Conversational Response Generation (OMCRG), a novel task designed to produce synchronized verbal and non-verbal listener feedback onli…
cs.CV2023
Task-Robust Pre-Training for Worst-Case Downstream Adaptation
Jianghui Wang, Yang Chen, Xingyu Xie +2
Pre-training has achieved remarkable success when transferred to downstream tasks. In machine learning, we care about not only the good performance of a model but also its behavior…
cs.CV2023
MoviePuzzle: Visual Narrative Reasoning through Multimodal Order Learning
Jianghui Wang, Yuxuan Wang, Dongyan Zhao +1
We introduce MoviePuzzle, a novel challenge that targets visual narrative reasoning and holistic movie understanding. Despite the notable progress that has been witnessed in the re…