1 citations · 1 across the 1 of their papers we have counts for
1 paper
Yuxuan Wang, Zilong Zheng, Xueliang Zhao +3
Video-grounded dialogue understanding is a challenging problem that requires machine to perceive, parse and reason over situated semantics extracted from weakly aligned video and d…