167 citations · 207 across the 8 of their papers we have counts for
1 paper · 1 filter
Sijie Mai, Songlong Xing, Jiaxuan He +2
In this paper, we study the task of multimodal sequence analysis which aims to draw inferences from visual, language and acoustic sequences. A majority of existing works generally…