10 citations · 16 across the 17 of their papers we have counts for
1 paper · 1 filter
Saghir Alfasly, Jian Lu, Chen Xu +1
With the assumption that a video dataset is multimodality annotated in which auditory and visual modalities both are labeled or class-relevant, current multimodal methods apply mod…