5 citations · 5 across the 5 of their papers we have counts for
1 paper · 1 filter
Fenglin Liu, Xian Wu, Shen Ge +4
Vision-and-language (V-L) tasks require the system to understand both vision content and natural language, thus learning fine-grained joint representations of vision and language (…