10 citations · 13 across the 3 of their papers we have counts for
1 paper · 1 filter
Xiangrui Su, Qi Zhang, Chongyang Shi +2
Visual Question Answering (VQA) aims to automatically answer natural language questions related to given image content. Existing VQA methods integrate vision modeling and language…