18 citations · 21 across the 5 of their papers we have counts for
1 paper · 1 filter
Rakesh Vaideeswaran, Feng Gao, Abhinav Mathur +1
The domain of joint vision-language understanding, especially in the context of reasoning in Visual Question Answering (VQA) models, has garnered significant attention in the recen…