8 citations · 9 across the 2 of their papers we have counts for
1 paper · 1 filter
Yan Zhang, Jonathon Hare, Adam Prügel-Bennett
Visual Question Answering (VQA) models have struggled with counting objects in natural images so far. We identify a fundamental problem due to soft attention in these models as a c…