28 citations · 78 across the 41 of their papers we have counts for
1 paper · 1 filter
Sheng Zhou, Dan Guo, Jia Li +2
Text-based visual question answering (TextVQA) faces the significant challenge of avoiding redundant relational inference. To be specific, a large number of detected objects and op…