1 paper · 1 filter
Sheng Zhou, Dan Guo, Jia Li +2
Text-based visual question answering (TextVQA) faces the significant challenge of avoiding redundant relational inference. To be specific, a large number of detected objects and op…