28 citations · 28 across the 1 of their papers we have counts for
1 paper
Hao Li, Jinfa Huang, Peng Jin +3
Text-based Visual Question Answering~(TextVQA) aims to produce correct answers for given questions about the images with multiple scene texts. In most cases, the texts naturally at…