3 citations · 3 across the 1 of their papers we have counts for
1 paper
Yixuan Qiao, Hao Chen, Jun Wang +7
TextVQA requires models to read and reason about text in images to answer questions about them. Specifically, models need to incorporate a new modality of text present in the image…