2 citations · 3 across the 2 of their papers we have counts for
1 paper · 1 filter
Ryota Tanaka, Kyosuke Nishida, Kosuke Nishida +3
Visual question answering on document images that contain textual, visual, and layout information, called document VQA, has received much attention recently. Although many datasets…