430 citations · 1.8k across the 26 of their papers we have counts for
1 paper · 2 filters
Amanpreet Singh, Vivek Natarajan, Meet Shah +5
Studies have shown that a dominant class of questions asked by visually impaired users on images of their surroundings involves reading text in the image. But today's VQA models ca…