2 citations · 2 across the 1 of their papers we have counts for
1 paper
Darryl Hannan, Akshay Jain, Mohit Bansal
We present a new multimodal question answering challenge, ManyModalQA, in which an agent must answer a question by considering three distinct modalities: text, images, and tables.…