2 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.CL2021★ 1 cited
Improving Generation and Evaluation of Visual Stories via Semantic Consistency
Adyasha Maharana, Darryl Hannan, Mohit Bansal
Story visualization is an under-explored task that falls at the intersection of many important research directions in both computer vision and natural language processing. In this…
cs.CL2020★ 2 cited
ManyModalQA: Modality Disambiguation and QA over Diverse Inputs
Darryl Hannan, Akshay Jain, Mohit Bansal
We present a new multimodal question answering challenge, ManyModalQA, in which an agent must answer a question by considering three distinct modalities: text, images, and tables.…