335 citations · 621 across the 12 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2019★ 3 cited
Contextual Grounding of Natural Language Entities in Images
Farley Lai, Ning Xie, Derek Doran +1
In this paper, we introduce a contextual grounding approach that captures the context in corresponding text entities and image regions to improve the grounding accuracy. Specifical…
cs.CV2019★ 162 cited
Visual Entailment: A Novel Task for Fine-Grained Image Understanding
Ning Xie, Farley Lai, Derek Doran +1
Existing visual reasoning datasets such as Visual Question Answering (VQA), often suffer from biases conditioned on the question, image or answer distributions. The recently propos…
cs.CV2018
Visual Entailment Task for Visually-Grounded Language Learning
Ning Xie, Farley Lai, Derek Doran +1
We introduce a new inference task - Visual Entailment (VE) - which differs from traditional Textual Entailment (TE) tasks whereby a premise is defined by an image, rather than a na…