2 papers
cs.CV2019
Align2Ground: Weakly Supervised Phrase Grounding Guided by Image-Caption Alignment
Samyak Datta, Karan Sikka, Anirban Roy +3
We address the problem of grounding free-form textual phrases by using weak supervision from image-caption pairs. We propose a novel end-to-end model that uses caption-to-image ret…
cs.CV2018
Understanding Visual Ads by Aligning Symbols and Objects using Co-Attention
Karuna Ahuja, Karan Sikka, Anirban Roy +1
We tackle the problem of understanding visual ads where given an ad image, our goal is to rank appropriate human generated statements describing the purpose of the ad. This problem…