1 paper
Yue Zhang, Xiangyu Li, Wanshu Fan +2
Answer grounding in visual question answering aims to locate the region from a given natural language question associated with the visual content of an image, which has garnered si…