17 citations · 17 across the 1 of their papers we have counts for
1 paper
Ruilin Yao, Shengwu Xiong, Yichen Zhao +1
Visual grounding is the task of locating objects specified by natural language expressions. Existing methods extend generic object detection frameworks to tackle this task. They ty…