6 citations · 6 across the 1 of their papers we have counts for
1 paper
Ming Dai, Lingfeng Yang, Yihao Xu +2
Visual grounding is a common vision task that involves grounding descriptive sentences to the corresponding regions of an image. Most existing methods use independent image-text en…