1 paper · 1 filter
Clarence Lee, M Ganesh Kumar, Cheston Tan
State-of-the-art visual grounding models can achieve high detection accuracy, but they are not designed to distinguish between all objects versus only certain objects of interest.…