1 paper · 1 filter
Yifan Xu, Mengdan Zhang, Xiaoshan Yang +1
In this paper, we for the first time explore helpful multi-modal contextual knowledge to understand novel categories for open-vocabulary object detection (OVD). The multi-modal con…