1 paper
Yitong Chen, Wenhao Yao, Lingchen Meng +3
Enabling models to recognize vast open-world categories has been a longstanding pursuit in object detection. By leveraging the generalization capabilities of vision-language models…