1 paper
Yongjian Wu, Yang Zhou, Jiya Saiyin +4
Large-scale visual-language pre-trained models (VLPMs) have demonstrated exceptional performance in downstream object detection through text prompts for natural scenes. However, th…