2 papers
cs.CV2024
MarvelOVD: Marrying Object Recognition and Vision-Language Models for Robust Open-Vocabulary Object Detection
Kuo Wang, Lechao Cheng, Weikai Chen +4
Learning from pseudo-labels that generated with VLMs~(Vision Language Models) has been shown as a promising solution to assist open vocabulary detection (OVD) in recent studies. Ho…
cs.CV2024
Other Tokens Matter: Exploring Global and Local Features of Vision Transformers for Object Re-Identification
Yingquan Wang, Pingping Zhang, Dong Wang +1
Object Re-Identification (Re-ID) aims to identify and retrieve specific objects from images captured at different places and times. Recently, object Re-ID has achieved great succes…