2 papers
cs.CV2026
VPTracker: Global Vision-Language Tracking via Visual Prompt
Jingchao Wang, Kaiwen Zhou, Zhijian Wu +3
Vision-Language Tracking aims to continuously localize objects described by a visual template and a language description. Existing methods, however, are typically limited to local…
cs.CV2025
Progressive Language-guided Visual Learning for Multi-Task Visual Grounding
Jingchao Wang, Hong Wang, Wenlong Zhang +3
Multi-task visual grounding (MTVG) includes two sub-tasks, i.e., Referring Expression Comprehension (REC) and Referring Expression Segmentation (RES). The existing representative a…