4 papers · 1 filter
An Efficient Token Compression Framework for Visual Object Tracking
Weijing Wu, Qihua Liang, Bineng Zhong +3
Refining visual representations by eliminating their internal feature-level redundancy is crucial for simultaneously optimizing the performance and computational cost of models in…
Learning to Track Instance from Single Nature Language Description
Yaozong Zheng, Bineng Zhong, Qihua Liang +3
How to achieve vision-language (VL) tracking using natural language descriptions from a video sequence \textbf{without relying on any bounding-box ground truth}? In this work, we a…
Explicit Context Reasoning with Supervision for Visual Tracking
Fansheng Zeng, Bineng Zhong, Haiying Xia +4
Contextual reasoning with constraints is crucial for enhancing temporal consistency in cross-frame modeling for visual tracking. However, mainstream tracking algorithms typically a…
Exploring Decoupled Spatio-Temporal Consistency Learning and Self-Prompting Evolution for Self-Supervised Tracking
Yaozong Zheng, Bineng Zhong, Qihua Liang +4
The success of visual tracking has been largely driven by datasets with manual box annotations. However, these box annotations require tremendous human effort, limiting the scale a…