13 papers
Dual-branch Distilled Transformer for Efficient Asymmetric UAV Tracking
Hongtao Yang, Bineng Zhong, Qihua Liang +4
Given the real-time demands of UAV tracking, many methods simplify the backbone to reduce computation, but this often weakens feature representation and degrades performance in com…
An Efficient Token Compression Framework for Visual Object Tracking
Weijing Wu, Qihua Liang, Bineng Zhong +3
Refining visual representations by eliminating their internal feature-level redundancy is crucial for simultaneously optimizing the performance and computational cost of models in…
Learning to Track Instance from Single Nature Language Description
Yaozong Zheng, Bineng Zhong, Qihua Liang +3
How to achieve vision-language (VL) tracking using natural language descriptions from a video sequence \textbf{without relying on any bounding-box ground truth}? In this work, we a…
Boosting Self-Supervised Tracking with Contextual Prompts and Noise Learning
Yaozong Zheng, Qihua Liang, Bineng Zhong +4
Learning robust contextual knowledge from unlabeled videos is essential for advancing self-supervised tracking. However, conventional self-supervised trackers lack effective contex…
UBATrack: Spatio-Temporal State Space Model for General Multi-Modal Tracking
Qihua Liang, Liang Chen, Yaozong Zheng +3
Multi-modal object tracking has attracted considerable attention by integrating multiple complementary inputs (e.g., thermal, depth, and event data) to achieve outstanding performa…
Robust RGB-T Tracking via Learnable Visual Fourier Prompt Fine-tuning and Modality Fusion Prompt Generation
Hongtao Yang, Bineng Zhong, Qihua Liang +3
Recently, visual prompt tuning is introduced to RGB-Thermal (RGB-T) tracking as a parameter-efficient finetuning (PEFT) method. However, these PEFT-based RGB-T tracking methods typ…