4 papers
Exploiting Lightweight Hierarchical ViT and Dynamic Framework for Efficient Visual Tracking
Ben Kang, Xin Chen, Jie Zhao +3
Transformer-based visual trackers have demonstrated significant advancements due to their powerful modeling capabilities. However, their practicality is limited on resource-constra…
Exploring Enhanced Contextual Information for Video-Level Object Tracking
Ben Kang, Xin Chen, Simiao Lai +3
Contextual information at the video level has become increasingly crucial for visual object tracking. However, existing methods typically use only a few tokens to convey this infor…
MambaVT: Spatio-Temporal Contextual Modeling for robust RGB-T Tracking
Simiao Lai, Chang Liu, Jiawen Zhu +4
Existing RGB-T tracking algorithms have made remarkable progress by leveraging the global interaction capability and extensive pre-trained models of the Transformer architecture. N…
3rd Place Solution for PVUW2023 VSS Track: A Large Model for Semantic Segmentation on VSPW
Shijie Chang, Zeqi Hao, Ben Kang +6
In this paper, we introduce 3rd place solution for PVUW2023 VSS track. Semantic segmentation is a fundamental task in computer vision with numerous real-world applications. We have…