activity
20242026
collaborators

7 papers

cs.CV2026

Dual-branch Distilled Transformer for Efficient Asymmetric UAV Tracking

Hongtao Yang, Bineng Zhong, Qihua Liang +4

Given the real-time demands of UAV tracking, many methods simplify the backbone to reduce computation, but this often weakens feature representation and degrades performance in com…

cs.NI2026

Digital Twin-Enabled Mobility-Aware Cooperative Caching in Vehicular Edge Computing

Jiahao Zeng, Zhenkui Shi, Chunpei Li +5

With the advancement of vehicle-to-vehicle (V2V) ad hoc networks and wireless communication technologies, mobile edge caching has become a key enabler for enhancing network perform…

cs.CV2025

Explicit Context Reasoning with Supervision for Visual Tracking

Fansheng Zeng, Bineng Zhong, Haiying Xia +4

Contextual reasoning with constraints is crucial for enhancing temporal consistency in cross-frame modeling for visual tracking. However, mainstream tracking algorithms typically a…

cs.CV2025

SwimVG: Step-wise Multimodal Fusion and Adaption for Visual Grounding

Liangtao Shi, Ting Liu, Xiantao Hu +3

Visual grounding aims to ground an image region through natural language, which heavily relies on cross-modal alignment. Most existing methods transfer visual/linguistic knowledge…

cs.CV2025

Adaptive Perception for Unified Visual Multi-modal Object Tracking

Xiantao Hu, Bineng Zhong, Qihua Liang +4

Recently, many multi-modal trackers prioritize RGB as the dominant modality, treating other modalities as auxiliary, and fine-tuning separately various multi-modal tasks. This imba…

cs.CV2024

Exploiting Multimodal Spatial-temporal Patterns for Video Object Tracking

Xiantao Hu, Ying Tai, Xu Zhao +5

Multimodal tracking has garnered widespread attention as a result of its ability to effectively address the inherent limitations of traditional RGB tracking. However, existing mult…