42 citations · 43 across the 2 of their papers we have counts for
3 papers
cs.CV2026
Focus on What Really Matters in Low-Altitude Governance: A Management-Centric Multi-Modal Benchmark with Implicitly Coordinated Vision-Language Reasoning Framework
Hao Chang, Zhihui Wang, Lingxiang Wu +5
Low-altitude vision systems are becoming a critical infrastructure for smart city governance. However, existing object-centric perception paradigms and loosely coupled vision-langu…
cs.CV2024★ 1 cited
Heterogeneous Graph Transformer for Multiple Tiny Object Tracking in RGB-T Videos
Qingyu Xu, Longguang Wang, Weidong Sheng +4
Tracking multiple tiny objects is highly challenging due to their weak appearance and limited features. Existing multi-object tracking algorithms generally focus on single-modality…
cs.CV2024★ 42 cited
Highly Efficient and Unsupervised Framework for Moving Object Detection in Satellite Videos
C. Xiao, W. An, Y. Zhang +5
Moving object detection in satellite videos (SVMOD) is a challenging task due to the extremely dim and small target characteristics. Current learning-based methods extract spatio-t…