31 citations · 45 across the 10 of their papers we have counts for
13 papers · 1 filter
DCPT: Darkness Clue-Prompted Tracking in Nighttime UAVs
Jiawen Zhu, Huayi Tang, Zhi-Qi Cheng +5
Existing nighttime unmanned aerial vehicle (UAV) trackers follow an "Enhance-then-Track" architecture - first using a light enhancer to brighten the nighttime video, then employing…
Refined Temporal Pyramidal Compression-and-Amplification Transformer for 3D Human Pose Estimation
Hanbing Liu, Wangmeng Xiang, Jun-Yan He +4
Accurately estimating the 3D pose of humans in video sequences requires both accuracy and a well-structured architecture. With the success of transformers, we introduce the Refined…
Towards Deeply Unified Depth-aware Panoptic Segmentation with Bi-directional Guidance Learning
Junwen He, Yifan Wang, Lijun Wang +6
Depth-aware panoptic segmentation is an emerging topic in computer vision which combines semantic and geometric understanding for more robust scene interpretation. Recent works pur…
PoSynDA: Multi-Hypothesis Pose Synthesis Domain Adaptation for Robust 3D Human Pose Estimation
Hanbing Liu, Jun-Yan He, Zhi-Qi Cheng +8
Existing 3D human pose estimators face challenges in adapting to new datasets due to the lack of 2D-3D pose pairs in training sets. To overcome this issue, we propose \textit{Multi…
Convolutions Die Hard: Open-Vocabulary Segmentation with Single Frozen Convolutional CLIP
Qihang Yu, Ju He, Xueqing Deng +2
Open-vocabulary segmentation is a challenging task requiring segmenting and recognizing objects from an open set of categories. One way to address this challenge is to leverage mul…
Tracking Anything in High Quality
Jiawen Zhu, Zhenyu Chen, Zeqi Hao +9
Visual object tracking is a fundamental video task in computer vision. Recently, the notably increasing power of perception algorithms allows the unification of single/multiobject…