23 citations · 29 across the 6 of their papers we have counts for
6 papers
Cyclic Refiner: Object-Aware Temporal Representation Learning for Multi-View 3D Detection and Tracking
Mingzhe Guo, Zhipeng Zhang, Liping Jing +3
We propose a unified object-aware temporal learning framework for multi-view 3D detection and tracking tasks. Having observed that the efficacy of the temporal fusion strategy in r…
VastTrack: Vast Category Visual Object Tracking
Liang Peng, Junyuan Gao, Xinran Liu +5
In this paper, we introduce a novel benchmark, dubbed VastTrack, towards facilitating the development of more general visual tracking via encompassing abundant classes and videos.…
Divert More Attention to Vision-Language Object Tracking
Mingzhe Guo, Zhipeng Zhang, Liping Jing +2
Multimodal vision-language (VL) learning has noticeably pushed the tendency toward generic intelligence owing to emerging large foundation models. However, tracking, as a fundament…
Augment and Criticize: Exploring Informative Samples for Semi-Supervised Monocular 3D Object Detection
Zhenyu Li, Zhipeng Zhang, Heng Fan +4
In this paper, we improve the challenging monocular 3D object detection problem with a general semi-supervised framework. Specifically, having observed that the bottleneck of this…
One for All: One-stage Referring Expression Comprehension with Dynamic Reasoning
Zhipeng Zhang, Zhimin Wei, Zhongzhen Huang +2
Referring Expression Comprehension (REC) is one of the most important tasks in visual reasoning that requires a model to detect the target object referred by a natural language exp…
Divert More Attention to Vision-Language Tracking
Mingzhe Guo, Zhipeng Zhang, Heng Fan +1
Relying on Transformer for complex visual feature learning, object tracking has witnessed the new standard for state-of-the-arts (SOTAs). However, this advancement accompanies by l…