3 citations · 6 across the 6 of their papers we have counts for
6 papers · 1 filter
FlowAD: Ego-Scene Interactive Modeling for Autonomous Driving
Mingzhe Guo, Yixiang Yang, Chuanrong Han +4
Effective environment modeling is the foundation for autonomous driving, underpinning tasks from perception to planning. However, current paradigms often inadequately consider the…
Cyclic Refiner: Object-Aware Temporal Representation Learning for Multi-View 3D Detection and Tracking
Mingzhe Guo, Zhipeng Zhang, Liping Jing +3
We propose a unified object-aware temporal learning framework for multi-view 3D detection and tracking tasks. Having observed that the efficacy of the temporal fusion strategy in r…
End-to-End Autonomous Driving without Costly Modularization and 3D Manual Annotation
Mingzhe Guo, Zhipeng Zhang, Yuan He +2
We propose UAD, a method for vision-based end-to-end autonomous driving (E2EAD), achieving the best open-loop evaluation performance in nuScenes, meanwhile showing robust closed-lo…
Divert More Attention to Vision-Language Object Tracking
Mingzhe Guo, Zhipeng Zhang, Liping Jing +2
Multimodal vision-language (VL) learning has noticeably pushed the tendency toward generic intelligence owing to emerging large foundation models. However, tracking, as a fundament…
ActionPrompt: Action-Guided 3D Human Pose Estimation With Text and Pose Prompting
Hongwei Zheng, Han Li, Bowen Shi +5
Recent 2D-to-3D human pose estimation (HPE) utilizes temporal consistency across sequences to alleviate the depth ambiguity problem but ignore the action related prior knowledge hi…
Learning Target-aware Representation for Visual Tracking via Informative Interactions
Mingzhe Guo, Zhipeng Zhang, Heng Fan +4
We introduce a novel backbone architecture to improve target-perception ability of feature representation for tracking. Specifically, having observed that de facto frameworks perfo…