activity
20222026
most citedLearning Target-aware Representation for Visual Tracking via Informative Interactions

3 citations · 6 across the 6 of their papers we have counts for

collaborators
Showing cs.CVShow all

6 papers · 1 filter

cs.CV2026

FlowAD: Ego-Scene Interactive Modeling for Autonomous Driving

Mingzhe Guo, Yixiang Yang, Chuanrong Han +4

Effective environment modeling is the foundation for autonomous driving, underpinning tasks from perception to planning. However, current paradigms often inadequately consider the…

cs.CV20241 cited

Cyclic Refiner: Object-Aware Temporal Representation Learning for Multi-View 3D Detection and Tracking

Mingzhe Guo, Zhipeng Zhang, Liping Jing +3

We propose a unified object-aware temporal learning framework for multi-view 3D detection and tracking tasks. Having observed that the efficacy of the temporal fusion strategy in r…

cs.CV20242 cited

End-to-End Autonomous Driving without Costly Modularization and 3D Manual Annotation

Mingzhe Guo, Zhipeng Zhang, Yuan He +2

We propose UAD, a method for vision-based end-to-end autonomous driving (E2EAD), achieving the best open-loop evaluation performance in nuScenes, meanwhile showing robust closed-lo…

cs.CV2023

Divert More Attention to Vision-Language Object Tracking

Mingzhe Guo, Zhipeng Zhang, Liping Jing +2

Multimodal vision-language (VL) learning has noticeably pushed the tendency toward generic intelligence owing to emerging large foundation models. However, tracking, as a fundament…

cs.CV2023

ActionPrompt: Action-Guided 3D Human Pose Estimation With Text and Pose Prompting

Hongwei Zheng, Han Li, Bowen Shi +5

Recent 2D-to-3D human pose estimation (HPE) utilizes temporal consistency across sequences to alleviate the depth ambiguity problem but ignore the action related prior knowledge hi…

cs.CV20223 cited

Learning Target-aware Representation for Visual Tracking via Informative Interactions

Mingzhe Guo, Zhipeng Zhang, Heng Fan +4

We introduce a novel backbone architecture to improve target-perception ability of feature representation for tracking. Specifically, having observed that de facto frameworks perfo…