activity
20242026
collaborators

7 papers

cs.RO2026

SignNav: Leveraging Signage for Semantic Visual Navigation in Large-Scale Indoor Environments

Jian Sun, Yuming Huang, He Li +6

Humans routinely leverage semantic hints provided by signage to navigate to destinations within novel Large-Scale Indoor (LSI) environments, such as hospitals and airport terminals…

cs.CV2026

Bridging Scene Generation and Planning: Driving with World Model via Unifying Vision and Motion Representation

Xingtai Gui, Meijie Zhang, Tianyi Yan +5

End-to-end autonomous driving aims to generate safe and plausible planning policies from raw sensor input. Driving world models have shown great potential in learning rich represen…

cs.CV2025

TrajDiff: End-to-end Autonomous Driving without Perception Annotation

Xingtai Gui, Jianbo Zhao, Wencheng Han +5

End-to-end autonomous driving systems directly generate driving policies from raw sensor inputs. While these systems can extract effective environmental features for planning, rely…

cs.CV2025

RLGF: Reinforcement Learning with Geometric Feedback for Autonomous Driving Video Generation

Tianyi Yan, Wencheng Han, Xia Zhou +4

Synthetic data is crucial for advancing autonomous driving (AD) systems, yet current state-of-the-art video generation models, despite their visual realism, suffer from subtle geom…

cs.CV2025

Breaking Down Monocular Ambiguity: Exploiting Temporal Evolution for 3D Lane Detection

Huan Zheng, Wencheng Han, Tianyi Yan +2

Monocular 3D lane detection aims to estimate the 3D position of lanes from frontal-view (FV) images. However, existing methods are fundamentally constrained by the inherent ambigui…

cs.CV2024

OLiDM: Object-aware LiDAR Diffusion Models for Autonomous Driving

Tianyi Yan, Junbo Yin, Xianpeng Lang +3

To enhance autonomous driving safety in complex scenarios, various methods have been proposed to simulate LiDAR point cloud data. Nevertheless, these methods often face challenges…