works on

From the 1 of 7 linked papers with an AI index.

collaborators

7 papers

cs.CV2026

LARAD: Layout-Aware Road Anomaly Detection via Spatial-Logic Reasoning

Shiyi Mu, Xujie Chen, Shugong Xu

The paper introduces LARAD, a method for detecting road anomalies in autonomous driving by training models to recognize spatial‑logic violations rather than relying on texture diff…

cs.CV2026

DDStereo: Efficient Dual Decoder Transformers for Stereo 3D Road Anomaly Detection

Shiyi Mu, Zichong Gu, Zhiqi Ai +2

Stereo-based 3D obstacle perception for autonomous driving is currently constrained by an imbalanced triplet: deployment cost, detection accuracy, and open-set adaptability. While…

cs.CV2025

StereoDETR: Stereo-based Transformer for 3D Object Detection

Shiyi Mu, Zichong Gu, Zhiqi Ai +3

Compared to monocular 3D object detection, stereo-based 3D methods offer significantly higher accuracy but still suffer from high computational overhead and latency. The state-of-t…

cs.CV2025

Visual Bridge: Universal Visual Perception Representations Generating

Yilin Gao, Shuguang Dou, Junzhou Li +4

Recent advances in diffusion models have achieved remarkable success in isolated computer vision tasks such as text-to-image generation, depth estimation, and optical flow. However…

cs.RO2025

AnchDrive: Bootstrapping Diffusion Policies with Hybrid Trajectory Anchors for End-to-End Driving

Jinhao Chai, Anqing Jiang, Hao Jiang +4

End-to-end multi-modal planning has become a transformative paradigm in autonomous driving, effectively addressing behavioral multi-modality and the generalization challenge in lon…

cs.CV2025

Knowledge Transfer from Interaction Learning

Yilin Gao, Kangyi Chen, Zhongxing Peng +2

Current visual foundation models (VFMs) face a fundamental limitation in transferring knowledge from vision language models (VLMs), while VLMs excel at modeling cross-modal interac…