From the 1 of 5 linked papers with an AI index.
5 papers
FoMoVLA: Bridging Visual Foresight and Motion Guidance for Vision-Language-Action Models
Wei Li, Peijin Jia, Yuan Ma +9
FoMoVLA enhances vision-language-action models by jointly predicting future visual features and tracking sparse 2D points, providing both goal states and motion paths to improve co…
Realistic and Controllable 3D Gaussian-Guided Object Editing for Driving Video Generation
Jiusi Li, Jackson Jiang, Jinyu Miao +10
Corner cases are crucial for training and validating autonomous driving systems, yet collecting them from the real world is often costly and hazardous. Editing objects within captu…
PriorFusion: Unified Integration of Priors for Robust Road Perception in Autonomous Driving
Xuewei Tang, Mengmeng Yang, Tuopu Wen +7
With the growing interest in autonomous driving, there is an increasing demand for accurate and reliable road perception technologies. In complex environments without high-definiti…
Enhancing Lane Segment Perception and Topology Reasoning with Crowdsourcing Trajectory Priors
Peijin Jia, Ziang Luo, Tuopu Wen +4
In autonomous driving, recent advances in lane segment perception provide autonomous vehicles with a comprehensive understanding of driving scenarios. Moreover, incorporating prior…
DiffMap: Enhancing Map Segmentation with Map Prior Using Diffusion Model
Peijin Jia, Tuopu Wen, Ziang Luo +9
Constructing high-definition (HD) maps is a crucial requirement for enabling autonomous driving. In recent years, several map segmentation algorithms have been developed to address…