Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
A Collaborative Multi-Modality Interaction for VLA-based End-to-End Autonomous Driving
Jingtao Sun, Xiaohai He, Yike Zhang +4
Vision-Language-Action (VLA) models have emerged as a powerful paradigm for end-to-end autonomous driving by jointly integrating perception, reasoning, and decision making within a…
cs.CV2024
EADReg: Probabilistic Correspondence Generation with Efficient Autoregressive Diffusion Model for Outdoor Point Cloud Registration
Linrui Gong, Jiuming Liu, Junyi Ma +3
Diffusion models have shown the great potential in the point cloud registration (PCR) task, especially for enhancing the robustness to challenging cases. However, existing diffusio…
cs.CV2024
Diffusion-Driven Self-Supervised Learning for Shape Reconstruction and Pose Estimation
Jingtao Sun, Yaonan Wang, Mingtao Feng +3
Fully-supervised category-level pose estimation aims to determine the 6-DoF poses of unseen instances from known categories, requiring expensive mannual labeling costs. Recently, v…