activity
20242026
collaborators

13 papers

cs.CV2026

FUSER: Feed-Forward MUltiview 3D Registration Transformer and SE(3) Diffusion Refinement

Haobo Jiang, Jin Xie, Jian Yang +2

Registration of multiview point clouds conventionally relies on extensive pairwise matching to build a pose graph for global synchronization, which is computationally expensive and…

cs.CV2026

LiAuto-GeoX: Efficient Grounded Driving Transformer

Jiawei Lian, Haoyi Sun, Yang Wu +8

Dense 3D reconstruction has demonstrated immense potential for spatial understanding, yet its viability as a real-time, onboard representation for autonomous driving remains an ope…

cs.CV2026

VGGT-Long: Chunk it, Loop it, Align it -- Pushing VGGT's Limits on Kilometer-scale Long RGB Sequences

Kai Deng, Zexin Ti, Jiawei Xu +2

Foundation models for 3D vision have recently demonstrated remarkable capabilities in 3D perception. However, extending these models to large-scale RGB stream 3D reconstruction rem…

cs.CV2026

Geometry-to-Image Synthesis-Driven Generative Point Cloud Registration

Haobo Jiang, Jin Xie, Jian Yang +2

In this paper, we propose a novel 3D registration paradigm, Generative Point Cloud Registration, which bridges advanced 2D generative models with 3D matching tasks to enhance regis…

cs.CV2025

AD-GS: Object-Aware B-Spline Gaussian Splatting for Self-Supervised Autonomous Driving

Jiawei Xu, Kai Deng, Zexin Fan +3

Modeling and rendering dynamic urban driving scenes is crucial for self-driving simulation. Current high-quality methods typically rely on costly manual object tracklet annotations…

cs.CV2025

MaskHOI: Robust 3D Hand-Object Interaction Estimation via Masked Pre-training

Yuechen Xie, Haobo Jiang, Jian Yang +2

In 3D hand-object interaction (HOI) tasks, estimating precise joint poses of hands and objects from monocular RGB input remains highly challenging due to the inherent geometric amb…