From the 1 of 6 linked papers with an AI index.
6 papers
Visual Geometry Foundation-Aware Gaussians for Single-Frame Surround-View Driving Reconstruction
Junhong Lin, Jinlong Wang, Xianda Guo +6
Single-frame surround-view reconstruction faces severe geometric instability and rendering artifacts due to minimal inter-camera overlap. While existing methods rely on complex dec…
SpatialQ: Understanding 3D Gaussian Splatting Scene Quality via Visual-based MLLM
Jingxuan Su, Shenglin Wang, Tiesong Zhao +2
The paper introduces SpatialQ, a multimodal framework that assesses the visual quality of 3D Gaussian Splatting scenes by combining view-specific image features, depth and point‑cl…
VGOcc: Learning Visual-Geometric Gaussians for Vision-Centric 3D Driving Occupancy Prediction
Junhong Lin, Xianda Guo, Kangli Wang +4
Vision-only occupancy prediction requires recovering a semantic 3D occupancy field from calibrated surround-view images, where each view provides observations with ambiguous depth…
CoSAG: Compact Semantic Anchor Gaussians via Training-Free Rate-Distortion Coding
Yuang Jia, Jinlong Wang, Junhong Lin +2
Open-vocabulary 3D scene understanding is commonly achieved by embedding 2D vision-language features such as CLIP into a 3D Gaussian Splatting scene, turning it into a text-queryab…
DriveExplorer: Images-Only Decoupled 4D Reconstruction with Progressive Restoration for Driving View Extrapolation
Yuang Jia, Jinlong Wang, Jiayi Zhao +3
This paper presents an effective solution for view extrapolation in autonomous driving scenarios. Recent approaches focus on generating shifted novel view images from given viewpoi…
VGD: Visual Geometry Gaussian Splatting for Feed-Forward Surround-view Driving Reconstruction
Junhong Lin, Kangli Wang, Shunzhou Wang +3
Feed-forward surround-view autonomous driving scene reconstruction offers fast, generalizable inference ability, which faces the core challenge of ensuring generalization while ele…