2 papers
cs.CV2026
USR-Drive: Unified Driving Scene Representation via Joint Denoising of 3D Gaussians and Boxes
Li-Heng Chen, Haokai Pang, Chengye Su +7
Spatial representation learning for autonomous driving aims to map raw visual signals into structured 3D scene representations, where object-centric bounding boxes and rendering-or…
cs.CV2026
VistaGEN: Consistent Driving Video Generation with Fine-Grained Control Using Multiview Visual-Language Reasoning
Li-Heng Chen, Ke Cheng, Yahui Liu +3
Driving video generation has achieved much progress in controllability, video resolution, and length, but fails to support fine-grained object-level controllability for diverse dri…