10 papers · 1 filter
RS2AD-LiDAR: End-to-End Autonomous Driving LiDAR Data Generation from Roadside Sensor Observations
Runyi Huang, Ni Ding, Ruidan Xing +3
End-to-end autonomous driving solutions, which directly process multimodal sensory data and output fine-grained control commands, have gradually become a mainstream direction with…
ArchSIBench: Benchmarking the Architectural Spatial Intelligence of Vision-Language Models
Qirui Shen, Wenda Wang, Jiachen Lu +5
Architectural spatial intelligence, the ability to recognize and infer architectural space, is fundamental to tasks such as robot navigation, embodied interaction, and 3D scene und…
Sketch2MinSurf: Vision-Language Guided Generation of Editable Minimal Surfaces from Hand-Drawn Sketches
Wenda Wang, Anqi Liu, Junqi Yang +4
Converting hand-drawn sketches into structured 3D geometries remains challenging due to the difficulty of representing non-Euclidean surfaces and maintaining topological consistenc…
V2X-DGPE: Addressing Domain Gaps and Pose Errors for Robust Collaborative 3D Object Detection
Sichao Wang, Ming Yuan, Chuang Zhang +3
In V2X collaborative perception, the domain gaps between heterogeneous nodes pose a significant challenge for effective information fusion. Pose errors arising from latency and GPS…
Unveiling the Black Box: Independent Functional Module Evaluation for Bird's-Eye-View Perception Model
Ludan Zhang, Xiaokang Ding, Yuqi Dai +2
End-to-end models are emerging as the mainstream in autonomous driving perception. However, the inability to meticulously deconstruct their internal mechanisms results in diminishe…
GS-Net: Heterogeneous Vehicle Data Reuse via Generalizable Plug-and-Play 3DGS Module
Yichen Zhang, Zihan Wang, Jiali Han +5
End-to-end autonomous driving is increasingly data-driven, yet data reuse across vehicles remains limited. Each new vehicle often requires additional data collection and retraining…