5 papers
RS2AD-LiDAR: End-to-End Autonomous Driving LiDAR Data Generation from Roadside Sensor Observations
Runyi Huang, Ni Ding, Ruidan Xing +3
End-to-end autonomous driving solutions, which directly process multimodal sensory data and output fine-grained control commands, have gradually become a mainstream direction with…
ArchSIBench: Benchmarking the Architectural Spatial Intelligence of Vision-Language Models
Qirui Shen, Wenda Wang, Jiachen Lu +5
Architectural spatial intelligence, the ability to recognize and infer architectural space, is fundamental to tasks such as robot navigation, embodied interaction, and 3D scene und…
Sketch2MinSurf: Vision-Language Guided Generation of Editable Minimal Surfaces from Hand-Drawn Sketches
Wenda Wang, Anqi Liu, Junqi Yang +4
Converting hand-drawn sketches into structured 3D geometries remains challenging due to the difficulty of representing non-Euclidean surfaces and maintaining topological consistenc…
Controllable Traffic Simulation through LLM-Guided Hierarchical Reasoning and Refinement
Zhiyuan Liu, Leheng Li, Yuning Wang +5
Evaluating autonomous driving systems in complex and diverse traffic scenarios through controllable simulation is essential to ensure their safety and reliability. However, existin…
V2X-DGPE: Addressing Domain Gaps and Pose Errors for Robust Collaborative 3D Object Detection
Sichao Wang, Ming Yuan, Chuang Zhang +3
In V2X collaborative perception, the domain gaps between heterogeneous nodes pose a significant challenge for effective information fusion. Pose errors arising from latency and GPS…