2 papers
cs.CV2025
BEVWorld: A Multimodal World Simulator for Autonomous Driving via Scene-Level BEV Latents
Yumeng Zhang, Shi Gong, Kaixin Xiong +7
World models have attracted increasing attention in autonomous driving for their ability to forecast potential future scenarios. In this paper, we propose BEVWorld, a novel framewo…
cs.CV2024
OPEN: Object-wise Position Embedding for Multi-view 3D Object Detection
Jinghua Hou, Tong Wang, Xiaoqing Ye +6
Accurate depth information is crucial for enhancing the performance of multi-view 3D object detection. Despite the success of some existing multi-view 3D detectors utilizing pixel-…