collaborators

5 papers

cs.CV2026

Speed3R: Sparse Feed-forward 3D Reconstruction Models

Weining Ren, Xiao Tan, Kai Han

While recent feed-forward 3D reconstruction models accelerate 3D reconstruction by jointly inferring dense geometry and camera poses in a single pass, their reliance on dense atten…

cs.CV2024

UniDet: Unified and Universal Framework for Prompt-Guided Multi-dataset 3D Detection

Yubin Wang, Zhikang Zou, Xiaoqing Ye +3

We present UniDet, a brand new framework for unified and universal multi-dataset training on 3D detection, enabling robust performance across diverse domains and generalization…

cs.CV2024

Explore the LiDAR-Camera Dynamic Adjustment Fusion for 3D Object Detection

Yiran Yang, Xu Gao, Tong Wang +5

Camera and LiDAR serve as informative sensors for accurate and robust autonomous driving systems. However, these sensors often exhibit heterogeneous natures, resulting in distribut…

cs.CV2024

OPEN: Object-wise Position Embedding for Multi-view 3D Object Detection

Jinghua Hou, Tong Wang, Xiaoqing Ye +6

Accurate depth information is crucial for enhancing the performance of multi-view 3D object detection. Despite the success of some existing multi-view 3D detectors utilizing pixel-…

cs.CV2024

BEVWorld: A Multimodal World Simulator for Autonomous Driving via Scene-Level BEV Latents

Yumeng Zhang, Shi Gong, Kaixin Xiong +7

World models have attracted increasing attention in autonomous driving for their ability to forecast potential future scenarios. In this paper, we propose BEVWorld, a novel framewo…