4 papers
DiffuBox: Refining 3D Object Detection with Point Diffusion
Xiangyu Chen, Zhenzhen Liu, Katie Z Luo +10
Ensuring robust 3D object detection and localization is crucial for many applications in robotics and autonomous driving. Recent models, however, face difficulties in maintaining h…
Language-Image Models with 3D Understanding
Jang Hyun Cho, Boris Ivanovic, Yulong Cao +8
Multi-modal large language models (MLLMs) have shown incredible capabilities in a variety of 2D vision and language tasks. We extend MLLMs' perceptual capabilities to ground and re…
Better Monocular 3D Detectors with LiDAR from the Past
Yurong You, Cheng Perng Phoo, Carlos Andres Diaz-Ruiz +5
Accurate 3D object detection is crucial to autonomous driving. Though LiDAR-based detectors have achieved impressive performance, the high cost of LiDAR sensors precludes their wid…
Pre-Training LiDAR-Based 3D Object Detectors Through Colorization
Tai-Yu Pan, Chenyang Ma, Tianle Chen +7
Accurate 3D object detection and understanding for self-driving cars heavily relies on LiDAR point clouds, necessitating large amounts of labeled data to train. In this work, we in…