Online Adaptation for Implicit Object Tracking and Shape Reconstruction in the Wild
arXiv:2111.12728 · doi:10.1109/LRA.2022.3189185
Abstract
Tracking and reconstructing 3D objects from cluttered scenes are the key components for computer vision, robotics and autonomous driving systems. While recent progress in implicit function has shown encouraging results on high-quality 3D shape reconstruction, it is still very challenging to generalize to cluttered and partially observable LiDAR data. In this paper, we propose to leverage the continuity in video data. We introduce a novel and unified framework which utilizes a neural implicit function to simultaneously track and reconstruct 3D objects in the wild. Our approach adapts the DeepSDF model (i.e., an instantiation of the implicit function) in the video online, iteratively improving the shape reconstruction while in return improving the tracking, and vice versa. We experiment with both Waymo and KITTI datasets and show significant improvements over state-of-the-art methods for both tracking and shape reconstruction tasks. Our project page is at https://jianglongye.com/implicit-tracking .
Accepted to RA-L 2022 & IROS 2022. Project page: https://jianglongye.com/implicit-tracking
References in corpus (6)
- Fast and Furious: Real Time End-to-End 3D Detection, Tracking and Motion Forecasting with a Single Convolutional Net
- 3D-SiamRPN: An End-to-End Learning Method for Real-Time 3D Single Object Tracking Using Raw Point Cloud
- SAMP: Shape and Motion Priors for 4D Vehicle Reconstruction
- Voxel R-CNN: Towards High Performance Voxel-based 3D Object Detection
- Pose Estimation and 3D Reconstruction of Vehicles from Stereo-Images Using a Subcategory-Aware Shape Prior
- Localization and Mapping using Instance-specific Mesh Models