Volumetric Instance-Aware Semantic Mapping and 3D Object Discovery
arXiv:1903.00268 · doi:10.1109/LRA.2019.2923960
Abstract
To autonomously navigate and plan interactions in real-world environments, robots require the ability to robustly perceive and map complex, unstructured surrounding scenes. Besides building an internal representation of the observed scene geometry, the key insight toward a truly functional understanding of the environment is the usage of higher-level entities during mapping, such as individual object instances. We propose an approach to incrementally build volumetric object-centric maps during online scanning with a localized RGB-D camera. First, a per-frame segmentation scheme combines an unsupervised geometric approach with instance-aware semantic object predictions. This allows us to detect and segment elements both from the set of known classes and from other, previously unseen categories. Next, a data association step tracks the predicted instances across the different frames. Finally, a map integration strategy fuses information about their 3D shape, location, and, if available, semantic class into a global volume. Evaluation on a publicly available dataset shows that the proposed approach for building instance-level semantic maps is competitive with state-of-the-art methods, while additionally able to discover objects of unseen categories. The system is further evaluated within a real-world robotic mapping setup, for which qualitative results highlight the online nature of the method.
8 pages, 4 figures. To be published in IEEE Robotics and Automation Letters (RA-L) and 2019 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS). Accompanying video material can be found at http://youtu.be/Jvl42VJmYxg
Cited by in corpus (24)
- EAO-SLAM: Monocular Semi-Dense Object SLAM Based on Ensemble Data Association
- Panoptic Multi-TSDFs: a Flexible Representation for Online Multi-resolution Volumetric Mapping and Long-term Dynamic Scene Consistency
- From SLAM to Situational Awareness: Challenges and Survey
- An Object SLAM Framework for Association, Mapping, and High-Level Tasks
- Real-Time Multi-Modal Semantic Fusion on Unmanned Aerial Vehicles with Label Propagation for Cross-Domain Adaptation
- Frontier Semantic Exploration for Visual Target Navigation
- 3D VSG: Long-term Semantic Scene Change Prediction through 3D Variable Scene Graphs
- Real-Time Metric-Semantic Mapping for Autonomous Navigation in Outdoor Environments
- Learning Densities in Feature Space for Reliable Segmentation of Indoor Scenes
- Multitask Learning for Scalable and Dense Multilayer Bayesian Map Inference
- Continual Adaptation of Semantic Segmentation using Complementary 2D-3D Data Representations
- Online Object-Oriented Semantic Mapping and Map Updating
- Evaluating the Impact of Semantic Segmentation and Pose Estimation on Dense Semantic SLAM
- Open Scene Graphs for Open-World Object-Goal Navigation
- Panoptic Vision-Language Feature Fields
- FM-Fusion: Instance-aware Semantic Mapping Boosted by Vision-Language Foundation Models
- Asynchronous Collaborative Autoscanning with Mode Switching for Multi-Robot Scene Reconstruction
- Active Exploration based on Information Gain by Particle Filter for Efficient Spatial Concept Formation
- Scrape, Cut, Paste and Learn: Automated Dataset Generation Applied to Parcel Logistics
- Kimera-Multi: a System for Distributed Multi-Robot Metric-Semantic Simultaneous Localization and Mapping
- Semantic 3D Grid Maps for Autonomous Driving
- Object Instance Retrieval in Assistive Robotics: Leveraging Fine-Tuned SimSiam with Multi-View Images Based on 3D Semantic Map
- Augmented Environment Representations with Complete Object Models
- Map as a By-product: Collective Landmark Mapping from IMU Data and User-provided Texts in Situated Tasks