DA-RNN: Semantic Mapping with Data Associated Recurrent Neural Networks
arXiv:1703.03098
Abstract
3D scene understanding is important for robots to interact with the 3D world in a meaningful way. Most previous works on 3D scene understanding focus on recognizing geometrical or semantic properties of the scene independently. In this work, we introduce Data Associated Recurrent Neural Networks (DA-RNNs), a novel framework for joint 3D scene mapping and semantic labeling. DA-RNNs use a new recurrent neural network architecture for semantic labeling on RGB-D videos. The output of the network is integrated with mapping techniques such as KinectFusion in order to inject semantic information into the reconstructed 3D scene. Experiments conducted on a real world dataset and a synthetic dataset with RGB-D videos demonstrate the ability of our method in semantic 3D scene mapping.
Published in RSS 2017
References in corpus (5)
- Very Deep Convolutional Networks for Large-Scale Image Recognition
- Sequence to Sequence Learning with Neural Networks
- TensorFlow: Large-Scale Machine Learning on Heterogeneous Distributed Systems
- On the Properties of Neural Machine Translation: Encoder-Decoder Approaches
- SemanticFusion: Dense 3D Semantic Mapping with Convolutional Neural Networks