ManiSkill: Generalizable Manipulation Skill Benchmark with Large-Scale Demonstrations
arXiv:2107.14483
Abstract
Object manipulation from 3D visual inputs poses many challenges on building generalizable perception and policy models. However, 3D assets in existing benchmarks mostly lack the diversity of 3D shapes that align with real-world intra-class complexity in topology and geometry. Here we propose SAPIEN Manipulation Skill Benchmark (ManiSkill) to benchmark manipulation skills over diverse objects in a full-physics simulator. 3D assets in ManiSkill include large intra-class topological and geometric variations. Tasks are carefully chosen to cover distinct types of manipulation challenges. Latest progress in 3D vision also makes us believe that we should customize the benchmark so that the challenge is inviting to researchers working on 3D deep learning. To this end, we simulate a moving panoramic camera that returns ego-centric point clouds or RGB-D images. In addition, we would like ManiSkill to serve a broad set of researchers interested in manipulation research. Besides supporting the learning of policies from interactions, we also support learning-from-demonstrations (LfD) methods, by providing a large number of high-quality demonstrations (~36,000 successful trajectories, ~1.5M point cloud/RGB-D frames in total). We provide baselines using 3D deep learning and LfD algorithms. All code of our benchmark (simulator, environment, SDK, and baselines) is open-sourced, and a challenge facing interdisciplinary researchers will be held based on the benchmark.
NeurIPS 2021 Track on Datasets and Benchmarks; code: https://github.com/haosulab/ManiSkill
References in corpus (9)
- Solving Rubik's Cube with a Robot Hand
- Submanifold Sparse Convolutional Networks
- Domain Randomization for Transferring Deep Neural Networks from Simulation to the Real World
- dm_control: Software and Tasks for Continuous Control
- Learning Invariant Feature Spaces to Transfer Skills with Reinforcement Learning
- Rearrangement: A Challenge for Embodied AI
- Asymmetric self-play for automatic goal discovery in robotic manipulation
- RoboNet: Large-Scale Multi-Robot Learning
- HRL4IN: Hierarchical Reinforcement Learning for Interactive Navigation with Mobile Manipulators