Benchmarking Structured Policies and Policy Optimization for Real-World Dexterous Object Manipulation
arXiv:2105.02087 · doi:10.1109/LRA.2021.3129139
Abstract
Dexterous manipulation is a challenging and important problem in robotics. While data-driven methods are a promising approach, current benchmarks require simulation or extensive engineering support due to the sample inefficiency of popular methods. We present benchmarks for the TriFinger system, an open-source robotic platform for dexterous manipulation and the focus of the 2020 Real Robot Challenge. The benchmarked methods, which were successful in the challenge, can be generally described as structured policies, as they combine elements of classical robotics and modern policy optimization. This inclusion of inductive biases facilitates sample efficiency, interpretability, reliability and high performance. The key aspects of this benchmarking is validation of the baselines across both simulation and the real system, thorough ablation study over the core features of each solution, and a retrospective analysis of the challenge as a manipulation benchmark. The code and demo videos for this work can be found on our website (https://sites.google.com/view/benchmark-rrc).
References in corpus (11)
- Practical Bayesian Optimization of Machine Learning Algorithms
- Solving Rubik's Cube with a Robot Hand
- Deep Dynamics Models for Learning Dexterous Manipulation
- Residual Policy Learning
- Benchmarking In-Hand Manipulation
- ReLMoGen: Leveraging Motion Generation in Reinforcement Learning for Mobile Manipulation
- On Policy Learning Robust to Irreversible Events: An Application to Robotic In-Hand Manipulation
- TriFinger: An Open-Source Robot for Learning Dexterity
- Neural Dynamic Policies for End-to-End Sensorimotor Learning
- Grasp and Motion Planning for Dexterous Manipulation for the Real Robot Challenge
- Solving Challenging Dexterous Manipulation Tasks With Trajectory Optimisation and Reinforcement Learning