Benchmarking Structured Policies and Policy Optimization for Real-World Dexterous Object Manipulation
Benchmarking Structured Policies and Policy Optimization for Real-World Dexterous Object Manipulation
复制标题
DOI:
10.1109/lra.2021.3129139
复制
发表时间:
2021-05
影响因子:
5.2
通讯作者:
Niklas Funk;Charles B. Schaff;Rishabh Madan;Takuma Yoneda;Julen Urain De Jesus;Joe Watson;E. Gordon;F. Widmaier;Stefan Bauer;S. Srinivasa;T. Bhattacharjee;Matthew R. Walter;Jan Peters
中科院分区:
文献类型:
--
作者:
Niklas Funk;Charles B. Schaff;Rishabh Madan;Takuma Yoneda;Julen Urain De Jesus;Joe Watson;E. Gordon;F. Widmaier;Stefan Bauer;S. Srinivasa;T. Bhattacharjee;Matthew R. Walter;Jan Peters
Dexterous manipulation is a challenging and important problem in robotics. While data-driven methods are a promising approach, current benchmarks require simulation or extensive engineering support due to the sample inefficiency of popular methods. We present benchmarks for the TriFinger system, an open-source robotic platform for dexterous manipulation and the focus of the 2020 Real Robot Challenge. The benchmarked methods, which were successful in the challenge, can be generally described as structured policies, as they combine elements of classical robotics and modern policy optimization. This inclusion of inductive biases facilitates sample efficiency, interpretability, reliability and high performance. The key aspects of this benchmarking is validation of the baselines across both simulation and the real system, thorough ablation study over the core features of each solution, and a retrospective analysis of the challenge as a manipulation benchmark.