Deep Imitation Learning of Sequential Fabric Smoothing From an Algorithmic Supervisor
Deep Imitation Learning of Sequential Fabric Smoothing From an Algorithmic Supervisor
复制标题
来自算法主管的顺序结构平滑的深度模仿学习
DOI:
--
复制
发表时间:
2019
期刊:
影响因子:
--
通讯作者:
Ken Goldberg
中科院分区:
文献类型:
--
作者:
Daniel Seita;Aditya Ganapathi;Ryan Hoque;M. Hwang;Edward Cen;A. Tanwani;A. Balakrishna;Brijen Thananjeyan;Jeffrey Ichnowski;Nawid Jamali;K. Yamane;Soshi Iba;J. Canny;Ken Goldberg
Sequential pulling policies to flatten and smooth fabrics have applications from surgery to manufacturing to home tasks such as bed making and folding clothes. Due to the complexity of fabric states and dynamics, we apply deep imitation learning to learn policies that, given color (RGB), depth (D), or combined color-depth (RGBD) images of a rectangular fabric sample, estimate pick points and pull vectors to spread the fabric to maximize coverage. To generate data, we develop a fabric simulator and an algorithmic supervisor that has access to complete state information. We train policies in simulation using domain randomization and dataset aggregation (DAgger) on three tiers of difficulty in the initial randomized configuration. We present results comparing five baseline policies to learned policies and report systematic comparisons of RGB vs D vs RGBD images as inputs. In simulation, learned policies achieve comparable or superior performance to analytic baselines. In 180 physical experiments with the da Vinci Research Kit (dVRK) surgical robot, RGBD policies trained in simulation attain coverage of 83% to 95% depending on difficulty tier, suggesting that effective fabric smoothing policies can be learned from an algorithmic supervisor and that depth sensing is a valuable addition to color alone. Supplementary material is available at https://sites.google.com/view/fabric-smoothing.
DOI:
--
发表时间:
2018-06
期刊:
ArXiv
影响因子:
--
作者:
J. Matas;Stephen James;A. Davison
通讯作者:
J. Matas;Stephen James;A. Davison