首页 /研究 /Adaptive t-Momentum-based Optimization for Unknown Ratio of Outliers in Amateur Data in Imitation Learning

MANIPULATION

Adaptive t-Momentum-based Optimization for Unknown Ratio of Outliers in Amateur Data in Imitation Learning

Wendyam Eric Lionel Ilboudo, Taisuke Kobayashi, Kenji Sugimoto

发表年份: 2021
访问权限: 开放获取

摘要

Behavioral cloning (BC) bears a high potential for safe and direct transfer of human skills to robots. However, demonstrations performed by human operators often contain noise or imperfect behaviors that can affect the efficiency of the imitator if left unchecked. In order to allow the imitators to effectively learn from imperfect demonstrations, we propose to employ the robust t-momentum optimization algorithm. This algorithm builds on the Student's t-distribution in order to deal with heavy-tailed data and reduce the effect of outlying observations. We extend the t-momentum algorithm to allow for an adaptive and automatic robustness and show empirically how the algorithm can be used to produce robust BC imitators against datasets with unknown heaviness. Indeed, the imitators trained with the t-momentum-based Adam optimizers displayed robustness to imperfect demonstrations on two different manipulation tasks with different robots and revealed the capability to take advantage of the additional data while reducing the adverse effect of non-optimal behaviors.

关键词

cs.LG

Adaptive t-Momentum-based Optimization for Unknown Ratio of Outliers in Amateur Data in Imitation Learning

摘要

关键词

相关论文

面向大型复杂构件的移动机器人辅助磨削技术综述

基于物理信息与机器学习的五轴铣削TC4钛合金刀具磨损融合预测模型

通过新型压电主动阻尼刀柄提升机器人铣削质量

一种利用磁致非线性宽带多向被动减振器抑制机器人铣削低频颤振的新方法