首页 /研究 /Imitation Learning with Inconsistent Demonstrations through Uncertainty-based Data Manipulation
MANIPULATION

Imitation Learning with Inconsistent Demonstrations through Uncertainty-based Data Manipulation

Peter Valletta, Rodrigo Pérez‐Dattari, Jens Kober

发表年份
2021
引用次数
2

摘要

Aleatoric uncertainty estimation, based on the observed training data, is applied for the detection of conflicts in a demonstration data set. The particular focus of this paper is the resolution of conflicting data resulting from scenarios with equivalent action choices, such as obstacle avoidance, path planning or multiple joint configurations. In terms of the estimated uncertainty, the proposed algorithm aims to decrease this otherwise irreducible value through direct alteration of the accrued data set and to provide data that a policy-learning neural network is able to fit appropriately. The proposed algorithm was validated with real robot scenarios while learning from inconsistent demonstrations, where the resulting policies consistently achieved their prescribed objectives. A video showing our method and experiments can be found at: https://youtu.be/oGYnzlW9Ncw.

关键词

Computer scienceObstacle avoidanceArtificial intelligenceImitationFocus (optics)Data setSet (abstract data type)Machine learningRobotPath (computing)

相关论文

查看 MANIPULATION 分类全部论文