Effective reward function in discernment behavior reinforcement learning based on categorization progress
Chyon Hae Kim, Yusuke Kon, Ricardo Navarro, Manabu Gouko, Yuichi Kobayashi
- 发表年份
- 2016
- 引用次数
- 4
摘要
In object categorization tasks, the behavior trough that a robot observes an object is important. In order to categorize an object well, multiple observations through multiple behaviors are required in many cases. In this paper, we propose a class of methods, Discernment Behavior Reinforcement Learning with Adaptive Reward, which allows a robot to learn the multiple behaviors. In DBRL-AR, the observation results in previously performed behaviors/observations are took into account in the latter behaviors/observations. The effectiveness of the proposed algorithms was validated using a humanoid robot under the criteria of Between-Class Non-coincidence (BCN) and Within-Class Coincidence (WCC) that were calculated from actual data that the humanoid robot had been observed through learned behaviors. In the experiment, DBRL-AR improved the values of BCN and WCC that were small in the first observation by the latter observation. This means that the robot was able to perform the behaviors that effectively standing out the features that was not observed in the first observation in the latter observation.
关键词
相关论文
Statistical Learning Theory
Yuhai Wu, Vladimir Vapnik
1999
Artificial intelligence: a modern approach
1995
Applied Nonlinear Control
Jean-Jacques Slotine, Weiping Li
1991
A new optimizer using particle swarm theory
R.C. Eberhart, James Kennedy
2002