Performance Evaluation of the Dyna-Q algorithm for Robot Navigation
Emanuele Vitolo, Alberto San-Miguel, Javier Civera, Cristian Mahulea
- 发表年份
- 2018
- 引用次数
- 4
摘要
This paper tackles the path planning of a team of robots moving in a partially known environment, with static obstacles within it. Given the initial positions of the set of robots and a set of destinations, the robots should safety reach them avoiding the obstacles. Our approach is based on Reinforcement Learning, which is suited to partial knowledge of the environment and its dynamics. We use specifically the Dyna-Q algorithm (based on the Dyna architecture), including Planning and Reinforcement Learning, initially developed to a single robot case, and extended here to a multi-robot system. We analyze the problem with extensive and thorough simulations, for single and multi-robot systems, using the Robot Motion Toolbox, with the goal of characterizing the behavior of the Dyna-Q algorithm with respect to its main parameters.
关键词
相关论文
Statistical Learning Theory
Yuhai Wu, Vladimir Vapnik
1999
Artificial intelligence: a modern approach
1995
Applied Nonlinear Control
Jean-Jacques Slotine, Weiping Li
1991
A new optimizer using particle swarm theory
R.C. Eberhart, James Kennedy
2002