Reinforcement learning with continuous vector output
Qiang Li, Zhu Hai, Lin Liang Ming, Yang Guo Zheng
- 发表年份
- 2002
- 引用次数
- 2
摘要
A new reinforcement learning algorithm with continuous vector output (CVRL) is proposed for a continuous process with multiple-input and multiple-output. CVRL is a generic hierarchically structured framework. The lower layer is composed with several groups of action units and continuous vector output can be produced based on action combination. The higher layer is a Q-learning unit defined on the space of combined action, its responsibility is the selection of properly combined actions. A detailed implementation of the CVRL is given, and the simulation on a mobile robot navigation problem demonstrates its effectiveness.
关键词
相关论文
Statistical Learning Theory
Yuhai Wu, Vladimir Vapnik
1999
Artificial intelligence: a modern approach
1995
Applied Nonlinear Control
Jean-Jacques Slotine, Weiping Li
1991
A new optimizer using particle swarm theory
R.C. Eberhart, James Kennedy
2002