首页 /研究 /An Improved Reinforcement Q-Learning Method with BP Neural Networks in Robot Soccer

LEARNING

An Improved Reinforcement Q-Learning Method with BP Neural Networks in Robot Soccer

Shi-chao Wang, Zhengxi Song, Hao Ding, Haobin Shi

发表年份: 2011
引用次数: 8

摘要

In traditional reinforcement Q-Learning method, there exists two problems: difficulty of dividing the state information, complexity of extreme large dimension input. To solve these two problems, this paper proposed an improved reinforcement Q-Learning method with BP neutral network. In this method, the large Q table is replaced by a BP neural network. Continuous environmental information is the input. The Q value is the output. The Q value and weight of the network are also adjusted by the action rewards. This paper presents an algorithm for single agent's action selection. Simulation shows proposed method is more stable and applicable for the agent's strategy selection.

关键词

Reinforcement learningArtificial neural networkComputer scienceArtificial intelligenceDimension (graph theory)Action (physics)Action selectionQ-learningSelection (genetic algorithm)Reinforcement

An Improved Reinforcement Q-Learning Method with BP Neural Networks in Robot Soccer

摘要

关键词

相关论文

Statistical Learning Theory

Artificial intelligence: a modern approach

Applied Nonlinear Control

A new optimizer using particle swarm theory