首页 /研究 /Heuristically Accelerated State Backtracking Q-Learning Based on Cost Analysis
LEARNING

Heuristically Accelerated State Backtracking Q-Learning Based on Cost Analysis

Fang Mi

发表年份
2013
引用次数
3

摘要

Since action strategy learning is time-consuming for the reinforcement learning algorithm,a heuristic reinforcement learning algorithm is presented based on state backtracking. By analyzing the repetitive states and comparing the action policies of the reinforcement learning,a cost function is defined to indicate the importance of repetitive actions. A probability-based heuristic function is presented by combining an action reward with an action cost. The proposed algorithm reinforces the importance of an action to speed up learning by the heuristic function and measures the feasibility of an action to reduce unnecessary exploration by the cost function at the same time,thus the learning efficiency is steadily improve. This cost-based action strategy is proved to be reasonable. Two simulation scenarios are built and the experimental results of robot games prove that the proposed algorithm can learn by the tradeoff between rewards and costs,and effectively improve the convergence of Q-learning.

关键词

Reinforcement learningBacktrackingComputer scienceHeuristicArtificial intelligenceAction (physics)Q-learningConvergence (economics)Function (biology)Machine learning

相关论文

查看 LEARNING 分类全部论文