首页 /研究 /Advice generation from observed execution: abstract Markov decision process learning
OTHER

Advice generation from observed execution: abstract Markov decision process learning

Patrick Riley, Manuela Veloso

发表年份
2004
引用次数
8

摘要

An advising agent, a coach, provides advice to other agents about how to act. In this paper we contribute an advice generation method using observations of agents acting in an environment. Given an abstract state definition and partially specified abstract actions, the algorithm extracts a Markov Chain, infers a Markov Decision Process, and then solves the MDP (given an arbitrary reward signal) to generate advice. We evaluate our work in a simulated robot soccer environment and experimental results show improved agent performance when using the advice generated from the MDP for both a sub-task and the full soccer game.

关键词

Advice (programming)Markov decision processComputer scienceMarkov chainMarkov processProcess (computing)Task (project management)Partially observable Markov decision processRobotArtificial intelligence

相关论文

查看 OTHER 分类全部论文