首页 /研究 /Robot Learning in Partially Observable, Noisy, Continuous Worlds
LEARNING

Robot Learning in Partially Observable, Noisy, Continuous Worlds

Robert Emer Broadbent, Todd Peterson

发表年份
2006
引用次数
3

摘要

Partially-observable Markov decision problems (POMDPs) pose special difficulties for the task of learning robot control policies, due to the need to disambiguate perceptually aliased states. Short-term memories of recent actions and/or percepts are required to provide context for the robot to perform such disambiguation. We introduce Variable-Resolution Percept Discretization (VRPD) as an extension to Utile Suffix Memory (USM), an algorithm designed to solve discrete POMDPs. This extension allows USM to function effectively in noisy, continuous worlds. We describe the extension in detail, then we demonstrate experimentally the improvements that it makes to USM in the context of continuous POMDPs.

关键词

Computer scienceExtension (predicate logic)Context (archaeology)ObservableArtificial intelligenceDiscretizationMarkov decision processRobotPartially observable Markov decision processVariable (mathematics)

相关论文

查看 LEARNING 分类全部论文