Multi-Task Policy Search

Marc Peter Deisenroth, Péter Englert, Jan Peters, Dieter Fox

发表年份: 2013
引用次数: 3
访问权限: 开放获取

摘要

Learning policies that generalize across multiple tasks is an important and challenging research topic in reinforcement learning and robotics. Training individual policies for every single potential task is often impractical, especially for continuous task variations, requiring more principled approaches to share and transfer knowledge among similar tasks. We present a novel approach for learning a nonlinear feedback policy that generalizes across multiple tasks. The key idea is to define a parametrized policy as a function of both the state and the task, which allows learning a single policy that generalizes across multiple known and unknown tasks. Applications of our novel approach to reinforcement and imitation learning in real-robot experiments are shown.

关键词

Reinforcement learningTask (project management)Computer scienceArtificial intelligencePolicy learningImitationKey (lock)Function (biology)Transfer of learningRobotics

Multi-Task Policy Search

摘要

关键词

相关论文

Statistical Learning Theory

Artificial intelligence: a modern approach

Applied Nonlinear Control

A new optimizer using particle swarm theory