首页 /研究 /RRTPI: Policy iteration on continuous domains using rapidly-exploring random trees
LEARNING

RRTPI: Policy iteration on continuous domains using rapidly-exploring random trees

Manimaran Sivasamy Sivamurugan, Balaraman Ravindran

发表年份
2014
引用次数
2

摘要

Path planning in continuous spaces has been a central problem in robotics. In the case of systems with complex dynamics, the performance of sampling based techniques relies on identifying a good approximation to the cost-to-go distance metric. We propose a technique that uses reinforcement learning to learn this distance metric on the fly from samples and combine it with existing sampling based planners to produce near optimal solutions. The resulting algorithm - RRTPI can solve problems with complex dynamics in a sample efficient manner while preserving asymptotic guarantees. We provide experimental evaluation of this technique on domains with underactuated and underpowered dynamics.

关键词

Metric (unit)Reinforcement learningComputer scienceSampling (signal processing)RoboticsMathematical optimizationPath (computing)Motion planningSample (material)Dynamics (music)

相关论文

查看 LEARNING 分类全部论文