Home /Research /RRTPI: Policy iteration on continuous domains using rapidly-exploring random trees
LEARNING

RRTPI: Policy iteration on continuous domains using rapidly-exploring random trees

Manimaran Sivasamy Sivamurugan, Balaraman Ravindran

Year
2014
Citations
2

Abstract

Path planning in continuous spaces has been a central problem in robotics. In the case of systems with complex dynamics, the performance of sampling based techniques relies on identifying a good approximation to the cost-to-go distance metric. We propose a technique that uses reinforcement learning to learn this distance metric on the fly from samples and combine it with existing sampling based planners to produce near optimal solutions. The resulting algorithm - RRTPI can solve problems with complex dynamics in a sample efficient manner while preserving asymptotic guarantees. We provide experimental evaluation of this technique on domains with underactuated and underpowered dynamics.

Keywords

Metric (unit)Reinforcement learningComputer scienceSampling (signal processing)RoboticsMathematical optimizationPath (computing)Motion planningSample (material)Dynamics (music)

Related papers

Browse all LEARNING papers