首页 /研究 /Optimizing a Continuum Manipulator’s Search Policy Through Model-Free Reinforcement Learning
MANIPULATION

Optimizing a Continuum Manipulator’s Search Policy Through Model-Free Reinforcement Learning

Chase G. Frazelle, Jonathan Rogers, Ioannis Karamouzas, Ian D. Walker

发表年份
2020
引用次数
8

摘要

Continuum robots have long held a great potential for applications in inspection of remote, hard-to-reach environments. In future environments such as the Deep Space Gateway, remote deployment of robotic solutions will require a high level of autonomy due to communication delays and unavailability of human crews. In this work, we explore the application of policy optimization methods through Actor-Critic gradient descent in order to optimize a continuum manipulator's search method for an unknown object. We show that we can deploy a continuum robot without prior knowledge of a goal object location and converge to a policy that finds the goal and can be reused in future deployments. We also show that the method can be quickly extended for multiple Degrees-of-Freedom and that we can restrict the policy with virtual and physical obstacles. These two scenarios are highlighted using a simulation environment with 15 and 135 unique states, respectively.

关键词

Reinforcement learningComputer scienceRobot manipulatorManipulator (device)Artificial intelligenceRobot

相关论文

查看 MANIPULATION 分类全部论文