Boosting Deep Reinforcement Learning-Based Path Planning for Robotic Manipulators With Egocentric State Space Descriptions
Sven Weishaupt, Ricus Husmann, Harald Aschemann
- 发表年份
- 2024
- 引用次数
- 2
摘要
In robotic path planning tasks, Reinforcement Learning agents typically receive global or relative Euclidean coordinates, e.g., with respect to a target reference point as direct state information. Nevertheless, a more egocentric view of the environment seems to be favorable – based on information in polar or spherical coordinates about objects surrounding the robot. Using the model-free, actor-critic algorithm Twin Delayed Deep Deterministic Policy Gradient in combination with Prioritized Experience Replay, the advantages of an alternative definition of states using egocentric TCP-coordinates is evaluated and compared in simulations to classical approaches within two typical environments. The training results indicate a tremendous potential of the egocentric state space definition that not only offers faster learning but also more successful trainings.
关键词
相关论文
Statistical Learning Theory
Yuhai Wu, Vladimir Vapnik
1999
Artificial intelligence: a modern approach
1995
Fractional Differential Equations
Igor Podlubný
2025
Applied Nonlinear Control
Jean-Jacques Slotine, Weiping Li
1991