首页 /研究 /Research on Multi-Robot Formation Control Based on MATD3 Algorithm
SWARM

Research on Multi-Robot Formation Control Based on MATD3 Algorithm

Conghang Zhou, Jianxing Li, Yujing Shi, Zhi-Rui Lin

发表年份
2023
引用次数
13
访问权限
开放获取

摘要

This paper investigates the problem of multi-robot formation control strategies in environments with obstacles based on deep reinforcement learning methods. To solve the problem of value function overestimation in the deep deterministic policy gradient (DDPG) algorithm, this paper proposes an improved multi-agent twin delayed deep deterministic policy gradient (MATD3) algorithm under the CTDE framework combined with the twin delayed deep deterministic policy gradient (TD3) algorithm, which adopts a prioritized experience replay strategy to improve the learning efficiency. For the problem of difficult obstacle avoidance for a robot formation, a hybrid reward mechanism is designed to use different formation maintenance strategies in obstacle areas and obstacle-free areas to achieve the control goal of obstacle avoidance by reasonably changing the formation. The simulation experiments verified the effectiveness of the multi-robot formation control strategy designed in this paper, and comparative simulations verified that the algorithm has a faster convergence speed and more stable performance.

关键词

Reinforcement learningObstacleComputer scienceObstacle avoidanceConvergence (economics)Control (management)RobotControl theory (sociology)Q-learningMathematical optimization

相关论文

查看 SWARM 分类全部论文