Multi-robot cooperative planning by consensus Q-learning
Arup Kumar Sadhu, Amit Konar, Bonny Banerjee, Atulya K. Nagar
- Year
- 2017
- Citations
- 3
Abstract
Multi-robot cooperation entails planning by multiple robots for a common objective, where each robot/agent actuates upon the environment-based on the sensory information received from the environment. Multi-robot cooperation employing equilibrium-based reinforcement learning is optimal in the sense of system resource (time and/or energy) utilization, because of the prior adaption of the environment by the robots. Unfortunately, robots cannot enjoy such benefit of reinforcement learning in presence of multiple types of equilibria (here Nash equilibrium or correlated equilibrium). In the above perspective, robots need to adapt with a strategy, so that robots can select the optimal equilibrium in each step of the learning. The paper proposes consensus-based multi-agent Q-learning to address the bottleneck of the optimal equilibrium selection among multiple types. An analysis reveals that a consensus (joint action) is coordination type pure strategy Nash equilibrium as well as pure strategy correlated equilibrium. The superiority of the proposed consensus-based multi-agent Q-learning algorithm over the traditional reference algorithms in terms of the average reward collection is shown in the experimental section. In addition, the proposed consensus-based planning algorithm is also verified considering multi-robot stick-carrying problem as a benchmark.
Keywords
Related papers
Statistical Learning Theory
Yuhai Wu, Vladimir Vapnik
1999
Artificial intelligence: a modern approach
1995
Applied Nonlinear Control
Jean-Jacques Slotine, Weiping Li
1991
A new optimizer using particle swarm theory
R.C. Eberhart, James Kennedy
2002