首页 /研究 /Learning How Pedestrians Navigate: A Deep Inverse Reinforcement Learning Approach
HRI

Learning How Pedestrians Navigate: A Deep Inverse Reinforcement Learning Approach

Muhammad Fahad, Zhuo Chen, Yi Guo

发表年份
2018
引用次数
49

摘要

Humans and mobile robots will be increasingly cohabiting in the same environments, which has lead to an increase in studies on human robot interaction (HRI). One important topic in these studies is the development of robot navigation algorithms that are socially compliant to humans navigating in the same space. In this paper, we present a method to learn human navigation behaviors using maximum entropy deep inverse reinforcement learning (MEDIRL). We use a large open dataset of pedestrian trajectories collected in an uncontrolled environment as the expert demonstrations. Human navigation behaviors are captured by a nonlinear reward function through deep neural network (DNN) approximation. The developed MEDIRL algorithm takes feature inputs including social affinity map (SAM) that are extracted from human motion trajectories. We perform simulation experiments using the learned reward function, and the performance is evaluated comparing it with the real measured pedestrian trajectories in the dataset. The evaluation results show that the proposed method has acceptable prediction accuracy compared to other state-of-the-art methods, and it can generate pedestrian trajectories similar to real human trajectories with natural social navigation behaviors such as collision avoidance, leader-follower, and split-and-rejoin.

关键词

Computer scienceReinforcement learningArtificial intelligencePedestrianRobotMobile robotArtificial neural networkMachine learningTrajectoryEntropy (arrow of time)

相关论文

查看 HRI 分类全部论文