Learning Implicit Social Navigation Behavior Using Deep Inverse Reinforcement Learning

  • Kathuria, Tribhi
  • Liu, Ke
  • Jang, Junwoo
  • Yang, X. Jessie
  • Ghaffari, Maani
Citations

WEB OF SCIENCE

1
Citations

SCOPUS

1

초록

This paper reports on learning a reward map for social navigation in dynamic environments where the robot can reason about its path at any time, given agent trajectories and scene geometry. Humans navigating in dense and dynamic indoor environments often work with several implied social rules. A rule-based approach fails to model all possible interactions between humans, robots, and scenes. We propose a novel Smooth Maximum Entropy Deep Inverse Reinforcement Learning (S-MEDIRL) algorithm that can extrapolate beyond expert demos to better encode scene navigability from few-shot demonstrations. The agent learns to predict the cost maps based on trajectory data as well as scene geometry. The trajectory sampled from the learned cost map is then executed using a local crowd navigation controller. We present results in a photo-realistic simulation environment, with a robot and a human navigating a narrow crossing scenario. The robot implicitly learns to exhibit social behaviors such as yielding to oncoming traffic and avoiding deadlocks. We compare the proposed approach to the popular model-based crowd navigation algorithm ORCA and a rule-based agent that exhibits yielding.

키워드

NavigationRobotsTrajectoryReinforcement learningGeometryPlanningTrainingSystem recoveryEntropyCostssocial HRIlearning from demonstrationdeep learning methods
제목
Learning Implicit Social Navigation Behavior Using Deep Inverse Reinforcement Learning
저자
Kathuria, TribhiLiu, KeJang, JunwooYang, X. JessieGhaffari, Maani
DOI
10.1109/LRA.2025.3557299
발행일
2025-05
유형
Article
저널명
IEEE Robotics and Automation Letters
10
5
페이지
5146 ~ 5153