Learning Implicit Social Navigation Behavior Using Deep Inverse Reinforcement Learning

Kathuria, Tribhi; Liu, Ke; Jang, Junwoo; Yang, X. Jessie; Ghaffari, Maani

doi:10.1109/LRA.2025.3557299

상세 보기

Learning Implicit Social Navigation Behavior Using Deep Inverse Reinforcement Learning

Kathuria, Tribhi;
Liu, Ke;
Jang, Junwoo;
Yang, X. Jessie;
Ghaffari, Maani

Citations

WEB OF SCIENCE

1

Citations

SCOPUS

1

초록

This paper reports on learning a reward map for social navigation in dynamic environments where the robot can reason about its path at any time, given agent trajectories and scene geometry. Humans navigating in dense and dynamic indoor environments often work with several implied social rules. A rule-based approach fails to model all possible interactions between humans, robots, and scenes. We propose a novel Smooth Maximum Entropy Deep Inverse Reinforcement Learning (S-MEDIRL) algorithm that can extrapolate beyond expert demos to better encode scene navigability from few-shot demonstrations. The agent learns to predict the cost maps based on trajectory data as well as scene geometry. The trajectory sampled from the learned cost map is then executed using a local crowd navigation controller. We present results in a photo-realistic simulation environment, with a robot and a human navigating a narrow crossing scenario. The robot implicitly learns to exhibit social behaviors such as yielding to oncoming traffic and avoiding deadlocks. We compare the proposed approach to the popular model-based crowd navigation algorithm ORCA and a rule-based agent that exhibits yielding.

키워드

Navigation; Robots; Trajectory; Reinforcement learning; Geometry; Planning; Training; System recovery; Entropy; Costs; social HRI; learning from demonstration; deep learning methods

제목: Learning Implicit Social Navigation Behavior Using Deep Inverse Reinforcement Learning

저자: Kathuria, Tribhi; Liu, Ke; Jang, Junwoo; Yang, X. Jessie; Ghaffari, Maani

DOI: 10.1109/LRA.2025.3557299

발행일: 2025-05

유형: Article

저널명: IEEE Robotics and Automation Letters

권: 10

호: 5

페이지: 5146 ~ 5153