paper-with-me

홈 › Papers

Multi-intention Inverse Q-learning for Interpretable Behavior Representation

2023-11-23 · Hao Zhu, Brice De La Crompe, Gabriel Kalweit, Artur Schneider, Maria Kalweit, Ilka Diester, Joschka Boedecker

In advancing the understanding of natural decision-making processes, inverse reinforcement learning (IRL) methods have proven instrumental in reconstructing animal's intentions underlying complex behaviors. Given the recent development of a continuous-time multi-intention IRL framework, there has been persistent inquiry into inferring discrete time-varying rewards with IRL. To address this challenge, we introduce the class of hierarchical inverse Q-learning (HIQL) algorithms. Through an unsupervised learning process, HIQL divides expert trajectories into multiple intention segments, and solves the IRL problem independently for each. Applying HIQL to simulated experiments and several real animal behavior datasets, our approach outperforms current benchmarks in behavior prediction and produces interpretable reward functions. Our results suggest that the intention transition dynamics underlying complex decision-making behavior is better modeled by a step function instead of a smoothly varying function. This advancement holds promise for neuroscience and cognitive science, contributing to a deeper understanding of decision-making and uncovering underlying brain mechanisms.

📄 PDF Abstract BibTeX arXiv:2311.13870

Code (1)

haozhu10015/hiql 공식 구현

Tasks

Decision MakingQ-Learning

Methods 이 논문이 사용한 방법론

Q-Learning Q-Learning is an off-policy temporal difference control algorithm: $$Q\left(S\_{t}, A\_{t}\right) \leftarrow Q\left(S\_{t}, A\_{t}\right) + \alpha\left[R_{t+1} +…

Similar Papers 제목 키워드 기반

CoMI-IRL: Contrastive Multi-Intention Inverse Reinforcement Learning

2026-02-07 · Antonio Mone, Frans A. Oliehoek, Luciano Cavalcante Siebert arxiv

Inverse Reinforcement Learning (IRL) seeks to infer reward functions from expert demonstrations. When demonstrations originate from multiple experts with different intentions, the problem is known as Multi-Intention IRL …

Reinforcement Learning

Foresight in Motion: Reinforcing Trajectory Prediction with Reward Heuristics

2025-07-16 · Muleilan Pei, Shaoshuai Shi, Xuesong Chen, Xu Liu 외 arxiv

Motion forecasting for on-road traffic agents presents both a significant challenge and a critical necessity for ensuring safety in autonomous driving systems. In contrast to most existing data-driven approaches that dir…

Reinforcement LearningTrajectory PredictionAutonomous DrivingMotion Forecasting

Inverse Decision Modeling: Learning Interpretable Representations of Behavior

2023-10-28 · Daniel Jarrett, Alihan Hüyük, Mihaela van der Schaar

Decision analysis deals with modeling and enhancing decision processes. A principal challenge in improving behavior is in obtaining a transparent description of existing behavior in the first place. In this paper, we dev…

Descriptive

Basis for Intentions: Efficient Inverse Reinforcement Learning using Past Experience

2022-08-09 · Marwa Abdulhai, Natasha Jaques, Sergey Levine

This paper addresses the problem of inverse reinforcement learning (IRL) -- inferring the reward function of an agent from observing its behavior. IRL can provide a generalizable and compact representation for apprentice…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Deep multi-intentional inverse reinforcement learning for cognitive multi-function radar inverse cognition

2024-08-16 · HanCong Feng, Kaili Jiang, Bin Tang

In recent years, radar systems have advanced significantly, offering environmental adaptation and multi-task capabilities. These developments pose new challenges for electronic intelligence (Elint) and electronic support…

reinforcement-learningReinforcement LearningTrajectory Clustering