paper-with-me

홈 › Papers

Trajectory Modeling via Random Utility Inverse Reinforcement Learning

2021-05-25 · Anselmo R. Pitombeira-Neto, Helano P. Santos, Ticiana L. Coelho da Silva, José Antonio F. de Macedo

We consider the problem of modeling trajectories of drivers in a road network from the perspective of inverse reinforcement learning. Cars are detected by sensors placed on sparsely distributed points on the street network of a city. As rational agents, drivers are trying to maximize some reward function unknown to an external observer. We apply the concept of random utility from econometrics to model the unknown reward function as a function of observed and unobserved features. In contrast to current inverse reinforcement learning approaches, we do not assume that agents act according to a stochastic policy; rather, we assume that agents act according to a deterministic optimal policy and show that randomness in data arises because the exact rewards are not fully observed by an external observer. We introduce the concept of extended state to cope with unobserved features and develop a Markov decision process formulation of drivers decisions. We present theoretical results which guarantee the existence of solutions and show that maximum entropy inverse reinforcement learning is a particular case of our approach. Finally, we illustrate Bayesian inference on model parameters through a case study with real trajectory data from a large city in Brazil.

📄 PDF Abstract BibTeX arXiv:2105.12092

Code (0)

등록된 구현이 없습니다.

Tasks

Bayesian InferenceEconometricsreinforcement-learningReinforcement LearningReinforcement Learning (RL)Trajectory Modeling

Similar Papers 제목 키워드 기반

Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization

2025-07-06 · Vikram Krishnamurthy arxiv

This monograph, spanning three chapters, explores Inverse Reinforcement Learning (IRL). The first two chapters view inverse reinforcement learning (IRL) through the lens of revealed preferences from microeconomics while …

Stochastic OptimizationReinforcement Learning

RIDM: Reinforced Inverse Dynamics Modeling for Learning from a Single Observed Demonstration

2019-06-18 · Brahma S. Pavse, Faraz Torabi, Josiah P. Hanna, Garrett Warnell 외

Augmenting reinforcement learning with imitation learning is often hailed as a method by which to improve upon learning from scratch. However, most existing methods for integrating these two techniques are subject to sev…

Imitation Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Inverse-Inverse Reinforcement Learning. How to Hide Strategy from an Adversarial Inverse Reinforcement Learner

2022-05-22 · Kunal Pattanayak, Vikram Krishnamurthy, Christopher Berry

Inverse reinforcement learning (IRL) deals with estimating an agent's utility function from its actions. In this paper, we consider how an agent can hide its strategy and mitigate an adversarial IRL attack; we call this …

Reinforcement Learning (RL)

Active Learning for Risk-Sensitive Inverse Reinforcement Learning

2019-09-14 · Rui Chen, Wenshuo Wang, Zirui Zhao, Ding Zhao

One typical assumption in inverse reinforcement learning (IRL) is that human experts act to optimize the expected utility of a stochastic cost with a fixed distribution. This assumption deviates from actual human behavio…

Active Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Cost Function Estimation Using Inverse Reinforcement Learning with Minimal Observations

2025-05-13 · Sarmad Mehrdad, Avadesh Meduri, Ludovic Righetti

We present an iterative inverse reinforcement learning algorithm to infer optimal cost functions in continuous spaces. Based on a popular maximum entropy criteria, our approach iteratively finds a weight improvement step…