paper-with-me

홈 › Papers

On Covariate Shift of Latent Confounders in Imitation and Reinforcement Learning

2021-10-13 · ICLR 2022 4 · Guy Tennenholtz, Assaf Hallak, Gal Dalal, Shie Mannor, Gal Chechik, Uri Shalit

We consider the problem of using expert data with unobserved confounders for imitation and reinforcement learning. We begin by defining the problem of learning from confounded expert data in a contextual MDP setup. We analyze the limitations of learning from such data with and without external reward, and propose an adjustment of standard imitation learning algorithms to fit this setup. We then discuss the problem of distribution shift between the expert data and the online environment when the data is only partially observable. We prove possibility and impossibility results for imitation learning under arbitrary distribution shift of the missing covariates. When additional external reward is provided, we propose a sampling procedure that addresses the unknown shift and prove convergence to an optimal solution. Finally, we validate our claims empirically on challenging assistive healthcare and recommender system simulation tasks.

📄 PDF Abstract BibTeX arXiv:2110.06539

Code (0)

등록된 구현이 없습니다.

Tasks

Imitation LearningRecommendation Systemsreinforcement-learningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Proxy Methods for Domain Adaptation

2024-03-12 · Katherine Tsai, Stephen R. Pfohl, Olawale Salaudeen, Nicole Chiou 외

We study the problem of domain adaptation under distribution shift, where the shift is due to a change in the distribution of an unobserved, latent variable that confounds both the covariates and the labels. In this sett…

Domain Adaptation

Point-Identification of a Robust Predictor Under Latent Shift with Imperfect Proxies

2026-03-16 · Zahra Rahiminasab, Reza Soumi, Arto Klami, Samuel Kaski arxiv

Addressing the domain adaptation problem becomes more challenging when distribution shifts across domains stem from latent confounders that affect both covariates and outcomes. Existing proxy-based approaches that addres…

Domain AdaptationActive Learning

DITTO: Offline Imitation Learning with World Models

2023-02-06 · Branton DeMoss, Paul Duckworth, Nick Hawes, Ingmar Posner

We propose DITTO, an offline imitation learning algorithm which uses world models and on-policy reinforcement learning to addresses the problem of covariate shift, without access to an oracle or any additional online int…

Imitation Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Causal Inference using Gaussian Processes with Structured Latent Confounders

2020-07-14 · ICML 2020 1 · Sam Witty, Kenta Takatsu, David Jensen, Vikash Mansinghka

Latent confounders---unobserved variables that influence both treatment and outcome---can bias estimates of causal effects. In some cases, these confounders are shared across observations, e.g. all students taking a cour…

Causal InferenceGaussian Processes

Towards Backwards-Compatible Data with Confounded Domain Adaptation

2022-03-23 · Calvin Mccarter

Most current domain adaptation methods address either covariate shift or label shift, but are not applicable where they occur simultaneously and are confounded with each other. Domain adaptation approaches which do accou…

Domain Adaptation