paper-with-me

홈 › Papers

Efficient Imitation under Misspecification

2025-03-17 · Nicolas Espinosa-Dice, Sanjiban Choudhury, Wen Sun, Gokul Swamy

We consider the problem of imitation learning under misspecification: settings where the learner is fundamentally unable to replicate expert behavior everywhere. This is often true in practice due to differences in observation space and action space expressiveness (e.g. perceptual or morphological differences between robots and humans). Given the learner must make some mistakes in the misspecified setting, interaction with the environment is fundamentally required to figure out which mistakes are particularly costly and lead to compounding errors. However, given the computational cost and safety concerns inherent in interaction, we'd like to perform as little of it as possible while ensuring we've learned a strong policy. Accordingly, prior work has proposed a flavor of efficient inverse reinforcement learning algorithms that merely perform a computationally efficient local search procedure with strong guarantees in the realizable setting. We first prove that under a novel structural condition we term reward-agnostic policy completeness, these sorts of local-search based IRL algorithms are able to avoid compounding errors. We then consider the question of where we should perform local search in the first place, given the learner may not be able to "walk on a tightrope" as well as the expert in the misspecified setting. We prove that in the misspecified setting, it is beneficial to broaden the set of states on which local search is performed to include those reachable by good policies the learner can actually play. We then experimentally explore a variety of sources of misspecification and how offline data can be used to effectively broaden where we perform local search from.

📄 PDF Abstract BibTeX arXiv:2503.13162

Code (0)

등록된 구현이 없습니다.

Tasks

Imitation Learning

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Robust PAC$^m$: Training Ensemble Models Under Misspecification and Outliers

2022-03-03 · Matteo Zecchin, Sangwoo Park, Osvaldo Simeone, Marios Kountouris 외

Standard Bayesian learning is known to have suboptimal generalization capabilities under misspecification and in the presence of outliers. PAC-Bayes theory demonstrates that the free energy criterion minimized by Bayesia…

Online Incident Response Planning under Model Misspecification through Bayesian Learning and Belief Quantization

2025-08-20 · Kim Hammar, Tao Li arxiv

Effective responses to cyberattacks require fast decisions, even when information about the attack is incomplete or inaccurate. However, most decision-support frameworks for incident response rely on a detailed system mo…

Federated Generalised Variational Inference: A Robust Probabilistic Federated Learning Framework

2025-02-02 · Terje Mildner, Oliver Hamelijnck, Paris Giampouras, Theodoros Damoulas

We introduce FedGVI, a probabilistic Federated Learning (FL) framework that is robust to both prior and likelihood misspecification. FedGVI addresses limitations in both frequentist and Bayesian FL by providing unbiased …

Federated LearningUncertainty QuantificationVariational Inference

Expectation Alignment: Handling Reward Misspecification in the Presence of Expectation Mismatch

2024-04-12 · Malek Mechergui, Sarath Sreedharan

Detecting and handling misspecified objectives, such as reward functions, has been widely recognized as one of the central challenges within the domain of Artificial Intelligence (AI) safety research. However, even with …

AI Agent

Dynamic treatment effects: high-dimensional inference under model misspecification

2021-11-12 · Yuqian Zhang, Weijie Ji, Jelena Bradic

Estimating dynamic treatment effects is crucial across various disciplines, providing insights into the time-dependent causal impact of interventions. However, this estimation poses challenges due to time-varying confoun…