paper-with-me

Papers

Imitation Learning from Observations under Transition Model Disparity

2022-04-25 · ICLR 2022 4 · Tanmay Gangwani, Yuan Zhou, Jian Peng

Learning to perform tasks by leveraging a dataset of expert observations, also known as imitation learning from observations (ILO), is an important paradigm for learning skills without access to the expert reward function or the expert actions. We consider ILO in the setting where the expert and the learner agents operate in different environments, with the source of the discrepancy being the transition dynamics model. Recent methods for scalable ILO utilize adversarial learning to match the state-transition distributions of the expert and the learner, an approach that becomes challenging when the dynamics are dissimilar. In this work, we propose an algorithm that trains an intermediary policy in the learner environment and uses it as a surrogate expert for the learner. The intermediary policy is learned such that the state transitions generated by it are close to the state transitions in the expert dataset. To derive a practical and scalable algorithm, we employ concepts from prior work on estimating the support of a probability distribution. Experiments using MuJoCo locomotion tasks highlight that our method compares favorably to the baselines for ILO with transition dynamics mismatch.

📄 PDF Abstract BibTeX arXiv:2204.11446

Code (1)

tgangwani/ailo 공식 구현

Tasks

Imitation LearningmodelMuJoCo

Similar Papers 제목 키워드 기반

A Novel Factor Graph-Based Optimization Technique for Stereo Correspondence Estimation

2021-09-22 · Hanieh Shabanian, Madhusudhanan Balasubramanian

Dense disparities among multiple views is essential for estimating the 3D architecture of a scene based on the geometrical relationship among the scene and the views or cameras. Scenes with larger extents of heterogeneou…

3D ArchitectureDisparity EstimationOptical Flow Estimation

Evaluating Model Retraining under Drift: Paired Comparisons of Cumulative Subgroup Disparity

2026-09-09 · Aaron Ceross arxiv

Choosing when to retrain a deployed classifier requires assessing subgroup error rates across the sequence of models used, including periods between updates. We compare complete scheduled, loss-triggered, and subgroup-ga…

Scale-Consistent Fusion: from Heterogeneous Local Sampling to Global Immersive Rendering

2021-06-17 · Wenpeng Xing, Jie Chen, Zaifeng Yang, Qiang Wang

Image-based geometric modeling and novel view synthesis based on sparse, large-baseline samplings are challenging but important tasks for emerging multimedia applications such as virtual reality and immersive telepresenc…

Novel View Synthesis

MAP Disparity Estimation Using Hidden Markov Trees

2015-12-01 · ICCV 2015 12 · Eric T. Psota, Jedrzej Kowalczuk, Mateusz Mittek, Lance C. Perez

A new method is introduced for stereo matching that operates on minimum spanning trees (MSTs) generated from the images. Disparity maps are represented as a collection of hidden states on MSTs, and each MST is modeled as…

Disparity EstimationStereo MatchingStereo Matching Hand

Object Disparity

2021-08-18 · Ynjiun Paul Wang

Most of stereo vision works are focusing on computing the dense pixel disparity of a given pair of left and right images. A camera pair usually required lens undistortion and stereo calibration to provide an undistorted …

Object