paper-with-me

홈 › Papers

Learning to Search via Retrospective Imitation

2018-04-03 · Jialin Song, Ravi Lanka, Albert Zhao, Aadyot Bhatnagar, Yisong Yue, Masahiro Ono

We study the problem of learning a good search policy for combinatorial search spaces. We propose retrospective imitation learning, which, after initial training by an expert, improves itself by learning from \textit{retrospective inspections} of its own roll-outs. That is, when the policy eventually reaches a feasible solution in a combinatorial search tree after making mistakes and backtracks, it retrospectively constructs an improved search trace to the solution by removing backtracks, which is then used to further train the policy. A key feature of our approach is that it can iteratively scale up, or transfer, to larger problem sizes than those solved by the initial expert demonstrations, thus dramatically expanding its applicability beyond that of conventional imitation learning. We showcase the effectiveness of our approach on a range of tasks, including synthetic maze solving and combinatorial problems expressed as integer programs.

📄 PDF Abstract BibTeX arXiv:1804.00846

Code (0)

등록된 구현이 없습니다.

Tasks

Imitation Learning

Similar Papers 제목 키워드 기반

Transdisciplinary AI Observatory -- Retrospective Analyses and Future-Oriented Contradistinctions

2020-11-26 · Nadisha-Marie Aliman, Leon Kester, Roman Yampolskiy

In the last years, AI safety gained international recognition in the light of heterogeneous safety-critical and ethical issues that risk overshadowing the broad beneficial impacts of AI. In this context, the implementati…

counterfactualDescriptive

Retrospective Causal Inference with Machine Learning Ensembles: An Application to Anti-Recidivism Policies in Colombia

2016-07-11 · Cyrus Samii, Laura Paler, Sarah Zukerman Daly

We present new methods to estimate causal effects retrospectively from micro data with the assistance of a machine learning ensemble. This approach overcomes two important limitations in conventional methods like regress…

BIG-bench Machine LearningCausal Inferenceregression

Reinforcement Learning for Branch-and-Bound Optimisation using Retrospective Trajectories

2022-05-28 · Christopher W. F. Parsonson, Alexandre Laterre, Thomas D. Barrett

Combinatorial optimisation problems framed as mixed integer linear programmes (MILPs) are ubiquitous across a range of real-world applications. The canonical branch-and-bound algorithm seeks to exactly solve MILPs by con…

Imitation Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Clinical Trial Active Learning

2023-07-20 · Zoe Fowler, Kiran Kokilepersaud, Mohit Prabhushankar, Ghassan AlRegib

This paper presents a novel approach to active learning that takes into account the non-independent and identically distributed (non-i.i.d.) structure of a clinical trial setting. There exists two types of clinical trial…

Active Learning

Temporal Leakage in Search-Engine Date-Filtered Web Retrieval: A Retrospective Forecasting Case Study

2026-01-31 · Ali El Lahib, Ying-Jieh Xia, Zehan Li, Yuxuan Wang 외 arxiv

Search-engine date filters are widely used to enforce pre-cutoff retrieval in retrospective evaluations of search-augmented forecasters. We show this approach is unreliable across two major search engines: auditing Googl…