paper-with-me

홈 › Papers

Meta-Learning surrogate models for sequential decision making

2019-03-28 · Alexandre Galashov, Jonathan Schwarz, Hyunjik Kim, Marta Garnelo, David Saxton, Pushmeet Kohli, S. M. Ali Eslami, Yee Whye Teh

We introduce a unified probabilistic framework for solving sequential decision making problems ranging from Bayesian optimisation to contextual bandits and reinforcement learning. This is accomplished by a probabilistic model-based approach that explains observed data while capturing predictive uncertainty during the decision making process. Crucially, this probabilistic model is chosen to be a Meta-Learning system that allows learning from a distribution of related problems, allowing data efficient adaptation to a target task. As a suitable instantiation of this framework, we explore the use of Neural processes due to statistical and computational desiderata. We apply our framework to a broad range of problem domains, such as control problems, recommender systems and adversarial attacks on RL agents, demonstrating an efficient and general black-box learning approach.

📄 PDF Abstract BibTeX arXiv:1903.11907

Code (0)

등록된 구현이 없습니다.

Tasks

Bayesian OptimisationDecision MakingMeta-LearningMulti-Armed BanditsRecommendation SystemsReinforcement LearningReinforcement Learning (RL)Sequential Decision Making

Similar Papers 제목 키워드 기반

Model-based Meta Reinforcement Learning using Graph Structured Surrogate Models

2021-02-16 · Qi Wang, Herke van Hoof

Reinforcement learning is a promising paradigm for solving sequential decision-making problems, but low data efficiency and weak generalization across tasks are bottlenecks in real-world applications. Model-based meta re…

Decision MakingMeta Reinforcement Learningreinforcement-learningReinforcement Learning+3

Meta-Prompt Optimization for LLM-Based Sequential Decision Making

2025-02-02 · Mingze Kong, Zhiyong Wang, Yao Shu, Zhongxiang Dai

Large language models (LLMs) have recently been employed as agents to solve sequential decision-making tasks such as Bayesian optimization and multi-armed bandits (MAB). These works usually adopt an LLM for sequential ac…

Bayesian OptimizationDecision MakingMulti-Armed BanditsSequential Decision Making

Effective Reinforcement Learning through Evolutionary Surrogate-Assisted Prescription

2020-02-13 · Olivier Francon, Santiago Gonzalez, Babak Hodjat, Elliot Meyerson 외

There is now significant historical data available on decision making in organizations, consisting of the decision problem, what decisions were made, and how desirable the outcomes were. Using this data, it is possible t…

Decision Makingreinforcement-learningReinforcement LearningReinforcement Learning (RL)+1

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments

2025-04-27 · Yun Qu, Qi Cheems Wang, Yixiu Mao, Yiqin Lv 외

Task robust adaptation is a long-standing pursuit in sequential decision-making. Some risk-averse strategies, e.g., the conditional value-at-risk principle, are incorporated in domain randomization or meta reinforcement …

Decision MakingDiversityMeta Reinforcement LearningSequential Decision Making

Unifying and Optimizing Data Values for Selection via Sequential-Decision-Making

2025-02-06 · Hongliang Chi, Qiong Wu, Zhengyi Zhou, Jonathan Light 외

Data selection has emerged as a crucial downstream application of data valuation. While existing data valuation methods have shown promise in selection tasks, the theoretical foundations and full potential of using data …

Data ValuationDecision MakingSequential Decision Making