paper-with-me

홈 › Papers

Amortising Bayesian Experimental Design for Sequential Information Gathering in LLMs

2026-07-03 · Jakob Hartmann, James Harvey, Jhonathan Navott, Erik Y. Wang, Luckeciano C. Melo, Flaviu Cipcigan, Cheng Zhang, Alessandro Abate arxiv

Large language models (LLMs) exhibit strong reasoning and world-knowledge capabilities, yet often struggle to gather information effectively across the multi-turn interactions required in sequential decision-making settings. We introduce Amortised Sequential Information Gathering (ASIG), a fine-tuning approach that amortises Bayesian Experimental Design (BED) into LLM policies via a multi-turn extension of Group Relative Policy Optimisation with an Expected Information Gain reward. Evaluated on the 20 Questions task, ASIG more than doubles the success rate of the 7B base model and reduces inference cost by over $25\times$ relative to BED-LLM, a competitive inference-time baseline. Applied to MediQ, a medical diagnosis benchmark unseen during training, ASIG improves information-seeking performance at the 7B scale, suggesting that the learned strategies can transfer out of distribution. Our findings show that amortising BED into LLM policies provides an effective and computationally efficient approach to sequential information gathering.

📄 PDF Abstract BibTeX arXiv:2607.03426

Code (0)

등록된 구현이 없습니다.

Tasks

Medical Diagnosis

Similar Papers 제목 키워드 기반

Deep Adaptive Design: Amortizing Sequential Bayesian Experimental Design

2021-03-03 · Adam Foster, Desi R. Ivanova, Ilyas Malik, Tom Rainforth

We introduce Deep Adaptive Design (DAD), a method for amortizing the cost of adaptive Bayesian experimental design that allows experiments to be run in real-time. Traditional sequential Bayesian optimal experimental desi…

Experimental Design

Sequential Bayesian Experimental Design for Implicit Models via Mutual Information

2020-03-20 · Steven Kleinegesse, Christopher Drovandi, Michael U. Gutmann

Bayesian experimental design (BED) is a framework that uses statistical models and decision making under uncertainty to optimise the cost and performance of a scientific experiment. Sequential BED, as opposed to static B…

Bayesian OptimisationDecision MakingDecision Making Under UncertaintyExperimental Design+1

Sequential Bayesian experimental designs via reinforcement learning

2022-02-14 · Hikaru Asano

Bayesian experimental design (BED) has been used as a method for conducting efficient experiments based on Bayesian inference. The existing methods, however, mostly focus on maximizing the expected information gain (EIG)…

Bayesian InferenceDecision MakingExperimental Designreinforcement-learning+3

Gradient-Free Sequential Bayesian Experimental Design via Interacting Particle Systems

2025-04-17 · Robert Gruhlke, Matei Hanu, Claudia Schillings, Philipp Wacker

We introduce a gradient-free framework for Bayesian Optimal Experimental Design (BOED) in sequential settings, aimed at complex systems where gradient information is unavailable. Our method combines Ensemble Kalman Inver…

Experimental Design

Amortising Inference and Meta-Learning Priors in Neural Networks

2026-02-09 · Tommy Rochussen, Vincent Fortuin arxiv

One of the core facets of Bayesianism is in the updating of prior beliefs in light of new evidence$\text{ -- }$so how can we maintain a Bayesian approach if we have no prior beliefs in the first place? This is one of the…