paper-with-me

홈 › Papers

GFlowNets for AI-Driven Scientific Discovery

2023-02-01 · Moksh Jain, Tristan Deleu, Jason Hartford, Cheng-Hao Liu, Alex Hernandez-Garcia, Yoshua Bengio

Tackling the most pressing problems for humanity, such as the climate crisis and the threat of global pandemics, requires accelerating the pace of scientific discovery. While science has traditionally relied on trial and error and even serendipity to a large extent, the last few decades have seen a surge of data-driven scientific discoveries. However, in order to truly leverage large-scale data sets and high-throughput experimental setups, machine learning methods will need to be further improved and better integrated in the scientific discovery pipeline. A key challenge for current machine learning methods in this context is the efficient exploration of very large search spaces, which requires techniques for estimating reducible (epistemic) uncertainty and generating sets of diverse and informative experiments to perform. This motivated a new probabilistic machine learning framework called GFlowNets, which can be applied in the modeling, hypotheses generation and experimental design stages of the experimental science loop. GFlowNets learn to sample from a distribution given indirectly by a reward function corresponding to an unnormalized probability, which enables sampling diverse, high-reward candidates. GFlowNets can also be used to form efficient and amortized Bayesian posterior estimators for causal models conditioned on the already acquired experimental data. Having such posterior models can then provide estimators of epistemic uncertainty and information gain that can drive an experimental design policy. Altogether, here we will argue that GFlowNets can become a valuable tool for AI-driven scientific discovery, especially in scenarios of very large candidate spaces where we have access to cheap but inaccurate measurements or to expensive but accurate measurements. This is a common setting in the context of drug and material discovery, which we use as examples throughout the paper.

📄 PDF Abstract BibTeX arXiv:2302.00615

Code (0)

등록된 구현이 없습니다.

Tasks

Efficient ExplorationExperimental Designscientific discovery

Similar Papers 제목 키워드 기반

Multi-Fidelity Active Learning with GFlowNets

2023-06-20 · Alex Hernandez-Garcia, Nikita Saxena, Moksh Jain, Cheng-Hao Liu 외

In the last decades, the capacity to generate large amounts of data in science and engineering applications has been growing steadily. Meanwhile, machine learning has progressed to become a suitable tool to process and u…

Active LearningBayesian Optimizationscientific discovery

An Empirical Study of the Effectiveness of Using a Replay Buffer on Mode Discovery in GFlowNets

2023-07-15 · Nikhil Vemgal, Elaine Lau, Doina Precup

Reinforcement Learning (RL) algorithms aim to learn an optimal policy by iteratively sampling actions to learn how to maximize the total expected return, $R(x)$. GFlowNets are a special class of algorithms designed to ge…

Drug DiscoveryReinforcement Learning (RL)

Routing by Reaching: Composition of Pre-trained GFlowNets for Multi-Objective Generation

2026-02-25 · Seokwon Yoon, Youngbin Choi, Seunghyuk Cho, Seungbeom Lee 외 arxiv

Generative Flow Networks (GFlowNets) learn to sample diverse candidates in proportion to a reward function, making them well-suited for scientific discovery, where exploring multiple promising solutions is crucial. Furth…

Planning-Augmented Sampling with Early Guidance for High-Reward Discovery

2025-10-01 · Rui Zhu, Yudong Zhang, Xuan Yu, Chen Zhang 외 arxiv

Generative Flow Networks (GFlowNets) enable structured generation with inherent diversity, but existing sampling strategies often rely on weak guided exploration, slowing early discovery of high-reward candidates. In tas…

Loss-Guided Auxiliary Agents for Overcoming Mode Collapse in GFlowNets

2025-05-21 · Idriss Malek, Abhijit Sharma, Salem Lahlou

Although Generative Flow Networks (GFlowNets) are designed to capture multiple modes of a reward function, they often suffer from mode collapse in practice, getting trapped in early discovered modes and requiring prolong…

Diversityvalid