paper-with-me

홈 › Papers

Enhancing Solution Efficiency in Reinforcement Learning: Leveraging Sub-GFlowNet and Entropy Integration

2024-10-01 · Siyi He

Traditional reinforcement learning often struggles to generate diverse, high-reward solutions, especially in domains like drug design and black-box function optimization. Markov Chain Monte Carlo (MCMC) methods provide an alternative method of RL in candidate selection but suffer from high computational costs and limited candidate diversity exploration capabilities. In response, GFlowNet, a novel neural network architecture, was introduced to model complex system dynamics and generate diverse high-reward trajectories. To further enhance this approach, this paper proposes improvements to GFlowNet by introducing a new loss function and refining the training objective associated with sub-GFlowNet. These enhancements aim to integrate entropy and leverage network structure characteristics, improving both candidate diversity and computational efficiency. We demonstrated the superiority of the refined GFlowNet over traditional methods by empirical results from hypergrid experiments and molecule synthesis tasks. The findings underscore the effectiveness of incorporating entropy and exploiting network structure properties in solution generation in molecule synthesis as well as diverse experimental designs.

📄 PDF Abstract BibTeX arXiv:2410.00461

Code (0)

등록된 구현이 없습니다.

Tasks

Computational EfficiencyDiversityDrug Design

Similar Papers 제목 키워드 기반

EMERGENT: Efficient and Manipulation-resistant Matching using GFlowNets

2025-05-22 · Mayesha Tasnim, Erman Acar, Sennay Ghebreab

The design of fair and efficient algorithms for allocating public resources, such as school admissions, housing, or medical residency, has a profound social impact. In one-sided matching problems, where individuals are a…

Improving GFlowNets with Monte Carlo Tree Search

2024-06-19 · Nikita Morozov, Daniil Tiapkin, Sergey Samsonov, Alexey Naumov 외

Generative Flow Networks (GFlowNets) treat sampling from distributions over compositional discrete spaces as a sequential decision-making problem, training a stochastic policy to construct objects step by step. Recent st…

Accurate and Diverse LLM Mathematical Reasoning via Automated PRM-Guided GFlowNets

2025-04-28 · Adam Younsi, Abdalgader Abubaker, Mohamed El Amine Seddik, Hakim Hacid 외

Achieving both accuracy and diverse reasoning remains challenging for Large Language Models (LLMs) in complex domains like mathematics. A key bottleneck is evaluating intermediate reasoning steps to guide generation with…

Data AugmentationDiversityMathMathematical Reasoning

GFlowNet Training by Policy Gradients

2024-08-12 · Puhua Niu, Shili Wu, Mingzhou Fan, Xiaoning Qian

Generative Flow Networks (GFlowNets) have been shown effective to generate combinatorial objects with desired properties. We here propose a new GFlowNet training framework, with policy-dependent rewards, that bridges kee…

Reinforcement Learning (RL)

Rectifying Reinforcement Learning for Reward Matching

2024-06-04 · Haoran He, Emmanuel Bengio, Qingpeng Cai, Ling Pan

The Generative Flow Network (GFlowNet) is a probabilistic framework in which an agent learns a stochastic policy and flow functions to sample objects with probability proportional to an unnormalized reward function. GFlo…

Decision Makingreinforcement-learningReinforcement LearningReinforcement Learning (RL)+1