paper-with-me

Papers

Nuclear Microreactor Control with Deep Reinforcement Learning

2025-03-31 · Leo Tunkle, Kamal Abdulraheem, Linyu Lin, Majdi I. Radaideh

The economic feasibility of nuclear microreactors will depend on minimizing operating costs through advancements in autonomous control, especially when these microreactors are operating alongside other types of energy systems (e.g., renewable energy). This study explores the application of deep reinforcement learning (RL) for real-time drum control in microreactors, exploring performance in regard to load-following scenarios. By leveraging a point kinetics model with thermal and xenon feedback, we first establish a baseline using a single-output RL agent, then compare it against a traditional proportional-integral-derivative (PID) controller. This study demonstrates that RL controllers, including both single- and multi-agent RL (MARL) frameworks, can achieve similar or even superior load-following performance as traditional PID control across a range of load-following scenarios. In short transients, the RL agent was able to reduce the tracking error rate in comparison to PID. Over extended 300-minute load-following scenarios in which xenon feedback becomes a dominant factor, PID maintained better accuracy, but RL still remained within a 1% error margin despite being trained only on short-duration scenarios. This highlights RL's strong ability to generalize and extrapolate to longer, more complex transients, affording substantial reductions in training costs and reduced overfitting. Furthermore, when control was extended to multiple drums, MARL enabled independent drum control as well as maintained reactor symmetry constraints without sacrificing performance -- an objective that standard single-agent RL could not learn. We also found that, as increasing levels of Gaussian noise were added to the power measurements, the RL controllers were able to maintain lower error rates than PID, and to do so with less control effort.

📄 PDF Abstract BibTeX arXiv:2504.00156

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Multistep Criticality Search and Power Shaping in Microreactors with Reinforcement Learning

2024-06-22 · Majdi I. Radaideh, Leo Tunkle, Dean Price, Kamal Abdulraheem 외

Reducing operation and maintenance costs is a key objective for advanced reactors in general and microreactors in particular. To achieve this reduction, developing robust autonomous control algorithms is essential to ens…

energy managementReinforcement Learning (RL)

Evaluation of Nuclear Microreactor Cost-competitiveness in Current Electricity Markets Considering Reactor Cost Uncertainties

2025-06-16 · Muhammad R. Abdusammi, Ikhwan Khaleb, Fei Gao, Aditi Verma

This paper evaluates the cost competitiveness of microreactors in today's electricity markets, with a focus on uncertainties in reactor costs. A Genetic Algorithm (GA) is used to optimize key technical parameters, such a…

Techno-economic optimization of a heat-pipe microreactor, part II: multi-objective optimization analysis

2026-01-27 · Paul Seurin, Dean Price arxiv

Heat-pipe microreactors (HPMRs) are compact and transportable nuclear power systems exhibiting inherent safety, well-suited for deployment in remote regions where access is limited and reliance on costly fossil fuels is …

Reinforcement Learning

Techno-economic optimization of a heat-pipe microreactor, part I: theory and cost optimization

2025-12-17 · Paul Seurin, Dean Price, Luis Nunez arxiv

Microreactors, particularly heat-pipe microreactors (HPMRs), are compact, transportable, self-regulated power systems well-suited for access-challenged remote areas where costly fossil fuels dominate. However, they suffe…

Reinforcement LearningGaussian Processes

Offline Reinforcement Learning for Plasma Control in Nuclear Fusion: Codebase and Benchmark

2026-05-19 · Yang Fu, Haomin Bao, Rohit Sonker, Xiaoyan Hu 외 arxiv

Offline reinforcement learning (RL) offers a promising route for developing plasma controllers from historical tokamak data, since online trial-and-error on real devices is costly and risky. However, progress in this dir…

Reinforcement LearningOffline RL