Enhancing reinforcement learning for population setpoint tracking in co-cultures
Efficient multiple setpoint tracking can enable advanced biotechnological applications, such as maintaining desired population levels in co-cultures for optimal metabolic division of labor. In this study, we employ reinforcement learning as a control method for population setpoint tracking in co-cultures, focusing on policy-gradient techniques where the control policy is parameterized by neural networks. However, achieving accurate tracking across multiple setpoints is a significant challenge in reinforcement learning, as the agent must effectively balance the contributions of various setpoints to maximize the expected system performance. Traditional return functions, such as those based on a quadratic cost, often yield suboptimal performance due to their inability to efficiently guide the agent toward the simultaneous satisfaction of all setpoints. To overcome this, we propose a novel return function that rewards the simultaneous satisfaction of multiple setpoints and diminishes overall reward gains otherwise, accounting for both stage and terminal system performance. This return function includes parameters to fine-tune the desired smoothness and steepness of the learning process. We demonstrate our approach considering an $\textit{Escherichia coli}$ co-culture in a chemostat with optogenetic control over amino acid synthesis pathways, leveraging auxotrophies to modulate growth.
Code (0)
등록된 구현이 없습니다.
Tasks
reinforcement-learningReinforcement LearningSimilar Papers 제목 키워드 기반
Reinforcement learning for efficient and robust multi-setpoint and multi-trajectory tracking in bioprocesses
Efficient and robust bioprocess control is essential for maximizing performance and adaptability in advanced biotechnological systems. In this work, we present a reinforcement-learning framework for multi-setpoint and mu…
reinforcement-learningReinforcement LearningReach and hold flexibility characterization and trade-off analysis for aggregations of thermostatically controlled loads
Thermostatically controlled loads (TCLs) have the potential to be flexible and responsive loads to be used in demand response (DR) schemes. With increasing renewable penetration, DR is playing an increasingly important r…
Data-Driven Tracking MPC for Changing Setpoints
We propose a data-driven tracking model predictive control (MPC) scheme to control unknown discrete-time linear time-invariant systems. The scheme uses a purely data-driven system parametrization to predict future trajec…
Model Predictive ControlControl-Informed Reinforcement Learning for Chemical Processes
This work proposes a control-informed reinforcement learning (CIRL) framework that integrates proportional-integral-derivative (PID) control components into the architecture of deep reinforcement learning (RL) policies. …
Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)A Unified Power-Setpoint Tracking Algorithm for Utility-Scale PV Systems with Power Reserves and Fast Frequency Response Capabilities
This paper presents a fast power-setpoint tracking algorithm to enable utility-scale photovoltaic (PV) systems to provide high quality grid services such as power reserves and fast frequency response. The algorithm unite…
Point Tracking