Deep reinforcement learning for feedback control in a collective flashing ratchet
A collective flashing ratchet transports Brownian particles using a spatially periodic, asymmetric, and time-dependent on-off switchable potential. The net current of the particles in this system can be substantially increased by feedback control based on the particle positions. Several feedback policies for maximizing the current have been proposed, but optimal policies have not been found for a moderate number of particles. Here, we use deep reinforcement learning (RL) to find optimal policies, with results showing that policies built with a suitable neural network architecture outperform the previous policies. Moreover, even in a time-delayed feedback situation where the on-off switching of the potential is delayed, we demonstrate that the policies provided by deep RL provide higher currents than the previous strategies.
Code (1)
Tasks
Deep Reinforcement Learningreinforcement-learningReinforcement Learning (RL)Methods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Leveraging Environmental Correlations: The Thermodynamics of Requisite Variety
Key to biological success, the requisite variety that confronts an adaptive organism is the set of detectable, accessible, and controllable states in its environment. We analyze its role in the thermodynamic functioning …
Learning entropy production via neural networks
This Letter presents a neural estimator for entropy production, or NEEP, that estimates entropy production (EP) from trajectories of relevant variables without detailed information on the system dynamics. For steady stat…
Optimal ratcheting of dividend payout under Brownian motion surplus
This paper is concerned with a long standing optimal dividend payout problem subject to the so-called ratcheting constraint, that is, the dividend payout rate shall be non-decreasing over time and is thus self-path-depen…
Optimal Tracking Portfolio with A Ratcheting Capital Benchmark
This paper studies the finite horizon portfolio management by optimally tracking a ratcheting capital benchmark process. It is assumed that the fund manager can dynamically inject capital into the portfolio account such …
ManagementFoundation Inference Models for Markov Jump Processes
Markov jump processes are continuous-time stochastic processes which describe dynamical systems evolving in discrete state spaces. These processes find wide application in the natural sciences and machine learning, but t…
Protein Folding