paper-with-me

홈 › Papers

Reinforcement Learning with Function-Valued Action Spaces for Partial Differential Equation Control

2018-06-13 · ICML 2018 7 · Yangchen Pan, Amir-Massoud Farahmand, Martha White, Saleh Nabi, Piyush Grover, Daniel Nikovski

Recent work has shown that reinforcement learning (RL) is a promising approach to control dynamical systems described by partial differential equations (PDE). This paper shows how to use RL to tackle more general PDE control problems that have continuous high-dimensional action spaces with spatial relationship among action dimensions. In particular, we propose the concept of action descriptors, which encode regularities among spatially-extended action dimensions and enable the agent to control high-dimensional action PDEs. We provide theoretical evidence suggesting that this approach can be more sample efficient compared to a conventional approach that treats each action dimension separately and does not explicitly exploit the spatial regularity of the action space. The action descriptor approach is then used within the deep deterministic policy gradient algorithm. Experiments on two PDE control problems, with up to 256-dimensional continuous actions, show the advantage of the proposed approach over the conventional one.

📄 PDF Abstract BibTeX arXiv:1806.06931

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Rational kernel-based interpolation for complex-valued frequency response functions

2023-07-25 · Julien Bect, Niklas Georg, Ulrich Römer, Sebastian Schöps

This work is concerned with the kernel-based approximation of a complex-valued function from data, where the frequency response function of a partial differential equation in the frequency domain is of particular interes…

Model Selection

Learning Functional Transduction

2023-02-01 · NeurIPS 2023 11

Research in machine learning has polarized into two general approaches for regression tasks: Transductive methods construct estimates directly from available data but are usually problem unspecific. Inductive methods can…

regression

Conditioning of Banach Space Valued Gaussian Random Variables: An Approximation Approach Based on Martingales

2024-04-04 · Ingo Steinwart

We investigate the conditional distributions of two Banach space valued, jointly Gaussian random variables. In particular, we show that these conditional distributions are again Gaussian and that their means and covarian…

Gaussian Processes

Value Functions as Supermartingale Certificates

2026-05-29 · Alessandro Abate, Daniel Contro, Mirco Giacobbe, Agustín Martínez-Suñé 외 arxiv

Certification methods for stochastic systems provide sufficient proof rules, based on real-valued supermartingale certificates, to determine the almost-sure satisfaction of $ω$-regular properties (and therefore of linear…

Reinforcement Learning

REValueD: Regularised Ensemble Value-Decomposition for Factorisable Markov Decision Processes

2024-01-16 · David Ireland, Giovanni Montana

Discrete-action reinforcement learning algorithms often falter in tasks with high-dimensional discrete action spaces due to the vast number of possible actions. A recent advancement leverages value-decomposition, a conce…

Multi-agent Reinforcement LearningQ-Learningreinforcement-learningReinforcement Learning