Computationally Efficient Methods for Solving Discrete-time Dynamic models with Continuous Actions
This study investigates computationally efficient algorithms for solving discrete-time infinite-horizon single-agent/multi-agent dynamic models with continuous actions. It shows that we can easily reduce the computational costs by slightly changing basic algorithms using value functions, such as the Value Function Iteration (VFI) and the Policy Iteration (PI). The PI method with a Krylov iterative method (GMRES), which can be easily implemented using built-in packages, works much better than VFI-based algorithms even when considering continuous state models. Concerning the VFI algorithm, we can largely speed up the convergence by introducing acceleration methods of fixed-point iterations. The current study also proposes the VF-PGI-Spectral (Value Function-Policy Gradient Iteration Spectral) algorithm, which is a slight modification of the VFI. It shows numerical results where the VF-PGI-Spectral performs much better than the VFI- and PI-based algorithms especially in multi-agent dynamic games. Finally, it shows that using relative value functions further reduces the computational cost of these methods.
Code (0)
등록된 구현이 없습니다.
Methods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Learning to Stop: Deep Learning for Mean Field Optimal Stopping
Optimal stopping is a fundamental problem in optimization with applications in risk management, finance, robotics, and machine learning. We extend the standard framework to a multi-agent setting, named multi-agent optima…
Deep LearningModel-Adaptive Approach to Dynamic Discrete Choice Models with Large State Spaces
Estimation and counterfactual experiments in dynamic discrete choice models with large state spaces pose computational difficulties. This paper develops a novel model-adaptive approach to solve the linear system of fixed…
Computational EfficiencycounterfactualDiscrete Choice ModelsLearning Compact Physics-Aware Delayed Photocurrent Models Using Dynamic Mode Decomposition
Radiation-induced photocurrent in semiconductor devices can be simulated using complex physics-based models, which are accurate, but computationally expensive. This presents a challenge for implementing device characteri…
Time SeriesTime Series AnalysisControllability and Tracking of Ensembles: An Optimal Transport Theory Viewpoint
This paper explores the controllability and state tracking of ensembles from the perspective of optimal transport theory. Ensembles, characterized as collections of systems evolving under the same dynamics but with varyi…
Predictive and Proactive Power Allocation For Energy Efficiency in Dynamic OWC Networks
Driven by the exponential growth in data traffic and the limitations of Radio Frequency (RF) networks, Optical Wireless Communication (OWC) has emerged as a promising solution for high data rate communication. However, t…