paper-with-me

Papers

Sparse Markov Decision Processes with Causal Sparse Tsallis Entropy Regularization for Reinforcement Learning

2017-09-19 · Kyungjae Lee, Sungjoon Choi, Songhwai Oh

In this paper, a sparse Markov decision process (MDP) with novel causal sparse Tsallis entropy regularization is proposed.The proposed policy regularization induces a sparse and multi-modal optimal policy distribution of a sparse MDP. The full mathematical analysis of the proposed sparse MDP is provided.We first analyze the optimality condition of a sparse MDP. Then, we propose a sparse value iteration method which solves a sparse MDP and then prove the convergence and optimality of sparse value iteration using the Banach fixed point theorem. The proposed sparse MDP is compared to soft MDPs which utilize causal entropy regularization. We show that the performance error of a sparse MDP has a constant bound, while the error of a soft MDP increases logarithmically with respect to the number of actions, where this performance error is caused by the introduced regularization term. In experiments, we apply sparse MDPs to reinforcement learning problems. The proposed method outperforms existing methods in terms of the convergence speed and performance.

📄 PDF Abstract BibTeX arXiv:1709.06293

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…
Entropy Regularization 설명 없음

Similar Papers 제목 키워드 기반

R-AIF: Solving Sparse-Reward Robotic Tasks from Pixels with Active Inference and World Models

2024-09-21 · Viet Dung Nguyen, Zhizhuo Yang, Christopher L. Buckley, Alexander Ororbia

Although research has produced promising results demonstrating the utility of active inference (AIF) in Markov decision processes (MDPs), there is relatively less work that builds AIF models in the context of environment…

Learning the Markov Decision Process in the Sparse Gaussian Elimination

2021-09-30 · Yingshi Chen

We propose a learning-based approach for the sparse Gaussian Elimination. There are many hard combinatorial optimization problems in modern sparse solver. These NP-hard problems could be handled in the framework of Marko…

Combinatorial OptimizationQ-LearningScheduling

Ordering-Based Causal Structure Learning in the Presence of Latent Variables

2019-10-20 · Daniel Irving Bernstein, Basil Saeed, Chandler Squires, Caroline Uhler

We consider the task of learning a causal graph in the presence of latent confounders given i.i.d.~samples from the model. While current algorithms for causal structure discovery in the presence of latent confounders are…

Reversible MCMC on Markov equivalence classes of sparse directed acyclic graphs

2012-09-26 · Yangbo He, Jinzhu Jia, Bin Yu

Graphical models are popular statistical tools which are used to represent dependent or causal complex systems. Statistically equivalent causal or directed graphical models are said to belong to a Markov equivalent class…

Relaxed Sparsest-Permutation Formulation for Causal Discovery at Scale

2026-05-07 · Sunmin Oh, Sang-Yun Oh, Gunwoong Park arxiv

Despite the growing availability of large datasets, causal structure learning remains computationally prohibitive at scale. We revisit sparsest-permutation learning for linear structural equation models and show that exa…