paper-with-me

홈 › Papers

Goal-oriented inference of environment from redundant observations

2023-05-08 · Kazuki Takahashi, Tomoki Fukai, Yutaka Sakai, Takashi Takekawa

The agent learns to organize decision behavior to achieve a behavioral goal, such as reward maximization, and reinforcement learning is often used for this optimization. Learning an optimal behavioral strategy is difficult under the uncertainty that events necessary for learning are only partially observable, called as Partially Observable Markov Decision Process (POMDP). However, the real-world environment also gives many events irrelevant to reward delivery and an optimal behavioral strategy. The conventional methods in POMDP, which attempt to infer transition rules among the entire observations, including irrelevant states, are ineffective in such an environment. Supposing Redundantly Observable Markov Decision Process (ROMDP), here we propose a method for goal-oriented reinforcement learning to efficiently learn state transition rules among reward-related "core states'' from redundant observations. Starting with a small number of initial core states, our model gradually adds new core states to the transition diagram until it achieves an optimal behavioral strategy consistent with the Bellman equation. We demonstrate that the resultant inference model outperforms the conventional method for POMDP. We emphasize that our model only containing the core states has high explainability. Furthermore, the proposed method suits online learning as it suppresses memory consumption and improves learning speed.

📄 PDF Abstract BibTeX arXiv:2305.04432

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement Learning

Similar Papers 제목 키워드 기반

Learning Spatial and Temporal Hierarchies: Hierarchical Active Inference for navigation in Multi-Room Maze Environments

2023-09-18 · Daria de Tinguy, Toon Van de Maele, Tim Verbelen, Bart Dhoedt

Cognitive maps play a crucial role in facilitating flexible behaviour by representing spatial and conceptual relationships within an environment. The ability to learn and infer the underlying structure of the environment…

Efficient Exploration

Inferring Hierarchical Structure in Multi-Room Maze Environments

2023-06-23 · Daria de Tinguy, Toon Van de Maele, Tim Verbelen, Bart Dhoedt

Cognitive maps play a crucial role in facilitating flexible behaviour by representing spatial and conceptual relationships within an environment. The ability to learn and infer the underlying structure of the environment…

Efficient Exploration

Common Language for Goal-Oriented Semantic Communications: A Curriculum Learning Framework

2021-11-15 · Mohammad Karimzadeh Farshbafan, Walid Saad, Merouane Debbah

Semantic communications will play a critical role in enabling goal-oriented services over next-generation wireless systems. However, most prior art in this domain is restricted to specific applications (e.g., text or ima…

Reinforcement Learning (RL)

Regioned Episodic Reinforcement Learning

2021-01-01 · Jiarui Jin, Cong Chen, Ming Zhou, Weinan Zhang 외

Goal-oriented reinforcement learning algorithms are often good at exploration, not exploitation, while episodic algorithms excel at exploitation, not exploration. As a result, neither of these approaches alone can lead t…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

6G goal-oriented communications: How to coexist with legacy systems?

2023-08-25 · Mattia Merluzzi, Miltiadis C. Filippou, Leonardo Gomes Baltar, Markus D. Muek 외

6G will connect heterogeneous intelligent agents to make them operate complex cooperative tasks. When connecting intelligence, two main research questions arise to identify how AI and ML models behave depending on: i) th…