paper-with-me

Papers

Dr. Strategy: Model-Based Generalist Agents with Strategic Dreaming

2024-02-29 · Hany Hamed, Subin Kim, Dongyeong Kim, Jaesik Yoon, Sungjin Ahn

Model-based reinforcement learning (MBRL) has been a primary approach to ameliorating the sample efficiency issue as well as to make a generalist agent. However, there has not been much effort toward enhancing the strategy of dreaming itself. Therefore, it is a question whether and how an agent can "dream better" in a more structured and strategic way. In this paper, inspired by the observation from cognitive science suggesting that humans use a spatial divide-and-conquer strategy in planning, we propose a new MBRL agent, called Dr. Strategy, which is equipped with a novel Dreaming Strategy. The proposed agent realizes a version of divide-and-conquer-like strategy in dreaming. This is achieved by learning a set of latent landmarks and then utilizing these to learn a landmark-conditioned highway policy. With the highway policy, the agent can first learn in the dream to move to a landmark, and from there it tackles the exploration and achievement task in a more focused way. In experiments, we show that the proposed model outperforms prior pixel-based MBRL methods in various visually complex and partially observable navigation tasks.

📄 PDF Abstract BibTeX arXiv:2402.18866

Code (0)

등록된 구현이 없습니다.

Tasks

Model-based Reinforcement Learning

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

AdaDemo: Data-Efficient Demonstration Expansion for Generalist Robotic Agent

2024-04-11 · Tongzhou Mu, Yijie Guo, Jie Xu, Ankit Goyal 외

Encouraged by the remarkable achievements of language and vision foundation models, developing generalist robotic agents through imitation learning, using large demonstration datasets, has become a prominent area of inte…

Imitation Learning

Dreaming in Code for Curriculum Learning in Open-Ended Worlds

2026-02-09 · Konstantinos Mitsides, Maxence Faldor, Antoine Cully arxiv

Open-ended learning frames intelligence as emerging from continual interaction with an ever-expanding space of environments. While recent advances have utilized foundation models to programmatically generate diverse envi…

Strategic Analysis of Fair Rank-Minimizing Mechanisms with Agent Refusal Option

2024-08-03 · Yasunori Okumura

This study examines strategic issues in fair rank-minimizing mechanisms, which choose an assignment that minimizes the average rank of object types to which agents are assigned and satisfy a fairness property called equa…

Fairness

AgentStore: Scalable Integration of Heterogeneous Agents As Specialized Generalist Computer Assistant

2024-10-24 · Chengyou Jia, Minnan Luo, Zhuohang Dang, Qiushi Sun 외

Digital agents capable of automating complex computer tasks have attracted considerable attention due to their immense potential to enhance human-computer interaction. However, existing agent methods exhibit deficiencies…

Reliable Intersection Control in Non-cooperative Environments

2018-02-22 · Muhammed O. Sayin, Chung-Wei Lin, Shinichi Shiraishi, Tamer Başar

We propose a reliable intersection control mechanism for strategic autonomous and connected vehicles (agents) in non-cooperative environments. Each agent has access to his/her earliest possible and desired passing times,…