paper-with-me

홈 › Papers

Work in Progress: Temporally Extended Auxiliary Tasks

2020-04-01 · Craig Sherstan, Bilal Kartal, Pablo Hernandez-Leal, Matthew E. Taylor

Predictive auxiliary tasks have been shown to improve performance in numerous reinforcement learning works, however, this effect is still not well understood. The primary purpose of the work presented here is to investigate the impact that an auxiliary task's prediction timescale has on the agent's policy performance. We consider auxiliary tasks which learn to make on-policy predictions using temporal difference learning. We test the impact of prediction timescale using a specific form of auxiliary task in which the input image is used as the prediction target, which we refer to as temporal difference autoencoders (TD-AE). We empirically evaluate the effect of TD-AE on the A2C algorithm in the VizDoom environment using different prediction timescales. While we do not observe a clear relationship between the prediction timescale on performance, we make the following observations: 1) using auxiliary tasks allows us to reduce the trajectory length of the A2C algorithm, 2) in some cases temporally extended TD-AE performs better than a straight autoencoder, 3) performance with auxiliary tasks is sensitive to the weight placed on the auxiliary loss, 4) despite this sensitivity, auxiliary tasks improved performance without extensive hyper-parameter tuning. Our overall conclusions are that TD-AE increases the robustness of the A2C algorithm to the trajectory length and while promising, further study is required to fully understand the relationship between auxiliary task prediction timescale and the agent's performance.

📄 PDF Abstract BibTeX arXiv:2004.00600

Code (0)

등록된 구현이 없습니다.

Tasks

PredictionReinforcement Learning

Methods 이 논문이 사용한 방법론

A2C A2C, or Advantage Actor Critic, is a synchronous version of the A3C policy gradient method. As an alternative to the asynchronous…

Similar Papers 제목 키워드 기반

Temporally Extended Successor Representations

2022-09-25 · Matthew J. Sargent, Peter J. Bentley, Caswell Barry, William de Cothi

We present a temporally extended variation of the successor representation, which we term t-SR. t-SR captures the expected state transition dynamics of temporally extended actions by constructing successor representation…

Diffusion Meets Options: Hierarchical Generative Skill Composition for Temporally-Extended Tasks

2024-10-03 · Zeyu Feng, Hao Luan, Kevin Yuchen Ma, Harold Soh

Safe and successful deployment of robots requires not only the ability to generate complex plans but also the capacity to frequently replan and correct execution errors. This paper addresses the challenge of long-horizon…

DiversityHierarchical Reinforcement LearningRobot NavigationTrajectory Planning

HRL4IN: Hierarchical Reinforcement Learning for Interactive Navigation with Mobile Manipulators

2019-10-24 · Chengshu Li, Fei Xia, Roberto Martin-Martin, Silvio Savarese

Most common navigation tasks in human environments require auxiliary arm interactions, e.g. opening doors, pressing buttons and pushing obstacles away. This type of navigation tasks, which we call Interactive Navigation,…

Hierarchical Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Learning to Follow Instructions in Text-Based Games

2022-11-08 · Mathieu Tuli, Andrew C. Li, Pashootan Vaezipoor, Toryn Q. Klassen 외

Text-based games present a unique class of sequential decision making problem in which agents interact with a partially observable, simulated environment via actions and observations conveyed through natural language. Su…

Decision MakingInstruction FollowingReinforcement Learning (RL)Sequential Decision Making+1

Temporally Correlated Task Scheduling for Sequence Learning

2020-07-10 · Xueqing Wu, Lewen Wang, Yingce Xia, Weiqing Liu 외

Sequence learning has attracted much research attention from the machine learning community in recent years. In many applications, a sequence learning task is usually associated with multiple temporally correlated auxili…

Machine TranslationSchedulingTranslation