paper-with-me

홈 › Papers

Learning State Representations from Random Deep Action-conditional Predictions

2021-02-09 · NeurIPS 2021 12 · Zeyu Zheng, Vivek Veeriah, Risto Vuorio, Richard Lewis, Satinder Singh

Our main contribution in this work is an empirical finding that random General Value Functions (GVFs), i.e., deep action-conditional predictions -- random both in what feature of observations they predict as well as in the sequence of actions the predictions are conditioned upon -- form good auxiliary tasks for reinforcement learning (RL) problems. In particular, we show that random deep action-conditional predictions when used as auxiliary tasks yield state representations that produce control performance competitive with state-of-the-art hand-crafted auxiliary tasks like value prediction, pixel control, and CURL in both Atari and DeepMind Lab tasks. In another set of experiments we stop the gradients from the RL part of the network to the state representation learning part of the network and show, perhaps surprisingly, that the auxiliary tasks alone are sufficient to learn state representations good enough to outperform an end-to-end trained actor-critic baseline. We opensourced our code at https://github.com/Hwhitetooth/random_gvfs.

📄 PDF Abstract BibTeX arXiv:2102.04897

Code (1)

Hwhitetooth/random_gvfs 공식 구현 jax

Tasks

Atari GamesReinforcement Learning (RL)Representation LearningValue prediction

Similar Papers 제목 키워드 기반

Pairwise Conditional Random Forests for Facial Expression Recognition

2015-12-01 · ICCV 2015 12 · Arnaud Dapogny, Kevin Bailly, Severine Dubuisson

Facial expression can be seen as the dynamic variation of one's appearance over time. Successful recognition thus involves finding representations of high-dimensional spatiotemporal patterns that can be generalized to un…

Facial Expression RecognitionFacial Expression Recognition (FER)

Temporal-Difference Networks

2015-04-21 · NeurIPS 2004 · Richard S. Sutton, Brian Tanner

We introduce a generalization of temporal-difference (TD) learning to networks of interrelated predictions. Rather than relating a single prediction to itself at a later time, as in conventional TD methods, a TD network …

World Knowledge

Binary Expansion Group Intersection Network

2026-03-25 · Sicheng Zhou, Kai Zhang arxiv

Conditional independence is central to modern statistics, but beyond special parametric families it rarely admits an exact covariance characterization. We introduce the binary expansion group intersection network (BEGIN)…

Conditional Adversarial Domain Adaptation

2017-05-26 · NeurIPS 2018 12 · Mingsheng Long, Zhangjie Cao, Jian-Min Wang, Michael. I. Jordan

Adversarial learning has been embedded into deep networks to learn disentangled and transferable representations for domain adaptation. Existing adversarial domain adaptation methods may not effectively align different d…

Domain AdaptationGeneral Classification

Dynamic Pose-Robust Facial Expression Recognition by Multi-View Pairwise Conditional Random Forests

2016-07-21 · Arnaud Dapogny, Kévin Bailly, Séverine Dubuisson

Automatic facial expression classification (FER) from videos is a critical problem for the development of intelligent human-computer interaction systems. Still, it is a challenging problem that involves capturing high-di…

Facial Expression RecognitionFacial Expression Recognition (FER)Head Pose EstimationPose Estimation