paper-with-me

홈 › Papers

Impact-driven Exploration with Contrastive Unsupervised Representations

2021-01-01 · Min Jae Song, Dan Kushnir

Procedurally-generated sparse reward environments pose significant challenges for many RL algorithms. The recently proposed impact-driven exploration method (RIDE) by Raileanu & Rocktäschel (2020), which rewards actions that lead to large changes (measured by $\ell_2$-distance) in the observation embedding, achieves state-of-the-art performance on such procedurally-generated MiniGrid tasks. Yet, the definition of "impact" in RIDE is not conceptually clear because its learned embedding space is not inherently equipped with any similarity measure, let alone $\ell_2$-distance. We resolve this issue in RIDE via contrastive learning. That is, we train the embedding with respect to cosine similarity, where we define two observations to be similar if the agent can reach one observation from the other within a few steps, and define impact in terms of this similarity measure. Experimental results show that our method performs similarly to RIDE on the MiniGrid benchmarks while learning a conceptually clear embedding space equipped with the cosine similarity measure. Our modification of RIDE also provides a new perspective which connects RIDE and episodic curiosity (Savinov et al., 2019), a different exploration method which rewards the agent for visiting states that are unfamiliar to the agent's episodic memory. By incorporating episodic memory into our method, we outperform RIDE on the MiniGrid benchmarks.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Contrastive Learning

Similar Papers 제목 키워드 기반

OPEn: An Open-ended Physics Environment for Learning Without a Task

2021-10-13 · Chuang Gan, Abhishek Bhandwaldar, Antonio Torralba, Joshua B. Tenenbaum 외

Humans have mental models that allow them to plan, experiment, and reason in the physical world. How should an intelligent agent go about learning such models? In this paper, we will study if models of the world learned …

Contrastive LearningRepresentation Learning

Unsupervised Contrastive Learning of Sound Event Representations

2020-11-15 · Eduardo Fonseca, Diego Ortego, Kevin McGuinness, Noel E. O'Connor 외

Self-supervised representation learning can mitigate the limitations in recognition tasks with few manually labeled data but abundant unlabeled data---a common scenario in sound event research. In this work, we explore u…

Contrastive LearningLinear evaluationRepresentation Learning

Demystifying the Mechanisms Behind Emergent Exploration in Goal-conditioned RL

2025-10-15 · Mahsa Bastankhah, Grace Liu, Dilip Arumugam, Thomas L. Griffiths 외 arxiv

In this work, we take a first step toward elucidating the mechanisms behind emergent exploration in unsupervised reinforcement learning. We study Single-Goal Contrastive Reinforcement Learning (SGCRL), a self-supervised …

Reinforcement Learning

UNICON: Unsupervised Intent Discovery via Semantic-level Contrastive Learning

2021-11-16 · ACL ARR November 2021 11 · Anonymous

Discovering new intents is crucial for expanding domains in dialogue systems or natural language understanding (NLU) systems. A typical approach is to leverage unsupervised and semi-supervised learning to train a neural …

ClusteringContrastive LearningData AugmentationIntent Discovery+2

Prototype and Instance Contrastive Learning for Unsupervised Domain Adaptation in Speaker Verification

2024-10-22 · Wen Huang, Bing Han, Zhengyang Chen, Shuai Wang 외

Speaker verification system trained on one domain usually suffers performance degradation when applied to another domain. To address this challenge, researchers commonly use feature distribution matching-based methods in…

Contrastive LearningDomain AdaptationSpeaker VerificationUnsupervised Domain Adaptation