paper-with-me

Papers

Information Content Exploration

2023-10-10 · Jacob Chmura, Hasham Burhani, Xiao Qi Shi

Sparse reward environments are known to be challenging for reinforcement learning agents. In such environments, efficient and scalable exploration is crucial. Exploration is a means by which an agent gains information about the environment. We expand on this topic and propose a new intrinsic reward that systemically quantifies exploratory behavior and promotes state coverage by maximizing the information content of a trajectory taken by an agent. We compare our method to alternative exploration based intrinsic reward techniques, namely Curiosity Driven Learning and Random Network Distillation. We show that our information theoretic reward induces efficient exploration and outperforms in various games, including Montezuma Revenge, a known difficult task for reinforcement learning. Finally, we propose an extension that maximizes information content in a discretely compressed latent space which boosts sample efficiency and generalizes to continuous state spaces.

📄 PDF Abstract BibTeX arXiv:2310.06777

Code (0)

등록된 구현이 없습니다.

Tasks

Efficient Explorationreinforcement-learningReinforcement Learning

Similar Papers 제목 키워드 기반

Breaking Information Cocoons: A Hyperbolic Graph-LLM Framework for Exploration and Exploitation in Recommender Systems

2024-11-21 · Qiyao Ma, Menglin Yang, Mingxuan Ju, Tong Zhao 외

Modern recommender systems often create information cocoons, restricting users' exposure to diverse content. A key challenge lies in balancing content exploration and exploitation while allowing users to adjust their rec…

Recommendation SystemsRepresentation Learning

Stylistic Dialogue Generation via Information-Guided Reinforcement Learning Strategy

2020-04-05 · Yixuan Su, Deng Cai, Yan Wang, Simon Baker 외

Stylistic response generation is crucial for building an engaging dialogue system for industrial use. While it has attracted much research interest, existing methods often generate stylistic responses at the cost of the …

Dialogue Generationreinforcement-learningReinforcement LearningReinforcement Learning (RL)+1

Watch Less and Uncover More: Could Navigation Tools Help Users Search and Explore Videos?

2022-01-10 · Maria Perez-Ortiz, Sahan Bulathwela, Claire Dormann, Meghana Verma 외

Prior research has shown how 'content preview tools' improve speed and accuracy of user relevance judgements across different information retrieval tasks. This paper describes a novel user interface tool, the Content Flo…

Information RetrievalRetrievalTime SeriesTime Series Analysis+1

Long-Term Value of Exploration: Measurements, Findings and Algorithms

2023-05-12 · Yi Su, Xiangyu Wang, Elaine Ya Le, Liang Liu 외

Effective exploration is believed to positively influence the long-term user experience on recommendation platforms. Determining its exact benefits, however, has been challenging. Regular A/B tests on exploration often m…

Recommendation Systems

Unveiling User Satisfaction and Creator Productivity Trade-Offs in Recommendation Platforms

2024-10-31 · Fan Yao, Yiming Liao, Jingzhou Liu, Shaoliang Nie 외

On User-Generated Content (UGC) platforms, recommendation algorithms significantly impact creators' motivation to produce content as they compete for algorithmically allocated user traffic. This phenomenon subtly shapes …

Diversity