paper-with-me

홈 › Papers

Opinion-Guided Reinforcement Learning

2024-05-27 · Kyanna Dagenais, Istvan David

Human guidance is often desired in reinforcement learning to improve the performance of the learning agent. However, human insights are often mere opinions and educated guesses rather than well-formulated arguments. While opinions are subject to uncertainty, e.g., due to partial informedness or ignorance about a problem, they also emerge earlier than hard evidence can be produced. Thus, guiding reinforcement learning agents by way of opinions offers the potential for more performant learning processes, but comes with the challenge of modeling and managing opinions in a formal way. In this article, we present a method to guide reinforcement learning agents through opinions. To this end, we provide an end-to-end method to model and manage advisors' opinions. To assess the utility of the approach, we evaluate it with synthetic (oracle) and human advisors, at different levels of uncertainty, and under multiple advice strategies. Our results indicate that opinions, even if uncertain, improve the performance of reinforcement learning agents, resulting in higher rewards, more efficient exploration, and a better reinforced policy. Although we demonstrate our approach through a two-dimensional topological running example, our approach is applicable to complex problems with higher dimensions as well.

📄 PDF Abstract BibTeX arXiv:2405.17287

Code (0)

등록된 구현이 없습니다.

Tasks

Efficient Explorationreinforcement-learningReinforcement Learning

Similar Papers 제목 키워드 기반

Facilitating on-line opinion dynamics by mining expressions of causation. The case of climate change debates on The Guardian

2019-12-03 · Tom Willaert, Sven Banisch, Paul Van Eecke, Katrien Beuls

News website comment sections are spaces where potentially conflicting opinions and beliefs are voiced. Addressing questions of how to study such cultural and societal conflicts through technological means, the present a…

Articles

Opinion shaping in social networks using reinforcement learning

2019-10-19 · Vivek Borkar, Alexandre Reiffers-Masson

In this paper, we study how to shape opinions in social networks when the matrix of interactions is unknown. We consider classical opinion dynamics with some stubborn agents and the possibility of continuously influencin…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Coevolution of Opinion Dynamics and Recommendation System: Modeling Analysis and Reinforcement Learning Based Manipulation

2024-11-18 · Yuhong Chen, Xiaobing Dai, Martin Buss, Fangzhou Liu

In this work, we develop an analytical framework that integrates opinion dynamics with a recommendation system. By incorporating elements such as collaborative filtering, we provide a precise characterization of how reco…

Collaborative FilteringRecommendation Systems

How social reinforcement learning can lead to metastable polarisation and the voter model

2024-06-12 · Benedikt V. Meylahn, Janusz M. Meylahn

Previous explanations for the persistence of polarization of opinions have typically included modelling assumptions that predispose the possibility of polarization (i.e., assumptions allowing a pair of agents to drift ap…

reinforcement-learningReinforcement Learning

Aspect-Sentiment-Multiple-Opinion Triplet Extraction

2021-10-14 · Fang Wang, Yuncong Li, Sheng-hua Zhong, Cunxiang Yin 외

Aspect Sentiment Triplet Extraction (ASTE) aims to extract aspect term (aspect), sentiment and opinion term (opinion) triplets from sentences and can tell a complete story, i.e., the discussed aspect, the sentiment towar…

Aspect Sentiment Triplet ExtractionExtract AspectSentenceSentiment Analysis+2