paper-with-me

홈 › Papers

Learning to Share and Hide Intentions using Information Regularization

2018-08-06 · NeurIPS 2018 12 · DJ Strouse, Max Kleiman-Weiner, Josh Tenenbaum, Matt Botvinick, David Schwab

Learning to cooperate with friends and compete with foes is a key component of multi-agent reinforcement learning. Typically to do so, one requires access to either a model of or interaction with the other agent(s). Here we show how to learn effective strategies for cooperation and competition in an asymmetric information game with no such model or interaction. Our approach is to encourage an agent to reveal or hide their intentions using an information-theoretic regularizer. We consider both the mutual information between goal and action given state, as well as the mutual information between goal and state. We show how to optimize these regularizers in a way that is easy to integrate with policy gradient reinforcement learning. Finally, we demonstrate that cooperative (competitive) policies learned with our approach lead to more (less) reward for a second agent in two simple asymmetric information games.

📄 PDF Abstract BibTeX arXiv:1808.02093

Code (1)

djstrouse/InfoMARL 공식 구현

Tasks

Multi-agent Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Uncovering the Unseen: Discover Hidden Intentions by Micro-Behavior Graph Reasoning

2023-08-29 · Zhuo Zhou, Wenxuan Liu, Danni Xu, Zheng Wang 외

This paper introduces a new and challenging Hidden Intention Discovery (HID) task. Unlike existing intention recognition tasks, which are based on obvious visual representations to identify common intentions for normal b…

Intent Detection

What is Proxy Discrimination?

2022-05-11 · Michael Carl Tschantz

The near universal condemnation of proxy discrimination hides a disagreement over what it is. This work surveys various notions of proxy and proxy discrimination found in prior work and represents them in a common framew…

Explicability? Legibility? Predictability? Transparency? Privacy? Security? The Emerging Landscape of Interpretable Agent Behavior

2018-11-23 · Tathagata Chakraborti, Anagha Kulkarni, Sarath Sreedharan, David E. Smith 외

There has been significant interest of late in generating behavior of agents that is interpretable to the human (observer) in the loop. However, the work in this area has typically lacked coherence on the topic, with pro…

TextHide: Tackling Data Privacy in Language Understanding Tasks

2020-10-12 · Findings of the Association for Computational Linguistics 2020 · Yangsibo Huang, Zhao Song, Danqi Chen, Kai Li 외

An unsolved challenge in distributed or federated learning is to effectively mitigate privacy risks without slowing down training or reducing accuracy. In this paper, we propose TextHide aiming at addressing this challen…

Federated LearningNatural Language UnderstandingSentence

Shared intentions and the advance of cumulative culture in hunter-gatherers

2015-03-24

It has been hypothesized that the evolution of modern human cognition was catalyzed by the development of jointly intentional modes of behaviour. From an early age (1-2 years), human infants outperform apes at tasks that…

Cultural Vocal Bursts Intensity Prediction