paper-with-me

홈 › Papers

A General Framework for Interacting Bayes-Optimally with Self-Interested Agents using Arbitrary Parametric Model and Model Prior

2013-04-07 · Trong Nghia Hoang, Kian Hsiang Low

Recent advances in Bayesian reinforcement learning (BRL) have shown that Bayes-optimality is theoretically achievable by modeling the environment's latent dynamics using Flat-Dirichlet-Multinomial (FDM) prior. In self-interested multi-agent environments, the transition dynamics are mainly controlled by the other agent's stochastic behavior for which FDM's independence and modeling assumptions do not hold. As a result, FDM does not allow the other agent's behavior to be generalized across different states nor specified using prior domain knowledge. To overcome these practical limitations of FDM, we propose a generalization of BRL to integrate the general class of parametric models and model priors, thus allowing practitioners' domain knowledge to be exploited to produce a fine-grained and compact representation of the other agent's behavior. Empirical evaluation shows that our approach outperforms existing multi-agent reinforcement learning algorithms.

📄 PDF Abstract BibTeX arXiv:1304.2024

Code (0)

등록된 구현이 없습니다.

Tasks

modelMulti-agent Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Bayesian mechanics of self-organising systems

2023-11-16 · Takuya Isomura

Bayesian mechanics provides a framework that addresses dynamical systems that can be conceptualised as Bayesian inference. However, elucidating the requisite generative models is essential for empirical applications to r…

Bayesian Inference

Modified Meta-Thompson Sampling for Linear Bandits and Its Bayes Regret Analysis

2024-09-10 · Hao Li, Dong Liang, Zheng Xie

Meta-learning is characterized by its ability to learn how to learn, enabling the adaptation of learning strategies across different tasks. Recent research introduced the Meta-Thompson Sampling (Meta-TS), which meta-lear…

Meta-LearningMulti-Armed BanditsThompson Sampling

Neural Bayes: A Generic Parameterization Method for Unsupervised Representation Learning

2020-02-20 · Devansh Arpit, Huan Wang, Caiming Xiong, Richard Socher 외

We introduce a parameterization method called Neural Bayes which allows computing statistical quantities that are in general difficult to compute and opens avenues for formulating new objectives for unsupervised represen…

ClusteringFormRepresentation Learning

Neural Bayes: A Generic Parameterization Method for Unsupervised Learning

2021-01-01 · Devansh Arpit, Huan Wang, Caiming Xiong, Richard Socher 외

We introduce a parameterization method called Neural Bayes which allows computing statistical quantities that are in general difficult to compute and opens avenues for formulating new objectives for unsupervised represen…

ClusteringFormRepresentation Learning

Interacting Large Language Model Agents. Interpretable Models and Social Learning

2024-11-02 · Adit Jain, Vikram Krishnamurthy

This paper discusses the theory and algorithms for interacting large language model agents (LLMAs) using methods from statistical signal processing and microeconomics. While both fields are mature, their application to d…

Bayesian InferenceLanguage ModelingLanguage ModellingLarge Language Model+2