paper-with-me

Papers

Cooperative Bayesian Optimization for Imperfect Agents

2024-03-07 · Ali Khoshvishkaie, Petrus Mikkola, Pierre-Alexandre Murena, Samuel Kaski

We introduce a cooperative Bayesian optimization problem for optimizing black-box functions of two variables where two agents choose together at which points to query the function but have only control over one variable each. This setting is inspired by human-AI teamwork, where an AI-assistant helps its human user solve a problem, in this simplest case, collaborative optimization. We formulate the solution as sequential decision-making, where the agent we control models the user as a computationally rational agent with prior knowledge about the function. We show that strategic planning of the queries enables better identification of the global maximum of the function as long as the user avoids excessive exploration. This planning is made possible by using Bayes Adaptive Monte Carlo planning and by endowing the agent with a user model that accounts for conservative belief updates and exploratory sampling of the points to query.

📄 PDF Abstract BibTeX arXiv:2403.04442

Code (0)

등록된 구현이 없습니다.

Tasks

Bayesian OptimizationDecision MakingSequential Decision Making

Similar Papers 제목 키워드 기반

Learning to Cooperate and Communicate Over Imperfect Channels

2023-11-24 · Jannis Weil, Gizem Ekinci, Heinz Koeppl, Tobias Meuser

Information exchange in multi-agent systems improves the cooperation among agents, especially in partially observable settings. In the real world, communication is often carried out over imperfect channels. This requires…

Q-Learning

Sequential Cooperative Bayesian Inference

2020-02-13 · ICML 2020 1 · Junqi Wang, Pei Wang, Patrick Shafto

Cooperation is often implicitly assumed when learning from other agents. Cooperation implies that the agent selecting the data, and the agent learning from the data, have the same goal, that the learner infer the intende…

Bayesian Inference

STBC-Aided Cooperative NOMA with Timing Offsets, Imperfect Successive Interference Cancellation, and Imperfect Channel State Information

2020-07-07 · Muhammad Waseem Akhtar, Syed Ali Hassan, Sajid Saleem, Haejoon Jung

The combination of non-orthogonal multiple access(NOMA) and cooperative communications can be a suitable solution for fifth generation (5G) and beyond 5G (B5G) wireless systems with massive connectivity, because it can p…

Fairness

Cooperative Bayesian and variance networks disentangle aleatoric and epistemic uncertainties

2025-05-05 · Jiaxiang Yi, Miguel A. Bessa

Real-world data contains aleatoric uncertainty - irreducible noise arising from imperfect measurements or from incomplete knowledge about the data generation process. Mean variance estimation (MVE) networks can learn thi…

Bayesian Inference

ORION: Option-Regularized Deep Reinforcement Learning for Cooperative Multi-Agent Online Navigation

2026-01-03 · Shizhe Zhang, Jingsong Liang, Zhitao Zhou, Shuhan Ye 외 arxiv

Existing methods for multi-agent navigation typically assume fully known environments, offering limited support for partially known scenarios with outdated or imperfect prior maps, such as warehouses or factory floors. T…

Reinforcement Learning