paper-with-me

홈 › Papers

Sample-Efficient Multi-Agent RL: An Optimization Perspective

2023-10-10 · Nuoya Xiong, Zhihan Liu, Zhaoran Wang, Zhuoran Yang

We study multi-agent reinforcement learning (MARL) for the general-sum Markov Games (MGs) under the general function approximation. In order to find the minimum assumption for sample-efficient learning, we introduce a novel complexity measure called the Multi-Agent Decoupling Coefficient (MADC) for general-sum MGs. Using this measure, we propose the first unified algorithmic framework that ensures sample efficiency in learning Nash Equilibrium, Coarse Correlated Equilibrium, and Correlated Equilibrium for both model-based and model-free MARL problems with low MADC. We also show that our algorithm provides comparable sublinear regret to the existing works. Moreover, our algorithm combines an equilibrium-solving oracle with a single objective optimization subprocedure that solves for the regularized payoff of each deterministic joint policy, which avoids solving constrained optimization problems within data-dependent constraints (Jin et al. 2020; Wang et al. 2023) or executing sampling procedures with complex multi-objective optimization problems (Foster et al. 2023), thus being more amenable to empirical implementation.

📄 PDF Abstract BibTeX arXiv:2310.06243

Code (0)

등록된 구현이 없습니다.

Tasks

Multi-agent Reinforcement Learning

Similar Papers 제목 키워드 기반

Order Matters: Agent-by-agent Policy Optimization

2023-02-13 · Xihuai Wang, Zheng Tian, Ziyu Wan, Ying Wen 외

While multi-agent trust region algorithms have achieved great success empirically in solving coordination tasks, most of them, however, suffer from a non-stationarity problem since agents update their policies simultaneo…

MuJoCo

CAF-I: A Collaborative Multi-Agent Framework for Enhanced Irony Detection with Large Language Models

2025-06-10 · Ziqi. Liu, Ziyang. Zhou, Mingxuan. Hu

Large language model (LLM) have become mainstream methods in the field of sarcasm detection. However, existing LLM methods face challenges in irony detection, including: 1. single-perspective limitations, 2. insufficient…

Language ModelingLanguage ModellingLarge Language ModelSarcasm Detection

Model-Based Decentralized Policy Optimization

2023-02-16 · Hao Luo, Jiechuan Jiang, Zongqing Lu

Decentralized policy optimization has been commonly used in cooperative multi-agent tasks. However, since all agents are updating their policies simultaneously, from the perspective of individual agents, the environment …

model

A brief note on learning problem with global perspectives

2026-01-09 · Getachew K. Befekadu arxiv

This brief note considers the problem of learning with dynamic-optimizing principal-agent setting, in which the agents are allowed to have global perspectives about the learning process, i.e., the ability to view things …

ProCrit: Self-Elicited Multi-Perspective Reasoning with Critic-Guided Revision for Multimodal Sarcasm Detection

2026-05-20 · Yingjia Xu, Jiulong Wu, Bowen Zhang, Baokui Guo 외 arxiv

Multimodal sarcasm detection requires reasoning over cross-modal incongruities between literal expression and intended meaning, yet the specific analytical perspectives needed vary across samples due to the diversity of …

Reinforcement LearningSarcasm Detection