paper-with-me

홈 › Papers

Multi-Scenario Combination Based on Multi-Agent Reinforcement Learning to Optimize the Advertising Recommendation System

2024-07-03 · Yang Zhao, Chang Zhou, Jin Cao, Yi Zhao, Shaobo Liu, Chiyu Cheng, Xingchen Li

This paper explores multi-scenario optimization on large platforms using multi-agent reinforcement learning (MARL). We address this by treating scenarios like search, recommendation, and advertising as a cooperative, partially observable multi-agent decision problem. We introduce the Multi-Agent Recurrent Deterministic Policy Gradient (MARDPG) algorithm, which aligns different scenarios under a shared objective and allows for strategy communication to boost overall performance. Our results show marked improvements in metrics such as click-through rate (CTR), conversion rate, and total sales, confirming our method's efficacy in practical settings.

📄 PDF Abstract BibTeX arXiv:2407.02759

Code (0)

등록된 구현이 없습니다.

Tasks

Multi-agent Reinforcement Learning

Similar Papers 제목 키워드 기반

Episodic Future Thinking Mechanism for Multi-agent Reinforcement Learning

2024-10-22 · Dongsu Lee, Minhae Kwon

Understanding cognitive processes in multi-agent interactions is a primary goal in cognitive science. It can guide the direction of artificial intelligence (AI) research toward social decision-making in multi-agent syste…

Autonomous DrivingMulti-agent Reinforcement Learningreinforcement-learningReinforcement Learning+1

Monotonic Value Function Factorisation for Deep Multi-Agent Reinforcement Learning

2020-03-19 · Tabish Rashid, Mikayel Samvelyan, Christian Schroeder de Witt, Gregory Farquhar 외

In many real-world settings, a team of agents must coordinate its behaviour while acting in a decentralised fashion. At the same time, it is often possible to train the agents in a centralised fashion where global state …

Multi-agent Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)+3

MACC: Cross-Layer Multi-Agent Congestion Control with Deep Reinforcement Learning

2022-06-04 · Jianing Bai, Tianhao Zhang, Guangming Xie

Congestion Control (CC), as the core networking task to efficiently utilize network capacity, received great attention and widely used in various Internet communication applications such as 5G, Internet-of-Things, UAN, a…

Deep Reinforcement LearningManagementMulti-agent Reinforcement Learningreinforcement-learning+2

Universal Policies to Learn Them All

2019-08-24 · Hassam Ullah Sheikh, Ladislau Bölöni

We explore a collaborative and cooperative multi-agent reinforcement learning setting where a team of reinforcement learning agents attempt to solve a single cooperative task in a multi-scenario setting. We propose a nov…

AllMulti-agent Reinforcement Learningreinforcement-learningReinforcement Learning+1

SeeUPO: Sequence-Level Agentic-RL with Convergence Guarantees

2026-02-06 · Tianyi Hu, Qingxu Fu, Yanxi Chen, Zhaoyang Liu 외 arxiv

Reinforcement learning (RL) has emerged as the predominant paradigm for training large language model (LLM)-based AI agents. However, existing backbone RL algorithms lack verified convergence guarantees in agentic scenar…

Reinforcement Learning