paper-with-me

Papers

Learning to Bid Long-Term: Multi-Agent Reinforcement Learning with Long-Term and Sparse Reward in Repeated Auction Games

2022-04-05 · Jing Tan, Ramin Khalili, Holger Karl

We propose a multi-agent distributed reinforcement learning algorithm that balances between potentially conflicting short-term reward and sparse, delayed long-term reward, and learns with partial information in a dynamic environment. We compare different long-term rewards to incentivize the algorithm to maximize individual payoff and overall social welfare. We test the algorithm in two simulated auction games, and demonstrate that 1) our algorithm outperforms two benchmark algorithms in a direct competition, with cost to social welfare, and 2) our algorithm's aggressive competitive behavior can be guided with the long-term reward signal to maximize both individual payoff and overall social welfare.

📄 PDF Abstract BibTeX arXiv:2204.02268

Code (1)

dracosource/biddinggame 공식 구현 pytorch

Tasks

Multi-agent Reinforcement Learningreinforcement-learningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Long-Term Fairness in Sequential Multi-Agent Selection with Positive Reinforcement

2024-07-10 · Bhagyashree Puranik, Ozgur Guldogan, Upamanyu Madhow, Ramtin Pedarsani

While much of the rapidly growing literature on fair decision-making focuses on metrics for one-shot decisions, recent work has raised the intriguing possibility of designing sequential decision-making to positively impa…

Decision MakingFairnessSequential Decision Making

Hebbian Synaptic Modifications in Spiking Neurons that Learn

2019-11-17 · Peter L. Bartlett, Jonathan Baxter

In this paper, we derive a new model of synaptic plasticity, based on recent algorithms for reinforcement learning (in which an agent attempts to learn appropriate actions to maximize its long-term average reward). We sh…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Long-Term Mapping of the Douro River Plume with Multi-Agent Reinforcement Learning

2025-10-03 · Nicolò Dal Fabbro, Milad Mesbahi, Renato Mendes, João Borges de Sousa 외 arxiv

We study the problem of long-term (multiple days) mapping of a river plume using multiple autonomous underwater vehicles (AUVs), focusing on the Douro river representative use-case. We propose an energy - and communicati…

Multi-agent Reinforcement Learning

Multi-Agent Craftax: Benchmarking Open-Ended Multi-Agent Reinforcement Learning at the Hyperscale

2025-11-07 · Bassel Al Omari, Michael Matthews, Alexander Rutherford, Jakob Nicolaus Foerster arxiv

Progress in multi-agent reinforcement learning (MARL) requires challenging benchmarks that assess the limits of current methods. However, existing benchmarks often target narrow short-horizon challenges that do not adequ…

Multi-agent Reinforcement Learning

Explore with Long-term Memory: A Benchmark and Multimodal LLM-based Reinforcement Learning Framework for Embodied Exploration

2026-01-11 · Sen Wang, Bangwei Liu, Zhenkun Gao, Lizhuang Ma 외 arxiv

An ideal embodied agent should possess lifelong learning capabilities to handle long-horizon and complex tasks, enabling continuous operation in general environments. This not only requires the agent to accurately accomp…

Reinforcement LearningQuestion Answering