Multi-Agent Reinforcement Learning for Long-Term Network Resource Allocation through Auction: a V2X Application
We formulate offloading of computational tasks from a dynamic group of mobile agents (e.g., cars) as decentralized decision making among autonomous agents. We design an interaction mechanism that incentivizes such agents to align private and system goals by balancing between competition and cooperation. In the static case, the mechanism provably has Nash equilibria with optimal resource allocation. In a dynamic environment, this mechanism's requirement of complete information is impossible to achieve. For such environments, we propose a novel multi-agent online learning algorithm that learns with partial, delayed and noisy state information, thus greatly reducing information need. Our algorithm is also capable of learning from long-term and sparse reward signals with varying delay. Empirical results from the simulation of a V2X application confirm that through learning, agents with the learning algorithm significantly improve both system and individual performance, reducing up to 30% of offloading failure rate, communication overhead and load variation, increasing computation resource utilization and fairness. Results also confirm the algorithm's good convergence and generalization property in different environments.
Code (0)
등록된 구현이 없습니다.
Tasks
Decision MakingFairnessMulti-agent Reinforcement LearningMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Multi-Agent Reinforcement Learning Based Resource Allocation for UAV Networks
Unmanned aerial vehicles (UAVs) are capable of serving as aerial base stations (BSs) for providing both cost-effective and on-demand wireless communications. This article investigates dynamic resource allocation of multi…
Multi-agent Reinforcement LearningQ-Learningreinforcement-learningReinforcement Learning+1Long-Term Fairness in Sequential Multi-Agent Selection with Positive Reinforcement
While much of the rapidly growing literature on fair decision-making focuses on metrics for one-shot decisions, recent work has raised the intriguing possibility of designing sequential decision-making to positively impa…
Decision MakingFairnessSequential Decision MakingCache-Aided NOMA Mobile Edge Computing: A Reinforcement Learning Approach
A novel non-orthogonal multiple access (NOMA) based cache-aided mobile edge computing (MEC) framework is proposed. For the purpose of efficiently allocating communication and computation resources to users' computation t…
Edge-computingQ-Learningreinforcement-learningReinforcement Learning+1Adaptive Resource Management for Edge Network Slicing using Incremental Multi-Agent Deep Reinforcement Learning
Multi-access edge computing provides local resources in mobile networks as the essential means for meeting the demands of emerging ultra-reliable low-latency communications. At the edge, dynamic computing requests requir…
Deep Reinforcement LearningEdge-computingIncremental LearningManagement+1Deep Reinforcement Learning for Distributed and Uncoordinated Cognitive Radios Resource Allocation
This paper presents a novel deep reinforcement learning-based resource allocation technique for the multi-agent environment presented by a cognitive radio network where the interactions of the agents during learning may …
Deep Reinforcement LearningQ-Learningreinforcement-learningReinforcement Learning+1