paper-with-me

Papers

Bayesian Nonparametric Reinforcement Learning in LTE and Wi-Fi Coexistence

2021-05-25 · Po-Kan Shih

With the formation of next generation wireless communication, a growing number of new applications like internet of things, autonomous car, and drone is crowding the unlicensed spectrum. Licensed network such as the long-term evolution (LTE) also comes to the unlicensed spectrum for better providing high-capacity contents with low cost. However, LTE was not designed for sharing spectrum with others. A cooperation center for these networks is costly because they possess heterogeneous properties and everyone can enter and leave the spectrum unrestrictedly, so the design will be challenging. Since it is infeasible to incorporate potentially infinite scenarios with one unified design, an alternative solution is to let each network learn its own coexistence policy. Previous solutions only work on fixed scenarios. In this work a reinforcement learning algorithm is presented to cope with the coexistence between Wi-Fi and LTE agents in 5 GHz unlicensed spectrum. The coexistence problem was modeled as a decentralized partially observable Markov decision process (Dec-POMDP) and Bayesian approach was adopted for policy learning with nonparametric prior to accommodate the uncertainty of policy for different agents. A fairness measure was introduced in the reward function to encourage fair sharing between agents. The reinforcement learning was turned into an optimization problem by transforming the value function as likelihood and variational inference for posterior approximation. Simulation results demonstrate that this algorithm can reach high value with compact policy representations, and stay computationally efficient when applying to agent set.

📄 PDF Abstract BibTeX arXiv:2105.12249

Code (0)

등록된 구현이 없습니다.

Tasks

Fairnessreinforcement-learningReinforcement LearningReinforcement Learning (RL)Variational Inference

Methods 이 논문이 사용한 방법론

Variational Inference 설명 없음

Similar Papers 제목 키워드 기반

Bayesian Nonparametric Modelling for Model-Free Reinforcement Learning in LTE-LAA and Wi-Fi Coexistence

2021-07-06 · Po-Kan Shih, Bahman Moraffah

With the arrival of next generation wireless communication, a growing number of new applications like internet of things, autonomous driving systems, and drone are crowding the unlicensed spectrum. Licensed network such …

Autonomous DrivingBayesian InferenceFairnessVariational Inference

Nonparametric Bayesian Policy Priors for Reinforcement Learning

2010-12-01 · NeurIPS 2010 12 · Finale Doshi-Velez, David Wingate, Nicholas Roy, Joshua B. Tenenbaum

We consider reinforcement learning in partially observable domains where the agent can query an expert for demonstrations. Our nonparametric Bayesian approach combines model knowledge, inferred from expert information an…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Reinforcement Learning with Budget-Constrained Nonparametric Function Approximation for Opportunistic Spectrum Access

2017-06-14 · Theodoros Tsiligkaridis, David Romero

Opportunistic spectrum access is one of the emerging techniques for maximizing throughput in congested bands and is enabled by predicting idle slots in spectrum. We propose a kernel-based reinforcement learning approach …

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Nonparametric Bayesian Inverse Reinforcement Learning for Multiple Reward Functions

2012-12-01 · NeurIPS 2012 12 · Jaedeug Choi, Kee-Eung Kim

We present a nonparametric Bayesian approach to inverse reinforcement learning (IRL) for multiple reward functions. Most previous IRL algorithms assume that the behaviour data is obtained from an agent who is optimizing …

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Bayesian Nonparametrics for Offline Skill Discovery

2022-02-09 · Valentin Villecroze, Harry J. Braviner, Panteha Naderian, Chris J. Maddison 외

Skills or low-level policies in reinforcement learning are temporally extended actions that can speed up learning and enable complex behaviours. Recent work in offline reinforcement learning and imitation learning has pr…

Imitation Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)+1