paper-with-me

홈 › Papers

Stability Enforced Bandit Algorithms for Channel Selection in Remote State Estimation of Gauss-Markov Processes

2022-05-20 · Alex S. Leong, Daniel E. Quevedo, Wanchun Liu

In this paper we consider the problem of remote state estimation of a Gauss-Markov process, where a sensor can, at each discrete time instant, transmit on one out of M different communication channels. A key difficulty of the situation at hand is that the channel statistics are unknown. We study the case where both learning of the channel reception probabilities and state estimation is carried out simultaneously. Methods for choosing the channels based on techniques for multi-armed bandits are presented, and shown to provide stability. Furthermore, we define the performance notion of estimation regret, and derive bounds on how it scales with time for the considered algorithms.

📄 PDF Abstract BibTeX arXiv:2205.09923

Code (0)

등록된 구현이 없습니다.

Tasks

channel selectionMulti-Armed BanditsState Estimation

Similar Papers 제목 키워드 기반

Upper Confidence Bounds for Combining Stochastic Bandits

2020-12-24 · Ashok Cutkosky, Abhimanyu Das, Manish Purohit

We provide a simple method to combine stochastic bandit algorithms. Our approach is based on a "meta-UCB" procedure that treats each of $N$ individual bandit algorithms as arms in a higher-level $N$-armed bandit problem …

Model Selection

Dynamic Rate and Channel Selection in Cognitive Radio Systems

2014-02-23 · Richard Combes, Alexandre Proutiere

In this paper, we investigate dynamic channel and rate selection in cognitive radio systems which exploit a large number of channels free from primary users. In such systems, transmitters may rapidly change the selected …

channel selection

Multi-user lax communications: a multi-armed bandit approach

2015-04-30 · Orly Avner, Shie Mannor

Inspired by cognitive radio networks, we consider a setting where multiple users share several channels modeled as a multi-user multi-armed bandit (MAB) problem. The characteristics of each channel are unknown and are di…

Structured Exploration vs. Generative Flexibility: A Field Study Comparing Bandit and LLM Architectures for Personalised Health Behaviour Interventions

2026-03-06 · Dominik P. Hofer, Haochen Song, Rania Islambouli, Laura Hawkins 외 arxiv

Behaviour Change Techniques (BCTs) are central to digital health interventions, yet selecting and delivering effective techniques remains challenging. Contextual bandits enable statistically grounded optimisation of BCT …

On Instability of Minimax Optimal Optimism-Based Bandit Algorithms

2025-11-24 · Samya Praharaj, Koulik Khamaru arxiv

Statistical inference from data generated by multi-armed bandit (MAB) algorithms is challenging due to their adaptive, non-i.i.d. nature. A classical manifestation is that sample averages of arm rewards under bandit samp…