paper-with-me

홈 › Papers

Geometry Meets Incentives: Sample-Efficient Incentivized Exploration with Linear Contexts

2025-06-02 · Benjamin Schiffer, Mark Sellke

In the incentivized exploration model, a principal aims to explore and learn over time by interacting with a sequence of self-interested agents. It has been recently understood that the main challenge in designing incentive-compatible algorithms for this problem is to gather a moderate amount of initial data, after which one can obtain near-optimal regret via posterior sampling. With high-dimensional contexts, however, this \emph{initial exploration} phase requires exponential sample complexity in some cases, which prevents efficient learning unless initial data can be acquired exogenously. We show that these barriers to exploration disappear under mild geometric conditions on the set of available actions, in which case incentive-compatibility does not preclude regret-optimality. Namely, we consider the linear bandit model with actions in the Euclidean unit ball, and give an incentive-compatible exploration algorithm with sample complexity that scales polynomially with the dimension and other parameters.

📄 PDF Abstract BibTeX arXiv:2506.01685

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Incentivized Lipschitz Bandits

2025-08-26 · Sourav Chakraborty, Amit Kiran Rege, Claire Monteleoni, Lijun Chen arxiv

We study incentivized exploration in multi-armed bandit (MAB) settings with infinitely many arms modeled as elements in continuous metric spaces. Unlike classical bandit models, we consider scenarios where the decision-m…

The Price of Incentivizing Exploration: A Characterization via Thompson Sampling and Sample Complexity

2020-02-03 · Mark Sellke, Aleksandrs Slivkins

We consider incentivized exploration: a version of multi-armed bandits where the choice of arms is controlled by self-interested agents, and the algorithm can only issue recommendations. The algorithm controls the flow o…

Multi-Armed BanditsThompson Sampling

Exploration and Persuasion

2024-10-22 · Aleksandrs Slivkins

How to incentivize self-interested agents to explore when they prefer to exploit? Consider a population of self-interested agents that make decisions under uncertainty. They "explore" to acquire new information and "expl…

Extracting Incentives from Black-Box Decisions

2019-10-13 · Yonadav Shavit, William S. Moses

An algorithmic decision-maker incentivizes people to act in certain ways to receive better decisions. These incentives can dramatically influence subjects' behaviors and lives, and it is important that both decision-make…

Reinforcement Learning

Diversity-Incentivized Exploration for Versatile Reasoning

2025-09-30 · Zican Hu, Shilin Zhang, Yafu Li, Jianhao Yan 외 arxiv

Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a crucial paradigm for incentivizing reasoning capabilities in Large Language Models (LLMs). Due to vast state-action spaces and reward sparsity in rea…

Reinforcement Learning