paper-with-me

Papers

Robust Experimentation in the Continuous Time Bandit Problem

2021-03-31 · Farzad Pourbabaee

We study the experimentation dynamics of a decision maker (DM) in a two-armed bandit setup (Bolton and Harris (1999)), where the agent holds ambiguous beliefs regarding the distribution of the return process of one arm and is certain about the other one. The DM entertains Multiplier preferences a la Hansen and Sargent (2001), thus we frame the decision making environment as a two-player differential game against nature in continuous time. We characterize the DM value function and her optimal experimentation strategy that turns out to follow a cut-off rule with respect to her belief process. The belief threshold for exploring the ambiguous arm is found in closed form and is shown to be increasing with respect to the ambiguity aversion index. We then study the effect of provision of an unambiguous information source about the ambiguous arm. Interestingly, we show that the exploration threshold rises unambiguously as a result of this new information source, thereby leading to more conservatism. This analysis also sheds light on the efficient time to reach for an expert opinion.

📄 PDF Abstract BibTeX arXiv:2104.00102

Code (0)

등록된 구현이 없습니다.

Tasks

Decision Making

Similar Papers 제목 키워드 기반

An Opportunistic Bandit Approach for User Interface Experimentation

2020-06-21 · Nader Bouacida, Amit Pande, Xin Liu

Facing growing competition from online rivals, the retail industry is increasingly investing in their online shopping platforms to win the high-stake battle of customer' loyalty. User experience is playing an essential r…

Overcoming Free-Riding in Bandit Games

2019-10-20 · Johannes Hörner, Nicolas Klein, Sven Rady

This paper considers a class of experimentation games with L\'{e}vy bandits encompassing those of Bolton and Harris (1999) and Keller, Rady and Cripps (2005). Its main result is that efficient (perfect Bayesian) equilibr…

Committing Bandits

2011-12-01 · NeurIPS 2011 12 · Loc X. Bui, Ramesh Johari, Shie Mannor

We consider a multi-armed bandit problem where there are two phases. The first phase is an experimentation phase where the decision maker is free to explore multiple options. In the second phase the decision maker has to…

Undiscounted Bandit Games

2020-08-25

We analyze undiscounted continuous-time games of strategic experimentation with two-armed bandits. The risky arm generates payoffs according to a L\'{e}vy process with an unknown average payoff per unit of time which nat…

Using Adaptive Bandit Experiments to Increase and Investigate Engagement in Mental Health

2023-10-13 · Harsh Kumar, Tong Li, Jiakai Shi, Ilya Musabirov 외

Digital mental health (DMH) interventions, such as text-message-based lessons and activities, offer immense potential for accessible mental health support. While these interventions can be effective, real-world experimen…

Thompson Sampling