paper-with-me

홈 › Papers

Bandit Models of Human Behavior: Reward Processing in Mental Disorders

2017-06-07 · Djallel Bouneffouf, Irina Rish, Guillermo A. Cecchi

Drawing an inspiration from behavioral studies of human decision making, we propose here a general parametric framework for multi-armed bandit problem, which extends the standard Thompson Sampling approach to incorporate reward processing biases associated with several neurological and psychiatric conditions, including Parkinson's and Alzheimer's diseases, attention-deficit/hyperactivity disorder (ADHD), addiction, and chronic pain. We demonstrate empirically that the proposed parametric approach can often outperform the baseline Thompson Sampling on a variety of datasets. Moreover, from the behavioral modeling perspective, our parametric framework can be viewed as a first step towards a unifying computational model capturing reward processing abnormalities across multiple mental conditions.

📄 PDF Abstract BibTeX arXiv:1706.02897

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingThompson Sampling

Similar Papers 제목 키워드 기반

Unified Models of Human Behavioral Agents in Bandits, Contextual Bandits and RL

2020-05-10 · Baihan Lin, Guillermo Cecchi, Djallel Bouneffouf, Jenna Reinen 외

Artificial behavioral agents are often evaluated based on their consistent behaviors and performance to take sequential actions in an environment to maximize some notion of cumulative reward. However, human decision maki…

Decision MakingLifelong learningMulti-Armed BanditsReinforcement Learning (RL)+1

Learning interactions to boost human creativity with bandits and GPT-4

2023-11-16 · Ara Vartanian, Xiaoxi Sun, Yun-Shiuan Chuang, Siddharth Suresh 외

This paper considers how interactions with AI algorithms can boost human creative thought. We employ a psychological task that demonstrates limits on human creativity, namely semantic feature generation: given a concept …

Modeling Human Decision-making in Generalized Gaussian Multi-armed Bandits

2013-07-23 · Paul Reverdy, Vaibhav Srivastava, Naomi E. Leonard

We present a formal model of human decision-making in explore-exploit tasks using the context of multi-armed bandit problems, where the decision-maker must choose among multiple options with uncertain rewards. We address…

Bayesian InferenceDecision MakingMulti-Armed Bandits

Reinforcement Learning Models of Human Behavior: Reward Processing in Mental Disorders

2019-09-11 · NeurIPS Workshop Neuro_AI 2019 12 · Baihan Lin, Guillermo Cecchi, Djallel Bouneffouf, Jenna Reinen 외

Drawing an inspiration from behavioral studies of human decision making, we propose here a general parametric framework for a reinforcement learning problem, which extends the standard Q-learning approach to incorporate …

Decision MakingQ-LearningRecommendation Systemsreinforcement-learning+1

Semi-Parametric Contextual Bandits with Graph-Laplacian Regularization

2022-05-17 · Young-Geun Choi, Gi-Soo Kim, Seunghoon Paik, Myunghee Cho Paik

Non-stationarity is ubiquitous in human behavior and addressing it in the contextual bandits is challenging. Several works have addressed the problem by investigating semi-parametric contextual bandits and warned that ig…

Multi-Armed BanditsThompson Sampling