paper-with-me

홈 › Papers

Leveraging Language Models and Bandit Algorithms to Drive Adoption of Battery-Electric Vehicles

2024-10-30 · Keiichi Namikoshi, David A. Shamma, Rumen Iliev, Jingchao Fang, Alexandre Filipowicz, Candice L Hogan, Charlene Wu, Nikos Arechiga

Behavior change interventions are important to coordinate societal action across a wide array of important applications, including the adoption of electrified vehicles to reduce emissions. Prior work has demonstrated that interventions for behavior must be personalized, and that the intervention that is most effective on average across a large group can result in a backlash effect that strengthens opposition among some subgroups. Thus, it is important to target interventions to different audiences, and to present them in a natural, conversational style. In this context, an important emerging application domain for large language models (LLMs) is conversational interventions for behavior change. In this work, we leverage prior work on understanding values motivating the adoption of battery electric vehicles. We leverage new advances in LLMs, combined with a contextual bandit, to develop conversational interventions that are personalized to the values of each study participant. We use a contextual bandit algorithm to learn to target values based on the demographics of each participant. To train our bandit algorithm in an offline manner, we leverage LLMs to play the role of study participants. We benchmark the persuasive effectiveness of our bandit-enhanced LLM against an unaided LLM generating conversational interventions without demographic-targeted values.

📄 PDF Abstract BibTeX arXiv:2410.23371

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

LLMs-augmented Contextual Bandit

2023-11-03 · Ali Baheri, Cecilia O. Alm

Contextual bandits have emerged as a cornerstone in reinforcement learning, enabling systems to make decisions with partial feedback. However, as contexts grow in complexity, traditional bandit algorithms can face challe…

Multi-Armed Banditsreinforcement-learningReinforcement Learning

Competing Bandits: Learning under Competition

2017-02-27 · Yishay Mansour, Aleksandrs Slivkins, Zhiwei Steven Wu

Most modern systems strive to learn from interactions with users, and many engage in exploration: making potentially suboptimal choices for the sake of acquiring new information. We initiate a study of the interplay betw…

Lifelong Learning in Multi-Armed Bandits

2020-12-28 · Matthieu Jedor, Jonathan Louëdec, Vianney Perchet

Continuously learning and leveraging the knowledge accumulated from prior tasks in order to improve future performance is a long standing machine learning problem. In this paper, we study the problem in the multi-armed b…

Lifelong learningMulti-Armed Bandits

Bandit-Driven Batch Selection for Robust Learning under Label Noise

2023-10-31 · Michal Lisicki, Mihai Nica, Graham W. Taylor

We introduce a novel approach for batch selection in Stochastic Gradient Descent (SGD) training, leveraging combinatorial bandit algorithms. Our methodology focuses on optimizing the learning process in the presence of l…

Computational Efficiency

Directional Optimism for Safe Linear Bandits

2023-08-29 · Spencer Hutchinson, Berkay Turan, Mahnoosh Alizadeh

The safe linear bandit problem is a version of the classical stochastic linear bandit problem where the learner's actions must satisfy an uncertain constraint at all rounds. Due its applicability to many real-world setti…