paper-with-me

홈 › Papers

AgentBuddy: A Contextual Bandit based Decision Support System for Customer Support Agents

2019-02-24 · Hrishikesh Ganu, Mithun Ghosh, Shashi Roshan

In this short paper, we present early insights from a Decision Support System for Customer Support Agents (CSAs) serving customers of a leading accounting software. The system is under development and is designed to provide suggestions to CSAs to make them more productive. A unique aspect of the solution is the use of bandit algorithms to create a tractable human-in-the-loop system that can learn from CSAs in an online fashion. In addition to discussing the ML aspects, we also bring out important insights we gleaned from early feedback from CSAs. These insights motivate our future work and also might be of wider interest to ML practitioners.

📄 PDF Abstract BibTeX arXiv:1903.03512

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Making Contextual Decisions with Low Technical Debt

2016-06-13 · Alekh Agarwal, Sarah Bird, Markus Cozowicz, Luong Hoang 외

Applications and systems are constantly faced with decisions that require picking from a set of actions based on contextual information. Reinforcement-based learning algorithms such as contextual bandits can be very effe…

Multi-Armed Bandits

Human-in-the-Loop Multi-Agent Ventilator Decision Support with Contextual Bandit Preference Learning

2026-05-22 · Sijia Li, Xiaoyu Tan, Qixing Wang, Weiyi Zhao 외 arxiv

Ventilator decision support requires sequential decisions that track evolving physiology and disease trajectories while respecting safety boundaries and clinician specific tuning styles. Rule based approaches rarely gene…

Reinforcement Learning

MABWiser: A Parallelizable Contextual Multi-Armed Bandit Library for Python

2019-10-04 · IEEE 31th International Conference on Tools with Artificial Intelligence, ICTAI 2019 2019 10 · Emily Strong, Bernard Kleynhans, Serdar Kadioglu

Contextual multi-armed bandit algorithms serve as an effective technique to address online sequential decision-making problems. Despite their popularity, when it comes to off-the-shelf tools the library support remains l…

Decision MakingSequential Decision Making

LLMs-augmented Contextual Bandit

2023-11-03 · Ali Baheri, Cecilia O. Alm

Contextual bandits have emerged as a cornerstone in reinforcement learning, enabling systems to make decisions with partial feedback. However, as contexts grow in complexity, traditional bandit algorithms can face challe…

Multi-Armed Banditsreinforcement-learningReinforcement Learning

Contextual memory bandit for pro-active dialog engagement

2018-01-01 · ICLR 2018 1 · julien perez, Tomi Silander

An objective of pro-activity in dialog systems is to enhance the usability of conversational agents by enabling them to initiate conversation on their own. While dialog systems have become increasingly popular during the…

Multi-Armed Bandits