paper-with-me

Papers

Reducing conversational agents' overconfidence through linguistic calibration

2020-12-30 · Sabrina J. Mielke, Arthur Szlam, Emily Dinan, Y-Lan Boureau

While improving neural dialogue agents' factual accuracy is the object of much research, another important aspect of communication, less studied in the setting of neural dialogue, is transparency about ignorance. In this work, we analyze to what extent state-of-the-art chit-chat models are linguistically calibrated in the sense that their verbalized expression of doubt (or confidence) matches the likelihood that the model's responses are factually incorrect (or correct). We find that these models are poorly calibrated, yet we show that likelihood of correctness can accurately be predicted. By incorporating such metacognitive features into the training of a controllable generation model, we obtain a dialogue agent with greatly improved linguistic calibration. While improving neural dialogue agents' factual accuracy is the object of much research, another important aspect of communication, less studied in the setting of neural dialogue, is transparency about ignorance. In this work, we analyze to what extent state-of-the-art chit-chat models are linguistically calibrated in the sense that their verbalized expression of doubt (or confidence) matches the likelihood that the model's responses are factually incorrect (or correct). We find that these models are poorly calibrated, yet we show that likelihood of correctness can accurately be predicted. By incorporating such metacognitive features into the training of a controllable generation model, we obtain a dialogue agent with greatly improved linguistic calibration.

📄 PDF Abstract BibTeX arXiv:2012.14983

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

In conversation with Artificial Intelligence: aligning language models with human values

2022-09-01 · Atoosa Kasirzadeh, Iason Gabriel

Large-scale language technologies are increasingly used in various forms of communication with humans across different contexts. One particular use case for these technologies is conversational agents, which output natur…

Simplification Is All You Need against Out-of-Distribution Overconfidence

2025-01-01 · CVPR 2025 1 · Keke Tang, Chao Hou, Weilong Peng, Xiang Fang 외

Deep neural networks (DNNs) often exhibit out-of-distribution (OOD) overconfidence, producing overly confident predictions on OOD samples. We attribute this issue to the inherent over-complexity of DNNs and investiga…

AllAttributeKnowledge Distillation

Can You be More Social? Injecting Politeness and Positivity into Task-Oriented Conversational Agents

2020-12-29 · Yi-Chia Wang, Alexandros Papangelis, Runze Wang, Zhaleh Feizollahi 외

Goal-oriented conversational agents are becoming prevalent in our daily lives. For these systems to engage users and achieve their goals, they need to exhibit appropriate social behavior as well as provide informative re…

Toward zero-shot Entity Recognition in Task-oriented Conversational Agents

2018-07-01 · WS 2018 7 · Marco Guerini, Simone Magnolini, Vevake Balaraman, Bernardo Magnini

We present a domain portable zero-shot learning approach for entity recognition in task-oriented conversational agents, which does not assume any annotated sentences at training time. Rather, we derive a neural model of …

Zero-Shot Learning

The Bots of Persuasion: Examining How Conversational Agents' Linguistic Expressions of Personality Affect User Perceptions and Decisions

2026-02-19 · Uğur Genç, Heng Gu, Chadha Degachi, Evangelos Niforatos 외 arxiv

Large Language Model-powered conversational agents (CAs) are increasingly capable of projecting sophisticated personalities through language, but how these projections affect users is unclear. We thus examine how CA pers…