paper-with-me

홈 › Papers

Reference Games as a Testbed for the Alignment of Model Uncertainty and Clarification Requests

2026-01-12 · Manar Ali, Judith Sieker, Sina Zarrieß, Hendrik Buschmeier arxiv

In human conversation, both interlocutors play an active role in maintaining mutual understanding. When listeners are uncertain about what speakers mean, for example, they can request clarification. It is an open question for language models whether they can assume a similar listener role, recognizing and expressing their own uncertainty through clarification. We argue that reference games are a suitable testbed to approach this question as they are controlled, self-contained, and make clarification needs explicit and measurable. To test this, we evaluate three vision-language models comparing a baseline reference resolution task to an experiment where the models are instructed to request clarification when uncertain. The results suggest that even in such simple tasks, models often struggle to recognize internal uncertainty and translate it into adequate clarification behavior. This demonstrates the value of reference games as testbeds for interaction qualities of (vision and) language models.

📄 PDF Abstract BibTeX arXiv:2601.07820

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Asking the Right Question at the Right Time: Human and Model Uncertainty Guidance to Ask Clarification Questions

2024-02-09 · Alberto Testoni, Raquel Fernández

Clarification questions are an essential dialogue tool to signal misunderstanding, ambiguities, and under-specification in language use. While humans are able to resolve uncertainty by asking questions since childhood, m…

Referential ambiguity and clarification requests: comparing human and LLM behaviour

2025-07-14 · Chris Madge, Matthew Purver, Massimo Poesio arxiv

In this work we examine LLMs' ability to ask clarification questions in task-oriented dialogues that follow the asynchronous instruction-giver/instruction-follower format. We present a new corpus that combines two existi…

Games of Knightian Uncertainty as AGI testbeds

2024-06-26 · Spyridon Samothrakis, Dennis J. N. J. Soemers, Damian Machlanski

Arguably, for the latter part of the late 20th and early 21st centuries, games have been seen as the drosophila of AI. Games are a set of exciting testbeds, whose solutions (in terms of identifying optimal players) would…

Board Games

Clarification Is Not Enough: Post-Clarification Answering Remains the Bottleneck in Multi-Turn QA

2026-05-24 · Jinyan Su, Jennifer Healey arxiv

Pluralistic alignment requires systems to adapt to diverse user values, communication styles, and contextual assumptions. We believe that a foundational prerequisite for such alignment enabling accurate preference elicit…

Question Answering

Frictive Policy Optimization for LLMs: Epistemic Intervention, Risk-Sensitive Control, and Reflective Alignment

2026-04-28 · James Pustejovsky, Nikhil Krishnaswamy arxiv

We propose Frictive Policy Optimization (FPO), a framework for learning language model policies that regulate not only what to say, but when and how to intervene in order to manage epistemic and normative risk. Unlike st…