paper-with-me

홈 › Papers

Beyond Single Models: Enhancing LLM Detection of Ambiguity in Requests through Debate

2025-07-16 · Ana Davila, Jacinto Colan, Yasuhisa Hasegawa arxiv

Large Language Models (LLMs) have demonstrated significant capabilities in understanding and generating human language, contributing to more natural interactions with complex systems. However, they face challenges such as ambiguity in user requests processed by LLMs. To address these challenges, this paper introduces and evaluates a multi-agent debate framework designed to enhance detection and resolution capabilities beyond single models. The framework consists of three LLM architectures (Llama3-8B, Gemma2-9B, and Mistral-7B variants) and a dataset with diverse ambiguities. The debate framework markedly enhanced the performance of Llama3-8B and Mistral-7B variants over their individual baselines, with Mistral-7B-led debates achieving a notable 76.7% success rate and proving particularly effective for complex ambiguities and efficient consensus. While acknowledging varying model responses to collaborative strategies, these findings underscore the debate framework's value as a targeted method for augmenting LLM capabilities. This work offers important insights for developing more robust and adaptive language understanding systems by showing how structured debates can lead to improved clarity in interactive systems.

📄 PDF Abstract BibTeX arXiv:2507.12370

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

When Search Agents Should Ask: DiscoBench for Clarification-Aware Deep Search

2026-06-26 · Yiling Tao, Shihan Deng, Meiling Tao, Pengzhi Wei 외 hf

Search agents powered by large language models (LLMs) are increasingly used to solve complex information-seeking tasks, requiring multi-step retrieval and reasoning to fulfill user goals. However, existing benchmarks oft…

Referential ambiguity and clarification requests: comparing human and LLM behaviour

2025-07-14 · Chris Madge, Matthew Purver, Massimo Poesio arxiv

In this work we examine LLMs' ability to ask clarification questions in task-oriented dialogues that follow the asynchronous instruction-giver/instruction-follower format. We present a new corpus that combines two existi…

It Depends: Resolving Referential Ambiguity in Minimal Contexts with Commonsense Knowledge

2025-09-19 · Lukas Ellinger, Georg Groh arxiv

Ambiguous words or underspecified references require interlocutors to resolve them, often by relying on shared context and commonsense knowledge. Therefore, we systematically investigate whether Large Language Models (LL…

Distributionally Robust Optimization via Generative Ambiguity Modeling

2026-02-09 · Jiaqi Wen, Jianyi Yang arxiv

This paper studies Distributionally Robust Optimization (DRO), a fundamental framework for enhancing the robustness and generalization of statistical learning and optimization. An effective ambiguity set for DRO must inv…

Distributionally Robust Optimization via Diffusion Ambiguity Modeling

2025-10-26 · Jiaqi Wen, Jianyi Yang arxiv

This paper studies Distributionally Robust Optimization (DRO), a fundamental framework for enhancing the robustness and generalization of statistical learning and optimization. An effective ambiguity set for DRO must inv…