paper-with-me

Papers

DialDefer: A Framework for Detecting and Mitigating LLM Dialogic Deference

2026-01-15 · Parisa Rabbani, Priyam Sahoo, Ruben Mathew, Aishee Mondal, Harshita Ketharaman, Nimet Beyza Bozdag, Dilek Hakkani-Tür arxiv

LLMs are increasingly used as third-party judges, yet their reliability when evaluating speakers in dialogue remains poorly understood. We show that LLMs judge identical claims differently depending on framing: the same content receives different verdicts when presented as a statement to verify ("Is this statement correct?") versus attributed to a speaker ("Is this speaker correct?"). We call this dialogic deference and introduce DialDefer, a framework for detecting and mitigating these framing-induced judgment shifts. Our Dialogic Deference Score (DDS) captures directional shifts that aggregate accuracy obscures. Across ten domains, 3k+ instances, and five models, conversational framing induces large shifts (mean|DDS|=15.9 percentage points (pp) across models, p < .0001) while accuracy remains stable (<2 pp), with effects amplifying 2--5x on naturalistic Reddit conversations. This effect is domain-dependent: a single model can shift toward disagreement (skepticism) on graduate-level science and toward agreement (deference) on social judgment. Ablations reveal that human-vs-LLM attribution drives the largest shifts (17.7 pp swing), suggesting models treat disagreement with humans as more costly than with AI. Mitigation attempts can reduce deference but over-correct into skepticism, revealing a calibration problem beyond accuracy optimization.

📄 PDF Abstract BibTeX arXiv:2601.10896

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

KNOWCOMP POKEMON Team at DialAM-2024: A Two-Stage Pipeline for Detecting Relations in Dialogical Argument Mining

2024-07-29 · Zihao Zheng, Zhaowei Wang, Qing Zong, Yangqiu Song

Dialogical Argument Mining(DialAM) is an important branch of Argument Mining(AM). DialAM-2024 is a shared task focusing on dialogical argument mining, which requires us to identify argumentative relations and illocutiona…

Argument MiningPrediction

Chevron's Sliding Scale in Wyeth v. Levine, 129 S. Ct. 1187 (2009)

2023-06-05 · Gregory M. Dickinson

In Wyeth v. Levine the Supreme Court once again failed to reconcile the interpretive presumption against preemption with the sometimes competing Chevron doctrine of deference to agencies' reasonable statutory interpretat…

Analysis of Dialogical Argumentation via Finite State Machines

2014-04-29 · Anthony Hunter

Dialogical argumentation is an important cognitive activity by which agents exchange arguments and counterarguments as part of some process such as discussion, debate, persuasion and negotiation. Whilst numerous formal s…

Detecting Argumentative Discourse Acts with Linguistic Alignment

2019-08-01 · WS 2019 8 · Timothy Niven, Hung-Yu Kao

We report the results of preliminary investigations into the relationship between linguistic alignment and dialogical argumentation at the level of discourse acts. We annotated a proof of concept dataset with illocutions…

Epistemic Deference to AI

2025-10-23 · Benjamin Lange arxiv

When should we defer to AI outputs over human expert judgment? Drawing on recent work in social epistemology, I motivate the idea that some AI systems qualify as Artificial Epistemic Authorities (AEAs) due to their demon…