paper-with-me

홈 › Papers

DiscoTrace: Representing and Comparing Answering Strategies of Humans and LLMs in Information-Seeking Question Answering

2026-04-16 · Neha Srikanth, Jordan Boyd-Graber, Rachel Rudinger arxiv

We introduce DiscoTrace, a method to identify the rhetorical strategies that answerers use when responding to information-seeking questions. DiscoTrace represents answers as a sequence of question-related discourse acts paired with interpretations of the original question, annotated on top of rhetorical structure theory parses. Applying DiscoTrace to answers from nine different human communities reveals that communities have diverse preferences for answer construction. In contrast, LLMs do not exhibit rhetorical diversity in their answers, even when prompted to mimic specific human community answering guidelines. LLMs also systematically opt for breadth, addressing interpretations of questions that human answerers choose not to address. Our findings can guide the development of pragmatic LLM answerers that consider a range of strategies informed by context in QA.

📄 PDF Abstract BibTeX arXiv:2604.15140

Code (0)

등록된 구현이 없습니다.

Tasks

Question Answering

Similar Papers 제목 키워드 기반

Cheater's Bowl: Human vs. Computer Search Strategies for Open-Domain Question Answering

2022-11-15 · Wanrong He, Andrew Mao, Jordan Boyd-Graber

For humans and computers, the first step in answering an open-domain question is retrieving a set of relevant documents from a large corpus. However, the strategies that computers use fundamentally differ from those of h…

Open-Domain Question AnsweringQuestion AnsweringWorld Knowledge

Dimensions underlying the representational alignment of deep neural networks with humans

2024-06-27 · Florian P. Mahner, Lukas Muttenthaler, Umut Güçlü, Martin N. Hebart

Determining the similarities and differences between humans and artificial intelligence (AI) is an important goal both in computational cognitive neuroscience and machine learning, promising a deeper understanding of hum…

Robusto-1 Dataset: Comparing Humans and VLMs on real out-of-distribution Autonomous Driving VQA from Peru

2025-03-10 · Dunant Cusipuma, David Ortega, Victor Flores-Benites, Arturo Deza

As multimodal foundational models start being deployed experimentally in Self-Driving cars, a reasonable question we ask ourselves is how similar to humans do these systems respond in certain driving situations -- especi…

Autonomous DrivingQuestion AnsweringSelf-Driving CarsVisual Question Answering+1

Performance Gains of LLMs With Humans in a World of LLMs Versus Humans

2025-05-13 · Lucas McCullum, Pelagie Ami Agassi, Leo Anthony Celi, Daniel K. Ebner 외

Currently, a considerable research effort is devoted to comparing LLMs to a group of human experts, where the term "expert" is often ill-defined or variable, at best, in a state of constantly updating LLM releases. Witho…

Comparing Humans and Models on a Similar Scale: Towards Cognitive Gender Bias Evaluation in Coreference Resolution

2023-05-24 · Gili Lior, Gabriel Stanovsky

Spurious correlations were found to be an important factor explaining model performance in various NLP tasks (e.g., gender or racial artifacts), often considered to be ''shortcuts'' to the actual task. However, humans te…

coreference-resolutionCoreference ResolutionDecision MakingQuestion Answering