paper-with-me

Papers

Context Does Matter: Implications for Crowdsourced Evaluation Labels in Task-Oriented Dialogue Systems

2024-04-15 · Clemencia Siro, Mohammad Aliannejadi, Maarten de Rijke

Crowdsourced labels play a crucial role in evaluating task-oriented dialogue systems (TDSs). Obtaining high-quality and consistent ground-truth labels from annotators presents challenges. When evaluating a TDS, annotators must fully comprehend the dialogue before providing judgments. Previous studies suggest using only a portion of the dialogue context in the annotation process. However, the impact of this limitation on label quality remains unexplored. This study investigates the influence of dialogue context on annotation quality, considering the truncated context for relevance and usefulness labeling. We further propose to use large language models (LLMs) to summarize the dialogue context to provide a rich and short description of the dialogue context and study the impact of doing so on the annotator's performance. Reducing context leads to more positive ratings. Conversely, providing the entire dialogue context yields higher-quality relevance ratings but introduces ambiguity in usefulness ratings. Using the first user utterance as context leads to consistent ratings, akin to those obtained using the entire dialogue, with significantly reduced annotation effort. Our findings show how task design, particularly the availability of dialogue context, affects the quality and consistency of crowdsourced evaluation labels.

📄 PDF Abstract BibTeX arXiv:2404.09980

Code (1)

clemenciah/effects-of-dialogue-context 공식 구현

Tasks

Task-Oriented Dialogue Systems

Similar Papers 제목 키워드 기반

Investigating the Benefits of Free-Form Rationales

2022-01-16 · ACL ARR January 2022 1 · Anonymous

Free-form rationales aim to aid model interpretability by supplying the background knowledge that can help understand model decisions. Crowdsourced rationales are provided for commonsense QA instances in popular datasets…

Form

Investigating the Benefits of Free-Form Rationales

2022-05-25 · Jiao Sun, Swabha Swayamdipta, Jonathan May, Xuezhe Ma

Free-form rationales aim to aid model interpretability by supplying the background knowledge that can help understand model decisions. Crowdsourced rationales are provided for commonsense QA instances in popular datasets…

Form

IERL: Interpretable Ensemble Representation Learning -- Combining CrowdSourced Knowledge and Distributed Semantic Representations

2023-06-24 · Yuxin Zi, Kaushik Roy, Vignesh Narayanan, Manas Gaur 외

Large Language Models (LLMs) encode meanings of words in the form of distributed semantics. Distributed semantics capture common statistical patterns among language tokens (words, phrases, and sentences) from large amoun…

Ensemble LearningHallucinationKnowledge GraphsRepresentation Learning+1

Context Matters: An Empirical Study of the Impact of Contextual Information in Temporal Question Answering Systems

2024-06-27 · Dan Schumacher, Fatemeh Haji, Tara Grey, Niharika Bandlamudi 외

Large language models (LLMs) often struggle with temporal reasoning, crucial for tasks like historical event analysis and time-sensitive information retrieval. Despite advancements, state-of-the-art models falter in hand…

Information RetrievalQuestion AnsweringRetrieval

Can Crowdsourcing Survive the LLM Era? A Community Survey on Human Data Collection

2026-06-03 · Aswathy Velutharambath, Neele Falk, Sofie Labat, Tarun Tater 외 arxiv

The widespread use of Large Language Models (LLMs) as writing tools challenges the validity of crowdsourced data, as crowdworkers may outsource tasks to models. To better understand how this is addressed, we surveyed 155…