paper-with-me

홈 › Papers

Practical Annotation Strategies for Question Answering Datasets

2020-03-06 · Bernhard Kratzwald, Xiang Yue, Huan Sun, Stefan Feuerriegel

Annotating datasets for question answering (QA) tasks is very costly, as it requires intensive manual labor and often domain-specific knowledge. Yet strategies for annotating QA datasets in a cost-effective manner are scarce. To provide a remedy for practitioners, our objective is to develop heuristic rules for annotating a subset of questions, so that the annotation cost is reduced while maintaining both in- and out-of-domain performance. For this, we conduct a large-scale analysis in order to derive practical recommendations. First, we demonstrate experimentally that more training samples contribute often only to a higher in-domain test-set performance, but do not help the model in generalizing to unseen datasets. Second, we develop a model-guided annotation strategy: it makes a recommendation with regard to which subset of samples should be annotated. Its effectiveness is demonstrated in a case study based on domain customization of QA to a clinical setting. Here, remarkably, annotating a stratified subset with only 1.2% of the original training set achieves 97.7% of the performance as if the complete dataset was annotated. Hence, the labeling effort can be reduced immensely. Altogether, our work fulfills a demand in practice when labeling budgets are limited and where thus recommendations are needed for annotating QA datasets more cost-effectively.

📄 PDF Abstract BibTeX arXiv:2003.03235

Code (0)

등록된 구현이 없습니다.

Tasks

Question Answering

Similar Papers 제목 키워드 기반

Detecting Temporal Ambiguity in Questions

2024-09-25 · Bhawna Piryani, Abdelrahman Abdallah, Jamshid Mozafari, Adam Jatowt

Detecting and answering ambiguous questions has been a challenging task in open-domain question answering. Ambiguous questions have different answers depending on their interpretation and can take diverse forms. Temporal…

Open-Domain Question AnsweringQuestion Answering

Fine-tuning Strategies for Domain Specific Question Answering under Low Annotation Budget Constraints

2022-01-16 · ACL ARR January 2022 1 · Anonymous

The progress introduced by pre-trained language models and their fine-tuning has resulted in significant improvements in most downstream NLP tasks. The unsupervised fine-tuning of a language model combined with further t…

Language ModelingLanguage ModellingQuestion Answering

Coarse-grained decomposition and fine-grained interaction for multi-hop question answering

2021-01-15 · Xing Cao, Yun Liu

Recent advances regarding question answering and reading comprehension have resulted in models that surpass human performance when the answer is contained in a single, continuous passage of text, requiring only single-ho…

Multi-hop Question AnsweringQuestion AnsweringReading Comprehension

Consensus or Conflict? Fine-Grained Evaluation of Conflicting Answers in Question-Answering

2025-08-17 · Eviatar Nachshoni, Arie Cattan, Shmuel Amar, Ori Shapira 외 arxiv

Large Language Models (LLMs) have demonstrated strong performance in question answering (QA) tasks. However, Multi-Answer Question Answering (MAQA), where a question may have several valid answers, remains challenging. T…

Question Answering

MedThink: Explaining Medical Visual Question Answering via Multimodal Decision-Making Rationale

2024-04-18 · Xiaotang Gai, Chenyi Zhou, Jiaxiang Liu, Yang Feng 외

Medical Visual Question Answering (MedVQA), which offers language responses to image-based medical inquiries, represents a challenging task and significant advancement in healthcare. It assists medical experts to swiftly…

Decision MakingMedical Visual Question AnsweringQuestion AnsweringVisual Question Answering+1