paper-with-me

Papers

ARTICLE: Annotator Reliability Through In-Context Learning

2024-09-18 · Sujan Dutta, Deepak Pandita, Tharindu Cyril Weerasooriya, Marcos Zampieri, Christopher M. Homan, Ashiqur R. KhudaBukhsh

Ensuring annotator quality in training and evaluation data is a key piece of machine learning in NLP. Tasks such as sentiment analysis and offensive speech detection are intrinsically subjective, creating a challenging scenario for traditional quality assessment approaches because it is hard to distinguish disagreement due to poor work from that due to differences of opinions between sincere annotators. With the goal of increasing diverse perspectives in annotation while ensuring consistency, we propose \texttt{ARTICLE}, an in-context learning (ICL) framework to estimate annotation quality through self-consistency. We evaluate this framework on two offensive speech datasets using multiple LLMs and compare its performance with traditional methods. Our findings indicate that \texttt{ARTICLE} can be used as a robust method for identifying reliable annotators, hence improving data quality.

📄 PDF Abstract BibTeX arXiv:2409.12218

Code (0)

등록된 구현이 없습니다.

Tasks

In-Context LearningSentiment Analysis

Similar Papers 제목 키워드 기반

Fact or Fiction? Can LLMs be Reliable Annotators for Political Truths?

2024-11-08 · Veronica Chatrath, Marcelo Lotif, Shaina Raza

Political misinformation poses significant challenges to democratic processes, shaping public opinion and trust in media. Manual fact-checking methods face issues of scalability and annotator bias, while machine learning…

ArticlesFact CheckingMisinformation

Modelling Instance-Level Annotator Reliability for Natural Language Labelling Tasks

2019-05-13 · NAACL 2019 6 · Maolin Li, Arvid Fahlström Myrman, Tingting Mu, Sophia Ananiadou

When constructing models that learn from noisy labels produced by multiple annotators, it is important to accurately estimate the reliability of annotators. Annotators may provide labels of inconsistent quality due to th…

Natural Language Inferencetext-classificationText Classification

Efficient Annotator Reliability Assessment with EffiARA

2025-04-01 · Owen Cook, Jake Vasilakes, Ian Roberts, Xingyi Song

Data annotation is an essential component of the machine learning pipeline; it is also a costly and time-consuming process. With the introduction of transformer-based models, annotation at the document level is increasin…

QUORUM: QUality-Optimized Routing Using Multiple annotators

2026-08-28 · Antonio Purificato, Maria Sofia Bucarelli, Andrea Bacciu, Amin Mantrach 외 arxiv

Data annotation remains a central bottleneck in natural language processing, requiring human effort to obtain high-quality labels at scale. While Large Language Models (LLMs) offer a fast and cost-effective alternative, …

On User Interfaces for Large-Scale Document-Level Human Evaluation of Machine Translation Outputs

2021-04-21 · EACL (HumEval) 2021 4 · Roman Grundkiewicz, Marcin Junczys-Dowmunt, Christian Federmann, Tom Kocmi

Recent studies emphasize the need of document context in human evaluation of machine translations, but little research has been done on the impact of user interfaces on annotator productivity and the reliability of asses…

Machine TranslationTranslation