paper-with-me

Papers

A Comparative Study on Annotation Quality of Crowdsourcing and LLM via Label Aggregation

2024-01-18 · Jiyi Li

Whether Large Language Models (LLMs) can outperform crowdsourcing on the data annotation task is attracting interest recently. Some works verified this issue with the average performance of individual crowd workers and LLM workers on some specific NLP tasks by collecting new datasets. However, on the one hand, existing datasets for the studies of annotation quality in crowdsourcing are not yet utilized in such evaluations, which potentially provide reliable evaluations from a different viewpoint. On the other hand, the quality of these aggregated labels is crucial because, when utilizing crowdsourcing, the estimated labels aggregated from multiple crowd labels to the same instances are the eventually collected labels. Therefore, in this paper, we first investigate which existing crowdsourcing datasets can be used for a comparative study and create a benchmark. We then compare the quality between individual crowd labels and LLM labels and make the evaluations on the aggregated labels. In addition, we propose a Crowd-LLM hybrid label aggregation method and verify the performance. We find that adding LLM labels from good LLMs to existing crowdsourcing datasets can enhance the quality of the aggregated labels of the datasets, which is also higher than the quality of LLM labels themselves.

📄 PDF Abstract BibTeX arXiv:2401.09760

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Efficient Online Crowdsourcing with Complex Annotations

2024-01-25 · Reshef Meir, Viet-An Nguyen, Xu Chen, Jagdish Ramakrishnan 외

Crowdsourcing platforms use various truth discovery algorithms to aggregate annotations from multiple labelers. In an online setting, however, the main challenge is to decide whether to ask for more annotations for each …

Rethinking Crowdsourcing Annotation: Partial Annotation with Salient Labels for Multi-Label Image Classification

2021-09-06 · Jianzhe Lin, Tianze Yu, Z. Jane Wang

Annotated images are required for both supervised model training and evaluation in image classification. Manually annotating images is arduous and expensive, especially for multi-labeled images. A recent trend for conduc…

Active Learningimage-classificationImage ClassificationMulti-Label Image Classification

Crowdsourcing Natural Language Data at Scale: A Hands-On Tutorial

2021-06-01 · NAACL 2021 4 · Alexey Drutsa, Dmitry Ustalov, Valentina Fedorova, Olga Megorskaya 외

In this tutorial, we present a portion of unique industry experience in efficient natural language data annotation via crowdsourcing shared by both leading researchers and engineers from Yandex. We will make an introduct…

Improve Learning from Crowds via Generative Augmentation

2021-07-22 · Zhendong Chu, Hongning Wang

Crowdsourcing provides an efficient label collection schema for supervised machine learning. However, to control annotation cost, each instance in the crowdsourced data is typically annotated by a small number of annotat…

BIG-bench Machine LearningData Augmentation

Controlled Crowdsourcing for High-Quality QA-SRL Annotation

2019-11-08 · ACL 2020 6 · Paul Roit, Ayal Klein, Daniela Stepanov, Jonathan Mamou 외

Question-answer driven Semantic Role Labeling (QA-SRL) was proposed as an attractive open and natural flavour of SRL, potentially attainable from laymen. Recently, a large-scale crowdsourced QA-SRL corpus and a trained p…

Semantic Role LabelingVocal Bursts Intensity Prediction