paper-with-me

홈 › Papers

Multi-Label Annotation Aggregation in Crowdsourcing

2017-06-19 · Xuan Wei, Daniel Dajun Zeng, Junming Yin

As a means of human-based computation, crowdsourcing has been widely used to annotate large-scale unlabeled datasets. One of the obvious challenges is how to aggregate these possibly noisy labels provided by a set of heterogeneous annotators. Another challenge stems from the difficulty in evaluating the annotator reliability without even knowing the ground truth, which can be used to build incentive mechanisms in crowdsourcing platforms. When each instance is associated with many possible labels simultaneously, the problem becomes even harder because of its combinatorial nature. In this paper, we present new flexible Bayesian models and efficient inference algorithms for multi-label annotation aggregation by taking both annotator reliability and label dependency into account. Extensive experiments on real-world datasets confirm that the proposed methods outperform other competitive alternatives, and the model can recover the type of the annotators with high accuracy.

📄 PDF Abstract BibTeX arXiv:1706.06120

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Comparative Study on Annotation Quality of Crowdsourcing and LLM via Label Aggregation

2024-01-18 · Jiyi Li

Whether Large Language Models (LLMs) can outperform crowdsourcing on the data annotation task is attracting interest recently. Some works verified this issue with the average performance of individual crowd workers and L…

Human-LLM Hybrid Text Answer Aggregation for Crowd Annotations

2024-10-22 · Jiyi Li

The quality is a crucial issue for crowd annotations. Answer aggregation is an important type of solution. The aggregated answers estimated from multiple crowd answers to the same instance are the eventually collected an…

Towards Long-term Annotators: A Supervised Label Aggregation Baseline

2023-11-15 · Haoyu Liu, Fei Wang, Minmin Lin, Runze Wu 외

Relying on crowdsourced workers, data crowdsourcing platforms are able to efficiently provide vast amounts of labeled data. Due to the variability in the annotation quality of crowd workers, modern techniques resort to r…

Optimizing the Wisdom of the Crowd: Inference, Learning, and Teaching

2018-06-23 · Yao Zhou, Jingrui He

The unprecedented demand for large amount of data has catalyzed the trend of combining human insights with machine learning techniques, which facilitate the use of crowdsourcing to enlist label information both effective…

MiSC: Mixed Strategies Crowdsourcing

2019-05-17 · Ching-Yun Ko, Rui Lin, Shu Li, Ngai Wong

Popular crowdsourcing techniques mostly focus on evaluating workers' labeling quality before adjusting their weights during label aggregation. Recently, another cohort of models regard crowdsourced annotations as incompl…