paper-with-me

Papers

A Light-weight, Effective and Efficient Model for Label Aggregation in Crowdsourcing

2022-11-19 · Yi Yang, Zhong-Qiu Zhao, Quan Bai, Qing Liu, Weihua Li

Due to the noises in crowdsourced labels, label aggregation (LA) has emerged as a standard procedure to post-process crowdsourced labels. LA methods estimate true labels from crowdsourced labels by modeling worker qualities. Most existing LA methods are iterative in nature. They need to traverse all the crowdsourced labels multiple times in order to jointly and iteratively update true labels and worker qualities until convergence. Consequently, these methods have high space and time complexities. In this paper, we treat LA as a dynamic system and model it as a Dynamic Bayesian network. From the dynamic model we derive two light-weight algorithms, LA\textsuperscript{onepass} and LA\textsuperscript{twopass}, which can effectively and efficiently estimate worker qualities and true labels by traversing all the labels at most twice. Due to the dynamic nature, the proposed algorithms can also estimate true labels online without re-visiting historical data. We theoretically prove the convergence property of the proposed algorithms, and bound the error of estimated worker qualities. We also analyze the space and time complexities of the proposed algorithms and show that they are equivalent to those of majority voting. Experiments conducted on 20 real-world datasets demonstrate that the proposed algorithms can effectively and efficiently aggregate labels in both offline and online settings even if they traverse all the labels at most twice.

📄 PDF Abstract BibTeX arXiv:2212.00007

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Error Rate Bounds and Iterative Weighted Majority Voting for Crowdsourcing

2014-11-15 · Hongwei Li, Bin Yu

Crowdsourcing has become an effective and popular tool for human-powered computation to label large datasets. Since the workers can be unreliable, it is common in crowdsourcing to assign multiple workers to one task, and…

MiSC: Mixed Strategies Crowdsourcing

2019-05-17 · Ching-Yun Ko, Rui Lin, Shu Li, Ngai Wong

Popular crowdsourcing techniques mostly focus on evaluating workers' labeling quality before adjusting their weights during label aggregation. Recently, another cohort of models regard crowdsourced annotations as incompl…

Optimizing the Wisdom of the Crowd: Inference, Learning, and Teaching

2018-06-23 · Yao Zhou, Jingrui He

The unprecedented demand for large amount of data has catalyzed the trend of combining human insights with machine learning techniques, which facilitate the use of crowdsourcing to enlist label information both effective…

A Comparative Study on Annotation Quality of Crowdsourcing and LLM via Label Aggregation

2024-01-18 · Jiyi Li

Whether Large Language Models (LLMs) can outperform crowdsourcing on the data annotation task is attracting interest recently. Some works verified this issue with the average performance of individual crowd workers and L…

Mitigating Observation Biases in Crowdsourced Label Aggregation

2023-02-25 · Ryosuke Ueda, Koh Takeuchi, Hisashi Kashima

Crowdsourcing has been widely used to efficiently obtain labeled datasets for supervised learning from large numbers of human resources at low cost. However, one of the technical challenges in obtaining high-quality resu…

Causal Inference