paper-with-me

Papers

Active Multi-Label Crowd Consensus

2019-11-07 · Jinzheng Tu, Guoxian Yu, Carlotta Domeniconi, Jun Wang, Xiangliang Zhang

Crowdsourcing is an economic and efficient strategy aimed at collecting annotations of data through an online platform. Crowd workers with different expertise are paid for their service, and the task requester usually has a limited budget. How to collect reliable annotations for multi-label data and how to compute the consensus within budget is an interesting and challenging, but rarely studied, problem. In this paper, we propose a novel approach to accomplish Active Multi-label Crowd Consensus (AMCC). AMCC accounts for the commonality and individuality of workers, and assumes that workers can be organized into different groups. Each group includes a set of workers who share a similar annotation behavior and label correlations. To achieve an effective multi-label consensus, AMCC models workers' annotations via a linear combination of commonality and individuality, and reduces the impact of unreliable workers by assigning smaller weights to the group. To collect reliable annotations with reduced cost, AMCC introduces an active crowdsourcing learning strategy that selects sample-label-worker triplets. In a triplet, the selected sample and label are the most informative for the consensus model, and the selected worker can reliably annotate the sample with low cost. Our experimental results on multi-label datasets demonstrate the advantages of AMCC over state-of-the-art solutions on computing crowd consensus and on reducing the budget by choosing cost-effective triplets.

📄 PDF Abstract BibTeX arXiv:1911.02789

Code (0)

등록된 구현이 없습니다.

Tasks

Triplet

Similar Papers 제목 키워드 기반

CROWDLAB: Supervised learning to infer consensus labels and quality scores for data with multiple annotators

2022-10-13 · Hui Wen Goh, Ulyana Tkachenko, Jonas Mueller

Real-world data for classification is often labeled by multiple annotators. For analyzing such data, we introduce CROWDLAB, a straightforward approach to utilize any trained classifier to estimate: (1) A consensus label …

C2A: Crowd Consensus Analytics for Virtual Colonoscopy

2018-10-21 · Ji Hwan Park, Saad Nadeem, Seyedkoosha Mirhosseini, Arie Kaufman

We present a medical crowdsourcing visual analytics platform called C{$^2$}A to visualize, classify and filter crowdsourced clinical data. More specifically, C$^2$A is used to build consensus on a clinical diagnosis by v…

Ranking annotators for crowdsourced labeling tasks

2011-12-01 · NeurIPS 2011 12 · Vikas C. Raykar, Shipeng Yu

With the advent of crowdsourcing services it has become quite cheap and reasonably effective to get a dataset labeled by multiple annotators in a short amount of time. Various methods have been proposed to estimate the c…

Crowd disagreement about medical images is informative

2018-06-21 · Veronika Cheplygina, Josien P. W. Pluim

Classifiers for medical image analysis are often trained with a single consensus label, based on combining labels given by experts or crowds. However, disagreement between annotators may be informative, and thus removing…

Medical Image Analysis

Fairness and Bias in Truth Discovery Algorithms: An Experimental Analysis

2023-04-25 · Simone Lazier, Saravanan Thirumuruganathan, Hadis Anahideh

Machine learning (ML) based approaches are increasingly being used in a number of applications with societal impact. Training ML models often require vast amounts of labeled data, and crowdsourcing is a dominant paradigm…

Fairness