paper-with-me

Papers

Evaluating Complex Task through Crowdsourcing: Multiple Views Approach

2017-03-30 · Lingyu Lyu, Mehmed Kantardzic

With the popularity of massive open online courses, grading through crowdsourcing has become a prevalent approach towards large scale classes. However, for getting grades for complex tasks, which require specific skills and efforts for grading, crowdsourcing encounters a restriction of insufficient knowledge of the workers from the crowd. Due to knowledge limitation of the crowd graders, grading based on partial perspectives becomes a big challenge for evaluating complex tasks through crowdsourcing. Especially for those tasks which not only need specific knowledge for grading, but also should be graded as a whole instead of being decomposed into smaller and simpler subtasks. We propose a framework for grading complex tasks via multiple views, which are different grading perspectives defined by experts for the task, to provide uniformity. Aggregation algorithm based on graders variances are used to combine the grades for each view. We also detect bias patterns of the graders, and debias them regarding each view of the task. Bias pattern determines how the behavior is biased among graders, which is detected by a statistical technique. The proposed approach is analyzed on a synthetic data set. We show that our model gives more accurate results compared to the grading approaches without different views and debiasing algorithm.

📄 PDF Abstract BibTeX arXiv:1703.10579

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Evaluating LLM-corrupted Crowdsourcing Data Without Ground Truth

2025-06-08 · Yichi Zhang, Jinlong Pang, Zhaowei Zhu, Yang Liu

The recent success of generative AI highlights the crucial role of high-quality human feedback in building trustworthy AI systems. However, the increasing use of large language models (LLMs) by crowdsourcing workers pose…

Multiple-choice

Evaluating Crowdsourcing Participants in the Absence of Ground-Truth

2016-05-30 · Ramanathan Subramanian, Romer Rosales, Glenn Fung, Jennifer Dy

Given a supervised/semi-supervised learning scenario where multiple annotators are available, we consider the problem of identification of adversarial or unreliable annotators.

Data Quality in Crowdsourcing and Spamming Behavior Detection

2024-04-04 · Yang Ba, Michelle V. Mancenido, Erin K. Chiou, Rong pan

As crowdsourcing emerges as an efficient and cost-effective method for obtaining labels for machine learning datasets, it is important to assess the quality of crowd-provided data, so as to improve analysis performance a…

Face Verification

``Fingers in the Nose'': Evaluating Speakers' Identification of Multi-Word Expressions Using a Slightly Gamified Crowdsourcing Platform

2018-08-01 · COLING 2018 8 · Kar{\"e}n Fort, Bruno Guillaume, Matthieu Constant, Nicolas Lef{\`e}bvre 외

This article presents the results we obtained in crowdsourcing French speakers{'} intuition concerning multi-work expressions (MWEs). We developed a slightly gamified crowdsourcing platform, part of which is designed to …

Crowdsourcing for Evaluating Machine Translation Quality

2014-05-01 · LREC 2014 5 · Shinsuke Goto, Donghui Lin, Toru Ishida

The recent popularity of machine translation has increased the demand for the evaluation of translations. However, the traditional evaluation approach, manual checking by a bilingual professional, is too expensive and to…

Machine TranslationSentenceTranslation