paper-with-me

홈 › Papers

Crowdsourced judgement elicitation with endogenous proficiency

2013-05-13 · WWW 2013 5 · Anirban Dasgupta, Arpita Ghosh

Crowdsourcing is now widely used to replace judgement or evaluation by an expert authority with an aggregate evaluation from a number of non-experts, in applications ranging from rating and categorizing online content all the way to evaluation of student assignments in massively open online courses (MOOCs) via peer grading. A key issue in these settings, where direct monitoring of both effort and accuracy is infeasible, is incentivizing agents in the 'crowd' to put in effort to make good evaluations, as well as to truthfully report their evaluations. We study the design of mechanisms for crowdsourced judgement elicitation when workers strategically choose both their reports and the effort they put into their evaluations. This leads to a new family of information elicitation problems with unobservable ground truth, where an agent's proficiency--- the probability with which she correctly evaluates the underlying ground truth--- is endogenously determined by her strategic choice of how much effort to put into the task. Our main contribution is a simple, new, mechanism for binary information elicitation for multiple tasks when agents have endogenous proficiencies, with the following properties: (i) Exerting maximum effort followed by truthful reporting of observations is a Nash equilibrium. (ii) This is the equilibrium with maximum payoff to all agents, even when agents have different maximum proficiencies, can use mixed strategies, and can choose a different strategy for each of their tasks. Our information elicitation mechanism requires only minimal bounds on the priors, asks agents to only report their own evaluations, and does not require any conditions on a diverging number of agent reports per task to achieve its incentive properties. The main idea behind our mechanism is to use the presence of multiple tasks and ratings to estimate a reporting statistic to identify and penalize low-effort agreement--- the mechanism rewards agents for agreeing with another 'reference' report on the same task, but also penalizes for blind agreement by subtracting out this statistic term, designed so that agents obtain rewards only when they put in effort into their observations.

📄 PDF Abstract BibTeX

Code (1)

knaoe/peer-prediction

Similar Papers 제목 키워드 기반

Reducing annotator bias by belief elicitation

2024-10-21 · Terne Sasha Thorn Jakobsen, Andreas Bjerre-Nielsen, Robert Böhm

Crowdsourced annotations of data play a substantial role in the development of Artificial Intelligence (AI). It is broadly recognised that annotations of text data can contain annotator bias, where systematic disagreemen…

Filling in the Blanks in Understanding Discourse Adverbials: Consistency, Conflict, and Context-Dependence in a Crowdsourced Elicitation Task

2016-08-01 · WS 2016 8 · Hannah Rohde, Anna Dickinson, Nathan Schneider, Christopher N. L. Clark 외

Human acceptability judgements for extractive sentence compression

2019-02-01 · Abram Handler, Brian Dillon, Brendan O'Connor

Recent approaches to English-language sentence compression rely on parallel corpora consisting of sentence-compression pairs. However, a sentence may be shortened in many different ways, which each might be suited to the…

SentenceSentence Compression

Eliciting judgements about dependent quantities of interest: The SHELF extension and copula methods illustrated using an asthma case study

2021-02-04 · Björn Holzhauer, Lisa V. Hampson, John Paul Gosling, Björn Bornkamp 외

Pharmaceutical companies regularly need to make decisions about drug development programs based on the limited knowledge from early stage clinical trials. In this situation, eliciting the judgements of experts is an attr…

Decoding moral judgement from text: a pilot study

2024-05-28 · Diana E. Gherman, Thorsten O. Zander

Moral judgement is a complex human reaction that engages cognitive and emotional dimensions. While some of the morality neural correlates are known, it is currently unclear if we can detect moral violation at a single-tr…

Attribute