paper-with-me

홈 › Papers

Evaluation Measures for Relevance and Credibility in Ranked Lists

2017-08-23 · Lioma Christina, Simonsen Jakob Grue, Larsen Birger

Recent discussions on alternative facts, fake news, and post truth politics have motivated research on creating technologies that allow people not only to access information, but also to assess the credibility of the information presented to them by information retrieval systems. Whereas technology is in place for filtering information according to relevance and/or credibility, no single measure currently exists for evaluating the accuracy or precision (and more generally effectiveness) of both the relevance and the credibility of retrieved results. One obvious way of doing so is to measure relevance and credibility effectiveness separately, and then consolidate the two measures into one. There at least two problems with such an approach: (I) it is not certain that the same criteria are applied to the evaluation of both relevance and credibility (and applying different criteria introduces bias to the evaluation); (II) many more and richer measures exist for assessing relevance effectiveness than for assessing credibility effectiveness (hence risking further bias). Motivated by the above, we present two novel types of evaluation measures that are designed to measure the effectiveness of both relevance and credibility in ranked lists of retrieval results. Experimental evaluation on a small human-annotated dataset (that we make freely available to the research community) shows that our measures are expressive and intuitive in their interpretation.

📄 PDF Abstract BibTeX arXiv:1708.07157

Code (1)

diku-irlab/A66 공식 구현

Tasks

Information RetrievalRetrieval

Similar Papers 제목 키워드 기반

Online Learning to Rank with Top-k Feedback

2016-08-23 · Sougata Chaudhuri, Ambuj Tewari

We consider two settings of online learning to rank where feedback is restricted to top ranked items. The problem is cast as an online game between a learner and sequence of users, over $T$ rounds. In both settings, the …

Learning-To-Rank

A Versatile Framework for Evaluating Ranked Lists in terms of Group Fairness and Relevance

2022-04-01 · Tetsuya Sakai, Jin Young Kim, Inho Kang

We present a simple and versatile framework for evaluating ranked lists in terms of group fairness and relevance, where the groups (i.e., possible attribute values) can be either nominal or ordinal in nature. First, we d…

AttributeFairness

Principled Multi-Aspect Evaluation Measures of Rankings

2022-12-01 · Maria Maistro, Lucas Chaves Lima, Jakob Grue Simonsen, Christina Lioma

Information Retrieval evaluation has traditionally focused on defining principled ways of assessing the relevance of a ranked list of documents with respect to a query. Several methods extend this type of evaluation beyo…

Document RankingInformation RetrievalRetrieval

Graded Relevance Assessments and Graded Relevance Measures of NTCIR: A Survey of the First Twenty Years

2019-03-27 · Tetsuya Sakai

NTCIR was the first large-scale IR evaluation conference to construct test collections with graded relevance assessments: the NTCIR-1 test collections from 1998 already featured relevant and partially relevant documents.…

RetrievalSurvey

How Relevant is the Long Tail? A Relevance Assessment Study on Million Short

2016-06-20 · Schaer Philipp, Mayr Philipp, Sünkler Sebastian, Lewandowski Dirk

Users of web search engines are known to mostly focus on the top ranked results of the search engine result page. While many studies support this well known information seeking pattern only few studies concentrate on the…