paper-with-me

홈 › Papers

LongEval-Retrieval: French-English Dynamic Test Collection for Continuous Web Search Evaluation

2023-03-06 · Petra Galuščáková Romain Deveaud, Gabriela Gonzalez-Saez, Philippe Mulhem, Lorraine Goeuriot, Florina Piroi, Martin Popel

LongEval-Retrieval is a Web document retrieval benchmark that focuses on continuous retrieval evaluation. This test collection is intended to be used to study the temporal persistence of Information Retrieval systems and will be used as the test collection in the Longitudinal Evaluation of Model Performance Track (LongEval) at CLEF 2023. This benchmark simulates an evolving information system environment - such as the one a Web search engine operates in - where the document collection, the query distribution, and relevance all move continuously, while following the Cranfield paradigm for offline evaluation. To do that, we introduce the concept of a dynamic test collection that is composed of successive sub-collections each representing the state of an information system at a given time step. In LongEval-Retrieval, each sub-collection contains a set of queries, documents, and soft relevance assessments built from click models. The data comes from Qwant, a privacy-preserving Web search engine that primarily focuses on the French market. LongEval-Retrieval also provides a 'mirror' collection: it is initially constructed in the French language to benefit from the majority of Qwant's traffic, before being translated to English. This paper presents the creation process of LongEval-Retrieval and provides baseline runs and analysis.

📄 PDF Abstract BibTeX arXiv:2303.03229

Code (0)

등록된 구현이 없습니다.

Tasks

Information RetrievalPrivacy PreservingRetrieval

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

LongEval at CLEF 2025: Longitudinal Evaluation of IR Model Performance

2025-03-11 · Matteo Cancellieri, Alaa El-Ebshihy, Tobias Fink, Petra Galuščáková 외

This paper presents the third edition of the LongEval Lab, part of the CLEF 2025 conference, which continues to explore the challenges of temporal persistence in Information Retrieval (IR). The lab features two tasks des…

Information RetrievalRetrieval

Cross-Lingual Training with Dense Retrieval for Document Retrieval

2021-09-03 · Peng Shi, Rui Zhang, He Bai, Jimmy Lin

Dense retrieval has shown great success in passage ranking in English. However, its effectiveness in document retrieval for non-English languages remains unexplored due to the limitation in training resources. In this wo…

Document RankingPassage RankingRetrieval

Submitted and Diagnostic Analysis of Full-Text Temporal Retrieval for LongEval-Sci

2026-07-05 · Yingdong Yang, Haijian Wu arxiv

LongEval-Sci evaluates scientific retrieval under collection change, where a system should be effective on the current corpus and remain usable as documents accumulate over time. This paper reports both official Task 1 r…

Text Retrieval

CURE: A dataset for Clinical Understanding & Retrieval Evaluation

2024-12-09 · Nadia Sheikh, Anne-Laure Jousse, Daniel Buades Marcos, Akintunde Oladipo 외

Given the dominance of dense retrievers that do not generalize well beyond their training dataset distributions, domain-specific test sets are essential in evaluating retrieval. There are few test datasets for retrieval …

Passage RankingRetrieval

Analyzing the Effectiveness of Listwise Reranking with Positional Invariance on Temporal Generalizability

2024-07-09 · Soyoung Yoon, Jongyoon Kim, Seung-won Hwang

This working note outlines our participation in the retrieval task at CLEF 2024. We highlight the considerable gap between studying retrieval performance on static knowledge documents and understanding performance in rea…

BenchmarkingDecoderInformation RetrievalReranking+1