paper-with-me

홈 › Papers

LongEval at CLEF 2025: Longitudinal Evaluation of IR Model Performance

2025-03-11 · Matteo Cancellieri, Alaa El-Ebshihy, Tobias Fink, Petra Galuščáková, Gabriela Gonzalez-Saez, Lorraine Goeuriot, David Iommi, Jüri Keller, Petr Knoth, Philippe Mulhem, Florina Piroi, David Pride, Philipp Schaer

This paper presents the third edition of the LongEval Lab, part of the CLEF 2025 conference, which continues to explore the challenges of temporal persistence in Information Retrieval (IR). The lab features two tasks designed to provide researchers with test data that reflect the evolving nature of user queries and document relevance over time. By evaluating how model performance degrades as test data diverge temporally from training data, LongEval seeks to advance the understanding of temporal dynamics in IR systems. The 2025 edition aims to engage the IR and NLP communities in addressing the development of adaptive models that can maintain retrieval quality over time in the domains of web search and scientific retrieval.

📄 PDF Abstract BibTeX arXiv:2503.08541

Code (0)

등록된 구현이 없습니다.

Tasks

Information RetrievalRetrieval

Similar Papers 제목 키워드 기반

LongEval-Retrieval: French-English Dynamic Test Collection for Continuous Web Search Evaluation

2023-03-06 · Petra Galuščáková Romain Deveaud, Gabriela Gonzalez-Saez, Philippe Mulhem, Lorraine Goeuriot 외

LongEval-Retrieval is a Web document retrieval benchmark that focuses on continuous retrieval evaluation. This test collection is intended to be used to study the temporal persistence of Information Retrieval systems and…

Information RetrievalPrivacy PreservingRetrieval

Keeping in Time: Adding Temporal Context to Sentiment Analysis Models

2023-09-24 · Dean Ninalga

This paper presents a state-of-the-art solution to the LongEval CLEF 2023 Lab Task 2: LongEval-Classification. The goal of this task is to improve and preserve the performance of sentiment analysis models across shorter …

Language ModelingLanguage ModellingSentiment AnalysisTask 2

DS@GT ARC at LongEval: Citation Integrity and Factual Grounding in Scientific QA

2026-07-15 · Brandon Michaels, Brendon Johnson arxiv

This paper describes DS@GT ARC's submission to the CLEF 2026 LongEval Task 4 on Retrieval-Augmented Generation (RAG). In this submission, we examine a divergence between traditional natural language evaluation metrics an…

Answer Generation

Evaluating Temporal Persistence Using Replicability Measures

2023-08-21 · Jüri Keller, Timo Breuer, Philipp Schaer

In real-world Information Retrieval (IR) experiments, the Evaluation Environment (EE) is exposed to constant change. Documents are added, removed, or updated, and the information need and the search behavior of users is …

Information RetrievalRetrieval

Submitted and Diagnostic Analysis of Full-Text Temporal Retrieval for LongEval-Sci

2026-07-05 · Yingdong Yang, Haijian Wu arxiv

LongEval-Sci evaluates scientific retrieval under collection change, where a system should be effective on the current corpus and remain usable as documents accumulate over time. This paper reports both official Task 1 r…

Text Retrieval