paper-with-me

홈 › Papers

Evaluating Research Novelty Detection: Counterfactual Approaches

2019-11-01 · WS 2019 11 · Reinald Kim Amplayo, Seung-won Hwang, Min Song

In this paper, we explore strategies to evaluate models for the task research paper novelty detection: Given all papers released at a given date, which of the papers discuss new ideas and influence future research? We find the novelty is not a singular concept, and thus inherently lacks of ground truth annotations with cross-annotator agreement, which is a major obstacle in evaluating these models. Test-of-time award is closest to such annotation, which can only be made retrospectively and is extremely scarce. We thus propose to compare and evaluate models using counterfactual simulations. First, we ask models if they can differentiate papers at time $t$ and counterfactual paper from future time $t+d$. Second, we ask models if they can predict test-of-time award at $t+d$. These are proxies that can be agreed by human annotators and easily augmented by correlated signals, using which evaluation can be done through four tasks: classification, ranking, correlation and feature selection. We show these proxy evaluation methods complement each other regarding error handling, coverage, interpretability, and scope, and thus altogether contribute to the observation of the relative strength of existing models.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

counterfactualfeature selectionNovelty Detection

Similar Papers 제목 키워드 기반

NovAScore: A New Automated Metric for Evaluating Document Level Novelty

2024-09-14 · Lin Ai, Ziwei Gong, Harshsaiprasad Deshpande, Alexander Johnson 외

The rapid expansion of online content has intensified the issue of information redundancy, underscoring the need for solutions that can identify genuinely new information. Despite this challenge, the research community h…

Novelty Detection

Evaluating Novelty in AI-Generated Research Plans Using Multi-Workflow LLM Pipelines

2025-12-24 · Devesh Saraogi, Rohit Singhee, Dhruv Kumar arxiv

The integration of Large Language Models (LLMs) into the scientific ecosystem raises fundamental questions about the creativity and originality of AI-generated research. Recent work has identified ``smart plagiarism'' as…

OpenNovelty: An LLM-powered Agentic System for Verifiable Scholarly Novelty Assessment

2026-01-04 · Ming Zhang, Kexin Tan, Yueyuan Huang, Yujiong Shen 외 arxiv

Evaluating novelty is critical yet challenging in peer review, as reviewers must assess submissions against a vast, rapidly evolving literature. This report presents OpenNovelty, an LLM-powered agentic system for transpa…

Continual Novelty Detection

2021-06-24 · Rahaf Aljundi, Daniel Olmeda Reino, Nikolay Chumerin, Richard E. Turner

Novelty Detection methods identify samples that are not representative of a model's training set thereby flagging misleading predictions and bringing a greater flexibility and transparency at deployment time. However, re…

Continual LearningNovelty Detection

An Improved System for Sentence-level Novelty Detection in Textual Streams

2016-04-30 · Xinyu Fu, Eugene Ch'ng, Uwe Aickelin, Lanyun Zhang

Novelty detection in news events has long been a difficult problem. A number of models performed well on specific data streams but certain issues are far from being solved, particularly in large data streams from the WWW…

Event DetectionNovelty DetectionSentence