paper-with-me

Papers

NovAScore: A New Automated Metric for Evaluating Document Level Novelty

2024-09-14 · Lin Ai, Ziwei Gong, Harshsaiprasad Deshpande, Alexander Johnson, Emmy Phung, Ahmad Emami, Julia Hirschberg

The rapid expansion of online content has intensified the issue of information redundancy, underscoring the need for solutions that can identify genuinely new information. Despite this challenge, the research community has seen a decline in focus on novelty detection, particularly with the rise of large language models (LLMs). Additionally, previous approaches have relied heavily on human annotation, which is time-consuming, costly, and particularly challenging when annotators must compare a target document against a vast number of historical documents. In this work, we introduce NovAScore (Novelty Evaluation in Atomicity Score), an automated metric for evaluating document-level novelty. NovAScore aggregates the novelty and salience scores of atomic information, providing high interpretability and a detailed analysis of a document's novelty. With its dynamic weight adjustment scheme, NovAScore offers enhanced flexibility and an additional dimension to assess both the novelty level and the importance of information within a document. Our experiments show that NovAScore strongly correlates with human judgments of novelty, achieving a 0.626 Point-Biserial correlation on the TAP-DLND 1.0 dataset and a 0.920 Pearson correlation on an internal human-annotated dataset.

📄 PDF Abstract BibTeX arXiv:2409.09249

Code (0)

등록된 구현이 없습니다.

Tasks

Novelty Detection

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

Automated Metrics for Medical Multi-Document Summarization Disagree with Human Evaluations

2023-05-23 · Lucy Lu Wang, Yulia Otmakhova, Jay DeYoung, Thinh Hung Truong 외

Evaluating multi-document summarization (MDS) quality is difficult. This is especially true in the case of MDS for biomedical literature reviews, where models must synthesize contradicting evidence reported across differ…

Document SummarizationMulti-Document Summarization

Pointwise Mutual Information Based Metric and Decoding Strategy for Faithful Generation in Document Grounded Dialogs

2023-05-20 · Yatin Nandwani, Vineet Kumar, Dinesh Raghu, Sachindra Joshi 외

A major concern in using deep learning based generative models for document-grounded dialogs is the potential generation of responses that are not \textit{faithful} to the underlying document. Existing automated metrics …

Response Generation

Towards Automated Document Revision: Grammatical Error Correction, Fluency Edits, and Beyond

2022-05-23 · Masato Mita, Keisuke Sakaguchi, Masato Hagiwara, Tomoya Mizumoto 외

Natural language processing technology has rapidly improved automated grammatical error correction tasks, and the community begins to explore document-level revision as one of the next challenges. To go beyond sentence-l…

Grammatical Error CorrectionLanguage ModellingSentence

A Quantitative Method for Shoulder Presentation Evaluation in Biometric Identity Documents

2025-11-18 · Alfonso Pedro Ridao arxiv

International standards for biometric identity documents mandate strict compliance with pose requirements, including the square presentation of a subject's shoulders. However, the literature on automated quality assessme…

Pose Estimation

VDE Bench: Evaluating The Capability of Image Editing Models to Modify Visual Documents

2026-01-27 · Hongzhu Yi, Yujia Yang, Yuanxiang Wang, Tong Li 외 arxiv

In recent years, image editing models have made significant progress, enabling users to manipulate visual content in a flexible and interactive manner through natural language instructions. However, an important yet unde…

Image Editing