paper-with-me

Papers

On Finding Inconsistencies in Documents

2025-12-21 · Charles J. Lovering, Seth Ebner, Brandon Smock, Michael Krumdick, Saad Rabbani, Ahmed Muhammad, Varshini Reddy, Chris Tanner arxiv

Professionals in academia, law, and finance audit their documents because inconsistencies can result in monetary, reputational, and scientific costs. Language models (LMs) have the potential to dramatically speed up this auditing process. To understand their abilities, we introduce a benchmark, FIND (Finding INconsistencies in Documents), where each example is a document with an inconsistency inserted manually by a domain expert. Despite the documents being long, technical, and complex, the best-performing model (gpt-5) recovered 64% of the inserted inconsistencies. Surprisingly, gpt-5 also found undiscovered inconsistencies present in the original documents. For example, on 50 arXiv papers, we judged 136 out of 196 of the model's suggestions to be legitimate inconsistencies missed by the original authors. However, despite these findings, even the best models miss almost half of the inconsistencies in FIND, demonstrating that inconsistency detection is still a challenging task.

📄 PDF Abstract BibTeX arXiv:2512.18601

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

NLP-based Regulatory Compliance -- Using GPT 4.0 to Decode Regulatory Documents

2024-12-29 · Bimal Kumar, Dmitri Roussinov

Large Language Models (LLMs) such as GPT-4.0 have shown significant promise in addressing the semantic complexities of regulatory documents, particularly in detecting inconsistencies and contradictions. This study evalua…

Contradictions in Context: Challenges for Retrieval-Augmented Generation in Healthcare

2025-11-10 · Saeedeh Javadi, Sara Mirabi, Manan Gangar, Bahadorreza Ofoghi arxiv

In high-stakes information domains such as healthcare, where large language models (LLMs) can produce hallucinations or misinformation, retrieval-augmented generation (RAG) has been proposed as a mitigation strategy, gro…

Mind the Gap: Analyzing Lacunae with Transformer-Based Transcription

2024-06-28 · Jaydeep Borkar, David A. Smith

Historical documents frequently suffer from damage and inconsistencies, including missing or illegible text resulting from issues such as holes, ink problems, and storage damage. These missing portions or gaps are referr…

Optical Character RecognitionOptical Character Recognition (OCR)

Directed Criteria Citation Recommendation and Ranking Through Link Prediction

2024-03-18 · William Watson, Lawrence Yong

We explore link prediction as a proxy for automatically surfacing documents from existing literature that might be topically or contextually relevant to a new document. Our model uses transformer-based graph embeddings t…

Citation RecommendationLink PredictionPrediction

Enhancing Cohesion and Coherence of Fake Text to Improve Believability for Deceiving Cyber Attackers

2018-08-01 · COLING 2018 8 · Prakruthi Karuna, Hemant Purohit, {\"O}zlem Uzuner, Sushil Jajodia 외

Ever increasing ransomware attacks and thefts of intellectual property demand cybersecurity solutions to protect critical documents. One emerging solution is to place fake text documents in the repository of critical doc…

Intrusion DetectionText Generation