paper-with-me

홈 › Papers

When Auditors Fabricate: Batch-Size Degradation and Confident Hallucination in LLM Detection of Planted Document Contamination

2026-09-09 · Karan Parekh, Sanjana Pendyala Ravinder, Sana Mhapsekar, Medina Maloku arxiv

Large language models are increasingly proposed as automated auditors of document quality, yet their reliability as detectors of planted errors is poorly characterised. We construct a contaminated corpus of 150 academic papers spanning supply chain management and medical research, injecting 450 known contaminants of three types: typographical corruption, semantic reversal, and absurd out-of-context insertion. We then evaluate Google Gemini 3.0 Pro's ability to recover a 180-contaminant answer-key subset across 60 documents under three prompting regimes of increasing scale: single document, small batch, and large batch. Detection holds at small scale and then collapses: 50% recovery on single documents, 60% on small batches, and 2.8% on large batches. The failure mode at scale is not abstention but fabrication. Rather than reporting incomplete processing, the model produced confident findings including invented contaminants of its own, absurdities such as "telepathic squirrel" and "quantum-powered toaster" that mimic the style of the planted material but do not appear in any document. Detection also varies by contamination type: absurd insertions were recovered at 75% in completed evaluations, while semantic reversals and typographical corruptions were each recovered at only 50%. The corruptions most likely to occur in the wild, plausible ones, are the ones most often missed. We conclude that LLM document auditing degrades not gracefully but deceptively, and outline the harness such systems require: bounded batch sizes, direct content injection, and mechanical verification of every reported finding against source text.

📄 PDF Abstract BibTeX arXiv:2609.09696

Code (3)

Tavish9/awesome-daily-AI-arxiv ★ 115
arxivsub/arXivSub_daily_arxiv ★ 4
exopoiesis/arxiv-radar-chemistry ★ 1

Similar Papers 제목 키워드 기반

Sequential Auditing for f-Differential Privacy

2026-02-06 · Tim Kutta, Martin Dunsche, Yu Wei, Vassilis Zikas arxiv

We present new auditors to assess Differential Privacy (DP) of an algorithm based on output samples. Such empirical auditors are common to check for algorithmic correctness and implementation bugs. Most existing auditors…

Impact of Batch Size on Stopping Active Learning for Text Classification

2018-01-24 · Garrett Beatty, Ethan Kochis, Michael Bloodgood

When using active learning, smaller batch sizes are typically more efficient from a learning efficiency perspective. However, in practice due to speed and human annotator considerations, the use of larger batch sizes is …

Active LearningGeneral ClassificationOpen-Ended Question Answeringtext-classification+1

A New Look at Ghost Normalization

2020-07-16 · Neofytos Dimitriou, Ognjen Arandjelovic

Batch normalization (BatchNorm) is an effective yet poorly understood technique for neural network optimization. It is often assumed that the degradation in BatchNorm performance to smaller batch sizes stems from it havi…

A Batch Sequential Halving Algorithm without Performance Degradation

2024-06-01 · Sotetsu Koyamada, Soichiro Nishimori, Shin Ishii

In this paper, we investigate the problem of pure exploration in the context of multi-armed bandits, with a specific focus on scenarios where arms are pulled in fixed-size batches. Batching has been shown to enhance comp…

Computational EfficiencyMulti-Armed Bandits

CUAAudit: Meta-Evaluation of Vision-Language Models as Auditors of Autonomous Computer-Use Agents

2026-03-11 · Marta Sumyk, Oleksandr Kosovan arxiv

Computer-Use Agents (CUAs) are emerging as a new paradigm in human-computer interaction, enabling autonomous execution of tasks in desktop environment by perceiving high-level natural-language instructions. As such agent…