paper-with-me

홈 › Papers

When Models Disagree: Rethinking LLM Evaluation for Public Comment Analysis

2026-05-27 · Aisha Najera, Alvin Moon, Vedant Srinivasan, Rajesh Veeraraghavan arxiv

Federal agencies are deploying large language models (LLMs) to categorize public comment corpora, where the model's organization of the record shapes what policymakers see and which arguments register. Standard evaluation, anchored on stance accuracy against a small validated set, cannot detect when different models produce materially different categorizations of the same public input. We propose an Interpretive Audit Pipeline that treats multi-model disagreement as diagnostic of interpretive complexity and directs human review toward genuinely ambiguous public input. Analyzing 1,260 public comments on a federal USDA docket across four LLMs, we find that inter-model thematic divergence exceeds within-model prompt variation, and that an expert rubric suppresses deep interpretive disagreement without resolving it. In a two-stage labeling study on a stratified 40-comment subsample, four LLMs and a human annotator labeled independently and then revised after seeing the others' labels. Revision behavior varied across labelers, and the human annotator's revisions frequently introduced framings absent from the ensemble's collective output. We argue disagreement-based evaluation is a necessary complement to accuracy metrics for LLM-assisted interpretive coding.

📄 PDF Abstract BibTeX arXiv:2605.29025

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

When Reviewers Lock Horn: Finding Disagreement in Scientific Peer Reviews

2023-10-28 · Sandeep Kumar, Tirthankar Ghosal, Asif Ekbal

To this date, the efficacy of the scientific publishing enterprise fundamentally rests on the strength of the peer review process. The journal editor or the conference chair primarily relies on the expert reviewers' asse…

Decoding Climate Disagreement: A Graph Neural Network-Based Approach to Understanding Social Media Dynamics

2024-07-09 · Ruiran Su, Janet B. Pierrehumbert

This work introduces the ClimateSent-GAT Model, an innovative method that integrates Graph Attention Networks (GATs) with techniques from natural language processing to accurately identify and predict disagreements withi…

Graph AttentionGraph Neural Network

Darshana Graph: A Parallel Commentary Corpus for Comparative Indian Philosophy, with Stylometric and Exploratory Graph Analyses

2026-06-16 · Joy Bose arxiv

We introduce Darshana Graph, a corpus of over 125,000 text records spanning classical Hindu, Buddhist, and Jain philosophical traditions, drawn from public-domain and openly licensed translations of sources including the…

A Collaborative Content Moderation Framework for Toxicity Detection based on Conformalized Estimates of Annotation Disagreement

2024-11-06 · Guillermo Villate-Castillo, Javier Del Ser, Borja Sanz

Content moderation typically combines the efforts of human moderators and machine learning models. However, these systems often rely on data where significant disagreement occurs during moderation, reflecting the subject…

Conformal Prediction

Rethinking Ground Truth: A Case Study on Human Label Variation in MLLM Benchmarking

2026-03-20 · Tomas Ruiz, Tanalp Agustoslu, Carsten Schwemmer arxiv

Human Label Variation (HLV), i.e. systematic differences among annotators' judgments, remains underexplored in benchmarks despite rapid progress in large language model (LLM) development. We address this gap by introduci…