paper-with-me

홈 › Papers

LLM-as-a-Reviewer: Benchmarking Their Ability, Divergence, and Prompt Injection Resistance as Paper Reviewers

2026-05-25 · Lingyao Li, Junjie Xiong, Changjia Zhu, Runlong Yu, Chen Chen, Junyu Wang, Renkai Ma, Zhicong Lu arxiv

Large language models (LLMs) are increasingly used in academic peer review, yet their reliability, alignment with human judgment, and robustness to adversarial attacks remain poorly understood. We present a systematic benchmark of LLM-as-a-Reviewer on 898 papers stratified from NeurIPS and ICLR, evaluating 12 LLMs along three axes: rating calibration, divergence from human reviewers, and resistance to prompt injection embedded via an invisible font-mapping attack. We find that LLMs systematically overrate weaker submissions and diverge from humans in topical emphasis, under-flagging Clarity and over-flagging Reproducibility, while producing reviews two to three times longer with lower lexical diversity and a more standardized vocabulary. Prompt injection remains highly effective. Simple hidden instructions can promote low-scoring papers to acceptance-level ratings in a substantial fraction of cases, with effectiveness varying sharply across model families. While LLMs offer utility in structuring evaluations, their integration into peer review requires safeguards against both intrinsic biases and adversarial risks.

📄 PDF Abstract BibTeX arXiv:2605.25415

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

PRISM: A Multi-Dimensional Benchmark for Evaluating LLM Peer Reviewers

2026-05-26 · Ngoc Phan Phuoc Loc, Toan Huynh La Viet, Thanh Tran Khanh, Duy A Nguyen 외 arxiv

The rapid growth in submissions to machine learning venues has strained the scientific peer-review system and intensified interest in LLM-based automated peer reviewers. However, how good these systems are actually, espe…

Argument Mining

PRAIB: Peer Review AI Benchmark of Behaviour of LLM-Assisted Reviewing

2026-05-28 · Krzysztof Żurawicki, Julia Farganus, Arkadiusz Gaweł, Mateusz Bystroński 외 arxiv

The growing number of submitted papers has motivated the exploration of Large Language Models (LLMs) as a means to support and augment the peer review process, particularly in terms of improving its speed and scalability…

"Give a Positive Review Only": An Early Investigation Into In-Paper Prompt Injection Attacks and Defenses for AI Reviewers

2025-11-03 · Qin Zhou, Zhexin Zhang, Zhi Li, Limin Sun arxiv

With the rapid advancement of AI models, their deployment across diverse tasks has become increasingly widespread. A notable emerging application is leveraging AI models to assist in reviewing scientific papers. However,…

Evoke: Evoking Critical Thinking Abilities in LLMs via Reviewer-Author Prompt Editing

2023-10-20 · Xinyu Hu, Pengfei Tang, Simiao Zuo, Zihan Wang 외

Large language models (LLMs) have made impressive progress in natural language processing. These models rely on proper human instructions (or prompts) to generate suitable responses. However, the potential of LLMs are no…

Logical Fallacy Detection

Reviewer2: Optimizing Review Generation Through Prompt Generation

2024-02-16 · Zhaolin Gao, Kianté Brantley, Thorsten Joachims

Recent developments in LLMs offer new opportunities for assisting authors in improving their work. In this paper, we envision a use case where authors can receive LLM-generated reviews that uncover weak points in the cur…

Review Generation