paper-with-me

Papers

Stop Automating Peer Review Without Rigorous Evaluation

2026-05-04 · Joachim Baumann, Jiaxin Pei, Sanmi Koyejo, Dirk Hovy arxiv

Large language models offer a tempting solution to address the peer review crisis. This position paper argues that today's AI systems should not be used to produce paper reviews. We ground this position in an empirical comparison of human- versus AI-generated ICLR 2026 reviews and an evaluation of the effect of automated paper rewriting on different AI reviewers. We identify two critical issues: 1) AI reviewers exhibit a hivemind effect of excessive agreement within and across papers that reduces perspective diversity. 2) AI review scores are trivially gameable through paper laundering: prompting an LLM to rewrite a paper could significantly increase the scores from AI reviewers, demonstrating that LLM reviewers are easy to game through stylistic changes rather than scientific results. However, non-gameability and review diversity are necessary but not sufficient conditions for automation. We argue that addressing the peer review crisis requires a science of peer review automation -- not general-purpose LLMs deployed without rigorous evaluation.

📄 PDF Abstract BibTeX arXiv:2605.03202

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Pre-review to Peer review: Pitfalls of Automating Reviews using Large Language Models

2025-12-14 · Akhil Pandey Akella, Harish Varma Siravuri, Shaurya Rohatgi arxiv

Large Language Models are versatile general-task solvers, and their capabilities can truly assist people with scholarly peer review as \textit{pre-review} agents, if not as fully autonomous \textit{peer-review} agents. W…

ReviewGuard: Aligning LLM-Assisted Peer Review with Long-Term Scientific Impact

2026-05-29 · Abdur Rasool, Xiaohui Huang, Yanqing Hu, Linyi Yang arxiv

Peer review is central to scientific quality control, yet it can undervalue papers that later achieve substantial citation impact. While frontier large language models have shown promise in automating aspects of peer rev…

Reinforcement Learning

ScholarPeer: A Context-Aware Multi-Agent Framework for Automated Peer Review

2026-01-30 · Palash Goyal, Mihir Parmar, Yiwen Song, Hamid Palangi 외 arxiv

The exponential growth of machine learning submissions has strained the traditional peer review process, resulting in slow feedback loops for authors and an immense burden on reviewers to rigorously audit technical sound…

From Passive Generation to Investigation: A Proactive Scientific Peer Review Agent

2026-06-11 · Haishuo Fang, Yue Feng, Iryna Gurevych arxiv

Large language models (LLMs) have shown promise in automating scientific peer review. However, existing approaches often struggle to generate in-depth reviews supported by concrete evidence. We argue that a key limitatio…

Reinforcement Learning

AstroReview: An LLM-driven Multi-Agent Framework for Telescope Proposal Peer Review and Refinement

2025-12-31 · Yutong Wang, Yunxiang Xiao, Yonglin Tian, Junyong Li 외 arxiv

Competitive access to modern observatories has intensified as proposal volumes outpace available telescope time, making timely, consistent, and transparent peer review a critical bottleneck for the advancement of astrono…