paper-with-me

Papers

AI-Assisted Peer Review at Scale: The AAAI-26 AI Review Pilot

2026-04-15 · Joydeep Biswas, Sheila Schoepp, Gautham Vasan, Anthony Opipari, Arthur Zhang, Zichao Hu, Sebastian Joseph, Matthew Lease, Junyi Jessy Li, Peter Stone, Kiri L. Wagstaff, Matthew E. Taylor, Odest Chadwicke Jenkins arxiv

Scientific peer review faces mounting strain as submission volumes surge, making it increasingly difficult to sustain review quality, consistency, and timeliness. Recent advances in AI have led the community to consider its use in peer review, yet a key unresolved question is whether AI can generate technically sound reviews at real-world conference scale. Here we report the first large-scale field deployment of AI-assisted peer review: every main-track submission at AAAI-26 received one clearly identified AI review from a state-of-the-art system. The system combined frontier models, tool use, and safeguards in a multi-stage process to generate reviews for all 22,977 full-review papers in less than a day. A large-scale survey of AAAI-26 authors and program committee members showed that participants not only found AI reviews useful, but actually preferred them to human reviews on key dimensions such as technical accuracy and research suggestions. We also introduce a novel benchmark and find that our system substantially outperforms a simple LLM-generated review baseline at detecting a variety of scientific weaknesses. Together, these results show that state-of-the-art AI methods can already make meaningful contributions to scientific peer review at conference scale, opening a path toward the next generation of synergistic human-AI teaming for evaluating research.

📄 PDF Abstract BibTeX arXiv:2604.13940

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Matching Papers and Reviewers at Large Conferences

2022-02-24 · Kevin Leyton-Brown, Mausam, Yatin Nandwani, Hedayat Zarkoob 외

Peer-reviewed conferences, the main publication venues in CS, rely critically on matching highly qualified reviewers for each paper. Because of the growing scale of these conferences, the tight timelines on which they op…

Do LLMs Favor LLMs? Quantifying Interaction Effects in Peer Review

2026-01-28 · Vibhhu Sharma, Thorsten Joachims, Sarah Dean arxiv

There are increasing indications that LLMs are not only used for producing scientific papers, but also as part of the peer review process. In this work, we provide the first comprehensive analysis of LLM use across the p…

Position on LLM-Assisted Peer Review: Addressing Reviewer Gap through Mentoring and Feedback

2026-01-14 · JungMin Yun, JuneHyoung Kwon, MiHyeon Kim, YoungBin Kim arxiv

The rapid expansion of AI research has intensified the Reviewer Gap, threatening the peer-review sustainability and perpetuating a cycle of low-quality evaluations. This position paper critiques existing LLM approaches t…

The AI Imperative: Scaling High-Quality Peer Review in Machine Learning

2025-06-09 · Qiyao Wei, Samuel Holt, Jing Yang, Markus Wulfmeier 외

Peer review, the bedrock of scientific advancement in machine learning (ML), is strained by a crisis of scale. Exponential growth in manuscript submissions to premier ML venues such as NeurIPS, ICML, and ICLR is outpacin…

Group Fairness in Peer Review

2024-10-04 · NeurIPS 2023 11 · Haris Aziz, Evi Micha, Nisarg Shah

Large conferences such as NeurIPS and AAAI serve as crossroads of various AI fields, since they attract submissions from a vast number of communities. However, in some cases, this has resulted in a poor reviewing experie…

Fairness