paper-with-me

홈 › Papers

Integrating Rankings into Quantized Scores in Peer Review

2022-04-05 · Yusha Liu, Yichong Xu, Nihar B. Shah, Aarti Singh

In peer review, reviewers are usually asked to provide scores for the papers. The scores are then used by Area Chairs or Program Chairs in various ways in the decision-making process. The scores are usually elicited in a quantized form to accommodate the limited cognitive ability of humans to describe their opinions in numerical values. It has been found that the quantized scores suffer from a large number of ties, thereby leading to a significant loss of information. To mitigate this issue, conferences have started to ask reviewers to additionally provide a ranking of the papers they have reviewed. There are however two key challenges. First, there is no standard procedure for using this ranking information and Area Chairs may use it in different ways (including simply ignoring them), thereby leading to arbitrariness in the peer-review process. Second, there are no suitable interfaces for judicious use of this data nor methods to incorporate it in existing workflows, thereby leading to inefficiencies. We take a principled approach to integrate the ranking information into the scores. The output of our method is an updated score pertaining to each review that also incorporates the rankings. Our approach addresses the two aforementioned challenges by: (i) ensuring that rankings are incorporated into the updates scores in the same manner for all papers, thereby mitigating arbitrariness, and (ii) allowing to seamlessly use existing interfaces and workflows designed for scores. We empirically evaluate our method on synthetic datasets as well as on peer reviews from the ICLR 2017 conference, and find that it reduces the error by approximately 30% as compared to the best performing baseline on the ICLR 2017 data.

📄 PDF Abstract BibTeX arXiv:2204.03505

Code (1)

myusha/rankings_and_quantized_scores 공식 구현

Tasks

Decision Making

Similar Papers 제목 키워드 기반

The ICML 2023 Ranking Experiment: Examining Author Self-Assessment in ML/AI Peer Review

2024-08-24 · Buxin Su, Jiayao Zhang, Natalie Collina, Yuling Yan 외

We conducted an experiment during the review process of the 2023 International Conference on Machine Learning (ICML), asking authors with multiple submissions to rank their papers based on perceived quality. In total, we…

PeerRank: Autonomous LLM Evaluation Through Web-Grounded, Bias-Controlled Peer Review

2026-02-01 · Yanki Margalit, Erni Avram, Ran Taig, Oded Margalit 외 arxiv

Evaluating large language models typically relies on human-authored benchmarks, reference answers, and human or single-model judgments, approaches that scale poorly, become quickly outdated, and mismatch open-world deplo…

How to Find Fantastic AI Papers: Self-Rankings as a Powerful Predictor of Scientific Impact Beyond Peer Review

2025-10-02 · Buxin Su, Natalie Collina, Garrett Wen, Didong Li 외 arxiv

Peer review in academic research aims not only to ensure factual correctness but also to identify work of high scientific potential that can shape future research directions. This task is especially critical in fast-movi…

Isotonic Mechanism for Exponential Family Estimation in Machine Learning Peer Review

2023-04-21 · Yuling Yan, Weijie J. Su, Jianqing Fan

In 2023, the International Conference on Machine Learning (ICML) required authors with multiple submissions to rank their submissions based on perceived quality. In this paper, we aim to employ these author-specified ran…

Too Polite to Disagree: Understanding Sycophancy Propagation in Multi-Agent Systems

2026-04-03 · Vira Kasprova, Amruta Parulekar, Abdulrahman AlRabah, Krishna Agaram 외 arxiv

Large language models (LLMs) often exhibit sycophancy: agreement with user stance even when it conflicts with the model's opinion. While prior work has mostly studied this in single-agent settings, it remains underexplor…