paper-with-me

홈 › Papers

Most Swiss-system tournaments are unfair: Evidence from chess

2024-10-25 · László Csató

Swiss-system is an increasingly popular tournament format as it provides an attractive trade-off between the number of matches and ranking accuracy. However, few empirical research consider the optimal design of the Swiss-system. We contribute to this issue by investigating the fairness of Swiss-system chess competitions with an odd number of rounds, where half of the players have an extra game with white pieces. They are proven to enjoy a significant advantage and to be overrepresented among both the highest-ranked and outperforming players. Therefore, Swiss-system tournaments should have an even number of rounds and use a pairing mechanism that guarantees a balanced colour assignment.

📄 PDF Abstract BibTeX arXiv:2410.19333

Code (0)

등록된 구현이 없습니다.

Tasks

Fairness

Similar Papers 제목 키워드 기반

Investigating Non-Transitivity in LLM-as-a-Judge

2025-02-19 · Yi Xu, Laura Ruis, Tim Rocktäschel, Robert Kirk

Automatic evaluation methods based on large language models (LLMs) are emerging as the standard tool for assessing the instruction-following abilities of LLM-based agents. The most common method in this paradigm, pairwis…

ChatbotComputational EfficiencyInstruction Following

Revealing Unfair Models by Mining Interpretable Evidence

2022-07-12 · Mohit Bajaj, Lingyang Chu, Vittorio Romaniello, Gursimran Singh 외

The popularity of machine learning has increased the risk of unfair models getting deployed in high-stake applications, such as justice system, drug/vaccination design, and medical diagnosis. Although there are effective…

BIG-bench Machine LearningMedical Diagnosis

PsychePass: Calibrating LLM Therapeutic Competence via Trajectory-Anchored Tournaments

2026-01-28 · Zhuang Chen, Dazhen Wan, Zhangkai Zheng, Guanqun Bi 외 arxiv

While large language models show promise in mental healthcare, evaluating their therapeutic competence remains challenging due to the unstructured and longitudinal nature of counseling. We argue that current evaluation p…

Reinforcement Learning

Swiss-Bench 003: Evaluating LLM Reliability and Adversarial Security for Swiss Regulatory Contexts

2026-04-07 · Fatih Uenal arxiv

The deployment of large language models (LLMs) in Swiss financial and regulatory contexts demands empirical evidence of both production reliability and adversarial security, dimensions not jointly operationalized in exis…

Selection on moral hazard in the Swiss market for mandatory health insurance: Empirical evidence from Swiss Household Panel data

2022-08-07 · Francetic Igor

Selection on moral hazard represents the tendency to select a specific health insurance coverage depending on the heterogeneity in utilisation ''slopes''. I use data from the Swiss Household Panel and from publicly avail…