paper-with-me

홈 › Papers

Correlated Errors in Large Language Models

2025-06-09 · Elliot Kim, Avi Garg, Kenny Peng, Nikhil Garg

Diversity in training data, architecture, and providers is assumed to mitigate homogeneity in LLMs. However, we lack empirical evidence on whether different LLMs differ meaningfully. We conduct a large-scale empirical evaluation on over 350 LLMs overall, using two popular leaderboards and a resume-screening task. We find substantial correlation in model errors -- on one leaderboard dataset, models agree 60% of the time when both models err. We identify factors driving model correlation, including shared architectures and providers. Crucially, however, larger and more accurate models have highly correlated errors, even with distinct architectures and providers. Finally, we show the effects of correlation in two downstream tasks: LLM-as-judge evaluation and hiring -- the latter reflecting theoretical predictions regarding algorithmic monoculture.

📄 PDF Abstract BibTeX arXiv:2506.07962

Code (0)

등록된 구현이 없습니다.

Tasks

Diversity

Similar Papers 제목 키워드 기반

Hidden Clones: Exposing and Fixing Family Bias in Vision-Language Model Ensembles

2026-03-17 · Zacharie Bugaud arxiv

Ensembling Vision-Language Models (VLMs) from different providers maximizes benchmark accuracy, yet models from the same architectural family share correlated errors that standard voting ignores. We study this structure …

The Oracle's Fingerprint: Correlated AI Forecasting Errors and the Limits of Bias Transmission

2026-04-07 · Theodor Spiro arxiv

When large language models (LLMs) are consulted as forecasting tools, the independence of individual errors -- the foundation of collective intelligence -- may collapse. We test three conditions necessary for this "epist…

Adjusting for Autocorrelated Errors in Neural Networks for Time Series

2021-01-28 · NeurIPS 2021 12 · Fan-Keng Sun, Christopher I. Lang, Duane S. Boning

An increasing body of research focuses on using neural networks to model time series. A common assumption in training neural networks via maximum likelihood estimation on time series is that the errors across time steps …

Time SeriesTime Series AnalysisTime Series ForecastingTime Series Regression

Revisiting Split Covariance Intersection: Correlated Components and Optimality

2025-01-14 · Colin Cros, Pierre-Olivier Amblard, Christophe Prieur, Jean-François da Rocha

Linear fusion is a cornerstone of estimation theory. Implementing optimal linear fusion requires knowledge of the covariance of the vector of errors associated with all the estimators. In distributed or cooperative syste…

A Note on the Asymptotic Properties of the GLS Estimator in Multivariate Regression with Heteroskedastic and Autocorrelated Errors

2025-03-18 · Koichiro Moriya, Akihiko Noda

We study the asymptotic properties of the GLS estimator in multivariate regression with heteroskedastic and autocorrelated errors. We derive Wald statistics for linear restrictions and assess their performance. The stati…

regression