paper-with-me

홈 › Papers

Towards AI epidemiology: a measurement standardisation framework for prospective risk detection

2025-12-15 · Kit Tempest-Walters arxiv

This paper proposes a measurement standardisation framework that compresses expert-AI interactions into structured, comparable fields for prospective risk detection in deployed AI systems, without access to model internals. The main aim of this concept paper is to define the scope of the framework, both semantically and statistically, and to specify a protocol for its empirical testing in future work. The population-level claims the framework is designed to support are therefore the subject of a staged research programme rather than results claimed in this paper. Measurement standardisation underpins all three claims that follow. The first is a reliability claim: under bounded conditions, large language models can produce reliable, standardised assessments of the evidential and policy alignment of expert-AI interactions. The second is a governance claim: alignment scores give experts an immediate signal during deployment and give institutions a basis for monitoring alignment patterns across mission types, models, and domains. The third is an epidemiological claim: once measurement standardisation is established, aggregate alignment scores could be used to study associations with downstream outcomes in regulated professional settings. This introduces the possibility of an "AI epidemiology" that detects risk based on correlated variables instead of mechanistic analysis. This paper addresses the first claim and specifies protocols for investigating the second and third. To enable empirical evaluation in future studies, this paper sets out a defined grammar, together with a statistical protocol based on paired bootstrap inference, DeLong's test for paired AUCs as a sensitivity check, a pre-specified one-sided non-inferiority margin of 0.05, and Holm-Bonferroni correction.

📄 PDF Abstract BibTeX arXiv:2512.15783

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Vertical Standardisation for High-Risk AI Systems under the EU AI Act: A Domain-Specific Framework for Algorithmic Hiring

2026-07-14 · Anna Gatzioura, Vrettos Moulos, Nina Baranowska arxiv

According to the recent European legislation, high-risk AI systems will have to adapt in order to comply with requirements related to specific areas, like risk management, data quality and governance, logging and traceab…

Considering Fundamental Rights in the European Standardisation of Artificial Intelligence: Nonsense or Strategic Alliance?

2024-01-23 · Marion Ho-Dac

In the European context, both the EU AI Act proposal and the draft Standardisation Request on safe and trustworthy AI link standardisation to fundamental rights. However, these texts do not provide any guidelines that sp…

valid

Prospective Learning: Learning for a Dynamic Future

2024-10-31 · Ashwin De Silva, Rahul Ramesh, Rubing Yang, Siyu Yu 외

In real-world applications, the distribution of the data, and our goals, evolve over time. The prevailing theoretical framework for studying machine learning, namely probably approximately correct (PAC) learning, largely…

PAC learning

Measurement Error in Nutritional Epidemiology: A Survey

2020-04-14 · Huimin Peng

This article reviews bias-correction models for measurement error of exposure variables in the field of nutritional epidemiology. Measurement error usually attenuates estimated slope towards zero. Due to the influence of…

EpidemiologyregressionSurvey

Prospective Dynamic 3D MRI Reconstruction via Latent-Space Motion Tracking from Single Measurement

2026-06-02 · Lixuan Chen, Zhongnan Liu, Jesse Hamilton, James M. Balter 외 arxiv

Prospective reconstruction is crucial in many clinical applications such as MRI-guided radiotherapy, which demands accurate image reconstruction and fast motion estimation from currently acquired measurements. However, p…

Image ReconstructionMRI Reconstruction