paper-with-me

홈 › Papers

Mapping AI Benchmark Data to Quantitative Risk Estimates Through Expert Elicitation

2025-03-06 · Malcolm Murray, Henry Papadatos, Otter Quarks, Pierre-François Gimenez, Simeon Campos

The literature and multiple experts point to many potential risks from large language models (LLMs), but there are still very few direct measurements of the actual harms posed. AI risk assessment has so far focused on measuring the models' capabilities, but the capabilities of models are only indicators of risk, not measures of risk. Better modeling and quantification of AI risk scenarios can help bridge this disconnect and link the capabilities of LLMs to tangible real-world harm. This paper makes an early contribution to this field by demonstrating how existing AI benchmarks can be used to facilitate the creation of risk estimates. We describe the results of a pilot study in which experts use information from Cybench, an AI benchmark, to generate probability estimates. We show that the methodology seems promising for this purpose, while noting improvements that can be made to further strengthen its application in quantitative AI risk assessment.

📄 PDF Abstract BibTeX arXiv:2503.04299

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Network Risk Estimation: A Risk Estimation Paradigm for Cyber Networks

2025-01-27 · Arda Bayer, David Maluf, Behnaam Aazhang

Cyber networks are fundamental to many organization's infrastructure, and the size of cyber networks is increasing rapidly. Risk measurement of the entities/endpoints that make up the network via available knowledge abou…

Uncertainty-Aware Mapping from 3D Keypoints to Anatomical Landmarks for Markerless Biomechanics

2026-03-27 · Cesare Davide Pace, Alessandro Marco De Nunzio, Claudio De Stefano, Francesco Fontanella 외 arxiv

Markerless biomechanics increasingly relies on 3D skeletal keypoints extracted from video, yet downstream biomechanical mappings typically treat these estimates as deterministic, providing no principled mechanism for fra…

Outlier Detection

NLCG-Net: A Model-Based Zero-Shot Learning Framework for Undersampled Quantitative MRI Reconstruction

2024-01-22 · Xinrui Jiang, Yohan Jun, Jaejin Cho, Mengze Gao 외

Typical quantitative MRI (qMRI) methods estimate parameter maps after image reconstructing, which is prone to biases and error propagation. We propose a Nonlinear Conjugate Gradient (NLCG) optimizer for model-based T2/T1…

MRI ReconstructionQuantitative MRIZero-Shot Learning

Comparison and calibration of MP2RAGE quantitative T1 values to multi-TI inversion recovery T1 values

2024-09-20 · Adam M. Saunders, Michael E. Kim, Chenyu Gao, Lucas W. Remedios 외

While typical qualitative T1-weighted magnetic resonance images reflect scanner and protocol differences, quantitative T1 mapping aims to measure T1 independent of these effects. Changes in T1 in the brain reflect struct…

A physics-informed foundation model for quantitative diffusion MRI

2026-05-29 · Zihan Li, Jialan Zheng, Ziyu Li, Xun Yuan 외 arxiv

Understanding the human brain requires access to its microscopic tissue architecture. Diffusion magnetic resonance imaging (MRI) provides the only noninvasive window into whole-brain microstructure in vivo, yet reliable …