paper-with-me

Papers

Cross-Audit Projection for Model Risk Prediction

2026-07-02 · Yijian Huang arxiv

For training-data-based model risk prediction, $K$-fold cross-validation~(CV) is widely used to mitigate the well-known over-optimism of the empirical risk and is often regarded as reliable. However, for binary classification via empirical risk minimization, our numerical studies reveal a surprising phenomenon: $K$-fold CV may perform poorly in estimating class-specific risks, even worse than the empirical estimator. We perform a higher-order asymptotic analysis showing that $K$-fold CV may converge at a slower rate, whereas the empirical estimator exhibits a second-order asymptotic bias that explains its over-optimism. These findings motivate a novel two-step procedure for model risk prediction, termed cross-audit projection (CAP). The cross-audit step adopts the same resampling scheme as $K$-fold CV to estimate over-optimism in subsamples, while the asymptotic-theory-informed projection step adjusts for the reduced sample size in bias correction of the empirical risk. The resulting CAP estimator is first-order asymptotically equivalent to the empirical risk while achieving second-order asymptotic unbiasedness. An accompanying inference procedure is also developed. Simulation studies support theoretical advantages of CAP and demonstrate favorable finite-sample performance. An application to breast cancer detection further illustrates the proposed method.

📄 PDF Abstract BibTeX arXiv:2607.02328

Code (0)

등록된 구현이 없습니다.

Tasks

Breast Cancer DetectionBinary Classification

Similar Papers 제목 키워드 기반

Exposing the Illusion of Fairness: Auditing Vulnerabilities to Distributional Manipulation Attacks

2025-07-28 · Valentin Lafargue, Adriana Laurindo Monteiro, Emmanuelle Claeys, Laurent Risser 외 arxiv

The rapid deployment of AI systems in high-stakes domains, including those classified as high-risk under the The EU AI Act (Regulation (EU) 2024/1689), has intensified the need for reliable compliance auditing. For binar…

Bias Detection

Interpretable Fine-Gray Deep Survival Model for Competing Risks: Predicting Post-Discharge Foot Complications for Diabetic Patients in Ontario

2025-11-16 · Dhanesh Ramachandram, Anne Loefler, Surain Roberts, Amol Verma 외 arxiv

Model interpretability is crucial for establishing AI safety and clinician trust in medical applications for example, in survival modelling with competing risks. Recent deep learning models have attained very good predic…

Feature Importance

TRACES: Proactive Safety Auditing for Multi-Turn LLM Agents via Trajectory-State Modeling

2026-05-26 · Jiaqian Li, Yanshu Li, Boxuan Zhang, Ruixiang Tang 외 arxiv

LLM agents increasingly operate through multi-turn tool use and environment interaction, where safety risks often emerge from intermediate steps long before they surface in the final outcome. Reactive auditing is therefo…

Fairness Audits of Institutional Risk Models in Deployed ML Pipelines

2026-04-21 · Kelly McConvey, Dipto Das, Maya Ghai, Angelina Zhai 외 arxiv

Fairness audits of institutional risk models are critical for understanding how deployed machine learning pipelines allocate resources. Drawing on multi-year collaboration with Centennial College, where our prior ethnogr…

Machine Learning based Enterprise Financial Audit Framework and High Risk Identification

2025-07-08 · Tingyu Yuan, Xi Zhang, Xuanjing Chen arxiv

In the face of global economic uncertainty, financial auditing has become essential for regulatory compliance and risk mitigation. Traditional manual auditing methods are increasingly limited by large data volumes, compl…

Feature ImportanceFraud Detection