paper-with-me

홈 › Papers

The Payment Heterogeneity Index: An Integrated Unsupervised Framework for High-Volume Procurement Oversight and Decision Support

2026-05-09 · Kyriakos Christodoulides arxiv

Public procurement is vulnerable to error, fraud, and corruption, particularly as high transaction volumes overwhelm oversight. While research often focuses on tender-stage anomalies, post-award payment monitoring remains underexplored. Since labelled datasets are rare and methods like Benford's Law face restrictive assumptions, there is a need for interpretable, unsupervised frameworks for high-volume procurement oversight and decision support. This paper introduces the Structural Heterogeneity Index (SHI), a composite statistic for one-dimensional samples, and its payment-specific instantiation, the Payment Heterogeneity Index (PHI), characterising payment structure and latent regimes. It incorporates Gaussian Mixture Model (GMM) parameters alongside non-parametric statistics, integrating four interpretable components: modality, asymmetry, tail behaviour, and structural dispersion. Uniquely, the tail-behaviour component captures both distributional heaviness and extreme-value concentration, while structural-dispersion combines the variability, prevalence, and separation of latent payment regimes. Applied to UK municipal procurement data, PHI identifies a financially significant cohort (0.6\% of suppliers; 10.1\% of high-volume vendors) with structurally distinct payment patterns. Statistical testing further supports these differences, and targeted human verification confirms the plausibility of prioritised cases. Comparative analysis shows PHI reveals regime separation obscured by the Coefficient of Variation ($ρ= 0.310$). PHI provides a transparent, decomposable, and computationally lightweight framework for procurement integrity oversight and targeted audit prioritisation.

📄 PDF Abstract BibTeX arXiv:2605.12547

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

DiviK: Divisive intelligent K-Means for hands-free unsupervised clustering in big biological data

2020-09-22 · Grzegorz Mrukwa, Joanna Polanska

Investigating molecular heterogeneity provides insights about tumor origin and metabolomics. The increasing amount of data gathered makes manual analyses infeasible - therefore, automated unsupervised learning approaches…

ClusteringFeature Engineering

Incentivizing Truthful Collaboration in Heterogeneous Federated Learning

2024-12-01 · Dimitar Chakarov, Nikita Tsoy, Kristian Minchev, Nikola Konstantinov

Federated learning (FL) is a distributed collaborative learning method, where multiple clients learn together by sharing gradient updates instead of raw data. However, it is well-known that FL is vulnerable to manipulate…

Federated Learning

Impact of social factors on loan delinquency in microfinance

2024-10-17 · Cedric H. A. Koffi, Viani Biatat Djeundje, Olivier Menoukeu Pamen

This paper develops multistate models to analyse loan delinquency in the microfinance sector, using data from Ghana. The models are designed to account for both partial repayments and the short repayment durations typica…

Healthcare cost prediction for heterogeneous patient profiles using deep learning models with administrative claims data

2025-02-17 · Mohammad Amin Morid, Olivia R. Liu Sheng

Problem: How can we design patient cost prediction models that effectively address the challenges of heterogeneity in administrative claims (AC) data to ensure accurate, fair, and generalizable predictions, especially fo…

FairnessPredictionRepresentation Learning

Capability Gates Are Not Authorization: Confused-Deputy Failures in LLM Agent Frameworks

2026-06-27 · David Mellafe Zuvic arxiv

Tool-using LLM agents increasingly read untrusted content while holding side-effecting tools such as payments, email, CRM, and infrastructure APIs, yet common framework defaults still conflate tool exposure with authoriz…