paper-with-me

홈 › Papers

Using massive health insurance claims data to predict very high-cost claimants: a machine learning approach

2019-12-30 · José M. Maisog, Wenhong Li, Yanchun Xu, Brian Hurley, Hetal Shah, Ryan Lemberg, Tina Borden, Stephen Bandeian, Melissa Schline, Roxanna Cross, Alan Spiro, Russ Michael, Alexander Gutfraind

Due to escalating healthcare costs, accurately predicting which patients will incur high costs is an important task for payers and providers of healthcare. High-cost claimants (HiCCs) are patients who have annual costs above $\$250,000$ and who represent just 0.16% of the insured population but currently account for 9% of all healthcare costs. In this study, we aimed to develop a high-performance algorithm to predict HiCCs to inform a novel care management system. Using health insurance claims from 48 million people and augmented with census data, we applied machine learning to train binary classification models to calculate the personal risk of HiCC. To train the models, we developed a platform starting with 6,006 variables across all clinical and demographic dimensions and constructed over one hundred candidate models. The best model achieved an area under the receiver operating characteristic curve of 91.2%. The model exceeds the highest published performance (84%) and remains high for patients with no prior history of high-cost status (89%), who have less than a full year of enrollment (87%), or lack pharmacy claims data (88%). It attains an area under the precision-recall curve of 23.1%, and precision of 74% at a threshold of 0.99. A care management program enrolling 500 people with the highest HiCC risk is expected to treat 199 true HiCCs and generate a net savings of $\$7.3$ million per year. Our results demonstrate that high-performing predictive models can be constructed using claims data and publicly available data alone, even for rare high-cost claimants exceeding $\$250,000$. Our model demonstrates the transformational power of machine learning and artificial intelligence in care management, which would allow healthcare payers and providers to introduce the next generation of care management programs.

📄 PDF Abstract BibTeX arXiv:1912.13032

Code (0)

등록된 구현이 없습니다.

Tasks

BIG-bench Machine LearningBinary ClassificationManagement

Similar Papers 제목 키워드 기반

Self-supervision for health insurance claims data: a Covid-19 use case

2021-07-19 · Emilia Apostolova, Fazle Karim, Guido Muscioni, Anubhav Rana 외

In this work, we modify and apply self-supervision techniques to the domain of medical health insurance claims. We model patients' healthcare claims history analogous to free-text narratives, and introduce pre-trained `p…

Markov model with machine learning integration for fraud detection in health insurance

2021-02-11 · Rohan Yashraj Gupta, Satya Sai Mudigonda, Pallav Kumar Baruah, Phani Krishna Kandala

Fraud has led to a huge addition of expenses in health insurance sector in India. The work is aimed to provide methods applied to health insurance fraud detection. The work presents two approaches - a markov model and an…

BIG-bench Machine LearningFraud Detection

Construction of extra-large scale screening tools for risks of severe mental illnesses using real world healthcare data

2022-12-20 · Dianbo Liu, Karmel W. Choi, Paulo Lizano, William Yuan 외

Importance: The prevalence of severe mental illnesses (SMIs) in the United States is approximately 3% of the whole population. The ability to conduct risk screening of SMIs at large scale could inform early prevention an…

A Self-Attention Network for Hierarchical Data Structures with an Application to Claims Management

2018-08-30 · Leander Löw, Martin Spindler, Eike Brechmann

Insurance companies must manage millions of claims per year. While most of these claims are non-fraudulent, fraud detection is core for insurance companies. The ultimate goal is a predictive model to single out the fraud…

Fraud DetectionManagement

Sequence embeddings help to identify fraudulent cases in healthcare insurance

2019-10-07 · I. Fursov, A. Zaytsev, R. Khasyanov, M. Spindler 외

Fraud causes substantial costs and losses for companies and clients in the finance and insurance industries. Examples are fraudulent credit card transactions or fraudulent claims. It has been estimated that roughly $10$ …

Fraud DetectionManagement