paper-with-me

홈 › Papers

Sequence embeddings help to identify fraudulent cases in healthcare insurance

2019-10-07 · I. Fursov, A. Zaytsev, R. Khasyanov, M. Spindler, E. Burnaev

Fraud causes substantial costs and losses for companies and clients in the finance and insurance industries. Examples are fraudulent credit card transactions or fraudulent claims. It has been estimated that roughly $10$ percent of the insurance industry's incurred losses and loss adjustment expenses each year stem from fraudulent claims. The rise and proliferation of digitization in finance and insurance have lead to big data sets, consisting in particular of text data, which can be used for fraud detection. In this paper, we propose architectures for text embeddings via deep learning, which help to improve the detection of fraudulent claims compared to other machine learning methods. We illustrate our methods using a data set from a large international health insurance company. The empirical results show that our approach outperforms other state-of-the-art methods and can help make the claims management process more efficient. As (unstructured) text data become increasingly available to economists and econometricians, our proposed methods will be valuable for many similar applications, particularly when variables have a large number of categories as is typical for example of the International Classification of Disease (ICD) codes in health economics and health services.

📄 PDF Abstract BibTeX arXiv:1910.03072

Code (0)

등록된 구현이 없습니다.

Tasks

Fraud DetectionManagement

Similar Papers 제목 키워드 기반

A Simple Baseline for Predicting Events with Auto-Regressive Tabular Transformers

2024-10-14 · Alex Stein, Samuel Sharpe, Doron Bergman, Senthil Kumar 외

Many real-world applications of tabular data involve using historic events to predict properties of new ones, for example whether a credit card transaction is fraudulent or what rating a customer will assign a product on…

Causal Language ModelingLanguage ModelingLanguage ModellingMissing Values

Identifying Linked Fraudulent Activities Using GraphConvolution Network

2021-06-05 · Sharmin Pathan, Vyom Shrivastava

In this paper, we present a novel approach to identify linked fraudulent activities or actors sharing similar attributes, using Graph Convolution Network (GCN). These linked fraudulent activities can be visualized as gra…

Community Detection

EmDT: Embedding Diffusion Transformer for Tabular Data Generation in Fraud Detection

2026-03-13 · En-Ya Kuo, Sebastien Motsch arxiv

Imbalanced datasets pose a difficulty in fraud detection, as classifiers are often biased toward the majority class and perform poorly on rare fraudulent transactions. Synthetic data generation is therefore commonly used…

Synthetic Data GenerationTabular Data GenerationFraud Detection

VERB: Visualizing and Interpreting Bias Mitigation Techniques for Word Representations

2021-04-06 · Archit Rathore, Sunipa Dev, Jeff M. Phillips, Vivek Srikumar 외

Word vector embeddings have been shown to contain and amplify biases in data they are extracted from. Consequently, many techniques have been proposed to identify, mitigate, and attenuate these biases in word representat…

Decision MakingDimensionality ReductionEthicsFairness+1

FaceTracer: Unveiling Source Identities from Swapped Face Images and Videos for Fraud Prevention

2024-12-11 · Zhongyi Zhang, Jie Zhang, Wenbo Zhou, Xinghui Zhou 외

Face-swapping techniques have advanced rapidly with the evolution of deep learning, leading to widespread use and growing concerns about potential misuse, especially in cases of fraud. While many efforts have focused on …

DisentanglementFace Swapping