paper-with-me

Papers

Improve Fidelity and Utility of Synthetic Credit Card Transaction Time Series from Data-centric Perspective

2024-01-01 · Din-Yin Hsieh, Chi-Hua Wang, Guang Cheng

Exploring generative model training for synthetic tabular data, specifically in sequential contexts such as credit card transaction data, presents significant challenges. This paper addresses these challenges, focusing on attaining both high fidelity to actual data and optimal utility for machine learning tasks. We introduce five pre-processing schemas to enhance the training of the Conditional Probabilistic Auto-Regressive Model (CPAR), demonstrating incremental improvements in the synthetic data's fidelity and utility. Upon achieving satisfactory fidelity levels, our attention shifts to training fraud detection models tailored for time-series data, evaluating the utility of the synthetic data. Our findings offer valuable insights and practical guidelines for synthetic data practitioners in the finance sector, transitioning from real to synthetic datasets for training purposes, and illuminating broader methodologies for synthesizing credit card transaction time series.

📄 PDF Abstract BibTeX arXiv:2401.00965

Code (0)

등록된 구현이 없습니다.

Tasks

Fraud DetectionTime Series

Similar Papers 제목 키워드 기반

HiMPO: Hindsight-Informed Memory Policy Optimization for Less-Entangled Credit in Long-Horizon Agents

2026-06-15 · Jiangze Yan, Yi Shen, Wenjing Zhang, Jieyun Huang 외 arxiv

Long-horizon agents rely on memory mechanisms to compress interaction history, but optimizing memory writing faces a distinct credit assignment challenge: a memory update may be rewarded or penalized due to downstream to…

Balancing Fidelity, Utility, and Privacy in Synthetic Cardiac MRI Generation: A Comparative Study

2026-03-04 · Madhura Edirisooriya, Dasuni Kawya, Ishan Kumarasinghe, Isuri Devindi 외 arxiv

Deep learning in cardiac MRI (CMR) is fundamentally constrained by both data scarcity and privacy regulations. This study systematically benchmarks three generative architectures: Denoising Diffusion Probabilistic Models…

Domain GeneralizationData Augmentation

Generating High-quality Privacy-preserving Synthetic Data

2026-02-06 · David Yavo, Richard Khoury, Christophe Pere, Sadoune Ait Kaci Azzou arxiv

Synthetic tabular data enables sharing and analysis of sensitive records, but its practical deployment requires balancing distributional fidelity, downstream utility, and privacy protection. We study a simple, model agno…

GraphGuard: Contrastive Self-Supervised Learning for Credit-Card Fraud Detection in Multi-Relational Dynamic Graphs

2024-07-17 · Kristófer Reynisson, Marco Schreyer, Damian Borth

Credit card fraud has significant implications at both an individual and societal level, making effective prevention essential. Current methods rely heavily on feature engineering and labeled information, both of which h…

Feature EngineeringFraud DetectionSelf-Supervised Learning

Synthetic Cardiac MRI Image Generation using Deep Generative Models

2026-03-25 · Ishan Kumarasinghe, Dasuni Kawya, Madhura Edirisooriya, Isuri Devindi 외 arxiv

Synthetic cardiac MRI (CMRI) generation has emerged as a promising strategy to overcome the scarcity of annotated medical imaging data. Recent advances in GANs, VAEs, diffusion probabilistic models, and flow-matching tec…

Domain GeneralizationImage Generation