paper-with-me

홈 › Papers

Towards a Taxonomy for the Use of Synthetic Data in Advanced Analytics

2022-12-05 · Peter Kowalczyk, Giacomo Welsch, Frédéric Thiesse

The proliferation of deep learning techniques led to a wide range of advanced analytics applications in important business areas such as predictive maintenance or product recommendation. However, as the effectiveness of advanced analytics naturally depends on the availability of sufficient data, an organization's ability to exploit the benefits might be restricted by limited data or likewise data access. These challenges could force organizations to spend substantial amounts of money on data, accept constrained analytics capacities, or even turn into a showstopper for analytics projects. Against this backdrop, recent advances in deep learning to generate synthetic data may help to overcome these barriers. Despite its great potential, however, synthetic data are rarely employed. Therefore, we present a taxonomy highlighting the various facets of deploying synthetic data for advanced analytics systems. Furthermore, we identify typical application scenarios for synthetic data to assess the current state of adoption and thereby unveil missed opportunities to pave the way for further research.

📄 PDF Abstract BibTeX arXiv:2212.02622

Code (0)

등록된 구현이 없습니다.

Tasks

Deep LearningProduct Recommendation

Similar Papers 제목 키워드 기반

Amalgam: Hybrid LLM-PGM Synthesis Algorithm for Accuracy and Realism

2026-03-28 · Antheas Kapenekakis, Bent Thomsen, Katja Hose, Michele Albano arxiv

To generate synthetic datasets, e.g., in domains such as healthcare, the literature proposes approaches of two main types: Probabilistic Graphical Models (PGMs) and Deep Learning models, such as LLMs. While PGMs produce …

In-RDBMS Hardware Acceleration of Advanced Analytics

2018-01-08 · Divya Mahajan, Joon Kyung Kim, Jacob Sacks, Adel Ardalan 외

The data revolution is fueled by advances in machine learning, databases, and hardware design. Programmable accelerators are making their way into each of these areas independently. As such, there is a void of solutions …

Optimizing the Privacy-Utility Balance using Synthetic Data and Configurable Perturbation Pipelines

2025-04-24 · Anantha Sharma, Swetha Devabhaktuni, Eklove Mohan

This paper explores the strategic use of modern synthetic data generation and advanced data perturbation techniques to enhance security, maintain analytical utility, and improve operational efficiency when managing large…

Privacy PreservingSynthetic Data Generation

Beyond Weights and Gradients: A Taxonomy of Federated Learning Messages

2026-06-15 · Alvaro Javier Vargas Guerrero, Xinguang Wang, Quang Manh Doan, Guy Nagels arxiv

Federated Learning is rapidly evolving beyond the exchange of traditional model weights and gradients, yet existing definitions fail to capture the full scope of modern payloads like synthetic data and federated analytic…

Federated Learning

Lightweight Knowledge Representations for Automating Data Analysis

2023-10-15 · Marko Sterbentz, Cameron Barrie, Donna Hooshmand, Shubham Shahi 외

The principal goal of data science is to derive meaningful information from data. To do this, data scientists develop a space of analytic possibilities and from it reach their information goals by using their knowledge o…