paper-with-me

홈 › Papers

Investigating Bias with a Synthetic Data Generator: Empirical Evidence and Philosophical Interpretation

2022-09-13 · Alessandro Castelnovo, Riccardo Crupi, Nicole Inverardi, Daniele Regoli, Andrea Cosentini

Machine learning applications are becoming increasingly pervasive in our society. Since these decision-making systems rely on data-driven learning, risk is that they will systematically spread the bias embedded in data. In this paper, we propose to analyze biases by introducing a framework for generating synthetic data with specific types of bias and their combinations. We delve into the nature of these biases discussing their relationship to moral and justice frameworks. Finally, we exploit our proposed synthetic data generator to perform experiments on different scenarios, with various bias combinations. We thus analyze the impact of biases on performance and fairness metrics both in non-mitigated and mitigated machine learning models.

📄 PDF Abstract BibTeX arXiv:2209.05889

Code (1)

rcrupiisp/isparity 공식 구현

Tasks

Decision MakingFairness

Similar Papers 제목 키워드 기반

On the effects of biased quantum random numbers on the initialization of artificial neural networks

2021-08-30 · Raoul Heese, Moritz Wolter, Sascha Mücke, Lukas Franken 외

Recent advances in practical quantum computing have led to a variety of cloud-based quantum computing platforms that allow researchers to evaluate their algorithms on noisy intermediate-scale quantum (NISQ) devices. A co…

Toward Annotator Group Bias in Crowdsourcing

2021-10-08 · ACL 2022 5 · Haochen Liu, Joseph Thekinen, Sinem Mollaoglu, Da Tang 외

Crowdsourcing has emerged as a popular approach for collecting annotated data to train supervised machine learning models. However, annotator bias can lead to defective annotations. Though there are a few works investiga…

When Pretty Isn't Useful: Investigating Why Modern Text-to-Image Models Fail as Reliable Training Data Generators

2026-02-23 · Krzysztof Adamkiewicz, Brian Bernhard Moser, Stanislav Frolov, Tobias Christian Nauen 외 arxiv

Recent text-to-image (T2I) diffusion models produce visually stunning images and demonstrate excellent prompt following. But do they perform well as synthetic vision data generators? In this work, we revisit the promise …

Twinning Complex Networked Systems: Data-Driven Calibration of the mABCD Synthetic Graph Generator

2026-02-02 · Piotr Bródka, Michał Czuba, Bogumił Kamiński, Łukasz Kraiński 외 arxiv

The increasing availability of relational data has contributed to a growing reliance on network-based representations of complex systems. Over time, these models have evolved to capture more nuanced properties, such as t…

DECAF: Generating Fair Synthetic Data Using Causally-Aware Generative Networks

2021-10-25 · NeurIPS 2021 12 · Boris van Breugel, Trent Kyono, Jeroen Berrevoets, Mihaela van der Schaar

Machine learning models have been criticized for reflecting unfair biases in the training data. Instead of solving for this by introducing fair learning algorithms directly, we focus on generating fair synthetic data, su…

Fairness