Causal Feature Selection for Responsible Machine Learning
Machine Learning (ML) has become an integral aspect of many real-world applications. As a result, the need for responsible machine learning has emerged, focusing on aligning ML models to ethical and social values, while enhancing their reliability and trustworthiness. Responsible ML involves many issues. This survey addresses four main issues: interpretability, fairness, adversarial robustness, and domain generalization. Feature selection plays a pivotal role in the responsible ML tasks. However, building upon statistical correlations between variables can lead to spurious patterns with biases and compromised performance. This survey focuses on the current study of causal feature selection: what it is and how it can reinforce the four aspects of responsible ML. By identifying features with causal impacts on outcomes and distinguishing causality from correlation, causal feature selection is posited as a unique approach to ensuring ML models to be ethically and socially responsible in high-stakes applications.
Code (0)
등록된 구현이 없습니다.
Tasks
Adversarial RobustnessDomain GeneralizationFairnessfeature selectionSurveyMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Causality-based Feature Selection: Methods and Evaluations
Feature selection is a crucial preprocessing step in data analytics and machine learning. Classical feature selection algorithms select features based on the correlations between predictive features and the class variabl…
feature selectionSelecting Robust Features for Machine Learning Applications using Multidata Causal Discovery
Robust feature selection is vital for creating reliable and interpretable Machine Learning (ML) models. When designing statistical prediction models in cases where domain knowledge is limited and underlying interactions …
Causal DiscoveryDimensionality ReductionExplainable artificial intelligencefeature selection+1Causal Feature Selection via Transfer Entropy
Machine learning algorithms are designed to capture complex relationships between features. In this context, the high dimensionality of data often results in poor model performance, with the risk of overfitting. Feature …
Causal Discoveryfeature selectionregressionTime SeriesTowards Robust Classification Model by Counterfactual and Invariant Data Generation
Despite the success of machine learning applications in science, industry, and society in general, many approaches are known to be non-robust, often relying on spurious correlations to make predictions. Spuriousness occu…
Classificationcounterfactualimage-classificationImage Classification+2Causality and Interpretability for Electrical Distribution System faults
Causal analysis helps us understand variables that are responsible for system failures. This improves fault detection and makes system more reliable. In this work, we present a new method that combines causal inference w…
Causal Inference