paper-with-me

홈 › Papers

Towards the Formalization of a Trustworthy AI for Mining Interpretable Models explOiting Sophisticated Algorithms

2025-10-23 · Riccardo Guidotti, Martina Cinquini, Marta Marchiori Manerba, Mattia Setzu, Francesco Spinnato arxiv

Interpretable-by-design models are crucial for fostering trust, accountability, and safe adoption of automated decision-making models in real-world applications. In this paper we formalize the ground for the MIMOSA (Mining Interpretable Models explOiting Sophisticated Algorithms) framework, a comprehensive methodology for generating predictive models that balance interpretability with performance while embedding key ethical properties. We formally define here the supervised learning setting across diverse decision-making tasks and data types, including tabular data, time series, images, text, transactions, and trajectories. We characterize three major families of interpretable models: feature importance, rule, and instance based models. For each family, we analyze their interpretability dimensions, reasoning mechanisms, and complexity. Beyond interpretability, we formalize three critical ethical properties, namely causality, fairness, and privacy, providing formal definitions, evaluation metrics, and verification procedures for each. We then examine the inherent trade-offs between these properties and discuss how privacy requirements, fairness constraints, and causal reasoning can be embedded within interpretable pipelines. By evaluating ethical measures during model generation, this framework establishes the theoretical foundations for developing AI systems that are not only accurate and interpretable but also fair, privacy-preserving, and causally aware, i.e., trustworthy.

📄 PDF Abstract BibTeX arXiv:2510.20621

Code (0)

등록된 구현이 없습니다.

Tasks

Feature Importance

Similar Papers 제목 키워드 기반

OpinionRank: Extracting Ground Truth Labels from Unreliable Expert Opinions with Graph-Based Spectral Ranking

2021-02-11 · Glenn Dawson, Robi Polikar

As larger and more comprehensive datasets become standard in contemporary machine learning, it becomes increasingly more difficult to obtain reliable, trustworthy label information with which to train sophisticated model…

From Correlation to Causation: Formalizing Interpretable Machine Learning as a Statistical Process

2022-07-11 · Lukas Klein, Mennatallah El-Assady, Paul F. Jäger

Explainable AI (XAI) is a necessity in safety-critical systems such as in clinical diagnostics due to a high risk for fatal decisions. Currently, however, XAI resembles a loose collection of methods rather than a well-de…

BIG-bench Machine LearningExplainable Artificial Intelligence (XAI)Interpretable Machine Learning

Revealing Unfair Models by Mining Interpretable Evidence

2022-07-12 · Mohit Bajaj, Lingyang Chu, Vittorio Romaniello, Gursimran Singh 외

The popularity of machine learning has increased the risk of unfair models getting deployed in high-stake applications, such as justice system, drug/vaccination design, and medical diagnosis. Although there are effective…

BIG-bench Machine LearningMedical Diagnosis

ReasonOps: A Unified Operational Paradigm for Trustworthy Verified LLM Reasoning

2026-05-26 · Adnan Rashid arxiv

Large Language Models (LLMs) have transformed artificial intelligence from primarily generative systems into increasingly capable reasoning agents. Recent advances in theorem proving, autoformalization, symbolic reasonin…

Conformal Prediction for Trustworthy Detection of Railway Signals

2023-01-26 · Léo Andéol, Thomas Fel, Florence De Grancey, Luca Mossina

We present an application of conformal prediction, a form of uncertainty quantification with guarantees, to the detection of railway signals. State-of-the-art architectures are tested and the most promising one undergoes…

Conformal PredictionPredictionUncertainty Quantification