paper-with-me

홈 › Papers

Are Data-driven Explanations Robust against Out-of-distribution Data?

2023-03-29 · CVPR 2023 1 · Tang Li, Fengchun Qiao, Mengmeng Ma, Xi Peng

As black-box models increasingly power high-stakes applications, a variety of data-driven explanation methods have been introduced. Meanwhile, machine learning models are constantly challenged by distributional shifts. A question naturally arises: Are data-driven explanations robust against out-of-distribution data? Our empirical results show that even though predict correctly, the model might still yield unreliable explanations under distributional shifts. How to develop robust explanations against out-of-distribution data? To address this problem, we propose an end-to-end model-agnostic learning framework Distributionally Robust Explanations (DRE). The key idea is, inspired by self-supervised learning, to fully utilizes the inter-distribution information to provide supervisory signals for the learning of explanations without human annotation. Can robust explanations benefit the model's generalization capability? We conduct extensive experiments on a wide range of tasks and data types, including classification and regression on image and scientific tabular data. Our results demonstrate that the proposed method significantly improves the model's performance in terms of explanation and prediction robustness against distributional shifts.

📄 PDF Abstract BibTeX arXiv:2303.16390

Code (1)

tangli-udel/dre 공식 구현 pytorch

Tasks

Self-Supervised Learning

Similar Papers 제목 키워드 기반

Investigating the Effect of Natural Language Explanations on Out-of-Distribution Generalization in Few-shot NLI

2021-10-12 · EMNLP (insights) 2021 11 · Yangqiaoyu Zhou, Chenhao Tan

Although neural models have shown strong performance in datasets such as SNLI, they lack the ability to generalize out-of-distribution (OOD). In this work, we formulate a few-shot learning setup and examine the effects o…

Few-Shot LearningFew-Shot NLIOut-of-Distribution Generalization

Fair and Explainable Credit-Scoring under Concept Drift: Adaptive Explanation Frameworks for Evolving Populations

2025-11-05 · Shivogo John arxiv

Evolving borrower behaviors, shifting economic conditions, and changing regulatory landscapes continuously reshape the data distributions underlying modern credit-scoring systems. Conventional explainability techniques, …

Variable Detection

Detecting Answer-Driven Reasoning in LLM-Based Educational Tutors via Truncated Chain-of-Thought Auditing

2026-07-06 · Bonan Shen, Dingyan Shang, Youting Wang, Tao Ning arxiv

Large language model (LLM) tutors often produce fluent step-by-step explanations, but a correct and pedagogically formatted response does not guarantee that the answer was derived from the student-facing problem. In real…

DISCOVER: A Solver for Distributional Counterfactual Explanations

2026-03-17 · Yikai Gu, Lele Cao, Bo Zhao, Lei Lei 외 arxiv

Counterfactual explanations (CE) explain model decisions by identifying input modifications that lead to different predictions. Most existing methods operate at the instance level. Distributional Counterfactual Explanati…

Deceptive AI systems that give explanations are more convincing than honest AI systems and can amplify belief in misinformation

2024-07-31 · Valdemar Danry, Pat Pataranutaporn, Matthew Groh, Ziv Epstein 외

Advanced Artificial Intelligence (AI) systems, specifically large language models (LLMs), have the capability to generate not just misinformation, but also deceptive explanations that can justify and propagate false info…

Logical ReasoningMisinformationPersuasiveness