paper-with-me

Papers

ER-TEST Evaluating Explanation Regularization Methods for NLP Models

2022-07-01 · NAACL (TrustNLP) 2022 7 · Brihi Joshi, Aaron Chan, Ziyi Liu, Xiang Ren

Neural language models’ (NLMs’) reasoning processes are notoriously hard to explain. Recently, there has been much progress in automatically generating machine rationales of NLM behavior, but less in utilizing the rationales to improve NLM behavior. For the latter, explanation regularization (ER) aims to improve NLM generalization by pushing the machine rationales to align with human rationales. Whereas prior works primarily evaluate such ER models via in-distribution (ID) generalization, ER’s impact on out-of-distribution (OOD) is largely underexplored. Plus, little is understood about how ER model performance is affected by the choice of ER criteria or by the number/choice of training instances with human rationales. In light of this, we propose ER-TEST, a protocol for evaluating ER models’ OOD generalization along three dimensions: (1) unseen datasets, (2) contrast set tests, and (3) functional tests. Using ER-TEST, we study two key questions: (A) Which ER criteria are most effective for the given OOD setting? (B) How is ER affected by the number/choice of training instances with human rationales? ER-TEST enables comprehensive analysis of these questions by considering a diverse range of tasks and datasets. Through ER-TEST, we show that ER has little impact on ID performance, but can yield large gains on OOD performance w.r.t. (1)-(3). Also, we find that the best ER criterion is task-dependent, while ER can improve OOD performance even with limited human rationales.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

ER-Test: Evaluating Explanation Regularization Methods for Language Models

2022-05-25 · Brihi Joshi, Aaron Chan, Ziyi Liu, Shaoliang Nie 외

By explaining how humans would solve a given task, human rationales can provide strong learning signal for neural language models (LMs). Explanation regularization (ER) aims to improve LM generalization by pushing the LM…

Sanity Checks for Explanation Uncertainty

2024-03-25 · Matias Valdenegro-Toro, Mihir Mulye

Explanations for machine learning models can be hard to interpret or be wrong. Combining an explanation method with an uncertainty estimation method produces explanation uncertainty. Evaluating explanation uncertainty is…

Shapley Explanation Networks

2021-04-06 · ICLR 2021 1 · Rui Wang, Xiaoqian Wang, David I. Inouye

Shapley values have become one of the most popular feature attribution explanation methods. However, most prior work has focused on post-hoc Shapley explanations, which can be computationally demanding due to its exponen…

Evaluating the overall sensitivity of saliency-based explanation methods

2023-06-21 · Harshinee Sriram, Cristina Conati

We address the need to generate faithful explanations of "black box" Deep Learning models. Several tests have been proposed to determine aspects of faithfulness of explanation methods, but they lack cross-domain applicab…

Sensitivity

Evaluating and Improving Graph-based Explanation Methods for Multi-Agent Coordination

2025-02-14 · Siva Kailas, Shalin Jain, Harish Ravichandar

Graph Neural Networks (GNNs), developed by the graph learning community, have been adopted and shown to be highly effective in multi-robot and multi-agent learning. Inspired by this successful cross-pollination, we inves…

Graph Learning