paper-with-me

홈 › Papers

On the Challenges of Evaluating Compositional Explanations in Multi-Hop Inference: Relevance, Completeness, and Expert Ratings

2021-09-07 · EMNLP 2021 11 · Peter Jansen, Kelly Smith, Dan Moreno, Huitzilin Ortiz

Building compositional explanations requires models to combine two or more facts that, together, describe why the answer to a question is correct. Typically, these "multi-hop" explanations are evaluated relative to one (or a small number of) gold explanations. In this work, we show these evaluations substantially underestimate model performance, both in terms of the relevance of included facts, as well as the completeness of model-generated explanations, because models regularly discover and produce valid explanations that are different than gold explanations. To address this, we construct a large corpus of 126k domain-expert (science teacher) relevance ratings that augment a corpus of explanations to standardized science exam questions, discovering 80k additional relevant facts not rated as gold. We build three strong models based on different methodologies (generation, ranking, and schemas), and empirically show that while expert-augmented ratings provide better estimates of explanation quality, both original (gold) and expert-augmented automatic evaluations still substantially underestimate performance by up to 36% when compared with full manual expert judgements, with different models being disproportionately affected. This poses a significant methodological challenge to accurately evaluating explanations produced by compositional reasoning models.

📄 PDF Abstract BibTeX arXiv:2109.03334

Code (0)

등록된 구현이 없습니다.

Tasks

valid

Similar Papers 제목 키워드 기반

Detection Accuracy for Evaluating Compositional Explanations of Units

2021-09-16 · Sayo M. Makinwa, Biagio La Rosa, Roberto Capobianco

The recent success of deep learning models in solving complex problems and in different domains has increased interest in understanding what they learn. Therefore, different approaches have been employed to explain these…

Compositional Explanations of Neurons

2020-06-24 · NeurIPS 2020 12 · Jesse Mu, Jacob Andreas

We describe a procedure for explaining neurons in deep representations by identifying compositional logical concepts that closely approximate neuron behavior. Compared to prior work that uses atomic labels as explanation…

image-classificationImage ClassificationNatural Language Inference

Towards a fuller understanding of neurons with Clustered Compositional Explanations

2023-10-27 · NeurIPS 2023 11 · Biagio La Rosa, Leilani H. Gilpin, Roberto Capobianco

Compositional Explanations is a method for identifying logical formulas of concepts that approximate the neurons' behavior. However, these explanations are linked to the small spectrum of neuron activations (i.e., the hi…

Open Vocabulary Compositional Explanations for Neuron Alignment

2025-11-25 · Biagio La Rosa, Leilani H. Gilpin arxiv

Neurons are the fundamental building blocks of deep neural networks, and their interconnections allow AI to achieve unprecedented results. Motivated by the goal of understanding how neurons encode information, compositio…

Open Vocabulary Semantic Segmentation

C2C: Component-to-Composition Learning for Zero-Shot Compositional Action Recognition

2024-07-08 · Rongchang Li, ZhenHua Feng, Tianyang Xu, Linze Li 외

Compositional actions consist of dynamic (verbs) and static (objects) concepts. Humans can easily recognize unseen compositions using the learned concepts. For machines, solving such a problem requires a model to recogni…

Action Recognition