paper-with-me

홈 › Papers

The FIX Benchmark: Extracting Features Interpretable to eXperts

2024-09-20 · Helen Jin, Shreya Havaldar, Chaehyeon Kim, Anton Xue, Weiqiu You, Helen Qu, Marco Gatti, Daniel A Hashimoto, Bhuvnesh Jain, Amin Madani, Masao Sako, Lyle Ungar, Eric Wong

Feature-based methods are commonly used to explain model predictions, but these methods often implicitly assume that interpretable features are readily available. However, this is often not the case for high-dimensional data, and it can be hard even for domain experts to mathematically specify which features are important. Can we instead automatically extract collections or groups of features that are aligned with expert knowledge? To address this gap, we present FIX (Features Interpretable to eXperts), a benchmark for measuring how well a collection of features aligns with expert knowledge. In collaboration with domain experts, we propose FIXScore, a unified expert alignment measure applicable to diverse real-world settings across cosmology, psychology, and medicine domains in vision, language, and time series data modalities. With FIXScore, we find that popular feature-based explanation methods have poor alignment with expert-specified knowledge, highlighting the need for new methods that can better identify features interpretable to experts.

📄 PDF Abstract BibTeX arXiv:2409.13684

Code (1)

BrachioLab/exlib 공식 구현 pytorch

Tasks

Time Series

Similar Papers 제목 키워드 기반

Extracting Lexical Features from Dialects via Interpretable Dialect Classifiers

2024-02-27 · Roy Xie, Orevaoghene Ahia, Yulia Tsvetkov, Antonios Anastasopoulos

Identifying linguistic differences between dialects of a language often requires expert knowledge and meticulous human analysis. This is largely due to the complexity and nuance involved in studying various dialects. We …

Interpretable and Personalized Apprenticeship Scheduling: Learning Interpretable Scheduling Policies from Heterogeneous User Demonstrations

2019-06-14 · NeurIPS 2020 12 · Rohan Paleja, Andrew Silva, Letian Chen, Matthew Gombolay

Resource scheduling and coordination is an NP-hard optimization requiring an efficient allocation of agents to a set of tasks with upper- and lower bound temporal and resource constraints. Due to the large-scale and dyna…

Decision MakingScheduling

The Need for Interpretable Features: Motivation and Taxonomy

2022-02-23 · Alexandra Zytek, Ignacio Arnaldo, Dongyu Liu, Laure Berti-Equille 외

Through extensive experience developing and explaining machine learning (ML) applications for real-world domains, we have learned that ML models are only as interpretable as their features. Even simple, highly interpreta…

Decision Making

Discussion: Effective and Interpretable Outcome Prediction by Training Sparse Mixtures of Linear Experts

2024-07-18 · Francesco Folino, Luigi Pontieri, Pietro Sabatino

Process Outcome Prediction entails predicting a discrete property of an unfinished process instance from its partial trace. High-capacity outcome predictors discovered with ensemble and deep learning methods have been sh…

feature selectionMixture-of-Experts

ERMoE: Eigen-Reparameterized Mixture-of-Experts for Stable Routing and Interpretable Specialization

2025-11-14 · Anzhe Cheng, Shukai Duan, Shixuan Li, Chenzhong Yin 외 arxiv

Mixture-of-Experts (MoE) architectures expand model capacity by sparsely activating experts but face two core challenges: misalignment between router logits and each expert's internal structure leads to unstable routing …

Text Retrieval