From Pixels to Perception: Interpretable Predictions via Instance-wise Grouped Feature Selection
Understanding the decision-making process of machine learning models provides valuable insights into the task, the data, and the reasons behind a model's failures. In this work, we propose a method that performs inherently interpretable predictions through the instance-wise sparsification of input images. To align the sparsification with human perception, we learn the masking in the space of semantically meaningful pixel regions rather than on pixel-level. Additionally, we introduce an explicit way to dynamically determine the required level of sparsity for each instance. We show empirically on semi-synthetic and natural image datasets that our inherently interpretable classifier produces more meaningful, human-understandable predictions than state-of-the-art benchmarks.
Code (1)
Tasks
Decision Makingfeature selectionMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
LIMIS: Locally Interpretable Modeling using Instance-wise Subsampling
Understanding black-box machine learning models is crucial for their widespread adoption. Learning globally interpretable models is one approach, but achieving high performance with them is challenging. An alternative ap…
Reinforcement LearningReinforcement Learning (RL)L2P: Learning to Place for Estimating Heavy-Tailed Distributed Outcomes
Many real-world prediction tasks have outcome variables that have characteristic heavy-tail distributions. Examples include copies of books sold, auction prices of art pieces, demand for commodities in warehouses, etc. B…
Layer-Wise Modality Decomposition for Interpretable Multimodal Sensor Fusion
In autonomous driving, transparency in the decision-making of perception models is critical, as even a single misperception can be catastrophic. Yet with multi-sensor inputs, it is difficult to determine how each modalit…
Autonomous DrivingA Dynamic Feature Interaction Framework for Multi-task Visual Perception
Multi-task visual perception has a wide range of applications in scene understanding such as autonomous driving. In this work, we devise an efficient unified framework to solve multiple common perception tasks, including…
Autonomous DrivingDepth EstimationInstance SegmentationScene Understanding+1Co-Win: Joint Object Detection and Instance Segmentation in LiDAR Point Clouds via Collaborative Window Processing
Accurate perception and scene understanding in complex urban environments is a critical challenge for ensuring safe and efficient autonomous navigation. In this paper, we present Co-Win, a novel bird's eye view (BEV) per…
Instance SegmentationScene UnderstandingAutonomous DrivingObject Detection