paper-with-me

홈 › Papers

Interpreting random forest classification models using a feature contribution method

2013-12-04 · Anna Palczewska, Jan Palczewski, Richard Marchese Robinson, Daniel Neagu

Model interpretation is one of the key aspects of the model evaluation process. The explanation of the relationship between model variables and outputs is relatively easy for statistical models, such as linear regressions, thanks to the availability of model parameters and their statistical significance. For "black box" models, such as random forest, this information is hidden inside the model structure. This work presents an approach for computing feature contributions for random forest classification models. It allows for the determination of the influence of each variable on the model prediction for an individual instance. By analysing feature contributions for a training dataset, the most significant variables can be determined and their typical contribution towards predictions made for individual classes, i.e., class-specific feature contribution "patterns", are discovered. These patterns represent a standard behaviour of the model and allow for an additional assessment of the model reliability for a new data. Interpretation of feature contributions for two UCI benchmark datasets shows the potential of the proposed methodology. The robustness of results is demonstrated through an extensive analysis of feature contributions calculated for a large number of generated random forest models.

📄 PDF Abstract BibTeX arXiv:1312.1121

Code (0)

등록된 구현이 없습니다.

Tasks

ClassificationGeneral Classification

Similar Papers 제목 키워드 기반

Interpreting Deep Forest through Feature Contribution and MDI Feature Importance

2023-05-01 · Yi-Xiao He, Shen-Huan Lyu, Yuan Jiang

Deep forest is a non-differentiable deep model which has achieved impressive empirical success across a wide variety of applications, especially on categorical/symbolic or mixed modeling tasks. Many of the application fi…

Explainable ModelsFeature Importance

Interpretation and Simplification of Deep Forest

2020-01-14 · Sangwon Kim, Mira Jeong, Byoung Chul Ko

This paper proposes a new method for interpreting and simplifying a black box model of a deep random forest (RF) using a proposed rule elimination. In deep RF, a large number of decision trees are connected to multiple l…

Explicating feature contribution using Random Forest proximity distances

2018-07-17 · Leanne S. Whitmore, Anthe George, Corey M. Hudson

In Random Forests, proximity distances are a metric representation of data into decision space. By observing how changes in input map to the movement of instances in this space we are able to determine the independent co…

Decision MakingGeneral Classification

Disentangled Attribution Curves for Interpreting Random Forests and Boosted Trees

2019-05-18 · Summer Devlin, Chandan Singh, W. James Murdoch, Bin Yu

Tree ensembles, such as random forests and AdaBoost, are ubiquitous machine learning models known for achieving strong predictive performance across a wide variety of domains. However, this strong performance comes at th…

Feature EngineeringFeature ImportanceInterpretable Machine Learning

Texture Feature Analysis for Classification of Early-Stage Prostate Cancer in mpMRI

2024-06-21 · Asmail Muftah, S M Schirmer, Frank C Langbein

Magnetic resonance imaging (MRI) has become a crucial tool in the diagnosis and staging of prostate cancer, owing to its superior tissue contrast. However, it also creates large volumes of data that must be assessed by t…

Classification