paper-with-me

홈 › Papers

Discriminatory Expressions to Produce Interpretable Models in Short Documents

2020-11-27 · Manuel Francisco, Juan Luis Castro

Social Networking Sites (SNS) are one of the most important ways of communication. In particular, microblogging sites are being used as analysis avenues due to their peculiarities (promptness, short texts...). There are countless researches that use SNS in novel manners, but machine learning has focused mainly in classification performance rather than interpretability and/or other goodness metrics. Thus, state-of-the-art models are black boxes that should not be used to solve problems that may have a social impact. When the problem requires transparency, it is necessary to build interpretable pipelines. Although the classifier may be interpretable, resulting models are too complex to be considered comprehensible, making it impossible for humans to understand the actual decisions. This paper presents a feature selection mechanism that is able to improve comprehensibility by using less but more meaningful features while achieving good performance in microblogging contexts where interpretability is mandatory. Moreover, we present a ranking method to evaluate features in terms of statistical relevance and bias. We conducted exhaustive tests with five different datasets in order to evaluate classification performance, generalisation capacity and complexity of the model. Results show that our proposal is better and the most stable one in terms of accuracy, generalisation and comprehensibility.

📄 PDF Abstract BibTeX arXiv:2012.02104

Code (1)

nutcrackerugr/discriminatory-expressions 공식 구현

Tasks

feature selection

Methods 이 논문이 사용한 방법론

Feature Selection Feature selection, also known as variable selection, attribute selection or variable subset selection, is the process of selecting a subset of relevant features (variables,…
Interpretability 설명 없음

Similar Papers 제목 키워드 기반

Morphology-based Entity and Relational Entity Extraction Framework for Arabic

2017-09-17 · Amin Jaber, Fadi A. Zaraket

Rule-based techniques to extract relational entities from documents allow users to specify desired entities with natural language questions, finite state automata, regular expressions and structured query language. They …

Entity Extraction using GANMorphological AnalysisTAG

Human-interpretable clustering of short-text using large language models

2024-05-12 · Justin K. Miller, Tristram J. Alexander

Clustering short text is a difficult problem, due to the low word co-occurrence between short text documents. This work shows that large language models (LLMs) can overcome the limitations of traditional clustering appro…

ClusteringShort Text ClusteringText Clustering

A case study on context-bound referring expression generation

2019-10-01 · WS 2019 10 · Maurice Langner

In recent years, Bayesian models of referring expression generation have gained prominence in order to produce situationally more adequate referring expressions. Basically, these models enable the integration of differen…

Referring ExpressionReferring expression generation

Thematic Cohesion: measuring terms discriminatory power toward themes

2014-05-01 · LREC 2014 5 · Cl{\'e}ment de Groc, Xavier Tannier, Claude de Loupy

We present a new measure of thematic cohesion. This measure associates each term with a weight representing its discriminatory power toward a theme, this theme being itself expressed by a list of terms (a thematic lexico…

Opinion MiningRetrievalText ClusteringTranslation

Interpretable Machine Learning-Derived Spectral Indices for Vegetation Monitoring

2025-12-26 · Ali Lotfi, Adam Carter, Thuan Ha, Mohammad Meysami 외 arxiv

Spectral indices such as NDVI have driven vegetation monitoring for decades, yet their design remains largely manual and ad hoc. Their usefulness stems not only from their empirical performance, but also from algebraic f…

Interpretable Machine Learning