paper-with-me

홈 › Papers

Inspecting Explainability of Transformer Models with Additional Statistical Information

2023-11-19 · Hoang C. Nguyen, Haeil Lee, Junmo Kim

Transformer becomes more popular in the vision domain in recent years so there is a need for finding an effective way to interpret the Transformer model by visualizing it. In recent work, Chefer et al. can visualize the Transformer on vision and multi-modal tasks effectively by combining attention layers to show the importance of each image patch. However, when applying to other variants of Transformer such as the Swin Transformer, this method can not focus on the predicted object. Our method, by considering the statistics of tokens in layer normalization layers, shows a great ability to interpret the explainability of Swin Transformer and ViT.

📄 PDF Abstract BibTeX arXiv:2311.11378

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Multi-Head Attention 설명 없음
Attention 설명 없음
Dropout Dropout is a regularization technique for neural networks that drops a unit (along with connections) at training time with a specified probability $p$ (a common value is…
Absolute Position Encodings Absolute Position Encodings are a type of position embeddings for [Transformer-based models] where positional encodings are…
Position-Wise Feed-Forward Layer 설명 없음
Label Smoothing Label Smoothing is a regularization technique that introduces noise for the labels. This accounts for the fact that datasets may have mistakes in them, so maximizing the…
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Adam 설명 없음

Similar Papers 제목 키워드 기반

Can human clinical rationales improve the performance and explainability of clinical text classification models?

2025-07-28 · Christoph Metzner, Shang Gao, Drahomira Herrmannova, Heidi A. Hanson arxiv

AI-driven clinical text classification is vital for explainable automated retrieval of population-level health information. This work investigates whether human-based clinical rationales can serve as additional supervisi…

Text Classification

The geometry of BERT

2025-02-17 · Matteo Bonino, Giorgia Ghione, Giansalvo Cirrincione

Transformer neural networks, particularly Bidirectional Encoder Representations from Transformers (BERT), have shown remarkable performance across various tasks such as classification, text summarization, and question an…

Question AnsweringText Summarization

R-Cut: Enhancing Explainability in Vision Transformers with Relationship Weighted Out and Cut

2023-07-18 · Yingjie Niu, Ming Ding, Maoning Ge, Robin Karlsson 외

Transformer-based models have gained popularity in the field of natural language processing (NLP) and are extensively utilized in computer vision tasks and multi-modal models such as GPT4. This paper presents a novel met…

image-classificationImage Classification

Evaluating Webcam-based Gaze Data as an Alternative for Human Rationale Annotations

2024-02-29 · Stephanie Brandl, Oliver Eberle, Tiago Ribeiro, Anders Søgaard 외

Rationales in the form of manually annotated input spans usually serve as ground truth when evaluating explainability methods in NLP. They are, however, time-consuming and often biased by the annotation process. In this …

valid

From Text to Graph: Leveraging Graph Neural Networks for Enhanced Explainability in NLP

2025-04-02 · Fabio Yáñez-Romero, Andrés Montoyo, Armando Suárez, Yoan Gutiérrez 외

Researchers have relegated natural language processing tasks to Transformer-type models, particularly generative models, because these models exhibit high versatility when performing generation and classification tasks. …