paper-with-me

홈 › Papers

WeNLEX: Weakly Supervised Natural Language Explanations for Multilabel Chest X-ray Classification

2026-03-19 · Isabel Rio-Torto, Jaime S. Cardoso, Luís F. Teixeira arxiv

Natural language explanations provide an inherently human-understandable way to explain black-box models, closely reflecting how radiologists convey their diagnoses in textual reports. Most works explicitly supervise the explanation generation process using datasets annotated with explanations. Thus, though plausible, the generated explanations are not faithful to the model's reasoning. In this work, we propose WeNLEX, a weakly supervised model for the generation of natural language explanations for multilabel chest X-ray classification. Faithfulness is ensured by matching images generated from their corresponding natural language explanations with original images, in the black-box model's feature space. Plausibility is maintained via distribution alignment with a small database of clinician-annotated explanations. We empirically demonstrate, through extensive validation on multiple metrics to assess faithfulness, simulatability, diversity, and plausibility, that WeNLEX is able to produce faithful and plausible explanations, using as little as 5 ground-truth explanations per diagnosis. Furthermore, WeNLEX can operate in both post-hoc and in-model settings. In the latter, i.e., when the multilabel classifier is trained together with the rest of the network, WeNLEX improves the classification AUC of the standalone classifier by 2.21%, thus showing that adding interpretability to the training process can actually increase the downstream task performance. Additionally, simply by changing the database, WeNLEX explanations are adaptable to any target audience, and we showcase this flexibility by training a layman version of WeNLEX, where explanations are simplified for non-medical users.

📄 PDF Abstract BibTeX arXiv:2603.18752

Code (0)

등록된 구현이 없습니다.

Tasks

Explanation Generation

Similar Papers 제목 키워드 기반

Weakly Supervised Explainable Phrasal Reasoning with Neural Fuzzy Logic

2021-09-18 · Zijun Wu, Zi Xuan Zhang, Atharva Naik, Zhijian Mei 외

Natural language inference (NLI) aims to determine the logical relationship between two sentences, such as Entailment, Contradiction, and Neutral. In recent years, deep learning models have become a prevailing approach t…

Explanation GenerationLogical ReasoningNatural Language InferenceSentence

Towards Interpretable Natural Language Understanding with Explanations as Latent Variables

2020-10-24 · NeurIPS 2020 12 · Wangchunshu Zhou, Jinyi Hu, HANLIN ZHANG, Xiaodan Liang 외

Recently generating natural language explanations has shown very promising results in not only offering interpretable explanations but also providing additional information and supervision for prediction. However, existi…

Explanation GenerationNatural Language Understanding

Weakly supervised one-stage vision and language disease detection using large scale pneumonia and pneumothorax studies

2020-07-31 · Leo K. Tam, Xiaosong Wang, Evrim Turkbey, Kevin Lu 외

Detecting clinically relevant objects in medical images is a challenge despite large datasets due to the lack of detailed labels. To address the label issue, we utilize the scene-level labels with a detection architectur…

Head DetectionReferring Expression

Learning Representations for Weakly Supervised Natural Language Processing Tasks

2014-03-01 · CL 2014 3 · Fei Huang, Arun Ahuja, Doug Downey, Yi Yang 외

Adapting Reinforcement Learning with Chain-of-Thought Supervision for Explainable Detection of Hateful and Propagandistic Memes

2026-06-13 · Mohamed Bayan Kmainasi, Mucahid Kutlu, Ali Ezzat Shahroor, Abul Hasnat 외 arxiv

Hateful and propagandistic memes exploit the interplay between images and text to convey harmful intent that neither modality reveals alone. Although thinking-based multimodal large language models (MLLMs) have advanced …

Reinforcement Learning