paper-with-me

홈 › Papers

VL-MedGuide: A Visual-Linguistic Large Model for Intelligent and Explainable Skin Disease Auxiliary Diagnosis

2025-08-08 · Kexin Yu, Zihan Xu, Jialei Xie, Carter Adams arxiv

Accurate diagnosis of skin diseases remains a significant challenge due to the complex and diverse visual features present in dermatoscopic images, often compounded by a lack of interpretability in existing purely visual diagnostic models. To address these limitations, this study introduces VL-MedGuide (Visual-Linguistic Medical Guide), a novel framework leveraging the powerful multi-modal understanding and reasoning capabilities of Visual-Language Large Models (LVLMs) for intelligent and inherently interpretable auxiliary diagnosis of skin conditions. VL-MedGuide operates in two interconnected stages: a Multi-modal Concept Perception Module, which identifies and linguistically describes dermatologically relevant visual features through sophisticated prompt engineering, and an Explainable Disease Reasoning Module, which integrates these concepts with raw visual information via Chain-of-Thought prompting to provide precise disease diagnoses alongside transparent rationales. Comprehensive experiments on the Derm7pt dataset demonstrate that VL-MedGuide achieves state-of-the-art performance in both disease diagnosis (83.55% BACC, 80.12% F1) and concept detection (76.10% BACC, 67.45% F1), surpassing existing baselines. Furthermore, human evaluations confirm the high clarity, completeness, and trustworthiness of its generated explanations, bridging the gap between AI performance and clinical utility by offering actionable, explainable insights for dermatological practice.

📄 PDF Abstract BibTeX arXiv:2508.06624

Code (0)

등록된 구현이 없습니다.

Tasks

Prompt Engineering

Similar Papers 제목 키워드 기반

MedGUIDE: Benchmarking Clinical Decision-Making in Large Language Models

2025-05-16 · Xiaomin Li, Mingye Gao, Yuexing Hao, Taoran Li 외

Clinical guidelines, typically structured as decision trees, are central to evidence-based medical practice and critical for ensuring safe and accurate diagnostic decision-making. However, it remains unclear whether Larg…

BenchmarkingDecision MakingDiagnosticMultiple-choice

MedGuideX: Internalizing Decision Logic from Executable Guidelines into Large Language Models for Clinical Reasoning

2026-05-26 · Yuhao Shen, Lang Cao, Simo Du, Yuqing Wang 외 arxiv

Clinical practice guidelines (CPGs) encode evidence-based decision logic that clinicians apply by evaluating patient variables, conditional criteria, and recommendation rules. However, existing methods often use CPGs as …

LAXARY: A Trustworthy Explainable Twitter Analysis Model for Post-Traumatic Stress Disorder Assessment

2020-03-16 · Mohammad Arif Ul Alam, Dhawal Kapadia

Veteran mental health is a significant national problem as large number of veterans are returning from the recent war in Iraq and continued military presence in Afghanistan. While significant existing works have investig…

BIG-bench Machine LearningExplainable Artificial Intelligence (XAI)Survey

Language Generation for Broad-Coverage, Explainable Cognitive Systems

2022-01-25 · Marjorie McShane, Ivan Leon

This paper describes recent progress on natural language generation (NLG) for language-endowed intelligent agents (LEIAs) developed within the OntoAgent cognitive architecture. The approach draws heavily from past work o…

Natural Language UnderstandingText Generation

Active Learning and Explainable AI for Multi-Objective Optimization of Spin Coated Polymers

2025-09-10 · Brendan Young, Brendan Alvey, Andreas Werbrouck, Will Murphy 외 arxiv

Spin coating polymer thin films to achieve specific mechanical properties is inherently a multi-objective optimization problem. We present a framework that integrates an active Pareto front learning algorithm (PyePAL) wi…

Active Learning