paper-with-me

홈 › Papers

SkinCAP: A Multi-modal Dermatology Dataset Annotated with Rich Medical Captions

2024-05-28 · Juexiao Zhou, Liyuan Sun, Yan Xu, wenbin liu, Shawn Afvari, Zhongyi Han, Jiaoyan Song, Yongzhi Ji, Xiaonan He, Xin Gao

With the widespread application of artificial intelligence (AI), particularly deep learning (DL) and vision-based large language models (VLLMs), in skin disease diagnosis, the need for interpretability becomes crucial. However, existing dermatology datasets are limited in their inclusion of concept-level meta-labels, and none offer rich medical descriptions in natural language. This deficiency impedes the advancement of LLM-based methods in dermatological diagnosis. To address this gap and provide a meticulously annotated dermatology dataset with comprehensive natural language descriptions, we introduce SkinCAP: a multi-modal dermatology dataset annotated with rich medical captions. SkinCAP comprises 4,000 images sourced from the Fitzpatrick 17k skin disease dataset and the Diverse Dermatology Images dataset, annotated by board-certified dermatologists to provide extensive medical descriptions and captions. Notably, SkinCAP represents the world's first such dataset and is publicly available at https://huggingface.co/datasets/joshuachou/SkinCAP.

📄 PDF Abstract BibTeX arXiv:2405.18004

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

DermaBench: A Clinician-Annotated Benchmark Dataset for Dermatology Visual Question Answering and Reasoning

2026-01-20 · Abdurrahim Yilmaz, Ozan Erdem, Ece Gokyayla, Ayda Acar 외 arxiv

Vision-language models (VLMs) are increasingly important in medical applications; however, their evaluation in dermatology remains limited by datasets that focus primarily on image-level classification tasks such as lesi…

Visual Question Answering

MM-Skin: Enhancing Dermatology Vision-Language Model with an Image-Text Dataset Derived from Textbooks

2025-05-09 · Wenqi Zeng, Yuqi Sun, Chenxi Ma, Weimin Tan 외

Medical vision-language models (VLMs) have shown promise as clinical assistants across various medical fields. However, specialized dermatology VLM capable of delivering professional and detailed diagnostic analysis rema…

DiagnosticInstruction FollowingLanguage ModelingLanguage Modelling+4

Enhanced Dermatology Image Quality Assessment via Cross-Domain Training

2025-06-19 · Ignacio Hernández Montilla, Alfonso Medela, Paola Pasquali, Andy Aguilar 외

Teledermatology has become a widely accepted communication method in daily clinical practice, enabling remote care while showing strong agreement with in-person visits. Poor image quality remains an unsolved problem in t…

Image Quality Assessment

DermaVQA-DAS: Dermatology Assessment Schema (DAS) & Datasets for Closed-Ended Question Answering & Segmentation in Patient-Generated Dermatology Images

2025-12-30 · Wen-wai Yim, Yujuan Fu, Asma Ben Abacha, Meliha Yetisgen 외 arxiv

Recent advances in dermatological image analysis have been driven by large-scale annotated datasets; however, most existing benchmarks focus on dermatoscopic images and lack patient-authored queries and clinical context,…

Lesion SegmentationQuestion Answering

Are Multimodal LLMs Ready for Clinical Dermatology? A Real-World Evaluation in Dermatology

2026-05-01 · Roy Jiang, Hyunjae Kim, Zhenyue Qin, Morten Lee 외 arxiv

Multimodal large language models (MLLMs) have demonstrated promise on publicly available dermatology benchmarks. However, benchmark performance may not generalize to real-world dermatologic decision-making. To quantify t…