paper-with-me

홈 › Papers

Can GPT-4V(ision) Serve Medical Applications? Case Studies on GPT-4V for Multimodal Medical Diagnosis

2023-10-15 · Chaoyi Wu, Jiayu Lei, Qiaoyu Zheng, Weike Zhao, Weixiong Lin, Xiaoman Zhang, Xiao Zhou, Ziheng Zhao, Ya zhang, Yanfeng Wang, Weidi Xie

Driven by the large foundation models, the development of artificial intelligence has witnessed tremendous progress lately, leading to a surge of general interest from the public. In this study, we aim to assess the performance of OpenAI's newest model, GPT-4V(ision), specifically in the realm of multimodal medical diagnosis. Our evaluation encompasses 17 human body systems, including Central Nervous System, Head and Neck, Cardiac, Chest, Hematology, Hepatobiliary, Gastrointestinal, Urogenital, Gynecology, Obstetrics, Breast, Musculoskeletal, Spine, Vascular, Oncology, Trauma, Pediatrics, with images taken from 8 modalities used in daily clinic routine, e.g., X-ray, Computed Tomography (CT), Magnetic Resonance Imaging (MRI), Positron Emission Tomography (PET), Digital Subtraction Angiography (DSA), Mammography, Ultrasound, and Pathology. We probe the GPT-4V's ability on multiple clinical tasks with or without patent history provided, including imaging modality and anatomy recognition, disease diagnosis, report generation, disease localisation. Our observation shows that, while GPT-4V demonstrates proficiency in distinguishing between medical image modalities and anatomy, it faces significant challenges in disease diagnosis and generating comprehensive reports. These findings underscore that while large multimodal models have made significant advancements in computer vision and natural language processing, it remains far from being used to effectively support real-world medical applications and clinical decision-making. All images used in this report can be found in https://github.com/chaoyi-wu/GPT-4V_Medical_Evaluation.

📄 PDF Abstract BibTeX arXiv:2310.09909

Code (1)

chaoyi-wu/gpt-4v_medical_evaluation 공식 구현

Tasks

AnatomyComputed Tomography (CT)Decision MakingMedical Diagnosis

Similar Papers 제목 키워드 기반

Ambiguous Dynamic Treatment Regimes: A Reinforcement Learning Approach

2021-12-08 · Soroush Saghafian

A main research goal in various studies is to use an observational data set and provide a new set of counterfactual guidelines that can yield causal improvements. Dynamic Treatment Regimes (DTRs) are widely studied to fo…

counterfactualDecision Makingreinforcement-learningReinforcement Learning+1

Affective Medical Estimation and Decision Making via Visualized Learning and Deep Learning

2022-05-09 · Mohammad Eslami, Solale Tabarestani, Ehsan Adeli, Glyn Elwyn 외

With the advent of sophisticated machine learning (ML) techniques and the promising results they yield, especially in medical applications, where they have been investigated for different tasks to enhance the decision-ma…

Decision MakingMemorizationSurveyUncertainty Visualization

Deep learning with noisy labels: exploring techniques and remedies in medical image analysis

2019-12-05 · Davood Karimi, Haoran Dou, Simon K. Warfield, Ali Gholipour

Supervised training of deep learning models requires large labeled datasets. There is a growing interest in obtaining such datasets for medical image analysis applications. However, the impact of label noise has not rece…

Deep LearningLearning with noisy labelsMedical Image Analysis

A review of deep learning in medical imaging: Imaging traits, technology trends, case studies with progress highlights, and future promises

2020-08-02 · S. Kevin Zhou, Hayit Greenspan, Christos Davatzikos, James S. Duncan 외

Since its renaissance, deep learning has been widely used in various medical imaging tasks and has achieved remarkable success in many medical imaging applications, thereby propelling us into the so-called artificial int…

Deep LearningSurveyUncertainty Quantification

Medical Adaptation of Large Language and Vision-Language Models: Are We Making Progress?

2024-11-06 · Daniel P. Jeong, Saurabh Garg, Zachary C. Lipton, Michael Oberst

Several recent works seek to develop foundation models specifically for medical applications, adapting general-purpose large language models (LLMs) and vision-language models (VLMs) via continued pretraining on publicly …

Medical Question AnsweringQuestion Answering