paper-with-me

홈 › Papers

Analysis of Blood Report Images Using General Purpose Vision-Language Models

2025-09-07 · Nadia Bakhsheshi, Hamid Beigy arxiv

The reliable analysis of blood reports is important for health knowledge, but individuals often struggle with interpretation, leading to anxiety and overlooked issues. We explore the potential of general-purpose Vision-Language Models (VLMs) to address this challenge by automatically analyzing blood report images. We conduct a comparative evaluation of three VLMs: Qwen-VL-Max, Gemini 2.5 Pro, and Llama 4 Maverick, determining their performance on a dataset of 100 diverse blood report images. Each model was prompted with clinically relevant questions adapted to each blood report. The answers were then processed using Sentence-BERT to compare and evaluate how closely the models responded. The findings suggest that general-purpose VLMs are a practical and promising technology for developing patient-facing tools for preliminary blood report analysis. Their ability to provide clear interpretations directly from images can improve health literacy and reduce the limitations to understanding complex medical information. This work establishes a foundation for the future development of reliable and accessible AI-assisted healthcare applications. While results are encouraging, they should be interpreted cautiously given the limited dataset size.

📄 PDF Abstract BibTeX arXiv:2509.06033

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Adaptive Gray World-Based Color Normalization of Thin Blood Film Images

2016-07-14 · F. Boray Tek, Andrew G. Dempster, İzzet Kale

This paper presents an effective color normalization method for thin blood film images of peripheral blood specimens. Thin blood film images can easily be separated to foreground (cell) and background (plasma) parts. The…

Color Normalization

Vision4PPG: Emergent PPG Analysis Capability of Vision Foundation Models for Vital Signs like Blood Pressure

2025-10-11 · Saurabh Kataria, Ayca Ermis, Lovely Yeswanth Panchumarthi, Minxiao Wang 외 arxiv

Photoplethysmography (PPG) sensor in wearable and clinical devices provides valuable physiological insights in a non-invasive and real-time fashion. Specialized Foundation Models (FM) or repurposed time-series FMs are us…

parameter-efficient fine-tuningBlood pressure estimation

Semantic Segmentation of Anaemic RBCs Using Multilevel Deep Convolutional Encoder-Decoder Network

2022-02-09 · Muhammad Shahzad, Arif Iqbal Umar, Syed Hamad Shirazi, Israr Ahmed Shaikh

Pixel-level analysis of blood images plays a pivotal role in diagnosing blood-related diseases, especially Anaemia. These analyses mainly rely on an accurate diagnosis of morphological deformities like shape, size, and p…

DecoderMorphological AnalysisSemantic Segmentation

A Labeled Ophthalmic Ultrasound Dataset with Medical Report Generation Based on Cross-modal Deep Learning

2024-07-26 · Jing Wang, Junyan Fan, Meng Zhou, Yanzhu Zhang 외

Ultrasound imaging reveals eye morphology and aids in diagnosing and treating eye diseases. However, interpreting diagnostic reports requires specialized physicians. We present a labeled ophthalmic dataset for the precis…

DiagnosticMedical Report Generation

Machine learning approach of automatic identification and counting of blood cells

2019-09-05 · Healthcare Technology Letters, IET 2019 9 · Mohammad Mahmudul Alam, Mohammad Tariqul Islam

A complete blood cell count is an important test in medical diagnosis to evaluate overall health condition. Traditionally blood cells are counted manually using haemocytometer along with other laboratory equipment’s and …

BIG-bench Machine LearningBlood Cell CountBlood Cell DetectionCBC TEST+3