A Labeled Ophthalmic Ultrasound Dataset with Medical Report Generation Based on Cross-modal Deep Learning
Ultrasound imaging reveals eye morphology and aids in diagnosing and treating eye diseases. However, interpreting diagnostic reports requires specialized physicians. We present a labeled ophthalmic dataset for the precise analysis and the automated exploration of medical images along with their associated reports. It collects three modal data, including the ultrasound images, blood flow information and examination reports from 2,417 patients at an ophthalmology hospital in Shenyang, China, during the year 2018, in which the patient information is de-identified for privacy protection. To the best of our knowledge, it is the only ophthalmic dataset that contains the three modal information simultaneously. It incrementally consists of 4,858 images with the corresponding free-text reports, which describe 15 typical imaging findings of intraocular diseases and the corresponding anatomical locations. Each image shows three kinds of blood flow indices at three specific arteries, i.e., nine parameter values to describe the spectral characteristics of blood flow distribution. The reports were written by ophthalmologists during the clinical care. The proposed dataset is applied to generate medical report based on the cross-modal deep learning model. The experimental results demonstrate that our dataset is suitable for training supervised models concerning cross-modal medical data.
Code (0)
등록된 구현이 없습니다.
Tasks
DiagnosticMedical Report GenerationSimilar Papers 제목 키워드 기반
Ophtha-LLaMA2: A Large Language Model for Ophthalmology
In recent years, pre-trained large language models (LLMs) have achieved tremendous success in the field of Natural Language Processing (NLP). Prior studies have primarily focused on general and generic domains, with rela…
DiagnosticLanguage ModelingLanguage ModellingLarge Language Model+1Cross-modal Clinical Graph Transformer for Ophthalmic Report Generation
Automatic generation of ophthalmic reports using data-driven neural networks has great potential in clinical practice. When writing a report, ophthalmologists make inferences with prior clinical knowledge. This knowledge…
Clinical KnowledgeDecoderMedical Report GenerationUltrasound Report Generation with Cross-Modality Feature Alignment via Unsupervised Guidance
Automatic report generation has arisen as a significant research area in computer-aided diagnosis, aiming to alleviate the burden on clinicians by generating reports automatically based on medical images. In this work, w…
Ultrasound Image Classification using ACGAN with Small Training Dataset
B-mode ultrasound imaging is a popular medical imaging technique. Like other image processing tasks, deep learning has been used for analysis of B-mode ultrasound images in the last few years. However, training deep lear…
ClassificationData AugmentationDeep LearningGeneral Classification+4OphGLM: Training an Ophthalmology Large Language-and-Vision Assistant based on Instructions and Dialogue
Large multimodal language models (LMMs) have achieved significant success in general domains. However, due to the significant differences between medical images and text and general web content, the performance of LMMs i…
Instruction FollowingLanguage ModelingLanguage ModellingLarge Language Model+1