paper-with-me

홈 › Papers

Unsupervised Multimodal Representation Learning across Medical Images and Reports

2018-11-21 · Tzu-Ming Harry Hsu, Wei-Hung Weng, Willie Boag, Matthew McDermott, Peter Szolovits

Joint embeddings between medical imaging modalities and associated radiology reports have the potential to offer significant benefits to the clinical community, ranging from cross-domain retrieval to conditional generation of reports to the broader goals of multimodal representation learning. In this work, we establish baseline joint embedding results measured via both local and global retrieval methods on the soon to be released MIMIC-CXR dataset consisting of both chest X-ray images and the associated radiology reports. We examine both supervised and unsupervised methods on this task and show that for document retrieval tasks with the learned representations, only a limited amount of supervision is needed to yield results comparable to those of fully-supervised methods.

📄 PDF Abstract BibTeX arXiv:1811.08615

Code (0)

등록된 구현이 없습니다.

Tasks

Representation LearningRetrieval

Similar Papers 제목 키워드 기반

Building RadiologyNET: Unsupervised annotation of a large-scale multimodal medical database

2023-07-27 · Mateja Napravnik, Franko Hržić, Sebastian Tschauner, Ivan Štajduhar

Background and objective: The usage of machine learning in medical diagnosis and treatment has witnessed significant growth in recent years through the development of computer-aided diagnosis systems that are often relyi…

ClusteringMedical DiagnosisSemantic SimilaritySemantic Textual Similarity

M-IDoL: Information Decomposition for Modality-Specific and Diverse Representation Learning in Medical Foundation Model

2026-04-10 · Yihang Liu, Longzhen Yang, Jiaxiong Yang, Ying Wen 외 arxiv

Medical foundation models (MFMs) aim to learn universal representations from multimodal medical images that can generalize effectively to diverse downstream clinical tasks. However, most existing MFMs suffer from informa…

Representation Learning

Unsupervised Multimodal 3D Medical Image Registration with Multilevel Correlation Balanced Optimization

2024-09-08 · Jiazheng Wang, Xiang Chen, Yuxi Zhang, Min Liu 외

Surgical navigation based on multimodal image registration has played a significant role in providing intraoperative guidance to surgeons by showing the relative position of the target area to critical anatomical structu…

global-optimizationImage RegistrationMedical Image Registration

Unsupervised Segmentation of 3D Medical Images Based on Clustering and Deep Representation Learning

2018-04-11 · Takayasu Moriya, Holger R. Roth, Shota NAKAMURA, Hirohisa ODA 외

This paper presents a novel unsupervised segmentation method for 3D medical images. Convolutional neural networks (CNNs) have brought significant advances in image segmentation. However, most of the recent methods rely o…

ClusteringImage SegmentationMedical Image SegmentationRepresentation Learning+2

Multimodal Logical Inference System for Visual-Textual Entailment

2019-06-10 · ACL 2019 7 · Riko Suzuki, Hitomi Yanaka, Masashi Yoshikawa, Koji Mineshima 외

A large amount of research about multimodal inference across text and vision has been recently developed to obtain visually grounded word and sentence representations. In this paper, we use logic-based representations as…

Automated Theorem ProvingNatural Language InferenceSemantic ParsingSentence