paper-with-me

홈 › Papers

Disruptive Autoencoders: Leveraging Low-level features for 3D Medical Image Pre-training

2023-07-31 · Jeya Maria Jose Valanarasu, Yucheng Tang, Dong Yang, Ziyue Xu, Can Zhao, Wenqi Li, Vishal M. Patel, Bennett Landman, Daguang Xu, Yufan He, Vishwesh Nath

Harnessing the power of pre-training on large-scale datasets like ImageNet forms a fundamental building block for the progress of representation learning-driven solutions in computer vision. Medical images are inherently different from natural images as they are acquired in the form of many modalities (CT, MR, PET, Ultrasound etc.) and contain granulated information like tissue, lesion, organs etc. These characteristics of medical images require special attention towards learning features representative of local context. In this work, we focus on designing an effective pre-training framework for 3D radiology images. First, we propose a new masking strategy called local masking where the masking is performed across channel embeddings instead of tokens to improve the learning of local feature representations. We combine this with classical low-level perturbations like adding noise and downsampling to further enable low-level representation learning. To this end, we introduce Disruptive Autoencoders, a pre-training framework that attempts to reconstruct the original image from disruptions created by a combination of local masking and low-level perturbations. Additionally, we also devise a cross-modal contrastive loss (CMCL) to accommodate the pre-training of multiple modalities in a single framework. We curate a large-scale dataset to enable pre-training of 3D medical radiology images (MRI and CT). The proposed pre-training framework is tested across multiple downstream tasks and achieves state-of-the-art performance. Notably, our proposed method tops the public test leaderboard of BTCV multi-organ segmentation challenge.

📄 PDF Abstract BibTeX arXiv:2307.16896

Code (1)

mrgiovanni/suprem pytorch

Tasks

Organ SegmentationRepresentation Learning

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

What Does a Chemical Language Model Know About Molecules?

2026-06-22 · Christian Kenneth, Etowah Adams, Liam Bai, Gerard JP van Westen arxiv

Chemical language models (cLMs) are widely assumed to learn surface-level syntactic patterns rather than learning meaningful molecular semantics. Here, we apply sparse autoencoders (SAEs) to MolFormer, an encoder-only cL…

MedVAE: Efficient Automated Interpretation of Medical Images with Large-Scale Generalizable Autoencoders

2025-02-20 · Maya Varma, Ashwin Kumar, Rogier van der Sluijs, Sophie Ostmeier 외

Medical images are acquired at high resolutions with large fields of view in order to capture fine-grained features necessary for clinical decision-making. Consequently, training deep learning models on medical images ca…

Computational Efficiency

Defocus Blur Synthesis and Deblurring via Interpolation and Extrapolation in Latent Space

2023-07-28 · Ioana Mazilu, Shunxin Wang, Sven Dummer, Raymond Veldhuis 외

Though modern microscopes have an autofocusing system to ensure optimal focus, out-of-focus images can still occur when cells within the medium are not all in the same focal plane, affecting the image quality for medical…

Data AugmentationDeblurringMedical Diagnosis

Improving Lesion Segmentation in Medical Images by Global and Regional Feature Compensation

2025-02-12 · Chuhan Wang, Zhenghao Chen, Jean Y. H. Yang, Jinman Kim

Automated lesion segmentation of medical images has made tremendous improvements in recent years due to deep learning advancements. However, accurately capturing fine-grained global and regional feature representations r…

Image SegmentationLesion SegmentationMedical Image SegmentationSegmentation+2

Attention-Based Chaotic Self-Supervision for Medical Image Classification

2026-05-06 · Joao Batista Florindo, Amanda Pontes de Oliveira Ornelas arxiv

Deep learning models for medical image classification usually achieve promising results but typically rely on large, annotated datasets or standard transfer learning from ImageNet. Self-Supervised Learning (SSL) has emer…

Medical Image ClassificationSelf-Supervised LearningTransfer Learning