paper-with-me

Papers

VolTA-3D: Self-Supervised Learning for Brain MRI using 3D Volumetric Token Alignment

2026-05-16 · Amy Makawana, Abhijeet Parida, Marius George Linguraru, Julia Ive, Syed Muhammad Anwar arxiv

Self-supervised learning (SSL) has advanced medical image analysis be enabling learning form large unlabelled data. However, in brain magnetic resonance imaging (MRI), most 3D models remain specialized for either segmentation of classification, limiting their ability to generalize across datasets, imaging protocols,, and downstream tasks. This lack of transferability constrains the clinical utility of 3D MRI models, despite the availability of unlabeled volumetric data. We present Volta-3D, a self-supervised 3D Vision Transformer framework designed to learn transferable volumetric representations. Volta-3D jointly aligns global class-style tokens and local patch tokens within a student-teacher paradigm and enforces fine-grained structural reconstruction. This combined global-local alignment addresses the limited semantic diversity and subtle anatomical characteristics of brain MRI, which challenges existing SSL approaches. We evaluate Volta-3D on multiple out-of-distribution downstream tasks, including hippocampal segmentation and classification of sex and Alzheimer's disease versus healthy controls. Across all tasks, representations learned by Volta-3D outperform randomly initialized baselines, demonstrating improved transferability and robustness under domain shift. Hence jointly enforcing global semantic consistency and local structural learning during pretraining enables broader concept learning from unlabeled brain MRI data. Overall VolTA-3D supports effective multi-task downstream performance with task-specific pertaining, a step towards generalizable and clinically viable 3D models.

📄 PDF Abstract BibTeX arXiv:2605.16775

Code (0)

등록된 구현이 없습니다.

Tasks

Self-Supervised Learning

Similar Papers 제목 키워드 기반

Self-supervised Feature Learning for 3D Medical Images by Playing a Rubik's Cube

2019-10-05 · Xinrui Zhuang, Yuexiang Li, Yifan Hu, Kai Ma 외

Witnessed the development of deep learning, increasing number of studies try to build computer aided diagnosis systems for 3D volumetric medical data. However, as the annotations of 3D medical data are difficult to acqui…

Brain Tumor SegmentationDeep LearningRubik's CubeSelf-Supervised Learning+1

Revisiting Rubik's Cube: Self-supervised Learning with Volume-wise Transformation for 3D Medical Image Segmentation

2020-07-17 · Xing Tao, Yuexiang Li, Wenhui Zhou, Kai Ma 외

Deep learning highly relies on the quantity of annotated data. However, the annotations for 3D volumetric medical data require experienced physicians to spend hours or even days for investigation. Self-supervised learnin…

Image SegmentationMedical Image SegmentationPancreas SegmentationRubik's Cube+2

Multi-Modal Brain Tumor Segmentation via 3D Multi-Scale Self-attention and Cross-attention

2025-04-12 · Yonghao Huang, Leiting Chen, Chuan Zhou

Due to the success of CNN-based and Transformer-based models in various computer vision tasks, recent works study the applicability of CNN-Transformer hybrid architecture models in 3D multi-modality medical segmentation …

Brain Tumor SegmentationDecoderImage SegmentationMedical Image Segmentation+3

BrainDINO: A Brain MRI Foundation Model for Generalizable Clinical Representation Learning

2026-04-30 · Yizhou Wu, Shansong Wang, Yuheng Li, Mojtaba Safari 외 arxiv

Brain MRI underpins a wide range of neuroscientific and clinical applications, yet most learning-based methods remain task-specific and require substantial labeled data. Here we show that a single self-supervised represe…

Self-Supervised LearningRepresentation LearningTumor SegmentationAge Estimation

Towards Generalisable Foundation Models for Brain MRI

2025-10-27 · Moona Mazher, Geoff J. M. Parker, Daniel C. Alexander arxiv

Foundation models in artificial intelligence (AI) are transforming medical imaging by enabling general-purpose feature learning from large-scale, unlabeled datasets. In this work, we introduce BrainFound, a self-supervis…

Image Segmentation