paper-with-me

홈 › Papers

FreMIM: Fourier Transform Meets Masked Image Modeling for Medical Image Segmentation

2023-04-21 · Wenxuan Wang, Jing Wang, Chen Chen, Jianbo Jiao, Yuanxiu Cai, Shanshan Song, Jiangyun Li

The research community has witnessed the powerful potential of self-supervised Masked Image Modeling (MIM), which enables the models capable of learning visual representation from unlabeled data. In this paper, to incorporate both the crucial global structural information and local details for dense prediction tasks, we alter the perspective to the frequency domain and present a new MIM-based framework named FreMIM for self-supervised pre-training to better accomplish medical image segmentation tasks. Based on the observations that the detailed structural information mainly lies in the high-frequency components and the high-level semantics are abundant in the low-frequency counterparts, we further incorporate multi-stage supervision to guide the representation learning during the pre-training phase. Extensive experiments on three benchmark datasets show the superior advantage of our FreMIM over previous state-of-the-art MIM methods. Compared with various baselines trained from scratch, our FreMIM could consistently bring considerable improvements to model performance. The code will be publicly available at https://github.com/Rubics-Xuan/FreMIM.

📄 PDF Abstract BibTeX arXiv:2304.10864

Code (1)

rubics-xuan/fremim 공식 구현 pytorch

Tasks

Image SegmentationMedical Image SegmentationRepresentation LearningSemantic Segmentation

Methods 이 논문이 사용한 방법론

MIM 설명 없음

Similar Papers 제목 키워드 기반

F2former: When Fractional Fourier Meets Deep Wiener Deconvolution and Selective Frequency Transformer for Image Deblurring

2024-09-03 · Subhajit Paul, Sahil Kumawat, Ashutosh Gupta, Deepak Mishra

Recent progress in image deblurring techniques focuses mainly on operating in both frequency and spatial domains using the Fourier transform (FT) properties. However, their performance is limited due to the dependency of…

DeblurringDecoderImage DeblurringImage Restoration

HiViT: Hierarchical Vision Transformer Meets Masked Image Modeling

2022-05-30 · Xiaosong Zhang, Yunjie Tian, Wei Huang, Qixiang Ye 외

Recently, masked image modeling (MIM) has offered a new methodology of self-supervised pre-training of vision transformers. A key idea of efficient implementation is to discard the masked image patches (or tokens) throug…

Transfer Learning

Compressive Hyperspectral Imaging: Fourier Transform Interferometry meets Single Pixel Camera

2018-09-04 · Amirafshar Moshtaghpour, José M. Bioucas-Dias, Laurent Jacques

This paper introduces a single-pixel HyperSpectral (HS) imaging framework based on Fourier Transform Interferometry (FTI). By combining a space-time coding of the light illumination with partial interferometric observati…

Compressive Sensing

ConvMAE: Masked Convolution Meets Masked Autoencoders

2022-05-08 · Peng Gao, Teli Ma, Hongsheng Li, Ziyi Lin 외

Vision Transformers (ViT) become widely-adopted architectures for various vision tasks. Masked auto-encoding for feature pretraining and multi-scale hybrid convolution-transformer architectures can further unleash the po…

Computational Efficiencyimage-classificationImage ClassificationObject Detection+1

RILS: Masked Visual Reconstruction in Language Semantic Space

2023-01-17 · CVPR 2023 1 · Shusheng Yang, Yixiao Ge, Kun Yi, Dian Li 외

Both masked image modeling (MIM) and natural language supervision have facilitated the progress of transferable visual pre-training. In this work, we seek the synergy between two paradigms and study the emerging properti…

Sentence