paper-with-me

홈 › Papers

VM-DDPM: Vision Mamba Diffusion for Medical Image Synthesis

2024-05-09 · Zhihan Ju, Wanting Zhou

In the realm of smart healthcare, researchers enhance the scale and diversity of medical datasets through medical image synthesis. However, existing methods are limited by CNN local perception and Transformer quadratic complexity, making it difficult to balance structural texture consistency. To this end, we propose the Vision Mamba DDPM (VM-DDPM) based on State Space Model (SSM), fully combining CNN local perception and SSM global modeling capabilities, while maintaining linear computational complexity. Specifically, we designed a multi-level feature extraction module called Multi-level State Space Block (MSSBlock), and a basic unit of encoder-decoder structure called State Space Layer (SSLayer) for medical pathological images. Besides, we designed a simple, Plug-and-Play, zero-parameter Sequence Regeneration strategy for the Cross-Scan Module (CSM), which enabled the S6 module to fully perceive the spatial features of the 2D image and stimulate the generalization potential of the model. To our best knowledge, this is the first medical image synthesis model based on the SSM-CNN hybrid architecture. Our experimental evaluation on three datasets of different scales, i.e., ACDC, BraTS2018, and ChestXRay, as well as qualitative evaluation by radiologists, demonstrate that VM-DDPM achieves state-of-the-art performance.

📄 PDF Abstract BibTeX arXiv:2405.05667

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderDiversityImage GenerationMamba

Methods 이 논문이 사용한 방법론

Attention 설명 없음
Dropout Dropout is a regularization technique for neural networks that drops a unit (along with connections) at training time with a specified probability $p$ (a common value is…
Label Smoothing Label Smoothing is a regularization technique that introduces noise for the labels. This accounts for the fact that datasets may have mistakes in them, so maximizing the…
Residual Connection 설명 없음
Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Position-Wise Feed-Forward Layer 설명 없음
Absolute Position Encodings Absolute Position Encodings are a type of position embeddings for [Transformer-based models] where positional encodings are…
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.

Similar Papers 제목 키워드 기반

Fast-DDPM: Fast Denoising Diffusion Probabilistic Models for Medical Image-to-Image Generation

2024-05-23 · Hongxu Jiang, Muhammad Imran, Linhai Ma, Teng Zhang 외

Denoising diffusion probabilistic models (DDPMs) have achieved unprecedented success in computer vision. However, they remain underutilized in medical imaging, a field crucial for disease diagnosis and treatment planning…

DenoisingImage DenoisingImage GenerationImage Super-Resolution+2

Measurement-conditioned Denoising Diffusion Probabilistic Model for Under-sampled Medical Image Reconstruction

2022-03-05 · Yutong Xie, Quanzheng Li

We propose a novel and unified method, measurement-conditioned denoising diffusion probabilistic model (MC-DDPM), for under-sampled medical image reconstruction based on DDPM. Different from previous works, MC-DDPM is de…

DenoisingImage ReconstructionMRI Reconstruction

Accelerating Diffusion Models via Pre-segmentation Diffusion Sampling for Medical Image Segmentation

2022-10-27 · Xutao Guo, Yanwu Yang, Chenfei Ye, Shang Lu 외

Based on the Denoising Diffusion Probabilistic Model (DDPM), medical image segmentation can be described as a conditional image generation task, which allows to compute pixel-wise uncertainty maps of the segmentation and…

Conditional Image GenerationDenoisingImage GenerationImage Segmentation+3

Computationally Efficient Diffusion Models in Medical Imaging: A Comprehensive Review

2025-05-09 · Abdullah, Tao Huang, Ickjai Lee, Euijoon Ahn

The diffusion model has recently emerged as a potent approach in computer vision, demonstrating remarkable performances in the field of generative artificial intelligence. Capable of producing high-quality synthetic imag…

Denoising

Soft Masked Mamba Diffusion Model for CT to MRI Conversion

2024-06-22 · Zhenbin Wang, Lei Zhang, Lituan Wang, Zhenwei Zhang

Magnetic Resonance Imaging (MRI) and Computed Tomography (CT) are the predominant modalities utilized in the field of medical imaging. Although MRI capture the complexity of anatomical structures with greater detail than…

Computed Tomography (CT)Image GenerationMambaMedical Image Generation