paper-with-me

홈 › Papers

MedGen3D: A Deep Generative Framework for Paired 3D Image and Mask Generation

2023-04-08 · Kun Han, Yifeng Xiong, Chenyu You, Pooya Khosravi, Shanlin Sun, Xiangyi Yan, James Duncan, Xiaohui Xie

Acquiring and annotating sufficient labeled data is crucial in developing accurate and robust learning-based models, but obtaining such data can be challenging in many medical image segmentation tasks. One promising solution is to synthesize realistic data with ground-truth mask annotations. However, no prior studies have explored generating complete 3D volumetric images with masks. In this paper, we present MedGen3D, a deep generative framework that can generate paired 3D medical images and masks. First, we represent the 3D medical data as 2D sequences and propose the Multi-Condition Diffusion Probabilistic Model (MC-DPM) to generate multi-label mask sequences adhering to anatomical geometry. Then, we use an image sequence generator and semantic diffusion refiner conditioned on the generated mask sequences to produce realistic 3D medical images that align with the generated masks. Our proposed framework guarantees accurate alignment between synthetic images and segmentation maps. Experiments on 3D thoracic CT and brain MRI datasets show that our synthetic data is both diverse and faithful to the original data, and demonstrate the benefits for downstream segmentation tasks. We anticipate that MedGen3D's ability to synthesize paired 3D medical images and masks will prove valuable in training deep learning models for medical imaging tasks.

📄 PDF Abstract BibTeX arXiv:2304.04106

Code (0)

등록된 구현이 없습니다.

Tasks

Image SegmentationMedical Image SegmentationSegmentationSemantic Segmentation

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…
Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

SynerMedGen: Synergizing Medical Multimodal Understanding with Generation via Task Alignment

2026-05-09 · Weiren Zhao, Yi Dong, Cheng Chen arxiv

Unifying multimodal understanding and generation is a compelling frontier that is beginning to emerge in the medical field. However, the limited existing unified medical models typically treat understanding and generatio…

MedGEN-Bench: Contextually entangled benchmark for open-ended multimodal medical generation

2025-11-17 · Junjie Yang, Yuhao Yan, Gang Wu, Yuxuan Wang 외 arxiv

As Vision-Language Models (VLMs) increasingly gain traction in medical applications, clinicians are progressively expecting AI systems not only to generate textual diagnoses but also to produce corresponding medical imag…

Visual Question Answeringmultimodal generationImage GenerationImage Editing

MRGen: Diffusion-based Controllable Data Engine for MRI Segmentation towards Unannotated Modalities

2024-12-04 · HaoNing Wu, Ziheng Zhao, Ya zhang, Weidi Xie 외

Medical image segmentation has recently demonstrated impressive progress with deep neural networks, yet the heterogeneous modalities and scarcity of mask annotations limit the development of segmentation models on unanno…

Image GenerationImage SegmentationMedical Image GenerationMedical Image Segmentation+3

MaskSketch: Unpaired Structure-guided Masked Image Generation

2023-02-10 · CVPR 2023 1 · Dina Bashkirova, Jose Lezama, Kihyuk Sohn, Kate Saenko 외

Recent conditional image generation methods produce images of remarkable diversity, fidelity and realism. However, the majority of these methods allow conditioning only on labels or text prompts, which limits their level…

Conditional Image GenerationDiversityImage GenerationImage-to-Image Translation+2

MaskedCLIP: Bridging the Masked and CLIP Space for Semi-Supervised Medical Vision-Language Pre-training

2025-07-23 · Lei Zhu, Jun Zhou, Rick Siow Mong Goh, Yong Liu arxiv

Foundation models have recently gained tremendous popularity in medical image analysis. State-of-the-art methods leverage either paired image-text data via vision-language pre-training or unpaired image data via self-sup…

Knowledge Distillation