paper-with-me

Papers

Memory-efficient High-resolution OCT Volume Synthesis with Cascaded Amortized Latent Diffusion Models

2024-05-26 · Kun Huang, Xiao Ma, Yuhan Zhang, Na Su, Songtao Yuan, Yong liu, Qiang Chen, Huazhu Fu

Optical coherence tomography (OCT) image analysis plays an important role in the field of ophthalmology. Current successful analysis models rely on available large datasets, which can be challenging to be obtained for certain tasks. The use of deep generative models to create realistic data emerges as a promising approach. However, due to limitations in hardware resources, it is still difficulty to synthesize high-resolution OCT volumes. In this paper, we introduce a cascaded amortized latent diffusion model (CA-LDM) that can synthesis high-resolution OCT volumes in a memory-efficient way. First, we propose non-holistic autoencoders to efficiently build a bidirectional mapping between high-resolution volume space and low-resolution latent space. In tandem with autoencoders, we propose cascaded diffusion processes to synthesize high-resolution OCT volumes with a global-to-local refinement process, amortizing the memory and computational demands. Experiments on a public high-resolution OCT dataset show that our synthetic data have realistic high-resolution and global features, surpassing the capabilities of existing methods. Moreover, performance gains on two down-stream fine-grained segmentation tasks demonstrate the benefit of the proposed method in training deep learning models for medical imaging tasks. The code is public available at: https://github.com/nicetomeetu21/CA-LDM.

📄 PDF Abstract BibTeX arXiv:2405.16516

Code (1)

nicetomeetu21/ca-ldm 공식 구현 pytorch

Methods 이 논문이 사용한 방법론

Latent Diffusion Model Diffusion models applied to latent spaces, which are normally built with (Variational) Autoencoders.
Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Memory-efficient Segmentation of High-resolution Volumetric MicroCT Images

2022-05-31 · YuAn Wang, Laura Blackie, Irene Miguel-Aliaga, Wenjia Bai

In recent years, 3D convolutional neural networks have become the dominant approach for volumetric medical image segmentation. However, compared to their 2D counterparts, 3D networks introduce substantially more training…

GPUImage SegmentationMedical Image SegmentationSegmentation+3

Cascaded Context Pyramid for Full-Resolution 3D Semantic Scene Completion

2019-08-01 · ICCV 2019 10 · Pingping Zhang, Wei Liu, Yinjie Lei, Huchuan Lu 외

Semantic Scene Completion (SSC) aims to simultaneously predict the volumetric occupancy and semantic category of a 3D scene. It helps intelligent devices to understand and interact with the surrounding scenes. Due to the…

3D Semantic Scene Completion

Cascaded 3D Diffusion Models for Whole-body 3D 18-F FDG PET/CT synthesis from Demographics

2025-05-28 · Siyeop Yoon, Sifan Song, Pengfei Jin, Matthew Tivnan 외

We propose a cascaded 3D diffusion model framework to synthesize high-fidelity 3D PET/CT volumes directly from demographic variables, addressing the growing need for realistic digital twins in oncologic imaging, virtual …

Data AugmentationSuper-Resolution

LUVE : Latent-Cascaded Ultra-High-Resolution Video Generation with Dual Frequency Experts

2026-02-12 · Chen Zhao, Jiawei Chen, Hongyu Li, Zhuoliang Kang 외 arxiv

Recent advances in video diffusion models have significantly improved visual quality, yet ultra-high-resolution (UHR) video generation remains a formidable challenge due to the compounded difficulties of motion modeling,…

Video Generation

Surf2CT: Cascaded 3D Flow Matching Models for Torso 3D CT Synthesis from Skin Surface

2025-05-28 · Siyeop Yoon, Yujin Oh, Pengfei Jin, Sifan Song 외

We present Surf2CT, a novel cascaded flow matching framework that synthesizes full 3D computed tomography (CT) volumes of the human torso from external surface scans and simple demographic data (age, sex, height, weight)…

AnatomyComputed Tomography (CT)Super-Resolution