paper-with-me

Papers

Conditioning 3D Diffusion Models with 2D Images: Towards Standardized OCT Volumes through En Face-Informed Super-Resolution

2024-10-13 · Coen de Vente, Mohammad Mohaiminul Islam, Philippe Valmaggia, Carel Hoyng, Adnan Tufail, Clara I. Sánchez

High anisotropy in volumetric medical images can lead to the inconsistent quantification of anatomical and pathological structures. Particularly in optical coherence tomography (OCT), slice spacing can substantially vary across and within datasets, studies, and clinical practices. We propose to standardize OCT volumes to less anisotropic volumes by conditioning 3D diffusion models with en face scanning laser ophthalmoscopy (SLO) imaging data, a 2D modality already commonly available in clinical practice. We trained and evaluated on data from the multicenter and multimodal MACUSTAR study. While upsampling the number of slices by a factor of 8, our method outperforms tricubic interpolation and diffusion models without en face conditioning in terms of perceptual similarity metrics. Qualitative results demonstrate improved coherence and structural similarity. Our approach allows for better informed generative decisions, potentially reducing hallucinations. We hope this work will provide the next step towards standardized high-quality volumetric imaging, enabling more consistent quantifications.

📄 PDF Abstract BibTeX arXiv:2410.09862

Code (0)

등록된 구현이 없습니다.

Tasks

Super-Resolution

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

CrossModalityDiffusion: Multi-Modal Novel View Synthesis with Unified Intermediate Representation

2025-01-16 · Alex Berian, Daniel Brignac, JhihYang Wu, Natnael Daba 외

Geospatial imaging leverages data from diverse sensing modalities-such as EO, SAR, and LiDAR, ranging from ground-level drones to satellite views. These heterogeneous inputs offer significant opportunities for scene unde…

Novel View SynthesisScene Understanding

Enhancing Person-to-Person Virtual Try-On with Multi-Garment Virtual Try-Off

2025-04-17 · Riza Velioglu, Petra Bevandic, Robin Chan, Barbara Hammer

Computer vision is transforming fashion through Virtual Try-On (VTON) and Virtual Try-Off (VTOFF). VTON generates images of a person in a specified garment using a target photo and a standardized garment image, while a m…

Garment ReconstructionImage GenerationVirtual Try-OffVirtual Try-on

Brain Imaging Generation with Latent Diffusion Models

2022-09-15 · Walter H. L. Pinaya, Petru-Daniel Tudosiu, Jessica Dafflon, Pedro F Da Costa 외

Deep neural networks have brought remarkable breakthroughs in medical image analysis. However, due to their data-hungry nature, the modest dataset sizes in medical imaging projects might be hindering their full potential…

Medical Image Analysis

MicroDiffusion: Implicit Representation-Guided Diffusion for 3D Reconstruction from Limited 2D Microscopy Projections

2024-03-16 · CVPR 2024 1 · Mude Hui, Zihao Wei, Hongru Zhu, Fei Xia 외

Volumetric optical microscopy using non-diffracting beams enables rapid imaging of 3D volumes by projecting them axially to 2D images but lacks crucial depth information. Addressing this, we introduce MicroDiffusion, a p…

3D ReconstructionDenoising

LAND: Lung and Nodule Diffusion for 3D Chest CT Synthesis with Anatomical Guidance

2025-10-21 · Anna Oliveras, Roger Marí, Rafael Redondo, Oriol Guardià 외 arxiv

This work introduces a new latent diffusion model to generate high-quality 3D chest CT scans conditioned on 3D anatomical masks. The method synthesizes volumetric images of size 256x256x256 at 1 mm isotropic resolution u…