paper-with-me

홈 › Papers

Temporal and Spatial Super Resolution with Latent Diffusion Model in Medical MRI images

2024-10-29 · Vishal Dubey

Super Resolution (SR) plays a critical role in computer vision, particularly in medical imaging, where hardware and acquisition time constraints often result in low spatial and temporal resolution. While diffusion models have been applied for both spatial and temporal SR, few studies have explored their use for joint spatial and temporal SR, particularly in medical imaging. In this work, we address this gap by proposing to use a Latent Diffusion Model (LDM) combined with a Vector Quantised GAN (VQGAN)-based encoder-decoder architecture for joint super resolution. We frame SR as an image denoising problem, focusing on improving both spatial and temporal resolution in medical images. Using the cardiac MRI dataset from the Data Science Bowl Cardiac Challenge, consisting of 2D cine images with a spatial resolution of 256x256 and 8-14 slices per time-step, we demonstrate the effectiveness of our approach. Our LDM model achieves Peak Signal to Noise Ratio (PSNR) of 30.37, Structural Similarity Index (SSIM) of 0.7580, and Learned Perceptual Image Patch Similarity (LPIPS) of 0.2756, outperforming simple baseline method by 5% in PSNR, 6.5% in SSIM, 39% in LPIPS. Our LDM model generates images with high fidelity and perceptual quality with 15 diffusion steps. These results suggest that LDMs hold promise for advancing super resolution in medical imaging, potentially enhancing diagnostic accuracy and patient outcomes. Code link is also shared.

📄 PDF Abstract BibTeX arXiv:2410.23898

Code (0)

등록된 구현이 없습니다.

Tasks

DenoisingDiagnosticImage DenoisingSSIMSuper-Resolution

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…
Latent Diffusion Model Diffusion models applied to latent spaces, which are normally built with (Variational) Autoencoders.

Similar Papers 제목 키워드 기반

Learning Spatial Adaptation and Temporal Coherence in Diffusion Models for Video Super-Resolution

2024-03-25 · CVPR 2024 1 · Zhikai Chen, Fuchen Long, Zhaofan Qiu, Ting Yao 외

Diffusion models are just at a tipping point for image super-resolution task. Nevertheless, it is not trivial to capitalize on diffusion models for video super-resolution which necessitates not only the preservation of v…

DecoderDenoisingImage Super-ResolutionSuper-Resolution+3

Semantic and Temporal Integration in Latent Diffusion Space for High-Fidelity Video Super-Resolution

2025-08-01 · Yiwen Wang, Xinning Chai, Yuhong Zhang, Zhengxue Cheng 외 arxiv

Recent advancements in video super-resolution (VSR) models have demonstrated impressive results in enhancing low-resolution videos. However, due to limitations in adequately controlling the generation process, achieving …

Video Super-Resolution

Geometry- and Relation-Aware Diffusion for EEG Super-Resolution

2026-02-02 · Laura Yao, Gengwei Zhang, Moajjem Chowdhury, Yunmei Liu 외 arxiv

Recent electroencephalography (EEG) spatial super-resolution (SR) methods, while showing improved quality by either directly predicting missing signals from visible channels or adapting latent diffusion-based generative …

Emotion RecognitionSeizure Detection

VISION-XL: High Definition Video Inverse Problem Solver using Latent Image Diffusion Models

2024-11-29 · Taesung Kwon, Jong Chul Ye

In this paper, we propose a novel framework for solving high-definition video inverse problems using latent image diffusion models. Building on recent advancements in spatio-temporal optimization for video inverse proble…

DeblurringGPUSuper-ResolutionVideo Reconstruction

STCDiT: Spatio-Temporally Consistent Diffusion Transformer for High-Quality Video Super-Resolution

2025-11-24 · Junyang Chen, Jiangxin Dong, Long Sun, Yixin Yang 외 arxiv

We present STCDiT, a video super-resolution framework built upon a pre-trained video diffusion model, aiming to restore structurally faithful and temporally stable videos from degraded inputs, even under complex camera m…

Video Super-Resolution