paper-with-me

Papers

Local Patches Meet Global Context: Scalable 3D Diffusion Priors for Computed Tomography Reconstruction

2025-12-20 · Taewon Yang, Jason Hu, Jeffrey A. Fessler, Liyue Shen arxiv

Diffusion models learn strong image priors that can be leveraged to solve inverse problems like medical image reconstruction. However, for real-world applications such as 3D Computed Tomography (CT) imaging, directly training diffusion models on 3D data presents significant challenges due to the high computational demands of extensive GPU resources and large-scale datasets. Existing works mostly reuse 2D diffusion priors to address 3D inverse problems, but fail to fully realize and leverage the generative capacity of diffusion models for high-dimensional data. In this study, we propose a novel 3D patch-based diffusion model that can learn a fully 3D diffusion prior from limited data, enabling scalable generation of high-resolution 3D images. Our core idea is to learn the prior of 3D patches to achieve scalable efficiency, while coupling local and global information to guarantee high-quality 3D image generation, by modeling the joint distribution of position-aware 3D local patches and downsampled 3D volume as global context. Our approach not only enables high-quality 3D generation, but also offers an unprecedentedly efficient and accurate solution to high-resolution 3D inverse problems. Experiments on 3D CT reconstruction across multiple datasets show that our method outperforms state-of-the-art methods in both performance and efficiency, notably achieving high-resolution 3D reconstruction of $512 \times 512 \times 256$ ($\sim$20 mins).

📄 PDF Abstract BibTeX arXiv:2512.18161

Code (0)

등록된 구현이 없습니다.

Tasks

Image Reconstruction3D ReconstructionImage Generation3D Generation

Similar Papers 제목 키워드 기반

U-Netmer: U-Net meets Transformer for medical image segmentation

2023-04-03 · Sheng He, Rina Bao, P. Ellen Grant, Yangming Ou

The combination of the U-Net based deep learning models and Transformer is a new trend for medical image segmentation. U-Net can extract the detailed local semantic and texture information and Transformer can learn the l…

Image SegmentationMedical Image SegmentationSegmentationSemantic Segmentation

Global-Local Transformer for Brain Age Estimation

2021-09-03 · Sheng He, P. Ellen Grant, Yangming Ou

Deep learning can provide rapid brain age estimation based on brain magnetic resonance imaging (MRI). However, most studies use one neural network to extract the global information from the whole input image, ignoring th…

Age Estimation

OCTOPUS: Enhancing the Spatial-Awareness of Vision SSMs with Multi-Dimensional Scans and Traversal Selection

2026-01-31 · Kunal Mahatha, Ali Bahri, Pierre Marza, Sahar Dastani 외 arxiv

State space models (SSMs) have recently emerged as an alternative to transformers due to their unique ability of modeling global relationships in text with linear complexity. However, their success in vision tasks has be…

Collaborative Global-Local Networks for Memory-Efficient Segmentation of Ultra-High Resolution Images

2019-05-15 · CVPR 2019 6 · Wuyang Chen, Ziyu Jiang, Zhangyang Wang, Kexin Cui 외

Segmentation of ultra-high resolution images is increasingly demanded, yet poses significant challenges for algorithm efficiency, in particular considering the (GPU) memory limits. Current approaches either downsample an…

GPULand Cover ClassificationSegmentationSemantic Segmentation

Neural-Schwarz Tiling for Geometry-Universal PDE Solving at Scale

2026-05-12 · Paolo Secchi, Daniel S. Balint, Marco Maurizi arxiv

Most learned PDE solvers follow a global-surrogate paradigm: a neural operator is trained to map full problem descriptions to full solution fields for a prescribed distribution of geometries, boundary conditions, and coe…