paper-with-me

Papers

Patch-PODiff-ViT: Structured Latent Diffusion with Patchwise POD for Super-Resolution and Uncertainty Quantification

2026-06-30 · Onkar Jadhav, Tim French, Matthew Rayson, Nicole L. Jones arxiv

Diffusion models enable probabilistic super-resolution and conditional generation, but pixel-space methods are computationally expensive and learned latent spaces often lack interpretable uncertainty quantification. We introduce Patch-PODiff-ViT, a structured latent diffusion framework in which the latent space is defined by patchwise Proper Orthogonal Decomposition (POD), a fixed linear orthonormal basis over local patches, rather than learned by a nonlinear autoencoder. This yields low-dimensional, variance-ordered tokens that preserve spatial structure and enable efficient diffusion in a structured low-dimensional latent space with a Vision Transformer. Because the decoder is fixed, linear, and orthonormal, latent coefficient uncertainty can be propagated directly to physical-space predictive variance, enabling analytic propagation of predictive variance through the linear decoder without Monte Carlo estimation in pixel space. Across sea surface temperature, medical imaging, and natural images, the method achieves strong reconstruction with fewer parameters and lower memory, while producing well-calibrated spatial uncertainty that closely matches empirical ensembles.

📄 PDF Abstract BibTeX arXiv:2606.31290

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

PODiff: Latent Diffusion in Proper Orthogonal Decomposition Space for Scientific Super-Resolution

2026-05-05 · Onkar Jadhav, Tim French, Matthew Rayson, Nicole L. Jones arxiv

Probabilistic super-resolution of high-dimensional spatial fields using diffusion models is often computationally prohibitive due to the cost of operating directly in pixel space. We propose PODiff, a structured conditio…

CompoDiff: Versatile Composed Image Retrieval With Latent Diffusion

2023-03-21 · Geonmo Gu, Sanghyuk Chun, Wonjae Kim, HeeJae Jun 외

This paper proposes a novel diffusion-based model, CompoDiff, for solving zero-shot Composed Image Retrieval (ZS-CIR) with latent diffusion. This paper also introduces a new synthetic dataset, named SynthTriplets18M, wit…

Composed Image Retrieval (CoIR)Image RetrievalRetrievalZero-Shot Composed Image Retrieval (ZS-CIR)

TopoDiffuser: A Diffusion-Based Multimodal Trajectory Prediction Model with Topometric Maps

2025-08-01 · Zehui Xu, Junhui Wang, Yongliang Shi, Chao Gao 외 arxiv

This paper introduces TopoDiffuser, a diffusion-based framework for multimodal trajectory prediction that incorporates topometric maps to generate accurate, diverse, and road-compliant future motion forecasts. By embeddi…

Trajectory Prediction

GeoTopoDiff: Learning Geometry--Topology Graph Priors through Boundary-Constrained Mixed Diffusion for Sparse-Slice 3D Porous Reconstruction

2026-05-05 · Yue Shi, Peng Wang, Mingzhe Yu, Yunlong Zhao 외 arxiv

Diffusion-based voxel prior modelling is challenging for the reconstruction of large-scale 3D porous microstructures. Due to the demanding requirements for simultaneously modelling both the continuous pore morphology and…

Geometry- and Relation-Aware Diffusion for EEG Super-Resolution

2026-02-02 · Laura Yao, Gengwei Zhang, Moajjem Chowdhury, Yunmei Liu 외 arxiv

Recent electroencephalography (EEG) spatial super-resolution (SR) methods, while showing improved quality by either directly predicting missing signals from visible channels or adapting latent diffusion-based generative …

Emotion RecognitionSeizure Detection