paper-with-me

Papers

Predicting 3D structure by latent posterior sampling

2026-05-11 · Azmi Haider, Dan Rosenbaum arxiv

The remarkable achievements of both generative models of 2D images and neural field representations for 3D scenes present a compelling opportunity to integrate the strengths of both approaches. In this work, we propose a methodology that combines a NeRF-based representation of 3D scenes with probabilistic modeling and reasoning using diffusion models. We view 3D reconstruction as a perception problem with inherent uncertainty that can thereby benefit from probabilistic inference methods. The core idea is to represent the 3D scene as a stochastic latent variable for which we can learn a prior and use it to perform posterior inference given a set of observations. We formulate posterior sampling using the score-based inference method of diffusion models in conjunction with a likelihood term computed from a reconstruction model that includes volumetric rendering. We train the model using a two-stage process: first we train the reconstruction model while auto-decoding the latent representations for a dataset of 3D scenes, and then we train the prior over the latents using a diffusion model. By using the model to generate samples from the posterior we demonstrate that various 3D reconstruction tasks can be performed, differing by the type of observation used as inputs. We showcase reconstruction from single-view, multi-view, noisy images, sparse pixels, and sparse depth data. These observations vary in the amount of information they provide for the scene and we show that our method can model the varying levels of inherent uncertainty associated with each task. Our experiments illustrate that this approach yields a comprehensive method capable of accurately predicting 3D structure from diverse types of observations.

📄 PDF Abstract BibTeX arXiv:2605.10830

Code (0)

등록된 구현이 없습니다.

Tasks

3D Reconstruction

Similar Papers 제목 키워드 기반

Select-and-Sample for Spike-and-Slab Sparse Coding

2016-12-01 · NeurIPS 2016 12 · Abdul-Saboor Sheikh, Jörg Lücke

Probabilistic inference serves as a popular model for neural processing. It is still unclear, however, how approximate probabilistic inference can be accurate and scalable to very high-dimensional continuous latent space…

Denoising

On the Shape of Latent Variables in a Denoising VAE-MoG: A Posterior Sampling-Based Study

2025-09-29 · Fernanda Zapata Bascuñán arxiv

In this work, we explore the latent space of a denoising variational autoencoder with a mixture-of-Gaussians prior (VAE-MoG), trained on gravitational wave data from event GW150914. To evaluate how well the model capture…

The Gaussian Latent Machine: Efficient Prior and Posterior Sampling for Inverse Problems

2025-05-19 · Muhamed Kuric, Martin Zach, Andreas Habring, Michael Unser 외

We consider the problem of sampling from a product-of-experts-type model that encompasses many standard prior and posterior distributions commonly found in Bayesian imaging. We show that this model can be easily lifted i…

Uncertainty Inspired RGB-D Saliency Detection

2020-09-07 · Jing Zhang, Deng-Ping Fan, Yuchao Dai, Saeed Anwar 외

We propose the first stochastic framework to employ uncertainty for RGB-D saliency detection by learning from the data labeling process. Existing RGB-D saliency detection models treat this task as a point estimation prob…

DecoderRGB-D Salient Object DetectionRGB Salient Object DetectionSaliency Detection+1

Discrete flow posteriors for variational inference in discrete dynamical systems

2018-05-28 · ICLR 2019 5 · Laurence Aitchison, Vincent Adam, Srinivas C. Turaga

Each training step for a variational autoencoder (VAE) requires us to sample from the approximate posterior, so we usually choose simple (e.g. factorised) approximate posteriors in which sampling is an efficient computat…

GPUVariational Inference