paper-with-me

Papers

ERUPT: Efficient Rendering with Unposed Patch Transformer

2025-03-31 · CVPR 2025 1 · Maxim V. Shugaev, Vincent Chen, Maxim Karrenbach, Kyle Ashley, Bridget Kennedy, Naresh P. Cuntoor

This work addresses the problem of novel view synthesis in diverse scenes from small collections of RGB images. We propose ERUPT (Efficient Rendering with Unposed Patch Transformer) a state-of-the-art scene reconstruction model capable of efficient scene rendering using unposed imagery. We introduce patch-based querying, in contrast to existing pixel-based queries, to reduce the compute required to render a target view. This makes our model highly efficient both during training and at inference, capable of rendering at 600 fps on commercial hardware. Notably, our model is designed to use a learned latent camera pose which allows for training using unposed targets in datasets with sparse or inaccurate ground truth camera pose. We show that our approach can generalize on large real-world data and introduce a new benchmark dataset (MSVS-1M) for latent view synthesis using street-view imagery collected from Mapillary. In contrast to NeRF and Gaussian Splatting, which require dense imagery and precise metadata, ERUPT can render novel views of arbitrary scenes with as few as five unposed input images. ERUPT achieves better rendered image quality than current state-of-the-art methods for unposed image synthesis tasks, reduces labeled data requirements by ~95\% and decreases computational requirements by an order of magnitude, providing efficient novel view synthesis for diverse real-world scenes.

📄 PDF Abstract BibTeX arXiv:2503.24374

Code (0)

등록된 구현이 없습니다.

Tasks

Image GenerationNeRFNovel View Synthesis

Similar Papers 제목 키워드 기반

PanoLAM: Large Avatar Model for Gaussian Full-Head Synthesis from One-shot Unposed Image

2025-09-09 · Peng Li, Yisheng He, Yingdong Hu, Yuan Dong 외 arxiv

We present a feed-forward framework for Gaussian full-head synthesis from a single unposed image. Unlike previous work that relies on time-consuming GAN inversion and test-time optimization, our framework can reconstruct…

Free360: Layered Gaussian Splatting for Unbounded 360-Degree View Synthesis from Extremely Sparse and Unposed Views

2025-03-31 · CVPR 2025 1 · Chong Bao, Xiyu Zhang, Zehao Yu, Jiale Shi 외

Neural rendering has demonstrated remarkable success in high-quality 3D neural reconstruction and novel view synthesis with dense input views and accurate poses. However, applying it to extremely sparse, unposed views in…

3D ReconstructionNeural RenderingNovel View SynthesisSurface Reconstruction

Pose-Free Generalizable Rendering Transformer

2023-10-05 · Zhiwen Fan, Panwang Pan, Peihao Wang, Yifan Jiang 외

In the field of novel-view synthesis, the necessity of knowing camera poses (e.g., via Structure from Motion) before rendering has been a common practice. However, the consistent acquisition of accurate camera poses rema…

Generalizable Novel View SynthesisNovel View Synthesis

Interpreting LSTM Prediction on Solar Flare Eruption with Time-series Clustering

2019-12-27 · Hu Sun, Ward Manchester, Zhenbang Jiao, Xiantong Wang 외

We conduct a post hoc analysis of solar flare predictions made by a Long Short Term Memory (LSTM) model employing data in the form of Space-weather HMI Active Region Patches (SHARP) parameters calculated from data in pro…

Binary ClassificationClusteringDimensionality ReductionTime Series+2

Learning Robust Multi-Scale Representation for Neural Radiance Fields from Unposed Images

2023-11-08 · Nishant Jain, Suryansh Kumar, Luc van Gool

We introduce an improved solution to the neural image-based rendering problem in computer vision. Given a set of images taken from a freely moving camera at train time, the proposed approach could synthesize a realistic …

Camera Pose EstimationDepth EstimationDepth PredictionGraph Neural Network+2