paper-with-me

Papers

VisDom: Sparse Novel View Synthesis with Visible Domain Constraint

2026-06-18 · Mariia Gladkova*, Tarun Yenamandra*, Edmond Boyer, Robert Maier, Tony Tung, Daniel Cremers arxiv

Sparse novel view synthesis (NVS) remains challenging due to the ambiguity of recovering 3D geometry from few input views. While NeRF- and Gaussian Splatting (GS)-based methods perform well with dense supervision, they often overfit in sparse settings, producing floating artifacts and inconsistent geometry. Silhouette consistency is commonly used as a regularizer, but it remains insufficient, as silhouette-consistent regions can extend beyond the true object geometry. We introduce VisDom, a learning-free geometric constraint that augments classical carving-based visual hull reconstruction by enforcing a minimum multi-view visibility requirement. Specifically, we define a visible domain as the subset of 3D space observed by at least $K$ views and use it as an additional filtering criterion on top of standard silhouette-based reconstruction. This provides a stronger spatial prior in sparse-view settings. We integrate VisDom into both implicit (NeRF) and explicit (GS) pipelines by restricting volumetric sampling and guiding Gaussian placement during optimization. Experiments on three challenging datasets show consistent improvements in sparse-view NVS, enabling high-quality object-centric reconstruction from as few as four input images. Our method is domain-agnostic, requires only silhouettes, and introduces no learned parameters, making it a simple complement to existing approaches. Applying VisDom on top of GaussianObject further improves performance on Omni3D and MipNeRF360, while matching or surpassing it at 22 $\times$ lower training cost.

📄 PDF Abstract BibTeX arXiv:2606.20531

Code (0)

등록된 구현이 없습니다.

Tasks

Novel View Synthesis

Similar Papers 제목 키워드 기반

VisDoM: Multi-Document QA with Visually Rich Elements Using Multimodal Retrieval-Augmented Generation

2024-12-14 · Manan Suri, Puneet Mathur, Franck Dernoncourt, Kanika Goswami 외

Understanding information from a collection of multiple documents, particularly those with visually rich elements, is important for document-grounded question answering. This paper introduces VisDoMBench, the first compr…

Question AnsweringRAGRetrievalRetrieval-augmented Generation

Heterogeneous Face Frontalization via Domain Agnostic Learning

2021-07-17 · Xing Di, Shuowen Hu, Vishal M. Patel

Recent advances in deep convolutional neural networks (DCNNs) have shown impressive performance improvements on thermal to visible face synthesis and matching problems. However, current DCNN-based synthesis models do not…

Face GenerationGenerative Adversarial Network

SRUG: Shadow-Guided Relightable Urban Scene with Generation Model

2026-05-23 · Yonghao Zhao, Zexin Yin, Jian Yang, Beibei Wang 외 arxiv

Creating relightable urban scenes from images or videos is widely useful but highly ill-posed. Urban environments are typically unbounded and extend beyond the visible regions. As a result, many portions of the scene rem…

Novel View Synthesis

TrackNeRF: Bundle Adjusting NeRF from Sparse and Noisy Views via Feature Tracks

2024-08-20 · Jinjie Mai, Wenxuan Zhu, Sara Rojas, Jesus Zarzar 외

Neural radiance fields (NeRFs) generally require many images with accurate poses for accurate novel view synthesis, which does not reflect realistic setups where views can be sparse and poses can be noisy. Previous solut…

NeRFNovel View Synthesis

Surgical Visual Domain Adaptation: Results from the MICCAI 2020 SurgVisDom Challenge

2021-02-26 · Aneeq Zia, Kiran Bhattacharyya, Xi Liu, Ziheng Wang 외

Surgical data science is revolutionizing minimally invasive surgery by enabling context-aware applications. However, many challenges exist around surgical data (and health data, more generally) needed to develop context-…

Domain Adaptation