paper-with-me

홈 › Papers

OrbitNVS: Harnessing Video Diffusion Priors for Novel View Synthesis

2026-03-20 · Jinglin Liang, Zijian Zhou, Rui Huang, Shuangping Huang, Yichen Gong arxiv

Novel View Synthesis (NVS) aims to generate unseen views of a 3D object given a limited number of known views. Existing methods often struggle to synthesize plausible views for unobserved regions, particularly under single-view input, and still face challenges in maintaining geometry- and appearance-consistency. To address these issues, we propose OrbitNVS, which reformulates NVS as an orbit video generation task. Through tailored model design and training strategies, we adapt a pre-trained video generation model to the NVS task, leveraging its rich visual priors to achieve high-quality view synthesis. Specifically, we incorporate camera adapters into the video model to enable accurate camera control. To enhance two key properties of 3D objects, geometry and appearance, we design a normal map generation branch and use normal map features to guide the synthesis of the target views via attention mechanism, thereby improving geometric consistency. Moreover, we apply a pixel-space supervision to alleviate blurry appearance caused by spatial compression in the latent space. Extensive experiments show that OrbitNVS significantly outperforms previous methods on the GSO and OmniObject3D benchmarks, especially in the challenging single-view setting (\eg, +2.9 dB and +2.4 dB PSNR).

📄 PDF Abstract BibTeX arXiv:2603.19613

Code (0)

등록된 구현이 없습니다.

Tasks

Novel View SynthesisVideo Generation

Similar Papers 제목 키워드 기반

NVS-Solver: Video Diffusion Model as Zero-Shot Novel View Synthesizer

2024-05-24 · Meng You, Zhiyu Zhu, Hui Liu, Junhui Hou

By harnessing the potent generative capabilities of pre-trained large video diffusion models, we propose NVS-Solver, a new novel view synthesis (NVS) paradigm that operates \textit{without} the need for training. NVS-Sol…

Novel View Synthesis

HIPPo: Harnessing Image-to-3D Priors for Model-free Zero-shot 6D Pose Estimation

2025-02-14 · Yibo Liu, Zhaodong Jiang, Binbin Xu, Guile Wu 외

This work focuses on model-free zero-shot 6D object pose estimation for robotics applications. While existing methods can estimate the precise 6D pose of objects, they heavily rely on curated CAD models or reference imag…

3D Reconstruction6D Pose Estimation6D Pose Estimation using RGBImage to 3D+2

BAGS: Building Animatable Gaussian Splatting from a Monocular Video with Diffusion Priors

2024-03-18 · Tingyang Zhang, Qingzhe Gao, Weiyu Li, Libin Liu 외

Animatable 3D reconstruction has significant applications across various fields, primarily relying on artists' handcraft creation. Recently, some studies have successfully constructed animatable 3D models from monocular …

3D Reconstruction

Zero4D: Training-Free 4D Video Generation From Single Video Using Off-the-Shelf Video Diffusion Model

2025-03-28 · Jangho Park, Taesung Kwon, Jong Chul Ye

Recently, multi-view or 4D video generation has emerged as a significant research topic. Nonetheless, recent approaches to 4D generation still struggle with fundamental limitations, as they primarily rely on harnessing m…

Video Generation

Denoise to Track: Harnessing Video Diffusion Priors for Robust Correspondence

2025-12-04 · Tianyu Yuan, Yuanbo Yang, Lin-Zhuo Chen, Yao Yao 외 arxiv

In this work, we introduce HeFT (Head-Frequency Tracker), a zero-shot point tracking framework that leverages the visual priors of pretrained video diffusion models. To better understand how they encode spatiotemporal in…

Point Tracking