paper-with-me

Papers

MetaView: Monocular Novel View Synthesis with Scale-Aware Implicit Geometry Priors

2026-07-13 · Yufei Cai, Xuesong Niu, Hao Lu, Kun Gai, Kai Wu, Guosheng Lin hf

Current visual generation models are capable of producing high-quality content, yet they lack a coherent perception of the spatial structure. Existing generative novel view synthesis methods typically introduce explicit geometry priors, which enforce spatial consistency but inherently restrict generalization in large view changes. In contrast, recent interactive generative methods favor implicit scene modeling, offering greater flexibility at the cost of precise camera control and geometry consistency. In this paper, we propose MetaView, a diffusion-based monocular novel view synthesis framework that enables rendering under large view changes from a single image. Our key insight is to combine implicit geometry modeling with minimal yet essential explicit 3D cues: we incorporate implicit geometry priors from a feed-forward geometry perception network to regularize structure without imposing restrictive reconstruction pipelines, while leveraging metric depth to anchor the generation to a metric scale. This design allows MetaView to achieve both geometry consistency and precise controllability. Extensive experiments demonstrate that, under challenging monocular large viewpoint changes, MetaView significantly outperforms existing methods and exhibits superior generalization. Our code is publicly available at https://github.com/KlingAIResearch/MetaView.

📄 PDF Abstract BibTeX arXiv:2607.12000

Code (0)

등록된 구현이 없습니다.

Tasks

Novel View Synthesis

Similar Papers 제목 키워드 기반

MetaViewer: Towards A Unified Multi-View Representation

2023-03-11 · CVPR 2023 1 · Ren Wang, Haoliang Sun, Yuling Ma, Xiaoming Xi 외

Existing multi-view representation learning methods typically follow a specific-to-uniform pipeline, extracting latent features from each view and then fusing or aligning them to obtain the unified object representation.…

MULTI-VIEW LEARNINGRepresentation Learning

DBMovi-GS: Dynamic View Synthesis from Blurry Monocular Video via Sparse-Controlled Gaussian Splatting

2025-06-26 · Yeon-Ji Song, Jaein Kim, Byung-Ju Kim, Byoung-Tak Zhang

Novel view synthesis is a task of generating scenes from unseen perspectives; however, synthesizing dynamic scenes from blurry monocular videos remains an unresolved challenge that has yet to be effectively addressed. Ex…

3D geometryNovel View Synthesis

Multi-Multi-View Learning: Multilingual and Multi-Representation Entity Typing

2018-10-24 · EMNLP 2018 10 · Yadollah Yaghoobzadeh, Hinrich Schütze

Knowledge bases (KBs) are paramount in NLP. We employ multiview learning for increasing accuracy and coverage of entity type information in KBs. We rely on two metaviews: language and representation. For language, we con…

Entity TypingMultiview LearningMULTI-VIEW LEARNING

Learning Detailed Radiance Manifolds for High-Fidelity and 3D-Consistent Portrait Synthesis from Monocular Image

2022-11-25 · CVPR 2023 1 · Yu Deng, Baoyuan Wang, Heung-Yeung Shum

A key challenge for novel view synthesis of monocular portrait images is 3D consistency under continuous pose variations. Most existing methods rely on 2D generative models which often leads to obvious 3D inconsistency a…

Image GenerationNovel View Synthesis

Few-shot Novel View Synthesis using Depth Aware 3D Gaussian Splatting

2024-10-14 · Raja Kumar, Vanshika Vats

3D Gaussian splatting has surpassed neural radiance field methods in novel view synthesis by achieving lower computational costs and real-time high-quality rendering. Although it produces a high-quality rendering with a …

3DGSDepth EstimationDepth PredictionNovel View Synthesis