paper-with-me

Papers

Geometry-guided Online 3D Video Synthesis with Multi-View Temporal Consistency

2025-05-25 · CVPR 2025 1 · Hyunho Ha, Lei Xiao, Christian Richardt, Thu Nguyen-Phuoc, Changil Kim, Min H. Kim, Douglas Lanman, Numair Khan

We introduce a novel geometry-guided online video view synthesis method with enhanced view and temporal consistency. Traditional approaches achieve high-quality synthesis from dense multi-view camera setups but require significant computational resources. In contrast, selective-input methods reduce this cost but often compromise quality, leading to multi-view and temporal inconsistencies such as flickering artifacts. Our method addresses this challenge to deliver efficient, high-quality novel-view synthesis with view and temporal consistency. The key innovation of our approach lies in using global geometry to guide an image-based rendering pipeline. To accomplish this, we progressively refine depth maps using color difference masks across time. These depth maps are then accumulated through truncated signed distance fields in the synthesized view's image space. This depth representation is view and temporally consistent, and is used to guide a pre-trained blending network that fuses multiple forward-rendered input-view images. Thus, the network is encouraged to output geometrically consistent synthesis results across multiple views and time. Our approach achieves consistent, high-quality video synthesis, while running efficiently in an online manner.

📄 PDF Abstract BibTeX arXiv:2505.18932

Code (0)

등록된 구현이 없습니다.

Tasks

Novel View Synthesis

Similar Papers 제목 키워드 기반

Vid-CamEdit: Video Camera Trajectory Editing with Generative Rendering from Estimated Geometry

2025-06-16 · Junyoung Seo, Jisang Han, Jaewoo Jung, Siyoon Jin 외

We introduce Vid-CamEdit, a novel framework for video camera trajectory editing, enabling the re-synthesis of monocular videos along user-defined camera paths. This task is challenging due to its ill-posed nature and the…

Novel View Synthesis

Geometry-Aware Single-Image 4D Synthesis via Dense Trajectory Generation

2025-12-04 · Yanran Zhang, Ziyi Wang, Wenzhao Zheng, Zheng Zhu 외 arxiv

Generating interactive and dynamic 4D scenes from a single static image remains a core challenge. Most existing generate-then-reconstruct and reconstruct-then-generate methods decouple geometry from motion, causing spati…

Boosting Text-Driven Video Segmentation via Geometry-Aware Distillation

2026-06-23 · Tianyu Zhu, Yingping Liang, Hesong Li, Ying Fu arxiv

Text-driven Referring Video Object Segmentation (RVOS) aims to locate and segment target objects in videos given natural language. However, existing models are typically trained on 2D image or video datasets with naive s…

Referring Video Object SegmentationZero-shot GeneralizationImage SegmentationVideo Segmentation

Diff4Splat: Controllable 4D Scene Generation with Latent Dynamic Reconstruction Models

2025-11-01 · Panwang Pan, Chenguo Lin, Jingjing Zhao, Chenxin Li 외 arxiv

We introduce Diff4Splat, a feed-forward method that synthesizes controllable and explicit 4D scenes from a single image. Our approach unifies the generative priors of video diffusion models with geometry and motion const…

Dynamic ReconstructionNovel View SynthesisScene GenerationVideo Generation

Portrait4D-v2: Pseudo Multi-View Data Creates Better 4D Head Synthesizer

2024-03-20 · Yu Deng, Duomin Wang, Baoyuan Wang

In this paper, we propose a novel learning approach for feed-forward one-shot 4D head avatar synthesis. Different from existing methods that often learn from reconstructing monocular videos guided by 3DMM, we employ pseu…