paper-with-me

Papers

Pseudo-Generalized Dynamic View Synthesis from a Video

2023-10-12 · Xiaoming Zhao, Alex Colburn, Fangchang Ma, Miguel Angel Bautista, Joshua M. Susskind, Alexander G. Schwing

Rendering scenes observed in a monocular video from novel viewpoints is a challenging problem. For static scenes the community has studied both scene-specific optimization techniques, which optimize on every test scene, and generalized techniques, which only run a deep net forward pass on a test scene. In contrast, for dynamic scenes, scene-specific optimization techniques exist, but, to our best knowledge, there is currently no generalized method for dynamic novel view synthesis from a given monocular video. To answer whether generalized dynamic novel view synthesis from monocular videos is possible today, we establish an analysis framework based on existing techniques and work toward the generalized approach. We find a pseudo-generalized process without scene-specific appearance optimization is possible, but geometrically and temporally consistent depth estimates are needed. Despite no scene-specific appearance optimization, the pseudo-generalized approach improves upon some scene-specific methods.

📄 PDF Abstract BibTeX arXiv:2310.08587

Code (0)

등록된 구현이 없습니다.

Tasks

Novel View Synthesis

Similar Papers 제목 키워드 기반

Broadening View Synthesis of Dynamic Scenes from Constrained Monocular Videos

2025-12-16 · Le Jiang, Shaotong Zhu, Yedi Luo, Shayda Moezzi 외 arxiv

In dynamic Neural Radiance Fields (NeRF) systems, state-of-the-art novel view synthesis methods often fail under significant viewpoint deviations, producing unstable and unrealistic renderings. To address this, we introd…

Novel View Synthesis

Portrait4D-v2: Pseudo Multi-View Data Creates Better 4D Head Synthesizer

2024-03-20 · Yu Deng, Duomin Wang, Baoyuan Wang

In this paper, we propose a novel learning approach for feed-forward one-shot 4D head avatar synthesis. Different from existing methods that often learn from reconstructing monocular videos guided by 3DMM, we employ pseu…

ReVISE: Self-Supervised Speech Resynthesis with Visual Input for Universal and Generalized Speech Enhancement

2022-12-21 · Wei-Ning Hsu, Tal Remez, Bowen Shi, Jacob Donley 외

Prior works on improving speech quality with visual input typically study each type of auditory distortion separately (e.g., separation, inpainting, video-to-speech) and present tailored algorithms. This paper proposes t…

Audio-Visual Speech RecognitionResynthesisSpeech Enhancementspeech-recognition+7

ReVISE: Self-Supervised Speech Resynthesis With Visual Input for Universal and Generalized Speech Regeneration

2023-01-01 · CVPR 2023 1 · Wei-Ning Hsu, Tal Remez, Bowen Shi, Jacob Donley 외

Prior works on improving speech quality with visual input typically study each type of auditory distortion separately (e.g., separation, inpainting, video-to-speech) and present tailored algorithms. This paper propos…

Audio-Visual Speech RecognitionResynthesisspeech-recognitionSpeech Recognition+6

GS-DiT: Advancing Video Generation with Pseudo 4D Gaussian Fields through Efficient Dense 3D Point Tracking

2025-01-05 · Weikang Bian, Zhaoyang Huang, Xiaoyu Shi, Yijin Li 외

4D video control is essential in video generation as it enables the use of sophisticated lens techniques, such as multi-camera shooting and dolly zoom, which are currently unsupported by existing methods. Training a vide…

Novel View SynthesisPoint TrackingVideo Generation