paper-with-me

홈 › Papers

Common Pets in 3D: Dynamic New-View Synthesis of Real-Life Deformable Categories

2022-11-07 · CVPR 2023 1 · Samarth Sinha, Roman Shapovalov, Jeremy Reizenstein, Ignacio Rocco, Natalia Neverova, Andrea Vedaldi, David Novotny

Obtaining photorealistic reconstructions of objects from sparse views is inherently ambiguous and can only be achieved by learning suitable reconstruction priors. Earlier works on sparse rigid object reconstruction successfully learned such priors from large datasets such as CO3D. In this paper, we extend this approach to dynamic objects. We use cats and dogs as a representative example and introduce Common Pets in 3D (CoP3D), a collection of crowd-sourced videos showing around 4,200 distinct pets. CoP3D is one of the first large-scale datasets for benchmarking non-rigid 3D reconstruction "in the wild". We also propose Tracker-NeRF, a method for learning 4D reconstruction from our dataset. At test time, given a small number of video frames of an unseen object, Tracker-NeRF predicts the trajectories of its 3D points and generates new views, interpolating viewpoint and time. Results on CoP3D reveal significantly better non-rigid new-view synthesis performance than existing baselines.

📄 PDF Abstract BibTeX arXiv:2211.03889

Code (0)

등록된 구현이 없습니다.

Tasks

3D Reconstruction4D reconstructionBenchmarkingNeRFObject Reconstruction

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

Artemis: Articulated Neural Pets with Appearance and Motion synthesis

2022-02-11 · Haimin Luo, Teng Xu, Yuheng Jiang, Chenglin Zhou 외

We, humans, are entering into a virtual era and indeed want to bring animals to the virtual world as well for companion. Yet, computer-generated (CGI) furry animals are limited by tedious off-line rendering, let alone in…

Motion Synthesis

Total-Recon: Deformable Scene Reconstruction for Embodied View Synthesis

2023-04-24 · ICCV 2023 1 · Chonghyuk Song, Gengshan Yang, Kangle Deng, Jun-Yan Zhu 외

We explore the task of embodied view synthesis from monocular videos of deformable scenes. Given a minute-long RGBD video of people interacting with their pets, we render the scene from novel camera trajectories derived …

FPETS : Fully Parallel End-to-End Text-to-Speech System

2018-12-12 · Dabiao Ma, Zhiba Su, Wenxuan Wang, Yuhao Lu

End-to-end Text-to-speech (TTS) system can greatly improve the quality of synthesised speech. But it usually suffers form high time latency due to its auto-regressive structure. And the synthesised speech may also suffer…

text-to-speechText to Speech

Advances in Feed-Forward 3D Reconstruction and View Synthesis: A Survey

2025-07-19 · Jiahui Zhang, Yuelei Li, Anpei Chen, Muyu Xu 외 arxiv

3D reconstruction and view synthesis are foundational problems in computer vision, graphics, and immersive technologies such as augmented reality (AR), virtual reality (VR), and digital twins. Traditional methods rely on…

3D Reconstruction

DBMovi-GS: Dynamic View Synthesis from Blurry Monocular Video via Sparse-Controlled Gaussian Splatting

2025-06-26 · Yeon-Ji Song, Jaein Kim, Byung-Ju Kim, Byoung-Tak Zhang

Novel view synthesis is a task of generating scenes from unseen perspectives; however, synthesizing dynamic scenes from blurry monocular videos remains an unresolved challenge that has yet to be effectively addressed. Ex…

3D geometryNovel View Synthesis