paper-with-me

홈 › Papers

PKU-DyMVHumans: A Multi-View Video Benchmark for High-Fidelity Dynamic Human Modeling

2024-03-24 · CVPR 2024 1 · Xiaoyun Zheng, Liwei Liao, Xufeng Li, Jianbo Jiao, Rongjie Wang, Feng Gao, Shiqi Wang, Ronggang Wang

High-quality human reconstruction and photo-realistic rendering of a dynamic scene is a long-standing problem in computer vision and graphics. Despite considerable efforts invested in developing various capture systems and reconstruction algorithms, recent advancements still struggle with loose or oversized clothing and overly complex poses. In part, this is due to the challenges of acquiring high-quality human datasets. To facilitate the development of these fields, in this paper, we present PKU-DyMVHumans, a versatile human-centric dataset for high-fidelity reconstruction and rendering of dynamic human scenarios from dense multi-view videos. It comprises 8.2 million frames captured by more than 56 synchronized cameras across diverse scenarios. These sequences comprise 32 human subjects across 45 different scenarios, each with a high-detailed appearance and realistic human motion. Inspired by recent advancements in neural radiance field (NeRF)-based scene representations, we carefully set up an off-the-shelf framework that is easy to provide those state-of-the-art NeRF-based implementations and benchmark on PKU-DyMVHumans dataset. It is paving the way for various applications like fine-grained foreground/background decomposition, high-quality human reconstruction and photo-realistic novel view synthesis of a dynamic scene. Extensive studies are performed on the benchmark, demonstrating new observations and challenges that emerge from using such high-fidelity dynamic data.

📄 PDF Abstract BibTeX arXiv:2403.16080

Code (1)

zhengxyun/PKU-DyMVHumans 공식 구현 pytorch

Tasks

NeRFNovel View Synthesis

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

4DAnyone: Create Anyone in 4D from a Casual Monocular Video

2026-08-20 · Yudong Jin, Tao Xie, Qihang Zhang, Zehong Shen 외 arxiv

We present 4DAnyone, a framework for reconstructing 4D humans from an uncalibrated monocular video by generating reconstruction-grade multiview-consistent videos and lifting them into 4D Gaussian Splatting (4DGS). Existi…

AnyView: Synthesizing Any Novel View in Dynamic Scenes

2026-01-23 · Basile Van Hoorick, Dian Chen, Shun Iwase, Pavel Tokmakov 외 arxiv

Modern generative video models excel at producing convincing, high-quality outputs, but struggle to maintain multi-view and spatiotemporal consistency in highly dynamic real-world environments. In this work, we introduce…

Video Generation

CAT4D: Create Anything in 4D with Multi-View Video Diffusion Models

2024-11-27 · CVPR 2025 1 · Rundi Wu, Ruiqi Gao, Ben Poole, Alex Trevithick 외

We present CAT4D, a method for creating 4D (dynamic 3D) scenes from monocular video. CAT4D leverages a multi-view video diffusion model trained on a diverse combination of datasets to enable novel view synthesis at any s…

4D reconstructionNovel View SynthesisScene Generation

A Multi-View Stereo Benchmark With High-Resolution Images and Multi-Camera Videos

2017-07-01 · CVPR 2017 7 · Thomas Schops, Johannes L. Schonberger, Silvano Galliani, Torsten Sattler 외

Motivated by the limitations of existing multi-view stereo benchmarks, we present a novel dataset for this task. Towards this goal, we recorded a variety of indoor and outdoor scenes using a high-precision laser scanner …

ImViD: Immersive Volumetric Videos for Enhanced VR Engagement

2025-03-18 · CVPR 2025 1 · Zhengxian Yang, Shi Pan, Shengqi Wang, Haoxiang Wang 외

User engagement is greatly enhanced by fully immersive multi-modal experiences that combine visual and auditory stimuli. Consequently, the next frontier in VR/AR technologies lies in immersive volumetric videos with comp…