paper-with-me

홈 › Papers

ImViD: Immersive Volumetric Videos for Enhanced VR Engagement

2025-03-18 · CVPR 2025 1 · Zhengxian Yang, Shi Pan, Shengqi Wang, Haoxiang Wang, Li Lin, Guanjun Li, Zhengqi Wen, Borong Lin, JianHua Tao, Tao Yu

User engagement is greatly enhanced by fully immersive multi-modal experiences that combine visual and auditory stimuli. Consequently, the next frontier in VR/AR technologies lies in immersive volumetric videos with complete scene capture, large 6-DoF interaction space, multi-modal feedback, and high resolution & frame-rate contents. To stimulate the reconstruction of immersive volumetric videos, we introduce ImViD, a multi-view, multi-modal dataset featuring complete space-oriented data capture and various indoor/outdoor scenarios. Our capture rig supports multi-view video-audio capture while on the move, a capability absent in existing datasets, significantly enhancing the completeness, flexibility, and efficiency of data capture. The captured multi-view videos (with synchronized audios) are in 5K resolution at 60FPS, lasting from 1-5 minutes, and include rich foreground-background elements, and complex dynamics. We benchmark existing methods using our dataset and establish a base pipeline for constructing immersive volumetric videos from multi-view audiovisual inputs for 6-DoF multi-modal immersive VR experiences. The benchmark and the reconstruction and interaction results demonstrate the effectiveness of our dataset and baseline method, which we believe will stimulate future research on immersive volumetric video production.

📄 PDF Abstract BibTeX arXiv:2503.14359

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

BASE 설명 없음

Similar Papers 제목 키워드 기반

Realizing Immersive Volumetric Video: A Multimodal Framework for 6-DoF VR Engagement

2026-04-10 · Zhengxian Yang, Shengqi Wang, Shi Pan, Hongshuai Li 외 arxiv

Fully immersive experiences that tightly integrate 6-DoF visual and auditory interaction are essential for virtual and augmented reality. While such experiences can be achieved through computer-generated content, constru…

RePerformer: Immersive Human-centric Volumetric Videos from Playback to Photoreal Reperformance

2025-01-01 · CVPR 2025 1 · Yuheng Jiang, Zhehao Shen, Chengcheng Guo, Yu Hong 외

Human-centric volumetric videos offer immersive free-viewpoint experiences, yet existing methods focus either on replaying general dynamic scenes or animating human avatars, limiting their ability to re-perform gener…

AttributePosition

Focus360: Guiding User Attention in Immersive Videos for VR

2026-01-29 · Paulo Vitor S. Silva, Lucas L. Neves, Rafael A. Goiás, Diogo F. C. Silva 외 arxiv

This demo introduces Focus360, a system designed to enhance user engagement in 360° VR videos by guiding attention to key elements within the scene. Using natural language descriptions, the system identifies important el…

FSVVD: A Dataset of Full Scene Volumetric Video

2023-03-07 · Kaiyuan Hu, Yili Jin, Haowen Yang, Junhua Liu 외

Recent years have witnessed a rapid development of immersive multimedia which bridges the gap between the real world and virtual space. Volumetric videos, as an emerging representative 3D video paradigm that empowers ext…

ImVideoEdit: Image-learning Video Editing via 2D Spatial Difference Attention Blocks

2026-04-09 · Jiayang Xu, Fan Zhuo, Majun Zhang, Changhao Pan 외 arxiv

Current video editing models often rely on expensive paired video data, which limits their practical scalability. In essence, most video editing tasks can be formulated as a decoupled spatiotemporal process, where the te…