paper-with-me

Papers

A Synchronized Audio-Visual Multi-View Capture System

2026-03-24 · Xiangwei Shi, Gara Dorta, Ruud de Jong, Ojas Shirekar, Chirag Raman arxiv

Multi-view capture systems have been an important tool in research for recording human motion under controlling conditions. Most existing systems are specified around video streams and provide little or no support for audio acquisition and rigorous audio-video alignment, despite both being essential for studying conversational interaction where timing at the level of turn-taking, overlap, and prosody matters. In this technical report, we describe an audio-visual multi-view capture system that addresses this gap by treating synchronized audio and synchronized video as first-class signals. The system combines a multi-camera pipeline with multi-channel microphone recording under a unified timing architecture and provides a practical workflow for calibration, acquisition, and quality control that supports repeatable recordings at scale. We quantify synchronization performance in deployment and show that the resulting recordings are temporally consistent enough to support fine-grained analysis and data-driven modeling of conversation behavior.

📄 PDF Abstract BibTeX arXiv:2603.23089

Code (0)

등록된 구현이 없습니다.

Tasks

Video Alignment

Similar Papers 제목 키워드 기반

ImViD: Immersive Volumetric Videos for Enhanced VR Engagement

2025-03-18 · CVPR 2025 1 · Zhengxian Yang, Shi Pan, Shengqi Wang, Haoxiang Wang 외

User engagement is greatly enhanced by fully immersive multi-modal experiences that combine visual and auditory stimuli. Consequently, the next frontier in VR/AR technologies lies in immersive volumetric videos with comp…

DiVa-360: The Dynamic Visual Dataset for Immersive Neural Fields

2023-07-31 · CVPR 2024 1 · Cheng-You Lu, Peisen Zhou, Angela Xing, Chandradeep Pokhariya 외

Advances in neural fields are enabling high-fidelity capture of the shape and appearance of dynamic 3D scenes. However, their capabilities lag behind those offered by conventional representations such as 2D videos becaus…

Audio-Synchronized Visual Animation

2024-03-08 · Lin Zhang, Shentong Mo, Yijing Zhang, Pedro Morgado

Current visual generation methods can produce high quality videos guided by texts. However, effectively controlling object dynamics remains a challenge. This work explores audio as a cue to generate temporally synchroniz…

Realizing Immersive Volumetric Video: A Multimodal Framework for 6-DoF VR Engagement

2026-04-10 · Zhengxian Yang, Shengqi Wang, Shi Pan, Hongshuai Li 외 arxiv

Fully immersive experiences that tightly integrate 6-DoF visual and auditory interaction are essential for virtual and augmented reality. While such experiences can be achieved through computer-generated content, constru…

Diff-Foley: Synchronized Video-to-Audio Synthesis with Latent Diffusion Models

2023-06-29 · NeurIPS 2023 11 · Simian Luo, Chuanhao Yan, Chenxu Hu, Hang Zhao

The Video-to-Audio (V2A) model has recently gained attention for its practical application in generating audio directly from silent videos, particularly in video/film production. However, previous methods in V2A have lim…

Audio Synthesis