paper-with-me

Papers

Novel-View Acoustic Synthesis

2023-01-20 · CVPR 2023 1 · Changan Chen, Alexander Richard, Roman Shapovalov, Vamsi Krishna Ithapu, Natalia Neverova, Kristen Grauman, Andrea Vedaldi

We introduce the novel-view acoustic synthesis (NVAS) task: given the sight and sound observed at a source viewpoint, can we synthesize the sound of that scene from an unseen target viewpoint? We propose a neural rendering approach: Visually-Guided Acoustic Synthesis (ViGAS) network that learns to synthesize the sound of an arbitrary point in space by analyzing the input audio-visual cues. To benchmark this task, we collect two first-of-their-kind large-scale multi-view audio-visual datasets, one synthetic and one real. We show that our model successfully reasons about the spatial cues and synthesizes faithful audio on both datasets. To our knowledge, this work represents the very first formulation, dataset, and approach to solve the novel-view acoustic synthesis task, which has exciting potential applications ranging from AR/VR to art and design. Unlocked by this work, we believe that the future of novel-view synthesis is in multi-modal learning from videos.

📄 PDF Abstract BibTeX arXiv:2301.08730

Code (0)

등록된 구현이 없습니다.

Tasks

Neural RenderingNovel View Synthesis

Similar Papers 제목 키워드 기반

AV-Surf: Surface-Enhanced Geometry-Aware Novel-View Acoustic Synthesis

2025-03-17 · Hadam Baek, Hannie Shin, Jiyoung Seo, Chanwoo Kim 외

Accurately modeling sound propagation with complex real-world environments is essential for Novel View Acoustic Synthesis (NVAS). While previous studies have leveraged visual perception to estimate spatial acoustics, the…

3DGS

Novel-View Acoustic Synthesis from 3D Reconstructed Rooms

2023-10-23 · Byeongjoo Ahn, Karren Yang, Brian Hamilton, Jonathan Sheaffer 외

We investigate the benefit of combining blind audio recordings with 3D scene information for novel-view acoustic synthesis. Given audio recordings from 2-4 microphones and the 3D geometry and material of a scene containi…

3D geometrySound Source Localization

SonarSplat: Novel View Synthesis of Imaging Sonar via Gaussian Splatting

2025-03-31 · Advaith V. Sethuraman, Max Rucker, Onur Bagoren, Pou-Chun Kung 외

In this paper, we present SonarSplat, a novel Gaussian splatting framework for imaging sonar that demonstrates realistic novel view synthesis and models acoustic streaking phenomena. Our method represents the scene as a …

3D ReconstructionImage GenerationNovel View Synthesis

SoundVista: Novel-View Ambient Sound Synthesis via Visual-Acoustic Binding

2025-04-08 · CVPR 2025 1 · Mingfei Chen, Israel D. Gebru, Ishwarya Ananthabhotla, Christian Richardt 외

We introduce SoundVista, a method to generate the ambient sound of an arbitrary scene at novel viewpoints. Given a pre-acquired recording of the scene from sparsely distributed microphones, SoundVista can synthesize the …

Speaker-Independent Speech-Driven Visual Speech Synthesis using Domain-Adapted Acoustic Models

2019-05-15 · Ahmed Hussen Abdelaziz, Barry-John Theobald, Justin Binder, Gabriele Fanelli 외

Speech-driven visual speech synthesis involves mapping features extracted from acoustic speech to the corresponding lip animation controls for a face model. This mapping can take many forms, but a powerful approach is to…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Face Modelspeech-recognition+2