paper-with-me

Papers

SOAF: Scene Occlusion-aware Neural Acoustic Field

2024-07-02 · Huiyu Gao, Jiahao Ma, David Ahmedt-Aristizabal, Chuong Nguyen, Miaomiao Liu

This paper tackles the problem of novel view audio-visual synthesis along an arbitrary trajectory in an indoor scene, given the audio-video recordings from other known trajectories of the scene. Existing methods often overlook the effect of room geometry, particularly wall occlusions on sound propagation, making them less accurate in multi-room environments. In this work, we propose a new approach called Scene Occlusion-aware Acoustic Field (SOAF) for accurate sound generation. Our approach derives a global prior for the sound field using distance-aware parametric sound-propagation modeling and then transforms it based on the scene structure learned from the input video. We extract features from the local acoustic field centered at the receiver using a Fibonacci Sphere to generate binaural audio for novel views with a direction-aware attention mechanism. Extensive experiments on the real dataset RWAVS and the synthetic dataset SoundSpaces demonstrate that our method outperforms previous state-of-the-art techniques in audio generation.

📄 PDF Abstract BibTeX arXiv:2407.02264

Code (0)

등록된 구현이 없습니다.

Tasks

Audio Generation

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음

Similar Papers 제목 키워드 기반

Scene-aware Far-field Automatic Speech Recognition

2021-04-21 · Zhenyu Tang, Dinesh Manocha

We propose a novel method for generating scene-aware training data for far-field automatic speech recognition. We use a deep learning-based estimator to non-intrusively compute the sub-band reverberation time of an envir…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Emotion and Theme Recognition in Music with Frequency-Aware RF-Regularized CNNs

2019-10-28 · Khaled Koutini, Shreyan Chowdhury, Verena Haunschmid, Hamid Eghbal-zadeh 외

We present CP-JKU submission to MediaEval 2019; a Receptive Field-(RF)-regularized and Frequency-Aware CNN approach for tagging music with emotion/mood labels. We perform an investigation regarding the impact of the RF o…

Acoustic Scene ClassificationScene Classification

Neural Rays for Occlusion-aware Image-based Rendering

2021-07-28 · CVPR 2022 1 · YuAn Liu, Sida Peng, Lingjie Liu, Qianqian Wang 외

We present a new neural representation, called Neural Ray (NeuRay), for the novel view synthesis task. Recent works construct radiance fields from image features of input views to render novel view images, which enables …

Neural RenderingNovel View SynthesisStereo Matching

AV-NeRF: Learning Neural Fields for Real-World Audio-Visual Scene Synthesis

2023-02-04 · NeurIPS 2023 11

Can machines recording an audio-visual scene produce realistic, matching audio-visual experiences at novel positions and novel view directions? We answer it by studying a new task -- real-world audio-visual scene synthes…

3D geometryAudio GenerationNeRF

Towards Nonlinear-Motion-Aware and Occlusion-Robust Rolling Shutter Correction

2023-03-31 · ICCV 2023 1 · Delin Qu, Yizhen Lao, Zhigang Wang, Dong Wang 외

This paper addresses the problem of rolling shutter correction in complex nonlinear and dynamic scenes with extreme occlusion. Existing methods suffer from two main drawbacks. Firstly, they face challenges in estimating …

Rolling Shutter Correction