paper-with-me

홈 › Papers

Past, Present, and Future of Spatial Audio and Room Acoustics

2025-03-17 · Shoichi Koyama, Enzo De Sena, Prasanga Samarasinghe, Mark R. P. Thomas, Fabio Antonacci

The study of spatial audio and room acoustics aims to create immersive audio experiences by modeling the physics and psychoacoustics of how sound behaves in space. In the long history of this research area, various key technologies have been developed based both on theoretical advancements and practical innovations. We highlight historical achievements, initiative activities, recent advancements, and future outlooks in the research area of spatial audio recording and reproduction, and room acoustic simulation, modeling, analysis, and control.

📄 PDF Abstract BibTeX arXiv:2503.12948

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Past and Future Motion Guided Network for Audio Visual Event Localization

2022-05-08 · Tingxiu Chen, Jianqin Yin, Jin Tang

In recent years, audio-visual event localization has attracted much attention. It's purpose is to detect the segment containing audio-visual events and recognize the event category from untrimmed videos. Existing methods…

audio-visual event localization

Loss functions incorporating auditory spatial perception in deep learning -- a review

2025-06-24 · Boaz Rafaely, Stefan Weinzierl, Or Berebi, Fabian Brinkmann

Binaural reproduction aims to deliver immersive spatial audio with high perceptual realism over headphones. Loss functions play a central role in optimizing and evaluating algorithms that generate binaural signals. Howev…

OWL: Geometry-Aware Spatial Reasoning for Audio Large Language Models

2025-09-30 · Subrata Biswas, Mohammad Nur Hossain Khan, Bashima Islam arxiv

Spatial reasoning is fundamental to auditory perception, yet current audio large language models (ALLMs) largely rely on unstructured binaural cues and single step inference. This limits both perceptual accuracy in direc…

Spatial Reasoning

DSpAST: Disentangled Representations for Spatial Audio Reasoning with Large Language Models

2025-09-17 · Kevin Wilkinghoff, Zheng-Hua Tan arxiv

Reasoning about spatial audio with large language models requires a spatial audio encoder as an acoustic front-end to obtain audio embeddings for further processing. Such an encoder needs to capture all information requi…

Spatial LibriSpeech: An Augmented Dataset for Spatial Audio Learning

2023-08-18 · Miguel Sarabia, Elena Menyaylenko, Alessandro Toso, Skyler Seto 외

We present Spatial LibriSpeech, a spatial audio dataset with over 650 hours of 19-channel audio, first-order ambisonics, and optional distractor noise. Spatial LibriSpeech is designed for machine learning model training,…

8kPosition