paper-with-me

Papers

AudioWorldSim: Realistic Binaural Audio Datasets For World Models

2026-08-21 · Luis Vitor Zerkowski, Luiz Velho arxiv

This technical report presents AudioWorldSim, an open-source platform designed to generate realistic binaural audio datasets and advance research in audio-based machine learning, particularly world models. Built as a custom extension of Meta's SoundSpaces 2.0 platform, AudioWorldSim leverages their comprehensive acoustics framework, but focuses on the automatic rollout of random agent navigations, as well as implements crucial fixes to how continuous sound is composed. AudioWorldSim is made publicly available to the research community at https://github.com/Luizerko/AudioWorldSim to facilitate reproducibility.

📄 PDF Abstract BibTeX arXiv:2608.21075

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Geometry-Aware Multi-Task Learning for Binaural Audio Generation from Video

2021-11-21 · Rishabh Garg, Ruohan Gao, Kristen Grauman

Binaural audio provides human listeners with an immersive spatial sound experience, but most existing videos lack binaural audio recordings. We propose an audio spatialization method that draws on visual information in v…

Audio GenerationMulti-Task LearningRoom Impulse Response (RIR)

Visually Informed Binaural Audio Generation without Binaural Audios

2021-04-13 · CVPR 2021 1 · Xudong Xu, Hang Zhou, Ziwei Liu, Bo Dai 외

Stereophonic audio, especially binaural audio, plays an essential role in immersive viewing environments. Recent research has explored generating visually guided stereophonic audios supervised by multi-channel audio coll…

Audio Generation

Neural Synthesis of Binaural Audio

2021-01-01 · ICLR 2021 1 · Alexander Richard, Dejan Markovic, Israel D. Gebru, Steven Krenn 외

We present a neural rendering approach for binaural sound synthesis that can produce realistic and spatially accurate binaural sound in realtime. The network takes, as input, a single-channel audio source and synthesizes…

Neural RenderingPosition

BinauralGrad: A Two-Stage Conditional Diffusion Probabilistic Model for Binaural Audio Synthesis

2022-05-30 · Yichong Leng, Zehua Chen, Junliang Guo, Haohe Liu 외

Binaural audio plays a significant role in constructing immersive augmented and virtual realities. As it is expensive to record binaural audio from the real world, synthesizing them from mono audio has attracted increasi…

Audio Synthesis

BinauralFlow: A Causal and Streamable Approach for High-Quality Binaural Speech Synthesis with Flow Matching Models

2025-05-28 · Susan Liang, Dejan Markovic, Israel D. Gebru, Steven Krenn 외

Binaural rendering aims to synthesize binaural audio that mimics natural hearing based on a mono audio and the locations of the speaker and listener. Although many methods have been proposed to solve this problem, they s…

Speech Synthesis