paper-with-me

Papers

Deep learning based spatial aliasing reduction in beamforming for audio capture

2025-05-26 · Mateusz Guzik, Giulio Cengarle, Daniel Arteaga

Spatial aliasing affects spaced microphone arrays, causing directional ambiguity above certain frequencies, degrading spatial and spectral accuracy of beamformers. Given the limitations of conventional signal processing and the scarcity of deep learning approaches to spatial aliasing mitigation, we propose a novel approach using a U-Net architecture to predict a signal-dependent de-aliasing filter, which reduces aliasing in conventional beamforming for spatial capture. Two types of multichannel filters are considered, one which treats the channels independently and a second one that models cross-channel dependencies. The proposed approach is evaluated in two common spatial capture scenarios: stereo and first-order Ambisonics. The results indicate a very significant improvement, both objective and perceptual, with respect to conventional beamforming. This work shows the potential of deep learning to reduce aliasing in beamforming, leading to improvements in multi-microphone setups.

📄 PDF Abstract BibTeX arXiv:2505.19781

Code (0)

등록된 구현이 없습니다.

Tasks

De-aliasingDeep Learning

Methods 이 논문이 사용한 방법론

Concatenated Skip Connection A Concatenated Skip Connection is a type of skip connection that seeks to reuse features by concatenating them to new layers, allowing more information to be retained from…
Max Pooling Max Pooling is a pooling operation that calculates the maximum value for patches of a feature map, and uses it to create a downsampled (pooled) feature map. It is usually…
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
U-Net 설명 없음

Similar Papers 제목 키워드 기반

EchoingPixels: Aliasing-Resistant Joint Token Reduction for Audio-Visual LLMs

2025-12-11 · Chao Gong, Depeng Wang, Zhipeng Wei, Ya Guo 외 arxiv

Audio-Visual Large Language Models (AV-LLMs) face prohibitive computational costs of processing massive, redundant audio-visual tokens. Existing unimodal compression techniques fail to capture the heterogeneous and mutua…

Sparse Learning

Beamforming-LLM: What, Where and When Did I Miss?

2025-09-07 · Vishal Choudhari arxiv

We present Beamforming-LLM, a system that enables users to semantically recall conversations they may have missed in multi-speaker environments. The system combines spatial audio capture using a microphone array with ret…

Natural Language QueriesMeeting Summarization

Aliasing Reduction in Neural Amp Modeling by Smoothing Activations

2025-05-07 · Ryota Sato, Julius O. Smith III

The increasing demand for high-quality digital emulations of analog audio hardware, such as vintage tube guitar amplifiers, led to numerous works on neural network-based black-box modeling, with deep learning architectur…

Aliasing Detection and Reduction in Plenoptic Imaging

2014-06-01 · CVPR 2014 6 · Zhaolin Xiao, Qing Wang, Guoqing Zhou, Jingyi Yu

When using plenoptic camera for digital refocusing, angular undersampling can cause severe (angular) aliasing artifacts. Previous approaches have focused on avoiding aliasing by pre-processing the acquired light field vi…

Demosaicking

A spherical harmonic-domain spatial audio signal enhancement method based on minimum variance distortionless response

2024-09-05 · Huawei Zhang, Jihui, Zhang, Huiyuan 외

Spatial audio signal enhancement aims to reduce interfering source contributions while preserving the desired sound field with its spatial cues intact. Existing methods generally rely on impractical assumptions (e.g. no …