paper-with-me

홈 › Papers

SANN-PSZ: Spatially Adaptive Neural Network for Head-Tracked Personal Sound Zones

2024-11-01 · Yue Qiao, Edgar Choueiri

A deep learning framework for dynamically rendering personal sound zones (PSZs) with head tracking is presented, utilizing a spatially adaptive neural network (SANN) that inputs listeners' head coordinates and outputs PSZ filter coefficients. The SANN model is trained using either simulated acoustic transfer functions (ATFs) with data augmentation for robustness in uncertain environments or a mix of simulated and measured ATFs for customization under known conditions. It is found that augmenting room reflections in the training data can more effectively improve the model robustness than augmenting the system imperfections, and that adding constraints such as filter compactness to the loss function does not significantly affect the model's performance. Comparisons of the best-performing model with traditional filter design methods show that, when no measured ATFs are available, the model yields equal or higher isolation in an actual room environment with fewer filter artifacts. Furthermore, the model achieves significant data compression (100x) and computational efficiency (10x) compared to the traditional methods, making it suitable for real-time rendering of PSZs that adapt to the listeners' head movements.

📄 PDF Abstract BibTeX arXiv:2411.00772

Code (0)

등록된 구현이 없습니다.

Tasks

Computational EfficiencyData AugmentationData Compression

Similar Papers 제목 키워드 기반

Multi-person Spatial Interaction in a Large Immersive Display Using Smartphones as Touchpads

2019-11-26 · Gyanendra Sharma, Richard J. Radke

In this paper, we present a multi-user interaction interface for a large immersive space that supports simultaneous screen interactions by combining (1) user input via personal smartphones and Bluetooth microphones, (2) …

Inject Where It Matters: Training-Free Spatially-Adaptive Identity Preservation for Text-to-Image Personalization

2026-02-15 · Guandong Li, Mengxia Ye arxiv

Personalized text-to-image generation aims to integrate specific identities into arbitrary contexts. However, existing tuning-free methods typically employ Spatially Uniform Visual Injection, causing identity features to…

Text-to-Image Generation

Stiffness-aware neural network for learning Hamiltonian systems

2021-09-29 · ICLR 2022 4 · Senwei Liang, Zhongzhan Huang, Hong Zhang

We propose stiffness-aware neural network (SANN), a new method for learning Hamiltonian dynamical systems from data. SANN identifies and splits the training data into stiff and nonstiff portions based on a stiffness-awar…

Wispy to Voluminous: Prior-free Multi-view Capture of Strand-level Facial Hair

2026-06-06 · Jaeseong Lee, Giljoo Nam, Adrian Jarabo, Carlos Aliaga arxiv

Facial hair is a defining trait of personal identity, yet remains a critical bottleneck for digital avatars. Recent volumetric methods achieve photorealism but bake hair into the underlying face geometry, preventing edit…

Rethinking Spatially-Adaptive Normalization

2020-04-06 · Zhentao Tan, Dongdong Chen, Qi Chu, Menglei Chai 외

Spatially-adaptive normalization is remarkably successful recently in conditional semantic image synthesis, which modulates the normalized activation with spatially-varying transformations learned from semantic layouts, …

Image Generation