paper-with-me

Papers Audio Signal Processing

“Audio Signal Processing” 태그가 달린 논문 70편 · 필터 해제

DiffusionRIR: Room Impulse Response Interpolation using Diffusion Models

2025-04-29 · Sagi Della Torre, Mirco Pezzoli, Fabio Antonacci, Sharon Gannot

Room Impulse Responses (RIRs) characterize acoustic environments and are crucial in multiple audio signal processing tasks. High-quality RIR estimates drive applications such as virtual microphones, sound source localiza…

Audio Signal ProcessingData AugmentationDenoisingImage Inpainting+1

TorchFX: A modern approach to Audio DSP with PyTorch and GPU acceleration

2025-04-11 · Matteo Spanio, Antonio Rodà

The burgeoning complexity and real-time processing demands of audio signals necessitate optimized algorithms that harness the computational prowess of Graphics Processing Units (GPUs). Existing Digital Signal Processing …

Audio Signal ProcessingBenchmarkingGPU

Sound Field Estimation: Theories and Applications

2025-03-13 · Natsuki Ueno, Shoichi Koyama

The spatial information of sound plays a crucial role in various situations, ranging from daily activities to advanced engineering technologies. To fully utilize its potential, numerous research studies on spatial audio …

Audio Signal Processing

Bridging The Multi-Modality Gaps of Audio, Visual and Linguistic for Speech Enhancement

2025-01-23 · Meng-Ping Lin, Jen-Cheng Hou, Chia-Wei Chen, Shao-Yi Chien 외

Speech enhancement (SE) aims to improve the quality and intelligibility of speech in noisy environments. Recent studies have shown that incorporating visual cues in audio signal processing can enhance SE performance. Giv…

Audio Signal ProcessingSpeech EnhancementTransfer Learning

Comparative Analysis of Mel-Frequency Cepstral Coefficients and Wavelet Based Audio Signal Processing for Emotion Detection and Mental Health Assessment in Spoken Speech

2024-12-12 · Idoko Agbo, Dr Hoda El-Sayed, M. D Kamruzzan Sarker

The intersection of technology and mental health has spurred innovative approaches to assessing emotional well-being, particularly through computational techniques applied to audio data analysis. This study explores the …

Audio Signal ProcessingData Augmentation

Detecting abnormal heart sound using mobile phones and on-device IConNet

2024-12-04 · Linh Vu, Thu Tran

Given the global prevalence of cardiovascular diseases, there is a pressing need for easily accessible early screening methods. Typically, this requires medical practitioners to investigate heart auscultations for irregu…

Audio Signal Processing

Manikin-Recorded Cardiopulmonary Sounds Dataset Using Digital Stethoscope

2024-10-04 · Yasaman Torabi, Shahram Shirani, James P. Reilly

Heart and lung sounds are crucial for healthcare monitoring. Recent improvements in stethoscope technology have made it possible to capture patient sounds with enhanced precision. In this dataset, we used a digital steth…

Audio Signal ProcessingSound Classification

Blind Localization of Early Room Reflections with Arbitrary Microphone Array

2024-09-23 · Yogev Hadadi, Vladimir Tourbabin, Zamir Ben-Hur, David Lou Alon 외

Blindly estimating the direction of arrival (DoA) of early room reflections without prior knowledge of the room impulse response or source signal is highly valuable in audio signal processing applications. The FF-PHALCOR…

Audio Signal Processing

Audio-Driven Reinforcement Learning for Head-Orientation in Naturalistic Environments

2024-09-16 · Wessel Ledder, Yuzhen Qin, Kiki van der Heijden

Although deep reinforcement learning (DRL) approaches in audio signal processing have seen substantial progress in recent years, audio-driven DRL for tasks such as navigation, gaze control and head-orientation control in…

Audio Signal ProcessingDeep Reinforcement LearningQ-Learningreinforcement-learning+1

Classification of Heart Sounds Using Multi-Branch Deep Convolutional Network and LSTM-CNN

2024-07-15 · Seyed Amir Latifi, Hassan Ghassemian, Maryam Imani

Cardiovascular diseases represent a leading cause of mortality worldwide, necessitating accurate and early diagnosis for improved patient outcomes. Current diagnostic approaches for cardiac abnormalities often present ch…

Audio Signal ProcessingBinary ClassificationDiagnostic

Spectral Mapping of Singing Voices: U-Net-Assisted Vocal Segmentation

2024-05-30 · Adam Sorrenti

Separating vocal elements from musical tracks is a longstanding challenge in audio signal processing. This study tackles the distinct separation of vocal components from musical spectrograms. We employ the Short Time Fou…

Audio Signal ProcessingAudio Source Separation

AudioSetMix: Enhancing Audio-Language Datasets with LLM-Assisted Augmentations

2024-05-17 · David Xu

Multi-modal learning in the audio-language domain has seen significant advancements in recent years. However, audio-language learning faces challenges due to limited and lower-quality data compared to image-language task…

Audio Signal ProcessingLanguage ModelingLanguage ModellingLarge Language Model

Comparative Study of State-based Neural Networks for Virtual Analog Audio Effects Modeling

2024-05-07 · Riccardo Simionato, Stefano Fasciani

Analog electronic circuits are at the core of an important category of musical devices, which includes a broad range of sound synthesizers and audio effects. The development of software that simulates analog musical devi…

Audio Effects ModelingAudio Signal ProcessingState Space Models

Overview of the L3DAS23 Challenge on Audio-Visual Extended Reality

2024-02-14 · Christian Marinoni, Riccardo Fosco Gramaccioni, Changan Chen, Aurelio Uncini 외

The primary goal of the L3DAS23 Signal Processing Grand Challenge at ICASSP 2023 is to promote and support collaborative research on machine learning for 3D audio signal processing, with a specific emphasis on 3D speech …

Audio Signal ProcessingSound Event Localization and DetectionSpeech Enhancement

A Survey on Data Augmentation in Large Model Era

2024-01-27 · Yue Zhou, Chenlu Guo, Xu Wang, Yi Chang 외

Large models, encompassing large language and diffusion models, have shown exceptional promise in approximating human-level intelligence, garnering significant interest from both academic and industrial spheres. However,…

Audio Signal ProcessingData AugmentationImage AugmentationSurvey+1

HAAQI-Net: A Non-intrusive Neural Music Audio Quality Assessment Model for Hearing Aids

2024-01-02 · Dyah A. M. G. Wisnu, Stefano Rini, Ryandhimas E. Zezario, Hsin-Min Wang 외

This paper introduces HAAQI-Net, a non-intrusive deep learning-based music audio quality assessment model for hearing aid users. Unlike traditional methods like the Hearing Aid Audio Quality Index (HAAQI) that require in…

Audio Quality AssessmentAudio Signal ProcessingComputational EfficiencyKnowledge Distillation+1

Unsupervised Harmonic Parameter Estimation Using Differentiable DSP and Spectral Optimal Transport

2023-12-22 · Bernardo Torres, Geoffroy Peeters, Gaël Richard

In neural audio signal processing, pitch conditioning has been used to enhance the performance of synthesizers. However, jointly training pitch estimators and synthesizers is a challenge when using standard audio-to-audi…

Audio Signal Processingparameter estimation

Study of speaker localization under dynamic and reverberant environments

2023-11-28 · Daniel A. Mitchell, Boaz Rafaely

Speaker localization in a reverberant environment is a fundamental problem in audio signal processing. Many solutions have been developed to tackle this problem. However, previous algorithms typically assume a stationary…

Audio Signal Processing

HPCNeuroNet: Advancing Neuromorphic Audio Signal Processing with Transformer-Enhanced Spiking Neural Networks

2023-11-21 · Murat Isik, Hiruna Vishwamith, Kayode Inadagbo, I. Can Dikmen

This paper presents a novel approach to neuromorphic audio processing by integrating the strengths of Spiking Neural Networks (SNNs), Transformers, and high-performance computing (HPC) into the HPCNeuroNet architecture. …

Audio Signal ProcessingCPUGPU

Neural Harmonium: An Interpretable Deep Structure for Nonlinear Dynamic System Identification with Application to Audio Processing

2023-10-10 · Karim Helwani, Erfan Soltanmohammadi, Michael M. Goodwin

Improving the interpretability of deep neural networks has recently gained increased attention, especially when the power of deep learning is leveraged to solve problems in physics. Interpretability helps us understand a…

Acoustic echo cancellationAudio Signal Processing
1–20 / 70 다음 →