Papers Audio Signal Processing
“Audio Signal Processing” 태그가 달린 논문 70편 · 필터 해제
DiffusionRIR: Room Impulse Response Interpolation using Diffusion Models
Room Impulse Responses (RIRs) characterize acoustic environments and are crucial in multiple audio signal processing tasks. High-quality RIR estimates drive applications such as virtual microphones, sound source localiza…
Audio Signal ProcessingData AugmentationDenoisingImage Inpainting+1TorchFX: A modern approach to Audio DSP with PyTorch and GPU acceleration
The burgeoning complexity and real-time processing demands of audio signals necessitate optimized algorithms that harness the computational prowess of Graphics Processing Units (GPUs). Existing Digital Signal Processing …
Audio Signal ProcessingBenchmarkingGPUSound Field Estimation: Theories and Applications
The spatial information of sound plays a crucial role in various situations, ranging from daily activities to advanced engineering technologies. To fully utilize its potential, numerous research studies on spatial audio …
Audio Signal ProcessingBridging The Multi-Modality Gaps of Audio, Visual and Linguistic for Speech Enhancement
Speech enhancement (SE) aims to improve the quality and intelligibility of speech in noisy environments. Recent studies have shown that incorporating visual cues in audio signal processing can enhance SE performance. Giv…
Audio Signal ProcessingSpeech EnhancementTransfer LearningComparative Analysis of Mel-Frequency Cepstral Coefficients and Wavelet Based Audio Signal Processing for Emotion Detection and Mental Health Assessment in Spoken Speech
The intersection of technology and mental health has spurred innovative approaches to assessing emotional well-being, particularly through computational techniques applied to audio data analysis. This study explores the …
Audio Signal ProcessingData AugmentationDetecting abnormal heart sound using mobile phones and on-device IConNet
Given the global prevalence of cardiovascular diseases, there is a pressing need for easily accessible early screening methods. Typically, this requires medical practitioners to investigate heart auscultations for irregu…
Audio Signal ProcessingManikin-Recorded Cardiopulmonary Sounds Dataset Using Digital Stethoscope
Heart and lung sounds are crucial for healthcare monitoring. Recent improvements in stethoscope technology have made it possible to capture patient sounds with enhanced precision. In this dataset, we used a digital steth…
Audio Signal ProcessingSound ClassificationBlind Localization of Early Room Reflections with Arbitrary Microphone Array
Blindly estimating the direction of arrival (DoA) of early room reflections without prior knowledge of the room impulse response or source signal is highly valuable in audio signal processing applications. The FF-PHALCOR…
Audio Signal ProcessingAudio-Driven Reinforcement Learning for Head-Orientation in Naturalistic Environments
Although deep reinforcement learning (DRL) approaches in audio signal processing have seen substantial progress in recent years, audio-driven DRL for tasks such as navigation, gaze control and head-orientation control in…
Audio Signal ProcessingDeep Reinforcement LearningQ-Learningreinforcement-learning+1Classification of Heart Sounds Using Multi-Branch Deep Convolutional Network and LSTM-CNN
Cardiovascular diseases represent a leading cause of mortality worldwide, necessitating accurate and early diagnosis for improved patient outcomes. Current diagnostic approaches for cardiac abnormalities often present ch…
Audio Signal ProcessingBinary ClassificationDiagnosticSpectral Mapping of Singing Voices: U-Net-Assisted Vocal Segmentation
Separating vocal elements from musical tracks is a longstanding challenge in audio signal processing. This study tackles the distinct separation of vocal components from musical spectrograms. We employ the Short Time Fou…
Audio Signal ProcessingAudio Source SeparationAudioSetMix: Enhancing Audio-Language Datasets with LLM-Assisted Augmentations
Multi-modal learning in the audio-language domain has seen significant advancements in recent years. However, audio-language learning faces challenges due to limited and lower-quality data compared to image-language task…
Audio Signal ProcessingLanguage ModelingLanguage ModellingLarge Language ModelComparative Study of State-based Neural Networks for Virtual Analog Audio Effects Modeling
Analog electronic circuits are at the core of an important category of musical devices, which includes a broad range of sound synthesizers and audio effects. The development of software that simulates analog musical devi…
Audio Effects ModelingAudio Signal ProcessingState Space ModelsOverview of the L3DAS23 Challenge on Audio-Visual Extended Reality
The primary goal of the L3DAS23 Signal Processing Grand Challenge at ICASSP 2023 is to promote and support collaborative research on machine learning for 3D audio signal processing, with a specific emphasis on 3D speech …
Audio Signal ProcessingSound Event Localization and DetectionSpeech EnhancementA Survey on Data Augmentation in Large Model Era
Large models, encompassing large language and diffusion models, have shown exceptional promise in approximating human-level intelligence, garnering significant interest from both academic and industrial spheres. However,…
Audio Signal ProcessingData AugmentationImage AugmentationSurvey+1HAAQI-Net: A Non-intrusive Neural Music Audio Quality Assessment Model for Hearing Aids
This paper introduces HAAQI-Net, a non-intrusive deep learning-based music audio quality assessment model for hearing aid users. Unlike traditional methods like the Hearing Aid Audio Quality Index (HAAQI) that require in…
Audio Quality AssessmentAudio Signal ProcessingComputational EfficiencyKnowledge Distillation+1Unsupervised Harmonic Parameter Estimation Using Differentiable DSP and Spectral Optimal Transport
In neural audio signal processing, pitch conditioning has been used to enhance the performance of synthesizers. However, jointly training pitch estimators and synthesizers is a challenge when using standard audio-to-audi…
Audio Signal Processingparameter estimationStudy of speaker localization under dynamic and reverberant environments
Speaker localization in a reverberant environment is a fundamental problem in audio signal processing. Many solutions have been developed to tackle this problem. However, previous algorithms typically assume a stationary…
Audio Signal ProcessingHPCNeuroNet: Advancing Neuromorphic Audio Signal Processing with Transformer-Enhanced Spiking Neural Networks
This paper presents a novel approach to neuromorphic audio processing by integrating the strengths of Spiking Neural Networks (SNNs), Transformers, and high-performance computing (HPC) into the HPCNeuroNet architecture. …
Audio Signal ProcessingCPUGPUNeural Harmonium: An Interpretable Deep Structure for Nonlinear Dynamic System Identification with Application to Audio Processing
Improving the interpretability of deep neural networks has recently gained increased attention, especially when the power of deep learning is leveraged to solve problems in physics. Interpretability helps us understand a…
Acoustic echo cancellationAudio Signal Processing