Papers Audio Signal Processing
“Audio Signal Processing” 태그가 달린 논문 70편 · 필터 해제
Speaker localization using direct path dominance test based on sound field directivity
Estimation of the direction-of-arrival (DoA) of a speaker in a room is important in many audio signal processing applications. Environments with reverberation that masks the DoA information are particularly challenging. …
Audio Signal ProcessingAudio signal based danger detection using signal processing and deep learning
Since there have been more and more incidents of women being harassed in the recent past, girls need to think twice before going out of their houses. Sometimes, they are not even safe in their house or workplace. These c…
Audio Signal ProcessingInstabilities in Convnets for Raw Audio
What makes waveform-based deep learning so hard? Despite numerous attempts at training convolutional neural networks (convnets) for filterbank design, they often fail to outperform hand-crafted baselines. These baselines…
Audio Signal ProcessingNeural Architectures Learning Fourier Transforms, Signal Processing and Much More....
This report will explore and answer fundamental questions about taking Fourier Transforms and tying it with recent advances in AI and neural architecture. One interpretation of the Fourier Transform is decomposing a sign…
Audio Signal ProcessingCompositional nonlinear audio signal processing with Volterra series
We present a compositional theory of nonlinear audio signal processing based on a categorification of the Volterra series. We begin by augmenting the classical definition of the Volterra series so that it is functorial w…
Audio Signal ProcessingMF-PAM: Accurate Pitch Estimation through Periodicity Analysis and Multi-level Feature Fusion
We introduce Multi-level feature Fusion-based Periodicity Analysis Model (MF-PAM), a novel deep learning-based pitch estimation model that accurately estimates pitch trajectory in noisy and reverberant acoustic environme…
Audio Signal ProcessingAcoustic Scene Clustering Using Joint Optimization of Deep Embedding Learning and Clustering Iteration
Recent efforts have been made on acoustic scene classification in the audio signal processing community. In contrast, few studies have been conducted on acoustic scene clustering, which is a newly emerging problem. Acous…
Acoustic Scene ClassificationAudio Signal ProcessingClusteringScene ClassificationSubspace-Configurable Networks
While the deployment of deep learning models on edge devices is increasing, these models often lack robustness when faced with dynamic changes in sensed data. This can be attributed to sensor drift, or variations in the …
Audio Signal ProcessingData AugmentationNovel features for the detection of bearing faults in railway vehicles
{In this paper, we address the challenging problem of detecting bearing faults from vibration signals. For this, several time- and frequency-domain features have been proposed already in the past. However, these features…
Audio Signal ProcessingFault DetectionHuman Behavior in the Time of COVID-19: Learning from Big Data
Since the World Health Organization (WHO) characterized COVID-19 as a pandemic in March 2020, there have been over 600 million confirmed cases of COVID-19 and more than six million deaths as of October 2022. The relation…
Audio Signal ProcessingContent Adaptive Front End For Audio Classification
We propose a learnable content adaptive front end for audio signal processing. Before the modern advent of deep learning, we used fixed representation non-learnable front-ends like spectrogram or mel-spectrogram with/wit…
Audio ClassificationAudio Signal ProcessingClassificationScene UnderstandingMYRiAD: A Multi-Array Room Acoustic Database
In the development of acoustic signal processing algorithms, their evaluation in various acoustic environments is of utmost importance. In order to advance evaluation in realistic and reproducible scenarios, several high…
Audio Signal ProcessingA Comparison of Audio Preprocessing Techniques and Deep Learning Algorithms for Raga Recognition
Ragas form the foundation for Indian Classical Music. The task of Raga Recognition has gained traction in the Music Information Retrieval community in the recent past, which can be attributed to the nuances of Indian Cla…
Audio Signal ProcessingInformation RetrievalMusic Information RetrievalRetrievalHigh Fidelity Neural Audio Compression
We introduce a state-of-the-art real-time, high-fidelity, audio codec leveraging neural networks. It consists in a streaming encoder-decoder architecture with quantized latent space trained in an end-to-end fashion. We s…
Audio CompressionAudio Signal ProcessingDecoderVocal Bursts Intensity PredictionA Unifying View on Blind Source Separation of Convolutive Mixtures based on Independent Component Analysis
In many daily-life scenarios, acoustic sources recorded in an enclosure can only be observed with other interfering sources. Hence, convolutive Blind Source Separation (BSS) is a central problem in audio signal processin…
Audio Signal Processingblind source separationRelationContext-sensitive neocortical neurons transform the effectiveness and efficiency of neural information processing
Deep learning (DL) has big-data processing capabilities that are as good, or even better, than those of humans in many real-world domains, but at the cost of high energy requirements that may be unsustainable in some app…
Audio Signal ProcessingSound2Synth: Interpreting Sound via FM Synthesizer Parameters Estimation
Synthesizer is a type of electronic musical instrument that is now widely used in modern music production and sound design. Each parameters configuration of a synthesizer produces a unique timbre and can be viewed as a u…
Audio ClassificationAudio Signal ProcessingDeclipping of Speech Signals Using Frequency Selective Extrapolation
The reconstruction of clipped speech signals is an important task in audio signal processing to achieve an enhanced audio quality for further processing. In this paper, Frequency Selective Extrapolation (FSE), which is c…
Audio Signal ProcessingBi-Sampling Approach to Classify Music Mood leveraging Raga-Rasa Association in Indian Classical Music
The impact of Music on the mood or emotion of the listener is a well-researched area in human psychology and behavioral science. In Indian classical music, ragas are the melodic structure that defines the various styles …
Audio Signal ProcessingMusic RecommendationManifold learning-supported estimation of relative transfer functions for spatial filtering
Many spatial filtering algorithms used for voice capture in, e.g., teleconferencing applications, can benefit from or even rely on knowledge of Relative Transfer Functions (RTFs). Accordingly, many RTF estimators have be…
Audio Signal Processing