paper-with-me

Papers Audio Denoising

“Audio Denoising” 태그가 달린 논문 26편 · 필터 해제

Scalable Operator Learning via Nyström Approximation With Denoising Applications

2026-06-25 · Naveen Gupta, Vaibhav Silmana, S. Sivananthan arxiv

In this paper, we study Nyström subsampling for vector-valued regression in vector-valued reproducing kernel Hilbert spaces. Standard kernel methods often suffer from prohibitive computational costs due to the constructi…

Image DenoisingAudio Denoising

Automatic Contextual Audio Denoising

2026-05-21 · Diep Luong, Konstantinos Drossos, Mikko Heikkinen, Tuomas Virtanen arxiv

Audio context determines which sound components and sources are relevant and which can be perceived as irrelevant (noise) by listeners. For example, traffic noise is informative in urban surveillance but noise for a phon…

Audio Denoising

Can Multimodal Large Language Models Understand Pathologic Movements? A Pilot Study on Seizure Semiology

2026-05-05 · Lina Zhang, Tonmoy Monsoor, Mehmet Efe Lorasdagi, Prateik Sinha 외 arxiv

Multimodal Large Language Models (MLLMs) have demonstrated robust capabilities in recognizing everyday human activities, yet their potential for analyzing clinically significant involuntary movements in neurological diso…

Audio DenoisingPose Estimation

SEE: Signal Embedding Energy for Quantifying Noise Interference in Large Audio Language Models

2026-01-12 · Yuanhe Zhang, Jiayu Tian, Yibo Zhang, Shilinlu Yan 외 arxiv

Large Audio Language Models (LALMs) have been widely applied in real-time scenarios, such as in-car assistants and online meeting comprehension. In practice, audio inputs are often corrupted by device and environmental n…

Audio Denoising

ADNAC: Audio Denoiser using Neural Audio Codec

2025-11-03 · Daniel Jimon, Mircea Vaida, Adriana Stan arxiv

Audio denoising is critical in signal processing, enhancing intelligibility and fidelity for applications like restoring musical recordings. This paper presents a proof-of-concept for adapting a state-of-the-art neural a…

Audio Denoising

On the Contribution of Lexical Features to Speech Emotion Recognition

2025-09-06 · David Combei arxiv

Although paralinguistic cues are often considered the primary drivers of speech emotion recognition (SER), we investigate the role of lexical content extracted from speech and show that it can achieve competitive and in …

Speech Emotion RecognitionAudio Denoising

Accelerating Linear Recurrent Neural Networks for the Edge with Unstructured Sparsity

2025-02-03 · Alessandro Pierro, Steven Abreu, Jonathan Timcheck, Philipp Stratmann 외

Linear recurrent neural networks enable powerful long-range sequence modeling with constant memory usage and time-per-token during inference. These architectures hold promise for streaming applications at the edge, but d…

Audio DenoisingDenoisingGPUModel Compression

CleanUMamba: A Compact Mamba Network for Speech Denoising using Channel Pruning

2024-10-14 · Sjoerd Groot, Qinyu Chen, Jan C. van Gemert, Chang Gao

This paper presents CleanUMamba, a time-domain neural network architecture designed for real-time causal audio denoising directly applied to raw waveforms. CleanUMamba leverages a U-Net encoder-decoder structure, incorpo…

Audio DenoisingDecoderDenoisingMamba+1

aTENNuate: Optimized Real-time Speech Enhancement with Deep SSMs on Raw Audio

2024-09-05 · Yan Ru Pei, Ritik Shrivastava, FNU Sidharth

We present aTENNuate, a simple deep state-space autoencoder configured for efficient online raw speech enhancement in an end-to-end fashion. The network's performance is primarily evaluated on raw speech denoising, with …

Audio DenoisingDenoisingSpeech DenoisingSpeech Enhancement+1

Diffusion Gaussian Mixture Audio Denoise

2024-06-13 · Pu Wang, Junhui Li, Jialu Li, Liangdong Guo 외

Recent diffusion models have achieved promising performances in audio-denoising tasks. The unique property of the reverse process could recover clean signals. However, the distribution of real-world noises does not compl…

Audio DenoisingDenoising

Complex Image Generation SwinTransformer Network for Audio Denoising

2023-10-24 · Youshan Zhang, Jialu Li

Achieving high-performance audio denoising is still a challenging task in real-world applications. Existing time-frequency methods often ignore the quality of generated frequency domain images. This paper converts the au…

Audio DenoisingDenoisingImage Generation

Efficient Video and Audio processing with Loihi 2

2023-10-05 · Sumit Bam Shrestha, Jonathan Timcheck, Paxon Frady, Leobardo Campos-Macias 외

Loihi 2 is an asynchronous, brain-inspired research processor that generalizes several fundamental elements of neuromorphic architecture, such as stateful neuron models communicating with event-driven spikes, in order to…

Audio DenoisingDenoising

Learning Spatial Features from Audio-Visual Correspondence in Egocentric Videos

2023-07-10 · CVPR 2024 1 · Sagnik Majumder, Ziad Al-Halah, Kristen Grauman

We propose a self-supervised method for learning representations based on spatial audio-visual correspondences in egocentric videos. Our method uses a masked auto-encoding framework to synthesize masked binaural (multi-c…

Active Speaker DetectionAudio DenoisingDenoising

The Intel Neuromorphic DNS Challenge

2023-03-16 · Jonathan Timcheck, Sumit Bam Shrestha, Daniel Ben Dayan Rubin, Adam Kupryjanow 외

A critical enabler for progress in neuromorphic computing research is the ability to transparently evaluate different neuromorphic solutions on important tasks and to compare them to state-of-the-art conventional solutio…

Audio DenoisingDenoising

Audio Denoising for Robust Audio Fingerprinting

2022-12-21 · Kamil Akesbi

Music discovery services let users identify songs from short mobile recordings. These solutions are often based on Audio Fingerprinting, and rely more specifically on the extraction of spectral peaks in order to be robus…

Audio DenoisingData AugmentationDenoising

BirdSoundsDenoising: Deep Visual Audio Denoising for Bird Sounds

2022-10-18 · Youshan Zhang, Jialu Li

Audio denoising has been explored for decades using both traditional and deep learning-based methods. However, these methods are still limited to either manually added artificial noise or lower denoised audio quality. To…

Audio DenoisingDenoisingImage SegmentationNoise Estimation+2

Self-Supervised Speech Denoising Using Only Noisy Audio Signals

2021-10-30 · Jiasong Wu, Qingchun Li, Guanyu Yang, Lei LI 외

In traditional speech denoising tasks, clean audio signals are often used as the training target, but absolutely clean signals are collected from expensive recording equipment or in studios with the strict environments. …

Audio DenoisingDenoisingSpeech Denoising

Self-Supervised Inference in State-Space Models

2021-07-28 · ICLR 2022 4 · David Ruhe, Patrick Forré

We perform approximate inference in state-space models with nonlinear state transitions. Without parameterizing a generative model, we apply Bayesian update formulas using a local linearity approximation parameterized by…

Audio DenoisingDenoisingState Space ModelsVariational Inference

Audio Attacks and Defenses against AED Systems -- A Practical Study

2021-06-14 · Rodrigo dos Santos, Shirin Nilizadeh

In this paper, we evaluate deep learning-enabled AED systems against evasion attacks based on adversarial examples. We test the robustness of multiple security critical AED tasks, implemented as CNNs classifiers, as well…

Audio DenoisingDenoisingEvent Detectionspeech-recognition+1

On the Design of Deep Priors for Unsupervised Audio Restoration

2021-04-14 · Vivek Sivaraman Narayanaswamy, Jayaraman J. Thiagarajan, Andreas Spanias

Unsupervised deep learning methods for solving audio restoration problems extensively rely on carefully tailored neural architectures that carry strong inductive biases for defining priors in the time or spectral domain.…

Audio DenoisingDenoising
1–20 / 26 다음 →