paper-with-me

Papers

FAConformer: Frequency-Aware Convolutional Transformer for Auditory Attention Decoding

2026-06-12 · Ziwei Wang, Xingyi He, Tianwang Jia, Hongbin Wang, Dongrui Wu arxiv

Auditory attention decoding (AAD) aims to infer the attended speaker from neural responses in multi-speaker acoustic environments and is a key problem for neuro-steered hearing systems. Although recent studies have achieved encouraging progress, existing AAD models still do not fully exploit frequency domain electroencephalography (EEG) information. In particular, most approaches introduce multi-band information through handcrafted feature extraction or direct cross-band feature concatenation, which mainly exploit frequency information at a shallow level and may overlook band-specific patterns and cross-band interactions. To address these limitations, this paper proposes FAConformer, a frequency-aware CNN-Transformer framework for AAD that explicitly integrates band-specific encoding and adaptive cross-band interaction. Specifically, FAConformer first decomposes EEG signals into multiple frequency bands and assigns each band to an independent CNN-Transformer encoder for band-specific modeling. The resulting band-wise features are then adaptively fused by a carefully designed frequency-aware attention (FAA) module that models cross-band dependencies by treating band-wise features as tokens. Further, band-wise auxiliary supervision (BAS) is introduced to prevent weakly contributing branches from being under-optimized during joint training. In this way, FAConformer performs frequency-aware modeling that more effectively exploits frequency domain information. Extensive experiments on two public AAD datasets with three decision-window lengths demonstrated that FAConformer consistently outperformed 12 competitive baselines, surpassing the current state-of-the-art model by 4.9%. Further analyses of band importance, ablation, and parameter sensitivity verify the effectiveness, robustness, and interpretability of the proposed framework. Code is available at https://github.com/wzwvv/FAConformer.

📄 PDF Abstract BibTeX arXiv:2606.14120

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

EEG-EMG FAConformer: Frequency Aware Conv-Transformer for the fusion of EEG and EMG

2024-09-12 · Zhengxiao He, Minghong Cai, Letian Li, Siyuan Tian 외

Motor pattern recognition paradigms are the main forms of Brain-Computer Interfaces(BCI) aimed at motor function rehabilitation and are the most easily promoted applications. In recent years, many researchers have sugges…

EEGElectromyography (EMG)

Slow-Fast Auditory Streams For Audio Recognition

2021-03-05 · Evangelos Kazakos, Arsha Nagrani, Andrew Zisserman, Dima Damen

We propose a two-stream convolutional network for audio recognition, that operates on time-frequency spectrogram inputs. Following similar success in visual recognition, we learn Slow-Fast auditory streams with separable…

Audio ClassificationHuman Interaction Recognition

How to train your ears: Auditory-model emulation for large-dynamic-range inputs and mild-to-severe hearing losses

2024-03-15 · Peter Leer, Jesper Jensen, Zheng-Hua Tan, Jan Østergaard 외

Advanced auditory models are useful in designing signal-processing algorithms for hearing-loss compensation or speech enhancement. Such auditory models provide rich and detailed descriptions of the auditory pathway, and …

Speech Enhancement

ISAC: An Invertible and Stable Auditory Filter Bank with Customizable Kernels for ML Integration

2025-05-12 · Daniel Haider, Felix Perfler, Peter Balazs, Clara Hollomey 외

This paper introduces ISAC, an invertible and stable, perceptually-motivated filter bank that is specifically designed to be integrated into machine learning paradigms. More precisely, the center frequencies and bandwidt…

ISAC

Auditory Neural Response Inspired Sound Event Detection Based on Spectro-temporal Receptive Field

2023-06-20 · Deokki Min, Hyeonuk Nam, Yong-Hwa Park

Sound event detection (SED) is one of tasks to automate function by human auditory system which listens and understands auditory scenes. Therefore, we were inspired to make SED recognize sound events in the way human aud…

Event DetectionSound Event Detection