Logarithmic Frequency Scaling and Consistent Frequency Coverage for the Selection of Auditory Filterbank Center Frequencies
This paper provides new insights into the problem of selecting filter center frequencies for the auditory filterbanks. We propose to use a constant frequency distance and a consistent frequency coverage as the two metrics that motivate the logarithmic frequency scaling and a regularized selection of center frequencies. The frequency scaling and the consistent frequency coverage have been derived based on a common harmonic speaker signal model. Furthermore, we have found that the existing linear equivalent rectangular bandwidth (ERB) function as well as any possible linear ERB approximation can also lead to a consistent frequency coverage. The results are verified and demonstrated using the gammatone filterbank.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
A Comparison of Audio Signal Preprocessing Methods for Deep Neural Networks on Music Tagging
In this paper, we empirically investigate the effect of audio preprocessing on music tagging with deep neural networks. We perform comprehensive experiments involving audio preprocessing using different time-frequency re…
Music TaggingMultidimensional Gabor-Like Filters Derived from Gaussian Functions on Logarithmic Frequency Axes
A novel wavelet-like function is presented that makes it convenient to create filter banks given mainly two parameters that influence the focus area and the filter count. This is accomplished by computing the inverse Fou…
Data Amplification: Instance-Optimal Property Estimation
The best-known and most commonly used distribution-property estimation technique uses a plug-in estimator, with empirical frequency replacing the underlying distribution. We present novel linear-time-computable estimator…
Natural Spectral Fusion: p-Exponent Cyclic Scheduling and Early Decision-Boundary Alignment in First-Order Optimization
Spectral behaviors have been widely discussed in machine learning, yet the optimizer's own spectral bias remains unclear. We argue that first-order optimizers exhibit an intrinsic frequency preference that significantly …
Dissonance Spectrum explicitly models perceptual frequency interactions for better music understanding
Conventional music representations describe acoustic energy over time and frequency but do not explicitly expose relations among simultaneous frequency components. We introduce the \emph{Dissonance Spectrum} (DS), a nonn…
Music Question AnsweringEmotion Recognition