paper-with-me

홈 › Papers

Human Voice Pitch Estimation: A Convolutional Network with Auto-Labeled and Synthetic Data

2023-08-14 · Jeremy Cochoy

In the domain of music and sound processing, pitch extraction plays a pivotal role. Our research presents a specialized convolutional neural network designed for pitch extraction, particularly from the human singing voice in acapella performances. Notably, our approach combines synthetic data with auto-labeled acapella sung audio, creating a robust training environment. Evaluation across datasets comprising synthetic sounds, opera recordings, and time-stretched vowels demonstrates its efficacy. This work paves the way for enhanced pitch extraction in both music and voice settings.

📄 PDF Abstract BibTeX arXiv:2308.07170

Code (1)

jeremycochoy/pitchnet 공식 구현 pytorch

Similar Papers 제목 키워드 기반

Estimation du pitch et d\'ecision de voisement par compression spectrale de l'autocorr\'elation du produit multi-\'echelle (Pitch estimation and voiced decision by spectral autocorrelation compression of multi-scale product) [in French]

2012-06-01 · JEPTALNRECITAL 2012 6 · Mohamed Anouar Ben Messaoud, A{\"\i}cha Bouzid, Noureddine Ellouze

Convolutional Speech Recognition with Pitch and Voice Quality Features

2020-09-02 · Guillermo Cámbara, Jordi Luque, Mireia Farrús

The effects of adding pitch and voice quality features such as jitter and shimmer to a state-of-the-art CNN model for Automatic Speech Recognition are studied in this work. Pitch features have been previously used for im…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Emotion Recognitionspeech-recognition+1

Comparing Conventional Pitch Detection Algorithms with a Neural Network Approach

2022-06-29 · Anja Kroon

Despite much research, traditional methods to pitch prediction are still not perfect. With the emergence of neural networks (NNs), researchers hope to create a NN-based pitch predictor that outperforms traditional method…

SPICE: Self-supervised Pitch Estimation

2019-10-25 · Beat Gfeller, Christian Frank, Dominik Roblek, Matt Sharifi 외

We propose a model to estimate the fundamental frequency in monophonic audio, often referred to as pitch estimation. We acknowledge the fact that obtaining ground truth annotations at the required temporal and frequency …

Self-Supervised LearningTranslation

Deep Autotuner: A Data-Driven Approach to Natural-Sounding Pitch Correction for Singing Voice in Karaoke Performances

2019-02-03 · Sanna Wager, George Tzanetakis, Cheng-i Wang, Lijiang Guo 외

We describe a machine-learning approach to pitch correcting a solo singing performance in a karaoke setting, where the solo voice and accompaniment are on separate tracks. The proposed approach addresses the situation wh…