paper-with-me

Papers

Tone Recognition Using Lifters and CTC

2018-07-06

In this paper, we present a new method for recognizing tones in continuous speech for tonal languages. The method works by converting the speech signal to a cepstrogram, extracting a sequence of cepstral features using a convolutional neural network, and predicting the underlying sequence of tones using a connectionist temporal classification (CTC) network. The performance of the proposed method is evaluated on a freely available Mandarin Chinese speech corpus, AISHELL-1, and is shown to outperform the existing techniques in the literature in terms of tone error rate (TER).

📄 PDF Abstract BibTeX arXiv:1807.02465

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Improving the Robustness of 3D Human Pose Estimation: A Benchmark and Learning from Noisy Input

2023-12-11 · Trung-Hieu Hoang, Mona Zehni, Huy Phan, Duc Minh Vo 외

Despite the promising performance of current 3D human pose estimation techniques, understanding and enhancing their generalization on challenging in-the-wild videos remain an open problem. In this work, we focus on the r…

3D Human Pose EstimationData AugmentationPose Estimation

Tone recognition in low-resource languages of North-East India: peeling the layers of SSL-based speech models

2025-06-04 · Parismita Gogoi, Sishir Kalita, Wendy Lalhminghlui, Viyazonuo Terhiija 외

This study explores the use of self-supervised learning (SSL) models for tone recognition in three low-resource languages from North Eastern India: Angami, Ao, and Mizo. We evaluate four Wav2vec2.0 base models that were …

Self-Supervised Learning

Automatic recognition of suprasegmentals in speech

2021-08-02 · Jiahong Yuan, Neville Ryant, Xingyu Cai, Kenneth Church 외

This study reports our efforts to improve automatic recognition of suprasegmentals by fine-tuning wav2vec 2.0 with CTC, a method that has been successful in automatic speech recognition. We demonstrate that the method ca…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Phoneme Recognitionspeech-recognition+1

Reference Points, Risk-Taking Behavior, and Competitive Outcomes in Sequential Settings

2024-09-20 · Masaya Nishihata, Suguru Otani

Understanding how competitive pressure affects risk-taking is crucial in sequential decision-making under uncertainty. This study examines these effects using bench press competition data, where individuals make risk-bas…

counterfactualDecision MakingDecision Making Under UncertaintySequential Decision Making

TrueSkin: Towards Fair and Accurate Skin Tone Recognition and Generation

2025-09-13 · Haoming Lu arxiv

Skin tone recognition and generation play important roles in model fairness, healthcare, and generative AI, yet they remain challenging due to the lack of comprehensive datasets and robust methodologies. Compared to othe…

Image Generation