paper-with-me

Papers

THCHS-30 : A Free Chinese Speech Corpus

2015-12-07 · Dong Wang, Xuewei Zhang

Speech data is crucially important for speech recognition research. There are quite some speech databases that can be purchased at prices that are reasonable for most research institutes. However, for young people who just start research activities or those who just gain initial interest in this direction, the cost for data is still an annoying barrier. We support the `free data' movement in speech recognition: research institutes (particularly supported by public funds) publish their data freely so that new researchers can obtain sufficient data to kick of their career. In this paper, we follow this trend and release a free Chinese speech database THCHS-30 that can be used to build a full- edged Chinese speech recognition system. We report the baseline system established with this database, including the performance under highly noisy conditions.

📄 PDF Abstract BibTeX arXiv:1512.01882

Code (1)

foamliu/Listen-Attend-and-Spell pytorch

Tasks

speech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

OC16-CE80: A Chinese-English Mixlingual Database and A Speech Recognition Baseline

2016-09-27 · Dong Wang, Zhiyuan Tang, Difei Tang, Qing Chen

We present the OC16-CE80 Chinese-English mixlingual speech database which was released as a main resource for training, development and test for the Chinese-English mixlingual speech recognition (MixASR-CHEN) challenge o…

speech-recognitionSpeech Recognition

PARCO: Phoneme-Augmented Robust Contextual ASR via Contrastive Entity Disambiguation

2025-09-04 · Jiajun He, Naoki Sawada, Koichi Miyazaki, Tomoki Toda arxiv

Automatic speech recognition (ASR) systems struggle with domain-specific named entities, especially homophones. Contextual ASR improves recognition but often fails to capture fine-grained phoneme variations due to limite…

Entity DisambiguationSpeech Recognition

Development of a Web-Scale Chinese Word N-gram Corpus with Parts of Speech Information

2012-05-01 · LREC 2012 5 · Chi-Hsin Yu, Yi-jie Tang, Hsin-Hsi Chen

Web provides a large-scale corpus for researchers to study the language usages in real world. Developing a web-scale corpus needs not only a lot of computation resources, but also great efforts to handle the large variat…

Information RetrievalLanguage Modelling

Data centric approach to Chinese Medical Speech Recognition

2021-10-01 · ROCLING 2021 10 · Sheng-Luen Chung, Yi-Shiuan Li, Hsien-Wei Ting

Concerning the development of Chinese medical speech recognition technology, this study re-addresses earlier encountered issues in accordance with the process of Machine Learning Engineering for Production (MLOps) from a…

Data Augmentationspeech-recognitionSpeech Recognition

PANDA -- Paired Anti-hate Narratives Dataset from Asia: Using an LLM-as-a-Judge to Create the First Chinese Counterspeech Dataset

2025-01-01 · Michael Bennie, Demi Zhang, Bushi Xiao, Jing Cao 외

Despite the global prevalence of Modern Standard Chinese language, counterspeech (CS) resources for Chinese remain virtually nonexistent. To address this gap in East Asian counterspeech research we introduce the a corpus…