paper-with-me

홈 › Papers

Developing Resources for Automated Speech Processing of Quebec French

2020-05-01 · LREC 2020 5 · M{\'e}lanie Lancien, Marie-H{\'e}l{\`e}ne C{\^o}t{\'e}, Brigitte Bigi

The analysis of the structure of speech nearly always rests on the alignment of the speech recording with a phonetic transcription. Nowadays several tools can perform this speech segmentation automatically. However, none of them allows the automatic segmentation of Quebec French (QF hereafter), the acoustics and phonotactics of QF differing widely from that of France French (FF hereafter). To adequately segment QF, features like diphthongization of long vowels and affrication of coronal stops have to be taken into account. Thus acoustic models for automatic segmentation must be trained on speech samples exhibiting those phenomena. Dictionaries and lexicons must also be adapted and integrate differences in lexical units and in the phonology of QF. This paper presents the development of linguistic resources to be included into SPPAS software tool in order to get Text normalization, Phonetization, Alignment and Syllabification. We adapted the existing French lexicon and developed a QF-specific pronunciation dictionary. We then created an acoustic model from the existing ones and adapted it with 5 minutes of manually time-aligned data. These new resources are all freely distributed with SPPAS version 2.7; they perform the full process of speech segmentation in Quebec French.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

SegmentationText Normalization

Similar Papers 제목 키워드 기반

Part of Speech Tagging (POST) of a Low-resource Language using another Language (Developing a POS-Tagged Lexicon for Kurdish (Sorani) using a Tagged Persian (Farsi) Corpus)

2022-01-30 · Hossein Hassani

Tagged corpora play a crucial role in a wide range of Natural Language Processing. The Part of Speech Tagging (POST) is essential in developing tagged corpora. It is time-and-effort-consuming and costly, and therefore, i…

Part-Of-Speech TaggingPOS

Challenges and Perspectives for Innu-Aimun within Indigenous Language Technologies

2022-05-01 · ComputEL (ACL) 2022 5 · Antoine Cadotte, Tan Le Ngoc, Mathieu Boivin, Fatiha Sadat

Innu-Aimun is an Algonquian language spoken in Eastern Canada. It is the language of the Innu, an Indigenous people that now lives for the most part in a dozen communities across Quebec and Labrador. Although it is alive…

Machine Translation

Benchmarking Large Language Models for Quebec Insurance: From Closed-Book to Retrieval-Augmented Generation

2026-03-08 · David Beauchemin, Richard Khoury arxiv

The digitization of insurance distribution in the Canadian province of Quebec, accelerated by legislative changes such as Bill 141, has created a significant "advice gap", leaving consumers to interpret complex financial…

An Adapter-Based Unified Model for Multiple Spoken Language Processing Tasks

2024-06-20 · Varsha Suresh, Salah Aït-Mokhtar, Caroline Brun, Ioan Calapodescu

Self-supervised learning models have revolutionized the field of speech processing. However, the process of fine-tuning these models on downstream tasks requires substantial computational resources, particularly when dea…

Automatic Speech RecognitionDecoderEmotion Recognitionintent-classification+7

Cost-effective Models for Detecting Depression from Speech

2023-02-18 · Mashrura Tasnim, Jekaterina Novikova

Depression is the most common psychological disorder and is considered as a leading cause of disability and suicide worldwide. An automated system capable of detecting signs of depression in human speech can contribute t…