paper-with-me

홈 › Papers

BulPhonC: Bulgarian Speech Corpus for the Development of ASR Technology

2016-05-01 · LREC 2016 5 · Neli Hateva, Petar Mitankin, Stoyan Mihov

In this paper we introduce a Bulgarian speech database, which was created for the purpose of ASR technology development. The paper describes the design and the content of the speech database. We present also an empirical evaluation of the performance of a LVCSR system for Bulgarian trained on the BulPhonC data. The resource is available free for scientific usage.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

The Political Speech Corpus of Bulgarian

2012-05-01 · LREC 2012 5 · Petya Osenova, Kiril Simov

The paper introduces the Political Speech Corpus of Bulgarian. First, its current state has been discussed with respect to its size, coverage, genre specification and related online services. Then, the focus goes to the …

LemmatizationMorphological AnalysisSentiment Analysis

Syntactic characteristics of emotive predicates in Bulgarian: A corpus-based study

2022-09-01 · CLIB 2022 9 · Yovka Tisheva, Marina Dzhonova

The paper presents a corpus-based study of emotive predicates (verbs and predicative constructions with adjectival, adverbial or noun phrases) in Bulgarian with respect to their syntactic characteristics. The sources of …

Feature-Rich Part-of-speech Tagging for Morphologically Complex Languages: Application to Bulgarian

2019-11-26 · EACL 2012 4 · Georgi Georgiev, Valentin Zhikov, Petya Osenova, Kiril Simov 외

We present experiments with part-of-speech tagging for Bulgarian, a Slavic language with rich inflectional and derivational morphology. Unlike most previous work, which has used a small number of grammatical categories, …

Part-Of-Speech TaggingPOS

Annotation of Clinical Narratives in Bulgarian language

2017-09-01 · RANLP 2017 9 · Ivajlo Radev, Kiril Simov, Galia Angelova, Svetla Boytcheva

In this paper we describe annotation process of clinical texts with morphosyntactic and semantic information. The corpus contains 1,300 discharge letters in Bulgarian language for patients with Endocrinology and Metaboli…

ChunkingDependency ParsingInformation Retrieval

Categorisation of Bulgarian Legislative Documents

2020-09-01 · CLIB 2020 9 · Nikola Obreshkov, Martin Yalamov, Svetla Koeva

The paper presents the categorisation of Bulgarian MARCELL corpus in toplevel EuroVoc domains. The Bulgarian MARCELL corpus is part of a recently developed multilingual corpus representing the national legislation in sev…

Term Extraction