paper-with-me

홈 › Papers

Creation of a Balanced State-of-the-Art Multilayer Corpus for NLU

2018-05-01 · LREC 2018 5 · Normunds Gruzitis, Lauma Pretkalnina, Baiba Saulite, Laura Rituma, Gunta Nespore-Berzkalne, Arturs Znotins, Peteris Paikens
📄 PDF Abstract BibTeX

Code (1)

LUMII-AILab/FullStack 공식 구현

Tasks

Abstractive Text SummarizationCoreference ResolutionEntity LinkingKnowledge Base PopulationNamed Entity Recognition (NER)Semantic ParsingSemantic Role LabelingText Summarization

Similar Papers 제목 키워드 기반

Designing the Latvian Speech Recognition Corpus

2014-05-01 · LREC 2014 5 · M{\=a}rcis Pinnis, Ilze Auzi{\c{n}}a, K{\=a}rlis Goba

In this paper the authors present the first Latvian speech corpus designed specifically for speech recognition purposes. The paper outlines the decisions made in the corpus designing process through analysis of related w…

speech-recognitionSpeech RecognitionSpeech Synthesis

AMALGUM -- A Free, Balanced, Multilayer English Web Corpus

2020-06-18 · LREC 2020 5 · Luke Gessler, Siyao Peng, Yang Liu, YIlun Zhu 외

We present a freely available, genre-balanced English web corpus totaling 4M tokens and featuring a large number of high-quality automatic annotation layers, including dependency trees, non-named entity annotations, core…

coreference-resolutionCoreference Resolution

Opera Graeca Adnotata: Building a 34M+ Token Multilayer Corpus for Ancient Greek

2024-03-31 · Giuseppe G. A. Celano

In this article, the beta version 0.1.0 of Opera Graeca Adnotata (OGA), the largest open-access multilayer corpus for Ancient Greek (AG) is presented. OGA consists of 1,687 literary works and 34M+ tokens coming from the …

LemmatizationSentenceSentence segmentation

Design of a Tigrinya Language Speech Corpus for Speech Recognition

2018-08-01 · COLING 2018 8 · Hafte Abera, Sebsibe H/mariam

In this paper, we describe the first Tigrinya Languages speech corpora designed and development for speech recognition purposes. Tigrinya, often written as Tigrigna (ትግርኛ) /tɪˈɡrinjə/ belongs to the Semitic branch of the…

speech-recognitionSpeech Recognition

Audiobook Dialogues as Training Data for Conversational Style Synthetic Voices

2022-06-01 · LREC 2022 6 · Liisi Piits, Hille Pajupuu, Heete Sahkai, Rene Altrov 외

Synthetic voices are increasingly used in applications that require a conversational speaking style, raising the question as to which type of training data yields the most suitable speaking style for such applications. T…

Sentencetext-to-speechText to Speech