The Use of Text Alignment in Semi-Automatic Error Analysis: Use Case in the Development of the Corpus of the Latvian Language Learners
Code (0)
등록된 구현이 없습니다.
Tasks
Language AcquisitionLemmatizationMorphological AnalysisPart-Of-Speech TaggingWord AlignmentSimilar Papers 제목 키워드 기반
Original-Transcribed Text Alignment for Manyosyu Written by Old Japanese Language
We are constructing an annotated diachronic corpora of the Japanese language. In part of thiswork, we construct a corpus of Manyosyu, which is an old Japanese poetry anthology. In thispaper, we describe how to align the …
Machine TranslationTranslationCreating a Corpus for Russian Data-to-Text Generation Using Neural Machine Translation and Post-Editing
In this paper, we propose an approach for semi-automatically creating a data-to-text (D2T) corpus for Russian that can be used to learn a D2T natural language generation model. An error analysis of the output of an Engli…
Data-to-Text GenerationMachine TranslationText GenerationTranslationBreaking the Script Barrier: Enabling Automatic Alignment for PoS-based ASR Error Analysis in Non-Latin Scripts
Automatic Speech Recognition (ASR) systems are commonly evaluated using aggregate metrics such as Word Error Rate (WER), which do not capture the linguistic structure of errors. Fine-grained analysis, such as Part-of-Spe…
Speech RecognitionCorrecting Errors in a New Gold Standard for Tagging Icelandic Text
In this paper, we describe the correction of PoS tags in a new Icelandic corpus, MIM-GOLD, consisting of about 1 million tokens sampled from the Tagged Icelandic Corpus, M{\'I}M, released in 2013. The goal is to use the …
Part-Of-Speech TaggingPOSNeural semi-Markov CRF for Monolingual Word Alignment
Monolingual word alignment is important for studying fine-grained editing operations (i.e., deletion, addition, and substitution) in text-to-text generation tasks, such as paraphrase generation, text simplification, neut…
Paraphrase GenerationSentenceSentence-Pair ClassificationText Generation+2