Large-Scale Native Language Identification with Cross-Corpus Evaluation
Code (0)
등록된 구현이 없습니다.
Tasks
Cross-corpusLanguage AcquisitionLanguage IdentificationNative Language IdentificationSimilar Papers 제목 키워드 기반
Native Language Identification using large scale lexical features
On the Development of a Large Scale Corpus for Native Language Identification
Native Language Identification (NLI) is the task of identifying an author’s native language from their writings in a second language. In this paper, we introduce a new corpus (italki), which is larger than the current co…
BIG-bench Machine LearningLanguage IdentificationNative Language IdentificationImproving Language Identification of Accented Speech
Language identification from speech is a common preprocessing step in many spoken language processing systems. In recent years, this field has seen fast progress, mostly due to the use of self-supervised models pretraine…
Language Identificationspeech-recognitionSpeech RecognitionSpoken language identificationA language model based approach towards large scale and lightweight language identification systems
Multilingual spoken dialogue systems have gained prominence in the recent past necessitating the requirement for a front-end Language Identification (LID) system. Most of the existing LID systems rely on modeling the lan…
Language IdentificationLanguage ModelingLanguage ModellingSpoken Dialogue SystemsCross-corpus Native Language Identification via Statistical Embedding
In this paper, we approach the task of native language identification in a realistic cross-corpus scenario where a model is trained with available data and has to predict the native language from data of a different corp…
Cross-corpusLanguage IdentificationNative Language Identification