Native Language Identification With Classifier Stacking and Ensembles
Ensemble methods using multiple classifiers have proven to be among the most successful approaches for the task of Native Language Identification (NLI), achieving the current state of the art. However, a systematic examination of ensemble methods for NLI has yet to be conducted. Additionally, deeper ensemble architectures such as classifier stacking have not been closely evaluated. We present a set of experiments using three ensemble-based models, testing each with multiple configurations and algorithms. This includes a rigorous application of meta-classification models for NLI, achieving state-of-the-art results on several large data sets, evaluated in both intra-corpus and cross-corpus modes.
Code (0)
등록된 구현이 없습니다.
Tasks
Cross-corpusGeneral ClassificationLanguage AcquisitionLanguage IdentificationNative Language IdentificationText ClassificationSimilar Papers 제목 키워드 기반
Classifier Stacking for Native Language Identification
This paper reports our contribution (team WLZ) to the NLI Shared Task 2017 (essay track). We first extract lexical and syntactic features from the essays, perform feature weighting and selection, and train linear support…
Language AcquisitionLanguage IdentificationNative Language IdentificationText Classification+1Native Language Identification using Stacked Generalization
Ensemble methods using multiple classifiers have proven to be the most successful approach for the task of Native Language Identification (NLI), achieving the current state of the art. However, a systematic examination o…
Language IdentificationNative Language IdentificationA Human-Centered Approach for Improving Supervised Learning
Supervised Learning is a way of developing Artificial Intelligence systems in which a computer algorithm is trained on labeled data inputs. Effectiveness of a Supervised Learning algorithm is determined by its performanc…
Ensemble LearningLanguage Identification using Classifier Ensembles
LTG-ST at NADI Shared Task 1: Arabic Dialect Identification using a Stacking Classifier
This paper presents our results for the Nuanced Arabic Dialect Identification (NADI) shared task of the Fifth Workshop for Arabic Natural Language Processing (WANLP 2020). We participated in the first sub-task for countr…
Dialect Identificationregression