paper-with-me

Papers

Classifier Stacking for Native Language Identification

2017-09-01 · WS 2017 9 · Wen Li, Liang Zou

This paper reports our contribution (team WLZ) to the NLI Shared Task 2017 (essay track). We first extract lexical and syntactic features from the essays, perform feature weighting and selection, and train linear support vector machine (SVM) classifiers each on an individual feature type. The output of base classifiers, as probabilities for each class, are then fed into a multilayer perceptron to predict the native language of the author. We also report the performance of each feature type, as well as the best features of a type. Our system achieves an accuracy of 86.55{\%}, which is among the best performing systems of this shared task.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Language AcquisitionLanguage IdentificationNative Language IdentificationText ClassificationVocal Bursts Type Prediction

Similar Papers 제목 키워드 기반

Native Language Identification With Classifier Stacking and Ensembles

2018-09-01 · CL 2018 9 · Shervin Malmasi, Mark Dras

Ensemble methods using multiple classifiers have proven to be among the most successful approaches for the task of Native Language Identification (NLI), achieving the current state of the art. However, a systematic exami…

Cross-corpusGeneral ClassificationLanguage AcquisitionLanguage Identification+2

Native Language Identification using Stacked Generalization

2017-03-19 · Shervin Malmasi, Mark Dras

Ensemble methods using multiple classifiers have proven to be the most successful approach for the task of Native Language Identification (NLI), achieving the current state of the art. However, a systematic examination o…

Language IdentificationNative Language Identification

LTG-ST at NADI Shared Task 1: Arabic Dialect Identification using a Stacking Classifier

2020-12-01 · COLING (WANLP) 2020 12 · Samia Touileb

This paper presents our results for the Nuanced Arabic Dialect Identification (NADI) shared task of the Fifth Workshop for Arabic Natural Language Processing (WANLP 2020). We participated in the first sub-task for countr…

Dialect Identificationregression

Native Language Identification on Text and Speech

2017-07-22 · WS 2017 9 · Marcos Zampieri, Alina Maria Ciobanu, Liviu P. Dinu

This paper presents an ensemble system combining the output of multiple SVM classifiers to native language identification (NLI). The system was submitted to the NLI Shared Task 2017 fusion track which featured students e…

Language IdentificationNative Language Identification

A Review of Standard Text Classification Practices for Multi-label Toxicity Identification of Online Content

2018-10-01 · WS 2018 10 · Isuru Gunasekara, Isar Nejadgholi

Language toxicity identification presents a gray area in the ethical debate surrounding freedom of speech and censorship. Today{'}s social media landscape is littered with unfiltered content that can be anywhere from sli…

ClassificationGeneral Classificationtext-classificationText Classification+1