ALL-IN-1: Short Text Classification with One Model for All Languages
We present ALL-IN-1, a simple model for multilingual text classification that does not require any parallel data. It is based on a traditional Support Vector Machine classifier exploiting multilingual word embeddings and character n-grams. Our model is simple, easily extendable yet very effective, overall ranking 1st (out of 12 teams) in the IJCNLP 2017 shared task on customer feedback analysis in four languages: English, French, Japanese and Spanish.
Code (1)
Tasks
AllGeneral ClassificationMultilingual text classificationMultilingual Word Embeddingstext-classificationText ClassificationWord EmbeddingsSimilar Papers 제목 키워드 기반
Izindaba-Tindzaba: Machine learning news categorisation for Long and Short Text for isiZulu and Siswati
Local/Native South African languages are classified as low-resource languages. As such, it is essential to build the resources for these languages so that they can benefit from advances in the field of natural language p…
ClassificationregressionTopic ClassificationWord EmbeddingsL3Cube-IndicNews: News-based Short Text and Long Document Classification Datasets in Indic Languages
In this work, we introduce L3Cube-IndicNews, a multilingual text classification corpus aimed at curating a high-quality dataset for Indian regional languages, with a specific focus on news headlines and articles. We have…
ArticlesClassificationDocument ClassificationMultilingual text classification+4All-In-1 at IJCNLP-2017 Task 4: Short Text Classification with One Model for All Languages
We present All-In-1, a simple model for multilingual text classification that does not require any parallel data. It is based on a traditional Support Vector Machine classifier exploiting multilingual word embeddings and…
AllGeneral ClassificationMultilingual text classificationMultilingual Word Embeddings+3Investigating Shallow and Deep Learning Techniques for Emotion Classification in Short Persian Texts
The identification of emotions in short texts of low-resource languages poses a significant challenge, requiring specialized frameworks and computational intelligence techniques. This paper presents a comprehensive explo…
Deep LearningDimensionality ReductionEmotion ClassificationTransfer LearningCross-Lingual Task-Specific Representation Learning for Text Classification in Resource Poor Languages
Neural network models have shown promising results for text classification. However, these solutions are limited by their dependence on the availability of annotated data. The prospect of leveraging resource-rich langu…
ClassificationGeneral ClassificationRepresentation LearningSentiment Analysis+2