Codeswitching language identification using Subword Information Enriched Word Vectors
Code (0)
등록된 구현이 없습니다.
Tasks
Language IdentificationNamed Entity Recognition (NER)Similar Papers 제목 키워드 기반
Word-Level Language Identification and Predicting Codeswitching Points in Swahili-English Language Data
2016-11-01 · WS 2016 11
· Mario Piergallini, Rouzbeh Shirvani, Gauri S. Gautam, Mohamed Chouikha
Language IdentificationSentiment AnalysisSpeech Recognition
The Howard University System Submission for the Shared Task in Language Identification in Spanish-English Codeswitching
2016-11-01 · WS 2016 11
· Rouzbeh Shirvani, Mario Piergallini, Gauri Shankar Gautam, Mohamed Chouikha
Language Identification
Borrowing or Codeswitching? Annotating for Finer-Grained Distinctions in Language Mixing
2022-06-10 · LREC 2022 6
· Elena Alvarez Mellado, Constantine Lignos
We present a new corpus of Twitter data annotated for codeswitching and borrowing between Spanish and English. The corpus contains 9,500 tweets annotated at the token level with codeswitches, borrowings, and named entiti…
Word-level Language Identification Using Subword Embeddings for Code-mixed Bangla-English Social Media Data
2022-06-01 · DCLRL (LREC) 2022 6
· Aparna Dutta
This paper reports work on building a word-level language identification (LID) model for code-mixed Bangla-English social media data using subword embeddings, with an ultimate goal of using this LID module as the first s…
Language IdentificationPOSStem-driven Language Models for Morphologically Rich Languages
2019-10-25
· Yash Shah, Ishan Tarunesh, Harsh Deshpande, Preethi Jyothi
Neural language models (LMs) have shown to benefit significantly from enhancing word vectors with subword-level information, especially for morphologically rich languages. This has been mainly tackled by providing subwor…
Multi-Task Learning