paper-with-me

Papers

CS-Embed at SemEval-2020 Task 9: The effectiveness of code-switched word embeddings for sentiment analysis

2020-06-08 · SEMEVAL 2020 · Frances Adriana Laureano De Leon, Florimond Guéniat, Harish Tayyar Madabushi

The growing popularity and applications of sentiment analysis of social media posts has naturally led to sentiment analysis of posts written in multiple languages, a practice known as code-switching. While recent research into code-switched posts has focused on the use of multilingual word embeddings, these embeddings were not trained on code-switched data. In this work, we present word-embeddings trained on code-switched tweets, specifically those that make use of Spanish and English, known as Spanglish. We explore the embedding space to discover how they capture the meanings of words in both languages. We test the effectiveness of these embeddings by participating in SemEval 2020 Task 9: ~\emph{Sentiment Analysis on Code-Mixed Social Media Text}. We utilised them to train a sentiment classifier that achieves an F-1 score of 0.722. This is higher than the baseline for the competition of 0.656, with our team (codalab username \emph{francesita}) ranking 14 out of 29 participating teams, beating the baseline.

📄 PDF Abstract BibTeX arXiv:2006.04597

Code (1)

francesita/CS-Embed-SemEval2020 공식 구현 tf

Tasks

Multilingual Word EmbeddingsSentiment AnalysisWord Embeddings

Similar Papers 제목 키워드 기반

GLUECoS: An Evaluation Benchmark for Code-Switched NLP

2020-07-01 · ACL 2020 6 · Simran Khanuja, D, S apat, ipan 외

Code-switching is the use of more than one language in the same conversation or utterance. Recently, multilingual contextual embedding models, trained on multiple monolingual corpora, have shown promising results on cros…

Language Identificationnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+5

Code-Switched Named Entity Recognition with Embedding Attention

2018-07-01 · WS 2018 7 · Changhan Wang, Kyunghyun Cho, Douwe Kiela

We describe our work for the CALCS 2018 shared task on named entity recognition on code-switched data. Our system ranked first place for MS Arabic-Egyptian named entity recognition and third place for English-Spanish.

Language Identificationnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+1

GLUECoS : An Evaluation Benchmark for Code-Switched NLP

2020-04-26 · Simran Khanuja, Sandipan Dandapat, Anirudh Srinivasan, Sunayana Sitaram 외

Code-switching is the use of more than one language in the same conversation or utterance. Recently, multilingual contextual embedding models, trained on multiple monolingual corpora, have shown promising results on cros…

Language Identificationnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+5

The Effectiveness of Intermediate-Task Training for Code-Switched Natural Language Understanding

2021-07-21 · EMNLP (MRL) 2021 11 · Archiki Prasad, Mohammad Ali Rehan, Shreya Pathak, Preethi Jyothi

While recent benchmarks have spurred a lot of new work on improving the generalization of pretrained multilingual language models on multilingual tasks, techniques to improve code-switched natural language understanding …

Language ModellingNatural Language InferenceNatural Language UnderstandingPretrained Multilingual Language Models+2

BLCU-ICALL at SemEval-2022 Task 1: Cross-Attention Multitasking Framework for Definition Modeling

2022-04-16 · SemEval (NAACL) 2022 7 · Cunliang Kong, Yujie Wang, Ruining Chong, Liner Yang 외

This paper describes the BLCU-ICALL system used in the SemEval-2022 Task 1 Comparing Dictionaries and Word Embeddings, the Definition Modeling subtrack, achieving 1st on Italian, 2nd on Spanish and Russian, and 3rd on En…

Language ModelingLanguage ModellingWord Embeddings