paper-with-me

홈 › Papers

LEBANONUPRISING: a thorough study of Lebanese tweets

2020-09-30 · Reda Khalaf, Mireille Makary

Recent studies showed a huge interest in social networks sentiment analysis. Twitter, which is a microblogging service, can be a great source of information on how the users feel about a certain topic, or what their opinion is regarding a social, economic and even political matter. On October 17, Lebanon witnessed the start of a revolution; the LebanonUprising hashtag became viral on Twitter. A dataset consisting of a 100,0000 tweets was collected between 18 and 21 October. In this paper, we conducted a sentiment analysis study for the tweets in spoken Lebanese Arabic related to the LebanonUprising hashtag using different machine learning algorithms. The dataset was manually annotated to measure the precision and recall metrics and to compare between the different algorithms. Furthermore, the work completed in this paper provides two more contributions. The first is related to building a Lebanese to Modern Standard Arabic mapping dictionary that was used for the preprocessing of the tweets and the second is an attempt to move from sentiment analysis to emotion detection using emojis, and the two emotions we tried to predict were the "sarcastic" and "funny" emotions. We built a training set from the tweets collected in October 2019 and then we used this set to predict sentiments and emotions of the tweets we collected between May and August 2020. The analysis we conducted shows the variation in sentiments, emotions and users between the two datasets. The results we obtained seem satisfactory especially considering that there was no previous or similar work done involving Lebanese Arabic tweets, to our knowledge.

📄 PDF Abstract BibTeX arXiv:2009.14459

Code (0)

등록된 구현이 없습니다.

Tasks

Sentiment Analysis

Similar Papers 제목 키워드 기반

Fine-Tuning LLMs for Low-Resource Dialect Translation: The Case of Lebanese

2025-04-30 · Silvana Yakhni, Ali Chehab

This paper examines the effectiveness of Large Language Models (LLMs) in translating the low-resource Lebanese dialect, focusing on the impact of culturally authentic data versus larger translated datasets. We compare th…

Translation

DiaLex: A Benchmark for Evaluating Multidialectal Arabic Word Embeddings

2020-11-22 · EACL (WANLP) 2021 4 · Muhammad Abdul-Mageed, Shady Elbassuoni, Jad Doughman, AbdelRahim Elmadany 외

Word embeddings are a core component of modern natural language processing systems, making the ability to thoroughly evaluate them a vital task. We describe DiaLex, a benchmark for intrinsic evaluation of dialectal Arabi…

Word Embeddings

Time-Aware Word Embeddings for Three Lebanese News Archives

2020-05-01 · LREC 2020 5 · Jad Doughman, Fatima Abu Salem, Shady Elbassuoni

Word embeddings have proven to be an effective method for capturing semantic relations among distinct terms within a large corpus. In this paper, we present a set of word embeddings learnt from three large Lebanese news …

Optical Character Recognition (OCR)Word Embeddings

Curras + Baladi: Towards a Levantine Corpus

2022-05-19 · LREC 2022 6 · Karim El Haff, Mustafa Jarrar, Tymaa Hammouda, Fadi Zaraket

The processing of the Arabic language is a complex field of research. This is due to many factors, including the complex and rich morphology of Arabic, its high degree of ambiguity, and the presence of several regional v…

Automatic Classification of Students on Twitter Using Simple Profile Information

2020-12-01 · Asian Chapter of the Association for Computational Linguistics 2020 · Lili-Michal Wilson, Christopher Wun

Obtaining social media demographic information using machine learning is important for efficient computational social science research. Automatic age classification has been accomplished with relative success and allows …

Age ClassificationClassification