Corpus Creation for Sentiment Analysis in Code-Mixed Tulu Text
Sentiment Analysis (SA) employing code-mixed data from social media helps in getting insights to the data and decision making for various applications. One such application is to analyze users’ emotions from comments of videos on YouTube. Social media comments do not adhere to the grammatical norms of any language and they often comprise a mix of languages and scripts. The lack of annotated code-mixed data for SA in a low-resource language like Tulu makes the SA a challenging task. To address the lack of annotated code-mixed Tulu data for SA, a gold standard trlingual code-mixed Tulu annotated corpus of 7,171 YouTube comments is created. Further, Machine Learning (ML) algorithms are employed as baseline models to evaluate the developed dataset and the performance of the ML algorithms are found to be encouraging.
Code (0)
등록된 구현이 없습니다.
Tasks
Decision MakingSentiment AnalysisSimilar Papers 제목 키워드 기반
Corpus Creation for Sentiment Analysis in Code-Mixed Tamil-English Text
Understanding the sentiment of a comment from a video or an image is an essential task in many applications. Sentiment analysis of a text can be useful for various decision-making processes. One such application is to an…
Decision MakingSentiment AnalysisCreation of Corpus and Analysis in Code-Mixed Kannada-English Social Media Data for POS Tagging
Part-of-Speech (POS) is one of the essential tasks for many Natural Language Processing (NLP) applications. There has been a significant amount of work done in POS tagging for resource-rich languages. POS tagging is an e…
coreference-resolutionCoreference Resolutionnamed-entity-recognitionNamed Entity Recognition+5A Sentiment Analysis Dataset for Code-Mixed Malayalam-English
There is an increasing demand for sentiment analysis of text from social media which are mostly code-mixed. Systems trained on monolingual data fail for code-mixed data due to the complexity of mixing at different levels…
Sentiment AnalysisPreparing Bengali-English Code-Mixed Corpus for Sentiment Analysis of Indian Languages
Analysis of informative contents and sentiments of social users has been attempted quite intensively in the recent past. Most of the systems are usable only for monolingual data and fails or gives poor results when used …
Sentiment AnalysisTAGUnsupervised Sentiment Analysis for Code-mixed Data
Code-mixing is the practice of alternating between two or more languages. Mostly observed in multilingual societies, its occurrence is increasing and therefore its importance. A major part of sentiment analysis research …
Sentiment AnalysisZero-Shot Learning