paper-with-me

Papers

KanCMD: Kannada CodeMixed Dataset for Sentiment Analysis and Offensive Language Detection

2020-12-01 · COLING (PEOPLES) 2020 12 · Adeep Hande, Ruba Priyadharshini, Bharathi Raja Chakravarthi

We introduce Kannada CodeMixed Dataset (KanCMD), a multi-task learning dataset for sentiment analysis and offensive language identification. The KanCMD dataset highlights two real-world issues from the social media text. First, it contains actual comments in code mixed text posted by users on YouTube social media, rather than in monolingual text from the textbook. Second, it has been annotated for two tasks, namely sentiment analysis and offensive language detection for under-resourced Kannada language. Hence, KanCMD is meant to stimulate research in under-resourced Kannada language on real-world code-mixed social media text and multi-task learning. KanCMD was obtained by crawling the YouTube, and a minimum of three annotators annotates each comment. We release KanCMD 7,671 comments for multitask learning research purpose.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Language IdentificationMulti-Task LearningSentiment Analysis

Similar Papers 제목 키워드 기반

Curriculum Learning Strategies for Hindi-English Codemixed Sentiment Analysis

2019-06-18 · Anirudh Dahiya, Neeraj Battan, Manish Shrivastava, Dipti Mishra Sharma

Sentiment Analysis and other semantic tasks are commonly used for social media textual analysis to gauge public opinion and make sense from the noise on social media. The language used on social media not only commonly d…

Sentiment Analysis

IRLab\_DAIICT at SemEval-2020 Task 9: Machine Learning and Deep Learning Methods for Sentiment Analysis of Code-Mixed Tweets

2020-12-01 · SEMEVAL 2020 · Apurva Parikh, Abhimanyu Singh Bisht, Prasenjit Majumder

The paper describes systems that our team IRLab{\_}DAIICT employed for the shared task Sentiment Analysis for Code-Mixed Social Media Text in SemEval 2020. We conducted our experiments on a Hindi-English CodeMixed Tweet …

regressionSentiment Analysis

Memotion 3: Dataset on Sentiment and Emotion Analysis of Codemixed Hindi-English Memes

2023-03-17 · Shreyash Mishra, S Suryavardan, Parth Patwa, Megha Chakraborty 외

Memes are the new-age conveyance mechanism for humor on social media sites. Memes often include an image and some text. Memes can be used to promote disinformation or hatred, thus it is crucial to investigate in details.…

Emotion Recognition

Building a Kannada POS Tagger Using Machine Learning and Neural Network Models

2018-08-09 · Ketan Kumar Todi, Pruthwik Mishra, Dipti Misra Sharma

POS Tagging serves as a preliminary task for many NLP applications. Kannada is a relatively poor Indian language with very limited number of quality NLP tools available for use. An accurate and reliable POS Tagger is ess…

BIG-bench Machine LearningDependency Parsingnamed-entity-recognitionNamed Entity Recognition+5

Offensive Language Identification in Low-resourced Code-mixed Dravidian languages using Pseudo-labeling

2021-08-27 · Adeep Hande, Karthik Puranik, Konthala Yasaswini, Ruba Priyadharshini 외

Social media has effectively become the prime hub of communication and digital marketing. As these platforms enable the free manifestation of thoughts and facts in text, images and video, there is an extensive need to sc…

Language IdentificationMarketing