Disambiguating False-Alarm Hashtag Usages in Tweets for Irony Detection
The reliability of self-labeled data is an important issue when the data are regarded as ground-truth for training and testing learning-based models. This paper addresses the issue of false-alarm hashtags in the self-labeled data for irony detection. We analyze the ambiguity of hashtag usages and propose a novel neural network-based model, which incorporates linguistic information from different aspects, to disambiguate the usage of three hashtags that are widely used to collect the training data for irony detection. Furthermore, we apply our model to prune the self-labeled training data. Experimental results show that the irony detection model trained on the less but cleaner training instances outperforms the models trained on all data.
Code (0)
등록된 구현이 없습니다.
Tasks
Opinion MiningSentiment AnalysisStance DetectionSimilar Papers 제목 키워드 기반
Segmenting Hashtags using Automatically Created Training Data
Hashtags, which are commonly composed of multiple words, are increasingly used to convey the actual messages in tweets. Understanding what tweets are saying is getting more dependent on understanding hashtags. Therefore,…
BIG-bench Machine LearningTwitter Trend Extraction: A Graph-based Approach for Tweet and Hashtag Ranking, Utilizing No-Hashtag Tweets
Twitter has become a major platform for users to express their opinions on any topic and engage in debates. User debates and interactions usually lead to massive content regarding a specific topic which is called a Trend…
Hashtag Processing for Enhanced Clustering of Tweets
Rich data provided by tweets have beenanalyzed, clustered, and explored in a variety of studies. Typically those studies focus on named entity recognition, entity linking, and entity disambiguation or clustering. Tweets …
ClusteringEntity DisambiguationEntity Linkingnamed-entity-recognition+4Hashtag-Guided Low-Resource Tweet Classification
Social media classification tasks (e.g., tweet sentiment analysis, tweet stance detection) are challenging because social media posts are typically short, informal, and ambiguous. Thus, training on tweets is challenging …
ClassificationSentiment AnalysisStance DetectionOn Identifying Hashtags in Disaster Twitter Data
Tweet hashtags have the potential to improve the search for information during disaster events. However, there is a large number of disaster-related tweets that do not have any user-provided hashtags. Moreover, only a sm…
Disaster ResponseMulti-Task Learning