No comments: Addressing commentary sections in websites' analyses
Removing or extracting the commentary sections from a series of websites is a tedious task, as no standard way to code them is widely adopted. This operation is thus very rarely performed. In this paper, we show that these commentary sections can induce significant biases in the analyses, especially in the case of controversial Highlights $\bullet$ Commentary sections can induce biases in the analysis of websites' contents $\bullet$ Analyzing these sections can be interesting per se. $\bullet$ We illustrate these points using a corpus of anti-vaccine websites. $\bullet$ We provide guidelines to remove or extract these sections.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Public Sphere 2.0: Targeted Commenting in Online News Media
With the increase in online news consumption, to maximize advertisement revenue, news media websites try to attract and retain their readers on their sites. One of the most effective tools for reader engagement is commen…
ArticlesVectorizing Classical Tamil: Representation Learning for Verse-Commentary Pairs
We construct a corpus of 1,262 verse-commentary (urai) pairs from five Classical Tamil source sections, ranging from technical grammatical prose to modern paraphrase, and ask what information representation learning can …
Representation LearningIdentifying Aggression and Toxicity in Comments using Capsule Network
Aggression and related activities like trolling, hate speech etc. involve toxic comments in various forms. These are common scenarios in today{'}s time and websites react by shutting down their comment sections. To tackl…
Data AugmentationTransliterationWord EmbeddingsStance Detection in Prediction Markets: Addressing Imbalanced Trader Commentary via Counterfactual Augmentation and Market Context
Prediction markets such as Polymarket aggregate crowd beliefs into real-time probability estimates, and the comments traders post beneath each market contain rich directional stance signals that prices alone cannot captu…
Stance DetectionMachine Learning Suites for Online Toxicity Detection
To identify and classify toxic online commentary, the modern tools of data science transform raw text into key features from which either thresholding or learning algorithms can make predictions for monitoring offensive …
BIG-bench Machine Learning