paper-with-me

홈 › Papers

Aggressive language in an online hacking forum

2018-10-01 · WS 2018 10 · Andrew Caines, Sergio Pastrana, Alice Hutchings, Paula Buttery

We probe the heterogeneity in levels of abusive language in different sections of the Internet, using an annotated corpus of Wikipedia page edit comments to train a binary classifier for abuse detection. Our test data come from the CrimeBB Corpus of hacking-related forum posts and we find that (a) forum interactions are rarely abusive, (b) the abusive language which does exist tends to be relatively mild compared to that found in the Wikipedia comments domain, and tends to involve aggressive posturing rather than hate speech or threats of violence. We observe that the purpose of conversations in online forums tend to be more constructive and informative than those in Wikipedia page edit comments which are geared more towards adversarial interactions, and that this may explain the lower levels of abuse found in our forum data than in Wikipedia comments. Further work remains to be done to compare these results with other inter-domain classification experiments, and to understand the impact of aggressive language in forum conversations.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Abuse DetectionAbusive Languagedomain classification

Similar Papers 제목 키워드 기반

Inferring Discussion Topics about Exploitation of Vulnerabilities from Underground Hacking Forums

2024-05-07 · Felipe Moreno-Vera

The increasing sophistication of cyber threats necessitates proactive measures to identify vulnerabilities and potential exploits. Underground hacking forums serve as breeding grounds for the exchange of hacking techniqu…

Cream Skimming the Underground: Identifying Relevant Information Points from Online Forums

2023-08-03 · Felipe Moreno-Vera, Mateus Nogueira, Cainã Figueiredo, Daniel Sadoc Menasché 외

This paper proposes a machine learning-based approach for detecting the exploitation of vulnerabilities in the wild by monitoring underground hacking forums. The increasing volume of posts discussing exploitation in the …

Predicting Cyber Events by Leveraging Hacker Sentiment

2018-04-14 · Ashok Deb, Kristina Lerman, Emilio Ferrara

Recent high-profile cyber attacks exemplify why organizations need better cyber defenses. Cyber threats are hard to accurately predict because attackers usually try to mask their traces. However, they often discuss explo…

Sentiment AnalysisTime SeriesTime Series Analysis

EUREKHA: Enhancing User Representation for Key Hackers Identification in Underground Forums

2024-11-08 · Abdoul Nasser Hassane Amadou, Anas Motii, Saida Elouardi, El Houcine Bergou

Underground forums serve as hubs for cybercriminal activities, offering a space for anonymity and evasion of conventional online oversight. In these hidden communities, malicious actors collaborate to exchange illicit kn…

Graph Neural NetworkLarge Language Model

Detecting Trending Terms in Cybersecurity Forum Discussions

2020-11-01 · EMNLP (WNUT) 2020 11 · Jack Hughes, Seth Aycock, Andrew Caines, Paula Buttery 외

We present a lightweight method for identifying currently trending terms in relation to a known prior of terms, using a weighted log-odds ratio with an informative prior. We apply this method to a dataset of posts from a…

Information RetrievalRetrieval