paper-with-me

홈 › Papers

Chinese Offensive Language Detection:Current Status and Future Directions

2024-03-27 · Yunze Xiao, Houda Bouamor, Wajdi Zaghouani

Despite the considerable efforts being made to monitor and regulate user-generated content on social media platforms, the pervasiveness of offensive language, such as hate speech or cyberbullying, in the digital space remains a significant challenge. Given the importance of maintaining a civilized and respectful online environment, there is an urgent and growing need for automatic systems capable of detecting offensive speech in real time. However, developing effective systems for processing languages such as Chinese presents a significant challenge, owing to the language's complex and nuanced nature, which makes it difficult to process automatically. This paper provides a comprehensive overview of offensive language detection in Chinese, examining current benchmarks and approaches and highlighting specific models and tools for addressing the unique challenges of detecting offensive language in this complex language. The primary objective of this survey is to explore the existing techniques and identify potential avenues for further research that can address the cultural and linguistic complexities of Chinese.

📄 PDF Abstract BibTeX arXiv:2403.18314

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

COLD: A Benchmark for Chinese Offensive Language Detection

2022-01-16 · Jiawen Deng, Jingyan Zhou, Hao Sun, Chujie Zheng 외

Offensive language detection is increasingly crucial for maintaining a civilized social media platform and deploying pre-trained language models. However, this task in Chinese is still under exploration due to the scarci…

Cross-Cultural Transfer Learning for Chinese Offensive Language Detection

2023-03-31 · Li Zhou, Laura Cabello, Yong Cao, Daniel Hershcovich

Detecting offensive language is a challenging task. Generalizing across different cultures and languages becomes even more challenging: besides lexical, syntactic and semantic differences, pragmatic aspects such as cultu…

Cultural Vocal Bursts Intensity PredictionFew-Shot LearningTransfer Learning

ToxiCloakCN: Evaluating Robustness of Offensive Language Detection in Chinese with Cloaking Perturbations

2024-06-18 · Yunze Xiao, Yujia Hu, Kenny Tsu Wei Choo, Roy Ka-Wei Lee

Detecting hate speech and offensive language is essential for maintaining a safe and respectful digital environment. This study examines the limitations of state-of-the-art large language models (LLMs) in identifying off…

Lost in Pronunciation: Detecting Chinese Offensive Language Disguised by Phonetic Cloaking Replacement

2025-07-10 · Haotan Guo, Jianfei He, Jiayuan Ma, Hongbin Na 외 arxiv

Phonetic Cloaking Replacement (PCR), defined as the deliberate use of homophonic or near-homophonic variants to hide toxic intent, has become a major obstacle to Chinese content moderation. While this problem is well-rec…

AustroTox: A Dataset for Target-Based Austrian German Offensive Language Detection

2024-06-12 · Pia Pachinger, Janis Goldzycher, Anna Maria Planitzer, Wojciech Kusa 외

Model interpretability in toxicity detection greatly profits from token-level annotations. However, currently such annotations are only available in English. We introduce a dataset annotated for offensive language detect…