Machine Generated Text: A Comprehensive Survey of Threat Models and Detection Methods
Machine generated text is increasingly difficult to distinguish from human authored text. Powerful open-source models are freely available, and user-friendly tools that democratize access to generative models are proliferating. ChatGPT, which was released shortly after the first edition of this survey, epitomizes these trends. The great potential of state-of-the-art natural language generation (NLG) systems is tempered by the multitude of avenues for abuse. Detection of machine generated text is a key countermeasure for reducing abuse of NLG models, with significant technical challenges and numerous open problems. We provide a survey that includes both 1) an extensive analysis of threat models posed by contemporary NLG systems, and 2) the most complete review of machine generated text detection methods to date. This survey places machine generated text within its cybersecurity and social context, and provides strong guidance for future work addressing the most critical threat models, and ensuring detection systems themselves demonstrate trustworthiness through fairness, robustness, and accountability.
Code (0)
등록된 구현이 없습니다.
Tasks
Abuse DetectionFairnessSurveyText DetectionText GenerationSimilar Papers 제목 키워드 기반
Adversarial Attacks in Multimodal Systems: A Practitioner's Survey
The introduction of multimodal models is a huge step forward in Artificial Intelligence. A single model is trained to understand multiple modalities: text, image, video, and audio. Open-source multimodal models have made…
Adversarial AttackSurveyMalicious URL Detection using Machine Learning: A Survey
Malicious URL, a.k.a. malicious website, is a common and serious threat to cybersecurity. Malicious URLs host unsolicited content (spam, phishing, drive-by exploits, etc.) and lure unsuspecting users to become victims of…
BIG-bench Machine LearningSurveyThe Science of Detecting LLM-Generated Texts
The emergence of large language models (LLMs) has resulted in the production of LLM-generated texts that is highly sophisticated and almost indistinguishable from texts written by humans. However, this has also sparked c…
LLM-generated Text DetectionMisinformationText DetectionText GenerationA Survey on Adversarial Robustness of LiDAR-based Machine Learning Perception in Autonomous Vehicles
In autonomous driving, the combination of AI and vehicular technology offers great potential. However, this amalgamation comes with vulnerabilities to adversarial attacks. This survey focuses on the intersection of Adver…
Adversarial RobustnessAutonomous DrivingAutonomous VehiclesA Survey of Privacy Threats and Defense in Vertical Federated Learning: From Model Life Cycle Perspective
Vertical Federated Learning (VFL) is a federated learning paradigm where multiple participants, who share the same set of samples but hold different features, jointly train machine learning models. Although VFL enables c…
Federated LearningSurveyVertical Federated Learning