TAD-Bench: A Comprehensive Benchmark for Embedding-Based Text Anomaly Detection
Text anomaly detection is crucial for identifying spam, misinformation, and offensive language in natural language processing tasks. Despite the growing adoption of embedding-based methods, their effectiveness and generalizability across diverse application scenarios remain under-explored. To address this, we present TAD-Bench, a comprehensive benchmark designed to systematically evaluate embedding-based approaches for text anomaly detection. TAD-Bench integrates multiple datasets spanning different domains, combining state-of-the-art embeddings from large language models with a variety of anomaly detection algorithms. Through extensive experiments, we analyze the interplay between embeddings and detection methods, uncovering their strengths, weaknesses, and applicability to different tasks. These findings offer new perspectives on building more robust, efficient, and generalizable anomaly detection systems for real-world applications.
Code (0)
등록된 구현이 없습니다.
Tasks
Anomaly DetectionMisinformationSimilar Papers 제목 키워드 기반
Text-ADBench: Text Anomaly Detection Benchmark based on LLMs Embedding
Text anomaly detection is a critical task in natural language processing (NLP), with applications spanning fraud detection, misinformation identification, spam detection and content moderation, etc. Despite significant a…
Anomaly DetectionFraud DetectionSpam detectionNLP-ADBench: NLP Anomaly Detection Benchmark
Anomaly detection (AD) is a critical machine learning task with diverse applications in web systems, including fraud detection, content moderation, and user behavior analysis. Despite its significance, AD in natural lang…
Anomaly DetectionFraud DetectionModel SelectionSLSG: Industrial Image Anomaly Detection by Learning Better Feature Embeddings and One-Class Classification
Industrial image anomaly detection under the setting of one-class classification has significant practical value. However, most existing models struggle to extract separable feature representations when performing featur…
Anomaly DetectionClassificationOne-Class ClassificationSelf-Supervised LearningAnomalyAgent: Training-Free Agentic Models for Zero-/Few-Shot Anomaly Detection
Benefiting from generalizability of vision-language models (VLMs) such as CLIP, many zero-/few-shot anomaly detection (AD) approaches have achieved impressive detection performance across various datasets. Nevertheless, …
Anomaly DetectionUncovering What Why and How: A Comprehensive Benchmark for Causation Understanding of Video Anomaly
Video anomaly understanding (VAU) aims to automatically comprehend unusual occurrences in videos thereby enabling various applications such as traffic surveillance and industrial manufacturing. While existing VAU ben…
Anomaly Detection