paper-with-me

Papers

DetoxBench: Benchmarking Large Language Models for Multitask Fraud & Abuse Detection

2024-09-09 · Joymallya Chakraborty, Wei Xia, Anirban Majumder, Dan Ma, Walid Chaabene, Naveed Janvekar

Large language models (LLMs) have demonstrated remarkable capabilities in natural language processing tasks. However, their practical application in high-stake domains, such as fraud and abuse detection, remains an area that requires further exploration. The existing applications often narrowly focus on specific tasks like toxicity or hate speech detection. In this paper, we present a comprehensive benchmark suite designed to assess the performance of LLMs in identifying and mitigating fraudulent and abusive language across various real-world scenarios. Our benchmark encompasses a diverse set of tasks, including detecting spam emails, hate speech, misogynistic language, and more. We evaluated several state-of-the-art LLMs, including models from Anthropic, Mistral AI, and the AI21 family, to provide a comprehensive assessment of their capabilities in this critical domain. The results indicate that while LLMs exhibit proficient baseline performance in individual fraud and abuse detection tasks, their performance varies considerably across tasks, particularly struggling with tasks that demand nuanced pragmatic reasoning, such as identifying diverse forms of misogynistic language. These findings have important implications for the responsible development and deployment of LLMs in high-risk applications. Our benchmark suite can serve as a tool for researchers and practitioners to systematically evaluate LLMs for multi-task fraud detection and drive the creation of more robust, trustworthy, and ethically-aligned systems for fraud and abuse detection.

📄 PDF Abstract BibTeX arXiv:2409.06072

Code (0)

등록된 구현이 없습니다.

Tasks

Abuse DetectionAbusive LanguageBenchmarkingFraud DetectionHate Speech Detection

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically
Focus 설명 없음

Similar Papers 제목 키워드 기반

Multi-task CNN Behavioral Embedding Model For Transaction Fraud Detection

2024-11-29 · Bo Qu, Zhurong Wang, Minghao Gu, Daisuke Yagi 외

The burgeoning e-Commerce sector requires advanced solutions for the detection of transaction fraud. With an increasing risk of financial information theft and account takeovers, deep learning methods have become integra…

Fraud DetectionInductive Bias

FraudSMSWalker: Benchmarking Agentic Large Language Models for SMS-to-Webpage Fraud Detection

2026-06-15 · Y. H. Zhou, Z. M. Ma, Y. J. Zhou, Y. T. Li 외 arxiv

SMS fraud is increasingly cross-channel: a message directs the user to a webpage, and the final risk depends on how the SMS claim aligns with the page content and requested user action. However, existing evaluations eith…

Fraud Detection

Multitask Prompted Training Enables Zero-Shot Task Generalization

2021-10-15 · ICLR 2022 4 · Victor Sanh, Albert Webson, Colin Raffel, Stephen H. Bach 외

Large language models have recently been shown to attain reasonable zero-shot generalization on a diverse set of tasks (Brown et al., 2020). It has been hypothesized that this is a consequence of implicit multitask learn…

BenchmarkingDecoderLanguage ModellingPrompt Engineering+1

Multitask finetuning and acceleration of chemical pretrained models for small molecule drug property prediction

2025-10-14 · Matthew Adrian, Yunsie Chung, Kevin Boyd, Saee Paliwal 외 arxiv

Chemical pretrained models, sometimes referred to as foundation models, are receiving considerable interest for drug discovery applications. The general chemical knowledge extracted from self-supervised training has the …

Graph Neural NetworkMulti-Task LearningDrug Discovery

Benchmarking Offline Reinforcement Learning Algorithms for E-Commerce Order Fraud Evaluation

2022-12-05 · Soysal Degirmenci, Chris Jones

Amazon and other e-commerce sites must employ mechanisms to protect their millions of customers from fraud, such as unauthorized use of credit cards. One such mechanism is order fraud evaluation, where systems evaluate o…

BenchmarkingBinary ClassificationOffline RLreinforcement-learning+1