paper-with-me

Papers

Exposing LLM Vulnerabilities: Adversarial Scam Detection and Performance

2024-12-01 · Chen-Wei Chang, Shailik Sarkar, Shutonu Mitra, Qi Zhang, Hossein Salemi, Hemant Purohit, Fengxiu Zhang, Michin Hong, Jin-Hee Cho, Chang-Tien Lu

Can we trust Large Language Models (LLMs) to accurately predict scam? This paper investigates the vulnerabilities of LLMs when facing adversarial scam messages for the task of scam detection. We addressed this issue by creating a comprehensive dataset with fine-grained labels of scam messages, including both original and adversarial scam messages. The dataset extended traditional binary classes for the scam detection task into more nuanced scam types. Our analysis showed how adversarial examples took advantage of vulnerabilities of a LLM, leading to high misclassification rate. We evaluated the performance of LLMs on these adversarial scam messages and proposed strategies to improve their robustness.

📄 PDF Abstract BibTeX arXiv:2412.00621

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

SCRIPTMIND: Crime Script Inference and Cognitive Evaluation for LLM-based Social Engineering Scam Detection System

2026-01-20 · Heedou Kim, Changsik Kim, Sanghwa Shin, Jaewoo Kang arxiv

Social engineering scams increasingly employ personalized, multi-turn deception, exposing the limits of traditional detection methods. While Large Language Models (LLMs) show promise in identifying deception, their cogni…

Can LLMs be Scammed? A Baseline Measurement Study

2024-10-14 · Udari Madhushani Sehwag, Kelly Patel, Francesca Mosca, Vineeth Ravi 외

Despite the importance of developing generative AI models that can effectively resist scams, current literature lacks a structured framework for evaluating their vulnerability to such threats. In this work, we address th…

Scam Shield: Multi-Model Voting and Fine-Tuned LLMs Against Adversarial Attacks

2025-11-03 · Chen-Wei Chang, Shailik Sarkar, Hossein Salemi, Hyungmin Kim 외 arxiv

Scam detection remains a critical challenge in cybersecurity as adversaries craft messages that evade automated filters. We propose a Hierarchical Scam Detection System (HSDS) that combines a lightweight multi-model voti…

Enhancing Trust and Safety in Digital Payments: An LLM-Powered Approach

2024-10-21 · Devendra Dahiphale, Naveen Madiraju, Justin Lin, Rutvik Karve 외

Digital payment systems have revolutionized financial transactions, offering unparalleled convenience and accessibility to users worldwide. However, the increasing popularity of these platforms has also attracted malicio…

AI-in-the-Loop: Privacy Preserving Real-Time Scam Detection and Conversational Scambaiting by Leveraging LLMs and Federated Learning

2025-09-04 · Ismail Hossain, Sai Puppala, Md Jahangir Alam, Sajedul Talukder arxiv

Scams exploiting real-time social engineering -- such as phishing, impersonation, and phone fraud -- remain a persistent and evolving threat across digital platforms. Existing defenses are largely reactive, offering limi…

Federated Learning