paper-with-me

홈 › Papers

Text Analysis in Adversarial Settings: Does Deception Leave a Stylistic Trace?

2019-02-24 · Tommi Gröndahl, N. Asokan

Textual deception constitutes a major problem for online security. Many studies have argued that deceptiveness leaves traces in writing style, which could be detected using text classification techniques. By conducting an extensive literature review of existing empirical work, we demonstrate that while certain linguistic features have been indicative of deception in certain corpora, they fail to generalize across divergent semantic domains. We suggest that deceptiveness as such leaves no content-invariant stylistic trace, and textual similarity measures provide superior means of classifying texts as potentially deceptive. Additionally, we discuss forms of deception beyond semantic content, focusing on hiding author identity by writing style obfuscation. Surveying the literature on both author identification and obfuscation techniques, we conclude that current style transformation methods fail to achieve reliable obfuscation while simultaneously ensuring semantic faithfulness to the original text. We propose that future work in style transformation should pay particular attention to disallowing semantically drastic changes.

📄 PDF Abstract BibTeX arXiv:1902.08939

Code (0)

등록된 구현이 없습니다.

Tasks

text-classificationText Classification

Similar Papers 제목 키워드 기반

AI Deception: Risks, Dynamics, and Controls

2025-11-27 · Boyuan Chen, Sitong Fang, Jiaming Ji, Yanxu Zhu 외 arxiv

As intelligence increases, so does its shadow. AI deception, in which systems induce false beliefs to secure self-beneficial outcomes, has evolved from a speculative concern to an empirically demonstrated risk across lan…

Can Adversarial Code Comments Fool AI Security Reviewers -- Large-Scale Empirical Study of Comment-Based Attacks and Defenses Against LLM Code Analysis

2026-02-18 · Scott Thornton arxiv

AI-assisted code review is widely used to detect vulnerabilities before production release. Prior work shows that adversarial prompt manipulation can degrade large language model (LLM) performance in code generation. We …

Vulnerability DetectionCode Generation

Adversarial Activation Patching: A Framework for Detecting and Mitigating Emergent Deception in Safety-Aligned Transformers

2025-07-12 · Santhosh Kumar Ravindran

Large language models (LLMs) aligned for safety through techniques like reinforcement learning from human feedback (RLHF) often exhibit emergent deceptive behaviors, where outputs appear compliant but subtly mislead or o…

Anomaly Detection

Optimal sensor deception in stochastic environments with partial observability to mislead a robot to a decoy goal

2025-03-07 · Hazhar Rahmani, Mukulika Ghosh, Syed Md Hasnayeen

Deception is a common strategy adapted by autonomous systems in adversarial settings. Existing deception methods primarily focus on increasing opacity or misdirecting agents away from their goal or itinerary. In this wor…

``I Don't Know Where He is Not'': Does Deception Research yet Offer a Basis for Deception Detectives?

2012-04-01 · WS 2012 4 · Anna Vartapetiance, Lee Gillam
Deception Detection