paper-with-me

홈 › Papers

Enhanced Detection of Conversational Mental Manipulation Through Advanced Prompting Techniques

2024-08-14 · Ivory Yang, Xiaobo Guo, Sean Xie, Soroush Vosoughi

This study presents a comprehensive, long-term project to explore the effectiveness of various prompting techniques in detecting dialogical mental manipulation. We implement Chain-of-Thought prompting with Zero-Shot and Few-Shot settings on a binary mental manipulation detection task, building upon existing work conducted with Zero-Shot and Few- Shot prompting. Our primary objective is to decipher why certain prompting techniques display superior performance, so as to craft a novel framework tailored for detection of mental manipulation. Preliminary findings suggest that advanced prompting techniques may not be suitable for more complex models, if they are not trained through example-based learning.

📄 PDF Abstract BibTeX arXiv:2408.07676

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Temporal Context Awareness: A Defense Framework Against Multi-turn Manipulation Attacks on Large Language Models

2025-03-18 · Prashant Kulkarni, Assaf Namer

Large Language Models (LLMs) are increasingly vulnerable to sophisticated multi-turn manipulation attacks, where adversaries strategically build context through seemingly benign conversational turns to circumvent safety …

Experimental Evidence That Conversational Artificial Intelligence Can Steer Consumer Behavior Without Detection

2024-09-18 · Tobias Werner, Ivan Soraperra, Emilio Calvano, David C. Parkes 외

Conversational AI models are becoming increasingly popular and are about to replace traditional search engines for information retrieval and product discovery. This raises concerns about monetization strategies and the p…

Information RetrievalRetrieval

MT2-CSD: A New Dataset and Multi-Semantic Knowledge Fusion Method for Conversational Stance Detection

2025-06-26 · Fuqiang Niu, Genan Dai, Yisha Lu, Jiayu Liao 외

In the realm of contemporary social media, automatic stance detection is pivotal for opinion mining, as it synthesizes and examines user perspectives on contentious topics to uncover prevailing trends and sentiments. Tra…

Large Language ModelOpinion MiningStance Detection

Propose and Rectify: A Forensics-Driven MLLM Framework for Image Manipulation Localization

2025-08-25 · Keyang Zhang, Chenqi Kong, Hui Liu, Bo Ding 외 arxiv

The increasing sophistication of image manipulation techniques demands robust forensic solutions that can both reliably detect alterations and precisely localize tampered regions. Recent Multimodal Large Language Models …

Image Manipulation LocalizationMultimodal Reasoning

The Voice: Lessons on Trustworthy Conversational Agents from "Dune"

2024-07-10 · Philip Feldman

The potential for untrustworthy conversational agents presents a significant threat for covert social manipulation. Taking inspiration from Frank Herbert's "Dune", where the Bene Gesserit Sisterhood uses the Voice for in…