paper-with-me

홈 › Papers

Beyond Context: Large Language Models' Failure to Grasp Users' Intent

2025-12-24 · Ahmed M. Hussain, Salahuddin Salahuddin arxiv

Current Large Language Models (LLMs) safety approaches focus on explicitly harmful content while overlooking a critical vulnerability: the inability to understand context and recognize user intent. This creates exploitable vulnerabilities that malicious users can systematically leverage to circumvent safety mechanisms. We empirically evaluate multiple state-of-the-art LLMs, including ChatGPT, Claude, Gemini, and DeepSeek. Our analysis demonstrates the circumvention of reliable safety mechanisms through emotional framing, progressive revelation, and academic justification techniques. Notably, reasoning-enabled configurations amplified rather than mitigated the effectiveness of exploitation, increasing factual precision while failing to interrogate the underlying intent. The exception was Claude Opus 4.1, which prioritized intent detection over information provision in some use cases. This pattern reveals that current architectural designs create systematic vulnerabilities. These limitations require paradigmatic shifts toward contextual understanding and intent recognition as core safety capabilities rather than post-hoc protective mechanisms.

📄 PDF Abstract BibTeX arXiv:2512.21110

Code (0)

등록된 구현이 없습니다.

Tasks

Intent RecognitionIntent Detection

Similar Papers 제목 키워드 기반

Beyond Visual Grasping: Benchmarking Complex Grasping from Detection to Execution

2026-07-15 · Hanyi Zhang, Khang Nguyen, Charith Munasinghe, Basu Hela 외 arxiv

Robust robotic grasping remains a fundamental challenge for complex real-world applications. Recent advances in large-scale models demonstrate promising capabilities for reasoning in robotic tasks. However, existing benc…

Robotic Grasping

A Real-World Grasping-in-Clutter Performance Evaluation Benchmark for Robotic Food Waste Sorting

2026-02-21 · Moniesha Thilakarathna, Xing Wang, Min Wang, David Hinwood 외 arxiv

Food waste management is critical for sustainability, yet inorganic contaminants hinder recycling potential. Robotic automation accelerates sorting through automated contaminant removal. Nevertheless, the diverse and unp…

Robotic GraspingPose Estimation

A Physical Agentic Loop for Language-Guided Grasping with Execution-State Monitoring

2026-04-08 · Wenze Wang, Mehdi Hosseinzadeh, Feras Dayoub arxiv

Robotic manipulation systems that follow language instructions often execute grasp primitives in a largely single-shot manner: a model proposes an action, the robot executes it, and failures such as empty grasps, slips, …

Grasp-Then-Plan with Failure Attribution: A Closed Two-Stage Framework for Precise and Generalizable Robotic Manipulation

2026-06-02 · Jiahao Xu, Peiyuan Wang, Hanzhuo Zhang, Zihao Yu 외 arxiv

In robotic manipulation, the tight coupling between grasping and motion planning often obscures the true source of failure, leading to inefficient trial-and-error. To enable efficient long-horizon manipulation, we propos…

Motion Planning

GraspCorrect: Robotic Grasp Correction via Vision-Language Model-Guided Feedback

2025-03-19 · Sungjae Lee, Yeonjoo Hong, Kwang In Kim

Despite significant advancements in robotic manipulation, achieving consistent and stable grasping remains a fundamental challenge, often limiting the successful execution of complex tasks. Our analysis reveals that even…

Language ModelingLanguage ModellingQuestion AnsweringVisual Question Answering