paper-with-me

Papers

Transient Turn Injection: Exposing Stateless Multi-Turn Vulnerabilities in Large Language Models

2026-04-23 · Naheed Rayhan, Sohely Jahan arxiv

Large language models (LLMs) are increasingly integrated into sensitive workflows, raising the stakes for adversarial robustness and safety. This paper introduces Transient Turn Injection(TTI), a new multi-turn attack technique that systematically exploits stateless moderation by distributing adversarial intent across isolated interactions. TTI leverages automated attacker agents powered by large language models to iteratively test and evade policy enforcement in both commercial and open-source LLMs, marking a departure from conventional jailbreak approaches that typically depend on maintaining persistent conversational context. Our extensive evaluation across state-of-the-art models-including those from OpenAI, Anthropic, Google Gemini, Meta, and prominent open-source alternatives-uncovers significant variations in resilience to TTI attacks, with only select architectures exhibiting substantial inherent robustness. Our automated blackbox evaluation framework also uncovers previously unknown model specific vulnerabilities and attack surface patterns, especially within medical and high stakes domains. We further compare TTI against established adversarial prompting methods and detail practical mitigation strategies, such as session level context aggregation and deep alignment approaches. Our study underscores the urgent need for holistic, context aware defenses and continuous adversarial testing to future proof LLM deployments against evolving multi-turn threats.

📄 PDF Abstract BibTeX arXiv:2604.21860

Code (0)

등록된 구현이 없습니다.

Tasks

Adversarial Robustness

Similar Papers 제목 키워드 기반

Benchmarking Factual Robustness of LLMs via Multi-conversation Persuasion

2026-09-15 · Zhuoang Cai arxiv

As Large Language Models (LLMs) increasingly serve as primary knowledge retrieval interfaces, their robustness against \textit{persuasion attacks}---attempts to inject misinformation or enforce counterfactuals---has beco…

DeepContext: Stateful Real-Time Detection of Multi-Turn Adversarial Intent Drift in LLMs

2026-02-18 · Justin Albrethsen, Yash Datta, Kunal Kumar, Sharath Rajasekar arxiv

While Large Language Model (LLM) capabilities have scaled, safety guardrails remain largely stateless, treating multi-turn dialogues as a series of disconnected events. This lack of temporal awareness facilitates a "Safe…

AGENTSERVESIM: A Hardware-aware Simulator for Multi-Turn LLM Agent Serving

2026-06-08 · Rakibul Hasan Rajib, Mengxin Zheng, Qian Lou arxiv

Multi-turn LLM agents interleave model calls with external tool invocations, shifting serving from stateless request processing to stateful program execution. Serving these workloads requires scheduling, KV-cache managem…

Fast voltage boosters to improve transient stability of power systems with 100% of grid-forming VSC-based generation

2021-06-08 · Régulo E. Ávila-Martínez, Javier Renedo, Luis Rouco, Aurelio García-Cerrada 외

Grid-forming voltage source converter (GF-VSC) has been identified as the key technology for the operation of future converter-dominated power systems. Among many other issues, transient stability of this type of power s…

Coordinated control in multi-terminal VSC-HVDC systems to improve transient stability: Impact on electromechanical-oscillation damping

2022-07-29 · Javier Renedo, Luis Rouco, Aurelio Garcia-Cerrada, Lukas Sigrist

Multi-terminal high-voltage Direct Current technology based on Voltage-Source Converter stations (VSC-MTDC) is expected to be one of the most important contributors to the future of electric power systems. In fact, among…