paper-with-me

홈 › Papers

Beyond Continuity: Challenges of Context Switching in Multi-Turn Dialogue with LLMs

2026-05-10 · Aditya Sinha, Harald Steck, Vito Ostuni, Matteo Rinaldi arxiv

Users interacting with Large Language Models (LLMs) in a multi-turn conversation routinely refine their requests or pivot to new topics. LLMs, however, often miss these topic shifts and carry over irrelevant context from previous turns, leading to inaccurate responses. In this paper, we stress-test the multi-turn understanding of LLMs and study the following two sub-tasks: (1) detecting whether the user pivots or refines in the current turn, and (2) shortlisting relevant context from previous turns. To this end, we construct synthetic benchmarks based on real-world datasets from varied domains, as to simulate context shifts of different levels of difficulty. We then evaluate the zero-shot performance of ten LLMs (open-weight, closed-source and reasoning), and demonstrate that only some reasoning and strongly instructed LLMs are accurate in detecting pivots; open-weight LLMs struggle with the task and frequently carry stale context even with explicit cues; and all models suffer from a position bias. Based on the results, we discuss key takeaways for improving long-term robustness in multi-turn capabilities for LLMs.

📄 PDF Abstract BibTeX arXiv:2605.09268

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

FastSwitch: Optimizing Context Switching Efficiency in Fairness-aware Large Language Model Serving

2024-11-27 · Ao Shen, Zhiyao Li, Mingyu Gao

Serving numerous users and requests concurrently requires good fairness in Large Language Models (LLMs) serving system. This ensures that, at the same cost, the system can meet the Service Level Objectives (SLOs) of more…

FairnessGPULanguage ModelingLanguage Modelling+2

Beyond Bilingual Transfer: Multilingual Code-Switching in Instruction Tuning

2026-05-28 · Shunta Asano, Jeonghun Baek, Toshihiko Yamasaki arxiv

Recent studies have shown that code-switching data (CSD), in which multiple languages are mixed within the same context, can improve cross-lingual transfer and multilingual alignment in large language models (LLMs). Howe…

Cross-Lingual Transfer

OLA: Output Language Alignment in Code-Switched LLM Interactions

2026-01-07 · Juhyun Oh, Haneul Yoo, Faiz Ghifari Haznitrama, Alice Oh arxiv

Code-switching, alternating between languages within a conversation, is natural for multilingual users, yet poses fundamental challenges for large language models (LLMs). When a user code-switches in their prompt to an L…

Multi-Agent Context Learning Strategy for Interference-Aware Beam Allocation in mmWave Vehicular Communications

2024-01-04 · Abdulkadir Kose, Haeyoung Lee, Chuan Heng Foh, Mohammad Shojafar

Millimeter wave (mmWave) has been recognized as one of key technologies for 5G and beyond networks due to its potential to enhance channel bandwidth and network capacity. The use of mmWave for various applications includ…

Fluidity Index: Next-Generation Super-intelligence Benchmarks

2025-10-23 · Eric Ngoiya, Tianshu Bao arxiv

This paper introduces the Fluidity Index (FI) to quantify model adaptability in dynamic, scaling environments. The benchmark evaluates response accuracy based on deviations in initial, current, and future environment sta…