paper-with-me

홈 › Papers

Real-Time Deadlines Reveal Temporal Awareness Failures in LLM Strategic Dialogues

2026-01-19 · Neil K. R. Sehgal, Sharath Chandra Guntuku, Lyle Ungar arxiv

Large Language Models (LLMs) generate text token-by-token in discrete time, yet real-world communication, from therapy sessions to business negotiations, critically depends on continuous time constraints. Current LLM architectures and evaluation protocols rarely test for temporal awareness under real-time deadlines. We use simulated negotiations between paired agents under strict deadlines to investigate how LLMs adjust their behavior in time-sensitive settings. In a control condition, agents know only the global time limit. In a time-aware condition, they receive remaining-time updates at each turn. Deal closure rates are substantially higher (32\% vs. 4\% for GPT-5.1) and offer acceptances are sixfold higher in the time-aware condition than in the control, suggesting LLMs struggle to internally track elapsed time. However, the same LLMs achieve near-perfect deal closure rates ($\geq$95\%) under turn-based limits, revealing the failure is in temporal tracking rather than strategic reasoning. These effects replicate across negotiation scenarios and models, illustrating a systematic lack of LLM time awareness that will constrain LLM deployment in many time-sensitive applications.

📄 PDF Abstract BibTeX arXiv:2601.13206

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

On Cost-Aware Sequential Hypothesis Testing with Random Costs and Action Cancellation

2025-12-22 · George Vershinin, Asaf Cohen, Omer Gurewitz arxiv

We study a variant of cost-aware sequential hypothesis testing in which a single active Decision Maker (DM) selects actions with positive, random costs to identify the true hypothesis under an average error constraint, w…

Tempora: Characterising the Time-Contingent Utility of Online Test-Time Adaptation

2026-02-05 · Sudarshan Sreeram, Young D. Kwon, Cecilia Mascolo arxiv

Test-time adaptation (TTA) offers a compelling remedy for machine learning (ML) models that degrade under domain shifts, improving generalisation on-the-fly with only unlabelled samples. This flexibility suits real deplo…

Test-time Adaptation

A Matter of Time: Revealing the Structure of Time in Vision-Language Models

2025-10-22 · Nidham Tekaya, Manuela Waldner, Matthias Zeppelzauer arxiv

Large-scale vision-language models (VLMs) such as CLIP have gained popularity for their generalizable and expressive multimodal representations. By leveraging large-scale training data with diverse textual metadata, VLMs…

Large Language Models Lack Temporal Awareness of Medical Knowledge

2026-05-13 · Zihan Guan, Qiao Jin, Guangzhi Xiong, Fangyuan Chen 외 arxiv

The existing methods for evaluating the medical knowledge of Large Language Models (LLMs) are largely based on atemporal examination-style benchmarks, while in reality, medical knowledge is inherently dynamic and continu…

Compiling Metric Temporal Answer Set Programming

2025-06-09 · Arvid Becker, Pedro Cabalar, Martin Diéguez, Javier Romero 외

We develop a computational approach to Metric Answer Set Programming (ASP) to allow for expressing quantitative temporal constrains, like durations and deadlines. A central challenge is to maintain scalability when deali…