paper-with-me

홈 › Papers

Temporal Referential Consistency: Do LLMs Favor Sequences Over Absolute Time References?

2025-10-17 · Ashutosh Bajpai, Tanmoy Chakraborty arxiv

The increasing acceptance of large language models (LLMs) as an alternative to knowledge sources marks a significant paradigm shift across various domains, including time-sensitive fields such as law, healthcare, and finance. To fulfill this expanded role, LLMs must not only be factually accurate but also demonstrate consistency across temporal dimensions, necessitating robust temporal reasoning capabilities. Despite this critical requirement, efforts to ensure temporal consistency in LLMs remain scarce including noticeable absence of endeavors aimed at evaluating or augmenting LLMs across temporal references in time-sensitive inquiries. In this paper, we seek to address this gap by introducing a novel benchmark entitled temporal referential consistency, accompanied by a resource TEMP-ReCon designed to benchmark a wide range of both open-source and closed-source LLMs with various linguistic contexts characterized by differing resource richness (including English, French, and Romanian). The findings emphasis that LLMs do exhibit insufficient temporal referent consistency. To address this, we propose \newmodel, a reasoning path alignment-based model that aims to enhance the temporal referential consistency of LLMs. Our empirical experiments substantiate the efficacy of UnTRaP compared to several baseline models.

📄 PDF Abstract BibTeX arXiv:2510.15513

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

SPARROW: Learning Spatial Precision and Temporal Referential Consistency in Pixel-Grounded Video MLLMs

2026-03-12 · Mohamad Alansari, Naufal Suryanto, Divya Velayudhan, Sajid Javed 외 arxiv

Multimodal large language models (MLLMs) have advanced from image-level reasoning to pixel-level grounding, but extending these capabilities to videos remains challenging as models must achieve spatial precision and temp…

Visual Grounding

Measuring the Inconsistency of Large Language Models in Preferential Ranking

2024-10-11 · Xiutian Zhao, Ke Wang, Wei Peng

Despite large language models' (LLMs) recent advancements, their bias and hallucination issues persist, and their ability to offer consistent preferential rankings remains underexplored. This study investigates the capac…

DiagnosticHallucination

CapRiCorn-1K: A Comprehensive Benchmark for Video Captioning and Subject Referential Consistency Across Temporal Scales

2026-06-20 · Xinlong Chen, Jiafu Tang, Yue Ding, Yizhuo Jia 외 arxiv

Accurate and comprehensive video captions with consistent subject references are critical for downstream understanding and generation tasks. However, few existing benchmarks can objectively and comprehensively evaluate t…

Video Captioning

Time Is a Feature: Exploiting Temporal Dynamics in Diffusion Language Models

2025-08-12 · Wen Wang, Bozhen Fang, Chenchen Jing, Yongliang Shen 외 arxiv

Diffusion large language models (dLLMs) generate text through iterative denoising, yet current decoding strategies discard rich intermediate predictions in favor of the final output. Our work here reveals a critical phen…

Network Formation and Dynamics Among Multi-LLMs

2024-02-16 · Marios Papachristou, Yuan Yuan

Social networks fundamentally shape human opinions, behaviors, and the dissemination of information. As large language models (LLMs) like GPT, Claude, and Llama increasingly integrate into social and professional setting…

Decision Making