paper-with-me

Papers

TiEBe: Tracking Language Model Recall of Notable Worldwide Events Through Time

2025-01-13 · Thales Sales Almeida, Giovana Kerche Bonás, João Guilherme Alves Santos, Hugo Abonizio, Rodrigo Nogueira

As the knowledge landscape evolves and large language models (LLMs) become increasingly widespread, there is a growing need to keep these models updated with current events. While existing benchmarks assess general factual recall, few studies explore how LLMs retain knowledge over time or across different regions. To address these gaps, we present the Timely Events Benchmark (TiEBe), a dataset of over 23,000 question-answer pairs centered on notable global and regional events, spanning more than 10 years of events, 23 regions, and 13 languages. TiEBe leverages structured retrospective data from Wikipedia to identify notable events through time. These events are then used to construct a benchmark to evaluate LLMs' understanding of global and regional developments, grounded in factual evidence beyond Wikipedia itself. Our results reveal significant geographic disparities in factual recall, emphasizing the need for more balanced global representation in LLM training. We also observe a Pearson correlation of more than 0.7 between models' performance in TiEBe and various countries' socioeconomic indicators, such as HDI. In addition, we examine the impact of language on factual recall by posing questions in the native language of the region where each event occurred, uncovering substantial performance gaps for low-resource languages.

📄 PDF Abstract BibTeX arXiv:2501.07482

Code (1)

timelyeventsbenchmark/tiebe 공식 구현

Tasks

Continual LearningLanguage ModelingLanguage Modelling

Similar Papers 제목 키워드 기반

Multi-modal Medical Image Fusion For Non-Small Cell Lung Cancer Classification

2024-09-27 · Salma Hassan, Hamad Al Hammadi, Ibrahim Mohammed, Muhammad Haris Khan

The early detection and nuanced subtype classification of non-small cell lung cancer (NSCLC), a predominant cause of cancer mortality worldwide, is a critical and complex issue. In this paper, we introduce an innovative …

Cancer Classification

Unveiling Factual Recall Behaviors of Large Language Models through Knowledge Neurons

2024-08-06 · Yifei Wang, YuHeng Chen, Wanting Wen, Yu Sheng 외

In this paper, we investigate whether Large Language Models (LLMs) actively recall or retrieve their internal repositories of factual knowledge when faced with reasoning tasks. Through an analysis of LLMs' internal factu…

WNTRAC: AI Assisted Tracking of Non-pharmaceutical Interventions Implemented Worldwide for COVID-19

2020-09-02 · Parthasarathy Suryanarayanan, Ching-Huei Tsou, Ananya Poddar, Diwakar Mahajan 외

The Coronavirus disease 2019 (COVID-19) global pandemic has transformed almost every facet of human society throughout the world. Against an emerging, highly transmissible disease with no definitive treatment or vaccine,…

Articles

Investigating Traffic Accident Detection Using Multimodal Large Language Models

2025-09-23 · Ilhan Skender, Kailin Tong, Selim Solmaz, Daniel Watzenig arxiv

Traffic safety remains a critical global concern, with timely and accurate accident detection essential for hazard reduction and rapid emergency response. Infrastructure-based vision sensors offer scalable and efficient …

Traffic Accident DetectionMulti-Object TrackingInstance SegmentationObject Detection

Significant changes in EEG neural oscillations during different phases of three-dimensional multiple object tracking task (3D-MOT) imply different roles for attention and working memory

2022-07-29 · Yannick Roy, Jocelyn Faubert

Our ability to track multiple objects in a dynamic environment enables us to perform everyday tasks such as driving, playing team sports, and walking in a crowded mall. Despite more than three decades of literature on mu…

EEGElectroencephalogram (EEG)Multiple Object TrackingObject Tracking