paper-with-me

홈 › Papers

Stuck in the Matrix: Probing Spatial Reasoning in Large Language Models

2025-10-23 · Maggie Bai, Ava Kim Cohen, Eleanor Koss, Charlie Lichtenbaum arxiv

This paper explores the spatial reasoning capability of large language models (LLMs) over textual input through a suite of five tasks aimed at probing their spatial understanding and computational abilities. The models were tested on both fundamental spatial reasoning and multi-step problem-solving within structured grid-based environments using tasks such as quadrant identification, geometric transformations, distance evaluation, word searches, and tile sliding. Each task was scaled in complexity through increasing grid dimensions, requiring models to extend beyond simple pattern recognition into abstract spatial reasoning. Our results reveal that while LLMs demonstrate moderate success in all tasks with small complexity and size, performance drops off rapidly as scale increases, with an average loss in accuracy of 42.7%, and reaching as high as 84%. Every test that began with over 50% accuracy showed a loss of at least 48%, illustrating the consistent nature of the deterioration. Furthermore, their struggles with scaling complexity hint at a lack of robust spatial representations in their underlying architectures. This paper underscores the gap between linguistic and spatial reasoning in LLMs, offering insights into their current limitations, and laying the groundwork for future integrative benchmarks at the intersection of language and geometry.

📄 PDF Abstract BibTeX arXiv:2510.20198

Code (0)

등록된 구현이 없습니다.

Tasks

Spatial Reasoning

Similar Papers 제목 키워드 기반

SpatialAct: Probing Spatial Reasoning-to-Action Capabilities of VLM Agents in 3D Scenes

2026-05-29 · Tianhui Liu, Jie Feng, Zhiheng Zheng, Shengyuan Wang 외 arxiv

Humans can effortlessly perceive spatial layouts, form cognitive representations, reason about spatial relations, and translate such reasoning into actions in everyday 3D environments. Although recent vision-language mod…

Spatial Reasoning

From Human Cognition to Neural Activations: Probing the Computational Primitives of Spatial Reasoning in LLMs

2026-03-27 · Jiyuan An, Liner Yang, Mengyan Wang, Luming Lu 외 arxiv

As spatial intelligence becomes an increasingly important capability for foundation models, it remains unclear whether large language models' (LLMs) performance on spatial reasoning benchmarks reflects structured interna…

Spatial Reasoning

Stuck-at Faults in ReRAM Neuromorphic Circuit Array and their Correction through Machine Learning

2024-02-15 · Vedant Sawal, Hiu Yung Wong

In this paper, we study the inference accuracy of the Resistive Random Access Memory (ReRAM) neuromorphic circuit due to stuck-at faults (stuck-on, stuck-off, and stuck at a certain resistive value). A simulation framewo…

Large Language Model-assisted Autonomous Vehicle Recovery from Immobilization

2025-10-29 · Zhipeng Bao, Qianwen Li arxiv

Despite significant advancements in recent decades, autonomous vehicles (AVs) continue to face challenges in navigating certain traffic scenarios where human drivers excel. In such situations, AVs often become immobilize…

Autonomous Vehicles

Are Large Language Models Geospatially Knowledgeable?

2023-10-09 · Prabin Bhandari, Antonios Anastasopoulos, Dieter Pfoser

Despite the impressive performance of Large Language Models (LLM) for various natural language processing tasks, little is known about their comprehension of geographic data and related ability to facilitate informed geo…

Decision Making