paper-with-me

홈 › Papers

Inherent limitations of LLMs regarding spatial information

2023-12-05 · He Yan, Xinyao Hu, Xiangpeng Wan, Chengyu Huang, Kai Zou, Shiqi Xu

Despite the significant advancements in natural language processing capabilities demonstrated by large language models such as ChatGPT, their proficiency in comprehending and processing spatial information, especially within the domains of 2D and 3D route planning, remains notably underdeveloped. This paper investigates the inherent limitations of ChatGPT and similar models in spatial reasoning and navigation-related tasks, an area critical for applications ranging from autonomous vehicle guidance to assistive technologies for the visually impaired. In this paper, we introduce a novel evaluation framework complemented by a baseline dataset, meticulously crafted for this study. This dataset is structured around three key tasks: plotting spatial points, planning routes in two-dimensional (2D) spaces, and devising pathways in three-dimensional (3D) environments. We specifically developed this dataset to assess the spatial reasoning abilities of ChatGPT. Our evaluation reveals key insights into the model's capabilities and limitations in spatial understanding.

📄 PDF Abstract BibTeX arXiv:2312.03042

Code (1)

protagolabs/SpatialEval 공식 구현

Tasks

Spatial Reasoning

Similar Papers 제목 키워드 기반

Spatial 3D-LLM: Exploring Spatial Awareness in 3D Vision-Language Models

2025-07-22 · Xiaoyan Wang, Zeju Li, Yifan Xu, Jiaxing Qi 외 arxiv

New era has unlocked exciting possibilities for extending Large Language Models (LLMs) to tackle 3D vision-language tasks. However, most existing 3D multimodal LLMs (MLLMs) rely on compressing holistic 3D scene informati…

ECG-aBcDe: Overcoming Model Dependence, Encoding ECG into a Universal Language for Any LLM

2025-09-16 · Yong Xia, Jingxuan Li, YeTeng Sun, Jiarui Bu arxiv

Large Language Models (LLMs) hold significant promise for electrocardiogram (ECG) analysis, yet challenges remain regarding transferability, time-scale information learning, and interpretability. Current methods suffer f…

Temporal Blind Spots in Large Language Models

2024-01-22 · Jonas Wallat, Adam Jatowt, Avishek Anand

Large language models (LLMs) have recently gained significant attention due to their unparalleled ability to perform various natural language processing tasks. These models, benefiting from their advanced natural languag…

Natural Language Understanding

Geode: A Zero-shot Geospatial Question-Answering Agent with Explicit Reasoning and Precise Spatio-Temporal Retrieval

2024-06-26 · Devashish Vikas Gupta, Azeez Syed Ali Ishaqui, Divya Kiran Kadiyala

Large language models (LLMs) have shown promising results in learning and contextualizing information from different forms of data. Recent advancements in foundational models, particularly those employing self-attention …

Question Answering

Semantic Density: Uncertainty Quantification for Large Language Models through Confidence Measurement in Semantic Space

2024-05-22 · Xin Qiu, Risto Miikkulainen

With the widespread application of Large Language Models (LLMs) to various domains, concerns regarding the trustworthiness of LLMs in safety-critical scenarios have been raised, due to their unpredictable tendency to hal…

MisinformationQuestion AnsweringUncertainty Quantification