paper-with-me

홈 › Papers

Distortions in Judged Spatial Relations in Large Language Models

2024-01-08 · Nir Fulman, Abdulkadir Memduhoğlu, Alexander Zipf

We present a benchmark for assessing the capability of Large Language Models (LLMs) to discern intercardinal directions between geographic locations and apply it to three prominent LLMs: GPT-3.5, GPT-4, and Llama-2. This benchmark specifically evaluates whether LLMs exhibit a hierarchical spatial bias similar to humans, where judgments about individual locations' spatial relationships are influenced by the perceived relationships of the larger groups that contain them. To investigate this, we formulated 14 questions focusing on well-known American cities. Seven questions were designed to challenge the LLMs with scenarios potentially influenced by the orientation of larger geographical units, such as states or countries, while the remaining seven targeted locations were less susceptible to such hierarchical categorization. Among the tested models, GPT-4 exhibited superior performance with 55 percent accuracy, followed by GPT-3.5 at 47 percent, and Llama-2 at 45 percent. The models showed significantly reduced accuracy on tasks with suspected hierarchical bias. For example, GPT-4's accuracy dropped to 33 percent on these tasks, compared to 86 percent on others. However, the models identified the nearest cardinal direction in most cases, reflecting their associative learning mechanism, thereby embodying human-like misconceptions. We discuss avenues for improving the spatial reasoning capabilities of LLMs.

📄 PDF Abstract BibTeX arXiv:2401.04218

Code (0)

등록된 구현이 없습니다.

Tasks

MisconceptionsSpatial Reasoning

Methods 이 논문이 사용한 방법론

American 설명 없음
{Dispute@FaQ-s}How to file a dispute with Expedia? How to file a dispute with Expedia? To file a complaint against Expedia, first try contacting their customer service directly. You can reach them by phone at…
Multi-Head Attention 설명 없음
15 Ways to Contact How can i speak to someone at Delta Airlines 설명 없음
Attention 설명 없음
Absolute Position Encodings Absolute Position Encodings are a type of position embeddings for [Transformer-based models] where positional encodings are…
Position-Wise Feed-Forward Layer 설명 없음
Label Smoothing Label Smoothing is a regularization technique that introduces noise for the labels. This accounts for the fact that datasets may have mistakes in them, so maximizing the…

Similar Papers 제목 키워드 기반

Distilling Spatially-Heterogeneous Distortion Perception for Blind Image Quality Assessment

2025-01-01 · CVPR 2025 1 · Xudong Li, Wenjie Nie, Yan Zhang, Runze Hu 외

In the Blind Image Quality Assessment (BIQA) field, accurately assessing the quality of authentically distorted images presents a substantial challenge due to the diverse distortion types in natural settings. Existin…

Blind Image Quality AssessmentImage Quality AssessmentKnowledge DistillationLocal Distortion

Disentangling Image Distortions in Deep Feature Space

2020-02-26 · Simone Bianco, Luigi Celona, Paolo Napoletano

Previous literature suggests that perceptual similarity is an emergent property shared across deep visual representations. Experiments conducted on a dataset of human-judged image distortions have proven that deep featur…

Image Quality Assessment

Hierarchical Graph Attention Network for No-Reference Omnidirectional Image Quality Assessment

2025-08-13 · Hao Yang, Xu Zhang, Jiaqi Ma, Linwei Zhu 외 arxiv

Current Omnidirectional Image Quality Assessment (OIQA) methods struggle to evaluate locally non-uniform distortions due to inadequate modeling of spatial variations in quality and ineffective feature representation capt…

Image Quality AssessmentGraph Neural NetworkLocal Distortion

Leveraging Spatial Uncertainty for Online Error Compensation in EMT

2020-04-28

Purpose: Electromagnetic Tracking (EMT) can potentially complement fluoroscopic navigation, reducing radiation exposure in a hybrid setting. Due to the susceptibility to external distortions, systematic error in EMT need…

Evaluation of Geographical Distortions in Language Models: A Crucial Step Towards Equitable Representations

2024-04-26 · Rémy Decoupes, Roberto Interdonato, Mathieu Roche, Maguelonne Teisseire 외

Language models now constitute essential tools for improving efficiency for many professional tasks such as writing, coding, or learning. For this reason, it is imperative to identify inherent biases. In the field of Nat…