paper-with-me

홈 › Papers

NoReGeo: Non-Reasoning Geometry Benchmark

2026-01-15 · Irina Abdullaeva, Anton Vasiliuk, Elizaveta Goncharova, Temurbek Rahmatullaev, Zagorulko Ivan, Maxim Kurkin, Andrey Kuznetsov arxiv

We present NoReGeo, a novel benchmark designed to evaluate the intrinsic geometric understanding of large language models (LLMs) without relying on reasoning or algebraic computation. Unlike existing benchmarks that primarily assess models' proficiency in reasoning-based geometry-where solutions are derived using algebraic methods-NoReGeo focuses on evaluating whether LLMs can inherently encode spatial relationships and recognize geometric properties directly. Our benchmark comprises 2,500 trivial geometric problems spanning 25 categories, each carefully crafted to be solvable purely through native geometric understanding, assuming known object locations. We assess a range of state-of-the-art models on NoReGeo, including frontier models like GPT-4, observing that even the most advanced systems achieve an overall maximum of 65% accuracy in binary classification tasks. Further, our ablation experiments demonstrate that such geometric understanding does not emerge through fine-tuning alone, indicating that effective training for geometric comprehension requires a specialized approach from the outset. Our findings highlight a significant gap in current LLMs' ability to natively grasp geometric concepts, providing a foundation for future research toward models with true geometric cognition.

📄 PDF Abstract BibTeX arXiv:2601.10254

Code (0)

등록된 구현이 없습니다.

Tasks

Binary Classification

Similar Papers 제목 키워드 기반

GeomVerse: A Systematic Evaluation of Large Models for Geometric Reasoning

2023-12-19 · Mehran Kazemi, Hamidreza Alvari, Ankit Anand, Jialin Wu 외

Large language models have shown impressive results for multi-hop mathematical reasoning when the input question is only textual. Many mathematical reasoning problems, however, contain both text and image. With the ever-…

Mathematical Reasoning

Make Geometry Matter for Spatial Reasoning

2026-03-27 · Shihua Zhang, Qiuhong Shen, Shizun Wang, Tianbo Pan 외 arxiv

Empowered by large-scale training, vision-language models (VLMs) achieve strong image and video understanding, yet their ability to perform spatial reasoning in both static scenes and dynamic videos remains limited. Rece…

Spatial Reasoning

Learning to Solve Geometry Problems via Simulating Human Dual-Reasoning Process

2024-05-10 · Tong Xiao, Jiayu Liu, Zhenya Huang, Jinze Wu 외

Geometry Problem Solving (GPS), which is a classic and challenging math problem, has attracted much attention in recent years. It requires a solver to comprehensively understand both text and diagram, master essential ge…

Geometry Problem SolvingMachine TranslationMath

DynaSolidGeo: A Dynamic Benchmark for Genuine Spatial Mathematical Reasoning of VLMs in Solid Geometry

2025-10-25 · Changti Wu, Shijie Lian, Zihao Liu, Lei Zhang 외 arxiv

Solid geometry problem solving demands spatial mathematical reasoning that integrates spatial intelligence and symbolic reasoning. However, most existing multimodal mathematical reasoning benchmarks focus primarily on 2D…

Mathematical ReasoningSpatial Reasoning

GeoWeaver: Grounding Visual Tokens with Geometric Evidence before Scene Reasoning

2026-05-21 · Deshui Miao, Xingsen Huang, Yameng Gu, Xin Li 외 arxiv

Spatio-temporal reasoning in vision-language models requires visual representations that preserve physical geometry rather than merely semantic appearance. Recent multimodal models incorporate geometric information throu…

Spatial Reasoning