paper-with-me

Papers

OBSR: Open Benchmark for Spatial Representations

2025-10-07 · Julia Moska, Oleksii Furman, Kacper Kozaczko, Szymon Leszkiewicz, Jakub Polczyk, Piotr Gramacki, Piotr Szymański arxiv

GeoAI is evolving rapidly, fueled by diverse geospatial datasets like traffic patterns, environmental data, and crowdsourced OpenStreetMap (OSM) information. While sophisticated AI models are being developed, existing benchmarks are often concentrated on single tasks and restricted to a single modality. As such, progress in GeoAI is limited by the lack of a standardized, multi-task, modality-agnostic benchmark for their systematic evaluation. This paper introduces a novel benchmark designed to assess the performance, accuracy, and efficiency of geospatial embedders. Our benchmark is modality-agnostic and comprises 7 distinct datasets from diverse cities across three continents, ensuring generalizability and mitigating demographic biases. It allows for the evaluation of GeoAI embedders on various phenomena that exhibit underlying geographic processes. Furthermore, we establish a simple and intuitive task-oriented model baselines, providing a crucial reference point for comparing more complex solutions.

📄 PDF Abstract BibTeX arXiv:2510.05879

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

FloorplanQA: A Benchmark for Spatial Reasoning in LLMs using Structured Representations

2025-07-10 · Fedor Rodionov, Abdelrahman Eldesokey, Michael Birsak, John Femiani 외 arxiv

We introduce FloorplanQA, a diagnostic benchmark for evaluating spatial reasoning in large language models (LLMs). FloorplanQA is grounded in structured representations of indoor scenes, such as (e.g., kitchens, living r…

Spatial Reasoning

SpatialReasoner: Towards Explicit and Generalizable 3D Spatial Reasoning

2025-04-28 · Wufei Ma, Yu-Cheng Chou, Qihao Liu, Xingrui Wang 외

Despite recent advances on multi-modal models, 3D spatial reasoning remains a challenging task for state-of-the-art open-source and proprietary models. Recent studies explore data-driven approaches and achieve enhanced s…

Question AnsweringSpatial ReasoningVisual Question Answering

Toward Memory-Aided World Models: Benchmarking via Spatial Consistency

2025-05-29 · Kewei Lian, Shaofei Cai, Yilun Du, Yitao Liang

The ability to simulate the world in a spatially consistent manner is a crucial requirements for effective world models. Such a model enables high-quality visual generation, and also ensures the reliability of world mode…

BenchmarkingMinecraft

ProxyCLIP: Proxy Attention Improves CLIP for Open-Vocabulary Segmentation

2024-08-09 · Mengcheng Lan, Chaofeng Chen, Yiping Ke, Xinjiang Wang 외

Open-vocabulary semantic segmentation requires models to effectively integrate visual representations with open-vocabulary semantic labels. While Contrastive Language-Image Pre-training (CLIP) models shine in recognizing…

Open Vocabulary Semantic SegmentationOpen-Vocabulary Semantic SegmentationSegmentationSemantic Segmentation+1

Concerto: Joint 2D-3D Self-Supervised Learning Emerges Spatial Representations

2025-10-27 · Yujia Zhang, Xiaoyang Wu, Yixing Lao, Chengyao Wang 외 arxiv

Humans learn abstract concepts through multisensory synergy, and once formed, such representations can often be recalled from a single modality. Inspired by this principle, we introduce Concerto, a minimalist simulation …

Self-Supervised LearningScene Understanding