paper-with-me

Papers

CHOICE: Benchmarking the Remote Sensing Capabilities of Large Vision-Language Models

2024-11-27 · Xiao An, Jiaxing Sun, Zihan Gui, wei he

The rapid advancement of Large Vision-Language Models (VLMs), both general-domain models and those specifically tailored for remote sensing, has demonstrated exceptional perception and reasoning capabilities in Earth observation tasks. However, a benchmark for systematically evaluating their capabilities in this domain is still lacking. To bridge this gap, we propose CHOICE, an extensive benchmark designed to objectively evaluate the hierarchical remote sensing capabilities of VLMs. Focusing on 2 primary capability dimensions essential to remote sensing: perception and reasoning, we further categorize 6 secondary dimensions and 23 leaf tasks to ensure a well-rounded assessment coverage. CHOICE guarantees the quality of all 10,507 problems through a rigorous process of data collection from 50 globally distributed cities, question construction and quality control. The newly curated data and the format of multiple-choice questions with definitive answers allow for an objective and straightforward performance assessment. Our evaluation of 3 proprietary and 21 open-source VLMs highlights their critical limitations within this specialized context. We hope that CHOICE will serve as a valuable resource and offer deeper insights into the challenges and potential of VLMs in the field of remote sensing. We will release CHOICE at https://github.com/ShawnAn-WHU/CHOICE.

📄 PDF Abstract BibTeX arXiv:2411.18145

Code (1)

shawnan-whu/choice 공식 구현 pytorch

Tasks

BenchmarkingEarth ObservationMultiple-choice

Methods 이 논문이 사용한 방법론

+ ( 1 ) ⟷ 888 ⟷ ( 829 ) ⟷ 0881||How do I resolve a dispute on Expedia? How do I resolve a dispute on Expedia contact their support at + ( 1 ) ⟷ 888 ⟷ ( 829 ) ⟷ 0881 or + ( 1 ) ⟷ 805 ⟷ ( 330 ) ⟷ 4056. Provide booking details and explain the issue…

Similar Papers 제목 키워드 기반

Think and Answer ME: Benchmarking and Exploring Multi-Entity Reasoning Grounding in Remote Sensing

2026-03-13 · Shuchang Lyu, Haiquan Wen, Guangliang Cheng, Meng Li 외 arxiv

Recent advances in reasoning language models and reinforcement learning with verifiable rewards have significantly enhanced multi-step reasoning capabilities. This progress motivates the extension of reasoning paradigms …

Reinforcement LearningVisual Grounding

Towards Efficient Benchmarking of Foundation Models in Remote Sensing: A Capabilities Encoding Approach

2025-05-06 · Pierre Adorni, Minh-Tan Pham, Stéphane May, Sébastien Lefèvre

Foundation models constitute a significant advancement in computer vision: after a single, albeit costly, training phase, they can address a wide array of tasks. In the field of Earth observation, over 75 remote sensing …

BenchmarkingEarth Observation

RS-Mamba for Large Remote Sensing Image Dense Prediction

2024-04-03 · Sijie Zhao, Hao Chen, Xueliang Zhang, Pengfeng Xiao 외

Context modeling is critical for remote sensing image dense prediction tasks. Nowadays, the growing size of very-high-resolution (VHR) remote sensing images poses challenges in effectively modeling context. While transfo…

Building change detection for remote sensing imagesChange DetectionMambaPrediction+2

Remote Sensing Image Classification with the SEN12MS Dataset

2021-04-01 · Michael Schmitt, Yu-Lun Wu

Image classification is one of the main drivers of the rapid developments in deep learning with convolutional neural networks for computer vision. So is the analogous task of scene classification in remote sensing. Howev…

BenchmarkingClassificationGeneral Classificationimage-classification+3

LithoBench: Benchmarking Large Multimodal Models for Remote-Sensing Lithology Interpretation

2026-05-08 · Jun Wang, Fengpeng Li, Hang Dong, Tianjin Huang 외 arxiv

Remote sensing lithology interpretation is fundamental to geological surveys, mineral exploration, and regional geological mapping. Unlike general land-cover recognition, lithology interpretation is a knowledge-intensive…