paper-with-me

홈 › Papers

AgriBench: A Hierarchical Agriculture Benchmark for Multimodal Large Language Models

2024-11-30 · Yutong Zhou, Masahiro Ryo

We introduce AgriBench, the first agriculture benchmark designed to evaluate MultiModal Large Language Models (MM-LLMs) for agriculture applications. To further address the agriculture knowledge-based dataset limitation problem, we propose MM-LUCAS, a multimodal agriculture dataset, that includes 1,784 landscape images, segmentation masks, depth maps, and detailed annotations (geographical location, country, date, land cover and land use taxonomic details, quality scores, aesthetic scores, etc), based on the Land Use/Cover Area Frame Survey (LUCAS) dataset, which contains comparable statistics on land use and land cover for the European Union (EU) territory. This work presents a groundbreaking perspective in advancing agriculture MM-LLMs and is still in progress, offering valuable insights for future developments and innovations in specific expert knowledge-based MM-LLMs.

📄 PDF Abstract BibTeX arXiv:2412.00465

Code (1)

yutong-zhou-cv/agribench 공식 구현

Similar Papers 제목 키워드 기반

AgriGPT-VL: Agricultural Vision-Language Understanding Suite

2025-10-05 · Bo Yang, Yunkui Chen, Lanfei Feng, Yu Zhang 외 arxiv

Despite rapid advances in multimodal large language models, agricultural applications remain constrained by the scarcity of domain-tailored models, curated vision-language corpora, and rigorous evaluation. To address the…

Reinforcement LearningMultimodal Reasoning

AgriGPT-Omni: A Unified Speech-Vision-Text Framework for Multilingual Agricultural Intelligence

2025-12-11 · Bo Yang, Lanfei Feng, Yunkui Chen, Yu Zhang 외 arxiv

Despite rapid advances in multimodal large language models, agricultural applications remain constrained by the lack of multilingual speech data, unified multimodal architectures, and comprehensive evaluation benchmarks.…

Reinforcement LearningMultimodal Reasoning

AgriGPT: a Large Language Model Ecosystem for Agriculture

2025-08-12 · Bo Yang, Yu Zhang, Lanfei Feng, Yunkui Chen 외 arxiv

Despite the rapid progress of Large Language Models (LLMs), their application in agriculture remains limited due to the lack of domain-specific models, curated datasets, and robust evaluation frameworks. To address these…

Domain Adaptation

AgroTools: A Benchmark for Tool-Augmented Multimodal Agents in Agriculture

2026-05-21 · Zi Ye, Yibin Wen, Xiaoya Fan, Xinyu Zhang 외 arxiv

Agricultural decision-making increasingly requires multimodal systems that can transform visual observations into reliable, executable actions. However, existing agricultural multimodal benchmarks mainly evaluate final-a…

Multi-label Instance-level Generalised Visual Grounding in Agriculture

2026-03-05 · Mohammadreza Haghighat, Alzayat Saleh, Mostafa Rahimi Azghadi arxiv

Understanding field imagery such as detecting plants and distinguishing individual crop and weed instances is a central challenge in precision agriculture. Despite progress in vision-language tasks like captioning and vi…

Visual Question AnsweringVisual Grounding