paper-with-me

홈 › Papers

Tomato Multi-Angle Multi-Pose Dataset for Fine-Grained Phenotyping

2025-07-15 · Yujie Zhang, Sabine Struckmeyer, Andreas Kolb, Sven Reichardt

Observer bias and inconsistencies in traditional plant phenotyping methods limit the accuracy and reproducibility of fine-grained plant analysis. To overcome these challenges, we developed TomatoMAP, a comprehensive dataset for Solanum lycopersicum using an Internet of Things (IoT) based imaging system with standardized data acquisition protocols. Our dataset contains 64,464 RGB images that capture 12 different plant poses from four camera elevation angles. Each image includes manually annotated bounding boxes for seven regions of interest (ROIs), including leaves, panicle, batch of flowers, batch of fruits, axillary shoot, shoot and whole plant area, along with 50 fine-grained growth stage classifications based on the BBCH scale. Additionally, we provide 3,616 high-resolution image subset with pixel-wise semantic and instance segmentation annotations for fine-grained phenotyping. We validated our dataset using a cascading model deep learning framework combining MobileNetv3 for classification, YOLOv11 for object detection, and MaskRCNN for segmentation. Through AI vs. Human analysis involving five domain experts, we demonstrate that the models trained on our dataset achieve accuracy and speed comparable to the experts. Cohen's Kappa and inter-rater agreement heatmap confirm the reliability of automated fine-grained phenotyping using our approach.

📄 PDF Abstract BibTeX arXiv:2507.11279

Code (0)

등록된 구현이 없습니다.

Tasks

Instance Segmentationobject-detectionObject DetectionPlant PhenotypingSemantic Segmentation

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…
Heatmap 설명 없음

Similar Papers 제목 키워드 기반

Rethinking Reasoning in LLMs: Neuro-Symbolic Local RetoMaton Beyond ICL and CoT

2025-08-22 · Rushitha Santhoshi Mamidala, Anshuman Chhabra, Ankur Mali arxiv

Prompt-based reasoning strategies such as Chain-of-Thought (CoT) and In-Context Learning (ICL) have become widely used for eliciting reasoning capabilities in large language models (LLMs). However, these methods rely on …

Reading Comprehension

Evaluating the Single-Shot MultiBox Detector and YOLO Deep Learning Models for the Detection of Tomatoes in a Greenhouse

2021-09-02 · Sandro A. Magalhães, Luís Castro, Germano Moreira, Filipe N. Santos 외

The development of robotic solutions for agriculture requires advanced perception capabilities that can work reliably in any crop stage. For example, to automatise the tomato harvesting process in greenhouses, the visual…

MATA: A Trainable Hierarchical Automaton System for Multi-Agent Visual Reasoning

2026-01-27 · Zhixi Cai, Fucai Ke, Kevin Leo, Sukai Huang 외 arxiv

Recent vision-language models have strong perceptual ability but their implicit reasoning is hard to explain and easily generates hallucinations on complex queries. Compositional methods improve interpretability, but mos…

Visual Reasoning

ToMAToMP: Robust and Multi-Parameter Topological Clustering

2026-05-14 · Ludo Andrianirina, Mathieu Carrière arxiv

Topological clustering, and its main algorithm ToMATo, is a clustering method from Topological Data Analysis (TDA) which has been applied successfully in several applications during the last few years. This is due to its…

Tomato Maturity Recognition with Convolutional Transformers

2023-07-04 · Asim Khan, Taimur Hassan, Muhammad Shafay, Israa Fahmy 외

Tomatoes are a major crop worldwide, and accurately classifying their maturity is important for many agricultural applications, such as harvesting, grading, and quality control. In this paper, the authors propose a novel…