paper-with-me

Papers

Automatic Fused Multimodal Deep Learning for Plant Identification

2024-06-03 · Alfreds Lapkovskis, Natalia Nefedova, Ali Beikmohammadi

Plant classification is vital for ecological conservation and agricultural productivity, enhancing our understanding of plant growth dynamics and aiding species preservation. The advent of deep learning (DL) techniques has revolutionized this field by enabling autonomous feature extraction, significantly reducing the dependence on manual expertise. However, conventional DL models often rely solely on single data sources, failing to capture the full biological diversity of plant species comprehensively. Recent research has turned to multimodal learning to overcome this limitation by integrating multiple data types, which enriches the representation of plant characteristics. This shift introduces the challenge of determining the optimal point for modality fusion. In this paper, we introduce a pioneering multimodal DL-based approach for plant classification with automatic modality fusion. Utilizing the multimodal fusion architecture search, our method integrates images from multiple plant organs -- flowers, leaves, fruits, and stems -- into a cohesive model. To address the lack of multimodal datasets, we contributed Multimodal-PlantCLEF, a restructured version of the PlantCLEF2015 dataset tailored for multimodal tasks. Our method achieves 82.61% accuracy on 979 classes of Multimodal-PlantCLEF, surpassing state-of-the-art methods and outperforming late fusion by 10.33%. Through the incorporation of multimodal dropout, our approach demonstrates strong robustness to missing modalities. We validate our model against established benchmarks using standard performance metrics and McNemar's test, further underscoring its superiority.

📄 PDF Abstract BibTeX arXiv:2406.01455

Code (1)

alfredslapkovskis/multimodalplantclassifier 공식 구현 tf

Tasks

Deep LearningDiversityMultimodal Deep Learning

Similar Papers 제목 키워드 기반

UniBO at SemEval-2022 Task 5: A Multimodal bi-Transformer Approach to the Binary and Fine-grained Identification of Misogyny in Memes

2022-07-01 · SemEval (NAACL) 2022 7 · Arianna Muti, Katerina Korre, Alberto Barrón-Cedeño

We present our submission to SemEval 2022 Task 5 on Multimedia Automatic Misogyny Identification. We address the two tasks: Task A consists of identifying whether a meme is misogynous. If so, Task B attempts to identify …

Position

Detecting total hip replacement prosthesis design on preoperative radiographs using deep convolutional neural network

2019-11-27 · Alireza Borjali, Antonia F. Chen, Orhun K. Muratoglu, Mohammad A. Morid 외

Identifying the design of a failed implant is a key step in preoperative planning of revision total joint arthroplasty. Manual identification of the implant design from radiographic images is time consuming and prone to …

SRCB at SemEval-2022 Task 5: Pretraining Based Image to Text Late Sequential Fusion System for Multimodal Misogynous Meme Identification

2022-07-01 · SemEval (NAACL) 2022 7 · Jing Zhang, Yujin Wang

Online misogyny meme detection is an image/text multimodal classification task, the complicated relation of image and text challenges the intelligent system’s modality fusion learning capability. In this paper, we invest…

Image to text

Digital Taxonomist: Identifying Plant Species in Community Scientists' Photographs

2021-06-07 · Riccardo de Lutio, Yihang She, Stefano D'Aronco, Stefania Russo 외

Automatic identification of plant specimens from amateur photographs could improve species range maps, thus supporting ecosystems research as well as conservation efforts. However, classifying plant specimens based on im…

Multimodal Deep Learning

Overview of PlantCLEF 2022: Image-based plant identification at global scale

2025-09-22 · Herve Goeau, Pierre Bonnet, Alexis Joly arxiv

It is estimated that there are more than 300,000 species of vascular plants in the world. Increasing our knowledge of these species is of paramount importance for the development of human civilization (agriculture, const…