paper-with-me

Papers

EnGraf-Net: Multiple Granularity Branch Network with Fine-Coarse Graft Grained for Classification Task

2025-09-25 · Riccardo La Grassa, Ignazio Gallo, Nicola Landro arxiv

Fine-grained classification models are designed to focus on the relevant details necessary to distinguish highly similar classes, particularly when intra-class variance is high and inter-class variance is low. Most existing models rely on part annotations such as bounding boxes, part locations, or textual attributes to enhance classification performance, while others employ sophisticated techniques to automatically extract attention maps. We posit that part-based approaches, including automatic cropping methods, suffer from an incomplete representation of local features, which are fundamental for distinguishing similar objects. While fine-grained classification aims to recognize the leaves of a hierarchical structure, humans recognize objects by also forming semantic associations. In this paper, we leverage semantic associations structured as a hierarchy (taxonomy) as supervised signals within an end-to-end deep neural network model, termed EnGraf-Net. Extensive experiments on three well-known datasets CIFAR-100, CUB-200-2011, and FGVC-Aircraft demonstrate the superiority of EnGraf-Net over many existing fine-grained models, showing competitive performance with the most recent state-of-the-art approaches, without requiring cropping techniques or manual annotations.

📄 PDF Abstract BibTeX arXiv:2509.21061

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

EnGraf-Net: Multiple Granularity Branch Network with Fine-Coarse Graft Grained for Classification Task

2021-10-21 · CAIP: Computer Analysis of Images and Patterns 2021 10 · Riccardo La Grassa, Ignazio Gallo, Nicola Landro

Fine-Grained classification models can expressly focus on the relevant details useful to distinguish highly similar classes typically when the intra-class variance is high and the inter-class variance is low given a data…

ClassificationFine-Grained Image ClassificationImage Classification

Dense Audio-Visual Event Localization under Cross-Modal Consistency and Multi-Temporal Granularity Collaboration

2024-12-17 · Ziheng Zhou, Jinxing Zhou, Wei Qian, Shengeng Tang 외

In the field of audio-visual learning, most research tasks focus exclusively on short videos. This paper focuses on the more practical Dense Audio-Visual Event Localization (DAVEL) task, advancing audio-visual scene unde…

audio-visual event localizationaudio-visual learningScene Understanding

MALT: Multi-scale Action Learning Transformer for Online Action Detection

2024-05-31 · Zhipeng Yang, Ruoyu Wang, Yang Tan, Liping Xie

Online action detection (OAD) aims to identify ongoing actions from streaming video in real-time, without access to future frames. Since these actions manifest at varying scales of granularity, ranging from coarse to fin…

Action DetectionDecoderOnline Action Detection

Assessing Engraftment Following Fecal Microbiota Transplant

2024-04-10 · Chloe Herman, Bridget M. Barker, Thais F. Bartelli, Vidhi Chandra 외

Fecal Microbiota Transplant (FMT) is an FDA approved treatment for recurrent Clostridium difficile infections, and is being explored for other clinical applications, from alleviating digestive and neurological disorders,…

Assessing microbiome engraftment extent following fecal microbiota transplant with q2-fmt

2024-11-26 · Chloe Herman, Evan Bolyen, Anthony Simard, Liz Gehret 외

We present q2-fmt, a QIIME 2 plugin that provides diverse methods for assessing the extent of microbiome engraftment following fecal microbiota transplant. The methods implemented here were informed by a recent literatur…