paper-with-me

Fine-Grained Image Recognition

4개 벤치마크 · 논문 79편 · 이 태스크의 논문 보기 →

Benchmarks

OVEN

결과 5개

CNFOOD-241-Chen

결과 2개

CUB Birds

결과 2개

CUB-200-2011

결과 2개

Most implemented

Papers

Structured-Condensed Prompt Tuning in Vision-Language Models for Fine-grained Image Recognition

2026-07-07 · Xinda Liu, Qinyu Zhang, Weiqing Min, Guohua Geng 외 arxiv

Fine-grained image recognition poses a significant challenge due to the substantial expertise and effort required for manual annotation. Vision-language models (VLMs) like CLIP provide a compelling zero-shot alternative,…

Fine-Grained Image Recognition

A Large-Scale Study on the Accuracy vs Cost Trade-offs of Training and Evaluation Settings in Fine-Grained Image Recognition

2026-05-18 · Edwin Arkel Rios, Augusto Christian Surya, Oswin Gosal, Fernando Mikael 외 arxiv

Prior work on fine-grained image recognition (FGIR) has established the importance of the backbone selection, but has neglected the accuracy-vs-cost trade-offs under different training and evaluation settings. In this wo…

Fine-Grained Image RecognitionData Augmentation

How to Choose Your Teacher for Fine Grained Image Recognition

2026-05-15 · Oswin Gosal, Edwin Arkel Rios, Augusto Christian Surya, Fernando Mikael 외 arxiv

Fine-grained image recognition classifies subcategories such as bird species or car models. While state-of-the-art (SOTA) models are accurate, they are often too resource-intensive for deployment on constrained devices. …

Fine-Grained Image RecognitionKnowledge Distillation

Rényi Attention Entropy for Patch Pruning

2026-04-04 · Hiroaki Aizawa, Yuki Igaue arxiv

Transformers are strong baselines in both vision and language because self-attention captures long-range dependencies across tokens. However, the cost of self-attention grows quadratically with the number of tokens. Patc…

Fine-Grained Image Recognition

Thinking Beyond Labels: Vocabulary-Free Fine-Grained Recognition using Reasoning-Augmented LMMs

2025-12-21 · Dmitry Demidov, Zaigham Zaheer, Zongyan Han, Omkar Thawakar 외 arxiv

Vocabulary-free fine-grained image recognition aims to distinguish visually similar categories within a meta-class without a fixed, human-defined label set. Existing solutions for this problem are limited by either the u…

Fine-Grained Visual RecognitionFine-Grained Image Recognition

DiVE-k: Differential Visual Reasoning for Fine-grained Image Recognition

2025-11-23 · Raja Kumar, Arka Sadhu, Ram Nevatia arxiv

Large Vision Language Models (LVLMs) possess extensive text knowledge but struggles to utilize this knowledge for fine-grained image recognition, often failing to differentiate between visually similar categories. Existi…

Fine-Grained Image RecognitionReinforcement LearningVisual Reasoning

전체 79편 보기 →