paper-with-me

Papers

On Utilizing Relationships for Transferable Few-Shot Fine-Grained Object Detection

2022-12-01 · Ambar Pal, Arnau Ramisa, Amit Kumar K C, René Vidal

State-of-the-art object detectors are fast and accurate, but they require a large amount of well annotated training data to obtain good performance. However, obtaining a large amount of training annotations specific to a particular task, i.e., fine-grained annotations, is costly in practice. In contrast, obtaining common-sense relationships from text, e.g., "a table-lamp is a lamp that sits on top of a table", is much easier. Additionally, common-sense relationships like "on-top-of" are easy to annotate in a task-agnostic fashion. In this paper, we propose a probabilistic model that uses such relational knowledge to transform an off-the-shelf detector of coarse object categories (e.g., "table", "lamp") into a detector of fine-grained categories (e.g., "table-lamp"). We demonstrate that our method, RelDetect, achieves performance competitive to finetuning based state-of-the-art object detector baselines when an extremely low amount of fine-grained annotations is available ($0.2\%$ of entire dataset). We also demonstrate that RelDetect is able to utilize the inherent transferability of relationship information to obtain a better performance ($+5$ mAP points) than the above baselines on an unseen dataset (zero-shot transfer). In summary, we demonstrate the power of using relationships for object detection on datasets where fine-grained object categories can be linked to coarse-grained categories via suitable relationships.

📄 PDF Abstract BibTeX arXiv:2212.00770

Code (0)

등록된 구현이 없습니다.

Tasks

Common Sense ReasoningObjectobject-detectionObject Detection

Similar Papers 제목 키워드 기반

NDPNet: A novel non-linear data projection network for few-shot fine-grained image classification

2021-06-13 · Weichuan Zhang, Xuefang Liu, Zhe Xue, Yongsheng Gao 외

Metric-based few-shot fine-grained image classification (FSFGIC) aims to learn a transferable feature embedding network by estimating the similarities between query images and support classes from very few examples. In t…

Few-Shot LearningFine-Grained Image Classificationimage-classificationImage Classification+1

Disentangled Ontology Embedding for Zero-shot Learning

2022-06-08 · Yuxia Geng, Jiaoyan Chen, Wen Zhang, Yajing Xu 외

Knowledge Graph (KG) and its variant of ontology have been widely used for knowledge representation, and have shown to be quite effective in augmenting Zero-shot Learning (ZSL). However, existing ZSL methods that utilize…

image-classificationImage ClassificationOntology EmbeddingZero-Shot Image Classification+1

AUCH-Net: Action Unit-Based Consistency-Aware Hypergraph Network for Cross-Domain Few-Shot Facial Expression Recognition

2026-07-23 · Xinhan Qiu, Yan Yan, Rui Zhu, Si Chen 외 arxiv

Recently, cross-domain few-shot facial expression recognition (CF-FER) has received considerable attention. However, the performance of existing CF-FER methods is still unsatisfactory due to inferior transferable feature…

Facial Expression RecognitionCross-Domain Few-Shot

Synthesize Diagnose and Optimize: Towards Fine-Grained Vision-Language Understanding

2024-01-01 · CVPR 2024 1 · Wujian Peng, Sicheng Xie, Zuyao You, Shiyi Lan 외

Vision language models (VLM) have demonstrated remarkable performance across various downstream tasks. However understanding fine-grained visual-linguistic concepts such as attributes and inter-object relationships r…

Attribute

Synthesize, Diagnose, and Optimize: Towards Fine-Grained Vision-Language Understanding

2023-11-30 · Wujian Peng, Sicheng Xie, Zuyao You, Shiyi Lan 외

Vision language models (VLM) have demonstrated remarkable performance across various downstream tasks. However, understanding fine-grained visual-linguistic concepts, such as attributes and inter-object relationships, re…

AttributeCompositional Zero-Shot LearningImage RetrievalImage-text matching+1