paper-with-me

홈 › Papers

Tree Structure-Aware Few-Shot Image Classification via Hierarchical Aggregation

2022-07-14 · Min Zhang, Siteng Huang, Wenbin Li, Donglin Wang

In this paper, we mainly focus on the problem of how to learn additional feature representations for few-shot image classification through pretext tasks (e.g., rotation or color permutation and so on). This additional knowledge generated by pretext tasks can further improve the performance of few-shot learning (FSL) as it differs from human-annotated supervision (i.e., class labels of FSL tasks). To solve this problem, we present a plug-in Hierarchical Tree Structure-aware (HTS) method, which not only learns the relationship of FSL and pretext tasks, but more importantly, can adaptively select and aggregate feature representations generated by pretext tasks to maximize the performance of FSL tasks. A hierarchical tree constructing component and a gated selection aggregating component is introduced to construct the tree structure and find richer transferable knowledge that can rapidly adapt to novel classes with a few labeled images. Extensive experiments show that our HTS can significantly enhance multiple few-shot methods to achieve new state-of-the-art performance on four benchmark datasets. The code is available at: https://github.com/remiMZ/HTS-ECCV22.

📄 PDF Abstract BibTeX arXiv:2207.06989

Code (1)

remiMZ/HTS-ECCV22 공식 구현 pytorch

Tasks

Few-Shot Image ClassificationFew-Shot Learningimage-classificationImage Classification

Similar Papers 제목 키워드 기반

Decomposing Visual Classification: Assessing Tree-Based Reasoning in VLMs

2025-09-10 · Sary Elmansoury, Islam Mesabah, Gerrit Großmann, Peter Neigel 외 arxiv

Vision language models (VLMs) excel at zero-shot visual classification, but their performance on fine-grained tasks and large hierarchical label spaces is understudied. This paper investigates whether structured, tree-ba…

Modality Alignment across Trees on Heterogeneous Hyperbolic Manifolds

2025-10-31 · Wei Wu, Xiaomeng Fan, Yuwei Wu, Zhi Gao 외 arxiv

Modality alignment is critical for vision-language models (VLMs) to effectively integrate information across modalities. However, existing methods extract hierarchical features from text while representing each image wit…

Dependency-aware Prototype Learning for Few-shot Relation Classification

2022-10-01 · COLING 2022 10 · Tianshu Yu, Min Yang, Xiaoyan Zhao

Few-shot relation classification aims to classify the relation type between two given entities in a sentence by training with a few labeled instances for each relation. However, most of existing models fail to distinguis…

ClassificationFew-Shot Relation ClassificationRelationRelation Classification+1

GlobalGeoTree: A Multi-Granular Vision-Language Dataset for Global Tree Species Classification

2025-05-18 · Yang Mu, Zhitong Xiong, Yi Wang, Muhammad Shahzad 외

Global tree species mapping using remote sensing data is vital for biodiversity monitoring, forest management, and ecological research. However, progress in this field has been constrained by the scarcity of large-scale,…

Benchmarking

Learning Context-Aware Representations of Subtrees

2021-11-08 · Cedric Cook

This thesis tackles the problem of learning efficient representations of complex, structured data with a natural application to web page and element classification. We hypothesise that the context around the element insi…

Classification