paper-with-me

홈 › Papers

Hybrid Feature Collaborative Reconstruction Network for Few-Shot Fine-Grained Image Classification

2024-07-02 · Shulei Qiu, Wanqi Yang, Ming Yang

Our research focuses on few-shot fine-grained image classification, which faces two major challenges: appearance similarity of fine-grained objects and limited number of samples. To preserve the appearance details of images, traditional feature reconstruction networks usually enhance the representation ability of key features by spatial feature reconstruction and minimizing the reconstruction error. However, we find that relying solely on a single type of feature is insufficient for accurately capturing inter-class differences of fine-grained objects in scenarios with limited samples. In contrast, the introduction of channel features provides additional information dimensions, aiding in better understanding and distinguishing the inter-class differences of fine-grained objects. Therefore, in this paper, we design a new Hybrid Feature Collaborative Reconstruction Network (HFCR-Net) for few-shot fine-grained image classification, which includes a Hybrid Feature Fusion Process (HFFP) and a Hybrid Feature Reconstruction Process (HFRP). In HFRP, we fuse the channel features and the spatial features. Through dynamic weight adjustment, we aggregate the spatial dependencies between arbitrary two positions and the correlations between different channels of each image to increase the inter-class differences. Additionally, we introduce the reconstruction of channel dimension in HFRP. Through the collaborative reconstruction of channel dimension and spatial dimension, the inter-class differences are further increased in the process of support-to-query reconstruction, while the intra-class differences are reduced in the process of query-to-support reconstruction. Ultimately, our extensive experiments on three widely used fine-grained datasets demonstrate the effectiveness and superiority of our approach.

📄 PDF Abstract BibTeX arXiv:2407.02123

Code (0)

등록된 구현이 없습니다.

Tasks

Fine-Grained Image Classificationimage-classificationImage Classification

Similar Papers 제목 키워드 기반

A Hybrid Quantum Neural Network for Split Learning

2024-09-25 · Hevish Cowlessur, Chandra Thapa, Tansu Alpcan, Seyit Camtepe

Quantum Machine Learning (QML) is an emerging field of research with potential applications to distributed collaborative learning, such as Split Learning (SL). SL allows resource-constrained clients to collaboratively tr…

Quantum Machine Learning

Modality-Collaborative Transformer with Hybrid Feature Reconstruction for Robust Emotion Recognition

2023-12-26 · Chengxin Chen, Pengyuan Zhang

As a vital aspect of affective computing, Multimodal Emotion Recognition has been an active research area in the multimedia community. Despite recent progress, this field still confronts two major challenges in real-worl…

Emotion RecognitionMultimodal Emotion Recognition

DyGLNet: Hybrid Global-Local Feature Fusion with Dynamic Upsampling for Medical Image Segmentation

2025-09-16 · Yican Zhao, Ce Wang, You Hao, Lei Li 외 arxiv

Medical image segmentation grapples with challenges including multi-scale lesion variability, ill-defined tissue boundaries, and computationally intensive processing demands. This paper proposes the DyGLNet, which achiev…

Medical Image SegmentationObject Segmentation

Few-shot point cloud reconstruction and denoising via learned Guassian splats renderings and fine-tuned diffusion features

2024-04-01 · Pietro Bonazzi, Marie-Julie Rakatosaona, Marco Cannici, Federico Tombari 외

Existing deep learning methods for the reconstruction and denoising of point clouds rely on small datasets of 3D shapes. We circumvent the problem by leveraging deep learning methods trained on billions of images. We pro…

3D ReconstructionDeep LearningDenoisingPoint cloud reconstruction

Free Lunch to Meet the Gap: Intermediate Domain Reconstruction for Cross-Domain Few-Shot Learning

2025-11-18 · Tong Zhang, Yifan Zhao, Liangyu Wang, Jia Li arxiv

Cross-Domain Few-Shot Learning (CDFSL) endeavors to transfer generalized knowledge from the source domain to target domains using only a minimal amount of training data, which faces a triplet of learning challenges in th…

cross-domain few-shot learning