paper-with-me

Papers

Mask-Guided Feature Extraction and Augmentation for Ultra-Fine-Grained Visual Categorization

2021-09-16 · Zicheng Pan, Xiaohan Yu, Miaohua Zhang, Yongsheng Gao

While the fine-grained visual categorization (FGVC) problems have been greatly developed in the past years, the Ultra-fine-grained visual categorization (Ultra-FGVC) problems have been understudied. FGVC aims at classifying objects from the same species (very similar categories), while the Ultra-FGVC targets at more challenging problems of classifying images at an ultra-fine granularity where even human experts may fail to identify the visual difference. The challenges for Ultra-FGVC mainly comes from two aspects: one is that the Ultra-FGVC often arises overfitting problems due to the lack of training samples; and another lies in that the inter-class variance among images is much smaller than normal FGVC tasks, which makes it difficult to learn discriminative features for each class. To solve these challenges, a mask-guided feature extraction and feature augmentation method is proposed in this paper to extract discriminative and informative regions of images which are then used to augment the original feature map. The advantage of the proposed method is that the feature detection and extraction model only requires a small amount of target region samples with bounding boxes for training, then it can automatically locate the target area for a large number of images in the dataset at a high detection accuracy. Experimental results on two public datasets and ten state-of-the-art benchmark methods consistently demonstrate the effectiveness of the proposed method both visually and quantitatively.

📄 PDF Abstract BibTeX arXiv:2109.07755

Code (0)

등록된 구현이 없습니다.

Tasks

Fine-Grained Visual Categorization

Similar Papers 제목 키워드 기반

Mask-Guided Attention U-Net for Enhanced Neonatal Brain Extraction and Image Preprocessing

2024-06-25 · Bahram Jafrasteh, Simon Pedro Lubian-Lopez, Emiliano Trimarco, Macarena Roman Ruiz 외

In this study, we introduce MGA-Net, a novel mask-guided attention neural network, which extends the U-net model for precision neonatal brain imaging. MGA-Net is designed to extract the brain from other structures and re…

DecoderImage ReconstructionImage SegmentationSemantic Segmentation

Diffusion Model-based Data Augmentation Method for Fetal Head Ultrasound Segmentation

2025-06-30 · Fangyijie Wang, Kevin Whelan, Félix Balado, Kathleen M. Curran 외

Medical image data is less accessible than in other domains due to privacy and regulatory constraints. In addition, labeling requires costly, time-intensive manual image annotation by clinical experts. To overcome these …

Data AugmentationSegmentation

Dual Co-Train: Cross-Dataset Ultrasound Tongue Segmentation Under Extreme Data Scarcity

2026-08-18 · Alisher Myrgyyassov, Zhen Song, Bruce Xiao Wang, Yu Sun 외 arxiv

Ultrasound tongue contour segmentation remains challenging under cross-dataset domain shift, where limited annotations, probe variability, and acquisition noise often degrade model generalization. We present a source-fre…

Source-Free Domain Adaptation

Standardisation of Convex Ultrasound Data Through Geometric Analysis and Augmentation

2025-02-13 · Alistair Weld, Giovanni Faoro, Luke Dixon, Sophie Camp 외

The application of ultrasound in healthcare has seen increased diversity and importance. Unlike other medical imaging modalities, ultrasound research and development has historically lagged, particularly in the case of a…

Benchmarking

EchoPilot: Training-Free Ultrasound Video Segmentation via Scale-Space Semantic Prompting and Reliability-Gated Memory

2026-05-25 · Ruiqiang Xiao, Zhaohu Xing, Yijun Yang, Zhenyan Han 외 arxiv

Ultrasound video segmentation is clinically valuable yet difficult due to speckle noise, weak boundaries, and rapid anatomical deformation. Recent promptable foundation models enable point-guided segmentation, but their …

Video Segmentation