paper-with-me

홈 › Papers

Geometry-Guided Self-Supervision for Ultra-Fine-Grained Recognition with Limited Data

2026-04-21 · Shijie Wang, Yadan Luo, Zijian Wang, Haojie Li, Zi Huang, Mahsa Baktashmotlagh arxiv

This paper investigates the intrinsic geometrical features of highly similar objects and introduces a general self-supervised framework called the Geometric Attribute Exploration Network (GAEor), which is designed to address the ultra-fine-grained visual categorization (Ultra-FGVC) task in data-limited scenarios. Unlike prior work that often captures subtle yet critical distinctions, GAEor generates geometric attributes as novel alternative recognition cues. These attributes are determined by various details within the object, aligned with its geometric patterns, such as the intricate vein structures in soybean leaves. Crucially, each category exhibits distinct geometric descriptors that serve as powerful cues, even among objects with minimal visual variation -- a factor largely overlooked in recent research. GAEor discovers these geometric attributes by first amplifying geometry-relevant details via visual feedback from a backbone network, then embedding the relative polar coordinates of these details into the final representation. Extensive experiments demonstrate that GAEor significantly sets new state-of-the-art records in five widely-used Ultra-FGVC benchmarks.

📄 PDF Abstract BibTeX arXiv:2604.19345

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

MAGE: View-guided Point Cloud Completion with Efficient Modality Alignment and Adaptive Geometry Enhancement

2026-06-30 · Weize Quan, Zhengwei Wu, Kai Wang, Dong-Ming Yan arxiv

View-based point cloud completion aims to recover a complete 3D shape from a partial point cloud, guided by a single-view image. However, existing approaches often suffer from limited performance due to weak modality ali…

Point Cloud Completion

Geometry Guided Convolutional Neural Networks for Self-Supervised Video Representation Learning

2018-06-01 · CVPR 2018 6 · Chuang Gan, Boqing Gong, Kun Liu, Hao Su 외

It is often laborious and costly to manually annotate videos for training high-quality video recognition models, so there has been some work and interest in exploring alternative, cheap, and yet often noisy and indirect,…

Action RecognitionRepresentation LearningScene RecognitionSelf-Supervised Learning+3

Exploring the Utility of Self-Supervised Pretraining Strategies for the Detection of Absent Lung Sliding in M-Mode Lung Ultrasound

2023-04-05 · Blake VanBerlo, Brian Li, Alexander Wong, Jesse Hoey 외

Self-supervised pretraining has been observed to improve performance in supervised learning tasks in medical imaging. This study investigates the utility of self-supervised pretraining prior to conducting supervised fine…

Data Augmentation

SG-WAM: Self-Guided World Modeling in Geometry-Aware Policy Space

2026-08-02 · Ruiteng Zhao, Zhengshen Zhang, Yue Su, Wenshuo Wang 외 hf

World Action Models (WAMs) couple action generation with prediction of future states. Their effectiveness depends on whether future dynamics are modeled in a space that is both aligned with action generation and sufficie…

Hunting Sparsity: Density-Guided Contrastive Learning for Semi-Supervised Semantic Segmentation

2023-01-01 · CVPR 2023 1 · Xiaoyang Wang, Bingfeng Zhang, Limin Yu, Jimin Xiao

Recent semi-supervised semantic segmentation methods combine pseudo labeling and consistency regularization to enhance model generalization from perturbation-invariant training. In this work, we argue that adequate s…

Contrastive LearningDensity EstimationSemantic SegmentationSemi-Supervised Semantic Segmentation