paper-with-me

Papers

Visual-Augmented Dynamic Semantic Prototype for Generative Zero-Shot Learning

2024-04-23 · CVPR 2024 1 · Wenjin Hou, Shiming Chen, Shuhuang Chen, Ziming Hong, Yan Wang, Xuetao Feng, Salman Khan, Fahad Shahbaz Khan, Xinge You

Generative Zero-shot learning (ZSL) learns a generator to synthesize visual samples for unseen classes, which is an effective way to advance ZSL. However, existing generative methods rely on the conditions of Gaussian noise and the predefined semantic prototype, which limit the generator only optimized on specific seen classes rather than characterizing each visual instance, resulting in poor generalizations (\textit{e.g.}, overfitting to seen classes). To address this issue, we propose a novel Visual-Augmented Dynamic Semantic prototype method (termed VADS) to boost the generator to learn accurate semantic-visual mapping by fully exploiting the visual-augmented knowledge into semantic conditions. In detail, VADS consists of two modules: (1) Visual-aware Domain Knowledge Learning module (VDKL) learns the local bias and global prior of the visual features (referred to as domain visual knowledge), which replace pure Gaussian noise to provide richer prior noise information; (2) Vision-Oriented Semantic Updation module (VOSU) updates the semantic prototype according to the visual representations of the samples. Ultimately, we concatenate their output as a dynamic semantic prototype, which serves as the condition of the generator. Extensive experiments demonstrate that our VADS achieves superior CZSL and GZSL performances on three prominent datasets and outperforms other state-of-the-art methods with averaging increases by 6.4\%, 5.9\% and 4.2\% on SUN, CUB and AWA2, respectively.

📄 PDF Abstract BibTeX arXiv:2404.14808

Code (0)

등록된 구현이 없습니다.

Tasks

Zero-Shot Learning

Similar Papers 제목 키워드 기반

Evolving Semantic Prototype Improves Generative Zero-Shot Learning

2023-06-12 · Shiming Chen, Wenjin Hou, Ziming Hong, Xiaohan Ding 외

In zero-shot learning (ZSL), generative methods synthesize class-related sample features based on predefined semantic prototypes. They advance the ZSL performance by synthesizing unseen class sample features for better t…

Zero-Shot Learning

Incentivizing Generative Zero-Shot Learning via Outcome-Reward Reinforcement Learning with Visual Cues

2026-03-22 · Wenjin Hou, Xiaoxiao Sun, Hehe Fan arxiv

Recent advances in zero-shot learning (ZSL) have demonstrated the potential of generative models. Typically, generative ZSL synthesizes visual features conditioned on semantic prototypes to model the data distribution of…

Reinforcement LearningZero-Shot Learning

State and Scene Enhanced Prototypes for Weakly Supervised Open-Vocabulary Object Detection

2025-11-22 · Jiaying Zhou, Qingchao Chen arxiv

Open-Vocabulary Object Detection (OVOD) aims to generalize object recognition to novel categories, while Weakly Supervised OVOD (WS-OVOD) extends this by combining box-level annotations with image-level labels. Despite r…

Object RecognitionObject Detection

Learning Semantic Ambiguities for Zero-Shot Learning

2022-01-05 · Celina Hanouti, Hervé Le Borgne

Zero-shot learning (ZSL) aims at recognizing classes for which no visual sample is available at training time. To address this issue, one can rely on a semantic description of each class. A typical ZSL model learns a map…

Zero-Shot Learning

Prototype-Guided Curriculum Learning for Zero-Shot Learning

2025-08-11 · Lei Wang, Shiming Chen, Guo-Sen Xie, Ziming Hong 외 arxiv

In Zero-Shot Learning (ZSL), embedding-based methods enable knowledge transfer from seen to unseen classes by learning a visual-semantic mapping from seen-class images to class-level semantic prototypes (e.g., attributes…

Zero-Shot Learning