paper-with-me

Papers

Prompt Ensemble Self-training for Open-Vocabulary Domain Adaptation

2023-06-29 · Jiaxing Huang, Jingyi Zhang, Han Qiu, Sheng Jin, Shijian Lu

Traditional domain adaptation assumes the same vocabulary across source and target domains, which often struggles with limited transfer flexibility and efficiency while handling target domains with different vocabularies. Inspired by recent vision-language models (VLMs) that enable open-vocabulary visual recognition by reasoning on both images and texts, we study open-vocabulary domain adaptation (OVDA), a new unsupervised domain adaptation framework that positions a pre-trained VLM as the source model and transfers it towards arbitrary unlabelled target domains. To this end, we design a Prompt Ensemble Self-training (PEST) technique that exploits the synergy between vision and language to mitigate the domain discrepancies in image and text distributions simultaneously. Specifically, PEST makes use of the complementary property of multiple prompts within and across vision and language modalities, which enables joint exploitation of vision and language information and effective learning of image-text correspondences in the unlabelled target domains. Additionally, PEST captures temporal information via temporal prompt ensemble which helps memorize previously learnt target information. Extensive experiments show that PEST outperforms the state-of-the-art consistently across 10 image recognition tasks.

📄 PDF Abstract BibTeX arXiv:2306.16658

Code (0)

등록된 구현이 없습니다.

Tasks

Domain AdaptationUnsupervised Domain Adaptation

Similar Papers 제목 키워드 기반

Fine-grained Visual-Text Prompt-Driven Self-Training for Open-Vocabulary Object Detection

2022-11-02 · Yanxin Long, Jianhua Han, Runhui Huang, Xu Hang 외

Inspired by the success of vision-language methods (VLMs) in zero-shot classification, recent works attempt to extend this line of work into object detection by leveraging the localization ability of pre-trained VLMs and…

Objectobject-detectionObject DetectionOpen-vocabulary object detection+6

Open-Vocabulary Object Detection with Meta Prompt Representation and Instance Contrastive Optimization

2024-03-14 · Zhao Wang, Aoxue Li, Fengwei Zhou, Zhenguo Li 외

Classical object detectors are incapable of detecting novel class objects that are not encountered before. Regarding this issue, Open-Vocabulary Object Detection (OVOD) is proposed, which aims to detect the objects in th…

Contrastive LearningKnowledge Distillationobject-detectionObject Detection+2

VOVTrack: Exploring the Potentiality in Videos for Open-Vocabulary Object Tracking

2024-10-11 · Zekun Qian, Ruize Han, Junhui Hou, Linqi Song 외

Open-vocabulary multi-object tracking (OVMOT) represents a critical new challenge involving the detection and tracking of diverse object categories in videos, encompassing both seen categories (base classes) and unseen c…

Multi-Object TrackingObjectobject-detectionObject Detection+4

Self-Prompting Diffusion Transformer for Open-Vocabulary Scene Text Editing via In-Context Learning

2026-05-15 · Hongxi Li, Tong Wang, Chengjing Wu, Tianbao Liu 외 arxiv

Scene text editing aims to modify text in a target region of an image while preserving surrounding background style and texture. Existing methods rely solely on image background information while neglecting the visual de…

Open-Vocabulary Object Detection via Language Hierarchy

2024-10-27 · Jiaxing Huang, Jingyi Zhang, Kai Jiang, Shijian Lu

Recent studies on generalizable object detection have attracted increasing attention with additional weak supervision from large-scale datasets with image-level labels. However, weakly-supervised detection learning often…

Objectobject-detectionObject DetectionOpen-vocabulary object detection+1