paper-with-me

홈 › Papers

Towards Efficient and General-Purpose Few-Shot Misclassification Detection for Vision-Language Models

2025-03-26 · Fanhu Zeng, Zhen Cheng, Fei Zhu, Xu-Yao Zhang

Reliable prediction by classifiers is crucial for their deployment in high security and dynamically changing situations. However, modern neural networks often exhibit overconfidence for misclassified predictions, highlighting the need for confidence estimation to detect errors. Despite the achievements obtained by existing methods on small-scale datasets, they all require training from scratch and there are no efficient and effective misclassification detection (MisD) methods, hindering practical application towards large-scale and ever-changing datasets. In this paper, we pave the way to exploit vision language model (VLM) leveraging text information to establish an efficient and general-purpose misclassification detection framework. By harnessing the power of VLM, we construct FSMisD, a Few-Shot prompt learning framework for MisD to refrain from training from scratch and therefore improve tuning efficiency. To enhance misclassification detection ability, we use adaptive pseudo sample generation and a novel negative loss to mitigate the issue of overconfidence by pushing category prompts away from pseudo features. We conduct comprehensive experiments with prompt learning methods and validate the generalization ability across various datasets with domain shift. Significant and consistent improvement demonstrates the effectiveness, efficiency and generalizability of our approach.

📄 PDF Abstract BibTeX arXiv:2503.20492

Code (0)

등록된 구현이 없습니다.

Tasks

Prompt Learning

Similar Papers 제목 키워드 기반

Misclassification Detection via Class Augmentation

2021-01-01 · Fei Zhu, Xu-Yao Zhang, Chuang Wang, Cheng-Lin Liu

Despite the impressive performance in various pattern recognition tasks, deep neural networks (DNNs) are typically overconfident in their predictions, making it difficult to determine whether a test example is misclassif…

Few-Shot Learning

Exploring the Limits of Zero Shot Vision Language Models for Hate Meme Detection: The Vulnerabilities and their Interpretations

2024-02-19 · Naquee Rizwan, Paramananda Bhaskar, Mithun Das, Swadhin Satyaprakash Majhi 외

There is a rapid increase in the use of multimedia content in current social media platforms. One of the highly popular forms of such multimedia content are memes. While memes have been primarily invented to promote funn…

Prompt EngineeringZero-Shot Learning

CRoF: CLIP-based Robust Few-shot Learning on Noisy Labels

2024-12-17 · Shizhuo Deng, Bowen Han, Jiaqi Chen, Hao Wang 외

Noisy labels threaten the robustness of few-shot learning (FSL) due to the inexact features in a new domain. CLIP, a large-scale vision-language model, performs well in FSL on image-text embedding similarities, but it is…

Domain GeneralizationFew-Shot Learningzero-shot-classificationZero-Shot Learning

Meta-DETR: Image-Level Few-Shot Object Detection with Inter-Class Correlation Exploitation

2021-03-22 · Gongjie Zhang, Zhipeng Luo, Kaiwen Cui, Shijian Lu

Few-shot object detection has been extensively investigated by incorporating meta-learning into region-based detection frameworks. Despite its success, the said paradigm is constrained by several factors, such as (i) low…

Few-Shot Object DetectionMeta-Learningobject-detectionObject Detection+1

Meta-DETR: Image-Level Few-Shot Detection with Inter-Class Correlation Exploitation

2022-07-30 · Gongjie Zhang, Zhipeng Luo, Kaiwen Cui, Shijian Lu 외

Few-shot object detection has been extensively investigated by incorporating meta-learning into region-based detection frameworks. Despite its success, the said paradigm is still constrained by several factors, such as (…

Few-Shot Object DetectionMeta-LearningObjectobject-detection+1