paper-with-me

홈 › Papers

Unlocking the Power of Open Set : A New Perspective for Open-Set Noisy Label Learning

2023-05-07 · Wenhai Wan, Xinrui Wang, Ming-Kun Xie, Shao-Yuan Li, Sheng-Jun Huang, Songcan Chen

Learning from noisy data has attracted much attention, where most methods focus on closed-set label noise. However, a more common scenario in the real world is the presence of both open-set and closed-set noise. Existing methods typically identify and handle these two types of label noise separately by designing a specific strategy for each type. However, in many real-world scenarios, it would be challenging to identify open-set examples, especially when the dataset has been severely corrupted. Unlike the previous works, we explore how models behave when faced with open-set examples, and find that \emph{a part of open-set examples gradually get integrated into certain known classes}, which is beneficial for the separation among known classes. Motivated by the phenomenon, we propose a novel two-step contrastive learning method CECL (Class Expansion Contrastive Learning) which aims to deal with both types of label noise by exploiting the useful information of open-set examples. Specifically, we incorporate some open-set examples into closed-set classes to enhance performance while treating others as delimiters to improve representative ability. Extensive experiments on synthetic and real-world datasets with diverse label noise demonstrate the effectiveness of CECL.

📄 PDF Abstract BibTeX arXiv:2305.04203

Code (0)

등록된 구현이 없습니다.

Tasks

Contrastive Learning

Methods 이 논문이 사용한 방법론

Contrastive Learning 설명 없음

Similar Papers 제목 키워드 기반

Open Data on GitHub: Unlocking the Potential of AI

2023-06-09 · Anthony Cintron Roman, Kevin Xu, Arfon Smith, Jehu Torres Vega 외

GitHub is the world's largest platform for collaborative software development, with over 100 million users. GitHub is also used extensively for open data collaboration, hosting more than 800 million open data files, tota…

Aspen Open Jets: Unlocking LHC Data for Foundation Models in Particle Physics

2024-12-13 · Oz Amram, Luca Anzalone, Joschka Birk, Darius A. Faroughy 외

Foundation models are deep learning models pre-trained on large amounts of data which are capable of generalizing to multiple datasets and/or downstream tasks. This work demonstrates how data collected by the CMS experim…

Digitizing Paper ECGs at Scale: An Open-Source Algorithm for Clinical Research

2025-10-22 · Elias Stenhede, Agnar Martin Bjørnstad, Arian Ranjbar arxiv

Millions of clinical ECGs exist only as paper scans, making them unusable for modern automated diagnostics. We introduce a fully automated, modular framework that converts scanned or photographed ECGs into digital signal…

MedPerf: Open Benchmarking Platform for Medical Artificial Intelligence using Federated Evaluation

2021-09-29 · Alexandros Karargyris, Renato Umeton, Micah J. Sheller, Alejandro Aristizabal 외

Medical AI has tremendous potential to advance healthcare by supporting the evidence-based practice of medicine, personalizing patient treatment, reducing costs, and improving provider and patient experience. We argue th…

BenchmarkingPhilosophy

Taming Self-Training for Open-Vocabulary Object Detection

2023-08-11 · CVPR 2024 1 · Shiyu Zhao, Samuel Schulter, Long Zhao, Zhixing Zhang 외

Recent studies have shown promising performance in open-vocabulary object detection (OVD) by utilizing pseudo labels (PLs) from pretrained vision and language models (VLMs). However, teacher-student self-training, a powe…

Objectobject-detectionObject DetectionOpen-vocabulary object detection+1