paper-with-me

홈 › Papers

Efficient Label Collection for Unlabeled Image Datasets

2015-06-01 · CVPR 2015 6 · Maggie Wigness, Bruce A. Draper, J. Ross Beveridge

Visual classifiers are part of many applications including surveillance, autonomous navigation and scene understanding. The raw data used to train these classifiers is abundant and easy to collect but lacks labels. Labels are necessary for training supervised classifiers, but the labeling process requires significant human effort. Techniques like active learning and group-based labeling have emerged to help reduce the labeling workload. However, the possibility of collecting label noise affects either the efficiency of these systems or the performance of the trained classifiers. Further, many introduce latency by iteratively re-training classifiers or re-clustering data. We introduce a technique that searches for structural change in hierarchically clustered data to identify a set of clusters that span a spectrum of visual concept granularities. This allows us to efficiently label clusters with less label noise and produce high performing classifiers. The data is hierarchically clustered only once, eliminating latency during the labeling process. Using benchmark data we show that collecting labels with our approach is more efficient than existing labeling techniques, and achieves higher classification accuracy. Finally, we demonstrate the speed and efficiency of our system using real-world data collected for an autonomous navigation task.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Active LearningAutonomous NavigationClusteringScene Understanding

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

Generating Synthetic Handwritten Historical Documents With OCR Constrained GANs

2021-03-15 · Lars Vögtlin, Manuel Drazyk, Vinaychandran Pondenkandath, Michele Alberti 외

We present a framework to generate synthetic historical documents with precise ground truth using nothing more than a collection of unlabeled historical images. Obtaining large labeled datasets is often the limiting fact…

Optical Character Recognition (OCR)Synthetic Data Generation

LaFTer: Label-Free Tuning of Zero-shot Classifier using Language and Unlabeled Image Collections

2023-05-29 · NeurIPS 2023 11

Recently, large-scale pre-trained Vision and Language (VL) models have set a new state-of-the-art (SOTA) in zero-shot visual classification enabling open-vocabulary recognition of potentially unlimited set of categories …

Language ModelingLanguage ModellingLarge Language Model

OSSGAN: Open-Set Semi-Supervised Image Generation

2022-04-29 · CVPR 2022 1 · Kai Katsumata, Duc Minh Vo, Hideki Nakayama

We introduce a challenging training scheme of conditional GANs, called open-set semi-supervised image generation, where the training dataset consists of two parts: (i) labeled data and (ii) unlabeled data with samples be…

Image Generation

Making Binary Classification from Multiple Unlabeled Datasets Almost Free of Supervision

2023-06-12 · Yuhao Wu, Xiaobo Xia, Jun Yu, Bo Han 외

Training a classifier exploiting a huge amount of supervised data is expensive or even prohibited in a situation, where the labeling cost is high. The remarkable progress in working with weaker forms of supervision is bi…

Binary ClassificationPseudo Label

Motion-Augmented Self-Training for Video Recognition at Smaller Scale

2021-05-04 · ICCV 2021 10 · Kirill Gavrilyuk, Mihir Jain, Ilia Karmanov, Cees G. M. Snoek

The goal of this paper is to self-train a 3D convolutional neural network on an unlabeled video collection for deployment on small-scale video collections. As smaller video datasets benefit more from motion than appearan…

Action RecognitionOptical Flow EstimationRetrievalTransfer Learning+1