paper-with-me

홈 › Papers

OOWL500: Overcoming Dataset Collection Bias in the Wild

2021-08-24 · Brandon Leung, Chih-Hui Ho, Amir Persekian, David Orozco, Yen Chang, Erik Sandstrom, Bo Liu, Nuno Vasconcelos

The hypothesis that image datasets gathered online "in the wild" can produce biased object recognizers, e.g. preferring professional photography or certain viewing angles, is studied. A new "in the lab" data collection infrastructure is proposed consisting of a drone which captures images as it circles around objects. Crucially, the control provided by this setup and the natural camera shake inherent to flight mitigate many biases. It's inexpensive and easily replicable nature may also potentially lead to a scalable data collection effort by the vision community. The procedure's usefulness is demonstrated by creating a dataset of Objects Obtained With fLight (OOWL). Denoted as OOWL500, it contains 120,000 images of 500 objects and is the largest "in the lab" image dataset available when both number of classes and objects per class are considered. Furthermore, it has enabled several of new insights on object recognition. First, a novel adversarial attack strategy is proposed, where image perturbations are defined in terms of semantic properties such as camera shake and pose. Indeed, experiments have shown that ImageNet has considerable amounts of pose and professional photography bias. Second, it is used to show that the augmentation of in the wild datasets, such as ImageNet, with in the lab data, such as OOWL500, can significantly decrease these biases, leading to object recognizers of improved generalization. Third, the dataset is used to study questions on "best procedures" for dataset collection. It is revealed that data augmentation with synthetic images does not suffice to eliminate in the wild datasets biases, and that camera shake and pose diversity play a more important role in object recognition robustness than previously thought.

📄 PDF Abstract BibTeX arXiv:2108.10992

Code (0)

등록된 구현이 없습니다.

Tasks

Adversarial AttackData AugmentationObjectObject Recognition

Similar Papers 제목 키워드 기반

Animal Wildlife Population Estimation Using Social Media Images Collections

2019-08-05 · Matteo Foglio, Lorenzo Semeria, Guido Muscioni, Riccardo Pressiani 외

We are losing biodiversity at an unprecedented scale and in many cases, we do not even know the basic data for the species. Traditional methods for wildlife monitoring are inadequate. Development of new computer vision t…

ContextLabeler Dataset: physical and virtual sensors data collected from smartphone usage in-the-wild

2023-07-07 · Mattia Giovanni Campana, Franca Delmastro

This paper describes a data collection campaign and the resulting dataset derived from smartphone sensors characterizing the daily life activities of 3 volunteers in a period of two weeks. The dataset is released as a co…

Contemplating Visual Emotions: Understanding and Overcoming Dataset Bias

2018-08-07 · ECCV 2018 9 · Rameswar Panda, Jianming Zhang, Haoxiang Li, Joon-Young Lee 외

While machine learning approaches to visual emotion recognition offer great promise, current methods consider training and testing models on small scale datasets covering limited visual emotion concepts. Our analysis ide…

Emotion Recognition

It is Okay to Not Be Okay: Overcoming Emotional Bias in Affective Image Captioning by Contrastive Data Collection

2022-04-15 · CVPR 2022 1 · Youssef Mohamed, Faizan Farooq Khan, Kilichbek Haydarov, Mohamed Elhoseiny

Datasets that capture the connection between vision, language, and affection are limited, causing a lack of understanding of the emotional aspect of human intelligence. As a step in this direction, the ArtEmis dataset wa…

Image Captioning

The iWildCam 2021 Competition Dataset

2021-05-07 · Sara Beery, Arushi Agarwal, Elijah Cole, Vighnesh Birodkar

Camera traps enable the automatic collection of large quantities of image data. Ecologists use camera traps to monitor animal populations all over the world. In order to estimate the abundance of a species from camera tr…

object-detectionObject Detection