paper-with-me

Papers

Deep Reinforcement Active Learning for Human-in-the-Loop Person Re-Identification

2019-10-01 · ICCV 2019 10 · Zimo Liu, Jingya Wang, Shaogang Gong, Huchuan Lu, Dacheng Tao

Most existing person re-identification(Re-ID) approaches achieve superior results based on the assumption that a large amount of pre-labelled data is usually available and can be put into training phrase all at once. However, this assumption is not applicable to most real-world deployment of the Re-ID task. In this work, we propose an alternative reinforcement learning based human-in-the-loop model which releases the restriction of pre-labelling and keeps model upgrading with progressively collected data. The goal is to minimize human annotation efforts while maximizing Re-ID performance. It works in an iteratively updating framework by refining the RL policy and CNN parameters alternately. In particular, we formulate a Deep Reinforcement Active Learning (DRAL) method to guide an agent (a model in a reinforcement learning process) in selecting training samples on-the-fly by a human user/annotator. The reinforcement learning reward is the uncertainty value of each human selected sample. A binary feedback (positive or negative) labelled by the human annotator is used to select the samples of which are used to fine-tune a pre-trained CNN Re-ID model. Extensive experiments demonstrate the superiority of our DRAL method for deep reinforcement learning based human-in-the-loop person Re-ID when compared to existing unsupervised and transfer learning models as well as active learning models.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Active LearningDeep Reinforcement LearningPerson Re-Identificationreinforcement-learningReinforcement LearningReinforcement Learning (RL)Transfer Learning

Similar Papers 제목 키워드 기반

Person Re-Identification in Identity Regression Space

2018-06-25 · Hanxiao Wang, Xiatian Zhu, Shaogang Gong, Tao Xiang

Most existing person re-identification (re-id) methods are unsuitable for real-world deployment due to two reasons: Unscalability to large population size, and Inadaptability over time. In this work, we present a unified…

BenchmarkingIncremental LearningPerson Re-Identificationregression

Highly Efficient Regression for Scalable Person Re-Identification

2016-12-05 · Hanxiao Wang, Shaogang Gong, Tao Xiang

Existing person re-identification models are poor for scaling up to large data required in real-world applications due to: (1) Complexity: They employ complex models for optimal performance resulting in high computationa…

Active LearningPerson Re-Identificationregression

FaiR-IoT: Fairness-aware Human-in-the-Loop Reinforcement Learning for Harnessing Human Variability in Personalized IoT

2021-03-30 · Salma Elmalaki

Thanks to the rapid growth in wearable technologies, monitoring complex human context becomes feasible, paving the way to develop human-in-the-loop IoT systems that naturally evolve to adapt to the human and environment …

FairnessGeneral Reinforcement Learningreinforcement-learningReinforcement Learning (RL)

Human Decision Makings on Curriculum Reinforcement Learning with Difficulty Adjustment

2022-08-04 · Yilei Zeng, Jiali Duan, Yang Li, Emilio Ferrara 외

Human-centered AI considers human experiences with AI performance. While abundant research has been helping AI achieve superhuman performance either by fully automatic or weak supervision learning, fewer endeavors are ex…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Human-In-The-Loop Person Re-Identification

2016-12-05 · Hanxiao Wang, Shaogang Gong, Xiatian Zhu, Tao Xiang

Current person re-identification (re-id) methods assume that (1) pre-labelled training data are available for every camera pair, (2) the gallery size for re-identification is moderate. Both assumptions scale poorly to re…

Ensemble LearningIncremental LearningPerson Re-Identification