Deep Reinforcement Active Learning for Human-in-the-Loop Person Re-Identification
Zimo Liu, Jingya Wang, Shaogang Gong, Dacheng Tao, Huchuan Lu
Abstract
Most existing person re-identification(Re-ID) approaches achieve superior results based on the assumption that a large amount of pre-labelled data is usually available and can be put into training phrase all at once. However, this assumption is not applicable to most real-world deployment of the Re-ID task. In this work, we propose an alternative reinforcement learning based human-in-the-loop model which releases the restriction of pre-labelling and keeps model upgrading with progressively collected data. The goal is to minimize human annotation efforts while maximizing Re-ID performance. It works in an iteratively updating framework by refining the RL policy and CNN parameters alternately. In particular, we formulate a Deep Reinforcement Active Learning (DRAL) method to guide an agent (a model in a reinforcement learning process) in selecting training samples on-the-fly by a human user/annotator. The reinforcement learning reward is the uncertainty value of each human selected sample. A binary feedback (positive or negative) labelled by the human annotator is used to select the samples of which are used to fine-tune a pre-trained CNN Re-ID model. Extensive experiments demonstrate the superiority of our DRAL method for deep reinforcement learning based human-in-the-loop person Re-ID when compared to existing unsupervised and transfer learning models as well as active learning models.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9dd33481-d8bf-48cf-9cc9-9c190c3a17bdCited by top-tier papers16
- Lifelong Person Re-identification by Pseudo Task Knowledge PreservationWenhang Ge, Junlong Du, Ancong Wu, Yuqiao Xian et al.AAAI 2022 · 54 citations
- Lifelong Person Re-identification via Knowledge Refreshing and ConsolidationChunlin Yu, Ye Shi, Zimo Liu, Shenghua Gao et al.AAAI 2023 · 52 citations
- Cross-Camera Feature Prediction for Intra-Camera Supervised Person Re-identification across Distant ScenesWenhang Ge, Chunyan Pan, Ancong Wu, Hongwei Zheng et al.ACM MM 2021 · 30 citations
- Prompting in the Dark: Assessing Human Performance in Prompt Engineering for Data Labeling When Gold Labels Are AbsentZeyu He, Saniya Naphade, Ting-Hao 'Kenneth' HuangCHI 2025 · 22 citations
- Semi-supervised Active Learning for Video Action DetectionAyush Singh, Aayush Jung Rana, Akash Kumar, Shruti Vyas et al.AAAI 2024 · 22 citations
Related papers
- PEBBLE: Feedback-Efficient Interactive Reinforcement Learning via Relabeling Experience and Unsupervised Pre-trainingKimin Lee, Laura M. Smith, Pieter AbbeelICML 2021 · 380 citations
- Reinforced active learning for image segmentationArantxa Casanova, Pedro O. Pinheiro, Negar Rostamzadeh, Christopher J. PalICLR 2020 · 127 citations
- CrowdRL: An End-to-End Reinforcement Learning Framework for Data LabellingKaiyu Li, Guoliang Li, Yong Wang, Yan Huang et al.ICDE 2021 · 17 citations
- Human-in-the-Loop Vehicle ReIDZepeng Li, Dongxiang Zhang, Yanyan Shen, Gang ChenAAAI 2023 · 3 citations
- BAR - A Reinforcement Learning Agent for Bounding-Box Automated RefinementMorgane Ayle, Jimmy Tekli, Julia El Zini, Boulos El Asmar et al.AAAI 2020 · 9 citations
