Deep Reinforcement Active Learning for Human-in-the-Loop Person Re-Identification
Zimo Liu, Jingya Wang, Shaogang Gong, Dacheng Tao, Huchuan Lu
摘要
Most existing person re-identification(Re-ID) approaches achieve superior results based on the assumption that a large amount of pre-labelled data is usually available and can be put into training phrase all at once. However, this assumption is not applicable to most real-world deployment of the Re-ID task. In this work, we propose an alternative reinforcement learning based human-in-the-loop model which releases the restriction of pre-labelling and keeps model upgrading with progressively collected data. The goal is to minimize human annotation efforts while maximizing Re-ID performance. It works in an iteratively updating framework by refining the RL policy and CNN parameters alternately. In particular, we formulate a Deep Reinforcement Active Learning (DRAL) method to guide an agent (a model in a reinforcement learning process) in selecting training samples on-the-fly by a human user/annotator. The reinforcement learning reward is the uncertainty value of each human selected sample. A binary feedback (positive or negative) labelled by the human annotator is used to select the samples of which are used to fine-tune a pre-trained CNN Re-ID model. Extensive experiments demonstrate the superiority of our DRAL method for deep reinforcement learning based human-in-the-loop person Re-ID when compared to existing unsupervised and transfer learning models as well as active learning models.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- Lifelong Person Re-identification by Pseudo Task Knowledge PreservationWenhang Ge, Junlong Du, Ancong Wu, Yuqiao Xian 等AAAI 2022 · 被引用 54 次
- Lifelong Person Re-identification via Knowledge Refreshing and ConsolidationChunlin Yu, Ye Shi, Zimo Liu, Shenghua Gao 等AAAI 2023 · 被引用 52 次
- Cross-Camera Feature Prediction for Intra-Camera Supervised Person Re-identification across Distant ScenesWenhang Ge, Chunyan Pan, Ancong Wu, Hongwei Zheng 等ACM MM 2021 · 被引用 30 次
- Prompting in the Dark: Assessing Human Performance in Prompt Engineering for Data Labeling When Gold Labels Are AbsentZeyu He, Saniya Naphade, Ting-Hao 'Kenneth' HuangCHI 2025 · 被引用 22 次
- Semi-supervised Active Learning for Video Action DetectionAyush Singh, Aayush Jung Rana, Akash Kumar, Shruti Vyas 等AAAI 2024 · 被引用 22 次
相关 Paper
- PEBBLE: Feedback-Efficient Interactive Reinforcement Learning via Relabeling Experience and Unsupervised Pre-trainingKimin Lee, Laura M. Smith, Pieter AbbeelICML 2021 · 被引用 380 次
- Reinforced active learning for image segmentationArantxa Casanova, Pedro O. Pinheiro, Negar Rostamzadeh, Christopher J. PalICLR 2020 · 被引用 127 次
- CrowdRL: An End-to-End Reinforcement Learning Framework for Data LabellingKaiyu Li, Guoliang Li, Yong Wang, Yan Huang 等ICDE 2021 · 被引用 17 次
- Human-in-the-Loop Vehicle ReIDZepeng Li, Dongxiang Zhang, Yanyan Shen, Gang ChenAAAI 2023 · 被引用 3 次
- BAR - A Reinforcement Learning Agent for Bounding-Box Automated RefinementMorgane Ayle, Jimmy Tekli, Julia El Zini, Boulos El Asmar 等AAAI 2020 · 被引用 9 次
