Dynamic Face Video Segmentation via Reinforcement Learning
Yujiang Wang, Mingzhi Dong, Jie Shen, Yang Wu, Shiyang Cheng, Maja Pantic
Abstract
For real-time semantic video segmentation, most recent works utilised a dynamic framework with a key scheduler to make online key/non-key decisions. Some works used a fixed key scheduling policy, while others proposed adaptive key scheduling methods based on heuristic strategies, both of which may lead to suboptimal global performance. To overcome this limitation, we model the online key decision process in dynamic video segmentation as a deep reinforcement learning problem and learn an efficient and effective scheduling policy from expert information about decision history and from the process of maximising global return. Moreover, we study the application of dynamic video segmentation on face videos, a field that has not been investigated before. By evaluating on the 300VW dataset, we show that the performance of our reinforcement key scheduler outperforms that of various baselines in terms of both effective key selections and running speed. Further results on the Cityscapes dataset demonstrate that our proposed method can also generalise to other scenarios. To the best of our knowledge, this is the first work to use reinforcement learning for online key-frame decision in dynamic video segmentation, and also the first work on its application on face videos.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- Do Smart Glasses Dream of Sentimental Visions?: Deep Emotionship Analysis for Eyewear DevicesYingying Zhao, Yuhu Chang, Yutian Lu, Yujiang Wang et al.UbiComp 2022 · 18 citations
- MemX: An Attention-Aware Smart Eyewear System for Personalized Moment Auto-captureYuhu Chang, Yingying Zhao, Mingzhi Dong, Yujiang Wang et al.UbiComp 2021 · 15 citations
- Enhancing Low-Rank Adaptation with Recoverability-Based Reinforcement Pruning for Object CountingHaojie Guo, Junyu Gao, Yuan YuanAAAI 2025 · 4 citations
Related papers
- Learning To Recommend Frame for Interactive Video Object Segmentation in the WildZhaoyuan Yin, Jia Zheng, Weixin Luo, Shenhan Qian et al.CVPR 2021
- Reinforced active learning for image segmentationArantxa Casanova, Pedro O. Pinheiro, Negar Rostamzadeh, Christopher J. PalICLR 2020 · 127 citations
- Fast Template Matching and Update for Video Object Tracking and SegmentationMingjie Sun, Jimin Xiao, Eng Gee Lim, Bingfeng Zhang et al.CVPR 2020
- VideoSeg-R1: Reasoning Video Object Segmentation via Reinforcement LearningZishan Xu, Yifu Guo, Yuquan Lu, Fengyu Yang et al.AAAI 2026
- STRONG: Spatio-Temporal Reinforcement Learning for Cross-Modal Video Moment LocalizationDa Cao, Yawen Zeng, Meng Liu, Xiangnan He et al.ACM MM 2020 · 47 citations
