Towards Learning Structure via Consensus for Face Segmentation and Parsing
Iacopo Masi, Joe Mathai, Wael AbdAlmageed
Abstract
Face segmentation is the task of densely labeling pixels on the face according to their semantics. While current methods place an emphasis on developing sophisticated architectures, use conditional random fields for smoothness, or rather employ adversarial training, we follow an alternative path towards robust face segmentation and parsing. Occlusions, along with other parts of the face, have a proper structure that needs to be propagated in the model during training. Unlike state-of-the-art methods that treat face segmentation as an independent pixel prediction problem, we argue instead that it should hold highly correlated outputs within the same object pixels. We thereby offer a novel learning mechanism to enforce structure in the prediction via consensus, guided by a robust loss function that forces pixel objects to be consistent with each other. Our face parser is trained by transferring knowledge from another model, yet it encourages spatial consistency while fitting the labels. Different than current practice, our method enjoys pixel-wise predictions, yet paves the way for fewer artifacts, less sparse masks, and spatially coherent outputs.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Automatic Animation of Hair Blowing in Still Portrait PhotosWenpeng Xiao, Wentao Liu, Yitong Wang, Bernard Ghanem et al.ICCV 2023 · 15 citations
- MSML: Enhancing Occlusion-Robustness by Multi-Scale Segmentation-Based Mask Learning for Face RecognitionGe Yuan, Huicheng Zheng, Jiayu DongAAAI 2022 · 12 citations
Related papers
- Parameter Efficient Local Implicit Image Function Network for Face SegmentationMausoom Sarkar, Nikitha S. R., Mayur Hemani, Rishabh Jain et al.CVPR 2023
- Interaction via Bi-directional Graph of Semantic Region Affinity for Scene ParsingHenghui Ding, Hui Zhang, Jun Liu, Jiaxin Li et al.ICCV 2021 · 14 citations
- Human Parsing Based Texture Transfer from Single Image to 3D Human via Cross-View ConsistencyFang Zhao, Shengcai Liao, Kaihao Zhang, Ling ShaoNeurIPS 2020 · 22 citations
- Dual-Structure Disentangling Variational Generation for Data-Limited Face ParsingPeipei Li, Yinglu Liu, Hailin Shi, Xiang Wu et al.ACM MM 2020 · 8 citations
- High Fidelity Face Swapping via Semantics Disentanglement and Structure EnhancementFengyuan Liu, Lingyun Yu, Hongtao Xie, Chuanbin Liu et al.ACM MM 2023 · 1 citation
