DisFaceRep: Representation Disentanglement for Co-occurring Facial Components in Weakly Supervised Face Parsing
Xiaoqin Wang, Xianxu Hou, Meidan Ding, Junliang Chen, Kaijun Deng, Jinheng Xie, Linlin Shen
摘要
Face parsing aims to segment facial images into key components such as eyes, lips, and eyebrows. While existing methods rely on dense pixel-level annotations, such annotations are expensive and labor-intensive to obtain. To reduce annotation cost, we introduce Weakly Supervised Face Parsing (WSFP), a new task setting that performs dense facial component segmentation using only weak supervision, such as image-level labels and natural language descriptions. WSFP introduces unique challenges due to the high co-occurrence and visual similarity of facial components, which lead to ambiguous activations and degraded parsing performance. To address this, we propose DisFaceRep, a representation disentanglement framework designed to separate co-occurring facial components through both explicit and implicit mechanisms. Specifically, we introduce a co-occurring component disentanglement strategy to explicitly reduce dataset-level bias, and a text-guided component disentanglement loss to guide component separation using language supervision implicitly. Extensive experiments on CelebAMask-HQ, LaPa, and Helen demonstrate the difficulty of WSFP and the effectiveness of DisFaceRep, which significantly outperforms existing weakly supervised semantic segmentation methods. The code will be released at https://github.com/CVI-SZU/DisFaceRep.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- FineXtrol: Controllable Motion Generation via Fine-Grained TextKeming Shen, Bizhu Wu, Junliang Chen, Xiaoqin Wang 等AAAI 2026 · 被引用 3 次
- SD-FSMIS: Adapting Stable Diffusion for Few-Shot Medical Image SegmentationMeihua Li, Yang Zhang, Weizhao He, Hu Qu 等CVPR 2026 · 被引用 1 次
- UniFace: A fied ine-grained Understanding and Generation ModelJunzhe Li, Sifan Zhou, Liya Guo, Xuerui Qiu 等ICLR 2026
它引用的顶会 Paper27
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- FSGAN: Subject Agnostic Face Swapping and ReenactmentYuval Nirkin, Yosi Keller, Tal HassnerICCV 2019 · 被引用 710 次
- Multi-class Token Transformer for Weakly Supervised Semantic SegmentationLian Xu, Wanli Ouyang, Mohammed Bennamoun, Farid Boussaïd 等CVPR 2022 · 被引用 275 次
- Integral Object Mining via Online Attention AccumulationPeng-Tao Jiang, Qibin Hou, Yang Cao, Ming-Ming Cheng 等ICCV 2019 · 被引用 246 次
相关 Paper
- Dual-Structure Disentangling Variational Generation for Data-Limited Face ParsingPeipei Li, Yinglu Liu, Hailin Shi, Xiang Wu 等ACM MM 2020 · 被引用 8 次
- A New Dataset and Boundary-Attention Semantic Segmentation for Face ParsingYinglu Liu, Hailin Shi, Hao Shen, Yue Si 等AAAI 2020 · 被引用 88 次
- Parameter Efficient Local Implicit Image Function Network for Face SegmentationMausoom Sarkar, Nikitha S. R., Mayur Hemani, Rishabh Jain 等CVPR 2023
- Decoupled Multi-task Learning with Cyclical Self-Regulation for Face ParsingQingping Zheng, Jiankang Deng, Zheng Zhu, Ying Li 等CVPR 2022 · 被引用 46 次
- Unsupervised Disentanglement of Linear-Encoded Facial SemanticsYutong Zheng, Yu-Kai Huang, Ran Tao, Zhiqiang Shen 等CVPR 2021
