FaceBench: A Multi-View Multi-Level Facial Attribute VQA Dataset for Benchmarking Face Perception MLLMs
Xiaoqin Wang, Xusen Ma, Xianxu Hou, Meidan Ding, Yudong Li, Junliang Chen, Wenting Chen, Xiaoyang Peng, Linlin Shen
2025Year
3Top-tier citations
Abstract
What is the shape of the person's face in the image? people Natural lighting Face-LLaVA What shape are the glasses worn by the person in the image? people Is the person in the image wearing a hat? people What type of hairline does the person in the image have? people What type of lighting is present in the image? people
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- Division of Labor and Collaboration Between Parents in Family EducationZiyi Wang, Congrong Zhang, Jingying Deng, Xiaofan Hu et al.CHI 2026 · 2 citations
- SD-FSMIS: Adapting Stable Diffusion for Few-Shot Medical Image SegmentationMeihua Li, Yang Zhang, Weizhao He, Hu Qu et al.CVPR 2026 · 1 citation
- UniFace: A fied ine-grained Understanding and Generation ModelJunzhe Li, Sifan Zhou, Liya Guo, Xuerui Qiu et al.ICLR 2026
Builds on16
- Visual Instruction TuningHaotian Liu, Chunyuan Li, Qingyang Wu, Yong Jae LeeNeurIPS 2023 · 11,349 citations
- BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language ModelsJunnan Li, Dongxu Li, Silvio Savarese, Steven C. H. HoiICML 2023 · 7,873 citations
- Flamingo: a Visual Language Model for Few-Shot LearningJean-Baptiste Alayrac, Jeff Donahue, Pauline Luc, Antoine Miech et al.NeurIPS 2022 · 6,707 citations
- MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language ModelsDeyao Zhu, Jun Chen, Xiaoqian Shen, Xiang Li et al.ICLR 2024 · 3,079 citations
- NExT-GPT: Any-to-Any Multimodal LLMShengqiong Wu, Hao Fei, Leigang Qu, Wei Ji et al.ICML 2024 · 786 citations
Related papers
- CelebV-Text: A Large-Scale Facial Text-Video DatasetJianhui Yu, Hao Zhu, Liming Jiang, Chen Change Loy et al.CVPR 2023
- Collaborative Diffusion for Multi-Modal Face Generation and EditingZiqi Huang, Kelvin C. K. Chan, Yuming Jiang, Ziwei LiuCVPR 2023
- DL2G: Degradation-guided Local-to-Global Restoration for Eyeglass Reflection RemovalZhilv Yi, Xiao Lu, Hong Ding, Jingbo Hu et al.CVPR 2025
- FaceInsight: A Multimodal Large Language Model for Face PerceptionJingzhi Li, Changjiang Luo, Ruoyu Chen, Hua Zhang et al.ACM MM 2025 · 3 citations
- FaceLit: Neural 3D Relightable FacesAnurag Ranjan, Kwang Moo Yi, Jen-Hao Rick Chang, Oncel TuzelCVPR 2023
