Scoot: A Perceptual Metric for Facial Sketches
Deng-Ping Fan, Shengchuan Zhang, Yu-Huan Wu, Yun Liu, Ming-Ming Cheng, Bo Ren, Paul L. Rosin, Rongrong Ji
摘要
While it is trivial for humans to quickly assess the perceptual similarity between two images, the underlying mechanism are thought to be quite complex. Despite this, the most widely adopted perceptual metrics today, such as SSIM and FSIM, are simple, shallow functions, and fail to consider many factors of human perception. Recently, the facial modeling community has observed that the inclusion of both structure and texture has a significant positive benefit for face sketch synthesis (FSS). But how perceptual are these so-called “perceptual features”? Which elements are critical for their success? In this paper, we design a perceptual metric, called Structure Co-Occurrence Texture (Scoot), which simultaneously considers the block-level spatial structure and co-occurrence texture statistics. To test the quality of metrics, we propose three novel meta-measures based on various reliable properties. Extensive experiments verify that our Scoot metric exceeds the performance of prior work. Besides, we built the first largest scale (152k judgments) human-perception-based sketch database that can evaluate how well a metric consistent with human perception. Our results suggest that “spatial structure” and “co-occurrence texture” are two generally applicable perceptual features in face sketch synthesis.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Beyond Domain Gap: Exploiting Subjectivity in Sketch-Based Person RetrievalKejun Lin, Zhixiang Wang, Zheng Wang, Yinqiang Zheng 等ACM MM 2023 · 被引用 16 次
- Human-Inspired Facial Sketch Synthesis with Dynamic AdaptationFei Gao, Yifan Zhu, Chang Jiang, Nannan WangICCV 2023 · 被引用 10 次
- Kaleido-BERT: Vision-Language Pre-Training on Fashion DomainMingchen Zhuge, Dehong Gao, Deng-Ping Fan, Linbo Jin 等CVPR 2021
- Deep Generative Model based Rate-Distortion for Image Downscaling AssessmentYuanbang Liang, Bhavesh Garg, Paul L. Rosin, Yipeng QinCVPR 2024
相关 Paper
- DreamSim: Learning New Dimensions of Human Visual Similarity using Synthetic DataStephanie Fu, Netanel Tamir, Shobhita Sundaram, Lucy Chai 等NeurIPS 2023 · 被引用 413 次
- ArtFRD: A Fisher-Rao Mixture Metric for Generative Model Aesthetic EvaluationChuanwei Huang, Zexi Jia, Hongyan Fei, Yeshuang Zhu 等ACM MM 2025 · 被引用 2 次
- Object-Centric Image Generation from LayoutsTristan Sylvain, Pengchuan Zhang, Yoshua Bengio, R. Devon Hjelm 等AAAI 2021 · 被引用 107 次
- Privacy Assessment on Reconstructed Images: Are Existing Evaluation Metrics Faithful to Human Perception?Xiaoxiao Sun, Nidham Gazagnadou, Vivek Sharma, Lingjuan Lyu 等NeurIPS 2023 · 被引用 26 次
- Multi-Dimensional Text-to-Face Image Quality Assessment Using LLM: Database and MethodYixuan Gao, Xiongkuo Min, Jinliang Han, Yuqin Cao 等ACM MM 2025 · 被引用 2 次
