Generic Perceptual Loss for Modeling Structured Output Dependencies
Yifan Liu, Hao Chen, Yu Chen, Wei Yin, Chunhua Shen
摘要
The perceptual loss has been widely used as an effective loss term in image synthesis tasks including image superresolution [16] , and style transfer [14] . It was believed that the success lies in the high-level perceptual feature representations extracted from CNNs pretrained with a large set of images. Here we reveal that, what matters is the network structure instead of the trained weights. Without any learning, the structure of a deep network is sufficient to capture the dependencies between multiple levels of variable statistics using multiple layers of CNNs. This insight removes the requirements of pre-training and a particular network structure (commonly, VGG) that are previously assumed for the perceptual loss, thus enabling a significantly wider range of applications. To this end, we demonstrate that a randomly-weighted deep CNN can be used to model the structured dependencies of outputs. On a few dense perpixel prediction tasks such as semantic segmentation, depth estimation and instance segmentation, we show improved results of using the extended randomized perceptual loss, compared to the baselines using pixel-wise loss alone. We hope that this simple, extended perceptual loss may serve as a generic structured-output loss that is applicable to most structured output learning tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper3
- FCOS: Fully Convolutional One-Stage Object DetectionZhi Tian, Chunhua Shen, Hao Chen, Tong HeICCV 2019 · 被引用 6,042 次
- Enforcing Geometric Constraints of Virtual Normal for Depth PredictionWei Yin, Yifan Liu, Chunhua Shen, Youliang YanICCV 2019 · 被引用 487 次
- BlendMask: Top-Down Meets Bottom-Up for Instance SegmentationHao Chen, Kunyang Sun, Zhi Tian, Chunhua Shen 等CVPR 2020
相关 Paper
- Understanding and Simplifying Perceptual DistancesDan Amir, Yair WeissCVPR 2021
- You Only Need Adversarial Supervision for Semantic Image SynthesisEdgar Schönfeld, Vadim Sushko, Dan Zhang, Juergen Gall 等ICLR 2021 · 被引用 219 次
- SROBB: Targeted Perceptual Loss for Single Image Super-ResolutionMohammad Saeed Rad, Behzad Bozorgtabar, Urs-Viktor Marti, Max Basler 等ICCV 2019 · 被引用 147 次
- Towards Interpretable Face RecognitionBangjie Yin, Luan Tran, Haoxiang Li, Xiaohui Shen 等ICCV 2019 · 被引用 92 次
- Context Prior for Scene SegmentationChangqian Yu, Jingbo Wang, Changxin Gao, Gang Yu 等CVPR 2020
