STAR: A Structure-aware Lightweight Transformer for Real-time Image Enhancement
Zhaoyang Zhang, Yitong Jiang, Jun Jiang, Xiaogang Wang, Ping Luo, Jinwei Gu
摘要
Image and video enhancement such as color constancy, low light enhancement, and tone mapping on smartphones is challenging, because high-quality images should be achieved efficiently with a limited resource budget. Unlike prior works that either used very deep CNNs or large Trans-former models, we propose a structure-aware lightweight Transformer, termed STAR, for real-time image enhancement. STAR is formulated to capture long-range dependencies between image patches, which naturally and implicitly captures the structural relationships of different regions in an image. STAR is a general architecture that can be easily adapted to different image enhancement tasks. Extensive experiments show that STAR can effectively boost the quality and efficiency of many tasks such as illumination enhancement, auto white balance, and photo retouching, which are indispensable components for image processing on smartphones. For example, STAR reduces model complexity and improves image quality compared to the recent state-of-the-art [19] on the MIT-Adobe FiveK dataset [7] (i.e., 1.8dB PSNR improvements with 25% parameters and 13% float operations.)
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- URetinex-Net: Retinex-based Deep Unfolding Network for Low-light Image EnhancementWenhui Wu, Jian Weng, Pingping Zhang, Xu Wang 等CVPR 2022 · 被引用 695 次
- FeatEnHancer: Enhancing Hierarchical Features for Object Detection and Beyond Under Low-Light VisionKhurram Azeem Hashmi, Goutham Kallempudi, Didier Stricker, Muhammad Zeshan AfzalICCV 2023 · 被引用 76 次
- Low-Light Image Enhancement with Illumination-Aware Gamma Correction and Complete Image Modelling NetworkYinglong Wang, Zhen Liu, Jianzhuang Liu, Songcen Xu 等ICCV 2023 · 被引用 70 次
- ShadowFormer: Global Context Helps Shadow RemovalLanqing Guo, Siyu Huang, Ding Liu, Hao Cheng 等AAAI 2023 · 被引用 60 次
- Low-Light Video Enhancement with Synthetic Event GuidanceLin Liu, Junfeng An, Jianzhuang Liu, Shanxin Yuan 等AAAI 2023 · 被引用 51 次
它引用的顶会 Paper15
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa 等ICML 2021 · 被引用 8,974 次
- Tokens-to-Token ViT: Training Vision Transformers from Scratch on ImageNetLi Yuan, Yunpeng Chen, Tao Wang, Weihao Yu 等ICCV 2021 · 被引用 2,462 次
- VL-BERT: Pre-training of Generic Visual-Linguistic RepresentationsWeijie Su, Xizhou Zhu, Yue Cao, Bin Li 等ICLR 2020 · 被引用 1,825 次
相关 Paper
- Transition-constant Normalization for Image EnhancementJie Huang, Man Zhou, Jinghao Zhang, Gang Yang 等NeurIPS 2023 · 被引用 3 次
- Structure- and Texture-Aware Learning for Low-Light Image EnhancementJinghao Zhang, Jie Huang, Mingde Yao, Man Zhou 等ACM MM 2022 · 被引用 18 次
- SNR-Aware Low-light Image EnhancementXiaogang Xu, Ruixing Wang, Chi-Wing Fu, Jiaya JiaCVPR 2022 · 被引用 552 次
- RT-VENet: A Convolutional Network for Real-time Video EnhancementMohan Zhang, Qiqi Gao, Jinglu Wang, Henrik Turbell 等ACM MM 2020 · 被引用 5 次
- Equivalent Transformation and Dual Stream Network Construction for Mobile Image Super-ResolutionJiahao Chao, Zhou Zhou, Hongfan Gao, Jiali Gong 等CVPR 2023
