Achieving on-Mobile Real-Time Super-Resolution with Neural Architecture and Pruning Search
Zheng Zhan, Yifan Gong, Pu Zhao, Geng Yuan, Wei Niu, Yushu Wu, Tianyun Zhang, Malith Jayaweera, David R. Kaeli, Bin Ren, Xue Lin, Yanzhi Wang
摘要
Though recent years have witnessed remarkable progress in single image super-resolution (SISR) tasks with the prosperous development of deep neural networks (DNNs), the deep learning methods are confronted with the computation and memory consumption issues in practice, especially for resource-limited platforms such as mobile devices. To overcome the challenge and facilitate the real-time deployment of SISR tasks on mobile, we combine neural architecture search with pruning search and propose an automatic search framework that derives sparse superresolution (SR) models with high image quality while satisfying the real-time inference requirement. To decrease the search cost, we leverage the weight sharing strategy by introducing a supernet and decouple the search problem into three stages, including supernet construction, compiler-aware architecture and pruning search, and compiler-aware pruning ratio search. With the proposed framework, we are the first to achieve real-time SR inference (with only tens of milliseconds per frame) for implementing 720p resolution with competitive image quality (in terms of PSNR and SSIM) on mobile platforms (Samsung Galaxy S20). * Equal contribution * We believe targeting sub 100ms can be reasonably called real-time [49] and we require the real-time implementation to be faster than 50ms.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- SparCL: Sparse Continual Learning on the EdgeZifeng Wang, Zheng Zhan, Yifan Gong, Geng Yuan 等NeurIPS 2022 · 被引用 97 次
- 4KAgent: Agentic Any Image to 4K Super-ResolutionYushen Zuo, Qi Zheng, Mingyang Wu, Xinrui Jiang 等NeurIPS 2025 · 被引用 51 次
- LazyDiT: Lazy Learning for the Acceleration of Diffusion TransformersXuan Shen, Zhao Song, Yufa Zhou, Bo Chen 等AAAI 2025 · 被引用 40 次
- Exploring Token Pruning in Vision State Space ModelsZheng Zhan, Zhenglun Kong, Yifan Gong, Yushu Wu 等NeurIPS 2024 · 被引用 34 次
- Search for Efficient Large Language ModelsXuan Shen, Pu Zhao, Yifan Gong, Zhenglun Kong 等NeurIPS 2024 · 被引用 23 次
它引用的顶会 Paper6
- PatDNN: Achieving Real-Time DNN Execution on Mobile Devices with Pattern-based Weight PruningWei Niu, Xiaolong Ma, Sheng Lin, Shihao Wang 等ASPLOS 2020 · 被引用 214 次
- AutoCompress: An Automatic DNN Structured Pruning Framework for Ultra-High Compression RatesNing Liu, Xiaolong Ma, Zhiyuan Xu, Yanzhi Wang 等AAAI 2020 · 被引用 204 次
- PCONV: The Missing but Desirable Sparsity in DNN Weight Pruning for Real-Time Execution on Mobile DevicesXiaolong Ma, Fu-Ming Guo, Wei Niu, Xue Lin 等AAAI 2020 · 被引用 201 次
- Efficient Residual Dense Block Search for Image Super-ResolutionDehua Song, Chang Xu, Xu Jia, Yiyi Chen 等AAAI 2020 · 被引用 145 次
- RTMobile: Beyond Real-Time Mobile Acceleration of RNNs for Speech RecognitionPeiyan Dong, Siyue Wang, Wei Niu, Chengming Zhang 等DAC 2020 · 被引用 50 次
相关 Paper
- Searching Lightweight Neural Network for Image Signal ProcessingHaojia Lin, Lijiang Li, Xiawu Zheng, Fei Chao 等ACM MM 2022 · 被引用 2 次
- MobileIE: An Extremely Lightweight and Effective ConvNet for Real-Time Image Enhancement on Mobile DevicesHailong Yan, Ao Li, Xiangtao Zhang, Zhe Liu 等ICCV 2025 · 被引用 12 次
- Achieving Lightweight Super-Resolution for Real-Time Computer GraphicsYu Wen, Chen Zhang, Chenhao Xie, Xin FuAAAI 2025 · 被引用 1 次
- DySR: Adaptive Super-Resolution via Algorithm and System Co-designSyed Zawad, Cheng Li, Zhewei Yao, Elton Zheng 等ICLR 2023
- Neural Pruning Search for Real-Time Object Detection of Autonomous VehiclesPu Zhao, Geng Yuan, Yuxuan Cai, Wei Niu 等DAC 2021 · 被引用 23 次
