Scaling Laws For Deep Learning Based Image Reconstruction
Tobit Klug, Reinhard Heckel
Abstract
Deep neural networks trained end-to-end to map a measurement of a (noisy) image to a clean image perform excellent for a variety of linear inverse problems. Current methods are only trained on a few hundreds or thousands of images as opposed to the millions of examples deep networks are trained on in other domains. In this work, we study whether major performance gains are expected from scaling up the training set size. We consider image denoising, accelerated magnetic resonance imaging, and super-resolution and empirically determine the reconstruction quality as a function of training set size, while simultaneously scaling the network size. For all three tasks we find that an initially steep power-law scaling slows significantly already at moderate training set sizes. Interpolating those scaling laws suggests that even training on millions of images would not significantly improve performance. To understand the expected behavior, we analytically characterize the performance of a linear estimator learned with early stopped gradient descent. The result formalizes the intuition that once the error induced by learning the signal model is small relative to the error floor, more training examples do not improve performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5eb1924b-0e70-48d3-bd8c-df88efee8bdaCited by top-tier papers5
- HUMUS-Net: Hybrid Unrolled Multi-scale Network Architecture for Accelerated MRI ReconstructionZalan Fabian, Berk Tinaz, Mahdi SoltanolkotabiNeurIPS 2022 · 83 citations
- Depthwise Hyperparameter Transfer in Residual Networks: Dynamics and Scaling LimitBlake Bordelon, Lorenzo Noci, Mufan Bill Li, Boris Hanin et al.ICLR 2024 · 54 citations
- Analyzing the Sample Complexity of Self-Supervised Image Reconstruction MethodsTobit Klug, Dogukan Atik, Reinhard HeckelNeurIPS 2023 · 13 citations
- Robustness of Deep Learning for Accelerated MRI: Benefits of Diverse Training DataKang Lin, Reinhard HeckelICML 2024 · 12 citations
- Model Merging Scaling Laws in Large Language ModelsYuanyi Wang, Yanggan Gu, Yiming Zhang, Qi ZHOU et al.ICML 2026 · 8 citations
Builds on14
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- MLP-Mixer: An all-MLP Architecture for VisionIlya O. Tolstikhin, Neil Houlsby, Alexander Kolesnikov, Lucas Beyer et al.NeurIPS 2021 · 3,862 citations
- Scaling Vision TransformersXiaohua Zhai, Alexander Kolesnikov, Neil Houlsby, Lucas BeyerCVPR 2022 · 767 citations
- Accuracy on the Line: on the Strong Correlation Between Out-of-Distribution and In-Distribution GeneralizationJohn Miller, Rohan Taori, Aditi Raghunathan, Shiori Sagawa et al.ICML 2021 · 323 citations
Related papers
- Data augmentation for deep learning based accelerated MRI reconstruction with limited dataZalan Fabian, Reinhard Heckel, Mahdi SoltanolkotabiICML 2021 · 60 citations
- Learning Provably Robust Estimators for Inverse Problems via JitteringAnselm Krainovic, Mahdi Soltanolkotabi, Reinhard HeckelNeurIPS 2023 · 10 citations
- Stochastic Deep Restoration Priors for Imaging Inverse ProblemsYuyang Hu, Albert Peng, Weijie Gan, Peyman Milanfar et al.ICML 2025
- Stochastic Solutions for Linear Inverse Problems using the Prior Implicit in a DenoiserZahra Kadkhodaie, Eero P. SimoncelliNeurIPS 2021 · 202 citations
- Measuring Robustness in Deep Learning Based Compressive SensingMohammad Zalbagi Darestani, Akshay S. Chaudhari, Reinhard HeckelICML 2021 · 94 citations
