How Low Can We Go: Trading Memory for Error in Low-Precision Training
Chengrun Yang, Ziyang Wu, Jerry Chee, Christopher De Sa, Madeleine Udell
摘要
Low-precision arithmetic trains deep learning models using less energy, less memory and less time. However, we pay a price for the savings: lower precision may yield larger round-off error and hence larger prediction error. As applications proliferate, users must choose which precision to use to train a new model, and chip manufacturers must decide which precisions to manufacture. We view these precision choices as a hyperparameter tuning problem, and borrow ideas from meta-learning to learn the tradeoff between memory and error. In this paper, we introduce Pareto Estimation to Pick the Perfect Precision (PEPPP). We use matrix factorization to find non-dominated configurations (the Pareto frontier) with a limited number of network evaluations. For any given memory budget, the precision that minimizes error is a point on this frontier. Practitioners can use the frontier to trade memory for error and choose the best precision for their goals. How should we choose this hyperparameter? Many believe the highest allowable precision (given a memory budget) generally produces the lowest error model. However, there are typically many ways * Equal contribution.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper3
- Ultra-Low Precision 4-bit Training of Deep Neural NetworksXiao Sun, Naigang Wang, Chia-Yu Chen, Jiamin Ni 等NeurIPS 2020 · 被引用 227 次
- AutoML Pipeline Selection: Efficiently Navigating the Combinatorial SpaceChengrun Yang, Jicong Fan, Ziyang Wu, Madeleine UdellKDD 2020 · 被引用 29 次
- Rethinking Differentiable Search for Mixed-Precision Neural NetworksZhaowei Cai, Nuno VasconcelosCVPR 2020
相关 Paper
- CPT: Efficient Deep Neural Network Training via Cyclic PrecisionYonggan Fu, Han Guo, Meng Li, Xin Yang 等ICLR 2021 · 被引用 36 次
- Learning the Pareto Front with HypernetworksAviv Navon, Aviv Shamsian, Ethan Fetaya, Gal ChechikICLR 2021 · 被引用 189 次
- Hypernetwork-based Meta-Learning for Low-Rank Physics-Informed Neural NetworksWoojin Cho, Kookjin Lee, Donsub Rim, Noseong ParkNeurIPS 2023 · 被引用 62 次
- Any-Precision Deep Neural NetworksHaichao Yu, Haoxiang Li, Humphrey Shi, Thomas S. Huang 等AAAI 2021 · 被引用 79 次
- Beyond Uniformity: Sample and Frequency Meta Weighting for Post-Training Quantization of Diffusion ModelsVan Cuong Pham, Anh Hoang, Cuong Nguyen, Trung Le 等ICLR 2026
