Improving Energy Saving of One-Sided Matrix Decompositions on CPU-GPU Heterogeneous Systems
Jieyang Chen, Xin Liang, Kai Zhao, Hadi Zamani Sabzi, Laxmi N. Bhuyan, Zizhong Chen
摘要
One-sided dense matrix decompositions (e.g., Cholesky, LU, and QR) are the key components in scientific computing in many different fields. Although their design has been highly optimized for modern processors, they still consume a considerable amount of energy. As CPU-GPU heterogeneous systems are commonly used for matrix decompositions, in this work, we aim to further improve the energy saving of onesided matrix decompositions on CPU-GPU heterogeneous systems. We first build an Algorithm-Based Fault Tolerance protected overclocking technique (ABFT-OC) to enable us to exploit reliable overclocking for key matrix decomposition operations. Then, we design an energy-saving matrix decomposition framework, Bi-directional Slack Reclamation (BSR), that can intelligently combine the capability provided by ABFT-OC and DVFS to maximize energy saving and maintain performance and reliability. Experiments show that BSR is able to save up to 11.7% more energy compared with the current best energy saving optimization approach with no performance degradation and up to 14.1% 𝐸𝑛𝑒𝑟𝑔𝑦×𝐷𝑒𝑙𝑎𝑦 2 reduction. Also, BSR enables the Pareto efficient performanceenergy trade-off, which is able to provide up to 1.43× performance improvement without costing extra energy.
• Hardware → Power and energy; • Computer systems organization → Dependable and fault-tolerant systems and networks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- Addressing Irregular Patterns of Matrix Computations on GPUs and Their Impact on Applications Powered by Sparse Direct SolversAhmad Abdelfattah, Pieter Ghysels, Wajih Boukaram, Stanimire Tomov 等SC 2022 · 被引用 4 次
- Towards Energy-Efficient Real-Time Scheduling of Heterogeneous Multi-GPU SystemsYidi Wang, Mohsen Karimi, Hyoseung KimRTSS 2022 · 被引用 8 次
- HeteroSVD: Efficient SVD Accelerator on Versal ACAP with Algorithm-Hardware Co-DesignXinya Luan, Zhe Lin, Kai Shi, Jianwang Zhai 等DAC 2025 · 被引用 1 次
- Minimizing Power Waste in Heterogenous Computing via Adaptive Uncore ScalingZhong Zheng, Seyfal Sultanov, Michael E. Papka, Zhiling LanSC 2025 · 被引用 2 次
- HP-MDR: High-performance and Portable Data Refactoring and Progressive Retrieval with Advanced GPUsYanliang Li, Wenbo Li, Qian Gong, Qing Liu 等SC 2025 · 被引用 2 次
