A High-Precision and Low-Cost Approximate Transform Accelerator for Video Coding
Zhijian Hao, Jiaming Liu, Chenlong He, Qi Zheng, Shushi Chen, Jinchang Xu, Yue Hao, Xiao Yan, Xiaohua Ma
Abstract
The introduction of multiple transform types in the Versatile Video Coding (VVC) standard has yielded notable encoding gains but also imposed considerable computational burdens. Existing transform circuits of different types are typically implemented separately due to their independence, leading to substantial hardware overhead. To address this, we explore the relationship between Discrete Cosine Transform Type-2 (DCT2) and Discrete Sine Transform Type-7 (DST7) matrices and reveal a prominent diagonal aggregation phenomenon in their transfer matrix. Based on this insight, the least-squares method is applied to optimize the transfer matrix sparsity, achieving a high-precision, low-cost approximate conversion from DCT2 to DST7. Furthermore, we optimize DCT2 computation by proposing an elaborate matrix decomposition approach that allows a lightweight shift-adder unit to efficiently generate all required product terms across varying sizes. Leveraging these algorithmic optimizations, we implement a highly reusable and area-efficient approximate transform accelerator that supports sizes from 4 to 32 points and accommodates three types in VVC. Experimental results demonstrate that the proposed accelerator achieves over 44% reduction in circuit resource consumption with negligible BD-BR performance loss of just , maintaining processing capabilities up to .
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 2afd39a9-e7d4-4ba4-8ee9-68edb67aed9fRelated papers
- A Hardware-efficient Unified Motion Estimation for Video CodingXizhong Zhu, Guoqing Xiang, Peng Zhang, Huizhu Jia et al.ACM MM 2023 · 2 citations
- Two-Stage Octave Residual Network for End-to-End Image CompressionFangdong Chen, Yumeng Xu, Li WangAAAI 2022 · 47 citations
- Learned Bi-Resolution Image Coding using Generalized Octave ConvolutionsMohammad Akbari, Jie Liang, Jingning Han, Chengjie TuAAAI 2021 · 21 citations
- InterArch: Video Transformer Acceleration via Inter-Feature Deduplication with Cube-based DataflowXuhang Wang, Zhuoran Song, Xiaoyao LiangDAC 2024 · 2 citations
- Enhanced Invertible Encoding for Learned Image CompressionYueqi Xie, Ka Leong Cheng, Qifeng ChenACM MM 2021 · 195 citations
