An Efficient Deep Learning Accelerator for Compressed Video Analysis
Yongchen Wang, Ying Wang, Huawei Li, Yinhe Han, Xiaowei Li
摘要
Previous neural network accelerators tailored to video analysis only accept data of RGB/YUV domain, requiring decompressing the video that are often compressed before transmitted from the edge sensors. A compressed video processing accelerator can remove the decoding overhead, and gain performance speedup by operating on more compact input data. This work proposes a novel deep learning accelerator architecture, Alchemist, which predicts results directly from the compressed video bitstream instead of reconstructing the full RGB images. By utilizing the metadata of motion vector and critical blocks extracted from bitstream, Alchemist contributes to remarkable performance speedup of 5x with negligible accuracy loss.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- LeCA: In-Sensor Learned Compressive Acquisition for Efficient Machine Vision on the EdgeTianrui Ma, Adith Jagadish Boloor, Xiangxing Yang, Weidong Cao 等ISCA 2023 · 被引用 27 次
- A Bit is All You Need! Efficient Video Capture via Single Bit ImagingKanchana Vaishnavi Gandikota, Michael Moeller, Andreas Kolb, Bhaskar Choubey 等CVPR 2026
- AdaStreamer: Machine-Centric High-Accuracy Multi-Video Analytics with Adaptive Neural CodecsAndong Zhu, Sheng Zhang, Ke Cheng, Xiaohang Shi 等INFOCOM 2024 · 被引用 8 次
- Co-Via: A Video Frame Interpolation Accelerator Exploiting Codec Information ReuseHaishuang Fan, Qichu Sun, Jingya Wu, Wenyan Lu 等DAC 2024 · 被引用 3 次
- Accurate and Fast Compressed Video CaptioningYaojie Shen, Xin Gu, Kai Xu, Heng Fan 等ICCV 2023 · 被引用 53 次
