Quark: Implementing Convolutional Neural Networks Entirely on Programmable Data Plane
Mai Zhang, Lin Cui, Xiaoquan Zhang, Fung Po Tso, Zhen Zhang, Yuhui Deng, Zhetao Li
Abstract
The rapid development of programmable network devices and the widespread use of machine learning (ML) in networking have facilitated efficient research into intelligent data plane (IDP). Offloading ML to programmable data plane (PDP) enables quick analysis and responses to network traffic dynamics, and efficient management of network links. However, PDP hardware pipeline has significant resource limitations. For instance, Intel Tofino ASIC has only 10Mb SRAM in each stage, and lacks support for multiplication, division and floating-point operations. These constraints significantly hinder the development of IDP. This paper presents Quark, a framework that fully offloads convolutional neural network (CNN) inference onto PDP. Quark employs model pruning to simplify the CNN model, and uses quantization to support floating-point operations. Additionally, Quark divides the CNN into smaller units to improve resource utilization on the PDP. We have implemented a testbed prototype of Quark on both P4 hardware switch (Intel Tofino ASIC) and software switch (i.e., BMv2). Extensive evaluation results demonstrate that Quark achieves 97.3% accuracy in anomaly detection task while using only 22.7% of the SRAM resources on the Intel Tofino ASIC switch, completing inference tasks at line rate with an average latency of 42.66µs.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 031d04ac-3fe4-49f3-a388-927de3521f92Cited by top-tier papers1
Ask how each one uses itBuilds on5
- Re-architecting Traffic Analysis with Neural Network Interface CardsGiuseppe Siracusano, Salvator Galea, Davide Sanvito, Mohammad Malekzadeh et al.NSDI 2022 · 99 citations
- Taurus: a data plane architecture for per-packet MLTushar Swamy, Alexander Rucker, Muhammad Shahbaz, Ishan Gaur et al.ASPLOS 2022 · 94 citations
- Programmable Switches for in-Networking ClassificationBruno Missi Xavier, Rafael Silva Guimarães, Giovanni Comarela, Magnos MartinelloINFOCOM 2021 · 79 citations
- Brain-on-Switch: Towards Advanced Intelligent Network Data Plane via NN-Driven Traffic Analysis at Line-SpeedJinzhu Yan, Haotian Xu, Zhuotao Liu, Qi Li et al.NSDI 2024 · 60 citations
- FlowLens: Enabling Efficient Flow Classification for ML-based Network Security ApplicationsDiogo Barradas, Nuno Santos, Luís Rodrigues, Salvatore Signorello et al.NDSS 2021
Related papers
- Flowrest: Practical Flow-Level Inference in Programmable Switches with Random ForestsAristide Tanyi-Jong Akem, Michele Gucciardo, Marco FioreINFOCOM 2023 · 63 citations
- Monic: In-Network Mixture-of-Experts Inference on Programmable Data PlanesXiaoquan Zhang, Bowen Liang, Fung Po Tso, Yuhui Deng et al.INFOCOM 2026
- FENIX: Enabling In-Network DNN Inference with FPGA-Enhanced Programmable SwitchesXiangyu Gao, Tong Li, Yinchao Zhang, Ziqiang Wang et al.NSDI 2026 · 12 citations
- Pegasus: A Universal Framework for Scalable Deep Learning Inference on the DataplaneYinchao Zhang, Su Yao, Yong Feng, Kang Chen et al.SIGCOMM 2025 · 10 citations
- Unlocking the Power of Inline Floating-Point Operations on Programmable SwitchesYifan Yuan, Omar Alama, Jiawei Fei, Jacob Nelson et al.NSDI 2022 · 33 citations
