Mitigating CPU Frontend for Complex Data Plane Applications
Yihan Dang, Hao Li, Ze Xia, Jiajun Luan, Peng Zhang
Abstract
The functionality and requirement of modern networks are becoming increasingly complex, giving rise to Complex Data Plane Applications (CDPA) with rich semantics but often limited performance. However, many existing optimizations fail to improve the performance of CDPAs. This is because CDPAs usually come with excessively large code size, which is often two orders of magnitude larger than today's L1 instruction cache (I-cache) size, causing the CPU to frequently stall on accessing instructions, thus presenting a distinct performance profile that bounds on CPU frontend. This paper proposes NanoPL, an I-cache-friendly new execution model that shuffles the packet processing logic to efficiently mitigate CPU frontend for CDPAs. Stemming from the common execution pattern of CDPAs, NanoPL analyzes its code to ensure semantic consistency after shuffling. By collecting performance profile of CDPAs over underlying traffic, NanoPL partitions CDPAs into execution stages and conducts I-cachefriendly shuffling policy. Experiments show that NanoPL can achieve 17.2% 30.2% higher throughput over real world CD-PAs due to the reduction of I-cache misses by up to 86.4%.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on11
- Propeller: A Profile Guided, Relinking Optimizer for Warehouse-Scale ApplicationsHan Shen, Krzysztof Pszeniczny, Rahman Lavaee, Snehasish Kumar et al.ASPLOS 2023 · 37 citations
- Twig: Profile-Guided BTB Prefetching for Data Center ApplicationsTanvir Ahmed Khan, Nathan Brown, Akshitha Sriraman, Niranjan K. Soundararajan et al.MICRO 2021 · 33 citations
- Batchy: Batch-scheduling Data Flow Graphs with Service-level ObjectivesTamás Lévai, Felicián Németh, Barath Raghavan, Gábor RétváriNSDI 2020 · 30 citations
- Domain specific run time optimization for software data planesSebastiano Miano, Alireza Sanaee, Fulvio Risso, Gábor Rétvári et al.ASPLOS 2022 · 20 citations
- Compacting points-to sets through object clusteringMohamad Barbar, Yulei SuiOOPSLA 2021 · 15 citations
Related papers
- hXDP: Efficient Software Packet Processing on FPGA NICsMarco Spaziani Brunella, Giacomo Belocchi, Marco Bonola, Salvatore Pontarelli et al.OSDI 2020 · 25 citations
- PacketMill: toward per-Core 100-Gbps networkingAlireza Farshin, Tom Barbette, Amir Roozbeh, Gerald Q. Maguire Jr. et al.ASPLOS 2021 · 46 citations
- NetCL: A Unified Programming Framework for In-Network ComputingGeorge Karlos, Henri E. Bal, Lin WangSC 2024 · 3 citations
- P4LRU: Towards An LRU Cache Entirely in Programmable Data PlaneYikai Zhao, Wenrui Liu, Fenghao Dong, Tong Yang et al.SIGCOMM 2023 · 18 citations
- Hoda: a High-performance Open vSwitch Dataplane with Multiple Specialized Data PathsHeng Pan, Peng He, Zhenyu Li, Pan Zhang et al.EuroSys 2024 · 6 citations
