A RISC-V in-network accelerator for flexible high-performance low-power packet processing
Salvatore Di Girolamo, Andreas Kurth, Alexandru Calotoiu, Thomas Benz, Timo Schneider, Jakub Beránek, Luca Benini, Torsten Hoefler
Abstract
The capacity of offloading data and control tasks to the network is becoming increasingly important, especially if we consider the faster growth of network speed when compared to CPU frequencies. In-network compute alleviates the host CPU load by running tasks directly in the network, enabling additional computation/communication overlap and potentially improving overall application performance. However, sustaining bandwidths provided by next-generation networks, e.g., 400 Gbit/s, can become a challenge. sPIN is a programming model for in-NIC compute, where users specify handler functions that are executed on the NIC, for each incoming packet belonging to a given message or flow. It enables a CUDA-like acceleration, where the NIC is equipped with lightweight processing elements that process network packets in parallel. We investigate the architectural specialties that a sPIN NIC should provide to enable high-performance, low-power, and flexible packet processing. We introduce PsPIN, a first open-source sPIN implementation, based on a multi-cluster RISC-V architecture and designed according to the identified architectural specialties. We investigate the performance of PsPIN with cycle-accurate simulations, showing that it can process packets at 400 Gbit/s for several use cases, introducing minimal latencies (26 ns for 64 B packets) and occupying a total area of 18.5 mm 2 (22 nm FDSOI).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b5a7330a-810b-43d5-9ccd-ae3a18d95096Cited by top-tier papers7
- Flare: flexible in-network allreduceDaniele De Sensi, Salvatore Di Girolamo, Saleh Ashkboos, Shigang Li et al.SC 2021 · 49 citations
- OSMOSIS: Enabling Multi-Tenancy in Datacenter SmartNICsMikhail Khalilov, Marcin Chrapek, Siyuan Shen, Alessandro Vezzu et al.USENIX ATC 2024 · 14 citations
- Turbo: SmartNIC-enabled Dynamic Load Balancing of µs-scale RPCsHamed Seyedroudbari, Srikar Vanavasam, Alexandros DaglisHPCA 2023 · 12 citations
- Building Blocks for Network-Accelerated Distributed File SystemsSalvatore Di Girolamo, Daniele De Sensi, Konstantin Taranov, Milos Malesevic et al.SC 2022 · 7 citations
- HADES: Hardware-Assisted Distributed Transactions in the Age of Fast Networks and SmartNICsApostolos Kokolis, Antonis Psistakis, Benjamin Reidys, Jian Huang et al.ISCA 2024 · 3 citations
Builds on4
- An in-depth analysis of the slingshot interconnectDaniele De Sensi, Salvatore Di Girolamo, Kim H. McMahon, Duncan Roweth et al.SC 2020 · 122 citations
- PANIC: A High-Performance Programmable NIC for Multi-tenant NetworksJiaxin Lin, Kiran Patel, Brent E. Stephens, Anirudh Sivaraman et al.OSDI 2020 · 104 citations
- StRoM: smart remote memoryDavid Sidler, Zeke Wang, Monica Chiosa, Amit Kulkarni et al.EuroSys 2020 · 83 citations
- hXDP: Efficient Software Packet Processing on FPGA NICsMarco Spaziani Brunella, Giacomo Belocchi, Marco Bonola, Salvatore Pontarelli et al.OSDI 2020 · 25 citations
Related papers
- The nanoPU: A Nanosecond Network Stack for DatacentersStephen Ibanez, Alex Mallery, Serhat Arslan, Theo Jepsen et al.OSDI 2021 · 74 citations
- NetCL: A Unified Programming Framework for In-Network ComputingGeorge Karlos, Henri E. Bal, Lin WangSC 2024 · 3 citations
- FpgaNIC: An FPGA-based Versatile 100Gb SmartNIC for GPUsZeke Wang, Hongjing Huang, Jie Zhang, Fei Wu et al.USENIX ATC 2022 · 58 citations
- Fast, Scalable, and Accurate Rate Limiter for RDMA NICsZilong Wang, Xinchen Wan, Luyang Li, Yijun Sun et al.SIGCOMM 2024 · 17 citations
- How to Hardware Accelerate Your 5G CUXin Zhe Khooi, Satis Kumar Permal, Cha Hwan Song, Nishant Budhdev et al.INFOCOM 2026
