ShRing: Networking with Shared Receive Rings
Boris Pismenny, Adam Morrison, Dan Tsafrir
摘要
Multicore systems parallelize to accommodate incoming Ethernet traffic, allocating one receive (Rx) ring with ≥1Ki entries per core by default. This ring size is sufficient to absorb packet bursts of single-core workloads. But the combined size of all Rx buffers (pointed to by all Rx rings) can exceed the size of the last-level cache. We observe that, in this case, NIC and CPU memory accesses are increasingly served by main memory, which might incur nonnegligible overheads when scaling to hundreds of incoming gigabits per second.
To alleviate this problem, we propose "shRing," which shares each Rx ring among several cores when networking memory bandwidth consumption is high. ShRing thus adds software synchronization costs, but this overhead is offset by the smaller memory footprint. We show that, consequently, shRing increases the throughput of NFV workloads by up to 1.27x, and that it reduces their latency by up to 38x. The substantial latency reduction occurs when shRing shortens the per-packet processing time to a value smaller than the packet interarrival time, thereby preventing overload conditions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Making Kernel Bypass Practical for the Cloud with JunctionJoshua Fried, Gohar Irfan Chaudhry, Enrique Saurez, Esha Choukse 等NSDI 2024 · 被引用 57 次
- CEIO: A Cache-Efficient Network I/O Architecture for NIC-CPU Data PathsBowen Liu, Xinyang Huang, Qijing Li, Zhuobin Huang 等SIGCOMM 2025 · 被引用 10 次
- A4: Microarchitecture-Aware LLC Management for Datacenter Servers with Emerging I/O DevicesHaneul Park, Jiaqi Lou, Sangjin Lee, Yifan Yuan 等ISCA 2025 · 被引用 2 次
- Disentangling the Dual Role of NIC Receive RingsBoris Pismenny, Adam Morrison, Dan TsafrirOSDI 2025 · 被引用 2 次
- Achieving Wire-Latency Storage Systems by Exploiting Hardware ACKsQing Wang, Jiwu Shu, Jing Wang, Yuhao ZhangNSDI 2025 · 被引用 1 次
它引用的顶会 Paper13
- Caladan: Mitigating Interference at Microsecond TimescalesJoshua Fried, Zhenyuan Ruan, Amy Ousterhout, Adam BelayOSDI 2020 · 被引用 213 次
- Understanding host network stack overheadsQizhe Cai, Shubham Chaudhary, Midhul Vuppalapati, Jaehyun Hwang 等SIGCOMM 2021 · 被引用 150 次
- Contention-Aware Performance Prediction For Virtualized Network FunctionsAntonis Manousis, Rahul Anand Sharma, Vyas Sekar, Justine SherrySIGCOMM 2020 · 被引用 57 次
- Don't Forget the I/O When Allocating Your LLCYifan Yuan, Mohammad Alian, Yipeng Wang, Ren Wang 等ISCA 2021 · 被引用 37 次
- Autonomous NIC offloadsBoris Pismenny, Haggai Eran, Aviad Yehezkel, Liran Liss 等ASPLOS 2021 · 被引用 32 次
相关 Paper
- The benefits of general-purpose on-NIC memoryBoris Pismenny, Liran Liss, Adam Morrison, Dan TsafrirASPLOS 2022 · 被引用 28 次
- State-Compute Replication: Parallelizing High-Speed Stateful Packet ProcessingQiongwen Xu, Sebastiano Miano, Xiangyu Gao, Tao Wang 等NSDI 2025
- TiNA: Tiered Network Buffer Architecture for Fast Networking in Chiplet-based CPUsSiddharth Agarwal, Tianchen Wang, Jinghan Huang, Saksham Agarwal 等ASPLOS 2026 · 被引用 1 次
- RECANS: Low-Latency Network Function Chains with Hierarchical State SharingJian Zhao, Shujun Zhuang, Jian Li, Haibing GuanHPDC 2020 · 被引用 2 次
- PeRF: Preemption-enabled RDMA FrameworkSugi Lee, Mingyu Choi, Ikjun Yeom, Younghoon KimUSENIX ATC 2024 · 被引用 3 次
