Miresga: Accelerating Layer-7 Load Balancing with Programmable Switches
Xiaoyi Shi, Lin He, Jiasheng Zhou, Yifan Yang, Ying Liu
摘要
As online cloud services expand rapidly, layer-7 load balancing has become indispensable for maintaining service availability and performance. The emergence of programmable switches with both high performance and a certain degree of flexibility has made it possible to apply programmable switches to load balancing. Nevertheless, the limited memory capacity and the relatively sluggish speed of table entry insertion and deletion of programmable switches have severely constrained their performance. To this end, we introduce Miresga, a hybrid and high-performance layer-7 load balancing system by co-designing hardware and software. The core idea of Miresga is to maximize the utilization of hardware and software resources by rationally partitioning the layer-7 load balancing task, thereby improving performance. To achieve this, Miresga offloads the elephant flows, which account for the majority of traffic, to programmable switches that excel at packet processing, and Miresga utilizes general-purpose servers with stronger computational capabilities to parse application layer protocols and apply load balancing rules. To alleviate memory pressure on the programmable switch, Miresga employs a back-end agent to handle memory-intensive tasks, working in conjunction with the programmable switch to complete the offloaded tasks. This design leverages the performance advantages of the programmable switch while avoiding bottlenecks caused by its limited memory and table insertion speed. We implement the Miresga prototype with a 3.2 Tbps Intel Tofino switch and general-purpose servers. The evaluation results show that Miresga achieves 3.9× throughput and 0.4× latency compared to software load balancing solutions. Compared to the state-of-the-art design employing programmable switches, Miresga achieves almost the same throughput and latency for delivering large objects and 5.0× throughput and 0.2× latency when transmitting small objects.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Hermes: Enhancing Layer-7 Cloud Load Balancers with Userspace-Directed I/O Event NotificationTian Pan, Enge Song, Yueshang Zuo, Shaokai Zhang 等SIGCOMM 2025 · 被引用 5 次
- Remote TCP Connection Offload and ApplicationsShuo Li, Steven W. D. Chien, Tianyi Gao, Michio HondaNSDI 2026 · 被引用 3 次
它引用的顶会 Paper14
- ATP: In-network Aggregation for Multi-tenant LearningChonLam Lao, Yanfang Le, Kshiteej Mahajan, Yixi Chen 等NSDI 2021 · 被引用 359 次
- AccelTCP: Accelerating Network Applications with Stateful TCP OffloadingYoungGyoun Moon, SeungEon Lee, Muhammad Asim Jamshed, KyoungSoo ParkNSDI 2020 · 被引用 121 次
- PowerTCP: Pushing the Performance Limits of Datacenter NetworksVamsi Addanki, Oliver Michel, Stefan SchmidNSDI 2022 · 被引用 116 次
- Sailfish: accelerating cloud-scale multi-tenant multi-service gateways with programmable switchesTian Pan, Nianbing Yu, Chenhao Jia, Jianwen Pi 等SIGCOMM 2021 · 被引用 111 次
- Tiara: A Scalable and Efficient Hardware Acceleration Architecture for Stateful Layer-4 Load BalancingChaoliang Zeng, Layong Luo, Teng Zhang, Zilong Wang 等NSDI 2022 · 被引用 97 次
相关 Paper
- Capybara: Dynamic Load Balancing with Microsecond-Scale TCP MigrationInho Choi, Nimish Wadekar, Guangda Sun, Raj Joshi 等SIGCOMM 2026
- QDSR: Accelerating Layer-7 Load Balancing by Direct Server Return with QUICZiqi Wei, Zhiqiang Wang, Qing Li, Yuan Yang 等USENIX ATC 2024 · 被引用 10 次
- Network Load Balancing with In-network Reordering Support for RDMACha Hwan Song, Xin Zhe Khooi, Raj Joshi, Inho Choi 等SIGCOMM 2023 · 被引用 110 次
- Pegasus: Tolerating Skewed Workloads in Distributed Storage with In-Network Coherence DirectoriesJialin Li, Jacob Nelson, Ellis Michael, Xin Jin 等OSDI 2020 · 被引用 96 次
- TEA: Enabling State-Intensive Network Functions on Programmable SwitchesDaehyeok Kim, Zaoxing Liu, Yibo Zhu, Changhoon Kim 等SIGCOMM 2020 · 被引用 121 次
