Miresga: Accelerating Layer-7 Load Balancing with Programmable Switches
Xiaoyi Shi, Lin He, Jiasheng Zhou, Yifan Yang, Ying Liu
Abstract
As online cloud services expand rapidly, layer-7 load balancing has become indispensable for maintaining service availability and performance. The emergence of programmable switches with both high performance and a certain degree of flexibility has made it possible to apply programmable switches to load balancing. Nevertheless, the limited memory capacity and the relatively sluggish speed of table entry insertion and deletion of programmable switches have severely constrained their performance. To this end, we introduce Miresga, a hybrid and high-performance layer-7 load balancing system by co-designing hardware and software. The core idea of Miresga is to maximize the utilization of hardware and software resources by rationally partitioning the layer-7 load balancing task, thereby improving performance. To achieve this, Miresga offloads the elephant flows, which account for the majority of traffic, to programmable switches that excel at packet processing, and Miresga utilizes general-purpose servers with stronger computational capabilities to parse application layer protocols and apply load balancing rules. To alleviate memory pressure on the programmable switch, Miresga employs a back-end agent to handle memory-intensive tasks, working in conjunction with the programmable switch to complete the offloaded tasks. This design leverages the performance advantages of the programmable switch while avoiding bottlenecks caused by its limited memory and table insertion speed. We implement the Miresga prototype with a 3.2 Tbps Intel Tofino switch and general-purpose servers. The evaluation results show that Miresga achieves 3.9× throughput and 0.4× latency compared to software load balancing solutions. Compared to the state-of-the-art design employing programmable switches, Miresga achieves almost the same throughput and latency for delivering large objects and 5.0× throughput and 0.2× latency when transmitting small objects.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 561f8f79-1e3f-47be-bfbc-a519ad19248dCited by top-tier papers2
- Hermes: Enhancing Layer-7 Cloud Load Balancers with Userspace-Directed I/O Event NotificationTian Pan, Enge Song, Yueshang Zuo, Shaokai Zhang et al.SIGCOMM 2025 · 5 citations
- Remote TCP Connection Offload and ApplicationsShuo Li, Steven W. D. Chien, Tianyi Gao, Michio HondaNSDI 2026 · 3 citations
Builds on14
- ATP: In-network Aggregation for Multi-tenant LearningChonLam Lao, Yanfang Le, Kshiteej Mahajan, Yixi Chen et al.NSDI 2021 · 359 citations
- AccelTCP: Accelerating Network Applications with Stateful TCP OffloadingYoungGyoun Moon, SeungEon Lee, Muhammad Asim Jamshed, KyoungSoo ParkNSDI 2020 · 121 citations
- PowerTCP: Pushing the Performance Limits of Datacenter NetworksVamsi Addanki, Oliver Michel, Stefan SchmidNSDI 2022 · 116 citations
- Sailfish: accelerating cloud-scale multi-tenant multi-service gateways with programmable switchesTian Pan, Nianbing Yu, Chenhao Jia, Jianwen Pi et al.SIGCOMM 2021 · 111 citations
- Tiara: A Scalable and Efficient Hardware Acceleration Architecture for Stateful Layer-4 Load BalancingChaoliang Zeng, Layong Luo, Teng Zhang, Zilong Wang et al.NSDI 2022 · 97 citations
Related papers
- Capybara: Dynamic Load Balancing with Microsecond-Scale TCP MigrationInho Choi, Nimish Wadekar, Guangda Sun, Raj Joshi et al.SIGCOMM 2026
- QDSR: Accelerating Layer-7 Load Balancing by Direct Server Return with QUICZiqi Wei, Zhiqiang Wang, Qing Li, Yuan Yang et al.USENIX ATC 2024 · 10 citations
- Network Load Balancing with In-network Reordering Support for RDMACha Hwan Song, Xin Zhe Khooi, Raj Joshi, Inho Choi et al.SIGCOMM 2023 · 110 citations
- Pegasus: Tolerating Skewed Workloads in Distributed Storage with In-Network Coherence DirectoriesJialin Li, Jacob Nelson, Ellis Michael, Xin Jin et al.OSDI 2020 · 96 citations
- TEA: Enabling State-Intensive Network Functions on Programmable SwitchesDaehyeok Kim, Zaoxing Liu, Yibo Zhu, Changhoon Kim et al.SIGCOMM 2020 · 121 citations
