Protego: Overload Control for Applications with Unpredictable Lock Contention
Inho Cho, Ahmed Saeed, Seo Jin Park, Mohammad Alizadeh, Adam Belay
Abstract
Modern datacenter applications are concurrent, so they require synchronization to control access to shared data. Requests can contend for different combinations of locks, depending on application and request state. In this paper, we show that locks, especially blocking synchronization, can squander throughput and harm tail latency, even when the CPU is underutilized. Moreover, the presence of a large number of contention points, and the unpredictability in knowing which locks a request will require, make it difficult to prevent contention through overload control using traditional signals such as queueing delay and CPU utilization.
We present Protego, a system that resolves these problems with two key ideas. First, it contributes a new admission control strategy that prevents compute congestion in the presence of lock contention. The key idea is to use marginal improvements in observed throughput, rather than CPU load or latency measurements, within a credit-based admission control algorithm that regulates the rate of incoming requests to a server. Second, it introduces a new latency-aware synchronization abstraction called Active Synchronization Queue Management (ASQM) that allows applications to abort requests if delays exceed latency objectives. We apply Protego to two real-world applications, Lucene and Memcached, and show that it achieves up to 3.3× more goodput and 12.2× lower 99th percentile latency than the state-of-the-art overload control systems while avoiding congestion collapse.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2d065586-d3ca-4349-bd4e-0c3b85a76495Cited by top-tier papers7
- Rajomon: Decentralized and Coordinated Overload Control for Latency-Sensitive MicroservicesJiali Xing, Akis Giannoukos, Paul Loh, Shuyue Wang et al.NSDI 2025 · 12 citations
- TopFull: An Adaptive Top-Down Overload Control for SLO-Oriented MicroservicesJinwoo Park, Jaehyeong Park, Youngmok Jung, Hwijoon Lim et al.SIGCOMM 2024 · 10 citations
- Pushing Performance Isolation Boundaries into Application with pBoxYigong Hu, Gongqi Huang, Peng HuangSOSP 2023 · 2 citations
- Towards Optimal Rack-scale μs-level CPU Scheduling through In-Network Workload ShapingXudong Liao, Han Tian, Xinchen Wan, Chaoliang Zeng et al.USENIX ATC 2025 · 1 citation
- Svalinn: Overload Control in Large-Scale Servers with Multiple Resource BottlenecksBhaskar Subhash Pardeshi, Peidi Song, Ahmed SaeedOSDI 2026
Builds on6
- FIRM: An Intelligent Fine-grained Resource Management Framework for SLO-Oriented MicroservicesHaoran Qiu, Subho S. Banerjee, Saurabh Jha, Zbigniew T. Kalbarczyk et al.OSDI 2020 · 350 citations
- Swift: Delay is Simple and Effective for Congestion Control in the DatacenterGautam Kumar, Nandita Dukkipati, Keon Jang, Hassan M. G. Wassel et al.SIGCOMM 2020 · 333 citations
- Autopilot: workload autoscaling at GoogleKrzysztof Rzadca, Pawel Findeisen, Jacek Swiderski, Przemyslaw Zych et al.EuroSys 2020 · 299 citations
- 1RMA: Re-envisioning Remote Memory Access for Multi-tenant DatacentersArjun Singhvi, Aditya Akella, Dan Gibson, Thomas F. Wenisch et al.SIGCOMM 2020 · 70 citations
- Overload Control for µs-scale RPCs with BreakwaterInho Cho, Ahmed Saeed, Joshua Fried, Seo Jin Park et al.OSDI 2020 · 61 citations
Related papers
- Mitigating Application Resource Overload with Targeted Task CancellationYigong Hu, Zeyin Zhang, Yicheng Liu, Yile Gu et al.SOSP 2025
- Achieving Microsecond-Scale Tail Latency Efficiently with Approximate Optimal SchedulingRishabh R. Iyer, Musa Unal, Marios Kogias, George CandeaSOSP 2023 · 17 citations
- Aequitas: admission control for performance-critical RPCs in datacentersYiwen Zhang, Gautam Kumar, Nandita Dukkipati, Xian Wu et al.SIGCOMM 2022 · 30 citations
- FlexGuard: Fast Mutual Exclusion Independent of SubscriptionVictor Laforet, Sanidhya Kashyap, Calin Iorgulescu, Julia Lawall et al.SOSP 2025
- Avoiding scheduler subversion using scheduler-cooperative locksYuvraj Patel, Leon Yang, Leo Prasath Arulraj, Andrea C. Arpaci-Dusseau et al.EuroSys 2020 · 6 citations
