Superways: A Datacenter Topology for Incast-heavy workloads
Hamed Rezaei, Balajee Vamanan
Abstract
Several important datacenter applications cause incast congestion, which severely degrades flow completion times of short flows and throughput of long flows. Further, because most flows are short and the incast duration is shorter than typical round-trip times, reactive mechanisms that rely on congestion control are not effective. While modern datacenter topologies provide high bisection bandwidth to support all-to-all traffic, incast is fundamentally a many-to-one traffic pattern, and therefore, requires deep buffers or high bandwidth at the network edge. We propose Superways, a heterogeneous datacenter topology that provides higher bandwidth for some servers to absorb incasts, as incasts occur only at a small number of servers that aggregate responses from other senders. Our design is based on the key observation that a small subset of servers which aggregate responses are likely to be network bound, whereas most other servers that communicate only with random servers are not. Superways can be implemented over many of the existing datacenter topologies and can be expanded flexibly without incurring high cost and cabling complexity. We also provide a heuristic for scheduling jobs in our topology to fully utilize the extra capacity. Using a real CloudLab implementation and using ns-3 simulations, we show that Superways significantly improves flow completion times and throughput over existing datacenter topologies. We also analyze cost and cabling complexity, and discuss how to expand our topology.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d92200fc-25c6-4eb9-b152-ee484c4ac6a1Cited by top-tier papers1
Ask how each one uses itRelated papers
- RateMP: Optimizing Bandwidth Utilization with High Burst Tolerance in Data Center NetworksJiangping Han, Kaiping Xue, Wentao Wang, Ruidong Li et al.INFOCOM 2024 · 8 citations
- Cutting Tail Latency in Commodity Datacenters with CloudburstGaoxiong Zeng, Li Chen, Bairen Yi, Kai ChenINFOCOM 2022 · 12 citations
- Preventing Network Bottlenecks: Accelerating Datacenter Services with Hotspot-Aware Placement for Compute and StorageHamid Hajabdolali Bazzaz, Yingjie Bi, Weiwu Pang, Minlan Yu et al.NSDI 2025 · 3 citations
- NegotiaToR: Towards A Simple Yet Effective On-demand Reconfigurable Datacenter NetworkCong Liang, Xiangli Song, Jing Cheng, Mowei Wang et al.SIGCOMM 2024 · 27 citations
- Annulus: A Dual Congestion Control Loop for Datacenter and WAN Traffic AggregatesAhmed Saeed, Varun Gupta, Prateesh Goyal, Milad Sharif et al.SIGCOMM 2020 · 72 citations
