High-throughput and Flexible Host Networking for Accelerated Computing
Athinagoras Skiadopoulos, Zhiqiang Xie, Mark Zhao, Qizhe Cai, Saksham Agarwal, Jacob Adelmann, David Ahern, Carlo Contavalli, Michael D. Goldflam, Vitaly Mayatskikh, Raghu Raja, Daniel Walton
2024年份
11被引次数
5顶会引用
摘要
Modern network hardware is able to meet the stringent bandwidth demands of applications like GPU-accelerated AI. However, existing host network stacks offer a hard tradeoff between performance (in terms of sustained throughput when compared to network hardware capacity) and flexibility (in terms of the ability to select, customize, and extend different network protocols).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Barre: Empowering Simplified and Versatile Programmable Congestion Control in High-Speed AI ClustersYajuan Peng, Haoran Wei, Xiaolong Zhong, Junkai Huang 等USENIX ATC 2025 · 被引用 5 次
- Remote TCP Connection Offload and ApplicationsShuo Li, Steven W. D. Chien, Tianyi Gao, Michio HondaNSDI 2026 · 被引用 3 次
- Presto: A Match-Action TCP Stack for the Terabit EraRajath Shashidhara, Antoine Kaufmann, Simon PeterSIGCOMM 2026 · 被引用 1 次
- UCCL-Tran: An Extensible Software Transport Layer for GPU NetworkingYang Zhou, Zhongjie Chen, Ziming Mao, ChonLam Lao 等OSDI 2026
- SmartNS: Enabling Line-rate and Flexible Network Stack with SmartNICXuzheng Chen, Jie Zhang, Baolin Zhu, Xueying Zhu 等EuroSys 2026
它引用的顶会 Paper22
- GShard: Scaling Giant Models with Conditional Computation and Automatic ShardingDmitry Lepikhin, HyoukJoong Lee, Yuanzhong Xu, Dehao Chen 等ICLR 2021 · 被引用 1,954 次
- Orca: A Distributed Serving System for Transformer-Based Generative ModelsGyeong-In Yu, Joo Seong Jeong, Geon-Woo Kim, Soojeong Kim 等OSDI 2022 · 被引用 690 次
- When Cloud Storage Meets RDMAYixiao Gao, Qiang Li, Lingbo Tang, Yongqing Xi 等NSDI 2021 · 被引用 228 次
- TopoOpt: Co-optimizing Network Topology and Parallelization Strategy for Distributed Training JobsWeiyang Wang, Moein Khazraee, Zhizhen Zhong, Manya Ghobadi 等NSDI 2023 · 被引用 215 次
- Caladan: Mitigating Interference at Microsecond TimescalesJoshua Fried, Zhenyuan Ruan, Amy Ousterhout, Adam BelayOSDI 2020 · 被引用 213 次
相关 Paper
- <u>G</u>PU <u>i</u>nitiated <u>O</u>penSHMEM: correct and efficient intra-kernel networking for dGPUsKhaled Hamidouche, Michael LeBeanePPoPP 2020 · 被引用 17 次
- Understanding host network stack overheadsQizhe Cai, Shubham Chaudhary, Midhul Vuppalapati, Jaehyun Hwang 等SIGCOMM 2021 · 被引用 150 次
- GPU-Ether: GPU-native Packet I/O for GPU Applications on Commodity EthernetChangue Jung, Suhwan Kim, Ikjun Yeom, Honguk Woo 等INFOCOM 2021 · 被引用 7 次
- F4T: A Fast and Flexible FPGA-based Full-stack TCP Acceleration FrameworkJunehyuk Boo, Yujin Chung, Eunjin Baek, Seongmin Na 等ISCA 2023 · 被引用 8 次
- Beehive: A Flexible Network Stack for Direct-Attached AcceleratorsKatie Lim, Matthew Giordano, Theano Stavrinos, Irene Zhang 等MICRO 2024 · 被引用 4 次
