Understanding Routable PCIe Performance for Composable Infrastructures
Wentao Hou, Jie Zhang, Zeke Wang, Ming Liu
摘要
Routable PCIe has become the predominant cluster interconnect to build emerging composable infrastructures. Empowered by PCIe non-transparent bridge devices, PCIe transactions can traverse multiple switching domains, enabling a server to elastically integrate a number of remote PCIe devices as local ones. However, it is unclear how to move data or perform communication efficiently over the routable PCIe fabric without understanding its capabilities and limitations.
This paper presents the design and implementation of rP-CIeBench 1 , a software-hardware co-designed benchmarking framework to systematically characterize the routable PCIe fabric. rPCIeBench provides flexible data communication primitives, exposes end-to-end PCIe transaction observability, and enables reconfigurable experiment deployment. Using rPCIeBench, we first analyze the communication characteristics of a routable PCIe path, quantify its performance tax, and compare it with the local PCIe link. We then use it to dissect in-fabric traffic orchestration behaviors and draw three interesting findings: approximate max-min bandwidth partition, fast end-to-end bandwidth synchronization, and interferencefree among orthogonal data paths. Finally, we encode gathered characterization insights as traffic orchestration rules and develop an edge constraints relaxing algorithm to estimate PCIe flow transmission performance over a shared fabric. We validate its accuracy and demonstrate its potential to provide an optimization guide to design efficient flow schedulers.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- CEIO: A Cache-Efficient Network I/O Architecture for NIC-CPU Data PathsBowen Liu, Xinyang Huang, Qijing Li, Zhuobin Huang 等SIGCOMM 2025 · 被引用 10 次
- Building an Elastic Block Storage over EBOFs Using Shadow ViewsSheng Jiang, Ming LiuNSDI 2025 · 被引用 10 次
- Understanding and Profiling NVMe-over-TCP Using ntprofYuyuan Kang, Ming LiuNSDI 2025 · 被引用 9 次
- Building Massive MIMO Baseband Processing on a Single-Node SupercomputerXincheng Xie, Wentao Hou, Zerui Guo, Ming LiuNSDI 2025 · 被引用 8 次
- Understanding and Profiling CXL.mem Using PathFinderXiao Li, Zerui Guo, Yuebin Bai, Mahesh Ketkar 等SIGCOMM 2025 · 被引用 6 次
它引用的顶会 Paper11
- Pond: CXL-Based Memory Pooling Systems for Cloud PlatformsHuaicheng Li, Daniel S. Berger, Lisa Hsu, Daniel Ernst 等ASPLOS 2023 · 被引用 328 次
- TPP: Transparent Page Placement for CXL-Enabled Tiered-MemoryHasan Al Maruf, Hao Wang, Abhishek Dhanotia, Johannes Weiner 等ASPLOS 2023 · 被引用 255 次
- Demystifying CXL Memory with Genuine CXL-Ready Systems and DevicesYan Sun, Yifan Yuan, Zeduo Yu, Reese Kuper 等MICRO 2023 · 被引用 133 次
- Overcoming the Memory Wall with CXL-Enabled SSDsShao-Peng Yang, Minjae Kim, Sanghyun Nam, Juhyung Park 等USENIX ATC 2023 · 被引用 75 次
- CXL-ANNS: Software-Hardware Collaborative Memory Disaggregation and Computation for Billion-Scale Approximate Nearest Neighbor SearchJunhyeok Jang, Hanjin Choi, Hanyeoreum Bae, Seungjun Lee 等USENIX ATC 2023 · 被引用 75 次
相关 Paper
- weBurst can be Harmless: Achieving Line-rate Software Traffic Shaping by Inter-flow BatchingDanfeng Shan, Shihao Hu, Yuqi Liu, Wanchun Jiang 等INFOCOM 2023 · 被引用 2 次
- RpcNIC: Enabling Efficient Datacenter RPC Offloading on PCIe-attached SmartNICsJie Zhang, Hongjing Huang, Xuzheng Chen, Xiang Li 等HPCA 2025 · 被引用 6 次
- NetTLP: A Development Platform for PCIe devices in Software Interacting with HardwareYohei Kuga, Ryo Nakamura, Takeshi Matsuya, Yuji SekiyaNSDI 2020 · 被引用 9 次
- 1Pipe: scalable total order communication in data center networksBojie Li, Gefei Zuo, Wei Bai, Lintao ZhangSIGCOMM 2021 · 被引用 5 次
- Powerful GPUs or Fast Interconnects: Analyzing Relational Workloads on Modern GPUsMarko Kabic, Bowen Wu, Jonas Dann, Gustavo AlonsoVLDB 2025 · 被引用 7 次
