Using trio: juniper networks' programmable chipset - for emerging in-network applications
Mingran Yang, Alex Baban, Valery Kugel, Jeff Libby, Scott Mackie, Swamy Sadashivaiah Renu Kananda, Chang-Hong Wu, Manya Ghobadi
摘要
This paper describes Trio, a programmable chipset used in Juniper Networks' MX-series routers and switches. Trio's architecture is based on a multi-threaded programmable packet processing engine and a hierarchy of high-capacity memory systems, making it fundamentally different from pipeline-based architectures. Trio gracefully handles non-homogeneous packet processing rates for a wide range of networking use cases and protocols, making it an ideal platform for emerging in-network applications. We begin by describing the Trio chipset's fundamental building blocks, including its multi-threaded Packet Forwarding and Packet Processing Engines. We then discuss Trio's programming language, called Microcode. To showcase Trio's flexible Microcode-based programming environment, we describe two use cases. First, we demonstrate Trio's ability to perform in-network aggregation for distributed machine learning. Second, we propose and design an in-network straggler mitigation technique using Trio's timer threads. We prototype both use cases on a testbed using three real DNN models (ResNet50, DenseNet161, and VGG11) to demonstrate Trio's ability to mitigate stragglers while performing in-network aggregation. Our evaluations show that when stragglers occur in the cluster, Trio outperforms today's pipeline-based solutions by up to 1.8x.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper20
- CASSINI: Network-Aware Job Scheduling in Machine Learning ClustersSudarsanan Rajasekaran, Manya Ghobadi, Aditya AkellaNSDI 2024 · 被引用 144 次
- THC: Accelerating Distributed Deep Learning Using Tensor Homomorphic CompressionMinghao Li, Ran Ben Basat, Shay Vargaftik, ChonLam Lao 等NSDI 2024 · 被引用 44 次
- A Generic Service to Provide In-Network Aggregation for Key-Value StreamsYongchao He, Wenfei Wu, Yanfang Le, Ming Liu 等ASPLOS 2023 · 被引用 37 次
- ClickINC: In-network Computing as a Service in Heterogeneous Programmable Data-center NetworksWenquan Xu, Zijian Zhang, Yong Feng, Haoyu Song 等SIGCOMM 2023 · 被引用 34 次
- Pyrrha: Congestion-Root-Based Flow Control to Eliminate Head-of-Line Blocking in DatacenterKexin Liu, Zhaochen Zhang, Chang Liu, Yizhi Wang 等NSDI 2025 · 被引用 13 次
它引用的顶会 Paper11
- ATP: In-network Aggregation for Multi-tenant LearningChonLam Lao, Yanfang Le, Kshiteej Mahajan, Yixi Chen 等NSDI 2021 · 被引用 359 次
- PINT: Probabilistic In-band Network TelemetryRan Ben Basat, Sivaramakrishnan Ramanathan, Yuliang Li, Gianni Antichi 等SIGCOMM 2020 · 被引用 268 次
- TEA: Enabling State-Intensive Network Functions on Programmable SwitchesDaehyeok Kim, Zaoxing Liu, Yibo Zhu, Changhoon Kim 等SIGCOMM 2020 · 被引用 121 次
- Taurus: a data plane architecture for per-packet MLTushar Swamy, Alexander Rucker, Muhammad Shahbaz, Ishan Gaur 等ASPLOS 2022 · 被引用 94 次
- An In-Network Architecture for Accelerating Shared-Memory Multiprocessor CollectivesBenjamin Klenk, Nan Jiang, Greg Thorson, Larry DennisonISCA 2020 · 被引用 67 次
相关 Paper
- OptimusPrime: Unleash Dataplane Programmability through a Transformable ArchitectureZhikang Chen, Yong Feng, Shuxin Liu, Haoyu Song 等SIGCOMM 2024 · 被引用 6 次
- Training Job Placement in Clusters with Statistical In-Network AggregationBohan Zhao, Wei Xu, Shuo Liu, Yang Tian 等ASPLOS 2024 · 被引用 17 次
- FENIX: Enabling In-Network DNN Inference with FPGA-Enhanced Programmable SwitchesXiangyu Gao, Tong Li, Yinchao Zhang, Ziqiang Wang 等NSDI 2026 · 被引用 12 次
- Rearchitecting Programmable Networks For In-Network Computing: From Hardware To LanguageHaifeng Sun, Bing Liu, Taixu Tian, Jinbo Sun 等EuroSys 2026 · 被引用 2 次
- Constrained In-network Computing with Low Congestion in Datacenter NetworksRaz Segal, Chen Avin, Gabriel ScalosubINFOCOM 2022 · 被引用 19 次
