Efficient and scalable core multiplexing with M³v
Nils Asmussen, Sebastian Haas, Carsten Weinhold, Till Miemietz, Michael Roitzsch
摘要
The M 3 system (ASPLOS '16) proposed a hardware/software codesign that simplifies integration between general-purpose cores and special-purpose accelerators, allowing users to easily utilize them in a unified manner. M 3 is a tiled architecture, whose tiles (cores and accelerators) are partitioned between applications, such that each tile is dedicated to its own application.
The M 3 x system (ATC '19) extended M 3 by trading off some isolation to enable coarse-grained multiplexing of tiles among multiple applications. With M 3 x, if source tile 𝑡 1 runs code of application 𝑝 and sends a message 𝑚 to destination tile 𝑡 2 while 𝑡 2 is currently not associated with 𝑝, then 𝑚 is forwarded to the right place through a łslow pathž, via some special OS tile.
In this paper, we present M 3 v, which extends M 3 x by further trading off some isolation between applications to support łfast pathž communication that does not require the said OS tile's involvement. Thus, with M 3 v, a tile can be efficiently multiplexed between applications provided it is a general-purpose core. M 3 v achieves this goal by 1) adding a local multiplexer to each such core, and by 2) virtualizing the core's hardware component responsible for cross-tile communications. We prototype M 3 v using RISC-V cores on an FPGA platform and show that it significantly outperforms M 3 x and may achieve competitive performance to Linux.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper2
- The Demikernel Datapath OS Architecture for Microsecond-scale Datacenter SystemsIrene Zhang, Amanda Raybuck, Pratyush Patel, Kirk Olynyk 等SOSP 2021 · 被引用 83 次
- LeapIO: Efficient and Portable Virtual NVMe Storage on ARM SoCsHuaicheng Li, Mingzhe Hao, Stanko Novakovic, Vaibhav Gogte 等ASPLOS 2020 · 被引用 58 次
相关 Paper
- Software-Defined Vector Processing on Manycore FabricsPhilip Bedoukian, Neil Adit, Edwin Peguero, Adrian SampsonMICRO 2021 · 被引用 3 次
- μShell: A Microkernel-based FPGA Shell ArchitectureJiyang Chen, Anubhav Panda, Harshavardhan Unnibhavi, Atsushi Koshiba 等OSDI 2026 · 被引用 1 次
- M³ViT: Mixture-of-Experts Vision Transformer for Efficient Multi-task Learning with Model-Accelerator Co-designHanxue Liang, Zhiwen Fan, Rishov Sarkar, Ziyu Jiang 等NeurIPS 2022 · 被引用 152 次
- The nanoPU: A Nanosecond Network Stack for DatacentersStephen Ibanez, Alex Mallery, Serhat Arslan, Theo Jepsen 等OSDI 2021 · 被引用 74 次
- Venus: A Versatile Deep Neural Network Accelerator Architecture Design for Multiple ApplicationsJiaqi Yang, Hao Zheng, Ahmed LouriDAC 2023 · 被引用 11 次
