Coyote v2: Raising the Level of Abstraction for Data Center FPGAs
Benjamin Ramhorst, Dario Korolija, Maximilian Jakob Heer, Jonas Dann, Luhao Liu, Gustavo Alonso
Abstract
In the trend towards hardware specialization, FPGAs play a dual role as accelerators for offloading, e.g., network virtualization, and as a vehicle for prototyping and exploring hardware designs. While FPGAs offer versatility and performance, integrating them in larger systems remains challenging. Thus, recent efforts have focused on raising the level of abstraction through better interfaces and high-level programming languages. Yet, there is still quite some room for improvement. In this paper, we present Coyote v2, an open-source FPGA shell built with a novel, three-layer hierarchical design supporting dynamic partial reconfiguration of services and user logic, with a unified logic interface, and high-level software abstractions which facilitate application deployment, multi-tenancy and transparent workload pipelining. Experimental results indicate Coyote v2 reduces synthesis times between 15% and 20% and run-time reconfiguration times by an order of magnitude, when compared to existing systems. We also demonstrate the advantages of Coyote v2 by deploying several realistic applications, including HyperLogLog cardinality estimation, AES encryption, and neural network inference. Finally, Coyote v2 places a great deal of emphasis on integration with real systems through reusable and reconfigurable services, including a fully RoCE v2-compliant networking stack, a shared virtual memory model with the host, and a DMA engine between FPGAs and GPUs. We demonstrate these features by, e.g., seamlessly deploying an FPGA-accelerated neural network from Python.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- NutCracker: A Compilation Framework for Hybrid DPU ArchitecturesYihan Yang, Haifeng Sun, Antoine Kaufmann, Jialin LiEuroSys 2026 · 2 citations
- μShell: A Microkernel-based FPGA Shell ArchitectureJiyang Chen, Anubhav Panda, Harshavardhan Unnibhavi, Atsushi Koshiba et al.OSDI 2026 · 1 citation
- RoCE BALBOA: Service-Enhanced RDMA Offload Engine for Data Center SmartNICsMaximilian Jakob Heer, Benjamin Ramhorst, Yu Zhu, Luhao Liu et al.OSDI 2026
Builds on14
- Do OS abstractions make sense on FPGAs?Dario Korolija, Timothy Roscoe, Gustavo AlonsoOSDI 2020 · 114 citations
- DFX: A Low-latency Multi-FPGA Appliance for Accelerating Transformer-based Text GenerationSeongmin Hong, Seungjae Moon, Junsoo Kim, Sungjae Lee et al.MICRO 2022 · 107 citations
- Virtualizing FPGAs in the CloudYue Zha, Jing LiASPLOS 2020 · 92 citations
- StRoM: smart remote memoryDavid Sidler, Zeke Wang, Monica Chiosa, Amit Kulkarni et al.EuroSys 2020 · 83 citations
- FpgaNIC: An FPGA-based Versatile 100Gb SmartNIC for GPUsZeke Wang, Hongjing Huang, Jie Zhang, Fei Wu et al.USENIX ATC 2022 · 58 citations
Related papers
- Beehive: A Flexible Network Stack for Direct-Attached AcceleratorsKatie Lim, Matthew Giordano, Theano Stavrinos, Irene Zhang et al.MICRO 2024 · 4 citations
- Reticle: a virtual machine for programming modern FPGAsLuis Vega, Joseph McMahan, Adrian Sampson, Dan Grossman et al.PLDI 2021 · 8 citations
- Harmonia: A Unified Framework for Heterogeneous FPGA Acceleration in the CloudLuyang Li, Heng Pan, Xinchen Wan, Kai Lv et al.ASPLOS 2025 · 3 citations
- Reconfigurable Virtual Memory for FPGA-Driven I/OJoshua Landgraf, Matthew Giordano, Esther Yoon, Christopher J. RossbachASPLOS 2023 · 11 citations
- Rosebud: Making FPGA-Accelerated Middlebox Development More PleasantMoein Khazraee, Alex Forencich, George C. Papen, Alex C. Snoeren et al.ASPLOS 2023 · 8 citations
