Wavel: A Fast and Efficient Compilation System for Wafer-Scale Accelerators
Yeqi Huang, Congjie He, Haocheng Xiao, Yanwei Ye, Yi-Chieh Wang, Boyao Song, Yangshen Deng, Ziming Miao, Lingxiao Ma, Fan Yang, Luo Mai
2026年份
摘要
Wafer-scale accelerators offer a new scaling point for AI infrastructure, but they also create a new compilation regime: communication cost varies sharply with location, and the space of possible placements and execution schedules is enormous. Existing GPU, distributed, and vendor compilation systems largely retain a sharding-oriented view and therefore fail to fully leverage these emerging accelerators, leaving the dominant physical scheduling decisions unresolved.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- An MLIR Lowering Pipeline for Stencils at Wafer-ScaleNicolai Stawinoga, David Katz, Anton Lydike, Justs Zarins 等ASPLOS 2026
- MeshRT: Compile-Time Governed Wafer-Scale Runtime for Low-Latency High-Throughput InferenceCongjie He, Le Xu, Zhan Lu, Yeqi Huang 等SOSP 2026
- WaferLLM: Large Language Model Inference at Wafer ScaleCongjie He, Yeqi Huang, Pei Mu, Ziming Miao 等OSDI 2025 · 被引用 20 次
- Matrix Is All You Need: Rearchitecting Quantum Chemistry to Scale on AI AcceleratorsHaozhi Han, Kun Li, Fusong Ju, Qi Li 等SC 2025 · 被引用 2 次
- WSC-LLM: Efficient LLM Service and Architecture Co-exploration for Wafer-scale ChipsZheng Xu, Dehao Kong, Jiaxin Liu, Jinxi Li 等ISCA 2025 · 被引用 22 次
