Wavel: A Fast and Efficient Compilation System for Wafer-Scale Accelerators
Yeqi Huang, Congjie He, Haocheng Xiao, Yanwei Ye, Yi-Chieh Wang, Boyao Song, Yangshen Deng, Ziming Miao, Lingxiao Ma, Fan Yang, Luo Mai
Abstract
Wafer-scale accelerators offer a new scaling point for AI infrastructure, but they also create a new compilation regime: communication cost varies sharply with location, and the space of possible placements and execution schedules is enormous. Existing GPU, distributed, and vendor compilation systems largely retain a sharding-oriented view and therefore fail to fully leverage these emerging accelerators, leaving the dominant physical scheduling decisions unresolved.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get bdd70603-ffe4-4b9b-a5d1-7c166dc8e302Related papers
- An MLIR Lowering Pipeline for Stencils at Wafer-ScaleNicolai Stawinoga, David Katz, Anton Lydike, Justs Zarins et al.ASPLOS 2026
- MeshRT: Compile-Time Governed Wafer-Scale Runtime for Low-Latency High-Throughput InferenceCongjie He, Le Xu, Zhan Lu, Yeqi Huang et al.SOSP 2026
- WaferLLM: Large Language Model Inference at Wafer ScaleCongjie He, Yeqi Huang, Pei Mu, Ziming Miao et al.OSDI 2025 · 20 citations
- Matrix Is All You Need: Rearchitecting Quantum Chemistry to Scale on AI AcceleratorsHaozhi Han, Kun Li, Fusong Ju, Qi Li et al.SC 2025 · 2 citations
- WSC-LLM: Efficient LLM Service and Architecture Co-exploration for Wafer-scale ChipsZheng Xu, Dehao Kong, Jiaxin Liu, Jinxi Li et al.ISCA 2025 · 22 citations
