Neural Topological Ordering for Computation Graphs
Mukul Gagrani, Corrado Rainone, Yang Yang, Harris Teague, Wonseok Jeon, Roberto Bondesan, Herke van Hoof, Christopher Lott, Weiliang Will Zeng, Piero Zappi
摘要
Recent works on machine learning for combinatorial optimization have shown that learning based approaches can outperform heuristic methods in terms of speed and performance. In this paper, we consider the problem of finding an optimal topological order on a directed acyclic graph with focus on the memory minimization problem which arises in compilers. We propose an end-to-end machine learning based approach for topological ordering using an encoder-decoder framework. Our encoder is a novel attention based graph neural network architecture called Topoformer which uses different topological transforms of a DAG for message passing. The node embeddings produced by the encoder are converted into node priorities which are used by the decoder to generate a probability distribution over topological orders. We train our model on a dataset of synthetically generated graphs called layered graphs. We show that our model outperforms, or is on-par, with several topological ordering baselines while being significantly faster on synthetic graphs with up to 2k nodes. We also train and test our model on a set of real-world computation graphs, showing performance improvements.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Transformers over Directed Acyclic GraphsYuankai Luo, Veronika Thost, Lei ShiNeurIPS 2023 · 被引用 43 次
- Learning to Scale Logits for Temperature-Conditional GFlowNetsMinsu Kim, Joohwan Ko, Taeyoung Yun, Dinghuai Zhang 等ICML 2024 · 被引用 31 次
- A Theory of Non-acyclic Generative Flow NetworksLeo Maxime Brunswic, Yinchuan Li, Yushun Xu, Yijun Feng 等AAAI 2024 · 被引用 9 次
- Differentiable Combinatorial Scheduling at ScaleMingju Liu, Yingjie Li, Jiaqi Yin, Zhiru Zhang 等ICML 2024 · 被引用 7 次
- Moccasin: Efficient Tensor Rematerialization for Neural NetworksBurak Bartan, Haoming Li, Harris Teague, Christopher Lott 等ICML 2023 · 被引用 3 次
它引用的顶会 Paper8
- Do Transformers Really Perform Badly for Graph Representation?Chengxuan Ying, Tianle Cai, Shengjie Luo, Shuxin Zheng 等NeurIPS 2021 · 被引用 1,632 次
- On Layer Normalization in the Transformer ArchitectureRuibin Xiong, Yunchang Yang, Di He, Kai Zheng 等ICML 2020 · 被引用 1,388 次
- NeuroLKH: Combining Deep Learning Model with Lin-Kernighan-Helsgaun Heuristic for Solving the Traveling Salesman ProblemLiang Xin, Wen Song, Zhiguang Cao, Jie ZhangNeurIPS 2021 · 被引用 202 次
- Directed Acyclic Graph Neural NetworksVeronika Thost, Jie ChenICLR 2021 · 被引用 134 次
- Reinforced Genetic Algorithm Learning for Optimizing Computation GraphsAditya Paliwal, Felix Gimeno, Vinod Nair, Yujia Li 等ICLR 2020 · 被引用 70 次
相关 Paper
- Transferable Graph Optimizers for ML CompilersYanqi Zhou, Sudip Roy, AmirAli Abdolrashidi, Daniel Wong 等NeurIPS 2020 · 被引用 63 次
- TopoFormer: Topology Meets Attention for Graph LearningMd Joshem Uddin, Astrit Tola, Cuneyt Gurcan Akcora, Baris CoskunuzerICLR 2026 · 被引用 2 次
- Directed Graph Grammars for Sequence-based LearningMichael Sun, Orion Foo, Gang Liu, Wojciech Matusik 等ICML 2025
- FlowerFormer: Empowering Neural Architecture Encoding Using a Flow-Aware Graph TransformerDongyeong Hwang, Hyunju Kim, Sunwoo Kim, Kijung ShinCVPR 2024
- PACE: A Parallelizable Computation Encoder for Directed Acyclic GraphsZehao Dong, Muhan Zhang, Fuhai Li, Yixin ChenICML 2022 · 被引用 24 次
