Exponential Quantum Communication Advantage in Distributed Inference and Learning
Dar Gilboa, Hagay Michaeli, Daniel Soudry, Jarrod R. McClean
Abstract
Training and inference with large machine learning models that far exceed the memory capacity of individual devices necessitates the design of distributed architectures, forcing one to contend with communication constraints. We present a framework for distributed computation over a quantum network in which data is encoded into specialized quantum states. We prove that for models within this framework, inference and training using gradient descent can be performed with exponentially less communication compared to their classical analogs, and with relatively modest overhead relative to standard gradient-based methods. We show that certain graph neural networks are particularly amenable to implementation within this framework, and moreover present empirical evidence that they perform well on standard benchmarks. To our knowledge, this is the first example of exponential quantum advantage for a generic class of machine learning problems that hold regardless of the data encoding cost. Moreover, we show that models in this class can encode highly nonlinear features of their inputs, and their expressivity increases exponentially with model depth. We also delineate the space of models for which exponential communication advantages hold by showing that they cannot hold for linear classification. Our results can be combined with natural privacy advantages in the communicated quantum states that limit the amount of information that can be extracted from them about the data and model parameters. Taken as a whole, these findings form a promising foundation for distributed machine learning over quantum networks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f43d534a-b8aa-47eb-902b-c05ad67231d7Cited by top-tier papers1
Ask how each one uses itBuilds on8
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong et al.NeurIPS 2020 · 3,935 citations
- Transformers are RNNs: Fast Autoregressive Transformers with Linear AttentionAngelos Katharopoulos, Apoorv Vyas, Nikolaos Pappas, François FleuretICML 2020 · 2,665 citations
- PaLM-E: An Embodied Multimodal Language ModelDanny Driess, Fei Xia, Mehdi S. M. Sajjadi, Corey Lynch et al.ICML 2023 · 2,601 citations
- DropGNN: Random Dropouts Increase the Expressiveness of Graph Neural NetworksPál András Papp, Karolis Martinkus, Lukas Faber, Roger WattenhoferNeurIPS 2021 · 182 citations
Related papers
- Concentration of Data Encoding in Parameterized Quantum CircuitsGuangxi Li, Ruilin Ye, Xuanqiang Zhao, Xin WangNeurIPS 2022 · 42 citations
- Equivariant Quantum Graph CircuitsPéter Mernyei, Konstantinos Meichanetzidis, Ismail Ilkan CeylanICML 2022 · 9 citations
- The Inductive Bias of Quantum KernelsJonas M. Kübler, Simon Buchholz, Bernhard SchölkopfNeurIPS 2021 · 190 citations
- On Designing General and Expressive Quantum Graph Neural Networks with Applications to MILP Instance RepresentationXinyu Ye, Hao Xiong, Jianhao Huang, Ziang Chen et al.ICLR 2025
- Distributed Learning of Fully Connected Neural Networks using Independent Subnet TrainingBinhang Yuan, Cameron R. Wolfe, Chen Dun, Yuxin Tang et al.VLDB 2022 · 42 citations
