Deep Tree Tensor Networks
Chang Nie
Abstract
Originating in quantum physics, tensor networks (TNs) have been widely adopted as exponential machines and parametric decomposers for recognition tasks. Typical TN models, such as Matrix Product States (MPS), have not yet achieved successful application in natural image recognition. When employed, they primarily serve to compress parameters within pre-existing networks, thereby losing their distinctive capability to capture exponential-order feature interactions. This paper introduces a novel architecture named Deep Tree Tensor Network (DTTN), which captures -order multiplicative interactions across features through multilinear operations, while essentially unfolding into a tree-like TN topology with the parameter-sharing property. DTTN is stacked with multiple antisymmetric interaction modules (AIMs), and this design facilitates efficient implementation. Furthermore, our theoretical analysis demonstrates the equivalence between quantum-inspired TN models and polynomial/multilinear networks under specific conditions. We posit that the DTTN could catalyze more interpretable research within this field. The proposed model is evaluated across multiple benchmarks and domains, demonstrating superior performance compared to both peer methods and state-of-the-art architectures. Our code is publicly available at https://github.com/NieCha/deep_tree_tensor_network.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ff2ffea9-5b33-4f71-b421-ccc2efbbe120Cited by top-tier papers1
Ask how each one uses itBuilds on14
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa et al.ICML 2021 · 8,974 citations
- MLP-Mixer: An all-MLP Architecture for VisionIlya O. Tolstikhin, Neil Houlsby, Alexander Kolesnikov, Lucas Beyer et al.NeurIPS 2021 · 3,862 citations
- Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space ModelLianghui Zhu, Bencheng Liao, Qian Zhang, Xinlong Wang et al.ICML 2024 · 1,725 citations
Related papers
- TensorNet: Cartesian Tensor Representations for Efficient Learning of Molecular PotentialsGuillem Simeon, Gianni De FabritiisNeurIPS 2023 · 100 citations
- High-Order Pooling for Graph Neural Networks with Tensor DecompositionChenqing Hua, Guillaume Rabusseau, Jian TangNeurIPS 2022 · 45 citations
- Tensor Decomposition Networks for Fast Machine Learning Interatomic Potential ComputationsYuchao Lin, Cong Fu, Zachary Krueger, Haiyang Yu et al.NeurIPS 2025 · 1 citation
- MTNL: A Unified Modeling Perspective for Enhancing Tensor Network LearningJunhua Zeng, Yuning Qiu, Binghua Li, Chao Li et al.ICML 2026
- ANTN: Bridging Autoregressive Neural Networks and Tensor Networks for Quantum Many-Body SimulationZhuo Chen, Laker Newhouse, Eddie Chen, Di Luo et al.NeurIPS 2023 · 19 citations
