Scaling Recurrent Models via Orthogonal Approximations in Tensor Trains
Ronak Mehta, Rudrasis Chakraborty, Vikas Singh, Yunyang Xiong
Abstract
Modern deep networks have proven to be very effective for analyzing real world images. However, their application in medical imaging is still in its early stages, primarily due to the large size of three-dimensional images, requiring enormous convolutional or fully connected layers – if we treat an image (and not image patches) as a sample. These issues only compound when the focus moves towards longitudinal analysis of 3D image volumes through recurrent structures, and when a point estimate of model parameters is insufficient in scientific applications where a reliability measure is necessary. Using insights from differential geometry, we adapt the tensor train decomposition to construct networks with significantly fewer parameters, allowing us to train powerful recurrent networks on whole brain image volume sequences. We describe the “orthogonal” tensor train, and demonstrate its ability to express a standard network layer both theoretically and empirically. We show its ability to effectively reconstruct whole brain volumes with faster convergence and stronger confidence intervals compared to the standard tensor train decomposition. We provide code and show experiments on the ADNI dataset using image sequences to regress on a cognition related outcome.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- Equivariance Allows Handling Multiple Nuisance Variables When Analyzing Pooled Neuroimaging DatasetsVishnu Suresh Lokhande, Rudrasis Chakraborty, Sathya N. Ravi, Vikas SinghCVPR 2022 · 1 citation
- Transformed Low-rank Adaptation via Tensor Decomposition and Its Applications to Text-to-image ModelsZerui Tao, Yuhta Takida, Naoki Murata, Qibin Zhao et al.ICCV 2025
- MobileDets: Searching for Object Detection Architectures for Mobile AcceleratorsYunyang Xiong, Hanxiao Liu, Suyog Gupta, Berkin Akin et al.CVPR 2021
Related papers
- Towards Efficient Tensor Decomposition-Based DNN Model Compression With Optimization FrameworkMiao Yin, Yang Sui, Siyu Liao, Bo YuanCVPR 2021
- Dilated Convolutional Neural Networks for Sequential Manifold-Valued DataRudrasis Chakraborty, Xingjian Zhen, Nicholas Vogt, Barbara B. Bendlin et al.ICCV 2019 · 43 citations
- Compact Autoregressive NetworkDi Wang, Feiqing Huang, Jingyu Zhao, Guodong Li et al.AAAI 2020 · 5 citations
- Cherry-Picking Gradients: Learning Low-Rank Embeddings of Visual Data via Differentiable Cross-ApproximationMikhail Usvyatsov, Anastasia Makarova, Rafael Ballester-Ripoll, Maxim V. Rakhuba et al.ICCV 2021 · 6 citations
- Fully-Connected Tensor Network Decomposition and Its Application to Higher-Order Tensor CompletionYu-Bang Zheng, Ting-Zhu Huang, Xi-Le Zhao, Qibin Zhao et al.AAAI 2021 · 183 citations
