Global Identifiability of Overcomplete Dictionary Learning via L1 and Volume Minimization
Yuchen Sun, Kejun Huang
Abstract
We propose a novel formulation for dictionary learning with an overcomplete dictionary, i.e., when the number of atoms is larger than the dimension of the dictionary. The proposed formulation consists of a weighted sum of โ 1 norms of the rows of the sparse coefficient matrix plus the log of the matrix volume of the dictionary matrix. The main contribution of this work is to show that this novel formulation guarantees global identifiability of the overcomplete dictionary, under a mild condition that the sparse coefficient matrix satisfies a strong scattering condition in the hypercube. Furthermore, if every column of the coefficient matrix is sparse and the dictionary guarantees โ 1 recovery, then the coefficient matrix is identifiable as well. This is a major breakthrough for not only dictionary learning but also general matrix factorization models as identifiability is guaranteed even when the latent dimension is higher than the ambient dimension. We also provide a probabilistic analysis and show that if the sparse coefficient matrix is generated from the widely adopted sparse-Gaussian model, then the ๐ ร ๐ overcomplete dictionary is globally identifiable if the sample size is bigger than a constant times (๐ 2 /๐) log(๐ 2 /๐) with overwhelming probability. Finally, we propose an algorithm based on alternating minimization to solve the new proposed formulation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fe620086-aa53-4b1f-b259-a80a87606543Cited by top-tier papers3
- Diverse Influence Component Analysis: A Geometric Approach to Nonlinear Mixture IdentifiabilityHoang-Son Nguyen, Xiao FuNeurIPS 2025 ยท 6 citations
- Diverse Dictionary LearningYujia Zheng, Zijian Li, Shunxing Fan, Andrew Gordon Wilson et al.ICLR 2026
- Mechanistic Interpretability Should Prioritize Feature Consistency in Sparse AutoencodersXiangchen Song, Aashiq Muhamed, Yujia Zheng, Lingjing Kong et al.ACL 2026
Builds on1
Related papers
- Global Identifiability of ๐1-based Dictionary Learning via Matrix Volume OptimizationJingzhou Hu, Kejun HuangNeurIPS 2023 ยท 19 citations
- Hiding Data Helps: On the Benefits of Masking for Sparse CodingMuthu Chidambaram, Chenwei Wu, Yu Cheng, Rong GeICML 2023
- The Power of Preconditioning in Overparameterized Low-Rank Matrix SensingXingyu Xu, Yandi Shen, Yuejie Chi, Cong MaICML 2023 ยท 51 citations
- Rank Overspecified Robust Matrix Recovery: Subgradient Method and Exact RecoveryLijun Ding, Liwei Jiang, Yudong Chen, Qing Qu et al.NeurIPS 2021 ยท 30 citations
- Deep Network Classification by Scattering and Homotopy Dictionary LearningJohn Zarka, Louis Thiry, Tomรกs Angles, Stรฉphane MallatICLR 2020 ยท 43 citations
