Learning of Discrete Graphical Models with Neural Networks
Abhijith Jayakumar, Andrey Y. Lokhov, Sidhant Misra, Marc Vuffray
Abstract
Graphical models are widely used in science to represent joint probability distributions with an underlying conditional dependence structure. The inverse problem of learning a discrete graphical model given i.i.d samples from its joint distribution can be solved with near-optimal sample complexity using a convex optimization method known as Generalized Regularized Interaction Screening Estimator (GRISE). But the computational cost of GRISE becomes prohibitive when the energy function of the true graphical model has higher order terms. We introduce NN-GRISE, a neural net based algorithm for graphical model learning, to tackle this limitation of GRISE. We use neural nets as function approximators in an interaction screening objective function. The optimization of this objective then produces a neural-net representation for the conditionals of the graphical model. NN-GRISE algorithm is seen to be a better alternative to GRISE when the energy function of the true model has a high order with a high degree of symmetry. In these cases, NN-GRISE is able to find the correct parsimonious representation for the conditionals without being fed any prior information about the true model. NN-GRISE can also be used to learn the underlying structure of the true model with some simple modifications to its training procedure. In addition, we also show a variant of NN-GRISE that can be used to learn a neural net representation for the full energy function of the true model.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 55bf5807-599e-4dde-93b9-1dad269c17ebCited by top-tier papers2
- Efficient Learning of Discrete Graphical ModelsMarc Vuffray, Sidhant Misra, Andrey Y. LokhovNeurIPS 2020 · 46 citations
- Concurrent Multi-Label Prediction in Event StreamsXiao Shou, Tian Gao, Dharmashankar Subramanian, Debarun Bhattacharjya et al.AAAI 2023 · 13 citations
Builds on2
Related papers
- Exponential Reduction in Sample Complexity with Learning of Ising Model DynamicsArkopal Dutt, Andrey Y. Lokhov, Marc Vuffray, Sidhant MisraICML 2021 · 8 citations
- Recovering Causal Structures from Low-Order Conditional IndependenciesMarcel Wienöbst, Maciej LiskiewiczAAAI 2020 · 13 citations
- Generalized Precision Matrix for Scalable Estimation of Nonparametric Markov NetworksYujia Zheng, Ignavier Ng, Yewen Fan, Kun ZhangICLR 2023
- Bilevel Network Learning via Hierarchically Structured SparsityJiayi Fan, Jingyuan Yang, Shuangge Ma, Mengyun WuNeurIPS 2025 · 1 citation
- GLAD: Learning Sparse Graph RecoveryHarsh Shrivastava, Xinshi Chen, Binghong Chen, Guanghui Lan et al.ICLR 2020 · 39 citations
