Learning Structured Latent Factors from Dependent Data:A Generative Model Framework from Information-Theoretic Perspective
Ruixiang Zhang, Masanori Koyama, Katsuhiko Ishiguro
Abstract
Learning controllable and generalizable representation of multivariate data with desired structural properties remains a fundamental problem in machine learning. In this paper, we present a novel framework for learning generative models with various underlying structures in the latent space. We represent the inductive bias in the form of mask variables to model the dependency structure in the graphical model and extend the theory of multivariate information bottleneck (Friedman et al., 2001) to enforce it. Our model provides a principled approach to learn a set of semantically meaningful latent factors that reflect various types of desired structures like capturing correlation or encoding invariance, while also offering the flexibility to automatically estimate the dependency structure from data. We show that our framework unifies many existing generative models and can be applied to a variety of tasks, including multimodal data modeling, algorithmic fairness, and out-of-distribution generalization.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- Neural Fourier Transform: A General Approach to Equivariant Representation LearningMasanori Koyama, Kenji Fukumizu, Kohei Hayashi, Takeru MiyatoICLR 2024 · 9 citations
- Flexible Language Modeling in Continuous Space with Transformer-based Autoregressive FlowsRuixiang Zhang, Shuangfei Zhai, Jiatao Gu, Yizhe Zhang et al.NeurIPS 2025 · 8 citations
- Learning Representation from Neural Fisher Kernel with Low-rank ApproximationRuixiang Zhang, Shuangfei Zhai, Etai Littwin, Joshua M. SusskindICLR 2022 · 5 citations
- Target Concrete Score Matching: A Holistic Framework for Discrete DiffusionRuixiang Zhang, Shuangfei Zhai, Yizhe Zhang, James Thornton et al.ICML 2025
Related papers
- IBMA: Information Bottleneck-Based Multimodal AlignmentYancheng Wang, Zeyu Dong, Dongfang Sun, Alvin Silva et al.ICML 2026
- A Bayesian Nonparametric Framework For Learning Disentangled RepresentationsVaishnavi Patil, Siddhi Patil, Matthew Evanusa, Amit Kumar Kundu et al.ICLR 2026
- Minimum Description Length and Generalization Guarantees for Representation LearningMilad Sefidgaran, Abdellatif Zaidi, Piotr KrasnowskiNeurIPS 2023 · 17 citations
- Identifiable Exchangeable Mechanisms for Causal Structure and Representation LearningPatrik Reizinger, Siyuan Guo, Ferenc Huszár, Bernhard Schölkopf et al.ICLR 2025
- Generative Interventions for Causal LearningChengzhi Mao, Augustine Cha, Amogh Gupta, Hao Wang et al.CVPR 2021
