Multilinear Latent Conditioning for Generating Unseen Attribute Combinations
Markos Georgopoulos, Grigorios Chrysos, Maja Pantic, Yannis Panagakis
摘要
Deep generative models rely on their inductive bias to facilitate generalization, especially for problems with high dimensional data, like images. However, empirical studies have shown that variational autoencoders (VAE) and generative adversarial networks (GAN) lack the generalization ability that occurs naturally in human perception. For example, humans can visualize a woman smiling after only seeing a smiling man. On the contrary, the standard conditional VAE (cVAE) is unable to generate unseen attribute combinations. To this end, we extend cVAE by introducing a multilinear latent conditioning framework that captures the multiplicative interactions between the attributes. We implement two variants of our model and demonstrate their efficacy on MNIST, Fashion-MNIST and CelebA. Altogether, we design a novel conditioning framework that can be used with any architecture to synthesize unseen attribute combinations.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Multilinear Mixture of Experts: Scalable Expert Specialization through FactorizationJames Oldfield, Markos Georgopoulos, Grigorios Chrysos, Christos Tzelepis 等NeurIPS 2024 · 被引用 41 次
- Polynomial Neural Fields for Subband Decomposition and ManipulationGuandao Yang, Sagie Benaim, Varun Jampani, Kyle Genova 等NeurIPS 2022 · 被引用 26 次
- Conditional Generation Using Polynomial ExpansionsGrigorios Chrysos, Markos Georgopoulos, Yannis PanagakisNeurIPS 2021 · 被引用 13 次
- Going beyond Compositions, DDPMs Can Produce Zero-Shot InterpolationsJustin Deschenaux, Igor Krawczuk, Grigorios Chrysos, Volkan CevherICML 2024 · 被引用 6 次
- Why DDIM Hallucinates More Than DDPM: A Theoretical Analysis of Reverse DynamicsMuhammad H Ashiq, Samanyu Arora, Abhinav Narayan Harish, Ishaan Kharbanda 等ICML 2026
相关 Paper
- MAGANet: Achieving Combinatorial Generalization by Modeling a Group ActionGeonho Hwang, Jaewoong Choi, Hyunsoo Cho, Myungjoo KangICML 2023 · 被引用 4 次
- Multi-Facet Clustering Variational AutoencodersFabian Falck, Haoting Zhang, Matthew Willetts, George Nicholson 等NeurIPS 2021 · 被引用 57 次
- Perceptual Generative AutoencodersZijun Zhang, Ruixiang Zhang, Zongpeng Li, Yoshua Bengio 等ICML 2020 · 被引用 31 次
- The role of Disentanglement in GeneralisationMilton Llera Montero, Casimir J. H. Ludwig, Rui Ponte Costa, Gaurav Malhotra 等ICLR 2021 · 被引用 97 次
- Information-Theoretic Generalization Bounds for VAEs: A Role of Encoder and Latent VariableFutoshi Futami, Masahiro FujisawaICML 2026
