Tripod: Three Complementary Inductive Biases for Disentangled Representation Learning
Kyle Hsu, Jubayer Ibn Hamid, Kaylee Burns, Chelsea Finn, Jiajun Wu
摘要
Inductive biases are crucial in disentangled representation learning for narrowing down an underspecified solution set. In this work, we consider endowing a neural network autoencoder with three select inductive biases from the literature: data compression into a grid-like latent space via quantization, collective independence amongst latents, and minimal functional influence of any latent on how other latents determine data generation. In principle, these inductive biases are deeply complementary: they most directly specify properties of the latent space, encoder, and decoder, respectively. In practice, however, naively combining existing techniques instantiating these inductive biases fails to yield significant benefits. To address this, we propose adaptations to the three techniques that simplify the learning problem, equip key regularization terms with stabilizing invariances, and quash degenerate incentives. The resulting model, Tripod, achieves state-of-the-art results on a suite of four image disentanglement benchmarks. We also verify that Tripod significantly improves upon its naive incarnation and that all three of its "legs" are necessary for best performance.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper13
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- Finite Scalar Quantization: VQ-VAE Made SimpleFabian Mentzer, David Minnen, Eirikur Agustsson, Michael TschannenICLR 2024 · 被引用 442 次
- Independent mechanism analysis, a new concept?Luigi Gresele, Julius von Kügelgen, Vincent Stimper, Bernhard Schölkopf 等NeurIPS 2021 · 被引用 133 次
- Disentanglement by Nonlinear ICA with General Incompressible-flow Networks (GIN)Peter Sorrenson, Carsten Rother, Ullrich KötheICLR 2020 · 被引用 132 次
- On the Identifiability of Nonlinear ICA: Sparsity and BeyondYujia Zheng, Ignavier Ng, Kun ZhangNeurIPS 2022 · 被引用 104 次
相关 Paper
- Disentanglement via Latent QuantizationKyle Hsu, William Dorrell, James C. R. Whittington, Jiajun Wu 等NeurIPS 2023 · 被引用 54 次
- Geometric Inductive Biases for Identifiable Unsupervised Learning of Disentangled RepresentationsZiqi Pan, Li Niu, Liqing ZhangAAAI 2023 · 被引用 3 次
- Orthogonality-Enforced Latent Space in Autoencoders: An Approach to Learning Disentangled RepresentationsJaehoon Cha, Jeyan ThiyagalingamICML 2023 · 被引用 1 次
- On Incorporating Inductive Biases into VAEsNing Miao, Emile Mathieu, Siddharth N, Yee Whye Teh 等ICLR 2022 · 被引用 12 次
- An Identifiable Double VAE For Disentangled RepresentationsGraziano Mita, Maurizio Filippone, Pietro MichiardiICML 2021 · 被引用 39 次
