Variational Interaction Information Maximization for Cross-domain Disentanglement
HyeongJoo Hwang, Geon-Hyeong Kim, Seunghoon Hong, Kee-Eung Kim
Abstract
Cross-domain disentanglement is the problem of learning representations partitioned into domain-invariant and domain-specific representations, which is a key to successful domain transfer or measuring semantic distance between two domains. Grounded in information theory, we cast the simultaneous learning of domain-invariant and domain-specific representations as a joint objective of multiple information constraints, which does not require adversarial training or gradient reversal layers. We derive a tractable bound of the objective and propose a generative model named Interaction Information Auto-Encoder (IIAE). Our approach reveals insights on the desirable representation for cross-domain disentanglement and its connection to Variational Auto-Encoder (VAE). We demonstrate the validity of our model in the image-to-image translation and the cross-domain retrieval tasks. We further show that our model achieves the state-of-the-art performance in the zero-shot sketch based image retrieval task, even without external knowledge. Our implementation is publicly available at: https://github.com/gr8joo/IIAE 34th Conference on Neural Information Processing Systems (NeurIPS 2020),
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 14216e5a-a73c-4745-a213-2a580bbfb298Cited by top-tier papers20
- InfoGCN: Representation Learning for Human Skeleton-based Action RecognitionHyung-Gun Chi, Myoung Hoon Ha, Seung-geun Chi, Sang Wan Lee et al.CVPR 2022 · 383 citations
- Cross-Domain Recommendation to Cold-Start Users via Variational Information BottleneckJiangxia Cao, Jiawei Sheng, Xin Cong, Tingwen Liu et al.ICDE 2022 · 112 citations
- Multi-View Representation Learning via Total Correlation ObjectiveHyeongJoo Hwang, Geon-Hyeong Kim, Seunghoon Hong, Kee-Eung KimNeurIPS 2021 · 63 citations
- TVT: Three-Way Vision Transformer through Multi-Modal Hypersphere Learning for Zero-Shot Sketch-Based Image RetrievalJialin Tian, Xing Xu, Fumin Shen, Yang Yang et al.AAAI 2022 · 54 citations
- Regularized Behavior Cloning for Blocking the Leakage of Past Action InformationSeokin Seo, HyeongJoo Hwang, Hongseok Yang, Kee-Eung KimNeurIPS 2023 · 14 citations
Builds on2
- Learning Robust Representations via Multi-View Information BottleneckMarco Federici, Anjan Dutta, Patrick Forré, Nate Kushman et al.ICLR 2020 · 330 citations
- Learning Cross-Aligned Latent Embeddings for Zero-Shot Cross-Modal RetrievalKaiyi Lin, Xing Xu, Lianli Gao, Zheng Wang et al.AAAI 2020 · 50 citations
Related papers
- Multimodal Disentanglement Variational AutoEncoders for Zero-Shot Cross-Modal RetrievalJialin Tian, Kai Wang, Xing Xu, Zuo Cao et al.SIGIR 2022 · 19 citations
- Learning Invariant Representation for Unsupervised Image RestorationWenchao Du, Hu Chen, Hongyu YangCVPR 2020
- Unsupervised Video Domain Adaptation for Action Recognition: A Disentanglement PerspectivePengfei Wei, Lingdong Kong, Xinghua Qu, Yi Ren et al.NeurIPS 2023 · 39 citations
- DisenCDR: Learning Disentangled Representations for Cross-Domain RecommendationJiangxia Cao, Xixun Lin, Xin Cong, Jing Ya et al.SIGIR 2022 · 119 citations
- Domain Disentangled Generative Adversarial Network for Zero-Shot Sketch-Based 3D Shape RetrievalRui Xu, Zongyan Han, Le Hui, Jianjun Qian et al.AAAI 2022 · 25 citations
