Adaptive Discrete Communication Bottlenecks with Dynamic Vector Quantization for Heterogeneous Representational Coarseness
Dianbo Liu, Alex Lamb, Xu Ji, Pascal Tikeng Notsawo Jr., Michael Mozer, Yoshua Bengio, Kenji Kawaguchi
摘要
Vector Quantization (VQ) is a method for discretizing latent representations and has become a major part of the deep learning toolkit. It has been theoretically and empirically shown that discretization of representations leads to improved generalization, including in reinforcement learning where discretization can be used to bottleneck multi-agent communication to promote agent specialization and robustness. The discretization tightness of most VQ-based methods is defined by the number of discrete codes in the representation vector and the codebook size, which are fixed as hyperparameters. In this work, we propose learning to dynamically select discretization tightness conditioned on inputs, based on the hypothesis that data naturally contains variations in complexity that call for different levels of representational coarseness which is observed in many heterogeneous data sets. We show that dynamically varying tightness in communication bottlenecks can improve model performance on visual reasoning and reinforcement learning tasks with heterogeneity in representations.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper3
- Recurrent Independent MechanismsAnirudh Goyal, Alex Lamb, Jordan Hoffmann, Shagun Sodhani 等ICLR 2021 · 被引用 357 次
- Contrastive Learning of Structured World ModelsThomas N. Kipf, Elise van der Pol, Max WellingICLR 2020 · 被引用 322 次
- Coordination Among Neural Modules Through a Shared Global WorkspaceAnirudh Goyal, Aniket Rajiv Didolkar, Alex Lamb, Kartikeya Badola 等ICLR 2022 · 被引用 114 次
相关 Paper
- Trading off Utility, Informativeness, and Complexity in Emergent CommunicationMycal Tucker, Roger Levy, Julie A. Shah, Noga ZaslavskyNeurIPS 2022 · 被引用 34 次
- Discrete-Valued Neural CommunicationDianbo Liu, Alex Lamb, Kenji Kawaguchi, Anirudh Goyal 等NeurIPS 2021 · 被引用 55 次
- The Variational Bandwidth Bottleneck: Stochastic Evaluation on an Information BudgetAnirudh Goyal, Yoshua Bengio, Matthew M. Botvinick, Sergey LevineICLR 2020 · 被引用 26 次
- Discrete Compositional Representations as an Abstraction for Goal Conditioned Reinforcement LearningRiashat Islam, Hongyu Zang, Anirudh Goyal, Alex Lamb 等NeurIPS 2022 · 被引用 10 次
- A Consciousness-Inspired Planning Agent for Model-Based Reinforcement LearningMingde Zhao, Zhen Liu, Sitao Luan, Shuyuan Zhang 等NeurIPS 2021 · 被引用 41 次
