Competing Mutual Information Constraints with Stochastic Competition-Based Activations for Learning Diversified Representations
Konstantinos P. Panousis, Anastasios Antoniadis, Sotirios Chatzis
摘要
This work aims to address the long-established problem of learning diversified representations. To this end, we combine information-theoretic arguments with stochastic competition-based activations, namely Stochastic Local Winner-Takes-All (LWTA) units. In this context, we ditch the conventional deep architectures commonly used in Representation Learning, that rely on non-linear activations; instead, we replace them with sets of locally and stochastically competing linear units. In this setting, each network layer yields sparse outputs, determined by the outcome of the competition between units that are organized into blocks of competitors. We adopt stochastic arguments for the competition mechanism, which perform posterior sampling to determine the winner of each block. We further endow the considered networks with the ability to infer the sub-part of the network that is essential for modeling the data at hand; we impose appropriate stick-breaking priors to this end. To further enrich the information of the emerging representations, we resort to information-theoretic principles, namely the Information Competing Process (ICP). Then, all the components are tied together under the stochastic Variational Bayes framework for inference. We perform a thorough experimental investigation for our approach using benchmark datasets on image classification. As we experimentally show, the resulting networks yield significant discriminative representation learning abilities. In addition, the introduced paradigm allows for a principled investigation mechanism of the emerging intermediate network representations.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper1
相关 Paper
- Stochastic Deep Networks with Linear Competing Units for Model-Agnostic Meta-LearningKonstantinos Kalais, Sotirios ChatzisICML 2022 · 被引用 10 次
- Bayesian Posterior Approximation With Stochastic EnsemblesOleksandr Balabanov, Bernhard Mehlig, Hampus LinanderCVPR 2023
- What should a neuron aim for? Designing local objective functions based on information theoryAndreas Christian Schneider, Valentin Neuhaus, David Alexander Ehrlich, Abdullah Makkeh 等ICLR 2025
- Usable Information and Evolution of Optimal Representations During TrainingMichael Kleinman, Alessandro Achille, Daksh Idnani, Jonathan C. KaoICLR 2021 · 被引用 16 次
- Stochastic Filter Groups for Multi-Task CNNs: Learning Specialist and Generalist Convolution KernelsFelix J. S. Bragman, Ryutaro Tanno, Sébastien Ourselin, Daniel C. Alexander 等ICCV 2019 · 被引用 97 次
