Learning Interpretable Low-dimensional Representation via Physical Symmetry
Xuanjie Liu, Daniel Chin, Yichen Huang, Gus Xia
摘要
We have recently seen great progress in learning interpretable music representations, ranging from basic factors, such as pitch and timbre, to high-level concepts, such as chord and texture. However, most methods rely heavily on music domain knowledge. It remains an open question what general computational principles give rise to interpretable representations, especially low-dim factors that agree with human perception. In this study, we take inspiration from modern physics and use physical symmetry as a self consistency constraint for the latent space of time-series data. Specifically, it requires the prior model that characterises the dynamics of the latent states to be equivariant with respect to certain group transformations. We show that physical symmetry leads the model to learn a linear pitch factor from unlabelled monophonic music audio in a self-supervised fashion. In addition, the same methodology can be applied to computer vision, learning a 3D Cartesian space from videos of a simple moving object without labels. Furthermore, physical symmetry naturally leads to counterfactual representation augmentation, a new technique which improves sample efficiency.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper9
- Spatio-temporal Self-Supervised Representation Learning for 3D Point CloudsSiyuan Huang, Yichen Xie, Song-Chun Zhu, Yixin ZhuICCV 2021 · 被引用 259 次
- GRF: Learning a General Radiance Field for 3D Representation and RenderingAlex Trevithick, Bo YangICCV 2021 · 被引用 258 次
- Towards Nonlinear Disentanglement in Natural Data with Temporal Sparse CodingDavid A. Klindt, Lukas Schott, Yash Sharma, Ivan Ustyuzhaninov 等ICLR 2021 · 被引用 156 次
- Equivariant Neural RenderingEmilien Dupont, Miguel Bautista Martin, Alex Colburn, Aditya Sankar 等ICML 2020 · 被引用 69 次
- Contrastively Disentangled Sequential Variational AutoencoderJunwen Bai, Weiran Wang, Carla P. GomesNeurIPS 2021 · 被引用 60 次
相关 Paper
- Learning Disentangled Representations and Group Structure of Dynamical EnvironmentsRobin Quessard, Thomas D. Barrett, William R. ClementsNeurIPS 2020 · 被引用 53 次
- Unsupervised Learning of Equivariant Structure from SequencesTakeru Miyato, Masanori Koyama, Kenji FukumizuNeurIPS 2022 · 被引用 17 次
- Structuring Representations Using Group InvariantsMehran Shakerinava, Arnab Kumar Mondal, Siamak RavanbakhshNeurIPS 2022 · 被引用 23 次
- Latent Mixture of Symmetries for Sample-Efficient Dynamic LearningHaoran Li, Chenhan Xiao, Muhao Guo, Yang WengNeurIPS 2025 · 被引用 7 次
- Homomorphism AutoEncoder - Learning Group Structured Representations from Observed TransitionsHamza Keurti, Hsiao-Ru Pan, Michel Besserve, Benjamin F. Grewe 等ICML 2023 · 被引用 22 次
