Learning the Predictability of the Future
Didac Suris, Ruoshi Liu, Carl Vondrick
Abstract
We introduce a framework for learning from unlabeled video what is predictable in the future. Instead of committing up front to features to predict, our approach learns from data which features are predictable. Based on the observation that hyperbolic geometry naturally and compactly encodes hierarchical structure, we propose a predictive model in hyperbolic space. When the model is most confident, it will predict at a concrete level of the hierarchy, but when the model is not confident, it learns to automatically select a higher level of abstraction. Experiments on two established datasets show the key role of hierarchical representations for action prediction. Although our representation is trained with unlabeled video, visualizations show that action hierarchies emerge in the representation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext eaa33219-4e8a-45e3-bd15-1f7faa4c9c42Cited by top-tier papers26
- Hyperbolic Image SegmentationMina Ghadimi Atigh, Julian Schoep, Erman Acar, Nanne van Noord et al.CVPR 2022 · 70 citations
- Hyperbolic Busemann Learning with Ideal PrototypesMina Ghadimi Atigh, Martin Keller-Ressel, Pascal MettesNeurIPS 2021 · 68 citations
- Visual Abductive ReasoningChen Liang, Wenguan Wang, Tianfei Zhou, Yi YangCVPR 2022 · 50 citations
- MOMA: Multi-Object Multi-Actor Activity ParsingZelun Luo, Wanze Xie, Siddharth Kapoor, Yiyun Liang et al.NeurIPS 2021 · 34 citations
- The Euclidean Space is Evil: Hyperbolic Attribute Editing for Few-shot Image GenerationLingxiao Li, Yi Zhang, Shuhui WangICCV 2023 · 27 citations
Builds on9
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Hyperbolic Neural Networks++Ryohei Shimizu, Yusuke Mukuta, Tatsuya HaradaICLR 2021 · 791 citations
- Improved Conditional VRNNs for Video PredictionLluís Castrejón, Nicolas Ballas, Aaron C. CourvilleICCV 2019 · 177 citations
- VideoFlow: A Conditional Flow-Based Model for Stochastic Video GenerationManoj Kumar, Mohammad Babaeizadeh, Dumitru Erhan, Chelsea Finn et al.ICLR 2020 · 142 citations
- Predicting the Future: A Jointly Learnt Model for Action AnticipationHarshala Gammulle, Simon Denman, Sridha Sridharan, Clinton FookesICCV 2019 · 93 citations
Related papers
- Searching for Actions on the HyperboleTeng Long, Pascal Mettes, Heng Tao Shen, Cees G. M. SnoekCVPR 2020
- Unsupervised Hyperbolic Representation Learning via Message Passing Auto-EncodersJiwoong Park, Junho Cho, Hyung Jin Chang, Jin Young ChoiCVPR 2021
- Beyond Euclidean: Dual-Space Representation Learning for Weakly Supervised Video Violence DetectionJiaxu Leng, Zhanjie Wu, Mingpi Tan, Yiran Liu et al.NeurIPS 2024 · 22 citations
- Revisiting Hierarchical Approach for Persistent Long-Term Video PredictionWonkwang Lee, Whie Jung, Han Zhang, Ting Chen et al.ICLR 2021 · 29 citations
- Hyperbolic Contrastive Learning for Visual Representations beyond ObjectsSongwei Ge, Shlok Mishra, Simon Kornblith, Chun-Liang Li et al.CVPR 2023
