Relax, it doesn't matter how you get there: A new self-supervised approach for multi-timescale behavior analysis
Mehdi Azabou, Michael Mendelson, Nauman Ahad, Maks Sorokin, Shantanu Thakoor, Carolina Urzay, Eva L. Dyer
摘要
Natural behavior consists of dynamics that are complex and unpredictable, especially when trying to predict many steps into the future. While some success has been found in building representations of behavior under constrained or simplified task-based conditions, many of these models cannot be applied to free and naturalistic settings where behavior becomes increasingly hard to model. In this work, we develop a multi-task representation learning model for behavior that combines two novel components: (i) an action-prediction objective that aims to predict the distribution of actions over future timesteps, and (ii) a multi-scale architecture that builds separate latent spaces to accommodate short-and long-term dynamics. After demonstrating the ability of the method to build representations of both local and global dynamics in realistic robots in varying environments and terrains, we apply our method to the MABe 2022 Multi-agent behavior challenge, where our model ranks first overall (top rank over all 13 tasks), first on all sequence-level tasks, and 1st or 2nd on 7 out of 9 frame-level tasks. In all of these cases, we show that our model can build representations that capture the many different factors that drive behavior and solve a wide range of downstream tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Animal behavioral analysis and neural encoding with transformer-based self-supervised pretrainingYanchen Wang, Han Yu, Ari Blau, Yizi Zhang 等ICLR 2026 · 被引用 8 次
- BaRISTA: Brain Scale Informed Spatiotemporal Representation of Human Intracranial Neural ActivityLucine L. Oganesian, Saba Hashemi, Maryam M. ShanechiNeurIPS 2025 · 被引用 7 次
- Disentangling 3D Animal Pose Dynamics with Scrubbed Conditional Latent VariablesJoshua Huang Wu, Hari Koneru, James Russell Ravenel, Anshuman Sabath 等ICLR 2025
- Dynamic Compression Flows for Neuroscience DataGanchao Wei, Daniela de Albuquerque, Miles Martinez, Shiyang Pan 等ICML 2026
它引用的顶会 Paper16
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- Contrastive Multi-View Representation Learning on GraphsKaveh Hassani, Amir Hosein Khas AhmadiICML 2020 · 被引用 1,663 次
- Perceiver: General Perception with Iterative AttentionAndrew Jaegle, Felix Gimeno, Andy Brock, Oriol Vinyals 等ICML 2021 · 被引用 1,399 次
相关 Paper
- Object-Oriented Dynamics Learning through Multi-Level AbstractionGuangxiang Zhu, Jianhao Wang, Zhizhou Ren, Zichuan Lin 等AAAI 2020 · 被引用 6 次
- MABe22: A Multi-Species Multi-Task Benchmark for Learned Representations of BehaviorJennifer J. Sun, Markus Marks, Andrew Wesley Ulmer, Dipam Chakraborty 等ICML 2023 · 被引用 20 次
- From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action ModelBing Hu, Zaijing Li, Rui Shao, Junda Chen 等ICML 2026 · 被引用 4 次
- A System for Morphology-Task Generalization via Unified Representation and Behavior DistillationHiroki Furuta, Yusuke Iwasawa, Yutaka Matsuo, Shixiang Shane GuICLR 2023
- Bootstrap Latent-Predictive Representations for Multitask Reinforcement LearningZhaohan Daniel Guo, Bernardo Ávila Pires, Bilal Piot, Jean-Bastien Grill 等ICML 2020 · 被引用 153 次
