Self-Supervised Keypoint Discovery in Behavioral Videos
Jennifer J. Sun, Serim Ryou, Roni H. Goldshmid, Brandon Weissbourd, John O. Dabiri, David J. Anderson, Ann Kennedy, Yisong Yue, Pietro Perona
摘要
We propose a method for learning the posture and structure of agents from unlabelled behavioral videos. Starting from the observation that behaving agents are generally the main sources of movement in behavioral videos, our method, Behavioral Keypoint Discovery (B-KinD), uses an encoder-decoder architecture with a geometric bottleneck to reconstruct the spatiotemporal difference between video frames. By focusing only on regions of movement, our approach works directly on input videos without requiring manual annotations. Experiments on a variety of agent types (mouse, fly, human, jellyfish, and trees) demonstrate the generality of our approach and reveal that our discovered keypoints represent semantically meaningful body parts, which achieve state-of-the-art performance on keypoint regression among self-supervised methods. Additionally, B-KinD achieve comparable performance to supervised keypoints on downstream tasks, such as behavior classification, suggesting that our method can dramatically reduce model training costs vis-a-vis supervised methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- AutoLink: Self-supervised Learning of Human Skeletons and Object Outlines by Linking KeypointsXingzhe He, Bastian Wandt, Helge RhodinNeurIPS 2022 · 被引用 28 次
- 3D Implicit Transporter for Temporally Consistent Keypoint DiscoveryChengliang Zhong, Yuhang Zheng, Yupeng Zheng, Hao Zhao 等ICCV 2023 · 被引用 23 次
- Relax, it doesn't matter how you get there: A new self-supervised approach for multi-timescale behavior analysisMehdi Azabou, Michael Mendelson, Nauman Ahad, Maks Sorokin 等NeurIPS 2023 · 被引用 18 次
- Latent Particle World Models: Self-supervised Object-centric Stochastic Dynamics ModelingTal Daniel, Carl Qi, Dan Haramati, Amir Zadeh 等ICLR 2026 · 被引用 12 次
- Pose Prior Learner: Unsupervised Categorical Prior Learning for Pose EstimationZiyu Wang, Shuangpeng Han, Mengmi ZhangICLR 2026 · 被引用 3 次
它引用的顶会 Paper6
- Anchor Loss: Modulating Loss Scale Based on Prediction DifficultySerim Ryou, Seong-Gyun Jeong, Pietro PeronaICCV 2019 · 被引用 46 次
- Normalized Human Pose Features for Human Action Video AlignmentJingyuan Liu, Mingyi Shi, Qifeng Chen, Hongbo Fu 等ICCV 2021 · 被引用 16 次
- Unsupervised Human Pose Estimation Through Transforming Shape TemplatesLuca Schmidtke, Athanasios Vlontzos, Simon Ellershaw, Anna Lukens 等CVPR 2021
- Self-Supervised Learning of Interpretable Keypoints From Unlabelled VideosTomas Jakab, Ankush Gupta, Hakan Bilen, Andrea VedaldiCVPR 2020
- Task Programming: Learning Data Efficient Behavior RepresentationsJennifer J. Sun, Ann Kennedy, Eric Zhan, David J. Anderson 等CVPR 2021
相关 Paper
- BKinD-3D: Self-Supervised 3D Keypoint Discovery from Multi-View VideosJennifer J. Sun, Lili Karashchuk, Amil Dravid, Serim Ryou 等CVPR 2023
- PREDICT & CLUSTER: Unsupervised Skeleton Based Action RecognitionKun Su, Xiulong Liu, Eli ShlizermanCVPR 2020
- Unsupervised Volumetric AnimationAliaksandr Siarohin, Willi Menapace, Ivan Skorokhodov, Kyle Olszewski 等CVPR 2023
- Motion Representations for Articulated AnimationAliaksandr Siarohin, Oliver J. Woodford, Jian Ren, Menglei Chai 等CVPR 2021
- Deep Graph Pose: a semi-supervised deep graphical model for improved animal pose trackingAnqi Wu, Estefany Kelly Buchanan, Matthew R. Whiteway, Michael Schartner 等NeurIPS 2020 · 被引用 61 次
