Task Programming: Learning Data Efficient Behavior Representations
Jennifer J. Sun, Ann Kennedy, Eric Zhan, David J. Anderson, Yisong Yue, Pietro Perona
Abstract
Specialized domain knowledge is often necessary to accurately annotate training sets for in-depth analysis, but can be burdensome and time-consuming to acquire from domain experts. This issue arises prominently in automated behavior analysis, in which agent movements or actions of interest are detected from video tracking data. To reduce annotation effort, we present TREBA: a method to learn annotation-sample efficient trajectory embedding for behavior analysis, based on multi-task self-supervised learning. The tasks in our method can be efficiently engineered by domain experts through a process we call "task programming", which uses programs to explicitly encode structured knowledge from domain experts. Total domain expert effort can be reduced by exchanging data annotation time for the construction of a small number of programmed tasks. We evaluate this trade-off using data from behavioral neuroscience, in which specialized domain knowledge is used to identify behaviors. We present experimental results in three datasets across two domains: mice and fruit flies. Using embeddings from TREBA, we reduce annotation burden by up to a factor of 10 without compromising accuracy compared to state-of-the-art features. Our results thus suggest that task programming and self-supervision can be an effective way to reduce annotation effort for domain experts. 2 × 10 1 3 × 10 1 4 × 10 1 Error (Log Scale) MARS Keypoints with Pre-Training Variations Keypoints + TVAE (MARS) Keypoints + TVAE (Mouse100) Keypoints + Programs (MARS) Keypoints + Programs (Mouse100)
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6c2c5749-6dfa-4a41-8978-c0bac6f11bd8Cited by top-tier papers14
- VideoPrism: A Foundational Visual Encoder for Video UnderstandingLong Zhao, Nitesh Bharadwaj Gundavarapu, Liangzhe Yuan, Hao Zhou et al.ICML 2024 · 91 citations
- AmadeusGPT: a natural language interface for interactive animal behavioral analysisShaokai Ye, Jessy Lauer, Mu Zhou, Alexander Mathis et al.NeurIPS 2023 · 39 citations
- Utilizing Expert Features for Contrastive Learning of Time-Series RepresentationsManuel T. Nonnenmacher, Lukas Oldenburg, Ingo Steinwart, David ReebICML 2022 · 27 citations
- Self-Supervised Keypoint Discovery in Behavioral VideosJennifer J. Sun, Serim Ryou, Roni H. Goldshmid, Brandon Weissbourd et al.CVPR 2022 · 24 citations
- MABe22: A Multi-Species Multi-Task Benchmark for Learned Representations of BehaviorJennifer J. Sun, Markus Marks, Andrew Wesley Ulmer, Dipam Chakraborty et al.ICML 2023 · 20 citations
Builds on8
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Supervised Contrastive LearningPrannay Khosla, Piotr Teterwak, Chen Wang, Aaron Sarna et al.NeurIPS 2020 · 7,049 citations
- Big Self-Supervised Models are Strong Semi-Supervised LearnersTing Chen, Simon Kornblith, Kevin Swersky, Mohammad Norouzi et al.NeurIPS 2020 · 2,611 citations
- VideoBERT: A Joint Model for Video and Language Representation LearningChen Sun, Austin Myers, Carl Vondrick, Kevin Murphy et al.ICCV 2019 · 1,396 citations
- Scaling and Benchmarking Self-Supervised Visual Representation LearningPriya Goyal, Dhruv Mahajan, Abhinav Gupta, Ishan MisraICCV 2019 · 429 citations
Related papers
- Automatic Synthesis of Diverse Weak Supervision Sources for Behavior AnalysisAlbert Tseng, Jennifer J. Sun, Yisong YueCVPR 2022 · 6 citations
- Mouse2Vec: Learning Reusable Semantic Representations of Mouse BehaviourGuanhua Zhang, Zhiming Hu, Mihai Bâce, Andreas BullingCHI 2024 · 5 citations
- A Systematic Study of Behavioral Cloning for Scientific Data AnnotationIshaan Singh Chandok, Core Francisco ParkICML 2026
- TrackMAE: Video Representation Learning via Track Mask and PredictRenaud Vandeghen, Fida Mohammad Thoker, Marc Van Droogenbroeck, Bernard GhanemCVPR 2026 · 3 citations
- Learning Disentangled Behavior EmbeddingsChanghao Shi, Sivan Schwartz, Shahar Levy, Shay Achvat et al.NeurIPS 2021 · 13 citations
