InfoCon: Concept Discovery with Generative and Discriminative Informativeness
Ruizhe Liu, Qian Luo, Yanchao Yang
摘要
We focus on the self-supervised discovery of manipulation concepts that can be adapted and reassembled to address various robotic tasks. We propose that the decision to conceptualize a physical procedure should not depend on how we name it (semantics) but rather on the significance of the informativeness in its representation regarding the low-level physical state and state changes. We model manipulation concepts -discrete symbols -as generative and discriminative goals and derive metrics that can autonomously link them to meaningful sub-trajectories from noisy, unlabeled demonstrations. Specifically, we employ a trainable codebook containing encodings (concepts) capable of synthesizing the end-state of a sub-trajectory given the current state -generative informativeness. Moreover, the encoding corresponding to a particular sub-trajectory should differentiate the state within and outside it and confidently predict the subsequent action based on the gradient of its discriminative score -discriminative informativeness. These metrics, which do not rely on human annotation, can be seamlessly integrated into a VQ-VAE framework, enabling the partitioning of demonstrations into semantically consistent sub-trajectories, fulfilling the purpose of discovering manipulation concepts and the corresponding sub-goal (key) states. We evaluate the effectiveness of the learned concepts by training policies that utilize them as guidance, demonstrating superior performance compared to other baselines. Additionally, our discovered manipulation concepts compare favorably to human-annotated ones while saving much manual effort. Our code is available at: https://zrllrz.github.io/InfoCon / * Work done as a research assistant at HKU.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- HiMaCon: Discovering Hierarchical Manipulation Concepts from Unlabeled Multi-Modal DataRuizhe Liu, Pei Zhou, Qian Luo, Li Sun 等NeurIPS 2025 · 被引用 2 次
- Neural Concept BinderWolfgang Stammer, Antonia Wüst, David Steinmann, Kristian KerstingNeurIPS 2024
- HuMoCon: Concept Discovery for Human Motion UnderstandingQihang Fang, Chengcheng Tang, Bugra Tekin, Shugao Ma 等CVPR 2025
- AutoCGP: Closed-Loop Concept-Guided Policies from Unlabeled DemonstrationsPei Zhou, Ruizhe Liu, Qian Luo, Fan Wang 等ICLR 2025
它引用的顶会 Paper14
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma 等NeurIPS 2022 · 被引用 22,562 次
- Decision Transformer: Reinforcement Learning via Sequence ModelingLili Chen, Kevin Lu, Aravind Rajeswaran, Kimin Lee 等NeurIPS 2021 · 被引用 2,557 次
- Learning Robot Skills with Temporal Variational InferenceTanmay Shankar, Abhinav GuptaICML 2020 · 被引用 80 次
- Masked Autoencoding for Scalable and Generalizable Decision MakingFangchen Liu, Hao Liu, Aditya Grover, Pieter AbbeelNeurIPS 2022 · 被引用 63 次
相关 Paper
- Latent Action Pretraining from VideosSeonghyeon Ye, Joel Jang, Byeongguk Jeon, Se June Joo 等ICLR 2025
- Hierarchical Goal Abstractions via Learned Subset RelationsFabian Wurzberger, Sebastian Gottwald, Zeqiang Zhang, Daniel A BraunICML 2026
- Theory and Evaluation Metrics for Learning Disentangled RepresentationsKien Do, Truyen TranICLR 2020 · 被引用 107 次
- Decoupling Meta-Reinforcement Learning with Gaussian Task Contexts and SkillsHongcai He, Anjie Zhu, Shuang Liang, Feiyu Chen 等AAAI 2024 · 被引用 6 次
- From Observation to Action: Latent Action-based Primitive Segmentation for VLA Pre-training in Industrial SettingsJiajie Zhang, Sören Schwertfeger, Alexander KleinerCVPR 2026
