PIPHEN: Physical Interaction Prediction with Hamiltonian Energy Networks
Kewei Chen, Yayu Long, Mingsheng Shang
摘要
Multi-robot systems in complex physical collaborations face a "shared brain dilemma": transmitting high-dimensional multimedia data (e.g., video streams at 30MB/s) creates severe bandwidth bottlenecks and decision-making latency. To address this, we propose PIPHEN, an innovative distributed physical cognition-control framework. Its core idea is to replace "raw data communication" with "semantic communication" by performing "semantic distillation" at the robot edge, reconstructing high-dimensional perceptual data into compact, structured physical representations. This idea is primarily realized through two key components: (1) a novel Physical Interaction Prediction Network (PIPN), derived from large model knowledge distillation, to generate this representation; and (2) a Hamiltonian Energy Network (HEN) controller, based on energy conservation, to precisely translate this representation into coordinated actions. Experiments show that, compared to baseline methods, PIPHEN can compress the information representation to less than 5% of the original data volume and reduce collaborative decision-making latency from 315ms to 76ms, while significantly improving task success rates. This work provides a fundamentally efficient paradigm for resolving the "shared brain dilemma" in resource-constrained multi-robot systems.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper5
- Contrastive Learning of Structured World ModelsThomas N. Kipf, Elise van der Pol, Max WellingICLR 2020 · 被引用 322 次
- SimVP: Simpler yet Better Video PredictionZhangyang Gao, Cheng Tan, Lirong Wu, Stan Z. LiCVPR 2022 · 被引用 313 次
- Building Cooperative Embodied Agents Modularly with Large Language ModelsHongxin Zhang, Weihua Du, Jiaming Shan, Qinhong Zhou 等ICLR 2024 · 被引用 303 次
- Hermes: an efficient federated learning framework for heterogeneous mobile clientsAng Li, Jingwei Sun, Pengcheng Li, Yu Pu 等MobiCom 2021 · 被引用 167 次
- 3D-IntPhys: Towards More Generalized 3D-grounded Visual Intuitive Physics under Challenging ScenesHaotian Xue, Antonio Torralba, Josh Tenenbaum, Dan Yamins 等NeurIPS 2023 · 被引用 19 次
相关 Paper
- Pixel2Phys: Distilling Governing Laws from Visual DynamicsRuikun Li, Jun Yao, Yingfan Hua, Shixiang Tang 等CVPR 2026 · 被引用 2 次
- A Multi-Agent View of Wireless Video Streaming with Delayed Client-FeedbackNouman Khan, Ujwal Dinesha, Subrahmanyam Arunachalam, Dheeraj Narasimha 等INFOCOM 2024 · 被引用 1 次
- Neurosymbolic Transformers for Multi-Agent CommunicationJeevana Priya Inala, Yichen Yang, James Paulos, Yewen Pu 等NeurIPS 2020 · 被引用 29 次
- ComAI: Enabling Lightweight, Collaborative Intelligence by Retrofitting Vision DNNsKasthuri Jayarajah, Dhanuja Wanniarachchige, Tarek F. Abdelzaher, Archan MisraINFOCOM 2022 · 被引用 9 次
- Learning Distilled Collaboration Graph for Multi-Agent PerceptionYiming Li, Shunli Ren, Pengxiang Wu, Siheng Chen 等NeurIPS 2021 · 被引用 464 次
