End-to-End Egospheric Spatial Memory
Daniel James Lenton, Stephen James, Ronald Clark, Andrew J. Davison
摘要
Spatial memory, or the ability to remember and recall specific locations and objects, is central to autonomous agents' ability to carry out tasks in real environments. However, most existing artificial memory modules are not very adept at storing spatial information. We propose a parameter-free module, Egospheric Spatial Memory (ESM), which encodes the memory in an ego-sphere around the agent, enabling expressive 3D representations. ESM can be trained end-to-end via either imitation or reinforcement learning, and improves both training efficiency and final performance against other memory baselines on both drone and manipulator visuomotor control tasks. The explicit egocentric geometry also enables us to seamlessly combine the learned controller with other non-learned modalities, such as local obstacle avoidance. We further show applications to semantic segmentation on the ScanNet dataset, where ESM naturally combines image-level and map-level inference modalities. Through our broad set of experiments, we show that ESM provides a general computation graph for embodied spatial reasoning, and the module forms a bridge between real-time mapping systems and differentiable memory architectures. Implementation at: https://github.com/ivy-dl/memory .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper2
相关 Paper
- Semantic MapNet: Building Allocentric Semantic Maps and Representations from Egocentric ViewsVincent Cartillier, Zhile Ren, Neha Jain, Stefan Lee 等AAAI 2021 · 被引用 89 次
- Scalable Spatial Memory for Scene Rendering and NavigationWen-Cheng Chen, Chu-Song Chen, Wei-Chen Chiu, Min-Chun HuAAAI 2023 · 被引用 1 次
- Keep It in Mind: User Centric Continual Spatial Intelligence Reasoning in Egocentric Video StreamsYun Wang, Junbin Xiao, Han Lyu, Yifan Wang 等ICML 2026
- SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Modelsjing wu, Jianhua Wu, Jiayi Guan, Jiahong Chen 等ICML 2026
- Multi-Object Navigation with dynamically learned neural implicit representationsPierre Marza, Laëtitia Matignon, Olivier Simonin, Christian WolfICCV 2023 · 被引用 32 次
