TopoMA: Topology-Guided Multi-Agent Dense RGB 3D Reconstruction via Distributed Inference
Xuanxuan Zhang, Shuhui Shi, Tianxiang Zhang, Zhetao Guo, Zixuan Huang, You Li
Abstract
Multi-agent 3D reconstruction, as a key technology for large-scale VR/AR, robot swarms, and digital twins, has attracted growing attention. Recent end-to-end 3D reconstruction methods achieve strong performance in single-agent scenarios, but they are difficult to directly extend to multi-agent collaborative settings, where they often suffer from unstable tracking, excessive memory consumption, and frequent loop-closure failures, thus failing to meet real-time and large-scale deployment requirements. To address these issues, we propose TOPOMA, a real-time end-to-end 3D reconstruction framework tailored for multi-agent collaboration. TOPOMA explicitly models the spatial topological structure of the scene and tightly couples it with end-to-end representation learning, thereby jointly solving core challenges such as inter-agent spatial alignment and submap fusion. Concretely, we introduce topology skeleton modeling and optimization, decentralized loop closure, and topology-guided residual transport, and build upon them a fully distributed inference architecture in which each agent can independently store, reconstruct, and incrementally optimize its map while collaborating through lightweight topological information. Extensive experiments demonstrate that, compared with existing methods, TOPOMA achieves consistently higher trajectory accuracy, reconstruction quality, robustness, and topological consistency, showing superior adaptability and scalability.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1d40f82b-35b9-4bb2-8e6b-faec7fc14ca4Builds on18
- DUSt3R: Geometric 3D Vision Made EasyShuzhe Wang, Vincent Leroy, Yohann Cabon, Boris Chidlovskii et al.CVPR 2024 · 302 citations
- VGGT-SLAM: Dense RGB SLAM Optimized on the SL(4) ManifoldDominic Maggio, Hyungtae Lim, Luca CarloneNeurIPS 2025 · 176 citations
- TTT3R: 3D Reconstruction as Test-Time TrainingXingyu Chen, Yue Chen, Yuliang Xiu, Andreas Geiger et al.ICLR 2026 · 139 citations
- FastVGGT: Fast Visual Geometry TransformerYou Shen, Zhipeng Zhang, Yansong Qu, Xiawu Zheng et al.ICLR 2026 · 73 citations
- CP-SLAM: Collaborative Neural Point-based SLAM SystemJiarui Hu, Mao Mao, Hujun Bao, Guofeng Zhang et al.NeurIPS 2023 · 65 citations
Related papers
- MAC-Ego3D: Multi-Agent Gaussian Consensus for Real-Time Collaborative Ego-Motion and Photorealistic 3D ReconstructionXiaohao Xu, Feng Xue, Shibo Zhao, Yike Pan et al.CVPR 2025
- HoloScene: Simulation-Ready Interactive 3D Worlds from a Single VideoHongchi Xia, Chih-Hao Lin, Hao-Yu Hsu, Quentin Leboutet et al.NeurIPS 2025 · 18 citations
- Topologically-Aware Deformation Fields for Single-View 3D ReconstructionShivam Duggal, Deepak PathakCVPR 2022 · 30 citations
- RoboTAG: End-to-end Robot Pose Estimation via Topological Alignment GraphYifan Liu, Fangneng Zhan, Wanhua Li, Haowen Sun et al.CVPR 2026
- VGGTFace: Topologically Consistent Facial Geometry Reconstruction in the WildXin Ming, Yuxuan Han, Tianyu Huang, Feng XuAAAI 2026 · 2 citations
