MACARONS: Mapping and Coverage Anticipation with RGB Online Self-Supervision
Antoine Guédon, Tom Monnier, Pascal Monasse, Vincent Lepetit
Abstract
https://imagine.enpc.fr/ ˜guedona/MACARONS/ (a) NBV methods with a depth sensor (e.g., [27]) (b) Our approach MACARONS with an RGB sensor Figure 1. Mapping and Coverage Anticipation with RGB Online Self-Supervision. (a) NBV methods such as [27] rely on a depth sensor to perform path planning (bottom) and scan the environment (top). They need to be trained with explicit 3D supervision, generally on small-scale meshes. (b) Our approach MACARONS instead simultaneously learns to efficiently explore the scene and to reconstruct it (top) using an RGB sensor only. Its self-supervised, online learning process scales to large-scale and complex scenes.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- ActiveGrasp: Information-Guided Active Grasping with Calibrated Energy-based ModelBoshu Lei, Wen Jiang, Kostas DaniilidisCVPR 2026 · 5 citations
- MAGICIAN: Efficient Long-Term Planning with Imagined Gaussians for Active MappingShiyao Li, Antoine Guédon, Shizhe Chen, Vincent LepetitCVPR 2026 · 5 citations
- Active Learning of 3D Gaussian Splatting with Consistent Region Partition and Robust Pose EstimationRuiqi Li, Yiu-ming CheungICLR 2026
- Dr. Splat: Directly Referring 3D Gaussian Splatting via Direct Language Embedding RegistrationKim Jun-Seong, GeonU Kim, Kim Yu-Ji, Yu-Chiang Frank Wang et al.CVPR 2025
- Paparazzo: Active Mapping of Moving 3D ObjectsDavide Allegro, Shiyao Li, Stefano Ghidoni, Vincent LepetitCVPR 2026
Builds on14
- Digging Into Self-Supervised Monocular Depth EstimationClément Godard, Oisin Mac Aodha, Michael Firman, Gabriel J. BrostowICCV 2019 · 2,416 citations
- Depth From Videos in the Wild: Unsupervised Monocular Depth Learning From Unknown CamerasAriel Gordon, Hanhan Li, Rico Jonschkowski, Anelia AngelovaICCV 2019 · 397 citations
- Consistent video depth estimationXuan Luo, Jia-Bin Huang, Richard Szeliski, Kevin Matzen et al.SIGGRAPH 2020 · 321 citations
- Self-Supervised Learning With Geometric Constraints in Monocular Video: Connecting Flow, Depth, and CameraYuhua Chen, Cordelia Schmid, Cristian SminchisescuICCV 2019 · 265 citations
- Exploiting Temporal Consistency for Real-Time Video Depth EstimationHaokui Zhang, Ying Li, Yuanzhouhan Cao, Yu Liu et al.ICCV 2019 · 137 citations
Related papers
- SCONE: Surface Coverage Optimization in Unknown Environments by Volumetric IntegrationAntoine Guédon, Pascal Monasse, Vincent LepetitNeurIPS 2022 · 26 citations
- ScanBot: Autonomous Reconstruction via Deep Reinforcement LearningHezhi Cao, Xi Xia, Guan Wu, Ruizhen Hu et al.SIGGRAPH 2023 · 13 citations
- SelfOcc: Self-Supervised Vision-Based 3D Occupancy PredictionYuanhui Huang, Wenzhao Zheng, Borui Zhang, Jie Zhou et al.CVPR 2024
- Self-Supervised Learning of Depth Inference for Multi-View StereoJiayu Yang, José M. Álvarez, Miaomiao LiuCVPR 2021
- MonoDream: Monocular Vision-Language Navigation with Panoramic DreamingShuo Wang, Yongcai Wang, Zhaoxin Fan, Yucheng Wang et al.AAAI 2026 · 11 citations
