Dual-Agent Reinforcement Learning for Adaptive and Cost-Aware Visual-Inertial Odometry
Feiyang Pan, Shenghe Zheng, Chunyan Yin, Guangbin Dou
摘要
Visual-Inertial Odometry (VIO) is a critical component for robust ego-motion estimation, enabling foundational capabilities such as autonomous navigation in robotics and real-time 6-DoF tracking for augmented reality. Existing methods face a well-known trade-off: filter-based approaches are efficient but prone to drift, while optimization-based methods, though accurate, rely on computationally prohibitive Visual-Inertial Bundle Adjustment (VIBA) that is difficult to run on resource-constrained platforms. Rather than removing VIBA altogether, we aim to reduce how often and how heavily it must be invoked. To this end, we cast two key design choices in modern VIO, when to run the visual frontend and how strongly to trust its output, as sequential decision problems, and solve them with lightweight reinforcement learning (RL) agents. Our framework introduces a lightweight, dual-pronged RL policy that serves as our core contribution: (1) a Select Agent intelligently gates the entire VO pipeline based only on high-frequency IMU data; and (2) a composite Fusion Agent that first estimates a robust velocity state via a supervised network, before an RL policy adaptively fuses the full (p, v, q) state. Experiments on the EuRoC MAV and TUM-VI datasets show that, in our unified evaluation, the proposed method achieves a more favorable accuracy-efficiency-memory trade-off than prior GPU-based VO/VIO systems: it attains the best average ATE while running up to 1.77 times faster and using less GPU memory. Compared to classical optimization-based VIO systems, our approach maintains competitive trajectory accuracy while substantially reducing computational load.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- DROID-SLAM: Deep Visual SLAM for Monocular, Stereo, and RGB-D CamerasZachary Teed, Jia DengNeurIPS 2021 · 被引用 1,248 次
- Learning To Explore Using Active Neural SLAMDevendra Singh Chaplot, Dhiraj Gandhi, Saurabh Gupta, Abhinav Gupta 等ICLR 2020 · 被引用 603 次
- Deep Patch Visual OdometryZachary Teed, Lahav Lipson, Jia DengNeurIPS 2023 · 被引用 323 次
- Adaptive VIO: Deep Visual-Inertial Odometry with Online Continual LearningYouqi Pan, Wugen Zhou, Yingdian Cao, Hongbin ZhaCVPR 2024 · 被引用 17 次
- Robust Tightly-Coupled Visual-Inertial Odometry with Pre-built Maps in High Latency SituationsHujun Bao, Weijian Xie, Quanhao Qian, Danpeng Chen 等IEEE VR 2022 · 被引用 16 次
相关 Paper
- A Rotation-Translation-Decoupled Solution for Robust and Efficient Visual-Inertial InitializationYijia He, Bo Xu, Zhanpeng Ouyang, Hongdong LiCVPR 2023
- AVA-VLA: Improving Vision-Language-Action models with Active Visual AttentionLei Xiao, Jifeng Li, Juntao Gao, Feiyang Ye 等CVPR 2026 · 被引用 26 次
- UAV: A Unified and Adaptive Scheduling Framework for UAV Autopilot System with Reinforcement LearningZeying Li, shuai zhao, Chaowen Wu, Boyang Li 等ICML 2026
- Information-Driven Direct RGB-D OdometryAlejandro Fontán, Javier Civera, Rudolph TriebelCVPR 2020
- 100-Phones: A Large VI-SLAM Dataset for Augmented Reality Towards Mass Deployment on Mobile PhonesGuofeng Zhang, Jin Yuan, Haomin Liu, Zhen Peng 等IEEE VR 2024 · 被引用 5 次
