CorrectNav: Self-Correction Flywheel Empowers Vision-Language-Action Navigation Model
Zhuoyuan Yu, Yuxing Long, Zihan Yang, Chengyan Zeng, Hongwei Fan, Jiyao Zhang, Hao Dong
摘要
PKU-Agibot Lab *Equal contribution, † Project Leader, ‡ Corresponding author https://correctnav.github.io Crowded Objects Avoidance Open-vocabulary Landmark Move Forward and Turn right at the human-like robot. Continue moving to stop near the yellow box. … Z-Shape Building Structure Walk down the corridor hallway in front of you and you will see an opened meeting room. Enter the … … … Walk straight along the hallway until you reach the red fire extinguisher box at the end and stop when you reach … Pedestrian Avoidance … Error Correction … in front of a white wall, turn right. Walk forward. When you see a green plant on your right front, stop. Drift Correction Walk straight and turn left in front of a wall. Walk straight and turn right at the opened door. Enter and walk to the wooden table. … … Walk until you reach the plant and turn left. Walk straight, turn left at the next corner, walk forward to the … Instruction Across Rooms Walk out of the kitchen room you are in and turn left. Move across the living room, walk to the end of the hallway and turn right .Walk into the bedroom and stop by the bed. Landmark State Change Move forward and turn right to walk through an opened doorway. … … … … … … C rrectNa Figure 1: Diverse Capabilities of CorrectNav. The model takes only monocular RGB video and language instructions as inputs, predicting navigation actions. Empowered by the Self-correction Flywheel post-training, CorrectNav not only maintains outstanding multimodal reasoning (Blue), but also displays improved deviation correction (Red), obstacle avoidance (Green), and complex action execution (Yellow).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- OmniNav: A Unified Framework for Prospective Exploration and Visual-Language NavigationXinda Xue, Junjun Hu, Minghua Luo, Xie Shichao 等ICLR 2026 · 被引用 51 次
- NavForesee: A Unified Vision-Language World Model for Hierarchical Planning and Dual-Horizon Navigation PredictionFei Liu, Shichao Xie, Minghua Luo, Zedong Chu 等CVPR 2026 · 被引用 16 次
- AwareVLN: Reasoning with Self-awareness for Vision-Language NavigationWenxuan Guo, Xiuwei Xu, Yichen Liu, Xiangyu Li 等CVPR 2026 · 被引用 7 次
- DecoVLN: Decoupling Observation, Reasoning, and Correction for Vision-and-Language NavigationZihao Xin, Wentong Li, Yixuan Jiang, Bin Wang 等CVPR 2026 · 被引用 6 次
- AdaNav: Adaptive Reasoning with Uncertainty for Vision-Language NavigationXin Ding, Jianyu Wei, Yifan Yang, Shiqi Jiang 等ICML 2026 · 被引用 6 次
它引用的顶会 Paper16
- Sigmoid Loss for Language Image Pre-TrainingXiaohua Zhai, Basil Mustafa, Alexander Kolesnikov, Lucas BeyerICCV 2023 · 被引用 2,932 次
- Room-Across-Room: Multilingual Vision-and-Language Navigation with Dense Spatiotemporal GroundingAlexander Ku, Peter Anderson, Roma Patel, Eugene Ie 等EMNLP 2020 · 被引用 208 次
- Waypoint Models for Instruction-guided Navigation in Continuous EnvironmentsJacob Krantz, Aaron Gokaslan, Dhruv Batra, Stefan Lee 等ICCV 2021 · 被引用 153 次
- Weakly-Supervised Multi-Granularity Map Learning for Vision-and-Language NavigationPeihao Chen, Dongyu Ji, Kunyang Lin, Runhao Zeng 等NeurIPS 2022 · 被引用 143 次
- Scaling Data Generation in Vision-and-Language NavigationZun Wang, Jialu Li, Yicong Hong, Yi Wang 等ICCV 2023 · 被引用 136 次
相关 Paper
- CapNav: Benchmarking Vision Language Models on Capability-conditioned Indoor NavigationXia Su, Ruiqi Chen, Benlin Liu, Jingwei Ma 等CVPR 2026 · 被引用 8 次
- LANA: A Language-Capable Navigator for Instruction Following and GenerationXiaohan Wang, Wenguan Wang, Jiayi Shao, Yi YangCVPR 2023
- OmniManip: Towards General Robotic Manipulation via Object-Centric Interaction Primitives as Spatial ConstraintsMingjie Pan, Jiyao Zhang, Tianshu Wu, Yinghao Zhao 等CVPR 2025
- Narrowing the Gap between Vision and Action in NavigationYue Zhang, Parisa KordjamshidiACM MM 2024 · 被引用 2 次
- VLN-ChEnv: Vision-language Navigation in Changeable EnvironmentsShubo Liu, Hongsheng Zhang, Qian Qiao, Qi Wu 等ACM MM 2025 · 被引用 2 次
