Pushing It Out of the Way: Interactive Visual Navigation
Kuo-Hao Zeng, Luca Weihs, Ali Farhadi, Roozbeh Mottaghi
摘要
Figure 1 : Visual navigation may require interactions that go beyond moving forward/backward, and turning left/right. For example, the agent in the top row needs to push the chair out of its way to reach the target. Interactive navigation entails deeper understanding of the outcome of agents actions on objects in the scene. In this paper, we introduce Neural Interaction Engine (NIE) to explicitly predict the effect of actions on objects poses. By integrating NIE with our policy network we show that we can perform long-horizon planning while predicting the outcome of the actions. We evaluate NIE for visual navigation where the path to the goal is obstructed, and moving objects to specific locations in the scene and show major improvements over state of the art in these tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Selective Visual Representations Improve Convergence and Generalization for Embodied AIAinaz Eftekhar, Kuo-Hao Zeng, Jiafei Duan, Ali Farhadi 等ICLR 2024 · 被引用 28 次
- GridToPix: Training Embodied Agents with Minimal SupervisionUnnat Jain, Iou-Jen Liu, Svetlana Lazebnik, Aniruddha Kembhavi 等ICCV 2021 · 被引用 25 次
- CaMP: Causal Multi-policy Planning for Interactive Navigation in Multi-room ScenesXiaohan Wang, Yuehu Liu, Xinhang Song, Beibei Wang 等NeurIPS 2023 · 被引用 16 次
- An Interactive Navigation Method with Effect-oriented AffordanceXiaohan Wang, Yuehu Liu, Xinhang Song, Yuyi Liu 等CVPR 2024 · 被引用 1 次
- MoMa-Kitchen: A 100K+ Benchmark for Affordance-Grounded Last-Mile Navigation in Mobile ManipulationPingrui Zhang, Xianqiang Gao, Yuhan Wu, Kehui Liu 等ICCV 2025
它引用的顶会 Paper6
- Habitat: A Platform for Embodied AI ResearchManolis Savva, Jitendra Malik, Devi Parikh, Dhruv Batra 等ICCV 2019 · 被引用 1,863 次
- Object Goal Navigation using Goal-Oriented Semantic ExplorationDevendra Singh Chaplot, Dhiraj Gandhi, Abhinav Gupta, Ruslan SalakhutdinovNeurIPS 2020 · 被引用 857 次
- DD-PPO: Learning Near-Perfect PointGoal Navigators from 2.5 Billion FramesErik Wijmans, Abhishek Kadian, Ari Morcos, Stefan Lee 等ICLR 2020 · 被引用 608 次
- Learning To Explore Using Active Neural SLAMDevendra Singh Chaplot, Dhiraj Gandhi, Saurabh Gupta, Abhinav Gupta 等ICLR 2020 · 被引用 603 次
- Visual Grounding of Learned Physical ModelsYunzhu Li, Toru Lin, Kexin Yi, Daniel Bear 等ICML 2020 · 被引用 88 次
相关 Paper
- Learning Long-term Visual Dynamics with Region Proposal Interaction NetworksHaozhi Qi, Xiaolong Wang, Deepak Pathak, Yi Ma 等ICLR 2021 · 被引用 63 次
- Object-Goal Visual Navigation via Effective Exploration of Relations Among Historical Navigation StatesHeming Du, Lincheng Li, Zi Huang, Xin YuCVPR 2023
- ManipulaTHOR: A Framework for Visual Object ManipulationKiana Ehsani, Winson Han, Alvaro Herrasti, Eli VanderBilt 等CVPR 2021
- Narrowing the Gap between Vision and Action in NavigationYue Zhang, Parisa KordjamshidiACM MM 2024 · 被引用 2 次
- VTNet: Visual Transformer Network for Object Goal NavigationHeming Du, Xin Yu, Liang ZhengICLR 2021 · 被引用 36 次
