Guided Distillation and Risk Adaptive Evolution for Multi-Robot Navigation
Xuyang Li, Jianwu Fang, Lin Li, Boyuan Chen, Guangliang Li, Jianru Xue
Abstract
Recent advancements in multi-robot navigation have explored methods that combine Large Language Models (LLMs) for tasks like scene understanding or high-level decision-making. However, these approaches face challenges with high inference latency and potential hallucinations. To address these challenges, we propose a knowledge-driven Reinforcement Learning (RL) framework, GUIDER, that utilizes an LLM in two different offline roles. First, we leverage the LLM as an offline knowledge source. Its expertise is distilled into a compact model, which is applied only when the RL agent is uncertain about its own value estimates and the model itself is confident in its prediction. Additionally, we utilize the LLM as an offline semantic engine. This process translates the LLM's high-level understanding of situational risk into a dynamic adjustment of the RL agent's behavioral style, evolving a function that optimally balances conservative and aggressive actions. We conduct extensive experiments in both terrestrial and maritime settings. Across all maritime scenarios (3–12 robots), GUIDER improves the task success rate and reduces the collision rate significantly compared to the state-of-the-art RL-based multi-robot navigation methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 10dcb726-de9c-40d5-bc01-55969c9f9847Builds on3
- DiLu: A Knowledge-Driven Approach to Autonomous Driving with Large Language ModelsLicheng Wen, Daocheng Fu, Xin Li, Xinyu Cai et al.ICLR 2024 · 255 citations
- Describe, Explain, Plan and Select: Interactive Planning with LLMs Enables Open-World Multi-Task AgentsZihao Wang, Shaofei Cai, Guanzhou Chen, Anji Liu et al.NeurIPS 2023 · 178 citations
- Enhancing Multi-Robot Semantic Navigation Through Multimodal Chain-of-Thought Score CollaborationZhixuan Shen, Haonan Luo, Kexun Chen, Fengmao Lv et al.AAAI 2025 · 21 citations
Related papers
- KALM: Knowledgeable Agents by Offline Reinforcement Learning from Large Language Model RolloutsJing-Cheng Pang, Si-Hang Yang, Kaiyuan Li, Jiaji Zhang et al.NeurIPS 2024 · 12 citations
- Hybrid-Driving: An Autonomous Driving Decision Framework Integrating Large Language Models, Knowledge Graphs and Driving RulesJiabao Wang, Zepeng Wu, Qian Dong, Lingzhong Meng et al.AAAI 2025 · 3 citations
- AutoGuide: Automated Generation and Selection of Context-Aware Guidelines for Large Language Model AgentsYao Fu, Dong-Ki Kim, Jaekyeom Kim, Sungryull Sohn et al.NeurIPS 2024 · 80 citations
- Divide and Conquer: Grounding LLMs as Efficient Decision-Making Agents via Offline Hierarchical Reinforcement LearningZican Hu, Wei Liu, Xiaoye Qu, Xiangyu Yue et al.ICML 2025
- Mitigating Object Hallucination in Large Vision-Language Models via Image-Grounded GuidanceLinxi Zhao, Yihe Deng, Weitong Zhang, Quanquan GuICML 2025
