Towards Reliable Code-as-Policies: A Neuro-Symbolic Framework for Embodied Task Planning
Sanghyun Ahn, Wonje Choi, Junyong Lee, Jinwoo Park, Honguk Woo
Abstract
Recent advances in large language models (LLMs) have enabled the automatic generation of executable code for task planning and control in embodied agents such as robots, demonstrating the potential of LLM-based embodied intelligence. However, these LLM-based code-as-policies approaches often suffer from limited environmental grounding, particularly in dynamic or partially observable settings, leading to suboptimal task success rates due to incorrect or incomplete code generation. In this work, we propose a neuro-symbolic embodied task planning framework that incorporates explicit symbolic verification and interactive validation processes during code generation. In the validation phase, the framework generates exploratory code that actively interacts with the environment to acquire missing observations while preserving task-relevant states. This integrated process enhances the grounding of generated code, resulting in improved task reliability and success rates in complex environments. We evaluate our framework on RLBench and in real-world settings across dynamic, partially observable scenarios. Experimental results demonstrate that our framework improves task success rates by 46.2% over Code-as-Policies baselines and attains over 86.8% executability of task-relevant actions, thereby enhancing the reliability of task planning in dynamic environments.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b5d72b2f-41f4-4585-89b2-ca0746d26b0fCited by top-tier papers4
- Embodied Interpretability: Linking Causal Understanding to Generalization in Vision-Language-Action ModelsHanxin Zhang, Mingshuo Xu, Abdulqader Dhafer, Shigang Yue et al.ICML 2026 · 2 citations
- DecoVer: A Decompose-and-Verify Neuro-Symbolic Framework for Embodied Task Planning with BC+YiXiang Jiang, Binqian Xu, Xiangbo ShuICML 2026
- Cross-Domain Demo-to-Code via Neurosymbolic Counterfactual ReasoningJooyoung Kim, Wonje Choi, Younguk Song, Honguk WooCVPR 2026
- Efficient Skill Grounding via Code Refactoring with Small Language ModelsSera Choi, Wonje Choi, Saehun Chun, Daehee Lee et al.ICML 2026
Builds on17
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma et al.NeurIPS 2022 · 22,562 citations
- PaLM-E: An Embodied Multimodal Language ModelDanny Driess, Fei Xia, Mehdi S. M. Sajjadi, Corey Lynch et al.ICML 2023 · 2,601 citations
- Language Models as Zero-Shot Planners: Extracting Actionable Knowledge for Embodied AgentsWenlong Huang, Pieter Abbeel, Deepak Pathak, Igor MordatchICML 2022 · 1,539 citations
- LLM-Planner: Few-Shot Grounded Planning for Embodied Agents with Large Language ModelsChan Hee Song, Brian M. Sadler, Jiaman Wu, Wei-Lun Chao et al.ICCV 2023 · 685 citations
- TEACh: Task-Driven Embodied Agents That ChatAishwarya Padmakumar, Jesse Thomason, Ayush Shrivastava, Patrick Lange et al.AAAI 2022 · 251 citations
Related papers
- Hierarchical Planning for Complex Tasks with Knowledge Graph-RAG and Symbolic VerificationFlavio Petruzzellis, Cristina Cornelio, Pietro LioICML 2025
- NeSyPr: Neurosymbolic Proceduralization For Efficient Embodied ReasoningWonje Choi, Jooyoung Kim, Honguk WooNeurIPS 2025 · 4 citations
- NeSyC: A Neuro-symbolic Continual Learner For Complex Embodied Tasks in Open DomainsWonje Choi, Jinwoo Park, Sanghyun Ahn, Daehee Lee et al.ICLR 2025
- Grounding Generative Planners in Verifiable Logic: A Hybrid Architecture for Trustworthy Embodied AIFeiyu Wu, Xu Zheng, Yue Qu, Zhuocheng Wang et al.ICLR 2026 · 4 citations
- Grounding LLMs in Scientific Discovery via Embodied ActionsBo Zhang, Jinfeng Zhou, Yuxuan Chen, Jianing Yin et al.ICML 2026
