Bidirectional Decoding: Improving Action Chunking via Guided Test-Time Sampling
Yuejiang Liu, Jubayer Ibn Hamid, Annie Xie, Yoonho Lee, Max Du, Chelsea Finn
摘要
Predicting and executing a sequence of actions without intermediate replanning, known as action chunking, is increasingly used in robot learning from human demonstrations. Yet, its effects on the learned policy remain inconsistent: some studies find it crucial for achieving strong results, while others observe decreased performance. In this paper, we first dissect how action chunking impacts the divergence between a learner and a demonstrator. We find that action chunking allows the learner to better capture the temporal dependencies in demonstrations but at the cost of reduced reactivity to unexpected states. To address this tradeoff, we propose Bidirectional Decoding (BID), a test-time inference algorithm that bridges action chunking with closed-loop adaptation. At each timestep, BID samples multiple candidate predictions and searches for the optimal one based on two criteria: (i) backward coherence, which favors samples that align with previous decisions; (ii) forward contrast, which seeks samples of high likelihood for future plans. By coupling decisions within and across action chunks, BID promotes both long-term consistency and short-term reactivity. Experimental results show that our method boosts the performance of two state-of-the-art generative policies across seven simulation benchmarks and two real-world tasks. Code and videos are available at https://bid-robot.github.io .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Adaptive Action Chunking at Inference-time for Vision-Language-Action ModelsYuanchang Liang, Xiaobo Wang, Kai Wang, Shuo Wang 等CVPR 2026 · 被引用 30 次
- Real-Time Robot Execution with Masked Action ChunkingHaoxuan Wang, Gengyu Zhang, Yan Yan, Yuzhang Shang 等ICLR 2026 · 被引用 24 次
- Diffusion Forcing Planner: History-Annealed Planning with Time-Dependent Guidance for Autonomous DrivingZehan Zhang, Yaoyi Li, Neng Zhang, Jia CaiCVPR 2026 · 被引用 2 次
- TapSampling: Inference-Time Sampling with a Task-Progress-Understanding Verifier for Robotic ManipulationSizhe Zhao, Shengping Zhang, Shuo Yang, Weiyu Zhao 等ICML 2026 · 被引用 1 次
- FocalPolicy: Frequency-Optimized Chunking and Locally Anchored Flow Matching for Coherent Visuomotor PolicyQian He, Zhenshuo Yang, Wenqi Liang, Chunhui Hao 等ICML 2026 · 被引用 1 次
它引用的顶会 Paper18
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- Fast Inference from Transformers via Speculative DecodingYaniv Leviathan, Matan Kalman, Yossi MatiasICML 2023 · 被引用 1,472 次
- Planning with Diffusion for Flexible Behavior SynthesisMichael Janner, Yilun Du, Joshua B. Tenenbaum, Sergey LevineICML 2022 · 被引用 1,115 次
- Neural Text Generation With Unlikelihood TrainingSean Welleck, Ilia Kulikov, Stephen Roller, Emily Dinan 等ICLR 2020 · 被引用 683 次
- Self-Consistency Improves Chain of Thought Reasoning in Language ModelsXuezhi Wang, Jason Wei, Dale Schuurmans, Quoc V. Le 等ICLR 2023 · 被引用 681 次
相关 Paper
- Improving Generative Behavior Cloning via Self-Guidance and Adaptive ChunkingJunhyuk So, Chiwoong Lee, Shinyoung Lee, Jungseul Ok 等NeurIPS 2025 · 被引用 14 次
- Generative Trajectory Stitching through Diffusion CompositionYunhao Luo, Utkarsh A. Mishra, Yilun Du, Danfei XuNeurIPS 2025 · 被引用 48 次
- Time Optimal Execution of Action Chunk Policies Beyond Demonstration SpeedSunwoo Kim, Jeongjun Kim, Joseph J LimICLR 2026
- Action Chunking and Data Augmentation Yield Exponential Improvements in Behavior Cloning for Continuous SpacesThomas TCK Zhang, Daniel Pfrommer, Chaoyi Pan, Nikolai Matni 等ICLR 2026
- BRIC: Bridging Kinematic Plans and Physical Control at Test TimeDohun Lim, Minji Kim, Jaewoon Lim, Sungchan KimAAAI 2026
