NeRF in the Palm of Your Hand: Corrective Augmentation for Robotics via Novel-View Synthesis
Allan Zhou, Moo Jin Kim, Lirui Wang, Pete Florence, Chelsea Finn
Abstract
Expert demonstrations are a rich source of supervision for training visual robotic manipulation policies, but imitation learning methods often require either a large number of demonstrations or expensive online expert supervision to learn reactive closed-loop behaviors. In this work, we introduce SPARTN (Synthetic Perturbations for Augmenting Robot Trajectories via NeRF): a fully-offline data augmentation scheme for improving robot policies that use eye-inhand cameras. Our approach leverages neural radiance fields (NeRFs) to synthetically inject corrective noise into visual demonstrations, using NeRFs to generate perturbed viewpoints while simultaneously calculating the corrective actions. This requires no additional expert supervision or environment interaction, and distills the geometric information in NeRFs into a real-time reactive RGB-only policy. In a simulated 6-DoF visual grasping benchmark, SPARTN improves success rates by 2.8× over imitation learning without the corrective augmentations and even outperforms some methods that use online supervision. It additionally closes the gap between RGB-only and RGB-D success rates, eliminating the previous need for depth sensors. In realworld 6-DoF robotic grasping experiments from limited human demonstrations, our method improves absolute success rates by 22.5% on average, including objects that are traditionally challenging for depth-based methods. See video results at https://bland.website/spartn .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers12
- CCIL: Continuity-Based Data Augmentation for Corrective Imitation LearningLiyiming Ke, Yunchu Zhang, Abhay Deshpande, Siddhartha S. Srinivasa et al.ICLR 2024 · 33 citations
- Latent Policy Barrier: Learning Robust Visuomotor Policies by Staying In-DistributionZhanyi Sun, Shuran SongNeurIPS 2025 · 27 citations
- AE-NeRF: Augmenting Event-Based Neural Radiance Fields for Non-ideal Conditions and Larger ScenesChaoran Feng, Wangbo Yu, Xinhua Cheng, Zhenyu Tang et al.AAAI 2025 · 21 citations
- OXE-AugE: A Large-Scale Robot Augmentation of OXE for Scaling Cross-Embodiment Policy LearningGuanhua Ji, Harsha Polavaram, Lawrence Yunliang Chen, Sandeep Bajamahal et al.ICML 2026 · 14 citations
- RF4D: Neural Radar Fields for Novel View Synthesis in Outdoor Dynamic ScenesJiarui Zhang, Zhihao Li, Chong Wang, Bihan WenCVPR 2026 · 8 citations
Builds on11
- Instant neural graphics primitives with a multiresolution hash encodingThomas Müller, Alex Evans, Christoph Schied, Alexander KellerSIGGRAPH 2022 · 4,089 citations
- Image Augmentation Is All You Need: Regularizing Deep Reinforcement Learning from PixelsDenis Yarats, Ilya Kostrikov, Rob FergusICLR 2021 · 911 citations
- Fine-Tuning can Distort Pretrained Features and Underperform Out-of-DistributionAnanya Kumar, Aditi Raghunathan, Robbie Matthew Jones, Tengyu Ma et al.ICLR 2022 · 911 citations
- Reinforcement Learning with Augmented DataMichael Laskin, Kimin Lee, Adam Stooke, Lerrel Pinto et al.NeurIPS 2020 · 833 citations
- 6-DOF GraspNet: Variational Grasp Generation for Object ManipulationArsalan Mousavian, Clemens Eppner, Dieter FoxICCV 2019 · 673 citations
Related papers
- Aug-NeRF: Training Stronger Neural Radiance Fields with Triple-Level Physically-Grounded AugmentationsTianlong Chen, Peihao Wang, Zhiwen Fan, Zhangyang WangCVPR 2022 · 32 citations
- NeuRAD: Neural Rendering for Autonomous DrivingAdam Tonderski, Carl Lindström, Georg Hess, William Ljungbergh et al.CVPR 2024 · 58 citations
- Reinforcement Learning with Neural Radiance FieldsDanny Driess, Ingmar Schubert, Pete Florence, Yunzhu Li et al.NeurIPS 2022 · 72 citations
- LidaRF: Delving into Lidar for Neural Radiance Field on Street ScenesShanlin Sun, Bingbing Zhuang, Ziyu Jiang, Buyu Liu et al.CVPR 2024 · 11 citations
- Dense Depth Priors for Neural Radiance Fields from Sparse Input ViewsBarbara Roessle, Jonathan T. Barron, Ben Mildenhall, Pratul P. Srinivasan et al.CVPR 2022 · 319 citations
