Theoretically Principled Deep RL Acceleration via Nearest Neighbor Function Approximation
Junhong Shen, Lin F. Yang
Abstract
Recently, deep reinforcement learning (RL) has achieved remarkable empirical success by integrating deep neural networks into RL frameworks. However, these algorithms often require a large number of training samples and admit little theoretical understanding. To mitigate these issues, we propose a theoretically principled nearest neighbor (NN) function approximator that can replace the value networks in deep RL methods. Inspired by human similarity judgments, the NN approximator estimates the action values using rollouts on past observations and can provably obtain a small regret bound that depends only on the intrinsic complexity of the environment. We present (1) Nearest Neighbor Actor-Critic (NNAC), an online policy gradient algorithm that demonstrates the practicality of combining function approximation with deep RL, and (2) a plug-and-play NN update module that aids the training of existing deep RL methods. Experiments on classical control and MuJoCo locomotion tasks show that the NN-accelerated agents achieve higher sample efficiency and stability than the baseline agents. Based on its theoretical benefits, we believe that the NN approximator can be further applied to other complex domains to speed-up learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext df21004f-64b9-4cf6-b400-8a58f346f8c1Cited by top-tier papers4
- Thinking vs. Doing: Improving Agent Reasoning by Scaling Test-Time InteractionJunhong Shen, Hao Bai, Lunjun Zhang, Yifei Zhou et al.NeurIPS 2025 · 34 citations
- CAT: Content-Adaptive Image TokenizationJunhong Shen, Kushal Tirumala, Michihiro Yasunaga, Ishan Misra et al.NeurIPS 2025 · 17 citations
- HyRNN: Hybrid Recurrent Neural Networks for Approximating Hybrid Dynamical SystemsRicardo G. SanfeliceAAAI 2026
- Specialized Foundation Models Struggle to Beat Supervised BaselinesZongzhe Xu, Ritvik Gupta, Wenduo Cheng, Alexander Shen et al.ICLR 2025
Builds on1
Related papers
- Finite-Time Analysis of Actor-Critic Methods with Deep Neural Network ApproximationXuyang Chen, Fengzhuo Zhang, Keyu Yan, Lin ZhaoICLR 2026
- Maximize to Explore: One Objective Function Fusing Estimation, Planning, and ExplorationZhihan Liu, Miao Lu, Wei Xiong, Han Zhong et al.NeurIPS 2023 · 30 citations
- The Benefits of Model-Based Generalization in Reinforcement LearningKenny John Young, Aditya A. Ramesh, Louis Kirsch, Jürgen SchmidhuberICML 2023 · 18 citations
- Provably Correct Optimization and Exploration with Non-linear PoliciesFei Feng, Wotao Yin, Alekh Agarwal, Lin YangICML 2021 · 13 citations
- Single-Timescale Actor-Critic Provably Finds Globally Optimal PolicyZuyue Fu, Zhuoran Yang, Zhaoran WangICLR 2021 · 52 citations
