Functional Regularization for Reinforcement Learning via Learned Fourier Features
Alexander C. Li, Deepak Pathak
摘要
We propose a simple architecture for deep reinforcement learning by embedding inputs into a learned Fourier basis and show that it improves the sample efficiency of both state-based and image-based RL. We perform infinite-width analysis of our architecture using the Neural Tangent Kernel and theoretically show that tuning the initial variance of the Fourier basis is equivalent to functional regularization of the learned deep network. That is, these learned Fourier features allow for adjusting the degree to which networks underfit or overfit different frequencies in the training data, and hence provide a controlled mechanism to improve the stability and performance of RL optimization. Empirically, this allows us to prioritize learning low-frequency functions and speed up learning by reducing networks' susceptibility to noise in the optimization process, such as during Bellman updates. Experiments on standard state-based and image-based RL benchmarks show clear benefits of our architecture over the baselines. Website at https://alexanderli.com/learned-fourier-features
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- AnyMorph: Learning Transferable Polices By Inferring Agent MorphologyBrandon Trabucco, Mariano Phielipp, Glen BersethICML 2022 · 被引用 37 次
- Overcoming The Spectral Bias of Neural Value ApproximationGe Yang, Anurag Ajay, Pulkit AgrawalICLR 2022 · 被引用 30 次
- Learning Energy-Based Prior Model with Diffusion-Amortized MCMCPeiyu Yu, Yaxuan Zhu, Sirui Xie, Xiaojian Ma 等NeurIPS 2023 · 被引用 17 次
- Bridging the performance-gap between target-free and target-based reinforcement learningThéo Vincent, Yogesh Tripathi, Tim Lukas Faust, Abdullah Akgül 等ICLR 2026 · 被引用 6 次
- Use the Online Network If You Can: Towards Fast and Stable Reinforcement LearningAhmed Hendawy, Henrik Metternich, Théo Vincent, Mahdi Kallel 等ICLR 2026 · 被引用 4 次
它引用的顶会 Paper9
- Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsMatthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil 等NeurIPS 2020 · 被引用 4,036 次
- CURL: Contrastive Unsupervised Representations for Reinforcement LearningMichael Laskin, Aravind Srinivas, Pieter AbbeelICML 2020 · 被引用 1,261 次
- Reinforcement Learning with Augmented DataMichael Laskin, Kimin Lee, Adam Stooke, Lerrel Pinto 等NeurIPS 2020 · 被引用 833 次
- Deep learning versus kernel learning: an empirical study of loss landscape geometry and the time evolution of the Neural Tangent KernelStanislav Fort, Gintare Karolina Dziugaite, Mansheej Paul, Sepideh Kharaghani 等NeurIPS 2020 · 被引用 255 次
- Neural Tangents: Fast and Easy Infinite Neural Networks in PythonRoman Novak, Lechao Xiao, Jiri Hron, Jaehoon Lee 等ICLR 2020 · 被引用 254 次
相关 Paper
- Q-functionals for Value-Based Continuous ControlSamuel Lobel, Sreehari Rammohan, Bowen He, Shangqun Yu 等AAAI 2023 · 被引用 10 次
- Revisiting Data Augmentation in Deep Reinforcement LearningJianshu Hu, Yunpeng Jiang, Paul WengICLR 2024 · 被引用 9 次
- Deep Frequency Principle Towards Understanding Why Deeper Learning Is FasterZhiqin John Xu, Hanxu ZhouAAAI 2021 · 被引用 67 次
- State Sequences Prediction via Fourier Transform for Representation LearningMingxuan Ye, Yufei Kuang, Jie Wang, Rui Yang 等NeurIPS 2023 · 被引用 18 次
- Deep Sturm-Liouville: From Sample-Based to 1D Regularization with Learnable Orthogonal Basis FunctionsDavid Vigouroux, Joseba Dalmau, Louis Béthune, Victor BoutinICML 2025
