FLEX: an Adaptive Exploration Algorithm for Nonlinear Systems
Matthieu Blanke, Marc Lelarge
摘要
Model-based reinforcement learning is a powerful tool, but collecting data to fit an accurate model of the system can be costly. Exploring an unknown environment in a sample-efficient manner is hence of great importance. However, the complexity of dynamics and the computational limitations of real systems make this task challenging. In this work, we introduce FLEX, an exploration algorithm for nonlinear dynamics based on optimal experimental design. Our policy maximizes the information of the next step and results in an adaptive exploration algorithm, compatible with generic parametric learning models and requiring minimal resources. We test our method on a number of nonlinear environments covering different settings, including time-varying dynamics. Keeping in mind that exploration is intended to serve an exploitation objective, we also test our algorithm on downstream model-based classical control tasks and compare it to other state-of-the-art model-based and model-free approaches. The performance achieved by FLEX is competitive and its computational cost is low.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Stochastic Gradient Descent under Markovian Sampling SchemesMathieu EvenICML 2023 · 被引用 41 次
- Interpretable Meta-Learning of Physical SystemsMatthieu Blanke, Marc LelargeICLR 2024 · 被引用 10 次
- PAC-Bayes Generalisation Bounds for Dynamical Systems including Stable RNNsDeividas Eringis, John Leth, Zheng-Hua Tan, Rafael Wisniewski 等AAAI 2024 · 被引用 5 次
- PAC-Bayesian Error Bound, via Rényi Divergence, for a Class of Linear Time-Invariant State-Space ModelsDeividas Eringis, John Leth, Zheng-Hua Tan, Rafal Wisniewski 等ICML 2024 · 被引用 2 次
它引用的顶会 Paper2
相关 Paper
- Optimal Exploration for Model-Based RL in Nonlinear SystemsAndrew Wagenmaker, Guanya Shi, Kevin JamiesonNeurIPS 2023 · 被引用 29 次
- Exploration via Planning for Information about the Optimal TrajectoryViraj Mehta, Ian Char, Joseph Abbate, Rory Conlin 等NeurIPS 2022 · 被引用 12 次
- PC-MLP: Model-based Reinforcement Learning with Policy Cover Guided ExplorationYuda Song, Wen SunICML 2021 · 被引用 23 次
- SOMBRL: Scalable and Optimistic Model-Based RLBhavya Sukhija, Lenart Treven, Carmelo Sferrazza, Florian Dörfler 等NeurIPS 2025 · 被引用 9 次
- Bayesian Optimistic Optimization: Optimistic Exploration for Model-based Reinforcement LearningChenyang Wu, Tianci Li, Zongzhang Zhang, Yang YuNeurIPS 2022 · 被引用 9 次
