Lune

ICML2025Top-tier venue

Action-Dependent Optimality-Preserving Reward Shaping

Grant C. Forbes, Jianxun Wang, Leonardo Villalobos-Arias, Arnav Jhala, David L. Roberts

2025Year

Abstract

Recent RL research has utilized reward shapingparticularly complex shaping rewards such as intrinsic motivation (IM)-to encourage agent exploration in sparse-reward environments. While

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext 0502e3ed-9bfb-4cfc-8156-5c11b051f9a1

Builds on4

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines