Lune

NeurIPS2024Top-tier venue

The Limits of Transfer Reinforcement Learning with Latent Low-rank Structure

Tyler Sam, Yudong Chen, Christina Lee Yu

2024Year
1Citations
1Top-tier citations

Abstract

Many reinforcement learning (RL) algorithms are too costly to use in practice due to the large sizes S,AS, A of the problem's state and action space. To resolve this issue, we study transfer RL with latent low rank structure. We consider the problem of transferring a latent low rank representation when the source and target MDPs have transition kernels with Tucker rank (S,d,A)(S , d, A ), (S,S,d),(d,S,A)(S , S , d), (d, S, A ), or (d,d,d)(d , d , d ). In each setting, we introduce the transfer-ability coefficient α\alpha that measures the difficulty of representational transfer. Our algorithm learns latent representations in each source MDP and then exploits the linear structure to remove the dependence on S,AS, A , or SAS A in the target MDP regret bound. We complement our positive results with information theoretic lower bounds that show our algorithms (excluding the (d,d,dd, d, d) setting) are minimax-optimal with respect to α\alpha.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

Cited by top-tier papers1

Ask how each one uses it

Builds on12

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines