Lune

NeurIPS2021顶会

Provably efficient multi-task reinforcement learning with model transfer

Chicheng Zhang, Zhi Wang

2021年份
20被引次数
9顶会引用

摘要

We study multi-task reinforcement learning (RL) in tabular episodic Markov decision processes (MDPs). We formulate a heterogeneous multi-player RL problem, in which a group of players concurrently face similar but not necessarily identical MDPs, with a goal of improving their collective performance through inter-player information sharing. We design and analyze an algorithm based on the idea of model transfer, and provide gap-dependent and gap-independent upper and lower bounds that characterize the intrinsic complexity of the problem. Algorithm 1: MULTI-TASK-EULER Input :Failure probability δ ∈ (0, 1), dissimilarity parameter ǫ ≥ 0. Initialize: Set V p (⊥) = 0 for all p in [M ], where ⊥ is the only state in S H+1 ; 1 for k = 1, 2, . . . , K do 2 for p = 1, 2, . . . , M do // Construct optimal value estimates for player p Update optimal action value function upper and lower bound estimates: // All players p interact with their respective environments, and update reward and transition estimates

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

lune papers fulltext 527d9928-a469-4ee8-a5c0-e6fb47f29f95

引用它的顶会 Paper9

问问它们各自怎么用它

它引用的顶会 Paper3

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖