Model-Based Reinforcement Learning via Imagination with Derived Memory

Abstract

Model-based reinforcement learning aims to improve the sample efficiency of policy learning by modeling the dynamics of the environment. Recently, the latent dynamics model is further developed to enable fast planning in a compact space. It summarizes the high-dimensional experiences of an agent, which mimics the memory function of humans. Learning policies via imagination with the latent model shows great potential for solving complex tasks. However, only considering memories from the true experiences in the process of imagination could limit its advantages. Inspired by the memory prosthesis proposed by neuroscientists, we present a novel model-based reinforcement learning framework called Imagining with Derived Memory (IDM). It enables the agent to learn policy from enriched diverse imagination with prediction-reliability weight, thus improving sample efficiency and policy robustness. Experiments on various high-dimensional visual control tasks in the DMControl benchmark demonstrate that IDM outperforms previous state-of-the-art methods in terms of policy robustness and further improves the sample efficiency of the model-based method.

Publication
In Conference on Neural Information Processing Systems (NeurIPS), 2021
{{title}}

{{snippet}}

更多内容

公开信息按发布时间滚动。阅读「亚洲威廉注册优惠」后,可回到列表或查看相邻条目。

建议先扫读标题与摘要,再进入全文。同栏目条目通常按时间倒序排列。

列表适合快速定位,正文适合核对表述。两者都保留在站内即可形成完整阅读路径。

快速通道

网站首页 · Contact · {{title}} · 2021

正文、栏目列表与相关阅读构成完整路径,适合按主题持续查阅。