跳到主要导航 跳到搜索 跳到主要内容

Attentive multi-view reinforcement learning

  • Yueyue Hu
  • , Shiliang Sun*
  • , Xin Xu
  • , Jing Zhao
  • *此作品的通讯作者
  • East China Normal University
  • National University of Defense Technology

科研成果: 期刊稿件文章同行评审

摘要

The reinforcement learning process usually takes millions of steps from scratch, due to the limited observation experience. More precisely, the representation approximated by a single deep network is usually limited for reinforcement learning agents. In this paper, we propose a novel multi-view deep attention network (MvDAN), which introduces multi-view representation learning into the reinforcement learning framework for the first time. Based on the multi-view scheme of function approximation, the proposed model approximates multiple view-specific policy or value functions in parallel by estimating the middle-level representation and integrates these functions based on attention mechanisms to generate a comprehensive strategy. Furthermore, we develop the multi-view generalized policy improvement to jointly optimize all policies instead of a single one. Compared with the single-view function approximation scheme in reinforcement learning methods, experimental results on eight Atari benchmarks show that MvDAN outperforms the state-of-the-art methods and has faster convergence and training stability.

源语言英语
页(从-至)2461-2474
页数14
期刊International Journal of Machine Learning and Cybernetics
11
11
DOI
出版状态已出版 - 1 11月 2020

指纹

探究 'Attentive multi-view reinforcement learning' 的科研主题。它们共同构成独一无二的指纹。

引用此