Multi-criteria expertness based cooperative method for SARSA and eligibility trace algorithms

Pakizeh Esmat; Pedram Mir Mohsen<sup>*</sup>; Palhang Maziar

doi:10.1007/s10489-015-0665-y

登录

免费注册

赞收藏引用

科研之友

微信

新浪微博

Facebook

分享链接

Multi-criteria expertness based cooperative method for SARSA and eligibility trace algorithms

作者：Pakizeh Esmat; Pedram Mir Mohsen^*; Palhang Maziar

来源：Applied Intelligence, 2015, 43(3): 487-498.

DOI：10.1007/s10489-015-0665-y

摘要

Temporal difference and eligibility traces are of the most common approaches to solve reinforcement learning problems. However, except in the case of Q-learning, there are no studies about using these two approaches in a cooperative multi-agent learning setting. This paper addresses this shortcoming by using temporal difference and eligibility traces as the core learning method in multi-criteria expertness based cooperative learning (MCE). The experiments, performed on a sample maze world, show the results of an empirical study on temporal difference and eligibility trace methods in a MCE based cooperative learning setting.

出版日期2015-10

全文

访问全文

收藏分享被引(4) 浏览

更新时间：2021-04-10 14:03

相似论文
引用论文
参考文献

产品服务

科研之友科研之友机构版科创云

站内浏览

科研成果科研人员科研机构

服务支持

帮助中心隐私政策服务条款

联系方式

在线客服：【立即咨询】客户热线：400-1616-289 电子邮箱：support@scholarmate.com

微信公众号