基于变学习率的多agent学习算法的研究

The study of multi-agent learning algorithm based on variable learning rate

作　　者：李琳娜[1]

出　　处：《长春工程学院学报（自然科学版）》2009年第4期81-83,共3页Journal of Changchun Institute of Technology：Natural Sciences Edition

摘　　要：对在动态学习的环境中的IGA算法做了研究,改进了梯度方向上的步长恒定不变的不足,引入了变学习率,并介绍了调节学习率的方法——WoLF原则,加速其收敛。最后根据该方法,对Q学习算法做了改进,并通过仿真试验证明了算法的有效性。This paper studied the IGA algorithm in a dynamic learning environment,and improved the insufficiency of step constantly invariable in the gradient direction.The variable learning rate and the WoLF principle to adjust learning rate were introduced in order to accelerate its convergence.Finally the Q learning algorithm was improved based on this method and the validity of the algorithm was proved through the simulation testing.

关键词：多AGENT学习变学习率 Q学习

分类号：TP181[自动化与计算机技术—控制理论与控制工程]

参考文献：

正在载入数据...

二级参考文献：

正在载入数据...

耦合文献：

正在载入数据...

引证文献：

正在载入数据...

二级引证文献：

正在载入数据...

同被引文献：

正在载入数据...

高级检索 检索式检索

时间限定

期刊范围

学科限定全选

高级检索 检索式检索

时间限定

期刊范围

学科限定全选

基于变学习率的多agent学习算法的研究

我的收藏

参考文献：

二级参考文献：

耦合文献：

引证文献：

二级引证文献：

同被引文献：

相关期刊文献：

相关的主题

相关的作者对象

相关的机构对象

下载全文

高级检索检索式检索

时间限定

期刊范围

学科限定全选

高级检索 检索式检索

时间限定

期刊范围

学科限定全选

基于变学习率的多agent学习算法的研究

我的收藏

参考文献：

二级参考文献：

耦合文献：

引证文献：

二级引证文献：

同被引文献：

相关期刊文献：

相关的主题

相关的作者对象

相关的机构对象

下载全文

用户登录

高级检索检索式检索