NobleBlocks
Public

Exponential moving average Q-learning algorithm

Published • Apr 1, 2013
NobleIDNI6P94W06R53S10
Authors:
Mostafa D. Awheda
,
Howard M. Schwartz

Abstract

A multi-agent policy iteration learning algorithm is proposed in this work. The Exponential Moving Average (EMA) mechanism is used to update the policy for a Q-learning agent so that it converges to an optimal policy against the policies of the other agents. The proposed EMA Q-learning algorithm is ...

Finding related papers...

Discussions

(0)

No comments yet

Be the first to share your thoughts!