Exponential moving average Q-learning algorithm
Published • Apr 1, 2013
NobleIDNI6P94W06R53S10
Authors:,
Mostafa D. Awheda
Howard M. Schwartz
Abstract
A multi-agent policy iteration learning algorithm is proposed in this work. The Exponential Moving Average (EMA) mechanism is used to update the policy for a Q-learning agent so that it converges to an optimal policy against the policies of the other agents. The proposed EMA Q-learning algorithm is ...
Finding related papers...
Discussions
(0)No comments yet
Be the first to share your thoughts!