PATTERN-RECOGNIZING STOCHASTIC LEARNING AUTOMATA
PATTERN-RECOGNIZING STOCHASTIC LEARNING AUTOMATA
复制标题
DOI:
10.1109/tsmc.1985.6313371
复制
发表时间:
1985-01-01
期刊:
影响因子:
--
通讯作者:
ANANDAN, P
中科院分区:
文献类型:
--
作者:
BARTO, AG;ANANDAN, P
A class of learning tasks is described that combines aspects of learning automation tasks and supervised learning pattern-classification tasks. These tasks are called associative reinforcement learning tasks. An algorithm is presented, called the associative reward-penalty, or AR-Palgorithm for which a form of optimal performance is proved. This algorithm simultaneously generalizes a class of stochastic learning automata and a class of supervised learning pattern-classification methods related to the Robbins-Monro stochastic approximation procedure. The relevance of this hybrid algorithm is discussed with respect to the collective behaviour of learning automata and the behaviour of networks of pattern-classifying adaptive elements. Simulation results are presented that illustrate the associative reinforcement learning task and the performance of the AR-Palgorithm as compared with that of several existing algorithms.