Universality of Gradient Descent Neural Network Training
Universality of Gradient Descent Neural Network Training
复制标题
DOI:
10.1016/j.neunet.2022.02.016
复制
发表时间:
2020-07
期刊:
影响因子:
--
通讯作者:
G. Welper
中科院分区:
文献类型:
--
作者:
G. Welper
It has been observed that design choices of neural networks are often crucial for their successful optimization. In this article, we therefore discuss the question if it is always possible to redesign a neural network so that it trains well with gradient descent. This yields the following universality result: If, for a given network, there is any algorithm that can find good network weights for a classification task, then there exists an extension of this network that reproduces the same forward model by mere gradient descent training. The construction is not intended for practical computations, but it provides some orientation on the possibilities of pre-trained networks in meta-learning and related approaches.