A pruning based method to learn both weights and connections for LSTM
A pruning based method to learn both weights and connections for LSTM
复制标题
一种基于剪枝的方法来学习 LSTM 的权重和连接
DOI:
--
复制
发表时间:
2015
期刊:
影响因子:
--
通讯作者:
Jianglei Han
中科院分区:
文献类型:
--
作者:
Shijian Tang;Jianglei Han
This project is one of the research topics in Professor William Dally’s group. In this project, we developed a pruning based method to learn both weights and connections for Long Short Term Memory (LSTM). In this method, we discard the unimportant connections in a pretrained LSTM, and make the weight matrix sparse. Then, we retrain the remaining model. After we remaining model is converge, we prune this model again and retrain the remaining model iteratively, until we achieve the desired size of model and performance. This method will save the size of the LSTM as well as prevent overfitting. Our results retrained on NeuralTalk shows that we can discard nearly 90% of the weights without hurting the performance too much. Part of the results in this project will be posted in NIPS 2015.