Structuring Neural Networks for More Explainable Predictions
Structuring Neural Networks for More Explainable Predictions
复制标题
DOI:
10.1007/978-3-319-98131-4_5
复制
发表时间:
2018
期刊:
影响因子:
--
通讯作者:
Laura Rieger;Pattarawat Chormai;G. Montavon;L. K. Hansen;K. Müller
中科院分区:
文献类型:
--
作者:
Laura Rieger;Pattarawat Chormai;G. Montavon;L. K. Hansen;K. Müller
Machine learning algorithms such as neural networks are more useful, when their predictions can be explained, e.g. in terms of input variables. Often simpler models are more interpretable than more complex models with higher performance. In practice, one can choose a readily interpretable (possibly less predictive) model. Another solution is to directly explain the original, highly predictive model. In this chapter, we present a middle-ground approach where the original neural network architecture is modified parsimoniously in order to reduce common biases observed in the explanations. Our approach leads to explanations that better separate classes in feed-forward networks, and that also better identify relevant time steps in recurrent neural networks.