Similarity Based Label Smoothing For Dialogue Generation
Similarity Based Label Smoothing For Dialogue Generation
复制标题
DOI:
--
复制
发表时间:
2021-07
期刊:
影响因子:
--
通讯作者:
Sougata Saha;Souvik Das;R. Srihari
中科院分区:
文献类型:
--
作者:
Sougata Saha;Souvik Das;R. Srihari
Generative neural conversational systems are typically trained by minimizing the entropy loss between the training “hard” targets and the predicted logits. Performance gains and improved generalization are often achieved by employing regularization techniques like label smoothing, which converts the training “hard” targets to soft targets. However, label smoothing enforces a data independent uniform distribution on the incorrect training targets, leading to a false assumption of equiprobability. In this paper, we propose and experiment with incorporating data-dependent word similarity-based weighing methods to transform the uniform distribution of the incorrect target probabilities in label smoothing to a more realistic distribution based on semantics. We introduce hyperparameters to control the incorrect target distribution and report significant performance gains over networks trained using standard label smoothing-based loss on two standard open-domain dialogue corpora.