Underreporting of errors in NLG output, and what to do about it
Underreporting of errors in NLG output, and what to do about it
复制标题
DOI:
10.18653/v1/2021.inlg-1.14
复制
发表时间:
2021-08
期刊:
影响因子:
--
通讯作者:
Emiel van Miltenburg;Miruna Clinciu;Ondrej Dusek;Dimitra Gkatzia;Stephanie Inglis;Leo Leppanen;Saad Mahamood;Emma Manning;S. Schoch;Craig Thomson;Luou Wen
中科院分区:
文献类型:
--
作者:
Emiel van Miltenburg;Miruna Clinciu;Ondrej Dusek;Dimitra Gkatzia;Stephanie Inglis;Leo Leppanen;Saad Mahamood;Emma Manning;S. Schoch;Craig Thomson;Luou Wen
We observe a severe under-reporting of the different kinds of errors that Natural Language Generation systems make. This is a problem, because mistakes are an important indicator of where systems should still be improved. If authors only report overall performance metrics, the research community is left in the dark about the specific weaknesses that are exhibited by ‘state-of-the-art’ research. Next to quantifying the extent of error under-reporting, this position paper provides recommendations for error identification, analysis and reporting.