Evaluating the Efficacy of Summarization Evaluation across Languages
Evaluating the Efficacy of Summarization Evaluation across Languages
复制标题
DOI:
10.18653/v1/2021.findings-acl.71
复制
发表时间:
2021-06
期刊:
影响因子:
--
通讯作者:
Fajri Koto;Jey Han Lau;Timothy Baldwin
中科院分区:
文献类型:
--
作者:
Fajri Koto;Jey Han Lau;Timothy Baldwin
While automatic summarization evaluation methods developed for English are routinely applied to other languages, this is the first attempt to systematically quantify their panlinguistic efficacy. We take a summarization corpus for eight different languages, and manually annotate generated summaries for focus (precision) and coverage (recall). Based on this, we evaluate 19 summarization evaluation metrics, and find that using multilingual BERT within BERTScore performs well across all languages, at a level above that for English.