SuperSim: a test set for word similarity and relatedness in Swedish
SuperSim: a test set for word similarity and relatedness in Swedish
复制标题
SuperSim:瑞典语单词相似性和相关性的测试集
DOI:
--
复制
发表时间:
2021
期刊:
影响因子:
--
通讯作者:
Nina Tahmasebi
中科院分区:
文献类型:
--
作者:
Simon Hengchen;Nina Tahmasebi
Language models are notoriously difficult to evaluate. We release SuperSim, a large-scale similarity and relatedness test set for Swedish built with expert human judgements. The test set is composed of 1,360 word-pairs independently judged for both relatedness and similarity by five annotators. We evaluate three different models (Word2Vec, fastText, and GloVe) trained on two separate Swedish datasets, namely the Swedish Gigaword corpus and a Swedish Wikipedia dump, to provide a baseline for future comparison. We will release the fully annotated test set, code, models, and data.
DOI:
10.18653/v1/n18-2027
发表时间:
2018
期刊:
影响因子:
--
作者:
Schlechtweg;Dominik;Sabine Schulte im Walde ;Stefanie Eckmann
通讯作者:
Stefanie Eckmann