Using Sequence Similarity Networks to Identify Partial Cognates in Multilingual Wordlists
Using Sequence Similarity Networks to Identify Partial Cognates in Multilingual Wordlists
复制标题
使用序列相似性网络识别多语言单词列表中的部分同源词
DOI:
--
复制
发表时间:
2016
期刊:
影响因子:
--
通讯作者:
Eric Bapteste
中科院分区:
文献类型:
--
作者:
Johann;P. Lopez;Eric Bapteste
Increasing amounts of digital data in historical linguistics necessitate the development of automatic methods for the detection of cognate words across languages. Recently developed methods work well on language families with moderate time depths, but they are not capable of identifying cognate morphemes in words which are only partially related. Partial cog-nacy, however, is a frequently recurring phenomenon, especially in language families with productive derivational morphology. This paper presents a pilot approach for partial cognate detection in which networks are used to represent similarities be-tween word parts and cognate morphemes are identified with help of state-of-the-art algorithms for network partitioning. The approach is tested on a newly created benchmark dataset with data from three sub-branches of Sino-Tibetan and yields very promising results, outperforming all algorithms which are not sensible to partial cognacy.
影响因子:
2.6
作者:
通讯作者:
--