NEIGHBORHOOD DENSITY AND FREQUENCY ACROSS LANGUAGES AND MODALITIES

NEIGHBORHOOD DENSITY AND FREQUENCY ACROSS LANGUAGES AND MODALITIES
复制标题

DOI:
10.1006/jmla.1993.1039
复制
发表时间:
1993-12-01
影响因子:
4.3
通讯作者:
HELLWIG, FM
HELLWIG, FM
中科院分区:
心理学2区
文献类型:
--
作者:
FRAUENFELDER, UH;BAAYEN, RH;HELLWIG, FM

文献摘要

被引文献

相似文献

本研究利用英语和荷兰语 CELEX 词汇数据库来研究单词之间的形式相似关系。词汇统计分析复制并扩展了 Landauer 和 Streeter (1973) 关于单词​​频率与其相似邻域的密度和频率之间关系的发现。荷兰语和英语的结果仅表明,高频书面和口语单词比稀有单词有更多邻居的趋势较弱,并且这些邻居比稀有单词更频繁。然而,我们发现邻居的数量与二元词频率的相关性比与词频的相关性更高。为了阐明这些属性之间的关系,提出了一个随机模型,该模型捕获了音规结构对邻域相似性的相关影响。考虑了这些发现对语言产生和理解模型的影响。
This research exploits the English and Dutch CELEX lexical database to investigate the form similarity relations between words. Lexical statistics analyses replicate and extend the findings of Landauer and Streeter (1973) concerning the relation between a word′s frequency and the density and frequency of its similarity neighborhood. The results for both Dutch and English reveal only a weak tendency for high-frequency written and spoken words to have more neighbors than rare words and for these neighbors to be more frequent than those of rare words. However, the number of neighbors was found to correlate more highly with bigram frequency than with word frequency. To clarify the relations between these properties, a stochastic model is presented which captures the relevant effects of phonotactic structure on neighborhood similarities. The implications of these findings for models of language production and comprehension are considered.