A statistical method for the identification and aggregation of regional linguistic variation

A statistical method for the identification and aggregation of regional linguistic variation
复制标题

DOI:
10.1017/s095439451100007x
复制
发表时间:
2011-01-01
影响因子:
1
通讯作者:
Geeraerts, Dirk
Geeraerts, Dirk
中科院分区:
人文科学3区
文献类型:
--
作者:
Grieve, Jack;Speelman, Dirk;Geeraerts, Dirk

文献摘要

被引文献

相似文献

本文介绍了一种分析区域语言变异的方法。该方法确定了一组语言变量的空间聚类的个人和共同的模式测量的一组位置的基础上的组合三种统计技术:空间自相关,因子分析和聚类分析。为了演示如何应用这种方法,它是用来分析40个连续测量的,高频词汇交替变量的值的区域变化在一个2600万字的信件语料库的编辑代表来自美国各地的206个城市。
This paper introduces a method for the analysis of regional linguistic variation. The method identifies individual and common patterns of spatial clustering in a set of linguistic variables measured over a set of locations based on a combination of three statistical techniques: spatial autocorrelation, factor analysis, and cluster analysis. To demonstrate how to apply this method, it is used to analyze regional variation in the values of 40 continuously measured, high-frequency lexical alternation variables in a 26-million-word corpus of letters to the editor representing 206 cities from across the United States.