Multi-Word Expressions in Second Language Writing: A Large-Scale Longitudinal Learner Corpus Study

Multi-Word Expressions in Second Language Writing: A Large-Scale Longitudinal Learner Corpus Study
复制标题

第二语言写作中的多词表达:大规模纵向学习者语料库研究

DOI:
10.1111/lang.12383
复制
发表时间:
2019
期刊:
影响因子:
4.4
通讯作者:
Stefania Spina
Stefania Spina
中科院分区:
人文科学1区
文献类型:
--
作者:
A. Siyanova;Stefania Spina

文献摘要

被引文献

相似文献

© 2019密歇根大学语言学习研究俱乐部在本研究中,我们试图通过跟踪在两个不同时间点产生的文章中短语词汇的发展来推进学习者语料库研究领域。为了这个目的,我们雇用了大量的第二语言(L2)学习者(N = 175)从三个熟练程度-初学者,小学和中级-并专注于代表性不足的L2(意大利语)。采用混合效应模型,一个灵活和强大的工具,语料库数据分析,我们分析了学习者的组合在五个不同的措施:短语频率,互信息,词汇重力,三角洲Pforward,和三角洲Pbackward。我们的研究结果表明,一个复杂的画面,在更高的熟练程度和更大的接触L2不会导致更多的习语和目标样的输出,并可能,事实上,导致更多的依赖于低频组合的组成词是非关联或相互吸引。
© 2019 Language Learning Research Club, University of Michigan In the present study, we sought to advance the field of learner corpus research by tracking the development of phrasal vocabulary in essays produced at two different points in time. To this aim, we employed a large pool of second language (L2) learners (N = 175) from three proficiency levels—beginner, elementary, and intermediate—and focused on an underrepresented L2 (Italian). Employing mixed-effects models, a flexible and powerful tool for corpus data analysis, we analyzed learner combinations in terms of five different measures: phrase frequency, mutual information, lexical gravity, delta Pforward, and delta Pbackward. Our findings suggest a complex picture, in which higher proficiency and greater exposure to the L2 do not result in more idiomatic and targetlike output, and may, in fact, result in greater reliance on low frequency combinations whose constituent words are non-associated or mutually attracted.