Lexicographer’s Lacunas or How to Deal with Missing Representative Dictionary Forms on the Example of Czech

Lexicographer’s Lacunas or How to Deal with Missing Representative Dictionary Forms on the Example of Czech
复制标题

词典编纂者的缺陷或如何处理捷克语示例中缺失的代表性词典形式

DOI:
--
复制
发表时间:
2019
影响因子:
0.5
通讯作者:
Jiří Milička
Jiří Milička
中科院分区:
人文科学3区
文献类型:
--
作者:
Dominika Kováríková;Michal Skrabal;V. Cvrček;L. Lukešová;Jiří Milička

文献摘要

被引文献

相似文献

在编纂词汇表时,每个词典编纂者都会遇到数据中具有未经验证的代表性词典形式的单词。这项研究的重点是如何区分由于缺乏数据而导致该表格丢失的情况和存在一些系统性或语言学原因的情况。基于我们对捷克语书面语料库的研究,我们为不同类型的‘空位’制定了词典推荐。作为先决条件,我们计算了频率阈值,以找到应该在数据中具有代表性形式的单词。基于对2700个名词、形容词和动词的手工分析,我们起草了一个空位分类。缺少词典形式的原因往往与搭配能力有限和对代表性语法类别的非偏好有关。对未经证实的单词形式的研究结果也对语言潜力有重大影响。
When compiling a list of headwords, every lexicographer comes across words with an unattested representative dictionary form in the data. This study focuses on how to distinguish between the cases when this form is missing due to a lack of data and when there are some systemic or linguistic reasons. We have formulated lexicographic recommendations for different types of such ‘lacunas’ based on our research carried out on Czech written corpora. As a prerequisite, we calculated a frequency threshold to find words that should have the representative form attested in the data. Based on a manual analysis of 2,700 nouns, adjectives and verbs that do not, we drew up a classification of lacunas. The reasons for a missing dictionary form are often associated with limited collocability and non-preference for the representative grammatical category. Findings on unattested word forms also have significant implications for language potentiality.