QuanTOR - Quantitative Analysis of Textual Organisation across Registers
QuanTOR - Quantitative Analysis of Textual Organisation across Registers
批准号:
528467412
负责人:
Professorin Dr. Stephanie Evert
金额:
$0.0万
依托单位国家:
德国
项目类别:
Research Grants
财政年份:
--
资助国家:
德国
项目状态:
未结题
起止时间:
中文摘要
点击翻译按钮获取中文摘要
英文摘要
In linguistic research, registers are usually analysed at the level of entire texts, despite the fact that common definitions of register are linked to the situational context. Situations unfold dynamically over time and language users make different choices at different points in this process. As a consequence, texts are linguistically different at the beginning, in the middle and at the end, and this dynamic organisation at the sub-textual level needs to be integrated in register studies. This project aims at enriching linguistic register studies with an account of the dynamic nature of situations and hence registers as well as quantitative methods for studying this phenomenon. Our focus on the temporal dynamics of text organisation requires an approach that is capable of extracting patterns of linguistic features and the underlying (latent) dimensions of variation from short text segments. It also necessitates automatic identification and classification of relevant segments in order to scale to the analysis of very large corpora. In order to achieve these goals, the project adopts a three-pronged approach to corpus analysis, leveraging the respective expertise of its three applicants. Our work programme combines linguistic interpretation, theory development and manual annotation with multivariate quantitative analysis as well as unsupervised and supervised machine learning. To this end, we develop a novel Bayesian version of Geometric Multivariate Analysis, a reliable and fine-grained approach to the investigation of linguistic variation, as well as machine-learning approaches for the segmentation and labelling of texts that apply state-of-the-art neural and statistical language models. In an iterative process, we develop a theory of the dynamics of language use in situational context, a gold standard of manually segmented and labelled texts, the BayesGMA approach for studying multivariate feature distributions in text time, as well as readily applicable language models for automatic text segmentation and labelling. All components and quantitative results are carefully evaluated and validated. The project uses components of the International Corpus of English (ICE), which not only facilitate the analysis of temporal dynamics across a range of different registers – in the spoken and written mode – but also allow us to submit theoretical claims to an empirical test, for instance, concerning the theoretical relationship between genre and register. The results of the computationally supported corpus analysis will provide a new perspective on a theory of register as a dynamic linguistic reflection of human behaviour in situational context, based on quantitative empirical insight, which, in turn, will feed into our understanding of the architecture of language.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
Reconstructing Arguments from Newsworthy Debates
-
批准号:377333057
-
项目类别:Priority Programmes
-
资助金额:$0.0万
-
财政年份:2017
-
负责人:Professorin Dr. Stephanie Evert
-
依托单位:
Reading concordances in the 21st century (RC21)
-
批准号:508235423
-
项目类别:Research Grants
-
资助金额:$0.0万
-
财政年份:--
-
负责人:Professorin Dr. Stephanie Evert
-
依托单位:
The Normalization of Right-wing Populist and New Right Discourses in Japan and Germany
-
批准号:466328567
-
项目类别:Research Grants
-
资助金额:$0.0万
-
财政年份:--
-
负责人:Professorin Dr. Stephanie Evert
-
依托单位:
海外基金