A Deep Learning Architecture for Corpus Creation for Telugu Language
A Deep Learning Architecture for Corpus Creation for Telugu Language
复制标题
用于泰卢固语语料库创建的深度学习架构
DOI:
10.1007/978-981-15-4029-5_1
复制
发表时间:
2020
期刊:
影响因子:
--
通讯作者:
Dhana L. Rao, Venkatesh R.
中科院分区:
文献类型:
--
作者:
Dhana L. Rao, Venkatesh R.
Many natural languages are on the decline due to the dominance of English as the language of the World Wide Web (WWW), globalized economy, socioeconomic, and political factors. Computational Linguistics offers unprecedented opportunities for preserving and promoting natural languages. However, availability of corpora is essential for leveraging the Computational Linguistics techniques. Only a handful of languages have corpora of diverse genre while most languages areresource-poorfrom the perspective of the availability of machine-readable corpora. Telugu is one such language, which is the official language of two southern states in India. In this paper, we provide an overview of techniques for assessing language vitality/endangerment, describe existing resources for developing corpora for the Telugu language, discuss our approach to developing corpora, and present preliminary results.
登录
查看更多内容
DOI:
--
发表时间:
2015
期刊:
影响因子:
--
作者:
Friederike Lüpke
通讯作者:
Friederike Lüpke
DOI:
--
发表时间:
2004
期刊:
影响因子:
--
作者:
M. Wynne
通讯作者:
M. Wynne
DOI:
--
发表时间:
2019
期刊:
影响因子:
--
作者:
Anita C. Faul
通讯作者:
Anita C. Faul
DOI:
--
发表时间:
2006
期刊:
影响因子:
--
作者:
H. Hughes
通讯作者:
H. Hughes
影响因子:
0.6
作者:
J. Fishman
通讯作者:
J. Fishman