Production by Computer of Index to Mahavastu-avasana
Production by Computer of Index to Mahavastu-avasana
批准号:
16320009
负责人:
OUSAKA Yumi
金额:
$4.74万
依托单位国家:
日本
项目类别:
Grant-in-Aid for Scientific Research (B)
财政年份:
2004
资助国家:
日本
项目状态:
已结题
起止时间:
2004 至 2006
中文摘要
印度-雅利安语通常分为古印度-雅利安语、中印度-雅利安语和现代印度-雅利安语。中印度雅利安由不同时代、不同地区的各种经典组成,如早期佛教的巴利文经典、早期耆那教的梵文经典、佛教-梵文混合经典等。中古印度雅利安语言与古典梵语非常不同,尽管它们在某些方面与古典梵语相关,并且具有非常复杂的语法结构,例如由于语音变化而完全改变特定单词,合并辅音的同化等。与古典梵语不同,中印度雅利安语的大多数语法特征尚未得到彻底研究。为了在中古印度雅利安语的研究中取得重大突破,现在需要在格律分析、语法结构、词汇和句法等方面对经典进行系统的研究。《圣歌》的格律分析是在…更不需要编辑一个批判版。《正典》词标的编制有助于翻译。倒序索引对于研究句子的语法结构是必要的,而调式索引或倒序调式索引对于确保文本的正确阅读是有用的。为此,需要处理大量数据。幸运的是,通过使用计算机来分析文本,可以在文本的系统研究方面取得相当大的进展。特别是,可以对诗进行分类,可以收集米,可以编制索引和分析语法结构。毫无疑问,这种系统的研究对所有从事印度学和佛学研究的学者都是极有价值的。佛教-梵文混合文本,Mahavastu-avadana是最重要的佛陀传记之一。我们已经在2003年和2006年完成了索引的出版,《马哈瓦斯图》的正字和反字索引。分别是I和II。此文在佛文杂文研究中具有十分重要的意义。我们现在正在编制第三卷的索引。在目前的节拍分析中,它的节拍名称是通过与基本节拍方案和文本数据的模式匹配来确定的。最重要的古老经典的米的识别率平均在70%到80%之间,而最古老的经典的米的识别率最高为20%。采用了一种不同于以往的新技术,即判别分析辅助的神经网络,可以大大提高识别率。我们可以获得目前还无法分析的有关仪表的信息。一种从标准梵文中提取出截然不同的韵律的新技术被开发出来。此外,这种方法可以非常有效和准确地将半段诗分成两段。本研究在神经网络中的应用取得了成功。通过与语言学和信息科学研究人员的合作,建立了实用的中印度雅利安语研究数据库。它的核心是三个组成、字体、编辑器和分析工具(文本数据、仪表分析和索引生成)。讨论了相关的数据库,包括如何在各个平台的新操作系统上设计分析工具。我们与语言学学者和计算机科学家合作,首先在Macintosh PC上制作了用于系统分析中印度雅利安(MIA)手稿的计算机工具,随后在Windows PC和Linux PC上进行了扩展。最近,Mac OS和Windows OS的计算机环境分别从旧的OS 9.2和Me等改为OSX和XP。因此,我们的几个计算机工具不能正常工作。由于我们的分析系统在MIA的手稿分析中起着重要的作用,因此我们基于JAVA重新构建了所有这些计算机工具,以便在每个平台上都能很好地工作。我们已经完成了编制索引的计算机工具手册的出版工作,其中的执行文件是用Java编写的光盘。少
英文摘要
The Indo-Aryan language is usually divided into Old Indo-Aryan, Middle Indo-Aryan (MIA) and Modern Indo-Aryan. Middle Indo-Aryan comprises various scriptures of different times and different localities, such as the Pali canon of early Buddhism, the Prakrit canon of the early Jainism and the Buddhist-Hybrid-Sanskrit canon, etc. Middle Indo-Aryan languages are very different from classical Sanskrit, although they are in some respects related to it, and have very complicated grammatical structures such as complete alteration of a particular word due to phonetic change, assimilation of conjunct consonants, etc. Unlike in the case of classical Sanskrit, most of the grammatical features of Middle Indo-Aryan have not yet been thoroughly investigated. In order to make a major breakthrough in the study of Middle Indo-Aryan, the systematic study of the canons is now required with respect to metrical analysis, grammatical structure, vocabulary and syntax. The metrical analysis of the canons is in … More dispensable to edit a critical edition. The compilation of word indexes of the canons is helpful for making a translation. A reverse index is necessary to investigate the grammatical structure of the sentences, and a pada index or a reverse pada index is useful to ensure a correct reading of the text. For these purposes much data needs to be processed. Fortunately, by using a computer to analyze texts, considerable advances in the systematic study of the texts could be made. In Particular, verses could be classified, metres could be collected, indexes could be compiled and the grammatical structure could be analysed. There is no doubt that this systematic study will be extremely valuable to all scholars engaged in the study of Indology and Buddhology.The Buddhist-Hybrid-Sanskrit text, Mahavastu-avadana is one of the biographies of most important Buddha. We have accomplished the publication of the indexes in 2003 and 2006, the word and reverse word index to the Mahavastu for Vols. I and II, respectively. This text is very important in the study of the Buddhist-Hybrid-Sanskrit. We are now preparing the index to the third volume.In the metre analysis so far, its metre name is identified by the pattern match with the basic metre scheme and the text data. The rate of the identification of the metre of the most important old canons runs from 70 to 80% on the average, and at most 20% level in the oldest one. The identification rate is able to be improved greatly by using a quite new technique different from the previous one, the neural network assisted by the discriminant analysis. We can obtain information on the metre that cannot be analyzed so far. A new technique for extracting a remarkably different metre from the standard Sanskrit is developed. Moreover, this technique can divide a half verse into two padas extremely efficiently and accurately. This research shows one application success in neural network.We have constructed practicable database of the study in middle Indo-Aryan by collaboration with the linguistics and the information scientific researchers. Its kernel is three of composition, the font, the editor, and analysis tools (text data, meter analysis, and index production). A database concerned is discussed, including the discussion how to devise the analysis tools on new OS in each platform.In collaboration with linguistic scholars and computer scientists, we have made the computer tools for the systematic analysis of the manuscripts in Middle Indo-Aryan (MIA) first on Macintosh PC, and subsequently extended it on Windows PC and on Linux PC. Recently computer environment on Mac OS and Windows OS has been changed into OSX and XP from old ones, OS 9.2 and Me etc., respectively. As a result, our several computer tools cannot work normally. Since our analysis system plays an important role in analysis of the manuscripts in MIA, we have rebuilt all these computer tools so that work well on every platforms, based on JAVA. We have accomplished the publication of the manual for computer tools of the compilation of the index, with the execution files on CD-ROM by Java. Less
期刊论文(12)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
Automatic Analysis of the Canon in Middle Indo-Aryan by Personal Computer II in both JAPANESE and ENGLISH with Jar Files and Their Programs by Java for Macintosh OSX, Windows XP, and Linux on CD-ROM
通过个人计算机 II 自动分析中印雅利安语的经典,包括 Jar 文件及其 Java 程序,适用于 Macintosh OSX、Windows XP 和 Linux,光盘上的日语和英语版本
DOI:
--
发表时间:
2005
期刊:
影响因子:
--
作者:
[M.Yamazaki, Y.Ousaka, イ・ヨンスク, 藤田 正勝, Y.Ousaka]
通讯作者:
Y.Ousaka
中期インド・アリアン古聖典解析ツールのジャバによる開発
使用Java开发中古印度-雅利安古典经文分析工具
DOI:
--
发表时间:
2004
期刊:
情報処理学会研究報告『人文科学とコンピュータ』 CH64
影响因子:
--
作者:
[Heisig, James W., Yoshiko Shibata, イ・ヨンスク, Fujita Masakatsu, 逢坂雄美]
通讯作者:
逢坂雄美
MAHAVASTU-AVADANA Vol. II : Word Index and Reverse Word Index
MAHAVASTU-AVADANA 卷。
DOI:
--
发表时间:
2006
期刊:
影响因子:
--
作者:
[E.Faure, B.Oguibenine, M.Yamazaki, Y.Ousaka]
通讯作者:
Y.Ousaka
Mahavastu-avadana Vol.II : Word Index and Reverse Word Index
Mahavastu-avadana Vol.II:单词索引和反向单词索引
DOI:
--
发表时间:
2006
期刊:
影响因子:
--
作者:
[E.Faure, B.Oguibenine, M.Yamazaki, Y.Ousaka]
通讯作者:
Y.Ousaka
DOI:
--
发表时间:
2004
期刊:
影响因子:
--
作者:
[Y.Ousaka, M.Yamazaki]
通讯作者:
M.Yamazaki
共 8 条
Production by Computer of Index to Samyutta-nikaya and Recompile of its Text
-
批准号:20320013
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$5.99万
-
财政年份:2008
-
负责人:OUSAKA Yumi
-
依托单位:
Production by Computer of Index to Majjhima-nikaya and Reconstruction of the Text
-
批准号:14310013
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$3.39万
-
财政年份:2002
-
负责人:OUSAKA Yumi
-
依托单位:
Production by Computer of Index to Visuddhi-magga and Reconstruction of the Text
-
批准号:12410006
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$3.58万
-
财政年份:2000
-
负责人:OUSAKA Yumi
-
依托单位:
Production by Computer of Indexes to Digha-nikaya and Jataka
-
批准号:09044013
-
项目类别:Grant-in-Aid for Scientific Research (B).
-
资助金额:$4.54万
-
财政年份:1997
-
负责人:OUSAKA Yumi
-
依托单位: