课题基金 / 基金详情

Robust Incremental Parsing based on Finite-State Approximation of Context Free Grammar

Robust Incremental Parsing based on Finite-State Approximation of Context Free Grammar
基于上下文无关文法有限状态逼近的鲁棒增量解析
批准号:
15300044
负责人:
OGAWA Yasuhiro
金额:
$8.58万
依托单位:
依托单位国家:
日本
项目类别:
Grant-in-Aid for Scientific Research (B)
财政年份:
2003
资助国家:
日本
项目状态:
已结题
起止时间:
2003 至 2004

项目摘要

项目成果

OGAWA Yasuhiro的其他基金

相似基金

相关文献

中文摘要
翻译
在本研究中,一直致力于开发与语音输入速度相同程度的解析技术的高速和渐进发展,以开发具有同声传译功能的多种语言之间的翻译技术。选择了开发基于统计方法的近似转换技术的方法,该方法能够通过使用具有句法树的语言语料库来适当地反映语法规则的使用频率等。具体推进了以下项目的研究。^*英语、日语、泰语语料库维护。使用了名古屋大学综合声音信息研究基地收集的同声传译会话语料库。在本研究中,它的翻译数据是利用日语的24小时数据和英语翻译与原始数据中的泰国。语法树数据被给予每个 ...更多信息 语言语料库中的短语结构语法的形式。^*利用句法树从大规模语料库中获取统计信息本文从给出句法树的语言语料库中,研究了获取各种统计信息的技术。通过搜索语法规则在语法树中出现的位置,并使用上下文信息对其进行描述,进行统计分析。^*有限自动机近似转换技术的发展:研究了将上下文无关文法转换为有限自动机的技术。在转化中。自动机是将上下文无关文法表示成自反迁移网络的形式,并将其向下发展而成的。根据概率计算,开发了优先开发使用频率高的区域的算法,作为开发方法。^*句法分析渐进系统的设计与实现:设计并实现了语法获取、有限自动机逼近和句法分析系统。有限的自动机,包括约5000万的面积,是为了实现一个可行的分析。句法分析的比较评价:使用该基准进行了英语、日语和泰语的句法分析实验。结果从精度、时间、语法树的数量和形式等多个角度对该分析技术进行了比较评价,通过两年的研究,验证了上下文无关文法的自动机近似的鲁棒增量解析的可实现性,并确认了分析速度的提高效果。少
英文摘要
In this research, it has been aiming at the development of high speed and gradual progress the parsing technology treatable by the same degree of the speed as the voice input to develop the translation technology between several languages that have the simultaneous interpreter function. The approach of developing the approximation conversion technique based on the statistical method that was able to reflect the use frequency etc. of the grammatical rule appropriately by using the language corpus with the syntax tree was selected. The research of the following items was concretely promoted.^*English, Japanese, and Thai corpus maintenance.The simultaneous interpreter conversation corpus that had been collected in the Nagoya University integration sound information research base was used. In this research, the translation data of it was made by using Japanese data of 24 hours and the English translation with the original data among these about Thai. The syntax tree data was given to each … More language corpus in the form of the phrase structure grammar.^*Acquisition of statistical information from large-scale corpus with syntax tree.The technique to acquire various statistical informations was examined from the language corpus from which the syntax tree was given. It was statistically analyzed by searching for the position in the syntax tree where the grammatical rule appeared, and describing it with the context information.^*Development of limited automata approximation conversion technique.The technique for converting it from the context-free grammar into limited automata was researched. In conversion. Automata were made by expressing the context-free grammar in the form of the reflexive transition network, and developing them descending. The algorithm that developed the are with high use frequency by priority was developed as a development method according to the probability calculation.^*Design and mounting of parsing gradual progress system.The system of grammatical acquisition, the limited automata approximation, and parsing was designed, and mounted. Limited automata that consisted of the are of about 50 million were made for the achievement of a practicable analysis.^*Evaluation for comparison of parsing.The parsing experiment in English, Japanese, and Thai was executed by using the bench mark. The evaluation for comparison of this analytical technique from a diversified viewpoint like accuracy, time, and the number and the form etc. of the syntax tree was executed as a result.The realizability of robust incremental parsing was verified by the automata approximation of the context-free grammar through the research on two years, and the effect on the speed-up of the analysis was able to be confirmed. Less
期刊论文(23)
专著(0)
科研奖励(0)
会议论文
Itsuki Kishida: "Construction of an Advanced In-Car Spoken Dialogue Corpus and its Characteristic Analysis"Proceedings of 8th European Conference on Speech Communication and Technology. (2003)
岸田Itsuki Kishida:《先进车载口语对话语料库的构建及其特征分析》第八届欧洲语音通信与技术会议论文集。
DOI: --
发表时间:
期刊:
影响因子: --
作者: []
通讯作者:
CIAIR In-car Speech Corpus-Influence of Driving Status-
CIAI​​R车载语音语料库-驾驶状态的影响-
DOI: --
发表时间: 2005
期刊: IEICE Transactions on Information and Systems E88-D・3
影响因子: --
作者: [Yoshihide Kato, Tomohiro Ohno, Nobuo Kawaguchi]
通讯作者: Nobuo Kawaguchi
CIAIR Simultaneous Interpretation Corpus
CIAI​​R同声传译语料库
DOI: --
发表时间: 2004
期刊: Proceedings of Oriental COCOSDA 2004
影响因子: --
作者: [Yoshihide Kato, Hitomi Tohyama]
通讯作者: Hitomi Tohyama
DOI: --
发表时间: 2005
期刊: Journal of Natural Language Processing Vol.11, No.5
影响因子: --
作者: [Yoshihide Kato, Tomohiro Ohno, Nobuo Kawaguchi, Makoto Tachibana, Yoshihide Kato, 能勢 隆, Tomohiro Ohno, 川島 啓吾, Nobuo Kawaguchi, Yasuhiro Ogawa]
通讯作者: Yasuhiro Ogawa
共 16 条
    Automatic generation of the Outlines of Japanese Statutes
    • 批准号:
      17K00460
    • 项目类别:
      Grant-in-Aid for Scientific Research (C)
    • 资助金额:
      $2.91万
    • 财政年份:
      2017
    • 负责人:
      OGAWA Yasuhiro
    • 依托单位:
    Relationship between spring water and slope failures occurred in a volcanic hillslope terrain under heavy rainfall
    Statistical inflectional processing of agglutinative languages
    • 批准号:
      22700143
    • 项目类别:
      Grant-in-Aid for Young Scientists (B)
    • 资助金额:
      $2.25万
    • 财政年份:
      2010
    • 负责人:
      OGAWA Yasuhiro
    • 依托单位:
    Identification of the therapeutic effect of new enzyme-targeting radiosensitization treatment, KORTUC for tumor stem cells
    • 批准号:
      21591610
    • 项目类别:
      Grant-in-Aid for Scientific Research (C)
    • 资助金额:
      $3.08万
    • 财政年份:
      2009
    • 负责人:
      OGAWA Yasuhiro
    • 依托单位:
    国内基金
    海外基金
    基于儿童心理分析的图解式汉语口语自动解析方法研究
    • 批准号:
      60175012
    • 项目类别:
      面上项目
    • 资助金额:
      18.0万元
    • 批准年份:
      2001
    • 负责人:
      宗成庆
    • 依托单位: