课题基金 / 基金详情

ITR/SY(CISE) Learning Syntactic/Semantic Information for Parsing

ITR/SY(CISE) Learning Syntactic/Semantic Information for Parsing
ITR/SY(CISE) 学习用于解析的句法/语义信息
批准号:
0112435
负责人:
Eugene Charniak
金额:
$44.94万
依托单位:
依托单位国家:
美国
项目类别:
Standard Grant
财政年份:
2001
资助国家:
美国
项目状态:
已结题
起止时间:
2001-08-15 至 2006-07-31

项目摘要

项目成果

Eugene Charniak的其他基金

相似基金

相关文献

中文摘要
翻译
这项研究关注的是对英语结构信息的无监督学习,这些信息在当前的树库(特别是各种Penn树库)中不存在。 也就是说,人们希望机器能够学习这些信息,而不必创建注释信息的语料库。 要学习的结构信息通常福尔斯在语法和语义之间的边界上;例如,“纽约证券交易所”的名称中有“纽约”这个位置,这一事实属于语法还是语义? 那么,“推销无用的东西”和“无用的东西的市场”这两种表达方式之间的相似性是什么呢? 目的是以当前统计解析器可以使用的形式学习这种信息,以便它们可以输出更精细结构的解析。 但这并不意味着解析是这类信息的唯一用途。 越来越多的用于从自由文本中自动提取信息的系统使用共指检测和“命名实体识别”(例如,认识到“纽约”是一个地点,但“纽约证券交易所”是一个组织)。 有证据表明,共指和命名实体的识别可以提高更精细的分析水平,使这项研究成为可能。 再一次,“语言模型”(为语言中的字符串分配概率的程序)是所有当前语音识别系统的标准部分;有证据表明,更细粒度的句法分析可以改进当前的语言模型。 因此,这项研究将使各种各样的系统,以更好地利用语言输入,使这些系统更容易访问不同的用户池。
英文摘要
This research concerns the unsupervised learning of structural information about English that is not present in current tree-banks (specifically the various Penn tree-banks). That is, one wants a machine to learn this information without having to create a corpus in which the information is annotated. The structural information to be learned often falls at the boundary between syntax and semantics; for example, does the fact that the "New York Stock Exchange" has as part of the name the location "New York" fall under syntax or semantics? What about the similarity between the expressions "[to] market useless items" and "the market for useless items"? The intention is to learn this kind of information in a form that current statistical parsers can use so that they can output more finely structured parses. But this is not meant to suggest that parsing is the sole use for this sort of information. More and more systems for automatically extracting information from free text use coreference detection and "named-entity recognition" (e.g., recognizing that "New York" is a location, but "New York Stock Exchange" is an organization). There is evidence to suggest that both coreference and named-entity recognition can be improved with the finer level of analysis to be made possible by this research. Or again, "language models" (programs that assign a probability to strings in a language) are standard parts of all current speech-recognition systems; there is evidence that suggests that finer grained syntactic analysis can improve current language models. Thus, this research will enable a wide variety of systems to make better use of language input and so make these systems more accessible to a diverse user pool.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
EAGER: Construction of Inter-Igbo
  • 批准号:
    1240178
  • 项目类别:
    Standard Grant
  • 资助金额:
    $8.3万
  • 财政年份:
    2012
  • 负责人:
    Eugene Charniak
  • 依托单位:
Improved Statistical Language Models
  • 批准号:
    9319516
  • 项目类别:
    Continuing Grant
  • 资助金额:
    $24.13万
  • 财政年份:
    1994
  • 负责人:
    Eugene Charniak
  • 依托单位:
Probability and Natural Language Processing
  • 批准号:
    8911122
  • 项目类别:
    Continuing Grant
  • 资助金额:
    $25.33万
  • 财政年份:
    1989
  • 负责人:
    Eugene Charniak
  • 依托单位:
Multiparadigm Design Environments
  • 批准号:
    8722809
  • 项目类别:
    Continuing Grant
  • 资助金额:
    $350.48万
  • 财政年份:
    1988
  • 负责人:
    Eugene Charniak
  • 依托单位:
国内基金
海外基金
基于Nurr1调节YAP-INF2-线粒体分裂途径探讨龙琥醒脑颗粒在SH-SY5Y细胞氧糖剥夺再灌注诱发的神经元损伤的保护作用研究
SY4835通过WEE1/DDR1双靶点抑制胰腺癌的作用及机制
  • 批准号:
    82373136
  • 项目类别:
    面上项目
  • 资助金额:
    48万元
  • 批准年份:
    2023
  • 负责人:
    张晓飞
  • 依托单位:
米糠黄酮抑制Aβ诱导的SH-SY5Y细胞中Tau蛋白过度磷酸化的分子机制研究
  • 批准号:
    2022JJ31009
  • 项目类别:
    省市级项目
  • 资助金额:
    --
  • 批准年份:
    2022
  • 负责人:
    张琳
  • 依托单位:
天目山来源链霉菌Streptomyces sp. SY1322中morindolestatin类新颖咔唑生物碱获取及其铁死亡抑制活性研究
  • 批准号:
    LY21H300001
  • 项目类别:
    省市级项目
  • 资助金额:
    --
  • 批准年份:
    2020
  • 负责人:
    马列峰
  • 依托单位: