课题基金 / 基金详情

CRI - Towards a Comprehensive Linguistic Annotation of Language

CRI - Towards a Comprehensive Linguistic Annotation of Language
CRI - 迈向语言的综合语言学注释
批准号:
0551615
负责人:
James Pustejovsky
金额:
$0.0万
依托单位:
依托单位国家:
美国
项目类别:
Continuing Grant
财政年份:
2006
资助国家:
美国
项目状态:
已结题
起止时间:
2006-04-01 至 2010-09-30

项目摘要

项目成果

James Pustejovsky的其他基金

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
This project, developing a Unified Linguistic Annotation (ULA) that integrates in one framework different layers of annotation (e.g., semantics, discourse, temporal, opinions), provides a large word corpus with balanced and annotated data. The effort involves the integration of several existing resources, including PropBank, NomBank, TimeBank, Penn Discourse Treebank, and coreference and opinion annotations. The work addresses the lack of progress in automatically producing semantic representations, a current major obstacle for natural language processing. The project aims at-Achieving an international consensus on a meta-specification framework allowing individual annotations to cohabit with one another (consistency) and specification components from different schemas to refer to merged information (integration);-Improving techniques for producing high performing systems for the reliable types of semantic annotation;-Producing a stable and language-independent methodology for the process of unified linguistic annotation, complete with widely accessible and broadly applicable tools and guidelines;-Validating (with workshops) generality and robustness of ULA by incorporating additional annotation schemas, new genres, and additional languages; and-Actively promoting dissemination of the techniques embodied in the ULA, as well as the resulting annotated corpora throughout the community, for use and further evaluation.The incorporation of automatic taggers into natural language processing (NLP) systems is expected to improve performance in question answering, information extraction, and machine translation, among others, by moving NLP to a new level of richer, deeper processing. The annotated data provides training material for automatic taggers and a wealth of data for further corpus linguistic studies. This project should bring us one step closer to understanding the nature of meaning.Broader Impact: The project offers the opportunity to produce an abstract representation of a vast amount of text, and therefore a vast amount of knowledge, which may be key to turning knowledge representation in general into a more tractable problem. The activity enhances infrastructure for research and education by providing a resource that could lead to major advances in robust, broad coverage semantic processing. The workshops provide tremendous learning opportunities for the students.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
EAGER: Integrating Dense Paraphrased-Enriched Representations with Large Language Models
  • 批准号:
    2326985
  • 项目类别:
    Standard Grant
  • 资助金额:
    $15.0万
  • 财政年份:
    2023
  • 负责人:
    James Pustejovsky
  • 依托单位:
Elements: Towards a Robust Cyberinfrastructure for NLP-based Search and Discoverability over Scientific Literature
  • 批准号:
    2104025
  • 项目类别:
    Standard Grant
  • 资助金额:
    $39.96万
  • 财政年份:
    2021
  • 负责人:
    James Pustejovsky
  • 依托单位:
Travel Support for North American Summer School for Logic, Language, and Information (NASSLLI)
  • 批准号:
    2002141
  • 项目类别:
    Standard Grant
  • 资助金额:
    $4.9万
  • 财政年份:
    2020
  • 负责人:
    James Pustejovsky
  • 依托单位:
Collaborative Research: NSF2026: EAGER: A Playground and Proposal for Growing an AGI
  • 批准号:
    2033932
  • 项目类别:
    Standard Grant
  • 资助金额:
    $10.0万
  • 财政年份:
    2020
  • 负责人:
    James Pustejovsky
  • 依托单位:
海外基金