课题基金 / 基金详情

RAPID: Rich and Accurate Auxiliary Databases for Supporting Virus Data Efforts

RAPID: Rich and Accurate Auxiliary Databases for Supporting Virus Data Efforts
RAPID:丰富、准确的辅助数据库,支持病毒数据工作
批准号:
2029556
负责人:
Michael Cafarella
金额:
$16.48万
依托单位国家:
美国
项目类别:
Standard Grant
财政年份:
2020
资助国家:
美国
项目状态:
已结题
起止时间:
2020-05-01 至 2022-04-30

项目摘要

项目成果

Michael Cafarella的其他基金

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
This COIVD-19 RAPID project will assist in the mitigation of the negative impacts of COVID-19 on public health, society, and the economy, by creating high-quality databases from highly distributed data about medical and governmental services related to COVID-19. The project will develop software tools to help in the creation of "auxiliary" databases with high-quality data to assist in making better decisions, avoiding fraud, and yielding high-quality analysis sooner in the urgent and rapidly evolving situation created by the coronavirus pandemic. The techniques that will be used to achieve high-quality include:(1) linking "background data" to the data sets to enable quality-checking and fraud detection. For example, ensuring that hospital information listed in the medical resource database is annotated with an accurate phone number so that a volunteer can contact the hospital and check on the accuracy of the data, and (2) creating new "join keys" to enable easy integration of data in the auxiliary database with other data. The project will work closely with other related COVID-19 RAPID efforts which are working on various aspects of data and information collection from the Web.The project will focus on creating two high-quality databases using these strategies: (1) A unified medical institution auxiliary database, which will be a database of all known US medical institutions and (2) A unified government office auxiliary database, which will be a database of all known government offices in the United States—city halls, courts, licensing offices, etc.—at any level of government. Both these data sets are crucial for ensuring that citizens receive a base level of medical aid and government assistance. These resources would be beneficial not only for this particular pandemic, but would become essential resources, in general, for the future. The proposed auxiliary data set creation infrastructure will include a rich schema of background information, used for quality-checking, and a set of join keys for data integration. While there is a huge array of medical institution data sets online, many of the data sets are misaligned due to lack of standard names and/or data integration keys since different projects make different local decisions in choosing these values that may not be universally compatible. As a result, the background information becomes less rich and makes integration with data from other institutions or analysis pipelines much more difficult. The strategies used to create this infrastructure would include: (1) synthesis of preliminary auxiliary datasets, which includes generating common, candidate attributes for all objects in the input set, for example, creating a helipad field for hospitals based on examining all hospital data in Wikidata; (2) identification of inputs with missing values, and filling in those values with a combination of Web extraction tasks and crowdsourcing tasks, and (3) flagging values that are suspected of being incorrect by, for example, automatically creating a set of machine-learned predictors for each column in the auxiliary data. The system could then run the predictor and identify outlier values.This RAPID award is made by the Convergence Accelerator program in the Office of Integrative Activities using funds from the Coronavirus Aid, Relief, and Economic Security (CARES) Act, and is associated with the Convergence Accelerator Track A: Open Knowledge Network.This award reflects NSF's statutory mission and has been deemed worthy of support through evaluation using the Foundation's intellectual merit and broader impacts review criteria.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
A1: Knowledge Network Development Infrastructure with Application to COVID-19 Science and Economics
  • 批准号:
    2132318
  • 项目类别:
    Cooperative Agreement
  • 资助金额:
    $499.45万
  • 财政年份:
    2021
  • 负责人:
    Michael Cafarella
  • 依托单位:
A1: Knowledge Network Development Infrastructure with Application to COVID-19 Science and Economics
Convergence Accelerator Phase I (RAISE): Simultaneous Knowledge Network Programming and Extraction
I-Corps: Explanation-Based Auditing: Improving the Security of Electronic Medical Records
国内基金
海外基金
Rich2通过调控自噬抑制炎症小体NLRP3通路在癫痫形成中的机制研 究
  • 批准号:
  • 项目类别:
    省市级项目
  • 资助金额:
    --
  • 批准年份:
    2024
  • 负责人:
    张小刚
  • 依托单位:
前扣带回GTP酶激活蛋白RICH2介导Shank3-/-孤独症小鼠社交行为障碍的机制研究
整合素β1/RICH1复合体感应细胞外基质硬度信号调控乳腺癌侵袭转移的机制研究
  • 批准号:
    82303462
  • 项目类别:
    青年科学基金项目
  • 资助金额:
    30万元
  • 批准年份:
    2023
  • 负责人:
    田琦
  • 依托单位:
转录因子NtMYB305通过AT-rich元件调控NtPMT表达及烟碱合成的分子机制研究
  • 批准号:
    32101643
  • 项目类别:
    青年科学基金项目(C类)
  • 资助金额:
    30.0万元
  • 批准年份:
    2021
  • 负责人:
    田田
  • 依托单位: