课题基金 / 基金详情

Data Repository and Management Core

Data Repository and Management Core
数据存储和管理核心
批准号:
10006827
负责人:
Patricia Kovatch
金额:
$108.66万
依托单位国家:
美国
项目类别:
财政年份:
--
资助国家:
美国
项目状态:
未结题
起止时间:
至

项目摘要

项目成果

Patricia Kovatch的其他基金

相似基金

相关文献

中文摘要
翻译
数据存储库和管理核心项目摘要 数据存储和管理核心(DRMC)将促进对环境的科学理解 通过将我们成功的CHEAR数据中心(DC)数据门户和存储库扩展到包括 所有生命阶段的人类。我们建议的HHEAR DC数据门户和存储库,以及我们的有效用户 支持团队,将帮助研究人员在他们的研究中添加全面的暴露分析。我们的演示 致力于公平原则和我们独特的协调数据集将提高研究效率和 一般公众理解,通过(1)提供对清洁和协调的数据和更大的数据集的访问 更大的统计能力和最大限度的再利用;(2)与相关的国家数据集,如 代谢组学工作台、ECHO和其他,以促进新数据类型的数据提交;以及(3) 方便访问现代协作数据分析工具,如Jupyter笔记本和谷歌的 合作。我们将采用高可靠性和安全性的最佳实践,并遵循HIPAA的指导原则。至 开发为CHEAR社区量身定做的高效且重点突出的基础设施、服务和流程,我们 跨学科团队与协调中心(CC)、实验室网络、 Echo DC、代谢组学工作台和其他机构。这些关系,以及我们现有的CHEAR 基础设施、独特的专业知识和既定的流程将延续到HHEAR DC的创建 并将有助于加快其推出速度。这些服务将使HHEAR社区能够发现新的 多尺度和多模式数据集之间的相关性和关系,从而朝着 大数据有望帮助解决全球人类环境健康研究面临的重大挑战 生命之路。如果将一组单独研究的数据组合在一起,可能需要大量的工作 在没有类似于HHEAR DMRC的资源的情况下进行尝试;数据存储库的设计将 促进现有数据集的可管理和高效组合;通用词汇表的可用性 将有助于最大限度地利用每项研究的可用数据;特别行政区将 最终接收到他们可以应用其曝光组相关分析方法来处理的数据集 关于汇集研究人群的环境健康的假设。我们最先进的DRMC一直是 在CHEAR计划中履行这些角色,并将建立和扩展我们作为HHEAR的能力 华盛顿特区。总而言之,利用我们现有的基础设施和专业知识将克服长期需要 执行过程充满挑战--我们已经遇到并克服了许多这样的挑战 在实施CHEAR DC方面的挑战,并将能够灵活地回应HHEAR的需求 网络。
英文摘要
DATA REPOSITORY AND MANAGEMENT CORE PROJECT SUMMARY The Data Repository and Management Core (DRMC) will advance scientific understanding of environmental exposures by expanding our successful CHEAR Data Center (DC) data portal and repository to encompass humans at all life stages. Our proposed HHEAR DC data portal and repository, along with our effective user support team, will help researchers add comprehensive exposure analysis to their studies. Our demonstrated commitment to FAIR principles and our unique harmonized data sets will enhance research productivity and general public understanding by (1) providing access to cleaned and harmonized data and larger data sets for greater statistical power and maximal reuse; (2) interoperating with relevant national data sets such as the Metabolomics Workbench, ECHO, and others to facilitate data submission of new data types; and (3) facilitating access to modern collaborative data analysis tools such as Jupyter notebooks and Google's Colaboratory. We will employ best practices for high reliability and security and follow HIPAA guidelines. To develop effective and focused infrastructure, services, and processes tailored for the CHEAR community, our interdisciplinary team developed strong partnerships with the Coordinating Center (CC), the Lab Network, the ECHO DC, the Metabolomics Workbench, and others. These relationships, along with our existing CHEAR infrastructure, singular expertise, and established processes, will carry over to the creation of the HHEAR DC and will help accelerate its rollout. These services will give the HHEAR community the ability to discover new correlations and relationships between multi-scale and multimodal data sets, thus progressing toward the promise of big data to help solve the major challenges of human environmental health research across the lifecourse. Combining data from a set of individual studies would likely require substantial work if it were attempted without resources similar to those of the HHEAR DMRC; the design of the data repository will facilitate manageable and efficient combining of existing data sets; the availability of common vocabularies developed by the DSR will contribute to maximizing the usable data from each study; and the SSAR will ultimately receive a dataset to which they can apply their exposome-related analytic methods to address hypotheses on the environmental health of the pooled study population. Our state-of-the-art DRMC has been fulfilling these roles within the CHEAR program, and will build on and extend our capabilities as the HHEAR DC. In sum, leveraging our existing infrastructure and expertise will overcome the need for a long implementation process fraught with challenges — we have already encountered and overcome many such challenges in implementing the CHEAR DC, and will be able to flexibly respond to the needs of HHEAR Network.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
COVID and Translational Science supercomputer (CATS)
Data Repository and Management Core
Data Repository and Management Core
Transforming Genomics with 5 PB Big Omics Data Engine Cray CS300-AC Supercomputer
国内基金
海外基金
Scalable Learning and Optimization: High-dimensional Models and Online Decision-Making Strategies for Big Data Analysis