课题基金 / 基金详情

A software tool to facilitate variable-level equivalency and harmonization in research data: Leveraging the NIH Common Data Elements Repository to link concepts and measures in an open format

A software tool to facilitate variable-level equivalency and harmonization in research data: Leveraging the NIH Common Data Elements Repository to link concepts and measures in an open format
促进研究数据中变量级别等效性和协调性的软件工具:利用 NIH 通用数据元素存储库以开放格式链接概念和测量
批准号:
10821517
负责人:
Dan Smith
金额:
$27.55万
依托单位:
依托单位国家:
美国
项目类别:
财政年份:
2023
资助国家:
美国
项目状态:
已结题
起止时间:
2023-09-18 至 2024-08-31

项目摘要

项目成果

Dan Smith的其他基金

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
Abstract The National Institute on Aging (NIA) supports numerous studies and archives that collect and disseminate critical data about the aging population of the United States. By supporting the collection and dissemination of longitudinal and multidisciplinary data, the NIA provides researchers the opportunity to measure change and stability in individuals over time, as well as to investigate aging phenomena from an integrated theoretical perspective. In both cases, equivalent or related variables must first be linked or merged before producing appropriately documented data products for eventual harmonization and analysis. The current aging research data environment provides many opportunities for linking similar topical datasets and harmonizing extant common variables, but few software tools are available to facilitate this resource-intensive task. The proposed project will demonstrate the feasibility of a guided harmonization software prototype by concording variables from three nationally representative NIA-funded studies (MIDUS, NHATS, NSHAP) and mapping them against extant data element concept sources such as the NIH Common Data Elements library to identify equivalent concepts and variables. The software prototype will use machine learning and advanced text analysis algorithms to guide the creation of concorded databases (variable crosswalks) that support harmonization and discoverability, both within and across aging-related statistical datasets. Additionally, the prototype will use an open-standards metadata framework to produce richly-described concordance databases that are interoperable, citable and FAIR. Colectica has a track record of creating open- standards based software tools that reduce data management burden by automatically extracting structured metadata from macro-level (study) and micro-level (variable) characteristics of aging studies. Specifically, the prototype will evaluate the feasibility of human-in-the-loop algorithms to operate as a “recommendation engine” to guide the concordance of potentially equivalent or similar variables among multiple datasets. The core hypothesis posits that the prototype will significantly decrease the labor, time, and resources required to create accurate and standardized concorded databases. To test this hypothesis, the research team will: construct and evaluate recommendation algorithms for variable concordance (Aim 1); establish metrics for measuring the accuracy and effectiveness of concordance (Aim 2); and create a user interface to test the recommendation engine, its functions, and associated inputs and outputs (Aim 3).
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
Open Standards-Based Data Extraction Web Tool for Complex Longitudinal Datasets
  • 批准号:
    8123227
  • 项目类别:
  • 资助金额:
    $15.0万
  • 财政年份:
    2011
  • 负责人:
    Dan Smith
  • 依托单位:
海外基金