课题基金 / 基金详情

Statistical methods and designs for correlated outcome and covariate errors in studies of HIV/AIDS

Statistical methods and designs for correlated outcome and covariate errors in studies of HIV/AIDS
HIV/艾滋病研究中相关结果和协变量误差的统计方法和设计
批准号:
10618614
负责人:
Pamela A Shaw
金额:
$89.35万
依托单位国家:
美国
项目类别:
财政年份:
2018
资助国家:
美国
项目状态:
未结题
起止时间:
2018-01-25 至 2028-01-31

项目摘要

项目成果

Pamela A Shaw的其他基金

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
PROJECT SUMMARY/ASBTRACT Electronic health record (EHR) and other routinely collected data are often used as cost-effective data sources for HIV/AIDS research. These data sources, however, are known to be prone to errors, typically across multiple variables, which can lead to biased study results and misleading conclusions. In addition, EHR data sources often lack gold-standard measurements that are needed to clearly define the presence or absence of co-morbidities (e.g., liver fibrosis). To address limitations of EHR data sources, researchers can validate or collect additional data on a subsample of their patient records. By combining the rich, but error-prone EHR data on all study subjects with the gold-standard / validated data collected on a subsample of subjects, researchers can improve study estimates. Specifically, researchers can eliminate the bias of estimates had they only used the EHR data, and they can improve the precision (e.g., narrower confidence intervals) of study estimates had they only used the subsample with gold-standard / validated data. In earlier research, we developed statistical methods and software to combine EHR data with validated sub-samples of data. We developed optimal, multi-wave designs for targeting records for data validation. Importantly, we applied these methods to multiple HIV studies using retrospective observational data from the International epidemiology Databases to Evaluate AIDS (IeDEA). However, in our applications, we have encountered additional challenges that have not yet been addressed. In particular, there is great potential in combining expensive, prospectively collected, gold-standard data that are sparsely measured (e.g., once per year) on a sub-sample of patients with EHR data that are collected much more frequently on a larger number of patients. We will develop methods to handle this setting, and we will develop statistical designs to better select which participants should be approached for prospective data collection and which patient records should be validated. We will also develop statistical methods to address other challenges encountered with using EHR data, including how to incorporate validation data into studies when inclusion in the study is error-prone, and methods to address more complex types of data (e.g., interval censored data), for which there are a lack of techniques to handle error-prone data. Our methods and designs will focus on extensions of multiple imputation, maximum likelihood, and generalized raking techniques. Open source tools and tutorials will be developed to help researchers to implement these novel methods and study designs. The methods and designs will be applied to data from the IeDEA network to estimate the incidence of and risk factors for liver fibrosis/steatosis and frailty among people living with HIV in East Africa and Latin America.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
Statistical methods for correlated outcome and covariate errors in studies of HIV/AIDS
海外基金