课题基金 / 基金详情

Statistical methods for population and family-based whole-genome sequence data

Statistical methods for population and family-based whole-genome sequence data
基于人群和家系的全基因组序列数据的统计方法
批准号:
9002848
负责人:
Wei Chen
金额:
$34.05万
依托单位国家:
美国
项目类别:
财政年份:
2014
资助国家:
美国
项目状态:
已结题
起止时间:
2014-04-25 至 2019-01-31

项目摘要

项目成果

Wei Chen的其他基金

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
DESCRIPTION (provided by applicant): Emerging sequencing technologies have made whole-genome sequencing become available for researches to study various phenotypes/diseases of interest, particularly focusing on rare variants sites. Although the first batch of sequencing projects has mainly focused on the analysis of unrelated individuals, numerous sequencing studies including related individuals have been carried out or launched recently as the sequencing cost reduces rapidly. However, the methodologies for analyzing family-based sequence data are largely falling behind partially due to the complexity of family structures and computational barrier. In this study, our primary goals are to efficiently and accurately infer individual genotypes and haplotypes - the key component of any sequencing project - by combining information from both family and population levels, and to study how differential sequencing errors will affect downstream association analysis. To achieve these goals, we propose specific aims as follows: 1) We will propose a novel statistical framework for genotyping calling and haplotype inference of sequence data including relative individuals. The new method takes advantages of both short stretches shared between unrelated individuals and long stretches shared between family members in a computationally feasible manner while retaining a high degree of accuracy via the synergy between two classic approaches: hidden Markov model (HMM) for linkage disequilibrium information and Lander-Green algorithm for inheritance vectors; 2) We will develop an exact algorithm for HMM computation to speed up a class of widely use genetics programs, including the method developed in Aim 1, without any sacrifice of accuracy; 3) We will assess the impact of sequencing errors on family-based association methods for rare variants and use the intrinsic stochastic nature of the proposed methods in Aim 1 to reduce the false positives under a framework of multiple imputation; 4) We will test and recalibrate our developed methods in collaboration with ongoing sequencing projects and systematically investigate different study designs. Successful completion of these aims will yield state-of-the-art statistical methods and software, which will facilitate the fast growing sequencing projects including family members and guide the design and analysis of future studies.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
An ensemble deep learning model for tumor bud detection and risk stratification in colorectal carcinoma.
  • 批准号:
    10564824
  • 项目类别:
  • 资助金额:
    $54.37万
  • 财政年份:
    2023
  • 负责人:
    Wei Chen
  • 依托单位:
Establishing translational neuroimaging tools for quantitative assessment of energy metabolism and metabolic reprogramming in healthy and diseased human brain at 7T
  • 批准号:
    10714863
  • 项目类别:
  • 资助金额:
    $63.02万
  • 财政年份:
    2023
  • 负责人:
    Wei Chen
  • 依托单位:
SCH: New Advanced Machine Learning Framework for Mining Heterogeneous Ocular Data to Accelerate
SCH: New Advanced Machine Learning Framework for Mining Heterogeneous Ocular Data to Accelerate
海外基金