New computational methods to dynamically pinpointing the subregions carrying disease-associated rare variants
New computational methods to dynamically pinpointing the subregions carrying disease-associated rare variants
批准号:
10709565
负责人:
Jichun Xie
金额:
$38.86万
依托单位:
依托单位国家:
美国
项目类别:
财政年份:
2022
资助国家:
美国
项目状态:
未结题
起止时间:
2022-09-23 至 2026-07-31
关键词:
AddressAmyotrophic Lateral SclerosisBioconductorChromosomesCodeComplexComputing MethodologiesData ProtectionData SetDiseaseEvaluationFoundationsGenesGenetic DiseasesGenomeGenomicsGoalsGrantHigh-Throughput Nucleotide SequencingIndividualLightningMathematicsMeasuresMeta-AnalysisMethodologyMethodsMutationNoiseOutputPathogenicityPatternPopulationPrivacyResearch PersonnelResolutionSecureSignal TransductionSiteSpeedStatistical Data InterpretationStatistical MethodsTechnologyTestingUntranslated RNAVariantVisualizationVisualization softwareanalysis pipelinecausal variantcomputerized toolsdata preservationdata privacydata sharingdesigndisease mechanisms studyexome sequencingexperiencegenetic varianthuman diseaseimprovedinsightmosaicnovelopen sourcepublic repositoryrare variantsimulationsoftware developmenttooltraittranslational studyvirulence gene
中文摘要
点击翻译按钮获取中文摘要
英文摘要
PROJECT SUMMARY/ABSTRACT
The high-throughput sequencing technology allows us to query both common and rare variants for complex
human diseases. When variants are rare, single variant association analyses suffer from low power. To increase
power, existing whole-exome sequencing studies often aggregate the rare variants (RVs) across an entire gene
to study their collective effect. Presumably, when a gene harbors many pathogenic RVs, the aggregation will
increase the signal-to-noise ratio and thus the power. However, a gene often carries many mutations, while only
a subset will lead to novel or altered activities. These mutations usually do not distribute uniformly across the
entire gene or domain. For genes whose functional mutations are localized or concentrated to the specific
subregions, aggregating all the RVs across the entire gene or domain will dilute the signal, resulting in a loss of
power. Besides, even if the gene- or domain-based analysis can identify the pathogenic genes, they cannot
pinpoint the pathogenic subregions. Pinpointing the pathogenic subregions is preferred because it is usually
more unified in function and will be more informative to the downstream disease mechanism and translational
studies. To address these concerns and needs, we propose a novel statistical and computational method for
rare-variant association analysis with the three main features. First, it automatically searches the GVSs with
different sizes for their disease associations to optimize power. Second, it can pinpoint the disease-associated
GVSs with high resolution to facilitate the downstream disease mechanism studies. Third, it can be easily
customized to fit the special needs, such as preserving data privacy, incorporating functional annotations, and
adjusting for varying ancestry loadings for admixed populations. We will establish a rigorous mathematical and
statistical foundation for the GVS analysis and develop the software to realize its implementation on high-
throughput sequencing studies. We will apply our method to an ongoing whole-exome sequencing study of
amyotrophic lateral sclerosis (ALS) to identify ALS-related genomic subregions.
期刊论文(1)
专著(0)
科研奖励(0)
会议论文
海外基金