Nonparametric methods for functional and translational genomics
Nonparametric methods for functional and translational genomics
批准号:
8532014
负责人:
James Bentley Brown
金额:
$10.3万
依托单位国家:
美国
项目类别:
财政年份:
2012
资助国家:
美国
项目状态:
已结题
起止时间:
2012-08-17 至 2014-07-31
关键词:
AlgorithmsAnimal Disease ModelsAnimal ModelAreaAutomobile DrivingAwardBase PairingBiochemicalBiologicalBiological AssayBiological ModelsBiological ProcessCellsChIP-seqCommunicationComplementary DNAComplexDataData AnalysesData SourcesDevelopmentDevelopmental BiologyDisease modelElementsGap JunctionsGene DeletionGenesGenomeGenomicsGoalsHumanHuman BiologyIndiumIndividualLeadLinkMapsMeasuresMentorsMethodsMetricModelingMolecularMutationOrphanOrthologous GenePathway AnalysisPharmaceutical PreparationsPhenotypePlayProblem SolvingPropertyProtein IsoformsRNAReadingResearchResearch PersonnelRunningSemanticsSystemTechniquesTechnologyToxic effectTrainingTraining ActivityTranscriptTranscriptional RegulationVariantWeightanalogbasecareer developmentdesigndriving forceexperiencefunctional genomicshigh throughput screeninghuman diseasenetwork modelsnext generation sequencingnovel strategiesstatisticsstem cell biologytheoriestooltranscription factortranscriptome sequencing
中文摘要
点击翻译按钮获取中文摘要
英文摘要
DESCRIPTION (provided by applicant): Next generation sequencing has revealed the molecular landscape of cells in unprecedented detail. However, for the massively large-scale data produced by assays based on these technologies, informativeness is not only a function of wet-lab technology, but is critically also a function of the analytical pipelines that interpret th data. Our group has developed four statistical tools designed maximize the informativeness of these assays: 1) the Genome Structural Correction (GSC), a nonparametric model of genomic annotations used to assess the significance of relationships between features; 2) the Irreproducible Discovery Rate (IDR), an analogue of the FDR that leverages information from biological replicates; 3) Statmap, a comprehensive analysis pipeline for ChIP-seq and CAGE data that propagates statistical confidence from base-calling to peak-calling; and 4) Sparse Linear Isoform Discovery and abundance Estimation (SLIDE), an integrative statistical framework for the analysis of RNA-seq, cDNA, and other RNA data aimed at obtaining and quantifying de novo transcript models. These tools are designed to identify and characterize functional elements in genomes; they make minimal assumptions about the data they analyze, and therefore draw reliable conclusions and measures of statistical confidence. During the K99, we will expand and integrate our tools to extend the reach of statistical confidence throughout data interpretation. During the R00, my research will progress toward the inference and assessment of biological networks. Just as ortholog identification has become an essential step in developing animal models of human disease, multi-species network analysis promises to become a key step in interpreting the relationship between genome variation and phenotype. Many mutations, even gene deletions, do not reveal an obvious phenotype. This is due to network robustness, which often differs between closely related species. To understand these phenomena, we aim to: 1) develop standard statistical tools for network inference, and 2) develop "meta models" of networks that will permit general measures of network orthology. These two aims are tightly linked: we will need critically to characterize the semantics of biological networks to model them. Currently, some models lack consistent definitions of edges and weights, resulting in untestable representations of genomics data. We will develop testable, quantitative models of biological processes, establishing a uniform semantics leveraging the rich theory of complex systems. Each of the tools above will play a key role, especially Statmap and the GSC, which will be needed to propagate statistical confidence into network analysis. Advances will have a transformative effect on our ability to map animal models of disease onto human biology. Nearly nine out of ten new drugs fail in human trials due to issues (e.g. toxicity) not present in animal models. Understanding the orthology not just of individual genes, but of entire biochemical networks will be essential to infer and correct for differences between models of disease and human biology. Solving this problem will be a major step forward in the march from "base-pairs to bedside".
期刊论文(5)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
DOI:
10.1101/gr.132159.111
发表时间:
2012-09
期刊:
Genome research
影响因子:
7
作者:
[Derrien T, Johnson R, Bussotti G, Tanzer A, Djebali S, Tilgner H, Guernec G, Martin D, Merkel A, Knowles DG, Lagarde J, Veeravalli L, Ruan X, Ruan Y, Lassmann T, Carninci P, Brown JB, Lipovich L, Gonzalez JM, Thomas M, Davis CA, Shiekhattar R, Gingeras TR, Hubbard TJ, Notredame C, Harrow J, Guigó R]
通讯作者:
Guigó R
DOI:
10.1101/gr.134767.111
发表时间:
2012-09
期刊:
Genome research
影响因子:
7
作者:
[Bánfai B, Jia H, Khatun J, Wood E, Risk B, Gundling WE Jr, Kundaje A, Gunawardena HP, Yu Y, Xie L, Krajewski K, Strahl BD, Chen X, Bickel P, Giddings MC, Brown JB, Lipovich L]
通讯作者:
Lipovich L
DOI:
10.1186/gb-2012-13-9-r48
发表时间:
2012-09-26
期刊:
Genome biology
影响因子:
12.3
作者:
[Yip KY, Cheng C, Bhardwaj N, Brown JB, Leng J, Kundaje A, Rozowsky J, Birney E, Bickel P, Snyder M, Gerstein M]
通讯作者:
Gerstein M
DOI:
10.1101/gr.136184.111
发表时间:
2012-09
期刊:
Genome research
影响因子:
7
作者:
[Landt SG, Marinov GK, Kundaje A, Kheradpour P, Pauli F, Batzoglou S, Bernstein BE, Bickel P, Brown JB, Cayting P, Chen Y, DeSalvo G, Epstein C, Fisher-Aylor KI, Euskirchen G, Gerstein M, Gertz J, Hartemink AJ, Hoffman MM, Iyer VR, Jung YL, Karmakar S, Kellis M, Kharchenko PV, Li Q, Liu T, Liu XS, Ma L, Milosavljevic A, Myers RM, Park PJ, Pazin MJ, Perry MD, Raha D, Reddy TE, Rozowsky J, Shoresh N, Sidow A, Slattery M, Stamatoyannopoulos JA, Tolstorukov MY, White KP, Xi S, Farnham PJ, Lieb JD, Wold BJ, Snyder M]
通讯作者:
Snyder M
Promoter analysis reveals globally differential regulation of human long non-coding RNA and protein-coding genes.
启动子分析揭示了人类长的非编码RNA和蛋白质编码基因的全球差异调节。
DOI:
10.1371/journal.pone.0109443
发表时间:
2014
期刊:
PloS one
影响因子:
3.7
作者:
[Alam T, Medvedeva YA, Jia H, Brown JB, Lipovich L, Bajic VB]
通讯作者:
Bajic VB
Nonparametric methods for functional and translational genomics
-
批准号:8916814
-
项目类别:
-
资助金额:$24.9万
-
财政年份:2014
-
负责人:James Bentley Brown
-
依托单位:
Nonparametric methods for functional and translational genomics
-
批准号:8280729
-
项目类别:
-
资助金额:$10.3万
-
财政年份:2012
-
负责人:James Bentley Brown
-
依托单位:
海外基金