Statistical Power Calculations for ChIP-seq experiments
Statistical Power Calculations for ChIP-seq experiments
批准号:
8284083
负责人:
Sunduz Keles
金额:
$18.41万
依托单位国家:
美国
项目类别:
财政年份:
2012
资助国家:
美国
项目状态:
已结题
起止时间:
2012-05-01 至 2014-03-31
关键词:
AllelesBase SequenceBindingBinomial ModelBioconductorBiologic CharacteristicBiologicalCellsChIP-seqCommunitiesComputer AnalysisComputer softwareDNADataData AnalysesData SetDatabasesDetectionDevelopmentDiagnosisDiseaseFamilyGene ExpressionGeneticGenomeGenomicsGoalsGuanine + Cytosine CompositionImmune SeraImmunoglobulin GIndividualLeadLettersLocationMapsMethodsModelingPlayPublic HealthReadingResearchResearch PersonnelResourcesRoleSamplingSimulateStagingStatistical ModelsTechnologyTissuesTrainingUnited States National Institutes of HealthValidationVariantbasechromatin immunoprecipitationdesignepigenomicsgenome sequencinggenome wide association studygenome-widehuman diseasenext generationnovelprogramsresearch studysimulationsoftware developmenttooltranscription factor
中文摘要
点击翻译按钮获取中文摘要
英文摘要
DESCRIPTION (provided by applicant): The advent of high throughput next generation sequencing (NGS) technologies have revolutionized the fields of genetics and genomics by allowing rapid and inexpensive sequencing of billions of bases. Among the NGS applications, ChIP-seq (chromatin immunoprecipitation followed by NGS) is perhaps the most successful to date. ChIP-seq technology enables investigators to study genome-wide binding of transcription factors and mapping of epigenomic marks. Both of these play crucial roles in programming of cell specific gene expression; therefore their genome-wide mapping can significantly advance our ability to understand and diagnose human diseases. Although basic analysis tools for ChIP-seq data are rapidly increasing, there has not been much progress on the design problems regarding ChIP-seq experiments. A challenging question that the researchers planning a ChIP-seq experiment need to answer is: how deeply should the ChIP and the control samples be sequenced? The answer depends on multiple factors some of which can be set by the experimenter based on pilot/preliminary data. The sequencing depth of a ChIP-seq experiment is one of the key factors that determine whether or not all the underlying targets (e.g., binding locations or epigenomic profiles) can be identified with a targeted power. This is especially important when the goal is the analysis of individual-to-individual and allele specific variation o transcription factor binding and epigenomic profiles. Insufficient sequencing depths may lead to spurious differences in binding or epigenome profiles. In this proposal, we aim to develop a general framework for power calculations in ChIP-seq experiments with three specific aims and by considering statistical models commonly used in ChIP-seq analysis: (1) Power calculations based on the conditional Binomial model; (2) Power calculations based on the Poisson and Negative Binomial regression models; (3) A power calculation tool for GALAXY and Bioconductor. This project will be accomplished through a combination of theoretical/methodological development, simulation, computational analysis, and experimental validation. Methods will be developed and evaluated using datasets from the ENCODE, modENCODE, and the RoadMap Epigenomics consortiums as well as novel datasets from collaborators. Statistical resources generated from the project, which will be disseminated in publicly available software, will provide essential tools for the efficient design of ChIP-seq experiments.
PUBLIC HEALTH RELEVANCE: The proposed research is relevant to public health because capturing genome-wide binding of transcription factors and epigenomic information by ChIP-seq technology is invaluable for comprehensively understanding development, differentiation, and disease. Design of ChIP-seq experiments present unprecedented challenges. We will develop a statistical framework for power calculations in designing ChIP-seq experiments and disseminate results and software to the research community.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
Statistical methods for co-expression network analysis of population-scale scRNA-seq data
-
批准号:10740240
-
项目类别:
-
资助金额:$40.76万
-
财政年份:2023
-
负责人:Sunduz Keles
-
依托单位:
Functionally relevant mapping of human GWAS SNPs on model organisms
-
批准号:10056966
-
项目类别:
-
资助金额:$40.05万
-
财政年份:2020
-
负责人:Sunduz Keles
-
依托单位:
High dimensional statistical data modeling and integration for studying regulatory variation
-
批准号:10413927
-
项目类别:
-
资助金额:$37.88万
-
财政年份:2007
-
负责人:Sunduz Keles
-
依托单位:
Statistical Analysis Methods and Software for ChIP-seq Data
-
批准号:8605900
-
项目类别:
-
资助金额:$29.95万
-
财政年份:2007
-
负责人:Sunduz Keles
-
依托单位:
Statistical Analysis Methods and Software for ChIP-seq Data
-
批准号:8785690
-
项目类别:
-
资助金额:$29.8万
-
财政年份:2007
-
负责人:Sunduz Keles
-
依托单位:
Statistical Methods for the Analysis of ChlP-chip Data
-
批准号:7253510
-
项目类别:
-
资助金额:$28.24万
-
财政年份:2007
-
负责人:Sunduz Keles
-
依托单位:
Statistical Analysis Methods and Software for ChIP-seq Data
-
批准号:8370723
-
项目类别:
-
资助金额:$29.52万
-
财政年份:2007
-
负责人:Sunduz Keles
-
依托单位:
Statistical Methods for the Analysis of ChlP-chip Data
-
批准号:7799293
-
项目类别:
-
资助金额:$28.19万
-
财政年份:2007
-
负责人:Sunduz Keles
-
依托单位:
High dimensional statistical data integration for studying regulatory variation
-
批准号:9344668
-
项目类别:
-
资助金额:$32.5万
-
财政年份:2007
-
负责人:Sunduz Keles
-
依托单位:
High dimensional statistical data modeling and integration for studying regulatory variation
-
批准号:10610872
-
项目类别:
-
资助金额:$37.88万
-
财政年份:2007
-
负责人:Sunduz Keles
-
依托单位:
Statistical Methods for the Analysis of ChlP-chip Data
-
批准号:7413330
-
项目类别:
-
资助金额:$28.47万
-
财政年份:2007
-
负责人:Sunduz Keles
-
依托单位:
Statistical Methods for the Analysis of ChlP-chip Data
-
批准号:7616521
-
项目类别:
-
资助金额:$28.47万
-
财政年份:2007
-
负责人:Sunduz Keles
-
依托单位:
High dimensional statistical data modeling and integration for studying regulatory variation
-
批准号:10213308
-
项目类别:
-
资助金额:$36.46万
-
财政年份:2007
-
负责人:Sunduz Keles
-
依托单位:
海外基金