A Poisson mixture model to identify changes in RNA polymerase II binding quantity using high-throughput sequencing technology.
A Poisson mixture model to identify changes in RNA polymerase II binding quantity using high-throughput sequencing technology.
复制标题
DOI:
10.1186/1471-2164-9-s2-s23
复制
发表时间:
2008-09-16
期刊:
影响因子:
4.4
通讯作者:
Li L
中科院分区:
文献类型:
--
作者:
Feng W;Liu Y;Wu J;Nephew KP;Huang TH;Li L
We present a mixture model-based analysis for identifying differences in the distribution of RNA polymerase II (Pol II) in transcribed regions, measured using ChIP-seq (chromatin immunoprecipitation following massively parallel sequencing technology). The statistical model assumes that the number of Pol II-targeted sequences contained within each genomic region follows a Poisson distribution. A Poisson mixture model was then developed to distinguish Pol II binding changes in transcribed region using an empirical approach and an expectation-maximization (EM) algorithm developed for estimation and inference. In order to achieve a global maximum in the M-step, a particle swarm optimization (PSO) was implemented. We applied this model to Pol II binding data generated from hormone-dependent MCF7 breast cancer cells and antiestrogen-resistant MCF7 breast cancer cells before and after treatment with 17β-estradiol (E2). We determined that in the hormone-dependent cells, ~9.9% (2527) genes showed significant changes in Pol II binding after E2 treatment. However, only ~0.7% (172) genes displayed significant Pol II binding changes in E2-treated antiestrogen-resistant cells. These results show that a Poisson mixture model can be used to analyze ChIP-seq data.