PairMotifChIP: A Fast Algorithm for Discovery of Patterns Conserved in Large ChIP-seq Data Sets.
PairMotifChIP: A Fast Algorithm for Discovery of Patterns Conserved in Large ChIP-seq Data Sets.
复制标题
PairMotifChIP:一种用于发现大型 ChIP-seq 数据集中保守模式的快速算法
DOI:
10.1155/2016/4986707
复制
发表时间:
2016
影响因子:
--
通讯作者:
Feng D
中科院分区:
文献类型:
--
作者:
Yu Q;Huo H;Feng D
Identifying conserved patterns in DNA sequences, namely, motif discovery, is an important and challenging computational task. With hundreds or more sequences contained, the high-throughput sequencing data set is helpful to improve the identification accuracy of motif discovery but requires an even higher computing performance. To efficiently identify motifs in large DNA data sets, a new algorithm called PairMotifChIP is proposed by extracting and combining pairs of l-mers in the input with relatively small Hamming distance. In particular, a method for rapidly extracting pairs of l-mers is designed, which can be used not only for PairMotifChIP, but also for other DNA data mining tasks with the same demand. Experimental results on the simulated data show that the proposed algorithm can find motifs successfully and runs faster than the state-of-the-art motif discovery algorithms. Furthermore, the validity of the proposed algorithm has been verified on real data.
登录
查看更多内容
影响因子:
14.9
作者:
GuhaThakurta D
通讯作者:
GuhaThakurta D
影响因子:
46.9
作者:
Tompa, M;Li, N;Zhu, Z
通讯作者:
Zhu, Z
影响因子:
3.9
作者:
Yu, Qiang;Huo, Hongwei;Huan, Jun
通讯作者:
Huan, Jun
影响因子:
1.7
作者:
Buhler, J;Tompa, M
通讯作者:
Tompa, M
影响因子:
46.9
作者:
D'haeseleer, P
通讯作者:
D'haeseleer, P