NestedMICA: sensitive inference of over-represented motifs in nucleic acid sequence.
NestedMICA: sensitive inference of over-represented motifs in nucleic acid sequence.
复制标题
嵌套:核酸序列中代表性过多的基序的敏感推断。
DOI:
10.1093/nar/gki282
复制
发表时间:
2005
影响因子:
14.9
通讯作者:
Hubbard, TJP
中科院分区:
文献类型:
--
作者:
Down, TA;Hubbard, TJP
NestedMICA is a new, scalable, pattern-discovery system for finding transcription factor binding sites and similar motifs in biological sequences. Like several previous methods, NestedMICA tackles this problem by optimizing a probabilistic mixture model to fit a set of sequences. However, the use of a newly developed inference strategy called Nested Sampling means NestedMICA is able to find optimal solutions without the need for a problematic initialization or seeding step. We investigate the performance of NestedMICA in a range scenario, on synthetic data and a well-characterized set of muscle regulatory regions, and compare it with the popular MEME program. We show that the new method is significantly more sensitive than MEME: in one case, it successfully extracted a target motif from background sequence four times longer than could be handled by the existing program. It also performs robustly on synthetic sequences containing multiple significant motifs. When tested on a real set of regulatory sequences, NestedMICA produced motifs which were good predictors for all five abundant classes of annotated binding sites.
登录
查看更多内容
影响因子:
5.8
作者:
Thijs, G;Lescot, M;Moreau, Y
通讯作者:
Moreau, Y
影响因子:
46.9
作者:
Tompa, M;Li, N;Zhu, Z
通讯作者:
Zhu, Z
影响因子:
14.9
作者:
Hubbard, T;Barker, D;Clamp, M
通讯作者:
Clamp, M
影响因子:
8
作者:
Saidi, SA;Holland, CM;Smith, SK
通讯作者:
Smith, SK
影响因子:
5.8
作者:
Bergman, CM;Carlson, JW;Celniker, SE
通讯作者:
Celniker, SE