FIMO: scanning for occurrences of a given motif.

FIMO: scanning for occurrences of a given motif.
复制标题

DOI:
10.1093/bioinformatics/btr064
复制
发表时间:
2011-04-01
期刊:
Bioinformatics (Oxford, England)
影响因子:
--
通讯作者:
Noble WS
Noble WS
中科院分区:
其他
文献类型:
--
作者:
Grant CE;Bailey TL;Noble WS

文献摘要

参考文献

被引文献

相似文献

总结:基序是一段短的DNA或蛋白质序列,有助于其所在序列的生物学功能。在过去的几十年里,已经描述了许多用于识别、表征和搜索序列模体的计算方法。对于几乎任何基于基序的序列分析流水线来说,关键的是扫描序列数据库以寻找由位置特异性频率矩阵描述的给定基序的出现的能力。结果如下:我们描述了查找单个基序出现(FIMO),一个软件工具,用于扫描DNA或蛋白质序列的基序描述为位置特异性评分矩阵。该程序计算给定序列数据库中每个位置的对数似然比得分,使用已建立的动态编程方法将该得分转换为P值,然后应用错误发现率分析来估计给定序列中每个位置的q值。FIMO提供多种格式的输出,包括HTML、XML和几种圣克鲁斯基因组浏览器格式。该程序是高效的,允许在单个CPU上以3.5 Mb/s的速率扫描DNA序列。可用性和实施:FIMO是MEME Suite软件工具包的一部分。Web服务器和源代码可在http://meme.sdsc.edu上获得。联系方式:t.bailey@ imb.uq.edu.au; t.bailey@ imb.uq.edu.au补充信息:补充数据可在生物信息学在线获得。
Summary: A motif is a short DNA or protein sequence that contributes to the biological function of the sequence in which it resides. Over the past several decades, many computational methods have been described for identifying, characterizing and searching with sequence motifs. Critical to nearly any motif-based sequence analysis pipeline is the ability to scan a sequence database for occurrences of a given motif described by a position-specific frequency matrix. Results: We describe Find Individual Motif Occurrences (FIMO), a software tool for scanning DNA or protein sequences with motifs described as position-specific scoring matrices. The program computes a log-likelihood ratio score for each position in a given sequence database, uses established dynamic programming methods to convert this score to a P-value and then applies false discovery rate analysis to estimate a q-value for each position in the given sequence. FIMO provides output in a variety of formats, including HTML, XML and several Santa Cruz Genome Browser formats. The program is efficient, allowing for the scanning of DNA sequences at a rate of 3.5 Mb/s on a single CPU. Availability and Implementation: FIMO is part of the MEME Suite software toolkit. A web server and source code are available at http://meme.sdsc.edu. Contact: t.bailey@imb.uq.edu.au; t.bailey@imb.uq.edu.au Supplementary information: Supplementary data are available at Bioinformatics online.
DOI: 10.1111/1467-9868.00346
发表时间: 2002-01-01
影响因子: 5.8
作者:
Storey, JD
通讯作者: Storey, JD
模因套件:用于发现和搜索的工具。
DOI: 10.1093/nar/gkp335
发表时间: 2009-07
影响因子: 14.9
作者:
Bailey TL;Boden M;Buske FA;Frith M;Grant CE;Clementi L;Ren J;Li WW;Noble WS
通讯作者: Noble WS
DOI: 10.1214/aos/1074290335
发表时间: 2003-12-01
影响因子: 4.5
作者:
Storey, JD
通讯作者: Storey, JD
DOI: 10.1093/bioinformatics/btg1054
发表时间: 2003-09-01
期刊: BIOINFORMATICS
影响因子: 5.8
作者:
Bailey, Timothy L.;Noble, William Stafford
通讯作者: Noble, William Stafford
DOI: 10.1093/bioinformatics/14.1.48
发表时间: 1998-01-01
期刊: BIOINFORMATICS
影响因子: 5.8
作者:
Bailey, TL;Gribskov, M
通讯作者: Gribskov, M