Base-calling of automated sequencer traces using phred.: II.: Error probabilities

Base-calling of automated sequencer traces using phred.: II.: Error probabilities
复制标题

DOI:
10.1101/gr.8.3.186
复制
发表时间:
1998-03-01
期刊:
影响因子:
7
通讯作者:
Green, P
Green, P
中科院分区:
生物学1区
文献类型:
--
作者:
Ewing, B;Green, P

文献摘要

被引文献

相似文献

消除高通量测序中的数据处理瓶颈将需要提高数据处理软件的准确性和该准确性的可靠测量。我们已经开发并实现了在我们的基地调用程序phred的能力,估计错误的概率为每个基地调用,作为从跟踪数据计算的某些参数的函数。这些错误概率在此显示为有效的(对应于实际错误率),并且对于在几种不同化学和电泳条件下收集的读取数据,具有区分正确碱基判定与不正确碱基判定的高能力。他们在我们的装配程序phrap和我们的整理程序consed中起着至关重要的作用。
Elimination of the data processing bottleneck in high-throughput sequencing will require both improved accuracy of data processing software and reliable measures of that accuracy. We have developed and implemented in our base-calling program phred the ability to estimate a probability of error for each base-call, as a function of certain parameters computed from the trace data. These error probabilities are shown here to be valid (correspond to actual error rates] and to have high power to discriminate correct base-calls from incorrect ones, For read data collected under several different chemistries and electrophoretic conditions. They play a critical role in our assembly program phrap and our finishing program consed.