HiTEC: accurate error correction in high-throughput sequencing data

HiTEC: accurate error correction in high-throughput sequencing data
复制标题

DOI:
10.1093/bioinformatics/btq653
复制
发表时间:
2011-02-01
期刊:
影响因子:
5.8
通讯作者:
Ilie, Silvana
Ilie, Silvana
中科院分区:
生物学3区
文献类型:
--
作者:
Ilie, Lucian;Fazayeli, Farideh;Ilie, Silvana

文献摘要

被引文献

相似文献

动机:高通量测序技术产生非常大量的数据,并且测序错误构成分析此类数据的主要问题之一。目前用于纠正这些错误的算法是不是很准确,不自动适应给定的data.Results:我们提出HiTEC,一种算法,提供了一个高度准确的,强大的和全自动的方法来纠正读取产生的高通量测序方法。我们的方法提供了显着更高的准确性比以前的方法。它具有时间和空间效率,并且对于所有读取长度、基因组大小和覆盖水平都非常有效。
Motivation: High-throughput sequencing technologies produce very large amounts of data and sequencing errors constitute one of the major problems in analyzing such data. Current algorithms for correcting these errors are not very accurate and do not automatically adapt to the given data.Results: We present HiTEC, an algorithm that provides a highly accurate, robust and fully automated method to correct reads produced by high-throughput sequencing methods. Our approach provides significantly higher accuracy than previous methods. It is time and space efficient and works very well for all read lengths, genome sizes and coverage levels.