Maximal Words in Sequence Comparisons Based on Subword Composition
Maximal Words in Sequence Comparisons Based on Subword Composition
复制标题
基于子词组合的序列比较中最大词数
DOI:
--
复制
发表时间:
2010
期刊:
影响因子:
--
通讯作者:
A. Apostolico
中科院分区:
文献类型:
--
作者:
A. Apostolico
Measures of sequence similarity and distance based more or less explicitly on subword composition are attracting an increasing interest driven by intensive applications such as massive document classification and genome-wide molecular taxonomy. A uniform character of such measures is in some underlying notion of relative compressibility, whereby two similar sequences are expected to share a larger number of common substrings than two distant ones. This paper reviews some of the approaches to sequence comparison based on subword composition and suggests that their common denominator may ultimately reside in special classes of subwords, the nature of which resonates in interesting ways with the structure of popular subword trees and graphs.
DOI:
10.1016/s0168-9525(00)89076-9
发表时间:
1995-07
期刊:
Trends in genetics : TIG
影响因子:
--
作者:
Samuel Kariin;C. Burge
通讯作者:
Samuel Kariin;C. Burge