Distribution and evolution of sequence characteristics in the E. coli genome.
Distribution and evolution of sequence characteristics in the E. coli genome.
复制标题
大肠杆菌基因组中序列特征的分布和进化。
DOI:
10.1080/07391102.1986.10506347
复制
发表时间:
1986
影响因子:
4.4
通讯作者:
Earley,S
中科院分区:
文献类型:
--
作者:
Blake,RD;Earley,S
The mean (G+C) composition (51.0%) and standard deviation (±3.8%) of published DNA sequences accounting for 10% of theE. coligenome is in excellent agreement with the principal overall distribution determined by high resolution melting. While differences in base and neighbor characteristics are small and uniform throughout all regions of the genome, it is found that the (G+C) content of sequences varies in segmented fashion within boundaries corresponding to coding (53% G+C) and noncoding (46% G+C) regions; with variances in the latter being six-fold greater than in coding regions. The variance in different regions shows a strong negative dependence on (G+C) content of the region, reflecting the condition that A-T and G-C base pairs are preferred neighbors of A-T and C-G pairs, respectively; with the bias increasing with decreasing (G+C) content. Neighbor analysis indicates the most extreme positive biases occur in AA, TT, GC and CG throughout all regions, but particularly in noncoding regions. Extraordinary numbers of oligomeric strings of (A)n, etc., are the further consequence of this bias. These and other characteristics point to the existence of inherent biases in neighbor frequencies levied during replication or repair, and which reflect, in turn, neighbor influences during mutation. The bias in codon usage noted by Grantham and others is seen here as due, in part, to the adaptation of coding sequences to this microenvironment through selection among synonymous codons so as to preserve inherent neighbor biases.