COMPLETE NUCLEOTIDE-SEQUENCE OF SV40 DNA

COMPLETE NUCLEOTIDE-SEQUENCE OF SV40 DNA
复制标题

DOI:
10.1038/273113a0
复制
发表时间:
1978-01-01
期刊:
影响因子:
64.8
通讯作者:
YSEBAERT, M
YSEBAERT, M
中科院分区:
综合性期刊1区
文献类型:
--
作者:
FIERS, W;CONTRERAS, R;YSEBAERT, M

文献摘要

被引文献

相似文献

通过对SV-40病毒5224碱基对DNA序列的测定,确定了已知基因在基因组上的精确位置。至少15.2%的基因组可能没有被翻译成多肽。完整序列揭示的特殊兴趣点是早期t和t[肿瘤]抗原在同一位置的起始,以及t抗原由基因组的两个不连续区域编码的事实;T抗原mRNA在编码区剪接。在后期区域,主要蛋白质VP1的基因与蛋白质VP2和VP3的基因重叠超过122个核苷酸,但在不同的框架中读取。由核苷酸序列推断出2个早期蛋白和晚期蛋白的几乎完整的氨基酸序列。后3种蛋白质的mRNA可能是从一个共同的初级RNA转录物中剪切出来的。简并密码子的使用显然是非随机的,但在早期和晚期区域是相似的。NUC、NCG和CGN [N = nucleotide base]型密码子缺失或非常罕见。
The determination of the total 5224 base-pair DNA sequence of the virus SV-40 gave the precise location of the known genes on the genome. At least 15.2% of the genome is presumably not translated into polypeptides. Particular points of interest revealed by the complete sequence are the initiation of the early t and T [tumor] antigens at the same position and the fact that the T antigen is coded by 2 non-contiguous regions of the genome; the T antigen mRNA is spliced in the coding region. In the late region the gene for the major protein VP1 overlaps those for proteins VP2 and VP3 over 122 nucleotides but is read in a different frame. The almost complete amino acid sequences of the 2 early proteins and those of the late proteins were deduced from the nucleotide sequence. The mRNA for the latter 3 proteins are presumably spliced out of a common primary RNA transcript. The use of degenerate codons is decidedly non-random, but is similar for the early and late regions. Codons of the type NUC, NCG and CGN [N = nucleotide base] are absent or very rare.