Ensembl Genomes 2016: more genomes, more complexity.
Ensembl Genomes 2016: more genomes, more complexity.
复制标题
DOI:
10.1093/nar/gkv1209
复制
发表时间:
2016-01-04
影响因子:
14.9
通讯作者:
Staines DM
中科院分区:
文献类型:
--
作者:
Kersey PJ;Allen JE;Armean I;Boddu S;Bolt BJ;Carvalho-Silva D;Christensen M;Davis P;Falin LJ;Grabmueller C;Humphrey J;Kerhornou A;Khobova J;Aranganathan NK;Langridge N;Lowy E;McDowall MD;Maheswari U;Nuhn M;Ong CK;Overduin B;Paulini M;Pedro H;Perry E;Spudich G;Tapanari E;Walts B;Williams G;Tello-Ruiz M;Stein J;Wei S;Ware D;Bolser DM;Howe KL;Kulesha E;Lawson D;Maslen G;Staines DM
Ensembl Genomes (http://www.ensemblgenomes.org) is an integrating resource for genome-scale data from non-vertebrate species, complementing the resources for vertebrate genomics developed in the context of the Ensembl project (http://www.ensembl.org). Together, the two resources provide a consistent set of programmatic and interactive interfaces to a rich range of data including reference sequence, gene models, transcriptional data, genetic variation and comparative analysis. This paper provides an update to the previous publications about the resource, with a focus on recent developments. These include the development of new analyses and views to represent polyploid genomes (of which bread wheat is the primary exemplar); and the continued up-scaling of the resource, which now includes over 23 000 bacterial genomes, 400 fungal genomes and 100 protist genomes, in addition to 55 genomes from invertebrate metazoa and 39 genomes from plants. This dramatic increase in the number of included genomes is one part of a broader effort to automate the integration of archival data (genome sequence, but also associated RNA sequence data and variant calls) within the context of reference genomes and make it available through the Ensembl user interfaces.