Multi-omic data integration enables discovery of hidden biological regularities.

Multi-omic data integration enables discovery of hidden biological regularities.
复制标题

DOI:
10.1038/ncomms13091
复制
发表时间:
2016-10-26
影响因子:
16.6
通讯作者:
Palsson, Bernhard O.
Palsson, Bernhard O.
中科院分区:
综合性期刊1区
文献类型:
--
作者:
Ebrahim, Ali;Brunk, Elizabeth;Tan, Justin;O'Brien, Edward J.;Kim, Donghyuk;Szubin, Richard;Lerman, Joshua A.;Lechner, Anna;Sastry, Anand;Bordbar, Aarash;Feist, Adam M.;Palsson, Bernhard O.

文献摘要

参考文献

被引文献

相似文献

生物数据集的规模和复杂性的快速增长导致了“大数据到知识”的挑战。我们开发了先进的数据集成方法,用于基因组,转录组,核糖体分析,蛋白质组和通量组数据的多层次分析。首先,我们表明,主要组学数据的成对整合揭示了在大肠杆菌中将细胞过程联系在一起的信息:每个mRNA转录物产生的蛋白质分子的数量和每个翻译的蛋白质分子所需的核糖体的数量。其次,我们表明,基因组规模的模型,基因组和书目数据的基础上,使不同的数据类型的定量同步。将组学数据与模型相结合,发现了两个新的问题:酶的条件不变体内周转率和蛋白质结构基序与翻译暂停的相关性。这些信息可以用一种可计算的格式来正式表示,从而可以对细胞生理学基础的适应性和选择进行连贯的解释和预测。 将组学数据集转化为生物学洞察力是我们这个时代的巨大挑战之一。在这里,作者通过跨条件的不变量同步组学数据类型对,并将数据集整合到E.大肠杆菌代谢和基因表达。
Rapid growth in size and complexity of biological data sets has led to the ‘Big Data to Knowledge' challenge. We develop advanced data integration methods for multi-level analysis of genomic, transcriptomic, ribosomal profiling, proteomic and fluxomic data. First, we show that pairwise integration of primary omics data reveals regularities that tie cellular processes together in Escherichia coli: the number of protein molecules made per mRNA transcript and the number of ribosomes required per translated protein molecule. Second, we show that genome-scale models, based on genomic and bibliomic data, enable quantitative synchronization of disparate data types. Integrating omics data with models enabled the discovery of two novel regularities: condition invariant in vivo turnover rates of enzymes and the correlation of protein structural motifs and translational pausing. These regularities can be formally represented in a computable format allowing for coherent interpretation and prediction of fitness and selection that underlies cellular physiology. Translating omics data sets into biological insight is one of the great challenges of our time. Here, the authors make headway by synchronising pairs of omics data types via invariants across conditions and by integrating datasets into a genome-scale model of E. coli metabolism and gene expression.
DOI: 10.1126/science.1234012
发表时间: 2013-06-07
期刊: Science (New York, N.Y.)
影响因子: --
作者:
Chang RL;Andrews K;Kim D;Li Z;Godzik A;Palsson BO
通讯作者: Palsson BO
DOI: 10.2144/000114302
发表时间: 2015-06-01
期刊: BIOTECHNIQUES
影响因子: 2.7
作者:
Latif, Haythem;Szubin, Richard;Palsson, Bernhard O.
通讯作者: Palsson, Bernhard O.
DOI: 10.1038/nmeth.1923
发表时间: 2012-03-04
期刊: NATURE METHODS
影响因子: 48
作者:
Langmead, Ben;Salzberg, Steven L.
通讯作者: Salzberg, Steven L.
Biopython:用于计算分子生物学和生物信息学的免费 Python 工具。
DOI: 10.1093/bioinformatics/btp163
发表时间: 2009-06-01
期刊: Bioinformatics (Oxford, England)
影响因子: --
作者:
Cock PJ;Antao T;Chang JT;Chapman BA;Cox CJ;Dalke A;Friedberg I;Hamelryck T;Kauff F;Wilczynski B;de Hoon MJ
通讯作者: de Hoon MJ
DOI: 10.1093/nar/gkx1094
发表时间: 2018-01-04
影响因子: 14.9
作者:
Benson DA;Cavanaugh M;Clark K;Karsch-Mizrachi I;Ostell J;Pruitt KD;Sayers EW
通讯作者: Sayers EW