Systems-Level Annotation of a Metabolomics Data Set Reduces 25 000 Features to Fewer than 1000 Unique Metabolites.

Systems-Level Annotation of a Metabolomics Data Set Reduces 25 000 Features to Fewer than 1000 Unique Metabolites.
复制标题

DOI:
10.1021/acs.analchem.7b02380
复制
发表时间:
2017-10-03
影响因子:
7.4
通讯作者:
Patti GJ
Patti GJ
中科院分区:
化学1区
文献类型:
--
作者:
Mahieu NG;Patti GJ

文献摘要

参考文献

被引文献

相似文献

当使用液相色谱/质谱 (LC/MS) 进行非靶向代谢组学时,现在可以常规检测生物样品中的数万个特征。然而,对数据的了解不足导致解释变得复杂,并掩盖了实验中实际测量的独特代谢物的数量。在这里,我们对使用一种非靶向代谢组学方法分析的大肠杆菌样品中检测到的独特代谢物的数量设定了上限。我们首先使用上下文驱动的注释方法对同一分析物产生的多个特征进行分组,我们将其称为“简并特征”。令人惊讶的是,这项分析揭示了数千种以前未报告的简并性,使独特分析物的数量减少到约 2,961 种。然后,我们采用正交方法,使用基于 13C 的认证技术从数据中删除非生物特征。这进一步将独特分析物的数量减少到 1,000 种以下。我们的数据减少了 90%,比之前发表的任何研究都要多五倍。根据结果​​,我们提出了一种依赖于彻底注释的参考数据集的非靶向代谢组学替代方法。为此,我们引入了 creDBle 数据库 (http://creDBle.wustl.edu),其中包含精确质量、保留时间和 MS/MS 碎片数据以及所有认证特征的注释。
When using liquid chromatography/mass spectrometry (LC/MS) to perform untargeted metabolomics, it is now routine to detect tens of thousands of features from biological samples. Poor understanding of the data, however, has complicated interpretation and masked the number of unique metabolites actually being measured in an experiment. Here we place an upper bound on the number of unique metabolites detected in Escherichia coli samples analyzed with one untargeted metabolomic method. We first group multiple features arising from the same analyte, which we call “degenerate features”, using a context-driven annotation approach. Surprisingly, this analysis revealed thousands of previously unreported degeneracies that reduced the number of unique analytes to ~2,961. We then applied an orthogonal approach to remove non-biological features from the data using the 13C-based credentialing technology. This further reduced the number of unique analytes to less than 1,000. Our 90% reduction in data is five fold greater than any previously published studies. On the basis of the results, we propose an alternative approach to untargeted metabolomics that relies on thoroughly annotated reference data sets. To this end, we introduce the creDBle database (http://creDBle.wustl.edu), which contains accurate mass, retention time, and MS/MS fragmentation data as well as annotations of all credentialed features.
DOI: 10.1093/bioinformatics/btr079
发表时间: 2011-04-15
期刊: Bioinformatics (Oxford, England)
影响因子: --
作者:
Brown M;Wedge DC;Goodacre R;Kell DB;Baker PN;Kenny LC;Mamas MA;Neyses L;Dunn WB
通讯作者: Dunn WB
DOI: 10.1021/ac400751j
发表时间: 2013-08-20
影响因子: 7.4
作者:
Nikolskiy, Igor;Mahieu, Nathaniel G.;Chen, Ying-Jr;Tautenhahn, Ralf;Patti, Gary J.
通讯作者: Patti, Gary J.
DOI: 10.1093/bioinformatics/btu370
发表时间: 2014-10
期刊: Bioinformatics (Oxford, England)
影响因子: --
作者:
Daly R;Rogers S;Wandy J;Jankevics A;Burgess KE;Breitling R
通讯作者: Breitling R
DOI: 10.1021/ac1021166
发表时间: 2010-12-01
影响因子: 7.4
作者:
Melamud E;Vastag L;Rabinowitz JD
通讯作者: Rabinowitz JD
DOI: 10.1101/gr.6.9.807
发表时间: 1996-09-01
期刊: GENOME RESEARCH
影响因子: 7
作者:
Hillier, L;Lennon, G;Marra, M
通讯作者: Marra, M