Detecting protein variants by mass spectrometry: a comprehensive study in cancer cell-lines.

Detecting protein variants by mass spectrometry: a comprehensive study in cancer cell-lines.
复制标题

DOI:
10.1186/s13073-017-0454-9
复制
发表时间:
2017-07-18
期刊:
影响因子:
12.3
通讯作者:
Kislinger T
Kislinger T
中科院分区:
生物学1区
文献类型:
--
作者:
Alfaro JA;Ignatchenko A;Ignatchenko V;Sinha A;Boutros PC;Kislinger T

文献摘要

参考文献

被引文献

相似文献

肿瘤蛋白基因组学旨在了解癌症基因组的变化如何影响其蛋白质组。整合这些分子数据的一个挑战是从质谱(MS)数据集中识别异常蛋白质产物,因为传统的蛋白质组学分析只能从参考序列数据库中识别蛋白质。我们建立了蛋白质组学工作流程来检测MS数据集中的肽变体。我们使用了公开可用的群体变异(dbSNP和UniProt)和癌症体细胞变异(COSMIC)的组合,沿着样本特异性基因组和转录组数据,以检查59种癌细胞系内和之间的蛋白质组变异。我们开发了一套使用三种搜索算法检测变体的建议,用于FDR估计的分裂目标诱饵方法和多个搜索后过滤器。我们检查了730万个在任何参考蛋白质组中未发现的独特变体胰蛋白酶肽,并在NCI 60细胞系蛋白质组中的2200个基因中鉴定了4771个与参考蛋白质组相对应的体细胞和种系偏差突变。我们详细讨论了通过MS识别变体肽的技术和计算挑战,并表明揭示这些变体可以识别重要癌症基因内的可药用突变。本文的在线版本(doi:10.1186/s13073-017-0454-9)包含补充材料,可供授权用户使用。
Onco-proteogenomics aims to understand how changes in a cancer’s genome influences its proteome. One challenge in integrating these molecular data is the identification of aberrant protein products from mass-spectrometry (MS) datasets, as traditional proteomic analyses only identify proteins from a reference sequence database. We established proteomic workflows to detect peptide variants within MS datasets. We used a combination of publicly available population variants (dbSNP and UniProt) and somatic variations in cancer (COSMIC) along with sample-specific genomic and transcriptomic data to examine proteome variation within and across 59 cancer cell-lines. We developed a set of recommendations for the detection of variants using three search algorithms, a split target-decoy approach for FDR estimation, and multiple post-search filters. We examined 7.3 million unique variant tryptic peptides not found within any reference proteome and identified 4771 mutations corresponding to somatic and germline deviations from reference proteomes in 2200 genes among the NCI60 cell-line proteomes. We discuss in detail the technical and computational challenges in identifying variant peptides by MS and show that uncovering these variants allows the identification of druggable mutations within important cancer genes. The online version of this article (doi:10.1186/s13073-017-0454-9) contains supplementary material, which is available to authorized users.
DOI: 10.1186/1471-2105-13-s16-s2
发表时间: 2012
期刊: BMC bioinformatics
影响因子: 3
作者:
Jeong K;Kim S;Bandeira N
通讯作者: Bandeira N
DOI: 10.1021/acs.jproteome.5b00817
发表时间: 2016-03-04
影响因子: 4.4
作者:
Cesnik AJ;Shortreed MR;Sheynkman GM;Frey BL;Smith LM
通讯作者: Smith LM
DOI: 10.1002/pmic.201200439
发表时间: 2013-01-01
期刊: PROTEOMICS
影响因子: 3.4
作者:
Eng, Jimmy K.;Jahan, Tahmina A.;Hoopmann, Michael R.
通讯作者: Hoopmann, Michael R.
DOI: 10.1021/ac025826t
发表时间: 2002-11-01
影响因子: 7.4
作者:
MacCoss, MJ;Wu, CC;Yates, JR
通讯作者: Yates, JR
DOI: 10.1093/bioinformatics/bth092
发表时间: 2004-06-12
期刊: BIOINFORMATICS
影响因子: 5.8
作者:
Craig, R;Beavis, RC
通讯作者: Beavis, RC