A method for assessing the statistical significance of mass spectrometry-based protein identifications using general scoring schemes

A method for assessing the statistical significance of mass spectrometry-based protein identifications using general scoring schemes
复制标题

DOI:
10.1021/ac0258709
复制
发表时间:
2003-02-15
影响因子:
7.4
通讯作者:
Beavis, RC
Beavis, RC
中科院分区:
化学1区
文献类型:
--
作者:
Fenyö, D;Beavis, RC

文献摘要

被引文献

相似文献

本文研究了使用生存函数和期望值来评估蛋白质鉴定实验的结果。这些函数是标准的统计测量,可用于将各种蛋白质鉴定评分方案简化为通用的、易于解释的表示。使用这种方法,以及改变主要识别参数的影响,评分系统的相对优点进行了探讨。我们提倡广泛使用这些简单的统计方法来简化和标准化蛋白质鉴定结果的置信度报告,允许不同鉴定算法的用户以直接和统计学显著的方式比较他们的结果。描述了一种使用被大多数蛋白质鉴定搜索引擎丢弃的信息来测量这些分布的方法,从而产生对评分算法、序列数据库和质谱的任何组合都特异的准确存活函数。
This paper investigates the use of survival functions and expectation values to evaluate the results of protein identification experiments. These functions are standard statistical measures that can be used to reduce various protein identification scoring schemes to a common, easily interpretably representation. The relative merits of scoring systems were explored using this approach, as well as the effects of altering primary identification parameters. We would advocate the widespread use of these simple statistical measures to simplify and standardize the reporting of the confidence of protein identification results, allowing the users of different identification algorithms to compare their results in a straightforward and statistically significant manner. A method is described for measuring these distributions using information that is being discarded by most protein identification search engines, resulting in accurate survival functions that are specific to any combination of scoring algorithms, sequence databases, and mass spectra.