Shape analysis of high-throughput transcriptomics experiment data

Shape analysis of high-throughput transcriptomics experiment data
复制标题

DOI:
10.1093/biostatistics/kxv018
复制
发表时间:
2015-10-01
期刊:
影响因子:
2.1
通讯作者:
Bravo, Hector Corrada
Bravo, Hector Corrada
中科院分区:
数学2区
文献类型:
--
作者:
Okrah, Kwame;Bravo, Hector Corrada

文献摘要

被引文献

相似文献

The recent growth of high-throughput transcriptome technology has been paralleled by the development of statistical methodologies to analyze the data they produce. Some of these newly developed methods are based on the assumption that the data observed or a transformation of the data are relatively symmetric with light tails, usually summarized by assuming a Gaussian random component. It is indeed very difficult to assess this assumption for small sample sizes. In this article, we utilize L-moments statistics as the basis of exploratory data analysis, the assessment of distributional assumptions, and the hypothesis testing of high-throughput transcriptomic data. In particular, we use L-moments ratios for assessing the shape (skewness and kurtosis) of high-throughput transcriptome data. Based on these statistics, we propose an algorithm for identifying genes with distributions that are markedly different from the majority in the data. In addition, we also illustrate the utility of this framework to characterize the robustness of distributional assumptions. We apply it to RNA-seq data and find that methods based on the simple t-test for differential expression analysis using L-moments as weights are robust.