Interpretable dimensionality reduction of single cell transcriptome data with deep generative models.

Interpretable dimensionality reduction of single cell transcriptome data with deep generative models.
复制标题

DOI:
10.1038/s41467-018-04368-5
复制
发表时间:
2018-05-21
影响因子:
16.6
通讯作者:
Shah SP
Shah SP
中科院分区:
综合性期刊1区
文献类型:
--
作者:
Ding J;Condon A;Shah SP

文献摘要

参考文献

被引文献

相似文献

单细胞RNA测序在发现细胞类型、识别细胞状态、追踪发育谱系和重建细胞空间组织方面具有巨大潜力。然而,降维以解释单细胞测序数据中的结构仍然是一个挑战。现有的算法要么不能发现数据中的聚类结构,要么丢失全局信息,例如彼此接近的聚类组。我们提出了一个强大的统计模型,scvis,捕捉和可视化的低维结构的单细胞基因表达数据。仿真结果表明,scvis学习的低维表示保留了数据中的局部和全局邻居结构。此外,scvis对数据点的数量具有鲁棒性,并学习一个概率参数映射函数来向现有的嵌入中添加新的数据点。然后,我们使用scvis来分析四个单细胞RNA测序数据集,举例说明高维单细胞RNA测序数据的可解释的二维表示。虽然单细胞转录组数据越来越多,但它们的解释仍然是一个挑战。在这里,作者提出了一种降维方法,保留了数据中的局部和全局邻域结构,从而提高了其可解释性。
Single-cell RNA-sequencing has great potential to discover cell types, identify cell states, trace development lineages, and reconstruct the spatial organization of cells. However, dimension reduction to interpret structure in single-cell sequencing data remains a challenge. Existing algorithms are either not able to uncover the clustering structures in the data or lose global information such as groups of clusters that are close to each other. We present a robust statistical model, scvis, to capture and visualize the low-dimensional structures in single-cell gene expression data. Simulation results demonstrate that low-dimensional representations learned by scvis preserve both the local and global neighbor structures in the data. In addition, scvis is robust to the number of data points and learns a probabilistic parametric mapping function to add new data points to an existing embedding. We then use scvis to analyze four single-cell RNA-sequencing datasets, exemplifying interpretable two-dimensional representations of the high-dimensional single-cell RNA-sequencing data. Although single-cell transcriptome data are increasingly available, their interpretation remains a challenge. Here, the authors present a dimensionality reduction approach that preserves both the local and global neighbourhood structures in the data thus enhancing its interpretability.
DOI: 10.1186/s12859-016-1176-5
发表时间: 2016-08-23
期刊: BMC bioinformatics
影响因子: 3
作者:
DeTomaso D;Yosef N
通讯作者: Yosef N
DOI: 10.1126/science.1247651
发表时间: 2014-02-14
期刊: Science (New York, N.Y.)
影响因子: --
作者:
Jaitin DA;Kenigsberg E;Keren-Shaul H;Elefant N;Paul F;Zaretsky I;Mildner A;Cohen N;Jung S;Tanay A;Amit I
通讯作者: Amit I
DOI: 10.12688/wellcomeopenres.11087.1
发表时间: 2017-03-15
影响因子: --
作者:
Campbell KR;Yau C
通讯作者: Yau C
DOI: 10.1126/science.1198704
发表时间: 2011-05-06
期刊: Science (New York, N.Y.)
影响因子: --
作者:
Bendall SC;Simonds EF;Qiu P;Amir el-AD;Krutzik PO;Finck R;Bruggner RV;Melamed R;Trejo A;Ornatsky OI;Balderas RS;Plevritis SK;Sachs K;Pe'er D;Tanner SD;Nolan GP
通讯作者: Nolan GP
DOI: 10.1214/aoms/1177729694
发表时间: 1951-01-01
影响因子: --
作者:
KULLBACK, S;LEIBLER, RA
通讯作者: LEIBLER, RA