Prediction of human disease genes by human-mouse conserved coexpression analysis.
Prediction of human disease genes by human-mouse conserved coexpression analysis.
复制标题
DOI:
10.1371/journal.pcbi.1000043
复制
发表时间:
2008-03-28
影响因子:
4.3
通讯作者:
Di Cunto F
中科院分区:
文献类型:
--
作者:
Ala U;Piro RM;Grassi E;Damasco C;Silengo L;Oti M;Provero P;Di Cunto F
Even in the post-genomic era, the identification of candidate genes within loci associated with human genetic diseases is a very demanding task, because the critical region may typically contain hundreds of positional candidates. Since genes implicated in similar phenotypes tend to share very similar expression profiles, high throughput gene expression data may represent a very important resource to identify the best candidates for sequencing. However, so far, gene coexpression has not been used very successfully to prioritize positional candidates. We show that it is possible to reliably identify disease-relevant relationships among genes from massive microarray datasets by concentrating only on genes sharing similar expression profiles in both human and mouse. Moreover, we show systematically that the integration of human-mouse conserved coexpression with a phenotype similarity map allows the efficient identification of disease genes in large genomic regions. Finally, using this approach on 850 OMIM loci characterized by an unknown molecular basis, we propose high-probability candidates for 81 genetic diseases. Our results demonstrate that conserved coexpression, even at the human-mouse phylogenetic distance, represents a very strong criterion to predict disease-relevant relationships among human genes. One of the most limiting aspects of biological research in the post-genomic era is the capability to integrate massive datasets on gene structure and function for producing useful biological knowledge. In this report we have applied an integrative approach to address the problem of identifying likely candidate genes within loci associated with human genetic diseases. Despite the recent progress in sequencing technologies, approaching this problem from an experimental perspective still represents a very demanding task, because the critical region may typically contain hundreds of positional candidates. We found that by concentrating only on genes sharing similar expression profiles in both human and mouse, massive microarray datasets can be used to reliably identify disease-relevant relationships among genes. Moreover, we found that integrating the coexpression criterion with systematic phenome analysis allows efficient identification of disease genes in large genomic regions. Using this approach on 850 OMIM loci characterized by unknown molecular basis, we propose high-probability candidates for 81 genetic diseases.
登录
查看更多内容
影响因子:
14.9
作者:
López-Bigas, N;Ouzounis, CA
通讯作者:
Ouzounis, CA
影响因子:
4
作者:
Jonsson, JJ;Renieri, A;Pober, BR
通讯作者:
Pober, BR
影响因子:
10.7
作者:
Jordan, IK;Mariño-Ramírez, L;Koonin, EV
通讯作者:
Koonin, EV
影响因子:
4.4
作者:
Fukuoka Y;Inaoka H;Kohane IS
通讯作者:
Kohane IS
影响因子:
46.9
作者:
Lage, Kasper;Karlberg, E. Olof;Brunak, Soren
通讯作者:
Brunak, Soren