Finding functional disease-associated non-coding variation using next-generation sequencing
Finding functional disease-associated non-coding variation using next-generation sequencing
复制标题
使用下一代测序寻找功能性疾病相关的非编码变异
DOI:
10.1101/060285
复制
发表时间:
2016
期刊:
影响因子:
--
通讯作者:
Devanna P
中科院分区:
文献类型:
--
作者:
Devanna P
Next generation sequencing has opened the way for the large scale interrogation of cohorts at the whole exome, or whole genome level. Currently, the field largely focuses on potential disease causing variants that fall within coding sequences and that are predicted to cause protein sequence changes, generally discarding non-coding variants. However non-coding DNA makes up~98% of the genome and contains a range of sequences essential for controlling the expression of protein coding genes. Thus, potentially causative non-coding variation is currently being overlooked. To address this, we have designed an approach to assess variation in one class of non-coding regulatory DNA; the 3′UTRome. Variants in the 3'UTR region of genes are of particular interest because 3'UTRs are responsible for modulating protein expression levels via their interactions with microRNAs. Furthermore they are amenable to large scale analysis as 3′UTR-microRNA interactions are based on complementary base pairing and as such can be predictedin silicoat the genome-wide level. We report a strategy for identifying and functionally testing variants in microRNA binding sites within the 3'UTRome and demonstrate the efficacy of this pipeline in a cohort of language impaired children. Using whole exome sequence data from 43 probands, we extracted variants that lay within 3'UTR microRNA binding sites. We identified a common variant (SNP) in a microRNA binding site and found this SNP to be associated with an endophenotype of language impairment (non-word repetition). We showed that this variant disrupted microRNA regulation in cells and was linked to altered gene expression in the brain, suggesting it may represent a risk factor contributing to SLI. This work demonstrates that biologically relevant variants are currently being under-investigated despite the wealth of next-generation sequencing data available and presents a simple strategy for interrogating non-coding regions of the genome. We propose that this strategy should be routinely applied to whole exome and whole genome sequence data in order to broaden our understanding of how non-coding genetic variation underlies complex phenotypes such as neurodevelopmental disorders.
登录
查看更多内容
影响因子:
64.8
作者:
通讯作者:
--
DOI:
--
发表时间:
2009
期刊:
影响因子:
--
作者:
D. Newbury;L. Winchester;L. Addis;S. Paracchini;Lyn;A. Clark;W. Cohen;H. Cowie;K. Dworzynski;A. Everitt;I. Goodyer;Elizabeth R. Hennessy;A. Kindley;L. Miller;J. Nasir;A. O'hare;Duncan J. Shaw;Z. Simkin;E. Simonoff;V. Slonims;J. Watson;J. Ragoussis;S. Fisher;J. Seckl;P. Helms;P. Bolton;A. Pickles;G. Conti;G. Baird;D. Bishop;A. Monaco
通讯作者:
A. Monaco
影响因子:
5.2
作者:
Jones, RW;Ring, S;Golding, J
通讯作者:
Golding, J
DOI:
--
发表时间:
2015
期刊:
影响因子:
--
作者:
M. Coret;Adam Mccrimmon
通讯作者:
Adam Mccrimmon
影响因子:
9.8
作者:
Newbury DF;Winchester L;Addis L;Paracchini S;Buckingham LL;Clark A;Cohen W;Cowie H;Dworzynski K;Everitt A;Goodyer IM;Hennessy E;Kindley AD;Miller LL;Nasir J;O'Hare A;Shaw D;Simkin Z;Simonoff E;Slonims V;Watson J;Ragoussis J;Fisher SE;Seckl JR;Helms PJ;Bolton PF;Pickles A;Conti-Ramsden G;Baird G;Bishop DV;Monaco AP
通讯作者:
Monaco AP