Application of geographic population structure (GPS) algorithm for biogeographical analyses of populations with complex ancestries: a case study of South Asians from 1000 genomes project.
Application of geographic population structure (GPS) algorithm for biogeographical analyses of populations with complex ancestries: a case study of South Asians from 1000 genomes project.
复制标题
DOI:
10.1186/s12863-017-0579-2
复制
发表时间:
2017-12-28
期刊:
影响因子:
2.9
通讯作者:
Upadhyai P
中科院分区:
文献类型:
--
作者:
Das R;Upadhyai P
The utilization of biological data to infer the geographic origins of human populations has been a long standing quest for biologists and anthropologists. Several biogeographical analysis tools have been developed to infer the geographical origins of human populations utilizing genetic data. However due to the inherent complexity of genetic information these approaches are prone to misinterpretations. The Geographic Population Structure (GPS) algorithm is an admixture based tool for biogeographical analyses and has been employed for the geo-localization of various populations worldwide. Here we sought to dissect its sensitivity and accuracy for localizing highly admixed groups. Given the complex history of population dispersal and gene flow in the Indian subcontinent, we have employed the GPS tool to localize five South Asian populations, Punjabi, Gujarati, Tamil, Telugu and Bengali from the 1000 Genomes project, some of whom were recent migrants to USA and UK, using populations from the Indian subcontinent available in Human Genome Diversity Panel (HGDP) and those previously described as reference. Our findings demonstrate reasonably high accuracy with regards to GPS assignment even for recent migrant populations sampled elsewhere, namely the Tamil, Telugu and Gujarati individuals, where 96%, 87% and 79% of the individuals, respectively, were positioned within 600 km of their native locations. While the absence of appropriate reference populations resulted in moderate-to-low levels of precision in positioning of Punjabi and Bengali genomes. Our findings reflect that the GPS approach is useful but likely overtly dependent on the relative proportions of admixture in the reference populations for determination of the biogeographical origins of test individuals. We conclude that further modifications are desired to make this approach more suitable for highly admixed individuals. The online version of this article (doi: 10.1186/s12863-017-0579-2) contains supplementary material, which is available to authorized users.
登录
查看更多内容
影响因子:
4.6
作者:
Flegontov P;Changmai P;Zidkova A;Logacheva MD;Altınışık NE;Flegontova O;Gelfand MS;Gerasimov ES;Khrameeva EE;Konovalova OP;Neretina T;Nikolsky YV;Starostin G;Stepanova VV;Travinsky IV;Tříska M;Tříska P;Tatarinova TV
通讯作者:
Tatarinova TV
影响因子:
3.3
作者:
Das R;Wexler P;Pirooznia M;Elhaik E
通讯作者:
Elhaik E
影响因子:
4
作者:
Chaubey, Gyaneshwer;Metspalu, Mait;Villems, Richard
通讯作者:
Villems, Richard
影响因子:
3.5
作者:
ArunKumar, GaneshPrasad;Tatarinova, Tatiana V.;Pitchappan, Ramasamy
通讯作者:
Pitchappan, Ramasamy
影响因子:
3.7
作者:
Gangal, Kavita;Sarson, Graeme R.;Shukurov, Anvar
通讯作者:
Shukurov, Anvar