Rxivist logo

Human demographic history impacts genetic risk prediction across diverse populations

By Alicia R. Martin, Christopher R Gignoux, Raymond K Walters, Genevieve L. Wojcik, Benjamin M Neale, Simon Gravel, M. Daly, Carlos D. Bustamante, Eimear Kenny

Posted 23 Aug 2016
bioRxiv DOI: 10.1101/070797 (published DOI: 10.1016/j.ajhg.2017.03.004)

The vast majority of genome-wide association studies are performed in Europeans, and their transferability to other populations is dependent on many factors (e.g. linkage disequilibrium, allele frequencies, genetic architecture). As medical genomics studies become increasingly large and diverse, gaining insights into population history and consequently the transferability of disease risk measurement is critical. Here, we disentangle recent population history in the widely-used 1000 Genomes Project reference panel, with an emphasis on populations underrepresented in medical studies. To examine the transferability of single-ancestry GWAS, we used published summary statistics to calculate polygenic risk scores for six well-studied traits and diseases. We identified directional inconsistencies in all scores; for example, height is predicted to decrease with genetic distance from Europeans, despite robust anthropological evidence that West Africans are as tall as Europeans on average. To gain deeper quantitative insights into GWAS transferability, we developed a complex trait coalescent-based simulation framework considering effects of polygenicity, causal allele frequency divergence, and heritability. As expected, correlations between true and inferred risk were typically highest in the population from which summary statistics were derived. We demonstrated that scores inferred from European GWAS were biased by genetic drift in other populations even when choosing the same causal variants, and that biases in any direction were possible and unpredictable. This work cautions that summarizing findings from large-scale GWAS may have limited portability to other populations using standard approaches, and highlights the need for generalized risk prediction methods and the inclusion of more diverse individuals in medical genomics.

Download data

  • Downloaded 3,111 times
  • Download rankings, all-time:
    • Site-wide: 3,891
    • In genomics: 445
  • Year to date:
    • Site-wide: 42,869
  • Since beginning of last month:
    • Site-wide: 31,718

Altmetric data

Downloads over time

Distribution of downloads per paper, site-wide


Sign up for the Rxivist weekly newsletter! (Click here for more details.)