Variable prediction accuracy of polygenic scores within an ancestry group
Abstract
Fields as diverse as human genetics and sociology are increasingly using polygenic scores based on genome-wide association studies (GWAS) for phenotypic prediction. However, recent work has shown that polygenic scores have limited portability across groups of different genetic ancestries, restricting the contexts in which they can be used reliably and potentially creating serious inequities in future clinical applications. Using the UK Biobank data, we demonstrate that even within a single ancestry group (i.e., when there are negligible differences in linkage disequilibrium or in causal alleles frequencies), the prediction accuracy of polygenic scores can depend on characteristics such as the socio-economic status, age or sex of the individuals in which the GWAS and the prediction were conducted, as well as on the GWAS design. Our findings highlight both the complexities of interpreting polygenic scores and underappreciated obstacles to their broad use.
Data availability
The GWAS summary statistics generated in this study have been uploaded to Dryad.
-
Variable prediction accuracy of polygenic scores within an ancestry groupDryad Digital Repository, 10.5061/dryad.66t1g1jxs.
Article and author information
Author details
Funding
National Institute of General Medical Sciences (GM121372)
- Molly Przeworski
National Human Genome Research Institute (HG008140)
- Jonathan K Pritchard
Robert Wood Johnson Foundation (84337817)
- Dalton Conley
Simons Foundation (633313)
- Arbel Harpak
The funders had no role in study design, data collection and interpretation, or the decision to submit the work for publication.
Ethics
Human subjects: This study has been conducted using the UK Biobank resource under application Number 11138, as approved by Columbia University Institutional Review Board, protocol AAAS2914.
Copyright
© 2020, Mostafavi et al.
This article is distributed under the terms of the Creative Commons Attribution License permitting unrestricted use and redistribution provided that the original author and source are credited.
Metrics
-
- 11,331
- views
-
- 1,592
- downloads
-
- 279
- citations
Views, downloads and citations are aggregated across all versions of this paper published by eLife.
Download links
Downloads (link to download the article as PDF)
Open citations (links to open the citations from this article in various online reference manager services)
Cite this article (links to download the citations from this article in formats compatible with various reference manager tools)
Further reading
-
- Genetics and Genomics
- Microbiology and Infectious Disease
Klebsiella pneumoniae is a global public health concern due to the rising myriad of hypervirulent and multidrug-resistant clones both alarmingly associated with high mortality. The molecular mechanisms underpinning these recalcitrant K. pneumoniae infection, and how virulence is coupled with the emergence of lineages resistant to nearly all present-day clinically important antimicrobials, are unclear. In this study, we performed a genome-wide screen in K. pneumoniae ECL8, a member of the endemic K2-ST375 pathotype most often reported in Asia, to define genes essential for growth in a nutrient-rich laboratory medium (Luria-Bertani [LB] medium), human urine, and serum. Through transposon directed insertion-site sequencing (TraDIS), a total of 427 genes were identified as essential for growth on LB agar, whereas transposon insertions in 11 and 144 genes decreased fitness for growth in either urine or serum, respectively. These studies not only provide further knowledge on the genetics of this pathogen but also provide a strong impetus for discovering new antimicrobial targets to improve current therapeutic options for K. pneumoniae infections.
-
- Developmental Biology
- Genetics and Genomics
New developmental programs can evolve through adaptive changes to gene expression. The annelid Streblospio benedicti has a developmental dimorphism, which provides a unique intraspecific framework for understanding the earliest genetic changes that take place during developmental divergence. Using comparative RNAseq through ontogeny, we find that only a small proportion of genes are differentially expressed at any time, despite major differences in larval development and life history. These genes shift expression profiles across morphs by either turning off any expression in one morph or changing the timing or amount of gene expression. We directly connect the contributions of these mechanisms to differences in developmental processes. We examine F1 offspring – using reciprocal crosses – to determine maternal mRNA inheritance and the regulatory architecture of gene expression. These results highlight the importance of both novel gene expression and heterochronic shifts in developmental evolution, as well as the trans-acting regulatory factors in initiating divergence.