Local adaptation and archaic introgression shape global diversity at human structural variant loci

Abstract

Large genomic insertions and deletions are a potent source of functional variation, but are challenging to resolve with short-read sequencing, limiting knowledge of the role of such structural variants (SVs) in human evolution. Here, we used a graph-based method to genotype long-read-discovered SVs in short-read data from diverse human genomes. We then applied an admixture-aware method to identify 220 SVs exhibiting extreme patterns of frequency differentiation—a signature of local adaptation. The top two variants traced to the immunoglobulin heavy chain locus, tagging a haplotype that swept to near fixation in certain Southeast Asian populations, but is rare in other global populations. Further investigation revealed evidence that the haplotype traces to gene flow from Neanderthals, corroborating the role of immune-related genes as prominent targets of adaptive introgression. Our study demonstrates how recent technical advances can help resolve signatures of key evolutionary events that remained obscured within technically challenging regions of the genome.

Data availability

All code necessary for reproducing our analysis is available on GitHub (https://github.com/mccoy-lab/sv_selection). SV genotypes, eQTL results, and selection scan results are available on Zenodo (doi: 10.5281/zenodo.4469976).

The following data sets were generated
The following previously published data sets were used

Article and author information

Author details

  1. Stephanie M Yan

    Department of Biology, Johns Hopkins University, Baltimore, United States
    Competing interests
    The authors declare that no competing interests exist.
  2. Rachel M Sherman

    Department of Biology, Johns Hopkins University, Baltimore, United States
    Competing interests
    The authors declare that no competing interests exist.
  3. Dylan J Taylor

    Department of Biology, Johns Hopkins University, Baltimore, United States
    Competing interests
    The authors declare that no competing interests exist.
    ORCID icon "This ORCID iD identifies the author of this article:" 0000-0001-5806-4494
  4. Divya R Nair

    Department of Biology, Johns Hopkins University, Baltimore, United States
    Competing interests
    The authors declare that no competing interests exist.
  5. Andrew N Bortvin

    Department of Biology, Johns Hopkins University, Baltimore, United States
    Competing interests
    The authors declare that no competing interests exist.
  6. Michael C Schatz

    Department of Computer Science, Johns Hopkins University, Baltimore, United States
    Competing interests
    The authors declare that no competing interests exist.
  7. Rajiv C McCoy

    Department of Biology, Johns Hopkins University, Baltimore, United States
    For correspondence
    rajiv.mccoy@jhu.edu
    Competing interests
    The authors declare that no competing interests exist.
    ORCID icon "This ORCID iD identifies the author of this article:" 0000-0003-0615-146X

Funding

National Institutes of Health (R35GM133747)

  • Rajiv C McCoy

National Science Foundation (DBI-1350041)

  • Michael C Schatz

The funders had no role in study design, data collection and interpretation, or the decision to submit the work for publication.

Copyright

© 2021, Yan et al.

This article is distributed under the terms of the Creative Commons Attribution License permitting unrestricted use and redistribution provided that the original author and source are credited.

Metrics

  • 3,880
    views
  • 423
    downloads
  • 38
    citations

Views, downloads and citations are aggregated across all versions of this paper published by eLife.

Download links

A two-part list of links to download the article, or parts of the article, in various formats.

Downloads (link to download the article as PDF)

Open citations (links to open the citations from this article in various online reference manager services)

Cite this article (links to download the citations from this article in formats compatible with various reference manager tools)

  1. Stephanie M Yan
  2. Rachel M Sherman
  3. Dylan J Taylor
  4. Divya R Nair
  5. Andrew N Bortvin
  6. Michael C Schatz
  7. Rajiv C McCoy
(2021)
Local adaptation and archaic introgression shape global diversity at human structural variant loci
eLife 10:e67615.
https://doi.org/10.7554/eLife.67615

Share this article

https://doi.org/10.7554/eLife.67615

Further reading

    1. Evolutionary Biology
    2. Microbiology and Infectious Disease
    Zach Hensel
    Short Report

    Accurate estimation of the effects of mutations on SARS-CoV-2 viral fitness can inform public-health responses such as vaccine development and predicting the impact of a new variant; it can also illuminate biological mechanisms including those underlying the emergence of variants of concern. Recently, Lan et al. reported a model of SARS-CoV-2 secondary structure and its underlying dimethyl sulfate reactivity data (Lan et al., 2022). I investigated whether base reactivities and secondary structure models derived from them can explain some variability in the frequency of observing different nucleotide substitutions across millions of patient sequences in the SARS-CoV-2 phylogenetic tree. Nucleotide basepairing was compared to the estimated ‘mutational fitness’ of substitutions, a measurement of the difference between a substitution’s observed and expected frequency that is correlated with other estimates of viral fitness (Bloom and Neher, 2023). This comparison revealed that secondary structure is often predictive of substitution frequency, with significant decreases in substitution frequencies at basepaired positions. Focusing on the mutational fitness of C→U, the most common type of substitution, I describe C→U substitutions at basepaired positions that characterize major SARS-CoV-2 variants; such mutations may have a greater impact on fitness than appreciated when considering substitution frequency alone.

    1. Evolutionary Biology
    Yiheng Zhang, Xing Wang ... Xiaoguang Yang
    Research Article

    Although fossil evidence suggests the existence of an early muscular system in the ancient cnidarian jellyfish from the early Cambrian Kuanchuanpu biota (ca. 535 Ma), south China, the mechanisms underlying the feeding and respiration of the early jellyfish are conjectural. Recently, the polyp inside the periderm of olivooids was demonstrated to be a calyx-like structure, most likely bearing short tentacles and bundles of coronal muscles at the edge of the calyx, thus presumably contributing to feeding and respiration. Here, we simulate the contraction and expansion of the microscopic periderm-bearing olivooid Quadrapyrgites via the fluid-structure interaction computational fluid dynamics (CFD) method to investigate their feeding and respiratory activities. The simulations show that the rate of water inhalation by the polyp subumbrella is positively correlated with the rate of contraction and expansion of the coronal muscles, consistent with the previous feeding and respiration hypothesis. The dynamic simulations also show that the frequent inhalation/exhalation of water through the periderm polyp expansion/contraction conducted by the muscular system of Quadrapyrgites most likely represents the ancestral feeding and respiration patterns of Cambrian sedentary medusozoans that predated the rhythmic jet-propelled swimming of the modern jellyfish. Most importantly for these Cambrian microscopic sedentary medusozoans, the increase of body size and stronger capacity of muscle contraction may have been indispensable in the stepwise evolution of active feeding and subsequent swimming in a higher flow (or higher Reynolds number) environment.