Mosaic cis-regulatory evolution drives transcriptional partitioning of HERVH endogenous retrovirus in the human embryo
Abstract
The human endogenous retrovirus type-H (HERVH) family is expressed in the preimplantation embryo. A subset of these elements are specifically transcribed in pluripotent stem cells where they appear to exert regulatory activities promoting self-renewal and pluripotency. How HERVH elements achieve such transcriptional specificity remains poorly understood. To uncover the sequence features underlying HERVH transcriptional activity, we performed a phyloregulatory analysis of the long terminal repeats (LTR7) of the HERVH family, which harbor its promoter, using a wealth of regulatory genomics data. We found that the family includes at least 8 previously unrecognized subfamilies that have been active at different timepoints in primate evolution and display distinct expression patterns during human embryonic development. Notably, nearly all HERVH elements transcribed in ESCs belong to one of the youngest subfamilies we dubbed LTR7up. LTR7 sequence evolution was driven by a mixture of mutational processes, including point mutations, duplications, and multiple recombination events between subfamilies, that led to transcription factor binding motif modules characteristic of each subfamily. Using a reporter assay, we show that one such motif, a predicted SOX2/3 binding site unique to LTR7up, is essential for robust promoter activity in induced pluripotent stem cells. Together these findings illuminate the mechanisms by which HERVH diversified its expression pattern during evolution to colonize distinct cellular niches within the human embryo.
Data availability
Scripts, data tables, and notes for figures 1-4,6a and supplemental figures 1-1,2-1,3-1,4-1,5-1,6-2 by TAC and JDC - https://github.com/LumpLord/Mosaic-cis-regulatory-evolution-drives-transcriptional-partitioning-of-HERVH-endogenous-retrovirus..Scripts and data tables by MS for figures 5,6c and supplemental figures 6-1,6-3,5-2 - https://github.com/Manu-1512/LTR7-up
-
Transcription factor binding dynamics during human ES cell differentiationNCBI Gene Expression Omnibus, GSE61475.
-
3D Chromosome Regulatory Landscape of Human Pluripotent CellsNCBI Gene Expression Omnibus, GSE69647.
-
ChIP-exo of human KRAB-ZNFs transduced in HEK 293T cells and KAP1 in hES H1 cellsNCBI Gene Expression Omnibus, GSE78099.
-
Repeat elements study in pluripotent stem cellsNCBI Gene Expression Omnibus, GSE54726.
-
Principles of Signalling Pathway Modulation for Enhancing Human Naïve Pluripotency Induction [ChIP-seq]NCBI Gene Expression Omnibus, GSE125553.
-
Tracing pluripotency of human early embryos and embryonic stem cells by single cell RNA-seqNCBI Gene Expression Omnibus, GSE36552.
-
Single-Cell RNA-seq Defines the Three Cell Lineages of the Human BlastocystNCBI Gene Expression Omnibus, GSE66507.
Article and author information
Author details
Funding
National Institutes of Health (GM112972)
- Cédric Feschotte
National Institutes of Health (HG009391)
- Cédric Feschotte
National Institutes of Health (GM122550)
- Cédric Feschotte
Cornell Center for Vertebrate Genomics
- Thomas Carter
Howard Hughes Medical Institute
- John L Rinn
National Institutes of Health (GM099117)
- John L Rinn
Cornell Presidential Fellow Program
- Manvendra Singh
The funders had no role in study design, data collection and interpretation, or the decision to submit the work for publication.
Reviewing Editor
- Mia T Levine, University of Pennsylvania, United States
Version history
- Preprint posted: July 8, 2021 (view preprint)
- Received: December 14, 2021
- Accepted: February 17, 2022
- Accepted Manuscript published: February 18, 2022 (version 1)
- Version of Record published: March 10, 2022 (version 2)
Copyright
© 2022, Carter et al.
This article is distributed under the terms of the Creative Commons Attribution License permitting unrestricted use and redistribution provided that the original author and source are credited.
Metrics
-
- 3,443
- views
-
- 410
- downloads
-
- 31
- citations
Views, downloads and citations are aggregated across all versions of this paper published by eLife.
Download links
Downloads (link to download the article as PDF)
Open citations (links to open the citations from this article in various online reference manager services)
Cite this article (links to download the citations from this article in formats compatible with various reference manager tools)
Further reading
-
- Chromosomes and Gene Expression
- Genetics and Genomics
Members of the diverse heterochromatin protein 1 (HP1) family play crucial roles in heterochromatin formation and maintenance. Despite the similar affinities of their chromodomains for di- and tri-methylated histone H3 lysine 9 (H3K9me2/3), different HP1 proteins exhibit distinct chromatin-binding patterns, likely due to interactions with various specificity factors. Previously, we showed that the chromatin-binding pattern of the HP1 protein Rhino, a crucial factor of the Drosophila PIWI-interacting RNA (piRNA) pathway, is largely defined by a DNA sequence-specific C2H2 zinc finger protein named Kipferl (Baumgartner et al., 2022). Here, we elucidate the molecular basis of the interaction between Rhino and its guidance factor Kipferl. Through phylogenetic analyses, structure prediction, and in vivo genetics, we identify a single amino acid change within Rhino’s chromodomain, G31D, that does not affect H3K9me2/3 binding but disrupts the interaction between Rhino and Kipferl. Flies carrying the rhinoG31D mutation phenocopy kipferl mutant flies, with Rhino redistributing from piRNA clusters to satellite repeats, causing pronounced changes in the ovarian piRNA profile of rhinoG31D flies. Thus, Rhino’s chromodomain functions as a dual-specificity module, facilitating interactions with both a histone mark and a DNA-binding protein.
-
- Genetics and Genomics
- Neuroscience
Cognitive decline is a significant health concern in our aging society. Here, we used the model organism C. elegans to investigate the impact of the IIS/FOXO pathway on age-related cognitive decline. The daf-2 Insulin/IGF-1 receptor mutant exhibits a significant extension of learning and memory span with age compared to wild-type worms, an effect that is dependent on the DAF-16 transcription factor. To identify possible mechanisms by which aging daf-2 mutants maintain learning and memory with age while wild-type worms lose neuronal function, we carried out neuron-specific transcriptomic analysis in aged animals. We observed downregulation of neuronal genes and upregulation of transcriptional regulation genes in aging wild-type neurons. By contrast, IIS/FOXO pathway mutants exhibit distinct neuronal transcriptomic alterations in response to cognitive aging, including upregulation of stress response genes and downregulation of specific insulin signaling genes. We tested the roles of significantly transcriptionally-changed genes in regulating cognitive functions, identifying novel regulators of learning and memory. In addition to other mechanistic insights, a comparison of the aged vs young daf-2 neuronal transcriptome revealed that a new set of potentially neuroprotective genes is upregulated; instead of simply mimicking a young state, daf-2 may enhance neuronal resilience to accumulation of harm and take a more active approach to combat aging. These findings suggest a potential mechanism for regulating cognitive function with age and offer insights into novel therapeutic targets for age-related cognitive decline.