Epigenetics: Making the most of methylation
DNA methylation is a key mechanism used by higher eukaryotes to regulate gene expression. The addition of a methyl group to carbon atom number 5 within cytosine bases in DNA is known to repress the transcription of genes into messenger RNA molecules, thus reducing the production of the proteins coded by these genes. Most methylation occurs at CpG dinucleotides—cytosines that are paired with guanines—and these often cluster together to form CpG islands in the promoter regions of genes. In the late 1990s, it was discovered that transcription was repressed when methyl CpG binding proteins were recruited to methylated CpG islands (Hendrich and Bird, 1998).
Subsequent studies have confirmed that the binding of these proteins throughout the genome is proportional to the density of DNA methylation (Baubec et al., 2013), and have identified additional proteins with a high affinity for methylated CpG sites (reviewed in Defossez and Stancheva, 2011). Moreover, in recent years, other screening approaches based mainly on mass spectrometry have revealed that more proteins bind to methylated DNA than previously thought (Mittler et al., 2009; Bartke et al., 2010; Bartels et al., 2011; Spruijt et al., 2013). Now, in eLife, Heng Zhu and co-workers at the Johns Hopkins University School of Medicine—including Shaohui Hu as first author—use a high-throughput screening method to show that many human transcription factors also interact with genomic DNA sequences containing methylated CpG sites (Hu et al., 2013).
To this end, the Johns Hopkins researchers made use of a published protein microarray consisting of 1,321 transcription factors and 210 co-factors (Hu et al., 2009). Hu et al. incubated the array with 154 distinct human promoter sequences, each of which contained at least one methylated CpG dinucleotide.Their results revealed that 150 (97%) of the 154 methylated human promoter sequences showed specific binding to at least one protein on the microarray. Moreover, of the 1531 proteins, 47 (3%) showed binding to methylated cytosines within the promoters. Most of the proteins bound to methylated DNA in a sequence-dependent manner; however, a minority bound to many different methylated DNA probes, indicating that binding can sometimes occur independent of DNA sequence (Figure 1).

Some human transcription factors can bind to both methylated and non-methylated DNA sequences. Hu et al. examined the ability of 17 human transcription factors to bind to 150 different DNA motifs containing methylated or non-methylated CpG islands. Each row represents one transcription factor. For each motif, some transcription factors bound only to the methylated version (red), some to only the non-methylated version (blue); some to both methylated and non-methylated versions (green), and some to neither (grey). From Figure 2a in Hu et al., 2013.
A number of transcription factors, including KLF4—a recently identified methyl-CpG binding protein (Spruijt et al., 2013)—interacted with methylated sequences that did not resemble their known consensus DNA binding motifs. Using a technique based on electrophoresis, Hu et al. showed that KLF4 binds methylated and non-methylated DNA in a non-competitive manner: this suggests that different domains of the protein may be responsible for each type of binding, which they duly confirmed using molecular modeling and mutagenesis studies.
The Johns Hopkins researchers then mined published ChIP-sequencing data from stem cells to identify the target DNA sequences of KLF4, and compared these with data on genome-wide DNA methylation. Strikingly, KLF4 binding appears to be bimodal in nature throughout the genome, with 38% of KLF4 binding sites showing less than 20% methylation, and 48% showing methylation levels over 80%. Finally, Hu et al. used ChIP-bisulfite sequencing, which makes it possible to determine the methylation status of each cytosine within a target DNA sequence, to confirm that KLF4 also binds to both methylated and non-methylated DNA in vivo.
Hu et al. only profiled a small fraction of the complete human methylome for interactions with transcription factors; further proteins capable of binding genomic methyl CpG sequences surely await identification. The same holds true for interactions with methylated non-CpG sequences such as methyl-CpA (cytosine adjacent to adenine), which are fairly abundant in embryonic stem cells (Ramsahoye et al., 2000; Lister et al., 2009). To determine the physiological relevance of these interactions, it will be important to deduce the affinity with which proteins bind these sequences compared to their known targets; initial experiments along these lines are presented in the current eLife paper. Furthermore, recent evidence suggests that non-methylated CpG islands recruit activator proteins, many of which contain a CXXC motif (reviewed in Long et al., 2013). The transcription factor microarray approach used by the Johns Hopkins team, combined with quantitative mass spectrometry-based technology (Spruijt et al., 2013), could thus be used to identify the complete cellular complement of proteins that bind specifically to non-methylated CpG islands.
Finally, this study and other recently published papers force us to reconsider the mechanism(s) via which CpG methylation regulates transcription. Although DNA methylation is generally considered to be a repressive epigenetic modification, experiments presented by Hu et al. suggest that in some cases, methylation of a given promoter sequence can result in activation of transcription. Moreover, other work has revealed a temporal uncoupling of DNA methylation and transcriptional repression during Xenopus embryogenesis (Bogdanovic et al., 2011). Further experiments are therefore required to determine whether the functional readout of CpG methylation is affected by the repertoire and abundance of different DNA methylation ‘readers’ acting at any given time in a cell or a developing organism.
References
-
Biological functions of methyl-CpG-binding proteinsProg Mol Biol Transi Sci 101:377–398.https://doi.org/10.1016/B978-0-12-387685-0.00012-3
-
Identification and characterization of a family of mammalian methyl-CpG binding proteinsMol Cell Biol 18:6538–6547.
-
ZF-CxxC domain-containing proteins, CpG islands and the chromatin connectionBiochem Soc trans 41:727–740.https://doi.org/10.1042/BST20130028
Article and author information
Author details
Publication history
Copyright
© 2013, Vermeulen
This article is distributed under the terms of the Creative Commons Attribution License, which permits unrestricted use and redistribution provided that the original author and source are credited.
Metrics
-
- 525
- views
-
- 57
- downloads
-
- 2
- citations
Views, downloads and citations are aggregated across all versions of this paper published by eLife.
Download links
Downloads (link to download the article as PDF)
Open citations (links to open the citations from this article in various online reference manager services)
Cite this article (links to download the citations from this article in formats compatible with various reference manager tools)
Further reading
-
- Biochemistry and Chemical Biology
In healthy cells, cyclin D1 is expressed during the G1 phase of the cell cycle, where it activates CDK4 and CDK6. Its dysregulation is a well-established oncogenic driver in numerous human cancers. The cancer-related function of cyclin D1 has been primarily studied by focusing on the phosphorylation of the retinoblastoma (RB) gene product. Here, using an integrative approach combining bioinformatic analyses and biochemical experiments, we show that GTSE1 (G-Two and S phases expressed protein 1), a protein positively regulating cell cycle progression, is a previously unrecognized substrate of cyclin D1–CDK4/6 in tumor cells overexpressing cyclin D1 during G1 and subsequent phases. The phosphorylation of GTSE1 mediated by cyclin D1–CDK4/6 inhibits GTSE1 degradation, leading to high levels of GTSE1 across all cell cycle phases. Functionally, the phosphorylation of GTSE1 promotes cellular proliferation and is associated with poor prognosis within a pan-cancer cohort. Our findings provide insights into cyclin D1’s role in cell cycle control and oncogenesis beyond RB phosphorylation.
-
- Biochemistry and Chemical Biology
- Microbiology and Infectious Disease
Teichoic acids (TA) are linear phospho-saccharidic polymers and important constituents of the cell envelope of Gram-positive bacteria, either bound to the peptidoglycan as wall teichoic acids (WTA) or to the membrane as lipoteichoic acids (LTA). The composition of TA varies greatly but the presence of both WTA and LTA is highly conserved, hinting at an underlying fundamental function that is distinct from their specific roles in diverse organisms. We report the observation of a periplasmic space in Streptococcus pneumoniae by cryo-electron microscopy of vitreous sections. The thickness and appearance of this region change upon deletion of genes involved in the attachment of TA, supporting their role in the maintenance of a periplasmic space in Gram-positive bacteria as a possible universal function. Consequences of these mutations were further examined by super-resolved microscopy, following metabolic labeling and fluorophore coupling by click chemistry. This novel labeling method also enabled in-gel analysis of cell fractions. With this approach, we were able to titrate the actual amount of TA per cell and to determine the ratio of WTA to LTA. In addition, we followed the change of TA length during growth phases, and discovered that a mutant devoid of LTA accumulates the membrane-bound polymerized TA precursor.