1. Michiel Vermeulen  Is a corresponding author
  1. University Medical Center Utrecht, Netherlands

DNA methylation is a key mechanism used by higher eukaryotes to regulate gene expression. The addition of a methyl group to carbon atom number 5 within cytosine bases in DNA is known to repress the transcription of genes into messenger RNA molecules, thus reducing the production of the proteins coded by these genes. Most methylation occurs at CpG dinucleotides—cytosines that are paired with guanines—and these often cluster together to form CpG islands in the promoter regions of genes. In the late 1990s, it was discovered that transcription was repressed when methyl CpG binding proteins were recruited to methylated CpG islands (Hendrich and Bird, 1998).

Subsequent studies have confirmed that the binding of these proteins throughout the genome is proportional to the density of DNA methylation (Baubec et al., 2013), and have identified additional proteins with a high affinity for methylated CpG sites (reviewed in Defossez and Stancheva, 2011). Moreover, in recent years, other screening approaches based mainly on mass spectrometry have revealed that more proteins bind to methylated DNA than previously thought (Mittler et al., 2009; Bartke et al., 2010; Bartels et al., 2011; Spruijt et al., 2013). Now, in eLife, Heng Zhu and co-workers at the Johns Hopkins University School of Medicine—including Shaohui Hu as first author—use a high-throughput screening method to show that many human transcription factors also interact with genomic DNA sequences containing methylated CpG sites (Hu et al., 2013).

To this end, the Johns Hopkins researchers made use of a published protein microarray consisting of 1,321 transcription factors and 210 co-factors (Hu et al., 2009). Hu et al. incubated the array with 154 distinct human promoter sequences, each of which contained at least one methylated CpG dinucleotide.Their results revealed that 150 (97%) of the 154 methylated human promoter sequences showed specific binding to at least one protein on the microarray. Moreover, of the 1531 proteins, 47 (3%) showed binding to methylated cytosines within the promoters. Most of the proteins bound to methylated DNA in a sequence-dependent manner; however, a minority bound to many different methylated DNA probes, indicating that binding can sometimes occur independent of DNA sequence (Figure 1).

Some human transcription factors can bind to both methylated and non-methylated DNA sequences. Hu et al. examined the ability of 17 human transcription factors to bind to 150 different DNA motifs containing methylated or non-methylated CpG islands. Each row represents one transcription factor. For each motif, some transcription factors bound only to the methylated version (red), some to only the non-methylated version (blue); some to both methylated and non-methylated versions (green), and some to neither (grey). From Figure 2a in Hu et al., 2013.

A number of transcription factors, including KLF4—a recently identified methyl-CpG binding protein (Spruijt et al., 2013)—interacted with methylated sequences that did not resemble their known consensus DNA binding motifs. Using a technique based on electrophoresis, Hu et al. showed that KLF4 binds methylated and non-methylated DNA in a non-competitive manner: this suggests that different domains of the protein may be responsible for each type of binding, which they duly confirmed using molecular modeling and mutagenesis studies.

The Johns Hopkins researchers then mined published ChIP-sequencing data from stem cells to identify the target DNA sequences of KLF4, and compared these with data on genome-wide DNA methylation. Strikingly, KLF4 binding appears to be bimodal in nature throughout the genome, with 38% of KLF4 binding sites showing less than 20% methylation, and 48% showing methylation levels over 80%. Finally, Hu et al. used ChIP-bisulfite sequencing, which makes it possible to determine the methylation status of each cytosine within a target DNA sequence, to confirm that KLF4 also binds to both methylated and non-methylated DNA in vivo.

Hu et al. only profiled a small fraction of the complete human methylome for interactions with transcription factors; further proteins capable of binding genomic methyl CpG sequences surely await identification. The same holds true for interactions with methylated non-CpG sequences such as methyl-CpA (cytosine adjacent to adenine), which are fairly abundant in embryonic stem cells (Ramsahoye et al., 2000; Lister et al., 2009). To determine the physiological relevance of these interactions, it will be important to deduce the affinity with which proteins bind these sequences compared to their known targets; initial experiments along these lines are presented in the current eLife paper. Furthermore, recent evidence suggests that non-methylated CpG islands recruit activator proteins, many of which contain a CXXC motif (reviewed in Long et al., 2013). The transcription factor microarray approach used by the Johns Hopkins team, combined with quantitative mass spectrometry-based technology (Spruijt et al., 2013), could thus be used to identify the complete cellular complement of proteins that bind specifically to non-methylated CpG islands.

Finally, this study and other recently published papers force us to reconsider the mechanism(s) via which CpG methylation regulates transcription. Although DNA methylation is generally considered to be a repressive epigenetic modification, experiments presented by Hu et al. suggest that in some cases, methylation of a given promoter sequence can result in activation of transcription. Moreover, other work has revealed a temporal uncoupling of DNA methylation and transcriptional repression during Xenopus embryogenesis (Bogdanovic et al., 2011). Further experiments are therefore required to determine whether the functional readout of CpG methylation is affected by the repertoire and abundance of different DNA methylation ‘readers’ acting at any given time in a cell or a developing organism.

References

    1. Hendrich B
    2. Bird A
    (1998)
    Identification and characterization of a family of mammalian methyl-CpG binding proteins
    Mol Cell Biol 18:6538–6547.

Article and author information

Author details

  1. Michiel Vermeulen

    Department of Molecular Cancer Research, University Medical Center Utrecht, Utrecht, Netherlands
    For correspondence
    m.vermeulen-3@umcutrecht.nl
    Competing interests
    The author declares that no competing interests exist.

Publication history

  1. Version of Record published:

Copyright

© 2013, Vermeulen

This article is distributed under the terms of the Creative Commons Attribution License, which permits unrestricted use and redistribution provided that the original author and source are credited.

Metrics

  • 522
    views
  • 56
    downloads
  • 2
    citations

Views, downloads and citations are aggregated across all versions of this paper published by eLife.

Download links

A two-part list of links to download the article, or parts of the article, in various formats.

Downloads (link to download the article as PDF)

Open citations (links to open the citations from this article in various online reference manager services)

Cite this article (links to download the citations from this article in formats compatible with various reference manager tools)

  1. Michiel Vermeulen
(2013)
Epigenetics: Making the most of methylation
eLife 2:e01387.
https://doi.org/10.7554/eLife.01387

Further reading

    1. Biochemistry and Chemical Biology
    2. Cell Biology
    Gina Partipilo, Yang Gao ... Benjamin K Keitz
    Feature Article

    Troubleshooting is an important part of experimental research, but graduate students rarely receive formal training in this skill. In this article, we describe an initiative called Pipettes and Problem Solving that we developed to teach troubleshooting skills to graduate students at the University of Texas at Austin. An experienced researcher presents details of a hypothetical experiment that has produced unexpected results, and students have to propose new experiments that will help identify the source of the problem. We also provide slides and other resources that can be used to facilitate problem solving and teach troubleshooting skills at other institutions.

    1. Biochemistry and Chemical Biology
    2. Plant Biology
    Hao Wang, Biying Zhu ... Zhaoliang Zhang
    Research Article

    Ethylamine (EA), the precursor of theanine biosynthesis, is synthesized from alanine decarboxylation by alanine decarboxylase (AlaDC) in tea plants. AlaDC evolves from serine decarboxylase (SerDC) through neofunctionalization and has lower catalytic activity. However, lacking structure information hinders the understanding of the evolution of substrate specificity and catalytic activity. In this study, we solved the X-ray crystal structures of AlaDC from Camellia sinensis (CsAlaDC) and SerDC from Arabidopsis thaliana (AtSerDC). Tyr341 of AtSerDC or the corresponding Tyr336 of CsAlaDC is essential for their enzymatic activity. Tyr111 of AtSerDC and the corresponding Phe106 of CsAlaDC determine their substrate specificity. Both CsAlaDC and AtSerDC have a distinctive zinc finger and have not been identified in any other Group II PLP-dependent amino acid decarboxylases. Based on the structural comparisons, we conducted a mutation screen of CsAlaDC. The results indicated that the mutation of L110F or P114A in the CsAlaDC dimerization interface significantly improved the catalytic activity by 110% and 59%, respectively. Combining a double mutant of CsAlaDCL110F/P114A with theanine synthetase increased theanine production 672% in an in vitro system. This study provides the structural basis for the substrate selectivity and catalytic activity of CsAlaDC and AtSerDC and provides a route to more efficient biosynthesis of theanine.