Comparative Genomics: One for all

Feb 11, 2016

Open access
Copyright information

Download
Cite
CommentOpen annotations (there are currently 0 annotations on this page).
Share

Duncan T Odom

University of Cambridge, United Kingdom

There is an old saying in computational circles that researchers in bioinformatics would rather use someone else’s toothbrush than use someone else’s code. One example of this adage being true can be seen in previous attempts to compare the rates at which differences in the mechanisms that control DNA accumulate in different species and lineages.

The information contained in DNA is first accessed by dedicated proteins called transcription factors (TF) that bind to preferred sequence of bases in the DNA. This sequence is typically short, between 8 and 20 bases in length (Vaquerizas et al., 2009), although some can be as long as 35 bases (Filippova et al., 1996). After transcription factor binding has taken place, the basal transcription machinery and its associated complexes open the region’s chromatin and begin transcribing DNA into RNA. These crude transcripts must undergo extensive processing and maturation before they can be exported to the cytoplasm as mature messenger RNA (mRNA). Understanding the rate at which all these steps (notably transcription factor binding and the production of mRNA) change during evolution is a long-standing goal in genetics (Wray, 2007; Wittkopp and Kalay, 2012).

Technically, it is (relatively) easy to map all the contacts between the transcription factors and the DNA, and also to map all the mRNA molecules, in a biological sample using high-throughput sequencing technologies. A number of research groups have compared the amount of transcription factor binding in many species of flies and mammals (He et al., 2011; Paris et al., 2013; Schmidt et al., 2010; Ballester et al., 2014). Based on this work it seemed as if transcription factor binding evolved rapidly in mammalian tissues (Weirauch and Hughes, 2010), but only very slowly in fruit flies (He et al., 2011). However, it can be difficult to compare the first results generated in an entirely novel field of study because different groups often use very different approaches. And in this case this difficulty is further compounded by the toothbrush issue.

Now, in eLife, Trey Ideker and colleagues at the University of California San Diego – including Anne-Ruxandra Carvunis, Tina Wang and Dylan Skola as joint first authors – report that they used a new analysis pipeline to study the raw data for more than 25 species of complex eukaryotes across three animal lineages (mammals, birds and insects) that previously had only been studied in isolation (Carvunis et al., 2015). In other words, they have cleaned everyone’s teeth with the same toothbrush. Moreover, their pipeline could be tweaked to vary the analysis parameters for all the datasets across three lineages at once, thus allowing them to make like-with-like comparisons.

This intellectual scrubbing resulted in two major insights. First, it appears that transcription factor binding (which dictates the function of the genome) and mRNA both evolve at a shared (and perhaps even fundamental) rate in complex eukaryotes. This result is somewhat surprising since most evolutionary geneticists think that the mechanisms that influence genome or functional evolution for the lineages studied by Carvunis et al. are radically different.

Second, particularly in mammals, the evolution of the genome sequence en masse is much more rapid than the evolution of transcription factor binding and transcription. This disconnect may be linked to the instability of the large number largely-silent repeat elements in mammalian genomes, and/or to the fact that insects and birds have more stable genomes.

Moreover, Carvunis et al. have powerfully demonstrated why it is important for all of us in the functional genomics community to meticulously curate our raw data and to make it readily available for others to analyse. None of the insights reported in this work would have been possible without easy access to carefully annotated sequencing reads from the original studies.

References

1. Ballester B
2. Medina-Rivera A
3. Schmidt D
4. Gonzàlez-Porta M
5. Carlucci M
6. Chen X
7. Chessman K
8. Faure AJ
9. Funnell APW
10. Goncalves A
11. Kutter C
12. Lukk M
13. Menon S
14. McLaren WM
15. Stefflova K
16. Watt S
17. Weirauch MT
18. Crossley M
19. Marioni JC
20. Odom DT
21. Flicek P
22. Wilson MD
(2014) Multi-species, multi-transcription factor binding highlights conserved control of tissue-specific biological pathways
eLife 3:e02626.

https://doi.org/10.7554/eLife.02626
- Google Scholar
1. Carvunis A-R
2. Wang T
3. Skola D
4. Yu A
5. Chen J
6. Kreisberg JF
7. Ideker T
(2015) Evidence for a common evolutionary rate in metazoan transcriptional networks
eLife 4:e11615.

https://doi.org/10.7554/eLife.11615
- Google Scholar
1. Filippova GN
2. Fagerlie S
3. Klenova EM
4. Myers C
5. Dehner Y
6. Goodwin G
7. Neiman PE
8. Collins SJ
9. Lobanenkov VV
(1996) An exceptionally conserved transcriptional repressor, CTCF, employs different combinations of zinc fingers to bind diverged promoter sequences of avian and mammalian c-myc oncogenes
Molecular and Cellular Biology 16:2802–2813.

https://doi.org/10.1128/MCB.16.6.2802
- Google Scholar
1. He Q
2. Bardet AF
3. Patton B
4. Purvis J
5. Johnston J
6. Paulson A
7. Gogol M
8. Stark A
9. Zeitlinger J
(2011) High conservation of transcription factor binding and evidence for combinatorial regulation across six drosophila species
Nature Genetics 43:414–420.

https://doi.org/10.1038/ng.808
- Google Scholar
1. Paris M
2. Kaplan T
3. Li XY
4. Villalta JE
5. Lott SE
6. Eisen MB
(2013) Extensive divergence of transcription factor binding in drosophila embryos with highly conserved gene expression
PLoS Genetics 9:e1003748.

https://doi.org/10.1371/journal.pgen.1003748
- Google Scholar
1. Schmidt D
2. Wilson MD
3. Ballester B
4. Schwalie PC
5. Brown GD
6. Marshall A
7. Kutter C
8. Watt S
9. Martinez-Jimenez CP
10. Mackay S
11. Talianidis I
12. Flicek P
13. Odom DT
(2010) Five-vertebrate ChIP-seq reveals the evolutionary dynamics of transcription factor binding
Science 328:1036–1040.

https://doi.org/10.1126/science.1186176
- Google Scholar
(2009) A census of human transcription factors: function, expression and evolution
Nature Reviews Genetics 10:252–263.

https://doi.org/10.1038/nrg2538
- Google Scholar
1. Weirauch MT
2. Hughes TR
(2010) Dramatic changes in transcription factor binding over evolutionary time
Genome Biology 11:122.

https://doi.org/10.1186/gb-2010-11-6-122
- Google Scholar
1. Wittkopp PJ
2. Kalay G
(2012) Cis-regulatory elements: molecular mechanisms and evolutionary processes underlying divergence
Nature Reviews Genetics 13:59–69.

https://doi.org/10.1038/nrg3095
- Google Scholar
1. Wray GA
(2007) The evolutionary significance of cis-regulatory mutations
Nature Reviews Genetics 8:206–216.

https://doi.org/10.1038/nrg2063
- Google Scholar

Article and author information

Author details

Duncan T Odom, Reviewing Editor

Cancer Research UK Cambridge Institute, University of Cambridge, Cambridge, United Kingdom

For correspondence
Duncan.Odom@cruk.cam.ac.uk

Competing interests
The author declares that no competing interests exist.

Publication history

Version of Record published: February 11, 2016

Copyright

This article is distributed under the terms of the Creative Commons Attribution License, which permits unrestricted use and redistribution provided that the original author and source are credited.