Computational and Systems Biology

Synthetic Biology: Minimal cells, maximal knowledge

Modeling all the chemical reactions that take place in a minimal cell will help us to understand the fundamental interactions that power life.

Mar 12, 2019

https://doi.org/10.7554/eLife.45379

Open access
Copyright information

Download
Cite
CommentOpen annotations (there are currently 0 annotations on this page).
Share

Université de Sherbrooke, Canada
University of California, United States
Technical University of Denmark, Denmark
University of California, Unites States

If we could map and understand every single molecular process in a cell, we would have a better grasp of the fundamental principles of life. We could ultimately use this knowledge to design and create artificial organisms. An obvious way to start this endeavor is to study minimal cells, natural or synthetic organisms that contain only the bare minimum of genetic information needed to survive. By building and studying these very simplified cells – so simple they have been described as the ‘hydrogen atoms of biology’ (Morowitz, 1984) – we may be able to dissect all the molecular mechanisms required to sustain cellular life.

The elucidation of the DNA double helix in 1953, and the subsequent cracking of the genetic code, made it possible to link molecular processes to DNA sequences (Figure 1). In turn, whole genome sequencing has revealed a collection of molecular roles encoded in the genomes of a great number of organisms, starting in 1995 with the first complete bacterial genomes (Fleischmann et al., 1995; Fraser et al., 1995), and then expanding thanks to next-generation sequencing methods (McGuire et al., 2008; Spencer, 2008). Yet, this has also showed that we do not know or can only guess the roles of many genes which are essential to life.

Figure 1

Download asset Open asset

Synthetic biology and minimal cells: an historical perspective.

Elucidating the DNA double helix marked the beginning of the molecular biology era, and it became possible to study molecular mechanisms that underpinned observable phenotypes. DNA sequencing methods improved, leading to whole-genome sequencing at the end of the 1990s. Methods for mathematical cell modeling were developed during the 1980s and 1990s, and computer simulations of metabolic networks (also known as genome-scale models of metabolism, or GEMs) could be reconstructed. A defining moment took place in 2008 (red), with the creation of the first artificial genome that mimicked the genetic information of *M. genitalium*, the free-living, non-synthetic organism with the smallest genome. Thanks to developments in next-generation sequencing methods, this was paired with the rise of large-scale genome sequencing ventures, such as the human microbiome and the 1000 genomes projects. Advances in whole-genome synthesis, assembly, and transplantation helped create the first cell living with an entirely synthetic genome shortly after. Taken together, these achievements marked the coming of age for synthetic biology.

In 2008, as large-scale sequencing projects were initiated, a group of scientists at the J. Craig Venter Institute (JCVI) artificially recreated the genome of a bacterium. The team made DNA fragments in the laboratory, and then used a combination of chemistry and biology techniques to assemble the pieces ‘in the right order’, using the genetic information of the Mycoplasma genitalium bacteria as a template (Gibson et al., 2008). This marked a significant branching point in the history of biology: while the previous decades had focused on acquiring as much knowledge as possible about natural organisms, creating a genome from scratch in a laboratory demonstrated the potential to design synthetic cells (Figure 1). This shifted synthetic biology, the field in which researchers try to build biological entities, towards an engineering discipline that could work at the scale of a genome. The same team then went on to build Mycoplasma mycoides JCVI-syn1.0, the first living cell with an entirely artificial chromosome (Gibson et al., 2010). In both cases, the artificial genetic information faithfully replicated that found in the wild-type bacteria.

The next goal was to piece together an artificial genome that contains only those genes that are absolutely necessary for life and growth. In 2016, after years of design and testing, the genetic information in JCVI-syn1.0 was whittled down to produce M. mycoides JCVI-syn3.0, which harbors the smallest genome of any free-living organism (Hutchison et al., 2016). Notably, JCVI-syn3.0 was originally reported to contain 149 genes whose roles were unknown. Since then this number has shrunk to 91, and further reducing this figure still represents the next challenge in synthetic biology (Danchin and Fang, 2016).

Now, in eLife, Zan Luthey-Schulten and colleagues at the JCVI, the University of Illinois at Urbana-Champaign, the University of California at San Diego, and the University of Florida – including Marian Breuer as first author – report the first computational or 'in silico' model for a synthetic minimal organism (Breuer et al., 2019). The team reconstructed the complete set of chemical reactions that take place in the organism (that is, its metabolism). This effort bridges the gap between DNA sequences and molecular processes at the level of an entire biological system.

Breuer et al. performed their modeling work on M. mycoides JCVI-syn3.0A, a robust variation of JCVI-syn3.0 that contains 11 more genes. This was required because genome reduction involves a high number of genetic modifications, which tend to produce weaker cells that are harder to grow under laboratory conditions (Choe et al., 2019). To create their computational model, the team used the biochemical knowledge readily available for the parent strain JCVI-syn1.0 and identified the remaining candidate genes that participate in metabolism in JCVI-syn3.0A. These genes were then associated with cellular chemical reactions and, step-by-step, the entire metabolic network was modeled. This approach regroups the extensive knowledge on the metabolism of JCVI-syn3.0A in a single, highly valuable community resource that can help interrogate missing roles in the metabolic network and integrate experimental data.

Once a genome-scale model was obtained, it became possible to use it to perform computer simulations of different cellular phenotypes. Briefly, the in silico model represents the optimal metabolic state of the cell as an optimization problem on which constraints are applied. For instance, the metabolic models are constrained by the balance of reactants and products in a given chemical reaction (stoichiometry), and the conversion rates of the metabolites (flux bounds). Breuer et al. simulated the growth phenotype of JCVI-syn3.0A by optimizing for the production of cellular biomass, and then juxtaposed the predictions with real-life data, such as results from quantitative proteomics studies. In particular, they compared the genes that the model deemed essential with those highlighted when systematically mutating the genome of JCVI-syn3.0A. This revealed 30 genes that are required for survival but whose role is unknown. Understanding what these genes do is the next priority in the effort to complete the characterization of all molecular processes in a cell.

Overall, the model and experimental data generally agreed on their identification of essential genes; yet, a perfect match was not achieved, as is also the case when similar computational models are applied to natural organisms. Still, one would imagine that if this standard were within reach, it would be achieved first for minimal cells. To improve the quality of prediction, constraints that are more accurate need to be applied, and this would require additional information. For example, a completely defined media that contains only the necessary nutrients for JCVI-syn3.0A should be generated. It would also prove useful to have a precise biomass composition, that is, a detailed report of the proportion of major molecules and metabolites in the cell. Finally, many biochemical processes, such as isozymes (when enzymes with different structures catalyze the same reaction) or promiscuous reactions (when an enzyme can participate in many reactions) would need to be carefully investigated.

Such constraint-based modeling may be key to help with the generation of working genomes from square one, and in this regard, the model generated by Breuer et al. is the first of many steps to perfectly mirror a synthetic cell in silico. Next, the simulation could be expanded beyond metabolism to include other sets of biological processes, such as the gene expression machinery. This would help identify key constraints and trade-offs that cells must deal with in the struggle for life. In turn, these constraints could become the framework required to artificially design increasingly complex organisms, much like the hydrogen atom paved the way to understanding the behavior of more complex elements.

References

1. Breuer M
2. Earnest TM
3. Merryman C
4. Wise KS
5. Sun L
6. Lynott MR
7. Hutchison CA
8. Smith HO
9. Lapek JD
10. Gonzalez DJ
11. de Crécy-Lagard V
12. Haas D
13. Hanson AD
14. Labhsetwar P
15. Glass JI
16. Luthey-Schulten Z
(2019) Essential metabolism for a minimal cell
eLife 8:e36842.

https://doi.org/10.7554/eLife.36842
- PubMed
- Google Scholar
1. Choe D
2. Lee JH
3. Yoo M
4. Hwang S
5. Sung BH
6. Cho S
7. Palsson B
8. Kim SC
9. Cho BK
(2019) Adaptive laboratory evolution of a genome-reduced Escherichia coli
Nature Communications 10:935.

https://doi.org/10.1038/s41467-019-08888-6
- PubMed
- Google Scholar
1. Danchin A
2. Fang G
(2016) Unknown unknowns: essential genes in quest for function
Microbial Biotechnology 9:530–540.

https://doi.org/10.1111/1751-7915.12384
- PubMed
- Google Scholar
1. Fleischmann RD
2. Adams MD
3. White O
4. Clayton RA
5. Kirkness EF
6. Kerlavage AR
7. Bult CJ
8. Tomb JF
9. Dougherty BA
10. Merrick JM
(1995) Whole-genome random sequencing and assembly of Haemophilus Influenzae Rd
Science 269:496–512.

https://doi.org/10.1126/science.7542800
- PubMed
- Google Scholar
1. Fraser CM
2. Gocayne JD
3. White O
4. Adams MD
5. Clayton RA
6. Fleischmann RD
7. Bult CJ
8. Kerlavage AR
9. Sutton G
10. Kelley JM
11. Fritchman RD
12. Weidman JF
13. Small KV
14. Sandusky M
15. Fuhrmann J
16. Nguyen D
17. Utterback TR
18. Saudek DM
19. Phillips CA
20. Merrick JM
21. Tomb JF
22. Dougherty BA
23. Bott KF
24. Hu PC
25. Lucier TS
26. Peterson SN
27. Smith HO
28. Hutchison CA
29. Venter JC
(1995) The minimal gene complement of Mycoplasma genitalium
Science 270:397–404.

https://doi.org/10.1126/science.270.5235.397
- PubMed
- Google Scholar
1. Gibson DG
2. Benders GA
3. Andrews-Pfannkoch C
4. Denisova EA
5. Baden-Tillson H
6. Zaveri J
7. Stockwell TB
8. Brownley A
9. Thomas DW
10. Algire MA
11. Merryman C
12. Young L
13. Noskov VN
14. Glass JI
15. Venter JC
16. Hutchison CA
17. Smith HO
(2008) Complete chemical synthesis, assembly, and cloning of a Mycoplasma genitalium genome
Science 319:1215–1220.

https://doi.org/10.1126/science.1151721
- PubMed
- Google Scholar
1. Gibson DG
2. Glass JI
3. Lartigue C
4. Noskov VN
5. Chuang RY
6. Algire MA
7. Benders GA
8. Montague MG
9. Ma L
10. Moodie MM
11. Merryman C
12. Vashee S
13. Krishnakumar R
14. Assad-Garcia N
15. Andrews-Pfannkoch C
16. Denisova EA
17. Young L
18. Qi ZQ
19. Segall-Shapiro TH
20. Calvey CH
21. Parmar PP
22. Hutchison CA
23. Smith HO
24. Venter JC
(2010) Creation of a bacterial cell controlled by a chemically synthesized genome
Science 329:52–56.

https://doi.org/10.1126/science.1190719
- PubMed
- Google Scholar
1. Hutchison CA
2. Chuang RY
3. Noskov VN
4. Assad-Garcia N
5. Deerinck TJ
6. Ellisman MH
7. Gill J
8. Kannan K
9. Karas BJ
10. Ma L
11. Pelletier JF
12. Qi ZQ
13. Richter RA
14. Strychalski EA
15. Sun L
16. Suzuki Y
17. Tsvetanova B
18. Wise KS
19. Smith HO
20. Glass JI
21. Merryman C
22. Gibson DG
23. Venter JC
(2016) Design and synthesis of a minimal bacterial genome
Science 351:aad6253.

https://doi.org/10.1126/science.aad6253
- PubMed
- Google Scholar
(2008) Ethical, legal, and social considerations in conducting the Human Microbiome Project
Genome Research 18:1861–1864.

https://doi.org/10.1101/gr.081653.108
- PubMed
- Google Scholar
1. Morowitz HJ
(1984)
Special guest lecture the completeness of molecular biology

Israel Journal of Medical Sciences 2:.
- Google Scholar
Website
1. Spencer G
(2008) International consortium announces the 1000 genomes project
Accessed February 25, 2019.

https://www.nih.gov/news-events/news-releases/international-consortium-announces-1000-genomes-project

Article and author information

Author details

Jean-Christophe Lachance

Jean-Christophe Lachance is in the Département de Biologie, Université de Sherbrooke, Sherbrooke, Canada

Competing interests
No competing interests declared

"This ORCID iD identifies the author of this article:" 0000-0002-3096-6995
Sébastien Rodrigue

Sébastien Rodrigue is in the Départment de Biologie, Université de Sherbrooke, Sherbrooke, Canada

Competing interests
No competing interests declared

"This ORCID iD identifies the author of this article:" 0000-0002-5366-7234
Bernhard O Palsson

Bernhard O Palsson is in the Department of Bioengineering, the Bioinformatics and Systems Biology Program and the Department of Pediatrics, University of California, San Diego, USA, and the Novo Nordisk Foundation Center for Biosustainability, Technical University of Denmark, Lyngby, Denmark

For correspondence
palsson@ucsd.edu

Competing interests
No competing interests declared

"This ORCID iD identifies the author of this article:" 0000-0003-2357-6785

Publication history

Version of Record published: March 12, 2019

Copyright

This article is distributed under the terms of the Creative Commons Attribution License, which permits unrestricted use and redistribution provided that the original author and source are credited.