The human gut chemical landscape predicts microbe-mediated biotransformation of foods and drugs

Abstract
eLife digest
Introduction
Results
Discussion
Materials and methods
Data availability
References
Article and author information
Metrics

Abstract

Microbes are nature’s chemists, capable of producing and metabolizing a diverse array of compounds. In the human gut, microbial biochemistry can be beneficial, for example vitamin production and complex carbohydrate breakdown; or detrimental, such as the reactivation of an inactive drug metabolite leading to patient toxicity. Identifying clinically relevant microbiome metabolism requires linking microbial biochemistry and ecology with patient outcomes. Here we present MicrobeFDT, a resource which clusters chemically similar drug and food compounds and links these compounds to microbial enzymes and known toxicities. We demonstrate that compound structural similarity can serve as a proxy for toxicity, enzyme sharing, and coarse-grained functional similarity. MicrobeFDT allows users to flexibly interrogate microbial metabolism, compounds of interest, and toxicity profiles to generate novel hypotheses of microbe-diet-drug-phenotype interactions that influence patient outcomes. We validate one such hypothesis experimentally, using MicrobeFDT to reveal unrecognized gut microbiome metabolism of the ovarian cancer drug altretamine.

https://doi.org/10.7554/eLife.42866.001

eLife digest

Microbes in the human gut can play helpful roles by producing vitamins or breaking down complex carbohydrates. Collectively, gut microbes carry out these roles using a large toolkit of enzymes that catalyze a diverse range of chemical reactions, some of which cannot be carried out by human enzymes. However, these microbial enzymes can also cause harm if they alter drugs in a way that makes them toxic or prevents them from working. Little is known about which microbial enzymes interact with which foods and drugs, or how these interactions affect human health.

Guthrie et al. have now developed and tested a tool called MicrobeFDT that can help researchers to understand these complex interactions. In MicrobeFDT, 10,000 compounds produced by the human body or found in food or drugs are grouped based on their structure. Compounds are linked to the microbial enzymes that interact with them and drugs are annotated with information on known toxicities. The result is a network where compounds with similar structure are linked to each other.

If a microbial enzyme interacts with one compound in a group, it may interact with related compounds as well, potentially causing similar effects on human health. The network makes it easier for researchers to work out which compounds are affected by particular gut microbes. For example, MicrobeFDT suggested how gut microbes might alter the structure of an ovarian cancer drug called altretamine, which can cause diarrhea and kidney damage as side effects. Experiments confirmed that the predicted structural change does occur in human feces.

MicrobeFDT may increase how quickly researchers can assess harmful interactions between gut microbes, food, and drugs. It also may help them to develop new strategies to improve human health based on how microbial enzymes interact with food and drugs.

https://doi.org/10.7554/eLife.42866.002

Introduction

Complex gut microbiome phenotypes shape human nutrition (Martens et al., 2014; Sonnenburg et al., 2016; Bretin et al., 2018), therapeutic drug responses (Guthrie et al., 2017; Haiser et al., 2013; Koppel et al., 2017) and disease susceptibility (Koeth et al., 2013). Multi’omic studies suggest that the human gut microbiota can be discretized at the resolution of microbial enzymes (Guthrie et al., 2017; Tang and Hazen, 2014), species (Haiser et al., 2013; Haiser et al., 2014), guilds (Joossens et al., 2011; Wu et al., 2013) or metabolites (Clayton et al., 2009) to characterize a range of human health and disease states. Gut microbial mediated biochemical transformations have consequences for drug treatment efficacy (Koppel et al., 2017; Spanogiannopoulos et al., 2016; Alexander et al., 2017; Wilson and Nicholson, 2017) and the etiology of inflammatory gastrointestinal diseases (Tilg et al., 2018; Arthur et al., 2014; Belcheva et al., 2014; Brennan and Garrett, 2016), however despite many examples there exist few unifying principles that govern microbiome impacts on human health.

Some microbiome/drug interactions have been characterized in detail. For example, the inactivation and decreased bioavailability of digoxin, a cardiac glycoside inhibitor, is linked to cgr operon expression levels in a single species, E. lenta (Haiser et al., 2013). Microbial β-glucuronidases mediate the reactivation of the key therapeutic metabolite of irinotecan, a chemotherapeutic prodrug used in the treatment of colorectal cancer, causing toxicity in some patients (Guthrie et al., 2017; Wallace et al., 2010). Notably, diet-derived compounds that are conjugated to glucuronic acid in the human liver and excreted via the biliary route into the GI tract are known substrates for microbial β-glucuronidases (O'Leary et al., 2003; Sakurama et al., 2014; Maathuis et al., 2012).

Many other gastrointestinally-routed drugs share overlapping chemical properties with diet-derived compounds. We understand in detail species-specific metabolism of some discrete chemical structures in dietary compounds, particularly polysaccharides (Martens et al., 2008); however we know little about the potential spectrum of drug metabolism by the microbiome.

Beyond the role of the microbiome in therapeutic drug treatment efficacy and polysaccharide metabolism, we have some mechanistic insight into how microbial metabolism contributes to host immunity. Microbial enzymes mediate the conversion of tryptophan into indole (Sasaki-Imamura et al., 2010) and indole derivatives (Arora and Bae, 2014) that shape human host immune responses (Levy et al., 2017; Blacher et al., 2017). Microbe produced indole 3-aldehyde functions as an activating ligand for human host aryl hydrocarbon receptors which are expressed by immune cells (Zelante et al., 2013). Indole binding induces IL-22 secretion by innate lymphoid cells, promoting the secretion of antimicrobial peptides that protects the host from pathogenic infection by Candida albicans (Zelante et al., 2013). Microbial production of short chain fatty acids (SCFAs) from dietary fiber also shapes host immunity, contributing to both innate and adaptive immune system functions (Fukuda et al., 2011; Donohoe et al., 2011; Smith et al., 2013).

Host-microbe interactions and phenotypes, ranging from host drug response to host immune response, are thus intimately connected to gut chemical signaling. Beyond these few well understood examples lie a vast space of uncharacterized microbe-drug-diet-phenotype interactions. We propose three key requirements to characterize the dynamics of the gut chemical space and its impact on health. The first is predicting which compounds microbes can metabolize, the second is connecting the chemistry of gut microbes to host phenotypes, and the third is linking gut chemistry to microbial ecology.

Towards the goal of systematically mapping the gut microbial chemistry that contributes to the metabolism of xenobiotics, including therapeutic drugs, recent efforts have used chemical structure-centric approaches to enable high-throughput computational predictions of gut microbe metabolism of drugs (Sharma et al., 2017; Mallory et al., 2018). These tools represent an important first step towards ecological and mechanistic insights into gut microbiota driven biotransformation of foods and drugs. The second requirement, which has not yet been achieved, is to connect the known and predicted chemistry of gut microbes to host phenotypes. To date, information on human responses to therapeutic drugs is available in disparate databases and formats including FDA Adverse Report System (FAERs) (Burkhart et al., 2015), the Side Effect Resource (SIDER) (Kuhn et al., 2016) and DrugBank (Law et al., 2014). The third requirement, also lacking, is to systematically link gut microbe chemistry to microbial ecology to understand how the distribution of enzymes in populations of microbes facilitates ecological interactions that structure the human gut.

Here, we develop MicrobeFDT, a resource encompassing this 3-step framework that connects compound structure, enzyme function, taxonomy, and toxicity to characterize microbe-diet-drug-phenotype interactions. We organize ~10,000 food, drug, and endogenous compounds by structural similarity. We then link toxicity, enzyme interactions, and the propensity for gut microbes to carry out metabolism on each compound to the structural similarity network. We validate MicrobeFDT computationally by demonstrating that structural similarity is a reasonable proxy for toxicity, enzyme sharing, and coarse-grained functional similarity. We propose, and experimentally validate, active gut microbiome demethylation of an ovarian cancer drug, altretamine, a metabolism that we propose may drive toxicity of this drug. All data is available in the MicrobeFDT database (MicrobeFDT; Guthrie, 2019; copy archived at https://github.com/elifesciences-publications/microbeFDT-neo4j).

Results

Structural similarity as a metric to organize enzyme/taxonomy/toxicity links between compounds

The foundation of the MicrobeFDT resource is a chemical similarity network linking 10,822 food, drug, and endogenous compounds with PubChem compound identifier (CIDs) (Kim et al., 2016). In the network, nodes designate compounds and edges are weighted by pairwise chemical substructure similarity quantified by comparing PubChem fingerprints (Kim et al., 2016) using the Tanimoto score (Bajusz et al., 2015) (Figure 1). The Tanimoto score prioritizes overlap between compounds that share substructures over compounds with shared co-absences (Bajusz et al., 2015). We hypothesized that compounds with overlapping substructure and physiochemical properties, in which one compound is a known substrate of an enzyme, will be more likely to serve as substrates for the same enzyme. Recent in silico approaches to predict enzymatic reactions of drugs in the context of human enzyme catalyzed reactions also employ this hypothesis (Niu et al., 2013; Yu et al., 2018). Substructure-based clustering thus serves as a first step towards synthesizing publicly available information on gut compound chemical diversity and gut microbiome biochemistry.

Figure 1

Download asset Open asset

MicrobeFDT is a searchable resource of gut microbiome food and drug metabolism with associated toxicities.

(1) Diet-derived, xenobiotic-derived and endogenous compounds were clustered based on the PubChem fingerprint system (Kim et al., 2016) and the Tanimoto coefficient (Bajusz et al., 2015). (2) The pairwise similarity matrix forms the basis of the (3) substructure similarity network in which nodes are compounds and links are weighted by substructure similarity. (4) A Z-score based threshold method was used to identify significant chemical similarity relationships between nodes (Baldi and Nasr, 2010). (5) The property graph model of nodes and relationships in the network highlights node-relationship pairs that can be queried. Node entities include compounds (blue), uses (orange) and enzymes (green). A compound node can have up to four types of directional relationships: compound pairwise substructure similarity, compound pairwise toxicity similarity, compound treatment use descriptor and compound microbial mediated metabolism descriptor.

https://doi.org/10.7554/eLife.42866.003

To validate that our network can identify shared metabolism, we developed an in silico prediction model to assign a probability of shared metabolism between compounds based on substructure overlap and the following physiochemical categories: geometry, functional groups, amino acid composition, polarity and hydrophobicity. We find that the probability estimates of compound-pairs sharing an enzyme based on substructure and physiochemical parameters, increase as the substructure overlap score between compound pairs increases (Figure 2). Weighting compound pair chemical similarity relationships based on substructure similarity is thus a reasonable filtering step to identify compounds that may share metabolism.

Figure 2

Download asset Open asset

Higher substructure similarity scores between pairs of compounds are associated with higher probability of sharing an enzyme.

Potential enzyme mediated metabolism of compound pairs is compared with substructure similarity to determine the probability that compounds have an experimentally determined shared enzyme (pink) or no known shared enzyme (blue). The gray vertical dashed line indicates the average cutoff for significance in substructure similarity neighborhood construction. Probability estimates are based on a Bayesian approach for support vector machines implemented in R using the probsvm package (Zhang et al., 2013).

https://doi.org/10.7554/eLife.42866.004

As an example of how the network can reveal shared metabolism we selected compounds in the network with substructure overlap with digoxin, a cardiac glycoside inhibitor. Reduction of digoxin by a human microbiome reductase inactivates the drug, contributing to poor bioavailability in some individuals (Haiser et al., 2013; Haiser et al., 2014; Lindenbaum et al., 1981). Koppel et al., biochemically characterized the capacity of a single flavin- and [4Fe-4S] cluster-dependent reductase, cgr2, to reduce various substrates with a range of substructure similarity to digoxin (Koppel et al., 2018). We identified the substructure overlap between digoxin and compounds in the Koppel et al. study that were evaluated as substrates of Cgr2 enzyme. Among the biochemically assayed compounds (Koppel et al., 2018) that are present in the MicrobeFDT network, compounds with substructure similarity scores greater than 0.8 are also substrates for Cgr2. This assessment suggests that for the cgr enzyme substructure based clustering can distinguish experimentally characterized substrates from non-substrates (Figure 3).

Figure 3

Download asset Open asset

Substructure similarity range of Cgr2 enzyme susceptible compounds.

Substructure based clustering distinguishes experimentally characterized substrates from non-substrates of the Cgr2 enzyme. Digoxin clusters with other cardenolides that are experimentally characterized substrates (Koppel et al., 2018) for Cgr2 at substructure similarity values greater than 0.8. Compounds that are not substrates of Cgr2 have lower substructure similarity with digoxin; compounds with minimal reduction (Koppel et al., 2018) include progesterone and cortisone (substructure similarity <= 0.63). Color bar intensity increases with compound overlap with digoxin.

https://doi.org/10.7554/eLife.42866.005

Previous studies have found that structural similarity predicts both toxicity and drug target similarity (Campillos et al., 2008). To evaluate whether our network also recapitulates shared drug toxicity we fit a linear regression and computed the effect size to assess the association between substructure similarity and toxicity similarity for therapeutic drugs in our network. We find that structural similarity moderately positively predicts toxicity similarity for therapeutic drug pairs linked by structural similarity overall in the network (r = 0.03116, p<2.2e-16) (Figure 4).

Figure 4

Download asset Open asset

Substructure similarity is predictive of toxicity similarity.

We evaluated the predictive power of substructure similarity to identify compounds with shared toxicity using a measure of pairwise toxicity defined by Campillos et al. (2008) and used a linear regression to determine the strength of the association. We find a modest positive correlation between substructure similarity and toxicity similarity that is stronger for more structurally similar compounds.

https://doi.org/10.7554/eLife.42866.006

Finally, we evaluated how well our compound clustering recapitulates structure-based chemical taxonomy as defined by the ClassyFire (Djoumbou Feunang et al., 2016) resource, a comprehensive chemical classification schema, at the level of superclass taxonomy. We found that substructure-based compound clustering, significantly groups compounds within a ClassyFire superclass based on a comparison of the MicrobeFDT network with a randomized network with the same number of nodes and edges (p<8.06×10–15, Wilcoxon rank-sum test). Compound-pairs at higher substructure similarity share Superclass membership at higher substructure values and at a greater frequency than randomized pairs, indicating that the MicrobeFDT substructure similarity metric can capture established chemical classifications (Figure 5).

Figure 5

Download asset Open asset

Compound-pairs share superclass annotation at a greater frequency as substructure similarity scores increase.

Ratio of compound-pairs substructure similarity with matched and unmatched superclass annotation for all compound pairs represented in MicrobeFDT. Within the hierarchical ClassyFire classification schema, the superclass level annotation represents the second level and includes 31 different structure-based categories (Djoumbou Feunang et al., 2016).

https://doi.org/10.7554/eLife.42866.007

Overlapping structural diversity of food, drug, and endogenous compounds

In the network, therapeutic drug structural diversity is embedded within food-derived chemical diversity. For example, drugs share structural similarity with food-derived compounds from a diverse range of classes including benzenoids, lipids, nucleosides and phenylpropanoids (Figure 6). Food derived compounds also contributed significantly greater molecular structure diversity (Figure 6—figure supplement 1) and higher self-similarity than therapeutic drug compounds (two-sample K-S test 0.49, p value=4.7395e-06).

Figure 6 with 1 supplement see all

Download asset Open asset

The chemical space of the gut microbiome.

(a) Chemical similarity network of food-derived or endogenous compounds (gray circles, "Other") and therapeutic drugs (black diamonds, "Drug"). Tan edges are weighted by substructure similarity where thicker edges indicate higher substructure similarity. The distribution of compounds in chemical similarity space illuminates regions of low and high chemical substructure overlap between drugs and other compounds. (b) Compounds from selected regions of the network are colored by their superclass level taxonomy based on the FooDB chemical structure classification (Wishart, 2012). Food-derived or endogenously produced compounds are identified with blue circles, therapeutic drugs with red diamonds. Within high-drug density, highlighted regions 1 and 2, drugs share substructure similarity with food-derived benzenoids, lipids, phenylpropanoids and polyketides. In the low-drug density highlighted region 3, drugs overlap with organonitrogen compounds and nucleosides. Region 4 includes organonitrogen compounds and nucleosides in addition to lipid-like molecules which have minimal overlap with therapeutic drugs.

https://doi.org/10.7554/eLife.42866.008

Figure 6—source data 1 Chemical similarity scores for drug and non-drug compounds.: https://doi.org/10.7554/eLife.42866.010
Download elife-42866-fig6-data1-v1.csv

Assessing the distribution of enzymatic functions across taxonomic groups

Metabolic functions are not necessarily equally distributed across microbes in the microbiome. For example, as described above, inactivation of digoxin, a cardiac glycoside inhibitor, is linked to cgr operon expression levels in a single species, E. lenta (Haiser et al., 2013; Koppel et al., 2018). In contrast, the deconjugation and resulting reactivation of SN-38, the active metabolite of the chemotherapeutic colorectal cancer drug irinotecan, is linked to a phylogenetically diverse guild of microbial β-glucuronidase carrying microbes (Guthrie et al., 2017; Pollet et al., 2017; Wallace et al., 2015).

The question arises, how many microbes can perform specific enzymatic functions? Knowing the taxonomic distribution of a function can guide approaches to validate hypotheses of microbiota driven modification of specific therapeutic drug or food compounds. More broadly, addressing this question informs therapeutic approaches for targeting specific enzymes to modulate patient responses to drugs and foods.

In MicrobeFDT, we quantify how many taxa have the capacity to carry out a specific function by applying a modified Simpson index function to compute an Enzyme Commission number-specific dominance (ECs_D) score for all enzymes present in the network. ECs_D scores are based on the abundance of enzymes annotated at the species level across healthy human metagenomes from the Integrative Human Microbiome Project (iHMP) (Proctor et al., 2014) and are normalized between 0 and 1. Functions carried out by small numbers of species have values closer to 0 while functions carried out by taxonomically diverse groups have functions closer to 1. Thus, the ECs_D indicates how broadly distributed a function is, a crucial metric for (1) understanding how to modify a function in the microbiome and (2) predicting how disruptive to the community modifying a function might be.

To validate ECs_D scores we first identified biochemical pathways containing enzymes with high and low taxonomic dominance in the literature. Bacterial synthesis of various B group vitamins including biotin, cobalamin and riboflavin vary in the number of potential producers at the Phylum level (Magnúsdóttir et al., 2015). The most commonly synthesized B vitamin across diverse microbial taxa is riboflavin while vitamin B12 is dominated by Fusobacteria (Magnúsdóttir et al., 2015). The ECs_D scores of cobalt-precorrin-2 C(20)-methyltransferase (0.305502) from the anaerobic Vitamin B12 synthesis pathway and riboflavin synthase (0.691618) from the riboflavin synthesis pathway in MicrobeFDT agree with the prior systematic genome assessment and experimental results of Magnúsdóttir et al. (2015) (Figure 7). While most bacteria do not synthesize sphingolipids, sphingolipid biosynthetic capacity has been identified in Sphingomonas spp, Bacteroides and human intestinal pathogens that synthesize and incorporate sphingolipids into their membranes or target host sphingolipids as a point of entry into host cell types (Heaver et al., 2018; Heung et al., 2006; Olsen and Jantzen, 2001). The low ECs_D score of phosphatidate phosphatase (0.007353), an enzyme involved in sphingolipid biosynthesis and metabolism (Olsen and Jantzen, 2001), mirrors the limited distribution of the sphingolipid biosynthetic capacity across gut microbes.

Figure 7

Download asset Open asset

Linking enzymatic functions with taxonomic diversity.

The Simpson index was adapted to describe enzyme-specific taxonomic dominance and diversity based on enzyme abundance in taxonomy-linked gene counts across healthy individuals in the Integrative Human Microbiome Project (Proctor et al., 2014). We define a microbial enzyme as high dominance and low taxonomic diversity if its Simpson index value falls below 0.46 (red dotted line), the mean value across all enzymes. Dominance-diversity values for gut microbiota functions that fall above or below the mean are highlighted by gray dashed lines and include the following enzymes and pathways: phosphatidate phosphatase (0.007353), cobalt-precorrin-2 C(20)-methyltransferase (0.305502) from the Vitamin B12 synthesis pathway, β-glucuronidase (0.691618), Acetyl-CoA synthase (0.718163) which is involved in the production of propionate from complex carbohydrates, riboflavin synthase (0.794781) from the riboflavin synthesis pathway and acetate kinase (0.931892) which is involved in acetate production. The shaded regions indicate the range of EDs_D values that are one standard deviation above and below the mean and reflect the most broadly distributed functions and most specialized functions.

https://doi.org/10.7554/eLife.42866.011

Combining chemical and toxicity similarity to predict microbial N-demethylase contribution to drug metabolism and toxicity

To provide a practical example of using multiple features of MicrobeFDT to identify uninvestigated microbiota-driven drug toxicity, we searched the network for compounds with high structural and toxicity similarity. Among these compounds were the ovarian cancer drug altretamine (Lee and Faulds, 1995) and the environmental contaminant melamine (Figure 8). Both melamine and altretamine have toxicity profiles that include diarrhea and renal toxicity (Rose et al., 1996; Zheng et al., 2013). Melamine, an industrial compound, has experimentally validated microbiome-mediated toxicity (Zheng et al., 2013). Altretamine toxicity, however, has not previously been linked to an individual’s gut microbiota. Approximately half of patients taking altretamine orally experience various forms of gastrointestinal toxicity including diarrhea, nausea and/or vomiting (Keldsen et al., 2003).

Figure 8 with 1 supplement see all

Download asset Open asset

Structure-toxicity relationship between melamine and altretamine suggests a role for microbial N-demethylases in altretamine toxicity.

(a) Substructure overlap between altretamine and its nearest neighbors in MicrobeFDT. A Z-score based threshold of significant overlap indicates that altretamine has both high substructure and (b) toxicity overlap with melamine. (c) The two compounds are distinguishable by the presence of N-methyl groups.

https://doi.org/10.7554/eLife.42866.012

Within the network altretamine is linked to microbial N-demethylase enzymes which may remove methyl groups from this compound, potentially leading to similar toxic effects as seen with melamine. We found no published experimental evidence of gut microbiota mediated conversion of altretamine. However, N-demethylases in Pseudomonas putida CBB5 enable this microbe to grow on caffeine and other purine alkaloids as the sole carbon and nitrogen source; thus annotated N-demethylases in P. putida CBB5 can act on compounds that are structurally similar to altretamine (Summers et al., 2012). Furthermore, we identify hypothetical proteins homologous to Pseudomonas putida CBB5 N-demethylases in a subset of healthy human guts (Figure 9—figure supplement 1). We hypothesized that gut microbial N-demethylases may partially or completely N-demethylate altretamine, converting it into metabolites that contribute to patient toxicity.

A first step in validating this hypothesis is to demonstrate that the gut microbiome can demethylate altretamine. We incubated altretamine in a pooled fecal slurry generated from three healthy individuals and monitored altretamine and potential metabolites using LC-MS. We controlled for the formation of spontaneous N-demethylation of altretamine, which has been reported in the literature (Damia and D'Incalci, 1995), and found that a metabolite that is structurally identical to pentamethylmelamine, a demethylated altretamine metabolite, increases in active fecal microcosms over 48 hr (Figure 9). In active fecal biotic conditions the metabolite continually increased between time 0 and 48 hr. Killed controls demonstrated an increase in metabolite between 0 and 24 hr, though to a lesser extent than in active fecal microcosms. Notably there was little metabolite formation after 24 hr, indicating that in addition to abiotic N-demethylation, active gut microbes demethylate altretamine to the putative metabolite pentamethylmelamine.

Figure 9 with 1 supplement see all

Download asset Open asset

Fecal microbiomes actively demethylate altretamine.

Liquid chromatography with tandem mass spectrometry (LC-MS) was used to quantify the formation of (a) pentamethylmelamine, an N-demethylated metabolite of altretamine identified in the pooled fecal microbiomes of three healthy unrelated individuals. (b) The formation of metabolite 1 at 24 and 48 hr was significantly increased under the experimental condition in comparison to the contribution of spontaneous N-demethylation by an unpaired two-sample Wilcoxon test (*=P < 0.05).

https://doi.org/10.7554/eLife.42866.014

Food derived compounds and non-antibiotic therapeutic drugs with potential antimicrobial properties

MicrobeFDT suggests an unrecognized role for bile acid-like foods and drugs in altering the composition of the human gut. Conjugated primary bile acids (BA) function as potent detergents and antimicrobial agents capable of dissolving microbial membranes and causing intracellular acidification; bile acid function is linked to specific structural features of these compounds (Jones et al., 2008; Begley et al., 2006). Taurochenodeoxycholic acid (TCDCA) is a taurine conjugated primary bile acid with a diet-tunable concentration in the gut (Ridlon et al., 2016). Energy drinks, animal protein and fish are rich sources of taurine while vegetarian and vegan diets dominated by fruits, vegetables, legumes and soy are poor sources (Ridlon et al., 2016). Taurine conjugated bile acids are hypothesized to contribute to the etiology of colorectal cancer by generating hydrogen sulfide during microbial mediated de-conjugation of taurine conjugates (Ridlon et al., 2016). Conjugated primary bile acids have demonstrated in vitro activity as antimicrobial compounds, for example glycocholic and taurocholic conjugated bile acids are bacteriostatic, inhibiting S. aureus growth by decreasing intracellular pH and disrupting the proton motor force (Sannasiddappa et al., 2017).

Using MicrobeFDT, we identified therapeutic drug and food compounds that are structurally similar to TCDCA; we propose these compounds might have similar antimicrobial effects on the microbiome and we discuss studies from other groups that support this hypothesis (Figure 10a).

Figure 10 with 1 supplement see all

Download asset Open asset

Food-drug compounds chemically similar to TCDCA are putative antimicrobials.

(a) Chemical structure of taurochenodeoxycholic acid (TCDCA). (b) TCDCA-like therapeutic drugs that are susceptible to bile salt hydrolases include finasteride and saxagliptin. (c) Non-susceptible TCDCA-like therapeutic drugs include betamethasone, dexamethasone and cortisone. (d) TCDCA-like food derived compounds include steviol, lanosterol and tomatidine.

https://doi.org/10.7554/eLife.42866.016

Bile salt hydrolase (BSH) mediated bile salt deconjugation is one mechanism that gut microbes use to detoxify conjugated primary bile acids (Begley et al., 2006); thus BSH activity may support gut bacterial persistence in face of frequent contact with primary BAs. We first subdivided TCDCA-like antimicrobial compounds based on BSH enzyme susceptibility. BSH enzymes are phylogenetically diverse and abundant across healthy human fecal metagenomes (Figure 10—figure supplement 1). Among the BSH-susceptible therapeutic drug compounds, we identified known antibiotics such as clindamycin and lincomycin, as well as non-antibiotic prescribed therapeutics such as finasteride, which is used for the treatment of androgenetic alopecia (Manabe et al., 2018) and benign prostatic hyperplasia (Chau et al., 2015), and the oral antidiabetic drug saxagliptin (Men et al., 2018) (Figure 10b). Notably, in a Wistar rat model of chronic bacterial prostatitis (CBP), finasteride reduces bacterial infection as a single agent and has a synergistic effect with ciprofloxacin through an unknown mechanism (Lee et al., 2011). Through in vitro studies, Chavex-Dozal and colleagues propose a role for finasteride in the prevention of Candida albicans biofilm formation and filamentation (Chavez-Dozal et al., 2014). These experimental results support the hypothesis that finasteride may have unrecognized off-target antibiotic effects.

Most TCDCA-like compounds in MicrobeFDT are non-BSH susceptible food-derived compounds. Among the TCDCA-like non-BSH susceptible compounds are oral steroid medications, including dexamethasone and betamethasone (Figure 10c). The immunomodulatory activities of glucocorticoids, including dexamethasone, involve the activation of genes related to anti-inflammatory cytokines such as IL-10 and proteins that inhibit the pro-inflammatory NFκB signaling pathway (Coutinho and Chapman, 2011; Huang et al., 2015). Dexamethasone has known anti-microbial properties. For example, dexamethasone has dose-dependent anti-microbial activity against clinically isolated Streptococcus milleri, Aspergillus flavus, and Aspergillus fumigatus in culture, while not killing Staphylococcus aureus (Neher et al., 2008). Pseudomonas aeruginosa was found to be susceptible to dexamethasone at high concentrations (Neher et al., 2008). Cortisone, which also has significant structural overlap with TCDCA, has been linked to a variety of opportunistic infections by enteric bacterial pathogens, for example an increase in gastrointestinal parasites (Nair et al., 1981) and reactivation of Chlamydia pneumoniae (Laitinen et al., 1996).

Food-derived TCDCA-like compounds include steviol, lanosterol and tomatidine. Steviol is a component of stevia which has antimicrobial properties against Borrelia burgdorferi in vitro (Theophilus et al., 2015), and lanosterol derivatives have antifungal activities (Shingate, 2013). Tomatidine was recently identified as an antibiotic molecule that inhibits ATP synthesis against Staphylococcus aureus (Lamontagne Boulet et al., 2018); we hypothesize that the antimicrobial activity of this compound may include intracellular acidification given its structural overlap with TCDCA (Figure 10d). The network thus identifies compounds with known anti-microbial properties in addition to proposing additional, structurally related compounds with uncharacterized effects. We propose that in addition to modulating immune responses, bile salt-like compounds may selectively alter human microbiomes, again, with unknown consequences for treatment outcomes and health.

MicrobeFDT identifies the diet-derived substrate pool for microbial BGs and candidates for nutritional competition with SN-38G

We next applied MicrobeFDT to identify diet-derived substrates of a gut carbohydrate active enzyme, β-glucuronidase. β-glucuronidases play a major role in the toxicity of the colorectal cancer chemotherapeutic prodrug irinotecan (CPT-11), whose active form, SN-38, is inactivated by hepatic glucuronidation and excreted into the gut as the inactive metabolite SN-38 glucuronide (SN-38G) (Wallace et al., 2010; Sparreboom et al., 1998). Microbial β-glucuronidases hydrolyze the glucuronide group, releasing the aglycone SN-38 into the intestinal environment (Figure 11a). Deconjugation promotes epithelial damage and severe diarrhea in some patients and in mouse models (Wallace et al., 2010; Sparreboom et al., 1998; Slatter et al., 2000).

Figure 11

Download asset Open asset

Microbial β-glucuronidase potential substrate pool of compounds structurally similar to SN-38G.

(a) SN-38G conversion to SN-38 in the gut is mediated by microbial β-glucuronidases. (b) The substrate pool for β-glucuronidases with above threshold substructure overlap with SN-38G are members of a diverse range of chemical structure superclasses as defined by FooDB chemical ontology (Wishart, 2018). (c) These compounds include glucuronidated food-derived compounds (purple), endogenous glucuronides (tan) and other non-glucuronides (blue).

https://doi.org/10.7554/eLife.42866.018

We previously demonstrated that individual human fecal samples have variable capacities to deconjugate SN-38G (Guthrie et al., 2017). Identifying the full substrate pool of β-glucuronidases is thus important for 1) understanding how diet contributes to β-glucuronidase abundance and expression levels in the gut and 2) to enable novel therapeutic strategies such as nutritional competition.

Some food compounds may be preferred substrates for microbiome β-glucuronidases which would otherwise deconjugate SN-38G. If true, one could potentially alleviate toxicity associated with the deconjugation of SN-38G via nutritional competition with a preferred substrate. Therefore, we scanned the chemical similarity module containing SN-38G for dietary compounds that may serve as alternative substrates for microbial β-glucuronidases. Most compounds identified as significantly similar to SN-38G were food derivatives or other constituents (Figure 11b). Among these targets were flavonoids such as baicalin and scutellarin which are widely distributed in plants (Kumar and Pandey, 2013) (Figure 11c). We propose that these compounds may compete with SN-38G for turnover by microbial β-glucuronidases and are a potential avenue for decreasing the adverse drug responses associated with irinotecan administration.

Discussion

The chemical space of the human gastrointestinal tract ecosystem is shaped by host dietary intake, xenobiotic exposure, and host and gut microbiome derived products. In turn, diet shapes the composition and potential niches of organisms within human gut microbiomes. A combination of compound, host, and microbiome features influence potential microbial metabolism. Examining these features individually cannot reliably infer clinical phenotypes associated with microbiome/compound interactions. Two molecules may have the same toxicity profile but very different biochemistry, for example. Automated enzyme annotation may be incorrect, and compound structural similarity is often insufficient to predict substrate preferences. Finally, enzymes that carry out a reaction associated with a patient phenotype may be unevenly distributed across microbes and across human microbiomes. MicrobeFDT is designed to overcome some of these limitations by enabling a more holistic analysis of toxicity, structure, metabolism and ecology. We used a combination of network features to successfully predict the novel microbial metabolism of the cancer drug altretamine.

Metabolomics data indicate active demethylation of altretamine by fecal slurries but cannot propose a mechanism by which microbial activity metabolizes this compound. MicrobeFDT suggests that altretamine is a putative substrate of microbial N-demethylases. Microbe-mediated N-demethylation reactions, and the subsequent release of N-methyl groups, occur as a part of amino acid and nucleotide metabolism (D’Mello and International, 2017). Notably, diet is a source of amino acids which are derived in part from metabolism of dietary choline, carnitine and legumes, and have physiological functions for bacteria including osmoprotection and incorporation into bacterial flagellin proteins and lipid membranes (Goldfine and Hagen, 1968). Amino acid-specific bacterial N-demethylases have been identified but are poorly characterized (Wargo, 2017). Additionally, fecal and species specific N-demethylation has been observed for other therapeutic drugs and commonly ingested compounds such as caffeine, which clusters with altretamine in the network due to its structural similarity (Summers et al., 2012; Caldwell and Hawksworth, 1973; Clark et al., 1983; Colombo et al., 1982). N-demethylases can act on chemically diverse substrates (Wargo, 2017; Burnet et al., 2000). Given this body of evidence, we propose N-demethylases may demethylate altretamine partially or completely, creating metabolites that are toxic to patients.

Human gut metagenomic data indicate that Rieske family oxidative N-demethylases are carried by a small, phylogenetically conserved set of gut taxa, with notable inter-personal variation. That these enzymes require oxygen may make them more relevant during disruptions to gut homeostasis when oxygen becomes available, such as colonic crypt hyperplasia caused by injuries to the intestinal epithelia (Litvak et al., 2018). Finally, we note that N-demethylation in the gut may be relevant for differences in individual metabolism of numerous other compounds such as the cancer drug tamoxifen, the widely used antihistamine diphenhydramine, and theobromine, a plant alkaloid found in foods. While is possible that N-demethylation is enzyme independent or that enzymes annotated with other functions are responsible for this activity, MicrobeFDT provides a clear path forward for mechanistic studies of N-demethylation in the gut.

Beyond predicting the toxicity or function of gut compounds, MicrobeFDT identifies the larger substrate pool for enzymes involved in drug metabolism. For example, shared conjugation patterns may represent a clinically relevant way to group compounds that share microbial enzymatic processing. As an example, compounds inactivated by glucuronidation are susceptible to microbial β-glucuronidase-mediated reactivation. We used MicrobeFDT to identify compounds structurally similar to the conjugated, detoxified irinotecan metabolite SN-38G and found dietary substrates that may interact with similar β-glucuronidases that this drug interacts with. Structurally similar compounds may act competitively – via inhibition of SN-38G turnover by higher priority β-glucuronidase substrates or synergistically – via substrate inducible transcriptional upregulation of β-glucuronidase enzymes. A person consuming a large amount of the plant-based compound scutellarin as part of a supplement, for example, might be inadvertently modulating the effects of their cancer therapy.

Outside of drug metabolism, β-glucuronidases mediate deconjugation and enterohepatic circulation of estrogens, impacting the human host total estrogen burden (Shapira et al., 2013; Kwa et al., 2016). It has been hypothesized that β-glucuronidase deconjugation may result in greater absorption of estrogens and thus influence the development of estrogen-driven cancers including breast, ovarian and endometrial cancers (Shapira et al., 2013; Kwa et al., 2016). Our network is useful for developing mechanistic hypotheses targeting how diet and the microbiome jointly act as moderators of estrogen-driven cancers, and to suggest opportunities for diet-based modulation of total estrogen levels.

An important step towards characterizing the role of the gut microbiome in shaping individual responses to foods and drugs is identifying how gut microbiome metabolism varies from compound to compound and how this metabolism relates to inter-personal variation in diet or drug responses to specific compounds. To tackle this challenge, we add the context of taxonomic diversity to the predicted impact of microbial on specific targets by quantifying enzyme specific taxonomic dominance and diversity with a novel metric, the ECs_D score. This score distinguishes enzymatic activities carried out by single species or few taxa, such as N-demethylase activity, from those where many taxa may contribute, such as β-glucuronidase and bile salt hydrolase activity. The ECs_D score is a readout of potential substrate metabolism at the community level that can be linked to inter-personal variation in gut function and phenotypic outcomes.

The structural similarity network that underlies MicrobeFDT could be improved by using compound atom and bond connectivity information as an additional filtering step for compounds of interest, for example by using information from the SMARTS molecular pattern matching language (Chepelev et al., 2012). SMARTS can be used to specify sub-structural patterns in molecules; these patterns could be added to MicrobeFDT as an additional information source indicating potential active moieties in compounds.

MicrobeFDT does not predict substrate specificity for microbiome enzymes; available data and methods are not sufficient to achieve this goal. Enzyme promiscuity also shapes the probability that two chemically overlapping compounds will be processed by the same enzyme. A future improvement to our resource could extract data from resources like RetroRules (Duigou et al., 2019), which uses SMARTS strings to define reaction rules, or utilize the Promis server measure of enzyme multifunctionality (Carbonell and Faulon, 2010) to further support a user’s ranking of hypothesized compound-enzyme interactions.

It must be noted that the set of diet-derived and xenobiotic compounds that form the basis of the network is a non-exhaustive representation of the gut chemical landscape. Efforts to characterize the gut chemical space using metabolomics approaches including mass spectrometry and nuclear magnetic resonance spectroscopy will play key roles in elucidating a fuller gut chemical landscape (Vernocchi et al., 2016; Wishart, 2012). MicrobeFDT does not address the issue of compound concentrations in the gut, which are vital to assess likely physiological effects. Lastly, MicrobeFDT is limited to enzymes in KEGG, and does not address the many hypothetical enzyme sequences identified through metagenomic sequencing. Despite these limitations, MicrobeFDT highlights areas of known gut chemical space for which our understanding of microbial processing is limited and is a powerful tool to guide mechanistic investigations into diet-drug-microbiota interactions.

Materials and methods

Key resources table

Reagent type (species) or resource	Designation	Source or reference	Identifiers	Additional information
Biological sample (community microbiota, feces)	fecal sample	other		fecal sample obtained from three healthy adults
Chemical compound, drug (altretamine)	altretamine	Sigma	Pubchem_ID:329748966; CAS_No:645-05-6	prepared in DMSO, 0.1 mM final concentration in fecal slurry
Other	Brain Heart Infusion broth	Himedia	Himedia:M210I
Chemical compound (Dimethyl sulfodixe)	DMSO	MP Biomedicals	MP:191418; CAS_No:67-68-5
Chemical compound (Melamine-triamine-(15N3))	Melamine-triamine-(¹⁵N₃)	Sigma	Pubchem:329758619; CAS_No:287476-11-3	prepared in DMSO, 400 nM final concentration in analytical sample

MicrobeFDT pipeline

Request a detailed protocol

The MicrobeFDT graph database encodes heterogeneous information on the interactions between compounds and microbial enzymes in the gut chemical landscape, highlighting the following four relationships across 13,440 nodes (10,822 xenobiotic, diet-derived and human gut endogenous compounds, 2062 microbial enzymes and 525 therapeutic drug use labels) defined from publicly available data directly or computed: (1) compound-compound substructure similarity; (2) compound-compound toxicity similarity; (3) microbial enzyme-compound interactions; and (4) drug-indication associations. The database is implemented in Neo4j (https://neo4j.com/) and can be queried through the Cypher Query Language. Through graph-based searches users can query the network based on node or relationship features. MicrobeFDT can be accessed here (Guthrie, 2019).

Publicly available datasets and resources used as inputs for the network

Request a detailed protocol

The SIDER 4.1 side effect resource is a database of approved medicines and their known adverse reactions (Kuhn et al., 2016). Drugs from this database with pharmacokinetic profiles that involve entry into the gastrointestinal tract were identified through literature mining and manual curation and indexed by their PubChem CID identifier (Kim et al., 2016). Drug use annotations were based on the WHO Anatomical Classification System (Skrbo et al., 2004). FooDB (http://foodb.ca/) (Wishart, 2018), a database containing raw food component structures, biological interactions and chemical properties was the source of food components linked to PubChem CID identifiers. ClassyFire was used to annotate all xenobiotic and food derived compounds with a shared chemical taxonomy (Djoumbou Feunang et al., 2016).

To link microbial enzymes to the set of compounds they metabolize we used KEGGREST (v1.14.1) to retrieve KEGG compound identifiers with links to Enzyme Commission numbers, metabolic modules and pathways, and presence in either organisms listed as microbial or Homo sapiens (Tenenbaum, 2019). Enzyme abundance data across human metagenomes were determined based on the total abundance of each enzyme in the healthy participants of the Human Microbiome Project. This data was extracted from the Integrated Microbial Genomes database (Markowitz et al., 2012). Enzyme specific dominance scores (ECs_D), which is a measure of the number of different species that carry a specific enzyme, were computed based on species-specific enzyme abundance data from healthy individuals from the Integrative Human Microbiome Project (Proctor et al., 2014).

Construction and assessment of the drug-food chemical similarity network

Chemical similarity calculation

Request a detailed protocol

To determine the pairwise chemical substructure similarity between all compounds we used the PubChem 2D molecular fingerprint (Kim et al., 2016). The fingerprint is an 881 dimension binary vector in which each bit represents a specific element, functional group, ring system or other discrete chemical entity (Kim et al., 2016). Similarity was defined by the Tanimoto coefficient of the molecular fingerprint representations present between two compounds (Bajusz et al., 2015).

Network construction

Request a detailed protocol

Similarity scores are percentages of substructure overlap between pairs of compounds and have values between 0 to 1. Similarity scores are filtered such that compound pairs with less than 0.3 substructure similarity were removed. These pairwise similarity scores formed the basis of the undirected chemical similarity network, where nodes represent compounds and edges represent substructure similarity score.

Network filtering

Request a detailed protocol

To cluster compounds in the network based on substructure similarity we used the Walktrap community detection method (Pons and Latapy, 2006) implemented in R/igraph v.1.1.1 (R Development Core Team, 2016). Within a community, significant similarity scores were defined as those with Z-scores of 1 standard deviation or greater away from the mean (Baldi and Nasr, 2010).

Assessment of compound substructure-based clustering recapitulation of chemical ontology

Request a detailed protocol

The MicrobeFDT substructure similarity network is defined by the Tanimoto coefficient of the PubChem 2D molecular fingerprint representations between two compounds (Kim et al., 2016; Bajusz et al., 2015). To assess how well compound substructure-based clustering recapitulates chemical ontology we compared network features between the MicrobeFDT substructure similarity network and randomized network with the same number of nodes, edges and labels. Each compound label includes a ClassyFire (Djoumbou Feunang et al., 2016) schema derived hierarchical set of chemical descriptors. The chemical similarity network was rendered in Cytoscape using NetworkRandomizer (Martens et al., 2014). Using the Wilcoxon rank-sum test we compared superclass level chemical descriptors across connected compounds between the real and random network. For the MicrobeFDT network, we also computed the ratio of compounds pairs with matched Superclass annotation to unmatched annotations for all pairs with the same substructure score to assess the relationship between substructure similarity and shared chemical ontology.

Predicting the probability of association of compound pairs both serving as substrates for an enzyme based on substructure and physiochemical parameters

Request a detailed protocol

Each compound pair was assigned one of two labels, associate or non-associated, based on whether both compounds are substrates for the same enzyme (associated) or not (non-associated), given compound-enzyme relationships in the KEGG database (Kanehisa and Goto, 2000). The DataWarrior program (Sander et al., 2015) was used to identify the following parameter categories for each compound: geometry, functional groups, aromaticity, amino acid composition, polarity and hydrophobicity. In order to translate compound pair substructure and physiochemical parameters into a probability of overlapping metabolism we used a machine learning approach for generating probability estimates for multi-class classification problems (Zhang et al., 2013; Wu et al., 2004). Briefly, this approach builds a multi-class prediction model by using pair-wise coupling. We then implemented the prediction model using the probsvm package in R using a one-vs-one decomposition scheme (Zhang et al., 2013).

Assessing toxicity similarity

Request a detailed protocol

Toxicity similarity was computed as described by Campillos and colleagues (Campillos et al., 2008) with three key steps: (1) extraction and standardization of side effect concepts across drugs of interest; (2) weighting of unique side effect concepts based on frequency of occurrence and correlation with other side effects; and (3) computation of pair-wise toxicity similarity between drugs based on weighted side-effect concept values. Briefly, Campillos et al. curated a dictionary of side-effects based on the Concepts of the Coding Symbols for Thesaurus of Adverse Reaction Terms (COSTART) ontology (US Food and Drug Administration, 1995). Side-effect information on therapeutic drug package labels was identified from publicly available sources and searched against this dictionary such that all unique side effect concepts per drug were based on COSTART ontology. For our analysis we used the side effect labels for therapeutic drugs of interest that were extracted from the Medical Dictionary for Regulatory Activities (Brown et al., 1999), which is an updated replacement of COSTART, and made publicly available at the download page for SIDER 4.1 which can be found here.

In Campillos et al., each side effect concept was given a rarity score which is the frequency at which it is found across all drug side effect lists. To account for co-dependence between side effects Campillos and colleagues also determined the correlation between all side effects based, using the Tanimoto score between pairs of side effects. This measure is based on how many drugs share a given side effect relative to the number of drugs that have either. The resulting matrix was used as input for the Gerstein-Sonnhammer-Chothia Algorithm (Gerstein et al., 1994), to output a score for each concept that down weights concepts that are redundant. We used a publicly available implementation of this algorithm in R available here. Pair-wise toxicity similarity between drugs was computed based on summing the products of weights over all shared side effect concepts between drug pairs. We fit a linear regression to determine whether there is a linear relationship between compound pair substructure similarity and toxicity similarity.

Taxonomic signatures of microbial enzymes

Request a detailed protocol

For each enzyme, we computed an enzyme commission number-specific dominance (ECs_D) score. This score is an application of the Simpson’s index, which is particularly sensitive to sample evenness (DeJong, 1975), and describes the dominance and diversity profile of species carrying the enzyme (Ofaim et al., 2017). The taxa-specific enzyme abundance information is based on data collected as a part of the integrative Human Microbiome Project (iHMP) (PRJNA306874) (Proctor et al., 2014). ECs_D scores are reported as Simpson index measure (Simpson, 1949) subtracted from one, as implemented in the phyloseq R package (McMurdie and Holmes, 2013). In this implementation, the Simpson dominance index per enzyme defined by its enzyme commission number (D(EC)) is computed such that n is number of individuals of each species that carry the enzyme and N is the total number of individuals of all species that carry the enzyme (1). For better interpretability, the dominance scores are subtracted from 1 (2).

D (E C) = \frac{Σ n (n - 1)}{N (N - 1)}

E C_{S D} = 1 - D (E C)

Thus, enzyme functions carried out by small numbers of microbes have values closer to 0 while functions carried out by taxonomically diverse groups have functions closer to 1.

Altretamine microbiome turnover validation

Collection and preparation of fecal samples

Request a detailed protocol

Fresh fecal samples were provided by three healthy adult men aged 23–30 with no history of antibiotics for 6 months prior to the study. The study was approved by the Albert Einstein College of Medicine Institutional Review Board. Samples were deposited, immediately stored on ice, and processed within 1 hr. One gram stool from each donor was added to 300 mL BHI supplemented with 0.5% glucose (weight/volume) and homogenized. The final fecal slurry was thus comprised of the pooled feces of the three donors at 1% w/v.

Altretamine metabolism

Request a detailed protocol

Fecal slurry cultures were incubated at 37°C in the dark under aerobic conditions. Altretamine stock was prepared in DMSO. Experimental cultures received a final concentration of 100 µM altretamine in DMSO and were prepared in triplicate. Triplicate heat-killed and denatured cultures were autoclaved three times on successive days and also received 100 µM altretamine in DMSO after the third autoclave. Background cultures received fecal slurry and DMSO but no altretamine. To determine matrix effects of altretamine in the media, a sterile media control was amended with 100 µM altretamine in DMSO. Cultures were sampled, immediately snap-frozen in liquid N₂ every 24 hr, and stored at −80°C until analysis.

Altretamine and metabolite quantification

Request a detailed protocol

Samples were thawed, centrifuged, and 100 µL aliquots were added to 900 µL 80% methanol. Melamine-triamine-(¹⁵N₃) was used as internal standard. Altretamine and metabolites were identified using LC/MS (Waters Acquity LC system and Waters Xevo TQ MS). Liquid samples were diluted 1:50 in 80% methanol with melamine-triamine-(¹⁵N₃) as internal standard. Each sample was injected 3 times at 5 mL/injection. Separation was performed on an ACE2 C18 column set to 45°C with 0.1% formic acid in 5% methanol (A) and 0.1% formic acid in methanol (B). Elution occurred at 0.35 ml/min with 100% A for 1 min, followed by a 1.5 min linear gradient from 100% A to 95% B, and finally 100% B for 1 min. The voltage was set to 0.044 kV.

Phylogenetic trees

N-demethylase phylogenetic tree

Request a detailed protocol

N-demethylases from Pseudomonas putida CBB5 (ndmABCD) (Summers et al., 2012) and Sphingobium sp. strain YBL2 (pdmAB) (Gu et al., 2013), both containing a Rieske non-heme iron oxygenase component, catalyze the N-demethylation of phenylurea herbicides and purine alkaloids, respectively; and range in size from 318 to 364 amino acids (Summers et al., 2012; Gu et al., 2013; Sharma et al., 2018). We clustered bacterial N-demethylase sequences described by Summers et al., and Tao et al., as well as protein sequences of >= 200 amino acids in length pulled based on text annotation from the RefSeq database (Pruitt et al., 2005) at 95% identity using the UCLUST algorithm (Edgar, 2010). The resulting 84 N-demethylase protein sequences served as a protein database which was mapped against the protein calls of healthy adult participants from the Human Microbiome Project (HMP) (PRJNA43021) using the UBLAST algorithm (Edgar, 2010) and e-value cutoff of e-40. N-demethylase hits of 200 amino acids or greater formed the basis of a phylogenetic tree which was constructed by aligning the protein sequences using MUSCLE with default parameters (Edgar, 2004). Aligned sequences were trimmed at 70% identity and phylogenetic trees were built with PhyML (Guindon et al., 2010) with 100 bootstrap replicates, a JTT model of substitution, and otherwise default parameters. The trees were visualized using the packages ggpplot2 (Wickham, 2016) and phyloseq (McMurdie and Holmes, 2013) in R (R Development Core Team, 2016). Each branch was colored based on the phylum level classification of the protein, marked by similarity to the experimentally characterized N-demethylase genes ndmABCD and pdmAB and by the normalized number of total hits found across individuals in the HMP. Black circles indicate bootstrap values of 80/100 or better.

Bile salt hydrolase phylogenetic tree

Request a detailed protocol

We identified bile salt hydrolase protein sequences based on text annotation from the RefSeq database (Pruitt et al., 2005) and developed a curated database of protein sequences that were clustered at 95% identity using the UCLUST algorithm (Edgar, 2010) resulting in 300 bile salt hydrolase protein sequences with a minimum amino acid length cutoff of 300. Bile salt hydrolase subunits can range in length up to 518 amino acids in the literature (Breton et al., 2002; Bron et al., 2006; Schmid and Roth, 1987). Sequence mapping against the HMP (Human et al., 2012), alignment and tree construction were carried out as described for the N-demethylases with the following exception: each branch representing a unique bile salt hydrolase sequence was marked by the presence or absence of reported activity in the literature.

Data availability

Data to use or reproduce MicrobeFDT can be found at https://github.com/kellylab/microbeFDT-neo4j (copy archived at https://github.com/elifesciences-publications/microbeFDT-neo4j).

References

(2017) Gut microbiota modulation of chemotherapy efficacy and toxicity
Nature Reviews Gastroenterology & Hepatology 14:356–365.

https://doi.org/10.1038/nrgastro.2017.20
- PubMed
- Google Scholar
1. Arora PK
2. Bae H
(2014) Identification of new metabolites of bacterial transformation of indole by gas chromatography-mass spectrometry and high performance liquid chromatography
International Journal of Analytical Chemistry 2014:1–5.

https://doi.org/10.1155/2014/239641
- PubMed
- Google Scholar
(2014) Microbial genomic analysis reveals the essential role of inflammation in bacteria-induced colorectal cancer
Nature Communications 5:4724.

https://doi.org/10.1038/ncomms5724
- PubMed
- Google Scholar
(2015) Why is Tanimoto index an appropriate choice for fingerprint-based similarity calculations?
Journal of Cheminformatics 7:20.

https://doi.org/10.1186/s13321-015-0069-3
- PubMed
- Google Scholar
1. Baldi P
2. Nasr R
(2010) When is Chemical Similarity Significant? The Statistical Distribution of Chemical Similarity Scores and Its Extreme Values
Journal of Chemical Information and Modeling 50:1205–1222.

https://doi.org/10.1021/ci100010v
- Google Scholar
1. Begley M
2. Hill C
3. Gahan CG
(2006) Bile salt hydrolase activity in probiotics
Applied and Environmental Microbiology 72:1729–1738.

https://doi.org/10.1128/AEM.72.3.1729-1738.2006
- PubMed
- Google Scholar
1. Belcheva A
2. Irrazabal T
3. Robertson SJ
4. Streutker C
5. Maughan H
6. Rubino S
7. Moriyama EH
8. Copeland JK
9. Surendra A
10. Kumar S
11. Green B
12. Geddes K
13. Pezo RC
14. Navarre WW
15. Milosevic M
16. Wilson BC
17. Girardin SE
18. Wolever TMS
19. Edelmann W
20. Guttman DS
21. Philpott DJ
22. Martin A
(2014) Gut microbial metabolism drives transformation of MSH2-deficient colon epithelial cells
Cell 158:288–299.

https://doi.org/10.1016/j.cell.2014.04.051
- PubMed
- Google Scholar
(2017) Microbiome-Modulated metabolites at the interface of host immunity
The Journal of Immunology 198:572–580.

https://doi.org/10.4049/jimmunol.1601247
- Google Scholar
1. Brennan CA
2. Garrett WS
(2016) Gut Microbiota, inflammation, and colorectal cancer
Annual Review of Microbiology 70:395–411.

https://doi.org/10.1146/annurev-micro-102215-095513
- PubMed
- Google Scholar
(2018) Microbiota and metabolism - What’s New in 2018
The American Journal of Physiology.

https://doi.org/10.1152/ajpendo.00014.2018
- Google Scholar
1. Breton YL
2. Mazé A
3. Hartke A
4. Lemarinier S
5. Auffray Y
6. Rincé A
(2002) Isolation and characterization of bile salts-sensitive mutants of Enterococcus faecalis
Current Microbiology 45:0434–0439.

https://doi.org/10.1007/s00284-002-3714-9
- PubMed
- Google Scholar
(2006) DNA micro-array-based identification of bile-responsive genes in Lactobacillus plantarum
Journal of Applied Microbiology 100:728–738.

https://doi.org/10.1111/j.1365-2672.2006.02891.x
- PubMed
- Google Scholar
1. Brown EG
2. Wood L
3. Wood S
(1999) The medical dictionary for regulatory activities (MedDRA)
Drug Safety 20:109–117.

https://doi.org/10.2165/00002018-199920020-00002
- PubMed
- Google Scholar
(2015) Data mining FAERS to analyze molecular targets of drugs highly associated with Stevens-Johnson syndrome
Journal of Medical Toxicology 11:265–273.

https://doi.org/10.1007/s13181-015-0472-1
- PubMed
- Google Scholar
1. Burnet MW
2. Goldmann A
3. Message B
4. Drong R
5. El Amrani A
6. Loreau O
7. Slightom J
8. Tepfer D
(2000) The stachydrine catabolism region in sinorhizobium meliloti encodes a multi-enzyme complex similar to the xenobiotic degrading systems in other bacteria
Gene 244:151–161.

https://doi.org/10.1016/S0378-1119(99)00554-5
- PubMed
- Google Scholar
1. Caldwell J
2. Hawksworth GM
(1973) The demethylation of methamphetamine by intestinal microflora
Journal of Pharmacy and Pharmacology 25:422–424.

https://doi.org/10.1111/j.2042-7158.1973.tb10043.x
- PubMed
- Google Scholar
1. Campillos M
2. Kuhn M
3. Gavin AC
4. Jensen LJ
5. Bork P
(2008) Drug target identification using side-effect similarity
Science 321:263–266.

https://doi.org/10.1126/science.1158140
- PubMed
- Google Scholar
1. Carbonell P
2. Faulon JL
(2010) Molecular signatures-based prediction of enzyme promiscuity
Bioinformatics 26:2012–2019.

https://doi.org/10.1093/bioinformatics/btq317
- PubMed
- Google Scholar
1. Chau CH
2. Price DK
3. Till C
4. Goodman PJ
5. Chen X
6. Leach RJ
7. Johnson-Pais TL
8. Hsing AW
9. Hoque A
10. Tangen CM
11. Chu L
12. Parnes HL
13. Schenk JM
14. Reichardt JK
15. Thompson IM
16. Figg WD
(2015) Finasteride concentrations and prostate cancer risk: results from the prostate cancer prevention trial
PLOS ONE 10:e0126672.

https://doi.org/10.1371/journal.pone.0126672
- PubMed
- Google Scholar
1. Chavez-Dozal AA
2. Lown L
3. Jahng M
4. Walraven CJ
5. Lee SA
(2014) In vitro Analysis of Finasteride Activity against Candida albicans Urinary Biofilm Formation and Filamentation
Antimicrobial Agents and Chemotherapy 58:5855–5862.

https://doi.org/10.1128/AAC.03137-14
- Google Scholar
(2012) Self-organizing ontology of biochemically relevant small molecules
BMC Bioinformatics 13:.

https://doi.org/10.1186/1471-2105-13-3
- PubMed
- Google Scholar
(1983) Demethylation of imipramine by enteric bacteria
Journal of Pharmaceutical Sciences 72:1288–1290.

https://doi.org/10.1002/jps.2600721113
- PubMed
- Google Scholar
(2009) Pharmacometabonomic identification of a significant host-microbiome metabolic interaction affecting human drug metabolism
PNAS 106:14728–14733.

https://doi.org/10.1073/pnas.0904489106
- PubMed
- Google Scholar
(1982) Routes of elimination of hexamethylmelamine and pentamethylmelamine in the rat
Xenobiotica 12:315–321.

https://doi.org/10.3109/00498258209052471
- PubMed
- Google Scholar
1. Coutinho AE
2. Chapman KE
(2011) The anti-inflammatory and immunosuppressive effects of glucocorticoids, recent developments and mechanistic insights
Molecular and Cellular Endocrinology 335:2–13.

https://doi.org/10.1016/j.mce.2010.04.005
- PubMed
- Google Scholar
1. Damia G
2. D'Incalci M
(1995) Clinical pharmacokinetics of altretamine
Clinical Pharmacokinetics 28:439–448.

https://doi.org/10.2165/00003088-199528060-00002
- PubMed
- Google Scholar
1. DeJong TM
(1975) A Comparison of Three Diversity Indices Based on Their Components of Richness and Evenness
Oikos 26:222.

https://doi.org/10.2307/3543712
- Google Scholar
1. Djoumbou Feunang Y
2. Eisner R
3. Knox C
4. Chepelev L
5. Hastings J
6. Owen G
7. Fahy E
8. Steinbeck C
9. Subramanian S
10. Bolton E
11. Greiner R
12. Wishart DS
(2016) ClassyFire: automated chemical classification with a comprehensive, computable taxonomy
Journal of Cheminformatics 8:61.

https://doi.org/10.1186/s13321-016-0174-y
- PubMed
- Google Scholar
1. Donohoe DR
2. Garge N
3. Zhang X
4. Sun W
5. O'Connell TM
6. Bunger MK
7. Bultman SJ
(2011) The microbiome and butyrate regulate energy metabolism and autophagy in the mammalian colon
Cell Metabolism 13:517–526.

https://doi.org/10.1016/j.cmet.2011.02.018
- PubMed
- Google Scholar
(2019) RetroRules: a database of reaction rules for engineering biology
Nucleic Acids Research 47:D1229–D1235.

https://doi.org/10.1093/nar/gky940
- PubMed
- Google Scholar
Book
1. D’Mello JPF
2. International CAB
(2017)
The Handbook of Microbial Metabolism of Amino Acids

CABI press.
- Google Scholar
1. Edgar RC
(2004) MUSCLE: multiple sequence alignment with high accuracy and high throughput
Nucleic Acids Research 32:1792–1797.

https://doi.org/10.1093/nar/gkh340
- PubMed
- Google Scholar
1. Edgar RC
(2010) Search and clustering orders of magnitude faster than BLAST
Bioinformatics 26:2460–2461.

https://doi.org/10.1093/bioinformatics/btq461
- PubMed
- Google Scholar
1. Fukuda S
2. Toh H
3. Hase K
4. Oshima K
5. Nakanishi Y
6. Yoshimura K
7. Tobe T
8. Clarke JM
9. Topping DL
10. Suzuki T
11. Taylor TD
12. Itoh K
13. Kikuchi J
14. Morita H
15. Hattori M
16. Ohno H
(2011) Bifidobacteria can protect from enteropathogenic infection through production of acetate
Nature 469:543–547.

https://doi.org/10.1038/nature09646
- PubMed
- Google Scholar
(1994) Volume changes in protein evolution
Journal of Molecular Biology 236:1067–1078.

https://doi.org/10.1016/0022-2836(94)90012-4
- PubMed
- Google Scholar
1. Goldfine H
2. Hagen P
(1968)
N-Methyl groups in bacterial lipids III. phospholipids of hyphomicrobia

Journal of Bacteriology 95:367–375.
- PubMed
- Google Scholar
1. Gu T
2. Zhou C
3. Sørensen SR
4. Zhang J
5. He J
6. Yu P
7. Yan X
8. Li S
(2013) The novel bacterial N-demethylase PdmAB is responsible for the initial step of N,N-dimethyl-substituted phenylurea herbicide degradation
Applied and Environmental Microbiology 79:7846–7856.

https://doi.org/10.1128/AEM.02478-13
- PubMed
- Google Scholar
(2010) New algorithms and methods to estimate maximum-likelihood phylogenies: assessing the performance of PhyML 3.0
Systematic Biology 59:307–321.

https://doi.org/10.1093/sysbio/syq010
- PubMed
- Google Scholar
1. Guthrie L
2. Gupta S
3. Daily J
4. Kelly L
(2017) Human microbiome signatures of differential colorectal cancer drug metabolism
Npj Biofilms and Microbiomes 3:27.

https://doi.org/10.1038/s41522-017-0034-1
- PubMed
- Google Scholar
Software
1. Guthrie L
(2019) MicrobeFDT-neo4j
Github.

https://github.com/kellylab/microbeFDT-neo4j
(2013) Predicting and manipulating cardiac drug inactivation by the human gut bacterium eggerthella lenta
Science 341:295–298.

https://doi.org/10.1126/science.1235872
- PubMed
- Google Scholar
(2014) Mechanistic insight into digoxin inactivation by eggerthella lenta augments our understanding of its pharmacokinetics
Gut Microbes 5:233–238.

https://doi.org/10.4161/gmic.27915
- PubMed
- Google Scholar
(2018) Sphingolipids in host-microbial interactions
Current Opinion in Microbiology 43:92–99.

https://doi.org/10.1016/j.mib.2017.12.011
- PubMed
- Google Scholar
(2006) Role of sphingolipids in microbial pathogenesis
Infection and Immunity 74:28–39.

https://doi.org/10.1128/IAI.74.1.28-39.2006
- PubMed
- Google Scholar
1. Huang EY
2. Inoue T
3. Leone VA
4. Dalal S
5. Touw K
6. Wang Y
7. Musch MW
8. Theriault B
9. Higuchi K
10. Donovan S
11. Gilbert J
12. Chang EB
(2015) Using corticosteroids to reshape the gut microbiome
Inflammatory Bowel Diseases 21:963–972.

https://doi.org/10.1097/MIB.0000000000000332
- Google Scholar
(2012) Structure, function and diversity of the healthy human microbiome
Nature 486:207–214.

https://doi.org/10.1038/nature11234
- PubMed
- Google Scholar
1. Jones BV
2. Begley M
3. Hill C
4. Gahan CG
5. Marchesi JR
(2008) Functional and comparative metagenomic analysis of bile salt hydrolase activity in the human gut microbiome
PNAS 105:13580–13585.

https://doi.org/10.1073/pnas.0804437105
- PubMed
- Google Scholar
1. Joossens M
2. Huys G
3. Cnockaert M
4. De Preter V
5. Verbeke K
6. Rutgeerts P
7. Vandamme P
8. Vermeire S
(2011) Dysbiosis of the faecal microbiota in patients with Crohn's disease and their unaffected relatives
Gut 60:631–637.

https://doi.org/10.1136/gut.2010.223263
- PubMed
- Google Scholar
1. Kanehisa M
2. Goto S
(2000) KEGG: kyoto encyclopedia of genes and genomes
Nucleic Acids Research 28:27–30.

https://doi.org/10.1093/nar/28.1.27
- PubMed
- Google Scholar
(2003) Altretamine (hexamethylmelamine) in the treatment of platinum-resistant ovarian cancer: a phase II study
Gynecologic Oncology 88:118–122.

https://doi.org/10.1016/S0090-8258(02)00103-8
- PubMed
- Google Scholar
1. Kim S
2. Thiessen PA
3. Bolton EE
4. Chen J
5. Fu G
6. Gindulyte A
7. Han L
8. He J
9. He S
10. Shoemaker BA
11. Wang J
12. Yu B
13. Zhang J
14. Bryant SH
(2016) PubChem substance and compound databases
Nucleic Acids Research 44:D1202–D1213.

https://doi.org/10.1093/nar/gkv951
- PubMed
- Google Scholar
1. Koeth RA
2. Wang Z
3. Levison BS
4. Buffa JA
5. Org E
6. Sheehy BT
7. Britt EB
8. Fu X
9. Wu Y
10. Li L
11. Smith JD
12. DiDonato JA
13. Chen J
14. Li H
15. Wu GD
16. Lewis JD
17. Warrier M
18. Brown JM
19. Krauss RM
20. Tang WH
21. Bushman FD
22. Lusis AJ
23. Hazen SL
(2013) Intestinal microbiota metabolism of L-carnitine, a nutrient in red meat, promotes atherosclerosis
Nature Medicine 19:576–585.

https://doi.org/10.1038/nm.3145
- PubMed
- Google Scholar
(2017) Chemical transformation of xenobiotics by the human gut microbiota
Science 356:eaag2770.

https://doi.org/10.1126/science.aag2770
- PubMed
- Google Scholar
(2018) Discovery and characterization of a prevalent human gut bacterial enzyme sufficient for the inactivation of a family of plant toxins
eLife 7:e33953.

https://doi.org/10.7554/eLife.33953
- PubMed
- Google Scholar
1. Kuhn M
2. Letunic I
3. Jensen LJ
4. Bork P
(2016) The SIDER database of drugs and side effects
Nucleic Acids Research 44:D1075–D1079.

https://doi.org/10.1093/nar/gkv1075
- PubMed
- Google Scholar
1. Kumar S
2. Pandey AK
(2013) Chemistry and biological activities of flavonoids: an overview
The Scientific World Journal 2013:1–16.

https://doi.org/10.1155/2013/162750
- Google Scholar
1. Kwa M
2. Plottel CS
3. Blaser MJ
4. Adams S
(2016) The intestinal microbiome and estrogen Receptor-Positive female breast cancer
Journal of the National Cancer Institute 108:djw029.

https://doi.org/10.1093/jnci/djw029
- Google Scholar
(1996)
Reactivation of chlamydia pneumoniae infection in mice by cortisone treatment

Infection and Immunity 64:1488–1490.
- PubMed
- Google Scholar
(2018) Tomatidine Is a Lead Antibiotic Molecule That Targets Staphylococcus aureus ATP Synthase Subunit C
Antimicrobial Agents and Chemotherapy 62:AAC.02197-17.

https://doi.org/10.1128/AAC.02197-17
- Google Scholar
1. Law V
2. Knox C
3. Djoumbou Y
4. Jewison T
5. Guo AC
6. Liu Y
7. Maciejewski A
8. Arndt D
9. Wilson M
10. Neveu V
11. Tang A
12. Gabriel G
13. Ly C
14. Adamjee S
15. Dame ZT
16. Han B
17. Zhou Y
18. Wishart DS
(2014) DrugBank 4.0: shedding new light on drug metabolism
Nucleic Acids Research 42:D1091–D1097.

https://doi.org/10.1093/nar/gkt1068
- PubMed
- Google Scholar
1. Lee CB
2. Ha US
3. Yim SH
4. Lee HR
5. Sohn DW
6. Han CH
7. Cho YH
(2011) Does Finasteride have a preventive effect on chronic bacterial prostatitis? pilot study using an animal model
Urologia Internationalis 86:204–209.

https://doi.org/10.1159/000320109
- PubMed
- Google Scholar
1. Lee CR
2. Faulds D
(1995) Altretamine
Drugs 49:932–953.

https://doi.org/10.2165/00003495-199549060-00007
- Google Scholar
(2017) Microbiome, metabolites and host immunity
Current Opinion in Microbiology 35:8–15.

https://doi.org/10.1016/j.mib.2016.10.003
- PubMed
- Google Scholar
1. Lindenbaum J
2. Rund DG
3. Butler VP
4. Tse-Eng D
5. Saha JR
(1981) Inactivation of digoxin by the gut flora: reversal by antibiotic therapy
New England Journal of Medicine 305:789–794.

https://doi.org/10.1056/NEJM198110013051403
- PubMed
- Google Scholar
(2018) Colonocyte metabolism shapes the gut microbiota
Science 362:eaat9076.

https://doi.org/10.1126/science.aat9076
- PubMed
- Google Scholar
(2012) Galacto-oligosaccharides have prebiotic activity in a dynamic in vitro colon model using a (13)C-labeling technique
The Journal of Nutrition 142:1205–1212.

https://doi.org/10.3945/jn.111.157420
- PubMed
- Google Scholar
(2015) Systematic genome assessment of B-vitamin biosynthesis suggests co-operation among gut microbes
Frontiers in Genetics 6:148.

https://doi.org/10.3389/fgene.2015.00148
- PubMed
- Google Scholar
(2018) Chemical reaction vector embeddings: towards predicting drug metabolism in the human gut microbiome
Pacific Symposium on Biocomputing. Pacific Symposium on Biocomputing 23:56–67.

https://doi.org/10.1142/9789813235533_0006
- PubMed
- Google Scholar
(2018) Guidelines for the diagnosis and treatment of male-pattern and female-pattern hair loss, 2017 version
The Journal of Dermatology 45:1031–1043.

https://doi.org/10.1111/1346-8138.14470
- PubMed
- Google Scholar
1. Markowitz VM
2. Chen IM
3. Palaniappan K
4. Chu K
5. Szeto E
6. Grechkin Y
7. Ratner A
8. Jacob B
9. Huang J
10. Williams P
11. Huntemann M
12. Anderson I
13. Mavromatis K
14. Ivanova NN
15. Kyrpides NC
(2012) IMG: the integrated microbial genomes database and comparative analysis system
Nucleic Acids Research 40:D115–D122.

https://doi.org/10.1093/nar/gkr1044
- PubMed
- Google Scholar
(2008) Mucosal glycan foraging enhances fitness and transmission of a saccharolytic human gut bacterial symbiont
Cell Host & Microbe 4:447–457.

https://doi.org/10.1016/j.chom.2008.09.007
- PubMed
- Google Scholar
(2014) The Devil lies in the details: how variations in polysaccharide fine-structure impact the physiology and evolution of gut microbes
Journal of Molecular Biology 426:3851–3865.

https://doi.org/10.1016/j.jmb.2014.06.022
- PubMed
- Google Scholar
1. McMurdie PJ
2. Holmes S
(2013) Phyloseq: an R package for reproducible interactive analysis and graphics of microbiome census data
PLOS ONE 8:e61217.

https://doi.org/10.1371/journal.pone.0061217
- Google Scholar
1. Men P
2. Li XT
3. Tang HL
4. Zhai SD
(2018) Efficacy and safety of saxagliptin in patients with type 2 diabetes: a systematic review and meta-analysis
PLOS ONE 13:e0197321.

https://doi.org/10.1371/journal.pone.0197321
- PubMed
- Google Scholar
(1981) Corticosteroid treatment increases parasite numbers in murine giardiasis
Gut 22:475–480.

https://doi.org/10.1136/gut.22.6.475
- PubMed
- Google Scholar
1. Neher A
2. Arnitz R
3. Gstöttner M
4. Schäfer D
5. Kröss E-M
6. Nagl M
(2008) Antimicrobial activity of dexamethasone and its combination with N-Chlorotaurine
Archives of Otolaryngology–Head & Neck Surgery 134:615.

https://doi.org/10.1001/archotol.134.6.615
- Google Scholar
1. Niu B
2. Huang G
3. Zheng L
4. Wang X
5. Chen F
6. Zhang Y
7. Huang T
(2013) Prediction of Substrate-Enzyme-Product interaction based on molecular descriptors and physicochemical properties
BioMed Research International 2013:1–7.

https://doi.org/10.1155/2013/674215
- Google Scholar
1. O'Leary KA
2. Day AJ
3. Needs PW
4. Mellon FA
5. O'Brien NM
6. Williamson G
(2003) Metabolism of quercetin-7- and quercetin-3-glucuronides by an in vitro hepatic model: the role of human beta-glucuronidase, Sulfotransferase, catechol-O-methyltransferase and multi-resistant protein 2 (MRP2) in flavonoid metabolism
Biochemical Pharmacology 65:479–491.

https://doi.org/10.1016/S0006-2952(02)01510-1
- PubMed
- Google Scholar
1. Ofaim S
2. Ofek-Lalzar M
3. Sela N
4. Jinag J
5. Kashi Y
6. Minz D
7. Freilich S
(2017) Analysis of Microbial Functions in the Rhizosphere Using a Metabolic-Network Based Framework for Metagenomics Interpretation
Frontiers in Microbiology 8:1606.

https://doi.org/10.3389/fmicb.2017.01606
- PubMed
- Google Scholar
1. Olsen I
2. Jantzen E
(2001) Sphingolipids in bacteria and fungi
Anaerobe 7:103–112.

https://doi.org/10.1006/anae.2001.0376
- Google Scholar
1. Pollet RM
2. D'Agostino EH
3. Walton WG
4. Xu Y
5. Little MS
6. Biernat KA
7. Pellock SJ
8. Patterson LM
9. Creekmore BC
10. Isenberg HN
11. Bahethi RR
12. Bhatt AP
13. Liu J
14. Gharaibeh RZ
15. Redinbo MR
(2017) An atlas of β-Glucuronidases in the human intestinal microbiome
Structure 25:967–977.

https://doi.org/10.1016/j.str.2017.05.003
- PubMed
- Google Scholar
1. Pons P
2. Latapy M
(2006) Computing communities in large networks using random walks
Journal of Graph Algorithms and Applications 10:191–218.

https://doi.org/10.7155/jgaa.00124
- Google Scholar
(2014) The integrative human microbiome project: dynamic analysis of microbiome-host omics profiles during periods of human health and disease
Cell Host & Microbe 16:276–289.

https://doi.org/10.1016/j.chom.2014.08.014
- PubMed
- Google Scholar
(2005) NCBI Reference Sequence (RefSeq): a curated non-redundant sequence database of genomes, transcripts and proteins
Nucleic Acids Research 33:D501–D504.

https://doi.org/10.1093/nar/gki025
- PubMed
- Google Scholar
1. R Development Core Team
(2016) R: A language and environment for statistical computing
R: A language and environment for statistical computing, 1.1.1, Vienna, Austria, http://www.r-project.org/.

http://www.r-project.org/
- Google Scholar
(2016) Taurocholic acid metabolism by gut microbes and colon cancer
Gut Microbes 7:201–215.

https://doi.org/10.1080/19490976.2016.1150414
- PubMed
- Google Scholar
(1996) Phase II evaluation of altretamine for advanced or recurrent squamous cell carcinoma of the cervix: a gynecologic oncology group study
Gynecologic Oncology 62:100–102.

https://doi.org/10.1006/gyno.1996.0196
- PubMed
- Google Scholar
1. Sakurama H
2. Kishino S
3. Uchibori Y
4. Yonejima Y
5. Ashida H
6. Kita K
7. Takahashi S
8. Ogawa J
(2014) β-Glucuronidase from lactobacillus brevis useful for baicalin hydrolysis belongs to glycoside hydrolase family 30
Applied Microbiology and Biotechnology 98:4021–4032.

https://doi.org/10.1007/s00253-013-5325-8
- PubMed
- Google Scholar
(2015) DataWarrior: an open-source program for chemistry aware data visualization and analysis
Journal of Chemical Information and Modeling 55:460–473.

https://doi.org/10.1021/ci500588j
- PubMed
- Google Scholar
(2017) In vitro Antibacterial Activity of Unconjugated and Conjugated Bile Salts on Staphylococcus aureus
Frontiers in Microbiology 8:1581.

https://doi.org/10.3389/fmicb.2017.01581
- PubMed
- Google Scholar
(2010) Production of indole from L-tryptophan and effects of these compounds on biofilm formation by fusobacterium nucleatum ATCC 25586
Applied and Environmental Microbiology 76:4260–4268.

https://doi.org/10.1128/AEM.00166-10
- PubMed
- Google Scholar
1. Schmid MB
2. Roth JR
(1987) Gene location affects expression level in Salmonella typhimurium
Journal of Bacteriology 169:2872–2875.

https://doi.org/10.1128/jb.169.6.2872-2875.1987
- PubMed
- Google Scholar
1. Shapira I
2. Sultan K
3. Lee A
4. Taioli E
(2013) Evolving concepts: how diet and the intestinal microbiome act as modulators of breast malignancy
ISRN Oncology 2013:1–10.

https://doi.org/10.1155/2013/693920
- Google Scholar
(2017) A novel approach for the prediction of species-specific biotransformation of xenobiotic/drug molecules by the human gut microbiota
Scientific Reports 7:9751.

https://doi.org/10.1038/s41598-017-10203-6
- PubMed
- Google Scholar
(2018) N-methylation in amino acids and peptides: Scope and limitations
Biopolymers 109:e23110.

https://doi.org/10.1002/bip.23110
- PubMed
- Google Scholar
Book
1. Shingate BB
(2013)
Synthesis and Antimicrobial Activity of Novel Oxysterols From Lanosterol

Elsevier.
- Google Scholar
1. Simpson EH
(1949) Measurement of diversity
Nature 163:688.

https://doi.org/10.1038/163688a0
- Google Scholar
(2004)
Classification of drugs using the ATC system (Anatomic, therapeutic, chemical classification) and the latest changes]

Medicinski Arhiv 58:138–141.
- PubMed
- Google Scholar
1. Slatter JG
2. Schaaf LJ
3. Sams JP
4. Feenstra KL
5. Johnson MG
6. Bombardt PA
7. Cathcart KS
8. Verburg MT
9. Pearson LK
10. Compton LD
11. Miller LL
12. Baker DS
13. Pesheck CV
14. Lord RS
(2000)
Pharmacokinetics, metabolism, and excretion of irinotecan (CPT-11) following I.V. infusion of [(14)C]CPT-11 in cancer patients

Drug Metabolism and Disposition: The Biological Fate of Chemicals 28:423–433.
- PubMed
- Google Scholar
1. Smith PM
2. Howitt MR
3. Panikov N
4. Michaud M
5. Gallini CA
6. Bohlooly-Y M
7. Glickman JN
8. Garrett WS
(2013) The microbial metabolites, short-chain fatty acids, regulate colonic treg cell homeostasis
Science 341:569–573.

https://doi.org/10.1126/science.1241165
- PubMed
- Google Scholar
(2016) Diet-induced extinctions in the gut microbiota compound over generations
Nature 529:212–215.

https://doi.org/10.1038/nature16504
- PubMed
- Google Scholar
(2016) The microbial pharmacists within Us: a metagenomic view of xenobiotic metabolism
Nature Reviews Microbiology 14:273–287.

https://doi.org/10.1038/nrmicro.2016.17
- PubMed
- Google Scholar
1. Sparreboom A
2. de Jonge MJ
3. de Bruijn P
4. Brouwer E
5. Nooter K
6. Loos WJ
7. van Alphen RJ
8. Mathijssen RH
9. Stoter G
10. Verweij J
(1998)
Irinotecan (CPT-11) metabolism and disposition in cancer patients

Clinical Cancer Research 4:2747–2754.
- PubMed
- Google Scholar
1. Summers RM
2. Louie TM
3. Yu CL
4. Gakhar L
5. Louie KC
6. Subramanian M
(2012) Novel, highly specific N-demethylases enable bacteria to live on caffeine and related purine alkaloids
Journal of Bacteriology 194:2041–2049.

https://doi.org/10.1128/JB.06637-11
- PubMed
- Google Scholar
1. Tang WH
2. Hazen SL
(2014) The contributory role of gut microbiota in cardiovascular disease
Journal of Clinical Investigation 124:4204–4211.

https://doi.org/10.1172/JCI72331
- PubMed
- Google Scholar
1. Tenenbaum D
(2019)
KEGGREST: Client-side REST access to KEGG

KEGGREST: Client-side REST access to KEGG, 1.20.0.
- Google Scholar
1. Theophilus PA
2. Victoria MJ
3. Socarras KM
4. Filush KR
5. Gupta K
6. Luecke DF
7. Sapi E
(2015) Effectiveness of stevia rebaudiana whole leaf extract against the various morphological forms of borrelia burgdorferi in vitro
European Journal of Microbiology and Immunology 5:268–280.

https://doi.org/10.1556/1886.2015.00031
- PubMed
- Google Scholar
1. Tilg H
2. Adolph TE
3. Gerner RR
4. Moschen AR
(2018) The intestinal microbiota in colorectal cancer
Cancer Cell 33:954–964.

https://doi.org/10.1016/j.ccell.2018.03.004
- PubMed
- Google Scholar
Book
1. US Food and Drug Administration
(1995)
COSTART: Coding Symbols for Thesaurus of Adverse Reaction Terms

NITS.
- Google Scholar
(2016) Gut microbiota profiling: metabolomics based approach to unravel compounds affecting human health
Frontiers in Microbiology 7:1144.

https://doi.org/10.3389/fmicb.2016.01144
- PubMed
- Google Scholar
1. Wallace BD
2. Wang H
3. Lane KT
4. Scott JE
5. Orans J
6. Koo JS
7. Venkatesh M
8. Jobin C
9. Yeh LA
10. Mani S
11. Redinbo MR
(2010) Alleviating cancer drug toxicity by inhibiting a bacterial enzyme
Science 330:831–835.

https://doi.org/10.1126/science.1191175
- PubMed
- Google Scholar
1. Wallace BD
2. Roberts AB
3. Pollet RM
4. Ingle JD
5. Biernat KA
6. Pellock SJ
7. Venkatesh MK
8. Guthrie L
9. O'Neal SK
10. Robinson SJ
11. Dollinger M
12. Figueroa E
13. McShane SR
14. Cohen RD
15. Jin J
16. Frye SV
17. Zamboni WC
18. Pepe-Ranney C
19. Mani S
20. Kelly L
21. Redinbo MR
(2015) Structure and inhibition of microbiome β-Glucuronidases essential to the alleviation of cancer drug toxicity
Chemistry & Biology 22:1238–1249.

https://doi.org/10.1016/j.chembiol.2015.08.005
- PubMed
- Google Scholar
Book
1. Wargo MJ
(2017)
The Handbook of Microbial Metabolism of Amino Acids

CABI.
- Google Scholar
Book
1. Wickham H
(2016)
Ggplot2: Elegant Graphics for Data Analysis

New York: Springer-Verlag.
- Google Scholar
1. Wilson ID
2. Nicholson JK
(2017) Gut microbiome interactions with drug metabolism, efficacy, and toxicity
Translational Research 179:204–222.

https://doi.org/10.1016/j.trsl.2016.08.002
- PubMed
- Google Scholar
Book
1. Wishart D
(2012)
Genetics Meets Metabolomics

New York: Springer.
- Google Scholar
Website
1. Wishart DS
(2018) FooDB: the food database
Accessed September 17, 2018.

http://foodb.ca/
1. Wu T-F
2. Lin C-J
3. Weng RC
(2004)
Probability estimates for Multi-class classification by pairwise coupling

Journal of Machine Learning Research 5:.
- Google Scholar
1. Wu N
2. Yang X
3. Zhang R
4. Li J
5. Xiao X
6. Hu Y
7. Chen Y
8. Yang F
9. Lu N
10. Wang Z
11. Luan C
12. Liu Y
13. Wang B
14. Xiang C
15. Wang Y
16. Zhao F
17. Gao GF
18. Wang S
19. Li L
20. Zhang H
21. Zhu B
(2013) Dysbiosis signature of fecal microbiota in colorectal cancer patients
Microbial Ecology 66:462–470.

https://doi.org/10.1007/s00248-013-0245-9
- PubMed
- Google Scholar
1. Yu MS
2. Lee HM
3. Park A
4. Park C
5. Ceong H
6. Rhee KH
7. Na D
(2018) In silico prediction of potential chemical reactions mediated by human enzymes
BMC Bioinformatics 19:207.

https://doi.org/10.1186/s12859-018-2194-2
- PubMed
- Google Scholar
1. Zelante T
2. Iannitti RG
3. Cunha C
4. De Luca A
5. Giovannini G
6. Pieraccini G
7. Zecchi R
8. D'Angelo C
9. Massi-Benedetti C
10. Fallarino F
11. Carvalho A
12. Puccetti P
13. Romani L
(2013) Tryptophan catabolites from microbiota engage aryl hydrocarbon receptor and balance mucosal reactivity via interleukin-22
Immunity 39:372–385.

https://doi.org/10.1016/j.immuni.2013.08.003
- PubMed
- Google Scholar
1. Zhang C
2. Shin SJ
3. Wang J
4. Wu Y
5. Zhang HH
(2013) probsvm: Class Probability Estimation for Support Vector Machines
probsvm: Class Probability Estimation for Support Vector Machines, 1.00, https://cran.r-project.org/web/packages/probsvm/index.html.

https://cran.r-project.org/web/packages/probsvm/index.html
- Google Scholar
1. Zheng X
2. Zhao A
3. Xie G
4. Chi Y
5. Zhao L
6. Li H
7. Wang C
8. Bao Y
9. Jia W
10. Luther M
11. Su M
12. Nicholson JK
13. Jia W
(2013) Melamine-Induced renal toxicity is mediated by the gut microbiota
Science Translational Medicine 5:172ra22.

https://doi.org/10.1126/scitranslmed.3005114
- Google Scholar

Article and author information

Author details

Leah Guthrie

Department of Systems and Computational Biology, Albert Einstein College of Medicine, New York, United States

Contribution
Conceptualization, Data curation, Formal analysis, Supervision, Funding acquisition, Visualization, Methodology, Writing—original draft, Writing—review and editing

For correspondence
leah.guthrie@phd.einstein.yu.edu

Competing interests
No competing interests declared

"This ORCID iD identifies the author of this article:" 0000-0002-9144-0110
Sarah Wolfson

Department of Systems and Computational Biology, Albert Einstein College of Medicine, New York, United States

Contribution
Validation, Methodology, Writing—review and editing

Competing interests
No competing interests declared

"This ORCID iD identifies the author of this article:" 0000-0001-9774-0977
Libusha Kelly
1. Department of Systems and Computational Biology, Albert Einstein College of Medicine, New York, United States
2. Department of Microbiology and Immunology, Albert Einstein College of Medicine, New York, United States
Contribution
Conceptualization, Supervision, Funding acquisition, Writing—original draft, Writing—review and editing

For correspondence
libusha.kelly@einstein.yu.edu

Competing interests
No competing interests declared

"This ORCID iD identifies the author of this article:" 0000-0002-7303-1022

Funding

National Institutes of Health (Training Program in Cellular and Molecular Biology and Genetics (5T32GM007491-41))

Leah Guthrie

National Institutes of Health (1R01CA222358)

Sarah Wolfson

United States Department of Defense (CA171019)

Libusha Kelly

The funders had no role in study design, data collection and interpretation, or the decision to submit the work for publication.

Acknowledgements

This work was supported by a Peer Reviewed Cancer Research Program Career Development Award from the United States Department of Defense to LK (CA171019). Leah Guthrie was supported in part by the predoctoral Training Program in Cellular and Molecular Biology and Genetics (5T32GM007491-41). Sarah Wolfson was supported in part by NIH/NCI funding (1R01CA222358). The Stable Isotope and Metabolomics Core Facility of the Diabetes Research and Training Center of the Albert Einstein College of Medicine is supported by NIH/NCI funding (P60DK020541). The authors thank the reviewers, members of the Kelly lab and Tyler Grove (Einstein) for helpful suggestions on the work.

Copyright

This article is distributed under the terms of the Creative Commons Attribution License, which permits unrestricted use and redistribution provided that the original author and source are credited.