<?xml version="1.0" encoding="UTF-8" standalone="no"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Microbiol.</journal-id>
<journal-title>Frontiers in Microbiology</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Microbiol.</abbrev-journal-title>
<issn pub-type="epub">1664-302X</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fmicb.2016.01979</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Microbiology</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>The Complete Genome Sequence of Hyperthermophile <italic>Dictyoglomus turgidum</italic> DSM 6724&#x02122; Reveals a Specialized Carbohydrate Fermentor</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name><surname>Brumm</surname> <given-names>Phillip J.</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="aff" rid="aff2"><sup>2</sup></xref>
<xref ref-type="author-notes" rid="fn001"><sup>&#x0002A;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/211959/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Gowda</surname> <given-names>Krishne</given-names></name>
<xref ref-type="aff" rid="aff2"><sup>2</sup></xref>
<xref ref-type="aff" rid="aff3"><sup>3</sup></xref>
</contrib>
<contrib contrib-type="author">
<name><surname>Robb</surname> <given-names>Frank T.</given-names></name>
<xref ref-type="aff" rid="aff4"><sup>4</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/19214/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Mead</surname> <given-names>David A.</given-names></name>
<xref ref-type="aff" rid="aff2"><sup>2</sup></xref>
<xref ref-type="aff" rid="aff5"><sup>5</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/127500/overview"/>
</contrib>
</contrib-group>
<aff id="aff1"><sup>1</sup><institution>C5-6 Technologies LLC</institution> <country>Fitchburg, WI, USA</country></aff>
<aff id="aff2"><sup>2</sup><institution>DOE Great Lakes Bioenergy Research Center, University of Wisconsin-Madison</institution> <country>Madison, WI, USA</country></aff>
<aff id="aff3"><sup>3</sup><institution>Lucigen Corporation</institution> <country>Middleton, WI, USA</country></aff>
<aff id="aff4"><sup>4</sup><institution>Department of Microbiology and Immunology, Institute of Marine and Environmental Technology, University of Maryland</institution> <country>Baltimore, MD, USA</country></aff>
<aff id="aff5"><sup>5</sup><institution>Varigen Biosciences Corporation</institution> <country>Madison, WI, USA</country></aff>
<author-notes>
<fn fn-type="edited-by"><p>Edited by: Kian Mau Goh, Universiti Teknologi Malaysia, Malaysia</p></fn>
<fn fn-type="edited-by"><p>Reviewed by: Biswarup Mukhopadhyay, Virginia Tech, USA; Ida Helene Steen, University of Bergen, Norway</p></fn>
<fn fn-type="corresp" id="fn001"><p>&#x0002A;Correspondence: Phillip J. Brumm <email>pbrumm&#x00040;c56technologies.com</email></p></fn>
<fn fn-type="other" id="fn002"><p>This article was submitted to Extreme Microbiology, a section of the journal Frontiers in Microbiology</p></fn>
</author-notes>
<pub-date pub-type="epub">
<day>20</day>
<month>12</month>
<year>2016</year>
</pub-date>
<pub-date pub-type="collection">
<year>2016</year>
</pub-date>
<volume>7</volume>
<elocation-id>1979</elocation-id>
<history>
<date date-type="received">
<day>28</day>
<month>07</month>
<year>2016</year>
</date>
<date date-type="accepted">
<day>25</day>
<month>11</month>
<year>2016</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x000A9; 2016 Brumm, Gowda, Robb and Mead.</copyright-statement>
<copyright-year>2016</copyright-year>
<copyright-holder>Brumm, Gowda, Robb and Mead</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/"><p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) or licensor are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p></license>
</permissions>
<abstract>
<p>Here we report the complete genome sequence of the chemoorganotrophic, extremely thermophilic bacterium, <italic>Dictyoglomus turgidum</italic>, which is a Gram negative, strictly anaerobic bacterium. <italic>D. turgidum</italic> and <italic>D. thermophilum</italic> together form the <italic>Dictyoglomi</italic> phylum. The two <italic>Dictyoglomus</italic> genomes are highly syntenic, and both are distantly related to <italic>Caldicellulosiruptor</italic> spp. <italic>D. turgidum</italic> is able to grow on a wide variety of polysaccharide substrates due to significant genomic commitment to glycosyl hydrolases, 16 of which were cloned and expressed in our study. The GH5, GH10, and GH42 enzymes characterized in this study suggest that <italic>D. turgidum</italic> can utilize most plant-based polysaccharides except crystalline cellulose. The DNA polymerase I enzyme was also expressed and characterized. The pure enzyme showed improved amplification of long PCR targets compared to Taq polymerase. The genome contains a full complement of DNA modifying enzymes, and an unusually high copy number (4) of a new, ancestral family of polB type nucleotidyltransferases designated as MNT (minimal nucleotidyltransferases). Considering its optimal growth at 72&#x000B0;C, <italic>D. turgidum</italic> has an anomalously low G&#x0002B;C content of 39.9% that may account for the presence of reverse gyrase, usually associated with hyperthermophiles.</p>
</abstract>
<kwd-group>
<kwd><italic>Dictyoglomus turgidum</italic></kwd>
<kwd>thermophile</kwd>
<kwd>biomass degradation</kwd>
<kwd>phage</kwd>
<kwd><italic>Dictyoglomi</italic></kwd>
<kwd>DNA polymerase</kwd>
<kwd>glucanase</kwd>
<kwd>reverse gyrase</kwd>
</kwd-group>
<contract-num rid="cn001">DE-FC02-07ER64494</contract-num>
<contract-num rid="cn001">DE-AC05-76RL01830</contract-num>
<contract-sponsor id="cn001">U.S. Department of Energy<named-content content-type="fundref-id">10.13039/100000015</named-content></contract-sponsor>
<counts>
<fig-count count="7"/>
<table-count count="6"/>
<equation-count count="0"/>
<ref-count count="76"/>
<page-count count="20"/>
<word-count count="14493"/>
</counts>
</article-meta>
</front>
<body>
<sec sec-type="intro" id="s1">
<title>Introduction</title>
<p><italic>Dictyoglomus</italic> species are genetically distinct and divergent from known taxa, and have been assigned to their own phylum, <italic>Dictyoglomi</italic> (Saiki et al., <xref ref-type="bibr" rid="B63">1985</xref>; Euz&#x000E9;by, <xref ref-type="bibr" rid="B17">2012</xref>). They have been cultivated from or detected in anaerobic, hyperthermophilic hot spring environments (Patel et al., <xref ref-type="bibr" rid="B57">1987</xref>; Svetlichny and Svetlichnaya, <xref ref-type="bibr" rid="B68">1988</xref>; Mathrani and Ahring, <xref ref-type="bibr" rid="B51">1991</xref>; Kublanov et al., <xref ref-type="bibr" rid="B45">2009</xref>; Gumerov et al., <xref ref-type="bibr" rid="B27">2011</xref>; Kochetkova et al., <xref ref-type="bibr" rid="B43">2011</xref>; Burgess et al., <xref ref-type="bibr" rid="B8">2012</xref>; Sahm et al., <xref ref-type="bibr" rid="B62">2013</xref>; Coil et al., <xref ref-type="bibr" rid="B11">2014</xref>; Menzel et al., <xref ref-type="bibr" rid="B53">2015</xref>) or isolated from paper-pulp factory effluent (Mathrani and Ahring, <xref ref-type="bibr" rid="B52">1992</xref>), but only two <italic>Dictyoglomus</italic> species have been validly described in the literature (Saiki et al., <xref ref-type="bibr" rid="B63">1985</xref>; Svetlichny and Svetlichnaya, <xref ref-type="bibr" rid="B68">1988</xref>). Both strains grow up to 80&#x000B0;C, are Gram negative, and exhibit unusual morphologies consisting of filaments, bundles, and spherical bodies. The first described <italic>Dictyoglomus</italic> species, <italic>Dictyoglomus thermophilum</italic> was isolated from Tsuetate Hot Spring in Kumamoto Prefecture, Japan (Saiki et al., <xref ref-type="bibr" rid="B63">1985</xref>). The genome of <italic>D. thermophilum</italic> has been sequenced (Coil et al., <xref ref-type="bibr" rid="B11">2014</xref>), and a number of potentially useful enzymes including amylase (Fukusumi et al., <xref ref-type="bibr" rid="B20">1988</xref>; Horinouchi et al., <xref ref-type="bibr" rid="B30">1988</xref>), xylanases (Gibbs et al., <xref ref-type="bibr" rid="B22">1995</xref>; Morris et al., <xref ref-type="bibr" rid="B54">1998</xref>), a mannanase (Gibbs et al., <xref ref-type="bibr" rid="B23">1999</xref>) and an endoglucanase (Shi et al., <xref ref-type="bibr" rid="B66">2013</xref>) have been cloned and characterized. The second described species, <italic>Dictyoglomus turgidus</italic>, was isolated from a hot spring in the Uzon Caldera, in eastern Kamchatka, Russia (Svetlichny and Svetlichnaya, <xref ref-type="bibr" rid="B68">1988</xref>). The name <italic>Dictyoglomus turgidus</italic> was subsequently corrected to <italic>Dictyoglomus turgidum</italic> (Euz&#x000E9;by, <xref ref-type="bibr" rid="B18">1998</xref>). Unlike <italic>D. thermophilum, D. turgidum</italic> was reported to grow on a wide range of substrates including starch, cellulose, pectin, carboxymethylcellulose, lignin, and humic acids, but not on pentose sugars such as xylose and arabinose (Svetlichny and Svetlichnaya, <xref ref-type="bibr" rid="B68">1988</xref>). Because of the wide range of substrates utilized, <italic>D. turgidum</italic> was selected for enzyme library construction and carbohydrase screening (Brumm et al., <xref ref-type="bibr" rid="B6">2011</xref>) as well as whole genome sequencing. Here we describe the complete genome sequence of <italic>D. turgidum</italic>, bioinformatic analysis of the metabolism of this unusual organism, and comparative analysis with the genome of <italic>D. thermophilum</italic>. We also present functional analysis of its DNA Pol I gene and a number of novel carbohydrases.</p>
</sec>
<sec sec-type="materials and methods" id="s2">
<title>Materials and methods</title>
<p><italic>D. turgidum</italic> strain 6724<sup>T</sup> was obtained from the Deutsche Sammlung von Mikroorganismen und Zellkulturen GmbH (DSMZ). 10G electrocompetent <italic>E. coli</italic> cells, pEZSeq (a lac promoter vector), Taq DNA polymerase and OmniAmp DNA polymerase were obtained from Lucigen, Middleton, WI. Azurine cross-linked-labeled polysaccharides were obtained from Megazyme International (Wicklow, Ireland). 4-methylumbelliferyl-&#x003B2;-D-cellobioside (MUC), 4-methylumbelliferyl-&#x003B2;-D -xylopyranoside (MUX), and 4-methylumbelliferyl-&#x003B2;-D- glucoyranoside (MUG) were obtained from Research Products International Corp. (Mt. Prospect, IL). CelLytic IIB reagent, pNP-&#x003B2;-glucoside, pNP-&#x003B2;-cellobioside, 4-methylumbelliferyl-&#x003B1;-D-arabinofuranoside (MUA), 4-methylumbelliferyl-&#x003B2;-D-lactopyranoside (MUL), 5-Bromo-4-chloro-3-indolyl &#x003B1;-D-galactopyranoside (X-&#x003B1;-Gal, XAG), and 5-Bromo-4-chloro-3-indolyl &#x003B2;-D-galactopyranoside (X-gal, XG) were purchased from Sigma-Aldrich (St. Louis, MO). All other chemicals were of analytical grade.</p>
<p><italic>D. turgidum</italic> DSM 6724&#x02122; was obtained from the DSMZ culture collection and maintained on DSM Medium 516 reduced with Na<sub>2</sub>S and N<sub>2</sub> at 75&#x000B0;C in Balch tubes with a headspace of N<sub>2</sub>. Cultures grown in 1 L stoppered flasks were harvested for DNA preparation. YT plate media (16 g/l tryptone, 10 g/l yeast extract, 5 g/l NaCl and 16 g/l agar) was used in all molecular biology screening experiments. Terrific Broth (12 g/l tryptone, 24 g/l yeast extract, 9.4 g/l K<sub>2</sub>HPO<sub>4</sub>, 2.2 g/l KH<sub>2</sub>PO<sub>4</sub>, and 4.0 g/l glycerol added after autoclaving) was used for liquid cultures.</p>
<p>A cell concentrate of <italic>D. turgidum</italic> strain 6724&#x02122; was lysed using a combination of SDS and proteinase (Sambrook et al., <xref ref-type="bibr" rid="B64">1989</xref>) and genomic DNA was purified using phenol/chloroform extraction. The genomic DNA was precipitated, treated with RNase to remove residual contaminating RNA, and fragmented by hydrodynamic shearing (HydroShear apparatus, GeneMachines, San Carlos, CA) to generate fragments of 2&#x02013;4 kb. The fragments were purified on an agarose gel, end-repaired, and ligated into pEZSeq (Lucigen Corp., Middleton WI). The recombinant plasmids were then used to transform electrocompetent cells. A copy of the library containing the <italic>Dictyoglomus turgidum</italic> genomic DNA was submitted to the Joint Genome Institute of the Department of Energy for whole genome sequencing; a second copy of the library was used for carbohydrase screening experiments.</p>
<p>The genome of <italic>D. turgidum</italic> DSM 6724&#x02122; was sequenced at the Joint Genome Institute (JGI) using a combination of 3 and 8 kb DNA libraries. In addition to 20x Sanger sequencing, 454 pyrosequencing was done to a depth of 20x coverage. Draft assemblies were based on 32,817 total reads. The Phred/Phrap/Consed software package was used for sequence assembly and quality assessment (Ewing and Green, <xref ref-type="bibr" rid="B19">1998</xref>; Gordon et al., <xref ref-type="bibr" rid="B24">1998</xref>). After the shotgun stage, reads were assembled with parallel phrap. Possible mis-assemblies were corrected with Dupfinisher or transposon bombing of bridging clones. Gaps between contigs were closed by editing in Consed, custom primer walking or PCR amplification. A total of 80 additional reactions were necessary to close gaps and to raise the quality of the finished sequence. The completed genome sequence of <italic>D. turgidum</italic> DSM 6724&#x02122; contains 34,756 reads, achieving an average of 17.3x coverage. The Accession number for the complete genome is <ext-link ext-link-type="DDBJ/EMBL/GenBank" xlink:href="NC_011661">NC_011661</ext-link>.</p>
<p>Genes were identified using Prodigal (Hyatt et al., <xref ref-type="bibr" rid="B34">2010</xref>) as part of the Oak Ridge National Laboratory genome annotation pipeline, followed by a round of manual curation using the JGI GenePRIMP pipeline. The predicted CDSs were translated and used to search the National Center for Biotechnology Information (NCBI) nonredundant database, UniProt, TIGRFam, Pfam, PRIAM, KEGG, COG, and InterPro databases. These data sources were combined to assert a product description for each predicted protein. Non-coding genes and miscellaneous features were predicted using tRNAscan-SE (Lowe and Eddy, <xref ref-type="bibr" rid="B50">1997</xref>), RNAMMer (Lagesen et al., <xref ref-type="bibr" rid="B47">2007</xref>), Rfam (Griffiths-Jones et al., <xref ref-type="bibr" rid="B25">2003</xref>), TMHMM (Krogh et al., <xref ref-type="bibr" rid="B44">2001</xref>), CRISPRFinder (Grissa et al., <xref ref-type="bibr" rid="B26">2007</xref>), and signalP (Krogh et al., <xref ref-type="bibr" rid="B44">2001</xref>). RAST annotations (Aziz et al., <xref ref-type="bibr" rid="B2">2008</xref>) of <italic>D. turgidum</italic> and <italic>D. thermophilum</italic> were carried out in parallel to further clarify genomic relationships using SEED genome comparison tools (Overbeek et al., <xref ref-type="bibr" rid="B56">2005</xref>).</p>
<p>The phylogeny of <italic>D. turgidum</italic> was determined using its 16S ribosomal RNA (rRNA) gene sequence as well as those of the most closely related 16S rRNA sequences identified by BLASTn. 16S rRNA gene sequences were aligned using MUSCLE (Edgar, <xref ref-type="bibr" rid="B16">2004</xref>), pairwise distances were estimated using the maximum composite likelihood (MCL) approach, and initial trees for heuristic search were obtained automatically by applying the neighbor-joining method in MEGA7 (Kumar et al., <xref ref-type="bibr" rid="B46">2016</xref>). The alignment and heuristic trees were then used to infer the phylogeny using the maximum likelihood method based on Tamura-Nei (Tamura and Nei, <xref ref-type="bibr" rid="B70">1993</xref>; Tamura et al., <xref ref-type="bibr" rid="B71">2011</xref>). The phylogeny of the reverse gyrase protein sequence was inferred using the Neighbor-Joining method. The optimal tree with the sum of branch length &#x0003D; 1.99686421 is shown. The percentage of replicate trees in which the associated taxa clustered together in the bootstrap test (1000 replicates) are shown next to the branches. The tree is drawn to scale, with branch lengths in the same units as those of the evolutionary distances used to infer the phylogenetic tree. The evolutionary distances were computed using the Maximum Composite Likelihood method and are in the units of the number of base substitutions per site. The analysis involved 7 nucleotide sequences. Codon positions included were 1st&#x0002B;2nd&#x0002B;3rd&#x0002B;Noncoding. All positions containing gaps and missing data were eliminated. There were a total of 3230 positions in the final dataset. Evolutionary analyses were conducted in MEGA7 (Kumar et al., <xref ref-type="bibr" rid="B46">2016</xref>).</p>
<p>The <italic>endo</italic>-glucanase specificity of enzymes was determined in 0.50 ml of 50 mM acetate buffer, pH 5.8, containing 0.2% azurine cross-linked-labeled (AZCL) insoluble substrates and 50 &#x003BC;l of clarified lysate. Each purified enzyme was evaluated for <italic>endo</italic>-activities using the following set of substrates: AZCL-arabinan (AR), AZCL-arabinoxylan (AX), AZCL-&#x003B2;-glucan (BG), AZCL-curdlan (CU), AZCL-galactan (GL), AZCL-galactomannan (GM), AZCL-hydroxyethyl cellulose (HEC), AZCL-pullulan (PUL), AZCL-rhamnogalacturonan (RH), and AZCL-xyloglucan (XG). Assays were performed at 70&#x000B0;C, with shaking at 1000 rpm, for 60 min in a Thermomixer R (Eppendorf, Hamburg, Germany). Tubes were clarified by centrifugation and absorbance values at 600 nm determined using a Bio-Tek EL<sub>x</sub>800 plate reader. The exo-glucanase specificity of enzymes was determined by spotting 2.0 &#x003BC;l of clarified lysate directly on agar plates containing 10 mM 4-methylumbelliferyl substrate. Plates were placed in a 70&#x000B0;C incubator for 60 min and then examined using a hand-held UV lamp and compared to negative and positive controls for fluorescence.</p>
<p>Amplification efficacy was compared between Dtur, Taq and OmniAmp DNA polymerases (DNAP) in side by side PCR reactions using four different sized amplicons (0.9, 2.8, 5.0, and 10.0 Kb). PCR reaction conditions contained 1&#x02013;20 ng of template DNA, 2.5U of Taq DNAP or 5U Dtur or OmniAmp DNAP (Lucigen Corp.), 200 &#x003BC;M dNTPs, and 0.5 &#x003BC;M primers in a 50 &#x003BC;l reaction. DNAP buffer (1X) contained10 mM Tris-HCl (pH 8.8), 10 mM KCl, 10 mM NH2SO4, 2 mM MgSO4, 0.1% tritonX-100, and 15% sucrose. Cycling conditions were 94&#x000B0;C 2 min and 30 cycles of 94&#x000B0;C for 15 s, 60&#x000B0;C for 30 s, and 72&#x000B0;C for 1 min per kb. The templates and PCR primers are as follows: pUC19 0.9 kb amplicon primers (CCC CTA TTT GTT TAT TTT TCT AAA ATT CAA TAT GTA TCC GCT and TTA CCA ATG CTT AAT CAG TGA GGC ACC TAT CT), <italic>E. coli</italic> 2.8 kb amplicon primers (TAC TGT CTG CCA TGG TTC AGA TCC CCC AAA ATC CAC TTA TCC TTG TAG A and TTA TCT GTG GTC GAC TTA GTG CGC CTG ATC CCA GTT TTC GCC ACT CCC CA), <italic>E. coli</italic> 5 kb amplicon primers (TCT CTC CGA CCA AAG AGT TG and GAA ACA TTG AGC GAA GAG GA), and <italic>E. coli</italic> 10 kb amplicon primers (CTA TGA TTA TCT AGG CTT AGG GTC AC and CAG TGT AGA GAG ATA GTC AGG AGT TA).</p>
<p>Functional screening for active carbohydrase enzymes involved plating transformed <italic>E. coli</italic> cells containing 2&#x02013;4 kb Dtur genomic DNA inserts in the pEZSeq vector on YT agar containing IPTG (for <italic>lac</italic>Z promoter induction) and one of the fluorescent substrates MUC, MUG or MUX. A long wavelength UV lamp was used to locate colonies that were fluorescent, which were sequenced by Sanger chemistry to identify the gene. Genes identified in the functional screen as well as additional genes of interest from the completed genome were amplified without their respective signal sequence, ligated into pET28A, and transformed into BL21(DE3) <italic>E. coli</italic> competent cells. Recombinant clones were cultured overnight at 37&#x000B0;C, 100 rpm, in 100 ml Luria Broth containing 50 mg/l kanamycin. Expression was induced using 1 mM IPTG, and cultures were harvested 18 h after induction. Cells were pelleted by centrifugation, and the pellets were lysed using Cellytic B reagent. Proteins were purified using standard methods for His-tagged proteins (Spriestersbach et al., <xref ref-type="bibr" rid="B67">2015</xref>), and their purity and identity verified by SDS PAGE.</p>
<p><italic>D. turdigum</italic> DNA polymerase I (Dtur DNAP) was cloned by PCR amplification using the proofreading enzyme Phusion (NEB, Waltham MA) and forward and reverse 24 base oligonucleotides that spanned the start and stop codons. The amplified DNA was inserted into the rhamnose promoter vector pRham containing an N terminal histidine tag and transformed into 10G competent <italic>E. coli</italic> cells (Lucigen Corp.). Recombinant Dtur DNAP production was induced by rhamnose and the enzyme was purified using standard methods for His-tagged proteins (Spriestersbach et al., <xref ref-type="bibr" rid="B67">2015</xref>).</p>
</sec>
<sec sec-type="results" id="s3">
<title>Results</title>
<sec>
<title>Genome of <italic>D. turgidum</italic></title>
<p>The genome of <italic>D. turgidum</italic> DSM 6724&#x02122; consists of a single chromosome of 1,855,560 bp and no plasmids or extrachromosomal elements. The GC content of the chromosome is 33.96% based on the genome sequence, slightly higher than the reported value of 32.5% (Svetlichny and Svetlichnaya, <xref ref-type="bibr" rid="B68">1988</xref>) and is predicted to contain 1813 protein-coding genes and 52 RNA genes (Figure <xref ref-type="fig" rid="F1">1</xref>). The completed genome sequence is available from GenBank (GenBank: <ext-link ext-link-type="DDBJ/EMBL/GenBank" xlink:href="CP001251.1">CP001251.1</ext-link>). Based on 16S rRNA gene sequence analysis, <italic>D. turgidum</italic> DSM 6724 and <italic>D. thermophilum</italic> are separate species. This is confirmed by average nucleotide analysis (ANI), where <italic>D. turgidum</italic> and <italic>D. thermophilum</italic> are calculated to have 82.4% average nucleotide identity, below the threshold for members of the same species.</p>
<fig id="F1" position="float">
<label>Figure 1</label>
<caption><p><bold>Genome map of <italic><bold>D. turgidum</bold></italic></bold>. From outside to the center: genes on forward strand (color by COG categories); genes on reverse strand (color by COG categories); RNA genes (tRNAs green, rRNAs red, other RNAs black); GC content; GC skew.</p></caption>
<graphic xlink:href="fmicb-07-01979-g0001.tif"/>
</fig>
<p>Of the 1813 protein-coding genes, 1354 genes (72.6%) were assigned to COGs categories (Table <xref ref-type="table" rid="T1">1</xref>). The fraction of the genes annotated as members of COG class G, carbohydrate transport and metabolism (highlighted in bold), 13.4%, is greater than the fraction observed for 95% of genomes in the MicrobesOnline database (Dehal et al., <xref ref-type="bibr" rid="B12">2010</xref>). This represents the lower limit of proteins involved in carbohydrate metabolism, because it does not include any proteins in categories R, S or not in COGS that were not identified by the algorithm as being involved in carbohydrate metabolism. A number of pectate lyases, for example, are not identified as members of COGs class G. No other COGs category had a significantly higher than average number of members, and no COGs category had a significantly lower than average percentage of members.</p>
<table-wrap position="float" id="T1">
<label>Table 1</label>
<caption><p><bold>Number of genes associated with general COG functional categories</bold>.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th valign="top" align="left"><bold>Code</bold></th>
<th valign="top" align="center"><bold>Value</bold></th>
<th valign="top" align="center"><bold>Percentage</bold></th>
<th valign="top" align="left"><bold>Description</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">J</td>
<td valign="top" align="center">168</td>
<td valign="top" align="center">11.0%</td>
<td valign="top" align="left">Translation, ribosomal structure and biogenesis</td>
</tr>
<tr>
<td valign="top" align="left">K</td>
<td valign="top" align="center">76</td>
<td valign="top" align="center">5.0%</td>
<td valign="top" align="left">Transcription</td>
</tr>
<tr>
<td valign="top" align="left">L</td>
<td valign="top" align="center">61</td>
<td valign="top" align="center">4.0%</td>
<td valign="top" align="left">Replication, recombination and repair</td>
</tr>
<tr>
<td valign="top" align="left">B</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">0.1%</td>
<td valign="top" align="left">Chromatin structure and dynamics</td>
</tr>
<tr>
<td valign="top" align="left">D</td>
<td valign="top" align="center">19</td>
<td valign="top" align="center">1.2%</td>
<td valign="top" align="left">Cell cycle control, Cell division, chromosome partitioning</td>
</tr>
<tr>
<td valign="top" align="left">V</td>
<td valign="top" align="center">40</td>
<td valign="top" align="center">2.6%</td>
<td valign="top" align="left">Defense mechanisms</td>
</tr>
<tr>
<td valign="top" align="left">T</td>
<td valign="top" align="center">48</td>
<td valign="top" align="center">3.1%</td>
<td valign="top" align="left">Signal transduction mechanisms</td>
</tr>
<tr>
<td valign="top" align="left">M</td>
<td valign="top" align="center">87</td>
<td valign="top" align="center">5.7%</td>
<td valign="top" align="left">Cell wall/membrane biogenesis</td>
</tr>
<tr>
<td valign="top" align="left">N</td>
<td valign="top" align="center">20</td>
<td valign="top" align="center">1.3%</td>
<td valign="top" align="left">Cell motility</td>
</tr>
<tr>
<td valign="top" align="left">U</td>
<td valign="top" align="center">18</td>
<td valign="top" align="center">1.2%</td>
<td valign="top" align="left">Intracellular trafficking and secretion</td>
</tr>
<tr>
<td valign="top" align="left">O</td>
<td valign="top" align="center">61</td>
<td valign="top" align="center">4.0%</td>
<td valign="top" align="left">Posttranslational modification, protein turnover, chaperones</td>
</tr>
<tr>
<td valign="top" align="left">C</td>
<td valign="top" align="center">79</td>
<td valign="top" align="center">5.2%</td>
<td valign="top" align="left">Energy production and conversion</td>
</tr>
<tr>
<td valign="top" align="left"><bold>G</bold></td>
<td valign="top" align="center"><bold>205</bold></td>
<td valign="top" align="center"><bold>13.4%</bold></td>
<td valign="top" align="left"><bold>Carbohydrate transport and metabolism</bold></td>
</tr>
<tr>
<td valign="top" align="left">E</td>
<td valign="top" align="center">170</td>
<td valign="top" align="center">11.1%</td>
<td valign="top" align="left">Amino acid transport and metabolism</td>
</tr>
<tr>
<td valign="top" align="left">F</td>
<td valign="top" align="center">60</td>
<td valign="top" align="center">3.9%</td>
<td valign="top" align="left">Nucleotide transport and metabolism</td>
</tr>
<tr>
<td valign="top" align="left">H</td>
<td valign="top" align="center">73</td>
<td valign="top" align="center">4.8%</td>
<td valign="top" align="left">Coenzyme transport and metabolism</td>
</tr>
<tr>
<td valign="top" align="left">I</td>
<td valign="top" align="center">44</td>
<td valign="top" align="center">2.9%</td>
<td valign="top" align="left">Lipid transport and metabolism</td>
</tr>
<tr>
<td valign="top" align="left">P</td>
<td valign="top" align="center">77</td>
<td valign="top" align="center">5.0%</td>
<td valign="top" align="left">Inorganic ion transport and metabolism</td>
</tr>
<tr>
<td valign="top" align="left">Q</td>
<td valign="top" align="center">18</td>
<td valign="top" align="center">1.2%</td>
<td valign="top" align="left">Secondary metabolites biosynthesis, transport and catabolism</td>
</tr>
<tr>
<td valign="top" align="left">R</td>
<td valign="top" align="center">130</td>
<td valign="top" align="center">8.5%</td>
<td valign="top" align="left">General function prediction only</td>
</tr>
<tr>
<td valign="top" align="left">S</td>
<td valign="top" align="center">58</td>
<td valign="top" align="center">3.8%</td>
<td valign="top" align="left">Function unknown</td>
</tr>
<tr>
<td valign="top" align="left">&#x02013;</td>
<td valign="top" align="center">511</td>
<td valign="top" align="center">27.4%</td>
<td valign="top" align="left">Not in COGs</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<p><italic>Highlighted in bold, COG class G. The fraction of the genes annotated as members of this class is greater than the fraction observed for 95% of genomes in the MicrobesOnline database</italic>.</p>
</table-wrap-foot>
</table-wrap>
</sec>
<sec>
<title>Genomic insights into the relationship of <italic>D. turgidum</italic> to <italic>D. thermophilum</italic> and other organisms</title>
<p>While being separate species, an in-depth comparison of the two <italic>Dictyoglomi</italic> genomes shows that <italic>D. turgidum</italic> is closely related to <italic>D. thermophilum</italic> on a number of levels. The genomes are similar in size, with <italic>D. turgidum</italic> being slightly smaller than the genome of <italic>D. thermophilum</italic> (1,855,560 bp vs. 1,959,987 bp) and containing approximately 100 fewer protein coding genes (1813 vs. 1912). The two organisms have a highly conserved set of genes present in their genomes. Over 95% of the proteins present in <italic>D. turgidum</italic> have orthologs in <italic>D. thermophilum</italic>. There are only 43 proteins of greater than 100 amino acids present in <italic>D. turgidum</italic> without orthologs in <italic>D. thermophilum</italic>, and there are only 109 proteins of greater than 100 amino acids present in <italic>D. thermophilum</italic> without orthologs in <italic>D. turgidum</italic>. Of the proteins with orthologs in both species, there are 614 proteins with &#x0003E;90% sequence identity.</p>
<p>Synteny plots were generated using both RAST and IMG annotation methods. The two annotation methods gave essentially identical plots, as did plots based on DNA or protein sequences. The plots show the genomes of <italic>D. turgidum</italic> and <italic>D. thermophilum</italic> have highly conserved large and small-scale organization (Figure <xref ref-type="fig" rid="F2">2A</xref>). This conserved organization appears to be an unusual phenomenon. Two sets of thermophilic organisms with similar ANI values, <italic>T. thermophilus</italic> and <italic>T. aquaticus</italic> (84.3% ANI, Figure <xref ref-type="fig" rid="F2">2B</xref>) and <italic>C. bescii</italic> and <italic>C. saccharolyticus</italic> (82.0% ANI, Figure <xref ref-type="fig" rid="F2">2C</xref>) show only limited short-range synteny and no extensive long-range synteny. It is unclear if this conserved genomic organization is limited to these two species, or is present in all <italic>Dictyoglomi</italic> genomes.</p>
<fig id="F2" position="float">
<label>Figure 2</label>
<caption><p><bold>Synteny plot of selected genomes</bold>. MUMmer (Delcher et al., <xref ref-type="bibr" rid="B13">2003</xref>) was used to generate the dotplot diagram between sets of two genomes. The six frame amino acid translation of the DNA input sequences were used for comparing genomes using PROmer software. Clockwise from top <bold>(A)</bold> genomes of <italic>D. turgidum</italic> and <italic>D. thermophilum</italic>; <bold>(B)</bold> genomes of <italic>T. thermophilus</italic> and <italic>T. aquaticus</italic>; <bold>(C)</bold> genomes of <italic>C. bescii</italic> and <italic>C. saccharolyticus</italic>. </p></caption>
<graphic xlink:href="fmicb-07-01979-g0002.tif"/>
</fig>
<p>The relationship of these two <italic>Dictyoglomus</italic> species to other organisms appears significantly more complicated, depending on the type of analysis and interpretation (Love et al., <xref ref-type="bibr" rid="B49">1993</xref>; Rees et al., <xref ref-type="bibr" rid="B60">1997</xref>; Takai et al., <xref ref-type="bibr" rid="B69">1999</xref>; Ding et al., <xref ref-type="bibr" rid="B15">2000</xref>; Wagner and Wiegel, <xref ref-type="bibr" rid="B74">2008</xref>). Phylogenetic analysis using 16S rRNA shows the two <italic>Dictyoglomus</italic> species appear most closely related to <italic>Thermotoga</italic> species before bootstrapping (data not shown). After bootstrapping, the relationship shifts dramatically, with the two <italic>Dictyoglomus</italic> species becoming most closely related to <italic>Caldicellulosiruptor</italic> species (Figure <xref ref-type="fig" rid="F3">3</xref>). Previous work using average nucleotide identity (ANI) calculations (Nishida et al., <xref ref-type="bibr" rid="B55">2011</xref>) identified <italic>Thermotoga</italic> species as the closest relatives to <italic>Dictyoglomus</italic>. ANI values were generated using the <italic>D. thermophilum</italic> genome, eight finished, closed <italic>Thermotoga</italic> genomes and three finished, closed <italic>Caldicellulosiruptor</italic> genomes. ANI values (Kim et al., <xref ref-type="bibr" rid="B41">2014</xref>) were computed as pairwise bidirectional best nSimScan hits of genes having 70% or more identity and at least 70% coverage of the shorter gene. ANI calculations performed as described above yielded 82.4% identity between the genomes of <italic>D. turgidum</italic> and <italic>D. thermophilum</italic>, based on 1584 proteins (87% of the genome) that met the criteria. The value of 82.4% is well below the cut-off value of 98% for strains of the same species, and confirms that <italic>D. turgidum</italic> and <italic>D. thermophilum</italic> are separate species. The ANI calculations found 67&#x02013;68% identity between <italic>D. turgidum</italic> and the three <italic>Caldicellulosiruptor</italic> species, based on 124&#x02013;129 proteins per genome that met the criteria for the calculation (approximately 7% of the genome). ANI calculations found 66&#x02013;68% identity between <italic>D. turgidum</italic> and the eight <italic>Thermotoga</italic> species, based on the 36&#x02013;64 proteins per genome that met the criteria (approximately 2&#x02013;4% of the genome). Rather than identifying relationships among these organisms, the low number of proteins in <italic>D. turgidum</italic> with at least 70% identity to the proteins in these 11 strains (on which these ANI values are calculated) further demonstrates the uniqueness of this organism.</p>
<fig id="F3" position="float">
<label>Figure 3</label>
<caption><p><bold>Molecular phylogenetic analysis of <italic><bold>Dictyoglomus turgidum</bold></italic> using 16S rDNA sequences</bold>. Molecular phylogenetic analysis by Maximum Likelihood method was detailed in the Material and Methods Section. The bootstrap consensus tree inferred from 550 replicates [2] is taken to represent the evolutionary history of the taxa analyzed. Branches corresponding to partitions reproduced in less than 50% bootstrap replicates are collapsed. The percentage of replicate trees in which the associated taxa clustered together in the bootstrap test (550 replicates) are shown next to the branches. Sequences used for the analysis are: <italic>Dictyoglomus turgidum</italic> strain DSM 6724; NR_074885; <italic>Dictyoglomus thermophilum</italic> strain H-6-12, NR_029235.1; <italic>Fervidicola ferrireducens</italic> strain Y170, NR_044504.1; <italic>Thermosediminibacter oceani</italic> strain DSM 16646; NR_074461.1; <italic>Caldicellulosiruptor saccharolyticus</italic> strain DSM 8903; NR_074845.1; <italic>Caldicellulosiruptor hydrothermalis</italic> strain 108, NR_074767.1; <italic>Caldicellulosiruptor bescii</italic> strain DSM 6725; NR_074788.1; <italic>Desulfotomaculum kuznetsovii</italic> strain DSM 6115; NR_075068.1; <italic>Thermovirga lienii</italic> strain DSM 17291; NR_074606.1; <italic>Thermotoga petrophila</italic> strain RKU-10, NR_042374.1; <italic>Thermotoga naphthophila</italic> strain RKU-10, NR_112092.1; <italic>Thermotoga maritima</italic> strain MSB-8, NR_029163.1; and <italic>Geobacillus thermoglucosidasius</italic> strain ATCC 43742; NR_112058.1.</p></caption>
<graphic xlink:href="fmicb-07-01979-g0003.tif"/>
</fig>
</sec>
<sec>
<title>Protein and amino acid metabolism</title>
<p>Based on the MEROPS database (Rawlings et al., <xref ref-type="bibr" rid="B59">2014</xref>), the <italic>D. turgidum</italic> genome codes for 55 potential peptidases. This value is within the range of peptidases reported in the database for <italic>Thermotoga</italic> species (52&#x02013;67) and <italic>Caldicellulosiruptor</italic> species (54&#x02013;74). Of the 55 potential peptidases, only a single peptidase, Dtur_0603, possesses an annotated signal sequence and is predicted to be secreted. While possessing only a single secreted peptidase to generate amino acids and peptides, <italic>D. turgidum</italic> possesses nine potential membrane transporter systems to transport amino acids and peptides into the cell. These nine transporters include seven annotated oligopeptide/dipeptide ABC transporter systems (Dtur_0082 through Dtur_0086; Dtur_0158 through Dtur_0162; Dtur_0214 through Dtur_0217; Dtur_0664 through Dtur_0668; Dtur_1061 through Dtur_1064; Dtur_1704 and Dtur_1707; Dtur_1719 through Dtur_1722) as well as two amino acid ABC transporter systems (Dtur_1051 through Dtur_1053 and Dtur_0932 through Dtur_0936).</p>
<p><italic>D. turgidum</italic> appears to utilize the amino acids and peptides taken up for protein synthesis, but it is unable to metabolize most amino acids as an energy or carbon source. Based on the BioCyc (Karp et al., <xref ref-type="bibr" rid="B37">2005</xref>; Caspi et al., <xref ref-type="bibr" rid="B9">2014</xref>) and SEED (Devoid et al., <xref ref-type="bibr" rid="B14">2013</xref>) metabolic reconstructions from the genome sequence, <italic>D. turgidum</italic> is lacking degradation pathways for the following 13 amino acids: aspartate, asparginine, cysteine, histidine, isoleucine, leucine, lysine, phenylalanine, proline, serine, tryptophan, tyrosine, and valine. Arginine is not metabolized, but may be converted to putrescine.</p>
<p>Only four amino acids appear to be metabolized by <italic>D. turgidum</italic>. Glutamate is converted to methyl aspartate using glutamate mutase (Dtur_1345 through Dtur_1347) and then to pyruvate and acetate. Threonine can be degraded to glycine and acetaldehyde via threonine aldolase (Dtur_0449), and the acetaldehyde generated is then converted to acetyl-CoenzymeA (acetyl-CoA) via aldehyde dehydrogenase (Dtur_0484). Alanine can be converted to pyruvate by alanine dehydrogenase (Dtur_1049), and glycine can be converted to ammonium 5,10-methylenetetrahydrofolate via glycine dehydrogenase and glycine cleavage system T protein (Dtur_1515 through Dtur_1518). The ability to utilize these four amino acids may be responsible for the observation of growth by <italic>D. turgidum</italic> on yeast extract, peptone, and casamino acids (Svetlichny and Svetlichnaya, <xref ref-type="bibr" rid="B68">1988</xref>).</p>
</sec>
<sec>
<title>Monosaccharide metabolism</title>
<p>Based on the genomic reconstruction of Dtur, the organism is able to metabolize most five and six carbon sugars, and the following pathways are predicted. Arabinose is utilized via isomerization to L-ribulose (Dtur_0379, or other isomerase), phosphorylation by L-ribulose kinase (Dtur_1748) to L-ribulose-5-phosphate, and isomerization by L-ribulose-5-phosphate-4-epimerase (Dtur_1734) to D-xylulose-5-phosphate, which is then metabolized via the pentose phosphate pathway. Rhamnose is utilized via isomerization by L-rhamnose isomerase to L-rhamulose (Dtur_0427), phosphorylation by L-rhamulose kinase (Dtur_1748) to L-rhamulose-1-phosphate, and cleavage into dhihydroxyacetone phosphate and L-lactaldehyde. Xylose is utilized via isomerization by xylose isomerase (Dtur_0036 or other sugar isomerase) to xylulose, and the xylulose is phosphorylated by xylulose kinase to (Dtur_0920) to D-xylulose-5-phosphate, which is then metabolized via the pentose phosphate pathway.</p>
<p>Fucose is utilized via isomerization by L-fucose isomerase to L-fuculose (Dtur_0410), phosphorylation by L-fuculokinase (Dtur_0920) to L-fuculose-1-phosphate, and cleavage into dhihydroxyacetone phosphate and L-lactaldehyde. Galactose is phosphorylated by galactose kinase (Dtur_1195) to galactose-1-phosphate, which is converted to UDP-galactose by galactose-1-phosphate uridyl transferase (Dtur_1196), isomerized by UDP-glucose-4-epimerase (Dtur_1352) to UDP-glucose, and finally to glucose-1-phosphate by UTP-glucose-1-phosphate uridylyltransferase (Dtur_1627). Mannose is phosphorylated by mannose kinase (Dtur_0176; Dtur_0716 or other annotated sugar kinase) to generate mannose-1-phosphate. The mannose-1-phosphate is isomerized to mannose-6-phosphate by phosphomannomutase/phosphoglucomutase (Dtur_0067) and then to fructose-6-phosphate by phosphoglucose/phosphomannose isomerase (Dtur_1271). UDP-glucose is either isomerized to fructose, or oxidized to UDP-glucuronate using either Dtur_575 or Dtur_718. The UDP-glucuronate can then be further oxidized to ribulose-5-phosphate by 6-phosphogluconate dehydrogenase (Dtur_0197).</p>
<p>Galacturonate generated by pectin degradation may be epimerized by one of the six UDP sugar epimerase genes found in the genome. Rarely-encountered sugars may be handled by any of a number of sugar isomerases. Dtur rhamnose isomerase (Dtur_0427) isomerizes seven monosaccharides: L-rhamnose, L-lyxose, L-mannose, L-xylulose, L-fructose, D-allose, and D-ribose (Kim et al., <xref ref-type="bibr" rid="B42">2013</xref>). The Dtur fucose isomerase (Dtur_0410) isomerizes L-fucose, D-arabinose, D-altrose, and L-galactose (Hong et al., <xref ref-type="bibr" rid="B29">2012</xref>). Dtur also possesses a cellobiose 2-epimerase that may isomerize non-metabolized disaccharides into easily-degradable ones (Kim et al., <xref ref-type="bibr" rid="B40">2012</xref>).</p>
</sec>
<sec>
<title>Polysaccharide degradation and transport</title>
<p>Polysaccharide degradation by <italic>D. turgidum</italic> is of interest for a number of reasons. Analysis of the <italic>D. turgidum</italic> genome shows an enrichment in COGS family members annotated as involved in carbohydrate transport and metabolism (Table <xref ref-type="table" rid="T1">1</xref>). <italic>D. turgidum</italic> is reported to utilize polysaccharides such as starch, cellulose, pectin, glycogen, and carboxymethyl cellulose (Svetlichny and Svetlichnaya, <xref ref-type="bibr" rid="B68">1988</xref>) while <italic>D. thermophilum</italic> is reported to utilize starch, but not cellulose. Finally, a number of carbohydrates with potential industrial applications have been identified in the two <italic>Dictyoglomus</italic> species including amylases and xylanases. A combination of genomic and enzymatic analyses was carried out to clarify the polysaccharide degradation potential of <italic>D. turgidum</italic>.</p>
<p>Analysis of the <italic>D. turgidum</italic> genome reveals a wide range of genes coding for annotated extracellular and intracellular polysaccharide degrading enzymes. The CAZy database (Lombard et al., <xref ref-type="bibr" rid="B48">2014</xref>) identifies 57 glycosyl hydrolases (GH), 3 polysaccharide lyases (PL) and 6 carbohydrate esterases (CE) in the Dtur genome. Based on signal sequence predictions (Petersen et al., <xref ref-type="bibr" rid="B58">2011</xref>), 20 of the polysaccharide-degrading enzymes are secreted into the medium (Table <xref ref-type="table" rid="T2">2</xref>), where they degrade polysaccharides into oligosaccharides and monosaccharides. After polysaccharide degradation, 18 annotated three-component ABC carbohydrate transporters are predicted to transport monosaccharides and oligosaccharides into the cell. <italic>D. turgidum</italic> is reported to utilize fructose, glucose, rhamnose, inositol, mannitol, and sorbitol (Svetlichny and Svetlichnaya, <xref ref-type="bibr" rid="B68">1988</xref>), indicating ABC carbohydrate transporters exist for these monosaccharides and sugar alcohols. <italic>D. turgidum</italic> cannot utilize arabinose, fucose, galactose, mannose, or xylose, indicating a lack of dedicated transport systems for these monosaccharides. These sugars may be transported into the cell as oligosaccharides by the oligosaccharide transporters and degraded to monosaccharides in the cytoplasm. Once inside the cell, oligosaccharides are degraded into monosaccharides by a combination of 46 <italic>exo</italic>-acting and <italic>endo</italic>-acting enzymes (Table <xref ref-type="table" rid="T3">3</xref>). Working together, these 46 enzymes appear capable of degrading oligosaccharides from most plant-based polysaccharides to monosaccharides.</p>
<table-wrap position="float" id="T2">
<label>Table 2</label>
<caption><p><bold>Annotated secreted polysaccharide-degrading enzymes</bold>.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th valign="top" align="left"><bold>Gene</bold></th>
<th valign="top" align="left"><bold>GH family</bold></th>
<th valign="top" align="left"><bold>Annotated activity</bold></th>
<th valign="top" align="left"><bold>Nearest ortholog</bold></th>
<th valign="top" align="center"><bold>Identity</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Dtur_0097</td>
<td valign="top" align="left">GH 44</td>
<td valign="top" align="left">&#x003B2;-mannanase</td>
<td valign="top" align="left">Calkro_0851</td>
<td valign="top" align="center">70.1%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0172</td>
<td valign="top" align="left">GH 28</td>
<td valign="top" align="left">pectinase</td>
<td valign="top" align="left">Cphy_3310</td>
<td valign="top" align="center">47.1%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0243</td>
<td valign="top" align="left">GH 11</td>
<td valign="top" align="left">xylanase</td>
<td valign="top" align="left">Calkro_0081</td>
<td valign="top" align="center">83.7%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0276</td>
<td valign="top" align="left">GH 5</td>
<td valign="top" align="left">cellulase</td>
<td valign="top" align="left">Mahau_0466</td>
<td valign="top" align="center">59.9%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0277</td>
<td valign="top" align="left">GH 26</td>
<td valign="top" align="left">&#x003B2;-mannanase</td>
<td valign="top" align="left">BG52_11385</td>
<td valign="top" align="center">52.7%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0430</td>
<td valign="top" align="left">PL 1</td>
<td valign="top" align="left">pectate lyase</td>
<td valign="top" align="left">SNOD_03765</td>
<td valign="top" align="center">42.1%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0431</td>
<td valign="top" align="left">PL 1</td>
<td valign="top" align="left">pectate lyase</td>
<td valign="top" align="left">M769_0111315</td>
<td valign="top" align="center">60.1%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0432</td>
<td valign="top" align="left">PLNC</td>
<td valign="top" align="left">pectate lyase</td>
<td valign="top" align="left">CSE_02370</td>
<td valign="top" align="center">57.3%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0433</td>
<td valign="top" align="left">CE 8</td>
<td valign="top" align="left">pectin esterase</td>
<td valign="top" align="left">Calkro_0154</td>
<td valign="top" align="center">56.0%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0628</td>
<td valign="top" align="left">GH 12</td>
<td valign="top" align="left">curdlanase</td>
<td valign="top" align="left">CTN_1107</td>
<td valign="top" align="center">48.4%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0669</td>
<td valign="top" align="left">GH 5</td>
<td valign="top" align="left">cellulase</td>
<td valign="top" align="left">Mahau_0466</td>
<td valign="top" align="center">54.7%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0675</td>
<td valign="top" align="left">GH 57</td>
<td valign="top" align="left">&#x003B1;-amylase</td>
<td valign="top" align="left">ANT_11030</td>
<td valign="top" align="center">41.3%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0676</td>
<td valign="top" align="left">CBM9</td>
<td valign="top" align="left">&#x003B1;-amylase</td>
<td valign="top" align="left">COCOR_00322</td>
<td valign="top" align="center">39.7%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0857</td>
<td valign="top" align="left">GH 53</td>
<td valign="top" align="left">&#x003B2;-galactanase</td>
<td valign="top" align="left">TRQ7_08325</td>
<td valign="top" align="center">56.5%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1586</td>
<td valign="top" align="left">GH 5</td>
<td valign="top" align="left">cellulase</td>
<td valign="top" align="left">BSONL12_10711</td>
<td valign="top" align="center">41.5%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1675</td>
<td valign="top" align="left">GH 13</td>
<td valign="top" align="left">&#x003B1;-amylase</td>
<td valign="top" align="left">CAAU_0986</td>
<td valign="top" align="center">51.6%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1715</td>
<td valign="top" align="left">GH 10</td>
<td valign="top" align="left">xylanase</td>
<td valign="top" align="left">Pmob_0231</td>
<td valign="top" align="center">46.9%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1729</td>
<td valign="top" align="left">GH 43</td>
<td valign="top" align="left">&#x003B2;-xylosidase</td>
<td valign="top" align="left">Csac_1560</td>
<td valign="top" align="center">67.9%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1739</td>
<td valign="top" align="left">GH 51</td>
<td valign="top" align="left">&#x003B2;-xylosidase</td>
<td valign="top" align="left">Calhy_1625</td>
<td valign="top" align="center">58.9%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1740</td>
<td valign="top" align="left">GH 39</td>
<td valign="top" align="left">&#x003B2;-xylosidase</td>
<td valign="top" align="left">TRQ7_03440</td>
<td valign="top" align="center">38.3%</td>
</tr>
</tbody>
</table>
</table-wrap>
<table-wrap position="float" id="T3">
<label>Table 3</label>
<caption><p><bold>Annotated intracellular polysaccharide-degrading enzymes</bold>.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th valign="top" align="left"><bold>Gene</bold></th>
<th valign="top" align="left"><bold>GH family</bold></th>
<th valign="top" align="left"><bold>Annotated activity</bold></th>
<th valign="top" align="left"><bold>Nearest ortholog</bold></th>
<th valign="top" align="center"><bold>Identity</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Dtur_0081</td>
<td valign="top" align="left">GH 2</td>
<td valign="top" align="left">&#x003B2;-galactosidase</td>
<td valign="top" align="left">Calhy_1828</td>
<td valign="top" align="center">60.9%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0157</td>
<td valign="top" align="left">GH 4</td>
<td valign="top" align="left">&#x003B1;-glucosidase</td>
<td valign="top" align="left">Mc24_02443</td>
<td valign="top" align="center">47.5%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0171</td>
<td valign="top" align="left">GH 31</td>
<td valign="top" align="left">&#x003B1;-glucosidase</td>
<td valign="top" align="left">A500_11654</td>
<td valign="top" align="center">44.9%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0219</td>
<td valign="top" align="left">GH 3</td>
<td valign="top" align="left">&#x003B2;-glucosidase</td>
<td valign="top" align="left"><italic>D. tunisiensis</italic> bglB3</td>
<td valign="top" align="center">67.3%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0222</td>
<td valign="top" align="left">GH 20</td>
<td valign="top" align="left">&#x003B2;-hexosaminidase</td>
<td valign="top" align="left">CDSM653_01797</td>
<td valign="top" align="center">67.2%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0242</td>
<td valign="top" align="left">CE NC</td>
<td valign="top" align="left">feruloyl esterase</td>
<td valign="top" align="left">TM_0033</td>
<td valign="top" align="center">55.1%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0265</td>
<td valign="top" align="left">CE 7</td>
<td valign="top" align="left">acetyl xylan esterase</td>
<td valign="top" align="left">Tmari_0074</td>
<td valign="top" align="center">66.4%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0289</td>
<td valign="top" align="left">GH 3</td>
<td valign="top" align="left">&#x003B2;-glucosidase</td>
<td valign="top" align="left">Cst_c03130</td>
<td valign="top" align="center">66.8%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0315</td>
<td valign="top" align="left">GH 29</td>
<td valign="top" align="left">&#x003B1;-fucosidase</td>
<td valign="top" align="left">Tthe_0662</td>
<td valign="top" align="center">60.7%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0320</td>
<td valign="top" align="left">GH 31</td>
<td valign="top" align="left">&#x003B1;-glucosidase</td>
<td valign="top" align="left">Csac_1354</td>
<td valign="top" align="center">65.9%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0321</td>
<td valign="top" align="left">GH 3</td>
<td valign="top" align="left">&#x003B2;-glucosidase</td>
<td valign="top" align="left">Cst_c12090</td>
<td valign="top" align="center">50.1%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0384</td>
<td valign="top" align="left">GH 4</td>
<td valign="top" align="left">&#x003B1;-glucosidase</td>
<td valign="top" align="left">CTER_5006</td>
<td valign="top" align="center">48.4%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0435</td>
<td valign="top" align="left">PL 1</td>
<td valign="top" align="left">pectate lyase</td>
<td valign="top" align="left">MB27_42800</td>
<td valign="top" align="center">36.0%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0440</td>
<td valign="top" align="left">GH 4</td>
<td valign="top" align="left">&#x003B1;-galacturonidase</td>
<td valign="top" align="left">BTS2_1711</td>
<td valign="top" align="center">61.6%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0450</td>
<td valign="top" align="left">CE 4</td>
<td valign="top" align="left">deacetylase</td>
<td valign="top" align="left">Tnap_0743</td>
<td valign="top" align="center">67.4%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0451</td>
<td valign="top" align="left">GH 16</td>
<td valign="top" align="left">curdlanase</td>
<td valign="top" align="left">TRQ7_04835</td>
<td valign="top" align="center">50.9%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0462</td>
<td valign="top" align="left">GH 1</td>
<td valign="top" align="left">&#x003B2;-glucosidase</td>
<td valign="top" align="left">CLDAP_02840</td>
<td valign="top" align="center">48.5%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0490</td>
<td valign="top" align="left">GH 31</td>
<td valign="top" align="left">&#x003B1;-glucosidase</td>
<td valign="top" align="left">Tbis_2416</td>
<td valign="top" align="center">45.4%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0502</td>
<td valign="top" align="left">GH 127</td>
<td valign="top" align="left">&#x003B2;-L-arabinofuranosidase</td>
<td valign="top" align="left">CTN_0404</td>
<td valign="top" align="center">56.3%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0505</td>
<td valign="top" align="left">GH 42</td>
<td valign="top" align="left">&#x003B2;-galactosidase</td>
<td valign="top" align="left">Mahau_1293</td>
<td valign="top" align="center">59.2%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0523</td>
<td valign="top" align="left">GH 18</td>
<td valign="top" align="left">chitinase</td>
<td valign="top" align="left">Bccel_2454</td>
<td valign="top" align="center">50.1%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0551</td>
<td valign="top" align="left">GH 32</td>
<td valign="top" align="left">invertase</td>
<td valign="top" align="left">Calhy_2186</td>
<td valign="top" align="center">47.6%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0629</td>
<td valign="top" align="left">GH 26</td>
<td valign="top" align="left">&#x003B2;-mannanase</td>
<td valign="top" align="left">Calkro_1144</td>
<td valign="top" align="center">54.5%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0650</td>
<td valign="top" align="left">GH 31</td>
<td valign="top" align="left">&#x003B1;-glucosidase</td>
<td valign="top" align="left">TheetDRAFT_1156</td>
<td valign="top" align="center">45.2%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0658</td>
<td valign="top" align="left">GH 130</td>
<td valign="top" align="left">&#x003B1;-D-mannosyltransferase</td>
<td valign="top" align="left">X274_02975</td>
<td valign="top" align="center">41.2%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0670</td>
<td valign="top" align="left">GH 5</td>
<td valign="top" align="left">cellulase</td>
<td valign="top" align="left">Mahau_0466</td>
<td valign="top" align="center">61.7%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0671</td>
<td valign="top" align="left">GH 5</td>
<td valign="top" align="left">cellulase</td>
<td valign="top" align="left">TM_1752</td>
<td valign="top" align="center">58.7%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0770</td>
<td valign="top" align="left">GH 57</td>
<td valign="top" align="left">&#x003B1;-amylase</td>
<td valign="top" align="left">BROSI_A0626</td>
<td valign="top" align="center">37.3%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0794</td>
<td valign="top" align="left">GH 13</td>
<td valign="top" align="left">&#x003B1;-amylase</td>
<td valign="top" align="left">AC812_10325</td>
<td valign="top" align="center">35.3%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0852</td>
<td valign="top" align="left">GH 3</td>
<td valign="top" align="left">&#x003B2;-glucosidase</td>
<td valign="top" align="left">M164_2324</td>
<td valign="top" align="center">58.2%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0895</td>
<td valign="top" align="left">GH 57</td>
<td valign="top" align="left">&#x003B1;-amylase</td>
<td valign="top" align="left">TSIB_1115</td>
<td valign="top" align="center">46.5%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0896</td>
<td valign="top" align="left">GH 57</td>
<td valign="top" align="left">&#x003B1;-amylase</td>
<td valign="top" align="left">Calab_2422</td>
<td valign="top" align="center">40.9%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1539</td>
<td valign="top" align="left">GH 2</td>
<td valign="top" align="left">&#x003B2;-glucuronidase</td>
<td valign="top" align="left">Calkro_0120</td>
<td valign="top" align="center">60.7%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1647</td>
<td valign="top" align="left">GH 10</td>
<td valign="top" align="left">xylanase</td>
<td valign="top" align="left">PaelaDRAFT_3013</td>
<td valign="top" align="center">51.2%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1670</td>
<td valign="top" align="left">GH 36</td>
<td valign="top" align="left">&#x003B1;-galactosidase</td>
<td valign="top" align="left">Calla_1244</td>
<td valign="top" align="center">77.7%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1677</td>
<td valign="top" align="left">GH 4</td>
<td valign="top" align="left">&#x003B2;-glucosidase</td>
<td valign="top" align="left">L21TH_1859</td>
<td valign="top" align="center">47.0%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1714</td>
<td valign="top" align="left">GH 67</td>
<td valign="top" align="left">&#x003B1;-glucuronidase</td>
<td valign="top" align="left">Mc24_01903</td>
<td valign="top" align="center">69.4%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1723</td>
<td valign="top" align="left">GH 3</td>
<td valign="top" align="left">&#x003B2;-glucosidase</td>
<td valign="top" align="left"><italic>C. polysaccharolyticus</italic> Xyl3A</td>
<td valign="top" align="center">46.7%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1735</td>
<td valign="top" align="left">GH 51</td>
<td valign="top" align="left">&#x003B2;-xylosidase</td>
<td valign="top" align="left">COB47_1422</td>
<td valign="top" align="center">70.2%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1749</td>
<td valign="top" align="left">GH 4</td>
<td valign="top" align="left">&#x003B1;-glucosidase</td>
<td valign="top" align="left">TRQ7_00895</td>
<td valign="top" align="center">68.7%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1758</td>
<td valign="top" align="left">GH 38</td>
<td valign="top" align="left">&#x003B1;-mannosidase</td>
<td valign="top" align="left">CTN_0786</td>
<td valign="top" align="center">41.3%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1799</td>
<td valign="top" align="left">GH 1</td>
<td valign="top" align="left">&#x003B2;-glucosidase</td>
<td valign="top" align="left">Hore_15280</td>
<td valign="top" align="center">57.7%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1800</td>
<td valign="top" align="left">GH 43</td>
<td valign="top" align="left">&#x003B2;-xylosidase</td>
<td valign="top" align="left">Athe_2555</td>
<td valign="top" align="center">82.9%</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1802</td>
<td valign="top" align="left">GH 2</td>
<td valign="top" align="left">&#x003B2;-galactosidase</td>
<td valign="top" align="left">Thewi_0408</td>
<td valign="top" align="center">42.2%</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>BLAST analysis was used to determine the closest orthologs of the 66 Dtur CAZymes. Of these 66 enzymes, 56 have their closest orthologs in <italic>D. thermophilum</italic>, with 80&#x02013;90% amino acid identity. The remaining 10 enzymes have no orthologs in <italic>D. thermophilum</italic>. Seven of the ten unique enzymes in <italic>D. turgidum</italic> are secreted enzymes, including three of the four predicted pectin-degrading enzymes, three of the four predicted xylan-degrading enzymes and the predicted <italic>endo</italic>-arabinase. Only two Dtur enzymes, Dtur_0243 and Dtur_1800, have non-<italic>Dictyoglomus</italic> orthologs with over 80% identity (Tables <xref ref-type="table" rid="T2">2</xref>, <xref ref-type="table" rid="T3">3</xref>). The nearest non-<italic>Dictyoglomus</italic> orthologs of most of the 66 have &#x0003C;60% identity, showing the uniqueness of the Dtur enzymes. The wide range of organisms these orthologs are found in further demonstrates the uniqueness of this organism. Of the 66 enzymes, 13 have nearest orthologs in <italic>Caldicellulosiruptor</italic> species and 13 have nearest orthologs in <italic>Thermotoga</italic> species. Five orthologs are found in mesophilic <italic>Clostridia</italic> species, four in mesophilic <italic>Mahella</italic> species, and three in thermophilic <italic>Thermoanaerobacter</italic> species. The remaining 28 orthologs are spread over a wide range of mesophilic and thermophilic organisms.</p>
</sec>
<sec>
<title>Degradation of polymeric substrates</title>
<sec>
<title>Substrates reported to be degraded for which genomic and enzymatic support exists</title>
<p>Both <italic>D. turgidum</italic> and <italic>D. thermophilum</italic> are reported to utilize starch, and a number of &#x003B1;-amylases have been cloned and characterized from <italic>D. thermophilum</italic>. The genome of <italic>D. turgidum</italic> codes for three annotated extracellular &#x003B1;-amylases (Dtur_0675; Dtur_0676, and Dtur_1675) as well as four annotated intracellular &#x003B1;-amylases (Dtur_0770; Dtur_0794; Dtur_0895, and Dtur_0896) and six annotated &#x003B2;-glucosidases (Dtur_0157; Dtur_0171; Dtur_0320; Dtur_0384; Dtur_0490, and Dtur_1749). These intracellular enzymes may function in both degradation of starch oligosaccharides transported into the cell as well as degradation of glycogen stored in the cell.</p>
<p><italic>D. turgidum</italic> is reported to utilize pectin, while no data on pectin utilization was reported for <italic>D. thermophilum</italic>. The genome of <italic>D. turgidum</italic> possesses three annotated secreted pectin lyases (Dtur_0430, Dtur_0431, and Dtur_0432), one secreted pectin esterase (Dtur_0433), an annotated intracellular pectin lyase (Dtur_0435), and an annotated intracellular &#x003B1;-galacturonidase (Dtur_0440).</p>
<p><italic>D. turgidum</italic> is reported to utilize carboxymethyl cellulose. Because carboxymethyl cellulose is a man-made chemically-modified derivative of cellulose, there are no specific annotated carboxymethyl cellulases present in nature. Unlike cellulose which is crystalline and insoluble in water, carboxymethyl cellulose is an amorphous polymer that is soluble in aqueous solutions. As a result of this solubility, carboxymethyl cellulose is used as a substrate in the assay of a number of enzyme families including xylanases, cellulases, and &#x003B2;-glucanases. The genome of <italic>D. turgidum</italic> has two annotated, secreted xylanases (Dtur_0243 and Dtur_1715) and three secreted annotated cellulases (Dtur_0276; Dtur_0669, and Dtur_1586). The organism also possesses one annotated intracellular xylanase (Dtur_1647), two intracellular cellulases (Dtur_0670 and Dtur_0671) as well as six annotated &#x003B2;-glucosidases (Dtur_0219; Dtur_0289; Dtur_0321; Dtur_0462; Dtur_1723, and Dtur_1799). Assay of the enzymes expressed and purified in this work showed three cellulases (Dtur_0276; Dtur_0670 and Dtur_0671) and two xylanases (Dtur_1647 and Dtur_1715) utilized carboxymethyl cellulose as substrate, producing high levels of reducing sugars from a carboxymethyl cellulose solution (data not shown). These enzymatic assay results confirm the genomic analyses indicating that <italic>D. turgidum</italic> can utilize carboxymethyl cellulose.</p>
</sec>
<sec>
<title>Substrates reported to be degraded for which genomic and enzymatic support does not exist</title>
<p><italic>D. turgidum</italic> is reported to utilize crystalline cellulose (Svetlichny and Svetlichnaya, <xref ref-type="bibr" rid="B68">1988</xref>), though the authors report &#x0201C;the organism grew markedly less readily on microcrystalline cellulose, lignin and humic acids.&#x0201D; This in in contrast to <italic>D. thermophilum</italic>, which is reported unable to utilize cellulose. Microbial degradation of crystalline cellulose requires the expression and secretion of multiple cellulases and accessory proteins to decrystallize the cellulose chains and generate soluble, low molecular weight cellodextrins (Brumm, <xref ref-type="bibr" rid="B5">2013</xref>). These cellodextrins are then taken up via membrane transporters and further degraded into glucose monomers in the cytoplasm. Genomic and enzymatic analysis of <italic>D. turgidum</italic> indicate the organism is most likely unable to degrade crystalline cellulose. Comparison of the two genomes indicates <italic>D. turgidum</italic> contains no additional annotated cellulases not found in <italic>D. thermophilum</italic>. All five of the <italic>D. turgidum</italic> cellulases (Dtur_0276; Dtur_0669; Dtur_0670; Dtur_0671 and Dtur_1586) have orthologs in <italic>D. thermophilum</italic> (Dicth_0008; Dicth_0505; Dicth_0506; Dicth_0508 and Dicth_1476, respectively). Analysis of the genome shows a lack of GH9, GH6, GH8, GH12, or GH48 cellulases found in truly cellulytic organisms (Brumm, <xref ref-type="bibr" rid="B5">2013</xref>). Close examination of the genome reveals no cellulases containing CBM2 or CBM3 modules present in cellulose-degrading <italic>Caldicellulosiruptor</italic> species or cellulosomal structures present in cellulose-degrading <italic>C. thermocellum</italic> species within the genome. Assay of the enzymes expressed and purified in this work showed three intracellular enzymes (Dtur_1647; Dtur_0670 and Dtur_0671) produced low levels of reducing sugars from crystalline cellulose (data not shown). The two secreted xylanases (Dtur_0276 and Dtur_1715) produced no reducing sugar from the crystalline cellulose. The lack of activity by the secreted enzymes confirms the genomic analyses indicating that <italic>D. turgidum</italic> most likely cannot utilize crystalline cellulose as a growth substrate. The microcrystalline cellulose preparation used in the original study may have contained glucan, xylan or mannan, resulting in the observed weak growth of the organism on cellulose used in the experiments (Svetlichny and Svetlichnaya, <xref ref-type="bibr" rid="B68">1988</xref>).</p>
<p>The ability of <italic>D. turgidum</italic> to utilize lignin and humic acids is questionable for many of the reasons described above. The authors do not describe the source, purification, and analysis of the lignin and humic acids used in the growth experiments. Depending on the method of purification, lignin is often contaminated with mannan, cellulose and hemicellulose. Humic acids and lignin also contain sugars chemically bonded via ester linkages. Utilization of these sugars may be responsible for the low-level growth seen with these substrates. Thermophilic organisms capable of degrading aliphatic and aromatic organic compounds such as <italic>Geobacillus</italic> species, contain clearly identifiable extended gene clusters with these functions. For example, <italic>Geobacillus</italic> species Y41MC52 possesses three clusters annotated for degradation of aromatic acid molecules, GYMC52_1956 through GYMC52_1962; GYM C52_1990 through GYMC52_2001, and GYMC52_ 3134 through GYMC52_3141. Three similar clusters are found in the related strain <italic>Geobacillus</italic> species Y41MC61. Manual annotation of the <italic>D. turgidum</italic> genome failed to identify orthologs of any of the genes present in the three clusters, confirming that <italic>D. turgidum</italic> cannot utilize the aromatic ring structures found in lignin and humic acids.</p>
</sec>
<sec>
<title>Substrates not reported to be degraded for which genomic and enzymatic support exists</title>
<p>No data was reported on xylan utilization by <italic>D. turgidum</italic>, however a xylanase has been cloned and expressed from <italic>D. thermophilum</italic>. The genome of <italic>D. turgidum</italic> has two annotated, secreted xylanases (Dtur_0243 and Dtur_1715) and three secreted &#x003B2;-xylosidases (Dtur_1729; Dtur_1739, and Dtur_1740). Genes for intracellular enzymes annotated as feruloyl esterase (Dtur_0242), acetyl xylan esterase (Dtur_0265), xylanase (Dtur_1647), &#x003B1;-glucuronidase (Dtur_1714), and two &#x003B2;-xylosidases (Dtur_1735 and Dtur_1800) may be involved in degradation of oligosaccharides derived from xylan.</p>
<p>Mannans and glucans comprise a diverse group of plant-based polysaccharides that share a &#x003B2;-linked hexose backbone. Among the members of these two groups are mannan, glucomannan, galactomannan, galactoglucomannan, &#x003B2;-glucan, curdlan, and xyloglucan. No data was reported on mannan or glucan utilization by either <italic>D. turgidum</italic> or <italic>D. thermophilum</italic>. The genome of <italic>D. turgidum</italic> codes for two annotated, secreted &#x003B2;-mannanases (Dtur_0097, and Dtur_0277) and one intracellular &#x003B2;-mannanase (Dtur_0629), three secreted annotated cellulases (Dtur_0276; Dtur_0669, and Dtur_1586) and two intracellular cellulases (Dtur_0670 and Dtur_0671), as well as six annotated &#x003B2;-glucosidases (Dtur_0219; Dtur_0289; Dtur_0321; Dtur_0462; Dtur_1723, and Dtur_1799).</p>
<p>Arabinogalactan is a polysaccharide found in many plants, with the highest concentration in larch wood. No data was reported on arabinogalactan utilization by either <italic>D. turgidum</italic> or <italic>D. thermophilum</italic>. The genome of <italic>D. turgidum</italic> codes for a secreted annotated &#x003B2;-galactanase (Dtur_0857) as well as four cytoplasmic &#x003B2;-galactosidases (Dtur_0081; Dtur_0081; Dtur_0505, and Dtur_1802) and one cytoplasmic &#x003B1;-galactosidase (Dtur_1670). Together these enzymes may be adequate for degradation of arabinogalactan as well as galactose-containing oligosaccharides. Annotated genes also code for intracellular fucosidase (Dtur_0315), invertase (Dtur_0551), galacturonidase (Dtur_0440) and &#x003B2;-glucuronidase (Dtur_1539), pectate lyase (Dtur_0435), and chitinase (Dtur_0523).</p>
<p>To verify the activities of some of these enzymes, cloning, expression, and purification was attempted for 30 of the annotated carbohydrase genes. Of these thirty, 16 Dtur enzymes were successfully expressed, purified, and characterized (Table <xref ref-type="table" rid="T4">4</xref>). The 16 included two each of GH1 and GH3, four GH5, two GH10, and one each of GH36, GH42, GH43, GH53, GH57, and GH67. The remaining genes either failed to give amplicons of the correct size, or failed to express a soluble protein of the correct molecular weight (Brumm et al., <xref ref-type="bibr" rid="B6">2011</xref>).</p>
<table-wrap position="float" id="T4">
<label>Table 4</label>
<caption><p><bold>Enzymatic activity of cloned gene products</bold>.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th valign="top" align="left"><bold>Gene</bold></th>
<th valign="top" align="left"><bold>GH family</bold></th>
<th valign="top" align="left"><bold>Annotated activity</bold></th>
<th valign="top" align="left"><bold>Substrates hydrolyzed</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Dtur_0462</td>
<td valign="top" align="left">GH 1</td>
<td valign="top" align="left">&#x003B2;-glucosidase</td>
<td valign="top" align="left">MUA, MUC, MUG, MUX, XG</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1799</td>
<td valign="top" align="left">GH 1</td>
<td valign="top" align="left">&#x003B2;-glucosidase</td>
<td valign="top" align="left">MUA, MUC, MUG, MUX, XG</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0852</td>
<td valign="top" align="left">GH 3</td>
<td valign="top" align="left">&#x003B2;-glucosidase</td>
<td valign="top" align="left">MUA, MUC, MUG, MUX</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1723</td>
<td valign="top" align="left">GH 3</td>
<td valign="top" align="left">&#x003B2;-glucosidase</td>
<td valign="top" align="left">MUA, MUG, MUX</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0276</td>
<td valign="top" align="left">GH 5</td>
<td valign="top" align="left">Cellulase</td>
<td valign="top" align="left">BG</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0669</td>
<td valign="top" align="left">GH 5</td>
<td valign="top" align="left">Cellulase</td>
<td valign="top" align="left">BG, GM</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0670</td>
<td valign="top" align="left">GH 5</td>
<td valign="top" align="left">Cellulase</td>
<td valign="top" align="left">AX, BG, GM, HEC, XG</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0671</td>
<td valign="top" align="left">GH 5</td>
<td valign="top" align="left">Cellulase</td>
<td valign="top" align="left">GM</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1647</td>
<td valign="top" align="left">GH 10</td>
<td valign="top" align="left">Xylanase</td>
<td valign="top" align="left">AX, ARA, BG, HEC, MUG, MUX</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1715</td>
<td valign="top" align="left">GH 10</td>
<td valign="top" align="left">Xylanase</td>
<td valign="top" align="left">AX</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1670</td>
<td valign="top" align="left">GH 36</td>
<td valign="top" align="left">&#x003B1;-galactosidase</td>
<td valign="top" align="left">XAG</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0505</td>
<td valign="top" align="left">GH 42</td>
<td valign="top" align="left">&#x003B2;-galactosidase</td>
<td valign="top" align="left">MUA, MUX</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1729</td>
<td valign="top" align="left">GH 43</td>
<td valign="top" align="left">&#x003B1;-arabinase</td>
<td valign="top" align="left">ARA</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0857</td>
<td valign="top" align="left">GH 53</td>
<td valign="top" align="left">&#x003B2;-galactanase</td>
<td valign="top" align="left">MUA, XAG</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0675</td>
<td valign="top" align="left">GH 57</td>
<td valign="top" align="left">&#x003B1;-amylase</td>
<td valign="top" align="left">PUL</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_1714</td>
<td valign="top" align="left">GH 67</td>
<td valign="top" align="left">&#x003B1;-glucuronidase</td>
<td valign="top" align="left">Xylan, xylooligosaccharides</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<p><italic>Legend: MUC, 4-methylumbelliferyl-&#x003B2;-D-cellobioside; MUX, 4-methylumbelliferyl-&#x003B2;-D&#x02013;xylopyranoside; MUG, 4-methylumbelliferyl-&#x003B2;-D-glucoyranoside; MUA, 4-methylumbelliferyl-&#x003B1;-D-arabinofuranoside; XAG, 5-Bromo-4-chloro-3-indolyl &#x003B1;-D-galactopyranoside; XG, 5-Bromo-4-chloro-3-indolyl &#x003B2;-D-galactopyranoside; AR, AZCL-arabinan; AX, AZCL-arabinoxylan; BG, AZCL-&#x003B2;-glucan; GL, AZCL-galactan; GM, AZCL-galactomannan; HEC, AZCL-hydroxyethyl cellulose; PUL, AZCL-pullulan; XG, AZCL-xyloglucan</italic>.</p>
</table-wrap-foot>
</table-wrap>
<p>The two GH1 family members, Dtur_0462 and Dtur_1799, are annotated as &#x003B2;-glucosidases. These two cloned enzymes possess not only the predicted &#x003B2;-glucosidase activity, but also possess &#x003B2;-cellobiosidase, &#x003B2;-galactosidase, &#x003B2;-xylosidase and &#x003B2;-arabinofuranosidase activities (Table <xref ref-type="table" rid="T4">4</xref>). GH 3 family members Dtur_0852 and Dtur_1723, also annotated as &#x003B2;-glucosidases, show &#x003B2;-glucosidase, &#x003B2;-xylosidase and &#x003B2;-arabinofuranosidase activity. Dtur_0852 also possesses &#x003B2;-cellobiosidase activity, which is absent in Dtur_1723 (Table <xref ref-type="table" rid="T4">4</xref>).</p>
<p>The four GH5 family annotated cellulases show a wide range of activities. Three of the GH5 family members hydrolyze a limited number of substrates. Dtur_0276 possesses only &#x003B2;-glucanase activity, while Dtur_0671 possesses only &#x003B2;-mannanase activity. Dtur_0669 possesses both &#x003B2;-mannanase and &#x003B2;-glucanase activities. None of the three possess &#x003B2;-glucosidase, &#x003B2;-cellobiosidase, &#x003B2;-galactosidase, or &#x003B2;-xylosidase activity. In contrast to these three enzymes, Dtur_0670 (Dtur CelA) possesses both endo-activity and exo-activity on a wide range of substrates (Table <xref ref-type="table" rid="T4">4</xref>). Dtur_0670 possesses endoglucanase activity on a number of insoluble chromogenic substrates including AZCL-HE cellulose, AZCL-&#x003B2;-glucan, and AZCL-xyloglucan, endomannanase activity on AZCL-glucomannan, endoxylanase activity on AZCL-arabinoxylan, as well as &#x003B2;-glucosidase and &#x003B2;-cellobiosidase activity (Brumm et al., <xref ref-type="bibr" rid="B6">2011</xref>). None of the GH5 family members released physiologically relevant amounts of sugar from crystalline cellulose even under prolonged incubation.</p>
<p>The two GH10 family xylanases show significantly different activities. Dtur_1715 possesses only endoxylanase activity, with no other detectable <italic>endo</italic>- or <italic>exo</italic>-activities. Dtur_1647 (XynA) displays <italic>endo</italic>-activity on &#x003B2;-(1,4)-linked pentose substrates such as xylan, arabinoxylan, and linear arabinan and &#x003B2;-(1,4)-linked hexose substrates such as &#x003B2;-glucan and hydroxyethyl cellulose. XynA also possesses &#x003B2;-glucosidase, &#x003B2;-xylosidase and &#x003B2;-cellobiosidase activity (Table <xref ref-type="table" rid="T4">4</xref>).</p>
<p>The GH42 family member, Dtur_0857 predicted to be a &#x003B2;-galactanase, possesses no <italic>endo</italic>-activity, but instead possesses &#x003B2;-galactosidase and &#x003B2;-arabinofuranosidase activities. The GH43 family member, Dtur_1729, does not possess &#x003B2;-xylosidase and &#x003B2;-arabinofuranosidase activity as expected from the annotation, but instead possesses only <italic>endo</italic>-&#x003B2;-arabinase activity. The GH57 family member, Dtur_0675, possesses &#x003B1;-amylase activity as predicted. The GH67 family member, Dtur_1714, shows strong &#x003B1;-glucuronidase activity on xylan and xylan oligosaccharides as predicted (Gao et al., <xref ref-type="bibr" rid="B21">2011</xref>). Comparison to structural orthologs indicate that all 16 enzymes possess a single active site, with the differences in substrate range being a function of active site accessibility for each enzyme. There is no evidence of multiple active sites in any of the enzymes examined. The properties of these 16 enzymes show <italic>D. turgidum</italic> possesses the enzymes capable of hydrolyzing arabinoxylan, arabinan and arabinogalactan, &#x003B2;-glucan, mannan, and glucomannan to usable sugars. Failure to obtain active pectinase clones prevented us from confirming the ability of the organism to utilize pectin.</p>
</sec>
</sec>
<sec>
<title>Energy generation</title>
<p>Dtur is predicted to utilize the Embden&#x02013;Meyerhof&#x02013;Parnas pathway to produce ATP, reducing equivalents, and fermentation products from monosaccharides. Predicted products from pyruvate include lactate (Dtur_0700), acetate via acetyl-CoenzymeA (acetyl-CoA) (Dtur_0260; Dtur_0261, and Dtur_0262), and ethanol via acetyl CoA (Dtur_0260; Dtur_0261, and Dtur_0262) and acetaldehyde (Dtur_0484 and Dtur_1632). Hydrogen production is predicted by the presence of three hydrogenase gene clusters. The annotation reveals a partial <italic>hypA</italic> operon (Dtur_0074 through Dtur_0080) upstream of a hydrogenase assembly cluster (Dtur_0086 through Dtur_0090). Additional hydrogenase genes are located in a downstream cluster (Dtur_0556 through Dtur_0561). These metabolic predictions are in agreement with the published microbiological studies (Svetlichny and Svetlichnaya, <xref ref-type="bibr" rid="B68">1988</xref>) showing the organism produces acetate, lactate, ethanol, CO<sub>2</sub>, and H<sub>2</sub> during fermentation on sugars.</p>
<p>Based on the genome annotation, Dtur possesses an incomplete, reductive TCA cycle. This cycle allows the organism to convert acetate to pyruvate, oxaloacetate, and eventually to &#x003B1;-ketoglutarate. The &#x003B1;-ketoglutarate generated in this pathway can then be utilized for production of glutamate and other amino acids. Based on the annotation, the organism is unable to synthesize citrate from oxaloacetate and acetyl CoA.</p>
<p>Dtur has an extremely simple respiratory system. The genome codes for no respiratory cytochromes. The pathways for production of aminolevulinic acid from either glycine and succinyl-CoA or glutamate are both absent in Dtur. No tetrapyrroles (hemes, sirohemes, or corrinoids) are synthesized by Dtur, and the organism has no ABC-type heme transporters to utilize exogenous heme. The organism is also lacking the pathway for ubiquinone biosynthesis, indicating either ubiquinone is scavenged from the environment, or an alternate electron acceptor is utilized, like ferridoxin (Dtur_0730) or ferredoxin-like proteins (Dtur_0076; Dtur_0457; Dtur_0556; Dtur_0730; Dtur_0774, and Dtur_1717). The proton gradient needed for ATP generation is produced by NADH oxidoreductase (Dtur_0558; Dtur_0559; Dtur_0916; Dtur_0919, and Dtur_1091), and succinate dehydrogenase (Dtur_0445). ATP is generated by proton translocation via an F0F1-type ATP synthase (Dtur_129 through Dtur_135). <italic>D</italic>. turgidum also possesses a V-type ATP synthase (Dtur_1499 through Dtur_1506), the function of which is unclear. The V-type ATP synthase may also be used for ATP generation, or may it may hydrolyze ATP to generate proton or ion gradients for transport.</p>
</sec>
<sec>
<title>Other metabolic pathways</title>
<p>As expected from its small genome size, <italic>D. turgidum</italic> does not possess a full set of biosynthetic and metabolic capabilities. <italic>D. turgidum</italic> appears able to synthesize all 20 amino acids from carbohydrate precursors, but as mentioned previously, is unable to metabolize the majority back to carbohydrates. Like the amino acid situation, <italic>D. turgidum</italic> is able to synthesize fatty acids, but not degrade them. The organism appears to synthesize folate, pyridoxal 5&#x00027;-phosphate, thiamine, NAD from aspartate, and ascorbate from glucose or galactose, but is lacking pathways for biosynthesis of biotin, pantothenate or flavins.</p>
<p>Carbon monoxide is utilized by strict anaerobes via the Wood-Ljungdahl pathway (Techtmann et al., <xref ref-type="bibr" rid="B72">2009</xref>), using the anaerobic CO dehydrogenase/acetyl CoA synthase complex. This complex catalyzes the complex multistep anaerobic reactions that include oxidizing CO to CO<sub>2</sub>, formation of H<sub>2</sub>, and biosynthesis of acetyl CoA. This pathway is found in thermophilic anaerobes such as <italic>T. tengcongensis</italic> and <italic>M. thermoacetica</italic> as well as in two <italic>G. thermoglucosidasius</italic> species. Manual curation of the genome indicates that D. turgidum does not possess the Wood-Ljungdahl pathway and is unable to utilize carbon monoxide as a carbon and energy source.</p>
</sec>
<sec>
<title>DNA replication, recombination, and repair</title>
<p>IMG (DOE JGI) annotation methods identified 61 COG functional category L members (Table <xref ref-type="table" rid="T1">1</xref>) for <italic>D. turdigm</italic>. Manual reannotation of each L category gene uncovered six mis-annotated genes. Five genes annotated as excinuclease ATPase subunits (Dtur_0247, 1011, 1053, 1153, 1667) are more likely ABC type transporters and an endonuclease (Dtur_0036) has supporting evidence to be annotated as a xylose isomerase. These genes were removed from the compilation shown in Table <xref ref-type="table" rid="T5">5</xref>. A complete review of the annotated genes for <italic>D. turdigum</italic> identified a number of missed and overlooked genes that properly belong in the L COG family which totals 85 members in our revised tabulation of the genome (Table <xref ref-type="table" rid="T5">5</xref>). Even though the 16S rRNA genes of the only two described species of <italic>Dictyoglomus, D. turgidum</italic> (CP001251) and <italic>D. thermophilum</italic> (CP001146), share 99% sequence identity, the divergence of their orthologous replication proteins is significant, as described below.</p>
<table-wrap position="float" id="T5">
<label>Table 5</label>
<caption><p><bold><italic><bold>D. turdigum</bold></italic> annotated DNA replication, recombination, and repair enzymes</bold>.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th valign="top" align="left"><bold>Dtur gene</bold></th>
<th valign="top" align="left"><bold>Annotation</bold></th>
<th valign="top" align="left"><bold>Nearest neighbor/% identity</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">0745</td>
<td valign="top" align="left">ATP-dependent DNA helicase</td>
<td valign="top" align="left"><italic>Thermoanaerobaculum aquaticum</italic>/43</td>
</tr>
<tr>
<td valign="top" align="left">1397</td>
<td valign="top" align="left">ATP-dependent DNA helicase PcrA</td>
<td valign="top" align="left"><italic>Caloranaerobacter azorensis</italic>/49</td>
</tr>
<tr>
<td valign="top" align="left">1140</td>
<td valign="top" align="left">ATP-dependent DNA helicase RecG</td>
<td valign="top" align="left"><italic>Halothermothrix orenii</italic>/52</td>
</tr>
<tr>
<td valign="top" align="left">0780</td>
<td valign="top" align="left">ATP-dependent DNA ligase</td>
<td valign="top" align="left">uncultured <italic>Acidobacteria</italic>/56</td>
</tr>
<tr>
<td valign="top" align="left">1514</td>
<td valign="top" align="left">Bacterial nucleoid DNA-binding protein</td>
<td valign="top" align="left"><italic>Symbiobacterium thermophilum</italic>/66</td>
</tr>
<tr>
<td valign="top" align="left">1151</td>
<td valign="top" align="left">Cell division protein FtsK</td>
<td valign="top" align="left"><italic>Thermoanaerobacterium aotearoense</italic>/41</td>
</tr>
<tr>
<td valign="top" align="left">0001</td>
<td valign="top" align="left">Chromosomal replication initiator protein DnaA</td>
<td valign="top" align="left"><italic>Pelotomaculum thermopropioni</italic>/59</td>
</tr>
<tr>
<td valign="top" align="left">1468</td>
<td valign="top" align="left">Chromosome segregation and condensation protein ScpA</td>
<td valign="top" align="left"><italic>Caldisalinibacter Kiritimati</italic>/28</td>
</tr>
<tr>
<td valign="top" align="left">1467</td>
<td valign="top" align="left">Chromosome segregation and condensation protein ScpB</td>
<td valign="top" align="left"><italic>Dethiobacter alkaliphilus</italic>/45</td>
</tr>
<tr>
<td valign="top" align="left">0947</td>
<td valign="top" align="left">Chromosome segregation protein SMC</td>
<td valign="top" align="left"><italic>Caloranaerobacter azorensis</italic>/27</td>
</tr>
<tr>
<td valign="top" align="left">1104</td>
<td valign="top" align="left">Competence protein ComE</td>
<td valign="top" align="left"><italic>Clostridium purinilyticum</italic>/29</td>
</tr>
<tr>
<td valign="top" align="left">1103</td>
<td valign="top" align="left">Competence protein ComEA</td>
<td valign="top" align="left"><italic>Clostridium</italic> sp. CAG:273/55</td>
</tr>
<tr>
<td valign="top" align="left">1614</td>
<td valign="top" align="left">Crossover junction endodeoxyribonuclease RuvC</td>
<td valign="top" align="left"><italic>Alkaliphilus oremlandii</italic>/50</td>
</tr>
<tr>
<td valign="top" align="left">0771</td>
<td valign="top" align="left">Deoxyinosine 3&#x00027;endonuclease (endonuclease V)</td>
<td valign="top" align="left">bacterium JGI-24/54</td>
</tr>
<tr>
<td valign="top" align="left">1268</td>
<td valign="top" align="left">DNA and RNA Helicase</td>
<td valign="top" align="left">candidate division <italic>Zixibacteria</italic>/57</td>
</tr>
<tr>
<td valign="top" align="left">1547</td>
<td valign="top" align="left">DNA gyrase B subunit</td>
<td valign="top" align="left"><italic>Thermoanaerobacter</italic> sp. YS13/66</td>
</tr>
<tr>
<td valign="top" align="left">1263</td>
<td valign="top" align="left">DNA gyrase, A subunit</td>
<td valign="top" align="left"><italic>Mahella australiensis</italic>/54</td>
</tr>
<tr>
<td valign="top" align="left">1478</td>
<td valign="top" align="left">DNA integrity scanning protein DisA</td>
<td valign="top" align="left"><italic>Bacillus</italic> sp. CMAA 1185/52</td>
</tr>
<tr>
<td valign="top" align="left">1078</td>
<td valign="top" align="left">DNA mismatch repair protein MutL</td>
<td valign="top" align="left"><italic>Caldanaerobacter subterraneus</italic>/42</td>
</tr>
<tr>
<td valign="top" align="left">1077</td>
<td valign="top" align="left">DNA mismatch repair protein MutS</td>
<td valign="top" align="left"><italic>Clostridium thermocellum</italic>/46</td>
</tr>
<tr>
<td valign="top" align="left">0884</td>
<td valign="top" align="left">DNA or RNA helicase of superfamily II (UvrB)</td>
<td valign="top" align="left"><italic>Caldanaerobacter subterraneus</italic>/68</td>
</tr>
<tr>
<td valign="top" align="left">0104</td>
<td valign="top" align="left">DNA polymerase beta domain-containing protein</td>
<td valign="top" align="left"><italic>Syntrophaceticus schinkii</italic>/57</td>
</tr>
<tr>
<td valign="top" align="left">0317</td>
<td valign="top" align="left">DNA polymerase beta domain-containing protein</td>
<td valign="top" align="left">marine sediment metagenome/38</td>
</tr>
<tr>
<td valign="top" align="left">0545</td>
<td valign="top" align="left">DNA polymerase beta domain-containing protein</td>
<td valign="top" align="left"><italic>Microgenomates</italic> bacterium/39</td>
</tr>
<tr>
<td valign="top" align="left">1295</td>
<td valign="top" align="left">DNA polymerase beta domain-containing protein<sup>&#x0002A;</sup></td>
<td valign="top" align="left"><italic>Chloracidobacterium thermophilum</italic>/60</td>
</tr>
<tr>
<td valign="top" align="left">0882</td>
<td valign="top" align="left">DNA polymerase I 3&#x00027;-5&#x00027; exonuclease &#x00026; polymerase domains</td>
<td valign="top" align="left"><italic>Tepidanaerobacter acetatoxydans</italic>/44</td>
</tr>
<tr>
<td valign="top" align="left">1391</td>
<td valign="top" align="left">DNA polymerase III alpha subunit</td>
<td valign="top" align="left"><italic>Caloranaerobacter</italic> sp. TR13/52</td>
</tr>
<tr>
<td valign="top" align="left">1105</td>
<td valign="top" align="left">DNA polymerase III delta subunit</td>
<td valign="top" align="left"><italic>Alkaliphilus oremlandii</italic>/29</td>
</tr>
<tr>
<td valign="top" align="left">0257</td>
<td valign="top" align="left">DNA polymerase III gamma/tau subunits</td>
<td valign="top" align="left"><italic>Thermincola potens</italic>/41</td>
</tr>
<tr>
<td valign="top" align="left">0789</td>
<td valign="top" align="left">DNA polymerase III gamma/tau subunits</td>
<td valign="top" align="left">marine sediment metagenome/39</td>
</tr>
<tr>
<td valign="top" align="left">1551</td>
<td valign="top" align="left">DNA polymerase III sliding clamp subunit beta</td>
<td valign="top" align="left"><italic>Clostridium acidurici</italic>/37</td>
</tr>
<tr>
<td valign="top" align="left">1600</td>
<td valign="top" align="left">DNA Polymerase X</td>
<td valign="top" align="left"><italic>Caldisericum exile</italic>/55</td>
</tr>
<tr>
<td valign="top" align="left">1316</td>
<td valign="top" align="left">DNA primase N</td>
<td valign="top" align="left"><italic>Lachnospiraceae</italic> bacterium/38</td>
</tr>
<tr>
<td valign="top" align="left">1527</td>
<td valign="top" align="left">DNA protecting protein DprA</td>
<td valign="top" align="left"><italic>Caldanaerobacter subterraneus</italic>/43</td>
</tr>
<tr>
<td valign="top" align="left">1549</td>
<td valign="top" align="left">DNA recombination protein RecF</td>
<td valign="top" align="left"><italic>Clostridium aceticum</italic>/36</td>
</tr>
<tr>
<td valign="top" align="left">1625</td>
<td valign="top" align="left">DNA repair exonuclease</td>
<td valign="top" align="left"><italic>Thermobaculum terrenum</italic>/33</td>
</tr>
<tr>
<td valign="top" align="left">0327</td>
<td valign="top" align="left">DNA repair photolyase</td>
<td valign="top" align="left"><italic>Parcubacteria</italic> bacterium/41</td>
</tr>
<tr>
<td valign="top" align="left">0463</td>
<td valign="top" align="left">DNA repair photolyase</td>
<td valign="top" align="left"><italic>Caldicellulosiruptor owensensis</italic>/66</td>
</tr>
<tr>
<td valign="top" align="left">1479</td>
<td valign="top" align="left">DNA repair protein RadA</td>
<td valign="top" align="left"><italic>Thermoanaerobacterium xylanolyticum</italic>/48</td>
</tr>
<tr>
<td valign="top" align="left">0881</td>
<td valign="top" align="left">DNA repair protein RADC</td>
<td valign="top" align="left"><italic>Paenibacillus</italic> sp./50</td>
</tr>
<tr>
<td valign="top" align="left">1047</td>
<td valign="top" align="left">DNA repair protein RecN</td>
<td valign="top" align="left"><italic>Caloramator australicus</italic>/42</td>
</tr>
<tr>
<td valign="top" align="left">1321</td>
<td valign="top" align="left">DNA repair protein RecO</td>
<td valign="top" align="left"><italic>Microgenomates</italic> bacterium/29</td>
</tr>
<tr>
<td valign="top" align="left">0015</td>
<td valign="top" align="left">DNA replication and repair protein RecF</td>
<td valign="top" align="left"><italic>Fervidobacterium nodosum</italic>/80</td>
</tr>
<tr>
<td valign="top" align="left">1526</td>
<td valign="top" align="left">DNA topoisomerase type I</td>
<td valign="top" align="left"><italic>Carboxydothermus hydrogenoformans</italic>/54</td>
</tr>
<tr>
<td valign="top" align="left">1522</td>
<td valign="top" align="left">Double-stranded DNA repair protein Rad50</td>
<td valign="top" align="left"><italic>Thermoanaerobacter siderophilus</italic>/26</td>
</tr>
<tr>
<td valign="top" align="left">1626</td>
<td valign="top" align="left">Double-stranded DNA repair protein Rad50</td>
<td valign="top" align="left"><italic>Thermofilum</italic> sp./25</td>
</tr>
<tr>
<td valign="top" align="left">0264</td>
<td valign="top" align="left">Endonuclease IV</td>
<td valign="top" align="left"><italic>Thermincola potens</italic>/36</td>
</tr>
<tr>
<td valign="top" align="left">1485</td>
<td valign="top" align="left">Excinuclease ABC subunit C</td>
<td valign="top" align="left"><italic>Acetohalobium arabaticum</italic>/45</td>
</tr>
<tr>
<td valign="top" align="left">0885</td>
<td valign="top" align="left">Excinuclease ATPase subunit (ABC-ATPase UvrA)</td>
<td valign="top" align="left"><italic>Thermosediminibacter oceani</italic>/64</td>
</tr>
<tr>
<td valign="top" align="left">1613</td>
<td valign="top" align="left">Holliday junction DNA helicase RuvA</td>
<td valign="top" align="left"><italic>Thermodesulfovibrio yellowstonii</italic>/34</td>
</tr>
<tr>
<td valign="top" align="left">1612</td>
<td valign="top" align="left">Holliday junction DNA helicase RuvB</td>
<td valign="top" align="left"><italic>Caldicellulosiruptor obsidiansis</italic>/54</td>
</tr>
<tr>
<td valign="top" align="left">0792</td>
<td valign="top" align="left">Holliday junction resolvasome helicase subunit</td>
<td valign="top" align="left"><italic>Carboxydothermus hydrogenoformans</italic>/54</td>
</tr>
<tr>
<td valign="top" align="left">1284</td>
<td valign="top" align="left">Integrase<sup>&#x0002A;</sup></td>
<td valign="top" align="left"><italic>Acetothermus autotrophicum</italic>/35</td>
</tr>
<tr>
<td valign="top" align="left">0886</td>
<td valign="top" align="left">Methylated DNA-protein cysteine methyltransferase</td>
<td valign="top" align="left"><italic>Anaerococcus lactolyticus</italic>/40</td>
</tr>
<tr>
<td valign="top" align="left">1227</td>
<td valign="top" align="left">Mg-dependent Dnase&#x02014;deoxyribonuclease TatD</td>
<td valign="top" align="left"><italic>Bacillus</italic> sp. SA1-12/47</td>
</tr>
<tr>
<td valign="top" align="left">1308</td>
<td valign="top" align="left">Mismatch repair ATPase (MutS family)</td>
<td valign="top" align="left"><italic>Peptococcaceae</italic> bacterium/40</td>
</tr>
<tr>
<td valign="top" align="left">1294</td>
<td valign="top" align="left">Modification methylase, type III R/M system<sup>&#x0002A;</sup></td>
<td valign="top" align="left"><italic>Sulfurihydrogenibium yellowstonense</italic>/65</td>
</tr>
<tr>
<td valign="top" align="left">0341</td>
<td valign="top" align="left">N6-adenine-specific methylase</td>
<td valign="top" align="left"><italic>Caloranaerobacter azorensis</italic>/65</td>
</tr>
<tr>
<td valign="top" align="left">1141</td>
<td valign="top" align="left">N6-adenine-specific methylase</td>
<td valign="top" align="left"><italic>Thermoanaerobacter ethanolicus</italic>/44</td>
</tr>
<tr>
<td valign="top" align="left">1497</td>
<td valign="top" align="left">Poly(A) polymerase</td>
<td valign="top" align="left"><italic>Peptococcaceae</italic> bacterium/36</td>
</tr>
<tr>
<td valign="top" align="left">0683</td>
<td valign="top" align="left">Predicted EndoIII-related endonuclease</td>
<td valign="top" align="left"><italic>Thermodesulfatator indicus</italic>/58</td>
</tr>
<tr>
<td valign="top" align="left">0846</td>
<td valign="top" align="left">Predicted EndoIII-related endonuclease</td>
<td valign="top" align="left"><italic>Candidatus Methanoperedens</italic> sp./32</td>
</tr>
<tr>
<td valign="top" align="left">1024</td>
<td valign="top" align="left">Predicted endonuclease involved in recombination</td>
<td valign="top" align="left"><italic>Veillonella atypica</italic>/44</td>
</tr>
<tr>
<td valign="top" align="left">1530</td>
<td valign="top" align="left">Predicted endonuclease related to Holliday junction resolvase</td>
<td valign="top" align="left"><italic>Caldicellulosiruptor saccharolyticus</italic>/42</td>
</tr>
<tr>
<td valign="top" align="left">1709</td>
<td valign="top" align="left">Predicted exonuclease</td>
<td valign="top" align="left">marine sediment metagenome/61</td>
</tr>
<tr>
<td valign="top" align="left">0102</td>
<td valign="top" align="left">Predicted nucleic acid-binding protein contains PIN domain</td>
<td valign="top" align="left"><italic>Leptospira wolbachii</italic>/47</td>
</tr>
<tr>
<td valign="top" align="left">1441</td>
<td valign="top" align="left">Primosomal protein N&#x00027; (replication factor Y)</td>
<td valign="top" align="left">human gut metagenome/29</td>
</tr>
<tr>
<td valign="top" align="left">1507</td>
<td valign="top" align="left">Putative chromosome partioning protein</td>
<td valign="top" align="left"><italic>Plasmodium knowlesi</italic>/38</td>
</tr>
<tr>
<td valign="top" align="left">1162</td>
<td valign="top" align="left">RecA/RadA recombinase</td>
<td valign="top" align="left"><italic>Thermosediminibacter oceani</italic>/73</td>
</tr>
<tr>
<td valign="top" align="left">1524</td>
<td valign="top" align="left">Recombinase XerD</td>
<td valign="top" align="left">Symbiobacterium thermophilum/47</td>
</tr>
<tr>
<td valign="top" align="left">0259</td>
<td valign="top" align="left">Recombination protein RecR</td>
<td valign="top" align="left"><italic>Oxobacter pfennigii</italic>/58</td>
</tr>
<tr>
<td valign="top" align="left">1358</td>
<td valign="top" align="left">Replicative DNA helicase</td>
<td valign="top" align="left"><italic>Acetohalobium arabaticum</italic>/56</td>
</tr>
<tr>
<td valign="top" align="left">0014</td>
<td valign="top" align="left">Reverse gyrase</td>
<td valign="top" align="left"><italic>Fervidobacterium nodosum</italic>/79</td>
</tr>
<tr>
<td valign="top" align="left">0708</td>
<td valign="top" align="left">Ribonuclease HII</td>
<td valign="top" align="left"><italic>Kosmotoga pacifica</italic>/35</td>
</tr>
<tr>
<td valign="top" align="left">1531</td>
<td valign="top" align="left">Ribonuclease HII</td>
<td valign="top" align="left"><italic>Clostridium</italic> sp./59</td>
</tr>
<tr>
<td valign="top" align="left">0118</td>
<td valign="top" align="left">Single stranded nucleic acid binding protein (R3H domain)</td>
<td valign="top" align="left"><italic>Butyrivibrio</italic> sp./45</td>
</tr>
<tr>
<td valign="top" align="left">1602</td>
<td valign="top" align="left">Single-stranded DNA exonuclease RecJ</td>
<td valign="top" align="left"><italic>Clostridium thermocellum</italic>/36</td>
</tr>
<tr>
<td valign="top" align="left">1362</td>
<td valign="top" align="left">Single-stranded DNA-binding protein</td>
<td valign="top" align="left"><italic>Syntrophothermus lipocalidus</italic>/52</td>
</tr>
<tr>
<td valign="top" align="left">0786</td>
<td valign="top" align="left">Single-stranded nucleic acid binding protein (R3H domain)</td>
<td valign="top" align="left"><italic>Syntrophobacter fumaroxidans</italic>/59</td>
</tr>
<tr>
<td valign="top" align="left">0880</td>
<td valign="top" align="left">Smc1 chromosome segregation protein, putative</td>
<td valign="top" align="left">marine sediment metagenome/28</td>
</tr>
<tr>
<td valign="top" align="left">0202</td>
<td valign="top" align="left">Thermostable 8-oxoguanine DNA glycosylase</td>
<td valign="top" align="left"><italic>Thermotoga maritima</italic>/59</td>
</tr>
<tr>
<td valign="top" align="left">1511</td>
<td valign="top" align="left">Transcription-repair coupling factor</td>
<td valign="top" align="left">marine sediment metagenome/40</td>
</tr>
<tr>
<td valign="top" align="left">1297</td>
<td valign="top" align="left">Type III restriction endonuclease subunit R<sup>&#x0002A;</sup></td>
<td valign="top" align="left">bacterium JGI-6/67</td>
</tr>
<tr>
<td valign="top" align="left">1473</td>
<td valign="top" align="left">Tyrosine recombinase XerD</td>
<td valign="top" align="left"><italic>Tepidanaerobacter acetatoxydans</italic>/51</td>
</tr>
<tr>
<td valign="top" align="left">1393</td>
<td valign="top" align="left">Uracil-DNA glycosylase</td>
<td valign="top" align="left"><italic>Nitrospira moscoviensis</italic>/54</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<p><italic>Prophage related genes are marked<sup>&#x0002A;</sup></italic>.</p>
</table-wrap-foot>
</table-wrap>
<p>The genome of <italic>D. turdigum</italic> possesses 85 annotated genes for DNA replication, recombination and repair (Table <xref ref-type="table" rid="T5">5</xref>), 78 of which have their closest ortholog to Dicth genes, whereas 6 have no orthologs in Dicth (Dtur_0102, 0317, 0545, 1284, 1294, 1297). The last four genes are part of a prophage that is not present in <italic>D. thermophilum</italic>. There are two genes in this category that are present in <italic>D. thermophilum</italic> but not <italic>turdigum</italic>, including a deoxyribodipyrimidine photolyase (Dicth_0072) and a DNA modification methylase (Dicth_0253). Only one gene (Dtur_1514), annotated as a bacterial nucleoid DNA-binding protein has 100% identity to its Dicth ortholog. On average the genes in this COG category share 85% identity with their Dicth counterparts, with the lowest at 66% identity (Dtur_1626, double-stranded DNA repair protein Rad50). Based on blastP analysis the most striking feature of this class of genes is how dissimilar they are to other sequenced genes and genomes in the database, other than <italic>D. thermophilum</italic>. This is apparent when the next nearest neighbors to <italic>D. turdigum</italic> DNA replication proteins are tabulated (Table <xref ref-type="table" rid="T5">5</xref>). On average the genes in this COG category share 48% amino acid identity to nearest neighbor genes (25&#x02013;80% range) and the cross section of homology is widespread among taxa that are primarily anaerobic, thermophilic, or halophilic. Only two of the genes share homology to Thermotogae (Dtur_0202 and 0708) and 3 to <italic>Caldicellulosiruptor</italic> (Dtur_0463, 1530, and 1612). The greatest frequency of nearest neighbor orthologs after <italic>D. thermophilum</italic> are to clostridial (9%) and <italic>Thermoanerobacterium</italic> (8%) genus members, followed by unknown metagenomic genes (6%). The low homology of <italic>Dictyoglomus</italic> replication proteins to orthologs in other organisms is another testament to how phylogenetical unique the genus is.</p>
<p>In spite of its preferred growth temperature of 72&#x000B0;C (Svetlichny and Svetlichnaya, <xref ref-type="bibr" rid="B68">1988</xref>), <italic>D. turdigum</italic> has an extremely low (33.96 mol%) G&#x0002B;C content, which seems counterintuitive to genome stability and repair (Ishino and Narumi, <xref ref-type="bibr" rid="B35">2015</xref>). Dtur does possess a reverse gyrase (Dtur_0014), a hallmark enzyme that is systematically present in all hyperthermophiles (Brochier-Armanet and Forterre, <xref ref-type="bibr" rid="B3">2007</xref>), which introduces positive supercoils in DNA and thereby protects it from unwinding. Dtur and Dicth do not appear to contain genes for exonuclease III, a DNA-repair enzyme that hydrolyzes the phosphodiester bond 5&#x02032; to an abasic site in DNA, which is commonly induced by heat. However, they both possess endonuclease IV, which has been shown to perform a similar abasic site processing function in <italic>Thermotoga maritima</italic> (Haas et al., <xref ref-type="bibr" rid="B28">1999</xref>).</p>
<p><italic>D. turgidum</italic> possesses 7 DNA replication and repair genes annotated to contain a nucleotidyltransferase (NT) domain, a superfamily that includes DNA polymerase beta domain-containing proteins (NT_Pol-beta), family X and poly-A DNA polymerases, as well as other proteins (Aravind and Koonin, <xref ref-type="bibr" rid="B1">1999</xref>). The majority of the NTs are characterized by a distinct amino acid residue pattern, namely hG[GS]x(9,13)Dh[DE]h (x indicates any amino acid and h indicates a hydrophobic amino acid) that are essential for catalysis, which is true for all 7 NT domain containing genes in <italic>D. turdigum</italic>. Three of the <italic>D. turdigum</italic> NT domain-containing DNA repair proteins are larger than 450 amino acids (Dtur_0257, 1497, and 1600), whereas four members further annotated as belonging to the Pol-beta subfamily only encode 99&#x02013;135 amino acids (Figure <xref ref-type="fig" rid="F4">4</xref>). DNA polymerase B is a proofreading-proficient enzyme thought to be involved with DNA repair activities in eubacteria (Wijffels et al., <xref ref-type="bibr" rid="B75">2005</xref>) and replication in archaea (Kelman and Kelman, <xref ref-type="bibr" rid="B39">2014</xref>), however there are no known thermophilic bacterial DNA polymerase B genes. A new, ancestral family of polB type nucleotidyltransferases designated as MNT (minimal nucleotidyltransferases) has been described (Aravind and Koonin, <xref ref-type="bibr" rid="B1">1999</xref>) that are one-half to one-third the size of the larger orthologs. They are not uncommon as 258 cases can be found in the protein NCBI database &#x0201C;dna polymerase beta domain-containing protein&#x0201D; in 129 different microbes as of January 2016. However, there are no known biochemical studies showing whether these diminutive NTs are catalytically functional monomeric enzymes or whether they are part of a larger multimeric complex. The four NT_Pol-beta genes found in <italic>D. turdigum</italic>, one of which is associated with the prophage element discussed below, is an unsolved mystery as to the function these diminutive proteins might play, particularly with regard to the lack of a PolB enzyme in this hyperthermophile.</p>
<fig id="F4" position="float">
<label>Figure 4</label>
<caption><p><bold>Multiple sequence alignment of all five <italic><bold>Dictyoglomus</bold></italic> (4 from <italic><bold>turgidum</bold></italic> and one from <italic><bold>thermophilum</bold></italic>) nucleotidyltransferase domains of DNA polymerase beta family protein sequences using MAFFT software (Katoh et al., <xref ref-type="bibr" rid="B38">2002</xref>)</bold>. From top to bottom: Dtur_0104, 114AA; Dtur_0317, 121AA, Dtur_0545, 135AA, Dtur_1295, 99AA, Dicth_0227, 114AA. Dtur_1295 is located in the prophage region of <italic>D. turdigum</italic>.</p></caption>
<graphic xlink:href="fmicb-07-01979-g0004.tif"/>
</fig>
</sec>
<sec>
<title>Functional analysis of <italic>D. turdigum</italic> DNA polymerase I</title>
<p><italic>D. turgidum</italic> possesses four sets of DNA polymerases: Pol X (Dtur_1600), Poly(A) polymerase (Dtur_1497), DNA polymerase I (Dtur_0882), and a minimal DNA polymerase III set of subunits (alpha, Dtur_1391; beta, Dtur_1551; delta, Dtur_1105, and gamma/tau, Dtur_0257 and _0789). In a survey looking for thermostable reverse transcriptase (RT) activity, the DNA polymerase I (PolI) from <italic>Dictyoglomus thermophilum</italic> strain Rt46B.1 has been cloned and expressed (Shandilya et al., <xref ref-type="bibr" rid="B65">2004</xref>). While the enzyme did not exhibit RT activity, it did show significant thermal stability at 85&#x000B0;C compared to eight other enzymes being studied. Presumably its ortholog behaves the same, as Dicth_0729 shares 90% identity (770/856) and 96% positives (826/856) at the amino acid level with Dtur_0882.</p>
<p>Dtur_0882 is a PolA type polymerase annotated to contain a 5&#x02032;-3&#x02032; and 3&#x02032;-5&#x02032; exonuclease domain in addition to the DNA-directed DNA polymerase domain. With the except of <italic>Rhodothermus marinus</italic> PolA (PMID:<underline>11483153</underline>) and the enzyme OmniAmp polymerase (Chander et al., <xref ref-type="bibr" rid="B10">2014</xref>) every thermostable PolA enzyme characterized thus far lacks a functional 3&#x02032;-5&#x02032; exonuclease domain associated with proofreading and enzyme fidelity. Because the high temperature growth conditions for <italic>D. turdigum</italic> are very similar to <italic>Thermus aquaticus</italic> (Taq), but the amino acid identities are so different between the two PolI enzymes (41%, 351/847), the utility of Dtur_0882 as a PCR enzyme was evaluated and compared with Taq DNAP (Figure <xref ref-type="fig" rid="F5">5</xref>). The 3&#x02032;-5&#x02032; exonuclease activity of both enzymes were also compared with the thermostable proof reading PolA enzyme called OmniAmp polymerase (Chander et al., <xref ref-type="bibr" rid="B10">2014</xref>). Dtur_0882 produced the same yield of amplicon for the 0.9 and 2.8 kb reactions as Taq DNAP (Figure <xref ref-type="fig" rid="F5">5A</xref> lanes 2, 3), but was more efficient at amplifying the 5 and 10 kb primer templates (Figure <xref ref-type="fig" rid="F5">5A</xref> lanes 4, 5). As with Taq DNAP, Dtur_0882 does not appear to have any measurable 3&#x02032;-5&#x02032; exonuclease activity (Figure <xref ref-type="fig" rid="F5">5B</xref> lanes 2/3 compared to 4/5) as opposed to a strong exonuclease activity demonstrated by the proofreader OmniAmp DNAP (Figure <xref ref-type="fig" rid="F5">5</xref> lanes 6/7).</p>
<fig id="F5" position="float">
<label>Figure 5</label>
<caption><p><bold>Comparison of PCR efficacy between Dtur_0882 and Taq DNAP (A)</bold> and 3&#x02032;-5&#x02032; exonuclease activity between Dtur_0882; Taq and OmniAmp DNAP <bold>(B)</bold>. PCR amplicons of 0.9 kb (lane A2), 2.8 kb (lane A3), 5 kb (lane A4), or 10 kb (lane A5) were produced by Taq (T) or Dtur (D) DNAP. To assess exonuclease activity (lanes B2-9) lambda DNA restriction digested with Hind III was incubated with 5U of Taq (T), Dtur (D) or OmniAmp (A) DNAP (in duplicate) overnight at 37&#x000B0;C in PCR buffer. Lane 1 is a 1 kb DNA ladder (Promega).</p></caption>
<graphic xlink:href="fmicb-07-01979-g0005.tif"/>
</fig>
</sec>
<sec>
<title>Prophage and CRISPR elements</title>
<p><italic>D. turgidum</italic> possesses two regions containing CRISPR repeats, suggesting the previous exposure to phage(s). The first CRISPR region, located between 470870 and 474424 nucleotides in the genome codes for 54 repeats. The repeat sequence is 30 nucleotides long, and the spacer average length is 36 nucleotides long. The second CRISPR region, located between nucleotides 6151530 and 617311 in the genome codes for 33 repeats. The repeat sequence is the same as the first CRISPR region, and the spacer average length is 36 nucleotides long. No CRISPR-associated proteins are in the vicinity of the first CRISPR region. Upstream of the second CRISPR region are eight CRISPR-associated proteins (Table <xref ref-type="table" rid="T5">5</xref>), while downstream of the second CRISPR region is a four-gene insert coding for biosynthetic enzymes (Dtur_0614&#x02014;Dtur_0617) and eight additional CRISPR-associated proteins. <italic>D. thermophilum</italic> shows a similar organization of its CRISPR-associated proteins (Table <xref ref-type="table" rid="T6">6</xref>), with a larger, 15-gene insert coding for biosynthetic enzymes (not shown).</p>
<table-wrap position="float" id="T6">
<label>Table 6</label>
<caption><p><bold><italic><bold>D. turgidum</bold></italic> CRISPR-associated proteins</bold>.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th valign="top" align="left"><bold>Gene</bold></th>
<th valign="top" align="left"><bold>Annotation</bold></th>
<th valign="top" align="left"><bold>Orthologs</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Dtur_0606</td>
<td valign="top" align="left">CRISPR-associated protein, TM1812 family</td>
<td valign="top" align="left">Dicth_0458</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0607</td>
<td valign="top" align="left">CRISPR-associated protein DxTHG motif protein</td>
<td valign="top" align="left">Dicth_0187</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0608</td>
<td valign="top" align="left">CRISPR-associated RAMP protein, Csm5 family</td>
<td valign="top" align="left">Dicth_0186</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0609</td>
<td valign="top" align="left">CRISPR-associated RAMP protein, Csm4 family</td>
<td valign="top" align="left">Dicth_0185</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0610</td>
<td valign="top" align="left">CRISPR-associated RAMP protein, Csm3 family</td>
<td valign="top" align="left">Dicth_0184</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0611</td>
<td valign="top" align="left">CRISPR-associated RAMP protein, Csm2 family</td>
<td valign="top" align="left">Dicth_0183</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0612</td>
<td valign="top" align="left">CRISPR-associated RAMP protein, Csm1 family</td>
<td valign="top" align="left">Dicth_0182</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0613</td>
<td valign="top" align="left">CRISPR-associated protein, Csx3 family</td>
<td valign="top" align="left">Dicth_0181</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0618</td>
<td valign="top" align="left">CRISPR-associated protein, Cas2 family</td>
<td valign="top" align="left">Dicth_0165</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0619</td>
<td valign="top" align="left">CRISPR-associated protein, Cas1 family</td>
<td valign="top" align="left">Dicth_0164</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0620</td>
<td valign="top" align="left">CRISPR-associated endonuclease, Cas4 family</td>
<td valign="top" align="left">Dicth_0163</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0621</td>
<td valign="top" align="left">CRISPR-associated helicase, Cas3 family</td>
<td valign="top" align="left">Dicth_0162</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0622</td>
<td valign="top" align="left">CRISPR-associated protein, Cas5h family</td>
<td valign="top" align="left">Dicth_0161</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0623</td>
<td valign="top" align="left">CRISPR-associated protein, Csh2 family</td>
<td valign="top" align="left">Dicth_0160</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0624</td>
<td valign="top" align="left">CRISPR-associated protein, Csh1 family</td>
<td valign="top" align="left">Dicth_0159</td>
</tr>
<tr>
<td valign="top" align="left">Dtur_0625</td>
<td valign="top" align="left">CRISPR-associated protein, Cas6 family</td>
<td valign="top" align="left">Dicth_0158</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>An incomplete prophage sequence was identified by PHAST (Hubisz et al., <xref ref-type="bibr" rid="B33">2011</xref>; Zhou et al., <xref ref-type="bibr" rid="B76">2011</xref>) as a 17,407 base insert from 1,301,419 to 1,318,825 (Dtur_1284-1300) and confirmed as foreign DNA by IslandViewer Software (Hsiao et al., <xref ref-type="bibr" rid="B31">2003</xref>) (1,298,422 to 1,317,048 containing genes Dtur_1282 through Dtur_1297). The prophage contains an integrase (Dtur_1284) followed by four annotated putative lipoproteins that potentially form part of a <italic>beta</italic>-barrel assembly machinery (Dtur_1285; Dtur_1286; Dtur_1288, and Dtur_1289) and a Type III restriction system (Dtur_1294 and Dtur_1297). The four annotated lipoprotein genes are closely related, as they share 79&#x02013;88% amino acid identity. Inspection of the amino acid sequence reveals a unique periodicity of aspartic (D) and glutamic (E) residues to hydrophobic residues. The reason for this novel periodicity is due to the seven back to back, nearly perfect 49&#x02013;52 amino acid tandem repeats found in these proteins (data not shown). Tandem repeat proteins are ubiquitous (Jernigan and Bordenstein, <xref ref-type="bibr" rid="B36">2015</xref>), but the unique signature sequence found here is only partially common to a handful of hypothetical proteins found in bacteria. Additional work is needed to clarify the function of these four repeated proteins with seven tandem internal repeats. This genomic island is unique to <italic>D. turgidum</italic>. <italic>D. thermophilum</italic> possesses an ortholog to only one of the four lipoproteins and no Type III restriction system proteins.</p>
</sec>
<sec>
<title>Thermophily, stress responses, and heat shock proteins</title>
<p><italic>D. turgidum</italic> in common with <italic>D. thermophilum</italic> has a complement of heat shock proteins typical of thermophilic bacteria, including a single GroEL/ES locus encoding Hsp60 and a DnaK/DnaJ locus encoding Hsp70 and cochaperones GrpE. Interestingly, in both <italic>D turgidum</italic> and <italic>D. thermophilum</italic>, the reverse gyrase is encoded in a gene cluster shared with recJ and revG, closely linked to the DnaK/DnaJ operon. Heat shock regulation is enigmatic since the genome lacks both CIRCE elements and sigma32 SOS regulation. It is tempting to speculate that conditional expression of the reverse gyrase under high temperature growth conditions might be a mechanism for regulating DNA positive supercoiling in concert with the heat shock response. The phylogeny of the reverse gyrase is extraordinary. The reverse gyrase is most closely related to orthologs in <italic>Fervidobacterium</italic> species as shown in Figure <xref ref-type="fig" rid="F6">6</xref>. This phylogenetic position of the reverse gyrase does not conform to the 16S rRNA phylogeny (Figure <xref ref-type="fig" rid="F3">3</xref>) where <italic>D. turgidum</italic> is most closely related to <italic>Caldicellulosiruptor</italic> species. This suggests that lateral gene transfer of the reverse gyrase may have taken place.</p>
<fig id="F6" position="float">
<label>Figure 6</label>
<caption><p><bold>Evolutionary relationships of reverse gyrase</bold>. The evolutionary history was inferred by using the Maximum Likelihood method based on the JTT matrix-based model (Tamura et al., <xref ref-type="bibr" rid="B71">2011</xref>) The tree with the highest log likelihood (&#x02212;22290.6302) is shown. Initial tree(s) for the heuristic search were obtained automatically by applying Neighbor-Join and BioNJ algorithms to a matrix of pairwise distances estimated using a JTT model, and then selecting the topology with superior log likelihood value. The tree is drawn to scale, with branch lengths measured in the number of substitutions per site. The analysis involved 7 amino acid sequences. All positions containing gaps and missing data were eliminated. There were a total of 1122 positions in the final dataset. Evolutionary analyses were conducted in MEGA7 (Kumar et al., <xref ref-type="bibr" rid="B46">2016</xref>). UniPpot sequences used for the analysis were: B8DYH3, <italic>Dictyoglomus turgidum</italic> strain DSM 6724; A7HMS7, <italic>Fervidobacterium nodosum</italic> strain DSM 5306; H9UDK4, <italic>Fervidobacterium pennivorans</italic> strain DSM 9078; C1DT23, <italic>Sulfurihydrogenibium azorense</italic> strain DSM 15241; B2V6S9, <italic>Sulfurihydrogenibium</italic> sp. strain YO3AOP; F8C2X1, <italic>Thermodesulfobacterium geofonti</italic>s strain OPF15, and P95479; <italic>Pyrococcus furiosus</italic> strain DSM 3638.</p></caption>
<graphic xlink:href="fmicb-07-01979-g0006.tif"/>
</fig>
</sec>
<sec>
<title>Morphological characteristics</title>
<p>Microbiological testing of <italic>D. turgidum</italic> indicate the organism stains Gram-negative (Svetlichny and Svetlichnaya, <xref ref-type="bibr" rid="B68">1988</xref>). The <italic>D. turgidum</italic> genome contains a cluster of 10 genes potentially coding for outer membrane proteins characteristic of Gram-negative organisms and outer membrane lipid biosynthesis (Dtur_0814 through Dtur_0824). This cluster contains genes coding for a TamB (Dtur_0814), two BamA orthologs (Dtur_0815 and Dtur_0816), and two outer membrane chaperone Skp (OmpH) orthologs (Dtur_0817 and Dtur_0818), all potentially involved in outer membrane transport and assembly. This cluster of genes provides genomic support for the observed Gram-negative membrane structure observed in electron micrographs (Svetlichny and Svetlichnaya, <xref ref-type="bibr" rid="B68">1988</xref>). Following these five genes, the <italic>D. turgidum</italic> genome contains a lipid biosynthesis cluster coding for the first four enzymes of Lipid A biosynthesis, LpxD (Dtur_0819), LpxC (Dtur_0820), LpxA (Dtur_0821), and (Dtur_0823) and an ortholog of FabZ (beta-hydroxyacyl-(acyl-carrier-protein) dehydratase, Dtur_0821). The final gene in the cluster (Dtur_0824) is a hypothetical protein related to LpxB. This 10-gene cluster has the identical organization in <italic>D. thermophilum</italic>, with individual genes averaging 80% identity to its Dtur ortholog. The cluster appears to be unique to <italic>Dictyoglomus</italic> species, as no similar cluster is found in any other sequenced organisms. The individual genes also show little homology to orthologs in other species, with observed amino acid identities being &#x02264;30% for all genes in the cluster.</p>
<p>Downstream of this ten-gene cluster is an annotated cluster of six proteins potentially involved in Gram-negative outer membrane efflux including orthologs of an ABC transporter ATP-binding protein (Dtur_0835), two of TolC (Dtur_0836 and Dtur_0837), HlyD (Dtur_0838) an ABC transporter permease protein (Dtur_0839) and a predicted transmembrane protein (Dtur_0840). An identical cluster is found in <italic>D. thermophilum</italic> (DICTH_0678 through DICTH_0683). A similar cluster is found in <italic>Meiothermus taiwanensis</italic> DSM 14542 as well as other <italic>Meiothermus</italic> and <italic>Thermus</italic> species. No orthologs of this cluster are found in any sequenced Thermotogales or Firmicutes species.</p>
<p>Many thermophilic bacteria possess complex morphologies, with varying shapes seen under different growth conditions. Examples of this include &#x0201C;rotund bodies&#x0201D; in <italic>Thermus aquaticus</italic> (Brumm P. J. et al., <xref ref-type="bibr" rid="B7">2015</xref>), the outer membrane &#x0201C;toga&#x0201D; of <italic>Thermotoga maritima</italic> (Huber et al., <xref ref-type="bibr" rid="B32">1986</xref>) and the multicellular spheres of <italic>D. turgidum</italic> (Svetlichny and Svetlichnaya, <xref ref-type="bibr" rid="B68">1988</xref>). Regulation of these morphologies may be controlled by the action of SpoVS (Brumm P. J. et al., <xref ref-type="bibr" rid="B7">2015</xref>). <italic>D. turgidum</italic> possesses a gene coding for SpoVS (Rigden and Galperin, <xref ref-type="bibr" rid="B61">2008</xref>) (Dtur_0800) similar to SpoVS proteins found in sporulating <italic>Firmicutes</italic> species as well as in the non-sporulating <italic>Thermus</italic>-<italic>Deinococcu</italic>s and <italic>Thermotoga</italic> groups, but not in non-sporulating <italic>Firmicutes</italic> species. Phylogenetic reconstruction indicates that <italic>D. turgidum</italic> SpoVS is most closely related to the <italic>D. thermophilum</italic> SpoVS, (Figure <xref ref-type="fig" rid="F7">7</xref>) followed by the SpoVS of <italic>Thermosediminibacter oceani</italic> DSM 16646. These three SpoVS molecules form a separate clade from the SpoVS of the <italic>Firmicutes</italic> and <italic>Deinococcus/Thermus</italic> species. The SpoVS orthologs show much higher homology than orthologs of other <italic>Dictyoglomus</italic> proteins, suggesting an important conserved function. <italic>Thermotoga</italic> SpoVS orthologs have 53&#x02013;78% identity, a <italic>Clostridium thermocellum</italic> ortholog has 74% identity, <italic>Caldicellulosiruptor</italic> orthologs have 65&#x02013;71% identity, and <italic>Thermus</italic> orthologs have 56% identity to Dtur_0800. SpoVS may be an important regulator of cell morphology and differentiation in both sporulating thermophiles where it regulates the transition from vegetative growth to spore formation as well as the non-sporulating thermophiles where it regulates the transition from vegetative growth to formation of multiple morphologies (Brumm P. J. et al., <xref ref-type="bibr" rid="B7">2015</xref>).</p>
<fig id="F7" position="float">
<label>Figure 7</label>
<caption><p><bold>Evolutionary relationships of SpoVS proteins</bold>. The evolutionary history was inferred by using the Maximum Likelihood method based on the JTT matrix-based model. The bootstrap consensus tree inferred from 550 replicates is taken to represent the evolutionary history of the taxa analyzed. Branches corresponding to partitions reproduced in less than 50% bootstrap replicates are collapsed. The percentage of replicate trees in which the associated taxa clustered together in the bootstrap test (550 replicates) are shown next to the branches. Initial tree(s) for the heuristic search were obtained automatically by applying Neighbor-Join and BioNJ algorithms to a matrix of pairwise distances estimated using a JTT model, and then selecting the topology with superior log likelihood value. The analysis involved 22 amino acid sequences. All positions containing gaps and missing data were eliminated. There were a total of 86 positions in the final dataset.</p></caption>
<graphic xlink:href="fmicb-07-01979-g0007.tif"/>
</fig>
</sec>
</sec>
<sec sec-type="discussion" id="s4">
<title>Discussion</title>
<p>We report here the genome sequence, sequence analysis, and cloning of key enzymes of <italic>D. turgidum</italic>, an anaerobic, thermophile reported to degrade a wide range of biomass components including starch, cellulose, pectin and lignin [14]. This hyperthermophile has a small 1.8 M bp genome with a G&#x0002B;C content of the 34%. COGS analysis shows the organism is enriched in genes coding for carbohydrate transport and metabolism. While <italic>Dictyoglomus</italic> make up 25% of the species identified by 16S rRNA sequencing in some environments currently only two species, <italic>D. thermophilum</italic> (Saiki et al., <xref ref-type="bibr" rid="B63">1985</xref>) and <italic>D. turgidus</italic> (Svetlichny and Svetlichnaya, <xref ref-type="bibr" rid="B68">1988</xref>), corrected to <italic>D. turgidum., D. turgidum</italic>, and <italic>D. thermophilum</italic>, have been sequenced and annotated. This first comparison of the two genomes shows that, while the two organisms are unique species, they show extremely high levels of orthologous genes, average nucleotide identity, and synteny. No unique metabolic pathways are present in either organism. Approximately 1/3 of the proteins present in Dtur have orthologs in <italic>D. thermophilum</italic> with over 90% amino acid identity, and less than 10% of the proteins present in either genome have no ortholog in the other genome. The two organisms show extensive short-range and long-range synteny. Genome sequences of additional <italic>Dictyoglomus</italic> species are needed to determine if this is coincidence or a conserved feature of the <italic>Dictyoglomi</italic>. Additional work is also needed to confirm that the differences in synteny observed between the two genomes are real and are not artifacts of the assembly of the genomes.</p>
<p>The genome of <italic>D. turgidum</italic> provides insights into an organism that is strangely foreign and vaguely familiar at the same time. At first glance, the genome is remarkably unremarkable, containing no novel pathways or secondary products. In fact, the organism is lacking many of the pathways normally associated with microbes, including amino acid and fatty acid degradation pathways and energy harvesting via proteins containing hemes, sirohemes, or quinones and appears to be the genome of a strict carbohydrate fermentor. Yet, at the same time, <italic>D. turgidum</italic> cannot be identified as similar to any one organism or Phylum. Results presented here and elsewhere (Nishida et al., <xref ref-type="bibr" rid="B55">2011</xref>; Vesth et al., <xref ref-type="bibr" rid="B73">2013</xref>) show the chameleon-like nature of the organism. Changing the method of comparison radically changes the resulting relationships between <italic>D. turgidum</italic> and other organisms.</p>
<p>The metabolic reconstruction based on the <italic>D. turgidum</italic> genome reveals two unusual features. Most evident is the importance of carbohydrate metabolism for the organism, because <italic>D. turgidum</italic> lacks the ability to metabolize fatty acids and most amino acids. The genome, while lacking in enzymes to degrade crystalline cellulose, possesses genes coding for utilization of most other biomass-derived polymers including xylans, glucans, pectins, arabinans and galactans. Utilization of these polysaccharides appears to involve secretion of enzymes that degrade the polysaccharides to oligosaccharides, transport of the oligosaccharides into the cytoplasm, and degradation of the oligosaccharides to monosaccharides in the cytoplasm. A similar strategy is utilized by thermophilic <italic>Geobacillus</italic> species (Brumm P. et al., <xref ref-type="bibr" rid="B4">2015</xref>). Genes for utilization of carbohydrates are distributed randomly throughout the <italic>D. turgidum</italic> genome, unlike the <italic>Geobacillus</italic> genomes, where genes for individual polysaccharide degradation pathways are organized into distinct operons. The second feature is the obligate fermentative nature of <italic>D. turgidum</italic>. Unlike many other thermophilic anaerobes including <italic>Thermus, Geobacillus, Caldicellulosiruptor</italic>, and <italic>Thermotoga</italic> species, <italic>D. turgidum</italic> possesses no genes for production or utilization of either cytochromes or quinones. Dtur is predicted to utilize the EMP pathway, to produce ATP, reducing equivalents, and fermentation products from monosaccharides. The predicted fermentation products, lactate, acetate, ethanol and hydrogen are in agreement with the published microbiological studies (Svetlichny and Svetlichnaya, <xref ref-type="bibr" rid="B68">1988</xref>) showing the organism produces these four products during fermentation on sugars. The proton gradient needed for ATP generation is produced by NADH oxidoreductase and succinate dehydrogenase, and the ATP is generated by an F0F1-type and a V-type ATP synthases.</p>
<p>Sixteen <italic>D. turdigum</italic> carbohydrases were cloned, expressed and characterized to better understand their function in the metabolism of the organism. The 16 included two each of GH1 and GH3, four GH5, two GH10, and one each of GH36, GH42, GH43, GH53, GH57, and GH67. Based on the proposed mechanism for polysaccharide utilization, <italic>D. turdigum</italic> produces oligosaccharides using secreted enzymes, and degrades the oligosaccharides using intracellular enzymes. The secreted enzymes would be expected to have high substrate specificity to generate oligosaccharides recognized by the transporter systems. The cloned enzymes predicted to be secreted showed activity only on one or two substrates, showing activity on xylan, arabinan, <italic>beta</italic>-glucan, starch, or mannan. Conversely, the intracellular enzymes would be expected to have low specificity, allowing them to degrade multiple substrates and linkages efficiently. The cloned intracellular enzymes typically showed a broader range of activities. GH1 and GH3 enzymes possess <italic>exo</italic>-activity on four or five different carbohydrate substrates. The cloned intracellular xylanase and cellulase both possess both <italic>exo</italic>-activity and <italic>endo</italic>-activity, as well as activity on multiple substrates.</p>
<p>Replication, recombination, and repair enzymes are critical to the genome maintenance and integrity of all cells. Many proteins from this COG functional category are expected to share conserved domains and motifs that could in theory be used to understand the phylogenetic relationship of <italic>D. turdigum</italic> to other organisms. The 16S rRNA genes of <italic>D. turgidum</italic> and <italic>D. thermophilum</italic> share 99% sequence identity. The fraction of replication proteins having 100&#x02013;90% similarity between the two species is 35%, with 42% sharing 89&#x02013;80% similarity, 17% sharing 79&#x02013;70% similarity and 5% with similarity below 69%. The similarity of <italic>D. turgidum</italic> replication proteins to other taxa drops off considerably from that with <italic>D. thermophilum</italic>. The fraction of replication proteins having 100&#x02013;90% similarity between <italic>D. turgidum</italic> and the next nearest neighbor species is 0%, with 1% sharing 89&#x02013;80% similarity, 3% sharing 79&#x02013;70% similarity, 13% sharing 69&#x02013;60 similarity, 32% sharing 59&#x02013;50% similarity, 28% sharing 49&#x02013;40 similarity, and 32% similarity below 40%. This informal comparison again demonstrates how unique <italic>Dictoglomi</italic> are compared to other species. The number and type of replication proteins found in <italic>D. turdigum</italic> is similar to those found in other hyperthermophilic bacteria using IMG tools at the Joint Genome Institute (data not shown). The phylogenetic position of the reverse gyrase does not conform to the 16S rRNA phylogeny suggesting that lateral gene transfer may have taken place.</p>
</sec>
<sec id="s5">
<title>Author contributions</title>
<p>FR analyzed data and contributed to manuscript preparation. PB wrote the manuscript, produced and purified Dtur proteins, and performed carbhohydrase analyses. KG produced and purified Dtur proteins. DM managed genome sequencing and analysis, performed DNAP analyses, and contributed to manuscript preparation.</p>
</sec>
<sec id="s6">
<title>Funding</title>
<p>This work was completely funded by the DOE Great Lakes Bioenergy Research Center (DOE BER Office of Science DE-FC02-07ER64494 and DOE OBP Office of Energy Efficiency and Renewable Energy DE-AC05-76RL01830). FR acknowledges support from the NASA Exobiology Program.</p>
<sec>
<title>Conflict of interest statement</title>
<p>At the time this work was performed, the authors PB and DM were employees and shareholders of C5-6 Technologies Inc. (WI, USA), a company that created bio-based solutions to efficiently convert biomass into five and six carbon sugars. The company ceased operation in December of 2014. PB has since purchased the assets of the company and started C5-6 Technologies LLC (WI, USA), a company focused on supplying reagent enzymes for carbohydrase research. The authors have no other relevant affiliations or financial involvement with any organization or entity with a financial interest in or financial conflict with the subject matter or materials discussed in the manuscript apart from those disclosed. No writing assistance was utilized in the production of this manuscript. The other authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
</sec>
</body>
<back>
<ack><p>The authors would like to thank Robb lab and Elizabeth O&#x00027;Connor for assistance in the growth and preparation of the cells used for genomic DNA isolation.</p>
</ack>
<ref-list>
<title>References</title>
<ref id="B1">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Aravind</surname> <given-names>L.</given-names></name> <name><surname>Koonin</surname> <given-names>E. V.</given-names></name></person-group> (<year>1999</year>). <article-title>DNA polymerase &#x003B2;-like nucleotidyltransferase superfamily: identification of three new families, classification and evolutionary history</article-title>. <source>Nucleic Acids Res.</source> <volume>27</volume>, <fpage>1609</fpage>&#x02013;<lpage>1618</lpage>. <pub-id pub-id-type="doi">10.1093/nar/27.7.1609</pub-id><pub-id pub-id-type="pmid">10075991</pub-id></citation>
</ref>
<ref id="B2">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Aziz</surname> <given-names>R. K.</given-names></name> <name><surname>Bartels</surname> <given-names>D.</given-names></name> <name><surname>Best</surname> <given-names>A. A.</given-names></name> <name><surname>DeJongh</surname> <given-names>M.</given-names></name> <name><surname>Disz</surname> <given-names>T.</given-names></name> <name><surname>Edwards</surname> <given-names>R. A.</given-names></name> <etal/></person-group>. (<year>2008</year>). <article-title>The RAST Server: rapid annotations using subsystems technology</article-title>. <source>BMC Genomics</source> <volume>9</volume>:<fpage>75</fpage>. <pub-id pub-id-type="doi">10.1186/1471-2164-9-75</pub-id><pub-id pub-id-type="pmid">18261238</pub-id></citation>
</ref>
<ref id="B3">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Brochier-Armanet</surname> <given-names>C.</given-names></name> <name><surname>Forterre</surname> <given-names>P.</given-names></name></person-group> (<year>2007</year>). <article-title>Widespread distribution of archaeal reverse gyrase in thermophilic bacteria suggests a complex history of vertical inheritance and lateral gene transfers</article-title>. <source>Archaea</source> <volume>2</volume>, <fpage>83</fpage>&#x02013;<lpage>93</lpage>. <pub-id pub-id-type="doi">10.1155/2006/582916</pub-id><pub-id pub-id-type="pmid">17350929</pub-id></citation>
</ref>
<ref id="B4">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Brumm</surname> <given-names>P. J.</given-names></name> <name><surname>De Maayer</surname> <given-names>P.</given-names></name> <name><surname>Cowan</surname> <given-names>D. A.</given-names></name> <name><surname>Mead</surname> <given-names>D. A.</given-names></name></person-group> (<year>2015</year>). <article-title>Genomic analysis of six new Geobacillus strains reveals highly conserved carbohydrate degradation architectures and strategies</article-title>. <source>Front. Microbiol.</source> <volume>6</volume>:<fpage>430</fpage>. <pub-id pub-id-type="doi">10.3389/fmicb.2015.00430</pub-id><pub-id pub-id-type="pmid">26029180</pub-id></citation>
</ref>
<ref id="B5">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Brumm</surname> <given-names>P. J.</given-names></name></person-group> (<year>2013</year>). <article-title>Bacterial genomes: what they teach us about cellulose degradation</article-title>. <source>Biofuels</source> <volume>4</volume>, <fpage>669</fpage>&#x02013;<lpage>681</lpage>. <pub-id pub-id-type="doi">10.4155/bfs.13.44</pub-id></citation>
</ref>
<ref id="B6">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Brumm</surname> <given-names>P. J.</given-names></name> <name><surname>Hermanson</surname> <given-names>S.</given-names></name> <name><surname>Luedtke</surname> <given-names>J.</given-names></name> <name><surname>Mead</surname> <given-names>D. A.</given-names></name></person-group> (<year>2011</year>). <article-title>Identification, cloning and characterization of <italic>Dictyoglomus turgidum</italic> CelA, an endoglucanase with cellulase and mannanase activity</article-title>. <source>J. Life Sci.</source> <volume>5</volume>, <fpage>488</fpage>&#x02013;<lpage>496</lpage>.</citation>
</ref>
<ref id="B7">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Brumm</surname> <given-names>P. J.</given-names></name> <name><surname>Monsma</surname> <given-names>S.</given-names></name> <name><surname>Keough</surname> <given-names>B.</given-names></name> <name><surname>Jasinovica</surname> <given-names>S.</given-names></name> <name><surname>Ferguson</surname> <given-names>E.</given-names></name> <name><surname>Schoenfeld</surname> <given-names>T.</given-names></name> <etal/></person-group>. (<year>2015</year>). <article-title>Complete genome sequence of thermus aquaticus Y51MC23</article-title>. <source>PLoS ONE</source> <volume>10</volume>:<fpage>e0138674</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pone.0138674</pub-id><pub-id pub-id-type="pmid">26465632</pub-id></citation>
</ref>
<ref id="B8">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Burgess</surname> <given-names>E. A.</given-names></name> <name><surname>Unrine</surname> <given-names>J. M.</given-names></name> <name><surname>Mills</surname> <given-names>G. L.</given-names></name> <name><surname>Romanek</surname> <given-names>C. S.</given-names></name> <name><surname>Wiegel</surname> <given-names>J.</given-names></name></person-group> (<year>2012</year>). <article-title>Comparative geochemical and microbiological characterization of two thermal pools in the Uzon Caldera, Kamchatka, Russia</article-title>. <source>Microb. Ecol.</source> <volume>63</volume>, <fpage>471</fpage>&#x02013;<lpage>489</lpage>. <pub-id pub-id-type="doi">10.1007/s00248-011-9979-4</pub-id><pub-id pub-id-type="pmid">22124570</pub-id></citation>
</ref>
<ref id="B9">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Caspi</surname> <given-names>R.</given-names></name> <name><surname>Altman</surname> <given-names>T.</given-names></name> <name><surname>Billington</surname> <given-names>R.</given-names></name> <name><surname>Dreher</surname> <given-names>K.</given-names></name> <name><surname>Foerster</surname> <given-names>H.</given-names></name> <name><surname>Fulcher</surname> <given-names>C. A.</given-names></name> <etal/></person-group>. (<year>2014</year>). <article-title>The MetaCyc database of metabolic pathways and enzymes and the BioCyc collection of pathway/genome databases</article-title>. <source>Nucleic Acids Res</source>. <volume>42</volume>, <fpage>D459</fpage>&#x02013;<lpage>D471</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gkt1103</pub-id><pub-id pub-id-type="pmid">24225315</pub-id></citation>
</ref>
<ref id="B10">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chander</surname> <given-names>Y.</given-names></name> <name><surname>Koelbl</surname> <given-names>J.</given-names></name> <name><surname>Puckett</surname> <given-names>J.</given-names></name> <name><surname>Moser</surname> <given-names>M. J.</given-names></name> <name><surname>Klingele</surname> <given-names>A. J.</given-names></name> <name><surname>Liles</surname> <given-names>M. R.</given-names></name> <etal/></person-group>. (<year>2014</year>). <article-title>A novel thermostable polymerase for RNA and DNA loop-mediated isothermal amplification (LAMP)</article-title>. <source>Front. Microbiol.</source> <volume>5</volume>:<fpage>395</fpage>. <pub-id pub-id-type="doi">10.3389/fmicb.2014.00395</pub-id><pub-id pub-id-type="pmid">25136338</pub-id></citation>
</ref>
<ref id="B11">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Coil</surname> <given-names>D. A.</given-names></name> <name><surname>Badger</surname> <given-names>J. H.</given-names></name> <name><surname>Forberger</surname> <given-names>H. C.</given-names></name> <name><surname>Riggs</surname> <given-names>F.</given-names></name> <name><surname>Madupu</surname> <given-names>R.</given-names></name> <name><surname>Fedorova</surname> <given-names>N.</given-names></name> <etal/></person-group>. (<year>2014</year>). <article-title>Complete genome sequence of the extreme thermophile <italic>Dictyoglomus thermophilum</italic> H-6-12</article-title>. <source>Genome Announc.</source> <volume>2</volume>:<fpage>e00109</fpage>&#x02013;<lpage>14</lpage>. <pub-id pub-id-type="doi">10.1128/genomeA.00109-14</pub-id><pub-id pub-id-type="pmid">24558247</pub-id></citation>
</ref>
<ref id="B12">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dehal</surname> <given-names>P. S.</given-names></name> <name><surname>Joachimiak</surname> <given-names>M. P.</given-names></name> <name><surname>Price</surname> <given-names>M. N.</given-names></name> <name><surname>Bates</surname> <given-names>J. T.</given-names></name> <name><surname>Baumohl</surname> <given-names>J. K.</given-names></name> <name><surname>Chivian</surname> <given-names>D.</given-names></name> <etal/></person-group>. (<year>2010</year>). <article-title>MicrobesOnline: an integrated portal for comparative and functional genomics</article-title>. <source>Nucleic Acids Res.</source> <volume>38</volume>, <fpage>D396</fpage>&#x02013;<lpage>D400</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gkp919</pub-id><pub-id pub-id-type="pmid">19906701</pub-id></citation>
</ref>
<ref id="B13">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Delcher</surname> <given-names>A. L.</given-names></name> <name><surname>Salzberg</surname> <given-names>S. L.</given-names></name> <name><surname>Phillippy</surname> <given-names>A. M.</given-names></name></person-group> (<year>2003</year>). <article-title>Using MUMmer to identify similar regions in large sequence sets</article-title>. <source>Curr. Protoc. Bioinformatics</source> Chapter 10, Unit 10.13. <pub-id pub-id-type="doi">10.1002/0471250953.bi1003s00</pub-id><pub-id pub-id-type="pmid">18428693</pub-id></citation>
</ref>
<ref id="B14">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Devoid</surname> <given-names>S.</given-names></name> <name><surname>Overbeek</surname> <given-names>R.</given-names></name> <name><surname>DeJongh</surname> <given-names>M.</given-names></name> <name><surname>Vonstein</surname> <given-names>V.</given-names></name> <name><surname>Best</surname> <given-names>A. A.</given-names></name> <name><surname>Henry</surname> <given-names>C.</given-names></name></person-group> (<year>2013</year>). <article-title>Automated genome annotation and metabolic model reconstruction in the SEED and Model SEED</article-title>. <source>Methods Mol. Biol.</source> <volume>985</volume>, <fpage>17</fpage>&#x02013;<lpage>45</lpage>. <pub-id pub-id-type="doi">10.1007/978-1-62703-299-5_2</pub-id><pub-id pub-id-type="pmid">23417797</pub-id></citation>
</ref>
<ref id="B15">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ding</surname> <given-names>Y. H.</given-names></name> <name><surname>Ronimus</surname> <given-names>R. S.</given-names></name> <name><surname>Morgan</surname> <given-names>H. W.</given-names></name></person-group> (<year>2000</year>). <article-title>Sequencing, cloning, and high-level expression of the pfp gene, encoding a PP(i)-dependent phosphofructokinase from the extremely thermophilic eubacterium <italic>Dictyoglomus thermophilum</italic></article-title>. <source>J. Bacteriol.</source> <volume>182</volume>, <fpage>4661</fpage>&#x02013;<lpage>4666</lpage>. <pub-id pub-id-type="doi">10.1128/JB.182.16.4661-4666.2000</pub-id><pub-id pub-id-type="pmid">10913106</pub-id></citation>
</ref>
<ref id="B16">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Edgar</surname> <given-names>R. C.</given-names></name></person-group> (<year>2004</year>). <article-title>MUSCLE: multiple sequence alignment with high accuracy and high throughput</article-title>. <source>Nucleic Acids Res.</source> <volume>32</volume>, <fpage>1792</fpage>&#x02013;<lpage>1797</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gkh340</pub-id><pub-id pub-id-type="pmid">15034147</pub-id></citation>
</ref>
<ref id="B17">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Euz&#x000E9;by</surname> <given-names>J.</given-names></name></person-group> (<year>2012</year>). <article-title>List of new names and new combinations previously effectively, but not validly, published</article-title>. <source>Int. J. Syst. Evol. Microbiol.</source> <volume>62</volume>, <fpage>1</fpage>&#x02013;<lpage>4</lpage>. <pub-id pub-id-type="doi">10.1099/ijs.0.039487-0</pub-id></citation>
</ref>
<ref id="B18">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Euz&#x000E9;by</surname> <given-names>J. P.</given-names></name></person-group> (<year>1998</year>). <article-title>Taxonomic note: necessary correction of specific and subspecific epithets according to rules 12c and 13b of the international code of nomenclature of bacteria (1990 Revision)</article-title>. <source>Int. J. Syst. Evol. Microbiol.</source> <volume>48</volume>, <fpage>1073</fpage>&#x02013;<lpage>1075</lpage>. <pub-id pub-id-type="doi">10.1099/00207713-48-3-1073</pub-id></citation>
</ref>
<ref id="B19">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ewing</surname> <given-names>B.</given-names></name> <name><surname>Green</surname> <given-names>P.</given-names></name></person-group> (<year>1998</year>). <article-title>Base-calling of automated sequencer traces using phred</article-title>. <source>II. Error probabilities. Genome Res.</source> <volume>8</volume>, <fpage>186</fpage>&#x02013;<lpage>194</lpage>. <pub-id pub-id-type="doi">10.1101/gr.8.3.186</pub-id><pub-id pub-id-type="pmid">9521922</pub-id></citation>
</ref>
<ref id="B20">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fukusumi</surname> <given-names>S.</given-names></name> <name><surname>Kamizono</surname> <given-names>A.</given-names></name> <name><surname>Horinouchi</surname> <given-names>S.</given-names></name> <name><surname>Beppu</surname> <given-names>T.</given-names></name></person-group> (<year>1988</year>). <article-title>Cloning and nucleotide sequence of a heat-stable amylase gene from an anaerobic thermophile, <italic>Dictyoglomus thermophilum</italic></article-title>. <source>Eur. J. Biochem.</source> <volume>174</volume>, <fpage>15</fpage>&#x02013;<lpage>21</lpage>. <pub-id pub-id-type="doi">10.1111/j.1432-1033.1988.tb14056.x</pub-id><pub-id pub-id-type="pmid">2453362</pub-id></citation>
</ref>
<ref id="B21">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gao</surname> <given-names>D.</given-names></name> <name><surname>Uppugundla</surname> <given-names>N.</given-names></name> <name><surname>Chundawat</surname> <given-names>S. P.</given-names></name> <name><surname>Yu</surname> <given-names>X.</given-names></name> <name><surname>Hermanson</surname> <given-names>S.</given-names></name> <name><surname>Gowda</surname> <given-names>K.</given-names></name> <etal/></person-group>. (<year>2011</year>). <article-title>Hemicellulases and auxiliary enzymes for improved conversion of lignocellulosic biomass to monosaccharides</article-title>. <source>Biotechnol. Biofuels</source> <volume>4</volume>:<fpage>5</fpage>. <pub-id pub-id-type="doi">10.1186/1754-6834-4-5</pub-id><pub-id pub-id-type="pmid">21342516</pub-id></citation>
</ref>
<ref id="B22">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gibbs</surname> <given-names>M. D.</given-names></name> <name><surname>Reeves</surname> <given-names>R. A.</given-names></name> <name><surname>Bergquist</surname> <given-names>P. L.</given-names></name></person-group> (<year>1995</year>). <article-title>Cloning, sequencing, and expression of a xylanase gene from the extreme thermophile <italic>Dictyoglomus thermophilum</italic> Rt46B.1 and activity of the enzyme on fiber-bound substrate</article-title>. <source>Appl. Environ. Microbiol.</source> <volume>61</volume>, <fpage>4403</fpage>&#x02013;<lpage>4408</lpage>. <pub-id pub-id-type="pmid">8534104</pub-id></citation>
</ref>
<ref id="B23">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gibbs</surname> <given-names>M. D.</given-names></name> <name><surname>Reeves</surname> <given-names>R. A.</given-names></name> <name><surname>Sunna</surname> <given-names>A.</given-names></name> <name><surname>Bergquist</surname> <given-names>P. L.</given-names></name></person-group> (<year>1999</year>). <article-title>Sequencing and expression of a &#x003B2;-mannanase gene from the extreme thermophile <italic>Dictyoglomus thermophilum</italic> Rt46B.1, and characteristics of the recombinant enzyme</article-title>. <source>Curr. Microbiol.</source> <volume>39</volume>, <fpage>351</fpage>&#x02013;<lpage>0357</lpage>. <pub-id pub-id-type="pmid">10525841</pub-id></citation>
</ref>
<ref id="B24">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gordon</surname> <given-names>D.</given-names></name> <name><surname>Abajian</surname> <given-names>C.</given-names></name> <name><surname>Green</surname> <given-names>P.</given-names></name></person-group> (<year>1998</year>). <article-title>Consed: a graphical tool for sequence finishing</article-title>. <source>Genome Res.</source> <volume>8</volume>, <fpage>195</fpage>&#x02013;<lpage>202</lpage>. <pub-id pub-id-type="doi">10.1101/gr.8.3.195</pub-id><pub-id pub-id-type="pmid">9521923</pub-id></citation>
</ref>
<ref id="B25">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Griffiths-Jones</surname> <given-names>S.</given-names></name> <name><surname>Bateman</surname> <given-names>A.</given-names></name> <name><surname>Marshall</surname> <given-names>M.</given-names></name> <name><surname>Khanna</surname> <given-names>A.</given-names></name> <name><surname>Eddy</surname> <given-names>S. R.</given-names></name></person-group> (<year>2003</year>). <article-title>Rfam: an RNA family database</article-title>. <source>Nucleic Acids Res.</source> <volume>31</volume>, <fpage>439</fpage>&#x02013;<lpage>441</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gkg006</pub-id><pub-id pub-id-type="pmid">12520045</pub-id></citation>
</ref>
<ref id="B26">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Grissa</surname> <given-names>I.</given-names></name> <name><surname>Vergnaud</surname> <given-names>G.</given-names></name> <name><surname>Pourcel</surname> <given-names>C.</given-names></name></person-group> (<year>2007</year>). <article-title>CRISPRFinder: a web tool to identify clustered regularly interspaced short palindromic repeats</article-title>. <source>Nucleic Acids Res.</source> <volume>35</volume>, <fpage>W52</fpage>&#x02013;<lpage>W57</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gkm360</pub-id><pub-id pub-id-type="pmid">17537822</pub-id></citation>
</ref>
<ref id="B27">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gumerov</surname> <given-names>V. M.</given-names></name> <name><surname>Mardanov</surname> <given-names>A. V.</given-names></name> <name><surname>Beletskii</surname> <given-names>A. V.</given-names></name> <name><surname>Bonch-Osmolovskaia</surname> <given-names>E. A.</given-names></name> <name><surname>Ravin</surname> <given-names>N. V.</given-names></name></person-group> (<year>2011</year>). <article-title>[Molecular analysis of microbial diversity in the Zavarzin Spring, the Uzon caldera]</article-title>. <source>Mikrobiologiia</source> <volume>80</volume>, <fpage>258</fpage>&#x02013;<lpage>265</lpage>. <pub-id pub-id-type="doi">10.1134/s002626171102007x</pub-id><pub-id pub-id-type="pmid">21774190</pub-id></citation>
</ref>
<ref id="B28">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Haas</surname> <given-names>B. J.</given-names></name> <name><surname>Sandigursky</surname> <given-names>M.</given-names></name> <name><surname>Tainer</surname> <given-names>J. A.</given-names></name> <name><surname>Franklin</surname> <given-names>W. A.</given-names></name> <name><surname>Cunningham</surname> <given-names>R. P.</given-names></name></person-group> (<year>1999</year>). <article-title>Purification and characterization of <italic>Thermotoga maritima</italic> endonuclease IV, a thermostable apurinic/apyrimidinic endonuclease and 3&#x02032;-repair diesterase</article-title>. <source>J. Bacteriol.</source> <volume>181</volume>, <fpage>2834</fpage>&#x02013;<lpage>2839</lpage>. <pub-id pub-id-type="pmid">10217775</pub-id></citation>
</ref>
<ref id="B29">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hong</surname> <given-names>S. H.</given-names></name> <name><surname>Lim</surname> <given-names>Y. R.</given-names></name> <name><surname>Kim</surname> <given-names>Y. S.</given-names></name> <name><surname>Oh</surname> <given-names>D. K.</given-names></name></person-group> (<year>2012</year>). <article-title>Molecular characterization of a thermostable L-fucose isomerase from <italic>Dictyoglomus turgidum</italic> that isomerizes L-fucose and D-arabinose</article-title>. <source>Biochimie</source> <volume>94</volume>, <fpage>1926</fpage>&#x02013;<lpage>1934</lpage>. <pub-id pub-id-type="doi">10.1016/j.biochi.2012.05.009</pub-id><pub-id pub-id-type="pmid">22627384</pub-id></citation>
</ref>
<ref id="B30">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Horinouchi</surname> <given-names>S.</given-names></name> <name><surname>Fukusumi</surname> <given-names>S.</given-names></name> <name><surname>Ohshima</surname> <given-names>T.</given-names></name> <name><surname>Beppu</surname> <given-names>T.</given-names></name></person-group> (<year>1988</year>). <article-title>Cloning and expression in <italic>Escherichia coli</italic> of two additional amylase genes of a strictly anaerobic thermophile, <italic>Dictyoglomus thermophilum</italic>, and their nucleotide sequences with extremely low guanine-plus-cytosine contents</article-title>. <source>Eur. J. Biochem.</source> <volume>176</volume>, <fpage>243</fpage>&#x02013;<lpage>253</lpage>. <pub-id pub-id-type="doi">10.1111/j.1432-1033.1988.tb14275.x</pub-id><pub-id pub-id-type="pmid">2458257</pub-id></citation>
</ref>
<ref id="B31">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hsiao</surname> <given-names>W.</given-names></name> <name><surname>Wan</surname> <given-names>I.</given-names></name> <name><surname>Jones</surname> <given-names>S. J.</given-names></name> <name><surname>Brinkman</surname> <given-names>F. S.</given-names></name></person-group> (<year>2003</year>). <article-title>IslandPath: aiding detection of genomic islands in prokaryotes</article-title>. <source>Bioinformatics</source> <volume>19</volume>, <fpage>418</fpage>&#x02013;<lpage>420</lpage>. <pub-id pub-id-type="doi">10.1093/bioinformatics/btg004</pub-id><pub-id pub-id-type="pmid">12584130</pub-id></citation>
</ref>
<ref id="B32">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Huber</surname> <given-names>R.</given-names></name> <name><surname>Langworthy</surname> <given-names>T. A.</given-names></name> <name><surname>K&#x000F6;nig</surname> <given-names>H.</given-names></name> <name><surname>Thomm</surname> <given-names>M.</given-names></name> <name><surname>Woese</surname> <given-names>C. R.</given-names></name> <name><surname>Sleytr</surname> <given-names>U. B.</given-names></name> <etal/></person-group>. (<year>1986</year>). <article-title><italic>Thermotoga maritima</italic> sp. nov. represents a new genus of unique extremely thermophilic eubacteria growing up to 90&#x000B0;C</article-title>. <source>Arch. Microbiol.</source> <volume>144</volume>, <fpage>324</fpage>&#x02013;<lpage>333</lpage>.</citation>
</ref>
<ref id="B33">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hubisz</surname> <given-names>M. J.</given-names></name> <name><surname>Pollard</surname> <given-names>K. S.</given-names></name> <name><surname>Siepel</surname> <given-names>A.</given-names></name></person-group> (<year>2011</year>). <article-title>PHAST and RPHAST: phylogenetic analysis with space/time models</article-title>. <source>Brief. Bioinformatics</source> <volume>12</volume>, <fpage>41</fpage>&#x02013;<lpage>51</lpage>. <pub-id pub-id-type="doi">10.1093/bib/bbq072</pub-id><pub-id pub-id-type="pmid">21278375</pub-id></citation>
</ref>
<ref id="B34">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hyatt</surname> <given-names>D.</given-names></name> <name><surname>Chen</surname> <given-names>G. L.</given-names></name> <name><surname>Locascio</surname> <given-names>P. F.</given-names></name> <name><surname>Land</surname> <given-names>M. L.</given-names></name> <name><surname>Larimer</surname> <given-names>F. W.</given-names></name> <name><surname>Hauser</surname> <given-names>L. J.</given-names></name></person-group> (<year>2010</year>). <article-title>Prodigal prokaryotic dynamic programming genefinding algorithm</article-title>. <source>BMC Bioinformatics</source> <volume>11</volume>:<fpage>119</fpage>. <pub-id pub-id-type="doi">10.1186/1471-2105-11-119</pub-id><pub-id pub-id-type="pmid">20211023</pub-id></citation>
</ref>
<ref id="B35">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ishino</surname> <given-names>Y.</given-names></name> <name><surname>Narumi</surname> <given-names>I.</given-names></name></person-group> (<year>2015</year>). <article-title>DNA repair in hyperthermophilic and hyperradioresistant microorganisms</article-title>. <source>Curr. Opin. Microbiol.</source> <volume>25</volume>, <fpage>103</fpage>&#x02013;<lpage>112</lpage>. <pub-id pub-id-type="doi">10.1016/j.mib.2015.05.010</pub-id><pub-id pub-id-type="pmid">26056771</pub-id></citation>
</ref>
<ref id="B36">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jernigan</surname> <given-names>K. K.</given-names></name> <name><surname>Bordenstein</surname> <given-names>S. R.</given-names></name></person-group> (<year>2015</year>). <article-title>Tandem-repeat protein domains across the tree of life</article-title>. <source>Peer J.</source> <volume>3</volume>:<fpage>e732</fpage>. <pub-id pub-id-type="doi">10.7717/peerj.732</pub-id><pub-id pub-id-type="pmid">25653910</pub-id></citation>
</ref>
<ref id="B37">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Karp</surname> <given-names>P. D.</given-names></name> <name><surname>Ouzounis</surname> <given-names>C. A.</given-names></name> <name><surname>Moore-Kochlacs</surname> <given-names>C.</given-names></name> <name><surname>Goldovsky</surname> <given-names>L.</given-names></name> <name><surname>Kaipa</surname> <given-names>P.</given-names></name> <name><surname>Ahr&#x000E9;n</surname> <given-names>D.</given-names></name> <etal/></person-group>. (<year>2005</year>). <article-title>Expansion of the BioCyc collection of pathway/genome databases to 160 genomes</article-title>. <source>Nucleic Acids Res.</source> <volume>33</volume>, <fpage>6083</fpage>&#x02013;<lpage>6089</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gki892</pub-id><pub-id pub-id-type="pmid">16246909</pub-id></citation>
</ref>
<ref id="B38">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Katoh</surname> <given-names>K.</given-names></name> <name><surname>Misawa</surname> <given-names>K.</given-names></name> <name><surname>Kuma</surname> <given-names>K.</given-names></name> <name><surname>Miyata</surname> <given-names>T.</given-names></name></person-group> (<year>2002</year>). <article-title>MAFFT: a novel method for rapid multiple sequence alignment based on fast Fourier transform</article-title>. <source>Nucleic Acids Res.</source> <volume>30</volume>, <fpage>3059</fpage>&#x02013;<lpage>3066</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gkf436</pub-id><pub-id pub-id-type="pmid">12136088</pub-id></citation>
</ref>
<ref id="B39">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kelman</surname> <given-names>L. M.</given-names></name> <name><surname>Kelman</surname> <given-names>Z.</given-names></name></person-group> (<year>2014</year>). <article-title>Archaeal DNA replication</article-title>. <source>Annu. Rev. Genet.</source> <volume>48</volume>, <fpage>71</fpage>&#x02013;<lpage>97</lpage>. <pub-id pub-id-type="doi">10.1146/annurev-genet-120213-092148</pub-id><pub-id pub-id-type="pmid">25421597</pub-id></citation>
</ref>
<ref id="B40">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kim</surname> <given-names>J. E.</given-names></name> <name><surname>Kim</surname> <given-names>Y. S.</given-names></name> <name><surname>Kang</surname> <given-names>L. W.</given-names></name> <name><surname>Oh</surname> <given-names>D. K.</given-names></name></person-group> (<year>2012</year>). <article-title>Characterization of a recombinant cellobiose 2-epimerase from <italic>Dictyoglomus turgidum</italic> that epimerizes and isomerizes &#x003B2;-1,4- and &#x003B1;-1,4-gluco-oligosaccharides</article-title>. <source>Biotechnol. Lett.</source> <volume>34</volume>, <fpage>2061</fpage>&#x02013;<lpage>2068</lpage>. <pub-id pub-id-type="doi">10.1007/s10529-012-0999-z</pub-id><pub-id pub-id-type="pmid">22782272</pub-id></citation>
</ref>
<ref id="B41">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kim</surname> <given-names>M.</given-names></name> <name><surname>Oh</surname> <given-names>H. S.</given-names></name> <name><surname>Park</surname> <given-names>S. C.</given-names></name> <name><surname>Chun</surname> <given-names>J.</given-names></name></person-group> (<year>2014</year>). <article-title>Towards a taxonomic coherence between average nucleotide identity and 16S rRNA gene sequence similarity for species demarcation of prokaryotes</article-title>. <source>Int. J. Syst. Evol. Microbiol.</source> <volume>64</volume>(<issue>Pt 2</issue>), <fpage>346</fpage>&#x02013;<lpage>351</lpage>. <pub-id pub-id-type="doi">10.1099/ijs.0.059774-0</pub-id><pub-id pub-id-type="pmid">24505072</pub-id></citation>
</ref>
<ref id="B42">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kim</surname> <given-names>Y. S.</given-names></name> <name><surname>Shin</surname> <given-names>K. C.</given-names></name> <name><surname>Lim</surname> <given-names>Y. R.</given-names></name> <name><surname>Oh</surname> <given-names>D. K.</given-names></name></person-group> (<year>2013</year>). <article-title>Characterization of a recombinant L-rhamnose isomerase from <italic>Dictyoglomus turgidum</italic> and its application for L-rhamnulose production</article-title>. <source>Biotechnol. Lett.</source> <volume>35</volume>, <fpage>259</fpage>&#x02013;<lpage>264</lpage>. <pub-id pub-id-type="doi">10.1007/s10529-012-1069-2</pub-id><pub-id pub-id-type="pmid">23070627</pub-id></citation>
</ref>
<ref id="B43">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kochetkova</surname> <given-names>T. V.</given-names></name> <name><surname>Rusanov</surname> <given-names>I. I.</given-names></name> <name><surname>Pimenov</surname> <given-names>N. V.</given-names></name> <name><surname>Kolganova</surname> <given-names>T. V.</given-names></name> <name><surname>Lebedinsky</surname> <given-names>A. V.</given-names></name> <name><surname>Bonch-Osmolovskaya</surname> <given-names>E. A.</given-names></name> <etal/></person-group>. (<year>2011</year>). <article-title>Anaerobic transformation of carbon monoxide by microbial communities of Kamchatka hot springs</article-title>. <source>Extremophiles</source> <volume>15</volume>, <fpage>319</fpage>&#x02013;<lpage>325</lpage>. <pub-id pub-id-type="doi">10.1007/s00792-011-0362-7</pub-id><pub-id pub-id-type="pmid">21387195</pub-id></citation>
</ref>
<ref id="B44">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Krogh</surname> <given-names>A.</given-names></name> <name><surname>Larsson</surname> <given-names>B.</given-names></name> <name><surname>von Heijne</surname> <given-names>G.</given-names></name> <name><surname>Sonnhammer</surname> <given-names>E. L.</given-names></name></person-group> (<year>2001</year>). <article-title>Predicting transmembrane protein topology with a hidden Markov model: application to complete genomes</article-title>. <source>J. Mol. Biol.</source> <volume>305</volume>, <fpage>567</fpage>&#x02013;<lpage>580</lpage>. <pub-id pub-id-type="doi">10.1006/jmbi.2000.4315</pub-id><pub-id pub-id-type="pmid">11152613</pub-id></citation>
</ref>
<ref id="B45">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kublanov</surname> <given-names>I. V.</given-names></name> <name><surname>Perevalova</surname> <given-names>A. A.</given-names></name> <name><surname>Slobodkina</surname> <given-names>G. B.</given-names></name> <name><surname>Lebedinsky</surname> <given-names>A. V.</given-names></name> <name><surname>Bidzhieva</surname> <given-names>S. K.</given-names></name> <name><surname>Kolganova</surname> <given-names>T. V.</given-names></name> <etal/></person-group>. (<year>2009</year>). <article-title>Biodiversity of thermophilic prokaryotes with hydrolytic activities in hot springs of Uzon Caldera, Kamchatka (Russia)</article-title>. <source>Appl. Environ. Microbiol.</source> <volume>75</volume>, <fpage>286</fpage>&#x02013;<lpage>291</lpage>. <pub-id pub-id-type="doi">10.1128/AEM.00607-08</pub-id><pub-id pub-id-type="pmid">18978089</pub-id></citation>
</ref>
<ref id="B46">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kumar</surname> <given-names>S.</given-names></name> <name><surname>Stecher</surname> <given-names>G.</given-names></name> <name><surname>Tamura</surname> <given-names>K.</given-names></name></person-group> (<year>2016</year>). <article-title>MEGA7: Molecular evolutionary genetics analysis version 7.0 for bigger datasets</article-title>. <source>Mol. Biol. Evol.</source> <volume>33</volume>, <fpage>1870</fpage>&#x02013;<lpage>1874</lpage>. <pub-id pub-id-type="doi">10.1093/molbev/msw054</pub-id><pub-id pub-id-type="pmid">27004904</pub-id></citation>
</ref>
<ref id="B47">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lagesen</surname> <given-names>K.</given-names></name> <name><surname>Hallin</surname> <given-names>P.</given-names></name> <name><surname>R&#x000F8;dland</surname> <given-names>E. A.</given-names></name> <name><surname>Staerfeldt</surname> <given-names>H. H.</given-names></name> <name><surname>Rognes</surname> <given-names>T.</given-names></name> <name><surname>Ussery</surname> <given-names>D. W.</given-names></name></person-group> (<year>2007</year>). <article-title>RNAmmer: consistent and rapid annotation of ribosomal RNA genes</article-title>. <source>Nucleic Acids Res.</source> <volume>35</volume>, <fpage>3100</fpage>&#x02013;<lpage>3108</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gkm160</pub-id><pub-id pub-id-type="pmid">17452365</pub-id></citation>
</ref>
<ref id="B48">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lombard</surname> <given-names>V.</given-names></name> <name><surname>Golaconda Ramulu</surname> <given-names>H.</given-names></name> <name><surname>Drula</surname> <given-names>E.</given-names></name> <name><surname>Coutinho</surname> <given-names>P. M.</given-names></name> <name><surname>Henrissat</surname> <given-names>B.</given-names></name></person-group> (<year>2014</year>). <article-title>The carbohydrate-active enzymes database (CAZy) in 2013</article-title>. <source>Nucleic Acids Res.</source> <volume>42</volume>, <fpage>D490</fpage>&#x02013;<lpage>D495</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gkt1178</pub-id><pub-id pub-id-type="pmid">24270786</pub-id></citation>
</ref>
<ref id="B49">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Love</surname> <given-names>C. A.</given-names></name> <name><surname>Patel</surname> <given-names>B. K. C.</given-names></name> <name><surname>Ludwig</surname> <given-names>W.</given-names></name> <name><surname>Stackebrandt</surname> <given-names>E.</given-names></name></person-group> (<year>1993</year>). <article-title>The phylogenetic position of <italic>Dictyoglomus thermophilum</italic> based on 16S rRNA sequence analysis</article-title>. <source>FEMS Microbiol. Lett.</source> <volume>107</volume>, <fpage>317</fpage>&#x02013;<lpage>320</lpage>. <pub-id pub-id-type="doi">10.1111/j.1574-6968.1993.tb06050.x</pub-id></citation>
</ref>
<ref id="B50">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lowe</surname> <given-names>T. M.</given-names></name> <name><surname>Eddy</surname> <given-names>S. R.</given-names></name></person-group> (<year>1997</year>). <article-title>tRNAscan-SE: a program for improved detection of transfer RNA genes in genomic sequence</article-title>. <source>Nucleic Acids Res.</source> <volume>25</volume>, <fpage>955</fpage>&#x02013;<lpage>964</lpage>. <pub-id pub-id-type="doi">10.1093/nar/25.5.0955</pub-id><pub-id pub-id-type="pmid">9023104</pub-id></citation>
</ref>
<ref id="B51">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mathrani</surname> <given-names>I.</given-names></name> <name><surname>Ahring</surname> <given-names>B.</given-names></name></person-group> (<year>1991</year>). <article-title>Isolation and characterization of a strictly xylan-degrading Dictyoglomus from a man-made, thermophilic anaerobic environment</article-title>. <source>Arch. Microbiol.</source> <volume>157</volume>, <fpage>13</fpage>&#x02013;<lpage>17</lpage>. <pub-id pub-id-type="doi">10.1007/BF00245328</pub-id></citation>
</ref>
<ref id="B52">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mathrani</surname> <given-names>I.</given-names></name> <name><surname>Ahring</surname> <given-names>B.</given-names></name></person-group> (<year>1992</year>). <article-title>Thermophilic and alkalophilic xylanases from several Dictyoglomus isolates</article-title>. <source>Appl. Microbiol. Biotechnol.</source> <volume>38</volume>, <fpage>23</fpage>&#x02013;<lpage>27</lpage>. <pub-id pub-id-type="doi">10.1007/BF00169413</pub-id></citation>
</ref>
<ref id="B53">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Menzel</surname> <given-names>P.</given-names></name> <name><surname>Gudbergsd&#x000F3;ttir</surname> <given-names>S. R.</given-names></name> <name><surname>Rike</surname> <given-names>A. G.</given-names></name> <name><surname>Lin</surname> <given-names>L.</given-names></name> <name><surname>Zhang</surname> <given-names>Q.</given-names></name> <name><surname>Contursi</surname> <given-names>P.</given-names></name> <etal/></person-group>. (<year>2015</year>). <article-title>Comparative metagenomics of eight geographically remote terrestrial hot springs</article-title>. <source>Microb. Ecol.</source> <volume>70</volume>, <fpage>411</fpage>&#x02013;<lpage>424</lpage>. <pub-id pub-id-type="doi">10.1007/s00248-015-0576-9</pub-id><pub-id pub-id-type="pmid">25712554</pub-id></citation>
</ref>
<ref id="B54">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Morris</surname> <given-names>D. D.</given-names></name> <name><surname>Gibbs</surname> <given-names>M. D.</given-names></name> <name><surname>Chin</surname> <given-names>C. W.</given-names></name> <name><surname>Koh</surname> <given-names>M. H.</given-names></name> <name><surname>Wong</surname> <given-names>K. K.</given-names></name> <name><surname>Allison</surname> <given-names>R. W.</given-names></name> <etal/></person-group>. (<year>1998</year>). <article-title>Cloning of the xynB gene from <italic>Dictyoglomus thermophilum</italic> Rt46B.1 and action of the gene product on kraft pulp</article-title>. <source>Appl. Environ. Microbiol.</source> <volume>64</volume>, <fpage>1759</fpage>&#x02013;<lpage>1765</lpage>. <pub-id pub-id-type="pmid">9572948</pub-id></citation>
</ref>
<ref id="B55">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Nishida</surname> <given-names>H.</given-names></name> <name><surname>Beppu</surname> <given-names>T.</given-names></name> <name><surname>Ueda</surname> <given-names>K.</given-names></name></person-group> (<year>2011</year>). <article-title>Whole-genome comparison clarifies close phylogenetic relationships between the phyla Dictyoglomi and Thermotogae</article-title>. <source>Genomics</source> <volume>98</volume>, <fpage>370</fpage>&#x02013;<lpage>375</lpage>. <pub-id pub-id-type="doi">10.1016/j.ygeno.2011.08.001</pub-id><pub-id pub-id-type="pmid">21851855</pub-id></citation>
</ref>
<ref id="B56">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Overbeek</surname> <given-names>R.</given-names></name> <name><surname>Begley</surname> <given-names>T.</given-names></name> <name><surname>Butler</surname> <given-names>R. M.</given-names></name> <name><surname>Choudhuri</surname> <given-names>J. V.</given-names></name> <name><surname>Chuang</surname> <given-names>H. Y.</given-names></name> <name><surname>Cohoon</surname> <given-names>M.</given-names></name> <etal/></person-group>. (<year>2005</year>). <article-title>The subsystems approach to genome annotation and its use in the project to annotate 1000 genomes</article-title>. <source>Nucleic Acids Res.</source> <volume>33</volume>, <fpage>5691</fpage>&#x02013;<lpage>5702</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gki866</pub-id><pub-id pub-id-type="pmid">16214803</pub-id></citation>
</ref>
<ref id="B57">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Patel</surname> <given-names>B. K.</given-names></name> <name><surname>Morgan</surname> <given-names>H. W.</given-names></name> <name><surname>Wiegel</surname> <given-names>J.</given-names></name> <name><surname>Daniel</surname> <given-names>R. M.</given-names></name></person-group> (<year>1987</year>). <article-title>Isolation of an extremely thermophilic chemoorganotrophic anaerobe similar to <italic>Dictyoglomus thermophilum</italic> from New Zealand hot springs</article-title>. <source>Arch. Microbiol.</source> <volume>147</volume>, <fpage>21</fpage>&#x02013;<lpage>24</lpage>. <pub-id pub-id-type="doi">10.1007/BF00492899</pub-id></citation>
</ref>
<ref id="B58">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Petersen</surname> <given-names>T. N.</given-names></name> <name><surname>Brunak</surname> <given-names>S.</given-names></name> <name><surname>von Heijne</surname> <given-names>G.</given-names></name> <name><surname>Nielsen</surname> <given-names>H.</given-names></name></person-group> (<year>2011</year>). <article-title>SignalP 4.0: discriminating signal peptides from transmembrane regions</article-title>. <source>Nat. Methods</source> <volume>8</volume>, <fpage>785</fpage>&#x02013;<lpage>786</lpage>. <pub-id pub-id-type="doi">10.1038/nmeth.1701</pub-id><pub-id pub-id-type="pmid">21959131</pub-id></citation>
</ref>
<ref id="B59">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Rawlings</surname> <given-names>N. D.</given-names></name> <name><surname>Barrett</surname> <given-names>A. J.</given-names></name> <name><surname>Bateman</surname> <given-names>A.</given-names></name></person-group> (<year>2014</year>). <article-title>Using the MEROPS database for proteolytic enzymes and their inhibitors and substrates</article-title>. <source>Curr. Protoc. Bioinformatics.</source> <volume>48</volume>, <fpage>1.25.1</fpage>&#x02013;<lpage>33</lpage>. <pub-id pub-id-type="doi">10.1002/0471250953.bi0125s48</pub-id><pub-id pub-id-type="pmid">25501939</pub-id></citation>
</ref>
<ref id="B60">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Rees</surname> <given-names>G. N.</given-names></name> <name><surname>Patel</surname> <given-names>B. K.</given-names></name> <name><surname>Grassia</surname> <given-names>G. S.</given-names></name> <name><surname>Sheehy</surname> <given-names>A. J.</given-names></name></person-group> (<year>1997</year>). <article-title>Anaerobaculum thermoterrenum gen. nov., sp. nov., a novel, thermophilic bacterium which ferments citrate</article-title>. <source>Int. J. Syst. Bacteriol.</source> <volume>47</volume>, <fpage>150</fpage>&#x02013;<lpage>154</lpage>. <pub-id pub-id-type="doi">10.1099/00207713-47-1-150</pub-id><pub-id pub-id-type="pmid">8995817</pub-id></citation>
</ref>
<ref id="B61">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Rigden</surname> <given-names>D. J.</given-names></name> <name><surname>Galperin</surname> <given-names>M. Y.</given-names></name></person-group> (<year>2008</year>). <article-title>Sequence analysis of GerM and SpoVS, uncharacterized bacterial &#x02018;sporulation&#x02019; proteins with widespread phylogenetic distribution</article-title>. <source>Bioinformatics</source> <volume>24</volume>, <fpage>1793</fpage>&#x02013;<lpage>1797</lpage>. <pub-id pub-id-type="doi">10.1093/bioinformatics/btn314</pub-id><pub-id pub-id-type="pmid">18562273</pub-id></citation>
</ref>
<ref id="B62">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sahm</surname> <given-names>K.</given-names></name> <name><surname>John</surname> <given-names>P.</given-names></name> <name><surname>Nacke</surname> <given-names>H.</given-names></name> <name><surname>Wemheuer</surname> <given-names>B.</given-names></name> <name><surname>Grote</surname> <given-names>R.</given-names></name> <name><surname>Daniel</surname> <given-names>R.</given-names></name> <etal/></person-group>. (<year>2013</year>). <article-title>High abundance of heterotrophic prokaryotes in hydrothermal springs of the azores as revealed by a network of 16S rRNA gene-based methods</article-title>. <source>Extremophiles</source> <volume>17</volume>, <fpage>649</fpage>&#x02013;<lpage>662</lpage>. <pub-id pub-id-type="doi">10.1007/s00792-013-0548-2</pub-id><pub-id pub-id-type="pmid">23708551</pub-id></citation>
</ref>
<ref id="B63">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Saiki</surname> <given-names>T.</given-names></name> <name><surname>Kobayashi</surname> <given-names>Y.</given-names></name> <name><surname>Kawagoe</surname> <given-names>K.</given-names></name> <name><surname>Beppu</surname> <given-names>T.</given-names></name></person-group> (<year>1985</year>). <article-title><italic>Dictyoglomus thermophilum</italic> gen. nov., sp. nov., a Chemoorganotrophic, Anaerobic, Thermophilic Bacterium</article-title>. <source>Int. J. Syst. Evol. Microbiol.</source> <volume>35</volume>, <fpage>253</fpage>&#x02013;<lpage>259</lpage>.</citation>
</ref>
<ref id="B64">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Sambrook</surname> <given-names>J.</given-names></name> <name><surname>Fritsch</surname> <given-names>E. F.</given-names></name> <name><surname>Maniatis</surname> <given-names>T.</given-names></name></person-group> (<year>1989</year>). <source>Molecular Cloning: A Laboratory Manual</source>. <publisher-loc>Cold Spring Harbor, NY</publisher-loc>: <publisher-name>Cold Spring Harbor Laboratory Press</publisher-name>.</citation>
</ref>
<ref id="B65">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Shandilya</surname> <given-names>H.</given-names></name> <name><surname>Griffiths</surname> <given-names>K.</given-names></name> <name><surname>Flynn</surname> <given-names>E. K.</given-names></name> <name><surname>Astatke</surname> <given-names>M.</given-names></name> <name><surname>Shih</surname> <given-names>P. J.</given-names></name> <name><surname>Lee</surname> <given-names>J. E.</given-names></name> <etal/></person-group>. (<year>2004</year>). <article-title>Thermophilic bacterial DNA polymerases with reverse-transcriptase activity</article-title>. <source>Extremophiles</source> <volume>8</volume>, <fpage>243</fpage>&#x02013;<lpage>251</lpage>. <pub-id pub-id-type="doi">10.1007/s00792-004-0384-5</pub-id><pub-id pub-id-type="pmid">15197605</pub-id></citation>
</ref>
<ref id="B66">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Shi</surname> <given-names>R.</given-names></name> <name><surname>Li</surname> <given-names>Z.</given-names></name> <name><surname>Ye</surname> <given-names>Q.</given-names></name> <name><surname>Xu</surname> <given-names>J.</given-names></name> <name><surname>Liu</surname> <given-names>Y.</given-names></name></person-group> (<year>2013</year>). <article-title>Heterologous expression and characterization of a novel thermo-halotolerant endoglucanase Cel5H from <italic>Dictyoglomus thermophilum</italic></article-title>. <source>Bioresour. Technol.</source> <volume>142</volume>, <fpage>338</fpage>&#x02013;<lpage>344</lpage>. <pub-id pub-id-type="doi">10.1016/j.biortech.2013.05.037</pub-id><pub-id pub-id-type="pmid">23747445</pub-id></citation>
</ref>
<ref id="B67">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Spriestersbach</surname> <given-names>A.</given-names></name> <name><surname>Kubicek</surname> <given-names>J.</given-names></name> <name><surname>Sch&#x000E4;fer</surname> <given-names>F.</given-names></name> <name><surname>Block</surname> <given-names>H.</given-names></name> <name><surname>Maertens</surname> <given-names>B.</given-names></name></person-group> (<year>2015</year>). <article-title>Purification of his-tagged proteins</article-title>. <source>Meth. Enzymol.</source> <volume>559</volume>, <fpage>1</fpage>&#x02013;<lpage>15</lpage>. <pub-id pub-id-type="doi">10.1016/bs.mie.2014.11.003</pub-id><pub-id pub-id-type="pmid">26096499</pub-id></citation>
</ref>
<ref id="B68">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Svetlichny</surname> <given-names>V. A.</given-names></name> <name><surname>Svetlichnaya</surname> <given-names>T. P.</given-names></name></person-group> (<year>1988</year>). <article-title><italic>Dictyoglomus turgidus</italic> sp. nov., a new extremely thermophilic eubacterium isolated from hot springs of the Uzon volcano caldera</article-title>. <source>Mikrobiologiya</source> <volume>57</volume>, <fpage>435</fpage>&#x02013;<lpage>441</lpage>.</citation>
</ref>
<ref id="B69">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Takai</surname> <given-names>K.</given-names></name> <name><surname>Inoue</surname> <given-names>A.</given-names></name> <name><surname>Horikoshi</surname> <given-names>K.</given-names></name></person-group> (<year>1999</year>). <article-title><italic>Thermaerobacter marianensis</italic> gen. nov., sp. nov., an aerobic extremely thermophilic marine bacterium from the 11,000 m deep Mariana Trench</article-title>. <source>Int. J. Syst. Bacteriol</source>. <volume>49</volume>(<issue>Pt 2</issue>), <fpage>619</fpage>&#x02013;<lpage>628</lpage>. <pub-id pub-id-type="doi">10.1099/00207713-49-2-619</pub-id><pub-id pub-id-type="pmid">10319484</pub-id></citation>
</ref>
<ref id="B70">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tamura</surname> <given-names>K.</given-names></name> <name><surname>Nei</surname> <given-names>M.</given-names></name></person-group> (<year>1993</year>). <article-title>Estimation of the number of nucleotide substitutions in the control region of mitochondrial DNA in humans and chimpanzees</article-title>. <source>Mol. Biol. Evol.</source> <volume>10</volume>, <fpage>512</fpage>&#x02013;<lpage>526</lpage>. <pub-id pub-id-type="pmid">8336541</pub-id></citation>
</ref>
<ref id="B71">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tamura</surname> <given-names>K.</given-names></name> <name><surname>Peterson</surname> <given-names>D.</given-names></name> <name><surname>Peterson</surname> <given-names>N.</given-names></name> <name><surname>Stecher</surname> <given-names>G.</given-names></name> <name><surname>Nei</surname> <given-names>M.</given-names></name> <name><surname>Kumar</surname> <given-names>S.</given-names></name></person-group> (<year>2011</year>). <article-title>MEGA5: molecular evolutionary genetics analysis using maximum likelihood, evolutionary distance, and maximum parsimony methods</article-title>. <source>Mol. Biol. Evol.</source> <volume>28</volume>, <fpage>2731</fpage>&#x02013;<lpage>2739</lpage>. <pub-id pub-id-type="doi">10.1093/molbev/msr121</pub-id><pub-id pub-id-type="pmid">21546353</pub-id></citation>
</ref>
<ref id="B72">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Techtmann</surname> <given-names>S. M.</given-names></name> <name><surname>Colman</surname> <given-names>A. S.</given-names></name> <name><surname>Robb</surname> <given-names>F. T.</given-names></name></person-group> (<year>2009</year>). <article-title>That which does not kill us only makes us stronger: the role of carbon monoxide in thermophilic microbial consortia</article-title>. <source>Environ. Microbiol.</source> <volume>11</volume>, <fpage>1027</fpage>&#x02013;<lpage>1037</lpage>. <pub-id pub-id-type="doi">10.1111/j.1462-2920.2009.01865.x</pub-id><pub-id pub-id-type="pmid">19239487</pub-id></citation>
</ref>
<ref id="B73">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Vesth</surname> <given-names>T.</given-names></name> <name><surname>Ozen</surname> <given-names>A.</given-names></name> <name><surname>Andersen</surname> <given-names>S. C.</given-names></name> <name><surname>Kaas</surname> <given-names>R. S.</given-names></name> <name><surname>Lukjancenko</surname> <given-names>O.</given-names></name> <name><surname>Bohlin</surname> <given-names>J.</given-names></name> <etal/></person-group>. (<year>2013</year>). <article-title>Veillonella, firmicutes: microbes disguised as Gram negatives</article-title>. <source>Stand. Genomic Sci.</source> <volume>9</volume>, <fpage>431</fpage>&#x02013;<lpage>448</lpage>. <pub-id pub-id-type="doi">10.4056/sigs.2981345</pub-id><pub-id pub-id-type="pmid">24976898</pub-id></citation>
</ref>
<ref id="B74">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wagner</surname> <given-names>I. D.</given-names></name> <name><surname>Wiegel</surname> <given-names>J.</given-names></name></person-group> (<year>2008</year>). <article-title>Diversity of thermophilic anaerobes</article-title>. <source>Ann. N.Y. Acad. Sci.</source> <volume>1125</volume>, <fpage>1</fpage>&#x02013;<lpage>43</lpage>. <pub-id pub-id-type="doi">10.1196/annals.1419.029</pub-id><pub-id pub-id-type="pmid">18378585</pub-id></citation>
</ref>
<ref id="B75">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wijffels</surname> <given-names>G.</given-names></name> <name><surname>Dalrymple</surname> <given-names>B.</given-names></name> <name><surname>Kongsuwan</surname> <given-names>K.</given-names></name> <name><surname>Dixon</surname> <given-names>N. E.</given-names></name></person-group> (<year>2005</year>). <article-title>Conservation of eubacterial replicases</article-title>. <source>IUBMB Life</source> <volume>57</volume>, <fpage>413</fpage>&#x02013;<lpage>419</lpage>. <pub-id pub-id-type="doi">10.1080/15216540500138246</pub-id><pub-id pub-id-type="pmid">16012050</pub-id></citation>
</ref>
<ref id="B76">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zhou</surname> <given-names>Y.</given-names></name> <name><surname>Liang</surname> <given-names>Y.</given-names></name> <name><surname>Lynch</surname> <given-names>K. H.</given-names></name> <name><surname>Dennis</surname> <given-names>J. J.</given-names></name> <name><surname>Wishart</surname> <given-names>D. S.</given-names></name></person-group> (<year>2011</year>). <article-title>PHAST: a fast phage search tool</article-title>. <source>Nucl. Acids Res</source>. <volume>39</volume>, <fpage>W347</fpage>&#x02013;<lpage>W352</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gkr485</pub-id><pub-id pub-id-type="pmid">21672955</pub-id></citation>
</ref>
</ref-list>
</back>
</article>