<?xml version="1.0" encoding="UTF-8" standalone="no"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="review-article">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Microbiol.</journal-id>
<journal-title>Frontiers in Microbiology</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Microbiol.</abbrev-journal-title>
<issn pub-type="epub">1664-302X</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fmicb.2014.00774</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Microbiology</subject>
<subj-group>
<subject>Review Article</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Next-generation sequencing approach for connecting secondary metabolites to biosynthetic gene clusters in fungi</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author">
<name><surname>Cacho</surname> <given-names>Ralph A.</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<uri xlink:href="http://community.frontiersin.org/people/u/194385"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Tang</surname> <given-names>Yi</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="aff" rid="aff2"><sup>2</sup></xref>
<uri xlink:href="http://community.frontiersin.org/people/u/126937"/>
</contrib>
<contrib contrib-type="author" corresp="yes">
<name><surname>Chooi</surname> <given-names>Yit-Heng</given-names></name>
<xref ref-type="aff" rid="aff3"><sup>3</sup></xref>
<xref ref-type="author-notes" rid="fn002"><sup>&#x0002A;</sup></xref>
<uri xlink:href="http://community.frontiersin.org/people/u/162932"/>
</contrib>
</contrib-group>
<aff id="aff1"><sup>1</sup><institution>Chemical and Biomolecular Engineering Department, University of California</institution> <country>Los Angeles, Los Angeles, CA, USA</country></aff>
<aff id="aff2"><sup>2</sup><institution>Chemistry and Biochemistry Department, University of California</institution> <country>Los Angeles, Los Angeles, CA, USA</country></aff>
<aff id="aff3"><sup>3</sup><institution>Plant Sciences Division, Research School of Biology, The Australian National University</institution> <country>Canberra, ACT, Australia</country></aff>
<author-notes>
<fn fn-type="edited-by"><p>Edited by: <italic>Jonathan Palmer, United States Department of Agriculture &#x02013; Forest Service, USA</italic></p></fn>
<fn fn-type="edited-by"><p>Reviewed by: <italic>Philipp Wiemann, University of Wisconsin&#x02013;Madison, USA; Liliana Losada, J. Craig Venter Institute, USA</italic></p></fn>
<fn fn-type="corresp" id="fn002"><p>&#x0002A;Correspondence: <italic>Yit-Heng Chooi, Plant Sciences Division, Research School of Biology, The Australian National University, Linnaeus Way, Canberra, Acton, ACT 2601, Australia e-mail: <email>yh.chooi@anu.edu.au</email></italic></p></fn>
<fn fn-type="other" id="fn001"><p>This article was submitted to Microbial Physiology and Metabolism, a section of the journal Frontiers in Microbiology.</p></fn>
</author-notes>
<pub-date pub-type="epub">
<day>14</day>
<month>01</month>
<year>2015</year>
</pub-date>
<pub-date pub-type="collection">
<year>2014</year>
</pub-date>
<volume>5</volume>
<elocation-id>774</elocation-id>
<history>
<date date-type="received">
<day>20</day>
<month>10</month>
<year>2014</year>
</date>
<date date-type="accepted">
<day>17</day>
<month>12</month>
<year>2014</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x000A9; 2015 Cacho, Tang and Chooi.</copyright-statement>
<copyright-year>2015</copyright-year>
<license license-type="open-access" xlink:href="http://creativecommons.org/licenses/by/4.0/"><p> This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) or licensor are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p></license>
</permissions>
<abstract>
<p>Genomics has revolutionized the research on fungal secondary metabolite (SM) biosynthesis. To elucidate the molecular and enzymatic mechanisms underlying the biosynthesis of a specific SM compound, the important first step is often to find the genes that responsible for its synthesis. The accessibility to fungal genome sequences allows the bypass of the cumbersome traditional library construction and screening approach. The advance in next-generation sequencing (NGS) technologies have further improved the speed and reduced the cost of microbial genome sequencing in the past few years, which has accelerated the research in this field. Here, we will present an example work flow for identifying the gene cluster encoding the biosynthesis of SMs of interest using an NGS approach. We will also review the different strategies that can be employed to pinpoint the targeted gene clusters rapidly by giving several examples stemming from our work.</p>
</abstract>
<kwd-group>
<kwd>filamentous fungi</kwd>
<kwd>secondary metabolites</kwd>
<kwd>gene clusters</kwd>
<kwd>next generation sequencing</kwd>
<kwd>genome mining</kwd>
</kwd-group>
<counts>
<fig-count count="7"/>
<table-count count="1"/>
<equation-count count="0"/>
<ref-count count="140"/>
<page-count count="16"/>
<word-count count="0"/>
</counts>
</article-meta>
</front>
<body>
<sec><title>INTRODUCTION</title>
<p>Human health has been benefited from the secondary metabolites (SMs) produced by fungi. These small molecules, also known as natural products, include important clinical drugs like the antibiotic penicillins (<xref ref-type="bibr" rid="B69">Kardos and Demain, 2011</xref>), the cholesterol-lowering statins (<xref ref-type="bibr" rid="B45">Endo, 2010</xref>), the immunosuppressive cyclosporins (<xref ref-type="bibr" rid="B14">Britton and Palacios, 1982</xref>) and the antifungal echinocandins (<xref ref-type="bibr" rid="B7">Balkovec et al., 2014</xref>). Microbial SMs, including those from bacteria and fungi, continue to serve as important sources of molecules for drug discovery. For many decades, the fascinating and diverse structures of microbial SMs have inspired the organic chemists to embark on a quest to elucidate their biosynthetic pathways. Many basic insights into SM pathways were obtained by organic chemists using isotopic tracers during the 1950s (<xref ref-type="bibr" rid="B9">Bentley, 1999</xref>). The research shifted to the molecular biology of SM biosynthesis with the availability of tools for DNA cloning and sequencing. This is marked by several landmark papers, which described the molecular cloning of whole SM biosynthetic pathway on a contiguous stretch of DNA from actinomycete bacteria (<xref ref-type="bibr" rid="B90">Malpartida and Hopwood, 1984</xref>; <xref ref-type="bibr" rid="B33">Cortes et al., 1990</xref>; <xref ref-type="bibr" rid="B44">Donadio et al., 1991</xref>). Around the same period, the first fungal SM gene cluster, the penicillin biosynthetic gene cluster with the core non-ribosomal peptide synthetase (NRPS) gene encoding <sc>L</sc>-&#x003B4;-(&#x003B1;-aminoadipoyl)-<sc>L</sc>-cysteinyl-<sc>D</sc>-valine (ACV) synthetase had been discovered in the fungus <italic>Penicillium chrysogenum</italic> (<xref ref-type="bibr" rid="B43">D&#x000ED;ez et al., 1990</xref>). This is followed by the discovery of the terpenoid gene cluster encoding trichothecenes (<xref ref-type="bibr" rid="B64">Hohn et al., 1993</xref>), and polyketide gene clusters encoding aflatoxin/sterigmatocystin biosynthesis in <italic>Aspergillus</italic> sp. (<xref ref-type="bibr" rid="B15">Brown et al., 1996</xref>; <xref ref-type="bibr" rid="B134">Yu et al., 2004</xref>) and lovastatin biosynthesis in <italic>A. terreus</italic> (<xref ref-type="bibr" rid="B72">Kennedy et al., 1999</xref>). This hallmark trait of gene clustering was then observed in almost all other classes of fungal SM pathways including indole alkaloids and terpenoids (<xref ref-type="bibr" rid="B71">Keller et al., 2005</xref>). The tendency for the biosynthetic genes in microbial SM pathways to cluster on a chromosomal locus greatly accelerated the elucidation of enzymatic steps involved in biosynthesis of individual SM compounds using molecular biology approaches. Consequently, identification of the gene cluster that encodes the production of a given SM is now becoming the common first step toward elucidating the molecular and enzymatic basis for the biosynthesis of a given SM. Subsequent verification of the predicted gene cluster is often achieved via targeted deletion and/or heterologous expression of key biosynthetic genes. Further characterization of the biosynthetic pathway can be done by deletion of the individual biosynthetic genes in the cluster or reconstruction of the whole pathway in heterologous systems.</p>
<p>Targeted SM gene cluster discovery in fungi in the pre-genomic era is a tedious and time-consuming process. Traditionally, this was done by either complementation of blocked mutants by cosmid libraries (e.g. <xref ref-type="bibr" rid="B91">Mayorga and Timberlake, 1990</xref>; <xref ref-type="bibr" rid="B63">Hendrickson et al., 1999</xref>), or insertional mutagenesis followed by plasmid rescue from the blocked mutant (e.g., <xref ref-type="bibr" rid="B133">Yang et al., 1996</xref>; <xref ref-type="bibr" rid="B31">Chung et al., 2003</xref>). Both of these aforementioned methods rely on screening of blocked/complementation mutants, which work well for pigment compounds but can be cumbersome if the phenotype, cannot be easily observed or assayed. For example, in the pioneering work to identify lovastatin gene cluster, 6000 mutants were screened for restored lovastatin production by HPLC/TLC after transformation of <italic>A. terreus</italic> with a cosmid library (<xref ref-type="bibr" rid="B63">Hendrickson et al., 1999</xref>). Other methods for identifying key biosynthetic gene include antibody screening of cDNA expression library (e.g., <xref ref-type="bibr" rid="B8">Beck et al., 1990</xref>), differential display reverse transcriptase-PCR (e.g., <xref ref-type="bibr" rid="B86">Linnemannstons et al., 2002</xref>), suppression subtractive hybridization-PCR (<xref ref-type="bibr" rid="B97">O&#x02019;Callaghan et al., 2003</xref>). However, these methods often lead to isolation of a single gene or partial gene cluster. Further cosmid library walking is required to obtain the whole gene cluster. Due to the relatively large genome size, genome scanning method, such as that demonstrated for discovery of enediyne antitumor antibiotic pathways in actinomycete bacteria (<xref ref-type="bibr" rid="B137">Zazopoulos et al., 2003</xref>), is less feasible for fungi.</p>
<p>Biosynthesis of polyketide SMs has been subjected to more intensive studies among other classes of SM pathways in fungi (<xref ref-type="bibr" rid="B29">Chooi and Tang, 2012</xref>). One of the earlier major advances in identification of fungal polyketide SM gene clusters is the development of degenerate primer PCR based on conserved ketosynthase (KS) domain of polyketide synthases (PKSs). The KS domain DNA fragments are then use as probes to identify the cosmid library clone carrying the whole or partial gene cluster. This method has been employed to localize the aflatoxin and fumonisin PKS genes (<xref ref-type="bibr" rid="B47">Feng and Leonard, 1995</xref>; <xref ref-type="bibr" rid="B99">Proctor et al., 1999</xref>). The subsequent primer sets developed to target KS domains of a non-reducing (NR-), partial-reducing (PR-), and highly-reducing polyketide synthases (HR-PKSs) are especially useful for localizing specific PKS gene clusters in cosmid libraries (<xref ref-type="bibr" rid="B11">Bingle et al., 1999</xref>; <xref ref-type="bibr" rid="B93">Nicholson et al., 2001</xref>). Pioneering work by <xref ref-type="bibr" rid="B79">Kroken et al. (2003)</xref> which performed a phylogenomic analysis of the PKS genes in the genomes of <italic>Gibberella</italic>, <italic>Neurospora</italic>, <italic>Cochliobolus</italic>, and <italic>Botrytis</italic> species revealed that fungal PKSs are likely to be derived from eight major lineages (<xref ref-type="bibr" rid="B79">Kroken et al., 2003</xref>). Nevertheless, despite the degeneracy of the primers and subdivision of fungal PKS genes in to different subclasses, some PKS genes in the genome can be still be missed by the degenerate primer PCR approach. Furthermore, it has to be bear in mind that, although most SM pathways are clustered on chromosome in fungi some of them can be split into two or three smaller subclusters, such as the pathways for dothistromin in <italic>Dothistroma septosporum</italic> (<xref ref-type="bibr" rid="B22">Chettri et al., 2013</xref>), tryptoquivaline in <italic>A. clavatus</italic> (<xref ref-type="bibr" rid="B56">Gao et al., 2011</xref>), echinocandin in <italic>Emericella rugulosa</italic> (<xref ref-type="bibr" rid="B19">Cacho et al., 2012</xref>), and prenylated xanthones in <italic>A. nidulans</italic> (<xref ref-type="bibr" rid="B103">Sanchez et al., 2011</xref>). There are also instances where the gene clusters of multiple SM pathways are intertwined together, such as the fumitremorgin, fumagillin, and pseurotin supercluster (<xref ref-type="bibr" rid="B123">Wiemann et al., 2013a</xref>). In such cases, the absence of whole genome sequence information can complicate the identification of the complete gene set for the target SM pathway.</p>
<p>Whole genome sequencing (WGS) of the target SM-producing fungus bypasses the need for the cumbersome library construction, screening, and chromosome walking. More importantly, the genome sequence can reveal the inventory of all the SM gene clusters in the producing fungus. Even though each fungus can harbor 30&#x02013;50 SM gene cluster, the number is still finite and one of them must encode the SM of interest. For example, WGS of <italic>G. zeae</italic> with the Sanger sequencing method has allowed the systematic deletion of all 15 PKS genes in the fungal genome, which lead to identification of the gene cluster for zearalenone, aurofusarin, fusarin C and an unidentified black perithecial pigment (<xref ref-type="bibr" rid="B55">Gaffoor et al., 2005</xref>). This opens up the opportunities for detailed characterization of these pathways for zearalenone (<xref ref-type="bibr" rid="B75">Kim et al., 2005</xref>; <xref ref-type="bibr" rid="B140">Zhou et al., 2010</xref>; <xref ref-type="bibr" rid="B80">Lee et al., 2011</xref>), aurofusarin (<xref ref-type="bibr" rid="B49">Frandsen et al., 2006</xref>; <xref ref-type="bibr" rid="B50">Frandsen et al., 2011</xref>), and fusarin C (<xref ref-type="bibr" rid="B95">Niehaus et al., 2013</xref>). The development of next-generation sequencing (NGS) technologies in the last decade has dramatically lowered the cost for DNA sequencing and put the power of microbial WGS in the hand of individual laboratories (<xref ref-type="bibr" rid="B117">van Dijk et al., 2014</xref>). This technology revolution has energized the natural product research field and sparked some exciting NGS-based targeted SM gene cluster discovery projects in fungi (<bold>Table <xref ref-type="table" rid="T1">1</xref></bold>). Our work that used such NGS approach includes the discovery of SM clusters encoding viridicatumtoxin and griseofulvin (<xref ref-type="bibr" rid="B24">Chooi et al., 2010</xref>), tryptoquialanine (<xref ref-type="bibr" rid="B56">Gao et al., 2011</xref>), echinocandin (<xref ref-type="bibr" rid="B19">Cacho et al., 2012</xref>), asperlicin (<xref ref-type="bibr" rid="B61">Haynes et al., 2012</xref>), ardeemin (<xref ref-type="bibr" rid="B62">Haynes et al., 2013</xref>), and brefeldin (<xref ref-type="bibr" rid="B135">Zabala et al., 2014</xref>). Other studies that have taken the advantage of NGS have identified the SM clusters for fungal bicyclo[2.2.2]diazaoctane indole alkaloids (<xref ref-type="bibr" rid="B83">Li et al., 2012</xref>), equisetin (<xref ref-type="bibr" rid="B68">Kakule et al., 2013</xref>), and pneumocandin (<xref ref-type="bibr" rid="B21">Chen et al., 2013</xref>).</p>
<table-wrap position="float" id="T1">
<label>Table 1</label>
<caption><p>Examples of biosynthetic gene clusters assigned to their respective compounds using next-generation sequencing technology.</p></caption>
<table cellspacing="5" cellpadding="5" frame="hsides" rules="groups">
<thead>
<tr>
<th valign="top" align="left">Species</th>
<th valign="top" align="left">Sequencing Method</th>
<th valign="top" align="left">Characterized SM biosynthetic gene clusters</th>
<th valign="top" align="left">Reference</th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left"><italic>Penicillium aethiopicum</italic></td>
<td valign="top" align="left">454</td>
<td valign="top" align="left">Griseofulvin</td>
<td valign="top" align="left"><xref ref-type="bibr" rid="B24">Chooi et al. (2010)</xref></td>
</tr>
<tr>
<td valign="top" align="left"></td>
<td valign="top" align="left"></td>
<td valign="top" align="left">Viridicatumtoxin</td>
<td valign="top" align="left"><xref ref-type="bibr" rid="B24">Chooi et al. (2010)</xref></td>
</tr>
<tr>
<td valign="top" align="left"></td>
<td valign="top" align="left"></td>
<td valign="top" align="left">Tryptoquialanine</td>
<td valign="top" align="left"><xref ref-type="bibr" rid="B56">Gao et al. (2011)</xref></td>
</tr>
<tr>
<td valign="top" align="left"><italic>Emericella rugulosa</italic></td>
<td valign="top" align="left">Illumina</td>
<td valign="top" align="left">Echinocandin B</td>
<td valign="top" align="left"><xref ref-type="bibr" rid="B19">Cacho et al. (2012)</xref></td>
</tr>
<tr>
<td valign="top" align="left"><italic>Aspergillus alliaceus</italic></td>
<td valign="top" align="left">Illumina</td>
<td valign="top" align="left">Asperlicin</td>
<td valign="top" align="left"><xref ref-type="bibr" rid="B61">Haynes et al. (2012)</xref></td>
</tr>
<tr>
<td valign="top" align="left"><italic>Aspergillus</italic> sp. MF 297-2</td>
<td valign="top" align="left">Illumina</td>
<td valign="top" align="left">(-)-notoamide A</td>
<td valign="top" align="left"><xref ref-type="bibr" rid="B83">Li et al. (2012)</xref></td>
</tr>
<tr>
<td valign="top" align="left"><italic>A. versicolor</italic> NRRL 35600</td>
<td valign="top" align="left">Illumina</td>
<td valign="top" align="left">(+)-notamide A</td>
<td valign="top" align="left"><xref ref-type="bibr" rid="B83">Li et al. (2012)</xref></td>
</tr>
<tr>
<td valign="top" align="left"><italic>P. fellutanum</italic> ATCC 20841</td>
<td valign="top" align="left">Illumina</td>
<td valign="top" align="left">Paraherquamide A</td>
<td valign="top" align="left"><xref ref-type="bibr" rid="B83">Li et al. (2012)</xref></td>
</tr>
<tr>
<td valign="top" align="left"><italic>Malbranchea aurantiaca</italic> RRC1813</td>
<td valign="top" align="left">Illumina</td>
<td valign="top" align="left">Malbrancheamide</td>
<td valign="top" align="left"><xref ref-type="bibr" rid="B83">Li et al. (2012)</xref></td>
</tr>
<tr>
<td valign="top" align="left"><italic>Glarea lozoyensis</italic></td>
<td valign="top" align="left">Illumina</td>
<td valign="top" align="left">Pneumocandin</td>
<td valign="top" align="left"><xref ref-type="bibr" rid="B21">Chen et al. (2013)</xref></td>
</tr>
<tr>
<td valign="top" align="left"><italic>A. fischeri</italic></td>
<td valign="top" align="left">Illumina</td>
<td valign="top" align="left">Ardeemin</td>
<td valign="top" align="left"><xref ref-type="bibr" rid="B62">Haynes et al. (2013)</xref></td>
</tr>
<tr>
<td valign="top" align="left"><italic>Fusarium fujikuroi</italic></td>
<td valign="top" align="left">454</td>
<td valign="top" align="left">Apicidin F</td>
<td valign="top" align="left"><xref ref-type="bibr" rid="B94">Niehaus et al. (2014)</xref></td>
</tr>
<tr>
<td valign="top" align="left"><italic>Eupenicillium brefeldianum</italic></td>
<td valign="top" align="left">454/Illumina</td>
<td valign="top" align="left">Brefeldin</td>
<td valign="top" align="left"><xref ref-type="bibr" rid="B135">Zabala et al. (2014)</xref></td>
</tr>
<tr>
<td valign="top" align="left"><italic>F. heterosporum</italic></td>
<td valign="top" align="left">Illumina</td>
<td valign="top" align="left">Equisetin and fusaridione A</td>
<td valign="top" align="left"><xref ref-type="bibr" rid="B68">Kakule et al. (2013)</xref></td>
</tr>
</tbody>
</table>
</table-wrap>
<p>Contemporary natural product research programs are now increasingly based on the understanding of the relationship between the SM molecules and the biosynthetic genes (<xref ref-type="bibr" rid="B120">Walsh and Fischbach, 2010</xref>). Although such studies are primarily motivated by the desire to understand the molecular and enzymatic basis of SM biosynthesis, the potential benefits and implications derived from such work are manifold. Firstly, the new chemical insights obtained in elucidating the metabolic pathway can be used to design more efficient total synthesis routes for complex natural products. Secondly, knowledge about the gene cluster can facilitate the metabolic engineering effort to increase the yield of useful SMs for commercial production (<xref ref-type="bibr" rid="B98">Pickens et al., 2011</xref>). The knowledge also forms the basis for generation of new SM analogs by mutasynthesis and combinatorial biosynthesis. A recent example is the combinatorial biosynthesis of benzenediol lactones (<xref ref-type="bibr" rid="B131">Xu et al., 2014b</xref>), which built on previous studies (<xref ref-type="bibr" rid="B101">Reeves et al., 2008</xref>; <xref ref-type="bibr" rid="B122">Wang et al., 2008</xref>; <xref ref-type="bibr" rid="B140">Zhou et al., 2010</xref>; <xref ref-type="bibr" rid="B129">Xu et al., 2013b</xref>). Novel biocatalysts useful for green chemistry and chemoenzymatic process development may also be discovered from SM pathways. For example, the characterization of the acyltransferase LovD in lovastatin pathway led to the development of green chemistry process for the semisynthetic cholesterol-lowering drug simvastatin (<xref ref-type="bibr" rid="B127">Xie et al., 2006</xref>, <xref ref-type="bibr" rid="B126">2009</xref>; <xref ref-type="bibr" rid="B58">Gao et al., 2009</xref>; <xref ref-type="bibr" rid="B128">Xu et al., 2013a</xref>). Furthermore, the gene cluster information will be useful for knowledge-based genome mining for structurally-related compound in other fungi, e.g., the discovery of the immunosuppressive neosartoricin based on viridicatumtoxin biosynthesis genes (<xref ref-type="bibr" rid="B30">Chooi et al., 2012</xref>, <xref ref-type="bibr" rid="B25">2013a</xref>). Importantly, these established links between genes and SMs are valuable knowledge that contributes to the overarching aims for (1) accurate prediction of SM structures based on DNA sequences and (2) rational design of biosynthetic pathways for synthesis of organic molecules. Finally, bridging the gaps between genes and molecules may facilitate our understanding of the natural functions of SMs using comparative genomics and transcriptomics tools (<xref ref-type="bibr" rid="B28">Chooi and Solomon, 2014</xref>).</p>
<p>It is important to note that a SM/natural product-motivated fungal genome sequencing project has different aims compared to the conventional WGS project coordinated by an international consortium of researchers associated with large sequencing centers. The major goal is to obtain the whole SM gene cluster that encode the targeted SM on a contiguous stretch of DNA sequence contig or scaffold. That means the completeness of the genome sequence coverage and whole-genome annotation is of lower priorities. The key question is how to rapidly narrow down and accurately pinpoint the correct SM cluster in the genome that encodes the production of a target SM. Getting an accurate initial prediction will significantly reduce the time spent on gene cluster verification. Below, we will provide the general work flow and guidelines for researchers who are considering adopting NGS technologies for targeted SM gene cluster discovery based on our own experience. We will also use several examples stemming from our work, two PKS pathways and two NRPS pathways, to illustrate the concepts and strategies.</p>
</sec>
<sec><title>NEXT-GENERATION <italic>DE NOVO</italic> FUNGAL GENOME SEQUENCING AND ASSEMBLY</title>
<p>The standard NGS-based targeted SM gene cluster discovery work flow used in our laboratory is presented in <bold>Figure <xref ref-type="fig" rid="F1">1</xref></bold>. The fungal strain acquired from culture collections or her sources are first verified for production of the targeted compound before its genomic DNA was sent for sequencing. A variety of culture media and conditions can also be tested to optimize the production of the target compound. Gene deletion or disruption followed by the detection of loss of target compound production is still the most common method for initial gene cluster verification. Alternatively, if genetic transformation is proven to be difficult on the fungus, expression of the backbone biosynthetic enzymes [e.g., PKS, NRPS, terpene synthase, or dimethylallyltryptophan synthase (DMATS)] in the candidate SM gene cluster in a heterologous system may be another way to verify the gene cluster.</p>
<fig id="F1" position="float">
<label>FIGURE 1</label>
<caption><p><bold>Next-generation sequencing (NGS) workflow for targeted secondary metabolite (SM) gene cluster discovery</bold>.</p></caption>
<graphic xlink:href="fmicb-05-00774-g001.tif"/>
</fig>
<p>The next thing to consider will be the choice of sequencing method. The choice will be mainly based on the cost, sequence quality, sequence read length, speed and project throughput (the number of strains to be sequenced together). There are currently several NGS technologies on the market suitable for <italic>de novo</italic> fungal WGS, with Illumina now dominating the market (<xref ref-type="bibr" rid="B96">Nowrousian, 2010</xref>; <xref ref-type="bibr" rid="B117">van Dijk et al., 2014</xref>). We have experience in both Roche/454 FLX Titanium and Illumina HiSeq2000, but other DNA sequencing platforms may work for this purpose as well. For an overview of the different latest NGS technologies, the readers are referred to <xref ref-type="bibr" rid="B117">van Dijk et al. (2014)</xref>. The <italic>de novo</italic> sequencing of <italic>P. aethiopicum</italic> (synonym <italic>P. lanosocoeruleum</italic>) IBT5753 genome is one of the earliest examples of SM-motivated fungal WGS undertaken by individual laboratories, which used the Roche/454 FLX Titanium platform (<xref ref-type="bibr" rid="B24">Chooi et al., 2010</xref>).</p>
<p>The introduction of Illumina HiSeq2000 (now superseded by HiSeq2500) with &#x0223C;100 bp paired end (PE) reads dramatically reduced the cost of sequencing and increased the sequencing output. The shorter read length of HiSeq2000 is compensated by the deeper coverage and the paired-end nature of the Illumina reads. The Illumina PE reads allowed the assembly of longer scaffolds with gaps, which can be filled in later using routine PCR and Sanger sequencing if any of the scaffold harbor interesting SM gene cluster. Longer PE information can be obtained by generating mate pair libraries with larger insert size (e.g., 3 kbp or 5 kbp). Indeed, it has been recently demonstrated that good fungal genome assembly can be obtained on platform with shorter sequence reads (50 bp) like SOLiD using mate pair libraries (<xref ref-type="bibr" rid="B115">Umemura et al., 2013b</xref>).</p>
<p>We first used HiSeq2000 for sequencing of the brefeldin-producing <italic>Eupenicillium brefeldianum</italic> ATCC 58665 and the echinocandin-producing <italic>E. rugulosa</italic> (<xref ref-type="bibr" rid="B19">Cacho et al., 2012</xref>; <xref ref-type="bibr" rid="B135">Zabala et al., 2014</xref>). The amount of sequence data acquired for assembly of the &#x0223C;32 Mbp <italic>E. rugulosa</italic> genome is more than sufficient despite the shorter read length (N50 = 235 kbp). In fact, we can get equally good quality assembly using half the amount of data. Now, we routinely perform HiSeq2000 sequencing multiplexed for four fungal genomic samples per lane (&#x0223C;200X depth of coverage each genome) and can routinely obtain good quality assembly for the purpose of SM gene cluster discovery. Using this arrangement, the cost of sequencing per fungal strain can be lower than the cost required for cosmid/fosmid library construction, screening, and chromosome walking, not to mention the significant time savings. With the use of mate pair libraries, more fungal genomic samples can likely be included into HiSeq2000 per lane yet still yield fine assembly. One recent study shows that the optimum sequencing depth for small bacteria to medium eukaryotic genomes with 2X 100 bp PE Illumina reads is 50-100X (<xref ref-type="bibr" rid="B41">Desai et al., 2013</xref>).</p>
<p>There is more than a dozen of software available for assembly of short reads generated by NGS, some are more memory intensive than the others, and there is differences in assembly speed as well. The popular ones including Velvet (<xref ref-type="bibr" rid="B138">Zerbino and Birney, 2008</xref>), SOAPdenovo (<xref ref-type="bibr" rid="B81">Li et al., 2010a</xref>), AllPATHS (<xref ref-type="bibr" rid="B17">Butler et al., 2008</xref>), and ABySS (<xref ref-type="bibr" rid="B107">Simpson et al., 2009</xref>). With optimization, some software may allow some small-medium size genomes to be assembled on a standalone workstation (<xref ref-type="bibr" rid="B76">Kleftogiannis et al., 2013</xref>). There are several studies that compare the efficiency and assembly quality of different assembly software (<xref ref-type="bibr" rid="B85">Lin et al., 2011</xref>; <xref ref-type="bibr" rid="B139">Zhang et al., 2011</xref>). We have experience mostly in using the SOAPdenovo assembler developed by BGI (<xref ref-type="bibr" rid="B81">Li et al., 2010a</xref>). Our SOAPdenovo assemblies were run on the UCLA Hoffman2 computer cluster. The assemblies can usually be completed with 32&#x02013;128 GB memory requested from the cluster, depending on the genome size and the amount of input data. For the standard 2X 100 bp PE reads, we often use a <italic>k</italic>-mer size of 63 or 79 for the SOAPdenovo assembly and were able to obtain good results. The latest version, SOAPdenovo2, promised to improve memory efficiency and assembly quality, and more optimized for the longer Illumina reads (rather than the 35&#x02013;50 bp reads from older Illumina platforms; <xref ref-type="bibr" rid="B87">Luo et al., 2012</xref>). Many NGS providers also provide service for sequence assembly with a fee.</p>
</sec>
<sec><title>LOCATING THE TARGET SECONDARY METABOLITE GENE CLUSTER</title>
<p>The first task after obtaining the scaffolds generated from the assembly is to narrow down the scaffolds containing the candidate SM gene clusters. Despite the enormous structural diversity, most fungal SMs can be divided into four major classes, polyketides, non-ribosomal peptides, terpenes, and indole alkaloids, based on the limited classes of carbon building blocks they derived from (<xref ref-type="bibr" rid="B71">Keller et al., 2005</xref>). Thus, an efficient way of narrowing down the scaffolds that may contain the target gene cluster is to search for genes encoding backbone biosynthetic enzymes that synthesized the specific class of compounds correspond to the target SM. This can be most easily achieved by performing a TBLASTN search against a database generated from the assembled fungal genome scaffolds using the &#x0201C;makeblastdb&#x0201D; command in the NCBI stand-alone BLAST application<sup><xref ref-type="fn" rid="fn01">1</xref></sup>. Depending on the class of SM, an arbitrary chosen conserved domain of the corresponding backbone enzymes can be used as a TBLASTN query. For example, a protein query sequence of a KS domain of PKS for polyketide SMs, an adenylation (A) domain of NRPS for non-ribosomal peptide SMs, terpene synthase for terpenoid SMs and DMATS for prenylated indole alkaloids. The TBLASTN will generate a list of scaffolds containing the SM gene clusters belongs the corresponding SM classes ranked based on homology to the query sequence. Alternatively, commercial bioinformatics software programs with intuitive graphic user interface (GUI) that support BLAST on local database are also available, e.g., CLC Genomics Workbench<sup><xref ref-type="fn" rid="fn02">2</xref></sup> (CLC Bio) and Geneious<sup><xref ref-type="fn" rid="fn03">3</xref></sup> (Biomatters). After locating the scaffolds containing gene clusters of the target SM classes, the number of candidate scaffolds can be further narrowed down with comparative genomics analysis (See Comparative Genomics Approach for Target SM Gene Cluster Prediction) followed by gene predictions and more in-depth knowledge-based bioinformatics analysis (see <bold>Figure <xref ref-type="fig" rid="F2">2</xref></bold> and Section &#x0201C;Pinpointing Target SM Gene Cluster with Knowledge-Based Analysis &#x02013; Retrobiosynthesis&#x0201D;).</p>
<fig id="F2" position="float">
<label>FIGURE 2</label>
<caption><p><bold>Comparative genomics and knowledge-based prediction strategies for targeted SM gene cluster discovery.</bold> New biosynthetic knowledge and insights into gene-to-molecular structure relationship would be useful for target SM gene cluster predictions and knowledge-based genome mining for discovery of novel SM compounds.</p></caption>
<graphic xlink:href="fmicb-05-00774-g002.tif"/>
</fig>
<p>For fungal gene predictions (the locations and exon&#x02013;intron structures of genes on individual scaffolds) we routinely use FGENESH (Softberry; <xref ref-type="bibr" rid="B108">Solovyev et al., 2006</xref>), which can yield relatively accurate predictions for fungi. The web server allows direct submission of FASTA nucleotide sequences with gene-finding parameters trained using datasets from several fungal species<sup><xref ref-type="fn" rid="fn04">4</xref></sup>. Alternatively, AUGUSTUS also provides relatively accurate fungal gene prediction<sup><xref ref-type="fn" rid="fn05">5</xref></sup> based on training sets from various fungal species (<xref ref-type="bibr" rid="B110">Stanke et al., 2004</xref>). Individual protein sequences predicted in a candidate scaffold are then submitted to NCBI BLASTP server for detailed conserved domain analysis (using the integrated NCBI Conserved Domain Search feature) and homologous sequence comparison. The EBI Interproscan also offers similar conserved domain prediction and protein functional analysis. As each gene cluster usually contain 3&#x02013;15 genes (>20 genes in some cases), it is often not too time consuming when the number of candidate scaffolds has been narrowed down significantly. Several SM gene cluster prediction software programs have also been developed, which can aid this process considerably (see Pinpointing Target SM Gene Cluster with Knowledge-Based Analysis &#x02013; Retrobiosynthesis).</p>
<sec><title>COMPARATIVE GENOMICS APPROACH FOR TARGET SM GENE CLUSTER PREDICTION</title>
<p>Comparative genomics can be a useful approach for filtering of candidate gene clusters (<bold>Figure <xref ref-type="fig" rid="F2">2</xref></bold>). The increasing number of sequenced fungal genomes in public databases, due in part to the lower sequencing cost enabled by NGS technologies, allows one to find a suitable genome of related sequenced organism for comparison with the organism of interest. Since the genomes of different organisms, even those belonging to the same genus, would encode different array of SMs, one can compare the SM backbone biosynthesis gene inventory of one organism with a closely related organism to rapidly narrow down the unique SM gene cluster. This approach is especially useful when there is pre-existing knowledge about the SM repertoire of the reference organism used for comparison. For example, in the search of SM gene clusters responsible for the biosynthesis of viridicatumtoxin and griseofulvin in <italic>P. aethiopicum</italic> (<xref ref-type="bibr" rid="B24">Chooi et al., 2010</xref>), we compared the PKS inventory of <italic>P. aethiopicum</italic> with that of <italic>P. chrysogenum</italic> based on prior knowledge from previous chemotaxonomy studies (<xref ref-type="bibr" rid="B52">Frisvad and Samson, 2004</xref>). Since it is known that <italic>P. chrysogenum</italic> produces neither viridicatumtoxin nor griseofulvin, we can first filter out the orthologous PKS genes shared between the two fungi (<bold>Figure <xref ref-type="fig" rid="F3">3A</xref></bold>; Section Case Study A: Griseofulvin and Case Study B: Viridicatumtoxin). Similar the genome-wide comparison of NRPS genes in <italic>A. nidulans</italic> and <italic>E. rugulosa</italic> was used to narrow down the candidate echinocandin gene cluster (<xref ref-type="bibr" rid="B19">Cacho et al., 2012</xref>). It has to be bear in mind that the absence of report about the presence of a target SM compound in a specific species does not always correlate with the absence of the gene cluster in the genome as there can be strain-strain variation and many SM gene clusters could be transcriptionally silent. For example, TAN-1612, reported in another <italic>A. niger</italic> strain as BMS-192548 (<xref ref-type="bibr" rid="B77">Kodukula et al., 1995</xref>; <xref ref-type="bibr" rid="B106">Shu et al., 1995</xref>), is not detected in the sequenced <italic>A. niger</italic> ATCC 1015. However, the corresponding gene cluster can be identified and activated by transcriptional regulator overexpression (<xref ref-type="bibr" rid="B84">Li et al., 2011</xref>). The abilities to produce fumitremorgin by <italic>A. fumigatus</italic> (<xref ref-type="bibr" rid="B70">Kato et al., 2013</xref>), and gibberrellin and beauvericin by <italic>G. fujikuroi</italic> (<xref ref-type="bibr" rid="B124">Wiemann et al., 2013b</xref>), also vary from strain to strain. Nonetheless, such a comparative genomics approach had served as a useful first-pass filter in our hands, especially for fungal species that have been characterized chemotaxonomically, such as those species in the <italic>Aspergillus</italic> and <italic>Penicillium</italic> genera (<xref ref-type="bibr" rid="B52">Frisvad and Samson, 2004</xref>; <xref ref-type="bibr" rid="B51">Frisvad et al., 2007</xref>).</p>
<fig id="F3" position="float">
<label>FIGURE 3</label>
<caption><p><bold>Comparative genomics approach for discovery of SM biosynthetic gene cluster.</bold> Comparison of the PKS genes in <italic>Penicillium aethiopicum</italic> and <italic>P. chrysogenum</italic> was utilized in searching for the griseofulvin and viridicatumtoxin gene cluster <bold>(A)</bold>. In order to narrow down possible candidates for the tryptoquialanine gene cluster in<italic> P. aethiopicum</italic>, NRPS genes that are non-orthologous to <italic>P. chrysogenum</italic> NRPS genes but are orthologous to the NRPS genes of the tryptoquivaline producer <italic>A. clavatus</italic> were found <bold>(B)</bold>. The echinocandin NRPS gene was found in <italic>Emericella rugulosa</italic> was found by searching for a hexamodule NRPS gene in <italic>E. rugulosa</italic> that is non-orthologous to <italic>A. nidulans</italic> NRPS genes (<bold>C</bold>).</p></caption>
<graphic xlink:href="fmicb-05-00774-g003.tif"/>
</fig>
<p>Conversely, comparative genomics can also be utilized for identifying the biosynthetic gene clusters of a target group of SMs bearing structural similarities but made by less-related fungal species (<bold>Figure <xref ref-type="fig" rid="F2">2</xref></bold>). In this case, orthologous gene clusters in two organisms that are capable of synthesizing the same family of compounds are targeted. An example of the latter strategy is demonstrated in our work on tryptoquialanine/tryptoquivaline gene cluster identification in <italic>P. aethiopicum</italic>/<italic>A. clavatus</italic> (<xref ref-type="bibr" rid="B56">Gao et al., 2011</xref>; <bold>Figure <xref ref-type="fig" rid="F3">3B</xref></bold>). The strategy is also well-demonstrated in the work by the Sherman group in investigating the biosynthesis of fungal bicyclo[2.2.2]diazaoctane indole alkaloids (-)- and (+)-notoamide A, paraherquamide A, and malbrancheamide A by <italic>Aspergillus</italic> sp. <italic>MF297-2, A. versicolor</italic> NRRL35600, <italic>P. fellutanum</italic> ATCC20841, and<italic> Malbranchea aurantiaca</italic> RRC1813, respectively (<xref ref-type="bibr" rid="B83">Li et al., 2012</xref>). More recently, NGS-driven comparative genomics of four <italic>Stachybotrys</italic> strains from two chemotypes producing either atranones or stratoxins have revealed two unique gene clusters, which possibly encode the biosynthesis of the two respective terpenoid-derived SMs (<xref ref-type="bibr" rid="B105">Semeiks et al., 2014</xref>). However, the identities of the two gene clusters are yet to be verified experimentally. Such comparative genomics approach is highly compatible with the high-throughput nature of Illumina sequencing as genomic samples from multiple fungal strains can be multiplexed on a single lane of an Illumina sequencer flow cell. In the case where the target SM gene cluster is split into more than one scaffold in one of the fungal genome assemblies, either due to disruption of assembly by repetitive sequence or the pathway is encoded by multiple loci, comparative genomics of the different strains can help localize the complete biosynthetic gene sets. This approach has facilitated the identification of genes for biosynthesis of tryptoquivaline in <italic>A. clavatus</italic>, which was separated into three genomic loci, by comparing with the tryptoquialanine-producing <italic>P. aethiopicum</italic> genome (<xref ref-type="bibr" rid="B56">Gao et al., 2011</xref>).</p>
<p>Lastly, one can use comparative genomics to make an educated guess on where the boundaries of the SM gene cluster are located. Oftentimes, SM gene clusters are flanked by syntenic blocks containing highly-conserved core genes (<xref ref-type="bibr" rid="B88">Machida et al., 2005</xref>; <xref ref-type="bibr" rid="B113">Tamano et al., 2008</xref>). Similarly, we have observed this trend in our studies toward the discovery of gene clusters encoding griseofulvin, viridicatumtoxin, cytochalasin, and tryptoquialanine. All four putative SM gene clusters are flanked by syntenic block of genes with high shared identity (>85%) in closely related ascomycetes species (<xref ref-type="bibr" rid="B24">Chooi et al., 2010</xref>; <xref ref-type="bibr" rid="B56">Gao et al., 2011</xref>; <xref ref-type="bibr" rid="B100">Qiao et al., 2011</xref>). In fact, a study reported a motif-independent bioinformatics approach for detection of SM gene cluster based on non-syntenic blocks in fungal genomes (<xref ref-type="bibr" rid="B114">Umemura et al., 2013a</xref>). One of the software that is useful for identifying and visualizing syntenic regions across multiple genomes is Mauve (<xref ref-type="bibr" rid="B40">Darling et al., 2004</xref>). While it does not guarantee that genes that are not highly conserved to a related organism are part of the cluster, this strategy is helpful toward minimizing the number of genes subjected to further functional analysis.</p>
</sec>
<sec><title>PINPOINTING TARGET SM GENE CLUSTER WITH KNOWLEDGE-BASED ANALYSIS &#x02013; RETROBIOSYNTHESIS</title>
<p>As mentioned above, gene predictions of individual scaffolds can be performed using software like FGENESH or AUGUSTUS after narrowing down the number of candidate SM gene clusters to a handful of scaffolds. Further detailed bioinformatics analysis of the individual SM gene clusters is then needed to identify the target gene cluster. Some software programs have been developed to aid SM gene cluster predictions. Two of the popular software programs available for prediction of SM gene clusters are antiSMASH (<xref ref-type="bibr" rid="B92">Medema et al., 2011</xref>; <xref ref-type="bibr" rid="B12">Blin et al., 2013</xref>) and SMURF (<xref ref-type="bibr" rid="B74">Khaldi et al., 2010</xref>). SMURF is specific for predicting SM gene clusters in fungal genomes, while antiSMASH can be used for both bacteria and fungi. Guidelines and detailed protocols for using these two and other related SM biosynthetic gene prediction programs can be found in <xref ref-type="bibr" rid="B46">Fedorova et al. (2012)</xref>. SMURF requires a protein FASTA file and a gene coordinate file as input data. On the other hand, antiSMASH can accepts single nucleotide FASTA file and can be a very useful tool for getting an initial idea of the composition of SM gene cluster on each candidate scaffold along with gene annotation suggestions. Besides providing the predicted domain architecture of multi-domain backbone enzymes (i.e., PKSs and NRPSs), the smCOG (SM Cluster of Orthologous Groups) analysis module in antiSMASH also predicts the function of probable tailoring enzymes encoded in the cluster based on conserved domain analysis. Unfortunately, the substrate prediction function of antiSMASH is yet to be as useful for fungal SM clusters compared to bacterial ones. Moreover, intron prediction using antiSMASH is not as accurate compared to FGENESH, which can potentially hamper efforts toward heterologous expression of fungal SM biosynthetic enzymes in <italic>E. coli</italic> or yeast. With increasing number of established connections between fungal SM gene clusters and molecular structures, this feature is likely to improve in the future.</p>
<p>Nonetheless, a good understanding of the biochemistry of SM biosynthesis and knowledge about the relationship between biosynthetic genes and molecular structures is often needed to accurately pinpoint the SM gene cluster of interest. An excellent introduction to the common building blocks, enzymes, and biochemical reactions involve in SM biosynthesis is available (<xref ref-type="bibr" rid="B42">Dewick, 2009</xref>). The specific question that will be asked by the researcher is &#x0201C;what kind (and combination) of enzymes and precursors are likely to be involved in the biosynthesis of the target SM compound?&#x0201D; Such analytic-deductive approach is sometimes referred as &#x0201C;retro-biosynthetic analysis&#x0201D; where the SM structure are taken apart into simpler intermediates and precursors to help determine the enzymes and biological building blocks required for the target SM biosynthesis. Here, we will focus on the general strategies that can be adopted for pinpointing target PKS and NRPS gene clusters.</p>
<sec><title>Polyketides</title>
<p>A great majority of fungal PKSs belong to the iterative type I PKSs whereupon the different PKS catalytic domains are juxtaposed on a single large polypeptide and the single set of PKS domains performs all the necessary catalytic activity during the biosynthesis of the polyketide. Fungal iterative type I PKSs are further classified into three major classes based on the degree of &#x003B2;-keto reduction performed by the PKS; namely the NR-PKSs, the PR-PKSs, and the HR-PKSs. The enzymology and classification of fungal iterative type I PKSs have been reviewed extensively (<xref ref-type="bibr" rid="B34">Cox, 2007</xref>; <xref ref-type="bibr" rid="B39">Crawford and Townsend, 2010</xref>; <xref ref-type="bibr" rid="B29">Chooi and Tang, 2012</xref>). As a general rule, aromatic polyketide compounds are synthesized by NR-PKSs, while HR-PKSs produce aliphatic compounds. PR-PKSs, on the other hand, has been shown to produce compounds lacking a phenolic hydroxyl group at the aromatic ring on the position where a &#x003B2;-keto group has been reduced to alcohol on the polyketide chain before cyclization, such as 6-methylsalicylic acid (<xref ref-type="bibr" rid="B8">Beck et al., 1990</xref>; <xref ref-type="bibr" rid="B53">Fujii et al., 1996</xref>) and (<italic>R</italic>)-mellein (<xref ref-type="bibr" rid="B27">Chooi et al., 2015</xref>). Thus, depending on the nature of the target polyketide compound, the number of candidate scaffolds can be further narrowed down by targeting NR-, PR-, or HR-PKS using corresponding KS domain. Querying the local BLAST database consist of the assembled genomic scaffolds with a corresponding KS domain will resulted in scaffolds containing the corresponding group of PKS genes appearing on top of the TBLASTN hit list. The specific group of PKSs encoded in the candidate scaffolds can be subjected to further scrutiny to pinpoint the PKS gene responsible for biosynthesis of the target SM. Type III PKSs, which present mainly in plants but can be found in some fungi as well, but are usually very limited in number (one or two) in most fungal genomes and are known to synthesize resorcylic acid-type compounds (<xref ref-type="bibr" rid="B60">Hashimoto et al., 2014</xref>).</p>
<p>In addition to the minimal PKS domains, a typical NR-PKS also contains the starter unit:ACP transacylase (SAT; <xref ref-type="bibr" rid="B36">Crawford et al., 2006</xref>); the product template (PT; <xref ref-type="bibr" rid="B38">Crawford et al., 2008</xref>, <xref ref-type="bibr" rid="B37">2009</xref>). An NR-PKS could also contain a thioesterase/Claisen-like cyclase (TE/CLC; <xref ref-type="bibr" rid="B54">Fujii et al., 2001</xref>; <xref ref-type="bibr" rid="B78">Korman et al., 2010</xref>) or a terminal reductive domain (<xref ref-type="bibr" rid="B6">Bailey et al., 2007</xref>) for product release. Although examples of NR-PKSs utilize an in-<italic>trans</italic> releasing domain (<xref ref-type="bibr" rid="B5">Awakawa et al., 2009</xref>; <xref ref-type="bibr" rid="B84">Li et al., 2011</xref>) and NR-PKS that do not require any releasing domain or enzyme (<xref ref-type="bibr" rid="B18">Cacho et al., 2013</xref>) have been characterized as well. Of these aforementioned NR-PKS domains, the PT domain, which controls the first ring cyclization of the incipient reactive polyketide backbone to form the aromatic product, was shown to be useful in predicting the product of the NR-PKS. In the work by <xref ref-type="bibr" rid="B82">Li et al. (2010b)</xref>, sequences of PT domains from characterized NR-PKS were subjected to phylogenetic analysis resulting in the grouping of the PT domains in accordance to their respective cyclization regiospecificity. The authors validated the model by demonstrating the previously uncharacterized PT domain of <italic>An03g05440</italic> from <italic>A. niger</italic>, predicted to catalyzed a C2-C7-type cyclization, does indeed perform the expected cyclization mode in a chimeric PKS <italic>in vitro</italic> system. <xref ref-type="bibr" rid="B2">Ahuja et al. (2012)</xref> has made similar observation for NR-PKSs in <italic>A. nidulans.</italic> Such PT domain analysis has facilitated the identification of viridicatumtoxin and TAN-1612 gene cluster (<xref ref-type="bibr" rid="B24">Chooi et al., 2010</xref>; <xref ref-type="bibr" rid="B84">Li et al., 2011</xref>).</p>
<p>Unlike NR-PKSs, HR-PKSs utilize the triad of the &#x003B2;-keto reductive domains ketoreductase (KR), dehydratase (DH), and enoylreductase (ER) domains to introduce complexity in the incipient polyketide during each extension cycle. However, in contrast to NR-PKSs, there is a lack of in-depth bioinformatic studies to predict the products of uncharacterized HR-PKS. While more studies in deciphering the biosynthetic rules programmed within these tailoring domains is needed to construct an accurate model of predicting HR-PKS product, important clues can be garnered about the possible product of the HR-PKS using rudimentary bioinformatic analysis. For instance, conserved domain prediction analysis tools can help reveal the presence of in-<italic>cis</italic> polyketide tailoring domains in the HR-PKS protein sequence such as a <italic>C</italic>-methyltransferase domain that would indicate &#x003B1;-methylation during one or more polyketide extension cycle. This feature has been exploited for identification of the PKS gene encoding the biosynthesis of the tetraketide side chain of squalestatin in a genomic FNA library using <italic>C</italic>-methyltransferase domain sequence probe (<xref ref-type="bibr" rid="B35">Cox et al., 2004</xref>). High sequence similarity and proximity of phylogenetic relationship with a characterized HR-PKS could, to a limited extend, suggest that the unknown PKS produce similar chain length and structure (<xref ref-type="bibr" rid="B79">Kroken et al., 2003</xref>).</p>
<p>Fungi use division of labor between NR-PKS and HR-PKS to generate compounds with an aromatic portion and highly-reduced aliphatic portion respectively (<xref ref-type="bibr" rid="B29">Chooi and Tang, 2012</xref>). The linear aliphatic chains generated by HR-PKSs are often used as starter units for NR-PKSs. Compounds generated by such NR/HR two-PKS systems include the resorcylic acid family of compounds (<xref ref-type="bibr" rid="B140">Zhou et al., 2010</xref>; <xref ref-type="bibr" rid="B129">Xu et al., 2013b</xref>) and asperfuranone (<xref ref-type="bibr" rid="B23">Chiang et al., 2009</xref>). The linear aliphatic acyl chain may also be attached to an aromatic polyketide portion as an ester, such as in the case of azanigerone (<xref ref-type="bibr" rid="B136">Zabala et al., 2012</xref>) and chaetoviridin (<xref ref-type="bibr" rid="B125">Winter et al., 2012</xref>). Thus, such structural features, if present in the target SM, will be a useful tell-tale for predicting the target SM gene cluster as there is limited (one or two) such two-PKS gene cluster(s) in filamentous fungal genomes surveyed to date. SM gene clusters encoding hybrid polyketide-nonribosomal peptide compounds (<xref ref-type="bibr" rid="B13">Boettger and Hertweck, 2013</xref>), such as cytochalasins (<xref ref-type="bibr" rid="B100">Qiao et al., 2011</xref>; <xref ref-type="bibr" rid="B65">Ishiuchi et al., 2013</xref>), equisetin (<xref ref-type="bibr" rid="B68">Kakule et al., 2013</xref>), pseurotin (<xref ref-type="bibr" rid="B89">Maiya et al., 2007</xref>), and fusarin C (<xref ref-type="bibr" rid="B95">Niehaus et al., 2013</xref>) can be identified easily in a genome as well as there is usually one or two PKS-NRPS hybrid gene(s) in most ascomycete fungal genomes. Further analysis of the tailoring enzymes encoding in the vicinity of these PKS genes can help pinpointing the exact target SM gene cluster (see Tailoring Enzymes).</p>
</sec>
<sec><title>Non-ribosomal peptides</title>
<p>Unlike iterative fungal PKSs, most fungal NRPSs are modular. Like the bacterial counterparts, fungal NRPSs are assembly-line-like protein complexes arranged in functional units known as modules (<xref ref-type="bibr" rid="B48">Finking and Marahiel, 2004</xref>, <xref ref-type="bibr" rid="B111">Strieker et al., 2010</xref>). Each NRPS module is minimally comprise of three domains: an adenylation (A) domain that selects and activates the amino acid (aa) substrate of the module, a thiolation (T) or peptidyl carrier protein (PCP) domain that serves as a covalent tether for the aa substrate or the growing peptide chain and a condensation (C) domain that catalyzes the peptide bond formation. One notable characteristic of NRPS enzymology is the ability of NRPSs to incorporate non-proteinogenic aa into the NRPS product (<xref ref-type="bibr" rid="B121">Walsh et al., 2013</xref>). In addition, NRPSs typically follow the collinearity rule such that the substrate specificity, the number and the linear arrangement of the module within the assembly determines the composition of the NRPS product. Thus, in most cases, the candidate NRPS-encoding genomic scaffolds can be narrowed down easily by first matching the number of modules in the NRPSs to the number of aa residues linked by peptide bonds in the target non-ribosomal peptide product. The presence of the in-line tailoring domain in the NRPS assembly line such as an epimerization (E) or N-methylation (M) domains can further indicate that the NRPS product undergoes respective modification.</p>
<p>There are some exceptions to the co-linearity rule for some fungal NRPSs, in which certain domains or modules on these NRPSs are used iteratively. Notable examples are the fungal siderophore NRPSs, which involved the iterative use of one of the A domains (activating an identical aa residue more than once in a complete NRPS catalytic cycle). These NRPSs harbor additional T&#x02013;C partial modules that extend the non-ribosomal peptide products beyond the number of complete A&#x02013;T&#x02013;C modules in the NRPSs (<xref ref-type="bibr" rid="B104">Schwecke et al., 2006</xref>; <xref ref-type="bibr" rid="B67">Johnson, 2008</xref>). There are also other non-canonical NRPSs, such as those that synthesize fungal cyclooligomer depsipetides (<xref ref-type="bibr" rid="B59">Glinski et al., 2002</xref>; <xref ref-type="bibr" rid="B132">Xu et al., 2008</xref>; <xref ref-type="bibr" rid="B112">Sussmuth et al., 2011</xref>) and the recently identified fungisporin NRPS from <italic>P. chrysogenum</italic> (<xref ref-type="bibr" rid="B3">Ali et al., 2014</xref>).</p>
<p>Pioneering studies by the Marahiel group led to the development of the A domain 10 aa code, which aided in prediction of the substrate specificities of adenylation of uncharacterized bacterial NRPSs via sequence alignment (<xref ref-type="bibr" rid="B32">Conti et al., 1997</xref>; <xref ref-type="bibr" rid="B109">Stachelhaus et al., 1999</xref>). Other algorithms that predict adenylation domain substrate specificity and online servers that utilized these algorithms were subsequently developed and are now in widespread use for preliminary bioinformatic characterization of NRPS genes (<xref ref-type="bibr" rid="B20">Challis et al., 2000</xref>; <xref ref-type="bibr" rid="B102">R&#x000F6;ttig et al., 2011</xref>). However, the utility of the A domain 10 aa code and NRPS analysis software programs for prediction of fungal NRPS domain is still limited for fungi compared to bacteria, as there remains a lack of sequence-substrate specificity relationship for fungal NRPSs. Nonetheless, the new NRPSpredictor2 software incorporates A domain substrate specificity prediction for fungal NRPSs and can be a useful starting point (<xref ref-type="bibr" rid="B102">R&#x000F6;ttig et al., 2011</xref>). Some progress have been made in this area for the A domain 10 aa codes for anthranilate (<xref ref-type="bibr" rid="B4">Ames and Walsh, 2010</xref>), &#x003B1;-keto acids (<xref ref-type="bibr" rid="B119">Wackler et al., 2012</xref>), and <sc>L</sc>-tryptophan (<xref ref-type="bibr" rid="B130">Xu et al., 2014a</xref>). Phylogenetic and structural analysis of fungal A domains may also yield some insights into the possible substrate of the A domain and the structural class of the final product (<xref ref-type="bibr" rid="B4">Ames and Walsh, 2010</xref>; <xref ref-type="bibr" rid="B16">Bushley and Turgeon, 2010</xref>). For instance, characterization of the anthranilate-activating fungal A domain in fumiquinazoline F biosynthesis (<xref ref-type="bibr" rid="B4">Ames and Walsh, 2010</xref>) led to the discovery of the gene clusters of fungal anthranilic acid-containing non-ribosomal peptides (<xref ref-type="bibr" rid="B56">Gao et al., 2011</xref>; <xref ref-type="bibr" rid="B61">Haynes et al., 2012</xref>, <xref ref-type="bibr" rid="B62">2013</xref>).</p>
</sec>
<sec><title>Tailoring enzymes</title>
<p>Since genes for the biosynthesis of SMs are typically clustered together, the types of tailoring reaction the polyketide or non-ribosomal backbone undergoes can be surmised based on the type of functional domains encoded in vicinity of the PKS or NRPS gene, respectively. Essentially, if the backbone biosynthetic gene analysis narrowed down the target SM gene cluster to a couple of possibilities, the combination of tailoring enzymes that matches the SM structure would allow one to confidently pinpoint the right cluster. Tailoring enzymes involved in functional-group transfer such as methyltransferases, acyltransferases, prenyltransferases, and halogenases are especially helpful in correlating a gene cluster to its corresponding compound since there is little ambiguity on what type of reaction these enzymes catalyzes. As described in Section 4 below, the co-localization of multiple methyltransferase genes with an NR-PKS gene was critical toward the discovery of the griseofulvin gene cluster, while the presence of a prenyltransferase gene and a methyltransferase gene flanking an NR-PKS gene pinpointed the correct viridicatumtoxin gene cluster (<xref ref-type="bibr" rid="B24">Chooi et al., 2010</xref>). Genes encoding redox enzymes, such as flavoenzymes, cytochrome P450s, and non-heme iron oxygenases, can suggest that the product of the cluster undergoes oxidative transformation. However, their role in biosynthesis can be ambiguous as these enzymes can catalyze diverse redox reactions including hydroxylation, epoxidation, oxidative cleavage, and rearrangement. This is illustrated by our work in functional elucidation of the redox enzymes in viridicatumtoxin, tryptoquialanine, and echinocandin B pathways. As described below, some of the functions of these redox enzymes turned out to be quite surprising. Thus, caution is advised on inferring the function of genes for redox enzymes based on conserved domain analysis, especially in cases where there is no known characterized enzyme that share high sequence similarity with the enzyme of interest.</p>
</sec>
</sec></sec>
<sec><title>EXAMPLES ILLUSTRATING THE STRATEGIES FOR TARGETED SM GENE CLUSTER DISCOVERY</title>
<sec><title>CASE STUDY A: GRISEOFULVIN</title>
<p>Griseofulvin is an antifungal polyketide made by several different <italic>Penicillium</italic> species (<xref ref-type="bibr" rid="B52">Frisvad and Samson, 2004</xref>). To identify the griseofulvin gene cluster, a curation of the PKS genes in <italic>P. aethiopicum</italic> was performed. Using a local BLAST search of the genome using an arbitrary KS as a query, 30 putative intact PKS genes were found within the genome of <italic>P. aethiopicum</italic>. Comparison of the PKS genes in <italic>P. aethiopicum</italic> and the closely related species <italic>P. chrysogenum</italic> revealed that the former contains six NR-PKS genes that are not orthologous to <italic>P. chrysogenum</italic> NR-PKS genes (<xref ref-type="bibr" rid="B116">van den Berg et al., 2008</xref>; <xref ref-type="bibr" rid="B24">Chooi et al., 2010</xref>). Since <italic>P. chrysogenum</italic> was not known to produce griseofulvin, it was inferred that one of the six non-orthologous NR-PKS genes in <italic>P. aethiopicum</italic> was responsible for the biosynthesis of griseofulvin (<bold>Figure <xref ref-type="fig" rid="F3">3A</xref></bold>).</p>
<p>Based on the distinct structural features of griseofulvin, the presence of multiple methyltransferase genes as well as a chlorinase gene in the gene cluster encoding griseofulvin is expected (<bold>Figure <xref ref-type="fig" rid="F4">4</xref></bold>). Using these search criteria, the <italic>gsf</italic> gene cluster, the candidate gene cluster for the biosynthesis of the compound, was found. Along with the expected three <italic>S</italic>-adenosyl-methionine (SAM) methyltransferase genes (<italic>gsfB-D</italic>) and a flavin-dependent halogenase gene (<italic>gsfI</italic>) flanking the NR-PKS gene<italic> gsfA</italic>, genes for the two redox enzymes GsfE and GsfF were also found within the cluster (<bold>Figure <xref ref-type="fig" rid="F4">4</xref></bold>). Deletion of the <italic>gsfA</italic> gene led to loss of production of griseofulvin in <italic>P. aethiopicum</italic> and thus confirming the role of the <italic>gsf</italic> cluster in the biosynthesis of the compound (<xref ref-type="bibr" rid="B24">Chooi et al., 2010</xref>). A subsequent paper by our group revealed the function of the genes in the cluster as well as the regioselectivity of the <italic>O</italic>-methyltransferases GsfB, C, and D (<xref ref-type="bibr" rid="B18">Cacho et al., 2013</xref>). A combination of gene deletion studies and reconstitution of the enzymatic reaction in the same study also revealed the role of GsfF and GsfE in the grisan ring formation and cyclohexadienone reduction, respectively.</p>
<fig id="F4" position="float">
<label>FIGURE 4</label>
<caption><p><bold>Griseofulvin biosynthetic gene clusterin <italic>P. aethiopicum</italic>.</bold> Shown are the roles of the genes in the biogenesis of the compound. The non-reducing polyketide synthase (NR-PKS) gene <italic>gsfA</italic> is found via comparison of PKS genes in <italic>P. aethiopicum</italic> and <italic>P. chrysogenum</italic>. Two structural features, <italic>O</italic>-methylation and chlorination (in red text), were used as criteria to find the candidate gene clusters.</p></caption>
<graphic xlink:href="fmicb-05-00774-g004.tif"/>
</fig>
</sec>
<sec><title>CASE STUDY B: VIRIDICATUMTOXIN</title>
<p>Viridicatumtoxin is another aromatic polyketide synthesized by <italic>P. aethiopicum</italic>. It contained a naphthacenedione core reminiscent of the core of the bacterial tetracyclines. Moreover, viridicatumtoxin also contained a cyclized geranyl moiety (<bold>Figure <xref ref-type="fig" rid="F5">5</xref></bold>), indicative of the possible presence of a prenyltransferase gene within the cluster. As in the case of griseofulvin, viridicatumtoxin was not known to be produced by <italic>P. chrysogenum</italic> and likewise, one of the non-orthologous NR-PKS genes in <italic>P. aethiopicum</italic> was presumably responsible for the biosynthesis of the compound (<bold>Figure <xref ref-type="fig" rid="F3">3A</xref></bold>). In addition to the aforementioned cyclized terpene group, viridicatumtoxin also contained an <italic>O-</italic>methyl group similar to what was found in griseofulvin. Using these two structural features as &#x0201C;landmarks,&#x0201D; the viridicatumtoxin biosynthetic (<italic>vrt</italic>) gene cluster is localized on a scaffold (<xref ref-type="bibr" rid="B24">Chooi et al., 2010</xref>). Flanking the NR-PKS gene <italic>vrtA</italic> are genes encoding a farnesyl diphosphate synthase analog (geranyl diphosphate synthase), a prenyltransferase gene <italic>vrtC</italic> and a SAM-dependent <italic>O</italic>-methyltransferase gene <italic>vrtF</italic>; all three are in agreement with the two distinctive chemical features found in the compound. Deletion of the NR-PKS gene <italic>vrtA</italic> confirmed the role of the cluster in the biogenesis of the compound (<xref ref-type="bibr" rid="B24">Chooi et al., 2010</xref>).</p>
<fig id="F5" position="float">
<label>FIGURE 5</label>
<caption><p><bold>Viridicatumtoxin biosynthetic gene cluster in <italic>P. aethiopicum</italic>.</bold> As with the case for the griseofulvin, the NR-PKS gene <italic>vrtA</italic> is non-orthologous to PKS genes in <italic>P. chrysogenum</italic>. The presence of the prenyltransferase gene <italic>vrtC</italic> and the <italic>O</italic>-methyltransferase gene <italic>vrtF</italic> (in red text) led to the identification of the gene cluster.</p></caption>
<graphic xlink:href="fmicb-05-00774-g005.tif"/>
</fig>
<p>Subsequent characterization of the prenyltransferase VrtC opened the doors toward the discovery of new fungal monoterpenoid biosynthetic gene cluster through genome mining (<xref ref-type="bibr" rid="B30">Chooi et al., 2012</xref>). In addition to the aforementioned genes, other genes encode five redox enzymes (<italic>vrtE</italic>, <italic>vrtG</italic>, <italic>vrtH</italic>, <italic>vrtI</italic>, and <italic>vrtK</italic>), a PLP-dependent threonine aldolase gene <italic>vrtJ</italic> and an acyl-CoA ligase gene <italic>vrtB</italic> were also found in the <italic>vrt</italic> gene cluster (<xref ref-type="bibr" rid="B24">Chooi et al., 2010</xref>). It was later found that VrtG and other related dimanganese-dependent thioesterase catalyze the Claisen-type cyclization in viridicatumtoxin and other fungal naphthacenediones (<xref ref-type="bibr" rid="B84">Li et al., 2011</xref>). Surprisingly, the cytochrome P450 VrtK, instead of a terpene cyclase, was shown to mediate the cyclization of the geranyl moiety found in viridicatumtoxin to afford a spirobicyclic structure fused to the tetracyclic core (<xref ref-type="bibr" rid="B26">Chooi et al., 2013b</xref>). Additionally, later studies revealed the role of VrtE and VrtH in the hydroxylation in the 5 and 12a positions in viridicatumtoxin, respectively (<xref ref-type="bibr" rid="B25">Chooi et al., 2013a</xref>,<xref ref-type="bibr" rid="B26">b</xref>).</p>
</sec>
<sec><title>CASE STUDY C: TRYPTOQUIALANINE</title>
<p>Tryptoquialanine (<bold>Figure <xref ref-type="fig" rid="F6">6</xref></bold>) is a quinazoline-containing indole alkaloid produced by <italic>P. aethiopicum</italic>. It is structurally similar to a known tremorgenic mycotoxin from <italic>A. clavatus</italic> tryptoquivaline. Due to the presence of the non-proteinogenic aa anthranilic acid in the scaffold of the peptide, it was inferred that the compound is assembled by a NRPS (<xref ref-type="bibr" rid="B4">Ames and Walsh, 2010</xref>). Comparative bioinformatic analysis of the NRPS genes between <italic>P. aethiopicum</italic> and <italic>P. chrysogenum</italic> revealed the presence of 10 NRPS genes in <italic>P. aethiopicum</italic> that are non-orthologous to the NRPS genes in <italic>P. chrysogenum</italic> and were therefore candidates for tryptoquialanine biosynthesis (<xref ref-type="bibr" rid="B56">Gao et al., 2011</xref>). On the other hand, comparative bioinformatic analysis of the NRPS genes in <italic>P. aethiopicum</italic> with NRPS genes with the tryptoquivaline-producing <italic>A. clavatus</italic>, led to the identification of two NRPS genes in <italic>P. aethiopicum</italic> that are orthologous with <italic>A. clavatus</italic> NRPS genes but not with <italic>P. chrysogenum</italic> NRPS genes (<bold>Figure <xref ref-type="fig" rid="F3">3B</xref></bold>). The two candidate NRPS genes, found on two separated scaffolds, encode a trimodule NRPS annotated as <italic>tqaA</italic> and a single-module NRPS annotated as <italic>tqaB</italic>. Primer walking and fosmid sequencing later revealed that both NRPS genes are co-localized in one segment of the genome. This demonstrates that cosmid library can be used to complement NGS-based targeted SM gene cluster discovery. Furthermore, the discovery of <italic>tqa</italic> cluster facilitated the identification of the tryptoquivaline biosynthetic genes, which are distributed on three separated genomic loci in <italic>A. clavatus</italic> (<xref ref-type="bibr" rid="B56">Gao et al., 2011</xref>).</p>
<fig id="F6" position="float">
<label>FIGURE 6</label>
<caption><p><bold>Tryptoquialanine biosynthetic gene cluster in <italic>P. aethiopicum.</italic></bold> The presence of acyltransferase gene <italic>tqaD</italic> and the anthranilic acid-activating A domain of TqaA (in red text) were clues to identification of the cluster.</p></caption>
<graphic xlink:href="fmicb-05-00774-g006.tif"/>
</fig>
<p>Sequence analysis of TqaA revealed high overall shared identity with Af12080, which was previously implicated to be involved in the biosynthesis of fumiquinazoline A, a related quinazoline-containing alkaloid from <italic>A. fumigatus</italic> (<xref ref-type="bibr" rid="B4">Ames and Walsh, 2010</xref><italic>)</italic>. Subsequent in-depth characterization of TqaA revealed its role in assembling fumiquinazoline F from anthranilate, <sc>L</sc>-tryptophan and <sc>L</sc>-alanine (<xref ref-type="bibr" rid="B57">Gao et al., 2012</xref>). The study also demonstrated the function of the terminal condensation domain in cyclization of the NRPS product and facilitated future genome mining of fungal cyclic non-ribosomal peptides. On the other hand, the second NRPS TqaB, based on the preliminary retrobiosynthetic analysis, presumably installs the 2-amino-isobutyric acid portion of the imidazolindolone pendant group (<xref ref-type="bibr" rid="B56">Gao et al., 2011</xref>). Interestingly, this reflects the arrangement of the aa building blocks, where anthranilic acid, <sc>L</sc>-tryptophan and <sc>L</sc>-alanine make up the main peptide chain while the fourth building block 2-amino-isobutyric acid is added to the modified <sc>L</sc>-tryptophan side chain.</p>
<p><bold>Figure <xref ref-type="fig" rid="F6">6</xref></bold> shows the other genes involved in the biosynthesis of tryptoquialanine (<xref ref-type="bibr" rid="B56">Gao et al., 2011</xref>). Based solely on the conserved domains of the enzymes encoded by the <italic>tqaA</italic> genes, only the function of TqaD in the installation of the acetyl group in the tryptoquialanine can be confidently assigned. The function of the remaining genes in the cluster could only be assigned by targeted deletion of individual genes (<xref ref-type="bibr" rid="B56">Gao et al., 2011</xref>). This revealed that <italic>tqaE</italic> and <italic>tqaH</italic> are responsible for <italic>N</italic>-hydroxylation and the 2,3-epoxidation of the pendant indole ring of tryptoquialanine, respectively. <italic>tqaC</italic> encodes a short-chain dehydrogenase that reduces the ketone in tryptoquialanone. Meanwhile, <italic>tqaG</italic> and <italic>tqaI</italic> were demonstrated to encode enzymes that mediate the formation of <italic>N</italic>-deoxytryptoquialanone. Finally, knockout of <italic>tqaL</italic> and <italic>tqaM</italic> demonstrated that both genes are required for the biosynthesis of the 2-amino-isobutyric acid.</p>
</sec>
<sec><title>CASE STUDY D: ECHINOCANDIN B</title>
<p>Echinocandin B is a cyclic lipopeptide made by the ascomycetes <italic>E. rugulosa.</italic> Due to their efficacy against a broad range of <italic>Candida</italic> species, semisynthetic derivatives of natural echinocandins such as anidulafungin, micafungin, and caspofungin are currently in use as frontline treatment against invasive candidiasis (<xref ref-type="bibr" rid="B73">Kett et al., 2011</xref>). Structurally, echinocandin B is consist of six aa: (4<italic>R</italic>, 5<italic>R</italic>)-4,5-dihydroxy-<sc>L</sc>-ornithine, two units of <sc>L</sc>-threonine, (3<italic>R</italic>)-3-hydroxy-<sc>L</sc>-proline, (3<italic>S</italic>, 4<italic>S</italic>)-3,4-dihydroxy-<sc>L</sc>-homotyrosine, and (3<italic>S</italic>, 4<italic>S</italic>)-3-methyl-4-hydroxy-<sc>L</sc>-proline. In addition, a linoleic acid is appended to the cyclic peptide comprising of the six-aa (<bold>Figure <xref ref-type="fig" rid="F7">7</xref></bold>). As with the examples given above, genome mining for the echinocandin biosynthetic gene cluster is initiated with comparative genomic analysis; in this case, between the NRPS genes in <italic>E. rugulosa</italic> and <italic>A. nidulans</italic> A4 (<xref ref-type="bibr" rid="B118">von Dohren, 2009</xref>; <xref ref-type="bibr" rid="B19">Cacho et al., 2012</xref>). Since <italic>A. nidulans</italic> A4 was not known to produce echinocandin B, NRPS genes orthologous to both species can be eliminated as candidates for echinocandin synthetase. Annotation of the unique NRPS genes in <italic>E. rugulosa</italic> revealed that only one of the ten non-orthologous NRPS genes encode for a six-module NRPS (one NRPS module is minimally consist of a C, A and T domain), which was the required number for the assembly of echinocandin B based on the collinearity rule (<bold>Figure <xref ref-type="fig" rid="F3">3C</xref></bold>). In addition, the six-module NRPS, annotated as EcdA, contained an additional C-terminal condensation domain that is expected to catalyze the cyclization of the full-length peptide product (<xref ref-type="bibr" rid="B57">Gao et al., 2012</xref>), in accordance with the cyclic nature of the compound.</p>
<fig id="F7" position="float">
<label>FIGURE 7</label>
<caption><p><bold>Echinocandin B biosynthetic gene cluster in <italic>E. rugulosa</italic>.</bold> The gene cluster was found in two loci (<italic>ecd</italic> and <italic>hty</italic>). The <italic>ecd</italic> cluster contained a six module NRPS gene <italic>ecdA</italic> and an acyl-CoA ligase gene <italic>ecdI</italic> (in red text), indicating that the gene cluster is involved in the biosynthesis of a lipo-hexapeptide SM. The separate homotyrosine biosynthesis <italic>hty</italic> gene cluster at a different locus was located by searching for an isopropylmalate synthase (IPMS) homolog (in red text).</p></caption>
<graphic xlink:href="fmicb-05-00774-g007.tif"/>
</fig>
<p>In addition to the NRPS gene <italic>ecdA</italic>, other genes found in the cluster in agreement to the chemical features of the compound were also found in the cluster such as the acyl-CoA ligase gene <italic>ecdI</italic> and the redox enzymes <italic>ecdG, H</italic> and <italic>K</italic> (<bold>Figure <xref ref-type="fig" rid="F7">7</xref></bold>; <xref ref-type="bibr" rid="B19">Cacho et al., 2012</xref>). The presence of <italic>ecdI</italic>, in conjunction with its co-localization with the NRPS gene <italic>ecdA</italic>, was highly indicative that the product of the <italic>ecd</italic> gene cluster is a lipopeptide since EcdI, based on its predicted conserved domain, belongs to a family of enzyme that can convert a carboxylic acid to the more reactive acyl-CoA. Enzymatic characterization of EcdI verified its role in the activation and transfer of linoleic acid to the N-terminal thiolation domain of EcdA.</p>
<p>While the presence of the multiple genes for redox enzymes in the locus suggested that the product of the <italic>ecd</italic> gene cluster undergoes a plethora of oxidative modification steps, the low shared identity of the protein sequences of EcdG, H and K with of then characterized protein in sequence databases hindered the deciphering of the exact roles of the enzymes solely by bioinformatic analysis. Later gene knockout and enzymatic reconstitution study revealed the regioselectivity of EcdG toward the C3 of <sc>L</sc>-homotyrosine and regioselectivity of EcdH toward C4 and C5 of <sc>L</sc>-ornithine in echinocandin B biosynthesis (<xref ref-type="bibr" rid="B66">Jiang et al., 2013</xref>). On the other hand, EcdK was revealed to perform the two-step oxidation of <sc>L</sc>-leucine to afford 5-hydroxy-<sc>L</sc>-leucine and &#x003B3;-methyl-glutamic acid-&#x003B3;-semialdehyde en route to the biosynthesis of (4<italic>R</italic>)-<italic>R</italic>-methyl-<sc>L</sc>-proline (<xref ref-type="bibr" rid="B66">Jiang et al., 2013</xref>).</p>
<p>Notably missing in the <italic>ecd</italic> gene cluster, however, were the genes required for the biosynthesis of <sc>L</sc>-homotyrosine, another non-proteinogenic aa building block of echinocandin B. Based on previous labeling studies on <sc>L</sc>-homotyrosine biosynthesis demonstrated that the latter originated from acetate and <sc>L</sc>-tyrosine (<xref ref-type="bibr" rid="B1">Adefarati et al., 1991</xref>). It was also proposed that the pathway is analogous to that of <sc>L</sc>-leucine biosynthesis, the first step of which is catalyzed by isopropylmalate synthase (IPMS). Thus, we predicted that a homolog of IPMS was involved in the first step of the biosynthesis of <sc>L</sc>-homotyrosine and found by genome mining the IPMS-like gene <italic>htyA</italic> in the <italic>E. rugulosa</italic> genome that is non-orthologous to the IPMS gene in <italic>A. nidulans</italic> A4 genome. Subsequent deletion of <italic>htyA</italic> implicated its role in the biosynthesis of <sc>L</sc>-homotyrosine (<xref ref-type="bibr" rid="B19">Cacho et al., 2012</xref>). Flanking the <italic>htyA</italic> gene were additional genes presumably involved in the biosynthesis of <sc>L</sc>-homotyrosine (<italic>htyB</italic>, <italic>C</italic>, and <italic>D</italic>) as well as two additional oxidase genes <italic>htyE</italic> and <italic>htyF</italic>.</p>
<p>Shortly after the studies on echinocandin B biosynthesis were reported, the biosynthetic gene cluster for the structurally-related pneumocandin B was described (<xref ref-type="bibr" rid="B21">Chen et al., 2013</xref>). The pneumocandin biosynthetic gene clusters contained gene homologs echinocandin gene cluster include <italic>ecdA</italic>, <italic>ecdI</italic>, cytochrome P450 genes (e<italic>cdH</italic> and <italic>htyF</italic>), <italic>ecdG</italic>, <italic>htyE</italic>, and <italic>ecdK</italic> and the <sc>L</sc>-homotyrosine biosynthetic genes. Interestingly, whereas the echinocandin B biosynthetic genes are located at multiple loci, the pneumocandin biosynthetic genes are situated in a single gene cluster. In addition, the pneumocandin biosynthetic gene cluster also contained a HR-PKS gene for the biosynthesis of the 10,12-dimethyl-myristic acid and an additional 2-oxoglutarate, non-heme iron-dependent oxygenase gene (GLAREA10042) which presumably is involved in the hydroxylation of <sc>L</sc>-glutamine at C3; both corresponding to the structural features of pneumocandin B but not of echinocandin B. The recent availability of several fungal genomes that encode production of the echinocandin family of lipopeptides has allowed bioinformatics comparison of these gene clusters and enables functional prediction of the unique pathway genes in correspondence to the unique structural features of individual echinocandin analogs (<xref ref-type="bibr" rid="B10">Bills et al., 2014</xref>).</p>
</sec>
</sec>
<sec><title>CONCLUSION AND OUTLOOK</title>
<p>Next-generation sequencing technologies have significantly accelerated the process of targeted SM gene cluster discovery. The whole process from extraction of genomic DNA of the producing-fungus, sequencing, to initial identification of the target SM gene cluster can occur in one to three months. Using the described strategies, we can often identify the target SM gene cluster in the genome, within 2 weeks from receiving the sequencing data, with high accuracy. Compared to the traditional genomic library screening approach, which can take years beginning from the process of identifying the first biosynthetic gene to finally obtaining the complete gene cluster, such pace of discovery is unimaginable as recent as half a decade ago. The more arduous task ahead is the subsequent gene cluster verification and characterization as fungal genetic transformation protocols need to be developed for individual organisms. The development of heterologous expression systems and <italic>in vitro</italic> recombinant enzyme characterization methods greatly complement the genetic approach and can provide biosynthetic insights unattainable by traditional knockout approach. More and more fungal natural product research labs are expected to take advantage of NGS technology for identifying the gene cluster encoding the biosynthesis of SMs of interest. The increasing throughput and lowering cost for NGS would also encourage the simultaneous sequencing of multiple strains that produce a same family of compounds, which enables the use of powerful comparative genomics tools in the gene cluster identification and promotes combinatorial biosynthesis of fungal natural products.</p>
</sec>
<sec><title>Conflict of Interest Statement</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
</body>
<back>
<ack>
<p>Ralph A. Cacho is supported by the NIH Chemistry-Biology Interface Training Grant (NRSA GM-08496) and UCLA Graduate Division. Yi Tang is supported by the NIH grant 1DP1GM106413. Yit-Heng Chooi is currently supported by an Australian Research Council (ARC) Discovery Early Career Researcher Award (DECRA).</p>
</ack>
<ref-list>
<title>REFERENCES</title>
<ref id="B1"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Adefarati</surname> <given-names>A. A.</given-names></name> <name><surname>Giacobbe</surname> <given-names>R. A.</given-names></name> <name><surname>Hensens</surname> <given-names>O. D.</given-names></name> <name><surname>Tkacz</surname> <given-names>J. S.</given-names></name></person-group> (<year>1991</year>). <article-title>Biosynthesis of L-671,329, an echinocandin-type antibiotic produced by Zalerion arboricola: origins of some of the unusual amino acids and the dimethylmyristic acid side chain.</article-title> <source><italic>J. Am. Chem. Soc.</italic></source> <volume>113</volume> <fpage>3542</fpage>&#x02013;<lpage>3545</lpage>. <pub-id pub-id-type="doi">10.1021/ja00009a048</pub-id></citation></ref>
<ref id="B2"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ahuja</surname> <given-names>M.</given-names></name> <name><surname>Chiang</surname> <given-names>Y. M.</given-names></name> <name><surname>Chang</surname> <given-names>S. L.</given-names></name> <name><surname>Praseuth</surname> <given-names>M. B.</given-names></name> <name><surname>Entwistle</surname> <given-names>R.</given-names></name> <name><surname>Sanchez</surname> <given-names>J. F.</given-names></name><etal/></person-group> (<year>2012</year>). <article-title>Illuminating the diversity of aromatic polyketide synthases in <italic>Aspergillus nidulans</italic>.</article-title> <source><italic>J. Am. Chem. Soc.</italic></source> <volume>134</volume> <fpage>8212</fpage>&#x02013;<lpage>8221</lpage>. <pub-id pub-id-type="doi">10.1021/ja3016395</pub-id></citation></ref>
<ref id="B3"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ali</surname> <given-names>H.</given-names></name> <name><surname>Ries</surname> <given-names>M. I.</given-names></name> <name><surname>Lankhorst</surname> <given-names>P. P.</given-names></name> <name><surname>Van Der Hoeven</surname> <given-names>R. A.</given-names></name> <name><surname>Schouten</surname> <given-names>O. L.</given-names></name> <name><surname>Noga</surname> <given-names>M.</given-names></name><etal/></person-group> (<year>2014</year>). <article-title>A non-canonical NRPS is involved in the synthesis of fungisporin and related hydrophobic cyclic tetrapeptides in <italic>Penicillium chrysogenum</italic>.</article-title> <source><italic>PLoS ONE</italic></source> <volume>9</volume>:<issue>e98212</issue>. <pub-id pub-id-type="doi">10.1371/journal.pone.0098212</pub-id></citation></ref>
<ref id="B4"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ames</surname> <given-names>B. D.</given-names></name> <name><surname>Walsh</surname> <given-names>C. T.</given-names></name></person-group> (<year>2010</year>). <article-title>Anthranilate-activating modules from fungal nonribosomal peptide assembly lines.</article-title> <source><italic>Biochemistry</italic></source> <volume>49</volume> <fpage>3351</fpage>&#x02013;<lpage>3365</lpage>. <pub-id pub-id-type="doi">10.1021/bi100198y</pub-id></citation></ref>
<ref id="B5"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Awakawa</surname> <given-names>T.</given-names></name> <name><surname>Yokota</surname> <given-names>K.</given-names></name> <name><surname>Funa</surname> <given-names>N.</given-names></name> <name><surname>Doi</surname> <given-names>F.</given-names></name> <name><surname>Mori</surname> <given-names>N.</given-names></name> <name><surname>Watanabe</surname> <given-names>H.</given-names></name><etal/></person-group> (<year>2009</year>). <article-title>Physically discrete beta-lactamase-type thioesterase catalyzes product release in atrochrysone synthesis by iterative type I polyketide synthase.</article-title> <source><italic>Chem. Biol.</italic></source> <volume>16</volume> <fpage>613</fpage>&#x02013;<lpage>623</lpage>. <pub-id pub-id-type="doi">10.1016/j.chembiol.2009.04.004</pub-id></citation></ref>
<ref id="B6"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bailey</surname> <given-names>A. M.</given-names></name> <name><surname>Cox</surname> <given-names>R. J.</given-names></name> <name><surname>Harley</surname> <given-names>K.</given-names></name> <name><surname>Lazarus</surname> <given-names>C. M.</given-names></name> <name><surname>Simpson</surname> <given-names>T. J.</given-names></name> <name><surname>Skellam</surname> <given-names>E.</given-names></name></person-group> (<year>2007</year>). <article-title>Characterisation of 3-methylorcinaldehyde synthase (MOS) in <italic>Acremonium strictum</italic>: first observation of a reductive release mechanism during polyketide biosynthesis.</article-title> <source><italic>Chem. Commun. (Camb.)</italic></source> <volume>39</volume> <fpage>4053</fpage>&#x02013;<lpage>4055</lpage>. <pub-id pub-id-type="doi">10.1039/b708614h</pub-id></citation></ref>
<ref id="B7"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Balkovec</surname> <given-names>J. M.</given-names></name> <name><surname>Hughes</surname> <given-names>D. L.</given-names></name> <name><surname>Masurekar</surname> <given-names>P. S.</given-names></name> <name><surname>Sable</surname> <given-names>C. A.</given-names></name> <name><surname>Schwartz</surname> <given-names>R. E.</given-names></name> <name><surname>Singh</surname> <given-names>S. B.</given-names></name></person-group> (<year>2014</year>). <article-title>Discovery and development of first in class antifungal caspofungin (CANCIDAS(R))&#x02013;a case study.</article-title> <source><italic>Nat. Prod. Rep.</italic></source> <volume>31</volume> <fpage>15</fpage>&#x02013;<lpage>34</lpage>. <pub-id pub-id-type="doi">10.1039/c3np70070d</pub-id></citation></ref>
<ref id="B8"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Beck</surname> <given-names>J.</given-names></name> <name><surname>Ripka</surname> <given-names>S.</given-names></name> <name><surname>Siegner</surname> <given-names>A.</given-names></name> <name><surname>Schiltz</surname> <given-names>E.</given-names></name> <name><surname>Schweizer</surname> <given-names>E.</given-names></name></person-group> (<year>1990</year>). <article-title>The multifunctional 6-methylsalicylic acid synthase gene of <italic>Penicillium patulum</italic>. Its gene structure relative to that of other polyketide synthases.</article-title> <source><italic>Eur. J. Biochem.</italic></source> <volume>192</volume> <fpage>487</fpage>&#x02013;<lpage>498</lpage>. <pub-id pub-id-type="doi">10.1111/j.1432-1033.1990.tb19252.x</pub-id></citation></ref>
<ref id="B9"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bentley</surname> <given-names>R.</given-names></name></person-group> (<year>1999</year>). <article-title>Secondary metabolite biosynthesis: the first century.</article-title> <source><italic>Crit. Rev. Biotechnol.</italic></source> <volume>19</volume> <fpage>1</fpage>&#x02013;<lpage>40</lpage>. <pub-id pub-id-type="doi">10.1080/0738-859991229189</pub-id></citation></ref>
<ref id="B10"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bills</surname> <given-names>G.</given-names></name> <name><surname>Li</surname> <given-names>Y.</given-names></name> <name><surname>Chen</surname> <given-names>L.</given-names></name> <name><surname>Yue</surname> <given-names>Q.</given-names></name> <name><surname>Niu</surname> <given-names>X. M.</given-names></name> <name><surname>An</surname> <given-names>Z.</given-names></name></person-group> (<year>2014</year>). <article-title>New insights into the echinocandins and other fungal non-ribosomal peptides and peptaibiotics.</article-title> <source><italic>Nat. Prod. Rep.</italic></source> <volume>31</volume> <fpage>1348</fpage>&#x02013;<lpage>1375</lpage>. <pub-id pub-id-type="doi">10.1039/c4np00046c</pub-id></citation></ref>
<ref id="B11"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bingle</surname> <given-names>L. E.</given-names></name> <name><surname>Simpson</surname> <given-names>T. J.</given-names></name> <name><surname>Lazarus</surname> <given-names>C. M.</given-names></name></person-group> (<year>1999</year>). <article-title>Ketosynthase domain probes identify two subclasses of fungal polyketide synthase genes.</article-title> <source><italic>Fungal Genet. Biol.</italic></source> <volume>26</volume> <fpage>209</fpage>&#x02013;<lpage>223</lpage>. <pub-id pub-id-type="doi">10.1006/fgbi.1999.1115</pub-id></citation></ref>
<ref id="B12"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Blin</surname> <given-names>K.</given-names></name> <name><surname>Medema</surname> <given-names>M. H.</given-names></name> <name><surname>Kazempour</surname> <given-names>D.</given-names></name> <name><surname>Fischbach</surname> <given-names>M. A.</given-names></name> <name><surname>Breitling</surname> <given-names>R.</given-names></name> <name><surname>Takano</surname> <given-names>E.</given-names></name><etal/></person-group> (<year>2013</year>). <article-title>AntiSMASH 2.0&#x02013;a versatile platform for genome mining of secondary metabolite producers.</article-title> <source><italic>Nucleic Acids Res.</italic></source> <volume>41</volume> <fpage>W204</fpage>&#x02013;<lpage>W212</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gkt449</pub-id></citation></ref>
<ref id="B13"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Boettger</surname> <given-names>D.</given-names></name> <name><surname>Hertweck</surname> <given-names>C.</given-names></name></person-group> (<year>2013</year>). <article-title>Molecular diversity sculpted by fungal PKS-NRPS hybrids.</article-title> <source><italic>Chembiochem</italic></source> <volume>14</volume> <fpage>28</fpage>&#x02013;<lpage>42</lpage>. <pub-id pub-id-type="doi">10.1002/cbic.201200624</pub-id></citation></ref>
<ref id="B14"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Britton</surname> <given-names>S.</given-names></name> <name><surname>Palacios</surname> <given-names>R.</given-names></name></person-group> (<year>1982</year>). <article-title>Cyclosporin A&#x02013;usefulness, risks and mechanism of action.</article-title> <source><italic>Immunol. Rev.</italic></source> <volume>65</volume> <fpage>5</fpage>&#x02013;<lpage>22</lpage>. <pub-id pub-id-type="doi">10.1111/j.1600-065X.1982.tb00425.x</pub-id></citation></ref>
<ref id="B15"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Brown</surname> <given-names>D. W.</given-names></name> <name><surname>Yu</surname> <given-names>J. H.</given-names></name> <name><surname>Kelkar</surname> <given-names>H. S.</given-names></name> <name><surname>Fernandes</surname> <given-names>M.</given-names></name> <name><surname>Nesbitt</surname> <given-names>T. C.</given-names></name> <name><surname>Keller</surname> <given-names>N. P.</given-names></name><etal/></person-group> (<year>1996</year>). <article-title>Twenty-five coregulated transcripts define a sterigmatocystin gene cluster in <italic>Aspergillus nidulans</italic>.</article-title> <source><italic>Proc. Natl. Acad. Sci. U.S.A.</italic></source> <volume>93</volume> <fpage>1418</fpage>&#x02013;<lpage>1422</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.93.4.1418</pub-id></citation></ref>
<ref id="B16"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bushley</surname> <given-names>K. E.</given-names></name> <name><surname>Turgeon</surname> <given-names>B. G.</given-names></name></person-group> (<year>2010</year>). <article-title>Phylogenomics reveals subfamilies of fungal nonribosomal peptide synthetases and their evolutionary relationships.</article-title> <source><italic>BMC Evol. Biol.</italic></source> <volume>10</volume>:<issue>26</issue>. <pub-id pub-id-type="doi">10.1186/1471-2148-10-26</pub-id></citation></ref>
<ref id="B17"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Butler</surname> <given-names>J.</given-names></name> <name><surname>Maccallum</surname> <given-names>I.</given-names></name> <name><surname>Kleber</surname> <given-names>M.</given-names></name> <name><surname>Shlyakhter</surname> <given-names>I. A.</given-names></name> <name><surname>Belmonte</surname> <given-names>M. K.</given-names></name> <name><surname>Lander</surname> <given-names>E. S.</given-names></name><etal/></person-group> (<year>2008</year>). <article-title>ALLPATHS: de novo assembly of whole-genome shotgun microreads.</article-title> <source><italic>Genome Res.</italic></source> <volume>18</volume> <fpage>810</fpage>&#x02013;<lpage>820</lpage>. <pub-id pub-id-type="doi">10.1101/gr.7337908</pub-id></citation></ref>
<ref id="B18"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cacho</surname> <given-names>R. A.</given-names></name> <name><surname>Chooi</surname> <given-names>Y. H.</given-names></name> <name><surname>Zhou</surname> <given-names>H.</given-names></name> <name><surname>Tang</surname> <given-names>Y.</given-names></name></person-group> (<year>2013</year>). <article-title>Complexity generation in fungal polyketide biosynthesis: a spirocycle-forming P450 in the concise pathway to the antifungal drug griseofulvin.</article-title> <source><italic>ACS Chem. Biol.</italic></source> <volume>8</volume> <fpage>2322</fpage>&#x02013;<lpage>2330</lpage>. <pub-id pub-id-type="doi">10.1021/cb400541z</pub-id></citation></ref>
<ref id="B19"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cacho</surname> <given-names>R. A.</given-names></name> <name><surname>Jiang</surname> <given-names>W.</given-names></name> <name><surname>Chooi</surname> <given-names>Y. H.</given-names></name> <name><surname>Walsh</surname> <given-names>C. T.</given-names></name> <name><surname>Tang</surname> <given-names>Y.</given-names></name></person-group> (<year>2012</year>). <article-title>Identification and characterization of the echinocandin B biosynthetic gene cluster from <italic>Emericella rugulosa</italic> NRRL 11440.</article-title> <source><italic>J. Am. Chem. Soc.</italic></source> <volume>134</volume> <fpage>16781</fpage>&#x02013;<lpage>16790</lpage>. <pub-id pub-id-type="doi">10.1021/ja307220z</pub-id></citation></ref>
<ref id="B20"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Challis</surname> <given-names>G. L.</given-names></name> <name><surname>Ravel</surname> <given-names>J.</given-names></name> <name><surname>Townsend</surname> <given-names>C. A.</given-names></name></person-group> (<year>2000</year>). <article-title>Predictive, structure-based model of amino acid recognition by nonribosomal peptide synthetase adenylation domains.</article-title> <source><italic>Chem. Biol.</italic></source> <volume>7</volume> <fpage>211</fpage>&#x02013;<lpage>224</lpage>. <pub-id pub-id-type="doi">10.1016/S1074-5521(00)00091-0</pub-id></citation></ref>
<ref id="B21"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chen</surname> <given-names>L.</given-names></name> <name><surname>Yue</surname> <given-names>Q.</given-names></name> <name><surname>Zhang</surname> <given-names>X.</given-names></name> <name><surname>Xiang</surname> <given-names>M.</given-names></name> <name><surname>Wang</surname> <given-names>C.</given-names></name> <name><surname>Li</surname> <given-names>S.</given-names></name><etal/></person-group> (<year>2013</year>). <article-title>Genomics-driven discovery of the pneumocandin biosynthetic gene cluster in the fungus Glarea lozoyensis.</article-title> <source><italic>BMC Genomics</italic></source> <volume>14</volume>:<issue>339</issue>. <pub-id pub-id-type="doi">10.1186/1471-2164-14-339</pub-id></citation></ref>
<ref id="B22"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chettri</surname> <given-names>P.</given-names></name> <name><surname>Ehrlich</surname> <given-names>K. C.</given-names></name> <name><surname>Cary</surname> <given-names>J. W.</given-names></name> <name><surname>Collemare</surname> <given-names>J.</given-names></name> <name><surname>Cox</surname> <given-names>M. P.</given-names></name> <name><surname>Griffiths</surname> <given-names>S. A.</given-names></name><etal/></person-group> (<year>2013</year>). <article-title>Dothistromin genes at multiple separate loci are regulated by AflR.</article-title> <source><italic>Fungal Genet. Biol.</italic></source> <volume>51</volume> <fpage>12</fpage>&#x02013;<lpage>20</lpage>. <pub-id pub-id-type="doi">10.1016/j.fgb.2012.11.006</pub-id></citation></ref>
<ref id="B23"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chiang</surname> <given-names>Y. M.</given-names></name> <name><surname>Szewczyk</surname> <given-names>E.</given-names></name> <name><surname>Davidson</surname> <given-names>A. D.</given-names></name> <name><surname>Keller</surname> <given-names>N.</given-names></name> <name><surname>Oakley</surname> <given-names>B. R.</given-names></name> <name><surname>Wang</surname> <given-names>C. C.</given-names></name></person-group> (<year>2009</year>). <article-title>A gene cluster containing two fungal polyketide synthases encodes the biosynthetic pathway for a polyketide, asperfuranone, in <italic>Aspergillus nidulans</italic>.</article-title> <source><italic>J. Am. Chem. Soc.</italic></source> <volume>131</volume> <fpage>2965</fpage>&#x02013;<lpage>2970</lpage>. <pub-id pub-id-type="doi">10.1021/ja8088185</pub-id></citation></ref>
<ref id="B24"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chooi</surname> <given-names>Y. H.</given-names></name> <name><surname>Cacho</surname> <given-names>R.</given-names></name> <name><surname>Tang</surname> <given-names>Y.</given-names></name></person-group> (<year>2010</year>). <article-title>Identification of the viridicatumtoxin and griseofulvin gene clusters from <italic>Penicillium aethiopicum</italic>.</article-title> <source><italic>Chem. Biol.</italic></source> <volume>17</volume> <fpage>483</fpage>&#x02013;<lpage>494</lpage>. <pub-id pub-id-type="doi">10.1016/j.chembiol.2010.03.015</pub-id></citation></ref>
<ref id="B25"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chooi</surname> <given-names>Y. H.</given-names></name> <name><surname>Fang</surname> <given-names>J.</given-names></name> <name><surname>Liu</surname> <given-names>H.</given-names></name> <name><surname>Filler</surname> <given-names>S. G.</given-names></name> <name><surname>Wang</surname> <given-names>P.</given-names></name> <name><surname>Tang</surname> <given-names>Y.</given-names></name></person-group> (<year>2013a</year>). <article-title>Genome mining of a prenylated and immunosuppressive polyketide from pathogenic fungi.</article-title> <source><italic>Org. Lett.</italic></source> <volume>15</volume> <fpage>780</fpage>&#x02013;<lpage>783</lpage>. <pub-id pub-id-type="doi">10.1021/ol303435y</pub-id></citation></ref>
<ref id="B26"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chooi</surname> <given-names>Y. H.</given-names></name> <name><surname>Hong</surname> <given-names>Y. J.</given-names></name> <name><surname>Cacho</surname> <given-names>R. A.</given-names></name> <name><surname>Tantillo</surname> <given-names>D. J.</given-names></name> <name><surname>Tang</surname> <given-names>Y.</given-names></name></person-group> (<year>2013b</year>). <article-title>A cytochrome P450 serves as an unexpected terpene cyclase during fungal meroterpenoid biosynthesis.</article-title> <source><italic>J. Am. Chem. Soc.</italic></source> <volume>135</volume> <fpage>16805</fpage>&#x02013;<lpage>16808</lpage>. <pub-id pub-id-type="doi">10.1021/ja408966t</pub-id></citation></ref>
<ref id="B27"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chooi</surname> <given-names>Y.-H.</given-names></name> <name><surname>Krill</surname> <given-names>C.</given-names></name> <name><surname>Barrow</surname> <given-names>R. A.</given-names></name> <name><surname>Chen</surname> <given-names>S.</given-names></name> <name><surname>Trengove</surname> <given-names>R.</given-names></name> <name><surname>Oliver</surname> <given-names>R. P.</given-names></name><etal/></person-group> (<year>2015</year>). <article-title>An <italic>in planta</italic>-expressed polyketide synthase produces (<italic>R</italic>)-mellein in the wheat pathogen <italic>Parastagonospora nodorum</italic>.</article-title> <source><italic>Appl. Environ. Microbiol.</italic></source> <volume>81</volume> <fpage>177</fpage>&#x02013;<lpage>186</lpage>. <pub-id pub-id-type="doi">10.1128/AEM.02745-14</pub-id></citation></ref>
<ref id="B28"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chooi</surname> <given-names>Y.-H.</given-names></name> <name><surname>Solomon</surname> <given-names>P. S.</given-names></name></person-group> (<year>2014</year>). <article-title>A chemical ecogenomics approach to understand the roles of secondary metabolites in fungal cereal pathogens.</article-title> <source><italic>Front. Microbiol.</italic></source> <volume>5</volume>:<issue>640</issue>. <pub-id pub-id-type="doi">10.3389/fmicb.2014.00640</pub-id></citation></ref>
<ref id="B29"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chooi</surname> <given-names>Y. H.</given-names></name> <name><surname>Tang</surname> <given-names>Y.</given-names></name></person-group> (<year>2012</year>). <article-title>Navigating the fungal polyketide chemical space: from genes to molecules.</article-title> <source><italic>J. Org. Chem.</italic></source> <volume>77</volume> <fpage>9933</fpage>&#x02013;<lpage>9953</lpage>. <pub-id pub-id-type="doi">10.1021/jo301592k</pub-id></citation></ref>
<ref id="B30"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chooi</surname> <given-names>Y. H.</given-names></name> <name><surname>Wang</surname> <given-names>P.</given-names></name> <name><surname>Fang</surname> <given-names>J.</given-names></name> <name><surname>Li</surname> <given-names>Y.</given-names></name> <name><surname>Wu</surname> <given-names>K.</given-names></name> <name><surname>Wang</surname> <given-names>P.</given-names></name><etal/></person-group> (<year>2012</year>). <article-title>Discovery and characterization of a group of fungal polycyclic polyketide prenyltransferases.</article-title> <source><italic>J. Am. Chem. Soc.</italic></source> <volume>134</volume> <fpage>9428</fpage>&#x02013;<lpage>9437</lpage>. <pub-id pub-id-type="doi">10.1021/ja3028636</pub-id></citation></ref>
<ref id="B31"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chung</surname> <given-names>K. R.</given-names></name> <name><surname>Ehrenshaft</surname> <given-names>M.</given-names></name> <name><surname>Wetzel</surname> <given-names>D. K.</given-names></name> <name><surname>Daub</surname> <given-names>M. E.</given-names></name></person-group> (<year>2003</year>). <article-title>Cercosporin-deficient mutants by plasmid tagging in the asexual fungus <italic>Cercospora nicotianae</italic>.</article-title> <source><italic>Mol. Genet. Genomics</italic></source> <volume>270</volume> <fpage>103</fpage>&#x02013;<lpage>113</lpage>. <pub-id pub-id-type="doi">10.1007/s00438-003-0902-7</pub-id></citation></ref>
<ref id="B32"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Conti</surname> <given-names>E.</given-names></name> <name><surname>Stachelhaus</surname> <given-names>T.</given-names></name> <name><surname>Marahiel</surname> <given-names>M. A.</given-names></name> <name><surname>Brick</surname> <given-names>P.</given-names></name></person-group> (<year>1997</year>). <article-title>Structural basis for the activation of phenylalanine in the non-ribosomal biosynthesis of gramicidin S.</article-title> <source><italic>EMBO J.</italic></source> <volume>16</volume> <fpage>4174</fpage>&#x02013;<lpage>4183</lpage>. <pub-id pub-id-type="doi">10.1093/emboj/16.14.4174</pub-id></citation></ref>
<ref id="B33"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cortes</surname> <given-names>J.</given-names></name> <name><surname>Haydock</surname> <given-names>S. F.</given-names></name> <name><surname>Roberts</surname> <given-names>G. A.</given-names></name> <name><surname>Bevitt</surname> <given-names>D. J.</given-names></name> <name><surname>Leadlay</surname> <given-names>P. F.</given-names></name></person-group> (<year>1990</year>). <article-title>An unusually large multifunctional polypeptide in the erythromycin-producing polyketide synthase of <italic>Saccharopolyspora erythraea</italic>.</article-title> <source><italic>Nature</italic></source> <volume>348</volume> <fpage>176</fpage>&#x02013;<lpage>178</lpage>. <pub-id pub-id-type="doi">10.1038/348176a0</pub-id></citation></ref>
<ref id="B34"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cox</surname> <given-names>R. J.</given-names></name></person-group> (<year>2007</year>). <article-title>Polyketides, proteins and genes in fungi: programmed nano-machines begin to reveal their secrets.</article-title> <source><italic>Org. Biomol. Chem.</italic></source> <volume>5</volume> <fpage>2010</fpage>&#x02013;<lpage>2026</lpage>. <pub-id pub-id-type="doi">10.1039/b704420h</pub-id></citation></ref>
<ref id="B35"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cox</surname> <given-names>R. J.</given-names></name> <name><surname>Glod</surname> <given-names>F.</given-names></name> <name><surname>Hurley</surname> <given-names>D.</given-names></name> <name><surname>Lazarus</surname> <given-names>C. M.</given-names></name> <name><surname>Nicholson</surname> <given-names>T. P.</given-names></name> <name><surname>Rudd</surname> <given-names>B. A.</given-names></name><etal/></person-group> (<year>2004</year>). <article-title>Rapid cloning and expression of a fungal polyketide synthase gene involved in squalestatin biosynthesis.</article-title> <source><italic>Chem. Commun. (Camb.)</italic></source> <volume>21</volume> <fpage>2260</fpage>&#x02013;<lpage>2261</lpage>. <pub-id pub-id-type="doi">10.1039/b411973h</pub-id></citation></ref>
<ref id="B36"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Crawford</surname> <given-names>J. M.</given-names></name> <name><surname>Dancy</surname> <given-names>B. C. R.</given-names></name> <name><surname>Hill</surname> <given-names>E. A.</given-names></name> <name><surname>Udwary</surname> <given-names>D. W.</given-names></name> <name><surname>Townsend</surname> <given-names>C. A.</given-names></name></person-group> (<year>2006</year>). <article-title>Identification of a starter unit acyl-carrier protein transacylase domain in an iterative type I polyketide synthase.</article-title> <source><italic>Proc. Natl. Acad. Sci. U.S.A.</italic></source> <volume>103</volume> <fpage>16728</fpage>&#x02013;<lpage>16733</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.0604112103</pub-id></citation></ref>
<ref id="B37"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Crawford</surname> <given-names>J. M.</given-names></name> <name><surname>Korman</surname> <given-names>T. P.</given-names></name> <name><surname>Labonte</surname> <given-names>J. W.</given-names></name> <name><surname>Vagstad</surname> <given-names>A. L.</given-names></name> <name><surname>Hill</surname> <given-names>E. A.</given-names></name> <name><surname>Kamari-Bidkorpeh</surname> <given-names>O.</given-names></name><etal/></person-group> (<year>2009</year>). <article-title>Structural basis for biosynthetic programming of fungal aromatic polyketide cyclization.</article-title> <source><italic>Nature</italic></source> <volume>461</volume> <fpage>1139</fpage>&#x02013;<lpage>1143</lpage>. <pub-id pub-id-type="doi">10.1038/nature08475</pub-id></citation></ref>
<ref id="B38"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Crawford</surname> <given-names>J. M.</given-names></name> <name><surname>Thomas</surname> <given-names>P. M.</given-names></name> <name><surname>Scheerer</surname> <given-names>J. R.</given-names></name> <name><surname>Vagstad</surname> <given-names>A. L.</given-names></name> <name><surname>Kelleher</surname> <given-names>N. L.</given-names></name> <name><surname>Townsend</surname> <given-names>C. A.</given-names></name></person-group> (<year>2008</year>). <article-title>Deconstruction of iterative multidomain polyketide synthase function.</article-title> <source><italic>Science</italic></source> <volume>320</volume> <fpage>243</fpage>&#x02013;<lpage>246</lpage>. <pub-id pub-id-type="doi">10.1126/science.1154711</pub-id></citation></ref>
<ref id="B39"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Crawford</surname> <given-names>J. M.</given-names></name> <name><surname>Townsend</surname> <given-names>C. A.</given-names></name></person-group> (<year>2010</year>). <article-title>New insights into the formation of fungal aromatic polyketides.</article-title> <source><italic>Nat. Rev. Microbiol.</italic></source> <volume>8</volume> <fpage>879</fpage>&#x02013;<lpage>889</lpage>. <pub-id pub-id-type="doi">10.1038/nrmicro2465</pub-id></citation></ref>
<ref id="B40"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Darling</surname> <given-names>A. C.</given-names></name> <name><surname>Mau</surname> <given-names>B.</given-names></name> <name><surname>Blattner</surname> <given-names>F. R.</given-names></name> <name><surname>Perna</surname> <given-names>N. T.</given-names></name></person-group> (<year>2004</year>). <article-title>Mauve: multiple alignment of conserved genomic sequence with rearrangements.</article-title> <source><italic>Genome Res.</italic></source> <volume>14</volume> <fpage>1394</fpage>&#x02013;<lpage>1403</lpage>. <pub-id pub-id-type="doi">10.1101/gr.2289704</pub-id></citation></ref>
<ref id="B41"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Desai</surname> <given-names>A.</given-names></name> <name><surname>Marwah</surname> <given-names>V. S.</given-names></name> <name><surname>Yadav</surname> <given-names>A.</given-names></name> <name><surname>Jha</surname> <given-names>V.</given-names></name> <name><surname>Dhaygude</surname> <given-names>K.</given-names></name> <name><surname>Bangar</surname> <given-names>U.</given-names></name><etal/></person-group> (<year>2013</year>). <article-title>Identification of optimum sequencing depth especially for de novo genome assembly of small genomes using next generation sequencing data.</article-title> <source><italic>PLoS ONE</italic></source> <volume>8</volume>:<issue>e60204</issue>. <pub-id pub-id-type="doi">10.1371/journal.pone.0060204</pub-id></citation></ref>
<ref id="B42"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dewick</surname> <given-names>P. M.</given-names></name></person-group> (<year>2009</year>). <source><italic>Medicinal natural products: a biosynthetic approach.</italic></source> <publisher-loc>Chichester</publisher-loc>: <publisher-name>John Wiley &#x00026; Sons</publisher-name>. <pub-id pub-id-type="doi">10.1002/9780470742761</pub-id></citation></ref>
<ref id="B43"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>D&#x000ED;ez</surname> <given-names>B.</given-names></name> <name><surname>Guti&#x000E9;rrez</surname> <given-names>S.</given-names></name> <name><surname>Barredo</surname> <given-names>J. L.</given-names></name> <name><surname>van Solingen</surname> <given-names>P.</given-names></name> <name><surname>van der Voort</surname> <given-names>L. H.</given-names></name> <name><surname>Mart&#x000ED;n</surname> <given-names>J. F.</given-names></name></person-group> (<year>1990</year>). <article-title>The cluster of penicillin biosynthetic genes. Identification and characterization of the pcbAB gene encoding the alpha-aminoadipyl-cysteinyl-valine synthetase and linkage to the pcbC and penDE genes.</article-title> <source><italic>J. Biol. Chem.</italic></source> <volume>265</volume> <fpage>16358</fpage>&#x02013;<lpage>16365</lpage>.</citation></ref>
<ref id="B44"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Donadio</surname> <given-names>S.</given-names></name> <name><surname>Staver</surname> <given-names>M. J.</given-names></name> <name><surname>Mcalpine</surname> <given-names>J. B.</given-names></name> <name><surname>Swanson</surname> <given-names>S. J.</given-names></name> <name><surname>Katz</surname> <given-names>L.</given-names></name></person-group> (<year>1991</year>). <article-title>Modular organization of genes required for complex polyketide biosynthesis.</article-title> <source><italic>Science</italic></source> <volume>252</volume> <fpage>675</fpage>&#x02013;<lpage>679</lpage>. <pub-id pub-id-type="doi">10.1126/science.2024119</pub-id></citation></ref>
<ref id="B45"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Endo</surname> <given-names>A.</given-names></name></person-group> (<year>2010</year>). <article-title>A historical perspective on the discovery of statins.</article-title> <source><italic>Proc. Jpn. Acad. Ser. B Phys. Biol. Sci.</italic></source> <volume>86</volume> <fpage>484</fpage>&#x02013;<lpage>493</lpage>. <pub-id pub-id-type="doi">10.2183/pjab.86.484</pub-id></citation></ref>
<ref id="B46"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fedorova</surname> <given-names>N. D.</given-names></name> <name><surname>Moktali</surname> <given-names>V.</given-names></name> <name><surname>Medema</surname> <given-names>M. H.</given-names></name></person-group> (<year>2012</year>). <article-title>Bioinformatics approaches and software for detection of secondary metabolic gene clusters.</article-title> <source><italic>Methods Mol. Biol.</italic></source> <volume>944</volume> <fpage>23</fpage>&#x02013;<lpage>45</lpage>. <pub-id pub-id-type="doi">10.1007/978-1-62703-122-6_2</pub-id></citation></ref>
<ref id="B47"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Feng</surname> <given-names>G. H.</given-names></name> <name><surname>Leonard</surname> <given-names>T. J.</given-names></name></person-group> (<year>1995</year>). <article-title>Characterization of the polyketide synthase gene (pksL1) required for aflatoxin biosynthesis in <italic>Aspergillus parasiticus</italic>.</article-title> <source><italic>J. Bacteriol.</italic></source> <volume>177</volume> <fpage>6246</fpage>&#x02013;<lpage>6254</lpage>.</citation></ref>
<ref id="B48"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Finking</surname> <given-names>R.</given-names></name> <name><surname>Marahiel</surname> <given-names>M. A.</given-names></name></person-group> (<year>2004</year>). <article-title>Biosynthesis of nonribosomal peptides1.</article-title> <source><italic>Annu. Rev. Microbiol.</italic></source> <volume>58</volume> <fpage>453</fpage>&#x02013;<lpage>488</lpage>. <pub-id pub-id-type="doi">10.1146/annurev.micro.58.030603.123615</pub-id></citation></ref>
<ref id="B49"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Frandsen</surname> <given-names>R. J.</given-names></name> <name><surname>Nielsen</surname> <given-names>N. J.</given-names></name> <name><surname>Maolanon</surname> <given-names>N.</given-names></name> <name><surname>Sorensen</surname> <given-names>J. C.</given-names></name> <name><surname>Olsson</surname> <given-names>S.</given-names></name> <name><surname>Nielsen</surname> <given-names>J.</given-names></name><etal/></person-group> (<year>2006</year>). <article-title>The biosynthetic pathway for aurofusarin in <italic>Fusarium graminearum</italic> reveals a close link between the naphthoquinones and naphthopyrones.</article-title> <source><italic>Mol. Microbiol.</italic></source> <volume>61</volume> <fpage>1069</fpage>&#x02013;<lpage>1080</lpage>. <pub-id pub-id-type="doi">10.1111/j.1365-2958.2006.05295.x</pub-id></citation></ref>
<ref id="B50"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Frandsen</surname> <given-names>R. J.</given-names></name> <name><surname>Schutt</surname> <given-names>C.</given-names></name> <name><surname>Lund</surname> <given-names>B. W.</given-names></name> <name><surname>Staerk</surname> <given-names>D.</given-names></name> <name><surname>Nielsen</surname> <given-names>J.</given-names></name> <name><surname>Olsson</surname> <given-names>S.</given-names></name><etal/></person-group> (<year>2011</year>). <article-title>Two novel classes of enzymes are required for the biosynthesis of aurofusarin in <italic>Fusarium graminearum</italic>.</article-title> <source><italic>J. Biol. Chem.</italic></source> <volume>286</volume> <fpage>10419</fpage>&#x02013;<lpage>10428</lpage>. <pub-id pub-id-type="doi">10.1074/jbc.M110.179853</pub-id></citation></ref>
<ref id="B51"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Frisvad</surname> <given-names>J. C.</given-names></name> <name><surname>Larsen</surname> <given-names>T. O.</given-names></name> <name><surname>De Vries</surname> <given-names>R.</given-names></name> <name><surname>Meijer</surname> <given-names>M.</given-names></name> <name><surname>Houbraken</surname> <given-names>J.</given-names></name> <name><surname>Cabanes</surname> <given-names>F. J.</given-names></name><etal/></person-group> (<year>2007</year>). <article-title>Secondary metabolite profiling, growth profiles and other tools for species recognition and important <italic>Aspergillus</italic> mycotoxins.</article-title> <source><italic>Stud. Mycol.</italic></source> <volume>59</volume> <fpage>31</fpage>&#x02013;<lpage>37</lpage>. <pub-id pub-id-type="doi">10.3114/sim.2007.59.04</pub-id></citation></ref>
<ref id="B52"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Frisvad</surname> <given-names>J. C.</given-names></name> <name><surname>Samson</surname> <given-names>R. A.</given-names></name></person-group> (<year>2004</year>). <article-title>Polyphasic taxonomy of <italic>Penicillium</italic> subgenus <italic>Penicillium</italic> - A guide to identification of food and air-borne terverticillate <italic>Penicillia</italic> and their mycotoxins.</article-title> <source><italic>Stud. Mycol.</italic></source> <volume>49</volume> <fpage>1</fpage>&#x02013;<lpage>173</lpage>.</citation></ref>
<ref id="B53"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fujii</surname> <given-names>I.</given-names></name> <name><surname>Ono</surname> <given-names>Y.</given-names></name> <name><surname>Tada</surname> <given-names>H.</given-names></name> <name><surname>Gomi</surname> <given-names>K.</given-names></name> <name><surname>Ebizuka</surname> <given-names>Y.</given-names></name> <name><surname>Sankawa</surname> <given-names>U.</given-names></name></person-group> (<year>1996</year>). <article-title>Cloning of the polyketide synthase gene <italic>atX</italic> from <italic>Aspergillus terreus</italic> and its identification as the 6-methylsalicylic acid synthase gene by heterologous expression.</article-title> <source><italic>Mol. Gen. Genet.</italic></source> <volume>253</volume> <fpage>1</fpage>&#x02013;<lpage>10</lpage>. <pub-id pub-id-type="doi">10.1007/s004380050289</pub-id></citation></ref>
<ref id="B54"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fujii</surname> <given-names>I.</given-names></name> <name><surname>Watanabe</surname> <given-names>A.</given-names></name> <name><surname>Sankawa</surname> <given-names>U.</given-names></name> <name><surname>Ebizuka</surname> <given-names>Y.</given-names></name></person-group> (<year>2001</year>). <article-title>Identification of Claisen cyclase domain in fungal polyketide synthase WA, a naphthopyrone synthase of <italic>Aspergillus nidulans</italic>.</article-title> <source><italic>Chem. Biol.</italic></source> <volume>8</volume> <fpage>189</fpage>&#x02013;<lpage>197</lpage>. <pub-id pub-id-type="doi">10.1016/S1074-5521(00)90068-1</pub-id></citation></ref>
<ref id="B55"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gaffoor</surname> <given-names>I.</given-names></name> <name><surname>Brown</surname> <given-names>D. W.</given-names></name> <name><surname>Plattner</surname> <given-names>R.</given-names></name> <name><surname>Proctor</surname> <given-names>R. H.</given-names></name> <name><surname>Qi</surname> <given-names>W.</given-names></name> <name><surname>Trail</surname> <given-names>F.</given-names></name></person-group> (<year>2005</year>). <article-title>Functional analysis of the polyketide synthase genes in the filamentous fungus <italic>Gibberella zeae</italic> (anamorph <italic>Fusarium graminearum</italic>).</article-title> <source><italic>Eukaryot. Cell</italic></source> <volume>4</volume> <fpage>1926</fpage>&#x02013;<lpage>1933</lpage>. <pub-id pub-id-type="doi">10.1128/EC.4.11.1926-1933.2005</pub-id></citation></ref>
<ref id="B56"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gao</surname> <given-names>X.</given-names></name> <name><surname>Chooi</surname> <given-names>Y. H.</given-names></name> <name><surname>Ames</surname> <given-names>B. D.</given-names></name> <name><surname>Wang</surname> <given-names>P.</given-names></name> <name><surname>Walsh</surname> <given-names>C. T.</given-names></name> <name><surname>Tang</surname> <given-names>Y.</given-names></name></person-group> (<year>2011</year>). <article-title>Fungal indole alkaloid biosynthesis: genetic and biochemical investigation of the tryptoquialanine pathway in <italic>Penicillium aethiopicum</italic>.</article-title> <source><italic>J. Am. Chem. Soc.</italic></source> <volume>133</volume> <fpage>2729</fpage>&#x02013;<lpage>2741</lpage>. <pub-id pub-id-type="doi">10.1021/ja1101085</pub-id></citation></ref>
<ref id="B57"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gao</surname> <given-names>X.</given-names></name> <name><surname>Haynes</surname> <given-names>S. W.</given-names></name> <name><surname>Ames</surname> <given-names>B. D.</given-names></name> <name><surname>Wang</surname> <given-names>P.</given-names></name> <name><surname>Vien</surname> <given-names>L.</given-names></name> <name><surname>Walsh</surname> <given-names>C. T.</given-names></name><etal/></person-group> (<year>2012</year>). <article-title>Cyclization of fungal nonribosomal peptides catalyzed by a terminal condensation-like domain.</article-title> <source><italic>Nat. Chem. Biol.</italic></source> <volume>8</volume> <fpage>823</fpage>&#x02013;<lpage>830</lpage>. <pub-id pub-id-type="doi">10.1038/nchembio.1047</pub-id></citation></ref>
<ref id="B58"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gao</surname> <given-names>X.</given-names></name> <name><surname>Xie</surname> <given-names>X.</given-names></name> <name><surname>Pashkov</surname> <given-names>I.</given-names></name> <name><surname>Sawaya</surname> <given-names>M. R.</given-names></name> <name><surname>Laidman</surname> <given-names>J.</given-names></name> <name><surname>Zhang</surname> <given-names>W.</given-names></name><etal/></person-group> (<year>2009</year>). <article-title>Directed evolution and structural characterization of a simvastatin synthase.</article-title> <source><italic>Chem. Biol.</italic></source> <volume>16</volume> <fpage>1064</fpage>&#x02013;<lpage>1074</lpage>. <pub-id pub-id-type="doi">10.1016/j.chembiol.2009.09.017</pub-id></citation></ref>
<ref id="B59"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Glinski</surname> <given-names>M.</given-names></name> <name><surname>Urbanke</surname> <given-names>C.</given-names></name> <name><surname>Hornbogen</surname> <given-names>T.</given-names></name> <name><surname>Zocher</surname> <given-names>R.</given-names></name></person-group> (<year>2002</year>). <article-title>Enniatin synthetase is a monomer with extended structure: evidence for an intramolecular reaction mechanism.</article-title> <source><italic>Arch. Microbiol.</italic></source> <volume>178</volume> <fpage>267</fpage>&#x02013;<lpage>273</lpage>. <pub-id pub-id-type="doi">10.1007/s00203-002-0451-1</pub-id></citation></ref>
<ref id="B60"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hashimoto</surname> <given-names>M.</given-names></name> <name><surname>Nonaka</surname> <given-names>T.</given-names></name> <name><surname>Fujii</surname> <given-names>I.</given-names></name></person-group> (<year>2014</year>). <article-title>Fungal type III polyketide synthases.</article-title> <source><italic>Nat. Prod. Rep.</italic></source> <volume>31</volume> <fpage>1306</fpage>&#x02013;<lpage>1317</lpage>. <pub-id pub-id-type="doi">10.1039/c4np00096j</pub-id></citation></ref>
<ref id="B61"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Haynes</surname> <given-names>S. W.</given-names></name> <name><surname>Gao</surname> <given-names>X.</given-names></name> <name><surname>Tang</surname> <given-names>Y.</given-names></name> <name><surname>Walsh</surname> <given-names>C. T.</given-names></name></person-group> (<year>2012</year>). <article-title>Assembly of asperlicin peptidyl alkaloids from anthranilate and tryptophan: a two-enzyme pathway generates heptacyclic scaffold complexity in asperlicin E.</article-title> <source><italic>J. Am. Chem. Soc.</italic></source> <volume>134</volume> <fpage>17444</fpage>&#x02013;<lpage>17447</lpage>. <pub-id pub-id-type="doi">10.1021/ja308371z</pub-id></citation></ref>
<ref id="B62"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Haynes</surname> <given-names>S. W.</given-names></name> <name><surname>Gao</surname> <given-names>X.</given-names></name> <name><surname>Tang</surname> <given-names>Y.</given-names></name> <name><surname>Walsh</surname> <given-names>C. T.</given-names></name></person-group> (<year>2013</year>). <article-title>Complexity generation in fungal peptidyl alkaloid biosynthesis: a two-enzyme pathway to the hexacyclic MDR export pump inhibitor ardeemin.</article-title> <source><italic>ACS Chem. Biol.</italic></source> <volume>8</volume> <fpage>741</fpage>&#x02013;<lpage>748</lpage>. <pub-id pub-id-type="doi">10.1021/cb3006787</pub-id></citation></ref>
<ref id="B63"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hendrickson</surname> <given-names>L.</given-names></name> <name><surname>Davis</surname> <given-names>C. R.</given-names></name> <name><surname>Roach</surname> <given-names>C.</given-names></name> <name><surname>Nguyen</surname> <given-names>D. K.</given-names></name> <name><surname>Aldrich</surname> <given-names>T.</given-names></name> <name><surname>Mcada</surname> <given-names>P. C.</given-names></name><etal/></person-group> (<year>1999</year>). <article-title>Lovastatin biosynthesis in <italic>Aspergillus terreus</italic>: characterization of blocked mutants, enzyme activities and a multifunctional polyketide synthase gene.</article-title> <source><italic>Chem. Biol.</italic></source> <volume>6</volume> <fpage>429</fpage>&#x02013;<lpage>439</lpage>. <pub-id pub-id-type="doi">10.1016/S1074-5521(99)80061-1</pub-id></citation></ref>
<ref id="B64"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hohn</surname> <given-names>T. M.</given-names></name> <name><surname>Mccormick</surname> <given-names>S. P.</given-names></name> <name><surname>Desjardins</surname> <given-names>A. E.</given-names></name></person-group> (<year>1993</year>). <article-title>Evidence for a gene cluster involving trichothecene-pathway biosynthetic genes in <italic>Fusarium sporotrichioides</italic>.</article-title> <source><italic>Curr. Genet.</italic></source> <volume>24</volume> <fpage>291</fpage>&#x02013;<lpage>295</lpage>. <pub-id pub-id-type="doi">10.1007/BF00336778</pub-id></citation></ref>
<ref id="B65"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ishiuchi</surname> <given-names>K.</given-names></name> <name><surname>Nakazawa</surname> <given-names>T.</given-names></name> <name><surname>Yagishita</surname> <given-names>F.</given-names></name> <name><surname>Mino</surname> <given-names>T.</given-names></name> <name><surname>Noguchi</surname> <given-names>H.</given-names></name> <name><surname>Hotta</surname> <given-names>K.</given-names></name><etal/></person-group> (<year>2013</year>). <article-title>Combinatorial generation of complexity by redox enzymes in the chaetoglobosin A biosynthesis.</article-title> <source><italic>J. Am. Chem. Soc.</italic></source> <volume>135</volume> <fpage>7371</fpage>&#x02013;<lpage>7377</lpage>. <pub-id pub-id-type="doi">10.1021/ja402828w</pub-id></citation></ref>
<ref id="B66"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jiang</surname> <given-names>W.</given-names></name> <name><surname>Cacho</surname> <given-names>R. A.</given-names></name> <name><surname>Chiou</surname> <given-names>G.</given-names></name> <name><surname>Garg</surname> <given-names>N. K.</given-names></name> <name><surname>Tang</surname> <given-names>Y.</given-names></name> <name><surname>Walsh</surname> <given-names>C. T.</given-names></name></person-group> (<year>2013</year>). <article-title>EcdGHK are three tailoring iron oxygenases for amino acid building blocks of the echinocandin scaffold.</article-title> <source><italic>J. Am. Chem. Soc.</italic></source> <volume>135</volume> <fpage>4457</fpage>&#x02013;<lpage>4466</lpage>. <pub-id pub-id-type="doi">10.1021/ja312572v</pub-id></citation></ref>
<ref id="B67"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Johnson</surname> <given-names>L.</given-names></name></person-group> (<year>2008</year>). <article-title>Iron and siderophores in fungal-host interactions.</article-title> <source><italic>Mycol. Res.</italic></source> <volume>112</volume> <fpage>170</fpage>&#x02013;<lpage>183</lpage>. <pub-id pub-id-type="doi">10.1016/j.mycres.2007.11.012</pub-id></citation></ref>
<ref id="B68"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kakule</surname> <given-names>T. B.</given-names></name> <name><surname>Sardar</surname> <given-names>D.</given-names></name> <name><surname>Lin</surname> <given-names>Z.</given-names></name> <name><surname>Schmidt</surname> <given-names>E. W.</given-names></name></person-group> (<year>2013</year>). <article-title>Two related pyrrolidinedione synthetase loci in <italic>Fusarium</italic> heterosporum ATCC 74349 produce divergent metabolites.</article-title> <source><italic>ACS Chem. Biol.</italic></source> <volume>8</volume> <fpage>1549</fpage>&#x02013;<lpage>1557</lpage>. <pub-id pub-id-type="doi">10.1021/cb400159f</pub-id></citation></ref>
<ref id="B69"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kardos</surname> <given-names>N.</given-names></name> <name><surname>Demain</surname> <given-names>A. L.</given-names></name></person-group> (<year>2011</year>). <article-title>Penicillin: the medicine with the greatest impact on therapeutic outcomes.</article-title> <source><italic>Appl. Microbiol. Biotechnol.</italic></source> <volume>92</volume> <fpage>677</fpage>&#x02013;<lpage>687</lpage>. <pub-id pub-id-type="doi">10.1007/s00253-011-3587-6</pub-id></citation></ref>
<ref id="B70"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kato</surname> <given-names>N.</given-names></name> <name><surname>Suzuki</surname> <given-names>H.</given-names></name> <name><surname>Okumura</surname> <given-names>H.</given-names></name> <name><surname>Takahashi</surname> <given-names>S.</given-names></name> <name><surname>Osada</surname> <given-names>H.</given-names></name></person-group> (<year>2013</year>). <article-title>A point mutation in ftmD blocks the fumitremorgin biosynthetic pathway in <italic>Aspergillus fumigatus strain</italic> Af293.</article-title> <source><italic>Biosci. Biotechnol. Biochem.</italic></source> <volume>77</volume> <fpage>1061</fpage>&#x02013;<lpage>1067</lpage>. <pub-id pub-id-type="doi">10.1271/bbb.130026</pub-id></citation></ref>
<ref id="B71"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Keller</surname> <given-names>N. P.</given-names></name> <name><surname>Turner</surname> <given-names>G.</given-names></name> <name><surname>Bennett</surname> <given-names>J. W.</given-names></name></person-group> (<year>2005</year>). <article-title>Fungal secondary metabolism - from biochemistry to genomics.</article-title> <source><italic>Nat. Rev. Microbiol.</italic></source> <volume>3</volume> <fpage>937</fpage>&#x02013;<lpage>947</lpage>. <pub-id pub-id-type="doi">10.1038/nrmicro1286</pub-id></citation></ref>
<ref id="B72"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kennedy</surname> <given-names>J.</given-names></name> <name><surname>Auclair</surname> <given-names>K.</given-names></name> <name><surname>Kendrew</surname> <given-names>S. G.</given-names></name> <name><surname>Park</surname> <given-names>C.</given-names></name> <name><surname>Vederas</surname> <given-names>J. C.</given-names></name> <name><surname>Hutchinson</surname> <given-names>C. R.</given-names></name></person-group> (<year>1999</year>). <article-title>Modulation of polyketide synthase activity by accessory proteins during lovastatin biosynthesis.</article-title> <source><italic>Science</italic></source> <volume>284</volume> <fpage>1368</fpage>&#x02013;<lpage>1372</lpage>. <pub-id pub-id-type="doi">10.1126/science.284.5418.1368</pub-id></citation></ref>
<ref id="B73"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kett</surname> <given-names>D. H.</given-names></name> <name><surname>Shorr</surname> <given-names>A. F.</given-names></name> <name><surname>Reboli</surname> <given-names>A. C.</given-names></name> <name><surname>Reisman</surname> <given-names>A. L.</given-names></name> <name><surname>Biswas</surname> <given-names>P.</given-names></name> <name><surname>Schlamm</surname> <given-names>H. T.</given-names></name></person-group> (<year>2011</year>). <article-title>Anidulafungin compared with fluconazole in severely ill patients with candidemia and other forms of invasive candidiasis: support for the 2009 IDSA treatment guidelines for candidiasis.</article-title> <source><italic>Crit. Care</italic></source> <volume>15</volume> R253. <pub-id pub-id-type="doi">10.1186/cc10514</pub-id></citation></ref>
<ref id="B74"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Khaldi</surname> <given-names>N.</given-names></name> <name><surname>Seifuddin</surname> <given-names>F. T.</given-names></name> <name><surname>Turner</surname> <given-names>G.</given-names></name> <name><surname>Haft</surname> <given-names>D.</given-names></name> <name><surname>Nierman</surname> <given-names>W. C.</given-names></name> <name><surname>Wolfe</surname> <given-names>K. H.</given-names></name><etal/></person-group> (<year>2010</year>). <article-title>SMURF: Genomic mapping of fungal secondary metabolite clusters.</article-title> <source><italic>Fungal Genet. Biol.</italic></source> <volume>47</volume> <fpage>736</fpage>&#x02013;<lpage>741</lpage>. <pub-id pub-id-type="doi">10.1016/j.fgb.2010.06.003</pub-id></citation></ref>
<ref id="B75"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kim</surname> <given-names>Y. T.</given-names></name> <name><surname>Lee</surname> <given-names>Y. R.</given-names></name> <name><surname>Jin</surname> <given-names>J.</given-names></name> <name><surname>Han</surname> <given-names>K. H.</given-names></name> <name><surname>Kim</surname> <given-names>H.</given-names></name> <name><surname>Kim</surname> <given-names>J. C.</given-names></name><etal/></person-group> (<year>2005</year>). <article-title>Two different polyketide synthase genes are required for synthesis of zearalenone in <italic>Gibberella zeae</italic>.</article-title> <source><italic>Mol. Microbiol.</italic></source> <volume>58</volume> <fpage>1102</fpage>&#x02013;<lpage>1113</lpage>. <pub-id pub-id-type="doi">10.1111/j.1365-2958.2005.04884.x</pub-id></citation></ref>
<ref id="B76"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kleftogiannis</surname> <given-names>D.</given-names></name> <name><surname>Kalnis</surname> <given-names>P.</given-names></name> <name><surname>Bajic</surname> <given-names>V. B.</given-names></name></person-group> (<year>2013</year>). <article-title>Comparing memory-efficient genome assemblers on stand-alone and cloud infrastructures.</article-title> <source><italic>PLoS ONE</italic></source> <volume>8</volume>:<issue>e75505</issue>. <pub-id pub-id-type="doi">10.1371/journal.pone.0075505</pub-id></citation></ref>
<ref id="B77"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kodukula</surname> <given-names>K.</given-names></name> <name><surname>Arcuri</surname> <given-names>M.</given-names></name> <name><surname>Cutrone</surname> <given-names>J. Q.</given-names></name> <name><surname>Hugill</surname> <given-names>R. M.</given-names></name> <name><surname>Lowe</surname> <given-names>S. E.</given-names></name> <name><surname>Pirnik</surname> <given-names>D. M.</given-names></name><etal/></person-group> (<year>1995</year>). <article-title>BMS-192548 a tetracyclic binding inhibitor of neuropeptide Y receptors, from <italic>Aspergillus niger</italic> WB2346. I. Taxonomy, fermentation, isolation and biological activity.</article-title> <source><italic>J. Antibiot. (Tokyo)</italic></source> <volume>48</volume> <fpage>1055</fpage>&#x02013;<lpage>1059</lpage>. <pub-id pub-id-type="doi">10.7164/antibiotics.48.1055</pub-id></citation></ref>
<ref id="B78"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Korman</surname> <given-names>T. P.</given-names></name> <name><surname>Crawford</surname> <given-names>J. M.</given-names></name> <name><surname>Labonte</surname> <given-names>J. W.</given-names></name> <name><surname>Newman</surname> <given-names>A. G.</given-names></name> <name><surname>Wong</surname> <given-names>J.</given-names></name> <name><surname>Townsend</surname> <given-names>C. A.</given-names></name><etal/></person-group> (<year>2010</year>). <article-title>Structure and function of an iterative polyketide synthase thioesterase domain catalyzing Claisen cyclization in aflatoxin biosynthesis.</article-title> <source><italic>Proc. Natl. Acad. Sci. U.S.A.</italic></source> <volume>107</volume> <fpage>6246</fpage>&#x02013;<lpage>6251</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.0913531107</pub-id></citation></ref>
<ref id="B79"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kroken</surname> <given-names>S.</given-names></name> <name><surname>Glass</surname> <given-names>N. L.</given-names></name> <name><surname>Taylor</surname> <given-names>J. W.</given-names></name> <name><surname>Yoder</surname> <given-names>O. C.</given-names></name> <name><surname>Turgeon</surname> <given-names>B. G.</given-names></name></person-group> (<year>2003</year>). <article-title>Phylogenomic analysis of type I polyketide synthase genes in pathogenic and saprobic ascomycetes.</article-title> <source><italic>Proc. Natl. Acad. Sci. U.S.A.</italic></source> <volume>100</volume> <fpage>15670</fpage>&#x02013;<lpage>15675</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.2532165100</pub-id></citation></ref>
<ref id="B80"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lee</surname> <given-names>S.</given-names></name> <name><surname>Son</surname> <given-names>H.</given-names></name> <name><surname>Lee</surname> <given-names>J.</given-names></name> <name><surname>Lee</surname> <given-names>Y. R.</given-names></name> <name><surname>Lee</surname> <given-names>Y. W.</given-names></name></person-group> (<year>2011</year>). <article-title>A putative ABC transporter gene, ZRA1, is required for zearalenone production in <italic>Gibberella zeae</italic>.</article-title> <source><italic>Curr. Genet.</italic></source> <volume>57</volume> <fpage>343</fpage>&#x02013;<lpage>351</lpage>. <pub-id pub-id-type="doi">10.1007/s00294-011-0352-4</pub-id></citation></ref>
<ref id="B81"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Li</surname> <given-names>R.</given-names></name> <name><surname>Zhu</surname> <given-names>H.</given-names></name> <name><surname>Ruan</surname> <given-names>J.</given-names></name> <name><surname>Qian</surname> <given-names>W.</given-names></name> <name><surname>Fang</surname> <given-names>X.</given-names></name> <name><surname>Shi</surname> <given-names>Z.</given-names></name><etal/></person-group> (<year>2010a</year>). <article-title>De novo assembly of human genomes with massively parallel short read sequencing.</article-title> <source><italic>Genome Res.</italic></source> <volume>20</volume> <fpage>265</fpage>&#x02013;<lpage>272</lpage>. <pub-id pub-id-type="doi">10.1101/gr.097261.109</pub-id></citation></ref>
<ref id="B82"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Li</surname> <given-names>Y.</given-names></name> <name><surname>Xu</surname> <given-names>W.</given-names></name> <name><surname>Tang</surname> <given-names>Y.</given-names></name></person-group> (<year>2010b</year>). <article-title>Classification, prediction, and verification of the regioselectivity of fungal polyketide synthase product template domains.</article-title> <source><italic>J. Biol. Chem.</italic></source> <volume>285</volume> <fpage>22764</fpage>&#x02013;<lpage>22773</lpage>. <pub-id pub-id-type="doi">10.1074/jbc.M110.128504</pub-id></citation></ref>
<ref id="B83"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Li</surname> <given-names>S.</given-names></name> <name><surname>Anand</surname> <given-names>K.</given-names></name> <name><surname>Tran</surname> <given-names>H.</given-names></name> <name><surname>Yu</surname> <given-names>F.</given-names></name> <name><surname>Finefield</surname> <given-names>J. M.</given-names></name> <name><surname>Sunderhaus</surname> <given-names>J. D.</given-names></name><etal/></person-group> (<year>2012</year>). <article-title>Comparative analysis of the biosynthetic systems for fungal bicyclo[2.2.2]diazaoctane indole alkaloids: the (+)/(-)-notoamide, paraherquamide and malbrancheamide pathways.</article-title> <source><italic>Medchemcomm.</italic></source> <volume>3</volume> <fpage>987</fpage>&#x02013;<lpage>996</lpage>. <pub-id pub-id-type="doi">10.1039/C2MD20029E</pub-id></citation></ref>
<ref id="B84"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Li</surname> <given-names>Y.</given-names></name> <name><surname>Chooi</surname> <given-names>Y. H.</given-names></name> <name><surname>Sheng</surname> <given-names>Y.</given-names></name> <name><surname>Valentine</surname> <given-names>J. S.</given-names></name> <name><surname>Tang</surname> <given-names>Y.</given-names></name></person-group> (<year>2011</year>). <article-title>Comparative characterization of fungal anthracenone and naphthacenedione biosynthetic pathways reveals an alpha-hydroxylation-dependent Claisen-like cyclization catalyzed by a dimanganese thioesterase.</article-title> <source><italic>J. Am. Chem. Soc.</italic></source> <volume>133</volume> <fpage>15773</fpage>&#x02013;<lpage>15785</lpage>. <pub-id pub-id-type="doi">10.1021/ja206906d</pub-id></citation></ref>
<ref id="B85"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lin</surname> <given-names>Y.</given-names></name> <name><surname>Li</surname> <given-names>J.</given-names></name> <name><surname>Shen</surname> <given-names>H.</given-names></name> <name><surname>Zhang</surname> <given-names>L.</given-names></name> <name><surname>Papasian</surname> <given-names>C. J.</given-names></name> <name><surname>Deng</surname> <given-names>H. W.</given-names></name></person-group> (<year>2011</year>). <article-title>Comparative studies of de novo assembly tools for next-generation sequencing technologies.</article-title> <source><italic>Bioinformatics</italic></source> <volume>27</volume> <fpage>2031</fpage>&#x02013;<lpage>2037</lpage>. <pub-id pub-id-type="doi">10.1093/bioinformatics/btr319</pub-id></citation></ref>
<ref id="B86"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Linnemannstons</surname> <given-names>P.</given-names></name> <name><surname>Schulte</surname> <given-names>J.</given-names></name> <name><surname>Del Mar Prado</surname> <given-names>M.</given-names></name> <name><surname>Proctor</surname> <given-names>R. H.</given-names></name> <name><surname>Avalos</surname> <given-names>J.</given-names></name> <name><surname>Tudzynski</surname> <given-names>B.</given-names></name></person-group> (<year>2002</year>). <article-title>The polyketide synthase gene pks4 from <italic>Gibberella fujikuroi</italic> encodes a key enzyme in the biosynthesis of the red pigment bikaverin.</article-title> <source><italic>Fungal Genet. Biol.</italic></source> <volume>37</volume> <fpage>134</fpage>&#x02013;<lpage>148</lpage>. <pub-id pub-id-type="doi">10.1016/S1087-1845(02)00501-7</pub-id></citation></ref>
<ref id="B87"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Luo</surname> <given-names>R.</given-names></name> <name><surname>Liu</surname> <given-names>B.</given-names></name> <name><surname>Xie</surname> <given-names>Y.</given-names></name> <name><surname>Li</surname> <given-names>Z.</given-names></name> <name><surname>Huang</surname> <given-names>W.</given-names></name> <name><surname>Yuan</surname> <given-names>J.</given-names></name><etal/></person-group> (<year>2012</year>). <article-title>SOAPdenovo2: an empirically improved memory-efficient short-read de novo assembler.</article-title> <source><italic>Gigascience</italic></source> <volume>1</volume> <issue>18</issue>. <pub-id pub-id-type="doi">10.1186/2047-217X-1-18</pub-id></citation></ref>
<ref id="B88"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Machida</surname> <given-names>M.</given-names></name> <name><surname>Asai</surname> <given-names>K.</given-names></name> <name><surname>Sano</surname> <given-names>M.</given-names></name> <name><surname>Tanaka</surname> <given-names>T.</given-names></name> <name><surname>Kumagai</surname> <given-names>T.</given-names></name> <name><surname>Terai</surname> <given-names>G.</given-names></name><etal/></person-group> (<year>2005</year>). <article-title>Genome sequencing and analysis of <italic>Aspergillus oryzae</italic>.</article-title> <source><italic>Nature</italic></source> <volume>438</volume> <fpage>1157</fpage>&#x02013;<lpage>1161</lpage>. <pub-id pub-id-type="doi">10.1038/nature04300</pub-id></citation></ref>
<ref id="B89"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Maiya</surname> <given-names>S.</given-names></name> <name><surname>Grundmann</surname> <given-names>A.</given-names></name> <name><surname>Li</surname> <given-names>X.</given-names></name> <name><surname>Li</surname> <given-names>S. M.</given-names></name> <name><surname>Turner</surname> <given-names>G.</given-names></name></person-group> (<year>2007</year>). <article-title>Identification of a hybrid PKS/NRPS required for pseurotin A biosynthesis in the human pathogen <italic>Aspergillus fumigatus</italic>.</article-title> <source><italic>Chembiochem</italic></source> <volume>8</volume> <fpage>1736</fpage>&#x02013;<lpage>1743</lpage>. <pub-id pub-id-type="doi">10.1002/cbic.200700202</pub-id></citation></ref>
<ref id="B90"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Malpartida</surname> <given-names>F.</given-names></name> <name><surname>Hopwood</surname> <given-names>D. A.</given-names></name></person-group> (<year>1984</year>). <article-title>Molecular cloning of the whole biosynthetic pathway of a <italic>Streptomyces</italic> antibiotic and its expression in a heterologous host.</article-title> <source><italic>Nature</italic></source> <volume>309</volume> <fpage>462</fpage>&#x02013;<lpage>464</lpage>. <pub-id pub-id-type="doi">10.1038/309462a0</pub-id></citation></ref>
<ref id="B91"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mayorga</surname> <given-names>M. E.</given-names></name> <name><surname>Timberlake</surname> <given-names>W. E.</given-names></name></person-group> (<year>1990</year>). <article-title>Isolation and molecular characterization of the <italic>Aspergillus nidulans</italic> wA gene.</article-title> <source><italic>Genetics</italic></source> <volume>126</volume> <fpage>73</fpage>&#x02013;<lpage>79</lpage>.</citation></ref>
<ref id="B92"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Medema</surname> <given-names>M. H.</given-names></name> <name><surname>Blin</surname> <given-names>K.</given-names></name> <name><surname>Cimermancic</surname> <given-names>P.</given-names></name> <name><surname>De Jager</surname> <given-names>V.</given-names></name> <name><surname>Zakrzewski</surname> <given-names>P.</given-names></name> <name><surname>Fischbach</surname> <given-names>M. A.</given-names></name><etal/></person-group> (<year>2011</year>). <article-title>AntiSMASH: rapid identification, annotation and analysis of secondary metabolite biosynthesis gene clusters in bacterial and fungal genome sequences.</article-title> <source><italic>Nucleic Acids Res.</italic></source> <volume>39</volume> <fpage>W339</fpage>&#x02013;<lpage>W346</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gkr466</pub-id></citation></ref>
<ref id="B93"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Nicholson</surname> <given-names>T. P.</given-names></name> <name><surname>Rudd</surname> <given-names>B. A.</given-names></name> <name><surname>Dawson</surname> <given-names>M.</given-names></name> <name><surname>Lazarus</surname> <given-names>C. M.</given-names></name> <name><surname>Simpson</surname> <given-names>T. J.</given-names></name> <name><surname>Cox</surname> <given-names>R. J.</given-names></name></person-group> (<year>2001</year>). <article-title>Design and utility of oligonucleotide gene probes for fungal polyketide synthases.</article-title> <source><italic>Chem. Biol.</italic></source> <volume>8</volume> <fpage>157</fpage>&#x02013;<lpage>178</lpage>. <pub-id pub-id-type="doi">10.1016/S1074-5521(00)90064-4</pub-id></citation></ref>
<ref id="B94"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Niehaus</surname> <given-names>E. M.</given-names></name> <name><surname>Janevska</surname> <given-names>S.</given-names></name> <name><surname>Von Bargen</surname> <given-names>K. W.</given-names></name> <name><surname>Sieber</surname> <given-names>C. M.</given-names></name> <name><surname>Harrer</surname> <given-names>H.</given-names></name> <name><surname>Humpf</surname> <given-names>H. U.</given-names></name><etal/></person-group> (<year>2014</year>). <article-title>Apicidin F: characterization and genetic manipulation of a new secondary metabolite gene cluster in the rice pathogen Fusarium fujikuroi.</article-title> <source><italic>PLoS ONE</italic></source> <volume>9</volume>:<issue>e103336</issue>. <pub-id pub-id-type="doi">10.1371/journal.pone.0103336</pub-id></citation></ref>
<ref id="B95"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Niehaus</surname> <given-names>E. M.</given-names></name> <name><surname>Kleigrewe</surname> <given-names>K.</given-names></name> <name><surname>Wiemann</surname> <given-names>P.</given-names></name> <name><surname>Studt</surname> <given-names>L.</given-names></name> <name><surname>Sieber</surname> <given-names>C. M.</given-names></name> <name><surname>Connolly</surname> <given-names>L. R.</given-names></name><etal/></person-group> (<year>2013</year>). <article-title>Genetic manipulation of the <italic>Fusarium fujikuroi</italic> fusarin gene cluster yields insight into the complex regulation and fusarin biosynthetic pathway.</article-title> <source><italic>Chem. Biol.</italic></source> <volume>20</volume> <fpage>1055</fpage>&#x02013;<lpage>1066</lpage>. <pub-id pub-id-type="doi">10.1016/j.chembiol.2013.07.004</pub-id></citation></ref>
<ref id="B96"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Nowrousian</surname> <given-names>M.</given-names></name></person-group> (<year>2010</year>). <article-title>Next-generation sequencing techniques for eukaryotic microorganisms: sequencing-based solutions to biological problems.</article-title> <source><italic>Eukaryot. Cell</italic></source> <volume>9</volume> <fpage>1300</fpage>&#x02013;<lpage>1310</lpage>. <pub-id pub-id-type="doi">10.1128/EC.00123-10</pub-id></citation></ref>
<ref id="B97"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>O&#x02019;Callaghan</surname> <given-names>J.</given-names></name> <name><surname>Caddick</surname> <given-names>M. X.</given-names></name> <name><surname>Dobson</surname> <given-names>A. D.</given-names></name></person-group> (<year>2003</year>). <article-title>A polyketide synthase gene required for ochratoxin A biosynthesis in <italic>Aspergillus ochraceus</italic>.</article-title> <source><italic>Microbiology</italic></source> <volume>149</volume> <fpage>3485</fpage>&#x02013;<lpage>3491</lpage>. <pub-id pub-id-type="doi">10.1099/mic.0.26619-0</pub-id></citation></ref>
<ref id="B98"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Pickens</surname> <given-names>L. B.</given-names></name> <name><surname>Tang</surname> <given-names>Y.</given-names></name> <name><surname>Chooi</surname> <given-names>Y. H.</given-names></name></person-group> (<year>2011</year>). <article-title>Metabolic engineering for the production of natural products.</article-title> <source><italic>Annu. Rev. Chem. Biomol. Eng.</italic></source> <volume>2</volume> <fpage>211</fpage>&#x02013;<lpage>236</lpage>. <pub-id pub-id-type="doi">10.1146/annurev-chembioeng-061010-114209</pub-id></citation></ref>
<ref id="B99"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Proctor</surname> <given-names>R. H.</given-names></name> <name><surname>Desjardins</surname> <given-names>A. E.</given-names></name> <name><surname>Plattner</surname> <given-names>R. D.</given-names></name> <name><surname>Hohn</surname> <given-names>T. M.</given-names></name></person-group> (<year>1999</year>). <article-title>A polyketide synthase gene required for biosynthesis of fumonisin mycotoxins in <italic>Gibberella fujikuroi</italic> mating population A.</article-title> <source><italic>Fungal Genet. Biol.</italic></source> <volume>27</volume> <fpage>100</fpage>&#x02013;<lpage>112</lpage>. <pub-id pub-id-type="doi">10.1006/fgbi.1999.1141</pub-id></citation></ref>
<ref id="B100"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Qiao</surname> <given-names>K.</given-names></name> <name><surname>Chooi</surname> <given-names>Y. H.</given-names></name> <name><surname>Tang</surname> <given-names>Y.</given-names></name></person-group> (<year>2011</year>). <article-title>Identification and engineering of the cytochalasin gene cluster from <italic>Aspergillus clavatus</italic> NRRL 1.</article-title> <source><italic>Metab. Eng.</italic></source> <volume>13</volume> <fpage>723</fpage>&#x02013;<lpage>732</lpage>. <pub-id pub-id-type="doi">10.1016/j.ymben.2011.09.008</pub-id></citation></ref>
<ref id="B101"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Reeves</surname> <given-names>C. D.</given-names></name> <name><surname>Hu</surname> <given-names>Z.</given-names></name> <name><surname>Reid</surname> <given-names>R.</given-names></name> <name><surname>Kealey</surname> <given-names>J. T.</given-names></name></person-group> (<year>2008</year>). <article-title>Genes for the biosynthesis of the fungal polyketides hypothemycin from <italic>Hypomyces subiculosus</italic> and radicicol from <italic>Pochonia chlamydosporia</italic>.</article-title> <source><italic>Appl. Environ. Microbiol.</italic></source> <volume>74</volume> <fpage>5121</fpage>&#x02013;<lpage>5129</lpage>. <pub-id pub-id-type="doi">10.1128/AEM.00478-08</pub-id></citation></ref>
<ref id="B102"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>R&#x000F6;ttig</surname> <given-names>M.</given-names></name> <name><surname>Medema</surname> <given-names>M. H.</given-names></name> <name><surname>Blin</surname> <given-names>K.</given-names></name> <name><surname>Weber</surname> <given-names>T.</given-names></name> <name><surname>Rausch</surname> <given-names>C.</given-names></name> <name><surname>Kohlbacher</surname> <given-names>O.</given-names></name></person-group> (<year>2011</year>). <article-title>NRPSpredictor2&#x02014;a web server for predicting NRPS adenylation domain specificity.</article-title> <source><italic>Nucleic Acids Res.</italic></source> <volume>39</volume> <fpage>W362</fpage>&#x02013;<lpage>W367</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gkr323</pub-id></citation></ref>
<ref id="B103"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sanchez</surname> <given-names>J. F.</given-names></name> <name><surname>Entwistle</surname> <given-names>R.</given-names></name> <name><surname>Hung</surname> <given-names>J. H.</given-names></name> <name><surname>Yaegashi</surname> <given-names>J.</given-names></name> <name><surname>Jain</surname> <given-names>S.</given-names></name> <name><surname>Chiang</surname> <given-names>Y. M.</given-names></name><etal/></person-group> (<year>2011</year>). <article-title>Genome-based deletion analysis reveals the prenyl xanthone biosynthesis pathway in <italic>Aspergillus nidulans</italic>.</article-title> <source><italic>J. Am. Chem. Soc.</italic></source> <volume>133</volume> <fpage>4010</fpage>&#x02013;<lpage>4017</lpage>. <pub-id pub-id-type="doi">10.1021/ja1096682</pub-id></citation></ref>
<ref id="B104"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Schwecke</surname> <given-names>T.</given-names></name> <name><surname>Gottling</surname> <given-names>K.</given-names></name> <name><surname>Durek</surname> <given-names>P.</given-names></name> <name><surname>Duenas</surname> <given-names>I.</given-names></name> <name><surname>Kaufer</surname> <given-names>N. F.</given-names></name> <name><surname>Zock-Emmenthal</surname> <given-names>S.</given-names></name><etal/></person-group> (<year>2006</year>). <article-title>Nonribosomal peptide synthesis in <italic>Schizosaccharomyces pombe</italic> and the architectures of ferrichrome-type siderophore synthetases in fungi.</article-title> <source><italic>Chembiochem</italic></source> <volume>7</volume> <fpage>612</fpage>&#x02013;<lpage>622</lpage>. <pub-id pub-id-type="doi">10.1002/cbic.200500301</pub-id></citation></ref>
<ref id="B105"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Semeiks</surname> <given-names>J.</given-names></name> <name><surname>Borek</surname> <given-names>D.</given-names></name> <name><surname>Otwinowski</surname> <given-names>Z.</given-names></name> <name><surname>Grishin</surname> <given-names>N. V.</given-names></name></person-group> (<year>2014</year>). <article-title>Comparative genome sequencing reveals chemotype-specific gene clusters in the toxigenic black mold <italic>Stachybotrys</italic>.</article-title> <source><italic>BMC Genomics</italic></source> <volume>15</volume>:<issue>590</issue>. <pub-id pub-id-type="doi">10.1186/1471-2164-15-590</pub-id></citation></ref>
<ref id="B106"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Shu</surname> <given-names>Y. Z.</given-names></name> <name><surname>Cutrone</surname> <given-names>J. Q.</given-names></name> <name><surname>Klohr</surname> <given-names>S. E.</given-names></name> <name><surname>Huang</surname> <given-names>S.</given-names></name></person-group> (<year>1995</year>). <article-title>BMS-192548 a tetracyclic binding inhibitor of neuropeptide Y receptors, from <italic>Aspergillus niger</italic> WB2346. II. Physico-chemical properties and structural characterization.</article-title> <source><italic>J. Antibiot. (Tokyo)</italic></source> <volume>48</volume> <fpage>1060</fpage>&#x02013;<lpage>1065</lpage>. <pub-id pub-id-type="doi">10.7164/antibiotics.48.1060</pub-id></citation></ref>
<ref id="B107"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Simpson</surname> <given-names>J. T.</given-names></name> <name><surname>Wong</surname> <given-names>K.</given-names></name> <name><surname>Jackman</surname> <given-names>S. D.</given-names></name> <name><surname>Schein</surname> <given-names>J. E.</given-names></name> <name><surname>Jones</surname> <given-names>S. J.</given-names></name> <name><surname>Birol</surname> <given-names>I.</given-names></name></person-group> (<year>2009</year>). <article-title>ABySS: a parallel assembler for short read sequence data.</article-title> <source><italic>Genome Res.</italic></source> <volume>19</volume> <fpage>1117</fpage>&#x02013;<lpage>1123</lpage>. <pub-id pub-id-type="doi">10.1101/gr.089532.108</pub-id></citation></ref>
<ref id="B108"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Solovyev</surname> <given-names>V.</given-names></name> <name><surname>Kosarev</surname> <given-names>P.</given-names></name> <name><surname>Seledsov</surname> <given-names>I.</given-names></name> <name><surname>Vorobyev</surname> <given-names>D.</given-names></name></person-group> (<year>2006</year>). <article-title>Automatic annotation of eukaryotic genes, pseudogenes and promoters.</article-title> <source><italic>Genome Biol.</italic></source> <volume>7(Suppl. 1)</volume> <fpage>S10.1</fpage>&#x02013;<lpage>S10.2</lpage>. <pub-id pub-id-type="doi">10.1186/gb-2006-7-s1-s10</pub-id></citation></ref>
<ref id="B109"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Stachelhaus</surname> <given-names>T.</given-names></name> <name><surname>Mootz</surname> <given-names>H. D.</given-names></name> <name><surname>Marahiel</surname> <given-names>M. A.</given-names></name></person-group> (<year>1999</year>). <article-title>The specificity-conferring code of adenylation domains in nonribosomal peptide synthetases.</article-title> <source><italic>Chem. Biol.</italic></source> <volume>6</volume> <fpage>493</fpage>&#x02013;<lpage>505</lpage>. <pub-id pub-id-type="doi">10.1016/S1074-5521(99)80082-9</pub-id></citation></ref>
<ref id="B110"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Stanke</surname> <given-names>M.</given-names></name> <name><surname>Steinkamp</surname> <given-names>R.</given-names></name> <name><surname>Waack</surname> <given-names>S.</given-names></name> <name><surname>Morgenstern</surname> <given-names>B.</given-names></name></person-group> (<year>2004</year>). <article-title>AUGUSTUS: a web server for gene finding in eukaryotes.</article-title> <source><italic>Nucleic Acids Res.</italic></source> <volume>32</volume> <fpage>W309</fpage>&#x02013;<lpage>W312</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gkh379</pub-id></citation></ref>
<ref id="B111"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Strieker</surname> <given-names>M.</given-names></name> <name><surname>Tanovic</surname> <given-names>A.</given-names></name> <name><surname>Marahiel</surname> <given-names>M. A.</given-names></name></person-group> (<year>2010</year>). <article-title>Nonribosomal peptide synthetases: structures and dynamics.</article-title> <source><italic>Curr. Opin. Struct. Biol</italic></source> <volume>20</volume> <fpage>234</fpage>&#x02013;<lpage>240</lpage>. <pub-id pub-id-type="doi">10.1016/j.sbi.2010.01.009</pub-id></citation></ref>
<ref id="B112"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sussmuth</surname> <given-names>R.</given-names></name> <name><surname>Muller</surname> <given-names>J.</given-names></name> <name><surname>Von Dohren</surname> <given-names>H.</given-names></name> <name><surname>Molnar</surname> <given-names>I.</given-names></name></person-group> (<year>2011</year>). <article-title>Fungal cyclooligomer depsipeptides: from classical biochemistry to combinatorial biosynthesis.</article-title> <source><italic>Nat. Prod. Rep.</italic></source> <volume>28</volume> <fpage>99</fpage>&#x02013;<lpage>124</lpage>. <pub-id pub-id-type="doi">10.1039/c001463j</pub-id></citation></ref>
<ref id="B113"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tamano</surname> <given-names>K.</given-names></name> <name><surname>Sano</surname> <given-names>M.</given-names></name> <name><surname>Yamane</surname> <given-names>N.</given-names></name> <name><surname>Terabayashi</surname> <given-names>Y.</given-names></name> <name><surname>Toda</surname> <given-names>T.</given-names></name> <name><surname>Sunagawa</surname> <given-names>M.</given-names></name><etal/></person-group> (<year>2008</year>). <article-title>Transcriptional regulation of genes on the non-syntenic blocks of <italic>Aspergillus oryzae</italic> and its functional relationship to solid-state cultivation.</article-title> <source><italic>Fungal Genet. Biol.</italic></source> <volume>45</volume> <fpage>139</fpage>&#x02013;<lpage>151</lpage>. <pub-id pub-id-type="doi">10.1016/j.fgb.2007.09.005</pub-id></citation></ref>
<ref id="B114"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Umemura</surname> <given-names>M.</given-names></name> <name><surname>Koike</surname> <given-names>H.</given-names></name> <name><surname>Nagano</surname> <given-names>N.</given-names></name> <name><surname>Ishii</surname> <given-names>T.</given-names></name> <name><surname>Kawano</surname> <given-names>J.</given-names></name> <name><surname>Yamane</surname> <given-names>N.</given-names></name><etal/></person-group> (<year>2013a</year>). <article-title>MIDDAS-M: Motif-independent de novo detection of secondary metabolite gene clusters through the integration of genome sequencing and transcriptome data.</article-title> <source><italic>PLoS ONE</italic></source> <volume>8</volume>:<issue>e84028</issue>. <pub-id pub-id-type="doi">10.1371/journal.pone.0084028</pub-id></citation></ref>
<ref id="B115"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Umemura</surname> <given-names>M.</given-names></name> <name><surname>Koyama</surname> <given-names>Y.</given-names></name> <name><surname>Takeda</surname> <given-names>I.</given-names></name> <name><surname>Hagiwara</surname> <given-names>H.</given-names></name> <name><surname>Ikegami</surname> <given-names>T.</given-names></name> <name><surname>Koike</surname> <given-names>H.</given-names></name><etal/></person-group> (<year>2013b</year>). <article-title>Fine de novo sequencing of a fungal genome using only SOLiD short read data: verification on <italic>Aspergillus oryzae</italic> RIB40.</article-title> <source><italic>PLoS ONE</italic></source> <volume>8</volume>:<issue>e63673</issue>. <pub-id pub-id-type="doi">10.1371/journal.pone.0063673</pub-id></citation></ref>
<ref id="B116"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>van den Berg</surname> <given-names>M. A.</given-names></name> <name><surname>Albang</surname> <given-names>R.</given-names></name> <name><surname>Albermann</surname> <given-names>K.</given-names></name> <name><surname>Badger</surname> <given-names>J. H.</given-names></name> <name><surname>Daran</surname> <given-names>J. M.</given-names></name> <name><surname>Driessen</surname> <given-names>A. J.</given-names></name><etal/></person-group> (<year>2008</year>). <article-title>Genome sequencing and analysis of the filamentous fungus <italic>Penicillium chrysogenum</italic>.</article-title> <source><italic>Nat. Biotechnol.</italic></source> <volume>26</volume> <fpage>1161</fpage>&#x02013;<lpage>1168</lpage>. <pub-id pub-id-type="doi">10.1038/nbt.1498</pub-id></citation></ref>
<ref id="B117"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>van Dijk</surname> <given-names>E. L.</given-names></name> <name><surname>Auger</surname> <given-names>H.</given-names></name> <name><surname>Jaszczyszyn</surname> <given-names>Y.</given-names></name> <name><surname>Thermes</surname> <given-names>C.</given-names></name></person-group> (<year>2014</year>). <article-title>Ten years of next-generation sequencing technology.</article-title> <source><italic>Trends Genet.</italic></source> <volume>30</volume> <fpage>418</fpage>&#x02013;<lpage>426</lpage>. <pub-id pub-id-type="doi">10.1016/j.tig.2014.07.001</pub-id></citation></ref>
<ref id="B118"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>von Dohren</surname> <given-names>H.</given-names></name></person-group> (<year>2009</year>). <article-title>A survey of nonribosomal peptide synthetase (NRPS) genes in <italic>Aspergillus nidulans</italic>.</article-title> <source><italic>Fungal Genet. Biol.</italic></source> <volume>46(Suppl. 1)</volume> <fpage>S45</fpage>&#x02013;<lpage>S52</lpage>. <pub-id pub-id-type="doi">10.1016/j.fgb.2008.08.008</pub-id></citation></ref>
<ref id="B119"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wackler</surname> <given-names>B.</given-names></name> <name><surname>Lackner</surname> <given-names>G.</given-names></name> <name><surname>Chooi</surname> <given-names>Y. H.</given-names></name> <name><surname>Hoffmeister</surname> <given-names>D.</given-names></name></person-group> (<year>2012</year>). <article-title>Characterization of the <italic>Suillus grevillei</italic> quinone synthetase GreA supports a nonribosomal code for aromatic alpha-keto acids.</article-title> <source><italic>Chembiochem</italic></source> <volume>13</volume> <fpage>1798</fpage>&#x02013;<lpage>1804</lpage>. <pub-id pub-id-type="doi">10.1002/cbic.201200187</pub-id></citation></ref>
<ref id="B120"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Walsh</surname> <given-names>C. T.</given-names></name> <name><surname>Fischbach</surname> <given-names>M. A.</given-names></name></person-group> (<year>2010</year>). <article-title>Natural products version 2.0: connecting genes to molecules</article-title> <source><italic>J. Am. Chem. Soc.</italic></source> <volume>132</volume> <fpage>2469</fpage>&#x02013;<lpage>2493</lpage>. <pub-id pub-id-type="doi">10.1021/ja909118a</pub-id></citation></ref>
<ref id="B121"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Walsh</surname> <given-names>C. T.</given-names></name> <name><surname>O&#x02019;brien</surname> <given-names>R. V.</given-names></name> <name><surname>Khosla</surname> <given-names>C.</given-names></name></person-group> (<year>2013</year>). <article-title>Nonproteinogenic amino acid building blocks for nonribosomal peptide and hybrid polyketide scaffolds.</article-title> <source><italic>Angew. Chem. Int. Ed. Engl.</italic></source> <volume>52</volume> <fpage>7098</fpage>&#x02013;<lpage>7124</lpage>. <pub-id pub-id-type="doi">10.1002/anie.201208344</pub-id></citation></ref>
<ref id="B122"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wang</surname> <given-names>S.</given-names></name> <name><surname>Xu</surname> <given-names>Y.</given-names></name> <name><surname>Maine</surname> <given-names>E. A.</given-names></name> <name><surname>Wijeratne</surname> <given-names>E. M.</given-names></name> <name><surname>Espinosa-Artiles</surname> <given-names>P.</given-names></name> <name><surname>Gunatilaka</surname> <given-names>A. A.</given-names></name><etal/></person-group> (<year>2008</year>). <article-title>Functional characterization of the biosynthesis of radicicol, an Hsp90 inhibitor resorcylic acid lactone from Chaetomium chiversii.</article-title> <source><italic>Chem. Biol.</italic></source> <volume>15</volume> <fpage>1328</fpage>&#x02013;<lpage>1338</lpage>. <pub-id pub-id-type="doi">10.1016/j.chembiol.2008.10.006</pub-id></citation></ref>
<ref id="B123"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wiemann</surname> <given-names>P.</given-names></name> <name><surname>Guo</surname> <given-names>C. J.</given-names></name> <name><surname>Palmer</surname> <given-names>J. M.</given-names></name> <name><surname>Sekonyela</surname> <given-names>R.</given-names></name> <name><surname>Wang</surname> <given-names>C. C.</given-names></name> <name><surname>Keller</surname> <given-names>N. P.</given-names></name></person-group> (<year>2013a</year>). <article-title>Prototype of an intertwined secondary-metabolite supercluster.</article-title> <source><italic>Proc. Natl. Acad. Sci. U.S.A.</italic></source> <volume>110</volume> <fpage>17065</fpage>&#x02013;<lpage>17070</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.1313258110</pub-id></citation></ref>
<ref id="B124"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wiemann</surname> <given-names>P.</given-names></name> <name><surname>Sieber</surname> <given-names>C. M.</given-names></name> <name><surname>Von Bargen</surname> <given-names>K. W.</given-names></name> <name><surname>Studt</surname> <given-names>L.</given-names></name> <name><surname>Niehaus</surname> <given-names>E. M.</given-names></name> <name><surname>Espino</surname> <given-names>J. J.</given-names></name><etal/></person-group> (<year>2013b</year>). <article-title>Deciphering the cryptic genome: genome-wide analyses of the rice pathogen <italic>Fusarium fujikuroi</italic> reveal complex regulation of secondary metabolism and novel metabolites.</article-title> <source><italic>PLoS Pathog.</italic></source> <volume>9</volume>:<issue>e1003475</issue>. <pub-id pub-id-type="doi">10.1371/journal.ppat.1003475</pub-id></citation></ref>
<ref id="B125"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Winter</surname> <given-names>J. M.</given-names></name> <name><surname>Sato</surname> <given-names>M.</given-names></name> <name><surname>Sugimoto</surname> <given-names>S.</given-names></name> <name><surname>Chiou</surname> <given-names>G.</given-names></name> <name><surname>Garg</surname> <given-names>N. K.</given-names></name> <name><surname>Tang</surname> <given-names>Y.</given-names></name><etal/></person-group> (<year>2012</year>). <article-title>Identification and characterization of the chaetoviridin and chaetomugilin gene cluster in <italic>Chaetomium globosum</italic> reveal dual functions of an iterative highly-reducing polyketide synthase.</article-title> <source><italic>J. Am. Chem. Soc.</italic></source> <volume>134</volume> <fpage>17900</fpage>&#x02013;<lpage>17903</lpage>. <pub-id pub-id-type="doi">10.1021/ja3090498</pub-id></citation></ref>
<ref id="B126"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Xie</surname> <given-names>X.</given-names></name> <name><surname>Meehan</surname> <given-names>M. J.</given-names></name> <name><surname>Xu</surname> <given-names>W.</given-names></name> <name><surname>Dorrestein</surname> <given-names>P. C.</given-names></name> <name><surname>Tang</surname> <given-names>Y.</given-names></name></person-group> (<year>2009</year>). <article-title>Acyltransferase mediated polyketide release from a fungal megasynthase.</article-title> <source><italic>J. Am. Chem. Soc.</italic></source> <volume>131</volume> <fpage>8388</fpage>&#x02013;<lpage>8389</lpage>. <pub-id pub-id-type="doi">10.1021/ja903203g</pub-id></citation></ref>
<ref id="B127"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Xie</surname> <given-names>X.</given-names></name> <name><surname>Watanabe</surname> <given-names>K.</given-names></name> <name><surname>Wojcicki</surname> <given-names>W. A.</given-names></name> <name><surname>Wang</surname> <given-names>C. C.</given-names></name> <name><surname>Tang</surname> <given-names>Y.</given-names></name></person-group> (<year>2006</year>). <article-title>Biosynthesis of lovastatin analogs with a broadly specific acyltransferase.</article-title> <source><italic>Chem. Biol.</italic></source> <volume>13</volume> <fpage>1161</fpage>&#x02013;<lpage>1169</lpage>. <pub-id pub-id-type="doi">10.1016/j.chembiol.2006.09.008</pub-id></citation></ref>
<ref id="B128"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Xu</surname> <given-names>W.</given-names></name> <name><surname>Chooi</surname> <given-names>Y. H.</given-names></name> <name><surname>Choi</surname> <given-names>J. W.</given-names></name> <name><surname>Li</surname> <given-names>S.</given-names></name> <name><surname>Vederas</surname> <given-names>J. C.</given-names></name> <name><surname>Da Silva</surname> <given-names>N. A.</given-names></name><etal/></person-group> (<year>2013a</year>). <article-title>LovG: the thioesterase required for dihydromonacolin L release and lovastatin nonaketide synthase turnover in lovastatin biosynthesis.</article-title> <source><italic>Angew. Chem. Int. Ed. Engl.</italic></source> <volume>52</volume> <fpage>6472</fpage>&#x02013;<lpage>6475</lpage>. <pub-id pub-id-type="doi">10.1002/anie.201302406</pub-id></citation></ref>
<ref id="B129"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Xu</surname> <given-names>Y.</given-names></name> <name><surname>Espinosa-Artiles</surname> <given-names>P.</given-names></name> <name><surname>Schubert</surname> <given-names>V.</given-names></name> <name><surname>Xu</surname> <given-names>Y. M.</given-names></name> <name><surname>Zhang</surname> <given-names>W.</given-names></name> <name><surname>Lin</surname> <given-names>M.</given-names></name><etal/></person-group> (<year>2013b</year>). <article-title>Characterization of the biosynthetic genes for 10,11-dehydrocurvularin, a heat shock response-modulating anticancer fungal polyketide from <italic>Aspergillus terreus</italic>.</article-title> <source><italic>Appl. Environ. Microbiol.</italic></source> <volume>79</volume> <fpage>2038</fpage>&#x02013;<lpage>2047</lpage>. <pub-id pub-id-type="doi">10.1128/AEM.03334-12</pub-id></citation></ref>
<ref id="B130"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Xu</surname> <given-names>W.</given-names></name> <name><surname>Gavia</surname> <given-names>D. J.</given-names></name> <name><surname>Tang</surname> <given-names>Y.</given-names></name></person-group> (<year>2014a</year>). <article-title>Biosynthesis of fungal indole alkaloids.</article-title> <source><italic>Nat. Prod. Rep.</italic></source> <volume>31</volume> <fpage>1474</fpage>&#x02013;<lpage>1487</lpage>. <pub-id pub-id-type="doi">10.1039/c4np00073k</pub-id></citation></ref>
<ref id="B131"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Xu</surname> <given-names>Y.</given-names></name> <name><surname>Zhou</surname> <given-names>T.</given-names></name> <name><surname>Zhang</surname> <given-names>S.</given-names></name> <name><surname>Espinosa-Artiles</surname> <given-names>P.</given-names></name> <name><surname>Wang</surname> <given-names>L.</given-names></name> <name><surname>Zhang</surname> <given-names>W.</given-names></name><etal/></person-group> (<year>2014b</year>). <article-title>Diversity-oriented combinatorial biosynthesis of benzenediol lactone scaffolds by subunit shu&#x0FB04;ing of fungal polyketide synthases.</article-title> <source><italic>Proc. Natl. Acad. Sci. U.S.A.</italic></source> <volume>111</volume> <fpage>12354</fpage>&#x02013;<lpage>12359</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.1406999111</pub-id></citation></ref>
<ref id="B132"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Xu</surname> <given-names>Y.</given-names></name> <name><surname>Orozco</surname> <given-names>R.</given-names></name> <name><surname>Wijeratne</surname> <given-names>E. M.</given-names></name> <name><surname>Gunatilaka</surname> <given-names>A. A.</given-names></name> <name><surname>Stock</surname> <given-names>S. P.</given-names></name> <name><surname>Molnar</surname> <given-names>I.</given-names></name></person-group> (<year>2008</year>). <article-title>Biosynthesis of the cyclooligomer depsipeptide beauvericin, a virulence factor of the entomopathogenic fungus <italic>Beauveria bassiana</italic>.</article-title> <source><italic>Chem. Biol.</italic></source> <volume>15</volume> <fpage>898</fpage>&#x02013;<lpage>907</lpage>. <pub-id pub-id-type="doi">10.1016/j.chembiol.2008.07.011</pub-id></citation></ref>
<ref id="B133"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Yang</surname> <given-names>G.</given-names></name> <name><surname>Rose</surname> <given-names>M. S.</given-names></name> <name><surname>Turgeon</surname> <given-names>B. G.</given-names></name> <name><surname>Yoder</surname> <given-names>O. C.</given-names></name></person-group> (<year>1996</year>). <article-title>A polyketide synthase is required for fungal virulence and production of the polyketide T-toxin.</article-title> <source><italic>Plant Cell</italic></source> <volume>8</volume> <fpage>2139</fpage>&#x02013;<lpage>2150</lpage>. <pub-id pub-id-type="doi">10.1105/tpc.8.11.2139</pub-id></citation></ref>
<ref id="B134"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Yu</surname> <given-names>J.</given-names></name> <name><surname>Chang</surname> <given-names>P. K.</given-names></name> <name><surname>Ehrlich</surname> <given-names>K. C.</given-names></name> <name><surname>Cary</surname> <given-names>J. W.</given-names></name> <name><surname>Bhatnagar</surname> <given-names>D.</given-names></name> <name><surname>Cleveland</surname> <given-names>T. E.</given-names></name><etal/></person-group> (<year>2004</year>). <article-title>Clustered pathway genes in aflatoxin biosynthesis.</article-title> <source><italic>Appl. Environ. Microbiol.</italic></source> <volume>70</volume> <fpage>1253</fpage>&#x02013;<lpage>1262</lpage>. <pub-id pub-id-type="doi">10.1128/AEM.70.3.1253-1262.2004</pub-id></citation></ref>
<ref id="B135"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zabala</surname> <given-names>A. O.</given-names></name> <name><surname>Chooi</surname> <given-names>Y. H.</given-names></name> <name><surname>Choi</surname> <given-names>M. S.</given-names></name> <name><surname>Lin</surname> <given-names>H. C.</given-names></name> <name><surname>Tang</surname> <given-names>Y.</given-names></name></person-group> (<year>2014</year>). <article-title>Fungal polyketide synthase product chain-length control by partnering thiohydrolase.</article-title> <source><italic>ACS Chem. Biol.</italic></source> <volume>9</volume> <fpage>1576</fpage>&#x02013;<lpage>1586</lpage>. <pub-id pub-id-type="doi">10.1021/cb500284t</pub-id></citation></ref>
<ref id="B136"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zabala</surname> <given-names>A. O.</given-names></name> <name><surname>Xu</surname> <given-names>W.</given-names></name> <name><surname>Chooi</surname> <given-names>Y. H.</given-names></name> <name><surname>Tang</surname> <given-names>Y.</given-names></name></person-group> (<year>2012</year>). <article-title>Characterization of a silent azaphilone gene cluster from <italic>Aspergillus niger</italic> ATCC 1015 reveals a hydroxylation-mediated pyran-ring formation.</article-title> <source><italic>Chem. Biol.</italic></source> <volume>19</volume> <fpage>1049</fpage>&#x02013;<lpage>1059</lpage>. <pub-id pub-id-type="doi">10.1016/j.chembiol.2012.07.004</pub-id></citation></ref>
<ref id="B137"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zazopoulos</surname> <given-names>E.</given-names></name> <name><surname>Huang</surname> <given-names>K.</given-names></name> <name><surname>Staffa</surname> <given-names>A.</given-names></name> <name><surname>Liu</surname> <given-names>W.</given-names></name> <name><surname>Bachmann</surname> <given-names>B. O.</given-names></name> <name><surname>Nonaka</surname> <given-names>K.</given-names></name><etal/></person-group> (<year>2003</year>). <article-title>A genomics-guided approach for discovering and expressing cryptic metabolic pathways.</article-title> <source><italic>Nat. Biotechnol.</italic></source> <volume>21</volume> <fpage>187</fpage>&#x02013;<lpage>190</lpage>. <pub-id pub-id-type="doi">10.1038/nbt784</pub-id></citation></ref>
<ref id="B138"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zerbino</surname> <given-names>D. R.</given-names></name> <name><surname>Birney</surname> <given-names>E.</given-names></name></person-group> (<year>2008</year>). <article-title>Velvet: algorithms for de novo short read assembly using de Bruijn graphs.</article-title> <source><italic>Genome Res.</italic></source> <volume>18</volume> <fpage>821</fpage>&#x02013;<lpage>829</lpage>. <pub-id pub-id-type="doi">10.1101/gr.074492.107</pub-id></citation></ref>
<ref id="B139"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zhang</surname> <given-names>W.</given-names></name> <name><surname>Chen</surname> <given-names>J.</given-names></name> <name><surname>Yang</surname> <given-names>Y.</given-names></name> <name><surname>Tang</surname> <given-names>Y.</given-names></name> <name><surname>Shang</surname> <given-names>J.</given-names></name> <name><surname>Shen</surname> <given-names>B.</given-names></name></person-group> (<year>2011</year>). <article-title>A practical comparison of de novo genome assembly software tools for next-generation sequencing technologies.</article-title> <source><italic>PLoS ONE</italic></source> <volume>6</volume>:<issue>e17915</issue>. <pub-id pub-id-type="doi">10.1371/journal.pone.0017915</pub-id></citation></ref>
<ref id="B140"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zhou</surname> <given-names>H.</given-names></name> <name><surname>Qiao</surname> <given-names>K.</given-names></name> <name><surname>Gao</surname> <given-names>Z.</given-names></name> <name><surname>Meehan</surname> <given-names>M. J.</given-names></name> <name><surname>Li</surname> <given-names>J. W.</given-names></name> <name><surname>Zhao</surname> <given-names>X.</given-names></name><etal/></person-group> (<year>2010</year>). <article-title>Enzymatic synthesis of resorcylic acid lactones by cooperation of fungal iterative polyketide synthases involved in hypothemycin biosynthesis.</article-title> <source><italic>J. Am. Chem. Soc.</italic></source> <volume>132</volume> <fpage>4530</fpage>&#x02013;<lpage>4531</lpage>. <pub-id pub-id-type="doi">10.1021/ja100060k</pub-id></citation></ref>
</ref-list>
<fn-group>
<fn id="fn01"><label>1</label><p><ext-link ext-link-type="uri" xlink:href="http://www.ncbi.nlm.nih.gov/guide/howto/run-blast-local/">http://www.ncbi.nlm.nih.gov/guide/howto/run-blast-local/</ext-link></p></fn>
<fn id="fn02"><label>2</label><p><ext-link ext-link-type="uri" xlink:href="http://www.clcbio.com/">http://www.clcbio.com/</ext-link></p></fn>
<fn id="fn03"><label>3</label><p><ext-link ext-link-type="uri" xlink:href="http://www.geneious.com/">http://www.geneious.com/</ext-link></p></fn>
<fn id="fn04"><label>4</label><p><ext-link ext-link-type="uri" xlink:href="http://www.softberry.com/berry.phtml?topic=index&#x0026;group=programs&#x0026;subgroup=gfind">http://www.softberry.com/berry.phtml?topic=index&#x0026;group=programs&#x0026;subgroup=gfind</ext-link></p></fn>
<fn id="fn05"><label>5</label><p><ext-link ext-link-type="uri" xlink:href="http://bioinf.uni-greifswald.de/webaugustus/predictiontutorial.gsp#exampledata">http://bioinf.uni-greifswald.de/webaugustus/predictiontutorial.gsp#exampledata</ext-link></p></fn>
</fn-group>
</back>
</article>