<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Archiving and Interchange DTD v2.3 20070202//EN" "archivearticle.dtd">
<article article-type="methods-article" dtd-version="2.3" xml:lang="EN" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Bioeng. Biotechnol.</journal-id>
<journal-title>Frontiers in Bioengineering and Biotechnology</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Bioeng. Biotechnol.</abbrev-journal-title>
<issn pub-type="epub">2296-4185</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="publisher-id">1217811</article-id>
<article-id pub-id-type="doi">10.3389/fbioe.2023.1217811</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Bioengineering and Biotechnology</subject>
<subj-group>
<subject>Methods</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>gRNA-SeqRET: a universal tool for targeted and genome-scale gRNA design and sequence extraction for prokaryotes and eukaryotes</article-title>
<alt-title alt-title-type="left-running-head">Simirenko et al.</alt-title>
<alt-title alt-title-type="right-running-head">
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/fbioe.2023.1217811">10.3389/fbioe.2023.1217811</ext-link>
</alt-title>
</title-group>
<contrib-group>
<contrib contrib-type="author">
<name>
<surname>Simirenko</surname>
<given-names>Lisa</given-names>
</name>
<xref ref-type="aff" rid="aff1">
<sup>1</sup>
</xref>
</contrib>
<contrib contrib-type="author">
<name>
<surname>Cheng</surname>
<given-names>Jan-Fang</given-names>
</name>
<xref ref-type="aff" rid="aff1">
<sup>1</sup>
</xref>
<xref ref-type="aff" rid="aff2">
<sup>2</sup>
</xref>
<uri xlink:href="https://loop.frontiersin.org/people/472518/overview"/>
</contrib>
<contrib contrib-type="author" corresp="yes">
<name>
<surname>Blaby</surname>
<given-names>Ian</given-names>
</name>
<xref ref-type="aff" rid="aff1">
<sup>1</sup>
</xref>
<xref ref-type="aff" rid="aff2">
<sup>2</sup>
</xref>
<xref ref-type="corresp" rid="c001">&#x2a;</xref>
<uri xlink:href="https://loop.frontiersin.org/people/2304083/overview"/>
</contrib>
</contrib-group>
<aff id="aff1">
<sup>1</sup>
<institution>US Department of Energy Joint Genome Institute</institution>, <institution>Lawrence Berkeley National Laboratory</institution>, <addr-line>Berkeley</addr-line>, <addr-line>CA</addr-line>, <country>United States</country>
</aff>
<aff id="aff2">
<sup>2</sup>
<institution>Environmental Genomics and Systems Biology Division</institution>, <institution>Lawrence Berkeley National Laboratory</institution>, <addr-line>Berkeley</addr-line>, <addr-line>CA</addr-line>, <country>United States</country>
</aff>
<author-notes>
<fn fn-type="edited-by">
<p>
<bold>Edited by:</bold> <ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/806360/overview">Anindya Bandyopadhyay</ext-link>, Reliance Industries, India</p>
</fn>
<fn fn-type="edited-by">
<p>
<bold>Reviewed by:</bold> <ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/225533/overview">Nicholas R. Sandoval</ext-link>, Tulane University, United States</p>
<p>
<ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/2372627/overview">Rahul Badhwar</ext-link>, Reliance Industries, India</p>
</fn>
<corresp id="c001">&#x2a;Correspondence: Ian Blaby, <email>ikblaby@lbl.gov</email>
</corresp>
</author-notes>
<pub-date pub-type="epub">
<day>29</day>
<month>08</month>
<year>2023</year>
</pub-date>
<pub-date pub-type="collection">
<year>2023</year>
</pub-date>
<volume>11</volume>
<elocation-id>1217811</elocation-id>
<history>
<date date-type="received">
<day>05</day>
<month>05</month>
<year>2023</year>
</date>
<date date-type="accepted">
<day>21</day>
<month>08</month>
<year>2023</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#xa9; 2023 Simirenko, Cheng and Blaby.</copyright-statement>
<copyright-year>2023</copyright-year>
<copyright-holder>Simirenko, Cheng and Blaby</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/">
<p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p>
</license>
</permissions>
<abstract>
<p>High-throughput genetic screening is frequently employed to rapidly associate gene with phenotype and establish sequence-function relationships. With the advent of CRISPR technology, and the ability to functionally interrogate previously genetically recalcitrant organisms, non-model organisms can be investigated using pooled guide RNA (gRNA) libraries and sequencing-based assays to quantitatively assess fitness of every targeted locus in parallel. To aid the construction of pooled gRNA assemblies, we have developed an <italic>in silico</italic> design workflow for gRNA selection using the gRNA Sequence Region Extraction Tool (gRNA-SeqRET). Built upon the previously developed CCTop, gRNA-SeqRET enables automated, scalable design of gRNA libraries that target user-specified regions or whole genomes of any prokaryote or eukaryote. Additionally, gRNA-SeqRET automates the bulk extraction of any regions of sequence relative to genes or other features, aiding in the design of homology arms for insertion or deletion constructs. We also assess <italic>in silico</italic> the application of a designed gRNA library to other closely related genomes and demonstrate that for very closely related organisms Average Nucleotide Identity (ANI) &#x3e; 95% a large fraction of the library may be of relevance. The gRNA-SeqRET web application pipeline can be accessed at <ext-link ext-link-type="uri" xlink:href="https://grna.jgi.doe.gov/">https://grna.jgi.doe.gov</ext-link>. The source code is comprised of freely available software tools and customized Python scripts, and is available at <ext-link ext-link-type="uri" xlink:href="https://bitbucket.org/berkeleylab/grnadesigner/src/master/">https://bitbucket.org/berkeleylab/grnadesigner/src/master/</ext-link> under a modified BSD open-source license (<ext-link ext-link-type="uri" xlink:href="https://bitbucket.org/berkeleylab/grnadesigner">https://bitbucket.org/berkeleylab/grnadesigner</ext-link>).</p>
</abstract>
<kwd-group>
<kwd>CRISPR</kwd>
<kwd>gRNA</kwd>
<kwd>design tool</kwd>
<kwd>batch DNA extraction</kwd>
<kwd>CRISPR screen</kwd>
</kwd-group>
<custom-meta-wrap>
<custom-meta>
<meta-name>section-at-acceptance</meta-name>
<meta-value>Synthetic Biology</meta-value>
</custom-meta>
</custom-meta-wrap>
</article-meta>
</front>
<body>
<sec id="s1">
<title>Introduction</title>
<p>CRISPR-based genome editing has rapidly become the targeted engineering technology of choice due to its programmability, scalability and near universal application (<xref ref-type="bibr" rid="B12">Jinek et al., 2012</xref>; <xref ref-type="bibr" rid="B32">Wiedenheft et al., 2012</xref>; <xref ref-type="bibr" rid="B11">Jiang et al., 2013</xref>; <xref ref-type="bibr" rid="B26">Ran et al., 2013</xref>). The core machinery enabling editing comprises a CRISPR-associated protein (Cas; an endonuclease) and a short guide RNA (a fusion of a variable, target specific sequence and a Cas-specific stem-loop forming sequence required for CRISPR nuclease maturation). Programmability is achieved by complementarity of this variable region to the target DNA, which for Cas binding and double-strand cleavage must be juxtaposed to a Cas-specific sequence termed the protospacer adjacent motif (PAM) (<xref ref-type="bibr" rid="B19">Mojica et al., 2009</xref>). This requirement, which for <italic>Streptococcus pyogenes</italic> Cas9, the first to be discovered and the most widely used, is 5&#x2032; NGG 3&#x2032;, constitutes the only limitation in targeting DNA. However, even here alternate Cas endonuclease&#x2019;s afford flexibility due to different PAMs (albeit with different activities; for example, Cas12a with a PAM of 5&#x2032; YTN 3&#x2032;, induces a 5&#x2032; overhang double strand cleavage) (<xref ref-type="bibr" rid="B35">Zetsche et al., 2015</xref>; <xref ref-type="bibr" rid="B34">Yamano et al., 2016</xref>).</p>
<p>Initial DNA-editing exploited the double-strand cutting induced by Cas9 followed by low efficiency non-homologous end joining (NHEJ) repair mechanisms in eukaryotes yielding loss-of-function mutants. In the absence of NHEJ, a DNA fragment comprising regions of homology on either side of the targeted cut site allows homologous recombination repair to generate a scarless mutation in the genome. Alternatively, CRISPR interference or activation (CRISPRi/a) can be employed to modulate transcription without inducing a break in the genome (<xref ref-type="bibr" rid="B25">Qi et al., 2013</xref>; <xref ref-type="bibr" rid="B17">Liu et al., 2019</xref>). As the technology has matured additional applications have been developed furthering CRISPR&#x2019;s editing utility (<xref ref-type="bibr" rid="B23">Pickar-Oliver and Gersbach, 2019</xref>; <xref ref-type="bibr" rid="B16">Liu et al., 2021</xref>; <xref ref-type="bibr" rid="B21">Nakamura et al., 2021</xref>; <xref ref-type="bibr" rid="B31">Wang and Doudna, 2023</xref>).</p>
<p>Many web-based and downloadable computational tools are available for gRNA design. These tools typically provide the user with 20 nucleotide sequences flanking a PAM for targeting the specified loci [(<xref ref-type="bibr" rid="B20">Naito et al., 2015</xref>; <xref ref-type="bibr" rid="B9">Doench et al., 2016</xref>; <xref ref-type="bibr" rid="B6">Concordet and Haeussler, 2018</xref>) and reviewed in (<xref ref-type="bibr" rid="B33">Wilson et al., 2018</xref>; <xref ref-type="bibr" rid="B1">Alipanahi et al., 2022</xref>)]. Automated design tools are particularly useful for the experimental design of constructs involving many, or genome-scale, targets such as needed for CRISPR-screening (<xref ref-type="bibr" rid="B2">Bock et al., 2022</xref>), or Perturb-Seq (<xref ref-type="bibr" rid="B8">Dixit et al., 2016</xref>). However, many preexisting tools are limited to single or pre-computed model organisms, or, where user-provided genomes can be provided, are specifically optimized for prokaryote genome architecture (i.e., input sequence format does not allow for structural annotations such as multiple chromosomes or specifying intron/exons coding/non-coding regions) (<xref ref-type="bibr" rid="B24">Poudel et al., 2022</xref>). CCTop and CHOPCHOP, for example, include the genomes for many organisms (<xref ref-type="bibr" rid="B29">Stemmer et al., 2015</xref>; <xref ref-type="bibr" rid="B14">Labun et al., 2019</xref>), and additional genomes can be requested by email. Other tools cater to specific communities or groups of organism (<xref ref-type="bibr" rid="B22">Peng and Tarleton, 2015</xref>; <xref ref-type="bibr" rid="B10">He et al., 2021</xref>).</p>
<p>To overcome these limitations, and to enable universal, organism-agnostic design, we developed the guide RNA Sequence Extraction Tool (gRNA-SeqRET), which is built upon the previously published tool CCTop. CCTop&#x2019;s standalone version was specifically selected due to its open-source licensing, allowing further development, and the options it provides for design. gRNA-SeqRET allows users to create their own accounts where genomes can be uploaded and securely saved. Designs are scoped by entering a series of criteria into the website, and the data packaged and piped into CCTop. Once the job is complete, the results are accessible for download from the tool&#x2019;s website. gRNA-SeqRET has two main advantages over other tools: 1) compatible with any user-provided input genome files in GenBank and GFF formats, allowing universal design for both prokaryote and eukaryote genome structures; and 2) functionality to bulk extract specified target DNA regions for scalable repair template design.</p>
</sec>
<sec sec-type="methods" id="s2">
<title>Methods</title>
<p>gRNA-SeqRET employs Flask and jQuery for the web-based user interface (UI) and a PostgreSQL database which maintains track of the user&#x2019;s genomes and submissions. The tool is composed in Python 3, and utilizes the following open source applications for the indicated tasks: CCTop standalone (<xref ref-type="bibr" rid="B29">Stemmer et al., 2015</xref>)&#x2014;generates the complete list of potential gRNAs for a given pre-processed genome file ranked by predicted cutting score, and a FASTA file with the extracted target region(s); BowTie v1.3.0 (<xref ref-type="bibr" rid="B15">Langmead et al., 2009</xref>)&#x2014;generates the indexes needed by CCTop; The ViennaRNA Package (<xref ref-type="bibr" rid="B18">Lorenz et al., 2011</xref>)&#x2014;generates RNA folding predictions used to evaluate the gRNAs in the CCTop output; BioPython (<xref ref-type="bibr" rid="B5">Cock et al., 2009</xref>)&#x2014;Bio.SeqIO is used to parse genome sequence from GenBank files and convert the sequence to FASTA format; GFFutils v0.11.1 (<ext-link ext-link-type="uri" xlink:href="https://daler.github.io/gffutils/index.html">https://daler.github.io/gffutils/index.html</ext-link>)&#x2014;provides methods for creating an SQLite database (which contains the processed genome files) from a GFF and searching features annotated in GFF files; GFFtools-GX (<ext-link ext-link-type="uri" xlink:href="https://github.com/vipints/GFFtools-GX">https://github.com/vipints/GFFtools-GX</ext-link>)&#x2014;converts GenBank annotations to GFF3 format; Cromwell&#x2014;Workflow engine for automating the back-end pipeline (<ext-link ext-link-type="uri" xlink:href="https://cromwell.readthedocs.io/en/stable/">https://cromwell.readthedocs.io/en/stable/</ext-link>).</p>
</sec>
<sec sec-type="results|discussion" id="s3">
<title>Results and Discussion</title>
<sec id="s3-1">
<title>gRNA-SeqRET is compatible with any prokaryote or eukaryote genome</title>
<p>The goal of gRNA-SeqRET is to provide an intuitive web-based user interface for the design of gRNA sequences and custom sequence extraction from user-provided genomes. <xref ref-type="fig" rid="F1">Figure 1</xref> provides overview schematics of the general tool workflow, focusing on both the processing of the uploaded genome data and the subsequent extraction and guide RNA designs.</p>
<fig id="F1" position="float">
<label>FIGURE 1</label>
<caption>
<p>Flowcharts demonstrating relationships between input files, logic choices and outputs for gRNA-SeqRET. Schemas are shown for saving target genomes <bold>(A)</bold> and sequence extraction and gRNA design <bold>(B)</bold>, where diamond boxes represent conditions and rounded rectangles represent tasks and files.</p>
</caption>
<graphic xlink:href="fbioe-11-1217811-g001.tif"/>
</fig>
<p>The gRNA-SeqRET application accepts either FASTA or GenBank file types for a given organism&#x2019;s genome. New users of gRNA-SeqRET are required to register an account, which serves two purposes. Firstly, this allows uploaded genomes to be securely maintained on the server (genome files need only be uploaded once, and no files are accessible to other investigators). Secondly, this approach enables asynchronous use of the tool, by which an uploaded genome can be saved and processing (which takes many minutes) can be performed in the background. Files can be uploaded in compressed format (.zip or .gz), though must only contain a single file. If a FASTA file is uploaded the user will be prompted to additionally provide annotations in the format of a GFF3 file (<xref ref-type="fig" rid="F1">Figure 1A</xref>; <xref ref-type="table" rid="T1">Table 1</xref>). The GFF format is preferred due to the complexities of converting GenBank to GFF, especially for eukaryotes. The uploaded files are processed in order to generate the required input files for CCTop: Bowtie indices, Exon and Gene BED and json files. Specifically, indices are generated using bowtie-build, and the BED files are created using a python script provided with CCTop&#x2019;s standalone source code. Additionally, a searchable SQLite database is created using GFFUtils, which contains the annotations extracted from the uploaded GenBank or GFF file. This is used to generate JSON files that list the annotated genes or locus IDs and the features (e.g., CDS, exons, <italic>etc.</italic>) enabling the population of genome feature menus (see below).</p>
<table-wrap id="T1" position="float">
<label>TABLE 1</label>
<caption>
<p>Input files, parameters and results for gRNASeq-RET.</p>
</caption>
<table>
<thead valign="top">
<tr>
<th align="center">Item (URL)</th>
<th align="center">Field</th>
<th align="center">Description, page location, default parameters and input options</th>
</tr>
</thead>
<tbody valign="top">
<tr>
<td align="center">Input file(s) (<ext-link ext-link-type="uri" xlink:href="https://grna.jgi.doe.gov/save_genome.html">https://grna.jgi.doe.gov/save_genome.html</ext-link>)</td>
<td align="center">Genome data</td>
<td align="center">Requisite genome data providing sequence data and feature coordinates. Can be provided as either a GenBank or both a FASTA and GFF (GFF is preferred, especially for eukaryotes)</td>
</tr>
<tr>
<td rowspan="12" align="center">Target region selection (<ext-link ext-link-type="uri" xlink:href="https://grna.jgi.doe.gov/create_step1.html">https://grna.jgi.doe.gov/create_step1.html</ext-link>)</td>
<td align="center">Design name</td>
<td align="center">Unique user provided name that will become the job ID</td>
</tr>
<tr>
<td align="center">Target genome</td>
<td align="center">Uploaded, processed genomes can be selected from the dropdown menu</td>
</tr>
<tr>
<td align="center">Target regions</td>
<td align="center">Individual, or groups of loci in the menu detected from the input annotations can be copied to select, or left blank for genome-scale targeting</td>
</tr>
<tr>
<td align="center">Limit regions</td>
<td align="center">Target regions can be limited to coding, non-coding or both</td>
</tr>
<tr>
<td align="center">Start and end point definitions</td>
<td align="center">Defines the number of base pairs upstream and downstream of a feature in the genome file</td>
</tr>
<tr>
<td align="center">Guide RNA design</td>
<td align="center">Opens options for gRNA design for the specified target regions</td>
</tr>
<tr>
<td align="center">PAM</td>
<td align="center">Specifies the Cas-specific protospacer adjacent motif (PAM)</td>
</tr>
<tr>
<td align="center">Scaffold sequence</td>
<td align="center">Defines the scaffold sequence; default is the canonical hybrid scaffold and contributes to the folding score generated by gRNA-SeqRET</td>
</tr>
<tr>
<td align="center">Target site length</td>
<td align="center">Specifies the spacer region length; default is 20</td>
</tr>
<tr>
<td align="center">Max mismatches offsite targets</td>
<td align="center">Specifies the maximum number of mismatches that will be considered an off-site target and will be excluded</td>
</tr>
<tr>
<td align="center">No. gRNAs</td>
<td align="center">Specifies the number of guide sequences returned to the output file per target region. Note, unlike CCTop, which returns all possible guides, gRNASeqRET will only output this specified number, ranked by CCTop&#x2019;s predicted cutting score (default is 3)</td>
</tr>
<tr>
<td align="center">Target strand</td>
<td align="center">Optionally filter results based on whether the gRNAs are on the coding or non-coding strand</td>
</tr>
<tr>
<td align="center">Job results (<ext-link ext-link-type="uri" xlink:href="https://grna.jgi.doe.gov/design_list">https://grna.jgi.doe.gov/design_list</ext-link>)</td>
<td align="left"/>
<td align="center">Lists all complete and running jobs</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>Processing newly uploaded genomes typically takes around a minute for a small prokaryotic genome, and approximately 5&#x2013;10&#xa0;min for a large complex eukaryote (e.g., a plant) genome. Once generated, the processed files are preserved in a user&#x2019;s individual account and are inaccessible to other users. Consequently, this step needs to be performed only once per genome, and once complete, is ready for initiating designs. All processed genomes will be available in the &#x201c;Your Genomes&#x201d; page of gRNA-SeqRET with the status confirmed as complete (<xref ref-type="fig" rid="F2">Figure 2</xref>).</p>
<fig id="F2" position="float">
<label>FIGURE 2</label>
<caption>
<p>Walkthrough of the gRNA-SeqRET design process.</p>
</caption>
<graphic xlink:href="fbioe-11-1217811-g002.tif"/>
</fig>
</sec>
<sec id="s3-2">
<title>gRNA-SeqRET automates feature batch extraction and guide RNA design</title>
<p>Once preprocessing is complete, designs and sequence extractions can be performed on genomes uploaded by selecting the Create New Design tool in the top left corner of the menu (<xref ref-type="fig" rid="F1">Figures 1B</xref>, <xref ref-type="fig" rid="F2">2</xref>). After entering a name for the design, which will become the job title, and selecting a genome from the drop-down menu, the user will specify the loci to be targeted. All locus IDs identified in the provided genome input files will be automatically populated in the menu on the right of the page; from here, individual loci can be selected or pasted, or if left blank all loci will be applied for genome-scale targeting.</p>
<p>Next, the precise regions to target must be provided, enabling batch programming for 1) extraction of genomic sequence data within defined regions and 2) design of gRNA sequences within these defined areas (described in <xref ref-type="table" rid="T1">Table 1</xref>). Specific use-case examples for various options are described below. Target region definitions are enabled at the nucleotide-level resolution relative to defined features in the input files by entering details into subsequent menus (<xref ref-type="fig" rid="F1">Figures 1B</xref>, <xref ref-type="fig" rid="F2">2</xref>). This is done by firstly stipulating whether coding, non-coding or both coding and non-coding regions should be included, and secondly by entering the number of bases upstream or downstream of feature reference points to target. The nature of features (defined as reference points in gRNA-SeqRET menus) available in these menus is dependent on those provided in the genome input file; for prokaryotes this will generally be limited to coding sequences (CDSs) and non-coding sequences, but eukaryote GFF files may provide broader options, such as introns and exon coordinates, depending on the extent of structural annotations in the uploaded annotations file. The tool warns if incompatible inputs are provided (for example, an error message will appear and the submit button will be disabled if in the first section the default &#x201c;coding regions only&#x201d; is selected as well as a region upstream of the loci coding region). To provide maximal flexibility of region selection, a user can define the number of bases up- or downstream of two independent reference points, constituting the start and end points of the target regions within the genome. Bases upstream of a feature are designated as being negative, and positive numbers indicate the number of bases downstream of this feature (e.g., a start codon). A dynamically generated schematic illustrating the selected area aids the user by updating in real time as entries are made (<xref ref-type="fig" rid="F2">Figure 2</xref>).</p>
<p>This sequence selection feature has multiple utilities, as it enables genome-scale, batch extraction of defined regions. One example would be for homology arm design, as required for CRISPR repair template design, which can then be bulk-downloaded as a FASTA file. Additionally, defining sequence regions also serves to specify the regions with which to target gRNA design, such as targeting specifically upstream (but not so far upstream as to enter the next 5&#x2019; coding genome feature) of a gene for promoter-region CRISPRa guides, as described below. Furthermore, batch extraction of defined regions of sequence has functions beyond CRISPR, for example, defining homology arms for homologous recombination (HR)-mediated gene disruptions, or for bulk extraction of sequence upstream of coding regions for promoter libraries. At this stage the defined sequences can be downloaded as a FASTA file, and/or the tool can proceed to gRNA design.</p>
<p>If the gRNA design option is selected, additional fields are made available (<xref ref-type="fig" rid="F2">Figure 2</xref>). Several of these parameters, including the specific PAM, scaffold sequence, target site length and number of allowed sequence mismatches are necessary inputs for CCTop, and detailed information is provided in the publication describing that tool (<xref ref-type="bibr" rid="B29">Stemmer et al., 2015</xref>) and summarized in <xref ref-type="table" rid="T1">Table 1</xref>. Beyond CCTop prerequisites, gRNA-SeqRET additionally asks for the desired number of gRNAs per target and a preference for targeted strand. These last two fields provide the criteria serving to limit the complete CCTop output to just those matching the user&#x2019;s requirements; i.e., the number of guides specified in the &#x201c;Number of guide RNAs per gene/custom feature&#x201d; input box. For example, entering &#x201c;3&#x201d; in this field will yield the top 3 gRNAs as determined by CCTop&#x2019;s predicted cutting score and RNA folding predictions for the protospacer and scaffold. Clicking &#x201c;submit&#x201d; executes the job, which can range from seconds for low numbers (e.g., &#x3c;10) of target loci to several minutes for genome-scale (i.e., all loci) in prokaryotic genomes, to several hours for genome-scale jobs in very large eukaryotic genomes (a submission comprising design of 3 guide RNAs for every coding region of the Araport11 annotation of the <italic>Arabidopsis thaliana</italic> genome (<xref ref-type="bibr" rid="B4">Cheng et al., 2017</xref>) ran &#x223c;12&#xa0;h). The &#x201c;View Your Designs&#x201d; page lists all completed and running jobs and provides a results link to the job output (<xref ref-type="fig" rid="F2">Figure 2</xref>). Clicking this link provides a summary of the job parameters as well as the output files for downloading. &#x201c;Results&#x201d; provides a comma separated values (CSV) file containing all guide sequences, providing a unique name (the locus name appended by the CCTop designation), the start and end chromosomal coordinates, strand orientation, sequence and specific PAM. &#x201c;Target regions&#x201d; provides a FASTA output of all targeted sequences, and the Report contains a list of loci where the desired number of gRNAs was not found, which can be used to run another design round with an altered sequence targeting criteria if desired. Clicking on &#x201c;All output files (.tar.gz)&#x201d; will download an archive of all files described above, as well as a file called &#x201c;scoring.log&#x201d; which contains all statistics and predicted cutting scores associated with each gRNA. Output files are maintained within the user&#x2019;s individual account for 3&#xa0;months. Additionally, users can make their designs public by clicking the &#x201c;Make Publicly Available&#x201d; option. Clicking this will open a new page where fields describing the purpose of the library and the target organism NCBI taxonomic ID can be completed, and the designs will be accessible to all users in the &#x201c;Public Designs&#x201d; link.</p>
</sec>
<sec id="s3-3">
<title>Applicability of a given gRNA library to other closely related genomes</title>
<p>Cost reductions and availability of high-variant oligonucleotide pools and high throughput sequencing has enabled the construction of gRNA libraries to become powerful and increasingly accessible approaches for rapid gene function interrogation (<xref ref-type="bibr" rid="B27">Schwartz et al., 2019</xref>; <xref ref-type="bibr" rid="B2">Bock et al., 2022</xref>; <xref ref-type="bibr" rid="B7">Cooper et al., 2022</xref>; <xref ref-type="bibr" rid="B28">Shi et al., 2022</xref>; <xref ref-type="bibr" rid="B30">Trivedi et al., 2023</xref>). Nevertheless, the library assembly and sequencing-based quality control (ensuring representation of all variants and possible skews in the population) can represent a significant endeavor and investment. Since a single constructed cloned library yields sufficient material for many tens of independent experiments, we wondered how applicable a given gRNA library would be to other closely related species. To address this, we first used gRNA-SeqRET to design a genome-scale gRNA library targeting all coding regions in the Actinomycetes <italic>Streptococcus lividans</italic> TK24. gRNA-SeqRET was employed to report 3 gRNAs per target region, resulting in 22,641 total spacer designs (<xref ref-type="sec" rid="s10">Supplementary Table S1</xref>). We turned to the Integrated Microbial Genomes and Microbiomes (IMG) (<xref ref-type="bibr" rid="B3">Chen et al., 2023</xref>) database and queried for all genomes characterized with an average nucleotide identity (ANI) &#x2265;80% relative to <italic>S. lividans</italic> TK24, yielding 192 genomes (<xref ref-type="sec" rid="s10">Supplementary Table S2</xref>). While an ANI is a statistic generated from the complete genome sequence, and we specifically searched for gRNA sequences in coding regions only, this approach enabled us to retrieve highly similar genome sequences. We then searched the genomes of each of these 192 microbes for perfect matches to each of the &#x223c;22 thousand 20mer sequences (plus PAM; <xref ref-type="sec" rid="s10">Supplementary Table S3</xref>), providing a sense as to how applicable a given gRNA library could be to alternate closely related organisms. Unsurprisingly, guide sequence alignment reduces dramatically with genome distance (<xref ref-type="fig" rid="F3">Figure 3</xref>). While the fraction of guide sequences with perfect alignment remains high (&#x223c;90%) for genomes with very high similarity (i.e., ANI &#x3e;98%), this drops to &#x223c;45% in genomes with and ANI of 95% (<xref ref-type="fig" rid="F3">Figure 3</xref>; <xref ref-type="sec" rid="s10">Supplementary Table S3</xref>). Nevertheless, since organisms belonging to the same species typically exhibit an ANI of 95%, this analysis demonstrates a pooled gRNA library designed to one organism could be of value to strains within a species, and perhaps to other closely related species (<xref ref-type="bibr" rid="B13">Konstantinidis and Tiedje, 2005</xref>), although this will depend on the specific species. Worth noting though is with increased taxonomic distance and genetic drift, off site targeting may also increase and confound the data. Additionally, we found that many of the 192 genomes contained several hundred occurrences of multiple matches per guide sequence, indicating a perfect sequence alignment in an off-target location (<xref ref-type="sec" rid="s10">Supplementary Table S3</xref>).</p>
<fig id="F3" position="float">
<label>FIGURE 3</label>
<caption>
<p>gRNA sequence applicability reduces rapidly with genome distance. Correlation between the percentage of 22,741 gRNAs designed to target <italic>S. lividans</italic> TK24 that align perfectly to 192 related organisms showing an average nucleotide identify (ANI) of &#x2265;80%.</p>
</caption>
<graphic xlink:href="fbioe-11-1217811-g003.tif"/>
</fig>
</sec>
</sec>
<sec id="s4">
<title>Future directions</title>
<p>During our internal use of the present version of gRNA-SeqRET, we have identified several areas we are considering for future enhancement. The first of these builds upon the possible re-use of guide RNAs in multiple organisms, by adding an option to provide multiple genomes so that the application can filter out possible gRNAs that occur in both genomes but without off-target events. Another is to enable a user to provide custom annotations allowing regions to be targeted or extracted in addition to those defined in the genome input file. Finally, we may develop an option to target specific alleles in polyploid genomes.</p>
</sec>
</body>
<back>
<sec sec-type="data-availability" id="s5">
<title>Data availability statement</title>
<p>The original contributions presented in the study are included in the article/<xref ref-type="sec" rid="s10">Supplementary Material</xref>, further inquiries can be directed to the corresponding author.</p>
</sec>
<sec id="s6">
<title>Author contributions</title>
<p>LS, J-FC, and IB conceived of the tool. LS designed and built the software and performed analysis. LS, J-FC, and IB tested and refined the software. IB wrote the manuscript with input and comments from J-FC and LS. All authors contributed to the article and approved the submitted version.</p>
</sec>
<sec id="s7">
<title>Funding</title>
<p>This work has been supported by the DOE Joint Genome Institute (<ext-link ext-link-type="uri" xlink:href="http://jgi.doe.gov">http://jgi.doe.gov</ext-link>) by the U.S. Department of Energy, Office of Science, Office of Biological and Environmental Research, through Contract DE-AC02-05CH11231 between Lawrence Berkeley National Laboratory and the U.S. Department of Energy.</p>
</sec>
<ack>
<p>The authors would like to thank Nicolas Grosjean for helpful feedback on the manuscript, Rekha Seshadri for technical assistance with IMG, as well as Ben Shen, Cameron Currie and Ashley Shade for access to their unpublished genomic data.</p>
</ack>
<sec sec-type="COI-statement" id="s8">
<title>Conflict of interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec sec-type="disclaimer" id="s9">
<title>Publisher&#x2019;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
<sec id="s10">
<title>Supplementary material</title>
<p>The Supplementary Material for this article can be found online at: <ext-link ext-link-type="uri" xlink:href="https://www.frontiersin.org/articles/10.3389/fbioe.2023.1217811/full#supplementary-material">https://www.frontiersin.org/articles/10.3389/fbioe.2023.1217811/full&#x23;supplementary-material</ext-link>
</p>
<supplementary-material xlink:href="Table2.XLSX" id="SM1" mimetype="application/XLSX" xmlns:xlink="http://www.w3.org/1999/xlink"/>
<supplementary-material xlink:href="Table3.XLSX" id="SM2" mimetype="application/XLSX" xmlns:xlink="http://www.w3.org/1999/xlink"/>
<supplementary-material xlink:href="Table1.XLSX" id="SM3" mimetype="application/XLSX" xmlns:xlink="http://www.w3.org/1999/xlink"/>
</sec>
<ref-list>
<title>References</title>
<ref id="B1">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Alipanahi</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Safari</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Khanteymoori</surname>
<given-names>A.</given-names>
</name>
</person-group> (<year>2022</year>). <article-title>CRISPR genome editing using computational approaches: a survey</article-title>. <source>Front. Bioinforma.</source> <volume>2</volume>, <fpage>1001131</fpage>. <pub-id pub-id-type="doi">10.3389/fbinf.2022.1001131</pub-id>
</citation>
</ref>
<ref id="B2">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Bock</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Datlinger</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Chardon</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Coelho</surname>
<given-names>M. A.</given-names>
</name>
<name>
<surname>Dong</surname>
<given-names>M. B.</given-names>
</name>
<name>
<surname>Lawson</surname>
<given-names>K. A.</given-names>
</name>
<etal/>
</person-group> (<year>2022</year>). <article-title>High-content CRISPR screening</article-title>. <source>Nat. Rev. Methods Prim.</source> <volume>2</volume> (<issue>1</issue>), <fpage>8</fpage>. <pub-id pub-id-type="doi">10.1038/s43586-021-00093-4</pub-id>
</citation>
</ref>
<ref id="B3">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Chen</surname>
<given-names>I.-M. A.</given-names>
</name>
<name>
<surname>Chu</surname>
<given-names>K.</given-names>
</name>
<name>
<surname>Palaniappan</surname>
<given-names>K.</given-names>
</name>
<name>
<surname>Ratner</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Huang</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Huntemann</surname>
<given-names>M.</given-names>
</name>
<etal/>
</person-group> (<year>2023</year>). <article-title>The IMG/M data management and analysis system v. 7: content updates and new features</article-title>. <source>Nucleic Acids Res.</source> <volume>51</volume> (<issue>1</issue>), <fpage>D723</fpage>&#x2013;<lpage>D732</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gkac976</pub-id>
</citation>
</ref>
<ref id="B4">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Cheng</surname>
<given-names>C. Y.</given-names>
</name>
<name>
<surname>Krishnakumar</surname>
<given-names>V.</given-names>
</name>
<name>
<surname>Chan</surname>
<given-names>A. P.</given-names>
</name>
<name>
<surname>Thibaud-Nissen</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Schobel</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Town</surname>
<given-names>C. D.</given-names>
</name>
</person-group> (<year>2017</year>). <article-title>Araport11: a complete reannotation of the <italic>Arabidopsis thaliana</italic> reference genome</article-title>. <source>Plant J.</source> <volume>89</volume> (<issue>4</issue>), <fpage>789</fpage>&#x2013;<lpage>804</lpage>. <pub-id pub-id-type="doi">10.1111/tpj.13415</pub-id>
</citation>
</ref>
<ref id="B5">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Cock</surname>
<given-names>P. J.</given-names>
</name>
<name>
<surname>Antao</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Chang</surname>
<given-names>J. T.</given-names>
</name>
<name>
<surname>Chapman</surname>
<given-names>B. A.</given-names>
</name>
<name>
<surname>Cox</surname>
<given-names>C. J.</given-names>
</name>
<name>
<surname>Dalke</surname>
<given-names>A.</given-names>
</name>
<etal/>
</person-group> (<year>2009</year>). <article-title>Biopython: freely available Python tools for computational molecular biology and bioinformatics</article-title>. <source>Bioinformatics</source> <volume>25</volume> (<issue>11</issue>), <fpage>1422</fpage>&#x2013;<lpage>1423</lpage>. <pub-id pub-id-type="doi">10.1093/bioinformatics/btp163</pub-id>
</citation>
</ref>
<ref id="B6">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Concordet</surname>
<given-names>J.-P.</given-names>
</name>
<name>
<surname>Haeussler</surname>
<given-names>M.</given-names>
</name>
</person-group> (<year>2018</year>). <article-title>CRISPOR: intuitive guide selection for CRISPR/Cas9 genome editing experiments and screens</article-title>. <source>Nucleic acids Res.</source> <volume>46</volume> (<issue>1</issue>), <fpage>W242</fpage>&#x2013;<lpage>W245</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gky354</pub-id>
</citation>
</ref>
<ref id="B7">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Cooper</surname>
<given-names>Y. A.</given-names>
</name>
<name>
<surname>Guo</surname>
<given-names>Q.</given-names>
</name>
<name>
<surname>Geschwind</surname>
<given-names>D. H.</given-names>
</name>
</person-group> (<year>2022</year>). <article-title>Multiplexed functional genomic assays to decipher the noncoding genome</article-title>. <source>Hum. Mol. Genet.</source> <volume>31</volume> (<issue>1</issue>), <fpage>R84</fpage>&#x2013;<lpage>R96</lpage>. <pub-id pub-id-type="doi">10.1093/hmg/ddac194</pub-id>
</citation>
</ref>
<ref id="B8">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Dixit</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Parnas</surname>
<given-names>O.</given-names>
</name>
<name>
<surname>Li</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Chen</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Fulco</surname>
<given-names>C. P.</given-names>
</name>
<name>
<surname>Jerby-Arnon</surname>
<given-names>L.</given-names>
</name>
<etal/>
</person-group> (<year>2016</year>). <article-title>Perturb-seq: dissecting molecular circuits with scalable single-cell RNA profiling of pooled genetic screens</article-title>. <source>Cell</source> <volume>167</volume> (<issue>7</issue>), <fpage>1853</fpage>&#x2013;<lpage>1866. e17</lpage>. <pub-id pub-id-type="doi">10.1016/j.cell.2016.11.038</pub-id>
</citation>
</ref>
<ref id="B9">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Doench</surname>
<given-names>J. G.</given-names>
</name>
<name>
<surname>Fusi</surname>
<given-names>N.</given-names>
</name>
<name>
<surname>Sullender</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Hegde</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Vaimberg</surname>
<given-names>E. W.</given-names>
</name>
<name>
<surname>Donovan</surname>
<given-names>K. F.</given-names>
</name>
<etal/>
</person-group> (<year>2016</year>). <article-title>Optimized sgRNA design to maximize activity and minimize off-target effects of CRISPR-Cas9</article-title>. <source>Nat. Biotechnol.</source> <volume>34</volume> (<issue>2</issue>), <fpage>184</fpage>&#x2013;<lpage>191</lpage>. <pub-id pub-id-type="doi">10.1038/nbt.3437</pub-id>
</citation>
</ref>
<ref id="B10">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>He</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Liu</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Chen</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Xie</surname>
<given-names>W. Z.</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Li</surname>
<given-names>Y.</given-names>
</name>
<etal/>
</person-group> (<year>2021</year>). <article-title>CRISPR-cereal: a guide RNA design tool integrating regulome and genomic variation for wheat, maize and rice</article-title>. <source>Plant Biotechnol. J.</source> <volume>19</volume> (<issue>11</issue>), <fpage>2141</fpage>&#x2013;<lpage>2143</lpage>. <pub-id pub-id-type="doi">10.1111/pbi.13675</pub-id>
</citation>
</ref>
<ref id="B11">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Jiang</surname>
<given-names>W.</given-names>
</name>
<name>
<surname>Bikard</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Cox</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Zhang</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Marraffini</surname>
<given-names>L. A.</given-names>
</name>
</person-group> (<year>2013</year>). <article-title>RNA-guided editing of bacterial genomes using CRISPR-Cas systems</article-title>. <source>Nat. Biotechnol.</source> <volume>31</volume> (<issue>3</issue>), <fpage>233</fpage>&#x2013;<lpage>239</lpage>. <pub-id pub-id-type="doi">10.1038/nbt.2508</pub-id>
</citation>
</ref>
<ref id="B12">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Jinek</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Chylinski</surname>
<given-names>K.</given-names>
</name>
<name>
<surname>Fonfara</surname>
<given-names>I.</given-names>
</name>
<name>
<surname>Hauer</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Doudna</surname>
<given-names>J. A.</given-names>
</name>
<name>
<surname>Charpentier</surname>
<given-names>E.</given-names>
</name>
</person-group> (<year>2012</year>). <article-title>A programmable dual-RNA&#x2013;guided DNA endonuclease in adaptive bacterial immunity</article-title>. <source>science</source> <volume>337</volume> (<issue>6096</issue>), <fpage>816</fpage>&#x2013;<lpage>821</lpage>. <pub-id pub-id-type="doi">10.1126/science.1225829</pub-id>
</citation>
</ref>
<ref id="B13">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Konstantinidis</surname>
<given-names>K. T.</given-names>
</name>
<name>
<surname>Tiedje</surname>
<given-names>J. M.</given-names>
</name>
</person-group> (<year>2005</year>). <article-title>Genomic insights that advance the species definition for prokaryotes</article-title>. <source>Proc. Natl. Acad. Sci.</source> <volume>102</volume> (<issue>7</issue>), <fpage>2567</fpage>&#x2013;<lpage>2572</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.0409727102</pub-id>
</citation>
</ref>
<ref id="B14">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Labun</surname>
<given-names>K.</given-names>
</name>
<name>
<surname>Montague</surname>
<given-names>T. G.</given-names>
</name>
<name>
<surname>Krause</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Torres Cleuren</surname>
<given-names>Y. N.</given-names>
</name>
<name>
<surname>Tjeldnes</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Valen</surname>
<given-names>E.</given-names>
</name>
</person-group> (<year>2019</year>). <article-title>CHOPCHOP v3: expanding the CRISPR web toolbox beyond genome editing</article-title>. <source>Nucleic acids Res.</source> <volume>47</volume> (<issue>1</issue>), <fpage>W171</fpage>&#x2013;<lpage>W174</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gkz365</pub-id>
</citation>
</ref>
<ref id="B15">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Langmead</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Trapnell</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Pop</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Salzberg</surname>
<given-names>S. L.</given-names>
</name>
</person-group> (<year>2009</year>). <article-title>Ultrafast and memory-efficient alignment of short DNA sequences to the human genome</article-title>. <source>Genome Biol.</source> <volume>10</volume> (<issue>3</issue>), <fpage>R25</fpage>&#x2013;<lpage>R10</lpage>. <pub-id pub-id-type="doi">10.1186/gb-2009-10-3-r25</pub-id>
</citation>
</ref>
<ref id="B16">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Liu</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Lin</surname>
<given-names>Q.</given-names>
</name>
<name>
<surname>Jin</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Gao</surname>
<given-names>C.</given-names>
</name>
</person-group> (<year>2021</year>). <article-title>The CRISPR-Cas toolbox and gene editing technologies</article-title>. <source>Mol. Cell</source> <volume>82</volume>, <fpage>333</fpage>&#x2013;<lpage>347</lpage>. <pub-id pub-id-type="doi">10.1016/j.molcel.2021.12.002</pub-id>
</citation>
</ref>
<ref id="B17">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Liu</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Wan</surname>
<given-names>X.</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>B.</given-names>
</name>
</person-group> (<year>2019</year>). <article-title>Engineered CRISPRa enables programmable eukaryote-like gene activation in bacteria</article-title>. <source>Nat. Commun.</source> <volume>10</volume> (<issue>1</issue>), <fpage>3693</fpage>. <pub-id pub-id-type="doi">10.1038/s41467-019-11479-0</pub-id>
</citation>
</ref>
<ref id="B18">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Lorenz</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Bernhart</surname>
<given-names>S. H.</given-names>
</name>
<name>
<surname>H&#xf6;ner zu Siederdissen</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Tafer</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Flamm</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Stadler</surname>
<given-names>P. F.</given-names>
</name>
<etal/>
</person-group> (<year>2011</year>). <article-title>ViennaRNA package 2.0</article-title>. <source>Algorithms Mol. Biol.</source> <volume>6</volume>, <fpage>26</fpage>&#x2013;<lpage>14</lpage>. <pub-id pub-id-type="doi">10.1186/1748-7188-6-26</pub-id>
</citation>
</ref>
<ref id="B19">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Mojica</surname>
<given-names>F. J.</given-names>
</name>
<name>
<surname>D&#xed;ez-Villase&#xf1;or</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Garc&#xed;a-Mart&#xed;nez</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Almendros</surname>
<given-names>C.</given-names>
</name>
</person-group> (<year>2009</year>). <article-title>Short motif sequences determine the targets of the prokaryotic CRISPR defence system</article-title>. <source>Microbiology</source> <volume>155</volume> (<issue>3</issue>), <fpage>733</fpage>&#x2013;<lpage>740</lpage>. <pub-id pub-id-type="doi">10.1099/mic.0.023960-0</pub-id>
</citation>
</ref>
<ref id="B20">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Naito</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Hino</surname>
<given-names>K.</given-names>
</name>
<name>
<surname>Bono</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Ui-Tei</surname>
<given-names>K.</given-names>
</name>
</person-group> (<year>2015</year>). <article-title>CRISPRdirect: software for designing CRISPR/Cas guide RNA with reduced off-target sites</article-title>. <source>Bioinformatics</source> <volume>31</volume> (<issue>7</issue>), <fpage>1120</fpage>&#x2013;<lpage>1123</lpage>. <pub-id pub-id-type="doi">10.1093/bioinformatics/btu743</pub-id>
</citation>
</ref>
<ref id="B21">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Nakamura</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Gao</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Dominguez</surname>
<given-names>A. A.</given-names>
</name>
<name>
<surname>Qi</surname>
<given-names>L. S.</given-names>
</name>
</person-group> (<year>2021</year>). <article-title>CRISPR technologies for precise epigenome editing</article-title>. <source>Nat. Cell Biol.</source> <volume>23</volume> (<issue>1</issue>), <fpage>11</fpage>&#x2013;<lpage>22</lpage>. <pub-id pub-id-type="doi">10.1038/s41556-020-00620-7</pub-id>
</citation>
</ref>
<ref id="B22">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Peng</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Tarleton</surname>
<given-names>R.</given-names>
</name>
</person-group> (<year>2015</year>). <article-title>EuPaGDT: a web tool tailored to design CRISPR guide RNAs for eukaryotic pathogens</article-title>. <source>Microb. genomics</source> <volume>1</volume> (<issue>4</issue>), <fpage>e000033</fpage>. <pub-id pub-id-type="doi">10.1099/mgen.0.000033</pub-id>
</citation>
</ref>
<ref id="B23">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Pickar-Oliver</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Gersbach</surname>
<given-names>C. A.</given-names>
</name>
</person-group> (<year>2019</year>). <article-title>The next generation of CRISPR&#x2013;Cas technologies and applications</article-title>. <source>Nat. Rev. Mol. Cell Biol.</source> <volume>20</volume> (<issue>8</issue>), <fpage>490</fpage>&#x2013;<lpage>507</lpage>. <pub-id pub-id-type="doi">10.1038/s41580-019-0131-5</pub-id>
</citation>
</ref>
<ref id="B24">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Poudel</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Rodriguez</surname>
<given-names>L. T.</given-names>
</name>
<name>
<surname>Reisch</surname>
<given-names>C. R.</given-names>
</name>
<name>
<surname>Rivers</surname>
<given-names>A. R.</given-names>
</name>
</person-group> (<year>2022</year>). <article-title>GuideMaker: software to design CRISPR-cas guide RNA pools in non-model genomes</article-title>. <source>GigaScience</source> <volume>11</volume>, <fpage>giac007</fpage>. <pub-id pub-id-type="doi">10.1093/gigascience/giac007</pub-id>
</citation>
</ref>
<ref id="B25">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Qi</surname>
<given-names>L. S.</given-names>
</name>
<name>
<surname>Larson</surname>
<given-names>M. H.</given-names>
</name>
<name>
<surname>Gilbert</surname>
<given-names>L. A.</given-names>
</name>
<name>
<surname>Doudna</surname>
<given-names>J. A.</given-names>
</name>
<name>
<surname>Weissman</surname>
<given-names>J. S.</given-names>
</name>
<name>
<surname>Arkin</surname>
<given-names>A. P.</given-names>
</name>
<etal/>
</person-group> (<year>2013</year>). <article-title>Repurposing CRISPR as an RNA-guided platform for sequence-specific control of gene expression</article-title>. <source>Cell</source> <volume>152</volume> (<issue>5</issue>), <fpage>1173</fpage>&#x2013;<lpage>1183</lpage>. <pub-id pub-id-type="doi">10.1016/j.cell.2013.02.022</pub-id>
</citation>
</ref>
<ref id="B26">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Ran</surname>
<given-names>F. A.</given-names>
</name>
<name>
<surname>Hsu</surname>
<given-names>P. D.</given-names>
</name>
<name>
<surname>Wright</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Agarwala</surname>
<given-names>V.</given-names>
</name>
<name>
<surname>Scott</surname>
<given-names>D. A.</given-names>
</name>
<name>
<surname>Zhang</surname>
<given-names>F.</given-names>
</name>
</person-group> (<year>2013</year>). <article-title>Genome engineering using the CRISPR-Cas9 system</article-title>. <source>Nat. Protoc.</source> <volume>8</volume> (<issue>11</issue>), <fpage>2281</fpage>&#x2013;<lpage>2308</lpage>. <pub-id pub-id-type="doi">10.1038/nprot.2013.143</pub-id>
</citation>
</ref>
<ref id="B27">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Schwartz</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Cheng</surname>
<given-names>J. F.</given-names>
</name>
<name>
<surname>Evans</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Schwartz</surname>
<given-names>C. A.</given-names>
</name>
<name>
<surname>Wagner</surname>
<given-names>J. M.</given-names>
</name>
<name>
<surname>Anglin</surname>
<given-names>S.</given-names>
</name>
<etal/>
</person-group> (<year>2019</year>). <article-title>Validating genome-wide CRISPR-Cas9 function improves screening in the oleaginous yeast Yarrowia lipolytica</article-title>. <source>Metab. Eng.</source> <volume>55</volume>, <fpage>102</fpage>&#x2013;<lpage>110</lpage>. <pub-id pub-id-type="doi">10.1016/j.ymben.2019.06.007</pub-id>
</citation>
</ref>
<ref id="B28">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Shi</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Doench</surname>
<given-names>J. G.</given-names>
</name>
<name>
<surname>Chi</surname>
<given-names>H.</given-names>
</name>
</person-group> (<year>2022</year>). <article-title>CRISPR screens for functional interrogation of immunity</article-title>. <source>Nat. Rev. Immunol.</source> <volume>23</volume>, <fpage>363</fpage>&#x2013;<lpage>380</lpage>. <pub-id pub-id-type="doi">10.1038/s41577-022-00802-4</pub-id>
</citation>
</ref>
<ref id="B29">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Stemmer</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Thumberger</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Del Sol Keyer</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Wittbrodt</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Mateo</surname>
<given-names>J. L.</given-names>
</name>
</person-group> (<year>2015</year>). <article-title>CCTop: an intuitive, flexible and reliable CRISPR/Cas9 target prediction tool</article-title>. <source>PLoS One</source> <volume>10</volume> (<issue>4</issue>), <fpage>e0124633</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pone.0124633</pub-id>
</citation>
</ref>
<ref id="B30">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Trivedi</surname>
<given-names>V.</given-names>
</name>
<name>
<surname>Ramesh</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Wheeldon</surname>
<given-names>I.</given-names>
</name>
</person-group> (<year>2023</year>). <article-title>Analyzing CRISPR screens in non-conventional microbes</article-title>. <source>J. Industrial Microbiol. Biotechnol.</source> <volume>50</volume>, <fpage>kuad006</fpage>. <pub-id pub-id-type="doi">10.1093/jimb/kuad006</pub-id>
</citation>
</ref>
<ref id="B31">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Wang</surname>
<given-names>J. Y.</given-names>
</name>
<name>
<surname>Doudna</surname>
<given-names>J. A.</given-names>
</name>
</person-group> (<year>2023</year>). <article-title>CRISPR technology: a decade of genome editing is only the beginning</article-title>. <source>Science</source> <volume>379</volume> (<issue>6629</issue>), <fpage>eadd8643</fpage>. <pub-id pub-id-type="doi">10.1126/science.add8643</pub-id>
</citation>
</ref>
<ref id="B32">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Wiedenheft</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Sternberg</surname>
<given-names>S. H.</given-names>
</name>
<name>
<surname>Doudna</surname>
<given-names>J. A.</given-names>
</name>
</person-group> (<year>2012</year>). <article-title>RNA-guided genetic silencing systems in bacteria and archaea</article-title>. <source>Nature</source> <volume>482</volume> (<issue>7385</issue>), <fpage>331</fpage>&#x2013;<lpage>338</lpage>. <pub-id pub-id-type="doi">10.1038/nature10886</pub-id>
</citation>
</ref>
<ref id="B33">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Wilson</surname>
<given-names>L. O.</given-names>
</name>
<name>
<surname>O&#x2019;Brien</surname>
<given-names>A. R.</given-names>
</name>
<name>
<surname>Bauer</surname>
<given-names>D. C.</given-names>
</name>
</person-group> (<year>2018</year>). <article-title>The current state and future of CRISPR-Cas9 gRNA design tools</article-title>. <source>Front. Pharmacol.</source> <volume>9</volume>, <fpage>749</fpage>. <pub-id pub-id-type="doi">10.3389/fphar.2018.00749</pub-id>
</citation>
</ref>
<ref id="B34">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Yamano</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Nishimasu</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Zetsche</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Hirano</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Slaymaker</surname>
<given-names>I. M.</given-names>
</name>
<name>
<surname>Li</surname>
<given-names>Y.</given-names>
</name>
<etal/>
</person-group> (<year>2016</year>). <article-title>Crystal structure of Cpf1 in complex with guide RNA and target DNA</article-title>. <source>Cell</source> <volume>165</volume> (<issue>4</issue>), <fpage>949</fpage>&#x2013;<lpage>962</lpage>. <pub-id pub-id-type="doi">10.1016/j.cell.2016.04.003</pub-id>
</citation>
</ref>
<ref id="B35">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Zetsche</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Gootenberg</surname>
<given-names>J. S.</given-names>
</name>
<name>
<surname>Abudayyeh</surname>
<given-names>O. O.</given-names>
</name>
<name>
<surname>Slaymaker</surname>
<given-names>I. M.</given-names>
</name>
<name>
<surname>Makarova</surname>
<given-names>K. S.</given-names>
</name>
<name>
<surname>Essletzbichler</surname>
<given-names>P.</given-names>
</name>
<etal/>
</person-group> (<year>2015</year>). <article-title>Cpf1 is a single RNA-guided endonuclease of a class 2 CRISPR-Cas system</article-title>. <source>Cell</source> <volume>163</volume> (<issue>3</issue>), <fpage>759</fpage>&#x2013;<lpage>771</lpage>. <pub-id pub-id-type="doi">10.1016/j.cell.2015.09.038</pub-id>
</citation>
</ref>
</ref-list>
</back>
</article>