<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article article-type="brief-report" dtd-version="2.3" xml:lang="EN" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Bioinform.</journal-id>
<journal-title>Frontiers in Bioinformatics</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Bioinform.</abbrev-journal-title>
<issn pub-type="epub">2673-7647</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="publisher-id">793309</article-id>
<article-id pub-id-type="doi">10.3389/fbinf.2022.793309</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Bioinformatics</subject>
<subj-group>
<subject>Brief Research Report</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>SingleCAnalyzer: Interactive Analysis of Single Cell RNA-Seq Data on the Cloud</article-title>
<alt-title alt-title-type="left-running-head">Prieto et al.</alt-title>
<alt-title alt-title-type="right-running-head">SingleCAnalyzer: scRNA-Seq Interactive Analysis</alt-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name>
<surname>Prieto</surname>
<given-names>Carlos</given-names>
</name>
<xref ref-type="corresp" rid="c001">&#x2a;</xref>
<uri xlink:href="https://loop.frontiersin.org/people/67483/overview"/>
</contrib>
<contrib contrib-type="author">
<name>
<surname>Barrios</surname>
<given-names>David</given-names>
</name>
</contrib>
<contrib contrib-type="author">
<name>
<surname>Villaverde</surname>
<given-names>Angela</given-names>
</name>
<uri xlink:href="https://loop.frontiersin.org/people/1663747/overview"/>
</contrib>
</contrib-group>
<aff>
<institution>Bioinformatics Service, Nucleus</institution>, <institution>University of Salamanca</institution>, <addr-line>Salamanca</addr-line>, <country>Spain</country>
</aff>
<author-notes>
<fn fn-type="edited-by">
<p>
<bold>Edited by:</bold> <ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/52111/overview">Yann Ponty</ext-link>, &#xc9;cole Polytechnique, France</p>
</fn>
<fn fn-type="edited-by">
<p>
<bold>Reviewed by:</bold> <ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/442768/overview">Lionel Spinelli</ext-link>, Aix-Marseille Universit&#xe9;, France</p>
<p>
<ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/52531/overview">Sebastian Will</ext-link>, &#xc9;cole Polytechnique, France</p>
</fn>
<corresp id="c001">&#x2a;Correspondence: Carlos Prieto, <email>cprietos@usal.es</email>
</corresp>
<fn fn-type="other">
<p>This article was submitted to Data Visualization, a section of the journal Frontiers in Bioinformatics</p>
</fn>
</author-notes>
<pub-date pub-type="epub">
<day>23</day>
<month>05</month>
<year>2022</year>
</pub-date>
<pub-date pub-type="collection">
<year>2022</year>
</pub-date>
<volume>2</volume>
<elocation-id>793309</elocation-id>
<history>
<date date-type="received">
<day>11</day>
<month>10</month>
<year>2021</year>
</date>
<date date-type="accepted">
<day>09</day>
<month>05</month>
<year>2022</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#xa9; 2022 Prieto, Barrios and Villaverde.</copyright-statement>
<copyright-year>2022</copyright-year>
<copyright-holder>Prieto, Barrios and Villaverde</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/">
<p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p>
</license>
</permissions>
<abstract>
<p>Single-cell RNA sequencing (scRNA-Seq) enables researchers to quantify the transcriptomes of individual cells. The capacity of researchers to perform this type of analysis has allowed researchers to undertake new scientific goals. The usefulness of scRNA-Seq has depended on the development of new computational biology methods, which have been designed to meeting challenges associated with scRNA-Seq analysis. However, the proper application of these computational methods requires extensive bioinformatics expertise. Otherwise, it is often difficult to obtain reliable and reproducible results. We have developed SingleCAnalyzer, a cloud platform that provides a means to perform full scRNA-Seq analysis from FASTQ within an easy-to-use and self-exploratory web interface. Its analysis pipeline includes the demultiplexing and alignment of FASTQ files, read trimming, sample quality control, feature selection, empty droplets detection, dimensional reduction, cellular type prediction, unsupervised clustering of cells, pseudotime/trajectory analysis, expression comparisons between groups, functional enrichment of differentially expressed genes and gene set expression analysis. Results are presented with interactive graphs, which provide exploratory and analytical features. SingleCAnalyzer is freely available at <ext-link ext-link-type="uri" xlink:href="https://singlecanalyzer.eu/">https://singleCAnalyzer.eu</ext-link>.</p>
</abstract>
<kwd-group>
<kwd>ScRNA-seq</kwd>
<kwd>data visualization</kwd>
<kwd>single cell</kwd>
<kwd>web server</kwd>
<kwd>data analysis</kwd>
</kwd-group>
</article-meta>
</front>
<body>
<sec id="s1">
<title>Introduction</title>
<p>Single-cell RNA sequencing (scRNA-seq) has allowed for the quantification of RNA transcripts within individual cells. These assays allow researchers to explore cell-to-cell variability and meet new scientific goals. In the last few years, scRNA-seq has been applied, for example, to differentiate tumor cells from healthy ones, deconvolute immune cells, describe states of cell differentiation and development, and to identify rare populations of cells that cause disease (<xref ref-type="bibr" rid="B9">Haque et al., 2017</xref>). Although experimental scRNA-seq assays are becoming increasingly user-friendly, the analysis of sequencing data is complex. Data analysis requires the application of complex computational pipelines and data analysis methods that require bioinformatics expertise (<xref ref-type="bibr" rid="B10">Hwang et al., 2018</xref>). The interpretation of scRNA-seq results is strongly influenced by its analysis pipeline, and the incorrect application of methods could lead to conclusions that are incorrect. Since data analysis is complex and very important for correctly interpreting results, the development of analysis tools that produce reliable results and minimize the possibility of error is essential for enhancing the usefulness of scRNA-seq data.</p>
<p>Throughout the last 5&#xa0;years, some software development projects have aimed to address the absence of software available for the analysis of scRNA-seq data (<xref ref-type="bibr" rid="B8">Guo et al., 2015</xref>; <xref ref-type="bibr" rid="B6">Gardeux et al., 2017</xref>; <xref ref-type="bibr" rid="B12">Kiselev et al., 2017</xref>; <xref ref-type="bibr" rid="B15">Lin et al., 2017</xref>; <xref ref-type="bibr" rid="B21">Perraudeau et al., 2017</xref>; <xref ref-type="bibr" rid="B37">Zhu et al., 2017</xref>; <xref ref-type="bibr" rid="B25">Scholz et al., 2018</xref>; <xref ref-type="bibr" rid="B33">Wagner and Yanai, 2018</xref>; <xref ref-type="bibr" rid="B4">Chen et al., 2019</xref>; <xref ref-type="bibr" rid="B18">Monier et al., 2019</xref>; <xref ref-type="bibr" rid="B31">Stuart et al., 2019</xref>). Designers of the projects have developed analysis pipelines that can be executed with R or Python function calls or with websites. Although the platforms have tremendous utility, they do possess some usability and functionality limitations that should be solved. For example, none of the applications are capable of analysing raw sequencing files (FASTQ), they do not allow for the interactive selection of groups and a few provide an integrated functional analysis of results (see <xref ref-type="sec" rid="s10">Supplementary Table S1</xref>).</p>
<p>We have developed SingleCAnalyzer to provide a Web application server that performs a fully interactive and comprehensive analysis of scRNA-Seq data with two simple steps. It provides an integrated and interactive platform which is able to process sequencing files (FASTQ) and perform full scRNA-seq analyses and the functional analysis of results. It was implemented as a cloud analysis platform that can be executed without installing any software. SingleCAnalyzer facilitates the analysis of scRNA-seq data to non-experienced users and provides quick exploratory analyses to computational biologists.</p>
</sec>
<sec sec-type="results" id="s2">
<title>Results</title>
<sec id="s2-1">
<title>The SingleCAnalyzer Website</title>
<p>The front-end of SingleCAnalyzer has been designed to provide a means to fully analyse scRNA-Seq data using the following two steps: 1) Setting input files and analysis parameters and 2) cluster determination and the execution of comparative analysis. In the first step, FASTQ/HDF5 files are uploaded or an ENA project identifier is provided by the user. Basic information regarding the species studied and type of sequencing performed, as well as optional parameters for the alignments of sequences can also be specified on the web. Once the files are uploaded, demultiplexed and aligned, users may perform further analysis including feature selection, empty droplet deletion, dimensional reduction, prediction of cellular type, analysis of trajectories/pseudotime and unsupervised clustering. These analyses can be performed and adjusted by selecting parameters in the &#x2018;analysis parameters&#x2019; section (<xref ref-type="fig" rid="F1">Figure 1</xref>).</p>
<fig id="F1" position="float">
<label>FIGURE 1</label>
<caption>
<p>SingleCAnalyzer Workflow. Schematic representation showing SingleCAnalyzer workflow example. Panel 1 and 2 show the web interface for setting input files and parameters. Panel 3 shows an example of visualization of dimensionality reduction, cellular type classification, pseudotime analysis and clustering results. Panels 4 and 5 show selected sections of differential expression and functional analysis results.</p>
</caption>
<graphic xlink:href="fbinf-02-793309-g001.tif"/>
</fig>
<p>Cluster determination and the execution of comparative analysis is accomplished through the website, which provides an interactive interface that allows the user to visualise cellular type prediction, pseudotime predictions or clustering results via six interconnected scatterplots generated using each dimensionality reduction technique. Point colour and type can be changed according to each analysis results. Users can also generate new representations of gene pairs and colour the points based on gene expression values. This interface specifies the most adequate aggrupation, cellular classification or time frame and is guided by the user&#x2019;s knowledge regarding the samples studied. On the interface, the user can also launch a comparative analysis of all groups, or manually determine which groups should be compared. The comparative analysis includes an analysis of differentially expressed groups of genes, and the functional analysis of gene ontology categories and pathways.</p>
<p>Results are displayed in tabular form, which reveal the execution status of each computational process and provide a link to final results. These are provided as static reports and interactive web pages. Results regarding the quantification of gene expression values are provided with a table of quantification statistics and downloadable files that contain information for aligned reads regarding the number of reads generated per transcript and the number of transcripts per million (TPM). The quality control page descriptively reveals the distribution patterns of expression using box plots, reveals estimated numbers of expressed genes using a bar plot and represents the first two components of a PCA analysis. The clustering results page integrates dimensionality reduction, clustering, pseudotime and cellular classification results within self-explanatory interface which can also generate static reports that incorporate user modifications and launch comparisons between groups. Reports containing results are generated for each comparison, which include differential expression, functional enrichment and GSEA analysis. Differential expression results are summarised in a table which is linked to the following means to visualise data: MA plot, volcano plot, box plot, line chart and heatmap. Functional analyses are also summarised in tables and interactive visual means to represent data such as bar plots, networks and symmetric heatmaps are provided.</p>
<p>
<xref ref-type="sec" rid="s10">Supplementary Table S1</xref> shows a comparison with 12 scRNA-Seq analysis platforms. The main features of the SingleCAnalyzer website are:<list list-type="simple">
<list-item>
<p>- scRNA-Seq analysis from raw FASTQ, HDF5 files or ENA project identifications</p>
</list-item>
<list-item>
<p>- Fully functional cloud platform that does not require the installation of software</p>
</list-item>
<list-item>
<p>- Semiautomated analysis which avoids the need for configuration using complex parameters</p>
</list-item>
<list-item>
<p>- User guided classification of cells within groups that is guided by interconnected graphs that integrate dimensional reduction, cellular type prediction, trajectory analysis and unsupervised clustering results</p>
</list-item>
<list-item>
<p>- Performs FASTQ processing, gene filtering, empty droplets detection, gene quantification, dimensionality reduction, unsupervised clustering, differential expression, functional overrepresentation and gene set expression analyses</p>
</list-item>
<list-item>
<p>- Straightforward presentation of results using interactive visual representations of data and provides a means to generate reports that are publication ready</p>
</list-item>
</list>
</p>
</sec>
<sec id="s2-2">
<title>Analysis Pipeline</title>
<p>
<xref ref-type="fig" rid="F2">Figure 2</xref> shows the analysis pipeline of SingleCAnalyzer. It integrates generally accepted tools used for the analysis of RNA-Seq data, which also perform well as computational resources. <xref ref-type="sec" rid="s10">Supplementary Table S2</xref> shows the computational time required to analyse nine scRNA-Seq public data sets. The complete analysis of 154 demultiplexed samples takes an average of 36 min, which allows for the real time analysis of low cell number scRNA-Seq experiments. The most time-consuming processes in the pipeline involves the upload, demultiplexing and alignment of samples, which are tasks that are performed in parallel. This parallelisation reduces the global analysis time by 57%, which makes the time requirement of our cloud infrastructure equal to virtual machine or local pipelines solutions. Moreover, SingleCAnalyzer does not store raw sequences or aligned files in order to avoid user disk space limitations, and the number of analysed samples of non-commercial cloud platforms.</p>
<fig id="F2" position="float">
<label>FIGURE 2</label>
<caption>
<p>SingleCAnalyzer Pipeline. Chart of SingleCAnalyzer pipeline with computational processes and output results.</p>
</caption>
<graphic xlink:href="fbinf-02-793309-g002.tif"/>
</fig>
<p>The next steps of the pipeline include feature selection, empty droplets detection, dimensionality reduction, cellular type prediction, trajectory/pseudotime analysis and unsupervised clustering. SingleCAnalyzer applies gene filtering, which is based on user input parameters to avoid non-informative output, noise or drop out events. Afterward, six dimensionality reduction methods are applied to the data and samples are visualised using interactive scatter plots. Simultaneously, four unsupervised clustering algorithms are applied to produce nine possible clustering divisions for each method, a cellular type prediction method is executed for each training dataset, and a pseudotime analysis is performed (see methods). These cluster types can be mapped on interactive plots at the request of the user.</p>
<p>Based on the unsupervised or manually curated clusters produced, users can identify gene characteristics and the functions of each group by launching comparison analysis. This feature incorporates the differential expression analysis of groups and the functional analysis of gene ontologies and pathways. The analysis pipeline also processes quality control, clustering, differential expression and functional analysis results, and integrates them in an interactive and self-explanatory web interface.</p>
<p>SingleCAnalyzer was conceived as an agile project, and new scRNA-Seq analysis methods can be integrated within its analysis pipeline. Only generally accepted methods that have been demonstrated to generate reliable and reproducible results that require reasonable quantities of computational resources will be considered for addition to our cloud platform. The increasing development of computational methods will inspire the adaptation of the platform to meet the needs of researchers as scientific trends regarding scRNA-SEQ data analysis emerge.</p>
</sec>
<sec id="s2-3">
<title>Interactive Visualization</title>
<p>Visualization is a key aspect on the interpretation of scRNA-Seq results (<xref ref-type="bibr" rid="B2">Cakir et al., 2020</xref>). Analysis pipelines performs scatter plots for the representation of dimensional reduction results where point colors represent clusters, cell types, gene expression or trajectory features of each cell (<xref ref-type="bibr" rid="B12">Kiselev et al., 2017</xref>; <xref ref-type="bibr" rid="B15">Lin et al., 2017</xref>; <xref ref-type="bibr" rid="B31">Stuart et al., 2019</xref>). These plots are adequate for publishing results, but not for explorative analyses. At present, new technologies based on JavaScript enable the generation of interactive graphs in a Web User Interface. They allow the connection between graphs and the use of HTML5 components which could control visualization aspects. SingleCAnalyzer adopts this technology to visualize information, interconnect graphs, show meta-information, calculate descriptive statistics, generate new graphs under user request and change the representation features interactively. SingleCAnalyzer includes six different graphical representations such as scatter plot, bar chart, heatmap, network, boxplot and density plot. These graphs are interactive, and the user can modify them by clicking on tables, html controls or other graphs.</p>
<p>The central result page is the representation of dimensional reduction and clustering of cells. It is composed by scatter plots where points represent cells, and the user can select the color and shape of points manually or by using clustering, cell population, pseudotime or gene expression results. The user can also explore group frequencies and define resulting groups based on meta-information or cell disposition on the graph. All graphs are interconnected, changes on graphical attributes or cell selections are synchronized on all displays. Cells can be located in all the graphs with a selection over one graph or by means of the locate samples menu. The application also allows the generation of new scatter plots which represent the expression of two genes in each cell.</p>
<p>Once the user defines the groups, he can launch a comparative expression analysis which results in two types of interactive reports. One is the differential expression report which are composed of interactive scatterplots, a boxplot, a line plot and a heatmap. All these graphs show information on mouse action and are connected with the table which summarizes the statistical analysis. They enable the comprehensive exploration of results and the query of information about expression changes of genes. The other report is the functional analysis which includes self-explanatory graphs such as bar plots, networks of terms and triangular heatmaps. Networks and heatmaps represent relations between gene sets which helps in the identification of related gene functions or pathways, while the bar plot shows the number of observed versus expected genes in each category.</p>
<p>Visualization features of SingleCAnalyzer enable the exploration and interpretation of results in an integrated platform which covers the main steps of scRNA-Seq analysis. The platform was presented and discussed at the VIZBI21 conference, where some improvements were suggested by attendants (<xref ref-type="bibr" rid="B32">VIZBI, 2022</xref>). Suggestions were focused on improving the usability of the platform and the adaptation of the analysis pipeline for their objectives. For example, an attendant required an adaptation for the analysis of RNA-Seq data which was developed and can be executed disabling the multiplexing process. SingleCAnalyzer is also distributed as a Docker machine and our graphical functions will be made public as R packages for its open use in analysis pipelines. All the representations performed with SingleCAnalyzer can be downloaded as graphical files ready for its inclusion in publications and analysis reports.</p>
</sec>
</sec>
<sec sec-type="materials|methods" id="s3">
<title>Material and Methods</title>
<sec id="s3-1">
<title>Implementation</title>
<p>The SingleCAnalyzer website runs using LAMP architecture (Linux, Apache, MySQL and PHP). The front-end of the website was developed using PHP, HTML5, JavaScript, D3, JQuery, AJAX and CSS3. Its implementation was based on the RaNA-Seq project, which contains similar alignment, differential expression and functional analysis tools (<xref ref-type="bibr" rid="B22">Prieto and Barrios, 2019</xref>). The analysis pipeline can be executed by a task manager that runs the analysis processes using R, Python or Linux Bash Shell. It also balances the computational load on our high-performance computing cluster. The analysis pipeline integrates cutting-edge tools which rapidly and reliably analyse scRNA-Seq data. <xref ref-type="fig" rid="F2">Figure 2</xref> shows a flowchart of the pipeline used. We have optimised the analysis processes in our pipeline by harnessing computational clustering. Most of the tasks of analysis can be executed in real time. This optimisation has facilitated the development of an open and free cloud-based system.</p>
</sec>
<sec id="s3-2">
<title>FASTQ Processing</title>
<p>Raw sequence files in FASTQ format can be demultiplexed with Alevin software (version 1.3.0) (<xref ref-type="bibr" rid="B30">Srivastava et al., 2019</xref>) or pre-processed using the Fastp tool (version 0.19.4) (<xref ref-type="bibr" rid="B3">Chen et al., 2018</xref>). Gene expression quantification of genes in the selected reference genome is performed using Alevin or Salmon software (<xref ref-type="bibr" rid="B20">Patro et al., 2017</xref>). The platform can be used to assess data generated from any organism. At present, we have downloaded the most popular genomes from Ensembl (164 genomes) and have incorporated their transcriptome indexes within our server (<xref ref-type="bibr" rid="B5">Cunningham et al., 2019</xref>). Quality control of samples is performed based on the alignment summary, descriptive statistics and the Alevin report of demultiplexed samples non-supervised clustering performed using AlevinQC package (version 1.4.0).</p>
</sec>
<sec id="s3-3">
<title>Gene Filtering</title>
<p>Gene filters based on the quantification of gene expression, which reduce the noise and computational costs are available on SingleCAnalyzer. The current version can filter genes with the lowest levels of expression or standard deviations. We have also integrated the function &#x2018;FindVariableFeatures&#x2019; within the Seurat package (version 3.2.2), which can identify variably genes by considering the strong relationship between variability and expression level (<xref ref-type="bibr" rid="B31">Stuart et al., 2019</xref>). Moreover, the user can also perform further dimensionality reduction and clustering processes by analysing the principal components obtained via principal component analysis (PCA). The optimum number of components used for the analyses can be determined using the <italic>calc_npc</italic> function of the CIDR package (version 0.1.5) (<xref ref-type="bibr" rid="B15">Lin et al., 2017</xref>). Empty droplets can be detected and removed with the application of the DropletUtils tool (version 1.8.0) (<xref ref-type="bibr" rid="B17">Lun et al., 2019</xref>).</p>
</sec>
<sec id="s3-4">
<title>Dimensionality Reduction</title>
<p>Interactive visualisation of samples in scatter plots requires a dimensionality reduction process, which is performed using the following methods: 1) PCA, which is generated with the <italic>prcomp</italic> function of the <italic>stats</italic> R package (version 4.0.3); 2) Classic multidimensional scaling (cMDS), which is performed with the <italic>cmdscale</italic> function of the <italic>stats</italic> R package using camberra as distance method; 3) Nonmetric multidimensional scaling (isoMDS), which is performed using the <italic>isoMDS function of the MASS</italic> R package (version 7.3); 4) t-distributed stochastic neighbor embedding (t-SNE), which is performed using the <italic>Rtsne</italic> function of the <italic>Rtsne</italic> R package (version 0.15); 5) Uniform manifold approximation and projection (UMAP), which is performed using the <italic>umap</italic> function of the <italic>uwot</italic> R package (version 0.1.9); 6) and Non-negative matrix factorisation (NMF), which is performed using the <italic>nnmf</italic> function of the <italic>NNLM</italic> package (version 0.4.3)<italic>.</italic> Collectively, application of these methods provides users with a multi-perspective assessment of the relationships between data.</p>
</sec>
<sec id="s3-5">
<title>Unsupervised Clustering</title>
<p>Determination of clusters within the interactive web interface is supported by the results provided by unsupervised clustering methods. At present, SingleCAnalyzer applies the following unsupervised clustering methods: 1) k-means, which is computed using the <italic>kmeans</italic> function of the <italic>stats</italic> R package (with iter_max &#x3d; 15); 2) partition around medoids (PAM), which is computed using the <italic>pam</italic> function of the <italic>cluster</italic> R package (version 2.1.0); 3) hierarchical clustering, which is performed using the <italic>hclust</italic> function of the <italic>stats</italic> R package; 4) leiden clustering and pseudotime analysis, which is performed using Monocle3 R package (version 0.2.3) (<xref ref-type="bibr" rid="B23">Qiu et al., 2017</xref>)<italic>.</italic> The user can specify input parameters such as the desired number of groups, the distance metric used by pam and hclust functions and the agglomeration parameter of hclust.</p>
</sec>
<sec id="s3-6">
<title>Pseudotime Analysis</title>
<p>Trajectory and pseudotime analyses are performed using the Monocle3 R package (<xref ref-type="bibr" rid="B23">Qiu et al., 2017</xref>). It calculates possible trajectories between leiden clusters over the UMAP projection. The pipeline calculates the pseudotime prediction for each cluster centroid and a scale colour which represent the time is applied over the points when an origin cluster is selected. The function preprocess_cds uses PCA or LSI output based on user options with the following parameters: norm_method &#x3d; log and scaling &#x3d; true. The function reduce_dimension uses the following parameters: max_components &#x3d; 2, reduction_method &#x3d; UMAP, umap. metric &#x3d; cosine, umap. min_dist &#x3d; 0.1, umap. n_neighbors &#x3d; 15L, umap. nn_method &#x3d; annoy. The function cluster_cells uses the following parameters: k &#x3d; 20, cluster_method &#x3d; Leiden, nunm_iter &#x3d; 2, partition_qval &#x3d; 0.05. The function learn_graph uses use_partitium and close_loop as true.</p>
</sec>
<sec id="s3-7">
<title>Comparison Between Clusters</title>
<p>Groups of samples can be compared by applying different methods to assess differential expression. Reviews of the use of methods have concluded that no single method outperforms the others under all circumstances, and suggest that it is necessary to determine the optimal method or pipeline for each analysis performed (<xref ref-type="bibr" rid="B26">Seyednasrollah et al., 2013</xref>; <xref ref-type="bibr" rid="B28">Soneson and Delorenzi, 2013</xref>). However, researchers have acknowledged that DESeq2 (version 1.28.1) (<xref ref-type="bibr" rid="B16">Love et al., 2014</xref>), EdgeR (version 3.30.3) (<xref ref-type="bibr" rid="B24">Robinson et al., 2010</xref>) and limma (version 3.44.3) (<xref ref-type="bibr" rid="B14">Law et al., 2014</xref>) are the most widely used methods and consistently performed well when their reliability was assessed. We have integrated all of the methods within a SingleCAnalyzer that can be adjusted to apply customised parameters to individual tests.</p>
<p>SingleCAnalyzer performs a functional enrichment analysis and a gene set enrichment analysis (GSEA) for each comparison result. The enrichment analysis is performed with the R package GOseq (version 1.40) (<xref ref-type="bibr" rid="B35">Young et al., 2010</xref>) and the GSEA is performed with the R package fgsea (version 1.14) (<xref ref-type="bibr" rid="B13">Korotkevich et al., 2016</xref>). Functional annotation database used by these methods was downloaded from the NCBI BioSystems repository (<xref ref-type="bibr" rid="B7">Geer et al., 2009</xref>). Resulting graphs are generated with the package RJSplot (version 2.6) (<xref ref-type="bibr" rid="B1">Barrios and Prieto, 2018</xref>).</p>
</sec>
<sec id="s3-8">
<title>Data Management</title>
<p>Analyses can be launched as anonymous or registered users. Anonymous accounts are regularly deleted, and registered users can require the cancellation of their account. Data of registered users are protected by their personal password which is encrypted on our system. Users can freely download or delete their processed data and analysis results without any limitation. Raw data files uploaded by users (FASTQ, HDF5) are deleted once they are processed. This deletion avoids storage limitations and the presence of sequences in our system.</p>
</sec>
</sec>
<sec sec-type="discussion" id="s4">
<title>Discussion</title>
<p>Single-cell platforms provide computational methods which enable the transformation of sequences into expression values of genes in each cell (<xref ref-type="bibr" rid="B36">Zheng et al., 2017</xref>; <xref ref-type="bibr" rid="B27">Shum et al., 2019</xref>). Further steps can be performed by the application of bioinformatics methods which are available on code repositories or analysis servers. These methods are connected in series to compose an analysis workflow. The development of pipelines is a complex work which involves the installation, test, setting up and integration of computational methods. In addition, full processing of scRNA-Seq data requires an intensive computational processing and the knowledge of programming languages for the execution of the pipeline. On the other hand, cloud servers are designed to avoid the development and execution of pipelines by the analysts, but its use also implies limitations such as additional data uploading time, uncertain server loads and limited customization of the analysis. Previous works have provided web servers for the analysis of scRNA-Seq data from a matrix with gene counts of cells (<xref ref-type="bibr" rid="B6">Gardeux et al., 2017</xref>; <xref ref-type="bibr" rid="B37">Zhu et al., 2017</xref>; <xref ref-type="bibr" rid="B25">Scholz et al., 2018</xref>; <xref ref-type="bibr" rid="B4">Chen et al., 2019</xref>; <xref ref-type="bibr" rid="B18">Monier et al., 2019</xref>). In this work we have developed the first cloud server which allow a complete analysis from sequences to pathways in a fully integrated platform. It was possible with the integration of low computational cost methods for the demultiplexing and quantification of reads which supports Drop-seq and 10x Chromium single-cell protocols (<xref ref-type="bibr" rid="B30">Srivastava et al., 2019</xref>).</p>
<p>Another approach for the analysis of scRNASeq sequences is the use of workflow management systems. A popular option is Galaxy which offers a web-based system for the pipeline construction and the execution of bioinformatic analyses (<xref ref-type="bibr" rid="B11">Jalili et al., 2021</xref>). A recent study has presented Galaxy workflows for the analysis of scRNASeq data (<xref ref-type="bibr" rid="B19">Moreno et al., 2021</xref>). One of the workflows allows the uploading of FASTQ files for processing into an annotated cell matrix with Alevin. Then, post processing is done with Scanpy (<xref ref-type="bibr" rid="B34">Wolf et al., 2018</xref>) and the interactive visualization with the UCSC CellBrowser (<xref ref-type="bibr" rid="B29">Speir et al., 2021</xref>). This workflow has similar limitations to cloud solutions, as customization and uploading time, and requires of a computational cluster account and training about Galaxy workflows. Regarding the integration of results, the application of standard visualization tools avoids the creation of custom interfaces which integrate different nature of results, and the execution of new analysis based on the user interaction with the graph cannot be performed.</p>
<p>Visualization is a key aspect on the interpretation of scRNA-Seq results (<xref ref-type="bibr" rid="B2">Cakir et al., 2020</xref>). An adequate and interactive representation facilitates the correct classification and characterization of cells. This issue has been extensively approached by analysis techniques of cytometry and visualization methods have been adapted to the specific characteristics of single-cell such as the lower number of cells and the increment on the number of variables (transcripts/proteins). Two dimensional plots have been traditionally used for the representation of fluorescent makers on Cytometry. At present, flow cytometry panels can include dozens of makers and its representation as scatterplots are performed by a dimensional reduction technique. Similar strategy is followed for single cell visualization, but the lower number of cells allows its representation with web-based technologies which avoids software installation and platform dependencies. SingleCAnalyzer has developed its graphical interface with D3 and JavaScript technologies which allows the user-graph interaction on a Web browser. This solution has efficiently tested for the representation of 6,000 cells on six simultaneous scatterplots and allows a full interaction with clustering, cell classification, transcript quantification and cell trajectory results. Regarding the differential expression interface, it can handle 60,000 transcripts and perform six interconnected representations (MA-plot, volcano plot, scatterplot, boxplot and heatmap) on user interaction. The scalability of the platform will depend on the optimization of Web Browsers in the storage, representation and processing of interactive HTML Canvas and Scalable Vector Graphics. Current browsers have memory management and multiprocessing limitations. However, these technologies are becoming popular, and browsers are adapting their rendering engines for improving their performace (e.g. RenderingNG technology of chrome).</p>
<p>Future implementations of SingleCAnalyzer will be directed to the integration of novel analysis methods for scRNA-Seq and to the compatibility with new platforms and experimental protocols. At present, we provide semi-automated analysis of scRNA-Seq data on the cloud with analytical and interactive graphs, which enable the comprehensive analysis of results. It is freely available for scientists to explore the potential of their scRNASeq studies running quick analysis on an easy-to-use interface.</p>
</sec>
</body>
<back>
<sec id="s5">
<title>Data Availability Statement</title>
<p>Publicly available datasets were analyzed in this study. This data can be found here: <ext-link ext-link-type="uri" xlink:href="https://singlecanalyzer.eu">https://singlecanalyzer.eu</ext-link>.</p>
</sec>
<sec id="s6">
<title>Author Contributions</title>
<p>CP contributed to conception and design of the study. DB developed the web server. CP and AV designed and implemented the analysis workflow. DB and CP performed the platform test and system optimization. CP wrote the first draft of the manuscript. CP, DB, and AV wrote sections of the manuscript. All authors contributed to manuscript revision, read, and approved the submitted version.</p>
</sec>
<sec id="s7">
<title>Funding</title>
<p>DB and AV was supported by Operational Programme of Youth Employment, European Social Fund (ESF), Junta de Castilla y Leon (JCyL). DB was supported by the PGC project (grant number PGC2018-093755-B-I00) of the Spanish Ministry of Science, Innovation and Universities. CP was supported by the PTA fellowship (grant number PTA2015-10483-I) of the Spanish Ministry of Economy, Industry and Competitiveness (MINECO).</p>
</sec>
<sec sec-type="COI-statement" id="s8">
<title>Conflict of Interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec sec-type="disclaimer" id="s9">
<title>Publisher&#x2019;s Note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
<sec id="s10">
<title>Supplementary Material</title>
<p>The Supplementary Material for this article can be found online at: <ext-link ext-link-type="uri" xlink:href="https://www.frontiersin.org/articles/10.3389/fbinf.2022.793309/full#supplementary-material">https://www.frontiersin.org/articles/10.3389/fbinf.2022.793309/full&#x23;supplementary-material</ext-link>
</p>
<supplementary-material xlink:href="DataSheet1.PDF" id="SM1" mimetype="application/PDF" xmlns:xlink="http://www.w3.org/1999/xlink"/>
</sec>
<ref-list>
<title>References</title>
<ref id="B1">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Barrios</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Prieto</surname>
<given-names>C.</given-names>
</name>
</person-group> (<year>2018</year>). <article-title>RJSplot: Interactive Graphs with R</article-title>. <source>Mol. Inf.</source> <volume>37</volume>, <fpage>1700090</fpage>. <pub-id pub-id-type="doi">10.1002/minf.201700090</pub-id> </citation>
</ref>
<ref id="B2">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Cakir</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Prete</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Huang</surname>
<given-names>N.</given-names>
</name>
<name>
<surname>van Dongen</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Pir</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Kiselev</surname>
<given-names>V. Y.</given-names>
</name>
</person-group> (<year>2020</year>). <article-title>Comparison of Visualization Tools for Single-Cell RNAseq Data</article-title>. <source>Nar. Genomics Bioinform.</source> <volume>2</volume>, <fpage>lqaa052</fpage>. <pub-id pub-id-type="doi">10.1093/nargab/lqaa052</pub-id> </citation>
</ref>
<ref id="B3">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Chen</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Zhou</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Chen</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Gu</surname>
<given-names>J.</given-names>
</name>
</person-group> (<year>2018</year>). <article-title>Fastp: An Ultra-fast All-In-One FASTQ Preprocessor</article-title>. <source>Bioinformatics</source> <volume>34</volume>, <fpage>i884</fpage>&#x2013;<lpage>i890</lpage>. <pub-id pub-id-type="doi">10.1093/bioinformatics/bty560</pub-id> </citation>
</ref>
<ref id="B4">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Chen</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Albergante</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Hsu</surname>
<given-names>J. Y.</given-names>
</name>
<name>
<surname>Lareau</surname>
<given-names>C. A.</given-names>
</name>
<name>
<surname>Lo Bosco</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Guan</surname>
<given-names>J.</given-names>
</name>
<etal/>
</person-group> (<year>2019</year>). <article-title>Single-cell Trajectories Reconstruction, Exploration and Mapping of Omics Data with STREAM</article-title>. <source>Nat. Commun.</source> <volume>10</volume>, <fpage>1903</fpage>. <pub-id pub-id-type="doi">10.1038/s41467-019-09670-4</pub-id> </citation>
</ref>
<ref id="B5">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Cunningham</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Achuthan</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Akanni</surname>
<given-names>W.</given-names>
</name>
<name>
<surname>Allen</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Amode</surname>
<given-names>M. R.</given-names>
</name>
<name>
<surname>Armean</surname>
<given-names>I. M.</given-names>
</name>
<etal/>
</person-group> (<year>2019</year>). <article-title>Ensembl 2019</article-title>. <source>Nucleic Acids Res.</source> <volume>47</volume>, <fpage>D745</fpage>. <pub-id pub-id-type="doi">10.1093/nar/gky1113</pub-id> </citation>
</ref>
<ref id="B6">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Gardeux</surname>
<given-names>V.</given-names>
</name>
<name>
<surname>David</surname>
<given-names>F. P. A.</given-names>
</name>
<name>
<surname>Shajkofci</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Schwalie</surname>
<given-names>P. C.</given-names>
</name>
<name>
<surname>Deplancke</surname>
<given-names>B.</given-names>
</name>
</person-group> (<year>2017</year>). <article-title>ASAP: A Web-Based Platform for the Analysis and Interactive Visualization of Single-Cell RNA-Seq Data</article-title>. <source>Bioinformatics</source> <volume>33</volume>, <fpage>3123</fpage>&#x2013;<lpage>3125</lpage>. <pub-id pub-id-type="doi">10.1093/bioinformatics/btx337</pub-id> </citation>
</ref>
<ref id="B7">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Geer</surname>
<given-names>L. Y.</given-names>
</name>
<name>
<surname>Marchler-Bauer</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Geer</surname>
<given-names>R. C.</given-names>
</name>
<name>
<surname>Han</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>He</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>He</surname>
<given-names>S.</given-names>
</name>
<etal/>
</person-group> (<year>2009</year>). <article-title>The NCBI BioSystems Database</article-title>. <source>Nucleic Acids Res.</source> <volume>38</volume>, <fpage>D492</fpage>&#x2013;<lpage>D496</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gkp858</pub-id> </citation>
</ref>
<ref id="B8">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Guo</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Potter</surname>
<given-names>S. S.</given-names>
</name>
<name>
<surname>Whitsett</surname>
<given-names>J. A.</given-names>
</name>
<name>
<surname>Xu</surname>
<given-names>Y.</given-names>
</name>
</person-group> (<year>2015</year>). <article-title>SINCERA: A Pipeline for Single-Cell RNA-Seq Profiling Analysis</article-title>. <source>PLoS Comput. Biol.</source> <volume>11</volume>, <fpage>e1004575</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pcbi.1004575</pub-id> </citation>
</ref>
<ref id="B9">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Haque</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Engel</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Teichmann</surname>
<given-names>S. A.</given-names>
</name>
<name>
<surname>L&#xf6;nnberg</surname>
<given-names>T.</given-names>
</name>
</person-group> (<year>2017</year>). <article-title>A Practical Guide to Single-Cell RNA-Sequencing for Biomedical Research and Clinical Applications</article-title>. <source>Genome Med.</source> <volume>9</volume>, <fpage>75</fpage>. <pub-id pub-id-type="doi">10.1186/s13073-017-0467-4</pub-id> </citation>
</ref>
<ref id="B10">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Hwang</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Lee</surname>
<given-names>J. H.</given-names>
</name>
<name>
<surname>Bang</surname>
<given-names>D.</given-names>
</name>
</person-group> (<year>2018</year>). <article-title>Single-cell RNA Sequencing Technologies and Bioinformatics Pipelines</article-title>. <source>Exp. Mol. Med.</source> <volume>50</volume>, <fpage>1</fpage>&#x2013;<lpage>14</lpage>. <pub-id pub-id-type="doi">10.1038/s12276-018-0071-8</pub-id> </citation>
</ref>
<ref id="B11">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Jalili</surname>
<given-names>V.</given-names>
</name>
<name>
<surname>Afgan</surname>
<given-names>E.</given-names>
</name>
<name>
<surname>Gu</surname>
<given-names>Q.</given-names>
</name>
<name>
<surname>Clements</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Blankenberg</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Goecks</surname>
<given-names>J.</given-names>
</name>
<etal/>
</person-group> (<year>2021</year>). <article-title>The Galaxy Platform for Accessible, Reproducible and Collaborative Biomedical Analyses: 2020 Update</article-title>. <source>Nucleic Acids Res.</source> <volume>48</volume>, <fpage>W395</fpage>. <pub-id pub-id-type="doi">10.1093/NAR/GKAA434</pub-id> </citation>
</ref>
<ref id="B12">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Kiselev</surname>
<given-names>V. Y.</given-names>
</name>
<name>
<surname>Kirschner</surname>
<given-names>K.</given-names>
</name>
<name>
<surname>Schaub</surname>
<given-names>M. T.</given-names>
</name>
<name>
<surname>Andrews</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Yiu</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Chandra</surname>
<given-names>T.</given-names>
</name>
<etal/>
</person-group> (<year>2017</year>). <article-title>SC3: Consensus Clustering of Single-Cell RNA-Seq Data</article-title>. <source>Nat. Methods</source> <volume>14</volume>, <fpage>483</fpage>&#x2013;<lpage>486</lpage>. <pub-id pub-id-type="doi">10.1038/nmeth.4236</pub-id> </citation>
</ref>
<ref id="B13">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Korotkevich</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Sukhov</surname>
<given-names>V.</given-names>
</name>
<name>
<surname>Budin</surname>
<given-names>N.</given-names>
</name>
<name>
<surname>Shpak</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Artyomov</surname>
<given-names>M. N.</given-names>
</name>
<name>
<surname>Sergushichev</surname>
<given-names>A.</given-names>
</name>
</person-group> (<year>2016</year>). <article-title>An Algorithm for Fast Preranked Gene Set Enrichment Analysis Using Cumulative Statistic Calculation</article-title>. <source>bioRxiv</source>, <fpage>60012</fpage>. <pub-id pub-id-type="doi">10.1101/060012</pub-id> </citation>
</ref>
<ref id="B14">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Law</surname>
<given-names>C. W.</given-names>
</name>
<name>
<surname>Chen</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Shi</surname>
<given-names>W.</given-names>
</name>
<name>
<surname>Smyth</surname>
<given-names>G. K.</given-names>
</name>
</person-group> (<year>2014</year>). <article-title>Voom: Precision Weights Unlock Linear Model Analysis Tools for RNA-Seq Read Counts</article-title>. <source>Genome Biol.</source> <volume>15</volume>, <fpage>R29</fpage>. <pub-id pub-id-type="doi">10.1186/gb-2014-15-2-r29</pub-id> </citation>
</ref>
<ref id="B15">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Lin</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Troup</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Ho</surname>
<given-names>J. W.</given-names>
</name>
</person-group> (<year>2017</year>). <article-title>CIDR: Ultrafast and Accurate Clustering through Imputation for Single-Cell RNA-Seq Data</article-title>. <source>Genome Biol.</source> <volume>18</volume>, <fpage>59</fpage>. <pub-id pub-id-type="doi">10.1186/s13059-017-1188-0</pub-id> </citation>
</ref>
<ref id="B16">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Love</surname>
<given-names>M. I.</given-names>
</name>
<name>
<surname>Anders</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Huber</surname>
<given-names>W.</given-names>
</name>
</person-group> (<year>2014</year>). <article-title>Differential Analysis of Count Data - The DESeq2 Package</article-title>. <source>Genome Biol.</source> <volume>15</volume>, <fpage>550</fpage>. <pub-id pub-id-type="doi">10.1186/s13059-014-0550-8</pub-id> </citation>
</ref>
<ref id="B17">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Lun</surname>
<given-names>A. T. L.</given-names>
</name>
<name>
<surname>Riesenfeld</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Andrews</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Dao</surname>
<given-names>T. P.</given-names>
</name>
<name>
<surname>Gomes</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Marioni</surname>
<given-names>J. C.</given-names>
</name>
<etal/>
</person-group> (<year>2019</year>). <article-title>EmptyDrops: Distinguishing Cells from Empty Droplets in Droplet-Based Single-Cell RNA Sequencing Data</article-title>. <source>Genome Biol.</source> <volume>20</volume>, <fpage>63</fpage>. <pub-id pub-id-type="doi">10.1186/s13059-019-1662-y</pub-id> </citation>
</ref>
<ref id="B18">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Monier</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>McDermaid</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Zhao</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Miller</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Fennell</surname>
<given-names>A.</given-names>
</name>
<etal/>
</person-group> (<year>2019</year>). <article-title>IRIS-EDA: An Integrated RNA-Seq Interpretation System for Gene Expression Data Analysis</article-title>. <source>PLoS Comput. Biol.</source> <volume>15</volume>, <fpage>e1006792</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pcbi.1006792</pub-id> </citation>
</ref>
<ref id="B19">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Moreno</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Huang</surname>
<given-names>N.</given-names>
</name>
<name>
<surname>Manning</surname>
<given-names>J. R.</given-names>
</name>
<name>
<surname>Mohammed</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Solovyev</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Polanski</surname>
<given-names>K.</given-names>
</name>
<etal/>
</person-group> (<year>2021</year>). <article-title>User-friendly, Scalable Tools and Workflows for Single-Cell RNA-Seq Analysis</article-title>. <source>Nat. Methods</source> <volume>18</volume>, <fpage>327</fpage>&#x2013;<lpage>328</lpage>. <pub-id pub-id-type="doi">10.1038/s41592-021-01102-w</pub-id> </citation>
</ref>
<ref id="B20">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Patro</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Duggal</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Love</surname>
<given-names>M. I.</given-names>
</name>
<name>
<surname>Irizarry</surname>
<given-names>R. A.</given-names>
</name>
<name>
<surname>Kingsford</surname>
<given-names>C.</given-names>
</name>
</person-group> (<year>2017</year>). <article-title>Salmon Provides Fast and Bias-Aware Quantification of Transcript Expression</article-title>. <source>Nat. Methods</source> <volume>14</volume>, <fpage>417</fpage>&#x2013;<lpage>419</lpage>. <pub-id pub-id-type="doi">10.1038/nmeth.4197</pub-id> </citation>
</ref>
<ref id="B21">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Perraudeau</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Risso</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Street</surname>
<given-names>K.</given-names>
</name>
<name>
<surname>Purdom</surname>
<given-names>E.</given-names>
</name>
<name>
<surname>Dudoit</surname>
<given-names>S.</given-names>
</name>
</person-group> (<year>2017</year>). <article-title>Bioconductor Workflow for Single-Cell RNA Sequencing: Normalization, Dimensionality Reduction, Clustering, and Lineage Inference</article-title>. <source>F1000Res</source> <volume>6</volume>, <fpage>1158</fpage>. <pub-id pub-id-type="doi">10.12688/f1000research.12122.1</pub-id> </citation>
</ref>
<ref id="B22">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Prieto</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Barrios</surname>
<given-names>D.</given-names>
</name>
</person-group> (<year>2019</year>). <article-title>RaNA-Seq: Interactive RNA-Seq Analysis from FASTQ Files to Functional Analysis</article-title>. <source>Bioinformatics</source> <volume>36</volume>, <fpage>1955</fpage>&#x2013;<lpage>1956</lpage>. <pub-id pub-id-type="doi">10.1093/bioinformatics/btz854</pub-id> </citation>
</ref>
<ref id="B23">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Qiu</surname>
<given-names>X.</given-names>
</name>
<name>
<surname>Mao</surname>
<given-names>Q.</given-names>
</name>
<name>
<surname>Tang</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Chawla</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Pliner</surname>
<given-names>H. A.</given-names>
</name>
<etal/>
</person-group> (<year>2017</year>). <article-title>Reversed Graph Embedding Resolves Complex Single-Cell Trajectories</article-title>. <source>Nat. Methods</source> <volume>14</volume>, <fpage>979</fpage>&#x2013;<lpage>982</lpage>. <pub-id pub-id-type="doi">10.1038/nmeth.4402</pub-id> </citation>
</ref>
<ref id="B24">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Robinson</surname>
<given-names>M. D.</given-names>
</name>
<name>
<surname>McCarthy</surname>
<given-names>D. J.</given-names>
</name>
<name>
<surname>Smyth</surname>
<given-names>G. K.</given-names>
</name>
</person-group> (<year>2010</year>). <article-title>edgeR: A Bioconductor Package for Differential Expression Analysis of Digital Gene Expression Data</article-title>. <source>Bioinformatics</source> <volume>26</volume>, <fpage>139</fpage>&#x2013;<lpage>140</lpage>. <pub-id pub-id-type="doi">10.1093/bioinformatics/btp616</pub-id> </citation>
</ref>
<ref id="B25">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Scholz</surname>
<given-names>C. J.</given-names>
</name>
<name>
<surname>Biernat</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Becker</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Ba&#xdf;ler</surname>
<given-names>K.</given-names>
</name>
<name>
<surname>G&#xfc;nther</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Balfer</surname>
<given-names>J.</given-names>
</name>
<etal/>
</person-group> (<year>2018</year>). <article-title>FASTGenomics: An Analytical Ecosystem for Single-Cell RNA Sequencing Data</article-title>. <source>bioRxiv</source>. <pub-id pub-id-type="doi">10.1101/272476</pub-id> </citation>
</ref>
<ref id="B26">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Seyednasrollah</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Laiho</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Elo</surname>
<given-names>L. L.</given-names>
</name>
</person-group> (<year>2013</year>). <article-title>Comparison of Software Packages for Detecting Differential Expression in RNA-Seq Studies</article-title>. <source>Brief. Bioinform.</source> <volume>16</volume>, <fpage>59</fpage>&#x2013;<lpage>70</lpage>. <pub-id pub-id-type="doi">10.1093/bib/bbt086</pub-id> </citation>
</ref>
<ref id="B27">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Shum</surname>
<given-names>E. Y.</given-names>
</name>
<name>
<surname>Walczak</surname>
<given-names>E. M.</given-names>
</name>
<name>
<surname>Chang</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Christina Fan</surname>
<given-names>H.</given-names>
</name>
</person-group> (<year>2019</year>). <article-title>Quantitation of mRNA Transcripts and Proteins Using the BD Rhapsody Single-Cell Analysis System</article-title>. <source>Adv. Exp. Med. Biol.</source> <volume>1129</volume>, <fpage>63</fpage>&#x2013;<lpage>79</lpage>. <pub-id pub-id-type="doi">10.1007/978-981-13-6037-4_5</pub-id> </citation>
</ref>
<ref id="B28">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Soneson</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Delorenzi</surname>
<given-names>M.</given-names>
</name>
</person-group> (<year>2013</year>). <article-title>A Comparison of Methods for Differential Expression Analysis of RNA-Seq Data</article-title>. <source>BMC Bioinform.</source> <volume>14</volume>, <fpage>91</fpage>. <pub-id pub-id-type="doi">10.1186/1471-2105-14-91</pub-id> </citation>
</ref>
<ref id="B29">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Speir</surname>
<given-names>M. L.</given-names>
</name>
<name>
<surname>Bhaduri</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Markov</surname>
<given-names>N. S.</given-names>
</name>
<name>
<surname>Moreno</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Nowakowski</surname>
<given-names>T. J.</given-names>
</name>
<name>
<surname>Papatheodorou</surname>
<given-names>I.</given-names>
</name>
<etal/>
</person-group> (<year>2021</year>). <article-title>UCSC Cell Browser: Visualize Your Single-Cell Data</article-title>. <source>Bioinformatics</source> <volume>37</volume>, <fpage>4578</fpage>&#x2013;<lpage>4580</lpage>. <pub-id pub-id-type="doi">10.1093/bioinformatics/btab503</pub-id> </citation>
</ref>
<ref id="B30">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Srivastava</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Malik</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Smith</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Sudbery</surname>
<given-names>I.</given-names>
</name>
<name>
<surname>Patro</surname>
<given-names>R.</given-names>
</name>
</person-group> (<year>2019</year>). <article-title>Alevin Efficiently Estimates Accurate Gene Abundances from dscRNA-Seq Data</article-title>. <source>Genome Biol.</source> <volume>20</volume>, <fpage>65</fpage>. <pub-id pub-id-type="doi">10.1186/s13059-019-1670-y</pub-id> </citation>
</ref>
<ref id="B31">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Stuart</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Butler</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Hoffman</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Hafemeister</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Papalexi</surname>
<given-names>E.</given-names>
</name>
<name>
<surname>Mauck</surname>
<given-names>W. M.</given-names>
</name>
<etal/>
</person-group> (<year>2019</year>). <article-title>Comprehensive Integration of Single-Cell Data</article-title>. <source>Cell</source> <volume>177</volume>, <fpage>1888</fpage>&#x2013;<lpage>1902</lpage>. <pub-id pub-id-type="doi">10.1016/j.cell.2019.05.031</pub-id> </citation>
</ref>
<ref id="B32">
<citation citation-type="web">
<collab>VIZBI</collab> (<year>2022</year>). <article-title>Posters</article-title>. <comment>Available at: <ext-link ext-link-type="uri" xlink:href="https://vizbi.org/Posters/2021/vC15">https://vizbi.org/Posters/2021/vC15</ext-link> (Accessed March 4, 2022)</comment>. </citation>
</ref>
<ref id="B33">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Wagner</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Yanai</surname>
<given-names>I.</given-names>
</name>
</person-group> (<year>2018</year>). <article-title>Moana: A Robust and Scalable Cell Type Classification Framework for Single-Cell RNA-Seq Data</article-title>. <source>bioRxiv</source>. <pub-id pub-id-type="doi">10.1101/456129</pub-id> </citation>
</ref>
<ref id="B34">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Wolf</surname>
<given-names>F. A.</given-names>
</name>
<name>
<surname>Angerer</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Theis</surname>
<given-names>F. J.</given-names>
</name>
</person-group> (<year>2018</year>). <article-title>SCANPY: Large-Scale Single-Cell Gene Expression Data Analysis</article-title>. <source>Genome Biol.</source> <volume>19</volume>, <fpage>15</fpage>. <pub-id pub-id-type="doi">10.1186/s13059-017-1382-0</pub-id> </citation>
</ref>
<ref id="B35">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Young</surname>
<given-names>M. D.</given-names>
</name>
<name>
<surname>Wakefield</surname>
<given-names>M. J.</given-names>
</name>
<name>
<surname>Smyth</surname>
<given-names>G. K.</given-names>
</name>
</person-group> (<year>2010</year>). <article-title>Goseq : Gene Ontology Testing for RNA-Seq Datasets Reading Data</article-title>. <source>Gene</source> <volume>11</volume>, <fpage>1</fpage>&#x2013;<lpage>21</lpage>. <comment>Available at: <ext-link ext-link-type="uri" xlink:href="http://cobra20.fhcrc.org/packages/release/bioc/vignettes/goseq/inst/doc/goseq.pdf">http://cobra20.fhcrc.org/packages/release/bioc/vignettes/goseq/inst/doc/goseq.pdf</ext-link>
</comment>. </citation>
</ref>
<ref id="B36">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Zheng</surname>
<given-names>G. X.</given-names>
</name>
<name>
<surname>Terry</surname>
<given-names>J. M.</given-names>
</name>
<name>
<surname>Belgrader</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Ryvkin</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Bent</surname>
<given-names>Z. W.</given-names>
</name>
<name>
<surname>Wilson</surname>
<given-names>R.</given-names>
</name>
<etal/>
</person-group> (<year>2017</year>). <article-title>Massively Parallel Digital Transcriptional Profiling of Single Cells</article-title>. <source>Nat. Commun.</source> <volume>8</volume>, <fpage>14049</fpage>. <pub-id pub-id-type="doi">10.1038/ncomms14049</pub-id> </citation>
</ref>
<ref id="B37">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Zhu</surname>
<given-names>X.</given-names>
</name>
<name>
<surname>Wolfgruber</surname>
<given-names>T. K.</given-names>
</name>
<name>
<surname>Tasato</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Arisdakessian</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Garmire</surname>
<given-names>D. G.</given-names>
</name>
<name>
<surname>Garmire</surname>
<given-names>L. X.</given-names>
</name>
</person-group> (<year>2017</year>). <article-title>Granatum: A Graphical Single-Cell RNA-Seq Analysis Pipeline for Genomics Scientists</article-title>. <source>Genome Med.</source> <volume>9</volume>, <fpage>108</fpage>. <pub-id pub-id-type="doi">10.1186/s13073-017-0492-3</pub-id> </citation>
</ref>
</ref-list>
</back>
</article>