<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Archiving and Interchange DTD v2.3 20070202//EN" "archivearticle.dtd">
<article article-type="methods-article" dtd-version="2.3" xml:lang="EN" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Genet.</journal-id>
<journal-title>Frontiers in Genetics</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Genet.</abbrev-journal-title>
<issn pub-type="epub">1664-8021</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="publisher-id">1120185</article-id>
<article-id pub-id-type="doi">10.3389/fgene.2023.1120185</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Genetics</subject>
<subj-group>
<subject>Methods</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Identifying colon cancer stage related genes and their cellular pathways</article-title>
<alt-title alt-title-type="left-running-head">Chen et al.</alt-title>
<alt-title alt-title-type="right-running-head">
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/fgene.2023.1120185">10.3389/fgene.2023.1120185</ext-link>
</alt-title>
</title-group>
<contrib-group>
<contrib contrib-type="author">
<name>
<surname>Chen</surname>
<given-names>Bolin</given-names>
</name>
<xref ref-type="aff" rid="aff1">
<sup>1</sup>
</xref>
<xref ref-type="aff" rid="aff2">
<sup>2</sup>
</xref>
<xref ref-type="aff" rid="aff3">
<sup>3</sup>
</xref>
<uri xlink:href="https://loop.frontiersin.org/people/833694/overview"/>
</contrib>
<contrib contrib-type="author">
<name>
<surname>Chakrobortty</surname>
<given-names>Nandita</given-names>
</name>
<xref ref-type="aff" rid="aff1">
<sup>1</sup>
</xref>
<uri xlink:href="https://loop.frontiersin.org/people/2136592/overview"/>
</contrib>
<contrib contrib-type="author">
<name>
<surname>Saha</surname>
<given-names>Apu Kumar</given-names>
</name>
<xref ref-type="aff" rid="aff1">
<sup>1</sup>
</xref>
<uri xlink:href="https://loop.frontiersin.org/people/2168265/overview"/>
</contrib>
<contrib contrib-type="author" corresp="yes">
<name>
<surname>Shang</surname>
<given-names>Xuequn</given-names>
</name>
<xref ref-type="aff" rid="aff1">
<sup>1</sup>
</xref>
<xref ref-type="aff" rid="aff2">
<sup>2</sup>
</xref>
<xref ref-type="aff" rid="aff3">
<sup>3</sup>
</xref>
<xref ref-type="corresp" rid="c001">&#x2a;</xref>
<uri xlink:href="https://loop.frontiersin.org/people/898708/overview"/>
</contrib>
</contrib-group>
<aff id="aff1">
<sup>1</sup>
<institution>School of Computer Science</institution>, <institution>Northwestern Polytechnical University</institution>, <addr-line>Xi&#x2019;an</addr-line>, <addr-line>Shaanxi</addr-line>, <country>China</country>
</aff>
<aff id="aff2">
<sup>2</sup>
<institution>MIIT Key Laboratory of Big Data Storage and Management</institution>, <institution>Northwestern Polytechnical University</institution>, <addr-line>Xi&#x2019;an</addr-line>, <addr-line>Shaanxi</addr-line>, <country>China</country>
</aff>
<aff id="aff3">
<sup>3</sup>
<institution>National Engineering Laboratory for Integrated Aero-Space-Ground-Ocean Big Data Application Technology</institution>, <institution>Northwestern Polytechnical University</institution>, <addr-line>Xi&#x2019;an</addr-line>, <addr-line>Shaanxi</addr-line>, <country>China</country>
</aff>
<author-notes>
<fn fn-type="edited-by">
<p>
<bold>Edited by:</bold> <ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/732609/overview">Fa Zhang</ext-link>, Institute of Computing Technology (CAS), China</p>
</fn>
<fn fn-type="edited-by">
<p>
<bold>Reviewed by:</bold> <ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/721147/overview">Tianjiao Zhang</ext-link>, Northeast Forestry University, China</p>
<p>
<ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/875818/overview">Jin-Xing Liu</ext-link>, Qufu Normal University, China</p>
</fn>
<corresp id="c001">&#x2a;Correspondence: Xuequn Shang, <email>npu_bioinf@hotmail.com</email>
</corresp>
<fn fn-type="other">
<p>This article was submitted to Computational Genomics, a section of the journal Frontiers in Genetics</p>
</fn>
</author-notes>
<pub-date pub-type="epub">
<day>19</day>
<month>01</month>
<year>2023</year>
</pub-date>
<pub-date pub-type="collection">
<year>2023</year>
</pub-date>
<volume>14</volume>
<elocation-id>1120185</elocation-id>
<history>
<date date-type="received">
<day>09</day>
<month>12</month>
<year>2022</year>
</date>
<date date-type="accepted">
<day>09</day>
<month>01</month>
<year>2023</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#xa9; 2023 Chen, Chakrobortty, Saha and Shang.</copyright-statement>
<copyright-year>2023</copyright-year>
<copyright-holder>Chen, Chakrobortty, Saha and Shang</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/">
<p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p>
</license>
</permissions>
<abstract>
<p>In the world, colon cancer is regarded as one of the most common deadly cancer. Due to the lack of a better understanding of its prognosis system, this prevailing cancer has the second-highest morbidity and mortality rate compared with other cancers. A variety of genes are responsible to participate in colon cancer and the molecular mechanism is almost unsure. In addition, various studies have been done to identify the differentially expressed genes to investigate the dysfunctions of the genes but most of them did it individually. In this study, we constructed a functional interaction network for identifying the group of genes that conduct cellular functions and Protein-Protein Interaction network, which aims to better understanding protein functions and their biological relationships. A functional evolution network was also generated to analyze the dysfunctions from initial stage to later stage of colon cancer by investigating the gene modules and their molecular functions. The results show that the proposed evolution network is able to detect the significant cellular functions, which can be used to explore the evolution process of colon cancer. Moreover, a total of 10 core genes associated with colon cancer were identified, which were INS, SNAP25, GRIA2, SST, GCG, PVALB, SLC17A7, SLC32A1, SLC17A6, and NPY, respectively. The responsible candidate genes and corresponding pathways presented in this study could be used to develop new tumor indicators and novel therapeutic targets for the prevention and treatment of colon cancer.</p>
</abstract>
<kwd-group>
<kwd>colon cancer</kwd>
<kwd>differentially expressed gene</kwd>
<kwd>cancer stage</kwd>
<kwd>functional evolution network</kwd>
<kwd>cancer evolution</kwd>
</kwd-group>
<contract-sponsor id="cn001">National Natural Science Foundation of China<named-content content-type="fundref-id">10.13039/501100001809</named-content>
</contract-sponsor>
<contract-sponsor id="cn002">National Key Research and Development Program of China<named-content content-type="fundref-id">10.13039/501100012166</named-content>
</contract-sponsor>
</article-meta>
</front>
<body>
<sec id="s1">
<title>Introduction</title>
<p>Colon cancer is a cancer with high morbidity and mortality. It is considered as the most common malignant cancer and the second commonest death cause in the modern world. Genome instability, epigenetic abnormalities, and gene expression disorders are typical molecular features of colon cancer. The increase in prevalence is related to an aging population as well as poor eating habits, smoking, lack of physical movements, and obesity in western countries (<xref ref-type="bibr" rid="B11">Kuipers et al., 20162015</xref>). A change in incidence is also observed in some familial cancer syndromes as well as in sporadic disease rates. The incidence of the disease is more common in urban areas compared with rural areas. More men are injured by this disease rather than women. Although older people are at high risk, however, in recent years a significant number of young generations are also victims of this cancer (<xref ref-type="bibr" rid="B2">Chen et al., 2020a</xref>).</p>
<p>As living standards around the world have improved and access to healthcare has increased, we have noticed a considerable improvement in the diagnosis and treatment of diseases. Despite these medical advances, even though the death rate has reduced in over the world, however, the mortality rate from colon cancer has increased, and overall survival is still poor. Many analyses based on the survival rate have demonstrated that metastasis can play a vital role in the reduction of survival rate (<xref ref-type="bibr" rid="B14">Cho-Chung et al., 2002</xref>; <xref ref-type="bibr" rid="B9">He et al., 2018</xref>). However, previous studies did the analysis where DEGs act alone. But genes are not isolated from each other. They worked together by creating modules and forming modules to undertake biological functions (<xref ref-type="bibr" rid="B3">Chen et al., 2020b</xref>). In this study, we analyzed depending on the modules of genes. Moreover, various relevant pathways have already been discovered for colon cancer development but still there remain so many and the main reason is the complex evolutionary process of CRC development.</p>
<p>As colon cancer is considered the leading cause of death cancer in the world and the molecular mechanism of colon cancer is almost unclear. It has been widely accepted that the earlier stage of the cancer is significantly different than the later stages, and if a patient can be diagnosised earlier, it may have higher probability to be cured. In this study, we have performed several investigations based on both data and networks. Traditional approaches have limitations as follows. They did not employ modules thoroughly to evaluate the differentially expressed genes, but directly utilized DEGs to make the functional enrichment analyses. But it is known that proteins rarely act alone, but often collaborate in groups to carry out biological tasks. In this study we performed a cluster analysis on the functional interactional network to identify the modules of genes. These modules were used to construct the PPI network and the cluster interaction network. Moreover, to better understand the relationships between modules and examine their biological functions, we created an evolutionary functional network using the most important DEGs. In addition, existing models also used directly the expression data for their analysis to find out DEGs and enrichment analysis, whereas in this study we work with modules, we use the data from the functional interaction network, cluster interaction network, and PPI network to get the higher precision result.</p>
<p>To be more specific, data and network analysis were performed in different ways to identify the stage-related genes and understand the different biological functions of colon cancer. Gene expression data and colon cancer clinical information data were obtained from the TCGA database and divided the expression data of colon cancer samples in four stages. The differentially expressed genes (DEGs) were obtained at each pair by analyzing the gene expression data, and the PPI network and functional interaction network were constructed for analyzing the interaction among genes. Then, the MCL graph clustering algorithm was applied to select the modules of proteins. By combining the seven-cluster networks, a functional evolutionary network was finally established to analyze the relationship between functional modules for each stage and adjacent stages. The most relevant pathways were identified by doing a KEGG pathway enrichment analysis. In addition, the stage-related hub genes of colon cancer were also identified. This study is expected to provide some significant information regarding colon cancer at different stages and the potential biomarkers that can be used for early diagnosis and make a contribution to colon cancer treatment. The overall framework is shown in <xref ref-type="fig" rid="F1">Figure 1</xref>.</p>
<fig id="F1" position="float">
<label>FIGURE 1</label>
<caption>
<p>The overall framework of the proposed method.</p>
</caption>
<graphic xlink:href="fgene-14-1120185-g001.tif"/>
</fig>
</sec>
<sec sec-type="materials|methods" id="s2">
<title>Materials and methods</title>
<sec id="s2-1">
<title>Gene expression data</title>
<p>Gene expression data were downloaded from The Cancer Genome Atlas (TCGA) for colon cancer. TCGA colon adenocarcinoma (TCGA-COAD) is the project name (<ext-link ext-link-type="uri" xlink:href="https://portal.gdc.cancer.gov/projects/TCGA-COAD">https://portal.gdc.cancer.gov/projects/TCGA-COAD</ext-link>). HTSeq-Counts data is included in this database, along with HTSeq-FPKM and HT-FPKM-UQ data. The RNA-seq samples of data type HTSeq-Counts were analyzed for this analysis. The total number of samples was 329. And the total number of 20,530 protein-coding genes were selected for the next analysis.</p>
</sec>
<sec id="s2-2">
<title>Clinical data and differentially expressed gene identification</title>
<p>Clinical data on colon cancer were also acquired from the TCGA dataset, which included 447 colon cancer samples. There was a variety of clinical information available in the original dataset for each sample, but for this study, only the sample number and four cancer stages information were retrieved here. After following the correlation of gene expression data with the corresponding clinical data and extracting the missing information data, got 312 overlapping samples. All these samples (overlapping sample numbers) were categorized into five groups. They are healthy human tissue or normal tissue, stage I, stage II, stage III, and stage IV. There are 41 normal tissues, 45 stage I samples, 109 stage II samples, 80 stage III samples, and 37 stage IV samples. For further analysis, all the normal samples were combined with the individual stage. After that, the final samples would be 86, 150, 121, and 78 respectively. These are considered the final four differentially expressed (DE) sets.</p>
<p>For detecting differentially expressed genes (DE), this study mainly used gene expression analysis. To avoid the noises of raw data, data preprocessing is crucial to minimize noise. In addition, performing high-level analysis is a must for quality assessment and preprocessing of sequencing data. Counts Per Million (CPM) value should be calculated for each group of datasets due to filtering lowly expressed genes. Four stages and one normal sample dataset were used for identifying Differentially Expressed Genes. We also considered the TMM (Trimmed Mean of the M-value) algorithm (<xref ref-type="bibr" rid="B16">Robinson and Oshlack, 2010</xref>) because it is more effective for comparisons between samples as it does not count gene length or library size.</p>
</sec>
<sec id="s2-3">
<title>Differential gene expression analysis</title>
<p>In this study, the edgeR package (<xref ref-type="bibr" rid="B15">Robinson et al., 2010</xref>) in R language was generated from Bioconductor and used to analyze and identify the differentially expressed genes (DEGs) list of the cancerous tissues for four DE sets. These analyses of DEGs were performed between control samples and CRC samples of four different stages. The <italic>p</italic>-values were calculated for all DEGs, which were further adjusted into false discovery rate (FDR) by using the Benjamini&#x2013;Hochberg method. Fold Change (FC) value for each group was also calculated and only the genes with <inline-formula id="inf1">
<mml:math id="m1">
<mml:mrow>
<mml:mi>F</mml:mi>
<mml:mi>D</mml:mi>
<mml:mi>R</mml:mi>
<mml:mtext>&#x2009;</mml:mtext>
<mml:mo>&#x3c;</mml:mo>
<mml:mn>0.05</mml:mn>
</mml:mrow>
</mml:math>
</inline-formula> and <inline-formula id="inf2">
<mml:math id="m2">
<mml:mrow>
<mml:mrow>
<mml:mfenced open="|" close="|" separators="|">
<mml:mrow>
<mml:mi>l</mml:mi>
<mml:mi>o</mml:mi>
<mml:mi>g</mml:mi>
<mml:mi>F</mml:mi>
<mml:mi>C</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>&#x2265;</mml:mo>
<mml:mn>1</mml:mn>
</mml:mrow>
</mml:math>
</inline-formula> were defined as differentially expressed (DE) genes.</p>
</sec>
<sec id="s2-4">
<title>All DEGs and intersection DEGs</title>
<p>An online tool, Vennplex (<xref ref-type="bibr" rid="B1">Cai et al., 2013</xref>), was used to identify not only all DEGs but also the intersection of DEGs between four DEG sets. It also showed the similarities and differences in expression changes between different groups. For this analysis, it is necessary that both gene symbol and <inline-formula id="inf3">
<mml:math id="m3">
<mml:mrow>
<mml:mfenced open="|" close="|" separators="|">
<mml:mrow>
<mml:mi mathvariant="italic">log</mml:mi>
<mml:mo>&#x2061;</mml:mo>
<mml:mn>2</mml:mn>
<mml:mi>F</mml:mi>
<mml:mi>C</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:math>
</inline-formula> should be inputted. All common DEGs in these datasets were selected for further study.</p>
</sec>
<sec id="s2-5">
<title>Filtering DEGs for stage-specific network generation</title>
<p>For filtering the remarkable genes, histogram analysis was performed on both clinical and PPI data. The DEGs with <inline-formula id="inf4">
<mml:math id="m4">
<mml:mrow>
<mml:mi>l</mml:mi>
<mml:mi>o</mml:mi>
<mml:mi>g</mml:mi>
<mml:mi>F</mml:mi>
<mml:mi>C</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> &#x3e; 1.5 and in the PPI network, connections of edge confidence &#x3e;0.75 were selected as the final DEGs data list as those are mostly connected to colon cancer.</p>
</sec>
<sec id="s2-6">
<title>Stage-specific cluster interaction network generation</title>
<p>In order to identify gene groups that conduct cancer related cellular functions, the functional interaction (FI) network was built. Because one of the motivations of this study is to analyze the functions of the group of genes and the FI network is unable to represent the relations between the coding genes. Four functional interaction networks were constructed using the 4 DE sets (filtered by histogram analysis). The nodes represented the genes, and the edge represented the functional associations and interaction associations between genes. The edge contains the implications of biological function, which provided the basis for the functional analysis. ReactomeFIViz (<xref ref-type="bibr" rid="B20">Wu et al., 2014</xref>), a tool in the Cytoscape software (<xref ref-type="bibr" rid="B17">Shannon et al., 2003</xref>), was utilized to analyze the functional interactions between all these DEGs.</p>
<p>In this study, the Markov Cluster Algorithm (MCL) was applied to each stage and whether the cluster has five nodes or above were connected, those clusters were selected to analyze the relationship between functional modules for each stage and adjacent stages. Finally, 25 clusters, 33 clusters, 27 clusters, and 29 clusters were chosen for stage I, stage II, stage III, and stage IV respectively. According to the connections of each cluster, seven networks were built, where four networks are for individual stage (stage I, stage II, stage III, and stage IV) and another three is between adjacent stages (between stage I and stage II, between stage II and stage III, between stage III and stage IV).</p>
</sec>
<sec id="s2-7">
<title>Functional evolutionary network generation</title>
<p>The main contribution of this study is the combination of seven networks with four groups (stage I, stage II, stage III, and stage IV) and seven connections (above seven networks). The Cytoscape software was used to make this combined network, called pathway interaction network, in order to analyze the relationship between functional modules. The final graphical view of the pathway interaction network was entitled as the functional evolutionary network. Where the correlation between four DEG sets was obtained and the staged genes involved in the pathway were used for further analysis.</p>
<p>In the functional evolutionary network analysis, the pathway enrichment analysis has been done. It was not necessary to perform a staged biological function analysis because the pathway enrichment provided a variety of the mixed results of pathway. In order to obtain results of functional analysis in stages, this study, therefore, considered using a method that represents pathways by a graph, which are enriched in different stages. This study analyzed two relationships between pathways and between adjacent stages, the first is to determine which pathways were significantly different between stages, and the second determination is to identify which pathways were associated with adjacent stages.</p>
<p>In this study, DAVID (<xref ref-type="bibr" rid="B4">Dennis et al., 2003</xref>) was used to perform the Kyoto Encyclopedia of Genes and Genomes (KEGG) analyses in order to identify the biological features of DEGs associated with biological functions as well as elucidate the functional annotation and pathway enrichment analysis to investigate the biological pathways of DEGs in each functional interaction module (criterion: FDR &#x3c;0.05 and <italic>p</italic>-value &#x3c;0.05).</p>
</sec>
<sec id="s2-8">
<title>Protein-protein interaction network of significant DEGs</title>
<p>Each biological system within a cell is controlled by proteins. Some proteins do their work on their own, but most of them rely on their interactions with others to perform their biological functions. To understand protein functions and the biological characteristics of the proteins, Protein-Protein Interaction is a must. PPI data became more available through high-throughput technologies and made it possible to visualize PPI data as networks, i.e., PPI networks.</p>
<p>In this study, all filtered DEGs were used to construct the PPI network based on the Search Tool for the Retrieval of Interacting Genes/Proteins (STRING) (<xref ref-type="bibr" rid="B19">Szklarczyk et al., 2018</xref>). Here, the species was homo and the score criterion was 0.75. By using cytoHubba, a Cytoscape plug-in, finally selected the hub genes and constructed a network among those candidate genes.</p>
</sec>
</sec>
<sec sec-type="results|discussion" id="s3">
<title>Results and discussion</title>
<sec id="s3-1">
<title>Differentially expressed genes</title>
<p>This table was made with the overlapping samples (323), where various clinical information (patient&#x2019;s age, gender, different stages information, and vital status) were concerned (<xref ref-type="table" rid="T1">Table 1</xref>).</p>
<table-wrap id="T1" position="float">
<label>TABLE 1</label>
<caption>
<p>Clinical parameters of colon cancer patients.</p>
</caption>
<table>
<thead valign="top">
<tr>
<th align="right">Group</th>
<th align="left">Subgroup</th>
<th align="center">Frequency</th>
<th align="center">Percent</th>
</tr>
</thead>
<tbody valign="top">
<tr>
<td rowspan="2" align="right">Age</td>
<td align="left">&#x3c;60</td>
<td align="center">101</td>
<td align="center">31.2</td>
</tr>
<tr>
<td align="left">&#x3e; &#x3d; 60</td>
<td align="center">222</td>
<td align="center">68.7</td>
</tr>
<tr>
<td rowspan="2" align="right">Gender</td>
<td align="left">Male</td>
<td align="center">173</td>
<td align="center">53.5</td>
</tr>
<tr>
<td align="left">Female</td>
<td align="center">150</td>
<td align="center">46.4</td>
</tr>
<tr>
<td rowspan="2" align="right">Histology</td>
<td align="left">Adenocarcinoma</td>
<td align="center">282</td>
<td align="center">87.3</td>
</tr>
<tr>
<td align="left">Mucinous adenocarcinoma</td>
<td align="center">41</td>
<td align="center">12.7</td>
</tr>
<tr>
<td rowspan="4" align="right">Stage</td>
<td align="left">Stage I</td>
<td align="center">49</td>
<td align="center">15.1</td>
</tr>
<tr>
<td align="left">Stage II</td>
<td align="center">132</td>
<td align="center">40.8</td>
</tr>
<tr>
<td align="left">Stage III</td>
<td align="center">87</td>
<td align="center">26.9</td>
</tr>
<tr>
<td align="left">Stage IV</td>
<td align="center">44</td>
<td align="center">13.6</td>
</tr>
<tr>
<td rowspan="2" align="right">T stage</td>
<td align="left">Tis, T1-3</td>
<td align="center">278</td>
<td align="center">86</td>
</tr>
<tr>
<td align="left">T4</td>
<td align="center">45</td>
<td align="center">13.9</td>
</tr>
<tr>
<td rowspan="2" align="right">N stage</td>
<td align="left">N0</td>
<td align="center">194</td>
<td align="center">60</td>
</tr>
<tr>
<td align="left">N1&#x2b;N2</td>
<td align="center">129</td>
<td align="center">39.9</td>
</tr>
<tr>
<td rowspan="2" align="right">M stage</td>
<td align="left">MX, M0</td>
<td align="center">274</td>
<td align="center">84.8</td>
</tr>
<tr>
<td align="left">M1</td>
<td align="center">44</td>
<td align="center">13.6</td>
</tr>
<tr>
<td rowspan="2" align="right">Vital status</td>
<td align="left">Alive</td>
<td align="center">244</td>
<td align="center">75.5</td>
</tr>
<tr>
<td align="left">Dead</td>
<td align="center">79</td>
<td align="center">24.4</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>A total of 312 samples were processed and normalized, where 271 colon cancer samples and 41 normal tissues were present. In the colon cancer samples, there were 45 samples for stage I, 109 for stage II, 80 for stage III, and 37 for stage IV. Finally, the DEGs were identified using adjusted and cut-off criteria. A total of 1,574, 1,607, 1,468, and 1,574 significant DEGs were found between healthy and colon cancer stages I, II, III, and IV, respectively.</p>
</sec>
<sec id="s3-2">
<title>All DEGs and intersection DEGs</title>
<p>There were a total of 2048 DE genes among 4 DE sets. 1,067 common DEGs were revealed through intersect function, including 1,185 upregulated and 863 downregulated DEGs in comparison with controls. And those common DEGs are defined as intersection DEGs. All the DEGs in the four sets were compared with each other by Vennplex (<xref ref-type="fig" rid="F2">Figure 2</xref>) and it is found that downregulated genes were less compared with upregulated genes in each set. There were no contra-regulated DEGs which suggested that the process (molecular function) of colon cancer at all stages is identical.</p>
<fig id="F2" position="float">
<label>FIGURE 2</label>
<caption>
<p>Comparison between 4 DE sets. List1: S1 and Normal, List2: S2 and Normal, List3: S3 and Normal, List4: S4 and Normal; where <italic>0</italic> indicates up-regulated genes, <inline-formula id="inf5">
<mml:math id="m5">
<mml:mrow>
<mml:munder>
<mml:mn>0</mml:mn>
<mml:mo>&#xaf;</mml:mo>
</mml:munder>
</mml:mrow>
</mml:math>
</inline-formula> represents down-regulated genes and <inline-formula id="inf6">
<mml:math id="m6">
<mml:mrow>
<mml:mstyle displaystyle="false" mathcolor="red">
<mml:mn>0</mml:mn>
</mml:mstyle>
</mml:mrow>
</mml:math>
</inline-formula> denotes contra-regulated genes (however there are no contra-regulated genes).</p>
</caption>
<graphic xlink:href="fgene-14-1120185-g002.tif"/>
</fig>
</sec>
<sec id="s3-3">
<title>Filtering DEGs for stage-specific network generation</title>
<p>Before the filtering process, there were 1,574, 1,607, 1,468, and 1,574 genes found for stages I to stage IV respectively. After making the histogram analysis based on the PPI network to filter these genes to get significant DE genes the number decreased in <xref ref-type="table" rid="T2">Table 2</xref>.</p>
<table-wrap id="T2" position="float">
<label>TABLE 2</label>
<caption>
<p>Filtered DEGs.</p>
</caption>
<table>
<thead valign="top">
<tr>
<th align="left">Stage</th>
<th align="left">All DEGs (before filtering)</th>
<th align="left">Filtered DEGs (logFC &#x3e;1.5 and PPI edge confidence &#x3e;0.75)</th>
</tr>
</thead>
<tbody valign="top">
<tr>
<td align="left">Stage I</td>
<td align="char" char=".">1,574</td>
<td align="char" char=".">589</td>
</tr>
<tr>
<td align="left">Stage II</td>
<td align="char" char=".">1,607</td>
<td align="char" char=".">569</td>
</tr>
<tr>
<td align="left">Stage III</td>
<td align="char" char=".">1,468</td>
<td align="char" char=".">484</td>
</tr>
<tr>
<td align="left">Stage IV</td>
<td align="char" char=".">1,574</td>
<td align="char" char=".">534</td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
<sec id="s3-4">
<title>Stage-specific cluster interaction network generation</title>
<p>According to the DEGs detected in each stage of CRC, we constructed four FI networks at each stage, respectively as the figure shows. Where each node represents a differentially expressed gene. Edge represents the interactions between genes. <xref ref-type="fig" rid="F3">Figures 3A&#x2013;D</xref> are the FI network constructed by DEG-stage1, DEG-stage2, DEG-stage3, DEG-stage4, respectively. These DEGs normally formed some modules to conduct their cellular functions. That&#x2019;s why MCL clustering was applied and 17, 15, 16, and 18 modules were found after the clustering at individual stages, respectively. These modules have a higher probability to enrich some functions of colon cancer.</p>
<fig id="F3" position="float">
<label>FIGURE 3</label>
<caption>
<p>FI network for each stage, where each node is a differentially expressed gene and edge represents the interaction between genes. <bold>(A)</bold> is the FI network for stage I, <bold>(B)</bold> represents the FI network built by DEG-Stage II, <bold>(C)</bold> and <bold>(D)</bold> are the FI network for stage III and stage IV.</p>
</caption>
<graphic xlink:href="fgene-14-1120185-g003.tif"/>
</fig>
</sec>
<sec id="s3-5">
<title>Cluster interaction networks</title>
<p>Then the cluster interaction networks are constructed for each stage and adjacent stage. There are 7 cluster interaction networks in total. These networks are constructed based on the FI network of each cancer stage by using the MCL cluster algorithm. Where each cluster represents the functions of biological mechanisms and for the investigation of the complex molecular mechanisms of colon cancer these clusters may play important roles. In order to uncover the associations between genes and functions, the interaction between clusters is constructed, where each cluster is treated as a vertex within the interaction and the number of connections between genes in two corresponding clusters as the edge weight. According to the FI network used here, the weights denote the degree of association between clusters. The wider the edges the more closely biology functions between clusters.</p>
</sec>
<sec id="s3-6">
<title>Functional evolutionary network analyses</title>
<p>Examining the functional modules of functional interaction networks at different stages of cancer can provide insight into the functional evolution of the disease and the strength of associations between genes at different stages. This is because these functional modules may contain overlapping genes, and analyzing them can provide valuable information about the progression of the disease.</p>
<p>
<xref ref-type="fig" rid="F4">Figure 4</xref> shows the pathway interaction network between the four stages. Red nodes or &#x201c;A&#x201d;: DE set I, green nodes or &#x201c;B&#x201d;: DE set II, purple nodes or &#x201c;C&#x201d;: DE set III, blue nodes or &#x201c;D&#x201d;: DE set IV. Thick edges were constructed by overlapping gene counts. This pathway interaction network is a major contribution to the field of pathway enrichment analysis. The network was constructed after doing many analyses to get the significant pathways. This network is the result of combining 7 cluster networks. There are four connected networks and some other pathways among the four stages. In this study, the large network is more significant because all the stages are connected in different ways and the other three are connected to one, two, or three stages. This network is the final functional evolutionary network. In terms of the connections between functional modules between adjacent stages, this study constructs this functional evolution network. Nodes indicate the modules. Edges reflect the connections between modules at colon cancer stages. The overlapping genes between functional clusters at adjacent stages are represented by the network&#x2019;s edges. The more overlapping genes there are between adjacent stages of functional modules, the thicker the edges. In the following sections, this network will be analyzed to investigate the molecular mechanism of colon cancer.</p>
<fig id="F4" position="float">
<label>FIGURE 4</label>
<caption>
<p>The Pathway Interaction Network. The correlation between 4 DE sets of pathways was obtained and expressed in this network. This network was constructed by combining the 7 cluster networks. Each node is a cluster of different stages which contained several numbers of genes and the edges represent the number of related genes. The thick edges were selected by overlapping genes counts between connected nodes. Red nodes or &#x201c;A&#x201d;: DE set I, green nodes or &#x201c;B&#x201d;: DE set II, purple nodes or &#x201c;C&#x201d;: DE set III, blue nodes or &#x201c;D&#x201d;: DE set IV.</p>
</caption>
<graphic xlink:href="fgene-14-1120185-g004.tif"/>
</fig>
</sec>
<sec id="s3-7">
<title>Pathway enrichment between edges of adjacent cluster</title>
<p>In this study, by analyzing the pathway interaction network and doing KEGG pathway enrichment analysis, 15 common pathways were found. They are: Neuroactive ligand-receptor interaction, Glutamatergic synapse, Circadian entrainment, Nicotine addiction, Retrograde endocannabinoid signaling, cAMP signaling pathway, Dopaminergic synapse, Adrenergic signaling in cardiomyocytes, Amphetamine addiction, Bile secretion, Wnt signaling pathway, Dilated cardiomyopathy, Long-term potentiation, Cardiac muscle contraction, Melanogenesis. Among them, cAMP signaling pathway controls essential physiological activities such as metabolism, secretion, calcium homeostasis, muscular contraction, cell destiny, and gene transcription (<xref ref-type="bibr" rid="B14">Cho-Chung et al., 2002</xref>). The neuroactive ligand-receptor interaction pathway is a group of receptors and ligands on the plasma membrane that are linked to intracellular and extracellular signaling pathways and this pathway is associated with prostate cancer (<xref ref-type="bibr" rid="B9">He et al., 2018</xref>). The APC mutant colon cancer cells remain dependent on Wnt and suppress the production of secreted Wnt antagonists epigenetically (<xref ref-type="bibr" rid="B8">He et al., 2005</xref>). Ke yang et al. demonstrated that the IWP inhibitors inhibit the WNT signaling pathway in colon cancer cells by disrupting the WNT ligand (<xref ref-type="bibr" rid="B21">Yang et al., 2016</xref>).</p>
</sec>
<sec id="s3-8">
<title>Pathway enrichment between edges of all stages</title>
<p>For analyzing the different pathways, KEGG pathway analysis was done and about five significantly different pathways were found. Pathways in cancer, Serotonergic synapse, Bile secretion, Hypertrophic cardiomyopathy (HCM), and Dilated cardiomyopathy are significantly different at all stages. The findings pathway namely Serotonergic synapse pathway is a neurotransmitter which is widely distributed in the vertebrate central nervous system, and it serves as a target for many physiologic regulations like modulators of gene transcription, steroids and neurotrophic factors. Ionotropic GluRs and 5-HT3 receptors mediate rapid synaptic transmission <italic>via</italic> serotonergic fibers making direct synaptic connections with GABAergic neurons (<xref ref-type="bibr" rid="B13">Maejima et al., 2013</xref>). There are active transport systems within hepatocytes and cholangiocytes. Hepatocytes secret their bile by secreting conjugate bilirubin, bile salts, cholesterol, phospholipids, and water into their canaliculi (<xref ref-type="bibr" rid="B10">Hundt et al., 2022</xref>). A primary myocardial disorder associated with the autosomal dominant pattern of inheritance is hypertrophic cardiomyopathy (HCM), which can be distinguished from other cardiovascular disorders primarily by the presence of myocyte hypertrophy, fibrillar disarrays, and interstitial fibrosis as histological characteristics. Dilated cardiomyopathy (DCM) is a heart muscle illness described by dilation and impaired contraction of the left or both ventricles that is responsible for progressive heart failure and sudden cardiac death from ventricular arrhythmia.</p>
</sec>
<sec id="s3-9">
<title>Heatmap analysis for pathways</title>
<p>To identify the significant pathways, heatmap analysis was done based on the <italic>p</italic>-values of each stage. <xref ref-type="table" rid="T3">Table 3</xref> shows a total of 15 pathways with the <italic>p</italic>-values of each stage.</p>
<table-wrap id="T3" position="float">
<label>TABLE 3</label>
<caption>
<p>Pathways table.</p>
</caption>
<table>
<thead valign="top">
<tr>
<th align="left">Pathways</th>
<th align="left">S1 <italic>p</italic>-value</th>
<th align="left">S2 <italic>p</italic>-value</th>
<th align="left">S3 <italic>p</italic>-value</th>
<th align="left">S4 <italic>p</italic>-value</th>
</tr>
</thead>
<tbody valign="top">
<tr>
<td align="left">Neuroactive ligand-receptor interaction</td>
<td align="left">2.80E-19</td>
<td align="left">2.60E-11</td>
<td align="left">2.90E-13</td>
<td align="left">1.50E-02</td>
</tr>
<tr>
<td align="left">Glutamatergic synapse</td>
<td align="left">5.40E-14</td>
<td align="left">1.40E-05</td>
<td align="left">2.20E-05</td>
<td align="left">3.30E-04</td>
</tr>
<tr>
<td align="left">Circadian entrainment</td>
<td align="left">4.50E-10</td>
<td align="left">3.20E-06</td>
<td align="left">5.40E-05</td>
<td align="left">1.40E-04</td>
</tr>
<tr>
<td align="left">Nicotine addiction</td>
<td align="left">3.90E-09</td>
<td align="left">4.80E-11</td>
<td align="left">1.90E-10</td>
<td align="left">5.30E-08</td>
</tr>
<tr>
<td align="left">Retrograde endocannabinoid signaling</td>
<td align="left">1.00E-08</td>
<td align="left">5.30E-06</td>
<td align="left">6.20E-04</td>
<td align="left">1.70E-02</td>
</tr>
<tr>
<td align="left">cAMP signaling pathway</td>
<td align="left">3.60E-07</td>
<td align="left">1.30E-06</td>
<td align="left">7.80E-07</td>
<td align="left">5.80E-04</td>
</tr>
<tr>
<td align="left">Dopaminergic synapse</td>
<td align="left">1.40E-06</td>
<td align="left">3.70E-05</td>
<td align="left">2.10E-03</td>
<td align="left">5.60E-04</td>
</tr>
<tr>
<td align="left">Adrenergic signaling in cardiomyocytes</td>
<td align="left">3.40E-06</td>
<td align="left">6.60E-05</td>
<td align="left">5.60E-04</td>
<td align="left">7.90E-04</td>
</tr>
<tr>
<td align="left">Amphetamine addiction</td>
<td align="left">4.60E-05</td>
<td align="left">1.80E-06</td>
<td align="left">5.90E-05</td>
<td align="left">2.40E-05</td>
</tr>
<tr>
<td align="left">Bile secretion</td>
<td align="left">6.40E-05</td>
<td align="left">2.70E-05</td>
<td align="left">7.30E-04</td>
<td align="left">5.90E-03</td>
</tr>
<tr>
<td align="left">Wnt signaling pathway</td>
<td align="left">9.30E-05</td>
<td align="left">1.00E-05</td>
<td align="left">5.60E-04</td>
<td align="left">5.10E-07</td>
</tr>
<tr>
<td align="left">Dilated cardiomyopathy</td>
<td align="left">2.60E-04</td>
<td align="left">1.10E-06</td>
<td align="left">2.20E-06</td>
<td align="left">4.80E-06</td>
</tr>
<tr>
<td align="left">Long-term potentiation</td>
<td align="left">3.30E-04</td>
<td align="left">1.90E-04</td>
<td align="left">4.90E-03</td>
<td align="left">4.00E-04</td>
</tr>
<tr>
<td align="left">Cardiac muscle contraction</td>
<td align="left">7.20E-04</td>
<td align="left">4.80E-06</td>
<td align="left">1.20E-05</td>
<td align="left">4.50E-05</td>
</tr>
<tr>
<td align="left">Melanogenesis</td>
<td align="left">8.40E-04</td>
<td align="left">2.90E-04</td>
<td align="left">2.00E-02</td>
<td align="left">1.90E-03</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>To investigate the significant biological functions of DEGs, heatmap analysis was done on each stage (<xref ref-type="fig" rid="F5">Figure 5</xref>) and among those pathways, some are significantly different such as neuroactive ligand-receptor interaction, Glutamatergic synapse, Circadian entrainment, and Nicotine addiction have low <italic>p</italic>-value (from <xref ref-type="table" rid="T3">Table 3</xref>) and a large number of genes.</p>
<fig id="F5" position="float">
<label>FIGURE 5</label>
<caption>
<p>Pathway analysis by clustered heatmap analysis of four stages where pathways were selected by the KEGG pathway enrichment analysis. The functional components identified from the KEGG pathway enrichment analysis are presented row-wise (right side) and each stage is presented column-wise (bottom). The color intensities indicate the enrichment score of each KEGG pathway with a color gradient that moves from light yellow to maroon. The dendrogram indicates the similarity between pathways as well as stages.</p>
</caption>
<graphic xlink:href="fgene-14-1120185-g005.tif"/>
</fig>
<p>From <xref ref-type="table" rid="T3">Table 3</xref>; <xref ref-type="fig" rid="F5">Figure 5</xref>, it is clear that the <italic>p</italic>-values of neuroactive ligand-receptor interaction, Glutamatergic synapse, Circadian entrainment, and Nicotine addiction change and increase from the initial stage to later stage and <xref ref-type="fig" rid="F5">Figure 5</xref> also shows that the color changes from light yellow to marron and it is said for heatmap analysis is that bright color indicates high activity and dark color is <italic>vice versa</italic>.</p>
<p>Fang et al. and Liu et al. has shown that the neuroactive ligand-receptor interaction signaling pathway is linked to bladder cancer and renal cell carcinoma progression (<xref ref-type="bibr" rid="B5">Fang et al., 2013</xref>; <xref ref-type="bibr" rid="B12">Liu et al., 2015</xref>). It is believed that glutamatergic synapse pathways play a crucial role in a large variety of normal physiological functions due to their links to many other neurotransmitter pathways and neurodevelopmental disorders and injuries are strongly associated with glutamate dysfunction (<xref ref-type="bibr" rid="B6">Glutamatergic Synapse Pathway, 2022</xref>). Circadian entrainment includes retinal sensitivity as well as circadian variations in the retina that contribute to the regulation of retinal diseases and these circadian disorders are related to entrainment deficits (<xref ref-type="bibr" rid="B7">Golombek and Rosenstein, 2010</xref>). It is shown in various studies that Nicotine addiction increased approximately 50% chances of colon cancer because smoking has been associated with adenomatous polyps (<xref ref-type="bibr" rid="B18">Slattery et al., 1997</xref>). The genes encoding these pathways may be important in the molecular development of these pathways since they are strongly associated with function.</p>
</sec>
<sec id="s3-10">
<title>PPI network of significant DEGs</title>
<p>A Protein-Protein Interaction network (PPI) is a visual framework for better understanding protein functional organization. When the iteration between genes was found by above analysis, the PPI network was also constructed to check their relationship because the PPI network is the widely used network to see whether there is a strong relationship or not. In the PPI network, all the significant DE genes were investigated and a PPI network was built. Among these DEGs, some DEGs also appeared in functional interaction network and some DEGs are also known colon cancer related genes, such as CXCL11, ADH1B, PYY, SLC17A7, and so on. The network involved 794 nodes, and 7,153 edges (<xref ref-type="fig" rid="F6">Figure 6</xref>).</p>
<fig id="F6" position="float">
<label>FIGURE 6</label>
<caption>
<p>Protein-Protein Interaction (PPI) Network of intersection DEGs.</p>
</caption>
<graphic xlink:href="fgene-14-1120185-g006.tif"/>
</fig>
</sec>
<sec id="s3-11">
<title>Identifying stage related hub genes</title>
<p>In this study, the top 10 hub genes (INS, SNAP25, GRIA2, SST, GCG, PVALB, SLC17A7, SLC32A1, SLC17A6, and NPY) were identified according to the degree of connectivity of the DEGs and arranged it in a descending order to find the highest degree of connectivity (<xref ref-type="table" rid="T4">Table 4</xref>). <xref ref-type="fig" rid="F7">Figure 7</xref> showed the network between the top 10 hub genes.</p>
<table-wrap id="T4" position="float">
<label>TABLE 4</label>
<caption>
<p>Top 10 hub genes with higher degree of connectivity and betweenness value.</p>
</caption>
<table>
<thead valign="top">
<tr>
<th align="center">Gene</th>
<th align="center">Degree of connectivity</th>
<th align="center">Betweenness value</th>
</tr>
</thead>
<tbody valign="top">
<tr>
<td align="center">INS</td>
<td align="center">147</td>
<td align="center">82257.16327</td>
</tr>
<tr>
<td align="center">SNAP25</td>
<td align="center">97</td>
<td align="center">10475.22985</td>
</tr>
<tr>
<td align="center">GRIA2</td>
<td align="center">96</td>
<td align="center">7915.71159</td>
</tr>
<tr>
<td align="center">SST</td>
<td align="center">90</td>
<td align="center">11508.22796</td>
</tr>
<tr>
<td align="center">GCG</td>
<td align="center">89</td>
<td align="center">12209.61485</td>
</tr>
<tr>
<td align="center">PVALB</td>
<td align="center">85</td>
<td align="center">17121.22796</td>
</tr>
<tr>
<td align="center">SLC17A7</td>
<td align="center">84</td>
<td align="center">7167.58773</td>
</tr>
<tr>
<td align="center">SLC32A1</td>
<td align="center">83</td>
<td align="center">5748.84911</td>
</tr>
<tr>
<td align="center">SLC17A6</td>
<td align="center">83</td>
<td align="center">8895.32674</td>
</tr>
<tr>
<td align="center">NPY</td>
<td align="center">81</td>
<td align="center">9424.62421</td>
</tr>
</tbody>
</table>
</table-wrap>
<fig id="F7" position="float">
<label>FIGURE 7</label>
<caption>
<p>This sub-network shows the connection between hub genes. Red color denotes the degree of connectivity is highest and yellow color means the lowest degree of connectivity and the color changes from red to yellow means the degree of connectivity decreases.</p>
</caption>
<graphic xlink:href="fgene-14-1120185-g007.tif"/>
</fig>
</sec>
</sec>
<sec sec-type="conclusion" id="s4">
<title>Conclusion</title>
<p>Gene expression profiling of four CRC stages and healthy colorectal tissue was investigated in this study to learn more about colon cancer mechanisms and stage-related genes of colon cancer. After selecting the DEGs for four stages FI network was constructed and an MCL graph clustering algorithm was performed on the FI network to extract some modules of colon cancer. Then we perform the cluster interaction network to get more specific and significant biological functions. After that, pathway enrichment analysis was done with the MCL modules and in order to improve the analysis, a functional evolutionary network was constructed which described the relationships among pathways at each stage. Finally, the PPI network was constructed using the strong common genes among 4 DE sets. Based on the degree of connectivity, 10 hub genes were chosen as the potential colon cancer stage-related genes and those are INS, SNAP25, GRIA2, SST, GCG, PVALB, SLC17A7, SLC32A1, SLC17A6, and NPY.</p>
<p>Comparing the relationships between the same pathways and different pathways in neighboring DE sets was a very useful way to analyze the staged biological functions of colon cancer. KEGG pathway enrichment analysis at adjacent stages showed that Pathways in cancer, Serotonergic synapse, Bile secretion, Hypertrophic cardiomyopathy (HCM), and Dilated cardiomyopathy are significantly different at all stages. And neuroactive ligand-receptor interaction, Glutamatergic synapse, Circadian entrainment, and Nicotine addiction are the significant pathways among each stage. Overall, this study has identified novel candidate biomarkers and pathways associated with colon cancer.</p>
<p>In conclusion, 10 potential colon cancer stage-related genes, four significant pathways, and some biological information were discovered on the disease in this study. These results might provide some significant information about the stages and the genes would serve as staged biomarkers of colon cancer. Though the current findings provided valuable information for early detection and prevention, as well as a viable therapeutic target for CRC. However, there are certain limitations to the research: i) No mRNAs or miRNAs were used with TCGA. ii) In order to verify the accuracy of the results, some biological experiments will still be necessary.</p>
</sec>
</body>
<back>
<sec sec-type="data-availability" id="s5">
<title>Data availability statement</title>
<p>The original contributions presented in the study are included in the article/Supplementary Material, further inquiries can be directed to the corresponding author.</p>
</sec>
<sec id="s6">
<title>Author contributions</title>
<p>BC initialized this study. NC and BC discussed many times to finalize the work plan. NC and AS prepared the datasets. XS and BC gave suggestions many times to modify this study. NC conducted the numerical experiments and drafted the manuscript. Everyone read the manuscript and revised it, and agreed with the final version.</p>
</sec>
<sec id="s7">
<title>Funding</title>
<p>This work was supported by the National Natural Science Foundation of China under Grant No. 61972320 and the National Key R&#x26;D Program of China (No. 2021YFA1000400).</p>
</sec>
<sec sec-type="COI-statement" id="s8">
<title>Conflict of interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec sec-type="disclaimer" id="s9">
<title>Publisher&#x2019;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
<ref-list>
<title>References</title>
<ref id="B1">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Cai</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Chen</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Yi</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Daimon</surname>
<given-names>C. M.</given-names>
</name>
<name>
<surname>Boyle</surname>
<given-names>J. P.</given-names>
</name>
<name>
<surname>Peers</surname>
<given-names>C.</given-names>
</name>
<etal/>
</person-group> (<year>2013</year>). <article-title>VennPlex &#x2013; a novel venn diagram Program for comparing and visualizing datasets with differentially regulated datapoints</article-title>. <source>PLoS One</source> <volume>8</volume> (<issue>1</issue>), <fpage>e53388</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pone.0053388</pub-id>
</citation>
</ref>
<ref id="B2">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Chen</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Shang</surname>
<given-names>X.</given-names>
</name>
</person-group> (<year>2020</year>). <article-title>Identification and analysis of genes involved in stages of colon cancer</article-title>. <source>Lect. Notes Comput. Sci. Incl. Subser. Lect. Notes Artif. Intell. Lect. Notes Bioinforma.</source> <volume>12464</volume>, <fpage>161</fpage>&#x2013;<lpage>172</lpage>. <pub-id pub-id-type="doi">10.1007/978-3-030-60802-6_15</pub-id>
</citation>
</ref>
<ref id="B3">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Chen</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Yang</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Gao</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Jiang</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Shang</surname>
<given-names>X.</given-names>
</name>
</person-group> (<year>2020</year>). <article-title>A functional network construction method to interpret the pathological process of colorectal cancer</article-title>. <source>Int. J. Data Min. Bioinform.</source> <volume>23</volume> (<issue>3</issue>), <fpage>251</fpage>&#x2013;<lpage>264</lpage>. <pub-id pub-id-type="doi">10.1504/IJDMB.2020.107879</pub-id>
</citation>
</ref>
<ref id="B14">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Cho-Chung</surname>
<given-names>Y. S.</given-names>
</name>
<name>
<surname>Nesterova</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Becker</surname>
<given-names>K. G.</given-names>
</name>
<name>
<surname>Srivastava</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Park</surname>
<given-names>Y. G.</given-names>
</name>
<name>
<surname>Lee</surname>
<given-names>Y. N.</given-names>
</name>
<etal/>
</person-group> (<year>2002</year>). <article-title>Dissecting the circuitry of protein kinase A and cAMP signaling in cancer genesis: Antisense, microarray, gene overexpression, and transcription factor decoy</article-title>. <source>Ann. N. Y. Acad. Sci.</source> <volume>36</volume>, <fpage>22</fpage>&#x2013;<lpage>36</lpage>. <pub-id pub-id-type="doi">10.1111/j.1749-6632.2002.tb04324.x</pub-id>
</citation>
</ref>
<ref id="B4">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Dennis</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Sherman</surname>
<given-names>B. T.</given-names>
</name>
<name>
<surname>Hosack</surname>
<given-names>D. A.</given-names>
</name>
<name>
<surname>Yang</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Gao</surname>
<given-names>W.</given-names>
</name>
<name>
<surname>Lane</surname>
<given-names>H. C.</given-names>
</name>
<etal/>
</person-group> (<year>2003</year>). <article-title>David: Database for annotation, visualization, and integrated discovery</article-title>. <source>Genome Biol.</source> <volume>4</volume> (<issue>5</issue>), <fpage>P3</fpage>. <pub-id pub-id-type="doi">10.1186/gb-2003-4-5-p3</pub-id>
</citation>
</ref>
<ref id="B5">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Fang</surname>
<given-names>Z.-Q.</given-names>
</name>
<name>
<surname>Zang</surname>
<given-names>W. D.</given-names>
</name>
<name>
<surname>Chen</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Ye</surname>
<given-names>B. W.</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>X. W.</given-names>
</name>
<name>
<surname>Yi</surname>
<given-names>S. H.</given-names>
</name>
<etal/>
</person-group> (<year>2013</year>). <article-title>Gene expression profile and enrichment pathways in different stages of bladder cancer</article-title>. <source>Genet. Mol. Res.</source> <volume>12</volume> (<issue>2</issue>), <fpage>1479</fpage>&#x2013;<lpage>1489</lpage>. <pub-id pub-id-type="doi">10.4238/2013.May.6.1</pub-id>
</citation>
</ref>
<ref id="B6">
<citation citation-type="other">&#x201c;<collab>Glutamatergic Synapse Pathway</collab>, &#x201d; <comment>CD Creat. Diagnositics</comment>, <year>2022</year>.</citation>
</ref>
<ref id="B7">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Golombek</surname>
<given-names>D. A.</given-names>
</name>
<name>
<surname>Rosenstein</surname>
<given-names>R. E.</given-names>
</name>
</person-group> (<year>2010</year>). <article-title>Physiology of circadian entrainment</article-title>. <source>Physiol. Rev.</source> <volume>90</volume> (<issue>3</issue>), <fpage>1063</fpage>&#x2013;<lpage>1102</lpage>. <pub-id pub-id-type="doi">10.1152/physrev.00009.2009</pub-id>
</citation>
</ref>
<ref id="B8">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>He</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Reguart</surname>
<given-names>N.</given-names>
</name>
<name>
<surname>You</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Mazieres</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Xu</surname>
<given-names>Z.</given-names>
</name>
<name>
<surname>Lee</surname>
<given-names>A. Y.</given-names>
</name>
<etal/>
</person-group> (<year>2005</year>). <article-title>Blockade of Wnt-1 signaling induces apoptosis in human colorectal cancer cells containing downstream mutations</article-title>. <source>Oncogene</source> <volume>24</volume> (<issue>18</issue>), <fpage>3054</fpage>&#x2013;<lpage>3058</lpage>. <pub-id pub-id-type="doi">10.1038/sj.onc.1208511</pub-id>
</citation>
</ref>
<ref id="B9">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>He</surname>
<given-names>Z.</given-names>
</name>
<name>
<surname>Tang</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Lu</surname>
<given-names>Z.</given-names>
</name>
<name>
<surname>Huang</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Lei</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Li</surname>
<given-names>Z.</given-names>
</name>
<etal/>
</person-group> (<year>2018</year>). <article-title>Analysis of differentially expressed genes, clinical value and biological pathways in prostate cancer</article-title>. <source>Am. J. Transl. Res.</source> <volume>10</volume> (<issue>5</issue>), <fpage>1444</fpage>&#x2013;<lpage>1456</lpage>.</citation>
</ref>
<ref id="B10">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Hundt</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Basit</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>John</surname>
<given-names>S.</given-names>
</name>
</person-group> (<year>2022</year>). <source>Physiology, bile secretion</source>. <publisher-loc>Treasure Island, FL</publisher-loc>: <publisher-name>Treasure Island FL</publisher-name>.</citation>
</ref>
<ref id="B11">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Kuipers</surname>
<given-names>E. J.</given-names>
</name>
<name>
<surname>Grady</surname>
<given-names>W. M.</given-names>
</name>
<name>
<surname>Lieberman</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Seufferlein</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Sung</surname>
<given-names>J. J.</given-names>
</name>
<name>
<surname>Boelens</surname>
<given-names>P. G.</given-names>
</name>
<etal/>
</person-group> (<year>2015</year>). <article-title>Colorectal cancer</article-title>. <source>Nat. Rev. Dis. Prim.</source> <volume>1</volume>, <fpage>15065</fpage>. <pub-id pub-id-type="doi">10.1038/nrdp.2015.65</pub-id>
</citation>
</ref>
<ref id="B12">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Liu</surname>
<given-names>X.</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Sun</surname>
<given-names>G.</given-names>
</name>
</person-group> (<year>2015</year>). <article-title>Identification of key genes and pathways in renal cell carcinoma through expression profiling data</article-title>, <source>Kidney Blood Press Res</source> <volume>40</volume>, <fpage>288</fpage>&#x2013;<lpage>297</lpage>. <pub-id pub-id-type="doi">10.1159/000368504</pub-id>
</citation>
</ref>
<ref id="B13">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Maejima</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Masseck</surname>
<given-names>O.</given-names>
</name>
<name>
<surname>Mark</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Herlitze</surname>
<given-names>S.</given-names>
</name>
</person-group> (<year>2013</year>). <article-title>Modulation of firing and synaptic transmission of serotonergic neurons by intrinsic G protein-coupled receptors and ion channels</article-title>. <source>Front. Integr. Neurosci.</source> <volume>7</volume>, <fpage>40</fpage>. <pub-id pub-id-type="doi">10.3389/fnint.2013.00040</pub-id>
</citation>
</ref>
<ref id="B15">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Robinson</surname>
<given-names>M. D.</given-names>
</name>
<name>
<surname>McCarthy</surname>
<given-names>D. J.</given-names>
</name>
<name>
<surname>Smyth</surname>
<given-names>G. K.</given-names>
</name>
</person-group> (<year>2010</year>). <article-title>edgeR: a Bioconductor package for differential expression analysis of digital gene expression data</article-title>. <source>Bioinformatics</source> <volume>26</volume> (<issue>1</issue>), <fpage>139</fpage>&#x2013;<lpage>140</lpage>. <pub-id pub-id-type="doi">10.1093/bioinformatics/btp616</pub-id>
</citation>
</ref>
<ref id="B16">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Robinson</surname>
<given-names>M. D.</given-names>
</name>
<name>
<surname>Oshlack</surname>
<given-names>A.</given-names>
</name>
</person-group> (<year>2010</year>). <article-title>A scaling normalization method for differential expression analysis of RNA-seq data</article-title>. <source>Genome Biol.</source> <volume>11</volume> (<issue>3</issue>), <fpage>R25</fpage>. <pub-id pub-id-type="doi">10.1186/gb-2010-11-3-r25</pub-id>
</citation>
</ref>
<ref id="B17">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Shannon</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Markiel</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Ozier</surname>
<given-names>O.</given-names>
</name>
<name>
<surname>Baliga</surname>
<given-names>N. S.</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>J. T.</given-names>
</name>
<name>
<surname>Ramage</surname>
<given-names>D.</given-names>
</name>
<etal/>
</person-group> (<year>2003</year>). <article-title>Cytoscape: A software environment for integrated models of biomolecular interaction networks</article-title>. <source>Genome Res.</source> <volume>13</volume> (<issue>11</issue>), <fpage>2498</fpage>&#x2013;<lpage>2504</lpage>. <pub-id pub-id-type="doi">10.1101/gr.1239303</pub-id>
</citation>
</ref>
<ref id="B18">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Slattery</surname>
<given-names>M. L.</given-names>
</name>
<name>
<surname>Potter</surname>
<given-names>J. D.</given-names>
</name>
<name>
<surname>Friedman</surname>
<given-names>G. D.</given-names>
</name>
<name>
<surname>Ma</surname>
<given-names>K. N.</given-names>
</name>
<name>
<surname>Edwards</surname>
<given-names>S.</given-names>
</name>
</person-group> (<year>1997</year>). <article-title>Tobacco use and colon cancer</article-title>. <source>Int. J. Cancer</source> <volume>70</volume> (<issue>3</issue>), <fpage>259</fpage>&#x2013;<lpage>264</lpage>. <pub-id pub-id-type="doi">10.1002/(SICI)1097-0215(19970127)70:3&#x3c;259::AID-IJC2&#x3e;3.0.CO;2-W</pub-id>
</citation>
</ref>
<ref id="B19">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Szklarczyk</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Gable</surname>
<given-names>A. L.</given-names>
</name>
<name>
<surname>Lyon</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Junge</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Wyder</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Huerta-Cepas</surname>
<given-names>J.</given-names>
</name>
<etal/>
</person-group> (<year>2018</year>). <article-title>STRING v11: Protein&#x2013;protein association networks with increased coverage, supporting functional discovery in genome-wide experimental datasets</article-title>. <source>Nucleic Acids Res.</source> <volume>47</volume> (<issue>D1</issue>), <fpage>D607</fpage>&#x2013;<lpage>D613</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gky1131</pub-id>
</citation>
</ref>
<ref id="B20">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Wu</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Dawson</surname>
<given-names>E.</given-names>
</name>
<name>
<surname>Duong</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Haw</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Stein</surname>
<given-names>L.</given-names>
</name>
</person-group> (<year>2014</year>). <article-title>ReactomeFIViz: A Cytoscape app for pathway and network-based data analysis</article-title>. <source>F1000Research</source> <volume>3</volume>, <fpage>146</fpage>. <pub-id pub-id-type="doi">10.12688/f1000research.4431.2</pub-id>
</citation>
</ref>
<ref id="B21">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Yang</surname>
<given-names>K.</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>X.</given-names>
</name>
<name>
<surname>Zhang</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>Z.</given-names>
</name>
<name>
<surname>Nan</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Li</surname>
<given-names>Y.</given-names>
</name>
<etal/>
</person-group> (<year>2016</year>). <article-title>The evolving roles of canonical WNT signaling in stem cells and tumorigenesis: Implications in targeted cancer therapies</article-title>. <source>Lab. Investig.</source> <volume>96</volume> (<issue>2</issue>), <fpage>116</fpage>&#x2013;<lpage>136</lpage>. <pub-id pub-id-type="doi">10.1038/labinvest.2015.144</pub-id>
</citation>
</ref>
</ref-list>
</back>
</article>