<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article article-type="research-article" dtd-version="2.3" xml:lang="EN" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Syst. Biol.</journal-id>
<journal-title>Frontiers in Systems Biology</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Syst. Biol.</abbrev-journal-title>
<issn pub-type="epub">2674-0702</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="publisher-id">1099413</article-id>
<article-id pub-id-type="doi">10.3389/fsysb.2023.1099413</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Systems Biology</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>MEMMAL: A tool for expanding large-scale mechanistic models with machine learned associations and big datasets</article-title>
<alt-title alt-title-type="left-running-head">Erdem and Birtwistle</alt-title>
<alt-title alt-title-type="right-running-head">
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/fsysb.2023.1099413">10.3389/fsysb.2023.1099413</ext-link>
</alt-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name>
<surname>Erdem</surname>
<given-names>Cemal</given-names>
</name>
<xref ref-type="aff" rid="aff1">
<sup>1</sup>
</xref>
<xref ref-type="corresp" rid="c001">&#x2a;</xref>
<xref ref-type="fn" rid="fn1">
<sup>&#x2020;</sup>
</xref>
<uri xlink:href="https://loop.frontiersin.org/people/1958050/overview"/>
</contrib>
<contrib contrib-type="author" corresp="yes">
<name>
<surname>Birtwistle</surname>
<given-names>Marc R.</given-names>
</name>
<xref ref-type="aff" rid="aff1">
<sup>1</sup>
</xref>
<xref ref-type="aff" rid="aff2">
<sup>2</sup>
</xref>
<xref ref-type="corresp" rid="c001">&#x2a;</xref>
<xref ref-type="fn" rid="fn1">
<sup>&#x2020;</sup>
</xref>
<uri xlink:href="https://loop.frontiersin.org/people/2197871/overview"/>
</contrib>
</contrib-group>
<aff id="aff1">
<sup>1</sup>
<institution>Department of Chemical and Biomolecular Engineering</institution>, <institution>Clemson University</institution>, <addr-line>Clemson</addr-line>, <addr-line>SC</addr-line>, <country>United States</country>
</aff>
<aff id="aff2">
<sup>2</sup>
<institution>Department of Bioengineering</institution>, <institution>Clemson University</institution>, <addr-line>Clemson</addr-line>, <addr-line>SC</addr-line>, <country>United States</country>
</aff>
<author-notes>
<fn fn-type="edited-by">
<p>
<bold>Edited by:</bold> <ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/44682/overview">Yoram Vodovotz</ext-link>, University of Pittsburgh, United States</p>
</fn>
<fn fn-type="edited-by">
<p>
<bold>Reviewed by:</bold> <ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/1076066/overview">Adrian Buganza Tepole</ext-link>, Purdue University, United States</p>
<p>
<ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/1343444/overview">Rahuman S. Malik-Sheriff</ext-link>, European Bioinformatics Institute (EMBL-EBI), United Kingdom</p>
</fn>
<corresp id="c001">&#x2a;Correspondence: Cemal Erdem, <email>cemalerdem@gmail.com</email>; Marc R. Birtwistle, <email>mbirtwi@clemson.edu</email>
</corresp>
<fn fn-type="equal" id="fn1">
<label>
<sup>&#x2020;</sup>
</label>
<p>These authors share last authorship</p>
</fn>
<fn fn-type="other">
<p>This article was submitted to Multiscale Mechanistic Modeling, a section of the journal Frontiers in Systems Biology</p>
</fn>
</author-notes>
<pub-date pub-type="epub">
<day>09</day>
<month>03</month>
<year>2023</year>
</pub-date>
<pub-date pub-type="collection">
<year>2023</year>
</pub-date>
<volume>3</volume>
<elocation-id>1099413</elocation-id>
<history>
<date date-type="received">
<day>15</day>
<month>11</month>
<year>2022</year>
</date>
<date date-type="accepted">
<day>06</day>
<month>02</month>
<year>2023</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#xa9; 2023 Erdem and Birtwistle.</copyright-statement>
<copyright-year>2023</copyright-year>
<copyright-holder>Erdem and Birtwistle</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/">
<p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p>
</license>
</permissions>
<abstract>
<p>Computational models that can explain and predict complex sub-cellular, cellular, and tissue-level drug response mechanisms could speed drug discovery and prioritize patient-specific treatments (i.e., precision medicine). Some models are mechanistic with detailed equations describing known (or supposed) physicochemical processes, while some are statistical or machine learning-based approaches, that explain datasets but have no mechanistic or causal guarantees. These two types of modeling are rarely combined, missing the opportunity to explore possibly causal but data-driven new knowledge while explaining what is already known. Here, we explore combining machine learned associations with mechanistic models to develop computational models that could more fully represent cellular behavior. In this proposed MEMMAL (MEchanistic Modeling with MAchine Learning) framework, machine learning/statistical models built using omics datasets provide predictions for new interactions between genes and proteins where there is physicochemical uncertainty. These interactions are used as a basis for new reactions in mechanistic models. As a test case, we focused on incorporating novel IFN&#x3b3;/PD-L1 related associations into a large-scale mechanistic model for cell proliferation and death to better recapitulate the recently released NIH LINCS Consortium MCF10A dataset and enable description of the cellular response to checkpoint inhibitor immunotherapies. This work is a template for combining big-data-inferred interactions with mechanistic models, which could be more broadly applicable for building multi-scale precision medicine and whole cell models.</p>
</abstract>
<kwd-group>
<kwd>mechanistic modeling</kwd>
<kwd>machine learning</kwd>
<kwd>SBML</kwd>
<kwd>multi-omics</kwd>
<kwd>data integration</kwd>
</kwd-group>
</article-meta>
</front>
<body>
<sec id="s1">
<title>Introduction</title>
<p>The molecular signaling mechanisms of cancer cells are highly heterogenous, leading to treatment resistance and recurrence. Thus, the need for personalized interventions to block tumor growth is high. The traditional drug discovery pipeline is comprised of extensive trial-and-error experiments, testing thousands of chemicals, refining their structure for safety and toxicity, and administering years of clinical trials. This burden might be reduced by understanding the underlying molecular mechanisms with the help of computational models (<xref ref-type="bibr" rid="B46">Yu et al., 2018</xref>; <xref ref-type="bibr" rid="B29">Saez-Rodriguez and Bl&#xfc;thgen, 2020</xref>).</p>
<p>Computational tools and models are becoming indispensable in medical research, where a cycle of experimentation and computation is used to learn about and test new hypotheses. The models guide experimental hypothesis generation, and experimental observations enable fine-tuning computational models to understand the biological phenomena. Owing to the advances in wet-lab experimental techniques and tools, &#x201c;Big Data&#x201d; repositories become more prominent each year. The knowledge base of these databases includes genomics, proteomics, epigenomics, and clinical information (<xref ref-type="bibr" rid="B3">Barrett et al., 2012</xref>; <xref ref-type="bibr" rid="B35">Uhlen et al., 2015</xref>; <xref ref-type="bibr" rid="B32">Subramanian et al., 2017</xref>; <xref ref-type="bibr" rid="B14">Hoadley et al., 2018</xref>; <xref ref-type="bibr" rid="B39">Wishart et al., 2018</xref>; <xref ref-type="bibr" rid="B27">Nusinow et al., 2020</xref>). To understand the underlying biological facts, analysis of the wealth of the aforementioned big datasets should become more practical and go beyond context-dependent and scope-limited biological events.</p>
<p>Building computational models that explain and predict such highly heterogenous and complex cellular responses is no easy task. The popular mechanistic models are sets of detailed equations describing curated knowledge of what is happening within the cells. Such models (<xref ref-type="bibr" rid="B4">Bouhaddou et al., 2018</xref>; <xref ref-type="bibr" rid="B8">Fr&#xf6;hlich et al., 2018</xref>; <xref ref-type="bibr" rid="B26">M&#xfc;nzner et al., 2019</xref>) are usually small in scale: tens of equations and 10s&#x2013;100s of model species (<xref ref-type="fig" rid="F1">Figure 1</xref>). Another popular class is machine learning based models, which are data-driven, descriptive, and mostly large-scale (genome-wide or exome-wide) (<xref ref-type="bibr" rid="B22">Malta et al., 2018</xref>; <xref ref-type="bibr" rid="B40">Wong and Yip, 2018</xref>; <xref ref-type="bibr" rid="B46">Yu et al., 2018</xref>; <xref ref-type="bibr" rid="B45">Yang et al., 2019</xref>). These types of models are generally coined as black-box models because although they perform well in precision/recall metrics, how they do so is blurry (<xref ref-type="fig" rid="F1">Figure 1</xref>). So far in the literature, these two types of models are rarely combined, missing the opportunity to generate new knowledge while explaining what is already known (<xref ref-type="bibr" rid="B2">Baker et al., 2018</xref>).</p>
<fig id="F1" position="float">
<label>FIGURE 1</label>
<caption>
<p>Different computational modeling types of biological data possess a variety of pros and cons and provide an opportunity for model merging. The mechanistic models are mostly curated, usually small-scale, causal networks of signaling pathways. Machine learning models are data-driven, large-scale, and usually correlative associations. Combining these two modes of modeling provides an opportunity for creating larger scale data-informed models to generate novel hypotheses for experimental validation. The merged model would include curated lists of pathway genes (species) as well as genes with new connections inferred <italic>via</italic> machine learning models. The final model structure could represent a collection of overlapping genes (and gene products) and interactions present in both lists.</p>
</caption>
<graphic xlink:href="fsysb-03-1099413-g001.tif"/>
</fig>
<p>Here, we explore a combination of both methods to develop better models that will more completely represent generated biological knowledge and introduce MEMMAL (MEchanistic Modeling with MAchine Learning) framework. MEMMAL processes connections inferred <italic>via</italic> machine-learning pipelines (i.e., MOBILE (<xref ref-type="bibr" rid="B5">Erdem et al., 2022a</xref>)) as new interactions into mechanistic models (i.e., SPARCED (<xref ref-type="bibr" rid="B6">Erdem et al., 2022b</xref>)) to better recapitulate available datasets (i.e., the recently-released MCF10A dataset (<xref ref-type="bibr" rid="B11">Gross et al., 2022</xref>)). The NIH-LINCS Consortium and MCF10A Common Project recently released this dataset, consisting of multiple omics assay types on breast epithelial MCF10A&#xa0;cell line. MOBILE is a new pipeline to integrate multi-omics datasets and identify context-specific interactions. SPARCED is one of the largest mechanistic models of mammalian cells and is an open-source, human-interpretable, and easy to alter modeling format. Here we focused on incorporating novel IFN&#x3b3;/PD-L1 related associations into the SPARCED model to enable description of the cellular response to checkpoint inhibitor immunotherapies. This work is a template for combining big data, machine-learning-inferred interactions with mechanistic models, which could be more broadly applicable towards building multi-scale precision medicine and whole cell models.</p>
</sec>
<sec sec-type="materials|methods" id="s2">
<title>Materials and methods</title>
<p>In this work, we use ligand-specific interactions between genes as new connections in a large-scale mechanistic model to study the effect of the newly added gene interactions in model responses. It is important to note that MEMMAL is agnostic to the specific tool used to nominate new associations, and the base mechanistic model used; the below are simply chosen as illustrative.</p>
<sec id="s3">
<title>MOBILE</title>
<p>MOBILE is a recent tool for finding context-specific network features by integrating pairs of omics datasets (<xref ref-type="bibr" rid="B5">Erdem et al., 2022a</xref>). In short, statistical associations are calculated between pairs of chromatin accessibility regions, mRNA expressions, and protein/phosphoprotein levels. Lasso (least absolute shrinkage and selection operator) regression models are run in replicate to select coefficients with high occurrence rates (<xref ref-type="bibr" rid="B34">Tibshirani, 1996</xref>; <xref ref-type="bibr" rid="B7">Erdem et al., 2016</xref>; <xref ref-type="bibr" rid="B5">Erdem et al., 2022a</xref>). The so-called Integrated Association Networks (IANs) are generated by combining the association networks inferred for RPPA (reverse phase protein array)&#x2b;RNAseq and RNAseq &#x2b; ATACseq data inputs. Finally, the IANs are coalesced into gene-level networks: nodes representing genes of the assay analytes and edges representing the inferred Lasso coefficients. From MOBILE generated IFN&#x03B3;-specific IAN, a sub-network of connections between canonical interferon genes, PD-L1, and PD-1 is filtered to obtain a 297 node &#x002B; 321 edge module. Then, only the interactions with IRF1, PD-L1, PD-1, and STAT1 are retained as input for MEMMAL.</p>
</sec>
<sec id="s4">
<title>SPARCED</title>
<p>The starting mechanistic model used in this work is obtained from the SPARCED repository (<ext-link ext-link-type="uri" xlink:href="http://github.com/birtwistlelab/SPARCED/tree/develop">github.com/birtwistlelab/SPARCED/tree/develop</ext-link>) (<xref ref-type="bibr" rid="B6">Erdem et al., 2022b</xref>). It is a recent framework for large-scale mechanistic modeling that enables model file creation using simple text files as input with minimal coding requirements. In short, a set of annotated text files are constructed to define model specifics. Then, Jupyter notebooks are used to process these files and create community-standard model file type called Systems Biology Markup Language (SBML) (<xref ref-type="bibr" rid="B15">Hucka et al., 2003</xref>; <xref ref-type="bibr" rid="B18">Keating et al., 2020</xref>). The software was first built to replicate the one of the largest mammalian single-cell mechanistic model of proliferation and death signaling (<xref ref-type="bibr" rid="B4">Bouhaddou et al., 2018</xref>; <xref ref-type="bibr" rid="B6">Erdem et al., 2022b</xref>). Then, an expanded SPARCED model was created to include IFN&#x03B3; signaling and SOCS1 crosstalk to growth pathways and the new model was named as SPARCED-IFNG-SOCS1 (<xref ref-type="bibr" rid="B6">Erdem et al., 2022b</xref>). This final model and its input files are used as the basic model in this work and is modified further with the MOBILE inferred set of new connections.</p>
</sec>
<sec id="s5">
<title>MEMMAL</title>
<sec id="s5-1">
<title>Jupyter notebooks</title>
<p>MEMMAL pipeline is composed of multiple Jupyter notebooks defined below and detailed steps given in <xref ref-type="sec" rid="s13">Supplementary Table S1</xref>.<list list-type="simple">
<list-item>
<p>1) <monospace>enlargeModel</monospace> notebook: As the core of MEMMAL, this Jupyter notebook processes the machine learning model inferred connections list and creates Species (genes, mRNAs, proteins, phosphoproteins), RateLaws (the reaction format and related parameters), Gene Regulatory Interactions (defining transcriptional activators and repressors) and finds relevant new omics data from LINCS datasets. The input files for SPARCED pipeline are then updated followed by model compilation and simulation steps.</p>
</list-item>
</list>
</p>
<p>The pipeline starts by finding the unique list of genes from the MOBILE associations input. Then, for each unique gene added we create species for the active gene, inactive gene, mRNA, and protein (phosphoproteins as well if the gene has corresponding phosphoprotein measurements). The species initial conditions are updated using LINCS (<xref ref-type="bibr" rid="B11">Gross et al., 2022</xref>), MCF10A (<xref ref-type="bibr" rid="B4">Bouhaddou et al., 2018</xref>), or other literature datasets (<xref ref-type="bibr" rid="B30">Schwanh&#xe4;usser et al., 2011</xref>). The experimental data in molecules per cell (mpc) are converted into nanomolar (nM) concentration and the corresponding values are updated. Next, first-order translation, transcription, and protein and mRNA degradation reactions are created and the rate laws are defined. The rate constants are set using literature data (<xref ref-type="bibr" rid="B30">Schwanh&#xe4;usser et al., 2011</xref>) or set to the mean value of the corresponding reaction parameter values for existing genes in SPARCED. The mRNA and protein degradation rate constants are set using literature half-life data (<inline-formula id="inf1">
<mml:math id="m1">
<mml:mrow>
<mml:mi>k</mml:mi>
<mml:mi>T</mml:mi>
<mml:mi>C</mml:mi>
<mml:mi>d</mml:mi>
<mml:mo>&#x3d;</mml:mo>
<mml:mfrac>
<mml:mrow>
<mml:mi>log</mml:mi>
<mml:mo>&#x2061;</mml:mo>
<mml:mo>&#x2061;</mml:mo>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mn>2</mml:mn>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
<mml:msub>
<mml:mrow>
<mml:mi>m</mml:mi>
<mml:mi>R</mml:mi>
<mml:mi>N</mml:mi>
<mml:mi>A</mml:mi>
</mml:mrow>
<mml:mrow>
<mml:mi>h</mml:mi>
<mml:mi>a</mml:mi>
<mml:mi>l</mml:mi>
<mml:mi>f</mml:mi>
<mml:mo>&#x2212;</mml:mo>
<mml:mi>l</mml:mi>
<mml:mi>i</mml:mi>
<mml:mi>f</mml:mi>
<mml:mi>e</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mfrac>
</mml:mrow>
</mml:math>
</inline-formula>; <inline-formula id="inf2">
<mml:math id="m2">
<mml:mrow>
<mml:mi>k</mml:mi>
<mml:mi>T</mml:mi>
<mml:mi>L</mml:mi>
<mml:mi>d</mml:mi>
<mml:mo>&#x3d;</mml:mo>
<mml:mfrac>
<mml:mrow>
<mml:mi>log</mml:mi>
<mml:mo>&#x2061;</mml:mo>
<mml:mo>&#x2061;</mml:mo>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mn>2</mml:mn>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
<mml:msub>
<mml:mrow>
<mml:mi>p</mml:mi>
<mml:mi>r</mml:mi>
<mml:mi>o</mml:mi>
<mml:mi>t</mml:mi>
<mml:mi>e</mml:mi>
<mml:mi>i</mml:mi>
<mml:mi>n</mml:mi>
</mml:mrow>
<mml:mrow>
<mml:mi>h</mml:mi>
<mml:mi>a</mml:mi>
<mml:mi>l</mml:mi>
<mml:mi>f</mml:mi>
<mml:mo>&#x2212;</mml:mo>
<mml:mi>l</mml:mi>
<mml:mi>i</mml:mi>
<mml:mi>f</mml:mi>
<mml:mi>e</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mfrac>
</mml:mrow>
</mml:math>
</inline-formula>), basal transcription rate constants using the equation (<inline-formula id="inf3">
<mml:math id="m3">
<mml:mrow>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>k</mml:mi>
<mml:mi>T</mml:mi>
<mml:mi>C</mml:mi>
<mml:mi>d</mml:mi>
<mml:mo>&#x2a;</mml:mo>
<mml:msub>
<mml:mrow>
<mml:mi>m</mml:mi>
<mml:mi>R</mml:mi>
<mml:mi>N</mml:mi>
<mml:mi>A</mml:mi>
</mml:mrow>
<mml:mrow>
<mml:mi>c</mml:mi>
<mml:mi>o</mml:mi>
<mml:mi>u</mml:mi>
<mml:mi>n</mml:mi>
<mml:mi>t</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>&#x2a;</mml:mo>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mrow>
<mml:mi>k</mml:mi>
<mml:mi>G</mml:mi>
</mml:mrow>
<mml:mrow>
<mml:mi>i</mml:mi>
<mml:mi>n</mml:mi>
</mml:mrow>
</mml:msub>
<mml:mo>&#x2b;</mml:mo>
<mml:msub>
<mml:mrow>
<mml:mi>k</mml:mi>
<mml:mi>G</mml:mi>
</mml:mrow>
<mml:mrow>
<mml:mi>a</mml:mi>
<mml:mi>c</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>/</mml:mo>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mrow>
<mml:mi>k</mml:mi>
<mml:mi>G</mml:mi>
</mml:mrow>
<mml:mrow>
<mml:mi>a</mml:mi>
<mml:mi>c</mml:mi>
</mml:mrow>
</mml:msub>
<mml:mo>&#x2a;</mml:mo>
<mml:mi>G</mml:mi>
<mml:mi>e</mml:mi>
<mml:mi>n</mml:mi>
<mml:mi>e</mml:mi>
<mml:mo>_</mml:mo>
<mml:mi>C</mml:mi>
<mml:mi>o</mml:mi>
<mml:mi>p</mml:mi>
<mml:mi>y</mml:mi>
<mml:mo>_</mml:mo>
<mml:mi>N</mml:mi>
<mml:mi>u</mml:mi>
<mml:mi>m</mml:mi>
<mml:mi>b</mml:mi>
<mml:mi>e</mml:mi>
<mml:mi>r</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula> where kG<sub>in</sub> and kG<sub>ac</sub> are rate of gene inactivation and activation, respectively. The translation rate constants are set using the equation (<inline-formula id="inf4">
<mml:math id="m4">
<mml:mrow>
<mml:msub>
<mml:mrow>
<mml:mi>p</mml:mi>
<mml:mi>r</mml:mi>
<mml:mi>o</mml:mi>
<mml:mi>t</mml:mi>
<mml:mi>e</mml:mi>
<mml:mi>i</mml:mi>
<mml:mi>n</mml:mi>
</mml:mrow>
<mml:mrow>
<mml:mi>c</mml:mi>
<mml:mi>o</mml:mi>
<mml:mi>n</mml:mi>
<mml:mi>c</mml:mi>
<mml:mi>e</mml:mi>
<mml:mi>n</mml:mi>
<mml:mi>t</mml:mi>
<mml:mi>r</mml:mi>
<mml:mi>a</mml:mi>
<mml:mi>t</mml:mi>
<mml:mi>i</mml:mi>
<mml:mi>o</mml:mi>
<mml:mi>n</mml:mi>
</mml:mrow>
</mml:msub>
<mml:mo>&#x2a;</mml:mo>
<mml:mi>k</mml:mi>
<mml:mi>T</mml:mi>
<mml:mi>L</mml:mi>
<mml:mi>d</mml:mi>
<mml:mo>/</mml:mo>
<mml:msub>
<mml:mrow>
<mml:mi>m</mml:mi>
<mml:mi>R</mml:mi>
<mml:mi>N</mml:mi>
<mml:mi>A</mml:mi>
</mml:mrow>
<mml:mrow>
<mml:mi>c</mml:mi>
<mml:mi>o</mml:mi>
<mml:mi>n</mml:mi>
<mml:mi>c</mml:mi>
<mml:mi>e</mml:mi>
<mml:mi>n</mml:mi>
<mml:mi>t</mml:mi>
<mml:mi>r</mml:mi>
<mml:mi>a</mml:mi>
<mml:mi>t</mml:mi>
<mml:mi>i</mml:mi>
<mml:mi>o</mml:mi>
<mml:mi>n</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula>).</p>
<p>Importantly, for this work we specify that all associations are gene regulatory mechanisms, and for each association, two transcriptional regulation connections are created: the protein species of gene1 activates/represses gene2 expression and protein of gene2 activates/represses gene1 expression. That however is because of the specific submodel of interest here being a gene regulatory subnetwork and future implementations would need to be considered case-by-case. These gene regulatory reactions are modeled as Hill equations as defined for other gene regulatory reactions in SPARCED (<xref ref-type="bibr" rid="B6">Erdem et al., 2022b</xref>). The Hill equation parameters are: i) n<sub>A</sub>: Hill coefficients set to &#x201c;4&#x201d; for all new reactions and ii) K<sub>A</sub> the concentration for half-maximal transcriptional output effect, initially set to half of the transcriptionally regulating protein concentration. The values of these K<sub>A</sub> parameters are fitted later, as described below. Finally, the updated input files are written into text files for model creation and compilation.<list list-type="simple">
<list-item>
<p>2) <monospace>createModel_o4a</monospace> notebook: The Jupyter notebook to create an integrated SBML version of the SPARCED type models (<xref ref-type="bibr" rid="B6">Erdem et al., 2022b</xref>). Creating the model file fully in SBML format provides extensive speed-up of simulations. The newly updated input files by <monospace>enlargeModel</monospace> notebook are used to create and compile the expanded model.</p>
</list-item>
<list-item>
<p>3) <monospace>runModel</monospace> notebook: This Jupyter notebook is used to simulate and explore multiple scenarios for the new model.</p>
</list-item>
<list-item>
<p>4) <monospace>enlargeSBMLModel</monospace> notebook: This Jupyter notebook contains an example to enlarge any SBML model using user defined lists of species, reactions, and parameters. We provide an example use of <monospace>enlargeModel</monospace> notebook created lists of model elements to expand the SBML file of IFN&#x3b3;/JAK/STAT signaling pathway (<xref ref-type="bibr" rid="B43">Yamada et al., 2003</xref>).</p>
</list-item>
<list-item>
<p>5) <monospace>testMEMMAL</monospace> notebook: This Jupyter notebook contains commands to run MEMMAL from start to finish. It calls the first three notebooks and plots the figure panels.</p>
</list-item>
</list>
</p>
</sec>
<sec id="s5-2">
<title>Input files</title>
<p>
<list list-type="simple">
<list-item>
<p>1) Compartments, GeneReg, OmicsData, RatelawsNoSM, Species, and Initializer text files: SPARCED input files for the SPARCED-IFNG-SOCS1 model from (<xref ref-type="bibr" rid="B6">Erdem et al., 2022b</xref>).</p>
</list-item>
<list-item>
<p>2) IRF1_PDL1sub: MOBILE derived associations list from (<xref ref-type="bibr" rid="B5">Erdem et al., 2022a</xref>). Steps to obtain the list are given in <xref ref-type="sec" rid="s13">Supplementary Figure S1</xref>.</p>
</list-item>
<list-item>
<p>3) RNAseqDataLINCS: RNAseq data in log2(fpkm&#x2b;1) format.</p>
</list-item>
<list-item>
<p>4) RPPADataLINCS, RPPADataStdLINCS, and RPPADataStdLINCSfc: Median normalized RPPA data in log2 format. &#x201c;Std&#x201d; refers to standard deviation of triplicate measurements. &#x201c;fc&#x201d; refers to fold-change with respect to time point zero.</p>
</list-item>
<list-item>
<p>5) Schwanhausser2011: Literature data on mRNA and protein half-lives (<xref ref-type="bibr" rid="B30">Schwanh&#xe4;usser et al., 2011</xref>).</p>
</list-item>
<list-item>
<p>6) Supplementary_Data_22: Transcriptomic and proteomic data for MCF10A&#xa0;cells (<xref ref-type="bibr" rid="B6">Erdem et al., 2022b</xref>).</p>
</list-item>
</list>
</p>
</sec>
<sec id="s5-3">
<title>Output files and folders</title>
<p>
<list list-type="simple">
<list-item>
<p>1) GeneReg_MM, OmicsData_MM, RatelawsNoSM_MM, and Species_MM text files: Updated/expanded input files with new connections and data.</p>
</list-item>
<list-item>
<p>2) &#x201c;Model name.txt&#x201d; [i.e., MEMMAL_orig.txt]: Model file in Antimony format (<xref ref-type="bibr" rid="B31">Smith et al., 2009</xref>).</p>
</list-item>
<list-item>
<p>3) &#x201c;Model name.xml&#x201d; [i.e., MEMMAL_orig.xml]: Model file in SBML format (<xref ref-type="bibr" rid="B18">Keating et al., 2020</xref>).</p>
</list-item>
<list-item>
<p>4) &#x201c;Model name folder&#x201d; [i.e., MEMMAL_orig]: Compiled model folder created by AMICI package (<xref ref-type="bibr" rid="B9">Fr&#xf6;hlich et al., 2020</xref>; <xref ref-type="bibr" rid="B38">Weindl et al., 2020</xref>).</p>
</list-item>
</list>
</p>
</sec>
<sec id="s5-4">
<title>Parameter fitting</title>
<p>The new parameter values were initially set using literature data or existing model parameters. We then estimated some of them in a semi-automated way. First, the basal transcription (mRNA production) rate constants of the new mRNAs species (eight in total) are fitted one at a time, in the order of species added to the model. If the mRNA level was not at steady state, degrading or accumulating in no ligand (growth factors or IFN&#x3b3;) stimulation simulations, the parameter value is estimated by varying it uniformly (15 points) within three orders of log10-magnitude of the default value. Then, the best-fit value that yields a constant level is manually adjusted for better fit if possible. Finally, such parameter values are kept constant and the next is explored. One of the mRNA degradation parameters (of FAM83D) was also fitted similarly.</p>
<p>The values for the K<sub>A</sub> (half-maximal) concentrations of the newly added gene regulatory reactions were adjusted using the LINCS mRNA (ACSL5, BST2, CLIC2, FAM83D, HIST2H2AA3, and METAP2) and protein (IRF1 and PD-L1) time course data with EGF and EGF &#x2b; IFN&#x3b3; stimulation. The model, starting from an initial steady-state condition in the absence of growth factors (from above), is simulated for 48&#xa0;h with EGF (1.5625&#xa0;nM) or EGF (1.5625&#xa0;nM) &#x2b; IFN&#x3b3; (1.1834&#xa0;nM) treatment. The K<sub>A</sub> for each new gene regulatory interaction (27 total) is varied uniformly (15 points) within three orders of log10-magnitude of the default value (half the regulating protein species concentration) and both stimulation conditions are simulated. The sum-of-squared errors between simulation and the data is evaluated for each, and the value giving minimum error is chosen. In some cases, the value with minimum error is manually adjusted between originally sampled values to achieve better fit. These fitted K<sub>A</sub> parameter values are reported in the <monospace>runModel</monospace> notebook.</p>
</sec>
<sec id="s5-5">
<title>Code availability</title>
<p>MEMMAL code is available at the GitHub repository github.com/cerdem12/MEMMAL.</p>
</sec>
</sec>
</sec>
<sec sec-type="results" id="s6">
<title>Results</title>
<sec id="s6-1">
<title>Large-scale mechanistic models can become larger and more precise by expansion using machine learned relationships</title>
<p>There are only a handful of large-scale (hundreds of genes, thousands of species) mechanistic signaling pathway models in the literature (<xref ref-type="bibr" rid="B8">Fr&#xf6;hlich et al., 2018</xref>). Usually, such big models are constructed by bottom-up modeling or by semi-manual stitching of previously published models (<xref ref-type="bibr" rid="B4">Bouhaddou et al., 2018</xref>). Both approaches are time consuming, manually curated, and biased for including/excluding model components: genes, proteins, post-translational modifications, interactions, or even cellular compartments. Here, we tackle this &#x201c;what-to-add&#x201d; problem by using association networks inferred <italic>via</italic> data-driven machine learning algorithms.</p>
<p>The Mechanistic Modeling with Machine Learning (MEMMAL) tool presented here (<xref ref-type="fig" rid="F2">Figure 2</xref>) is comprised of scripts to expand mechanistic models created using SPARCED pipeline (<xref ref-type="bibr" rid="B6">Erdem et al., 2022b</xref>) with candidate connections generated by the tool called MOBILE, a recent pipeline for multi-omics data integration (<xref ref-type="bibr" rid="B5">Erdem et al., 2022a</xref>). However, other tools and models could be used in their place; they are simply used to demonstrate the approach. For now, the MEMMAL Jupyter notebooks process these new connection candidates to update SPARCED input files, taking advantage of their modular structure for model building (<ext-link ext-link-type="uri" xlink:href="http://github.com/birtwistlelab/SPARCED/tree/develop">github.com/birtwistlelab/SPARCED/tree/develop</ext-link>). Here, we combine novel connections inferred <italic>via</italic> MOBILE with a large-scale mechanistic model called SPARCED to add an immune-checkpoint related sub-module to the existing pan-cancer model to study effects of the newly added gene products on the regulation of Interferon Regulatory Factor 1 (gene name IRF) and Programmed Death Ligand 1 (PD-L1, gene name CD274) upon interferon-gamma (IFN&#x3b3;, gene name IFNG) stimulation.</p>
<fig id="F2" position="float">
<label>FIGURE 2</label>
<caption>
<p>MEMMAL is a pipeline to merge mechanistic modeling with machine learning <bold>(A)</bold> The MEMMAL pipeline combines mechanistic models created by SPARCED with association networks generated <italic>via</italic> MOBILE pipeline <bold>(B)</bold> The recipe for MEMMAL pipeline starts by obtaining a set of connections not presented in the candidate mechanistic model. Here, the novel gene-level connections list is inferred <italic>via</italic> the MOBILE tool and then filtered for overlap with SPARCED model genes. Next, this candidate network is imported into SPARCED environment, where the MEMMAL <monospace>enlargeModel</monospace> Jupyter notebook processes the network file and updates SPARCED input files The nodes (genes) of the IFNG/PD-L1 subnetwork are used to create new genes and species (mRNAs, proteins, phosphoproteins) for SPARCED. The new genes can get activated/inactivated as described in SPARCED. The expanded MEMMAL model is created and compiled by default SPARCED model notebooks. The final step in MEMMAL is to run user defined exploratory simulations to gain insights on the effects of new connections added.</p>
</caption>
<graphic xlink:href="fsysb-03-1099413-g002.tif"/>
</fig>
</sec>
<sec id="s6-2">
<title>MOBILE pipeline integrated LINCS MCF10A multi-omics dataset to infer ligand-specific associations</title>
<p>The normal-like breast epithelial cell line MCF10A was recently profiled with multiple assay types under multiple ligand stimulation conditions (<xref ref-type="bibr" rid="B11">Gross et al., 2022</xref>). Using this newly released multi-omics dataset, our lab introduced the MOBILE pipeline for data integration and showed how ligand-specific associations can be inferred (<xref ref-type="bibr" rid="B5">Erdem et al., 2022a</xref>). One of the ligands included in the LINCS study that induced MCF10A growth inhibition was interferon-gamma (<xref ref-type="bibr" rid="B11">Gross et al., 2022</xref>). We previously analyzed the LINCS MCF10A dataset to find IFN&#x3b3;-specific associations that nominate novel connections with the PD-L1 (gene name CD274) axis (<xref ref-type="bibr" rid="B5">Erdem et al., 2022a</xref>). IFN&#x3b3; can induce transient PD-L1 expression, a transmembrane protein that binds to its receptor PD-1 on T-cells (<xref ref-type="bibr" rid="B1">Abiko et al., 2015</xref>; <xref ref-type="bibr" rid="B33">Thiem et al., 2019</xref>; <xref ref-type="bibr" rid="B17">Ju et al., 2020</xref>). This binding inhibits tumor clearance, where targeted therapies towards these proteins are a new class of anti-cancer drugs: the immune checkpoint inhibitors (<xref ref-type="bibr" rid="B10">Gong et al., 2018</xref>). However, inter- and intra-tumor variability of PD-L1 expression results in heterogeneous patient responses and makes the response predictions a challenge (<xref ref-type="bibr" rid="B41">Wu et al., 2019</xref>). A more thorough understanding of the regulatory mechanism of PD-L1 expression could help inform new immunotherapeutic drugs or treatment options.</p>
<p>Applying MOBILE, we generated a data-driven IFN&#x3b3;-specific integrated associations network, which had 297 nodes (genes) and 321 edges (connections) (<xref ref-type="fig" rid="F2">Figure 2B</xref> and <xref ref-type="sec" rid="s13">Supplementary Figure S1</xref>). We further filtered this network by looking for connections with STAT1 (the only overlapping gene with the mechanistic model). The final list of candidate connections had nine genes (ACSL5, BST2, CD274, CLIC2, FAM83D, HIST2H2AA3, IRF1, METAP2, and STAT1) and 14 connections. The list is imported into the SPARCED environment to start altering the existing mechanistic model structure (<xref ref-type="fig" rid="F2">Figure 2B</xref> and <xref ref-type="sec" rid="s13">Supplementary Figure S1</xref>).</p>
</sec>
<sec id="s6-3">
<title>SPARCED modeling makes mechanistic model expansions easy</title>
<p>SPARCED is a recent software (<xref ref-type="bibr" rid="B6">Erdem et al., 2022b</xref>) and modeling framework for large-scale mechanistic modeling. It enables SBML model file creation using simple text files as input with minimal coding requirements. Jupyter notebooks (<xref ref-type="bibr" rid="B19">Kluyver et al., 2016</xref>) are used to process the input files and to create the model files. The software was first built to replicate the largest mammalian single-cell mechanistic model of proliferation and death signaling (<xref ref-type="bibr" rid="B4">Bouhaddou et al., 2018</xref>) and was then expanded to include a new sub-module of IFN&#x3b3; signaling (<xref ref-type="bibr" rid="B43">Yamada et al., 2003</xref>). So, the starting mechanistic model in this work, SPARCED-IFNG-SOCS1 already includes an IFN&#x3b3; submodule (<xref ref-type="fig" rid="F3">Figure 3A</xref>, gray background), with a total of 149 genes, 1,302 species, and 3,584 ratelaws (<xref ref-type="fig" rid="F3">Figure 3B</xref>).</p>
<fig id="F3" position="float">
<label>FIGURE 3</label>
<caption>
<p>MOBILE inferred IFN&#x03B3;/PD-L1 network nodes and connections are inserted into the SPARCED-IFNG model using MEMMAL <bold>(A)</bold> The SPARCED network is enlarged to include a sub-network spanning innate immune response and PD-L1 regulation <bold>(B)</bold> The final MEMMAL model is 157 genes, 1,318 species, and 3,600 ratelaws, 60 TARs, and 3,885 parameters <bold>(C)</bold> The reactions added into SPARCED include translation (black arrows). The connections from MOBILE are modeled as transcriptional activation and repression (TAR) reactions in MEMMAL (gray arrows). The TAR reactions linking existing SPARCED species with the newly added species are represented as integrative links (red arrows) <bold>(D)</bold> The final MEMMAL model recapitulates canonical transient STAT1 and SOCS1 activation in response to IFN&#x3b3; stimulation in MCF10A cells (<xref ref-type="bibr" rid="B6">Erdem et al., 2022b</xref>). Normalized simulation trajectories of the activated nuclear STAT1 dimer (STAT1&#x2a;Dn), SOCS1 mRNA (mRNA_SOCS1), and free SOCS1 protein (SOCS1) are shown (solid gray lines).</p>
</caption>
<graphic xlink:href="fsysb-03-1099413-g003.tif"/>
</fig>
</sec>
<sec id="s6-4">
<title>MEMMAL incorporates MOBILE-inferred gene-level statistical associations into SPARCED as gene regulatory mechanisms</title>
<p>The list of candidate connections from MOBILE pipeline are processed <italic>via</italic> MEMMAL <monospace>enlargeModel</monospace> notebook to add rows and update SPARCED input files (<xref ref-type="fig" rid="F2">Figure 2B</xref>). As a default SPARCED requirement, each gene node from MOBILE list is interpreted to create active gene, inactive gene, mRNA, and protein species, with relevant basic reactions: gene switching, transcription, translation (<xref ref-type="fig" rid="F3">Figure 3C</xref>, black arrows), mRNA degradation, and protein degradation. Importantly, the MOBILE inferred connections are interpreted as transcriptional activator and repressor (TAR) reactions (<xref ref-type="fig" rid="F3">Figure 3C</xref>) because the MOBILE inferred connections are obtained by looking at pairs of mRNA-protein and chromatin region-mRNA dataset pairs. A logical way a protein affecting another mRNA&#x2019;s expression level is by transcriptional regulation. Additionally, a highly open chromatin region can permit transcription, which potentially yields higher mRNA expression and thus another gene regulatory connection. So, all the candidate associations are treated as TARs in the current MEMMAL pipeline. For future work, users should decide how to handle such connections.</p>
<p>The negative valued associations here are treated as inhibitory whereas the positive magnitude connections are added as activators (<xref ref-type="fig" rid="F3">Figure 3C</xref>, gray and red arrows). Some of the transcriptional activators are labeled as &#x201C;integrative links&#x201d; because they connect existing SPARCED model genes with the new gene species (<xref ref-type="fig" rid="F3">Figure 3C</xref>, red arrows). After all the input files are updated, <monospace>createModel_o4a</monospace> Jupyter notebook is used to create and compile the new SBML model file (<xref ref-type="fig" rid="F2">Figure 2B</xref>). The MEMMAL expansion of SPARCED <italic>via</italic> MOBILE inferred network resulted in the addition of eight genes, 16 species, 16 signaling reactions, and 27 transcriptional regulatory mechanisms (<xref ref-type="fig" rid="F3">Figures 3A, B</xref>). With the current addition, the SPARCED model now includes an IFN&#x3b3;-PD-L1 submodule (<xref ref-type="fig" rid="F3">Figure 3A</xref>, green background).</p>
<p>Following model expansion, we first verified the model can recapitulate previous observations (<xref ref-type="fig" rid="F3">Figure 3D</xref>). We show that inclusion of new species and reactions did not alter canonical STAT1-SOCS1 response to IFN&#x3b3; stimulation. Previous studies have shown that in response to IFN&#x3b3;, STAT1 and SOCS1 show transient activation over several hours followed by damped oscillations before reaching a steady state slightly higher than the baseline levels (<xref ref-type="bibr" rid="B43">Yamada et al., 2003</xref>). In the model, IFN&#x3b3; treatment leads to transient STAT1 activation by inducing its phosphorylation, dimerization, and translocation to nucleus (<xref ref-type="fig" rid="F3">Figure 3D</xref>, top panel). Nuclear STAT1 dimer acts as an activating transcription factor for SOCS1 and induces SOCS1 mRNA production (<xref ref-type="fig" rid="F3">Figure 3D</xref>, middle panel), which then causes SOCS1 protein levels to increase (<xref ref-type="fig" rid="F3">Figure 3D</xref>, bottom panel). Moreover, as reported previously in (<xref ref-type="bibr" rid="B6">Erdem et al., 2022b</xref>), IFN&#x3b3; does not induce significant changes in MAPK signaling but leads to a slight decrease in early AKT response (<xref ref-type="sec" rid="s13">Supplementary Figure S2</xref>).</p>
</sec>
<sec id="s6-5">
<title>MEMMAL model offers exploration of the effect of novel connector genes on the expression of PD-L1 expression in response to IFN&#x03B3;</title>
<p>Since the modified model passed these quality control checks, the next step was to fit new unknown parameters to recapitulate experimental time-course data for newly added genes (RNAseq: ACSL5, BST2, CLIC2, FAM83D, HIST2H2AA3, METAP2 and RPPA: IRF1, PD-L1) (<xref ref-type="fig" rid="F4">Figure 4A</xref>). These 27 &#x2b; 16 (43 total) unknown parameters were the half-maximal concentrations for the Hill functions underlying the new gene regulatory reactions and protein/mRNA degradation rate constants. The data show IFN&#x3b3; induces transcription of ACSL5, BST2, CLIC2, and HIST2H2AA3 and expression of both IRF1 and PD-L1 with no sustained induction of FAM83D and METAP2, and the fitted model captures these trends. There are only two discrepancies where the model could not capture: 24-h time point data of FAM83D and HIST2H2AA3 mRNA levels. However, the model can recapitulate the increasing trend of mRNA_ HIST2H2AA3 and fit the last time points for both species levels. The <monospace>runModel</monospace> Jupyter notebook reports the final updated parameter values and scripts to compare simulation trajectories with LINCS data (<xref ref-type="fig" rid="F4">Figure 4A</xref>).</p>
<fig id="F4" position="float">
<label>FIGURE 4</label>
<caption>
<p>MEMMAL can replicate the previous SPARCED-IFNG model and offers new insights into IFN&#x03B3; regulation of IRF1 and PD-L1 dynamics <bold>(A)</bold> MEMMAL model parameters are fitted to recapitulate experimental data from LINCS RNAseq and RPPA assays. Fold-changes are shown for data (dots, crosses, and error bars, STD) and simulations (solid lines). Most mRNAs and IRF1 and PD-L1 (gene name CD274) are induced by IFN&#x3b3; <bold>(B)</bold> Simulation scenarios to test the effects of newly added genes. The parameter fitted MEMMAL model is simulated with reported perturbation under IGF1 stimulation (basal growth condition) and then stimulated with additional EGF &#x2b; IFN&#x3b3; for 48&#xa0;h <bold>(C)</bold> Comparison of complete gene knock-out perturbation scenario (dotted lines) to wild-type (no perturbation, black lines) condition shows genes with induced IRF1 and PD-L1 changes. Among the newly added genes, METAP2 induces the greatest change: a complete recession of IRF1 response and decreased PD-L1 steady-state level. The network diagram (summary of <xref ref-type="fig" rid="F3">Figure 3C</xref>) shows the connections among functional genes and STAT1, with non-functional edges faded out.</p>
</caption>
<graphic xlink:href="fsysb-03-1099413-g004.tif"/>
</fig>
<p>After acceptable agreement was achieved between simulations and experimental mRNA and protein levels (<xref ref-type="fig" rid="F4">Figure 4A</xref>), we simulated scenarios (<xref ref-type="fig" rid="F4">Figure 4B</xref>) to explore the effects of new genes on the IRF1 and PD-L1 responses. We wanted to nominate the new connections predicted to be most important in regulating PD-L1 expression. To do this we compared wild-type simulations (new model with fit parameters) to single gene knock-out simulations (protein, gene, and mRNA levels set to zero) (<xref ref-type="fig" rid="F4">Figures 4B,C</xref>).</p>
<p>Only BST2, FAM83D, and METAP2 knock-outs had observable effects on simulated PD-L1 and/or IRF1 dynamics (<xref ref-type="fig" rid="F4">Figure 4C</xref>). Knocking out other newly added genes (ACSL5, CLIC2, HIST2H2AA3) had no significant effects and thus are not shown here. Perturbing BST2 caused a small decrease in initial PD-L1 levels, which later reaches to wild-type response levels (<xref ref-type="fig" rid="F4">Figure 4C</xref>, top row). Perturbing FAM83D only slightly increased steady-state IRF1 levels (<xref ref-type="fig" rid="F4">Figure 4C</xref>, middle row). Perturbing METAP2 caused a significant decrease in late IRF1 and PD-L1 responses (<xref ref-type="fig" rid="F4">Figure 4C</xref>, bottom row). We summarized all these knock-out response observations with the candidate gene regulatory network in <xref ref-type="fig" rid="F3">Figure 3C</xref> to show a functional network with possibly causal links only (<xref ref-type="fig" rid="F4">Figure 4C</xref>). These results demonstrate that mechanistic models with machine learning derived connections can nominate genes for follow-up experimental studies.</p>
</sec>
</sec>
<sec sec-type="discussion" id="s7">
<title>Discussion</title>
<p>Combining and synergizing machine learning with mechanistic modeling could bring clinically predictive computational models and personalized medicine closer to reality. To that end, here we introduced a recipe to expand a large-scale mechanistic model with machine learned connections between gene products. Because understanding PD-L1 regulation mechanisms would help us design better therapeutic interventions, we focused on exploring the IFN&#x3b3;/PD-L1 axis. We used the LINCS MCF10A dataset and added the recently inferred (<italic>via</italic> MOBILE pipeline) IFN&#x3b3;/PD-L1 connections to the existing SPARCED mechanistic model. We then were able to study the effects of new gene regulatory mechanisms. We showed that perturbing BST2, FAM83D, or METAP2 induces changes in PD-L1 and IRF1 dynamics.</p>
<p>MEMMAL could serve as an initial step towards combining mechanistic models with machine learnt potential connections by providing a rationale for such a merging protocol. MEMMAL protocol first creates genes and gene products (mRNA and protein) if MOBILE list nodes are not already present in SPARCED. It then updates -omics level information for the new genes and adds corresponding reactions. It also assigns transcriptional activator and repressors (based on MOBILE association coefficient sign) and related rate constant parameters. The updated SPARCED input files are then processed <italic>via</italic> modified default Jupyter notebooks to execute desired simulations. The current state of the MEMMAL assumes an overlap (genes) between the mechanistic model and machine learned associations. Although this is not a hard assumption, it also makes logical sense that the effects of added interactions can be explored <italic>via</italic> crosstalk mechanisms.</p>
<p>Although MEMMAL makes use of recent tools from our lab, the idea is applicable to other tools available in the literature. For instance, rule-based modeling software like BioNetGen (<xref ref-type="bibr" rid="B13">Harris et al., 2016</xref>) and PySB (<xref ref-type="bibr" rid="B21">Lopez et al., 2013</xref>) can also be used for mechanistic model creation and update if machine learning predicted associations are converted into new rules. Another possible application can include INDRA (<xref ref-type="bibr" rid="B12">Gyori et al., 2017</xref>) if the new connections are put into suitable sentence format. Such options will be valuable to expand the MEMMAL idea and its applications.</p>
<p>MEMMAL is agnostic to the approach or tool used to identify connections and to the base mechanistic model for expansion. MEMMAL can generate mechanistic ODE models by integrating connections inferred using MOBILE, databases, correlation studies (<xref ref-type="bibr" rid="B20">Lin et al., 2013</xref>; <xref ref-type="bibr" rid="B25">Min et al., 2021</xref>), kernel-based methods (<xref ref-type="bibr" rid="B23">Mariette and Villa-Vialaneix, 2018</xref>; <xref ref-type="bibr" rid="B44">Yang et al., 2018</xref>), other machine learning tools (<xref ref-type="bibr" rid="B28">Park et al., 2015</xref>; <xref ref-type="bibr" rid="B47">Zhang et al., 2018</xref>; <xref ref-type="bibr" rid="B16">Hulot et al., 2021</xref>), or direct experiments. For the base model any mechanistic model that can be modified programmatically could be used. To facilitate the use of other models, we have provided a Jupyter notebook (<monospace>enlargeSBMLmodel</monospace>) to expand any SBML model with MEMMAL generated lists of new species, reactions, and parameters.</p>
<p>The MOBILE pipeline was used to infer ligand-specific and statistically robust association networks (<xref ref-type="bibr" rid="B5">Erdem et al., 2022a</xref>). Here we used a filtered list of connections for interferon-gamma signaling and among them some genes were already shown to be associated with immunotherapeutic signatures including BST2, CLIC2, and FAM83D (<xref ref-type="bibr" rid="B37">Wang et al., 2013</xref>; <xref ref-type="bibr" rid="B36">Walian et al., 2016</xref>; <xref ref-type="bibr" rid="B42">Xu et al., 2020</xref>; <xref ref-type="bibr" rid="B48">Zhou et al., 2020</xref>; <xref ref-type="bibr" rid="B24">Mei et al., 2021</xref>). In short, BST2 is part of an anti-CTLA4 response in melanoma (<xref ref-type="bibr" rid="B24">Mei et al., 2021</xref>) and CLIC2 is a favorable prognosis biomarker (<xref ref-type="bibr" rid="B42">Xu et al., 2020</xref>). FAM83D functions in cell growth regulation and is a prognostic marker for multiple cancer types (<xref ref-type="bibr" rid="B37">Wang et al., 2013</xref>; <xref ref-type="bibr" rid="B36">Walian et al., 2016</xref>). In addition to such pieces of literature support, we can take a step further to explore their mechanistic functionalities by combining these genes and their predicted connections as new interactions in a computational model.</p>
<p>The investigation of the effects of new genes (<italic>via</italic> knock-out simulations) was carried out after fitting the new reaction parameter values to match experimental time course data. The simple semi-automated fitting procedure in this work resulted in a set of parameter values, reported in <monospace>runModel</monospace> notebook, but their identifiability is not guaranteed. Because the effects of single gene knock-outs simulations are dependent on such values, a more extensive parameter exploration would build confidence in the predictions of which genes are more important for PD-L1 regulation. Indeed, the AMICI package (<xref ref-type="bibr" rid="B9">Fr&#xf6;hlich et al., 2020</xref>) used by SPARCED enables users to do such high-level parameter estimation studies.</p>
<p>In conclusion, the MEMMAL pipeline provides a starting point for merging large-scale mechanistic models with big-data based association networks. We used MEMMAL to test novel candidate interactions for their effect on regulating IRF1 and PD-L1 expression and found that METAP2 is a good candidate yet to be studied experimentally. We believe combining big data, machine learning, and mechanistic models is a valuable direction to unravel novel context-specific mechanisms.</p>
</sec>
</body>
<back>
<sec sec-type="data-availability" id="s8">
<title>Data availability statement</title>
<p>The original contributions presented in the study are included in the article/<xref ref-type="sec" rid="s13">Supplementary Material</xref>, further inquiries can be directed to the corresponding authors. All the data used in this study are available within the MOBILE repository and adapted from (<xref ref-type="bibr" rid="B11">Gross et al., 2022</xref>; <xref ref-type="bibr" rid="B6">Erdem et al., 2022b</xref>).</p>
</sec>
<sec id="s10">
<title>Author contributions</title>
<p>Conceptualization, CE and MRB; Methodology, CE and MRB; Software, CE; Validation: CE; Formal analysis: CE; Resources: MRB; Writing&#x2013;Original Draft: CE and MRB; Writing&#x2013;Review and Editing: CE and MRB; Visualization: CE and MRB; Supervision: CE and MRB; Project administration: CE and MRB; Funding acquisition: MRB.</p>
</sec>
<sec id="s9">
<title>Funding</title>
<p>The authors acknowledge funding from the National Institutes of Health Grants 1R35GM141891 and U54HG008098-LINCS Center (MRB). CE was an NIH-LINCS Consortium Postdoctoral Fellow (2018&#x2013;2020).</p>
</sec>
<sec sec-type="COI-statement" id="s11">
<title>Conflict of interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec sec-type="disclaimer" id="s12">
<title>Publisher&#x2019;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
<sec id="s13">
<title>Supplementary material</title>
<p>The Supplementary Material for this article can be found online at: <ext-link ext-link-type="uri" xlink:href="https://www.frontiersin.org/articles/10.3389/fsysb.2023.1099413/full#supplementary-material">https://www.frontiersin.org/articles/10.3389/fsysb.2023.1099413/full&#x23;supplementary-material</ext-link>.</p>
<supplementary-material xlink:href="Image2.TIF" id="SM1" mimetype="application/TIF" xmlns:xlink="http://www.w3.org/1999/xlink"/>
<supplementary-material xlink:href="Image1.TIF" id="SM2" mimetype="application/TIF" xmlns:xlink="http://www.w3.org/1999/xlink"/>
<supplementary-material xlink:href="Table1.XLSX" id="SM3" mimetype="application/XLSX" xmlns:xlink="http://www.w3.org/1999/xlink"/>
</sec>
<ref-list>
<title>References</title>
<ref id="B1">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Abiko</surname>
<given-names>K.</given-names>
</name>
<name>
<surname>Matsumura</surname>
<given-names>N.</given-names>
</name>
<name>
<surname>Hamanishi</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Horikawa</surname>
<given-names>N.</given-names>
</name>
<name>
<surname>Murakami</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Yamaguchi</surname>
<given-names>K.</given-names>
</name>
<etal/>
</person-group> (<year>2015</year>). <article-title>IFN-&#x3b3; from lymphocytes induces PD-L1 expression and promotes progression of ovarian cancer</article-title>. <source>Br. J. Cancer</source> <volume>112</volume> (<issue>9</issue>), <fpage>1501</fpage>&#x2013;<lpage>1509</lpage>. <pub-id pub-id-type="doi">10.1038/bjc.2015.101</pub-id>
</citation>
</ref>
<ref id="B2">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Baker</surname>
<given-names>R. E.</given-names>
</name>
<name>
<surname>Pe&#xf1;a</surname>
<given-names>J. M.</given-names>
</name>
<name>
<surname>Jayamohan</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>J&#xe9;rusalem</surname>
<given-names>A.</given-names>
</name>
</person-group> (<year>2018</year>). <article-title>Mechanistic models versus machine learning, a fight worth fighting for the biological community?</article-title> <source>Biol. Lett.</source> <volume>14</volume> (<issue>5</issue>), <fpage>20170660</fpage>. <pub-id pub-id-type="doi">10.1098/rsbl.2017.0660</pub-id>
</citation>
</ref>
<ref id="B3">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Barrett</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Wilhite</surname>
<given-names>S. E.</given-names>
</name>
<name>
<surname>Ledoux</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Evangelista</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Kim</surname>
<given-names>I. F.</given-names>
</name>
<name>
<surname>Tomashevsky</surname>
<given-names>M.</given-names>
</name>
<etal/>
</person-group> (<year>2012</year>). <article-title>NCBI geo: Archive for functional genomics data sets&#x2014;update</article-title>. <source>Nucleic Acids Res.</source> <volume>41</volume> (<issue>1</issue>), <fpage>D991</fpage>&#x2013;<lpage>D995</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gks1193</pub-id>
</citation>
</ref>
<ref id="B4">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Bouhaddou</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Barrette</surname>
<given-names>A. M.</given-names>
</name>
<name>
<surname>Stern</surname>
<given-names>A. D.</given-names>
</name>
<name>
<surname>Koch</surname>
<given-names>R. J.</given-names>
</name>
<name>
<surname>DiStefano</surname>
<given-names>M. S.</given-names>
</name>
<name>
<surname>Riesel</surname>
<given-names>E. A.</given-names>
</name>
<etal/>
</person-group> (<year>2018</year>). <article-title>A mechanistic pan-cancer pathway model informed by multi-omics data interprets stochastic cell fate responses to drugs and mitogens</article-title>. <source>PLoS Comput. Biol.</source> <volume>14</volume> (<issue>3</issue>), <fpage>e1005985</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pcbi.1005985</pub-id>
</citation>
</ref>
<ref id="B5">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Erdem</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Gross</surname>
<given-names>S. M.</given-names>
</name>
<name>
<surname>Heiser</surname>
<given-names>L. M.</given-names>
</name>
<name>
<surname>Birtwistle</surname>
<given-names>M. R.</given-names>
</name>
</person-group> <article-title>Multi-Omics Binary Integration via Lasso Ensembles (MOBILE) for identification of context-specific networks and new regulatory mechanisms</article-title>. <comment>bioRxiv</comment>. <year>2022</year>.</citation>
</ref>
<ref id="B6">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Erdem</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Mutsuddy</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Bensman</surname>
<given-names>E. M.</given-names>
</name>
<name>
<surname>Dodd</surname>
<given-names>W. B.</given-names>
</name>
<name>
<surname>Saint-Antoine</surname>
<given-names>M. M.</given-names>
</name>
<name>
<surname>Bouhaddou</surname>
<given-names>M.</given-names>
</name>
<etal/>
</person-group> (<year>2022</year>). <article-title>A scalable, open-source implementation of a large-scale mechanistic model for single cell proliferation and death signaling</article-title>. <source>Nat. Commun.</source> <volume>13</volume> (<issue>1</issue>), <fpage>3555</fpage>&#x2013;<lpage>3618</lpage>. <pub-id pub-id-type="doi">10.1038/s41467-022-31138-1</pub-id>
</citation>
</ref>
<ref id="B7">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Erdem</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Nagle</surname>
<given-names>A. M.</given-names>
</name>
<name>
<surname>Casa</surname>
<given-names>A. J.</given-names>
</name>
<name>
<surname>Litzenburger</surname>
<given-names>B. C.</given-names>
</name>
<name>
<surname>Wangfen</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Taylor</surname>
<given-names>D. L.</given-names>
</name>
<etal/>
</person-group> (<year>2016</year>). <article-title>Proteomic screening and Lasso regression reveal differential signaling in insulin and insulin-like growth factor I (IGF1) pathways</article-title>. <source>Mol. Cell. Proteomics</source> <volume>15</volume> (<issue>9</issue>), <fpage>3045</fpage>&#x2013;<lpage>3057</lpage>. <pub-id pub-id-type="doi">10.1074/mcp.M115.057729</pub-id>
</citation>
</ref>
<ref id="B8">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Fr&#xf6;hlich</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Kessler</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Weindl</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Shadrin</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Schmiester</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Hache</surname>
<given-names>H.</given-names>
</name>
<etal/>
</person-group> (<year>2018</year>). <article-title>Efficient parameter estimation enables the prediction of drug response using a mechanistic pan-cancer pathway model</article-title>. <source>Cell Syst.</source> <volume>7</volume> (<issue>6</issue>), <fpage>567</fpage>&#x2013;<lpage>579.e6</lpage>. <pub-id pub-id-type="doi">10.1016/j.cels.2018.10.013</pub-id>
</citation>
</ref>
<ref id="B9">
<citation citation-type="web">
<person-group person-group-type="author">
<name>
<surname>Fr&#xf6;hlich</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Weindl</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Sch&#xe4;lte</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Pathirana</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Paszkowski</surname>
<given-names>&#x141;.</given-names>
</name>
<name>
<surname>Lines</surname>
<given-names>G. T.</given-names>
</name>
<etal/>
</person-group> (<year>2020</year>). <article-title>Amici: High-performance sensitivity analysis for large ordinary differential equation models</article-title>. <comment>arXiv:201209122 [q-bio] [Internet] Available from: <ext-link ext-link-type="uri" xlink:href="http://arxiv.org/abs/2012.09122">http://arxiv.org/abs/2012.09122</ext-link>.</comment>
</citation>
</ref>
<ref id="B10">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Gong</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Chehrazi-Raffle</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Reddi</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Salgia</surname>
<given-names>R.</given-names>
</name>
</person-group> (<year>2018</year>). <article-title>Development of PD-1 and PD-L1 inhibitors as a form of cancer immunotherapy: A comprehensive review of registration trials and future considerations</article-title>. <source>J. Immunother. cancer</source> <volume>6</volume> (<issue>1</issue>), <fpage>8</fpage>. <pub-id pub-id-type="doi">10.1186/s40425-018-0316-z</pub-id>
</citation>
</ref>
<ref id="B11">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Gross</surname>
<given-names>S. M.</given-names>
</name>
<name>
<surname>Dane</surname>
<given-names>M. A.</given-names>
</name>
<name>
<surname>Smith</surname>
<given-names>R. L.</given-names>
</name>
<name>
<surname>Devlin</surname>
<given-names>K. L.</given-names>
</name>
<name>
<surname>McLean</surname>
<given-names>I. C.</given-names>
</name>
<name>
<surname>Derrick</surname>
<given-names>D. S.</given-names>
</name>
<etal/>
</person-group> (<year>2022</year>). <article-title>A multi-omic analysis of MCF10A cells provides a resource for integrative assessment of ligand-mediated molecular and phenotypic responses</article-title>. <source>Commun. Biol.</source> <volume>5</volume> (<issue>1</issue>), <fpage>1066</fpage>. <pub-id pub-id-type="doi">10.1038/s42003-022-03975-9</pub-id>
</citation>
</ref>
<ref id="B12">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Gyori</surname>
<given-names>B. M.</given-names>
</name>
<name>
<surname>Bachman</surname>
<given-names>J. A.</given-names>
</name>
<name>
<surname>Subramanian</surname>
<given-names>K.</given-names>
</name>
<name>
<surname>Muhlich</surname>
<given-names>J. L.</given-names>
</name>
<name>
<surname>Galescu</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Sorger</surname>
<given-names>P. K.</given-names>
</name>
</person-group> (<year>2017</year>). <article-title>From word models to executable models of signaling networks using automated assembly</article-title>. <source>Mol. Syst. Biol.</source> <volume>13</volume> (<issue>11</issue>), <fpage>954</fpage>. <pub-id pub-id-type="doi">10.15252/msb.20177651</pub-id>
</citation>
</ref>
<ref id="B13">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Harris</surname>
<given-names>L. A.</given-names>
</name>
<name>
<surname>Hogg</surname>
<given-names>J. S.</given-names>
</name>
<name>
<surname>Tapia</surname>
<given-names>J. J.</given-names>
</name>
<name>
<surname>Sekar</surname>
<given-names>J. A.</given-names>
</name>
<name>
<surname>Gupta</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Korsunsky</surname>
<given-names>I.</given-names>
</name>
<etal/>
</person-group> (<year>2016</year>). <article-title>BioNetGen 2.2: Advances in rule-based modeling</article-title>. <source>Bioinformatics</source> <volume>32</volume>, <fpage>3366</fpage>&#x2013;<lpage>3368</lpage>. <pub-id pub-id-type="doi">10.1093/bioinformatics/btw469</pub-id>
</citation>
</ref>
<ref id="B14">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Hoadley</surname>
<given-names>K. A.</given-names>
</name>
<name>
<surname>Yau</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Hinoue</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Wolf</surname>
<given-names>D. M.</given-names>
</name>
<name>
<surname>Lazar</surname>
<given-names>A. J.</given-names>
</name>
<name>
<surname>Drill</surname>
<given-names>E.</given-names>
</name>
<etal/>
</person-group> (<year>2018</year>). <article-title>Cell-of-Origin patterns dominate the molecular classification of 10,000 tumors from 33 types of cancer</article-title>. <source>Cell</source> <volume>173</volume> (<issue>2</issue>), <fpage>291</fpage>&#x2013;<lpage>304.e6</lpage>. <pub-id pub-id-type="doi">10.1016/j.cell.2018.03.022</pub-id>
</citation>
</ref>
<ref id="B15">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Hucka</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Finney</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Sauro</surname>
<given-names>H. M.</given-names>
</name>
<name>
<surname>Bolouri</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Doyle</surname>
<given-names>J. C.</given-names>
</name>
<name>
<surname>Kitano</surname>
<given-names>H.</given-names>
</name>
<etal/>
</person-group> (<year>2003</year>). <article-title>The systems biology markup language (SBML): A medium for representation and exchange of biochemical network models</article-title>. <source>Bioinformatics</source> <volume>19</volume>, <fpage>524</fpage>&#x2013;<lpage>531</lpage>. <pub-id pub-id-type="doi">10.1093/bioinformatics/btg015</pub-id>
</citation>
</ref>
<ref id="B16">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Hulot</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Lalo&#xeb;</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Jaffr&#xe9;zic</surname>
<given-names>F.</given-names>
</name>
</person-group> (<year>2021</year>). <article-title>A unified framework for the integration of multiple hierarchical clusterings or networks from multi-source data</article-title>. <source>BMC Bioinforma.</source> <volume>22</volume> (<issue>1</issue>), <fpage>392</fpage>. <pub-id pub-id-type="doi">10.1186/s12859-021-04303-4</pub-id>
</citation>
</ref>
<ref id="B17">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Ju</surname>
<given-names>X.</given-names>
</name>
<name>
<surname>Zhang</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Zhou</surname>
<given-names>Z.</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>Q.</given-names>
</name>
</person-group> (<year>2020</year>). <article-title>Regulation of PD-L1 expression in cancer and clinical implications in immunotherapy</article-title>. <source>Am. J. Cancer Res.</source> <volume>10</volume> (<issue>1</issue>), <fpage>1</fpage>&#x2013;<lpage>11</lpage>.</citation>
</ref>
<ref id="B18">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Keating</surname>
<given-names>S. M.</given-names>
</name>
<name>
<surname>Waltemath</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>K&#xf6;nig</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Zhang</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Dr&#xe4;ger</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Chaouiya</surname>
<given-names>C.</given-names>
</name>
<etal/>
</person-group> (<year>2020</year>). <article-title>SBML level 3: An extensible format for the exchange and reuse of biological models</article-title>. <source>Mol. Syst. Biol.</source> <volume>16</volume> (<issue>8</issue>). <pub-id pub-id-type="doi">10.15252/msb.20199110</pub-id>
</citation>
</ref>
<ref id="B19">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>Kluyver</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Ragan-Kelley</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>P&#xe9;rez</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Granger</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Bussonnier</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Frederic</surname>
<given-names>J.</given-names>
</name>
<etal/>
</person-group> (<year>2016</year>). &#x201c;<article-title>Jupyter notebooks &#x2013; A publishing format for reproducible computational workflows</article-title>,&#x201d; in <source>Positioning and power in academic publishing: Players, agents and agendas</source>. Editors <person-group person-group-type="editor">
<name>
<surname>Loizides</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Schmidt</surname>
<given-names>B.</given-names>
</name>
</person-group> (<publisher-loc>Amsterdam</publisher-loc>: <publisher-name>IOS Press</publisher-name>), <fpage>87</fpage>&#x2013;<lpage>90</lpage>.</citation>
</ref>
<ref id="B20">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Lin</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Zhang</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Li</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Calhoun</surname>
<given-names>V. D.</given-names>
</name>
<name>
<surname>Deng</surname>
<given-names>H. W.</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>Y. P.</given-names>
</name>
</person-group> (<year>2013</year>). <article-title>Group sparse canonical correlation analysis for genomic data integration</article-title>. <source>BMC Bioinforma.</source> <volume>14</volume> (<issue>1</issue>), <fpage>245</fpage>. <pub-id pub-id-type="doi">10.1186/1471-2105-14-245</pub-id>
</citation>
</ref>
<ref id="B21">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Lopez</surname>
<given-names>C. F.</given-names>
</name>
<name>
<surname>Muhlich</surname>
<given-names>J. L.</given-names>
</name>
<name>
<surname>Bachman</surname>
<given-names>J. A.</given-names>
</name>
<name>
<surname>Sorger</surname>
<given-names>P. K.</given-names>
</name>
</person-group> (<year>2013</year>). <article-title>Programming biological models in Python using PySB</article-title>. <source>Mol. Syst. Biol.</source> <volume>9</volume>, <fpage>646</fpage>. <pub-id pub-id-type="doi">10.1038/msb.2013.1</pub-id>
</citation>
</ref>
<ref id="B22">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Malta</surname>
<given-names>T. M.</given-names>
</name>
<name>
<surname>Sokolov</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Gentles</surname>
<given-names>A. J.</given-names>
</name>
<name>
<surname>Burzykowski</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Poisson</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Weinstein</surname>
<given-names>J. N.</given-names>
</name>
<etal/>
</person-group> (<year>2018</year>). <article-title>Machine learning identifies stemness features associated with oncogenic dedifferentiation</article-title>. <source>Cell</source> <volume>173</volume> (<issue>2</issue>), <fpage>338</fpage>&#x2013;<lpage>354.e15</lpage>. <pub-id pub-id-type="doi">10.1016/j.cell.2018.03.034</pub-id>
</citation>
</ref>
<ref id="B23">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Mariette</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Villa-Vialaneix</surname>
<given-names>N.</given-names>
</name>
</person-group> (<year>2018</year>). <article-title>Unsupervised multiple kernel learning for heterogeneous data integration</article-title>. <source>Bioinformatics</source> <volume>34</volume> (<issue>6</issue>), <fpage>1009</fpage>&#x2013;<lpage>1015</lpage>. <pub-id pub-id-type="doi">10.1093/bioinformatics/btx682</pub-id>
</citation>
</ref>
<ref id="B24">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Mei</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Chen</surname>
<given-names>M. J. M.</given-names>
</name>
<name>
<surname>Liang</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Ma</surname>
<given-names>L.</given-names>
</name>
</person-group> (<year>2021</year>). <article-title>A four-gene signature predicts survival and anti-CTLA4 immunotherapeutic responses based on immune classification of melanoma</article-title>. <source>Commun. Biol.</source> <volume>4</volume> (<issue>1</issue>), <fpage>383</fpage>. <pub-id pub-id-type="doi">10.1038/s42003-021-01911-x</pub-id>
</citation>
</ref>
<ref id="B25">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Min</surname>
<given-names>W.</given-names>
</name>
<name>
<surname>Chang</surname>
<given-names>T. H.</given-names>
</name>
<name>
<surname>Zhang</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Wan</surname>
<given-names>X.</given-names>
</name>
</person-group> (<year>2021</year>). <article-title>Tscca: A tensor sparse cca method for detecting microRNA-gene patterns from multiple cancers</article-title>. <source>PLoS Comput. Biol.</source> <volume>17</volume> (<issue>6</issue>), <fpage>e1009044</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pcbi.1009044</pub-id>
</citation>
</ref>
<ref id="B26">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>M&#xfc;nzner</surname>
<given-names>U.</given-names>
</name>
<name>
<surname>Klipp</surname>
<given-names>E.</given-names>
</name>
<name>
<surname>Krantz</surname>
<given-names>M.</given-names>
</name>
</person-group> (<year>2019</year>). <article-title>A comprehensive, mechanistically detailed, and executable model of the cell division cycle in <italic>Saccharomyces cerevisiae</italic>
</article-title>. <source>Nat. Commun.</source> <volume>10</volume> (<issue>1</issue>), <fpage>1308</fpage>. <pub-id pub-id-type="doi">10.1038/s41467-019-08903-w</pub-id>
</citation>
</ref>
<ref id="B27">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Nusinow</surname>
<given-names>D. P.</given-names>
</name>
<name>
<surname>Szpyt</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Ghandi</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Rose</surname>
<given-names>C. M.</given-names>
</name>
<name>
<surname>McDonald</surname>
<given-names>E. R.</given-names>
</name>
<name>
<surname>Kalocsay</surname>
<given-names>M.</given-names>
</name>
<etal/>
</person-group> (<year>2020</year>). <article-title>Quantitative proteomics of the cancer cell line encyclopedia</article-title>. <source>Cell</source> <volume>180</volume> (<issue>2</issue>), <fpage>387</fpage>&#x2013;<lpage>402.e16</lpage>. <pub-id pub-id-type="doi">10.1016/j.cell.2019.12.023</pub-id>
</citation>
</ref>
<ref id="B28">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Park</surname>
<given-names>C. Y.</given-names>
</name>
<name>
<surname>Krishnan</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Zhu</surname>
<given-names>Q.</given-names>
</name>
<name>
<surname>Wong</surname>
<given-names>A. K.</given-names>
</name>
<name>
<surname>Lee</surname>
<given-names>Y. S.</given-names>
</name>
<name>
<surname>Troyanskaya</surname>
<given-names>O. G.</given-names>
</name>
</person-group> (<year>2015</year>). <article-title>Tissue-aware data integration approach for the inference of pathway interactions in metazoan organisms</article-title>. <source>Bioinformatics</source> <volume>31</volume> (<issue>7</issue>), <fpage>1093</fpage>&#x2013;<lpage>1101</lpage>. <pub-id pub-id-type="doi">10.1093/bioinformatics/btu786</pub-id>
</citation>
</ref>
<ref id="B29">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Saez-Rodriguez</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Bl&#xfc;thgen</surname>
<given-names>N.</given-names>
</name>
</person-group> (<year>2020</year>). <article-title>Personalized signaling models for personalized treatments</article-title>. <source>Mol. Syst. Biol.</source> <volume>16</volume> (<issue>1</issue>). <pub-id pub-id-type="doi">10.15252/msb.20199042</pub-id>
</citation>
</ref>
<ref id="B30">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Schwanh&#xe4;usser</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Busse</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Li</surname>
<given-names>N.</given-names>
</name>
<name>
<surname>Dittmar</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Schuchhardt</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Wolf</surname>
<given-names>J.</given-names>
</name>
<etal/>
</person-group> (<year>2011</year>). <article-title>Global quantification of mammalian gene expression control</article-title>. <source>Nature</source> <volume>473</volume> (<issue>7347</issue>), <fpage>337</fpage>&#x2013;<lpage>342</lpage>. <pub-id pub-id-type="doi">10.1038/nature10098</pub-id>
</citation>
</ref>
<ref id="B31">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Smith</surname>
<given-names>L. P.</given-names>
</name>
<name>
<surname>Bergmann</surname>
<given-names>F. T.</given-names>
</name>
<name>
<surname>Chandran</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Sauro</surname>
<given-names>H. M.</given-names>
</name>
</person-group> (<year>2009</year>). <article-title>Antimony: A modular model definition language</article-title>. <source>Bioinformatics</source> <volume>25</volume> (<issue>18</issue>), <fpage>2452</fpage>&#x2013;<lpage>2454</lpage>. <pub-id pub-id-type="doi">10.1093/bioinformatics/btp401</pub-id>
</citation>
</ref>
<ref id="B32">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Subramanian</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Narayan</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Corsello</surname>
<given-names>S. M.</given-names>
</name>
<name>
<surname>Peck</surname>
<given-names>D. D.</given-names>
</name>
<name>
<surname>Natoli</surname>
<given-names>T. E.</given-names>
</name>
<name>
<surname>Lu</surname>
<given-names>X.</given-names>
</name>
<etal/>
</person-group> (<year>2017</year>). <article-title>A next generation connectivity map: L1000 platform and the first 1,000,000 profiles</article-title>. <source>Cell</source> <volume>171</volume> (<issue>6</issue>), <fpage>1437</fpage>&#x2013;<lpage>1452.e17</lpage>. <pub-id pub-id-type="doi">10.1016/j.cell.2017.10.049</pub-id>
</citation>
</ref>
<ref id="B33">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Thiem</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Hesbacher</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Kneitz</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>di Primio</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Heppt</surname>
<given-names>M. V.</given-names>
</name>
<name>
<surname>Hermanns</surname>
<given-names>H. M.</given-names>
</name>
<etal/>
</person-group> (<year>2019</year>). <article-title>IFN-gamma-induced PD-L1 expression in melanoma depends on p53 expression</article-title>. <source>J. Exp. Clin. Cancer Res.</source> <volume>38</volume> (<issue>1</issue>), <fpage>397</fpage>. <pub-id pub-id-type="doi">10.1186/s13046-019-1403-9</pub-id>
</citation>
</ref>
<ref id="B34">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Tibshirani</surname>
<given-names>R.</given-names>
</name>
</person-group> (<year>1996</year>). <article-title>Regression shrinkage and selection via the Lasso</article-title>. <source>J. R. Stat. Soc. Ser. B-Methodological</source> <volume>58</volume>, <fpage>267</fpage>&#x2013;<lpage>288</lpage>. <pub-id pub-id-type="doi">10.1111/j.2517-6161.1996.tb02080.x</pub-id>
</citation>
</ref>
<ref id="B35">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Uhlen</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Fagerberg</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Hallstrom</surname>
<given-names>B. M.</given-names>
</name>
<name>
<surname>Lindskog</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Oksvold</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Mardinoglu</surname>
<given-names>A.</given-names>
</name>
<etal/>
</person-group> (<year>2015</year>). <article-title>Tissue-based map of the human proteome</article-title>. <source>Science</source> <volume>347</volume> (<issue>6220</issue>), <fpage>1260419</fpage>. <pub-id pub-id-type="doi">10.1126/science.1260419</pub-id>
</citation>
</ref>
<ref id="B36">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Walian</surname>
<given-names>P. J.</given-names>
</name>
<name>
<surname>Hang</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Mao</surname>
<given-names>J. H.</given-names>
</name>
</person-group> (<year>2016</year>). <article-title>Prognostic significance of FAM83D gene expression across human cancer types</article-title>. <source>Oncotarget</source> <volume>7</volume> (<issue>3</issue>), <fpage>3332</fpage>&#x2013;<lpage>3340</lpage>. <pub-id pub-id-type="doi">10.18632/oncotarget.6620</pub-id>
</citation>
</ref>
<ref id="B37">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Wang</surname>
<given-names>Z.</given-names>
</name>
<name>
<surname>Liu</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Zhang</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Zhang</surname>
<given-names>W.</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>W.</given-names>
</name>
<name>
<surname>Curr</surname>
<given-names>K.</given-names>
</name>
<etal/>
</person-group> (<year>2013</year>). <article-title>FAM83D promotes cell proliferation and motility by downregulating tumor suppressor gene FBXW7</article-title>. <source>Oncotarget</source> <volume>4</volume> (<issue>12</issue>), <fpage>2476</fpage>&#x2013;<lpage>2486</lpage>. <pub-id pub-id-type="doi">10.18632/oncotarget.1581</pub-id>
</citation>
</ref>
<ref id="B38">
<citation citation-type="web">
<person-group person-group-type="author">
<name>
<surname>Weindl</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Fr&#xf6;hlich</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Stapor</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Sch&#xe4;lte</surname>
<given-names>Y.</given-names>
</name>
</person-group> (<year>2020</year>). <article-title>ICB-DCM/AMICI: AMICI v0.11.2</article-title>. <comment>Zenodo[cited 2020 Jul 27] Available from: <ext-link ext-link-type="uri" xlink:href="https://zenodo.org/record/3949231">https://zenodo.org/record/3949231</ext-link>.</comment>
</citation>
</ref>
<ref id="B39">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Wishart</surname>
<given-names>D. S.</given-names>
</name>
<name>
<surname>Feunang</surname>
<given-names>Y. D.</given-names>
</name>
<name>
<surname>Guo</surname>
<given-names>A. C.</given-names>
</name>
<name>
<surname>Lo</surname>
<given-names>E. J.</given-names>
</name>
<name>
<surname>Marcu</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Grant</surname>
<given-names>J. R.</given-names>
</name>
<etal/>
</person-group> (<year>2018</year>). <article-title>DrugBank 5.0: A major update to the DrugBank database for 2018</article-title>. <source>Nucleic Acids Res.</source> <volume>46</volume> (<issue>D1</issue>), <fpage>D1074</fpage>&#x2013;<lpage>D1082</lpage>. <pub-id pub-id-type="doi">10.1093/nar/gkx1037</pub-id>
</citation>
</ref>
<ref id="B40">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Wong</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Yip</surname>
<given-names>S.</given-names>
</name>
</person-group> (<year>2018</year>). <article-title>Machine learning classifies cancer</article-title>. <source>Nature</source> <volume>555</volume> (<issue>7697</issue>), <fpage>446</fpage>&#x2013;<lpage>447</lpage>. <pub-id pub-id-type="doi">10.1038/d41586-018-02881-7</pub-id>
</citation>
</ref>
<ref id="B41">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Wu</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Chen</surname>
<given-names>W.</given-names>
</name>
<name>
<surname>Xu</surname>
<given-names>Z. P.</given-names>
</name>
<name>
<surname>Gu</surname>
<given-names>W.</given-names>
</name>
</person-group> (<year>2019</year>). <article-title>PD-L1 distribution and perspective for cancer immunotherapy&#x2014;blockade, knockdown, or inhibition</article-title>. <source>Front. Immunol.</source> <volume>10</volume>, <fpage>2022</fpage>. <pub-id pub-id-type="doi">10.3389/fimmu.2019.02022</pub-id>
</citation>
</ref>
<ref id="B42">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Xu</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>Z.</given-names>
</name>
<name>
<surname>Dong</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Wu</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Liao</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Li</surname>
<given-names>X.</given-names>
</name>
</person-group> (<year>2020</year>). <article-title>Chloride intracellular channel protein 2: Prognostic marker and correlation with PD-1/PD-L1 in breast cancer</article-title>. <source>Aging</source> <volume>12</volume> (<issue>17</issue>), <fpage>17305</fpage>&#x2013;<lpage>17327</lpage>. <pub-id pub-id-type="doi">10.18632/aging.103712</pub-id>
</citation>
</ref>
<ref id="B43">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Yamada</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Shiono</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Joo</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Yoshimura</surname>
<given-names>A.</given-names>
</name>
</person-group> (<year>2003</year>). <article-title>Control mechanism of JAK/STAT signal transduction pathway</article-title>. <source>FEBS Lett.</source> <volume>534</volume> (<issue>1&#x2013;3</issue>), <fpage>190</fpage>&#x2013;<lpage>196</lpage>. <pub-id pub-id-type="doi">10.1016/s0014-5793(02)03842-5</pub-id>
</citation>
</ref>
<ref id="B44">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Yang</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Cao</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>He</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Cui</surname>
<given-names>Y.</given-names>
</name>
</person-group> (<year>2018</year>). <article-title>Multilevel heterogeneous omics data integration with kernel fusion</article-title>. <source>Briefings Bioinforma.</source> <volume>2018</volume>, <fpage>bby115</fpage>. <pub-id pub-id-type="doi">10.1093/bib/bby115</pub-id>
</citation>
</ref>
<ref id="B45">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Yang</surname>
<given-names>J. H.</given-names>
</name>
<name>
<surname>Wright</surname>
<given-names>S. N.</given-names>
</name>
<name>
<surname>Hamblin</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>McCloskey</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Alcantar</surname>
<given-names>M. A.</given-names>
</name>
<name>
<surname>Schr&#xfc;bbers</surname>
<given-names>L.</given-names>
</name>
<etal/>
</person-group> (<year>2019</year>). <article-title>A white-box machine learning approach for revealing antibiotic mechanisms of action</article-title>. <source>Cell</source> <volume>177</volume> (<issue>6</issue>), <fpage>1649</fpage>&#x2013;<lpage>1661.e9</lpage>. <pub-id pub-id-type="doi">10.1016/j.cell.2019.04.016</pub-id>
</citation>
</ref>
<ref id="B46">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Yu</surname>
<given-names>M. K.</given-names>
</name>
<name>
<surname>Ma</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Fisher</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Kreisberg</surname>
<given-names>J. F.</given-names>
</name>
<name>
<surname>Raphael</surname>
<given-names>B. J.</given-names>
</name>
<name>
<surname>Ideker</surname>
<given-names>T.</given-names>
</name>
</person-group> (<year>2018</year>). <article-title>Visible machine learning for biomedicine</article-title>. <source>Cell</source> <volume>173</volume> (<issue>7</issue>), <fpage>1562</fpage>&#x2013;<lpage>1565</lpage>. <pub-id pub-id-type="doi">10.1016/j.cell.2018.05.056</pub-id>
</citation>
</ref>
<ref id="B47">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Zhang</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Lv</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Jin</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Cheng</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Fu</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Yuan</surname>
<given-names>D.</given-names>
</name>
<etal/>
</person-group> (<year>2018</year>). <article-title>Deep learning-based multi-omics data integration reveals two prognostic subtypes in high-risk neuroblastoma</article-title>. <source>Front. Genet.</source> <volume>9</volume>, <fpage>477</fpage>. <pub-id pub-id-type="doi">10.3389/fgene.2018.00477</pub-id>
</citation>
</ref>
<ref id="B48">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Zhou</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>X.</given-names>
</name>
<name>
<surname>Liu</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Meng</surname>
<given-names>Q.</given-names>
</name>
<name>
<surname>Yu</surname>
<given-names>Y.</given-names>
</name>
</person-group> (<year>2020</year>). <article-title>FAM83A drives PD-L1 expression via ERK signaling and FAM83A/PD-L1 co-expression correlates with poor prognosis in lung adenocarcinoma</article-title>. <source>Int. J. Clin. Oncol.</source> <volume>25</volume> (<issue>9</issue>), <fpage>1612</fpage>&#x2013;<lpage>1623</lpage>. <pub-id pub-id-type="doi">10.1007/s10147-020-01696-9</pub-id>
</citation>
</ref>
</ref-list>
</back>
</article>