<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article article-type="research-article" dtd-version="2.3" xml:lang="EN" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Mol. Biosci.</journal-id>
<journal-title>Frontiers in Molecular Biosciences</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Mol. Biosci.</abbrev-journal-title>
<issn pub-type="epub">2296-889X</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="publisher-id">1197154</article-id>
<article-id pub-id-type="doi">10.3389/fmolb.2023.1197154</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Molecular Biosciences</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Allosterically coupled conformational dynamics in solution prepare the sterol transfer protein StarD4 to release its cargo upon interaction with target membranes</article-title>
<alt-title alt-title-type="left-running-head">Xie and Weinstein</alt-title>
<alt-title alt-title-type="right-running-head">
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/fmolb.2023.1197154">10.3389/fmolb.2023.1197154</ext-link>
</alt-title>
</title-group>
<contrib-group>
<contrib contrib-type="author">
<name>
<surname>Xie</surname>
<given-names>Hengyi</given-names>
</name>
<xref ref-type="aff" rid="aff1">
<sup>1</sup>
</xref>
</contrib>
<contrib contrib-type="author" corresp="yes">
<name>
<surname>Weinstein</surname>
<given-names>Harel</given-names>
</name>
<xref ref-type="aff" rid="aff1">
<sup>1</sup>
</xref>
<xref ref-type="aff" rid="aff2">
<sup>2</sup>
</xref>
<xref ref-type="corresp" rid="c001">&#x2a;</xref>
<uri xlink:href="https://loop.frontiersin.org/people/22970/overview"/>
</contrib>
</contrib-group>
<aff id="aff1">
<sup>1</sup>
<institution>Department of Physiology and Biophysics</institution>, <institution>Weill Cornell Medicine</institution>, <addr-line>New York</addr-line>, <addr-line>NY</addr-line>, <country>United States</country>
</aff>
<aff id="aff2">
<sup>2</sup>
<institution>Institute for Computational Biomedicine</institution>, <institution>Weill Cornell Medicine</institution>, <addr-line>New York</addr-line>, <addr-line>NY</addr-line>, <country>United States</country>
</aff>
<author-notes>
<fn fn-type="edited-by">
<p>
<bold>Edited by:</bold> <ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/60221/overview">Lu Shaoyong</ext-link>, Shanghai Jiao Tong University, China</p>
</fn>
<fn fn-type="edited-by">
<p>
<bold>Reviewed by:</bold> <ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/807021/overview">Qin Xu</ext-link>, Shanghai Jiao Tong University, China</p>
<p>
<ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/759936/overview">Ruo-Xu Gu</ext-link>, Shanghai Jiao Tong University, China</p>
<p>
<ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/1786509/overview">Durba Sengupta</ext-link>, National Chemical Laboratory (CSIR), India</p>
<p>
<ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/715541/overview">Jinan Wang</ext-link>, University of Kansas, United States</p>
</fn>
<corresp id="c001">&#x2a;Correspondence: Harel Weinstein, <email>haw2002@med.cornell.edu</email>
</corresp>
</author-notes>
<pub-date pub-type="epub">
<day>18</day>
<month>05</month>
<year>2023</year>
</pub-date>
<pub-date pub-type="collection">
<year>2023</year>
</pub-date>
<volume>10</volume>
<elocation-id>1197154</elocation-id>
<history>
<date date-type="received">
<day>30</day>
<month>03</month>
<year>2023</year>
</date>
<date date-type="accepted">
<day>04</day>
<month>05</month>
<year>2023</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#xa9; 2023 Xie and Weinstein.</copyright-statement>
<copyright-year>2023</copyright-year>
<copyright-holder>Xie and Weinstein</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/">
<p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p>
</license>
</permissions>
<abstract>
<p>Complex mechanisms regulate the cellular distribution of cholesterol, a critical component of eukaryote membranes involved in regulation of membrane protein functions directly and through the physiochemical properties of membranes. StarD4, a member of the steroidogenic acute regulator-related lipid-transfer (StART) domain (StARD)-containing protein family, is a highly efficient sterol-specific transfer protein involved in cholesterol homeostasis. Its mechanism of cargo loading and release remains unknown despite recent insights into the key role of phosphatidylinositol phosphates in modulating its interactions with target membranes. We have used large-scale atomistic Molecular dynamics (MD) simulations to study how the dynamics of cholesterol bound to the StarD4 protein can affect interaction with target membranes, and cargo delivery. We identify the two major cholesterol (CHL) binding modes in the hydrophobic pocket of StarD4, one near S136&#x26;S147 (the Ser-mode), and another closer to the putative release gate located near W171, R92&#x26;Y117 (the Trp-mode). We show that conformational changes of StarD4 associated directly with the transition between these binding modes facilitate the opening of the gate. To understand the dynamics of this connection we apply a machine-learning algorithm for the detection of rare events in MD trajectories (RED), which reveals the structural motifs involved in the opening of a front gate and a back corridor in the StarD4 structure occurring together with the spontaneous transition of CHL from the Ser-mode of binding to the Trp-mode. Further analysis of MD trajectory data with the information-theory based NbIT method reveals the allosteric network connecting the CHL binding site to the functionally important structural components of the gate and corridor. Mutations of residues in the allosteric network are shown to affect the performance of the allosteric connection. These findings outline an allosteric mechanism which prepares the CHL-bound StarD4 to release and deliver the cargo when it is bound to the target membrane.</p>
</abstract>
<kwd-group>
<kwd>molecular dynamics (MD) simulations</kwd>
<kwd>cholesterol traffic and distribution in cells</kwd>
<kwd>N-body information theory analysis of MD trajectories</kwd>
<kwd>detection of rare events in MD trajectories</kwd>
<kwd>allosteric network coordination</kwd>
<kwd>machine learning analysis of MD trajectories</kwd>
<kwd>ligand-induced conformational changes</kwd>
<kwd>allosteric channel</kwd>
</kwd-group>
<custom-meta-wrap>
<custom-meta>
<meta-name>section-at-acceptance</meta-name>
<meta-value>Biological Modeling and Simulation</meta-value>
</custom-meta>
</custom-meta-wrap>
</article-meta>
</front>
<body>
<sec id="s1">
<title>Introduction</title>
<p>Cholesterol is a critical component of mammalian cell membranes, involved in the regulation of membrane protein function both through direct protein-CHL interactions and through the effect of CHL on the physiochemical properties of the host membranes. It is heterogeneously distributed among cellular organelles: the plasma membrane (PM) and the endocytic recycling compartment (ERC) are major pools of the total cellular cholesterol, whereas the endoplasmic reticulum (ER), where the cholesterol concentration is sensed and regulated, and the cellular sterol homoeostasis is maintained (<xref ref-type="bibr" rid="B25">Liscum and Munn, 1999</xref>; <xref ref-type="bibr" rid="B39">Radhakrishnan et al., 2008</xref>; <xref ref-type="bibr" rid="B16">Iaea and Maxfield, 2015</xref>), contains only a low amount of cellular cholesterol (0.1%&#x2013;2%).</p>
<p>Subcellular distribution of CHL occurs through both vesicular and non-vesicular transport mechanisms, the latter accounting for 70% of the cholesterol transport (<xref ref-type="bibr" rid="B16">Iaea and Maxfield, 2015</xref>; <xref ref-type="bibr" rid="B12">Hao et al., 2002</xref>; <xref ref-type="bibr" rid="B15">Iaea et al., 2017</xref>). Rapid non-vesicular transport requires lipid transfer proteins that provide the hydrophobic environment needed to accommodate lipid transferring across the aqueous phase (<xref ref-type="bibr" rid="B38">Prinz, 2007</xref>; <xref ref-type="bibr" rid="B53">Wong et al., 2019</xref>). One major family of sterol transfer proteins is the (START) domain family of the steroidogenic acute regulatory protein (StAR)-related lipid-transfer. A START domain (STARD) is composed of &#x223c;210&#xa0;amino acids that fold into an &#x3b1;/&#x3b2; helix-grip structure to create an internal hydrophobic cavity for lipid binding and recognition as illustrated in <xref ref-type="fig" rid="F1">Figure 1A</xref> (<xref ref-type="bibr" rid="B26">Mathieu et al., 2002</xref>; <xref ref-type="bibr" rid="B22">Letourneau et al., 2015</xref>). The mammalian STARD protein family comprises 15 proteins, subdivided into six subfamilies based on domain architecture and ligand specificity (<xref ref-type="bibr" rid="B1">Alpy and Tomasetto, 2005</xref>; <xref ref-type="bibr" rid="B28">Mesmin et al., 2011</xref>; <xref ref-type="bibr" rid="B53">Wong et al., 2019</xref>; <xref ref-type="bibr" rid="B5">Clark, 2020</xref>). The characteristic of the STARD4 subfamily is a single soluble sterol-binding START domain. The StarD4 protein is well-adapted to bind and transport sterol (<xref ref-type="bibr" rid="B42">Romanowski et al., 2002</xref>; <xref ref-type="bibr" rid="B14">Iaea et al., 2015</xref>) and is widely expressed in multiple tissues (<xref ref-type="bibr" rid="B48">Soccio et al., 2002</xref>; <xref ref-type="bibr" rid="B9">Elbadawy et al., 2011</xref>; <xref ref-type="bibr" rid="B40">Rodriguez-Agudo et al., 2011</xref>). StarD4 knockdown was shown to result in a decrease of (i) ER cholesterol concentration, (ii) acyl-CoA:cholesterol acyl-transferase 1 (ACAT1) activity, and (iii) cellular cholesterol ester concentration (<xref ref-type="bibr" rid="B11">Garbarino et al., 2012</xref>). The non-vesicular sterol transport kinetics between PM, ERC and ER was reduced by 33% (<xref ref-type="bibr" rid="B17">Iaea et al., 2020</xref>). Conversely, StarD4 overexpression resulted in facilitated intracellular cholesteryl ester accumulation in a ACAT1 dependent manner (<xref ref-type="bibr" rid="B41">Rodriguez-Agudo et al., 2008</xref>; <xref ref-type="bibr" rid="B28">Mesmin et al., 2011</xref>).</p>
<fig id="F1" position="float">
<label>FIGURE 1</label>
<caption>
<p>Two modes of CHL binding in the hydrophobic pocket of StarD4. The binding of cholesterol was modeled computationally in the crystal structure of StarD4 (PDB ID: 1jss) <bold>(A)</bold> The &#x3b1;/&#x3b2; helix-grip structure of the STARD domain in the crystal structure of StarD4. StarD4 is rendered in gray, with the C-terminal Helix (Helix4) in orange, the loop between &#x3b2;5 and &#x3b2;6 (&#x3a9;1 loop) in yellow, the &#x3b2;1&#x26;2 sheets in blue, and the &#x3b2;8 loop and &#x3b2;9 loop in pink. Cholesterol binding sites, residues S136, S147, W171, R92 and Y117, are rendered in &#x201c;licorice&#x201d; which draws the atoms as spheres and the bonds as cylinders, with oxygen colored in red, nitrogen in blue and carbons in cyan (hydrogens are not shown). <bold>(B, C)</bold> The &#x201c;cavity coordinate&#x201d; of CHL (rendered in VDW in tan color) and StarD4 residues reveal two modes of interaction. <bold>(B)</bold> Definition of the &#x201c;cavity coordinate&#x201d;: With the CHL-StarD4 complex rotated first to align the H4 along the x coordinate and position the Center of Mass (CoM) of the protein right above the CoM of H4, the &#x201c;cavity coordinate&#x201d; is defined by taking the center of mass of S136&#x26;S147 as the origin and measuring the distance along the direction of the angle bisector of x- and <italic>z</italic>-axes. <bold>(C)</bold> Frequency histogram of the cavity coordinate of the hydroxyl oxygen in CHL (filled), and of the sidechain oxygen or nitrogen atoms in S136, S147, W171, R92 and Y117 (hollow). <bold>(D&#x2013;G)</bold> Binding modes of CHL in the CHL-StarD4 complex: <bold>(D)</bold> the &#x201c;Ser-binding mode&#x201d;; <bold>(E)</bold> the &#x201c;water-bridged Ser-binding mode&#x201d;; <bold>(F)</bold> the &#x201c;Trp-binding mode&#x201d;; <bold>(G)</bold> &#x201c;water-bridged Trp-binding mode&#x201d;. StarD4 is shown in the same representation as in <bold>(A)</bold>. CHL is rendered in VDW with the oxygen in red, and other atoms in tan color. Water molecules directly involved in hydrogen bonds with the CHL and binding-sites are shown in licorice rendering. Hydrogen bonds between CHL and binding-sites are indicated with dashed lines.</p>
</caption>
<graphic xlink:href="fmolb-10-1197154-g001.tif"/>
</fig>
<p>Despite extensive studies of STARD protein function, no structure of a START domain complexed with a sterol has been resolved. Thus, the ligand binding modes of the START domain remain undetermined, and the molecular mechanism that underlies the lipid specificity and determines the lipid uptake and release pathways, is still unclear. Studies of sterol binding in START domains that employed primarily docking and short MD simulation have identified the cholesterol binding pocket and residues likely to be involved in ligand binding (<xref ref-type="fig" rid="F1">Figure 1A</xref>) (<xref ref-type="bibr" rid="B52">Tsujishita and andHurley, 2000</xref>; <xref ref-type="bibr" rid="B26">Mathieu et al., 2002</xref>; <xref ref-type="bibr" rid="B30">Murcia et al., 2006</xref>; <xref ref-type="bibr" rid="B50">Thorsell et al., 2011</xref>; <xref ref-type="bibr" rid="B49">Tan et al., 2019</xref>). In these studies, sterol was reported to favor a binding mode where the CHL inserts into the depths of the hydrophobic cavity and the hydroxyl group forms interactions with polar side chains or backbone carbonyls of residues at the bottom of the pocket. In <xref ref-type="fig" rid="F1">Figure 1A</xref> residues S136 and S147 are labeled in the binding pocket because they correspond to the predicted CHL binding sites in StarD3 (S362 and R351) and StarD5 (S132).</p>
<p>The recently determined structures of LAM proteins in the StARkin superfamily show that the sterol hydroxyl group makes hydrogen bonds with residues located in the upper two-thirds of the hydrophobic tunnel, and the sterol hydrocarbon tail is exposed to the solvent through a partially open lid (<xref ref-type="bibr" rid="B13">Horenkamp et al., 2018</xref>; <xref ref-type="bibr" rid="B18">Jentsch et al., 2018</xref>; <xref ref-type="bibr" rid="B51">Tong et al., 2018</xref>). Structural comparison between yeast LAM and human StarD4 also suggested that conformational changes at Helix4 and &#x3a9;1 loop (the loop between &#x3b2;5 and &#x3b2;6) (<xref ref-type="fig" rid="F1">Figure 1A</xref>) are required for the apo-StarD4 crystal structure to accommodate a cholesterol ligand at the same binding site as LAM protein (<xref ref-type="bibr" rid="B49">Tan et al., 2019</xref>).</p>
<p>In the present study we have used extensive (&#x223c;0.4 milliseconds) MD simulation trajectories to explore the binding modes of cholesterol in the pocket of StarD4 in solution in order to (1) assess the structural relationships between the <italic>apo</italic> and <italic>holo</italic> states of the protein, (2) identify the dynamic rearrangements required to accommodate the binding of CHL, and (3) the relation of these dynamic changes to the formation of the protein state required for its functionally productive membrane interactions. The long trajectories revealed a rich dynamic landscape of the protein structure in which the bound CHL adopts positions and configurations suggesting preparations for its CHL trafficking functions. Such function-related conformational changes of the StarD4 protein and its complex with CHL were revealed from the analysis of the MD trajectories with a machine learning-based Rare Event Detection (RED) protocol (<xref ref-type="bibr" rid="B36">Plante and Weinstein, 2021</xref>). To reveal the mechanisms underlying the function-related conformational changes we explored the allosteric pathways connecting the CHL repositioning in the binding site, with the identified conformational changes by applying the N-body Information Theory (NbIT) analysis (<xref ref-type="bibr" rid="B23">LeVine et al., 2014</xref>; <xref ref-type="bibr" rid="B24">LeVine and Weinstein, 2014</xref>) to the MD trajectories. We show here that the allosteric mechanism resulting from this analysis involves the coupling between the cholesterol translocation dynamics in the binding site and configurational changes of specific regions of the structure that are involved in the interaction of StarD4 with the membrane (<xref ref-type="bibr" rid="B54">Zhang et al., 2022</xref>). The detailed structure-based information about the conformational changes and the allosteric channel enabled the development of specific testable hypotheses for mutations that we used here to probe the molecular mechanisms of function of a CHL-loaded StarD4 diffusing in the aqueous medium of the cytosol, and mutations that affect the molecular rearrangements that prepare the pathway for sterol release when the protein is embedding in a target membrane as we have shown (<xref ref-type="bibr" rid="B54">Zhang et al., 2022</xref>).</p>
</sec>
<sec sec-type="results" id="s2">
<title>Results</title>
<sec id="s2-1">
<title>MD simulations reveal two modes of cholesterol binding in StarD4</title>
<p>The binding modes of CHL in the pocket identified previously (<xref ref-type="bibr" rid="B26">Mathieu et al., 2002</xref>; <xref ref-type="bibr" rid="B5">Clark, 2020</xref>) were explored with extensive atomistic MD simulations comprised of 12 statistically independent replicates of 33.6 &#x3bc;s each for a total of 403&#xa0;&#xb5;s, starting from the crystal structure of StarD4 (PDBid: 1JSS) with a CHL molecule docked in the binding site using the Schrodinger Induced Fit Docking protocol (<xref ref-type="bibr" rid="B45">Sherman et al., 2006a</xref>; <xref ref-type="bibr" rid="B44">Sherman et al., 2006b</xref>) (see Methods for details). The simulations produced two different modes of CHL binding in the hydrophobic pocket (<xref ref-type="fig" rid="F1">Figure 1B</xref>), one with the hydroxyl group located near Ser136/Ser147 (termed &#x201c;Ser-binding mode&#x201d;), the other with the CHL OH near Trp171 (&#x201c;Trp-binding mode&#x201d;). In the &#x201c;Ser-binding mode&#x201d;, the cholesterol hydroxyl group forms hydrogen bonds with the Ser136 and Ser147 sidechains either directly, or through a water bridge (<xref ref-type="fig" rid="F1">Figures 1C, D</xref>). In the &#x201c;Trp-binding mode&#x201d;, the cholesterol hydroxyl group interacts either directly with Trp171, or resides between Trp171, Arg92 and Tyr117 with which it interacts through H-bonds mediated by a water molecule (<xref ref-type="fig" rid="F1">Figures 1E, F</xref>). Spontaneous transitions of cholesterol between binding modes are observed in more than half of the trajectories (<xref ref-type="sec" rid="s9">Supplementary Figure S1</xref>).</p>
</sec>
<sec id="s2-2">
<title>Transitions between the two modes of cholesterol binding are concurrent with conformational changes of the StarD4 protein</title>
<p>To reveal the sequence of conformational events that accompany the transitions between the Ser and Trp binding modes, we applied the recently developed Rare Event Detection (RED) protocol (<xref ref-type="bibr" rid="B36">Plante and Weinstein, 2021</xref>) as described in Methods. The mechanistically important state-to-state transitions of StarD4 in response to the translocation of ligand cholesterol (see ref. 31 for a discussion of the relation between rare events and function) were extracted from the dynamic information in an ensemble of 6 trajectory stretches (4&#x2013;8 &#x3bc;s each, 34.4 &#x3bc;s total) from different trajectory replicas in which a stable cholesterol transition from one binding mode to another was observed. This ensemble of 6 trajectories served as the input training data for the RED protocol in identifying the rearrangement events shared among the different independent replicas.</p>
<p>The RED algorithm decomposed the trajectories into 5 components based on the evolution of the residue-residue contact map over time (see Methods). As described in ref (<xref ref-type="bibr" rid="B36">Plante and Weinstein, 2021</xref>), each component represents a state of StarD4 in which a particular conformation becomes dominant, at a particular time in the trajectory. This time-ordered series of events in which different components dominate the structural changes encoded in the trajectory, is shown in <xref ref-type="fig" rid="F2">Figures 2A,B</xref>. Several components identified by the RED algorithm are observed to be dominant concurrently with the repositioning of cholesterol in the binding site (<xref ref-type="fig" rid="F2">Figures 2A&#x2013;D</xref>). This is remarkable, because the information provided to the RED algorithm consists only of the contact matrices for each frame in the trajectory, without direct information about CHL and its position. The correspondence between the binding modes and particular RED components is illustrated in the time frames by the correspondence of the times when component A (&#x201c;comp A&#x201d; in <xref ref-type="fig" rid="F2">Figures 2A,B</xref>) is dominant, and CHL is in the Ser-binding mode of CHL (<xref ref-type="fig" rid="F2">Figures 2C,D</xref>). Moreover, as the dominance of the A component wanes and is replaced by that of comp-s B and C, CHL is in the Trp-binding mode. Thus, in every independent trajectory the transition events from comp A to either comp B or comp C coincide with the translocation of cholesterol (Traj0&#x26;3 shown in <xref ref-type="fig" rid="F2">Figures 2A&#x2013;D</xref>, and Traj1,4,6,7 in <xref ref-type="sec" rid="s9">Supplementary Figure S2A</xref>). This suggests a mechanism of coupled dynamics in the cholesterol-StarD4 complex, connecting the structural rearrangements of the protein frame with the transition of CHL in the binding pocket. The states associated with the transition of CHL to the Trp-binding mode are of particular interest because in this position the CHL is near a putative &#x201c;exit gate&#x201d; from the binding pocket.</p>
<fig id="F2" position="float">
<label>FIGURE 2</label>
<caption>
<p>RED analysis revealed a set of simultaneous conformational changes concurrent with the translocation of CHL. <bold>(A, B)</bold> The time evolution of the RED-detected normalized temporal weight of structural components over the simulation time of trajectory 0 <bold>(A)</bold> and 3 <bold>(B)</bold>. Color code: comp <bold>(A)</bold> (orange), comp <bold>(B)</bold> (blue), comp <bold>(C)</bold> (purple), comp <bold>(D)</bold> (green), comp <bold>(E)</bold> (red). At each time point, the color column represents the stacked values of the weights for all the components beneath. For example, at time &#x3d; 2700&#xa0;ns in traj03 [left dashed line in (B)] the weight of comp A is 0.67, comp C is 0.12, comp D is 0.22, and comp B&#x26;E are 0; at time &#x3d; 3300&#xa0;ns (right dashed line), the weight of comp A is 0.06, comp B is 0, comp C is 0.66, comp D is 0.23 and comp E is 0.05. <bold>(C, D)</bold> The time evolution of the cholesterol binding mode, represented by the change in the distances between CHL and the Ser binding site (in green) and to the Trp site (blue) for: <bold>(C)</bold> in trajectory 0; <bold>(D)</bold> in trajectory 3. <bold>(E&#x2013;G)</bold> Spatial arrays of components A, B, and C show the contact features characterized by the structure-differentiating contact pairs (SDCPs). Along the X-axis there are 3,880 data points, each representing one residue pair. The Y coordinate of each data point shows the contribution of the residue pair to the contact feature of the component (see Methods). The SDCPs (summarized in <xref ref-type="sec" rid="s9">Supplementary Table S1</xref>) are grouped by secondary structure location and highlighted in colors. Each E&#x2013;G shows a conformational state for which the determinant conformational features are encoded by the contact patterns of structure-differentiating contact pairs (SDCPs). For example, in comp B (in <bold>(F)</bold> residue pair groups 1&#x2013;5 lost contact, while groups 6&#x2013;8 have formed contacts. The text defines the structural elements containing the pairs collected in the SDCP groups. Together, these groups contain the SDCPs that define the structural feature in the time segment when comp B is prevalent in determining the conformational state.Note that the contributions in the RED spatial component analysis (<xref ref-type="fig" rid="F2">Figures 2E&#x2013;G</xref>) are not constrained to values lower than 1, because they are relative values indicating the contribution of residue pairs in comparison to the constructive residue pairs. Constructive residue pairs are those remaining in contact in all conformational states, and thus they are assigned a constant contribution value of 1 in every component. In contrast, the structure-differentiating contact pairs have highly varying contribution values across the different components. In some components, these pairs may contribute more strongly than the constructive residue pairs, resulting in contribution values &#x3e; 1.</p>
</caption>
<graphic xlink:href="fmolb-10-1197154-g002.tif"/>
</fig>
<p>We propose that the specific structural rearrangements of the protein frame constitute a preparatory step for a subsequent functional docking of the loaded StarD4 to the membrane for the release of its cargo. A key indication for the likelihood of this mechanism is the consistency in the different simulation replicas of the conformational changes coupled to the transitions between CHL binding modes.</p>
<p>
<italic>RED analysis reveals the specific conformational changes that occur simultaneously:</italic> The conformational changes that are coupled to the CHL translocation toward the vicinity of the &#x201c;gate&#x201d; are identified from comparisons of the spatial arrays in the RED analysis. For this comparison we identify structure-differentiating contact pairs (SDCPs) that underlie the conformational changes that occur simultaneously during the rare event of transition from one component to the next. While numerous residue contacts included in a component may change from &#x201c;on&#x201d; (i.e., in contact) to &#x201c;off&#x201d; (i.e., no longer in contact; see protocol in Methods) an SDCP must satisfy two simultaneous criteria: the pair makes a substantial contribution to the contact feature in one component and a minimal contribution in the subsequent component in the transition. Both the substantial contribution, and the minimal contributions, are determinant elements of a particular component&#x2019;s conformational features, because the change in structure depends on both the formation of new contacts and the release of contacts from a previous conformation. Thus, they collectively determine the conformational change characteristic to the rare dynamic event which makes them essential for interpreting the RED results.</p>
<p>Comparison of the time axis in panels A and B of <xref ref-type="fig" rid="F2">Figure 2</xref> with that in panels C and D shows the transition between components A and B, and A and C occurs at the same time as the CHL transitions from the Ser-binding mode to the Trp-binding mode. The structural motifs involved in this transition are identified by the quantification of contributions to the components shown in <xref ref-type="fig" rid="F2">Figure 2E-G</xref>, and defined structurally in <xref ref-type="fig" rid="F3">Figure 3</xref>.</p>
<fig id="F3" position="float">
<label>FIGURE 3</label>
<caption>
<p>The translocation of CHL is concurrent with the opening of the gate and corridor to the hydrophobic pocket. The elements of motif movements are compared in representative conformation in the closed state <bold>(A&#x2013;C)</bold> (corresponding to the <bold>left</bold> dashed line in <xref ref-type="fig" rid="F2">Figure 2B</xref>), and in the opened state (corresponding to the <bold>right</bold> dashed line in <xref ref-type="fig" rid="F2">Figure 2B</xref>). StarD4 is shown in gray, and the CHL is rendered the same way as in <xref ref-type="fig" rid="F1">Figure 1</xref>. The SDCPs are shown in red with licorice representations: <bold>(A, D)</bold> the contact forming SDCPs (group 6, summarized in <xref ref-type="sec" rid="s9">Supplementary Table S1</xref>) between &#x3b2;1-H4; <bold>(B, E)</bold> the contact breaking SDCPs (group 2,3) between &#x3b2;9, H4 and &#x3a9;1; <bold>(C, F)</bold>. The contact breaking SDCPs (group 4) between &#x3b2;8-&#x3b2;9.</p>
</caption>
<graphic xlink:href="fmolb-10-1197154-g003.tif"/>
</fig>
<p>The SDCPs in <xref ref-type="fig" rid="F2">Figures 2E&#x2013;G</xref> are grouped (from 1 to 8) by their location in a particular secondary structure motif. Specifically, group 1 includes contacts between &#x3b2;1 &#x3b2;2 &#x26; the C-terminal Helix (H4); group 2 includes contacts between &#x3a9;1 &#x26; H4 (see <xref ref-type="fig" rid="F3">Figures 3B, E</xref>); group 3 includes contacts between &#x3a9;1 &#x26; &#x3b2;9 (<xref ref-type="fig" rid="F3">Figures 3B, E</xref>); group 4 has the contacts between &#x3b2;9 &#x26; &#x3b2;8 (<xref ref-type="fig" rid="F3">Figures 3C, F</xref>), and group 5 includes the interactions in the first turn of H4 (<xref ref-type="sec" rid="s9">Supplementary Figure S4A</xref>).</p>
<p>The complete list of SDCPs is given in <xref ref-type="sec" rid="s9">Supplementary Table S1</xref>, which summarizes the structural characteristics of the system at different stages in the trajectory. The structural analysis shows that SDCPs in groups 1 to 5 are in contact in comp A in which the CHL is in the Ser-bound mode (orange, <xref ref-type="fig" rid="F2">Figure 2E</xref>), but most of them break contact during the transition from comp A to comp B or C, and thus have low feature values in comp B and comp C where CHL is in the Trp-binding mode (orange, <xref ref-type="fig" rid="F2">Figure 2F, G</xref>).</p>
<p>Another set of SDCPs, 6 to 8, have low values in comp A (blue and purple in <xref ref-type="fig" rid="F2">Figure 2E</xref>) but high values in comp B and C (blue in <xref ref-type="fig" rid="F2">Figure 2F</xref> and purple in <xref ref-type="fig" rid="F2">Figure 2G</xref>). They are found to represent contacts formed only after the repositioning of the CHL. More generally, comp B and comp C share a highly similar (but not identical) set of structural specifics represented in the salient SDCPs in groups 6 to 8. The highly similar pattern of motif interactions represented by these shared SDCPs include the following: in group 6: interactions between &#x3b2;1 &#x3b2;2 &#x26; H4 (<xref ref-type="fig" rid="F3">Figures 3A, D</xref>); in group 7: interactions between &#x3b2;2 &#x26; &#x3b2;3 &#x26; &#x3b2;9; in group 8: interactions of &#x3b2;8 or &#x3b2;9 with &#x3a9;1 or H4. Thus, the summary in <xref ref-type="sec" rid="s9">Supplementary Table S1</xref> shows that groups 6-to-8 in comp B (B6-B8) and groups C6-C8 in comp C share most of the key secondary structural elements and their conformations.</p>
<p>The major difference between comp B and comp C is the partial unfolding of the N-terminal head of H4 (group 5; residues 199&#x2013;201) which loses some contacts in comp B but remains structured in comp C. Together, the structural features of components B or C compared to those of comp A reveal the specific conformational changes that are captured as &#x201c;rare events&#x201d;.</p>
<p>The main rare events identified by the transition from a conformation in which one component is dominant to a new conformation with a different dominant component, are the A&#x2192;B and A&#x2192;C transitions which are concurrent with the CHL translocation (<xref ref-type="fig" rid="F2">Figures 2A&#x2013;D</xref>). The comparison of the structural features in comp A and in comp-s B&#x26;C shows that the conformational changes concurrent with the CHL translocation are the opening the access to the hydrophobic pocket. These conformational changes include (1) the opening of a front gate by the distancing of the H4 N-terminus from the &#x3a9;1-loop and strengthening interactions with &#x3b2;1 (see <xref ref-type="fig" rid="F3">Figures 3A, B, D, E</xref>, and the detailed residue interaction data in Figures 8A,B); the opening of a back corridor between the &#x3b2;9 tail and the &#x3b2;7&#x3b2;8-loop (<xref ref-type="fig" rid="F3">Figures 3C, F</xref>).</p>
</sec>
<sec id="s2-3">
<title>The predominant changes in the conformational space of StarD4 emerge from concurrent dynamics of CHL binding mode switching and gate opening</title>
<p>To facilitate analysis of the metastable states and transition pathways encoded in the MD trajectory data, we built the conformational landscape of the cholesterol-bound StarD4 by applying the dimensionality reduction approach tICA: time-structure based independent component analysis. To describe the dynamics of cholesterol binding modes and of the protein conformational changes (see Methods and <xref ref-type="sec" rid="s9">Supplementary Table S2</xref>) we chose a set of 7 collective variables (CVs) to define the tICA space. The analysis showed that the first 3 tICA vectors describe 81% of the total dynamics of the system in the 403 &#x3bc;s production run trajectories (<xref ref-type="fig" rid="F4">Figures 4A,B</xref>). The projections on the space spanned by these three vectors are shown in <xref ref-type="fig" rid="F4">Figure 4C</xref>.</p>
<fig id="F4" position="float">
<label>FIGURE 4</label>
<caption>
<p>Representation of the 3D tICA space built on CVs describing the CHL binding modes in the pocket, and the protein conformational changes detected with the RED analysis. <bold>(A)</bold> Contribution of tICA eigenvectors to the total variations of the system contributed, showing that 81% of the information is captured by the first three tICA eigenvectors. <bold>(B)</bold> Contributions of individual CVs (which are identified on the Y-axis, and defined in <xref ref-type="sec" rid="s9">Supplementary Table S2</xref>) to the first three tICA eigenvectors: tIC 1 (red), tIC 2 (blue), tIC 3 (green). <bold>(C)</bold> 2D projections of the 3D space defined by the 3 tICA eigenvectors, shown as population density maps: tIC2-tIC1 (top left), tIC3-tIC1 (bottom left), tIC3-tIC2 (bottom right). The population densities are colored according to the code shown on the right of the first panel.</p>
</caption>
<graphic xlink:href="fmolb-10-1197154-g004.tif"/>
</fig>
<p>To create a Markov State Model (MSM) and analyze the transitions among the metastable states visited by the StarD4 protein, the 3D tICA space was first discretized into 200 microstates using k-means clustering (see Methods and specifics in <xref ref-type="sec" rid="s9">Supplementary Figure S5</xref>). The 200 microstates were then segregated according to similarity in kinetic properties, resulting in the identification of 9 macrostates. The 5 macrostates with high population are shown in <xref ref-type="fig" rid="F5">Figure 5A</xref>, and representative conformations obtained from structural analysis are shown in <xref ref-type="fig" rid="F5">Figures 5B&#x2013;D</xref> and <xref ref-type="sec" rid="s9">Supplementary Figures S6, S8</xref>.</p>
<fig id="F5" position="float">
<label>FIGURE 5</label>
<caption>
<p>Structural elements in the conformational changes of the CHL translocation coupled to the opening of the H4-&#x3a9;1 gate and &#x3b2;8-&#x3b2;9 corridor. <bold>(A)</bold> Macrostates 0&#x2013;4 are outlined in the 2D projections of the 3D conformational space: the tIC2-tIC1 plane (upper left), tIC3-tIC1 plane (lower left), tIC3-tIC2 plane (lower right). The gray shade represents the population density in the conformational space as shown in <xref ref-type="fig" rid="F4">Figure 4C</xref>. The populations (% of total frames) are listed at the upper right for each microstate. <bold>(B&#x2013;D)</bold> Structural models representing macrostates 1 to 3 viewed from front (left column) and back (right column). The StarD4-CHL complex is rendered in the same way as in <xref ref-type="fig" rid="F1">Figure 1</xref>. Red arrows in the second and third rows <bold>(C, D)</bold> indicate the conformational changes of the H4-&#x3a9;1 gate and the &#x3b2;8-&#x3b2;9 corridor, as well as the movement of CHL. <bold>(E)</bold> Structural characteristics of macrostates 0 to 4 shown represented by probability density histograms of the characteristic CV values that determine the tICA space, as defined in <xref ref-type="sec" rid="s9">Supplementary Table S2</xref>.</p>
</caption>
<graphic xlink:href="fmolb-10-1197154-g005.tif"/>
</fig>
<p>The structural interpretation of macrostates 0&#x2013;4 in <xref ref-type="fig" rid="F5">Figures 5B&#x2013;D</xref> and <xref ref-type="sec" rid="s9">Supplementary Figures S6, S8</xref> indicated that the differences reflect the dynamics of gate, corridor, and cholesterol we detected in the stable transition events, as indicated by the data in <xref ref-type="fig" rid="F5">Figure 5E</xref> showing the relevant CVs, and underscoring that these differences prevail throughout the conformational space.</p>
<p>Specifically, macrostate 1 in <xref ref-type="fig" rid="F5">Figure 5B</xref> represents conformational states where cholesterol is bound in the Ser-site mode (with low &#x201c;CHL-Ser-dist&#x201d; and high &#x201c;CHL-Tyr-dist&#x201d;, seen in the first and second panel of <xref ref-type="fig" rid="F5">Figure 5E</xref>), StarD4 is in the crystal-like conformations with a closed H4-&#x3a9;1 gate (high &#x201c;&#x3b2;1-H4-dist&#x201d;, low &#x201c;H4-&#x3a9;1-dist&#x201d; in the third and fourth panel of <xref ref-type="fig" rid="F5">Figure 5E</xref>) in which the &#x3b2;8-&#x3b2;9 corridor is closed (shown by low &#x201c;&#x3b2;6-&#x3b2;9-dist&#x201d; and high &#x201c;&#x3b2;9-&#x3b2;2-&#x3b2;3-dist&#x201d;&#x26;&#x201c;&#x3b2;23loop-H1&#x3b2;8-dist&#x201d; in the fifth to seventh panel in <xref ref-type="fig" rid="F5">Figure 5E</xref>). Macrostate 0 also represents conformational states with Ser-binding CHL in crystal-like StarD4 with gates closed, but with a conformation of the &#x3b2;23 loop that is extended away (<xref ref-type="sec" rid="s9">Supplementary Figure S6A</xref>) (high &#x201c;&#x3b2;23loop-H1&#x3b2;8-dist&#x201d;, seventh panel, <xref ref-type="fig" rid="F5">Figure 5E</xref>) which may hinder the opening of &#x3b2;8-&#x3b2;9 corridor. These Ser-binding/gates-closed conformations in Macrostates 1 and 0 contribute 30% of the populations in the conformation space.</p>
<p>The most populated macrostate is Macrostate 3, with 47% of the population. It contains conformational states in which CHL is in the Trp-binding mode with both the H4-&#x3a9;1 gate and &#x3b2;8-&#x3b2;9 corridor open, and with &#x3b2;9 leaning towards &#x3b2;2&#x26;&#x3b2;3 which rearrange into narrower conformations (<xref ref-type="fig" rid="F5">Figures 5D, E</xref>). Notably, all these observed rearrangements are consistent with the conformational changes detected by the RED analysis of the corresponding RED event. In Macrostate 4 the cholesterol is in a Trp-binding binding mode or even further down towards the H4-&#x3a9;1 gate that opens wider (<xref ref-type="fig" rid="F5">Figure 5E</xref>, <xref ref-type="sec" rid="s9">Supplementary Figure S6B</xref>). As indicated by the analysis, Macrostates 0,1,3,4 constitute 88% of the conformational space containing the concurrent dynamic rearrangements of the cholesterol binding mode, the opening H4-&#x3a9;1 gate, and the &#x3b2;8-&#x3b2;9 corridor.</p>
<p>The other structural motif movements that were most important in the Comp A to Comp B/C transition event detected with RED are also evidenced by the comparisons of CVs related to the transition between the Ser- and Trp-binding modes of CHL. These are quantified in <xref ref-type="sec" rid="s9">Supplementary Figure S7</xref> and include the unfolding in H4 as well as the straightening of H4, the narrowing of &#x3b2;2 and &#x3b2;3, and residue interactions between &#x3b2;1-H4 and disassociations between &#x3b2;8-&#x3b2;9. Thus, the conformational changes of the StarD4 protein that we found to occur simultaneously with the transitions between CHL binding modes, yield conformational states that are specific to the position of CHL in the binding pocket.</p>
<p>The time sequence of cholesterol translocation in the binding site is embedded in the MSM transition pathways starting from Macrostate 1 (crystal-like StarD4 with Ser-binding cholesterol) and ending in Macrostate 3 (gates-opened StarD4 with Trp-binding cholesterol). The fluxes along the most probable pathways (contributing &#x223c;95% of the transition flux) are summarized in <xref ref-type="table" rid="T1">Table 1</xref>. Each of the top two most probable paths contributes &#x223c;30% of the total flux. One is the direct transition from Macrostate 1 to Macrostate 3 which encompasses the simultaneous translocation of CHL and conformational change of StarD4, and the other involves an intermediate, the less populated Macrostate 2 in which CHL is in Ser-binding modes but with both H4-&#x3a9;1 gate and &#x3b2;8-&#x3b2;9 corridor open (<xref ref-type="fig" rid="F5">Figures 5C, E</xref>). These results are consistent with the structural analysis in a report (<xref ref-type="bibr" rid="B49">Tan et al., 2019</xref>) showing that conformational changes at the H4-&#x3a9;1 gate are required for cholesterol to bind lower in the hydrophobic pocket, due to steric hindrance.</p>
<table-wrap id="T1" position="float">
<label>TABLE 1</label>
<caption>
<p>Top transition pathways in the cholesterol translocation process from Ser-binding state (State 1) to Trp-binding state (State 2).</p>
</caption>
<table>
<thead valign="top">
<tr>
<th align="left">Path</th>
<th align="left">Norm Flux</th>
<th align="left">Cumulative Flux</th>
</tr>
</thead>
<tbody valign="top">
<tr>
<td align="left">1 &#x2192; 2 &#x2192; 3</td>
<td align="left">0.3269</td>
<td align="left">0.3269</td>
</tr>
<tr>
<td align="left">1 &#x2192; 3</td>
<td align="left">0.2984</td>
<td align="left">0.6253</td>
</tr>
<tr>
<td align="left">1 &#x2192; 0 &#x2192; 7 &#x2192; 3</td>
<td align="left">0.1151</td>
<td align="left">0.7404</td>
</tr>
<tr>
<td align="left">1 &#x2192; 0 &#x2192; 2 &#x2192; 3</td>
<td align="left">0.0826</td>
<td align="left">0.8230</td>
</tr>
<tr>
<td align="left">1 &#x2192; 0 &#x2192; 3</td>
<td align="left">0.0464</td>
<td align="left">0.8754</td>
</tr>
<tr>
<td align="left">1 &#x2192; 5 &#x2192; 3</td>
<td align="left">0.0414</td>
<td align="left">0.9168</td>
</tr>
<tr>
<td align="left">1 &#x2192; 0 &#x2192; 2 &#x2192; 4 &#x2192; 3</td>
<td align="left">0.0258</td>
<td align="left">0.9426</td>
</tr>
<tr>
<td align="left">1 &#x2192; 5 &#x2192; 2 &#x2192; 4 &#x2192; 3</td>
<td align="left">0.0140</td>
<td align="left">0.9567</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>The less populated intermediate macrostates are summarized in <xref ref-type="sec" rid="s9">Supplementary Figure S8</xref>. They indicate the flexibility of the StarD4-cholesterol complex and support the relation between CHL position and the gate and corridor dynamics. Thus, with CHL in the Ser-binding mode, the gates-closed crystal-like conformations represent 30% of the population, much more than the ones with either the H4-&#x3a9;1 gate or the &#x3b2;8-&#x3b2;9 corridor open (10%). Most likely this is due to the energy cost of exposing the StarD4 hydrophobic pocket and cholesterol&#x2019;s hydrophobic tail to water. In contrast, when CHL is in the Trp-binding mode the H4-&#x3a9;1 gate is always open, and the conformations with both gate and corridor open amount to 58% compared to those with &#x3b2;8-&#x3b2;9 corridor closed, which are very unlikely (1.7%).</p>
</sec>
<sec id="s2-4">
<title>The allosteric pathway that couples cholesterol translocation in the hydrophobic pocket to the opening of access gates</title>
<p>Having identified a set of concurrent conformational changes in several structural motifs that include the cholesterol binding site, we applied NbIT analysis (see (<xref ref-type="bibr" rid="B23">LeVine et al., 2014</xref>; <xref ref-type="bibr" rid="B24">LeVine and Weinstein, 2014</xref>) and Methods) to reveal the allosteric network that connects them. NbIT analysis calculates the normalized coordination information (NCI), which describes how much information of the <italic>receiver</italic> system can be gained from information about the <italic>transmitter</italic> system (see (<xref ref-type="bibr" rid="B24">LeVine and Weinstein, 2014</xref>) and Method for details). Here we applied the NCI analysis to the trajectories from the StarD4-cholesterol complex simulations to quantify the information transmission between motifs that emerged from the RED analysis described above. The residue index in Method identifies the components of the structural motifs used in this analysis summarized in <xref ref-type="table" rid="T2">Table 2</xref>. In <xref ref-type="sec" rid="s9">Supplementary Table S3</xref> we report the structural motifs that were used as negative controls for which the NCI quantification is expected to reveal only background-level sharing of information with the cholesterol binding site.</p>
<table-wrap id="T2" position="float">
<label>TABLE 2</label>
<caption>
<p>Normalized coordination information between sites in cholesterol-StarD4 system.</p>
</caption>
<table>
<thead valign="top">
<tr>
<th align="center"/>
<th align="left"/>
<th colspan="10" align="center">Transmitter:</th>
</tr>
<tr>
<td align="left"/>
<td align="left"/>
<td align="center">&#x3b2;1</td>
<td align="center">&#x3b2;2&#x3b2;3</td>
<td align="center">&#x3b2;2&#x3b2;3 loop</td>
<td align="center">H4 head</td>
<td align="center">&#x3a9;1 loop</td>
<td align="center">&#x3b2;9</td>
<td align="center">&#x3b2;7&#x3b2;8loop-near&#x3b2;9</td>
<td align="center">&#x3b2;7&#x3b2;8loop-mid</td>
<td align="center">&#x3b2;7&#x3b2;8loop-near&#x3b2;6</td>
<td align="center">CHLsite</td>
</tr>
</thead>
<tbody valign="top">
<tr>
<td rowspan="10" align="center">Receiver</td>
<td align="center">&#x3b2;1</td>
<td align="center" style="background-color:#D9D9D9">19.48</td>
<td align="center" style="background-color:#FFCCCC">22.1%</td>
<td align="center">8.2%</td>
<td align="center">16.8%</td>
<td align="center">7.7%</td>
<td align="center">14.0%</td>
<td align="center">6.2%</td>
<td align="center">5.7%</td>
<td align="center">6.8%</td>
<td align="center">12.2%</td>
</tr>
<tr>
<td align="center">&#x3b2;2&#x3b2;3</td>
<td align="center" style="background-color:#FF6699">30.6%</td>
<td align="center" style="background-color:#D9D9D9">14.68</td>
<td align="center" style="background-color:#FF6699">31.1%</td>
<td align="center" style="background-color:#FFCCCC">24.3%</td>
<td align="center">13.5%</td>
<td align="center" style="background-color:#FFCCCC">20.8%</td>
<td align="center">7.7%</td>
<td align="center">6.5%</td>
<td align="center">7.3%</td>
<td align="center">18.4%</td>
</tr>
<tr>
<td align="center">&#x3b2;2&#x3b2;3loop</td>
<td align="center">10.0%</td>
<td align="center" style="background-color:#FF6699">35.4%</td>
<td align="center" style="background-color:#D9D9D9">10.11</td>
<td align="center">9.4%</td>
<td align="center">7.0%</td>
<td align="center">11.0%</td>
<td align="center">4.9%</td>
<td align="center">3.9%</td>
<td align="center">3.8%</td>
<td align="center">9.3%</td>
</tr>
<tr>
<td align="center">H4head</td>
<td align="center" style="background-color:#FFCCCC">26.5%</td>
<td align="center" style="background-color:#FFCCCC">27.9%</td>
<td align="center">9.9%</td>
<td align="center" style="background-color:#D9D9D9">16.50</td>
<td align="center">10.6%</td>
<td align="center">25.1%</td>
<td align="center">7.2%</td>
<td align="center">6.1%</td>
<td align="center">5.9%</td>
<td align="center">16.9%</td>
</tr>
<tr>
<td align="center">&#x3a9;1loop</td>
<td align="center">8.8%</td>
<td align="center">12.5%</td>
<td align="center">6.6%</td>
<td align="center">12.7%</td>
<td align="center" style="background-color:#D9D9D9">7.49</td>
<td align="center">11.8%</td>
<td align="center">6.0%</td>
<td align="center">6.6%</td>
<td align="center">6.4%</td>
<td align="center">18.5%</td>
</tr>
<tr>
<td align="center">&#x3b2;9</td>
<td align="center">16.0%</td>
<td align="center" style="background-color:#FFCCCC">20.6%</td>
<td align="center">12.6%</td>
<td align="center" style="background-color:#FFCCCC">22.0%</td>
<td align="center">8.8%</td>
<td align="center" style="background-color:#D9D9D9">6.45</td>
<td align="center">7.1%</td>
<td align="center">6.8%</td>
<td align="center">6.3%</td>
<td align="center">13.6%</td>
</tr>
<tr>
<td align="center">&#x3b2;7&#x3b2;8loop-near&#x3b2;9</td>
<td align="center">4.9%</td>
<td align="center">4.8%</td>
<td align="center">2.7%</td>
<td align="center">3.3%</td>
<td align="center">3.4%</td>
<td align="center">6.1%</td>
<td align="center" style="background-color:#D9D9D9">8.90</td>
<td align="center" style="background-color:#FF6699">32.9%</td>
<td align="center">18.8%</td>
<td align="center">10.9%</td>
</tr>
<tr>
<td align="center">&#x3b2;7&#x3b2;8loop-mid</td>
<td align="center">3.8%</td>
<td align="center">4.6%</td>
<td align="center">2.8%</td>
<td align="center">3.1%</td>
<td align="center">1.9%</td>
<td align="center">5.3%</td>
<td align="center" style="background-color:#FFCCCC">27.4%</td>
<td align="center" style="background-color:#D9D9D9">13.62</td>
<td align="center" style="background-color:#FF6699">38.1%</td>
<td align="center">5.9%</td>
</tr>
<tr>
<td align="center">&#x3b2;7&#x3b2;8loop-near&#x3b2;6</td>
<td align="center">4.5%</td>
<td align="center">4.2%</td>
<td align="center">3.4%</td>
<td align="center">3.5%</td>
<td align="center">1.5%</td>
<td align="center">5.2%</td>
<td align="center">18.4%</td>
<td align="center" style="background-color:#FF6699">38.0%</td>
<td align="center" style="background-color:#D9D9D9">9.57</td>
<td align="center">8.2%</td>
</tr>
<tr>
<td align="center">CHLsite</td>
<td align="center" style="background-color:#FF6699">35.4%</td>
<td align="center" style="background-color:#FF6699">36.0%</td>
<td align="center">19.7%</td>
<td align="center" style="background-color:#FF6699">29.4%</td>
<td align="center" style="background-color:#FF6699">29.8%</td>
<td align="center" style="background-color:#FFCCCC">21.4%</td>
<td align="center" style="background-color:#FFCCCC">22.6%</td>
<td align="center">19.1%</td>
<td align="center" style="background-color:#FF6699">31.6%</td>
<td align="center" style="background-color:#D9D9D9">2.13</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<fn id="Tfn1">
<label>
<sup>a</sup>
</label>
<p>The normalized coordination information (NCI) values are presented for NbIT, calculations with the <italic>Transmitter</italic> residues defining the columns acting and <italic>Receiver</italic> residues defining the rows. On the diagonal, the total correlation (TC) of the site is shown in gray.</p>
</fn>
</table-wrap-foot>
</table-wrap>
<p>The results from the NCI calculation show that the &#x3b2;1, &#x3b2;2&#x3b2;3, H4 head, and &#x3a9;1 loop motifs are strongly coordinated with the cholesterol binding site (S136, S147, W171, R92, Y117), while weaker coordination is found with the &#x3b2;9 and the &#x3b2;7&#x3b2;8 loop regions. Notably, the H4 head and &#x3a9;1 loop constitute the H4-&#x3a9;1 gate, and the &#x3b2;9 and the &#x3b2;7&#x3b2;8 loop constitute the &#x3b2;8-&#x3b2;9 corridor of the hydrophobic pocket. The stronger coordination of the CHL binding site by the H4-&#x3a9;1 gate than &#x3b2;8-&#x3b2;9 corridor is in accordance with the structural analysis in the previous section showing that the conformational changes and interaction changes at the H4-&#x3a9;1 gate always accompany CHL translocation to the Trp-binding state (<xref ref-type="fig" rid="F3">Figures 3</xref>, <xref ref-type="fig" rid="F5">5E</xref>)</p>
<p>Because &#x3b2;1 and &#x3b2;2&#x3b2;3 are also strong coordinators of the binding site despite the long distance separating these motifs, we calculated the <italic>normalized mutual coordination information</italic> (NMCI) to detect the coordination channel for information transmission from &#x3b2;1 to the CHL binding site. The calculated NMCI shows that &#x3b2;2&#x3b2;3, H4, and &#x3b2;9, all share the information with the binding site, which identifies them as components of the <italic>coordination channel</italic> shown in <xref ref-type="fig" rid="F6">Figure 6A</xref>. These results show how the coordination of cholesterol binding mode by &#x3b2;1 is mediated through the interaction of &#x3b2;1 with &#x3b2;2&#x3b2;3 and H4, which stabilizes the &#x3b2;9 and H4 in the gate opening conformation, as summarized in <xref ref-type="fig" rid="F6">Figure 6B</xref>
</p>
<fig id="F6" position="float">
<label>FIGURE 6</label>
<caption>
<p>The allosteric coordination channel between &#x3b2;1 and the CHL binding site involves Helix4, and the &#x3b2;2&#x3b2;3 and &#x3b2;9 sheets. <bold>(A)</bold> StarD4 is rendered in cartoon, with residues colored according to the normalized mutual coordination information (NCMI) value calculated with the NbIT information theory-based analysis (<xref ref-type="bibr" rid="B23">LeVine et al., 2014</xref>; <xref ref-type="bibr" rid="B24">LeVine and Weinstein, 2014</xref>). The quantification scale is at the lower right. The <italic>transmitter motif</italic> (&#x3b2;1 residues 47&#x2013;52) is colored in yellow, and the <italic>receiver motif</italic> (CHL binding site residues S136, S147, W171, R92 and Y117) is labeled in green cartoon on the secondary structure with side chains in green licorice. <bold>(B)</bold> An allosteric model of information transmission from &#x3b2;1 to the cholesterol binding site.</p>
</caption>
<graphic xlink:href="fmolb-10-1197154-g006.tif"/>
</fig>
</sec>
<sec id="s2-5">
<title>The probability and stability of cholestrol binding in the Trp-binding mode are reduced in the W171A mutant StarD4</title>
<p>The release pathway of CHL in the yeast sterol transport protein Osh4 (<xref ref-type="bibr" rid="B47">Singh et al., 2009</xref>) was shown to comprise step-wise interactions of the CHL hydroxyl with a sequence of hydrophilic residues inside the hydrophobic pocket. In StarD4 we observed a similar sequence of CHL occupying binding sites along the pocket, composed of the S136&#x26;S147-site (<xref ref-type="fig" rid="F1">Figure 1D</xref>), a water-mediated S136&#x26;S147-site (<xref ref-type="fig" rid="F1">Figure 1E</xref>), the W171-site (<xref ref-type="fig" rid="F1">Figure 1F</xref>), and a water-mediated W171, R92 &#x26; Y117 site (<xref ref-type="fig" rid="F1">Figure 1G</xref>). Given the coordination between CHL translocation in the binding site and specific changes in the conformation of the StarD4 protein identified from the detailed analysis of MD simulation trajectories we report, we probed the impact of impairing such a binding site by introducing a W171A mutation. We used a Free Energy Perturbation (FEP) protocol to introduce the mutation in representative conformations of the two most populated Macrostates (<xref ref-type="fig" rid="F5">Figure 5</xref>): State 1 with CHL in a Ser-binding mode, and State 3 where CHL interacts with W171 and is also involved in the water bridged interactions with R92 and Y117. From each binding mode we launched 24 replicate runs of 2,720ns/each, resulting in an ensemble of simulations totaling 131&#x3bc;s for the W171A-StarD4 mutant.</p>
<p>In simulations starting from the Ser-binding state, we find that the W171A mutation eliminated the translocation of CHL away from the Ser-sites, as the ligand is either in the Ser-binding mode or in the water-bridged Ser-binding mode. In the simulations started from the Trp-binding-like state where the CHL is interacting with R92 and Y117 (through a water-bridge), the W171A mutation destabilizes this binding mode and the CHL tends to translocate from the Trp-binding-like state back to the Ser-binding site. The backward transitions are observed in 8 out of the 24 replicas, with the Ser-binding conformation contributing 21% of the population in the ensemble started from the Trp-binding-like mode (<xref ref-type="sec" rid="s9">Supplementary Figure S9</xref>). In comparison, stable backward transitions were observed only 2 times in the 403 &#x3bc;s simulations of the WT StarD4 (<xref ref-type="sec" rid="s9">Supplementary Figure S1</xref>). Thus, W171 turns out to be an important intermediate site that enables accessibility and stability for CHL binding to the downstream binding sites around W171, R92 and Y117.</p>
<p>Notably, structural analysis suggests that the allosteric connection persists in the W171A StarD4-CHL complex and the conformation of StarD4 is compatible with the CHL in the Ser-binding mode. Thus, when CHL transitions back from the R92 Y117 water network to the Ser-binding site, the initially open &#x3a9;1-H4 gate closed, and the unraveled N-terminal of H4 folded back (<xref ref-type="fig" rid="F7">Figure 7</xref>; <xref ref-type="sec" rid="s9">Supplementary Figures S9B&#x2013;D</xref>). These conformational changes are consistent with the allosterically connected changes in WT StarD4 where the &#x3a9;1-H4 gate opening and the unfolding in H4 are found to be dominant features in the RED analysis only in the Trp-binding mode (<xref ref-type="fig" rid="F5">Figure 5E</xref>; <xref ref-type="sec" rid="s9">Supplementary Figure S7</xref>). Results from the NbIT analysis of the mutant confirm that the allosteric effect remains strong between the CHL and binding site to the peripheral transmitter motifs around H4 (<xref ref-type="sec" rid="s9">Supplementary Table S5</xref>).</p>
<fig id="F7" position="float">
<label>FIGURE 7</label>
<caption>
<p>&#x3b2;1 and H4 conformational changes are allosterically coupled with the translocation of CHL binding in the StarD4 W171A mutant. <bold>(A)</bold> Starting position of the CHL in the simulation of the W171A StarD4 structural model. CHL resides in the water-bridged binding site with R92 and Y117. <bold>(B)</bold> Shows that the CHL has transitioned back up the hydrophobic binding site to the Ser-binding site in which the gate and corridor are closed. StarD4 is rendered in gray, with the H4 in orange, the &#x3a9;1 loop in yellow, the &#x3b2;1&#x26;2 sheets in blue, and the &#x3b2;8 loop and &#x3b2;9 loop in pink. Residues S136 and S147 are rendered in &#x201c;licorice&#x201d; which draws the atoms as spheres and the bonds as cylinders, with oxygen colored in red, nitrogen in blue, carbons in cyan and hydrogens in white. CHL is rendered in tan color, in the VDW representation. The surface of residues 198 to 212 of H4 is rendered in transparent orange, and the surface of residues 121 to 130 on &#x3a9;1 loop is rendered in transparent yellow.</p>
</caption>
<graphic xlink:href="fmolb-10-1197154-g007.tif"/>
</fig>
</sec>
<sec id="s2-6">
<title>Mutations of K49 disrupt the allosteric network and slow down the dynamics of cholesterol transitions in the binding site</title>
<p>The mutation was chosen based on the results from analysis with the RED algorithm which revealed that the transitions of CHL in the binding site are coupled to a shift of the K49-S208 interaction to a K49-T204 configuration, and the establishment of a V51-A207 contact (<xref ref-type="sec" rid="s9">Supplementary Table S1</xref>). According to the allosteric model created with NbIT, &#x3b2;1 is an allosteric transmitter that coordinates the dynamics of bound CHL through its interaction with H4 and &#x3b2;2. The H4-&#x3b2;1 interaction involved in H4 repositioning is stabilized by both electrostatic and hydrophobic interactions. Thus, when the H4-&#x3b2;1gate is closed the K49 sidechain relocates near S208 and its hydrocarbon chain interacts with A207 (<xref ref-type="fig" rid="F8">Figure 8A</xref>). When the H4-&#x3b2;1gate opens, the K49 sidechain shifts closer to T204, enabling V51 to form a contact with A207 and remove this hydrophobic residue from an unfavorable aqueous environment (<xref ref-type="fig" rid="F8">Figures 8A, B</xref>). The NbIT coordination contribution analysis confirms that K49 is contributing the most to the allosteric interaction between &#x3b2;1 and the CHL binding site residues. Therefore, we mutated K49 expecting changes of the &#x3b2;1-H4 interaction mode that weaken the allosteric coordination.</p>
<fig id="F8" position="float">
<label>FIGURE 8</label>
<caption>
<p>The &#x3b2;1-H4 interactions that are allosterically coupled with CHL translocation in WT are destabilized in K49 mutants. Comparison of detailed changes in H4 and &#x3b2;1 residue interaction modes in <bold>(A)</bold> WT StarD4 in the initial Ser-binding state; <bold>(B)</bold> WT StarD4 after transition to the Trp-binding state, <bold>(C)</bold> K49W StarD4 after transition to the Trp-binding state, <bold>(D)</bold> K49A StarD4 after transition to the Trp-binding state. The StarD4 and CHL are rendered the same way as in <xref ref-type="fig" rid="F3">Figure 3</xref>. Residues K/W/A49, V51, D203, T204, A207 and S208 are rendered in &#x201c;licorice&#x201d; in blue, orange, red, green, purple and cyan, respectively.</p>
</caption>
<graphic xlink:href="fmolb-10-1197154-g008.tif"/>
</fig>
<p>Mutations K49A and K49W were introduced with the same FEP protocol as above into the StarD4-CHL complex in Microstates 1 (Ser-binding mode) and 3 (Trp-binding mode). Then, simulations were carried out for 24 replicas from each state for 2,320ns per replica of the K49A-StarD4 (111&#x3bc;s total), and 2,720ns per replica for K49W-StarD4 (131&#x3bc;s total). When started from the Ser-binding mode, the CHL transitioned to the Trp-binding site in 4 out of 24 trajectories for K49W-StarD4 (containing 7% of the population, <xref ref-type="sec" rid="s9">Supplementary Figure S10A</xref>), and in only 1 out of 24 trajectories for K49A-StarD4 (containing 2% of the population, <xref ref-type="sec" rid="s9">Supplementary Figure S10B</xref>). These results suggest that the dynamics of CHL are significantly slower in the mutants compared to the WT StarD4 in which 6 out of 12 trajectories sampled the transition within the first 2720ns. No backward transition from a Trp-binding state to a Ser-binding state was observed in the K49A and K49W mutant constructs. These results are consistent with decreased CHL dynamics in the K49W and in the K49A mutants.</p>
<p>To study the conformational changes and allosteric effect, adaptive sampling was carried out to expand the sampling along the reaction pathway of the K49A mutation for another 48 &#x3bc;s in 24 replicas (<xref ref-type="sec" rid="s9">Supplementary Figure S10C</xref>). The impact of the mutation on the allosteric mechanism of conformational changes in StarD4 is evident in the change of contact probability between residues in &#x3b2;1 and H4 as shown in <xref ref-type="table" rid="T3">Table 3</xref>.</p>
<table-wrap id="T3" position="float">
<label>TABLE 3</label>
<caption>
<p>Contact probability<sup>&#x2a;</sup> between &#x3b2;1 and H4 in WT, K49W and K49A StarD4 WT.</p>
</caption>
<table>
<thead valign="top">
<tr>
<th align="left" style="background-color:#8e8f92">Contact probability</th>
<th align="left" style="background-color:#8e8f92">K49&#x2013;T204 (%)</th>
<th align="left" style="background-color:#8e8f92">K49&#x2013;S208 (%)</th>
<th align="left" style="background-color:#8e8f92">V51&#x2013;A207 (%)</th>
</tr>
</thead>
<tbody valign="top">
<tr>
<td align="left">Ser-binding</td>
<td align="left">13.6</td>
<td align="left">68.8</td>
<td align="left">14.8</td>
</tr>
<tr>
<td align="left">Trp-binding</td>
<td align="left">83.3</td>
<td align="left">13.1</td>
<td align="left">75.4</td>
</tr>
<tr>
<td colspan="4" align="left">K49W</td>
</tr>
<tr>
<td align="left" style="background-color:#c6c7c9">Contact probability</td>
<td align="left" style="background-color:#c6c7c9">W49&#x2013;T204 (%)</td>
<td align="left" style="background-color:#c6c7c9">W49&#x2013;S208 (%)</td>
<td align="left" style="background-color:#c6c7c9">V51&#x2013;A207 (%)</td>
</tr>
<tr>
<td align="left">Ser-binding</td>
<td align="left">51.0</td>
<td align="left">77.1</td>
<td align="left">1.3</td>
</tr>
<tr>
<td align="left">Trp-binding</td>
<td align="left">91.4</td>
<td align="left">93.2</td>
<td align="left">0.3</td>
</tr>
<tr>
<td colspan="4" align="left">K49A</td>
</tr>
<tr>
<td align="left" style="background-color:#c6c7c9">Contact probability</td>
<td align="left" style="background-color:#c6c7c9">A49&#x2013;T204 (%)</td>
<td align="left" style="background-color:#c6c7c9">A49&#x2013;S208 (%)</td>
<td align="left" style="background-color:#c6c7c9">V51&#x2013;A207 (%)</td>
</tr>
<tr>
<td align="left">Ser-binding</td>
<td align="left">2.8</td>
<td align="left">10.9</td>
<td align="left">3.2</td>
</tr>
<tr>
<td align="left">Trp-binding</td>
<td align="left">76.5</td>
<td align="left">17.7</td>
<td align="left">39.6</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<fn id="Tfn2">
<label>
<sup>a</sup>
</label>
<p>Defined as the probability of the specified residue pairs having at least one atom in each at a distance &#x3c;3.5&#xc5;. The Probability is expressed as the percent occurrences of such contacts.</p>
</fn>
</table-wrap-foot>
</table-wrap>
<p>The resulting structures, compared in <xref ref-type="fig" rid="F8">Figure 8</xref>, show how the K49W and K49A mutations changed the mode of interaction between &#x3b2;1 and H4. In the WT StarD4, K49 establishes both hydrophobic and hydrophilic contact with A207, S208 and T204, and its translocation opens space for the formation of a V51-A207 contact pair (<xref ref-type="fig" rid="F8">Figures 8A, B</xref>). With the K49W mutation, the large W49 forms stable contacts with both S208 and T204 but hinders sterically the V51-A207 interaction (<xref ref-type="fig" rid="F8">Figure 8C</xref>; <xref ref-type="table" rid="T3">Table 3</xref>). The K49A mutation, on the other hand, resulted in a decreased contact rate with both S208 and T204, eliminated the hydrophilic interactions, and destabilized the V51-A207 contact (<xref ref-type="fig" rid="F8">Figure 8D</xref>;<xref ref-type="table" rid="T3">Table 3</xref>).</p>
<p>The interactions between &#x3b2;1 and H4 are weakened overall in both mutants but the interaction modes differ. The probability of contact is decreasing gradually from WT to K49W to K49A, and the electrostatic interaction is also gradually removed in the same order. Consistent with the weakening of interaction and the decrease of CHL dynamics, the NbIT results showed a decrease in coordination information (CI values) in the allosteric network between the &#x3b2;1, H4 and &#x3b2;9 motifs the K49W mutant (<xref ref-type="sec" rid="s9">Supplementary Table S6</xref>), and a loss of coordination effect from &#x3b2;1 to the CHL site in the K49A mutant (<xref ref-type="sec" rid="s9">Supplementary Table S7</xref>)&#x2013;as expected from our results from the analysis of WT StarD4.</p>
</sec>
</sec>
<sec sec-type="discussion" id="s3">
<title>Discussion</title>
<p>The function of StarD4 in vectorial transport of sterols between preferred organelles is reported to be regulated by anionic lipids (<xref ref-type="bibr" rid="B54">Zhang et al., 2022</xref>). To learn about the molecular mechanism that underlies the StarD4 sterol transport, and the role of phosphatidylinositol phosphate (PIP) anionic lipids in regulating its function, we explored the structural and dynamic properties of the protein and the binding modes of cholesterol (CHL). Our findings show that CHL can shuttle between two modes of stable binding in the hydrophobic pocket of StarD4, interacting as shown in <xref ref-type="fig" rid="F1">Figure 1</xref> either with Ser136 and Ser147 which are located at the upper end of the hydrophobic pocket, or with Trp171, Arg92 and Tyr117 which are lower in the hydrophobic pocket and thereby place the CHL closer to the exit gate (see also <xref ref-type="sec" rid="s9">Supplementary Figures S11A, B</xref>). An even lower position of the bound CHL was observed in the simulations (<xref ref-type="fig" rid="F1">Figure 1G</xref>), but it is visited only rarely because it triggers incipient water penetration from the aqueous environment which leads to energetically unfavorable interactions with the hydrophobic tail of CHL. Together, the three modes of CHL binding outline an energetically favorable path for the sterol in the functional steps of uptake and release when StarD4 is embedded in the membrane environment appropriate for these functions.</p>
<p>Indeed, our recent findings from a combined experimental/computational study (<xref ref-type="bibr" rid="B54">Zhang et al., 2022</xref>) showed that StarD4 can preferentially extract sterol from liposome membranes containing certain PIPs (especially, PI(4,5)P<sub>2</sub> and to a lesser degree PI(3,5)P<sub>2</sub>), while enhancement of transport was less effective when the same PIPs were present in the acceptor membranes. Moreover, we showed that StarD4 recognizes membrane-specific PIPs through specific interaction with the geometry of the PIP headgroup as well as the surrounding membrane environment.</p>
<p>Our key mechanistic finding of the connection between the movement of CHL among its binding modes in the hydrophobic pocket, and the movements of specific structural motifs of StarD4, emerged from the analysis of the very long MD trajectories with the time-resolved rare event detection algorithm RED (<xref ref-type="bibr" rid="B36">Plante and Weinstein, 2021</xref>). These trajectory analyses showed that the connected conformational changes at the &#x3b2;9 C-terminal and H4 N-terminal portions result in the opening of a gate between H4 and &#x3a9;1-loop, and a corridor between &#x3b2;9 and &#x3b2;7&#x3b2;8-loop. From the time-resolved RED analysis we were able to ascertain that these gate-opening/gate-closing conformational changes occur simultaneously with the transitions between CHL binding modes, with the opening corresponding to the position of CHL in the binding pocket near Trp171, Arg92 and Tyr117.</p>
<p>The opening of &#x3b2;8-&#x3b2;9 corridor and the partial unraveling of Helix4 N-terminal head (res199-201) in the Trp-binding state make the lower end of the hydrophobic pocket accessible to the external environment. Although being observed here directly for the first time, these function-related conformational changes are consistent with results from a previous study using a membrane penetration assay to study the environment of M206 (<xref ref-type="bibr" rid="B14">Iaea et al., 2015</xref>). M206 is a residue located at the lower half of the hydrophobic pocket adjacent to the gate. In the crystal structure (PDBid:1JSS), it is facing inwards to the hydrophobic environment. However, in the absence of membrane, the iodoacetamide-NBD (IANBD) probe at the M206C site gave a signal similar to the L124C site immersed in water, which suggests that the M206 at the inner side of H4 is accessible to the external water environment. With the introduction of liposomes, the probe at the M206C site was found to be inserted into the nonpolar environment of the membrane. The demonstrated accessibility of M206 to the external water or membrane environments supports the inference that the Trp-binding state with opened gates and local unraveling in H4 represent an intermediate state of StarD4 in an aqueous environment that is also likely to be an intermediate state along the CHL release/uptake pathways when StarD4 becomes embedded in a target membrane.</p>
<p>That conformational changes of the H4-&#x3a9;1gate are related to sterol release has long been proposed for StarD-protein family members (<xref ref-type="bibr" rid="B30">Murcia et al., 2006</xref>; <xref ref-type="bibr" rid="B14">Iaea et al., 2015</xref>; <xref ref-type="bibr" rid="B21">Letourneau et al., 2016</xref>; <xref ref-type="bibr" rid="B49">Tan et al., 2019</xref>), but the allosteric path and exact conformational changes remained unknown. While the &#x3a9;1 loop was suggested to be essential for the opening of the gate, the structural comparison analysis shown in (<xref ref-type="bibr" rid="B21">Letourneau et al., 2016</xref>; <xref ref-type="bibr" rid="B49">Tan et al., 2019</xref>) suggests as well a change in H4&#x2013;as seen in our study&#x2013;likely as a result of steric interaction with the CHL at the Trp-binding site.</p>
<p>To understand the dynamics of the mechanism observed in the RED analysis, we employed the information-theory-based analysis NbIT (<xref ref-type="bibr" rid="B23">LeVine et al., 2014</xref>; <xref ref-type="bibr" rid="B24">LeVine and Weinstein, 2014</xref>) to quantify the allosteric interactions that connect the changes in CHL binding modes to the function-related conformational changes. The results identified the paths of allosteric connectivity and showed that the gating element H4 is indeed a strong allosteric coordinator of the cholesterol binding mode dynamics. The coordination is achieved dynamically as cholesterol moves to a binding site lower in the pocket, by (1) the straightening of the H4 helix that makes space for the movement of the CHL, and (2) H4 dissociating from the interaction with the &#x3a9;1 loop which opens the gate further. These coordinated changes trigger yet more conformational rearrangement as shown in macrostates 4 and 6 (<xref ref-type="sec" rid="s9">Supplementary Figures S6B, 8D</xref>). Notably, this coordinated set of CHL translocation in the binding pocket and rearrangements of structural motifs result in a conformation similar to the structures of cholesterol-binding LAM protein complexes. There, the bound sterol is located close to the tunnel entrance rather than residing at the bottom of the hydrophobic cavity, and as the tunnel entrance opens slightly at the &#x3a9;1-loop, the ligand sterol is exposed to water (<xref ref-type="bibr" rid="B13">Horenkamp et al., 2018</xref>; <xref ref-type="bibr" rid="B18">Jentsch et al., 2018</xref>; <xref ref-type="bibr" rid="B51">Tong et al., 2018</xref>). This similarity is illustrated in the structural superposition of StarD4 and LAM2 (<xref ref-type="sec" rid="s9">Supplementary Figures S11C, D</xref>) where the location of the LAM ligand close to the entrance is shown to correspond to that of the CHL in the Trp-binding mode of StarD4, adjacent to the water network between W171, R92 and Y117 (<xref ref-type="bibr" rid="B55">Zhu and Weng, 2005</xref>).</p>
<p>Notably, the allosteric network revealed by the NbIT analysis in the StarD4-cholesterol complex shows that the allosteric coordination between the cholesterol binding and the peripheral motifs includes the &#x3b2;1-3 and &#x3b2;9 sheets. The &#x3b2;1 and &#x3b2;2 sheets are known to be essential for StarD4 membrane-binding that is mediated by their basic residues patch which interacts with anionic lipids (<xref ref-type="bibr" rid="B14">Iaea et al., 2015</xref>; <xref ref-type="bibr" rid="B54">Zhang et al., 2022</xref>). Thus, we have shown recently (<xref ref-type="bibr" rid="B54">Zhang et al., 2022</xref>) that upon membrane-embedding of StarD4 the &#x3b2;1 and &#x3b2;2 sheets bind PIP2 in a subtype-specific manner, which is likely to contribute to the recognition of target organelles based on the prevalent PIP2 subtype composition of their membranes. Also, membrane-embedded StarD4 adopts an orientation where the &#x3b2;9 and &#x3b2;7&#x3b2;8-loop motifs that are part of the allosteric network (<xref ref-type="fig" rid="F6">Figure 6</xref>), are interacting with lipid heads in the membrane (<xref ref-type="bibr" rid="B54">Zhang et al., 2022</xref>).</p>
<p>Together, the results from the RED and NbIT analyses of the extensive set of long MD trajectories of cholesterol-bound complexes of WT StarD4 and the three mutants have revealed a well-defined allosteric mechanism in CHL-bound StarD4. The conformational rearrangements determined with the RED algorithm to occur simultaneously with the relocation of the CHL from the Ser-binding mode to the Trp-binding mode in the hydrophobic pocket are proposed to constitute a preparation for CHL release into a membrane environment. The mechanism of this preparation is defined in specific detail by the NbIT-quantified allosteric network of interactions between the CHL binding site and the peripheral structural motifs involved in cholesterol translocation. As &#x3b2;1 was reported to be an essential membrane-binding motif (<xref ref-type="bibr" rid="B14">Iaea et al., 2015</xref>; <xref ref-type="bibr" rid="B54">Zhang et al., 2022</xref>), we propose that when the bound StarD4 interacts with the membrane, the movement of CHL between binding modes propagate trough the allosteric pathway to change the conformations of the peripheral motifs of StarD4 to enable delivery. We are currently evaluating this hypothesis in membrane-embedded StarD4 systems described in <xref ref-type="bibr" rid="B54">Zhang et al. (2022)</xref>.</p>
</sec>
<sec sec-type="materials|methods" id="s4">
<title>Materials and methods</title>
<sec id="s4-1">
<title>Modeling of the cholesterol StarD4 complex</title>
<p>The structure of mouse apo-StarD4 was taken from the X-ray crystal structure (PDB ID: 1jss), which includes residues 24 to 222 of the protein (<xref ref-type="bibr" rid="B42">Romanowski et al., 2002</xref>). Lys223 and Ala224 were introduced using Modeller 9.23 software (<xref ref-type="bibr" rid="B43">Sali, 1995</xref>), resulting in the conformation of mouse StarD4 22&#x2013;224 segment with N-terminus acetylation. StarD4 complexes with cholesterol (CHL) were obtained by docking the ligand in the hydrophobic pocket of StarD4 using Schrodinger&#x2019;s Induced Fit Docking protocol (<xref ref-type="bibr" rid="B45">Sherman et al., 2006a</xref>; <xref ref-type="bibr" rid="B44">Sherman et al., 2006b</xref>). The common features shared by the top 15 docking results are: the interaction of the cholesterol hydroxyl head that forms 2 or 3 hydrogen bonds with the sidechains of Ser136 &#x26; Ser147 (<xref ref-type="fig" rid="F1">Figure 1A</xref>) and with the backbone carbonyl of Cys148, and the sequestering of the cholesterol hydrocarbon tail inside the hydrophobic pocket, away from the water environment. This top docking mode was used as the initial structure for the MD simulations of the StarD4-CHL complex solvated in 0.15M&#xa0;K<sup>&#x2b;</sup>Cl<sup>&#x2212;</sup> ionic aqueous solution (&#x223c;32,600 atoms).</p>
</sec>
<sec id="s4-2">
<title>Long unbiased simulation of cholesterol-bound StarD4 in water</title>
<p>The cholesterol-bound StarD4 system was equilibrated first in MD simulations with NAMD version 2.12 (<xref ref-type="bibr" rid="B35">Phillips et al., 2005</xref>) using a multi-stage protocol. In the first stage the backbone of StarD4 and the heavy atoms of the ligand were harmonically constrained and gradually released in three steps of 0.5ns each, with restraining force constants changing from 1 kcal/(mol&#xb7;&#xc5;2) to 0.5, and 0.1, respectively. This stage was followed by 6ns unbiased MD simulation using NAMD. In the third (production) stage, the velocities of all the atoms were reset, and the system was simulated with openMM software (<xref ref-type="bibr" rid="B8">Eastman et al., 2017</xref>) in 12 independent replicates, each for 33.6&#x3bc;s, resulting in a cumulative simulation time of 403&#x3bc;s (&#x3e;0.4 milliseconds). The OpenMM simulations were conducted in NPT ensemble (T &#x3d; 310K, <italic>p</italic> &#x3d; 1&#xa0;atm) using a 4fs integration time-step and a Monte Carlo Membrane Barostat.</p>
</sec>
<sec id="s4-3">
<title>Mutant constructs of the StarD4-CHL complex</title>
<p>Mutations were introduced into the StarD4-CHL complex with the free-energy perturbation protocol in NAMD version 2.12 (<xref ref-type="bibr" rid="B35">Phillips et al., 2005</xref>). Accordingly, hybrid StarD4 models containing overlapped residues (e.g., overlapped K49 and A49 to introduce the K49A mutation) were constructed using representative conformations obtained in the most popular Ser-binding state 1 and Trp-binding state 3. Then, the mutations are introduced by gradual annihilation of the interactions between the WT residue and its surrounding, achieved by decreasing the coupling parameters from 1 to 0 in 10 steps, and concomitantly increasing the interactions between the mutant residue and its surrounding from 0 to 1 in 10 steps. The simulation was run for 5ns per window. After the last window, the annihilated original residues are removed from the hybrid StarD4 models, and the mutant StarD4 is equilibrated for another 10ns. For production runs, 24 replicas are made for each system and sampled in long unbiased simulation with openMM (<xref ref-type="bibr" rid="B8">Eastman et al., 2017</xref>).</p>
</sec>
<sec id="s4-4">
<title>Rare event detection (RED) protocol</title>
<p>Rare events in MD trajectory data were detected as described with our RED protocol (<xref ref-type="bibr" rid="B36">Plante and Weinstein, 2021</xref>) which utilizes a Non-negative Matrix Factorization (NMF) algorithm implemented in the Scikit-Learn python package (<xref ref-type="bibr" rid="B10">Fabian Pedregosa et al., 2011</xref>). The unsupervised machine learning algorithm learns to decompose a trajectory into a set of components that represent structural motifs that move simultaneously, and reports their appearance as dominant structural characteristics during a particular period in the course of the trajectory. NMF is a machine learning technique used to decompose high-dimensional non-negative data. It has been widely applied in computational biology research, where it has proven to be a powerful tool for tasks such as the molecular pattern discovery, class comparison and prediction, and functional characterization of genes as described in <xref ref-type="bibr" rid="B36">Plante and Weinstein (2021)</xref> and <xref ref-type="bibr" rid="B20">Lee and Seung (1999)</xref>. Briefly summarized, the event detection method in the RED protocol uses the NMF algorithm to learn a sparse, parts-based representation of the data (<xref ref-type="bibr" rid="B20">Lee and Seung, 1999</xref>). For trajectory analyses the input array (I) is a contact matrix constructed as described below from the data in each frame of the trajectory. By construction, this array encodes the information about the protein structure as it evolves over the trajectory time. The NMF algorithm decomposes the I array into a &#x201c;spatial&#x201d; array and a &#x201c;temporal&#x201d; array, which together are responsible for the integrated detection of the temporally defined rare events that provide the mechanistic information. Thus, the columns of the &#x201c;spatial&#x201d; output matrix produced by the NMF analysis are termed &#x201c;components&#x201d;, as each of them represents a set of conformational features that change simultaneously even if they are in different structural motifs. As the determinant features of a conformational state evolve over trajectory time, the different components containing them dominate the structural characteristics of the molecule at different times. Thus, the corresponding temporal array represents the contribution of each component along the trajectory, and thus identifies the time period where a particular component dominates the structural characteristics of the protein (<xref ref-type="bibr" rid="B36">Plante and Weinstein, 2021</xref>). A simple illustration provided in the full description of the method (<xref ref-type="bibr" rid="B36">Plante and Weinstein, 2021</xref>) refers to the analysis of a conformational transition in the unfolding of a fully folded alpha helical segment. The analysis results in one component that is the folded structure, and a second one that is the unfolded structure. Each frame of the simulation trajectory will be a mixing of these archetypes, with the folded component dominating the trajectory frames at the time segments before the transition, and the unfolded component dominating those after the transition.</p>
<p>
<italic>Building the time-evolution contact map for the input array (</italic>
<bold>
<italic>I</italic>
</bold>
<italic>)</italic>: A contact map is constructed between every pair of residues for each frame in the trajectory, with contact marked as 1 if any atom of one residue is &#x3c;3.5&#xc5; away from any atom in another residue. For the 201 residues of StarD4 this yields a tensor of 0 and 1 values with dimensions (t, 201,201), where t represents the number of frames. Subsequently, the tensor is rearranged to a matrix of size 201<sup>2</sup> columns and t rows, where 201<sup>2</sup> is the total number of residue pairs. The matrix columns, representing all the residue pairs, are further trimmed by excluding the contact pairs that remain invariant throughout the entire simulation, which yields a <bold>
<italic>I</italic>
</bold> matrix with n_pairs &#x3d; 3,880 columns. Each column represents a contact pair that has altered its contact state at least once during the entire simulation, encapsulating protein dynamics information for further analysis.</p>
<p>To focus on the conformational changes around the time of cholesterol relocation between its binding modes in the hydrophobic pocket, we excerpted 6 trajectory segments from the 6 independent runs as indicated by the time intervals (<bold>traj0</bold>, from 6.5 to 14.5&#x3bc;s; <bold>traj1</bold>, 0.1&#x2013;4.1&#x3bc;s<bold>; traj3</bold>, 0.1&#x2013;8.1&#x3bc;s; <bold>traj4</bold>, 0.1&#x2013;6.5&#x3bc;s; <bold>traj6</bold>, 0.1&#x2013;4.1&#x3bc;s; <bold>traj7</bold>, 0.1&#x2013;4.1&#x3bc;s; 34.4&#x3bc;s total). In order to obtain results that are generalizable among trajectories (which decreases the sensitivity to incidental structural fluctuations and benefits the reproducibility of RED analysis), the contact matrices of the 6 segments are inputted at the same time as the training data for the NMF algorithm. This input of multiple trajectories is constructed by concatenating the data matrices along the time-axis. This contact map is recorded every 0.8ns, which yields an <bold>
<italic>I</italic>
</bold> matrix with t &#x3d; 43,000 rows. Smoothing is applied to the contact matrix across the time dimension by averaging the array in a sliding window of 20 nanoseconds (25 frames).</p>
<p>
<italic>Details of the protocol</italic>: The algorithm decomposes the input contact matrix described above into multiple components, where each component has a structural state described by a &#x201c;spatial&#x201d; row array with length n_pairs (here n_pairs &#x3d; 3,880), and a corresponding &#x201c;temporal&#x201d; column array of length t (here t &#x3d; 43,000). The number of components (c) is a super-parameter chosen by the user according to the number of independent processes in the system and was set to be c &#x3d; 5 to study generalizable features shared in independent runs. Thus, the ensemble of the &#x201c;temporal&#x201d; column arrays is a temporal matrix of size (t, c) in this study, and the ensemble of the &#x201c;spatial&#x201d; row arrays is a spatial matrix of size (c, n_pairs &#x3d; 3,880), where the multiplication of the temporal matrix and the spatial matrix is fitted to resemble the time-evolution contact map data matrix. Each row in the c &#x3d; 5 rows of components in the spatial matrix is an array that encodes the determinant conformational features that evolve concurrently over time. The corresponding column in the temporal matrix represents the contribution of these features in determining the conformational state of the protein at each time. It and identifies the trajectory times at which this particular set of features dominates the structural characteristics of the protein by showing it to have the highest &#x201c;weight&#x201d; in this period of trajectory time.</p>
<p>
<italic>Interpretation of the results</italic>: As shown in <xref ref-type="fig" rid="F2">Figures 2E&#x2013;G</xref>, the spatial arrays scatter plots show that there are groups of pairs that form a horizonal line, suggesting that these pairs are constantly contributing similarly to every component. These groups of pairs that are salient in every component by being in contact throughout the StarD4 simulation, and contribute similarly to each component, are used for normalization of the spatial arrays. Thus, each spatial array is divided by a normalization factor that equals to the mean value of the constructive residue pairs contributing to this particular component, thereby placing the contact features of these constitutive residue pairs at 1 in <xref ref-type="fig" rid="F2">Figures 2E&#x2013;G</xref>; <xref ref-type="sec" rid="s9">Supplementary Figure S2B</xref> To preserve the relative weight between components, each corresponding temporal array is multiplied by the same normalization factor derived from the spatial array of this particular component, so that the multiplication result of spatial and temporal matrix remains the same and resembles the input contact matrix.</p>
<p>This normalization allows for direct comparison between components. Subtracting the spatial array of comp A from comp B, as represented in (<xref ref-type="sec" rid="s9">Supplementary Figure S3</xref>), cancels out the contributions of the constitutive residue pairs so that pairs with positive contributions are those newly established in the event corresponding to the A&#x2192;B transition, whereas the negatively contributing pairs are those breaking contact during this A&#x2192;B event (<xref ref-type="sec" rid="s9">Supplementary Figure S3</xref>). These pairs are salient in the transition as they contribute to the structural differences between components and are termed <italic>structure differentiating contact pairs</italic> (SDCPs). The SDCPs that are more salient than 4 times the standard deviation, are summarized in <xref ref-type="sec" rid="s9">Supplementary Table S1</xref> and are grouped based on the motif movement they represent.</p>
</sec>
<sec id="s4-5">
<title>Dimensionality reduction with the &#x201c;time-structure based independent component analysis&#x201d; (tICA) approach</title>
<p>The MD trajectory frames were projected onto a space of 7 collective variables described in <xref ref-type="sec" rid="s9">Supplementary Table S2</xref>, with the &#x201c;time-structure based independent component analysis&#x201d; (tICA) method. The tICA method identifies the slowest reaction coordinates of a system in the dynamic dataset (<xref ref-type="bibr" rid="B29">Molgedey and Schuster, 1994</xref>; <xref ref-type="bibr" rid="B31">Naritomi and Fuchigami, 2011</xref>; <xref ref-type="bibr" rid="B34">Perez-Hernandez et al., 2013</xref>) by solving for the generalized eigenvectors and eigenvalues of the time-lagged covariance matrix:<disp-formula id="equ1">
<mml:math id="m1">
<mml:mrow>
<mml:msup>
<mml:mi>C</mml:mi>
<mml:mi>x</mml:mi>
</mml:msup>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>&#x3c4;</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:msub>
<mml:mi mathvariant="normal">V</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
<mml:mo>&#x3d;</mml:mo>
<mml:msub>
<mml:mi mathvariant="normal">&#x3bb;</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
<mml:msup>
<mml:mi>C</mml:mi>
<mml:mi>x</mml:mi>
</mml:msup>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mn>0</mml:mn>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:msub>
<mml:mi mathvariant="normal">V</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:math>
</disp-formula>in which the <inline-formula id="inf1">
<mml:math id="m2">
<mml:mrow>
<mml:msup>
<mml:mi>C</mml:mi>
<mml:mi>x</mml:mi>
</mml:msup>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>&#x3c4;</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula> is the time-lagged covariance matrix <inline-formula id="inf2">
<mml:math id="m3">
<mml:mrow>
<mml:msup>
<mml:mi>C</mml:mi>
<mml:mi>x</mml:mi>
</mml:msup>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>&#x3c4;</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>&#x3d;</mml:mo>
<mml:msub>
<mml:mrow>
<mml:mfenced open="&#x2329;" close="&#x232a;" separators="|">
<mml:mrow>
<mml:mi>X</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>t</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:msup>
<mml:mi>X</mml:mi>
<mml:mi>T</mml:mi>
</mml:msup>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>t</mml:mi>
<mml:mo>&#x2b;</mml:mo>
<mml:mi>&#x3c4;</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mi>t</mml:mi>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula> between collective variables, <inline-formula id="inf3">
<mml:math id="m4">
<mml:mrow>
<mml:msup>
<mml:mi>C</mml:mi>
<mml:mi>x</mml:mi>
</mml:msup>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mn>0</mml:mn>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>&#x3d;</mml:mo>
<mml:msub>
<mml:mrow>
<mml:mfenced open="&#x2329;" close="&#x232a;" separators="|">
<mml:mrow>
<mml:mi>X</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>t</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:msup>
<mml:mi>X</mml:mi>
<mml:mi>T</mml:mi>
</mml:msup>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>t</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mi>t</mml:mi>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula> is the regular covariance matrix, <inline-formula id="inf4">
<mml:math id="m5">
<mml:mrow>
<mml:mi>X</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>t</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula> is the data vector at time <inline-formula id="inf5">
<mml:math id="m6">
<mml:mrow>
<mml:mi>t</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula>, <inline-formula id="inf6">
<mml:math id="m7">
<mml:mrow>
<mml:mi>&#x3c4;</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> is the lag-time, <inline-formula id="inf7">
<mml:math id="m8">
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="normal">V</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula> is the i<sup>th</sup> tICA eigenvector, and <inline-formula id="inf8">
<mml:math id="m9">
<mml:mrow>
<mml:msub>
<mml:mi>&#x3bb;</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula> is the i<sup>th</sup> tICA eigenvalue. In this study, the stride of the trajectories is 0.4 ns/frame, and the lag-time <inline-formula id="inf9">
<mml:math id="m10">
<mml:mrow>
<mml:mi>&#x3c4;</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> for the covariance matrix construction is 10ns (25 frames), thus ignoring any fluctuations faster than 10ns which are not relevant to the slow processes in the system. The first 1&#x3bc;s in every independent run are discarded to equilibrate the systems.</p>
</sec>
<sec id="s4-6">
<title>Construction of the Markov State Model (MSM) and assignment of macrostates</title>
<p>
<list list-type="simple">
<list-item>
<p>1. Construction of the Markov State Model</p>
</list-item>
</list>
</p>
<p>Markov State Model (MSM) is a powerful tool to study the kinetics of equilibrium state-transition processes in protein function or activation pathways (<xref ref-type="bibr" rid="B3">Beauchamp et al., 2012</xref>; <xref ref-type="bibr" rid="B19">Kohlhoff et al., 2014</xref>; <xref ref-type="bibr" rid="B46">Shukla et al., 2014</xref>). The implementation and validation of MSMs are well documented (<xref ref-type="bibr" rid="B32">Noe and andFischer, 2008</xref>; <xref ref-type="bibr" rid="B33">Pande et al., 2010</xref>; <xref ref-type="bibr" rid="B37">Prinz et al., 2011</xref>). To construct MSMs, the 3D conformational space of the first three tIC vectors was discretized into microstates. The transition count matrixes between the microstates were then recorded in a transition count matrix, which is then symmetrized by averaging with its transpose matrix in order to satisfy detailed balance and local equilibrium (<xref ref-type="bibr" rid="B2">Beauchamp et al., 2011</xref>; <xref ref-type="bibr" rid="B37">Prinz et al., 2011</xref>). Finally, the transition probability matrixes (TPMs) were built by normalizing the transition count matrix by the transitions departing from the same microstate.</p>
<p>
<list list-type="simple">
<list-item>
<p>2. Selection of parameters for markov model construction</p>
</list-item>
</list>
</p>
<p>The choice of super parameters: the lag-time of the transition, and the discretization of the conformational space (i.e., number of microstates), are optimized to ensure the quality of MSM (<xref ref-type="bibr" rid="B2">Beauchamp et al., 2011</xref>; <xref ref-type="bibr" rid="B37">Prinz et al., 2011</xref>). The lag-time is optimized with implied timescale test:<disp-formula id="equ2">
<mml:math id="m11">
<mml:mrow>
<mml:msub>
<mml:mi>&#x3c4;</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
<mml:mo>&#x3d;</mml:mo>
<mml:mfrac>
<mml:msup>
<mml:mi>&#x3c4;</mml:mi>
<mml:mo>&#x2032;</mml:mo>
</mml:msup>
<mml:mrow>
<mml:mi>ln</mml:mi>
<mml:mo>&#x2061;</mml:mo>
<mml:mo>&#x2061;</mml:mo>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>&#x3bb;</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:mfrac>
</mml:mrow>
</mml:math>
</disp-formula>in which <inline-formula id="inf10">
<mml:math id="m12">
<mml:mrow>
<mml:msup>
<mml:mi>&#x3c4;</mml:mi>
<mml:mo>&#x2032;</mml:mo>
</mml:msup>
</mml:mrow>
</mml:math>
</inline-formula> is the lag-time used for building the transition matrix, <inline-formula id="inf11">
<mml:math id="m13">
<mml:mrow>
<mml:msub>
<mml:mi>&#x3bb;</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula> is the i<sup>th</sup> eigenvector of the TPM, and <inline-formula id="inf12">
<mml:math id="m14">
<mml:mrow>
<mml:msub>
<mml:mi>&#x3c4;</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula> is the implied timescale corresponding to the i<sup>th</sup> relaxation mode of the system. In this study, the lag-time of 120 ns defined as shown in <xref ref-type="sec" rid="s9">Supplementary Figure S5A</xref> ensures the best available Markovian behavior with the minimal loss of data. The best discretization of the conformational space is chosen by comparing the resulted MSM with the generalized matrix Rayleigh quotient (GMRQ) method (<xref ref-type="bibr" rid="B2">Beauchamp et al., 2011</xref>; <xref ref-type="bibr" rid="B27">McGibbon and andPande, 2015</xref>), where the MSM is trained on a randomly picked training set comprised of half the set of trajectories, and then cross-validated on the test set which includes the other half of the trajectory set. The method tests how well the eigenvectors of the training set diagonalize the TPM of the test set, scored as:<disp-formula id="equ3">
<mml:math id="m15">
<mml:mrow>
<mml:mi>R</mml:mi>
<mml:mo>&#x3d;</mml:mo>
<mml:mi>T</mml:mi>
<mml:mi>r</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msup>
<mml:mi>V</mml:mi>
<mml:mi>T</mml:mi>
</mml:msup>
<mml:mi>S</mml:mi>
<mml:mi>T</mml:mi>
<mml:mi>V</mml:mi>
<mml:msup>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msup>
<mml:mi>V</mml:mi>
<mml:mi>T</mml:mi>
</mml:msup>
<mml:mi>S</mml:mi>
<mml:mi>V</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mrow>
<mml:mo>&#x2212;</mml:mo>
<mml:mn>1</mml:mn>
</mml:mrow>
</mml:msup>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:math>
</disp-formula>in which <inline-formula id="inf13">
<mml:math id="m16">
<mml:mrow>
<mml:mi>T</mml:mi>
<mml:mi>r</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> denotes trace of the inner matrix product, <inline-formula id="inf14">
<mml:math id="m17">
<mml:mrow>
<mml:mi>V</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> contains the first n eigenvectors of the training set, <inline-formula id="inf15">
<mml:math id="m18">
<mml:mrow>
<mml:mi>S</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> is the diagonal matrix composed of equilibrium populations of the test set (i.e., a diagonal matrix composed by the first eigenvector of the test set TPM), and <inline-formula id="inf16">
<mml:math id="m19">
<mml:mrow>
<mml:mi>T</mml:mi>
</mml:mrow>
</mml:math>
</inline-formula> is the test set TPM. The best MSM parameters yield the highest GMRQ score. In this study, the 200 k-means microstates were chosen so as to provide a good balance between the consistency shown by GMRQ result (<xref ref-type="sec" rid="s9">Supplementary Figure S5B</xref>) and the coverage of the 3D conformational space.</p>
<p>
<list list-type="simple">
<list-item>
<p>3. Interpretation of MSM, and macrostate assignment</p>
</list-item>
</list>
</p>
<p>To extract information from the MSM, TPM is decomposed into its eigenvalues and eigenvectors. The largest eigenvalue is 1, and its corresponding first eigenvector represents the equilibrium state population of the system. The remaining eigenvectors represent the population flow between microstates. The second largest eigenvalue corresponds to the eigenvector representing the population flow contributing to the slowest relaxation process of the system. In turn, the next largest eigenvalues and eigenvector correspond to the next slowest relaxation process, and so on. Here, the 200 microstates were grouped into 9 macrostates according to their kinetics similarities learnt from MSM, using the using the Robust Perron Cluster Analysis (PCCA&#x2b;) algorithm (<xref ref-type="bibr" rid="B6">Deuflhard and Weber, 2005</xref>).</p>
</sec>
<sec id="s4-7">
<title>Transition path theory (TPT) analysis</title>
<p>The transition paths among the 9 macrostates were obtained from the TPT analysis (<xref ref-type="bibr" rid="B4">Berezhkovskii et al., 2009</xref>), which is based on the flux matrix between states whose components are defined as:<disp-formula id="equ4">
<mml:math id="m20">
<mml:mrow>
<mml:msub>
<mml:mi>J</mml:mi>
<mml:mrow>
<mml:mi>i</mml:mi>
<mml:mi>j</mml:mi>
</mml:mrow>
</mml:msub>
<mml:mo>&#x3d;</mml:mo>
<mml:msub>
<mml:mi>&#x3c0;</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mn>1</mml:mn>
<mml:mo>&#x2212;</mml:mo>
<mml:msub>
<mml:mi>P</mml:mi>
<mml:mrow>
<mml:mi>t</mml:mi>
<mml:mi>a</mml:mi>
<mml:mi>r</mml:mi>
<mml:mi>g</mml:mi>
<mml:mi>e</mml:mi>
<mml:mi>t</mml:mi>
<mml:mo>,</mml:mo>
<mml:mi>i</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:msub>
<mml:mi>T</mml:mi>
<mml:mrow>
<mml:mi>i</mml:mi>
<mml:mi>j</mml:mi>
</mml:mrow>
</mml:msub>
<mml:msub>
<mml:mi>P</mml:mi>
<mml:mrow>
<mml:mi>t</mml:mi>
<mml:mi>a</mml:mi>
<mml:mi>r</mml:mi>
<mml:mi>g</mml:mi>
<mml:mi>e</mml:mi>
<mml:mi>t</mml:mi>
<mml:mo>,</mml:mo>
<mml:mi>j</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mrow>
</mml:math>
</disp-formula>where <inline-formula id="inf17">
<mml:math id="m21">
<mml:mrow>
<mml:msub>
<mml:mi>&#x3c0;</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula> is the equilibrium population of state I, <inline-formula id="inf18">
<mml:math id="m22">
<mml:mrow>
<mml:msub>
<mml:mi>T</mml:mi>
<mml:mrow>
<mml:mi>i</mml:mi>
<mml:mi>j</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula> is the transition probability from state i to state j, and <inline-formula id="inf19">
<mml:math id="m23">
<mml:mrow>
<mml:msub>
<mml:mi>P</mml:mi>
<mml:mrow>
<mml:mi>t</mml:mi>
<mml:mi>a</mml:mi>
<mml:mi>r</mml:mi>
<mml:mi>g</mml:mi>
<mml:mi>e</mml:mi>
<mml:mi>t</mml:mi>
<mml:mo>,</mml:mo>
<mml:mi>i</mml:mi>
</mml:mrow>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula> is the probability of visiting the target state i as the first state during the transition before visiting other states. The most probable transition pathways between specific states was obtained by iteratively finding the next pathway with the highest flux in the flux matrix using a graph theory algorithm, the Dijkstra algorithm (<xref ref-type="bibr" rid="B7">Dijkstra, 1959</xref>), implemented in the MSM builder software package (<xref ref-type="bibr" rid="B2">Beauchamp et al., 2011</xref>).</p>
</sec>
<sec id="s4-8">
<title>Quantifying allosteric effect using N-body information theory (NbIT)</title>
<p>
<list list-type="simple">
<list-item>
<p>1. NbIT analysis focused on multiple motifs in the StarD4-CHL complex</p>
</list-item>
</list>
</p>
<p>The information theory-based NbIT analysis was used to measure the coordination between distanced motifs and to quantify the allosteric effect (<xref ref-type="bibr" rid="B23">LeVine et al., 2014</xref>; <xref ref-type="bibr" rid="B24">LeVine and Weinstein, 2014</xref>). For the calculation of the coordination information we chose the following structural motifs: <bold>CHLsite:</bold> S136 S147 W171 R92 Y117; <bold>&#x3b2;1</bold>: res R46 V47 A48 K49 K50 V51 K52; <bold>&#x3b2;2&#x3b2;3</bold>: res R58 K59 P60 Y67 L68 Y69; <bold>&#x3b2;2&#x3b2;3loop</bold>: E62 E63 F64 N65; <bold>H4head</bold>: Q199 S200 A201 D203 T204 A207; <bold>&#x3a9;1loop</bold>: L124 N125 I126; &#x3b2;9: D192 R194 G195; <bold>&#x3b2;7&#x3b2;8loop-near&#x3b2;9</bold>: V162 R163 G164; <bold>&#x3b2;7&#x3b2;8loop-mid</bold>: T157 R158 P159 E160; <bold>&#x3b2;7&#x3b2;8loop-near&#x3b2;6</bold>: E153 W154 S155 E156; <bold>H4tail</bold>: R218 K219&#xa0;G220 L221; <bold>&#x3b2;3tail</bold>: G73 V74 M75 D76; <bold>&#x3b2;8&#x3b2;9loop</bold>: S179 P180 S181 Q182. The allosteric effect was then quantified from the coordination information as described in ref 32.</p>
<p>Briefly, the first step in the calculation of the coordination information was the alignment of the trajectories to a reference structure&#x2013;we used the StarD4 crystal structure (PDB ID: 1JSS). The alignment of snapshot (per 0.4ns) was carried out on the alpha carbons of residues in &#x3b2;3-9 and H4, excluding the motifs have been found to undergo conformational changes (70&#x2013;76, 79&#x2013;86, 103&#x2013;108, 113&#x2013;118, 131&#x2013;140, 146&#x2013;150, 170&#x2013;175, 184&#x2013;189, 211&#x2013;220). With the aligned trajectory, the entropy in each motif was calculated analytically through the differential entropy:<disp-formula id="equ5">
<mml:math id="m24">
<mml:mrow>
<mml:mi>H</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>X</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>&#x3d;</mml:mo>
<mml:mfrac>
<mml:mrow>
<mml:mn>1</mml:mn>
</mml:mrow>
<mml:mrow>
<mml:mn>2</mml:mn>
</mml:mrow>
</mml:mfrac>
<mml:mi>ln</mml:mi>
<mml:mo>&#x2061;</mml:mo>
<mml:mo>&#x2061;</mml:mo>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mfenced open="|" close="|" separators="|">
<mml:mrow>
<mml:mn>2</mml:mn>
<mml:mi>&#x3c0;</mml:mi>
<mml:mi>e</mml:mi>
<mml:mi>C</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>X</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:math>
</disp-formula>in which <inline-formula id="inf20">
<mml:math id="m25">
<mml:mrow>
<mml:mi>C</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>X</mml:mi>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula> is the covariance matrix describing all variables in X, and X is the displacement array describing the position of every atom in the motif by its distance to the ref conformation. Then, the <italic>Total Correlation</italic> (TC) describing the total amount of information share in a set of motifs was calculated as:<disp-formula id="equ6">
<mml:math id="m26">
<mml:mrow>
<mml:mi>T</mml:mi>
<mml:mi>C</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mn>1</mml:mn>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:mo>&#x2026;</mml:mo>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>N</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>&#x3d;</mml:mo>
<mml:mrow>
<mml:mstyle displaystyle="true">
<mml:munderover>
<mml:mo>&#x2211;</mml:mo>
<mml:mi>i</mml:mi>
<mml:mi>N</mml:mi>
</mml:munderover>
</mml:mstyle>
<mml:mrow>
<mml:mi>H</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:mrow>
<mml:mo>&#x2212;</mml:mo>
<mml:mi>H</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mn>1</mml:mn>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:mo>&#x2026;</mml:mo>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>N</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>.</mml:mo>
</mml:mrow>
</mml:math>
</disp-formula>
</p>
<p>And the <italic>Coordination Information</italic> (CI) describing the amount of information shared in the receptors that is also shared with another motif that works as the <italic>transmitter</italic>:<disp-formula id="equ7">
<mml:math id="m27">
<mml:mrow>
<mml:mi>C</mml:mi>
<mml:mi>I</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mrow>
<mml:mfenced open="{" close="}" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mn>1</mml:mn>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:mo>&#x2026;</mml:mo>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>N</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>m</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>&#x3d;</mml:mo>
<mml:mi>T</mml:mi>
<mml:mi>C</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mn>1</mml:mn>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:mo>&#x2026;</mml:mo>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>N</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>&#x2212;</mml:mo>
<mml:mi>T</mml:mi>
<mml:mi>C</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mn>1</mml:mn>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:mo>&#x2026;</mml:mo>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>N</mml:mi>
</mml:msub>
<mml:mtext>&#x2009;</mml:mtext>
<mml:mo>&#x7c;</mml:mo>
<mml:mtext>&#x2009;</mml:mtext>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>m</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:math>
</disp-formula>
</p>
<p>Here <inline-formula id="inf21">
<mml:math id="m28">
<mml:mrow>
<mml:mi>T</mml:mi>
<mml:mi>C</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mn>1</mml:mn>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:mo>&#x2026;</mml:mo>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>N</mml:mi>
</mml:msub>
<mml:mo>&#x7c;</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>m</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula> is the conditional total correlation between <inline-formula id="inf22">
<mml:math id="m29">
<mml:mrow>
<mml:mfenced open="{" close="}" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mn>1</mml:mn>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:mo>&#x2026;</mml:mo>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>N</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:math>
</inline-formula> conditioning on <inline-formula id="inf23">
<mml:math id="m30">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>m</mml:mi>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula>:<disp-formula id="equ8">
<mml:math id="m31">
<mml:mrow>
<mml:mi>T</mml:mi>
<mml:mi>C</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mn>1</mml:mn>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:mo>&#x2026;</mml:mo>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>N</mml:mi>
</mml:msub>
<mml:mtext>&#x2009;</mml:mtext>
<mml:mo>&#x7c;</mml:mo>
<mml:mtext>&#x2009;</mml:mtext>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>m</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>&#x3d;</mml:mo>
<mml:mrow>
<mml:mstyle displaystyle="true">
<mml:munderover>
<mml:mo>&#x2211;</mml:mo>
<mml:mi>i</mml:mi>
<mml:mi>N</mml:mi>
</mml:munderover>
</mml:mstyle>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>H</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>m</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>&#x2212;</mml:mo>
<mml:mi>H</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>m</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
<mml:mo>&#x2212;</mml:mo>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mi>H</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mn>1</mml:mn>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:mo>&#x2026;</mml:mo>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>N</mml:mi>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>m</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>&#x2212;</mml:mo>
<mml:mi>H</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>m</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>.</mml:mo>
</mml:mrow>
</mml:math>
</disp-formula>
</p>
<p>The <italic>Coordination Information</italic> was then normalized by the <italic>Total Correlation</italic> of the <italic>receiver</italic>, to represent the portion of dynamics of the receptor that is allosterically coupled with the <italic>transmitter</italic>.<disp-formula id="equ9">
<mml:math id="m32">
<mml:mrow>
<mml:mover accent="true">
<mml:mrow>
<mml:mi>C</mml:mi>
<mml:mi>I</mml:mi>
</mml:mrow>
<mml:mo>&#xaf;</mml:mo>
</mml:mover>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mrow>
<mml:mfenced open="{" close="}" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mn>1</mml:mn>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:mo>&#x2026;</mml:mo>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>N</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>m</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>&#x3d;</mml:mo>
<mml:mfrac>
<mml:mrow>
<mml:mi>C</mml:mi>
<mml:mi>I</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mrow>
<mml:mfenced open="{" close="}" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mn>1</mml:mn>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:mo>&#x2026;</mml:mo>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>N</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>m</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
<mml:mrow>
<mml:mi>T</mml:mi>
<mml:mi>C</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mn>1</mml:mn>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:mo>&#x2026;</mml:mo>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>N</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:mfrac>
<mml:mo>&#x2217;</mml:mo>
<mml:mn>100</mml:mn>
<mml:mo>%</mml:mo>
<mml:mo>.</mml:mo>
</mml:mrow>
</mml:math>
</disp-formula>
</p>
<p>In <xref ref-type="table" rid="T2">Table 2</xref> and the <xref ref-type="sec" rid="s9">Supplementary Figures S3, S5, S6, S7</xref>, the <italic>Total Correlation</italic> (TC) within each motif is displayed along the diagonal, while the <italic>Normalized Coordination Information</italic> is presented in the off-diagonal elements, with residues on the top (columns) acting as the <italic>Transmitter</italic> and residues on the left (rows) being the coordinated <italic>Receiver</italic>.</p>
<p>3. Coordination channel analysis based on mutual coordination information</p>
<p>To define the allosteric channels that mediate the coordination information (<xref ref-type="fig" rid="F6">Figure 6A</xref>), we calculated the <italic>mutual coordination information</italic> that measures the amount of coordination information shared between the <italic>receiver</italic> and the <italic>transmitter</italic> motifs that is also shared with another structural element:<disp-formula id="equ10">
<mml:math id="m33">
<mml:mrow>
<mml:mi>M</mml:mi>
<mml:mi>C</mml:mi>
<mml:mi>I</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mrow>
<mml:mfenced open="{" close="}" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mn>1</mml:mn>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:mo>&#x2026;</mml:mo>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>N</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>m</mml:mi>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>n</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>&#x3d;</mml:mo>
<mml:mi>C</mml:mi>
<mml:mi>I</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mrow>
<mml:mfenced open="{" close="}" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mn>1</mml:mn>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:mo>&#x2026;</mml:mo>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>N</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>m</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:math>
</disp-formula>
<disp-formula id="equ11">
<mml:math id="m34">
<mml:mrow>
<mml:mo>&#x2b;</mml:mo>
<mml:mi>C</mml:mi>
<mml:mi>I</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mrow>
<mml:mfenced open="{" close="}" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mn>1</mml:mn>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:mo>&#x2026;</mml:mo>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>N</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>n</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>&#x2212;</mml:mo>
<mml:mi>C</mml:mi>
<mml:mi>I</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mrow>
<mml:mfenced open="{" close="}" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mn>1</mml:mn>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:mo>&#x2026;</mml:mo>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>N</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>m</mml:mi>
</mml:msub>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>n</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:math>
</disp-formula>
</p>
<p>Here <inline-formula id="inf24">
<mml:math id="m35">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>m</mml:mi>
</mml:msub>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>n</mml:mi>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula> is the union set constituted by the residues in the <italic>Transmitter</italic> <inline-formula id="inf25">
<mml:math id="m36">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>m</mml:mi>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula> and the <italic>channel</italic> <inline-formula id="inf26">
<mml:math id="m37">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>n</mml:mi>
</mml:msub>
</mml:mrow>
</mml:math>
</inline-formula>. The <italic>mutual coordination information</italic> was then normalized by the <italic>Coordination Information</italic> between the receiver and the <italic>transmitter</italic> to obtain the NMCI values used in the determination of the allosteric pathway in <xref ref-type="fig" rid="F6">Figure 6B</xref>
<disp-formula id="equ12">
<mml:math id="m38">
<mml:mrow>
<mml:mi>N</mml:mi>
<mml:mi>M</mml:mi>
<mml:mi>C</mml:mi>
<mml:mi>I</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mrow>
<mml:mfenced open="{" close="}" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mn>1</mml:mn>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:mo>&#x2026;</mml:mo>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>N</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>m</mml:mi>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>n</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>&#x3d;</mml:mo>
<mml:mfrac>
<mml:mrow>
<mml:mi>M</mml:mi>
<mml:mi>C</mml:mi>
<mml:mi>I</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mrow>
<mml:mfenced open="{" close="}" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mn>1</mml:mn>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:mo>&#x2026;</mml:mo>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>N</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>m</mml:mi>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>n</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
<mml:mrow>
<mml:mi>C</mml:mi>
<mml:mi>I</mml:mi>
<mml:mrow>
<mml:mfenced open="(" close=")" separators="|">
<mml:mrow>
<mml:mrow>
<mml:mfenced open="{" close="}" separators="|">
<mml:mrow>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mn>1</mml:mn>
</mml:msub>
<mml:mo>,</mml:mo>
<mml:mo>&#x2026;</mml:mo>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>N</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
<mml:mo>,</mml:mo>
<mml:msub>
<mml:mi>X</mml:mi>
<mml:mi>m</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:mrow>
</mml:mfrac>
<mml:mo>.</mml:mo>
</mml:mrow>
</mml:math>
</disp-formula>
</p>
</sec>
</sec>
</body>
<back>
<sec sec-type="data-availability" id="s5">
<title>Data availability statement</title>
<p>The datasets presented in this study can be found in online repositories. The names of the repository/repositories and accession number(s) can be found below: Computational data presented in the manuscript will be managed in full accordance with the institution&#x2019;s Data Management and Sharing policy of Weill Cornell Medical College which is in full compliance with the NIH requirements. The raw data used to reach the inferences and conclusions are available upon reasonable request from (HW). Atomistic MD simulations were carried out with OpenMM 7.5. Computational analysis was carried out using a combination of python scripts, and in-house scripts available on GitHub <ext-link ext-link-type="uri" xlink:href="https://github.com/weinsteinlab/">https://github.com/weinsteinlab/</ext-link>.</p>
</sec>
<sec id="s6">
<title>Author contributions</title>
<p>Conceptualization: HW; Data curation: HW and HX; Formal analysis: HW and HX; Investigation: HX; Methodology: HW and HX; Resources: HW; Software: HX; Supervision: HW; Visualization: HX; Writing&#x2014;Original draft preparation: HX, HW; Writing&#x2014;Review and editing: HX, HW; Writing&#x2014;Final draft review: HX and HW. All authors contributed to the article and approved the submitted version.</p>
</sec>
<ack>
<p>We gratefully acknowledge helpful discussions with Drs. Ambrose Plante, Ekaterina D. Kots, and Prof. George Khelashvili. HX gratefully acknowledges their helpful guidance regarding the principles, applicability, implementation, and interpretation of the RED, NbIT and TPT analysis. We are grateful for discussions with Prof. Frederick R. Maxfield and his lab members. Support from the 1923 Foundation for the project &#x201c;<italic>How Needed Molecular Precision is Achieved for Addressing, Pickup, and Delivery in the Trafficking of Cholesterol Among Cell Membranes</italic>&#x201d; is gratefully acknowledged. The computational resources and technical help at the Center for Computational Innovations (CCI) at the Rensselaer Polytechnic Institute (RPI), and the efficient and sustained access to the AiMOS supercomputer at CCI generously awarded through the COVID-19 High Performance Computing Consortium, are gratefully acknowledged. We are grateful for the computational resources under Projects BIP225 and BIP109 at the Oak Ridge Leadership Computing Facility, which is a DOE Office of Science User Facility supported under Contract DE-AC05-00OR22725, and for the in-house computational resources of the David A. Cofrin Center for Biomedical Information in the Institute for Computational Biomedicine at Weill Cornell Medical College.</p>
</ack>
<sec sec-type="COI-statement" id="s7">
<title>Conflict of interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec sec-type="disclaimer" id="s8">
<title>Publisher&#x2019;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
<sec id="s9">
<title>Supplementary material</title>
<p>The Supplementary Material for this article can be found online at: <ext-link ext-link-type="uri" xlink:href="https://www.frontiersin.org/articles/10.3389/fmolb.2023.1197154/full#supplementary-material">https://www.frontiersin.org/articles/10.3389/fmolb.2023.1197154/full&#x23;supplementary-material</ext-link>
</p>
<supplementary-material xlink:href="DataSheet1.PDF" id="SM1" mimetype="application/PDF" xmlns:xlink="http://www.w3.org/1999/xlink"/>
</sec>
<ref-list>
<title>References</title>
<ref id="B1">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Alpy</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Tomasetto</surname>
<given-names>C.</given-names>
</name>
</person-group> (<year>2005</year>). <article-title>Give lipids a START: The StAR-related lipid transfer (START) domain in mammals</article-title>. <source>J. Cell Sci.</source> <volume>118</volume>, <fpage>2791</fpage>&#x2013;<lpage>2801</lpage>. <pub-id pub-id-type="doi">10.1242/jcs.02485</pub-id>
</citation>
</ref>
<ref id="B2">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Beauchamp</surname>
<given-names>K. A.</given-names>
</name>
<name>
<surname>Bowman</surname>
<given-names>G. R.</given-names>
</name>
<name>
<surname>Lane</surname>
<given-names>T. J.</given-names>
</name>
<name>
<surname>Maibaum</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Haque</surname>
<given-names>I. S.</given-names>
</name>
<name>
<surname>andPande</surname>
<given-names>V. S.</given-names>
</name>
</person-group> (<year>2011</year>). <article-title>MSMBuilder2: Modeling conformational dynamics at the picosecond to millisecond scale</article-title>. <source>J. Chem. Theory Comput.</source> <volume>7</volume>, <fpage>3412</fpage>&#x2013;<lpage>3419</lpage>. <pub-id pub-id-type="doi">10.1021/ct200463m</pub-id>
</citation>
</ref>
<ref id="B3">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Beauchamp</surname>
<given-names>K. A.</given-names>
</name>
<name>
<surname>McGibbon</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Lin</surname>
<given-names>Y. S.</given-names>
</name>
<name>
<surname>Pande</surname>
<given-names>V. S.</given-names>
</name>
</person-group> (<year>2012</year>). <article-title>Simple few-state models reveal hidden complexity in protein folding</article-title>. <source>Proc. Natl. Acad. Sci. U. S. A.</source> <volume>109</volume>, <fpage>17807</fpage>&#x2013;<lpage>17813</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.1201810109</pub-id>
</citation>
</ref>
<ref id="B4">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Berezhkovskii</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Hummer</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>andSzabo</surname>
<given-names>A.</given-names>
</name>
</person-group> (<year>2009</year>). <article-title>Reactive flux and folding pathways in network models of coarse-grained protein dynamics</article-title>. <source>J. Chem. Phys.</source> <volume>130</volume>, <fpage>205102</fpage>. <pub-id pub-id-type="doi">10.1063/1.3139063</pub-id>
</citation>
</ref>
<ref id="B5">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Clark</surname>
<given-names>B. J.</given-names>
</name>
</person-group> (<year>2020</year>). <article-title>The START-domain proteins in intracellular lipid transport and beyond</article-title>. <source>Mol. Cell Endocrinol.</source> <volume>504</volume>, <fpage>110704</fpage>. <pub-id pub-id-type="doi">10.1016/j.mce.2020.110704</pub-id>
</citation>
</ref>
<ref id="B6">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Deuflhard</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Weber</surname>
<given-names>M.</given-names>
</name>
</person-group> (<year>2005</year>). <article-title>Robust Perron cluster analysis in conformation dynamics linear algebra and its applications</article-title>. <source>Linear Algebra its Appl.</source> <volume>398</volume>, <fpage>161</fpage>&#x2013;<lpage>184</lpage>. <pub-id pub-id-type="doi">10.1016/j.laa.2004.10.026</pub-id>
</citation>
</ref>
<ref id="B7">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Dijkstra</surname>
<given-names>E. W.</given-names>
</name>
</person-group> (<year>1959</year>). <article-title>A note on two problems in connexion with graphs</article-title>. <source>Numer. Math.</source> <volume>1</volume>, <fpage>269</fpage>&#x2013;<lpage>271</lpage>. <pub-id pub-id-type="doi">10.1007/bf01386390</pub-id>
</citation>
</ref>
<ref id="B8">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Eastman</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Swails</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Chodera</surname>
<given-names>J. D.</given-names>
</name>
<name>
<surname>McGibbon</surname>
<given-names>R. T.</given-names>
</name>
<name>
<surname>Zhao</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Beauchamp</surname>
<given-names>K. A.</given-names>
</name>
<etal/>
</person-group> (<year>2017</year>). <article-title>OpenMM 7: Rapid development of high performance algorithms for molecular dynamics</article-title>. <source>PLoS Comput. Biol.</source> <volume>13</volume>, <fpage>e1005659</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pcbi.1005659</pub-id>
</citation>
</ref>
<ref id="B9">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Elbadawy</surname>
<given-names>H. M.</given-names>
</name>
<name>
<surname>Borthwick</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Wright</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Martin</surname>
<given-names>P. E.</given-names>
</name>
<name>
<surname>Graham</surname>
<given-names>A.</given-names>
</name>
</person-group> (<year>2011</year>). <article-title>Cytosolic StAR-related lipid transfer domain 4 (STARD4) protein influences keratinocyte lipid phenotype and differentiation status</article-title>. <source>Br. J. Dermatol</source> <volume>164</volume>, <fpage>628</fpage>&#x2013;<lpage>632</lpage>. <pub-id pub-id-type="doi">10.1111/j.1365-2133.2010.10102.x</pub-id>
</citation>
</ref>
<ref id="B10">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Fabian Pedregosa</surname>
<given-names>G. V.</given-names>
</name>
<name>
<surname>Gramfort</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Michel</surname>
<given-names>V.</given-names>
</name>
<name>
<surname>Bertrand</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Grisel</surname>
<given-names>O.</given-names>
</name>
<name>
<surname>Blondel</surname>
<given-names>M.</given-names>
</name>
</person-group> (<year>2011</year>). <article-title>Scikit-learn</article-title>. <source>Mach. Learn. Python JMLR</source> <volume>12</volume>, <fpage>2825</fpage>&#x2013;<lpage>2830</lpage>. <pub-id pub-id-type="doi">10.5555/1953048.2078195</pub-id>
</citation>
</ref>
<ref id="B11">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Garbarino</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Pan</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Chin</surname>
<given-names>H. F.</given-names>
</name>
<name>
<surname>Lund</surname>
<given-names>F. W.</given-names>
</name>
<name>
<surname>Maxfield</surname>
<given-names>F. R.</given-names>
</name>
<name>
<surname>andBreslow</surname>
<given-names>J. L.</given-names>
</name>
</person-group> (<year>2012</year>). <article-title>STARD4 knockdown in HepG2 cells disrupts cholesterol trafficking associated with the plasma membrane, ER, and ERC</article-title>. <source>J. Lipid Res.</source> <volume>53</volume>, <fpage>2716</fpage>&#x2013;<lpage>2725</lpage>. <pub-id pub-id-type="doi">10.1194/jlr.M032227</pub-id>
</citation>
</ref>
<ref id="B12">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Hao</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Lin</surname>
<given-names>S. X.</given-names>
</name>
<name>
<surname>Karylowski</surname>
<given-names>O. J.</given-names>
</name>
<name>
<surname>Wustner</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>McGraw</surname>
<given-names>T. E.</given-names>
</name>
<name>
<surname>Maxfield</surname>
<given-names>F. R.</given-names>
</name>
</person-group> (<year>2002</year>). <article-title>Vesicular and non-vesicular sterol transport in living cells. The endocytic recycling compartment is a major sterol storage organelle</article-title>. <source>J. Biol. Chem.</source> <volume>277</volume>, <fpage>609</fpage>&#x2013;<lpage>617</lpage>. <pub-id pub-id-type="doi">10.1074/jbc.M108861200</pub-id>
</citation>
</ref>
<ref id="B13">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Horenkamp</surname>
<given-names>F. A.</given-names>
</name>
<name>
<surname>Valverde</surname>
<given-names>D. P.</given-names>
</name>
<name>
<surname>Nunnari</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Reinisch</surname>
<given-names>K. M.</given-names>
</name>
</person-group> (<year>2018</year>). <article-title>Molecular basis for sterol transport by StART-like lipid transfer domains</article-title>. <source>EMBO J.</source> <volume>37</volume>, <fpage>e98002</fpage>. <pub-id pub-id-type="doi">10.15252/embj.201798002</pub-id>
</citation>
</ref>
<ref id="B14">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Iaea</surname>
<given-names>D. B.</given-names>
</name>
<name>
<surname>Dikiy</surname>
<given-names>I.</given-names>
</name>
<name>
<surname>Kiburu</surname>
<given-names>I.</given-names>
</name>
<name>
<surname>Eliezer</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Maxfield</surname>
<given-names>F. R.</given-names>
</name>
</person-group> (<year>2015</year>). <article-title>STARD4 membrane interactions and sterol binding</article-title>. <source>STARD4 Membr. Interact. Sterol Bind. Biochem.</source> <volume>54</volume>, <fpage>4623</fpage>&#x2013;<lpage>4636</lpage>. <pub-id pub-id-type="doi">10.1021/acs.biochem.5b00618</pub-id>
</citation>
</ref>
<ref id="B15">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Iaea</surname>
<given-names>D. B.</given-names>
</name>
<name>
<surname>Mao</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Lund</surname>
<given-names>F. W.</given-names>
</name>
<name>
<surname>Maxfield</surname>
<given-names>F. R.</given-names>
</name>
</person-group> (<year>2017</year>). <article-title>Role of STARD4 in sterol transport between the endocytic recycling compartment and the plasma membrane</article-title>. <source>Mol. Biol. Cell</source> <volume>28</volume>, <fpage>1111</fpage>&#x2013;<lpage>1122</lpage>. <pub-id pub-id-type="doi">10.1091/mbc.E16-07-0499</pub-id>
</citation>
</ref>
<ref id="B16">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Iaea</surname>
<given-names>D. B.</given-names>
</name>
<name>
<surname>Maxfield</surname>
<given-names>F. R.</given-names>
</name>
</person-group> (<year>2015</year>). <article-title>Cholesterol trafficking and distribution</article-title>. <source>Biochem.</source> <volume>57</volume>, <fpage>43</fpage>&#x2013;<lpage>55</lpage>. <pub-id pub-id-type="doi">10.1042/bse0570043</pub-id>
</citation>
</ref>
<ref id="B17">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Iaea</surname>
<given-names>D. B.</given-names>
</name>
<name>
<surname>Spahr</surname>
<given-names>Z. R.</given-names>
</name>
<name>
<surname>Singh</surname>
<given-names>R. K.</given-names>
</name>
<name>
<surname>Chan</surname>
<given-names>R. B.</given-names>
</name>
<name>
<surname>Zhou</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Bareja</surname>
<given-names>R.</given-names>
</name>
<etal/>
</person-group> (<year>2020</year>). <article-title>Stable reduction of STARD4 alters cholesterol regulation and lipid homeostasis</article-title>. <source>Biochim. Biophys. Acta Mol. Cell Biol. Lipids</source> <volume>1865</volume>, <fpage>158609</fpage>. <pub-id pub-id-type="doi">10.1016/j.bbalip.2020.158609</pub-id>
</citation>
</ref>
<ref id="B18">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Jentsch</surname>
<given-names>J. A.</given-names>
</name>
<name>
<surname>Kiburu</surname>
<given-names>I.</given-names>
</name>
<name>
<surname>Pandey</surname>
<given-names>K.</given-names>
</name>
<name>
<surname>Timme</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Ramlall</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Levkau</surname>
<given-names>B.</given-names>
</name>
<etal/>
</person-group> (<year>2018</year>). <article-title>Structural basis of sterol binding and transport by a yeast StARkin domain</article-title>. <source>J. Biol. Chem.</source> <volume>293</volume>, <fpage>5522</fpage>&#x2013;<lpage>5531</lpage>. <pub-id pub-id-type="doi">10.1074/jbc.RA118.001881</pub-id>
</citation>
</ref>
<ref id="B19">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Kohlhoff</surname>
<given-names>K. J.</given-names>
</name>
<name>
<surname>Shukla</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Lawrenz</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Bowman</surname>
<given-names>G. R.</given-names>
</name>
<name>
<surname>Konerding</surname>
<given-names>D. E.</given-names>
</name>
<name>
<surname>Belov</surname>
<given-names>D.</given-names>
</name>
<etal/>
</person-group> (<year>2014</year>). <article-title>Cloud-based simulations on Google Exacycle reveal ligand modulation of GPCR activation pathways</article-title>. <source>Nat. Chem.</source> <volume>6</volume>, <fpage>15</fpage>&#x2013;<lpage>21</lpage>. <pub-id pub-id-type="doi">10.1038/nchem.1821</pub-id>
</citation>
</ref>
<ref id="B20">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Lee</surname>
<given-names>D. D.</given-names>
</name>
<name>
<surname>Seung</surname>
<given-names>H. S.</given-names>
</name>
</person-group> (<year>1999</year>). <article-title>Learning the parts of objects by non-negative matrix factorization</article-title>. <source>Nature</source> <volume>401</volume>, <fpage>788</fpage>&#x2013;<lpage>791</lpage>. <pub-id pub-id-type="doi">10.1038/44565</pub-id>
</citation>
</ref>
<ref id="B21">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Letourneau</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Bedard</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Cabana</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Lefebvre</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>LeHoux</surname>
<given-names>J. G.</given-names>
</name>
<name>
<surname>Lavigne</surname>
<given-names>P.</given-names>
</name>
</person-group> (<year>2016</year>). <article-title>STARD6 on steroids: Solution structure, multiple timescale backbone dynamics and ligand binding mechanism</article-title>. <source>Sci. Rep.</source> <volume>6</volume>, <fpage>28486</fpage>. <pub-id pub-id-type="doi">10.1038/srep28486</pub-id>
</citation>
</ref>
<ref id="B22">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Letourneau</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Lefebvre</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Lavigne</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>LeHoux</surname>
<given-names>J. G.</given-names>
</name>
</person-group> (<year>2015</year>). <article-title>The binding site specificity of STARD4 subfamily: Breaking the cholesterol paradigm</article-title>. <source>Mol. Cell Endocrinol.</source> <volume>408</volume>, <fpage>53</fpage>&#x2013;<lpage>61</lpage>. <pub-id pub-id-type="doi">10.1016/j.mce.2014.12.016</pub-id>
</citation>
</ref>
<ref id="B23">
<citation citation-type="book">
<person-group person-group-type="author">
<name>
<surname>LeVine</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Perez-Aguilar</surname>
<given-names>J. M.</given-names>
</name>
<name>
<surname>Weinstein</surname>
<given-names>H.</given-names>
</name>
</person-group> (<year>2014</year>). <article-title>N-Body information theory (NbIT) analysis of rigid- body dynamics in intracellular loop 2 of the 5-HT 2A receptor</article-title>. <source>Proceedings International Work-Conference on Bioinformatics and Biomedical Engineering (IWBBIO- 2014)</source> <volume>2</volume>, <fpage>1190</fpage>&#x2013;<lpage>1200</lpage>.</citation>
</ref>
<ref id="B24">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>LeVine</surname>
<given-names>M. V.</given-names>
</name>
<name>
<surname>Weinstein</surname>
<given-names>H.</given-names>
</name>
</person-group> (<year>2014</year>). <article-title>NbIT--a new information theory-based analysis of allosteric mechanisms reveals residues that underlie function in the leucine transporter LeuT</article-title>. <source>PLoS Comput. Biol.</source> <volume>10</volume>, <fpage>e1003603</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pcbi.1003603</pub-id>
</citation>
</ref>
<ref id="B25">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Liscum</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Munn</surname>
<given-names>N. J.</given-names>
</name>
</person-group> (<year>1999</year>). <article-title>Intracellular cholesterol transport</article-title>. <source>Biochim. Biophys. Acta</source> <volume>1438</volume>, <fpage>19</fpage>&#x2013;<lpage>37</lpage>. <pub-id pub-id-type="doi">10.1016/s1388-1981(99)00043-8</pub-id>
</citation>
</ref>
<ref id="B26">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Mathieu</surname>
<given-names>A. P.</given-names>
</name>
<name>
<surname>Fleury</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Ducharme</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Lavigne</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>LeHoux</surname>
<given-names>J. G.</given-names>
</name>
</person-group> (<year>2002</year>). <article-title>Insights into steroidogenic acute regulatory protein (StAR)-dependent cholesterol transfer in mitochondria: Evidence from molecular modeling and structure-based thermodynamics supporting the existence of partially unfolded states of StAR</article-title>. <source>J. Mol. Endocrinol.</source> <volume>29</volume>, <fpage>327</fpage>&#x2013;<lpage>345</lpage>. <pub-id pub-id-type="doi">10.1677/jme.0.0290327</pub-id>
</citation>
</ref>
<ref id="B27">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>McGibbon</surname>
<given-names>R. T.</given-names>
</name>
<name>
<surname>andPande</surname>
<given-names>V. S.</given-names>
</name>
</person-group> (<year>2015</year>). <article-title>Variational cross-validation of slow dynamical modes in molecular kinetics</article-title>. <source>J. Chem. Phys.</source> <volume>142</volume>, <fpage>124105</fpage>. <pub-id pub-id-type="doi">10.1063/1.4916292</pub-id>
</citation>
</ref>
<ref id="B28">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Mesmin</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Pipalia</surname>
<given-names>N. H.</given-names>
</name>
<name>
<surname>Lund</surname>
<given-names>F. W.</given-names>
</name>
<name>
<surname>Ramlall</surname>
<given-names>T. F.</given-names>
</name>
<name>
<surname>Sokolov</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Eliezer</surname>
<given-names>D.</given-names>
</name>
<etal/>
</person-group> (<year>2011</year>). <article-title>STARD4 abundance regulates sterol transport and sensing</article-title>. <source>Mol. Biol. Cell</source> <volume>22</volume>, <fpage>4004</fpage>&#x2013;<lpage>4015</lpage>. <pub-id pub-id-type="doi">10.1091/mbc.E11-04-0372</pub-id>
</citation>
</ref>
<ref id="B29">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Molgedey</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Schuster</surname>
<given-names>H. G.</given-names>
</name>
</person-group> (<year>1994</year>). <article-title>Separation of a mixture of independent signals using time delayed correlations</article-title>. <source>Phys. Rev. Lett.</source> <volume>72</volume>, <fpage>3634</fpage>&#x2013;<lpage>3637</lpage>. <pub-id pub-id-type="doi">10.1103/PhysRevLett.72.3634</pub-id>
</citation>
</ref>
<ref id="B30">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Murcia</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Faraldo-Gomez</surname>
<given-names>J. D.</given-names>
</name>
<name>
<surname>Maxfield</surname>
<given-names>F. R.</given-names>
</name>
<name>
<surname>Roux</surname>
<given-names>B.</given-names>
</name>
</person-group> (<year>2006</year>). <article-title>Modeling the structure of the StART domains of MLN64 and StAR proteins in complex with cholesterol</article-title>. <source>J. Lipid Res.</source> <volume>47</volume>, <fpage>2614</fpage>&#x2013;<lpage>2630</lpage>. <pub-id pub-id-type="doi">10.1194/jlr.M600232-JLR200</pub-id>
</citation>
</ref>
<ref id="B31">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Naritomi</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Fuchigami</surname>
<given-names>S.</given-names>
</name>
</person-group> (<year>2011</year>). <article-title>Slow dynamics in protein fluctuations revealed by time-structure based independent component analysis: The case of domain motions</article-title>. <source>J. Chem. Phys.</source> <volume>134</volume>, <fpage>065101</fpage>. <pub-id pub-id-type="doi">10.1063/1.3554380</pub-id>
</citation>
</ref>
<ref id="B32">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Noe</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>andFischer</surname>
<given-names>S.</given-names>
</name>
</person-group> (<year>2008</year>). <article-title>Transition networks for modeling the kinetics of conformational change in macromolecules</article-title>. <source>Curr. Opin. Struct. Biol.</source> <volume>18</volume>, <fpage>154</fpage>&#x2013;<lpage>162</lpage>. <pub-id pub-id-type="doi">10.1016/j.sbi.2008.01.008</pub-id>
</citation>
</ref>
<ref id="B33">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Pande</surname>
<given-names>V. S.</given-names>
</name>
<name>
<surname>Beauchamp</surname>
<given-names>K.</given-names>
</name>
<name>
<surname>andBowman</surname>
<given-names>G. R.</given-names>
</name>
</person-group> (<year>2010</year>). <article-title>Everything you wanted to know about Markov State Models but were afraid to ask</article-title>. <source>Methods</source> <volume>52</volume>, <fpage>99</fpage>&#x2013;<lpage>105</lpage>. <pub-id pub-id-type="doi">10.1016/j.ymeth.2010.06.002</pub-id>
</citation>
</ref>
<ref id="B34">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Perez-Hernandez</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Paul</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Giorgino</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>De Fabritiis</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Noe</surname>
<given-names>F.</given-names>
</name>
</person-group> (<year>2013</year>). <article-title>Identification of slow molecular order parameters for Markov model construction</article-title>. <source>J. Chem. Phys.</source> <volume>139</volume>, <fpage>015102</fpage>. <pub-id pub-id-type="doi">10.1063/1.4811489</pub-id>
</citation>
</ref>
<ref id="B35">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Phillips</surname>
<given-names>J. C.</given-names>
</name>
<name>
<surname>Braun</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Wang</surname>
<given-names>W.</given-names>
</name>
<name>
<surname>Gumbart</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Tajkhorshid</surname>
<given-names>E.</given-names>
</name>
<name>
<surname>Villa</surname>
<given-names>E.</given-names>
</name>
<etal/>
</person-group> (<year>2005</year>). <article-title>Scalable molecular dynamics with NAMD</article-title>. <source>NAMD J. Comput. Chem.</source> <volume>26</volume>, <fpage>1781</fpage>&#x2013;<lpage>1802</lpage>. <pub-id pub-id-type="doi">10.1002/jcc.20289</pub-id>
</citation>
</ref>
<ref id="B36">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Plante</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>andWeinstein</surname>
<given-names>H.</given-names>
</name>
</person-group> (<year>2021</year>). <article-title>Ligand-dependent conformational transitions in molecular dynamics trajectories of GPCRs revealed by a new machine learning rare event detection protocol</article-title>. <source>Molecules</source> <volume>26</volume>, <fpage>3059</fpage>. <pub-id pub-id-type="doi">10.3390/molecules26103059</pub-id>
</citation>
</ref>
<ref id="B37">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Prinz</surname>
<given-names>J. H.</given-names>
</name>
<name>
<surname>Wu</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Sarich</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Keller</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Senne</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Held</surname>
<given-names>M.</given-names>
</name>
<etal/>
</person-group> (<year>2011</year>). <article-title>Markov models of molecular kinetics: Generation and validation</article-title>. <source>J. Chem. Phys.</source> <volume>134</volume>, <fpage>174105</fpage>. <pub-id pub-id-type="doi">10.1063/1.3565032</pub-id>
</citation>
</ref>
<ref id="B38">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Prinz</surname>
<given-names>W. A.</given-names>
</name>
</person-group> (<year>2007</year>). <article-title>Non-vesicular sterol transport in cells</article-title>. <source>Prog. Lipid Res.</source> <volume>46</volume>, <fpage>297</fpage>&#x2013;<lpage>314</lpage>. <pub-id pub-id-type="doi">10.1016/j.plipres.2007.06.002</pub-id>
</citation>
</ref>
<ref id="B39">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Radhakrishnan</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Goldstein</surname>
<given-names>J. L.</given-names>
</name>
<name>
<surname>McDonald</surname>
<given-names>J. G.</given-names>
</name>
<name>
<surname>Brown</surname>
<given-names>M. S.</given-names>
</name>
</person-group> (<year>2008</year>). <article-title>Switch-like control of SREBP-2 transport triggered by small changes in ER cholesterol: A delicate balance</article-title>. <source>Cell Metab.</source> <volume>8</volume>, <fpage>512</fpage>&#x2013;<lpage>521</lpage>. <pub-id pub-id-type="doi">10.1016/j.cmet.2008.10.008</pub-id>
</citation>
</ref>
<ref id="B40">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Rodriguez-Agudo</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Calderon-Dominguez</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Ren</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Marques</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Redford</surname>
<given-names>K.</given-names>
</name>
<name>
<surname>Medina-Torres</surname>
<given-names>M. A.</given-names>
</name>
<etal/>
</person-group> (<year>2011</year>). <article-title>Subcellular localization and regulation of StarD4 protein in macrophages and fibroblasts</article-title>. <source>Biochim. Biophys. Acta</source> <volume>1811</volume>, <fpage>597</fpage>&#x2013;<lpage>606</lpage>. <pub-id pub-id-type="doi">10.1016/j.bbalip.2011.06.028</pub-id>
</citation>
</ref>
<ref id="B41">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Rodriguez-Agudo</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Ren</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Wong</surname>
<given-names>E.</given-names>
</name>
<name>
<surname>Marques</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Redford</surname>
<given-names>K.</given-names>
</name>
<name>
<surname>Gil</surname>
<given-names>G.</given-names>
</name>
<etal/>
</person-group> (<year>2008</year>). <article-title>Intracellular cholesterol transporter StarD4 binds free cholesterol and increases cholesteryl ester formation</article-title>. <source>J. Lipid Res.</source> <volume>49</volume>, <fpage>1409</fpage>&#x2013;<lpage>1419</lpage>. <pub-id pub-id-type="doi">10.1194/jlr.M700537-JLR200</pub-id>
</citation>
</ref>
<ref id="B42">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Romanowski</surname>
<given-names>M. J.</given-names>
</name>
<name>
<surname>Soccio</surname>
<given-names>R. E.</given-names>
</name>
<name>
<surname>Breslow</surname>
<given-names>J. L.</given-names>
</name>
<name>
<surname>Burley</surname>
<given-names>S. K.</given-names>
</name>
</person-group> (<year>2002</year>). <article-title>Crystal structure of the <italic>Mus musculus</italic> cholesterol-regulated START protein 4 (StarD4) containing a StAR-related lipid transfer domain</article-title>. <source>Proc. Natl. Acad. Sci. U. S. A.</source> <volume>99</volume>, <fpage>6949</fpage>&#x2013;<lpage>6954</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.052140699</pub-id>
</citation>
</ref>
<ref id="B43">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Sali</surname>
<given-names>A.</given-names>
</name>
</person-group> (<year>1995</year>). <article-title>Comparative protein modeling by satisfaction of spatial restraints</article-title>. <source>Mol. Med. Today</source> <volume>1</volume>, <fpage>270</fpage>&#x2013;<lpage>277</lpage>. <pub-id pub-id-type="doi">10.1016/s1357-4310(95)91170-7</pub-id>
</citation>
</ref>
<ref id="B44">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Sherman</surname>
<given-names>W.</given-names>
</name>
<name>
<surname>Beard</surname>
<given-names>H. S.</given-names>
</name>
<name>
<surname>Farid</surname>
<given-names>R.</given-names>
</name>
</person-group> (<year>2006</year>). <article-title>Use of an induced fit receptor structure in virtual screening</article-title>. <source>Screen. Chem. Biol. Drug Des.</source> <volume>67</volume>, <fpage>83</fpage>&#x2013;<lpage>84</lpage>. <pub-id pub-id-type="doi">10.1111/j.1747-0285.2005.00327.x</pub-id>
</citation>
</ref>
<ref id="B45">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Sherman</surname>
<given-names>W.</given-names>
</name>
<name>
<surname>Day</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Jacobson</surname>
<given-names>M. P.</given-names>
</name>
<name>
<surname>Friesner</surname>
<given-names>R. A.</given-names>
</name>
<name>
<surname>Farid</surname>
<given-names>R.</given-names>
</name>
</person-group> (<year>2006</year>). <article-title>Novel procedure for modeling ligand/receptor induced fit effects</article-title>. <source>J. Med. Chem.</source> <volume>49</volume>, <fpage>534</fpage>&#x2013;<lpage>553</lpage>. <pub-id pub-id-type="doi">10.1021/jm050540c</pub-id>
</citation>
</ref>
<ref id="B46">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Shukla</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Meng</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Roux</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Pande</surname>
<given-names>V. S.</given-names>
</name>
</person-group> (<year>2014</year>). <article-title>Activation pathway of Src kinase reveals intermediate states as targets for drug design</article-title>. <source>Nat. Commun.</source> <volume>5</volume>, <fpage>3397</fpage>. <pub-id pub-id-type="doi">10.1038/ncomms4397</pub-id>
</citation>
</ref>
<ref id="B47">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Singh</surname>
<given-names>R. P.</given-names>
</name>
<name>
<surname>Brooks</surname>
<given-names>B. R.</given-names>
</name>
<name>
<surname>Klauda</surname>
<given-names>J. B.</given-names>
</name>
</person-group> (<year>2009</year>). <article-title>Binding and release of cholesterol in the Osh4 protein of yeast</article-title>. <source>Proteins</source> <volume>75</volume>, <fpage>468</fpage>&#x2013;<lpage>477</lpage>. <pub-id pub-id-type="doi">10.1002/prot.22263</pub-id>
</citation>
</ref>
<ref id="B48">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Soccio</surname>
<given-names>R. E.</given-names>
</name>
<name>
<surname>Adams</surname>
<given-names>R. M.</given-names>
</name>
<name>
<surname>Romanowski</surname>
<given-names>M. J.</given-names>
</name>
<name>
<surname>Sehayek</surname>
<given-names>E.</given-names>
</name>
<name>
<surname>Burley</surname>
<given-names>S. K.</given-names>
</name>
<name>
<surname>Breslow</surname>
<given-names>J. L.</given-names>
</name>
</person-group> (<year>2002</year>). <article-title>The cholesterol-regulated StarD4 gene encodes a StAR-related lipid transfer protein with two closely related homologues, StarD5 and StarD6</article-title>. <source>StarD5 StarD6 Proc Natl Acad Sci U. S. A.</source> <volume>99</volume>, <fpage>6943</fpage>&#x2013;<lpage>6948</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.052143799</pub-id>
</citation>
</ref>
<ref id="B49">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Tan</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Tong</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Chun</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Im</surname>
<given-names>Y. J.</given-names>
</name>
</person-group> (<year>2019</year>). <article-title>Structural analysis of human sterol transfer protein STARD4</article-title>. <source>Biochem. Biophys. Res. Commun.</source> <volume>520</volume>, <fpage>466</fpage>&#x2013;<lpage>472</lpage>. <pub-id pub-id-type="doi">10.1016/j.bbrc.2019.10.054</pub-id>
</citation>
</ref>
<ref id="B50">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Thorsell</surname>
<given-names>A. G.</given-names>
</name>
<name>
<surname>Lee</surname>
<given-names>W. H.</given-names>
</name>
<name>
<surname>Persson</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Siponen</surname>
<given-names>M. I.</given-names>
</name>
<name>
<surname>Nilsson</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Busam</surname>
<given-names>R. D.</given-names>
</name>
<etal/>
</person-group> (<year>2011</year>). <article-title>Comparative structural analysis of lipid binding START domains</article-title>. <source>PLoS One</source> <volume>6</volume>, <fpage>e19521</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pone.0019521</pub-id>
</citation>
</ref>
<ref id="B51">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Tong</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Manik</surname>
<given-names>M. K.</given-names>
</name>
<name>
<surname>Im</surname>
<given-names>Y. J.</given-names>
</name>
</person-group> (<year>2018</year>). <article-title>Structural basis of sterol recognition and nonvesicular transport by lipid transfer proteins anchored at membrane contact sites</article-title>. <source>Proc. Natl. Acad. Sci. U. S. A.</source> <volume>115</volume>, <fpage>E856-E865</fpage>&#x2013;<lpage>E865</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.1719709115</pub-id>
</citation>
</ref>
<ref id="B52">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Tsujishita</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>andHurley</surname>
<given-names>J. H.</given-names>
</name>
</person-group> (<year>2000</year>). <article-title>Structure and lipid transport mechanism of a StAR-related domain</article-title>. <source>Nat. Struct. Biol.</source> <volume>7</volume>, <fpage>408</fpage>&#x2013;<lpage>414</lpage>. <pub-id pub-id-type="doi">10.1038/75192</pub-id>
</citation>
</ref>
<ref id="B53">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Wong</surname>
<given-names>L. H.</given-names>
</name>
<name>
<surname>Gatta</surname>
<given-names>A. T.</given-names>
</name>
<name>
<surname>Levine</surname>
<given-names>T. P.</given-names>
</name>
</person-group> (<year>2019</year>). <article-title>Lipid transfer proteins: The lipid commute via shuttles, bridges and tubes</article-title>. <source>Nat. Rev. Mol. Cell Biol.</source> <volume>20</volume>, <fpage>85</fpage>&#x2013;<lpage>101</lpage>. <pub-id pub-id-type="doi">10.1038/s41580-018-0071-5</pub-id>
</citation>
</ref>
<ref id="B54">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Zhang</surname>
<given-names>X.</given-names>
</name>
<name>
<surname>Xie</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Iaea</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Khelashvili</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Weinstein</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Maxfield</surname>
<given-names>F. R.</given-names>
</name>
</person-group> (<year>2022</year>). <article-title>Phosphatidylinositol phosphates modulate interactions between the StarD4 sterol trafficking protein and lipid membranes</article-title>. <source>J. Biol. Chem.</source> <volume>298</volume>, <fpage>102058</fpage>. <pub-id pub-id-type="doi">10.1016/j.jbc.2022.102058</pub-id>
</citation>
</ref>
<ref id="B55">
<citation citation-type="journal">
<person-group person-group-type="author">
<name>
<surname>Zhu</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Weng</surname>
<given-names>Z.</given-names>
</name>
</person-group> (<year>2005</year>). <article-title>Fast: A novel protein structure alignment algorithm</article-title>. <source>Proteins</source> <volume>58</volume>, <fpage>618</fpage>&#x2013;<lpage>627</lpage>. <pub-id pub-id-type="doi">10.1002/prot.20331</pub-id>
</citation>
</ref>
</ref-list>
</back>
</article>