<?xml version="1.0" encoding="UTF-8" standalone="no"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Comput. Neurosci.</journal-id>
<journal-title>Frontiers in Computational Neuroscience</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Comput. Neurosci.</abbrev-journal-title>
<issn pub-type="epub">1662-5188</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fncom.2017.00104</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Neuroscience</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Conduction Delay Learning Model for Unsupervised and Supervised Classification of Spatio-Temporal Spike Patterns</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name><surname>Matsubara</surname> <given-names>Takashi</given-names></name>
<xref ref-type="author-notes" rid="fn001"><sup>&#x0002A;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/274653/overview"/>
</contrib>
</contrib-group>
<aff><institution>Computational Intelligence, Fundamentals of Computational Science, Department of Computational Science, Graduate School of System Informatics, Kobe University</institution>, <addr-line>Hyogo</addr-line>, <country>Japan</country></aff>
<author-notes>
<fn fn-type="edited-by"><p>Edited by: Florentin W&#x000F6;rg&#x000F6;tter, University of G&#x000F6;ttingen, Germany</p></fn>
<fn fn-type="edited-by"><p>Reviewed by: J. Michael Herrmann, University of Edinburgh, United Kingdom; Christian Leibold, Ludwig-Maximilians-Universit&#x000E4;t M&#x000FC;nchen, Germany</p></fn>
<fn fn-type="corresp" id="fn001"><p>&#x0002A;Correspondence: Takashi Matsubara <email>matsubara&#x00040;phoenix.kobe-u.ac.jp</email></p></fn>
</author-notes>
<pub-date pub-type="epub">
<day>21</day>
<month>11</month>
<year>2017</year>
</pub-date>
<pub-date pub-type="collection">
<year>2017</year>
</pub-date>
<volume>11</volume>
<elocation-id>104</elocation-id>
<history>
<date date-type="received">
<day>31</day>
<month>05</month>
<year>2017</year>
</date>
<date date-type="accepted">
<day>02</day>
<month>11</month>
<year>2017</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x000A9; 2017 Matsubara.</copyright-statement>
<copyright-year>2017</copyright-year>
<copyright-holder>Matsubara</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/"><p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) or licensor are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p></license>
</permissions>
<abstract>
<p>Precise spike timing is considered to play a fundamental role in communications and signal processing in biological neural networks. Understanding the mechanism of spike timing adjustment would deepen our understanding of biological systems and enable advanced engineering applications such as efficient computational architectures. However, the biological mechanisms that adjust and maintain spike timing remain unclear. Existing algorithms adopt a supervised approach, which adjusts the axonal conduction delay and synaptic efficacy until the spike timings approximate the desired timings. This study proposes a spike timing-dependent learning model that adjusts the axonal conduction delay and synaptic efficacy in both unsupervised and supervised manners. The proposed learning algorithm approximates the Expectation-Maximization algorithm, and classifies the input data encoded into spatio-temporal spike patterns. Even in the supervised classification, the algorithm requires no external spikes indicating the desired spike timings unlike existing algorithms. Furthermore, because the algorithm is consistent with biological models and hypotheses found in existing biological studies, it could capture the mechanism underlying biological delay learning.</p>
</abstract>
<kwd-group>
<kwd>spiking neural network</kwd>
<kwd>temporal coding</kwd>
<kwd>delay learning</kwd>
<kwd>activity-dependent myelination</kwd>
<kwd>spike timing-dependent plasticity</kwd>
<kwd>unsupervised learning</kwd>
</kwd-group>
<contract-num rid="cn001">16K12487</contract-num>
<contract-sponsor id="cn001">Japan Society for the Promotion of Science<named-content content-type="fundref-id">10.13039/501100001691</named-content></contract-sponsor>
<counts>
<fig-count count="7"/>
<table-count count="5"/>
<equation-count count="16"/>
<ref-count count="60"/>
<page-count count="16"/>
<word-count count="10249"/>
</counts>
</article-meta>
</front>
<body>
<sec sec-type="intro" id="s1">
<title>1. Introduction</title>
<p>As confirmed in biological studies, the precise timing of a neuronal spike plays a fundamental role in information processing in the central nervous system (Carr and Konishi, <xref ref-type="bibr" rid="B10">1990</xref>; Middlebrooks et al., <xref ref-type="bibr" rid="B35">1994</xref>; Seidl et al., <xref ref-type="bibr" rid="B47">2010</xref>). In information coding called <italic>temporal coding</italic>, the timing of at least one generated spike represents an output. In contrast, in rate coding, spiking neural network (SNN) generates spikes repeatedly over a certain period and the number of generated spikes represents an output (Brader et al., <xref ref-type="bibr" rid="B7">2007</xref>; Nessler et al., <xref ref-type="bibr" rid="B38">2009</xref>; Beyeler et al., <xref ref-type="bibr" rid="B4">2013</xref>; O&#x00027;Connor et al., <xref ref-type="bibr" rid="B39">2013</xref>; Diehl and Cook, <xref ref-type="bibr" rid="B11">2015</xref>; Zambrano and Bohte, <xref ref-type="bibr" rid="B60">2016</xref>). SNNs in rate coding are increasingly being investigated for efficient computational architectures in engineering applications. The SNN-based architectures consume less energy and require smaller hardware area than traditional artificial neural network architectures (Querlioz et al., <xref ref-type="bibr" rid="B43">2013</xref>; Neftci et al., <xref ref-type="bibr" rid="B37">2014</xref>; Cao et al., <xref ref-type="bibr" rid="B9">2015</xref>). However, repeated spike generation increases the computational time (VanRullen and Thorpe, <xref ref-type="bibr" rid="B55">2001</xref>). To improve the efficiency of SNN-based architectures, the SNN can be implemented in temporal coding rather than rate coding (Matsubara and Torikai, <xref ref-type="bibr" rid="B32">2013</xref>, <xref ref-type="bibr" rid="B33">2016</xref>).</p>
<p>When multiple pre-synaptic spikes simultaneously arrive at a post-synaptic neuron, they evoke a large excitatory post-synaptic potential (EPSP). In response, the post-synaptic neuron elicits a spike and thereby delivers the signal to the latter part of the SNN with a high probability (see Figure <xref ref-type="fig" rid="F1">1</xref>). In other words, a neuron behaves as a coincidence detector. As the membrane potential of the post-synaptic neuron increases rapidly with increasing efficacy of the projecting synapse, synaptic modification influences the post-synaptic spike timing. Many algorithms adjust the timing of post-synaptic spikes in a supervised manner by potentiating synapses that potentially evoke EPSPs at the desired timing, while depressing other synapses. Examples are the ReSuMe algorithm (Ponulak, <xref ref-type="bibr" rid="B42">2005</xref>; Sporea and Gr&#x000FC;ning, <xref ref-type="bibr" rid="B50">2013</xref>; Matsubara and Torikai, <xref ref-type="bibr" rid="B33">2016</xref>), the Tempotron algorithm (G&#x000FC;tig and Sompolinsky, <xref ref-type="bibr" rid="B19">2006</xref>; Yu et al., <xref ref-type="bibr" rid="B59">2014</xref>), and algorithms proposed in Pfister et al. (<xref ref-type="bibr" rid="B41">2006</xref>) and Paugam-Moisy et al. (<xref ref-type="bibr" rid="B40">2008</xref>). A pre-synaptic spike arrives at a post-synaptic neuron through the axon of the pre-synaptic neuron. The delay incurred by traveling through the axon is called the <italic>axonal conduction delay</italic> (Waxman and Swadlow, <xref ref-type="bibr" rid="B57">1976</xref>). Thus, temporal coding must consider the pre-synaptic spike timing plus the conduction delay (see left panel of Figure <xref ref-type="fig" rid="F1">1B</xref>). Izhikevich et al. (<xref ref-type="bibr" rid="B24">2004</xref>) and Izhikevich (<xref ref-type="bibr" rid="B23">2006b</xref>) demonstrated that even when the conduction delays are constant, an SNN with synaptic modification called <italic>spike timing-dependent plasticity</italic> (STDP) (Markram et al., <xref ref-type="bibr" rid="B30">1997</xref>; Bi and Poo, <xref ref-type="bibr" rid="B5">1998</xref>) self-organizes its characteristic responses to spatio-temporal spike patterns. In Gerstner et al. (<xref ref-type="bibr" rid="B17">1996</xref>), multiple pre-synaptic neurons driven by a single signal source deliver the signal to a post-synaptic neuron after various delays. Synaptic modification maintains the synapses that simultaneously evoke an EPSP by potentiating them and prunes other synapses by depressing them. SpikeProp (Bohte et al., <xref ref-type="bibr" rid="B6">2002</xref>) models a similar process in a supervised manner. However, under the assumption of multiple connections, the SNN requires numerous unused paths for future development. Alternatively, the SNN accepts limited spatio-temporal spike patterns depending on the initial network connections and conduction delays. Fixing the delay time limits the flexibility and efficiency of the SNN (see right panel of Figure <xref ref-type="fig" rid="F1">1B</xref>).</p>
<fig id="F1" position="float">
<label>Figure 1</label>
<caption><p><bold>(A)</bold> Schematic of a spiking neural network (SNN) consisting of multiple pre-synaptic neurons <italic>x</italic><sub><italic>i</italic></sub> and a post-synaptic neuron <italic>z. W</italic><sub>0</sub> denotes the weight of the synapse projecting from the pre-synaptic neuron <italic>x</italic><sub>0</sub> to the post-synaptic neuron <italic>z</italic>, and &#x003C4;<sub>0</sub> denotes the conduction delay of the axon corresponding to the synapse. <bold>(B,C)</bold> Post-synaptic spikes in response to spatio-temporal spike patterns. Each vertical line denotes a spike elicited by a neuron. Dotted lines show the transmission paths of the pre-synaptic spikes to the post-synaptic neuron <italic>z</italic>. <bold>(B)</bold> When the conduction delays are fixed, the post-synaptic neuron responds to the spatio-temporal pattern of pre-synaptic spikes and elicits a spike (left panel), but is unresponsive to a second pattern (right panel). <bold>(C)</bold> When the conduction delays are plastic, the optimized SNN generates post-synaptic spikes at times depending on the given spatio-temporal patterns of the pre-synaptic spikes.</p></caption>
<graphic xlink:href="fncom-11-00104-g0001.tif"/>
</fig>
<p>As mentioned above, information processing in a neural circuit requires the synchronous arrival of spikes elicited by multiple pre-synaptic neurons, and thus optimal conduction delay in the axons is critical. Axons are surrounded by <italic>myelin</italic>, which works as an electrical insulator and reduces the conduction delay (Rushton, <xref ref-type="bibr" rid="B45">1951</xref>). Myelination and demyelination of axons appear to depend on neuronal activity (Fields, <xref ref-type="bibr" rid="B14">2005</xref>; Bakkum et al., <xref ref-type="bibr" rid="B2">2008</xref>; Jamann et al., <xref ref-type="bibr" rid="B25">2017</xref>). Fields (<xref ref-type="bibr" rid="B15">2015</xref>) and Baraban et al. (<xref ref-type="bibr" rid="B3">2016</xref>) hypothesized that synaptic modification occurs only after the arrival timings of the pre-synaptic spikes have been adjusted by activity-dependent myelination.</p>
<p>The adaptation of the conduction delay is called <italic>delay learning</italic>. Supervised delay learning algorithms such as the DL-ReSuMe algorithm (Taherkhani et al., <xref ref-type="bibr" rid="B51">2015</xref>) directly adjust the conduction delay to suit the given spatio-temporal patterns of the pre- and post-synaptic spikes. After the adjustment, the SNN generates post-synaptic spikes at the desired timings (see Figure <xref ref-type="fig" rid="F1">1C</xref>). These algorithms answer the purpose to reproduce the spatio-temporal patterns, but are not suited for classification of the spatio-temporal patterns because the desired timings are generally unknown and should be manually adjusted with great care in the classification tasks. Hence, this approach reduces the flexibility in the context of machine learning and poorly represents biological systems that self-adapt to changing environments. In unsupervised delay learning, the SNN instead self-organizes in response to a given spike pattern (H&#x000FC;ning et al., <xref ref-type="bibr" rid="B21">1998</xref>; Eurich et al., <xref ref-type="bibr" rid="B12">1999</xref>, <xref ref-type="bibr" rid="B13">2000</xref>). However, none of the unsupervised delay learning algorithms tackled practical tasks such as classification and reproduction of given spike patterns (see Table <xref ref-type="table" rid="T1">1</xref> for comparison). Thus, a delay learning algorithm that classifies given spike patterns in both unsupervised and supervised manners is greatly desired.</p>
<table-wrap position="float" id="T1">
<label>Table 1</label>
<caption><p>Comparison of SNNs in different studies.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th valign="top" align="left"><bold>Studies</bold></th>
<th valign="top" align="left"><bold>Weight</bold></th>
<th valign="top" align="left"><bold>Delay</bold></th>
<th valign="top" align="left"><bold>Coding</bold></th>
<th valign="top" align="left"><bold>Learning manner</bold></th>
<th valign="top" align="left"><bold>Classification</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Brader et al., <xref ref-type="bibr" rid="B7">2007</xref> and others<xref ref-type="table-fn" rid="TN1"><sup>&#x0002A;</sup></xref></td>
<td valign="top" align="left">Plastic</td>
<td valign="top" align="left">No</td>
<td valign="top" align="left">Rate</td>
<td valign="top" align="left">Unsupervised and/or supervised</td>
<td valign="top" align="left">Yes</td>
</tr>
<tr>
<td valign="top" align="left">Ponulak, <xref ref-type="bibr" rid="B42">2005</xref> and others <xref ref-type="table-fn" rid="TN2"><sup>&#x0002A;&#x0002A;</sup></xref></td>
<td valign="top" align="left">Plastic</td>
<td valign="top" align="left">No</td>
<td valign="top" align="left">Temporal</td>
<td valign="top" align="left">Supervised</td>
<td valign="top" align="left">Yes</td>
</tr>
<tr>
<td valign="top" align="left">Izhikevich et al., <xref ref-type="bibr" rid="B24">2004</xref>; Izhikevich, <xref ref-type="bibr" rid="B23">2006b</xref></td>
<td valign="top" align="left">Plastic</td>
<td valign="top" align="left">Fixed</td>
<td valign="top" align="left">Temporal</td>
<td valign="top" align="left">Unsupervised</td>
<td valign="top" align="left">No</td>
</tr>
<tr>
<td valign="top" align="left">Gerstner et al., <xref ref-type="bibr" rid="B17">1996</xref></td>
<td valign="top" align="left">Plastic</td>
<td valign="top" align="left">Fixed/multiple</td>
<td valign="top" align="left">Temporal</td>
<td valign="top" align="left">Unsupervised</td>
<td valign="top" align="left">No</td>
</tr>
<tr>
<td valign="top" align="left">Bohte et al., <xref ref-type="bibr" rid="B6">2002</xref></td>
<td valign="top" align="left">Plastic</td>
<td valign="top" align="left">Fixed/multiple</td>
<td valign="top" align="left">Temporal</td>
<td valign="top" align="left">Supervised</td>
<td valign="top" align="left">Yes</td>
</tr>
<tr>
<td valign="top" align="left">Taherkhani et al., <xref ref-type="bibr" rid="B51">2015</xref></td>
<td valign="top" align="left">Plastic</td>
<td valign="top" align="left">Plastic</td>
<td valign="top" align="left">Temporal</td>
<td valign="top" align="left">Supervised</td>
<td valign="top" align="left">Yes</td>
</tr>
<tr>
<td valign="top" align="left">H&#x000FC;ning et al., <xref ref-type="bibr" rid="B21">1998</xref> and others <xref ref-type="table-fn" rid="TN3"><sup>&#x0002A;&#x0002A;&#x0002A;</sup></xref></td>
<td valign="top" align="left">Plastic</td>
<td valign="top" align="left">Plastic</td>
<td valign="top" align="left">Temporal</td>
<td valign="top" align="left">Unsupervised</td>
<td valign="top" align="left">No</td>
</tr>
<tr>
<td valign="top" align="left">This study</td>
<td valign="top" align="left">Plastic</td>
<td valign="top" align="left">Plastic</td>
<td valign="top" align="left">Temporal</td>
<td valign="top" align="left">Unsupervised and supervised</td>
<td valign="top" align="left">Yes</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<fn id="TN1">
<label>&#x0002A;</label>
<p><italic>Brader et al., <xref ref-type="bibr" rid="B7">2007</xref>; Nessler et al., <xref ref-type="bibr" rid="B38">2009</xref>; Beyeler et al., <xref ref-type="bibr" rid="B4">2013</xref>; O&#x00027;Connor et al., <xref ref-type="bibr" rid="B39">2013</xref>; Querlioz et al., <xref ref-type="bibr" rid="B43">2013</xref>; Neftci et al., <xref ref-type="bibr" rid="B37">2014</xref>; Diehl and Cook, <xref ref-type="bibr" rid="B11">2015</xref>; Zambrano and Bohte, <xref ref-type="bibr" rid="B60">2016</xref></italic>.</p></fn>
<fn id="TN2">
<label>&#x0002A;&#x0002A;</label>
<p><italic>Bohte et al., <xref ref-type="bibr" rid="B6">2002</xref>; Ponulak, <xref ref-type="bibr" rid="B42">2005</xref>; G&#x000FC;tig and Sompolinsky, <xref ref-type="bibr" rid="B19">2006</xref>; Pfister et al., <xref ref-type="bibr" rid="B41">2006</xref>; Paugam-Moisy et al., <xref ref-type="bibr" rid="B40">2008</xref>; Matsubara and Torikai, <xref ref-type="bibr" rid="B33">2016</xref></italic>.</p></fn>
<fn id="TN3">
<label>&#x0002A;&#x0002A;&#x0002A;</label>
<p><italic>H&#x000FC;ning et al., <xref ref-type="bibr" rid="B21">1998</xref>; Eurich et al., <xref ref-type="bibr" rid="B12">1999</xref>, <xref ref-type="bibr" rid="B13">2000</xref></italic>.</p></fn>
</table-wrap-foot>
</table-wrap>
<p>The present study proposes an unsupervised learning algorithm that adjusts the conduction delays and synaptic weights of an SNN. In a theoretical analysis, the proposed learning algorithm is confirmed to approximate the Expectation-Maximization (EM) algorithm. Optimization algorithms in probabilistic models including EM algorithm have been analyzed SNNs with no or fixed delay (e.g., Nessler et al., <xref ref-type="bibr" rid="B38">2009</xref>; Kappel et al., <xref ref-type="bibr" rid="B27">2014</xref>, <xref ref-type="bibr" rid="B26">2015</xref>; Rezende and Gerstner, <xref ref-type="bibr" rid="B44">2014</xref>). The proposed learning algorithm can be considered to extend the earlier methods to the conduction delay optimization. When evaluated on several practical classification tasks, the proposed algorithm successfully discriminated the given spatio-temporal spike patterns in an unsupervised manner. Furthermore, the algorithm was adaptable to supervised learning for improved classification accuracy. Remarkably, even in the supervised classification, the algorithm does give no external spikes indicating the timings at which the SNN should generate spikes. The adjustment of the synaptic weight in the proposed algorithm mimics that of STDP and occurs after the conduction delay was adjusted, supporting the hypothesis of Fields (<xref ref-type="bibr" rid="B15">2015</xref>) and Baraban et al. (<xref ref-type="bibr" rid="B3">2016</xref>). Therefore, the proposed learning algorithm presents as a good hypothetical model of biological delay learning, and might contribute to further investigations of SNNs in temporal coding. Preliminary and limited results of this study were presented in our conference paper (Matsubara, <xref ref-type="bibr" rid="B31">2017</xref>).</p>
</sec>
<sec id="s2">
<title>2. Spike timing-dependent conduction delay learning</title>
<sec>
<title>2.1. Spiking neural network (SNN)</title>
<p>An SNN consists of multiple pre-synaptic neurons <italic>x</italic><sub>0</sub>, &#x02026;, <italic>x</italic><sub><italic>N</italic>&#x02212;1</sub> connected to a post-synaptic neuron <italic>z</italic> (see Figure <xref ref-type="fig" rid="F1">1A</xref>). Let <italic>T</italic> be an experimental time period and &#x003B4; be the time step. The discrete times are denoted by <italic>t</italic> &#x0003D; 0, &#x003B4;, 2&#x003B4;, &#x02026;&#x000A0; (<italic>t</italic> &#x0003C; <italic>T</italic>). The binary variable <italic>x</italic><sub><italic>i,s</italic></sub> is set to 1 when a pre-synaptic neuron <italic>x</italic><sub><italic>i</italic></sub> generates a spike at time <italic>s</italic> and 0 otherwise. Owing to the discrete time, a spike has a width equal to the time step &#x003B4;. Hence, each set <bold><italic>x</italic></bold> &#x0003D; {<italic>x</italic><sub><italic>is</italic></sub>} represents a spatio-temporal spike pattern (see Figure <xref ref-type="fig" rid="F1">1B</xref>). The pre-synaptic neuron <italic>x</italic><sub><italic>i</italic></sub> is connected to the post-synaptic neuron <italic>z</italic> via a synapse with a synaptic weight <italic>W</italic><sub><italic>i</italic></sub> &#x02265; 0 and a conduction delay &#x003C4;<sub><italic>i</italic></sub> &#x02265; 0. A pre-synaptic spike <italic>x</italic><sub><italic>is</italic></sub> &#x0003D; 1 arrives at the post-synaptic neuron <italic>z</italic> after a conduction delay &#x003C4;<sub><italic>i</italic></sub>. After &#x003C4;<sub><italic>i</italic></sub>, the spike evokes an excitatory post-synaptic potential (EPSP) proportional to the synaptic weight <italic>W</italic><sub><italic>i</italic></sub>. The time course of the EPSP is expressed by a temporal relation function <italic>g</italic>(&#x00394;<italic>t</italic>) that depends on the temporal difference <italic>s</italic>&#x02212;<italic>t</italic> and the conduction delay &#x003C4;<sub><italic>i</italic></sub> (namely, &#x00394;<italic>t</italic> &#x0003D; <italic>s</italic> &#x0002B; &#x003C4;<sub><italic>i</italic></sub> &#x02212; <italic>t</italic>). For multiple pre-synaptic spikes, the EPSP is the linear sum of the EPSPs evoked by the pre-synaptic spikes. Then, the post-synaptic membrane potential <italic>v</italic><sub><italic>t</italic></sub> at time <italic>t</italic> is then given by</p>
<disp-formula id="E1"><label>(1)</label><mml:math id="M1"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:msub><mml:mrow><mml:mi>v</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mstyle displaystyle="true"><mml:munder class="msub"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder></mml:mstyle><mml:msub><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub><mml:msub><mml:mrow><mml:mi>W</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mi>g</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>&#x003C4;</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>-</mml:mo><mml:mi>t</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>The binary variable <italic>z</italic><sub><italic>t</italic></sub> is set to 1 when the post-synaptic neuron <italic>z</italic> generates a spike at time <italic>t</italic> and 0 otherwise. For the parameters <bold>&#x003C4;</bold> &#x0003D; {&#x003C4;<sub><italic>i</italic></sub>} and <bold><italic>W</italic></bold> &#x0003D; {<italic>W</italic><sub><italic>i</italic></sub>} and a given spatio-temporal spike pattern <bold><italic>x</italic></bold> &#x0003D; {<italic>x</italic><sub><italic>is</italic></sub>}, the post-synaptic neuron <italic>z</italic> is assumed to generate a single spike within the experimental time period 0 &#x02264; <italic>t</italic> &#x0003C; <italic>T</italic>, i.e., <inline-formula><mml:math id="M2"><mml:munder class="msub"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:munder><mml:msub><mml:mrow><mml:mi>z</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:math></inline-formula>. Such a constraint is important and is commonly applied in studies of temporal coding (e.g., Bohte et al., <xref ref-type="bibr" rid="B6">2002</xref>) where the output of the SNN in temporal coding is defined as the timing of the post-synaptic spike. The present study initially follows the previous studies, but later proposed an alternative model without the constraint (see section 2.5). This study also assumes that for a spatio-temporal spike pattern <bold><italic>x</italic></bold> &#x0003D; {<italic>x</italic><sub><italic>is</italic></sub>}, the probability of generating a post-synaptic spike <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 at time <italic>t</italic> is exponentially proportional to the EPSP normalized by a constant <inline-formula><mml:math id="M3"><mml:mi>Z</mml:mi><mml:mo>=</mml:mo><mml:munder class="msub"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:msup><mml:mrow><mml:mi>t</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x02032;</mml:mi></mml:mrow></mml:msup></mml:mrow></mml:munder><mml:mo class="qopname">exp</mml:mo><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>v</mml:mi></mml:mrow><mml:mrow><mml:msup><mml:mrow><mml:mi>t</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x02032;</mml:mi></mml:mrow></mml:msup></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> as follows:</p>
<disp-formula id="E2"><label>(2)</label><mml:math id="M4"><mml:mtable columnalign='left'><mml:mtr><mml:mtd><mml:mi>p</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mi>x</mml:mi></mml:mstyle><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mi>exp</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>v</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mrow><mml:mstyle displaystyle='true'><mml:msub><mml:mo>&#x02211;</mml:mo><mml:msup><mml:mi>t</mml:mi><mml:mo>&#x02032;</mml:mo></mml:msup></mml:msub><mml:mrow><mml:mi>exp</mml:mi></mml:mrow></mml:mstyle><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>v</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo stretchy='false'>)</mml:mo></mml:mrow></mml:mfrac></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mi>Z</mml:mi></mml:mfrac><mml:mi>exp</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>v</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo stretchy='false'>)</mml:mo></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>&#x0221D;</mml:mo><mml:mi>exp</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mstyle><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mo>)</mml:mo></mml:mrow><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>Note that <italic>p</italic>(<italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1|<bold><italic>x</italic></bold>) &#x0003E; 0 at any time <italic>t</italic>. In other words, the post-synaptic neuron <italic>z</italic> can elicit a spike even before the first pre-synaptic spike arrives. In biological neural networks, such spikes are induced by physiological noise. Note also that the normalizing constant <italic>Z</italic> includes a term <inline-formula><mml:math id="M7"><mml:msub><mml:mrow><mml:mi>v</mml:mi></mml:mrow><mml:mrow><mml:msup><mml:mrow><mml:mi>t</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x02032;</mml:mi></mml:mrow></mml:msup></mml:mrow></mml:msub></mml:math></inline-formula> denoting the membrane potential at a future time <italic>t</italic>&#x02032; &#x0003E; <italic>t</italic>. Therefore, sampling the post-synaptic spike <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 requires anti-causal information. The alternative model proposed in section 2.5 overcomes this limitation.</p>
</sec>
<sec>
<title>2.2. Multinoulli-Bernoulli mixture model</title>
<p>This subsection introduces the mixture probabilistic model. Let <bold><italic>x</italic></bold> be a set of visible binary random variables {<italic>x</italic><sub><italic>is</italic></sub>}, each following a Bernoulli distribution, and <bold><italic>z</italic></bold> be a set of latent binary random variables {<italic>z</italic><sub><italic>t</italic></sub>} following a Multinoulli distribution, i.e., <inline-formula><mml:math id="M8"><mml:munder class="msub"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:munder><mml:msub><mml:mrow><mml:mi>z</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:math></inline-formula>. Also let <bold>&#x003C9;</bold> &#x0003D; {&#x003C9;<sub><italic>t</italic></sub>} denote the mixture weights, and <bold>&#x003C0;</bold> &#x0003D; {&#x003C0;<sub><italic>ist</italic></sub>} denote the posterior probability of <italic>x</italic><sub><italic>is</italic></sub> &#x0003D; 1 given <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1. The generative Multinoulli-Bernoulli mixture model <italic>p</italic>(<bold><italic>z</italic></bold>, <bold><italic>x</italic></bold>; <bold>&#x003C9;</bold>, <bold>&#x003C0;</bold>) is then expressed as</p>
<disp-formula id="E5"><mml:math id="M9"><mml:mtable columnalign='left'><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mi>p</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>z</mml:mi></mml:mstyle><mml:mo>,</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>x</mml:mi></mml:mstyle><mml:mo>;</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>&#x003C9;</mml:mi></mml:mstyle><mml:mo>,</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>&#x003C0;</mml:mi></mml:mstyle><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x0220F;</mml:mo><mml:mi>t</mml:mi></mml:munder><mml:mrow><mml:msup><mml:mrow><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:msub><mml:mi>&#x003C9;</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x0220F;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:msup><mml:mrow><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>&#x003C0;</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:msup></mml:mrow></mml:mstyle><mml:msup><mml:mrow><mml:mo stretchy='false'>(</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x02212;</mml:mo><mml:msub><mml:mi>&#x003C0;</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x02212;</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:msup></mml:mrow><mml:mo>]</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:msup></mml:mrow></mml:mstyle><mml:mo>,</mml:mo></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>log&#x000A0;</mml:mtext><mml:mi>p</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>z</mml:mi></mml:mstyle><mml:mo>,</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>x</mml:mi></mml:mstyle><mml:mo>;</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>&#x003C9;</mml:mi></mml:mstyle><mml:mo>,</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>&#x003C0;</mml:mi></mml:mstyle><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mi>t</mml:mi></mml:munder><mml:mtext>&#x000A0;</mml:mtext></mml:mstyle><mml:mtext>&#x000A0;</mml:mtext><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtext>log&#x000A0;</mml:mtext><mml:msub><mml:mi>&#x003C9;</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo>&#x0002B;</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mstyle><mml:mtext>log&#x000A0;</mml:mtext><mml:mfrac><mml:mrow><mml:msub><mml:mi>&#x003C0;</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x02212;</mml:mo><mml:msub><mml:mi>&#x003C0;</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mfrac></mml:mrow></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mrow><mml:mrow><mml:mo>&#x0002B;</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:mtext>log&#x000A0;</mml:mtext></mml:mrow></mml:mstyle><mml:mo stretchy='false'>(</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x02212;</mml:mo><mml:msub><mml:mi>&#x003C0;</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>Here, the posterior probability &#x003C0;<sub><italic>ist</italic></sub> is substituted by sigm(<italic>V</italic><sub><italic>ist</italic></sub>) using the sigmoid function sigm(<italic>u</italic>) &#x0003D; (1&#x0002B;exp(&#x02212;<italic>u</italic>))<sup>&#x02212;1</sup>. In addition, <italic>V</italic><sub><italic>ist</italic></sub> is modeled by the function <italic>W</italic><sub><italic>i</italic></sub><italic>g</italic>(<italic>s</italic> &#x0002B; &#x003C4;<sub><italic>i</italic></sub> &#x02212; <italic>t</italic>)&#x02212;<italic>v</italic> where <italic>v</italic> is a bias parameter and the mixture weight <italic>w</italic><sub><italic>t</italic></sub> is assumed to have a constant value <italic>e</italic><sup><italic>b</italic></sup>. The log-probability is then rewritten as</p>
<disp-formula id="E8"><label>(3)</label><mml:math id="M12"><mml:mtable columnalign='left'><mml:mtr><mml:mtd><mml:mtext>log&#x000A0;</mml:mtext><mml:mi>p</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>z</mml:mi></mml:mstyle><mml:mo>,</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>x</mml:mi></mml:mstyle><mml:mo>;</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>&#x003C9;</mml:mi></mml:mstyle><mml:mo>,</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>&#x003C0;</mml:mi></mml:mstyle><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mstyle displaystyle='true'><mml:msub><mml:mo>&#x02211;</mml:mo><mml:mi>t</mml:mi></mml:msub><mml:mtext>&#x000A0;</mml:mtext></mml:mstyle><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtext>log&#x000A0;</mml:mtext><mml:msub><mml:mi>&#x003C9;</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo>&#x0002B;</mml:mo><mml:mstyle displaystyle='true'><mml:msub><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mstyle><mml:msub><mml:mi>V</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mstyle displaystyle='true'><mml:msub><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mtext>log&#x000A0;</mml:mtext></mml:mrow></mml:mstyle><mml:mo stretchy='false'>(</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0002B;</mml:mo><mml:msup><mml:mi>e</mml:mi><mml:mrow><mml:msub><mml:mi>V</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:msup><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mo>]</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mstyle displaystyle='true'><mml:msub><mml:mo>&#x02211;</mml:mo><mml:mi>t</mml:mi></mml:msub><mml:mtext>&#x000A0;</mml:mtext></mml:mstyle><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mi>b</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:mstyle displaystyle='true'><mml:msub><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mstyle><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x02212;</mml:mo><mml:mi>v</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>&#x02212;</mml:mo><mml:mstyle displaystyle='true'><mml:msub><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mtext>log&#x000A0;</mml:mtext></mml:mrow></mml:mstyle><mml:mrow><mml:mrow><mml:mo stretchy='false'>(</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0002B;</mml:mo><mml:msup><mml:mi>e</mml:mi><mml:mrow><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x02212;</mml:mo><mml:mi>v</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow></mml:msup><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mo>]</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mtext>log&#x000A0;</mml:mtext><mml:mi>p</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>z</mml:mi></mml:mstyle><mml:mo>,</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>x</mml:mi></mml:mstyle><mml:mo>;</mml:mo><mml:mi>W</mml:mi><mml:mo>,</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>&#x003C4;</mml:mi></mml:mstyle><mml:mo stretchy='false'>)</mml:mo><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>When the experimental time period <italic>T</italic> is sufficiently long and the time <italic>t</italic> is sufficiently distant from the temporal boundaries (0 and <italic>T</italic>), <inline-formula><mml:math id="M13"><mml:mover accent="true"><mml:mrow><mml:mi>b</mml:mi></mml:mrow><mml:mo>^</mml:mo></mml:mover></mml:math></inline-formula> is independent of the post-synaptic spike timing <italic>t</italic>; that is, <inline-formula><mml:math id="M14"><mml:mover accent="true"><mml:mrow><mml:mi>b</mml:mi></mml:mrow><mml:mo>^</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:mi>b</mml:mi><mml:mo>-</mml:mo><mml:munder class="msub"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mtext>log&#x000A0;</mml:mtext><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0002B;</mml:mo><mml:msup><mml:mrow><mml:mi>e</mml:mi></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>W</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mi>g</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>&#x003C4;</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>-</mml:mo><mml:mi>t</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>-</mml:mo><mml:mi>v</mml:mi></mml:mrow></mml:msup></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula>. The posterior probability of <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 given <bold><italic>x</italic></bold> &#x0003D; {<italic>x</italic><sub><italic>is</italic></sub>} is then equivalent to Equation (2);</p>
<disp-formula id="E9"><mml:math id="M15"><mml:mtable columnalign="left"><mml:mtr><mml:mtd><mml:mi>p</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>z</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mn>1</mml:mn><mml:mo>|</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mi>x</mml:mi></mml:mstyle><mml:mo>;</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mi>W</mml:mi></mml:mstyle><mml:mo>,</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mo>&#x003C4;</mml:mo></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mi>p</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>z</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mn>1</mml:mn><mml:mo>,</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mi>x</mml:mi></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mstyle displaystyle="true"><mml:msub><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:msup><mml:mrow><mml:mi>t</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x02032;</mml:mi></mml:mrow></mml:msup></mml:mrow></mml:msub></mml:mstyle><mml:mi>p</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>z</mml:mi></mml:mrow><mml:mrow><mml:msup><mml:mrow><mml:mi>t</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x02032;</mml:mi></mml:mrow></mml:msup></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mn>1</mml:mn><mml:mo>,</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mi>x</mml:mi></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow></mml:mfrac></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mo>&#x0221D;</mml:mo><mml:mo class="qopname">exp</mml:mo><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mover accent="true"><mml:mrow><mml:mi>b</mml:mi></mml:mrow><mml:mo class="qopname">^</mml:mo></mml:mover><mml:mo>&#x0002B;</mml:mo><mml:mstyle displaystyle="true"><mml:munder class="msub"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder></mml:mstyle><mml:msub><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>W</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mi>g</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>&#x003C4;</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>-</mml:mo><mml:mi>t</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>-</mml:mo><mml:mi>v</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mo>&#x0221D;</mml:mo><mml:mo class="qopname">exp</mml:mo><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle displaystyle="true"><mml:munder class="msub"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder></mml:mstyle><mml:msub><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub><mml:msub><mml:mrow><mml:mi>W</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mi>g</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>&#x003C4;</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>-</mml:mo><mml:mi>t</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
</sec>
<sec>
<title>2.3. Learning algorithm based on the EM algorithm</title>
<p>The SNN model Equation (2) can be optimized through optimization of the Multinoulli-Bernoulli mixture model Equation (4) under the certain constraints. Let &#x003B8; be the parameter set and <bold><italic>X</italic></bold> be a dataset of spatio-temporal spike patterns <bold><italic>x</italic></bold>. In general, training a generative model involves maximizing the model evidence &#x1D53C;<sub><bold><italic>x</italic></bold>&#x0007E;<bold><italic>X</italic></bold></sub>[log <italic>p</italic>(<bold><italic>x</italic></bold>; <bold>&#x003B8;</bold>)] (Murphy, <xref ref-type="bibr" rid="B36">2012</xref>). For an arbitrary distribution <italic>q</italic>(<bold><italic>z</italic></bold>), the evidence is expressed as</p>
<disp-formula id="E10"><mml:math id="M16"><mml:mtable columnalign="left"><mml:mtr><mml:mtd><mml:msub><mml:mrow><mml:mo>&#x1D53C;</mml:mo></mml:mrow><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>x</mml:mi></mml:mstyle><mml:mo>&#x0007E;</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mi>X</mml:mi></mml:mstyle></mml:mrow></mml:msub><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtext>log&#x000A0;</mml:mtext><mml:mi>p</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>x</mml:mi></mml:mstyle><mml:mo>;</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mo>&#x003B8;</mml:mo></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mo>&#x1D53C;</mml:mo></mml:mrow><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>x</mml:mi></mml:mstyle><mml:mo>&#x0007E;</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mi>X</mml:mi></mml:mstyle></mml:mrow></mml:msub><mml:msub><mml:mrow><mml:mo>&#x1D53C;</mml:mo></mml:mrow><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>z</mml:mi></mml:mstyle><mml:mo>&#x0007E;</mml:mo><mml:mi>q</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>z</mml:mi></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow></mml:msub><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtext>log&#x000A0;</mml:mtext><mml:mfrac><mml:mrow><mml:mi>p</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>z</mml:mi></mml:mstyle><mml:mo>,</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mi>x</mml:mi></mml:mstyle><mml:mo>;</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mo>&#x003B8;</mml:mo></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mi>q</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>z</mml:mi></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mi>p</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>z</mml:mi></mml:mstyle><mml:mo>|</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mi>x</mml:mi></mml:mstyle><mml:mo>;</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mo>&#x003B8;</mml:mo></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mi>q</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>z</mml:mi></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow></mml:mfrac></mml:mrow><mml:mo>]</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mo>&#x1D53C;</mml:mo></mml:mrow><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>x</mml:mi></mml:mstyle><mml:mo>&#x0007E;</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mi>X</mml:mi></mml:mstyle></mml:mrow></mml:msub><mml:msub><mml:mrow><mml:mo>&#x1D53C;</mml:mo></mml:mrow><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>z</mml:mi></mml:mstyle><mml:mo>&#x0007E;</mml:mo><mml:mi>q</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>z</mml:mi></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow></mml:msub><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtext>log&#x000A0;</mml:mtext><mml:mi>p</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>z</mml:mi></mml:mstyle><mml:mo>,</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mi>x</mml:mi></mml:mstyle><mml:mo>;</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mo>&#x003B8;</mml:mo></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>-</mml:mo><mml:mtext>log&#x000A0;</mml:mtext><mml:mi>q</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>z</mml:mi></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mo>]</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mo>&#x0002B;</mml:mo><mml:mtext>&#x000A0;</mml:mtext><mml:msub><mml:mrow><mml:mo>&#x1D53C;</mml:mo></mml:mrow><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>x</mml:mi></mml:mstyle><mml:mo>&#x0007E;</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mi>X</mml:mi></mml:mstyle></mml:mrow></mml:msub><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>D</mml:mi></mml:mrow><mml:mrow><mml:mi>K</mml:mi><mml:mi>L</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>q</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>z</mml:mi></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>|</mml:mo><mml:mo>|</mml:mo><mml:mi>p</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>z</mml:mi></mml:mstyle><mml:mo>|</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mi>x</mml:mi></mml:mstyle><mml:mo>;</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mo>&#x003B8;</mml:mo></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mo>]</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mo>=</mml:mo><mml:mrow><mml:mi mathvariant="-tex-caligraphic">L</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mo>&#x003B8;</mml:mo></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mo>&#x1D53C;</mml:mo></mml:mrow><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>x</mml:mi></mml:mstyle><mml:mo>&#x0007E;</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mi>X</mml:mi></mml:mstyle></mml:mrow></mml:msub><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>D</mml:mi></mml:mrow><mml:mrow><mml:mi>K</mml:mi><mml:mi>L</mml:mi></mml:mrow></mml:msub><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>q</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>z</mml:mi></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>|</mml:mo><mml:mo>|</mml:mo><mml:mi>p</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>z</mml:mi></mml:mstyle><mml:mo>|</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mi>x</mml:mi></mml:mstyle><mml:mo>;</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mo>&#x003B8;</mml:mo></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mo>,</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>where the evidence lower bound <inline-formula><mml:math id="M17"><mml:mrow><mml:mi mathvariant="-tex-caligraphic">L</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mo>&#x003B8;</mml:mo></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> equals &#x1D53C;<sub><bold><italic>x</italic></bold>&#x0007E;<bold><italic>X</italic></bold></sub>&#x1D53C;<sub><bold><italic>z</italic></bold>&#x0007E;<italic>q</italic>(<bold><italic>z</italic></bold>)</sub>[log <italic>p</italic>(<bold><italic>z</italic></bold>, <bold><italic>x</italic></bold>; <bold>&#x003B8;</bold>)&#x02212;log <italic>q</italic>(<bold><italic>z</italic></bold>)] and <italic>D</italic><sub><italic>KL</italic></sub>(&#x000B7;) is the Kullback-Leibler divergence. The model evidence &#x1D53C;<sub><bold><italic>x</italic></bold>&#x0007E;<bold><italic>X</italic></bold></sub>[log <italic>p</italic>(<bold><italic>x</italic></bold>; <bold>&#x003B8;</bold>)] can be maximized by the Expectation-Maximization (EM) algorithm. In the E-step, <italic>q</italic>(<bold><italic>z</italic></bold>) is substituted by the posterior probability <italic>p</italic>(<bold><italic>z</italic></bold>|<bold><italic>x</italic></bold>; <bold>&#x003B8;</bold><sup><italic>old</italic></sup>) of the mixture model with the currently estimated parameter set <bold>&#x003B8;</bold><sup><italic>old</italic></sup>. The M-step updates the estimated parameter set <bold>&#x003B8;</bold> to maximize the evidence lower bound <inline-formula><mml:math id="M18"><mml:mrow><mml:mi mathvariant="-tex-caligraphic">L</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mo>&#x003B8;</mml:mo></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula>. The present study employs a stochastic version of the EM algorithm (Sato, <xref ref-type="bibr" rid="B46">1999</xref>; Nessler et al., <xref ref-type="bibr" rid="B38">2009</xref>). Given a spatio-temporal spike pattern <inline-formula><mml:math id="M19"><mml:mover accent="true"><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>x</mml:mi></mml:mstyle></mml:mrow><mml:mo>^</mml:mo></mml:mover></mml:math></inline-formula>, the stochastic EM algorithm first generates a post-synaptic spike <inline-formula><mml:math id="M20"><mml:mover accent="true"><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>z</mml:mi></mml:mstyle></mml:mrow><mml:mo>^</mml:mo></mml:mover></mml:math></inline-formula> by Equation (2). This spike generation corresponds to the E-step of the EM algorithm. Next, the parameters &#x003B8; &#x0003D; {<bold><italic>W</italic></bold>, <bold>&#x003C4;</bold>} are updated to maximize the evidence lower bound <inline-formula><mml:math id="M21"><mml:mrow><mml:mi mathvariant="-tex-caligraphic">L</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mo>&#x003B8;</mml:mo></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> using a gradient ascent algorithm. This parameter update corresponds to the M-step of the EM algorithm. From an SNN perspective, the parameter update corresponds to synaptic plasticity with delay learning. Note that, unlike the original EM algorithm, the M-step of the stochastic EM algorithm is not guaranteed to maximize the model evidence &#x1D53C;<sub><bold><italic>x</italic></bold>&#x0007E;<bold><italic>X</italic></bold></sub>[log <italic>p</italic>(<bold><italic>x</italic></bold>; <bold>&#x003B8;</bold>)] but at least maximizes the evidence lower bound <inline-formula><mml:math id="M22"><mml:mrow><mml:mi mathvariant="-tex-caligraphic">L</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mo>&#x003B8;</mml:mo></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula>.</p>
<p>Based on Equation (4), the gradients of the evidence lower bound <inline-formula><mml:math id="M23"><mml:mrow><mml:mi mathvariant="-tex-caligraphic">L</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mo>&#x003B8;</mml:mo></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> w.r.t. the conduction delay &#x003C4;<sub><italic>i</italic></sub> and synaptic weight <italic>W</italic><sub><italic>i</italic></sub> are respectively given by;</p>
<disp-formula id="E19"><label>(4)</label><mml:math id="M65"><mml:mtable columnalign='left'><mml:mtr><mml:mtd><mml:mfrac><mml:mo>&#x02202;</mml:mo><mml:mrow><mml:mo>&#x02202;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:mfrac><mml:mi mathvariant="-tex-caligraphic">L</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>&#x003B8;</mml:mi></mml:mstyle><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mi>t</mml:mi></mml:munder><mml:mrow><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:mstyle><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mi>s</mml:mi></mml:munder><mml:mrow><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mtext>sigm</mml:mtext><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x02212;</mml:mo><mml:mi>v</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mo>)</mml:mo></mml:mrow></mml:mrow></mml:mstyle></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>&#x000D7;</mml:mo><mml:mtext>&#x000A0;</mml:mtext><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mfrac><mml:mrow><mml:mo>&#x02212;</mml:mo><mml:mo stretchy='false'>(</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo>&#x02212;</mml:mo><mml:mi>&#x003BC;</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mrow><mml:msup><mml:mi>&#x003C3;</mml:mi><mml:mn>2</mml:mn></mml:msup></mml:mrow></mml:mfrac></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mo>&#x02212;</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mi>t</mml:mi></mml:munder><mml:mrow><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:mstyle><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mi>s</mml:mi></mml:munder><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mstyle><mml:mo>&#x000B7;</mml:mo><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mfrac><mml:mrow><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo>&#x02212;</mml:mo><mml:mi>&#x003BC;</mml:mi></mml:mrow><mml:mrow><mml:msup><mml:mi>&#x003C3;</mml:mi><mml:mn>2</mml:mn></mml:msup></mml:mrow></mml:mfrac></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mo>&#x02212;</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mi>t</mml:mi></mml:munder><mml:mrow><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:mstyle><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mi>s</mml:mi></mml:munder><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mstyle><mml:mo>&#x000B7;</mml:mo><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x00394;</mml:mi><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mfrac><mml:mrow><mml:mi>&#x00394;</mml:mi><mml:mi>t</mml:mi><mml:mo>&#x02212;</mml:mo><mml:mi>&#x003BC;</mml:mi></mml:mrow><mml:mrow><mml:msup><mml:mi>&#x003C3;</mml:mi><mml:mn>2</mml:mn></mml:msup></mml:mrow></mml:mfrac><mml:mo>,</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<disp-formula id="E20"><label>(5)</label><mml:math id="M66"><mml:mtable columnalign='left'><mml:mtr><mml:mtd><mml:mfrac><mml:mo>&#x02202;</mml:mo><mml:mrow><mml:mo>&#x02202;</mml:mo><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:mfrac><mml:mi mathvariant="-tex-caligraphic">L</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>&#x003B8;</mml:mi></mml:mstyle><mml:mo stretchy='false'>)</mml:mo></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>t</mml:mi><mml:mo>,</mml:mo><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:mstyle><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mtext>sigm</mml:mtext><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x02212;</mml:mo><mml:mi>v</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>)</mml:mo><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>t</mml:mi><mml:mo>,</mml:mo><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:mstyle><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mo>&#x02212;</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>&#x00394;</mml:mi><mml:msup><mml:mi>t</mml:mi><mml:mo>&#x02032;</mml:mo></mml:msup></mml:mrow></mml:munder><mml:mtext>&#x000A0;</mml:mtext></mml:mstyle><mml:mtext>sigm</mml:mtext><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x00394;</mml:mi><mml:msup><mml:mi>t</mml:mi><mml:mo>&#x02032;</mml:mo></mml:msup><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x02212;</mml:mo><mml:mi>v</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x00394;</mml:mi><mml:msup><mml:mi>t</mml:mi><mml:mo>&#x02032;</mml:mo></mml:msup><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mo>]</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>t</mml:mi><mml:mo>,</mml:mo><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:mstyle><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x00394;</mml:mi><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mo>&#x02212;</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>&#x00394;</mml:mi><mml:msup><mml:mi>t</mml:mi><mml:mo>&#x02032;</mml:mo></mml:msup></mml:mrow></mml:munder><mml:mrow><mml:mtext>sigm</mml:mtext></mml:mrow></mml:mstyle><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x00394;</mml:mi><mml:msup><mml:mi>t</mml:mi><mml:mo>&#x02032;</mml:mo></mml:msup><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x02212;</mml:mo><mml:mi>v</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x00394;</mml:mi><mml:msup><mml:mi>t</mml:mi><mml:mo>&#x02032;</mml:mo></mml:msup><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>As Equation (4) and the first term in Equation (5) are independent of the time step &#x003B4;, the second term in Equation (5) (which does depend on &#x003B4;) is multiplied by &#x003B4;: This normalization makes Equation (5) robust to the size of the time step &#x003B4;. The parameters &#x003B8; are then updated as follows:</p>
<disp-formula id="E11"><mml:math id="M24"><mml:mtable columnalign="left"><mml:mtr><mml:mtd><mml:mi>&#x003B8;</mml:mi><mml:mo>&#x02190;</mml:mo><mml:mi>&#x003B8;</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:mi>&#x003B7;</mml:mi><mml:mfrac><mml:mrow><mml:mi>&#x02202;</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x02202;</mml:mi><mml:mi>&#x003B8;</mml:mi></mml:mrow></mml:mfrac><mml:mrow><mml:mi mathvariant="-tex-caligraphic">L</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mo>&#x003B8;</mml:mo></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:msub><mml:mrow><mml:mo stretchy="true">|</mml:mo></mml:mrow><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>z</mml:mi></mml:mstyle><mml:mo>=</mml:mo><mml:mover accent="true"><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>z</mml:mi></mml:mstyle></mml:mrow><mml:mo>^</mml:mo></mml:mover><mml:mo>,</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mi>x</mml:mi></mml:mstyle><mml:mo>=</mml:mo><mml:mover accent="true"><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>x</mml:mi></mml:mstyle></mml:mrow><mml:mo>^</mml:mo></mml:mover></mml:mrow></mml:msub><mml:mo>,</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>where &#x003B7; is a learning rate. The conditions are similar in a supervised manner, except that the post-synaptic spike <inline-formula><mml:math id="M25"><mml:mover accent="true"><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mi>z</mml:mi></mml:mstyle></mml:mrow><mml:mo>^</mml:mo></mml:mover></mml:math></inline-formula> is sourced externally.</p>
</sec>
<sec>
<title>2.4. Classification</title>
<p>A spatio-temporal spike pattern <bold><italic>x</italic></bold> is associated with one of the timings <italic>t</italic> &#x0003D; 0&#x003B4;, 1&#x003B4;, &#x02026;&#x000A0; depending on the timing <italic>t</italic> of the generated post-synaptic spike <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1. However, a given dataset <bold><italic>X</italic></bold> of spatio-temporal spike patterns <bold><italic>x</italic></bold> should be classifiable into smaller groups. In this study, the decision boundaries between <italic>N</italic> groups are defined as the <italic>n</italic>-th <italic>N</italic>-quantiles (<italic>n</italic> &#x0003D; 1, 2, &#x02026;, <italic>N</italic> &#x02212; 1) of the generated post-synaptic spike timings <italic>t</italic> for which <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1. For example, suppose that a dataset <bold><italic>X</italic></bold> contains 100 spatio-temporal spike patterns <bold><italic>x</italic></bold> falling into estimated two groups. The spatio-temporal spike patterns <bold><italic>x</italic></bold> are then sorted by the timings <italic>t</italic> of the post-synaptic spikes <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1. After sorting, patterns 1&#x02013;50 are classified into group 1 and the remainder are classified into group 2.</p>
</sec>
<sec>
<title>2.5. Alternative spiking neural network</title>
<p>In the SNN introduced above, the Multinoulli distribution <italic>p</italic>(<bold><italic>z</italic></bold>|<bold><italic>x</italic></bold>; <bold>&#x003C4;</bold>, <bold><italic>W</italic></bold>) must be sampled over time <italic>t</italic> to generate a post-synaptic spike <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1. In other words, whether a post-synaptic spike <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 is generated at time <italic>t</italic> anti-causally depends on the EPSP <inline-formula><mml:math id="M26"><mml:msub><mml:mrow><mml:mi>v</mml:mi></mml:mrow><mml:mrow><mml:msup><mml:mrow><mml:mi>t</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x02032;</mml:mi></mml:mrow></mml:msup></mml:mrow></mml:msub></mml:math></inline-formula> and post-synaptic spike <inline-formula><mml:math id="M27"><mml:msub><mml:mrow><mml:mi>z</mml:mi></mml:mrow><mml:mrow><mml:msup><mml:mrow><mml:mi>t</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x02032;</mml:mi></mml:mrow></mml:msup></mml:mrow></mml:msub></mml:math></inline-formula> at a future time <italic>t</italic>&#x02032; &#x0003E; <italic>t</italic>. To remove this unrealistic assumption, this subsection slightly modifies the SNN formulation. The new formulation replaces the Multinoulli distribution with multiple independent Bernoulli distributions and introduces the following Bernoulli&#x02013;Bernoulli mixture model with posterior probability <bold>&#x003C1;</bold> &#x0003D; {&#x003C1;<sub><italic>ist</italic></sub>} of <italic>x</italic><sub><italic>is</italic></sub> &#x0003D; 1 given <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 0;</p>
<disp-formula id="E12"><mml:math id="M28"><mml:mtable columnalign="left"><mml:mtr><mml:mtd><mml:mi>p</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>z</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>,</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mi>x</mml:mi></mml:mstyle><mml:mo>;</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mo>&#x003C9;</mml:mo></mml:mstyle><mml:mo>,</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mo>&#x003C0;</mml:mo></mml:mstyle><mml:mo>,</mml:mo><mml:mstyle mathvariant="bold-italic"><mml:mo>&#x003C1;</mml:mo></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:msup><mml:mrow><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>&#x003C9;</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mstyle displaystyle="true"><mml:munder class="msub"><mml:mrow><mml:mo>&#x0220F;</mml:mo></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder></mml:mstyle><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>&#x003C0;</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:msup><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>-</mml:mo><mml:msub><mml:mrow><mml:mi>&#x003C0;</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>1</mml:mn><mml:mo>-</mml:mo><mml:msub><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:msup></mml:mrow><mml:mo>]</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>z</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:msup></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mo>&#x000D7;</mml:mo><mml:msup><mml:mrow><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>-</mml:mo><mml:msub><mml:mrow><mml:mi>&#x003C9;</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mstyle displaystyle="true"><mml:munder class="msub"><mml:mrow><mml:mo>&#x0220F;</mml:mo></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder></mml:mstyle><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>&#x003C1;</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:msup><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>-</mml:mo><mml:msub><mml:mrow><mml:mi>&#x003C1;</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>1</mml:mn><mml:mo>-</mml:mo><mml:msub><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:msup></mml:mrow><mml:mo>]</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>1</mml:mn><mml:mo>-</mml:mo><mml:msub><mml:mrow><mml:mi>z</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:msup><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>In this case, given a spatio-temporal spike pattern <bold><italic>x</italic></bold>, the posterior probabilities <italic>p</italic>(<italic>z</italic><sub><italic>t</italic></sub>|<bold><italic>x</italic></bold>; <bold>&#x003C9;</bold>, <bold>&#x003C0;</bold>, <bold>&#x003C1;</bold>) of the post-synaptic spikes <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 at times <italic>t</italic> are completely independent. As described above, the prior probability &#x003C9;<sub><italic>t</italic></sub> of <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 is replaced with a time-independent term <italic>e</italic><sup><italic>b</italic></sup> and the posterior probability &#x003C0;<sub><italic>ist</italic></sub> of <italic>x</italic><sub><italic>is</italic></sub> &#x0003D; 1 given <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 is modeled as &#x003C0;<sub><italic>ist</italic></sub> &#x0003D; sigm(<italic>V</italic><sub><italic>ist</italic></sub>) &#x0003D; sigm(<italic>W</italic><sub><italic>i</italic></sub><italic>g</italic>(<italic>s</italic> &#x0002B; &#x003C4;<sub><italic>i</italic></sub> &#x02212; <italic>t</italic>)&#x02212;<italic>v</italic>). In addition, the posterior probability &#x003C1;<sub><italic>ist</italic></sub> is substituted by a constant sigm(<italic>v</italic>). The log-probability function becomes</p>
<disp-formula id="E13"><mml:math id="M29"><mml:mtable columnalign='left'><mml:mtr><mml:mtd><mml:mtext>log&#x000A0;</mml:mtext><mml:mi>p</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:mi>x</mml:mi><mml:mo>;</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>&#x003C9;</mml:mi></mml:mstyle><mml:mo>,</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>&#x003C0;</mml:mi></mml:mstyle><mml:mo>,</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>&#x003C1;</mml:mi></mml:mstyle><mml:mo stretchy='false'>)</mml:mo></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtext>log&#x000A0;</mml:mtext><mml:msub><mml:mi>&#x003C9;</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo>&#x0002B;</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mstyle><mml:mtext>log&#x000A0;</mml:mtext><mml:mfrac><mml:mrow><mml:msub><mml:mi>&#x003C0;</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x02212;</mml:mo><mml:msub><mml:mi>&#x003C0;</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mfrac><mml:mo>&#x0002B;</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:mtext>log&#x000A0;</mml:mtext></mml:mrow></mml:mstyle><mml:mo stretchy='false'>(</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x02212;</mml:mo><mml:msub><mml:mi>&#x003C0;</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mo>]</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>&#x0002B;</mml:mo><mml:mo stretchy='false'>(</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x02212;</mml:mo><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo stretchy='false'>)</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtext>log&#x000A0;</mml:mtext><mml:mo stretchy='false'>(</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x02212;</mml:mo><mml:msub><mml:mi>&#x003C9;</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x0002B;</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mstyle><mml:mtext>log&#x000A0;</mml:mtext><mml:mfrac><mml:mrow><mml:msub><mml:mi>&#x003C1;</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x02212;</mml:mo><mml:msub><mml:mi>&#x003C1;</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mfrac></mml:mrow></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mrow><mml:mrow><mml:mo>&#x0002B;</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:mtext>log&#x000A0;</mml:mtext></mml:mrow></mml:mstyle><mml:mo stretchy='false'>(</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x02212;</mml:mo><mml:msub><mml:mi>&#x003C1;</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mo>]</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtext>log&#x000A0;</mml:mtext><mml:mfrac><mml:mrow><mml:msub><mml:mi>&#x003C9;</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x02212;</mml:mo><mml:msub><mml:mi>&#x003C9;</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:mfrac><mml:mo>&#x0002B;</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mstyle><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:msub><mml:mi>V</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>&#x0002B;</mml:mo><mml:mi>v</mml:mi></mml:mrow><mml:mo>)</mml:mo></mml:mrow><mml:mo>&#x02212;</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:mtext>log&#x000A0;</mml:mtext></mml:mrow></mml:mstyle><mml:mfrac><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0002B;</mml:mo><mml:msup><mml:mi>e</mml:mi><mml:mrow><mml:msub><mml:mi>V</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:msup></mml:mrow><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0002B;</mml:mo><mml:msup><mml:mi>e</mml:mi><mml:mrow><mml:mo>&#x02212;</mml:mo><mml:mi>v</mml:mi></mml:mrow></mml:msup></mml:mrow></mml:mfrac></mml:mrow><mml:mo>]</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>&#x0002B;</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtext>log&#x000A0;</mml:mtext><mml:mo stretchy='false'>(</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x02212;</mml:mo><mml:msub><mml:mi>&#x003C9;</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x02212;</mml:mo><mml:mi>v</mml:mi><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mstyle><mml:mo>&#x0002B;</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:mtext>log&#x000A0;</mml:mtext></mml:mrow></mml:mstyle><mml:mo stretchy='false'>(</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0002B;</mml:mo><mml:msup><mml:mi>e</mml:mi><mml:mrow><mml:mo>&#x02212;</mml:mo><mml:mi>v</mml:mi></mml:mrow></mml:msup><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mo>]</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mover accent='true'><mml:mi>b</mml:mi><mml:mo>&#x0005E;</mml:mo></mml:mover><mml:mo>&#x0002B;</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mstyle><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mrow><mml:mrow><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>&#x02212;</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:mtext>log&#x000A0;</mml:mtext></mml:mrow></mml:mstyle><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0002B;</mml:mo><mml:msup><mml:mi>e</mml:mi><mml:mrow><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x02212;</mml:mo><mml:mi>v</mml:mi></mml:mrow></mml:msup></mml:mrow><mml:mo>)</mml:mo></mml:mrow></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mo>&#x0002B;</mml:mo><mml:mover accent='true'><mml:mi>c</mml:mi><mml:mo>&#x0005E;</mml:mo></mml:mover></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mtext>log&#x000A0;</mml:mtext><mml:mi>p</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo>,</mml:mo><mml:mi>x</mml:mi><mml:mo>;</mml:mo><mml:mi>W</mml:mi><mml:mo>,</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>&#x003C4;</mml:mi></mml:mstyle><mml:mo stretchy='false'>)</mml:mo><mml:mo>,</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>where <inline-formula><mml:math id="M30"><mml:mover accent="true"><mml:mrow><mml:mi>b</mml:mi></mml:mrow><mml:mo>^</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:mtext>log&#x000A0;</mml:mtext><mml:mfrac><mml:mrow><mml:msub><mml:mrow><mml:mi>&#x003C9;</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mrow><mml:mn>1</mml:mn><mml:mo>-</mml:mo><mml:msub><mml:mrow><mml:mi>&#x003C9;</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mfrac><mml:mo>&#x0002B;</mml:mo><mml:munder class="msub"><mml:mrow><mml:mo class="qopname">&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mtext>log&#x000A0;</mml:mtext><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0002B;</mml:mo><mml:msup><mml:mrow><mml:mi>e</mml:mi></mml:mrow><mml:mrow><mml:mo>-</mml:mo><mml:mi>v</mml:mi></mml:mrow></mml:msup></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M31"><mml:mi>&#x00109;</mml:mi><mml:mo>=</mml:mo><mml:mtext>log&#x000A0;</mml:mtext><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>-</mml:mo><mml:msub><mml:mrow><mml:mi>&#x003C9;</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>-</mml:mo><mml:mi>v</mml:mi><mml:munder class="msub"><mml:mrow><mml:mo class="qopname">&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:msub><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub><mml:mo>&#x0002B;</mml:mo><mml:munder class="msub"><mml:mrow><mml:mo class="qopname">&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mtext>log&#x000A0;</mml:mtext><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0002B;</mml:mo><mml:msup><mml:mrow><mml:mi>e</mml:mi></mml:mrow><mml:mrow><mml:mo>-</mml:mo><mml:mi>v</mml:mi></mml:mrow></mml:msup></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula>. The prior probability &#x003C9;<sub><italic>t</italic></sub> is replaced by an independent variable <inline-formula><mml:math id="M32"><mml:mover accent="true"><mml:mrow><mml:mi>b</mml:mi></mml:mrow><mml:mo>^</mml:mo></mml:mover></mml:math></inline-formula> and a variable &#x00109; that depends on <inline-formula><mml:math id="M33"><mml:mover accent="true"><mml:mrow><mml:mi>b</mml:mi></mml:mrow><mml:mo>^</mml:mo></mml:mover></mml:math></inline-formula> and <italic>v</italic>.</p>
<p>The posterior probability of <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 given <bold><italic>x</italic></bold> is expressed as</p>
<disp-formula id="E16"><label>(6)</label><mml:math id="M36"><mml:mtable columnalign='left'><mml:mtr><mml:mtd><mml:mi>p</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>x</mml:mi><mml:mo>;</mml:mo><mml:mi>W</mml:mi><mml:mo>,</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>&#x003C4;</mml:mi></mml:mstyle><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mtext>sigm</mml:mtext><mml:mo stretchy='false'>(</mml:mo><mml:mtext>log&#x000A0;</mml:mtext><mml:mi>p</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mn>1</mml:mn><mml:mo>,</mml:mo><mml:mi>x</mml:mi><mml:mo>;</mml:mo><mml:mi>W</mml:mi><mml:mo>,</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>&#x003C4;</mml:mi></mml:mstyle><mml:mo stretchy='false'>)</mml:mo></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>&#x02212;</mml:mo><mml:mtext>log&#x000A0;</mml:mtext><mml:mi>p</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mn>0</mml:mn><mml:mo>,</mml:mo><mml:mi>x</mml:mi><mml:mo>;</mml:mo><mml:mi>W</mml:mi><mml:mo>,</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>&#x003C4;</mml:mi></mml:mstyle><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>)</mml:mo></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mtext>sigm</mml:mtext><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mover accent='true'><mml:mi>b</mml:mi><mml:mo>&#x0005E;</mml:mo></mml:mover><mml:mo>&#x0002B;</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mstyle><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi></mml:mrow><mml:mo>)</mml:mo></mml:mrow></mml:mrow><mml:mo>)</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>&#x02212;</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:mtext>log&#x000A0;</mml:mtext></mml:mrow></mml:mstyle><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mrow><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0002B;</mml:mo><mml:msup><mml:mi>e</mml:mi><mml:mrow><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x02212;</mml:mo><mml:mi>v</mml:mi></mml:mrow></mml:msup></mml:mrow><mml:mo>)</mml:mo></mml:mrow></mml:mrow><mml:mo>)</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mtext>sigm</mml:mtext><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mstyle><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x0002B;</mml:mo><mml:mover accent='true'><mml:mi>b</mml:mi><mml:mo>&#x002DC;</mml:mo></mml:mover></mml:mrow><mml:mo>)</mml:mo></mml:mrow><mml:mo>,</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>where <inline-formula><mml:math id="M37"><mml:mover accent="true"><mml:mrow><mml:mi>b</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:mover accent="true"><mml:mrow><mml:mi>b</mml:mi></mml:mrow><mml:mo>^</mml:mo></mml:mover><mml:mo>-</mml:mo><mml:munder class="msub"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mtext>log&#x000A0;</mml:mtext><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0002B;</mml:mo><mml:msup><mml:mrow><mml:mi>e</mml:mi></mml:mrow><mml:mrow><mml:msub><mml:mrow><mml:mi>W</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mi>g</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mrow><mml:mi>&#x003C4;</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>-</mml:mo><mml:mi>t</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>-</mml:mo><mml:mi>v</mml:mi></mml:mrow></mml:msup></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula>. The variable <inline-formula><mml:math id="M38"><mml:mover accent="true"><mml:mrow><mml:mi>b</mml:mi></mml:mrow><mml:mo>^</mml:mo></mml:mover></mml:math></inline-formula> is replaced by a variable <inline-formula><mml:math id="M39"><mml:mover accent="true"><mml:mrow><mml:mi>b</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:math></inline-formula>, which is treated as an independent parameter hereafter. The independent parameter <italic>b</italic> in the Bernoulli&#x02013;Bernoulli mixture model is additional to the parameters in the Multinoulli-Bernoulli mixture model.</p>
<p>The evidence lower bound <inline-formula><mml:math id="M40"><mml:mrow><mml:mi mathvariant="-tex-caligraphic">L</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mo>&#x003B8;</mml:mo></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> is given by &#x1D53C;<sub><bold><italic>x</italic></bold>&#x0007E;<bold><italic>X</italic></bold></sub>&#x1D53C;<sub><italic>z</italic><sub><italic>t</italic></sub>&#x0007E;<italic>q</italic>(<italic>z</italic><sub><italic>t</italic></sub>)</sub>[log <italic>p</italic>(<italic>z</italic><sub><italic>t</italic></sub>, <bold><italic>x</italic></bold>; <bold>&#x003B8;</bold>)&#x02212;log <italic>q</italic>(<italic>z</italic><sub><italic>t</italic></sub>)]]. The gradients of <inline-formula><mml:math id="M41"><mml:mrow><mml:mi mathvariant="-tex-caligraphic">L</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mo>&#x003B8;</mml:mo></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> w.r.t. the parameters &#x003C4;<sub><italic>i</italic></sub> and <italic>W</italic><sub><italic>i</italic></sub> are, respectively given by</p>
<disp-formula id="E22"><label>(7)</label><mml:math id="M68"><mml:mtable columnalign='left'><mml:mtr><mml:mtd><mml:mfrac><mml:mo>&#x02202;</mml:mo><mml:mrow><mml:mo>&#x02202;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:mfrac><mml:mi mathvariant="-tex-caligraphic">L</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>&#x003B8;</mml:mi></mml:mstyle><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mi>s</mml:mi></mml:munder><mml:mrow><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mtext>sigm</mml:mtext><mml:mo stretchy='false'>(</mml:mo></mml:mrow></mml:mstyle><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x02212;</mml:mo><mml:mi>v</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>)</mml:mo></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>&#x000D7;</mml:mo><mml:mtext>&#x000A0;</mml:mtext><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mfrac><mml:mrow><mml:mo>&#x02212;</mml:mo><mml:mo stretchy='false'>(</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo>&#x02212;</mml:mo><mml:mi>&#x003BC;</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mrow><mml:msup><mml:mi>&#x003C3;</mml:mi><mml:mn>2</mml:mn></mml:msup></mml:mrow></mml:mfrac></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mo>&#x02212;</mml:mo><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mi>s</mml:mi></mml:munder><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mstyle><mml:mo>&#x000B7;</mml:mo><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mfrac><mml:mrow><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo>&#x02212;</mml:mo><mml:mi>&#x003BC;</mml:mi></mml:mrow><mml:mrow><mml:msup><mml:mi>&#x003C3;</mml:mi><mml:mn>2</mml:mn></mml:msup></mml:mrow></mml:mfrac></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mo>&#x02212;</mml:mo><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mi>s</mml:mi></mml:munder><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:mstyle><mml:mo>&#x000B7;</mml:mo><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x00394;</mml:mi><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mfrac><mml:mrow><mml:mi>&#x00394;</mml:mi><mml:mi>t</mml:mi><mml:mo>&#x02212;</mml:mo><mml:mi>&#x003BC;</mml:mi></mml:mrow><mml:mrow><mml:msup><mml:mi>&#x003C3;</mml:mi><mml:mn>2</mml:mn></mml:msup></mml:mrow></mml:mfrac><mml:mo>,</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<disp-formula id="E21"><label>(8)</label><mml:math id="M67"><mml:mtable columnalign='left'><mml:mtr><mml:mtd><mml:mfrac><mml:mo>&#x02202;</mml:mo><mml:mrow><mml:mo>&#x02202;</mml:mo><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:mfrac><mml:mi mathvariant="-tex-caligraphic">L</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mstyle mathvariant='bold' mathsize='normal'><mml:mi>&#x003B8;</mml:mi></mml:mstyle><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:mstyle><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mtext>sigm</mml:mtext><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x02212;</mml:mo><mml:mi>v</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>)</mml:mo><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:mstyle><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>s</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mi>&#x003C4;</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x02212;</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mo>&#x02212;</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>&#x00394;</mml:mi><mml:msup><mml:mi>t</mml:mi><mml:mo>&#x02032;</mml:mo></mml:msup></mml:mrow></mml:munder><mml:mtext>&#x000A0;</mml:mtext></mml:mstyle><mml:mtext>sigm</mml:mtext><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x00394;</mml:mi><mml:msup><mml:mi>t</mml:mi><mml:mo>&#x02032;</mml:mo></mml:msup><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x02212;</mml:mo><mml:mi>v</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x00394;</mml:mi><mml:msup><mml:mi>t</mml:mi><mml:mo>&#x02032;</mml:mo></mml:msup><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mo>]</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>s</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub></mml:mrow></mml:mstyle><mml:msub><mml:mi>x</mml:mi><mml:mrow><mml:mi>i</mml:mi><mml:mi>s</mml:mi></mml:mrow></mml:msub><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x00394;</mml:mi><mml:mi>t</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mo>&#x02212;</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>&#x00394;</mml:mi><mml:msup><mml:mi>t</mml:mi><mml:mo>&#x02032;</mml:mo></mml:msup></mml:mrow></mml:munder><mml:mrow><mml:mtext>sigm</mml:mtext></mml:mrow></mml:mstyle><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>W</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x00394;</mml:mi><mml:msup><mml:mi>t</mml:mi><mml:mo>&#x02032;</mml:mo></mml:msup><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x02212;</mml:mo><mml:mi>v</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mi>g</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x00394;</mml:mi><mml:msup><mml:mi>t</mml:mi><mml:mo>&#x02032;</mml:mo></mml:msup><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>Equations (7) and (8) are equivalent to Equations (4) and (5) respectively, but are summed over time <italic>t</italic> (they sum the post-synaptic spikes <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1). Therefore, given a spatio-temporal spike pattern <bold><italic>x</italic></bold> and a single post-synaptic spike <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1, the parameters are updated as in the Multinoulli-Bernoulli mixture model.</p>
<p>Unlike the SNN based on the Multinoulli-Bernoulli mixture model, the above SNN is unconstrained by the number of post-synaptic spikes (i.e., <inline-formula><mml:math id="M42"><mml:munder class="msub"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:munder><mml:msub><mml:mrow><mml:mi>z</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>&#x02208;</mml:mo><mml:mi>&#x02124;</mml:mi></mml:math></inline-formula>). To prevent an excessive number or the complete absence of post-synaptic spikes, this subsection introduces <italic>homeostatic plasticity</italic> (Turrigiano et al., <xref ref-type="bibr" rid="B52">1999</xref>; Turrigiano and Nelson, <xref ref-type="bibr" rid="B53">2000</xref>), a biological mechanism that presumably maintains the excitability of a neuron within a regular range (Van Rossum et al., <xref ref-type="bibr" rid="B54">2000</xref>; Abraham, <xref ref-type="bibr" rid="B1">2008</xref>; Watt and Desai, <xref ref-type="bibr" rid="B56">2010</xref>; Matsubara and Uehara, <xref ref-type="bibr" rid="B34">2016</xref>). The additional parameter <inline-formula><mml:math id="M43"><mml:mover accent="true"><mml:mrow><mml:mi>b</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:math></inline-formula> in Equation (6) embodies the intrinsic excitability of the Bernoulli neuron <italic>z</italic>. The intrinsic excitability <inline-formula><mml:math id="M44"><mml:mover accent="true"><mml:mrow><mml:mi>b</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:math></inline-formula> is adjusted after every receipt of the spatio-temporal spike pattern <bold><italic>x</italic></bold>. When post-synaptic neuron <italic>z</italic> generates at least one spike (i.e., <inline-formula><mml:math id="M45"><mml:munder class="msub"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:munder><mml:msub><mml:mrow><mml:mi>z</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>&#x0003E;</mml:mo><mml:mn>0</mml:mn></mml:math></inline-formula>), its intrinsic excitability <inline-formula><mml:math id="M46"><mml:mover accent="true"><mml:mrow><mml:mi>b</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:math></inline-formula> reduces by <inline-formula><mml:math id="M47"><mml:msub><mml:mrow><mml:mover accent="true"><mml:mrow><mml:mi>b</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mo>-</mml:mo></mml:mrow></mml:msub></mml:math></inline-formula>; conversely if no spikes are generated (i.e., <inline-formula><mml:math id="M48"><mml:munder class="msub"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:munder><mml:msub><mml:mrow><mml:mi>z</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mn>0</mml:mn></mml:math></inline-formula>), its intrinsic excitability <inline-formula><mml:math id="M49"><mml:mover accent="true"><mml:mrow><mml:mi>b</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:math></inline-formula> increases by <inline-formula><mml:math id="M50"><mml:msub><mml:mrow><mml:mover accent="true"><mml:mrow><mml:mi>b</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mo>&#x0002B;</mml:mo></mml:mrow></mml:msub></mml:math></inline-formula>. Algorithmically, this conditional expression is given by</p>
<disp-formula id="E17"><mml:math id="M51"><mml:mtable columnalign="left"><mml:mtr><mml:mtd><mml:mover accent='true'><mml:mi>b</mml:mi><mml:mo>&#x002DC;</mml:mo></mml:mover><mml:mo>&#x02190;</mml:mo><mml:mrow><mml:mo>{</mml:mo><mml:mrow><mml:mtable columnalign='left'><mml:mtr columnalign='left'><mml:mtd columnalign='left'><mml:mrow><mml:mover accent='true'><mml:mi>b</mml:mi><mml:mo>&#x002DC;</mml:mo></mml:mover><mml:mo>&#x02212;</mml:mo><mml:msub><mml:mover accent='true'><mml:mi>b</mml:mi><mml:mo>&#x002DC;</mml:mo></mml:mover><mml:mo>&#x02212;</mml:mo></mml:msub></mml:mrow></mml:mtd><mml:mtd columnalign='left'><mml:mrow><mml:mtext>if</mml:mtext><mml:mstyle displaystyle='true'><mml:msub><mml:mo>&#x02211;</mml:mo><mml:mi>t</mml:mi></mml:msub><mml:mrow><mml:msub><mml:mi>z</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mo>&#x0003E;</mml:mo><mml:mn>0</mml:mn></mml:mrow></mml:mstyle></mml:mrow></mml:mtd></mml:mtr><mml:mtr columnalign='left'><mml:mtd columnalign='left'><mml:mrow><mml:mover accent='true'><mml:mi>b</mml:mi><mml:mo>&#x002DC;</mml:mo></mml:mover><mml:mo>&#x0002B;</mml:mo><mml:msub><mml:mover accent='true'><mml:mi>b</mml:mi><mml:mo>&#x002DC;</mml:mo></mml:mover><mml:mo>&#x0002B;</mml:mo></mml:msub></mml:mrow></mml:mtd><mml:mtd columnalign='left'><mml:mrow><mml:mtext>otherwise</mml:mtext><mml:mo>.</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>The spatio-temporal spike patterns <bold><italic>x</italic></bold> are classified as described in section 2.4, with an important difference; each post-synaptic spike <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 is weighted by the inverse number of generated post-synaptic spikes, i.e., <inline-formula><mml:math id="M52"><mml:mn>1</mml:mn><mml:mo>/</mml:mo><mml:munder class="msub"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:munder><mml:msub><mml:mrow><mml:mi>z</mml:mi></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub></mml:math></inline-formula>. In other words, a spatio-temporal spike pattern <bold><italic>x</italic></bold> is associated with a group by majority voting of multiple post-synaptic spikes <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1.</p>
</sec>
</sec>
<sec id="s3">
<title>3. Experiments and results</title>
<sec>
<title>3.1. Windows of plasticity</title>
<p>In this study, the temporal relation function <italic>g</italic>(&#x00394;<italic>t</italic>) was expressed as the following Gaussian function:</p>
<disp-formula id="E18"><label>(9)</label><mml:math id="M53"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:mi>g</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>&#x00394;</mml:mi><mml:mi>t</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:msqrt><mml:mrow><mml:mn>2</mml:mn><mml:mi>&#x003C0;</mml:mi></mml:mrow></mml:msqrt><mml:mi>&#x003C3;</mml:mi></mml:mrow></mml:mfrac><mml:mo class="qopname">exp</mml:mo><mml:mrow><mml:mo stretchy="true">(</mml:mo><mml:mrow><mml:mo>-</mml:mo><mml:mfrac><mml:mrow><mml:msup><mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>&#x00394;</mml:mi><mml:mi>t</mml:mi><mml:mo>-</mml:mo><mml:mi>&#x003BC;</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mrow><mml:mrow><mml:mn>2</mml:mn><mml:msup><mml:mrow><mml:mi>&#x003C3;</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:mfrac></mml:mrow><mml:mo stretchy="true">)</mml:mo></mml:mrow><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>Here, the parameters &#x003BC; and &#x003C3; denote the mean and standard deviation of the distribution, respectively. Figure <xref ref-type="fig" rid="F2">2A</xref> plots the temporal relation function <italic>g</italic>(&#x00394;<italic>t</italic>) when &#x003BC; &#x0003D; 1.5 ms and &#x003C3; &#x0003D; 1 ms. The EPSP is commonly modeled by the double-exponential function (e.g., Shouval et al., <xref ref-type="bibr" rid="B49">2010</xref>), which has zero gradient for all &#x00394;<italic>t</italic> &#x0003C; 0. As the proposed SNN is trained according to the gradient of the temporal relation function <italic>g</italic>(&#x00394;<italic>t</italic>), the double-exponential function is unsuitable for the purpose. In this study, the causality was preserved by clamping the temporal relation function <italic>g</italic>(<italic>s</italic> &#x0002B; &#x003C4;<sub><italic>i</italic></sub> &#x02212; <italic>t</italic>) to 0 for all <italic>s</italic> &#x02212; <italic>t</italic> &#x0003C; 0.</p>
<fig id="F2" position="float">
<label>Figure 2</label>
<caption><p><bold>(A)</bold> A temporal relation function <italic>g</italic>(&#x00394;<italic>t</italic>), representing an excitatory post-synaptic potential (EPSP) or current (EPSC), when the pre-synaptic spike is elicited at <italic>s</italic> &#x0003D; 0 ms. The parameters were set to &#x003BC; &#x0003D; 1.5 ms and &#x003C3; &#x0003D; 1.0 ms and the conduction delay &#x003C4;<sub><italic>i</italic></sub> was set to &#x003C4; &#x0003D; 10 ms. <bold>(B&#x02013;D)</bold> The gradients of the lower bound <inline-formula><mml:math id="M55"><mml:mrow><mml:mi mathvariant="-tex-caligraphic">L</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mo>&#x003B8;</mml:mo></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> of the model evidence &#x1D53C;<sub><bold><italic>x</italic></bold>&#x0007E;<bold><italic>X</italic></bold></sub>[log <italic>p</italic>(<bold><italic>x</italic></bold>; <bold>&#x003B8;</bold>)] with respect to the parameters <italic>W</italic><sub><italic>i</italic></sub> and &#x003C4;<sub><italic>i</italic></sub>. The time step &#x003B4; was 0.05 ms, the experimental time period <italic>T</italic> was 50 ms, and the remaining parameters were set to <italic>v</italic> &#x0003D; 10, and &#x003B7; &#x0003D; 0.001. The gradient shapes depend on the shape of the temporal relation function <italic>g</italic>(&#x00394;<italic>t</italic>). <bold>(B)</bold> Temporal window of the delay learning, showing the relationship between the temporal difference &#x00394;<italic>t</italic> and the change &#x00394;&#x003C4;<sub><italic>i</italic></sub> in conduction delay &#x003C4;<sub><italic>i</italic></sub>, where &#x00394;<italic>t</italic> &#x0003D; <italic>s</italic> &#x0002B; &#x003C4;<sub><italic>i</italic></sub> &#x02212; <italic>t</italic>. <bold>(C)</bold> Temporal window of the synaptic plasticity, showing the relationship between the temporal difference &#x00394;<italic>t</italic> and the change &#x00394;<italic>W</italic><sub><italic>i</italic></sub> in synaptic weight <italic>W</italic><sub><italic>i</italic></sub>. <bold>(D)</bold> The change &#x00394;<italic>W</italic><sub><italic>i</italic></sub> in synaptic weight <italic>W</italic><sub><italic>i</italic></sub> versus the current synaptic weight <italic>W</italic><sub><italic>i</italic></sub> for &#x00394;<italic>t</italic> &#x0003D; 1.5 ms (orange) and &#x00394;<italic>t</italic> &#x0003D; -5.0 ms (blue).</p></caption>
<graphic xlink:href="fncom-11-00104-g0002.tif"/>
</fig>
<p>The experimental time period <italic>T</italic> was 50 ms, the time step &#x003B4; was set to 0.05 ms. The remaining parameters were set to &#x003BC; &#x0003D; 1.5 ms, &#x003C3; &#x0003D; 1.0 ms, <italic>v</italic> &#x0003D; 10, and &#x003B7; &#x0003D; 0.001, unless otherwise stated. The gradients of the evidence lower bound <inline-formula><mml:math id="M54"><mml:mrow><mml:mi mathvariant="-tex-caligraphic">L</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mstyle mathvariant="bold-italic"><mml:mo>&#x003B8;</mml:mo></mml:mstyle></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> w.r.t. the conduction delay &#x003C4;<sub><italic>i</italic></sub> and synaptic weight <italic>W</italic><sub><italic>i</italic></sub> were considered to be changes by the delay learning and synaptic plasticity, respectively. When plotted against the temporal difference &#x00394;<italic>t</italic> &#x0003D; <italic>s</italic>&#x0002B;&#x003C4;<sub><italic>i</italic></sub>&#x02212;<italic>t</italic>, the gradients represent the temporal windows of the delay learning and synaptic plasticity. As shown in Figure <xref ref-type="fig" rid="F2">2B</xref>, the conduction delay &#x003C4;<sub><italic>i</italic></sub> increases (decreases) when the temporal difference &#x00394;<italic>t</italic> is smaller (larger) than 1.5 ms. The EPSP peaks at 1.5 ms, indicating that a post-synaptic spike <italic>x</italic><sub><italic>t</italic></sub> &#x0003D; 1 attracts the nearby pre-synaptic spikes <italic>x</italic><sub><italic>is</italic></sub> &#x0003D; 1. The temporal window of the synaptic weight <italic>W</italic><sub><italic>i</italic></sub> is similar to that of the STDP (Figure <xref ref-type="fig" rid="F2">2C</xref>). Specifically, the synapse is potentiated (depressed) when the temporal difference &#x00394;<italic>t</italic> is positive (negative), but is always depressed when the temporal difference &#x00394;<italic>t</italic> is large and positive, as noted in previous electrophysiological studies (Markram et al., <xref ref-type="bibr" rid="B30">1997</xref>; Bi and Poo, <xref ref-type="bibr" rid="B5">1998</xref>) and theoretical studies (Shouval et al., <xref ref-type="bibr" rid="B48">2002</xref>; Wittenberg and Wang, <xref ref-type="bibr" rid="B58">2006</xref>; Shouval et al., <xref ref-type="bibr" rid="B49">2010</xref>). Figure <xref ref-type="fig" rid="F2">2D</xref> plots the amount of synaptic modification &#x00394;<italic>W</italic><sub><italic>i</italic></sub> versus the current synaptic weight <italic>W</italic><sub><italic>i</italic></sub>. The decreasing trend is again consistent with previous electrophysiological (Bi and Poo, <xref ref-type="bibr" rid="B5">1998</xref>) and theoretical (Van Rossum et al., <xref ref-type="bibr" rid="B54">2000</xref>; G&#x000FC;tig et al., <xref ref-type="bibr" rid="B18">2003</xref>; Shouval et al., <xref ref-type="bibr" rid="B49">2010</xref>; Matsubara and Uehara, <xref ref-type="bibr" rid="B34">2016</xref>) studies.</p>
</sec>
<sec>
<title>3.2. Results for toy spike patterns</title>
<p>First, the SNN and the proposed learning algorithm based on the Multinoulli-Bernoulli mixture model were evaluated on toy spike patterns generated by three pre-synaptic neurons. Figure <xref ref-type="fig" rid="F3">3A</xref> shows typical spatio-temporal spike patterns <bold><italic>x</italic></bold>. The pre-synaptic neurons <italic>x</italic><sub>0</sub>, <italic>x</italic><sub>1</sub>, and <italic>x</italic><sub>2</sub> elicited spikes at 1&#x0002B;&#x003BE;<sub>0</sub>, 5&#x0002B;&#x003BE;<sub>1</sub>, and 13&#x0002B;&#x003BE;<sub>2</sub> ms respectively in spike pattern A, and at 13&#x0002B;&#x003BE;<sub>0</sub>, 9&#x0002B;&#x003BE;<sub>1</sub>, and 1&#x0002B;&#x003BE;<sub>2</sub> ms, respectively in spike pattern B. Here, &#x003BE;<sub>0</sub>, &#x003BE;<sub>1</sub>, and &#x003BE;<sub>2</sub> are the noise terms following a uniform distribution <inline-formula><mml:math id="M56"><mml:mrow><mml:mi mathvariant="-tex-caligraphic">U</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mo>-</mml:mo><mml:mn>1</mml:mn><mml:mo>,</mml:mo><mml:mn>1</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula>. Fifty samples from each of spike patterns A and B were extracted as the training dataset; other fifty samples were reserved as the test dataset. All synaptic weights <italic>W</italic><sub><italic>i</italic></sub> were uniformly initialized to 1, and the conduction delays &#x003C4;<sub><italic>i</italic></sub> (ms) were sampled from the uniform distribution <inline-formula><mml:math id="M57"><mml:mrow><mml:mi mathvariant="-tex-caligraphic">U</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mn>5</mml:mn><mml:mo>,</mml:mo><mml:mn>15</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula>. The post-synaptic neuron <italic>z</italic> was repeatedly fed with the spatio-temporal spike patterns <bold><italic>x</italic></bold>, and generated spikes <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1. The gray shaded areas in Figure <xref ref-type="fig" rid="F3">3A</xref> depicts the posterior probability <italic>p</italic>(<italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1|<bold><italic>x</italic></bold>; <bold><italic>W</italic></bold>, <bold>&#x003C4;</bold>) of spiking each millisecond. Figure <xref ref-type="fig" rid="F3">3B</xref> shows the probability distribution of post-synaptic spike timings <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 for the test dataset before learning. The unlearned model cannot discriminate the spike patterns A and B using the post-synaptic spike timings <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1. Note that the proposed SNN is intrinsically probabilistic, meaning that a post-synaptic neuron can elicit a spike at an arbitrary time <italic>t</italic>. During the learning procedure, the synaptic weights <italic>W</italic><sub><italic>i</italic></sub> and the conduction delays &#x003C4;<sub><italic>i</italic></sub> were gradually changed by the proposed learning algorithm (Figure <xref ref-type="fig" rid="F3">3E</xref>). The conduction delays &#x003C4;<sub><italic>i</italic></sub> (ms) were clamped to the range 0&#x02013;20 ms but never reached the limits. Figure <xref ref-type="fig" rid="F3">3F</xref> shows random samples of the post-synaptic spike timings <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1, separated into spike patterns A and B by the decision boundary (the solid black line). After sufficiently many samples, the spike patterns converged into limited temporal ranges and the two groups were clearly demarcated. The post-synaptic spike timings <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 and their distributions after the learning procedure are exemplified in Figures <xref ref-type="fig" rid="F3">3C,D</xref>, respectively. In this trial, the classification accuracy of both the training and test datasets converged to 100 %. The average classification accuracy over 100 trials was 99.6 &#x000B1; 2.6% for the training dataset and 99.6 &#x000B1; 2.5% for the test dataset (Figure <xref ref-type="fig" rid="F3">3G</xref>). The results are summarized in Table <xref ref-type="table" rid="T2">2</xref>.</p>
<fig id="F3" position="float">
<label>Figure 3</label>
<caption><p>Results of SNN based on the Multinoulli-Bernoulli mixture model, evaluated on toy spike patterns. <bold>(A)</bold> Representative spatio-temporal spike patterns A (top panel) and B (bottom panel), and the post-synaptic spikes <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 before the learning procedure. Vertical lines represent pre-synaptic spikes <italic>x</italic><sub><italic>is</italic></sub> &#x0003D; 1 and post-synaptic spikes <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 (the left vertical axes). Dotted lines denote transmission of the pre-synaptic spikes <italic>x</italic><sub><italic>is</italic></sub> &#x0003D; 1 to the post-synaptic neuron <italic>z</italic> with conduction delays &#x003C4;<sub><italic>i</italic></sub>. Gray shaded areas denote the probability of eliciting a post-synaptic spikes <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 per 1 ms (the right vertical axes). <bold>(B)</bold> Test distributions of the post-synaptic spike timings <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 before learning. <bold>(C)</bold> Post-synaptic spike timings <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 after learning, given the spatio-temporal spike patterns <bold><italic>x</italic></bold> in <bold>(A)</bold>. <bold>(D)</bold> Test distributions of the post-synaptic spike timing <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 after learning. <bold>(E)</bold> Trajectories of the synaptic weights <italic>W</italic><sub><italic>i</italic></sub> and conduction delays &#x003C4;<sub><italic>i</italic></sub> during the learning procedure. The darkest lines denote the values of the parameters <italic>W</italic><sub>0</sub> and &#x003C4;<sub>0</sub>, and the lighter lines denote <italic>W</italic><sub><italic>i</italic></sub> and &#x003C4;<sub><italic>i</italic></sub> for <italic>i</italic> &#x0003D; 1, 2, &#x02026;&#x000A0;. <bold>(F)</bold> Random samples of the post-synaptic spike timings <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1. The black solid line marks the decision boundary between spike patterns A and B. <bold>(G)</bold> Classification accuracy in the training and test datasets, each averaged over 100 trials.</p></caption>
<graphic xlink:href="fncom-11-00104-g0003.tif"/>
</fig>
<table-wrap position="float" id="T2">
<label>Table 2</label>
<caption><p>Classification accuracy of the proposed SNN based on the Multinoulli-Bernoulli mixture model.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th valign="top" align="left"><bold>Dataset</bold></th>
<th valign="top" align="center" colspan="2" style="border-bottom: thin solid #000000;"><bold>Unsupervised</bold></th>
<th valign="top" align="center" colspan="2" style="border-bottom: thin solid #000000;"><bold>Supervised</bold></th>
</tr>
<tr>
<th/>
<th valign="top" align="center"><bold>Training</bold></th>
<th valign="top" align="center"><bold>Test</bold></th>
<th valign="top" align="center"><bold>Training</bold></th>
<th valign="top" align="center"><bold>Test</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Toy spike patterns</td>
<td valign="top" align="center">99.6 &#x000B1; 2.6</td>
<td valign="top" align="center">99.6 &#x000B1; 2.5</td>
<td valign="top" align="center">100.0 &#x000B1; 0.0</td>
<td valign="top" align="center">99.8 &#x000B1; 1.2</td>
</tr>
<tr>
<td valign="top" align="left">Iris flower dataset</td>
<td valign="top" align="center">89.5 &#x000B1; 5.3</td>
<td valign="top" align="center">89.5 &#x000B1; 7.5</td>
<td valign="top" align="center">93.9 &#x000B1; 5.3</td>
<td valign="top" align="center">89.4 &#x000B1; 9.1</td>
</tr>
<tr>
<td valign="top" align="left">MNIST dataset</td>
<td valign="top" align="center">88.7 &#x000B1; 3.4</td>
<td valign="top" align="center">88.7 &#x000B1; 3.7</td>
<td valign="top" align="center">90.2 &#x000B1; 4.0</td>
<td valign="top" align="center">90.1 &#x000B1; 4.1</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<p><italic>All the classification accuracies are expressed as ave. &#x000B1; std. (%)</italic>.</p>
</table-wrap-foot>
</table-wrap>
</sec>
<sec>
<title>3.3. Results for the iris flower dataset</title>
<p>Next, the proposed learning algorithm was evaluated on the iris flower dataset (Fisher, <xref ref-type="bibr" rid="B16">1936</xref>), which consists of the data of three species (Setosa, Versicolor, and Virginica). The dataset of each species comprises 50 data points, and each data point consists of four features (the lengths and widths of the sepals and petals). Accordingly, the number <italic>N</italic> of pre-synaptic neurons was set to 4. The spike timing of each feature was normalized to the range 0&#x02013;10 ms. As the sepal length <italic>k</italic><sub>0</sub> ranges from 4.3 to 7.9 cm, the pre-synaptic neuron <italic>x</italic><sub>0</sub> generated a single spike at <inline-formula><mml:math id="M58"><mml:msub><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>0</mml:mn></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mn>10</mml:mn><mml:mo>&#x000D7;</mml:mo><mml:mfrac><mml:mrow><mml:msub><mml:mrow><mml:mi>k</mml:mi></mml:mrow><mml:mrow><mml:mn>0</mml:mn></mml:mrow></mml:msub><mml:mo>-</mml:mo><mml:mn>4</mml:mn><mml:mo>.</mml:mo><mml:mn>3</mml:mn></mml:mrow><mml:mrow><mml:mn>7</mml:mn><mml:mo>.</mml:mo><mml:mn>9</mml:mn><mml:mo>-</mml:mo><mml:mn>4</mml:mn><mml:mo>.</mml:mo><mml:mn>3</mml:mn></mml:mrow></mml:mfrac></mml:math></inline-formula> ms. The generated post-synaptic spikes <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 were then classified into three groups corresponding to the three species. Fifteen randomly chosen data points were assigned as the test set; the remainder were used as the training dataset. The other conditions were those described in section 3.2. The results are summarized in Figure <xref ref-type="fig" rid="F4">4</xref>. The average classification accuracy over 100 trials was 89.5 &#x000B1; 5.3% for the training dataset and 89.5 &#x000B1; 7.5% for the test dataset.</p>
<fig id="F4" position="float">
<label>Figure 4</label>
<caption><p>Results of SNN based on the Multinoulli-Bernoulli mixture model, evaluated on the iris flower dataset (Fisher, <xref ref-type="bibr" rid="B16">1936</xref>). <bold>(A)</bold> Representative spatio-temporal spike patterns of Setosa (top panel), Versicolor (middle panel), and Virginica (bottom panel), and the post-synaptic spikes <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 before learning. Vertical lines represent pre-synaptic spikes <italic>x</italic><sub><italic>is</italic></sub> &#x0003D; 1 and post-synaptic spikes <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 (the left vertical axes). Dotted lines denote transmission of the pre-synaptic spikes <italic>x</italic><sub><italic>is</italic></sub> &#x0003D; 1 to the post-synaptic neuron <italic>z</italic> with conduction delays &#x003C4;<sub><italic>i</italic></sub>. Gray shaded areas denote the probability of eliciting a post-synaptic spikes <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 per 1 ms (the right vertical axes). <bold>(B)</bold> Test distributions of the post-synaptic spike timings <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 before learning. <bold>(C)</bold> Post-synaptic spike timings <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 after learning, given the spatio-temporal spike patterns <bold><italic>x</italic></bold> in <bold>(A)</bold>. <bold>(D)</bold> Test distributions of the post-synaptic spike timing <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 after learning. <bold>(E)</bold> Trajectories of the synaptic weights <italic>W</italic><sub><italic>i</italic></sub> and conduction delays &#x003C4;<sub><italic>i</italic></sub> during the learning procedure. The darkest lines denote the values of the parameters <italic>W</italic><sub>0</sub> and &#x003C4;<sub>0</sub>, and the lighter lines denote <italic>W</italic><sub><italic>i</italic></sub> and &#x003C4;<sub><italic>i</italic></sub> for <italic>i</italic> &#x0003D; 1, 2, &#x02026;&#x000A0;. <bold>(F)</bold> Random samples of the post-synaptic spike timings <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1. The black solid lines mark the decision boundaries between the three species. <bold>(G)</bold> Classification accuracy in the training and test datasets, each averaged over 100 trials.</p></caption>
<graphic xlink:href="fncom-11-00104-g0004.tif"/>
</fig>
<p>For comparison, this subsection introduces the results of the <italic>multilayer ReSuMe</italic> algorithm (Sporea and Gr&#x000FC;ning, <xref ref-type="bibr" rid="B50">2013</xref>), a supervised learning algorithm for multilayer SNNs in rate coding. The multilayer ReSuMe algorithm was also evaluated on the iris flower dataset. In this evaluation, a trial was deemed successful if the classification accuracy exceeded 95% on the training dataset. After weeding out the unsuccessful trials, the proposed learning algorithm and multilayer ReSuMe algorithm achieved a classification accuracy of 94.7 &#x000B1; 6.5 and 94.0 &#x000B1; 0.79% respectively, on the test dataset (see Table <xref ref-type="table" rid="T3">3</xref>). Despite its single-layer architecture, the SNN trained by the proposed learning algorithm classified the iris flower dataset at least as accurately as the multilayer SNN trained by the multilayer ReSuMe algorithm.</p>
<table-wrap position="float" id="T3">
<label>Table 3</label>
<caption><p>Classification accuracies in the iris flower dataset after selection.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th valign="top" align="left"><bold>Model</bold></th>
<th valign="top" align="center" colspan="2" style="border-bottom: thin solid #000000;"><bold>Unsupervised</bold></th>
<th valign="top" align="center" colspan="2" style="border-bottom: thin solid #000000;"><bold>Supervised</bold></th>
</tr>
<tr>
<th/>
<th valign="top" align="center"><bold>Training</bold></th>
<th valign="top" align="center"><bold>Test</bold></th>
<th valign="top" align="center"><bold>Training</bold></th>
<th valign="top" align="center"><bold>Test</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">MB model<xref ref-type="table-fn" rid="TN4"><sup>&#x0002A;</sup></xref></td>
<td valign="top" align="center">96.1 &#x000B1; 0.7</td>
<td valign="top" align="center">94.7 &#x000B1; 6.5</td>
<td valign="top" align="center">96.7 &#x000B1; 1.1</td>
<td valign="top" align="center">92.2 &#x000B1; 7.4</td>
</tr>
<tr>
<td valign="top" align="left">BB model<xref ref-type="table-fn" rid="TN5"><sup>&#x0002A;&#x0002A;</sup></xref></td>
<td valign="top" align="center">96.2 &#x000B1; 0.5</td>
<td valign="top" align="center">93.3 &#x000B1; 5.4</td>
<td valign="top" align="center">96.7 &#x000B1; 0.9</td>
<td valign="top" align="center">94.2 &#x000B1; 5.8</td>
</tr>
<tr>
<td valign="top" align="left">Multilayer ReSuMe<xref ref-type="table-fn" rid="TN6"><sup>&#x0002A;&#x0002A;&#x0002A;</sup></xref></td>
<td valign="top" align="center">&#x02013;</td>
<td valign="top" align="center">&#x02013;</td>
<td valign="top" align="center">96.0 &#x000B1; 0.0</td>
<td valign="top" align="center">94.0 &#x000B1; 0.8</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<fn id="TN4">
<label>&#x0002A;</label>
<p><italic>Proposed SNN based on the Multinoulli-Bernoulli model (section 3.3)</italic>.</p></fn>
<fn id="TN5">
<label>&#x0002A;&#x0002A;</label>
<p><italic>Proposed SNN based on the Bernoulli&#x02013;Bernoulli model (section 3.7)</italic>.</p></fn>
<fn id="TN6">
<label>&#x0002A;&#x0002A;&#x0002A;</label>
<p><italic>Sporea and Gr&#x000FC;ning, <xref ref-type="bibr" rid="B50">2013</xref></italic>.</p></fn>
</table-wrap-foot>
</table-wrap>
</sec>
<sec>
<title>3.4. Results for the MNIST dataset</title>
<p>In this subsection, the proposed learning algorithm was evaluated on the MNIST dataset (LeCun et al., <xref ref-type="bibr" rid="B28">1998</xref>), which contains 28 &#x000D7; 28 grayscale images of 70,000 handwritten digits. In standard rate coding, the pixel intensities are represented by 28 &#x000D7; 28 &#x0003D; 784 pre-synaptic neurons (Nessler et al., <xref ref-type="bibr" rid="B38">2009</xref>; Beyeler et al., <xref ref-type="bibr" rid="B4">2013</xref>; O&#x00027;Connor et al., <xref ref-type="bibr" rid="B39">2013</xref>; Querlioz et al., <xref ref-type="bibr" rid="B43">2013</xref>; Neftci et al., <xref ref-type="bibr" rid="B37">2014</xref>; Diehl and Cook, <xref ref-type="bibr" rid="B11">2015</xref>; Zambrano and Bohte, <xref ref-type="bibr" rid="B60">2016</xref>). To emphasize the temporal coding, this study instead represented the rows and columns of the images by 28 pre-synaptic neurons and their 28 corresponding spike timings, respectively. When a pre-synaptic neuron <italic>x</italic><sub><italic>i</italic></sub> generated a spike at time <italic>s</italic>, the intensity of the pixel in row <italic>i</italic> and column <italic>s</italic> of the image was considered to exceed 0.5 (see Figure <xref ref-type="fig" rid="F5">5A</xref>). For simplicity, only the digits 0 and 8 were included in the analysis. After removing the other digits, 13,728 images were available. This experiment employed a 10-fold cross-validation: 10 % of images were randomly chosen for the test set. The results are summarized in Figure <xref ref-type="fig" rid="F5">5</xref>. This experiment yielded much more pre-synaptic spikes <italic>x</italic><sub><italic>is</italic></sub> &#x0003D; 1 than the previous two experiments. During transmission, the pre-synaptic spikes <italic>x</italic><sub><italic>is</italic></sub> &#x0003D; 1 that largely contributed to the post-synaptic spike <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 were emphasized. More specifically, if &#x003BC;&#x02212;1 &#x0003C; &#x00394;<italic>t</italic> &#x0003D; <italic>s</italic>&#x0002B;&#x003C4;<sub><italic>i</italic></sub>&#x02212;<italic>t</italic> &#x0003C; &#x003BC;&#x0002B;1 for the post-synaptic spike <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1, the pre-synaptic spikes <italic>x</italic><sub><italic>is</italic></sub> &#x0003D; 1 was emphasized with a thick color line, whereas other pre-synaptic spikes <italic>x</italic><sub><italic>is</italic></sub> &#x0003D; 1 were attenuated (see Figures <xref ref-type="fig" rid="F5">5A,C</xref>). According to Figure <xref ref-type="fig" rid="F5">5C</xref>, the C-shaped edges are heightened, indicating that the SNN selectively detects and responds to these edges. In general, the C-shaped edge appears on the left side of digit 0 and in the center of digit 8, so the post-synaptic spikes <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 responded earlier to &#x0201C;0&#x0201D; than to &#x0201C;8&#x0201D;. In some trials, the SNN responded to reversed C-shaped edges (or to edges in digits such as &#x0201C;7&#x0201D;), and the post-synaptic spikes <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 responded earlier to &#x0201C;8&#x0201D; than to &#x0201C;0&#x0201D; (see Figure <xref ref-type="fig" rid="F6">6</xref>). The average classification accuracy over 100 trials was 88.7 &#x000B1; 3.4% for the training dataset and 88.7 &#x000B1; 3.7% for the test dataset.</p>
<fig id="F5" position="float">
<label>Figure 5</label>
<caption><p>Results of SNN based on the Multinoulli-Bernoulli mixture model, evaluated on the hand-written digits 0 and 8 in the MNIST dataset (LeCun et al., <xref ref-type="bibr" rid="B28">1998</xref>). <bold>(A)</bold> Representative spatio-temporal spike patterns of digit 0 (top panel) and digit 8 (bottom panel), and the post-synaptic spikes <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 before learning. Thick colored lines indicate the pre-synaptic spikes <italic>x</italic><sub><italic>is</italic></sub> &#x0003D; 1, elicited within &#x003BC; &#x02212; 1 &#x0003C; &#x00394;<italic>t</italic> &#x0003D; <italic>s</italic> &#x0002B; &#x003C4;<sub><italic>i</italic></sub> &#x02212; <italic>t</italic> &#x0003C; &#x003BC; &#x0002B; 1 for <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1; other pre-synaptic spikes <italic>x</italic><sub><italic>is</italic></sub> are depicted as thin gray lines (the left vertical axes). Gray shaded areas denote the probability of eliciting a post-synaptic spikes <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 per 1 ms (the right vertical axes). <bold>(B)</bold> Test distributions of the post-synaptic spike timings <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 before learning. <bold>(C)</bold> Post-synaptic spike timings <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 after learning, given the spatio-temporal spike patterns <bold><italic>x</italic></bold> in <bold>(A)</bold>. <bold>(D)</bold> Test distributions of the post-synaptic spike timing <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 after learning. <bold>(E)</bold> Trajectories of the synaptic weights <italic>W</italic><sub><italic>i</italic></sub> and conduction delays &#x003C4;<sub><italic>i</italic></sub> during the learning procedure. The darkest lines denote the values of the parameters <italic>W</italic><sub>0</sub> and &#x003C4;<sub>0</sub>, and the lighter lines denote <italic>W</italic><sub><italic>i</italic></sub> and &#x003C4;<sub><italic>i</italic></sub> for <italic>i</italic> &#x0003D; 1, 2, &#x02026;&#x000A0;. <bold>(F)</bold> Random samples of the post-synaptic spike timings <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1. The black solid line marks the decision boundary between spike patterns the hand-written digits 0 and 8. <bold>(G)</bold> Classification accuracy in the training and test datasets, each averaged over 100 trials.</p></caption>
<graphic xlink:href="fncom-11-00104-g0005.tif"/>
</fig>
<fig id="F6" position="float">
<label>Figure 6</label>
<caption><p>Other results on the hand-written digits 0 and 8 in the MNIST dataset (LeCun et al., <xref ref-type="bibr" rid="B28">1998</xref>). <bold>(A)</bold> Representative spatio-temporal spike patterns of digit 0 (top panel) and digit 8 (bottom panel), and the post-synaptic spikes <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 after the learning procedure as per Figure <xref ref-type="fig" rid="F5">5C</xref>. Thick colored lines indicate the pre-synaptic spikes <italic>x</italic><sub><italic>is</italic></sub> &#x0003D; 1, elicited within &#x003BC; &#x02212; 1 &#x0003C; &#x00394;<italic>t</italic> &#x0003D; <italic>s</italic> &#x0002B; &#x003C4;<sub><italic>i</italic></sub> &#x02212; <italic>t</italic> &#x0003C; &#x003BC;&#x0002B;1 for <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1; other pre-synaptic spikes <italic>x</italic><sub><italic>is</italic></sub> are depicted as thin gray lines (the left vertical axes). Gray shaded areas denote the probability of eliciting a post-synaptic spikes <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 per 1 ms (the right vertical axes). <bold>(B)</bold> Test distributions of the post-synaptic spike timing <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 after learning. The SNN selectively responds to the reversed C-shaped edges, so the post-synaptic spikes <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 responding to &#x0201C;8&#x0201D; are earlier than those responding to &#x0201C;0&#x0201D; (in contrast to Figure <xref ref-type="fig" rid="F5">5</xref>).</p></caption>
<graphic xlink:href="fncom-11-00104-g0006.tif"/>
</fig>
</sec>
<sec>
<title>3.5. Results without delay learning</title>
<p>The proposed learning algorithm was also evaluated without delay learning. More specifically, the synaptic weights <italic>W</italic><sub><italic>i</italic></sub> were updated by Equation (5), and the conduction delays &#x003C4;<sub><italic>i</italic></sub> were clamped to their initial values (i.e., were not updated by Equation 4). The other conditions were those described in previous subsections. All of the accuracies were drastically reduced (see Table <xref ref-type="table" rid="T4">4</xref>).</p>
<table-wrap position="float" id="T4">
<label>Table 4</label>
<caption><p>Classification accuracy of the proposed SNN without delay learning.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th valign="top" align="left"><bold>Dataset</bold></th>
<th valign="top" align="center" colspan="2" style="border-bottom: thin solid #000000;"><bold>Unsupervised (Fixed delays)</bold></th>
</tr>
<tr>
<th/>
<th valign="top" align="center"><bold>Training</bold></th>
<th valign="top" align="center"><bold>Test</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Toy spike patterns</td>
<td valign="top" align="center">94.9 &#x000B1; 0.6</td>
<td valign="top" align="center">94.4 &#x000B1; 0.9</td>
</tr>
<tr>
<td valign="top" align="left">Iris flower dataset</td>
<td valign="top" align="center">78.5 &#x000B1; 12.0</td>
<td valign="top" align="center">77.2 &#x000B1; 15.0</td>
</tr>
<tr>
<td valign="top" align="left">MNIST dataset</td>
<td valign="top" align="center">72.9 &#x000B1; 11.8</td>
<td valign="top" align="center">72.9 &#x000B1; 11.8</td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
<sec>
<title>3.6. Results with supervised learning</title>
<p>The proposed learning algorithm is adaptable to supervised learning, in which synaptic plasticity and delay learning are not self-driven by generated post-synaptic spikes, but are driven by external spikes <inline-formula><mml:math id="M59"><mml:msub><mml:mrow><mml:mover accent="true"><mml:mrow><mml:mi>z</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:math></inline-formula> generated at the desired timings (Bohte et al., <xref ref-type="bibr" rid="B6">2002</xref>; Ponulak, <xref ref-type="bibr" rid="B42">2005</xref>; G&#x000FC;tig and Sompolinsky, <xref ref-type="bibr" rid="B19">2006</xref>; Pfister et al., <xref ref-type="bibr" rid="B41">2006</xref>; Paugam-Moisy et al., <xref ref-type="bibr" rid="B40">2008</xref>; Taherkhani et al., <xref ref-type="bibr" rid="B51">2015</xref>; Matsubara and Torikai, <xref ref-type="bibr" rid="B33">2016</xref>). However, the desired timings <inline-formula><mml:math id="M60"><mml:msub><mml:mrow><mml:mover accent="true"><mml:mrow><mml:mi>z</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:math></inline-formula> are generally unknown. In this study, the spatio-temporal spike patterns <bold><italic>x</italic></bold> are classified by whether the post-synaptic spike timing <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 arrives earlier or later than the boundary decision. The group with the latest average post-synaptic spike timing should generate the most delayed post-synaptic spike. Therefore, the &#x0201C;late&#x0201D; group were given an external post-synaptic spike <inline-formula><mml:math id="M61"><mml:msub><mml:mrow><mml:mover accent="true"><mml:mrow><mml:mi>z</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>t</mml:mi><mml:mo>&#x0002B;</mml:mo><mml:mi>d</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:math></inline-formula> (where <italic>d</italic> &#x0003E; 0) after the generated post-synaptic spike <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1. Meanwhile, the &#x0201C;early&#x0201D; group were given a post-synaptic spike <inline-formula><mml:math id="M62"><mml:msub><mml:mrow><mml:mover accent="true"><mml:mrow><mml:mi>z</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>t</mml:mi><mml:mo>-</mml:mo><mml:mi>d</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:math></inline-formula> before the generated post-synaptic spike <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1. These post-synaptic spikes <italic>z</italic><sub><italic>t</italic>&#x0002B;<italic>d</italic></sub> &#x0003D; 1 and <italic>z</italic><sub><italic>t</italic>&#x02212;<italic>d</italic></sub> &#x0003D; 1 tend to delay or advance the generated post-synaptic spike <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1, respectively. In this study, the time parameter <italic>d</italic> was set to &#x003B4; &#x0003D; 0.05 ms. The results of supervised learning are summarized in Tables <xref ref-type="table" rid="T2">2</xref>, <xref ref-type="table" rid="T3">3</xref>. In almost every case, the classification accuracy was improved, confirming that the proposed learning algorithm can be adapted to supervised learning without manually adjusting the spike timings.</p>
</sec>
<sec>
<title>3.7. Results with multiple post-synaptic spikes</title>
<p>This subsection evaluates the SNN based on the Bernoulli&#x02013;Bernoulli mixture model described section 2.5. The homeostatic plasticity parameters were set to <italic>b</italic><sub>&#x0002B;</sub> &#x0003D; 0.01 and <italic>b</italic><sub>&#x02212;</sub> &#x0003D; 0.0001, indicating that, on average, one in every hundred spatio-temporal spike patterns induces no post-synaptic spike <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1. To determine the decision boundaries, the absence of the post-synaptic spike <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 was considered to indicate an exceptionally late post-synaptic spike (e.g., <italic>z</italic><sub><italic>T</italic></sub> &#x0003D; 1). As shown in Tables <xref ref-type="table" rid="T3">3</xref>, <xref ref-type="table" rid="T5">5</xref>, the classification accuracy was improved in the iris flower dataset but degraded in the MNIST dataset. As an example, Figure <xref ref-type="fig" rid="F7">7</xref> shows the results of applying multiple post-synaptic spikes in the MNIST dataset.</p>
<table-wrap position="float" id="T5">
<label>Table 5</label>
<caption><p>Classification accuracy of the proposed SNN based on thle Bernoulli&#x02013;Bernoulli mixture model.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th valign="top" align="left"><bold>Dataset</bold></th>
<th valign="top" align="center" colspan="2" style="border-bottom: thin solid #000000;"><bold>Unsupervised</bold></th>
<th valign="top" align="center" colspan="2" style="border-bottom: thin solid #000000;"><bold>Supervised</bold></th>
</tr>
<tr>
<th/>
<th valign="top" align="center"><bold>Training</bold></th>
<th valign="top" align="center"><bold>Test</bold></th>
<th valign="top" align="center"><bold>Training</bold></th>
<th valign="top" align="center"><bold>Test</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Toy spike patterns</td>
<td valign="top" align="center">98.9 &#x000B1; 3.0</td>
<td valign="top" align="center">98.9 &#x000B1; 2.8</td>
<td valign="top" align="center">99.3 &#x000B1; 2.3</td>
<td valign="top" align="center">98.1 &#x000B1; 5.7</td>
</tr>
<tr>
<td valign="top" align="left">Iris flower dataset</td>
<td valign="top" align="center">90.1 &#x000B1; 5.1</td>
<td valign="top" align="center">90.7 &#x000B1; 7.8</td>
<td valign="top" align="center">93.1 &#x000B1; 4.7</td>
<td valign="top" align="center">90.8 &#x000B1; 7.7</td>
</tr>
<tr>
<td valign="top" align="left">MNIST dataset</td>
<td valign="top" align="center">83.7 &#x000B1; 1.2</td>
<td valign="top" align="center">83.3 &#x000B1; 1.5</td>
<td valign="top" align="center">84.8 &#x000B1; 1.7</td>
<td valign="top" align="center">84.4 &#x000B1; 1.8</td>
</tr>
</tbody>
</table>
</table-wrap>
<fig id="F7" position="float">
<label>Figure 7</label>
<caption><p>Results of SNN based on the Bernoulli&#x02013;Bernoulli mixture model for the hand-written digits 0 and 8 in the MNIST dataset (LeCun et al., <xref ref-type="bibr" rid="B28">1998</xref>). <bold>(A)</bold> Representative spatio-temporal spike patterns of digit 0 (top panel) and digit 8 (bottom panel), and the post-synaptic spikes <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 after the learning procedure as per Figure <xref ref-type="fig" rid="F5">5C</xref>. Thick colored lines indicate the pre-synaptic spikes <italic>x</italic><sub><italic>is</italic></sub> &#x0003D; 1, elicited within &#x003BC; &#x02212; 1 &#x0003C; &#x00394;<italic>t</italic> &#x0003D; <italic>s</italic> &#x0002B; &#x003C4;<sub><italic>i</italic></sub> &#x02212; <italic>t</italic> &#x0003C; &#x003BC;&#x0002B;1 for <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1; other pre-synaptic spikes <italic>x</italic><sub><italic>is</italic></sub> are depicted as thin gray lines (the left vertical axes). Gray shaded areas denote the probability of eliciting a post-synaptic spikes <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 per 1 ms (the right vertical axes). <bold>(B)</bold> Test distribution of mean timings of the post-synaptic spikes <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 after learning. Like the Multinoulli-Bernoulli mixture model (see Figure <xref ref-type="fig" rid="F5">5</xref>), the SNN selectively responds to the C-shaped edges.</p></caption>
<graphic xlink:href="fncom-11-00104-g0007.tif"/>
</fig>
</sec>
</sec>
<sec sec-type="discussion" id="s4">
<title>4. Discussion</title>
<p>Rate-coding SNNs have already succeeded in various unsupervised and supervised learning tasks (Brader et al., <xref ref-type="bibr" rid="B7">2007</xref>; Nessler et al., <xref ref-type="bibr" rid="B38">2009</xref>; Beyeler et al., <xref ref-type="bibr" rid="B4">2013</xref>; O&#x00027;Connor et al., <xref ref-type="bibr" rid="B39">2013</xref>; Diehl and Cook, <xref ref-type="bibr" rid="B11">2015</xref>; Zambrano and Bohte, <xref ref-type="bibr" rid="B60">2016</xref>), and have provided the opportunity for efficient computational architectures (Querlioz et al., <xref ref-type="bibr" rid="B43">2013</xref>; Neftci et al., <xref ref-type="bibr" rid="B37">2014</xref>; Cao et al., <xref ref-type="bibr" rid="B9">2015</xref>). However, rate coding requires repeated sampling of the generated spikes, which increases the computational time (VanRullen and Thorpe, <xref ref-type="bibr" rid="B55">2001</xref>). The alternative approach is temporal coding, which encodes the information into a spatio-temporal spike pattern. Temporal coding requires the timing of at least one spike (Bohte et al., <xref ref-type="bibr" rid="B6">2002</xref>; Ponulak, <xref ref-type="bibr" rid="B42">2005</xref>; G&#x000FC;tig and Sompolinsky, <xref ref-type="bibr" rid="B19">2006</xref>; Pfister et al., <xref ref-type="bibr" rid="B41">2006</xref>; Paugam-Moisy et al., <xref ref-type="bibr" rid="B40">2008</xref>; Yu et al., <xref ref-type="bibr" rid="B59">2014</xref>; Taherkhani et al., <xref ref-type="bibr" rid="B51">2015</xref>), potentially enabling more efficient computational architectures than those based on rate coding (Matsubara and Torikai, <xref ref-type="bibr" rid="B32">2013</xref>, <xref ref-type="bibr" rid="B33">2016</xref>). These studies computed gradients of spike timing or excitatory post-synaptic potential with respect to synaptic weight and/or conduction delay, and adjusted to them to elicit post-synaptic spikes at the desired timings. However, they focused exclusively on supervised learning algorithms, which always require teacher spikes at the desired timings. In general, the desired timings are unknown and require careful manual adjustment. Therefore, this approach limits the flexibility in the context of machine learning and poorly represents biological systems that self-adapt to changing environments. Unsupervised delay learning algorithms have received far less attention than supervised learning. The few studies published in this area have not approached practical tasks such as classification and reproduction of given spike patterns (H&#x000FC;ning et al., <xref ref-type="bibr" rid="B21">1998</xref>; Eurich et al., <xref ref-type="bibr" rid="B12">1999</xref>, <xref ref-type="bibr" rid="B13">2000</xref>). In contrast, the proposed unsupervised learning algorithm approximates the EM algorithm (see section 2), and can therefore classify the given spatio-temporal spike patterns by inferring their hidden causes. Under the proposed learning algorithm, the SNN also detected frequent patterns in the given spatio-temporal spike patterns. In addition, the proposed learning algorithm can be adapted to supervised learning without the application of externally determined spike timings.</p>
<p>The proposed learning algorithm relates the synaptic weight modification &#x00394;<italic>W</italic> to the temporal difference &#x00394;<italic>t</italic> and the current synaptic weight <italic>W</italic> (see section 3.1). When the temporal difference &#x00394;<italic>t</italic> is positive, the synapse is potentiated; when &#x00394;<italic>t</italic> is negative or largely positive, it is depressed. This relationship is consistent with previous electrophysiological studies (Markram et al., <xref ref-type="bibr" rid="B30">1997</xref>; Bi and Poo, <xref ref-type="bibr" rid="B5">1998</xref>) and has been discussed in theoretical studies (Shouval et al., <xref ref-type="bibr" rid="B48">2002</xref>; Wittenberg and Wang, <xref ref-type="bibr" rid="B58">2006</xref>; Shouval et al., <xref ref-type="bibr" rid="B49">2010</xref>). The amount of synaptic modification &#x00394;<italic>W</italic><sub><italic>i</italic></sub> decreases with increasing current synaptic weight <italic>W</italic><sub><italic>i</italic></sub>. This relationship is also consistent with previous electrophysiological study (Bi and Poo, <xref ref-type="bibr" rid="B5">1998</xref>) and is considered to maintain neuronal activity by suppressing excessive potentiation (Van Rossum et al., <xref ref-type="bibr" rid="B54">2000</xref>; G&#x000FC;tig et al., <xref ref-type="bibr" rid="B18">2003</xref>; Shouval et al., <xref ref-type="bibr" rid="B49">2010</xref>; Matsubara and Uehara, <xref ref-type="bibr" rid="B34">2016</xref>). The Results of the iris flower dataset (summarized in Figure <xref ref-type="fig" rid="F4">4</xref>) demonstrate that during the early learning phase, the pre-synaptic spikes are not simultaneously delivered to the post-synaptic neuron, and the synaptic weights and conduction delays change only gradually. After 50,000 samples, the conduction delays rapidly changed and almost converged to certain values before 60,000 samples. At that time, the pre-synaptic spikes arrived at the post-synaptic neuron simultaneously. Following the changing conduction delays, the synaptic weights increase, post-synaptic spikes become clustered, and the classification accuracy becomes higher (see Figures <xref ref-type="fig" rid="F4">4E&#x02013;G</xref>). Similar time courses were observed for the toy spike patterns (see Figure <xref ref-type="fig" rid="F3">3</xref>). These results support Fields (<xref ref-type="bibr" rid="B15">2015</xref>) and Baraban et al. (<xref ref-type="bibr" rid="B3">2016</xref>), who hypothesized that conduction delay is adjusted for the synchronous arrival of multiple pre-synaptic spikes before the synaptic modification dominates. They also demonstrated that clustered post-synaptic spikes and accurate classification require the combined optimization of conduction delay and synaptic weight. Before the synaptic weights increase, the classification accuracy is poor. Therefore, delay learning and synaptic plasticity are inherently linked as mentioned by Jamann et al. (<xref ref-type="bibr" rid="B25">2017</xref>). Although, the detailed of biological delay learning (i.e., activity-dependent myelination) remains unclear, this study provides a good hypothetical explanation of the mechanism underlying this process.</p>
<p>With the MNIST dataset, the test accuracies of the proposed algorithm based on the Multinoulli-Bernoulli mixture model were almost equal to the corresponding train accuracies as summarized in Tables <xref ref-type="table" rid="T2">2</xref>, <xref ref-type="table" rid="T4">4</xref>. This could be because the MNIST has a limited number of local minima insensitively to the separation of the test dataset. Since the experiment employed a 10-fold cross-validation, the average test accuracy becomes equal to the average train accuracy if the proposed algorithm always converges to the same local minimum. Actually, in all the 100 trials, the proposed algorithm apparently converged to one of two local minima shown in Figures <xref ref-type="fig" rid="F5">5</xref>, <xref ref-type="fig" rid="F6">6</xref>.</p>
<p>When the conduction delays were fixed, the classification accuracy of the proposed learning algorithm drastically reduced (see Table <xref ref-type="table" rid="T4">4</xref>). This occurred because the classification depends on the timing <italic>t</italic> of the post-synaptic spike <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1, which remained almost static the delay modification. This result demonstrates one advantage of the proposed learning algorithm over other temporal coding learning algorithm such as the ReSuMe algorithm (Ponulak, <xref ref-type="bibr" rid="B42">2005</xref>; Sporea and Gr&#x000FC;ning, <xref ref-type="bibr" rid="B50">2013</xref>) and the Tempotron algorithm (G&#x000FC;tig and Sompolinsky, <xref ref-type="bibr" rid="B19">2006</xref>; Yu et al., <xref ref-type="bibr" rid="B59">2014</xref>), which learn by synaptic modification only. Gerstner et al. (<xref ref-type="bibr" rid="B17">1996</xref>) and Bohte et al. (<xref ref-type="bibr" rid="B6">2002</xref>) assumed multiple paths from a single source to a post-synaptic neuron with various delays. In such a situation, synaptic modification can adjust the post-synaptic spike timing by pruning the inappropriate paths leaving the appropriate ones only. However, an SNN based on this approach either requires numerous unused paths for future development, or its flexibility is limited by the initial network connections and conduction delays. SNNs optimized by delay learning algorithms such as the proposed learning algorithm are more efficient and flexible. A performance comparison between the two learning approaches is outside the scope of this paper, but is a worthwhile future task.</p>
<p>Despite the single-layer architecture, the classification accuracy of the SNN trained by the proposed learning algorithm is comparable to (even slightly superior to) that of multilayer ReSuMe, a supervised learning algorithm designed for multilayer SNNs (Sporea and Gr&#x000FC;ning, <xref ref-type="bibr" rid="B50">2013</xref>). Supervised learning algorithms such as variants of ReSuMe (Ponulak, <xref ref-type="bibr" rid="B42">2005</xref>; Sporea and Gr&#x000FC;ning, <xref ref-type="bibr" rid="B50">2013</xref>; Taherkhani et al., <xref ref-type="bibr" rid="B51">2015</xref>) and Tempotron (G&#x000FC;tig and Sompolinsky, <xref ref-type="bibr" rid="B19">2006</xref>; Yu et al., <xref ref-type="bibr" rid="B59">2014</xref>) require the desired timings of post-synaptic spikes, which must be determined before the learning procedure. As the desired timings are generally unknown, they should be adjusted manually and carefully in classification tasks. On the other hand, the proposed learning algorithm automatically determines the desired timing from the given dataset and the initial parameter values, even during a supervised learning. The removal of the need for external timing is another advantage of the proposed learning algorithm.</p>
<p>In temporal coding, the proposed learning algorithm requires fewer parameters and fewer spikes than existing learning algorithms in rate coding. Single-layer SNN in rate coding required 4 pre-synaptic and 3 post-synaptic neurons with 15 parameters (4 weight parameters and a bias parameter for each group) for the iris flower dataset, and more than 784 neurons and parameters for 2 digits in the MNIST dataset (e.g., Nessler et al., <xref ref-type="bibr" rid="B38">2009</xref>; Neftci et al., <xref ref-type="bibr" rid="B37">2014</xref>). In these algorithms, each pre-synaptic neuron corresponds to a feature or a pixel, and each post-synaptic neuron corresponds to a group. As a single neuron represents multiple values by its spike timing in temporal coding, the proposed neural network requires 4 pre-synaptic neurons and 1 post-synaptic neuron with 10 parameters (4 synaptic weights <italic>W</italic><sub><italic>i</italic></sub>, 4 conduction delays &#x003C4;<sub><italic>i</italic></sub>, and 2 decision boundaries) for the iris flower dataset, and only 29 neurons and 57 parameters for 2 digits in the MNIST dataset. In addition, each feature is the iris flower dataset was encoded into a single spike timing with in the range 0&#x02013;10 ms with a time step of &#x003B4; &#x0003D; 0.05 ms. Therefore, the resolution of each feature was 200 stages. To achieve the same resolution in rate coding, the pre-synaptic spikes must be generated up to 200 times. The proposed learning algorithm requires far fewer pre-synaptic spikes. In future works, rate coding in the proposed and existing learning algorithms should be compared with similar numbers of parameters and spikes. Note that, as the proposed SNNs have only one post-synaptic neuron, they can classify two or three groups at most. To classify more groups, the proposed SNN must be generalized to multiple interacting post-synaptic neurons.</p>
<p>This study proposed two types of SNNs; one based on the Multinoulli-Bernoulli mixture model, the other on the Bernoulli&#x02013;Bernoulli mixture model. The former takes a single post-synaptic spike from the temporal distribution, whereas the latter can independently elicit a post-synaptic spike at every time step and exhibits a burst-like behavior. The SNN based on the Bernoulli&#x02013;Bernoulli mixture model classified the iris flower dataset more accurately, and the MNIST dataset less accurately, than the SNN based on the Multinoulli-Bernoulli mixture model. Owing to the multiple generations of post-synaptic spikes, there was little variation among trials in the iris flower dataset, and the learning proceeded more robustly than in the SNN based on the Multinoulli-Bernoulli mixture model. Consequently, the accuracy was improved in this dataset. Conversely, the number of pre-synaptic spikes representing a single image varied widely in the MNIST dataset, generating a widely varying number of post-synaptic spikes. The SNN based on the Bernoulli&#x02013;Bernoulli mixture model sometimes detected too many edges similar to the template edge in various sub-regions, and sometimes detected no edge. Meanwhile, the SNN based on the Multinoulli-Bernoulli mixture model always detected the single edge that best fitted the template edge. Therefore, the learning of the SNN based on the Bernoulli&#x02013;Bernoulli mixture model proceeded less robustly, lowering the classification accuracy of the MNIST dataset.</p>
<p>The SNN proposed in section 2.1 can also be implemented in continuous time. However, the number of pre-synaptic spikes <italic>x</italic><sub><italic>is</italic></sub> &#x0003D; 1 is unlimited in this mode, and the joint probability <italic>p</italic>(<bold><italic>x</italic></bold>, <bold><italic>z</italic></bold>) of the Multinoulli-Bernoulli and Bernoulli&#x02013;Bernoulli mixture models is difficult to define. For this reason, the SNN was implemented in discrete time. The EPSP of the proposed SNN is the linear sum of the EPSPs induced by multiple pre-synaptic spikes <italic>x</italic><sub><italic>is</italic></sub> &#x0003D; 1, and the proposed learning algorithms in sections 2.3 and 2.5 were normalized by the time step &#x003B4;. Hence, the proposed SNN is robust to the time step &#x003B4; and can be approximated to continuous time as the time step &#x003B4; approaches 0.</p>
<p>Note that Equation (4) and the first term in Equation (5) are independent of the time step &#x003B4; but the second term in Equation (5) does depend on &#x003B4;. When adjusting the time step &#x003B4;,</p>
<p>In contrast, the behavior of dynamical spiking neurons described by an ordinary differential equation is sensitive to the time step and the numerical simulation algorithm (Hansel et al., <xref ref-type="bibr" rid="B20">1998</xref>). A continuous-time SNN and its corresponding generative model will be explored in future work.</p>
<p>The Bernoulli&#x02013;Bernoulli mixture model independently draws a post-synaptic spike <italic>z</italic><sub><italic>t</italic></sub> &#x0003D; 1 at each time step. However, biological studies have confirmed that the intervals between the generated spikes do not follow a Poisson distribution, implying that the spikes in biological neural networks are not independently generated (Burns and Webb, <xref ref-type="bibr" rid="B8">1976</xref>; Levine, <xref ref-type="bibr" rid="B29">1991</xref>). When a dynamical neuron generates a spike, it enters the relative refractory period. During this time, the neuron becomes less sensitive to stimuli and is less likely to generate a spike. After the relative refractory period, the neuron returns to its resting state and resumes its usual chance of generating a spike (Izhikevich, <xref ref-type="bibr" rid="B22">2006a</xref>, Chapter 8). Incorporated with these dynamics, the SNN based on the Bernoulli&#x02013;Bernoulli mixture model would be purged of its burst-like behavior in the MNIST dataset (which degraded the classification accuracy), and would be rendered more computationally efficient and more biologically realistic. A dynamic version of the proposed SNN is a further opportunity for future work.</p>
</sec>
<sec sec-type="conclusions" id="s5">
<title>5. Conclusion</title>
<p>This study proposed an unsupervised learning algorithm that adjusts the conduction delays and synaptic weights in an SNN. The proposed learning algorithm approximates the Expectation-Maximization (EM) algorithm and was confirmed to classify the spatio-temporal spike patterns in several practical problems. In addition, the proposed learning algorithm is adjustable to supervised learning, which improves its classification accuracy. The formulation of the proposed learning algorithm is partially consistent with the synaptic plasticity demonstrated in previous biological and theoretical studies. Therefore, the proposed learning algorithm is a strong candidate model of biological delay learning and will contribute to further investigations of SNNs in temporal coding.</p>
</sec>
<sec id="s6">
<title>Author contributions</title>
<p>The author confirms being the sole contributor of this work and approved it for publication.</p>
<sec>
<title>Conflict of interest statement</title>
<p>The author declares that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
</sec>
</body>
<back>
<ack><p>The author would like to thank Prof. Kuniaki Uehara at Kobe University, and Dr. Ferdinand Peper and Dr. Hiromasa Takemura at CiNet for valuable discussions.</p>
</ack>
<ref-list>
<title>References</title>
<ref id="B1">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Abraham</surname> <given-names>W. C.</given-names></name></person-group> (<year>2008</year>). <article-title>Metaplasticity: tuning synapses and networks for plasticity</article-title>. <source>Nat. Rev. Neurosci.</source> <volume>9</volume>:<fpage>387</fpage>. <pub-id pub-id-type="doi">10.1038/nrn2356</pub-id><pub-id pub-id-type="pmid">18401345</pub-id></citation>
</ref>
<ref id="B2">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bakkum</surname> <given-names>D. J.</given-names></name> <name><surname>Chao</surname> <given-names>Z. C.</given-names></name> <name><surname>Potter</surname> <given-names>S. M.</given-names></name></person-group> (<year>2008</year>). <article-title>Long-term activity-dependent plasticity of action potential propagation delay and amplitude in cortical networks</article-title>. <source>PLOS ONE</source> <volume>3</volume>:<fpage>e2088</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pone.0002088</pub-id><pub-id pub-id-type="pmid">18461127</pub-id></citation>
</ref>
<ref id="B3">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Baraban</surname> <given-names>M.</given-names></name> <name><surname>Mensch</surname> <given-names>S.</given-names></name> <name><surname>Lyons</surname> <given-names>D. A.</given-names></name></person-group> (<year>2016</year>). <article-title>Adaptive myelination from fish to man</article-title>. <source>Brain Res.</source> <volume>1641</volume>, <fpage>149</fpage>&#x02013;<lpage>161</lpage>. <pub-id pub-id-type="doi">10.1016/j.brainres.2015.10.026</pub-id><pub-id pub-id-type="pmid">26498877</pub-id></citation>
</ref>
<ref id="B4">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Beyeler</surname> <given-names>M.</given-names></name> <name><surname>Dutt</surname> <given-names>N. D.</given-names></name> <name><surname>Krichmar</surname> <given-names>J. L.</given-names></name></person-group> (<year>2013</year>). <article-title>Categorization and decision-making in a neurobiologically plausible spiking network using a STDP-like learning rule</article-title>. <source>Neural Netw.</source> <volume>48</volume>, <fpage>109</fpage>&#x02013;<lpage>124</lpage>. <pub-id pub-id-type="doi">10.1016/j.neunet.2013.07.012</pub-id><pub-id pub-id-type="pmid">23994510</pub-id></citation>
</ref>
<ref id="B5">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bi</surname> <given-names>G.-Q.</given-names></name> <name><surname>Poo</surname> <given-names>M.-M.</given-names></name></person-group> (<year>1998</year>). <article-title>Synaptic modifications in cultured hippocampal neurons: dependence on spike timing, synaptic strength, and postsynaptic cell type</article-title>. <source>J. Neurosci.</source> <volume>18</volume>, <fpage>10464</fpage>&#x02013;<lpage>10472</lpage>. <pub-id pub-id-type="pmid">9852584</pub-id></citation>
</ref>
<ref id="B6">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bohte</surname> <given-names>S. M.</given-names></name> <name><surname>Kok</surname> <given-names>J. N.</given-names></name> <name><surname>La Poutr&#x000E9;</surname> <given-names>H.</given-names></name></person-group> (<year>2002</year>). <article-title>Error-backpropagation in temporally encoded networks of spiking neurons</article-title>. <source>Neurocomputing</source> <volume>48</volume>, <fpage>17</fpage>&#x02013;<lpage>37</lpage>. <pub-id pub-id-type="doi">10.1016/S0925-2312(01)00658-0</pub-id></citation>
</ref>
<ref id="B7">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Brader</surname> <given-names>J. M.</given-names></name> <name><surname>Senn</surname> <given-names>W.</given-names></name> <name><surname>Fusi</surname> <given-names>S.</given-names></name></person-group> (<year>2007</year>). <article-title>Learning real-world stimuli in a neural network with spike-driven synaptic dynamics</article-title>. <source>Neural Comput.</source> <volume>19</volume>, <fpage>2881</fpage>&#x02013;<lpage>2912</lpage>. <pub-id pub-id-type="doi">10.1162/neco.2007.19.11.2881</pub-id><pub-id pub-id-type="pmid">17883345</pub-id></citation>
</ref>
<ref id="B8">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Burns</surname> <given-names>B. D.</given-names></name> <name><surname>Webb</surname> <given-names>A. C.</given-names></name></person-group> (<year>1976</year>). <article-title>The spontaneous activity of neurones in the cat&#x00027;s cerebral cortex</article-title>. <source>Proc. R. Soc. Lond. Ser. B Biol. Sci.</source> <volume>194</volume>, <fpage>211</fpage>&#x02013;<lpage>223</lpage>. <pub-id pub-id-type="doi">10.1098/rspb.1976.0074</pub-id><pub-id pub-id-type="pmid">11486</pub-id></citation>
</ref>
<ref id="B9">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cao</surname> <given-names>Y.</given-names></name> <name><surname>Chen</surname> <given-names>Y.</given-names></name> <name><surname>Khosla</surname> <given-names>D.</given-names></name></person-group> (<year>2015</year>). <article-title>Spiking deep convolutional neural networks for energy-efficient object recognition</article-title>. <source>Int. J. Comput. Vis.</source> <volume>113</volume>, <fpage>54</fpage>&#x02013;<lpage>66</lpage>. <pub-id pub-id-type="doi">10.1007/s11263-014-0788-3</pub-id></citation>
</ref>
<ref id="B10">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Carr</surname> <given-names>C. E.</given-names></name> <name><surname>Konishi</surname> <given-names>M.</given-names></name></person-group> (<year>1990</year>). <article-title>A circuit for detection of interaural time differences in the brain stem of the barn owl</article-title>. <source>J. Neurosci.</source> <volume>10</volume>, <fpage>3227</fpage>&#x02013;<lpage>3246</lpage>. <pub-id pub-id-type="pmid">2213141</pub-id></citation>
</ref>
<ref id="B11">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Diehl</surname> <given-names>P. U.</given-names></name> <name><surname>Cook</surname> <given-names>M.</given-names></name></person-group> (<year>2015</year>). <article-title>Unsupervised learning of digit recognition using spike-timing-dependent plasticity</article-title>. <source>Front. Comput. Neurosci.</source> <volume>52538</volume>, <fpage>1</fpage>&#x02013;<lpage>6</lpage>. <pub-id pub-id-type="doi">10.3389/fncom.2015.00099</pub-id></citation>
</ref>
<ref id="B12">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Eurich</surname> <given-names>C. W.</given-names></name> <name><surname>Pawelzik</surname> <given-names>K.</given-names></name> <name><surname>Ernst</surname> <given-names>U.</given-names></name> <name><surname>Cowan</surname> <given-names>J. D.</given-names></name> <name><surname>Milton</surname> <given-names>J. G.</given-names></name></person-group> (<year>1999</year>). <article-title>Dynamics of self-organized delay adaptation</article-title>. <source>Phys. Rev. Lett.</source> <volume>82</volume>, <fpage>1594</fpage>&#x02013;<lpage>1597</lpage>. <pub-id pub-id-type="doi">10.1103/PhysRevLett.82.1594</pub-id></citation>
</ref>
<ref id="B13">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Eurich</surname> <given-names>C. W.</given-names></name> <name><surname>Pawelzik</surname> <given-names>K.</given-names></name> <name><surname>Ernst</surname> <given-names>U.</given-names></name> <name><surname>Thiel</surname> <given-names>A.</given-names></name> <name><surname>Cowan</surname> <given-names>J. D.</given-names></name> <name><surname>Milton</surname> <given-names>J. G.</given-names></name></person-group> (<year>2000</year>). <article-title>Delay adaptation in the nervous system</article-title>. <source>Neurocomputing</source> 32&#x02013;<volume>33</volume>, <fpage>741</fpage>&#x02013;<lpage>748</lpage>. <pub-id pub-id-type="doi">10.1016/S0925-2312(00)00239-3</pub-id></citation>
</ref>
<ref id="B14">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fields</surname> <given-names>R. D.</given-names></name></person-group> (<year>2005</year>). <article-title>Myelination: an overlooked mechanism of synaptic plasticity?</article-title> <source>Neuroscientist</source> <volume>11</volume>, <fpage>528</fpage>&#x02013;<lpage>531</lpage>. <pub-id pub-id-type="doi">10.1177/1073858405282304</pub-id><pub-id pub-id-type="pmid">16282593</pub-id></citation>
</ref>
<ref id="B15">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fields</surname> <given-names>R. D.</given-names></name></person-group> (<year>2015</year>). <article-title>A new mechanism of nervous system plasticity: activity-dependent myelination</article-title>. <source>Nat. Rev. Neurosci.</source> <volume>16</volume>, <fpage>756</fpage>&#x02013;<lpage>767</lpage>. <pub-id pub-id-type="doi">10.1038/nrn4023</pub-id><pub-id pub-id-type="pmid">26585800</pub-id></citation>
</ref>
<ref id="B16">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fisher</surname> <given-names>R. A.</given-names></name></person-group> (<year>1936</year>). <article-title>The use of multiple measurements in taxonomic problems</article-title>. <source>Ann. Eugen.</source> <volume>7</volume>, <fpage>179</fpage>&#x02013;<lpage>188</lpage>. <pub-id pub-id-type="doi">10.1017/CBO9781107415324.004</pub-id></citation>
</ref>
<ref id="B17">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gerstner</surname> <given-names>W.</given-names></name> <name><surname>Kempter</surname> <given-names>R.</given-names></name> <name><surname>van Hemmen</surname> <given-names>J. L.</given-names></name> <name><surname>Wagner</surname> <given-names>H.</given-names></name></person-group> (<year>1996</year>). <article-title>A neuronal learning rule for sub-millisecond temporal coding</article-title>. <source>Nature</source> <volume>383</volume>, <fpage>76</fpage>&#x02013;<lpage>78</lpage>. <pub-id pub-id-type="doi">10.1038/383076a0</pub-id><pub-id pub-id-type="pmid">8779718</pub-id></citation>
</ref>
<ref id="B18">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>G&#x000FC;tig</surname> <given-names>R.</given-names></name> <name><surname>Aharonov</surname> <given-names>R.</given-names></name> <name><surname>Rotter</surname> <given-names>S.</given-names></name> <name><surname>Sompolinsky</surname> <given-names>H.</given-names></name></person-group> (<year>2003</year>). <article-title>Learning input correlations through nonlinear temporally asymmetric Hebbian plasticity</article-title>. <source>J. Neurosci.</source> <volume>23</volume>, <fpage>3697</fpage>&#x02013;<lpage>3714</lpage>. <pub-id pub-id-type="pmid">12736341</pub-id></citation>
</ref>
<ref id="B19">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>G&#x000FC;tig</surname> <given-names>R.</given-names></name> <name><surname>Sompolinsky</surname> <given-names>H.</given-names></name></person-group> (<year>2006</year>). <article-title>The tempotron: a neuron that learns spike timing-based decisions</article-title>. <source>Nat. Neurosci.</source> <volume>9</volume>, <fpage>420</fpage>&#x02013;<lpage>428</lpage>. <pub-id pub-id-type="doi">10.1038/nn1643</pub-id><pub-id pub-id-type="pmid">16474393</pub-id></citation>
</ref>
<ref id="B20">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hansel</surname> <given-names>D.</given-names></name> <name><surname>Mato</surname> <given-names>G.</given-names></name> <name><surname>Meunier</surname> <given-names>C.</given-names></name> <name><surname>Neltner</surname> <given-names>L.</given-names></name></person-group> (<year>1998</year>). <article-title>On numerical simulations of integrate-and-fire neural networks</article-title>. <source>Neural Comput.</source> <volume>10</volume>, <fpage>467</fpage>&#x02013;<lpage>483</lpage>. <pub-id pub-id-type="doi">10.1162/089976698300017845</pub-id><pub-id pub-id-type="pmid">9472491</pub-id></citation>
</ref>
<ref id="B21">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>H&#x000FC;ning</surname> <given-names>H.</given-names></name> <name><surname>Gl&#x000FC;nder</surname> <given-names>H.</given-names></name> <name><surname>Palm</surname> <given-names>G.</given-names></name></person-group> (<year>1998</year>). <article-title>Synaptic delay learning in pulse-coupled neurons</article-title>. <source>Neural Comput.</source> <volume>10</volume>, <fpage>555</fpage>&#x02013;<lpage>565</lpage>. <pub-id pub-id-type="doi">10.1162/089976698300017665</pub-id><pub-id pub-id-type="pmid">9527834</pub-id></citation>
</ref>
<ref id="B22">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Izhikevich</surname> <given-names>E. M.</given-names></name></person-group> (<year>2006a</year>). <source>Dynamical Systems in Neuroscience: The Geometry of Excitability and Bursting</source>. <publisher-name>The MIT Press</publisher-name>.</citation>
</ref>
<ref id="B23">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Izhikevich</surname> <given-names>E. M.</given-names></name></person-group> (<year>2006b</year>). <article-title>Polychronization: computation with spikes</article-title>. <source>Neural Comput.</source> <volume>18</volume>, <fpage>245</fpage>&#x02013;<lpage>282</lpage>. <pub-id pub-id-type="doi">10.1162/089976606775093882</pub-id><pub-id pub-id-type="pmid">16378515</pub-id></citation>
</ref>
<ref id="B24">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Izhikevich</surname> <given-names>E. M.</given-names></name> <name><surname>Gally</surname> <given-names>J. A.</given-names></name> <name><surname>Edelman</surname> <given-names>G. M.</given-names></name></person-group> (<year>2004</year>). <article-title>Spike-timing dynamics of neuronal groups</article-title>. <source>Cereb. Cortex</source> <volume>14</volume>, <fpage>933</fpage>&#x02013;<lpage>944</lpage>. <pub-id pub-id-type="doi">10.1093/cercor/bhh053</pub-id><pub-id pub-id-type="pmid">15142958</pub-id></citation>
</ref>
<ref id="B25">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jamann</surname> <given-names>N.</given-names></name> <name><surname>Jordan</surname> <given-names>M.</given-names></name> <name><surname>Engelhardt</surname> <given-names>M.</given-names></name></person-group> (<year>2017</year>). <article-title>Activity-dependent axonal plasticity in sensory systems</article-title>. <source>Neuroscience</source> [Epub ahead of print]. <pub-id pub-id-type="doi">10.1016/j.neuroscience.2017.07.035</pub-id><pub-id pub-id-type="pmid">28739523</pub-id></citation>
</ref>
<ref id="B26">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Kappel</surname> <given-names>D.</given-names></name> <name><surname>Habenschuss</surname> <given-names>S.</given-names></name> <name><surname>Legenstein</surname> <given-names>R.</given-names></name> <name><surname>Maass</surname> <given-names>W.</given-names></name></person-group> (<year>2015</year>). <article-title>Synaptic sampling: a Bayesian approach to neural network plasticity and rewiring</article-title>, in <source>Advances in Neural Information Processing Systems (NIPS)</source> (<publisher-loc>Montr&#x000E9;al, QC</publisher-loc>), <fpage>370</fpage>&#x02013;<lpage>378</lpage>.</citation>
</ref>
<ref id="B27">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kappel</surname> <given-names>D.</given-names></name> <name><surname>Nessler</surname> <given-names>B.</given-names></name> <name><surname>Maass</surname> <given-names>W.</given-names></name></person-group> (<year>2014</year>). <article-title>STDP installs in winner-take-all circuits an online approximation to hidden Markov model learning</article-title>. <source>PLoS Comput. Biol.</source> <volume>10</volume>:<fpage>e1003511</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pcbi.1003511</pub-id><pub-id pub-id-type="pmid">24675787</pub-id></citation>
</ref>
<ref id="B28">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>LeCun</surname> <given-names>Y.</given-names></name> <name><surname>Bottou</surname> <given-names>L.</given-names></name> <name><surname>Bengio</surname> <given-names>Y.</given-names></name> <name><surname>Haffner</surname> <given-names>P.</given-names></name></person-group> (<year>1998</year>). <article-title>Gradient-based learning applied to document recognition</article-title>. <source>Proc. IEEE</source> <volume>86</volume>, <fpage>2278</fpage>&#x02013;<lpage>2323</lpage>. <pub-id pub-id-type="doi">10.1109/5.726791</pub-id></citation>
</ref>
<ref id="B29">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Levine</surname> <given-names>M. W.</given-names></name></person-group> (<year>1991</year>). <article-title>The distribution of the intervals between neural impulses in the maintained discharges of retinal ganglion cells</article-title>. <source>Biol. Cybern.</source> <volume>65</volume>, <fpage>459</fpage>&#x02013;<lpage>467</lpage>. <pub-id pub-id-type="pmid">1958731</pub-id></citation>
</ref>
<ref id="B30">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Markram</surname> <given-names>H.</given-names></name> <name><surname>L&#x000FC;bke</surname> <given-names>J.</given-names></name> <name><surname>Frotscher</surname> <given-names>M.</given-names></name> <name><surname>Sakmann</surname> <given-names>B.</given-names></name></person-group> (<year>1997</year>). <article-title>Regulation of synaptic efficacy by coincidence of postsynaptic APs and EPSPs</article-title>. <source>Science</source> <volume>275</volume>, <fpage>213</fpage>&#x02013;<lpage>215</lpage>. <pub-id pub-id-type="doi">10.1126/science.275.5297.213</pub-id><pub-id pub-id-type="pmid">8985014</pub-id></citation>
</ref>
<ref id="B31">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Matsubara</surname> <given-names>T.</given-names></name></person-group> (<year>2017</year>). <article-title>Spike timing-dependent conduction delay learning model classifying spatio-temporal spike patterns</article-title>, in <source>The 2017 International Joint Conference on Neural Networks (IJCNN2017)</source> (<publisher-loc>Anchorage</publisher-loc>: <publisher-name>IEEE</publisher-name>), <fpage>1831</fpage>&#x02013;<lpage>1839</lpage>.</citation>
</ref>
<ref id="B32">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Matsubara</surname> <given-names>T.</given-names></name> <name><surname>Torikai</surname> <given-names>H.</given-names></name></person-group> (<year>2013</year>). <article-title>Asynchronous cellular automaton-based neuron: theoretical analysis and on-FPGA learning</article-title>. <source>IEEE Trans. Neural Netw. Learn. Syst.</source> <volume>24</volume>, <fpage>736</fpage>&#x02013;<lpage>748</lpage>. <pub-id pub-id-type="doi">10.1109/TNNLS.2012.2230643</pub-id><pub-id pub-id-type="pmid">24808424</pub-id></citation>
</ref>
<ref id="B33">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Matsubara</surname> <given-names>T.</given-names></name> <name><surname>Torikai</surname> <given-names>H.</given-names></name></person-group> (<year>2016</year>). <article-title>An asynchronous recurrent network of cellular automaton-based neurons and its reproduction of spiking neural network activities</article-title>. <source>IEEE Trans. Neural Netw. Learn. Syst.</source> <volume>27</volume>, <fpage>836</fpage>&#x02013;<lpage>852</lpage>. <pub-id pub-id-type="doi">10.1109/TNNLS.2015.2425893</pub-id><pub-id pub-id-type="pmid">25974951</pub-id></citation>
</ref>
<ref id="B34">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Matsubara</surname> <given-names>T.</given-names></name> <name><surname>Uehara</surname> <given-names>K.</given-names></name></person-group> (<year>2016</year>). <article-title>Homeostatic plasticity achieved by incorporation of random fluctuations and soft-bounded hebbian plasticity in excitatory synapses</article-title>. <source>Front. Neural Circ.</source> <volume>10</volume>:<fpage>42</fpage>. <pub-id pub-id-type="doi">10.3389/fncir.2016.00042</pub-id><pub-id pub-id-type="pmid">27313513</pub-id></citation>
</ref>
<ref id="B35">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Middlebrooks</surname> <given-names>J. C.</given-names></name> <name><surname>Clock</surname> <given-names>A. E.</given-names></name> <name><surname>Xu</surname> <given-names>L.</given-names></name> <name><surname>Green</surname> <given-names>D. M.</given-names></name></person-group> (<year>1994</year>). <article-title>A panoramic code for sound localization by cortical neurons</article-title>. <source>Science</source> <volume>264</volume>, <fpage>842</fpage>&#x02013;<lpage>844</lpage>.</citation>
</ref>
<ref id="B36">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Murphy</surname> <given-names>K.</given-names></name></person-group> (<year>2012</year>). <source>Machine Learning: A Probabilistic Perspective</source>. <publisher-name>The MIT Press</publisher-name>.</citation>
</ref>
<ref id="B37">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Neftci</surname> <given-names>E.</given-names></name> <name><surname>Das</surname> <given-names>S.</given-names></name> <name><surname>Pedroni</surname> <given-names>B.</given-names></name> <name><surname>Kreutz-Delgado</surname> <given-names>K.</given-names></name> <name><surname>Cauwenberghs</surname> <given-names>G.</given-names></name></person-group> (<year>2014</year>). <article-title>Event-driven contrastive divergence for spiking neuromorphic systems</article-title>. <source>Front. Neurosci.</source> <volume>7</volume>:<fpage>272</fpage>. <pub-id pub-id-type="doi">10.3389/fnins.2013.00272</pub-id><pub-id pub-id-type="pmid">24574952</pub-id></citation>
</ref>
<ref id="B38">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Nessler</surname> <given-names>B.</given-names></name> <name><surname>Pfeiffer</surname> <given-names>M.</given-names></name> <name><surname>Maass</surname> <given-names>W.</given-names></name></person-group> (<year>2009</year>). <article-title>STDP enables spiking neurons to detect hidden causes of their inputs</article-title>, in <source>Advances in Neural Information Processing Systems (NIPS)</source> (<publisher-loc>Whistler, BC</publisher-loc>), <fpage>1357</fpage>&#x02013;<lpage>1365</lpage>.</citation>
</ref>
<ref id="B39">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>O&#x00027;Connor</surname> <given-names>P.</given-names></name> <name><surname>Neil</surname> <given-names>D.</given-names></name> <name><surname>Liu</surname> <given-names>S. C.</given-names></name> <name><surname>Delbruck</surname> <given-names>T.</given-names></name> <name><surname>Pfeiffer</surname> <given-names>M.</given-names></name></person-group> (<year>2013</year>). <article-title>Real-time classification and sensor fusion with a spiking deep belief network</article-title>. <source>Front. Neurosci.</source> <volume>7</volume>:<fpage>178</fpage>. <pub-id pub-id-type="doi">10.3389/fnins.2013.00178</pub-id><pub-id pub-id-type="pmid">24115919</pub-id></citation>
</ref>
<ref id="B40">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Paugam-Moisy</surname> <given-names>H. H.</given-names></name> <name><surname>Martinez</surname> <given-names>R.</given-names></name> <name><surname>Bengio</surname> <given-names>S.</given-names></name></person-group> (<year>2008</year>). <article-title>Delay learning and polychronization for reservoir computing</article-title>. <source>Neurocomputing</source> <volume>71</volume>, <fpage>1143</fpage>&#x02013;<lpage>1158</lpage>. <pub-id pub-id-type="doi">10.1016/j.neucom.2007.12.027</pub-id></citation>
</ref>
<ref id="B41">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Pfister</surname> <given-names>J.-P.</given-names></name> <name><surname>Toyoizumi</surname> <given-names>T.</given-names></name> <name><surname>Barber</surname> <given-names>D.</given-names></name> <name><surname>Gerstner</surname> <given-names>W.</given-names></name></person-group> (<year>2006</year>). <article-title>Optimal spike-timing-dependent plasticity for precise action potential firing in supervised learning</article-title>. <source>Neural Comput.</source> <volume>18</volume>, <fpage>1318</fpage>&#x02013;<lpage>1348</lpage>. <pub-id pub-id-type="doi">10.1162/neco.2006.18.6.1318</pub-id><pub-id pub-id-type="pmid">16764506</pub-id></citation>
</ref>
<ref id="B42">
<citation citation-type="other"><person-group person-group-type="author"><name><surname>Ponulak</surname> <given-names>F.</given-names></name></person-group> (<year>2005</year>). <source>ReSuMe-New Supervised Learning Method for Spiking Neural Networks. Poznan University, Institute of Control and Information Engineering, Vol. 22</source>, <fpage>467</fpage>&#x02013;<lpage>510</lpage>.</citation>
</ref>
<ref id="B43">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Querlioz</surname> <given-names>D.</given-names></name> <name><surname>Bichler</surname> <given-names>O.</given-names></name> <name><surname>Dollfus</surname> <given-names>P.</given-names></name> <name><surname>Gamrat</surname> <given-names>C.</given-names></name></person-group> (<year>2013</year>). <article-title>Immunity to device variations in a spiking neural network with memristive nanodevices</article-title>. <source>IEEE Trans. Nanotechnol.</source> <volume>12</volume>, <fpage>288</fpage>&#x02013;<lpage>295</lpage>. <pub-id pub-id-type="doi">10.1109/TNANO.2013.2250995</pub-id></citation>
</ref>
<ref id="B44">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Rezende</surname> <given-names>J. D.</given-names></name> <name><surname>Gerstner</surname> <given-names>W.</given-names></name></person-group> (<year>2014</year>). <article-title>Stochastic variational learning in recurrent spiking networks</article-title>. <source>Front. Comput. Neurosci.</source> <volume>8</volume>:<fpage>38</fpage>. <pub-id pub-id-type="doi">10.3389/fncom.2014.00038</pub-id></citation>
</ref>
<ref id="B45">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Rushton</surname> <given-names>W. A. H.</given-names></name></person-group> (<year>1951</year>). <article-title>A theory of the effects of fibre size in medullated nerve</article-title>. <source>J. Physiol.</source> <volume>115</volume>, <fpage>101</fpage>&#x02013;<lpage>122</lpage>. <pub-id pub-id-type="doi">10.1113/jphysiol.1951.sp004655</pub-id><pub-id pub-id-type="pmid">14889433</pub-id></citation>
</ref>
<ref id="B46">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Sato</surname> <given-names>M.-A.</given-names></name></person-group> (<year>1999</year>). <source>Fast Learning of On-Line EM Algorithm</source>. Rapport Technique, <publisher-name>ATR Human Information Processing Research Laboratories</publisher-name>.</citation>
</ref>
<ref id="B47">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Seidl</surname> <given-names>A. H.</given-names></name> <name><surname>Rubel</surname> <given-names>E. W.</given-names></name> <name><surname>Harris</surname> <given-names>D. M.</given-names></name></person-group> (<year>2010</year>). <article-title>Mechanisms for adjusting interaural time differences to achieve binaural coincidence detection</article-title>. <source>J. Neurosci.</source> <volume>30</volume>, <fpage>70</fpage>&#x02013;<lpage>80</lpage>. <pub-id pub-id-type="doi">10.1523/JNEUROSCI.3464-09.2010</pub-id><pub-id pub-id-type="pmid">20053889</pub-id></citation>
</ref>
<ref id="B48">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Shouval</surname> <given-names>H. Z.</given-names></name> <name><surname>Bear</surname> <given-names>M. F.</given-names></name> <name><surname>Cooper</surname> <given-names>L. N.</given-names></name></person-group> (<year>2002</year>). <article-title>A unified model of NMDA receptor-dependent bidirectional synaptic plasticity</article-title>. <source>Proc. Natl. Acad. Sci. U.S.A.</source> <volume>99</volume>, <fpage>10831</fpage>&#x02013;<lpage>10836</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.152343099</pub-id><pub-id pub-id-type="pmid">12136127</pub-id></citation>
</ref>
<ref id="B49">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Shouval</surname> <given-names>H. Z.</given-names></name> <name><surname>Wang</surname> <given-names>S. S.-H.</given-names></name> <name><surname>Wittenberg</surname> <given-names>G. M.</given-names></name></person-group> (<year>2010</year>). <article-title>Spike timing dependent plasticity: a consequence of more fundamental learning rules</article-title>. <source>Front. Comput. Neurosci.</source> <volume>4</volume>:<fpage>19</fpage>. <pub-id pub-id-type="doi">10.3389/fncom.2010.00019</pub-id><pub-id pub-id-type="pmid">20725599</pub-id></citation>
</ref>
<ref id="B50">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sporea</surname> <given-names>I.</given-names></name> <name><surname>Gr&#x000FC;ning</surname> <given-names>A.</given-names></name></person-group> (<year>2013</year>). <article-title>Supervised learning in multilayer spiking neural networks</article-title>. <source>Neural Comput.</source> <volume>25</volume>, <fpage>473</fpage>&#x02013;<lpage>509</lpage>. <pub-id pub-id-type="doi">10.1162/NECO_a_00396</pub-id><pub-id pub-id-type="pmid">23148411</pub-id></citation>
</ref>
<ref id="B51">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Taherkhani</surname> <given-names>A.</given-names></name> <name><surname>Belatreche</surname> <given-names>A.</given-names></name> <name><surname>Li</surname> <given-names>Y.</given-names></name> <name><surname>Maguire</surname> <given-names>L. P.</given-names></name></person-group> (<year>2015</year>). <article-title>DL-ReSuMe: a delay learning-based remote supervised method for spiking neurons</article-title>. <source>IEEE Trans. Neural Netw. Learn. Syst.</source> <volume>26</volume>, <fpage>3137</fpage>&#x02013;<lpage>3149</lpage>. <pub-id pub-id-type="doi">10.1109/TNNLS.2015.2404938</pub-id><pub-id pub-id-type="pmid">25794401</pub-id></citation>
</ref>
<ref id="B52">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Turrigiano</surname> <given-names>G. G.</given-names></name> <name><surname>Desai</surname> <given-names>N. S.</given-names></name> <name><surname>Rutherford</surname> <given-names>L. C.</given-names></name></person-group> (<year>1999</year>). <article-title>Plasticity in the intrinsic excitability of cortical pyramidal neurons</article-title>. <source>Nat. Neurosci.</source> <volume>2</volume>, <fpage>515</fpage>&#x02013;<lpage>520</lpage>. <pub-id pub-id-type="doi">10.1038/9165</pub-id><pub-id pub-id-type="pmid">10448215</pub-id></citation>
</ref>
<ref id="B53">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Turrigiano</surname> <given-names>G. G.</given-names></name> <name><surname>Nelson</surname> <given-names>S. B.</given-names></name></person-group> (<year>2000</year>). <article-title>Hebb and homeostasis in neuronal plasticity</article-title>. <source>Curr. Opin. Neurobiol.</source> <volume>10</volume>, <fpage>358</fpage>&#x02013;<lpage>364</lpage>. <pub-id pub-id-type="doi">10.1016/S0959-4388(00)00091-X</pub-id><pub-id pub-id-type="pmid">10851171</pub-id></citation>
</ref>
<ref id="B54">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Van Rossum</surname> <given-names>M. C. W.</given-names></name> <name><surname>Bi</surname> <given-names>G.-Q. Q.</given-names></name> <name><surname>Turrigiano</surname> <given-names>G. G.</given-names></name></person-group> (<year>2000</year>). <article-title>Stable Hebbian learning from spike timing-dependent plasticity</article-title>. <source>J. Neurosci.</source> <volume>20</volume>, <fpage>8812</fpage>&#x02013;<lpage>8821</lpage>. <pub-id pub-id-type="pmid">11102489</pub-id></citation>
</ref>
<ref id="B55">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>VanRullen</surname> <given-names>R.</given-names></name> <name><surname>Thorpe</surname> <given-names>S. J.</given-names></name></person-group> (<year>2001</year>). <article-title>Rate coding versus temporal order coding: what the retinal ganglion cells tell the visual cortex</article-title>. <source>Neural Comput.</source> <volume>13</volume>, <fpage>1255</fpage>&#x02013;<lpage>1283</lpage>. <pub-id pub-id-type="doi">10.1162/08997660152002852</pub-id></citation>
</ref>
<ref id="B56">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Watt</surname> <given-names>A. J.</given-names></name> <name><surname>Desai</surname> <given-names>N. S.</given-names></name></person-group> (<year>2010</year>). <article-title>Homeostatic plasticity and STDP: keeping a neuron&#x00027;s cool in a fluctuating world</article-title>. <source>Front. Synaptic Neurosci.</source> <volume>2</volume>:<fpage>5</fpage>. <pub-id pub-id-type="doi">10.3389/fnsyn.2010.00005</pub-id><pub-id pub-id-type="pmid">21423491</pub-id></citation>
</ref>
<ref id="B57">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Waxman</surname> <given-names>S. G.</given-names></name> <name><surname>Swadlow</surname> <given-names>H. A.</given-names></name></person-group> (<year>1976</year>). <article-title>Ultrastructure of visual callosal axons in the rabbit</article-title>. <source>Exp. Neurol.</source> <volume>53</volume>, <fpage>115</fpage>&#x02013;<lpage>127</lpage>. <pub-id pub-id-type="doi">10.1016/0014-4886(76)90287-9</pub-id><pub-id pub-id-type="pmid">964332</pub-id></citation>
</ref>
<ref id="B58">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wittenberg</surname> <given-names>G. M.</given-names></name> <name><surname>Wang</surname> <given-names>S. S.-H.</given-names></name></person-group> (<year>2006</year>). <article-title>Malleability of spike-timing-dependent plasticity at the CA3-CA1 synapse</article-title>. <source>J. Neurosci.</source> <volume>26</volume>, <fpage>6610</fpage>&#x02013;<lpage>6617</lpage>. <pub-id pub-id-type="doi">10.1523/JNEUROSCI.5388-05.2006</pub-id><pub-id pub-id-type="pmid">16775149</pub-id></citation>
</ref>
<ref id="B59">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Yu</surname> <given-names>Q.</given-names></name> <name><surname>Tang</surname> <given-names>H.</given-names></name> <name><surname>Tan</surname> <given-names>K. C.</given-names></name> <name><surname>Yu</surname> <given-names>H.</given-names></name></person-group> (<year>2014</year>). <article-title>A brain-inspired spiking neural network model with temporal encoding and learning</article-title>. <source>Neurocomputing</source> <volume>138</volume>, <fpage>3</fpage>&#x02013;<lpage>13</lpage>. <pub-id pub-id-type="doi">10.1016/j.neucom.2013.06.052</pub-id></citation>
</ref>
<ref id="B60">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zambrano</surname> <given-names>D.</given-names></name> <name><surname>Bohte</surname> <given-names>S. M.</given-names></name></person-group> (<year>2016</year>). <article-title>Fast and efficient asynchronous neural computation with adapting spiking neural networks</article-title>. <source>arXiv:1609.02053</source>.</citation>
</ref>
</ref-list>
<fn-group>
<fn fn-type="financial-disclosure"><p><bold>Funding.</bold> This study was partially supported by the JSPS KAKENHI (16K12487), Kayamori Foundation of Information Science Advancement, The Nakajima Foundation, and SEI Group CSR Foundation.</p>
</fn>
</fn-group>
</back>
</article>