<?xml version="1.0" encoding="UTF-8" standalone="no"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xml:lang="EN" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Cognit.</journal-id>
<journal-title>Frontiers in Cognition</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Cognit.</abbrev-journal-title>
<issn pub-type="epub">2813-4532</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fcogn.2025.1623227</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Cognition</subject>
<subj-group>
<subject>Hypothesis and Theory</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Commutativity of probabilistic belief revision</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name><surname>Jacobs</surname> <given-names>Bart</given-names></name>
<xref ref-type="corresp" rid="c001"><sup>&#x0002A;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/2841372/overview"/>
<role content-type="https://credit.niso.org/contributor-roles/writing-original-draft/"/>
<role content-type="https://credit.niso.org/contributor-roles/methodology/"/>
<role content-type="https://credit.niso.org/contributor-roles/formal-analysis/"/>
<role content-type="https://credit.niso.org/contributor-roles/investigation/"/>
<role content-type="https://credit.niso.org/contributor-roles/conceptualization/"/>
<role content-type="https://credit.niso.org/contributor-roles/writing-review-editing/"/>
</contrib>
</contrib-group>
<aff><institution>iHub, Radboud University</institution>, <addr-line>Nijmegen</addr-line>, <country>Netherlands</country></aff>
<author-notes>
<fn fn-type="edited-by"><p>Edited by: Andrew Tolmie, University College London, United Kingdom</p></fn>
<fn fn-type="edited-by"><p>Reviewed by: Emmanuel E. Haven, Memorial University of Newfoundland, Canada</p>
<p>Pier Luigi Gentili, Universit&#x000E0; degli Studi di Perugia, Italy</p></fn>
<corresp id="c001">&#x0002A;Correspondence: Bart Jacobs <email>bart&#x00040;cs.ru.nl</email></corresp>
</author-notes>
<pub-date pub-type="epub">
<day>06</day>
<month>08</month>
<year>2025</year>
</pub-date>
<pub-date pub-type="collection">
<year>2025</year>
</pub-date>
<volume>4</volume>
<elocation-id>1623227</elocation-id>
<history>
<date date-type="received">
<day>05</day>
<month>05</month>
<year>2025</year>
</date>
<date date-type="accepted">
<day>25</day>
<month>06</month>
<year>2025</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x000A9; 2025 Jacobs.</copyright-statement>
<copyright-year>2025</copyright-year>
<copyright-holder>Jacobs</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/"><p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p></license>
</permissions>
<abstract>
<p>Bayesian updating, also known as belief revision or conditioning, is a core mechanism of probability theory, and of AI. The human mind is very sensitive to the order in which it is being &#x0201C;primed&#x0201D;, but Bayesian updating works commutatively: the order of the evidence does not matter. Thus, there is a mismatch. This paper develops Bayesian updating as an explicit operation on (discrete) probability distributions, so that the commutativity of Bayesian updating can be clearly formulated and made explicit in several examples. The commutativity mismatch is underexplored, but plays a fundamental role, for instance in the move to quantum cognition.</p></abstract>
<kwd-group>
<kwd>Bayesian updating</kwd>
<kwd>multiset</kwd>
<kwd>commutativity</kwd>
<kwd>notation</kwd>
<kwd>cognition</kwd>
</kwd-group>
<counts>
<fig-count count="1"/>
<table-count count="0"/>
<equation-count count="37"/>
<ref-count count="37"/>
<page-count count="9"/>
<word-count count="7162"/>
</counts>
<custom-meta-wrap>
<custom-meta>
<meta-name>section-at-acceptance</meta-name>
<meta-value>Reason and Decision-Making</meta-value>
</custom-meta>
</custom-meta-wrap>
</article-meta>
</front>
<body>
<sec sec-type="intro" id="s1">
<title>1 Introduction</title>
<p>In mathematics an operation is called commutative if swapping its arguments does not change the outcome, as in addition of numbers: <italic>n</italic>&#x0002B;<italic>m</italic> &#x0003D; <italic>m</italic>&#x0002B;<italic>n</italic>. Also, actions can be called commutative when the effect does not depend on the order in which they are taken: if I first give someone <italic>n</italic> Euros and then another <italic>m</italic> Euros, the financial effect is the same as first giving you <italic>m</italic> and then <italic>n</italic> Euros. However, the emotional effect on the receiver&#x00027;s side may be quite different, for instance when <italic>n</italic> is much greater than <italic>m</italic>: first giving the higher amount <italic>n</italic> then <italic>m</italic> may lead to disappointment, whereas after first giving the lower amount <italic>m</italic> and then <italic>n</italic> the receiver may end up in a more positive mood.</p>
<p>Similar differences are well-known in human cognition, especially when one looks at how the mind processes (new) information&#x02014;in what is called priming. The order of such priming (or updating) is highly relevant. Here is a simple example, involving two sentences <italic>p</italic> and <italic>q</italic>, about Alice and Bob, which will be presented below in two orders: (a) as <italic>p</italic> and <italic>q</italic>, and (b) as <italic>q</italic> and <italic>p</italic>. This makes no difference in (Boolean) logic, but watch carefully what effect the different orders has on your understanding of the situation.</p>
<list list-type="simple">
<list-item><p>(a) &#x0201C;Alice is sick&#x0201D; and &#x0201C;Bob visits Alice&#x0201D;;</p></list-item>
<list-item><p>(b) &#x0201C;Bob visits Alice&#x0201D; and &#x0201C;Alice is sick.&#x0201D;</p></list-item>
</list>
<p>In the first case (a) you may think that Bob is a nice guy, but maybe less so in the second case (b). It is surprising how quickly the human mind makes a (causal) connection. The strength of this effect depends on many factors, including one&#x00027;s own background (priors). Check for instance what the effect is of replacing in the above two sentences &#x0201C;sick&#x0201D; with &#x0201C;pregnant.&#x0201D; This dependence on the order is called the order effect in Uzan (<xref ref-type="bibr" rid="B36">2023</xref>). It is the main motivation to switch to quantum logic in cognition theory (see e.g., Busemeyer and Bruza, <xref ref-type="bibr" rid="B3">2012</xref>; Yearsley and Busemeyer, <xref ref-type="bibr" rid="B37">2016</xref>; Gentili, <xref ref-type="bibr" rid="B10">2021</xref>), since conjunction (&#x0201C;and&#x0201D;) in quantum logic is not commutative.</p>
<p>This paper offers reflections on commutativity in probability theory, and in particular in probabilistic updating. Bayesian updating is introduced as an explicit operation that takes a probability distribution &#x003C9; with some form of evidence <italic>p</italic> and produces a new, updated distribution &#x003C9;|<italic>p</italic>. This formulation generalizes the common approach. The (mathematical) details will appear later, but at this stage it is relevant to emphasize that this new formulation of Bayesian updating makes it possible to clearly express commutativity: for two pieces of evidence <italic>p</italic> and <italic>q</italic> one has:</p>
<disp-formula id="E1"><label>(1)</label><mml:math id="M1"><mml:mrow><mml:mi>&#x003C9;</mml:mi><mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mi>p</mml:mi></mml:msub><mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mi>q</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mi>&#x003C9;</mml:mi><mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mi>q</mml:mi></mml:msub><mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mi>p</mml:mi></mml:msub><mml:mo>.</mml:mo></mml:mrow></mml:math></disp-formula>
<p>In words: updating a distribution &#x003C9; first with <italic>p</italic> and then with <italic>q</italic> gives the same result as updating &#x003C9; first with <italic>q</italic> and then with <italic>p</italic>. This commutativity cannot be expressed using the traditional notation <italic>P</italic>(<italic>E</italic>&#x02223;<italic>D</italic>) of conditional probabilities, since it leaves the distribution implicit.</p>
<p>To add some terminology: this updating is also called conditioning or belief revision. The distribution &#x003C9;, before the update, is called the prior, and the distribution &#x003C9;|<italic>p</italic>, after the update, is called the posterior. The task of computing the posterior distribution is called inference, or also (probabilistic) learning.</p>
<p>Probabilistic updating is an essential part of the ongoing AI-revolution, in various forms, via learning and training. Generative and conversational AI are becoming part of professional and private environments. If these tools are meant to behave like humans, their updating should be non-commutative, as illustrated above, with the different effects resulting from different orders (a) and (b). Bayesian updating, however, is commutative (<xref ref-type="disp-formula" rid="E1">Equation 1</xref>). Hence there is a mismatch [as also emphasized in Uzan (<xref ref-type="bibr" rid="B36">2023</xref>)]. One aim of this paper is increase the awareness of this gap, by making the commutativity of Bayesian updating explicit, in a new form (<xref ref-type="disp-formula" rid="E1">Equation 1</xref>).</p>
<p>The paper starts with some simple observations about lists and multisets. One can have a list of letters, say (<italic>a, b, c, a, c, b, a</italic>). In a list, elements can occur multiple times and the order of their occurrence matters. Notice that in a subset like {<italic>a, c</italic>}, elements can occur at most once, and their order does not matter. Mulitsets are &#x0201C;inbetween&#x0201D; lists and subsets: elements may occur multiple times, but their order does not matter. The latter property makes them relevant in the current context. Multisets are a highly undervalued datatype. The fact that they are so little used may be part of our poor understanding of the role of commutativity. We can already make a connection with what we saw above: updating a distribution with two pieces of evidence&#x02014;as in <xref ref-type="disp-formula" rid="E1">Equation 1</xref>&#x02014;should not be done with a list (<italic>p, q</italic>) or (<italic>q, p</italic>), but with a multiset of evidence containing both <italic>p</italic> and <italic>q</italic>, where the order does not matter. Section 2 below starts with an informal introduction to multisets and ends with some notation and definitions that are relevant in this setting.</p>
<p>Multisets form a natural preparation for (discrete probability) distributions, in Section 3. Distributions keep count of elements via probabilities or weights (in the unit interval [0, 1]) that add up to one. Multisets can be turned into &#x0201C;fractional&#x0201D; distributions via normalization. A basic fact is that every distribution can be obtained as limit of such fractional distributions, just like every real number can be obtained as a limit of fractions. Mathematically, this is described as: the set <inline-formula><mml:math id="M2"><mml:mrow><mml:mstyle mathvariant="script"><mml:mi>D</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>X</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> of distributions on a finite set <italic>X</italic> is a compact complete metric space, with (normalized) multisets as dense subset, see Theorem 1. This denseness formalizes the &#x0201C;frequentist&#x0201D; perspective on probability distributions, as results of long-term experiments.</p>
<p>Section 4 introduces Bayesian updating &#x003C9;|<italic>p</italic> in concrete form, shows that traditional notation <italic>P</italic>(<italic>E</italic>&#x02223;<italic>D</italic>) is a special case, and illustrates usage of updates &#x003C9;|<italic>p</italic> in two examples. The first one is a rather straightforward application, where a prior bird distribution is updated after a bird count. If there is another bird count one year later, the bird distribution can be updated once again. It turns out that eventhough the counts are chronologically ordered, this order is irrelevant for the distribution updates. The second example is more challenging. It involves various inference questions about the sex and ages of children in a family (with a twin), after specific observations. The possibilities for the offspring are represented as multisets, based on Section 2. The prior distribution in this case is a distribution over these multisets. Again, the order of the multiple updates does not matter. This basic fact is proven in general form at the end of this section, see Proposition 1.</p>
<p>The commutativity of Bayesian updating may be seen as folklore knowledge but it is hardly made explicit in the literature. One reason is that the standard formulation of conditional probabilities <italic>P</italic>(<italic>E</italic>&#x02223;<italic>D</italic>) does not lend it self to a commutativity result, as above in <xref ref-type="disp-formula" rid="E1">Equation 1</xref>, since it hides the distribution, assuming there is only one implicit distribution. Hence one cannot express facts about different distributions via traditional notation. Our formulation of Bayesian updating &#x003C9;|<italic>p</italic> as an operation on distributions thus has advantages&#x02014;as hopefully also becomes clear from the illustrations in Section 4. The <xref ref-type="supplementary-material" rid="SM1">Appendix</xref> derives the update formulation &#x003C9;|<italic>p</italic> from the traditional formulation via Kadison duality. This is a new result. The derivation is mathematically sophisticated and is not necessary for the main line of the paper. This line is part of a new approach to probability theory using the language and methods of category theory. In the body of the paper the role of category theory remains in the background and no prior knowledge of that field is required.</p>
<p>Thus, this paper&#x00027;s contributions lie in putting the spotlight on the commutativity of Bayesian updating, in a new form (<xref ref-type="disp-formula" rid="E1">Equation 1</xref>), via a new formulation &#x003C9;|<italic>p</italic> of this update mechanism, which is both given in concrete form and derived from a fundamental duality result. At the same time, the paper provides a gentle introduction to a new approach to the area, in which multisets and explicitly written distributions play a central role.</p></sec>
<sec id="s2">
<title>2 Multisets, with multiplicities of elements</title>
<p>Suppose you check how much money you have in your pocket and you find that you have three 2-Euro coins <inline-graphic mimetype="image" mime-subtype="tiff" xlink:href="fcogn-04-1623227-i0002.tif"/> and two 1-Euro coins <inline-graphic mimetype="image" mime-subtype="tiff" xlink:href="fcogn-04-1623227-i0001.tif"/>. How would you describe this handful of coins mathematically? It is not a subset of coins {<inline-graphic mimetype="image" mime-subtype="tiff" xlink:href="fcogn-04-1623227-i0002.tif"/>, <inline-graphic mimetype="image" mime-subtype="tiff" xlink:href="fcogn-04-1623227-i0001.tif"/>}, since such a subset ignores the multiplicities of the coins that you have. One can describe the contents of your pocket as a list, for instance as, (<inline-graphic mimetype="image" mime-subtype="tiff" xlink:href="fcogn-04-1623227-i0002.tif"/>, <inline-graphic mimetype="image" mime-subtype="tiff" xlink:href="fcogn-04-1623227-i0001.tif"/>, <inline-graphic mimetype="image" mime-subtype="tiff" xlink:href="fcogn-04-1623227-i0002.tif"/>, <inline-graphic mimetype="image" mime-subtype="tiff" xlink:href="fcogn-04-1623227-i0001.tif"/>, <inline-graphic mimetype="image" mime-subtype="tiff" xlink:href="fcogn-04-1623227-i0002.tif"/>), when you lay them out in your hand. But the order of this list is arbitrary and does not reflect your answer: I have three of <inline-graphic mimetype="image" mime-subtype="tiff" xlink:href="fcogn-04-1623227-i0002.tif"/> and two of <inline-graphic mimetype="image" mime-subtype="tiff" xlink:href="fcogn-04-1623227-i0001.tif"/>.</p>
<p>The proper way to capture the situation mathematically is via a <italic>multiset</italic>. It can be understood as a subset in which elements may occur multiple times, or as a list in which the order does not matter. Unfortunately, there is no established notation for multisets. We use kets |&#x02212;&#x0232A;, where we put the elements of the multiset inside the ket and their multiplicity in front. Thus, the coins in your pocket, are properly described as multiset:</p>
<p><graphic xlink:href="fcogn-04-1623227-e0001.tif"/></p>
<p>These kets |&#x02212;&#x0232A; are borrowed from quantum theory. They have no mathematical meaning here and are used only to separate the elements of a multiset from their multiplicities.</p>
<p>As a practical example, consider the outcome of an election, say between two candidates Alice (<italic>A</italic>) and Bob (<italic>B</italic>). The outcome may be written as a multiset <inline-formula><mml:math id="M3"><mml:mrow><mml:mn>55</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>A</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>45</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow></mml:math></inline-formula>, indicating that 55 votes are for Alice and 45 votes for Bob. Describing the outcome of the election as a list of length 100, say with the chronological order of casted votes, is a bad idea, for two reasons: the list does not immediately tell what the outcome is, and the list may leak information about who voted for whom&#x02014;when the order of the voters is recorded.</p>
<p>In which ways can you break a note of 10 Euro into coins of 2 and 1 Euros? The six options can be described as multisets:</p>
<p><graphic xlink:href="fcogn-04-1623227-e0002.tif"/></p>
<p>When we are interested in the ways to break the note of 10, we do not care about the order of the coins. When we do describe the break-up options as lists we end up with 10 different lists of coins.</p>
<p>Multisets are a useful &#x0201C;datatype,&#x0201D; in the language of computer science, for keeping counts of elements. However, multisets are often not recognized or expounded as such. For instance, in mathematics, the solutions of a polynomial form a multiset, and not a set, since solutions may occur multiple times. For example, the multiset of solutions of the polynomial <italic>x</italic><sup>3</sup>&#x02212;7<italic>x</italic><sup>2</sup>&#x0002B;16<italic>x</italic>&#x02212;12 &#x0003D; (<italic>x</italic>&#x02212;2)(<italic>x</italic>&#x02212;2)(<italic>x</italic>&#x02212;3) takes the form 2|2&#x0232A;&#x0002B;1|3&#x0232A;, since the number 2 occurs twice as solution and 3 once. Similarly, the eigenvalues of a matrix form a multiset. In the notation 2|2&#x0232A;&#x0002B;1|3&#x0232A; the kets play a useful role, since they make clear which numbers are in the multiset and which numbers are the corresponding multiplicities.</p>
<p>Consider the following basic question. A friend of mine has three children, but I don&#x00027;t know if they are girls (<italic>G</italic>) or boys (<italic>B</italic>). How many offspring options are there? Many people will quickly say: <italic>eight</italic>, namely:</p>
<disp-formula id="E2"><label>(2)</label><mml:math id="M4"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mi>G</mml:mi><mml:mtext>&#x02003;&#x000A0;</mml:mtext><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mi>B</mml:mi><mml:mtext>&#x02003;&#x000A0;</mml:mtext><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mi>G</mml:mi><mml:mtext>&#x02003;&#x000A0;</mml:mtext><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mi>G</mml:mi></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mi>B</mml:mi><mml:mtext>&#x02003;&#x000A0;</mml:mtext><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mi>B</mml:mi><mml:mtext>&#x02003;&#x000A0;</mml:mtext><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mi>G</mml:mi><mml:mtext>&#x02003;&#x000A0;</mml:mtext><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mi>B</mml:mi><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>One can also say: there are <italic>four</italic> options, namely with three, two, one, or zero girls. These four options are described as multisets:</p>
<disp-formula id="E3"><label>(3)</label><mml:math id="M5"><mml:mrow><mml:mn>3</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mtext>&#x02003;</mml:mtext><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mtext>&#x02003;</mml:mtext><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mtext>&#x02003;</mml:mtext><mml:mn>3</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>.</mml:mo></mml:mrow></mml:math></disp-formula>
<p>The eight list options in <xref ref-type="disp-formula" rid="E2">Equation 2</xref> only make sense if there is an order&#x02014;<italic>e.g</italic>. by ascending or descending age&#x02014;but no order was specified in the question. Hence the multiset answer, with four options, seems most natural. It abstracts away from any ordering of the children.</p>
<p>In the next section we illustrate how multisets form the basis for discrete probability theory. Indeed, probabilities naturally arise from counting, possibly in a limit process. Urns filled with balls of different colors form a basic model in probability theory (see e.g. Johnson and Kotz, <xref ref-type="bibr" rid="B24">1977</xref>; Mahmoud, <xref ref-type="bibr" rid="B28">2008</xref>; Ross, <xref ref-type="bibr" rid="B33">2018</xref>) and many other references. Here, such an urn is identified with a multiset over the set of colors (see <xref ref-type="fig" rid="F1">Figure 1</xref>). The multiplicities of the different colors determine the probability of drawing a ball of a particular color. For instance, in this case, the probability of drawing a single red ball is <inline-formula><mml:math id="M6"><mml:mfrac><mml:mrow><mml:mn>4</mml:mn></mml:mrow><mml:mrow><mml:mn>9</mml:mn></mml:mrow></mml:mfrac></mml:math></inline-formula>. A general draw from such an urn is also a multiset, as on the right in <xref ref-type="fig" rid="F1">Figure 1</xref>.</p>
<fig id="F1" position="float">
<label>Figure 1</label>
<caption><p>An urn filled with colored balls on the left, written as multiset, and a possible draw from this urn on the right, also written as a multiset. One can ask what is the probability of this draw from the urn. This depends on the mode of drawing: draw-and-delete, draw-and-replace, draw-and-duplicate (see Jacobs, <xref ref-type="bibr" rid="B18">2022</xref>, <xref ref-type="bibr" rid="B20">2025</xref>) for details.</p></caption>
<alt-text>Illustration of two scenarios: on the left, a beaker contains balls labeled R, B, and G in circles, representing four R, three B, and two G. On the right, a hand holds three balls, labeled with two R, one B, and one G.</alt-text>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fcogn-04-1623227-g0001.tif"/>
</fig>
<p>These introductory observations illustrate that multisets are useful mathematical abstractions for keeping count of multiple occurrences of different objects (or data items). We have mentioned (<xref ref-type="disp-formula" rid="E1">Equation 1</xref>) that the order of data in Bayesian updating is irrelevant. This implies that data in Bayesian learning is best organized as a multiset. Indeed, a histogram of data&#x02014;with heights in natural numbers&#x02014;is another example of a multiset.</p>
<p>The next definition fixes the notation and terminology that we shall use in the sequel. In this setting, a multiset involves only finitely many elements from a given set. The multiplicities are natural numbers. One could also allow non-negative real numbers as multiplicities, but we don&#x00027;t need such generalizations. Here and in the sequel we shall use the sign := for definitions.</p>
<p>Definition 1. Let <italic>X</italic> be an arbitrary set.</p>
<list list-type="order">
<list-item><p>A <italic>multiset</italic> over <italic>X</italic> is a finite formal sum of the form:</p></list-item>
</list>
<disp-formula id="E4"><mml:math id="M7"><mml:mrow><mml:mtable><mml:mtr><mml:mtd><mml:mrow><mml:msub><mml:mi>n</mml:mi><mml:mn>1</mml:mn></mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mn>1</mml:mn></mml:msub></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mo>&#x022EF;</mml:mo><mml:mo>+</mml:mo><mml:msub><mml:mi>n</mml:mi><mml:mi>k</mml:mi></mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi>k</mml:mi></mml:msub></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:mtext>with&#x02003;</mml:mtext><mml:mrow><mml:mo>{</mml:mo> <mml:mrow><mml:mtable columnalign='left'><mml:mtr columnalign='left'><mml:mtd columnalign='left'><mml:mrow><mml:mtext>multiplicities&#x000A0;</mml:mtext><mml:msub><mml:mi>n</mml:mi><mml:mn>1</mml:mn></mml:msub><mml:mo>,</mml:mo><mml:mo>&#x02026;</mml:mo><mml:mo>,</mml:mo><mml:msub><mml:mi>n</mml:mi><mml:mi>k</mml:mi></mml:msub><mml:mo>&#x02208;</mml:mo><mml:mi>&#x02115;</mml:mi></mml:mrow></mml:mtd></mml:mtr><mml:mtr columnalign='left'><mml:mtd columnalign='left'><mml:mrow><mml:mtext>elements&#x000A0;</mml:mtext><mml:msub><mml:mi>x</mml:mi><mml:mn>1</mml:mn></mml:msub><mml:mo>,</mml:mo><mml:mo>&#x02026;</mml:mo><mml:mo>,</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi>k</mml:mi></mml:msub><mml:mo>&#x02208;</mml:mo><mml:mi>X</mml:mi><mml:mo>.</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow> </mml:mrow></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:math></disp-formula>
<list list-type="simple">
<list-item><p>Alternatively, a multiset over <italic>X</italic> may be described as a function &#x003C6;:<italic>X</italic> &#x02192; &#x02115; with finite support <italic>supp</italic>(&#x003C6;): &#x0003D; {<italic>x</italic> &#x02208; <italic>X</italic>|&#x003C6;(<italic>x</italic>) &#x02260; 0}.</p></list-item>
<list-item><p>2. The <italic>size</italic> ||&#x003C6;||&#x02208;&#x02115; of a multiset &#x003C6; is its total number of elements, including multiplicities. Explicitly, both in ket and function notation:</p></list-item>
</list>
<disp-formula id="E5"><mml:math id="M8"><mml:mrow><mml:mtable columnalign='right'><mml:mtr columnalign='right'><mml:mtd columnalign='right'><mml:mrow><mml:mrow><mml:mo>&#x02016;</mml:mo> <mml:mrow><mml:mstyle displaystyle='true'><mml:msub><mml:mo>&#x02211;</mml:mo><mml:mi>i</mml:mi></mml:msub><mml:mrow><mml:msub><mml:mi>n</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow></mml:mstyle></mml:mrow> <mml:mo>&#x02016;</mml:mo></mml:mrow><mml:mo>:</mml:mo><mml:mtext>&#x0200B;</mml:mtext><mml:mo>=</mml:mo><mml:mstyle displaystyle='true'><mml:msub><mml:mo>&#x02211;</mml:mo><mml:mi>i</mml:mi></mml:msub><mml:mrow><mml:msub><mml:mi>n</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:mstyle><mml:mtext>&#x000A0;and&#x000A0;&#x000A0;</mml:mtext><mml:mrow><mml:mo>&#x02016;</mml:mo> <mml:mi>&#x003C6;</mml:mi> <mml:mo>&#x02016;</mml:mo></mml:mrow><mml:mo>:</mml:mo><mml:mo>=</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>x</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mi>s</mml:mi><mml:mi>u</mml:mi><mml:mi>p</mml:mi><mml:mi>p</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x003C6;</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow></mml:munder><mml:mrow><mml:mi>&#x003C6;</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow></mml:mstyle></mml:mrow></mml:mtd></mml:mtr><mml:mtr columnalign='right'><mml:mtd columnalign='right'><mml:mrow><mml:mo>=</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>x</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mi>X</mml:mi></mml:mrow></mml:munder><mml:mi>&#x003C6;</mml:mi></mml:mstyle><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>.</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:math></disp-formula>
<list list-type="simple">
<list-item><p>3. We shall write <inline-formula><mml:math id="M9"><mml:mrow><mml:mstyle mathvariant="script"><mml:mi>M</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>X</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> for the set of all multisets over <italic>X</italic>. This operation <inline-formula><mml:math id="M10"><mml:mrow><mml:mstyle mathvariant="script"><mml:mi>M</mml:mi></mml:mstyle></mml:mrow></mml:math></inline-formula> is functorial: it works not only on sets <italic>X</italic> but also on functions <italic>f</italic>:<italic>X</italic>&#x02192;<italic>Y</italic>, and then yields a function <inline-formula><mml:math id="M11"><mml:mrow><mml:mstyle mathvariant="script"><mml:mi>M</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>f</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>:</mml:mo><mml:mrow><mml:mstyle mathvariant="script"><mml:mi>M</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>X</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x02192;</mml:mo><mml:mrow><mml:mstyle mathvariant="script"><mml:mi>M</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> given by:</p></list-item>
</list>
<disp-formula id="E6"><label>(4)</label><mml:math id="M12"><mml:mrow><mml:mstyle mathvariant="script"><mml:mi>M</mml:mi></mml:mstyle><mml:mtable><mml:mtr><mml:mtd><mml:mrow><mml:mo stretchy='false'>(</mml:mo><mml:mi>f</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='true'>(</mml:mo><mml:mstyle displaystyle='true'><mml:msub><mml:mo>&#x02211;</mml:mo><mml:mi>i</mml:mi></mml:msub><mml:mrow><mml:msub><mml:mi>n</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow></mml:mstyle><mml:mo stretchy='true'>)</mml:mo></mml:mrow><mml:mrow><mml:mo>:</mml:mo><mml:mo>=</mml:mo></mml:mrow><mml:mrow><mml:mstyle displaystyle='true'><mml:msub><mml:mo>&#x02211;</mml:mo><mml:mi>i</mml:mi></mml:msub><mml:mrow><mml:msub><mml:mi>n</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mrow><mml:mi>f</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow></mml:mstyle><mml:mo>.</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:math></disp-formula>
<p>Notice that in this definition the set <italic>X</italic> may be infinite, but a multiset over <italic>X</italic> has only finitely many elements from <italic>X</italic>, in its support. We freely switch between the ket and function notation for multisets and use whichever is most convenient in a particular situation.</p>
<p>The ket notation involves a formal sum, with some conventions: (1) terms 0|<italic>x</italic>&#x0232A; with multiplicity zero may be omitted; (2) a sum <italic>n</italic>|<italic>x</italic>&#x0232A;&#x0002B;<italic>m</italic>|<italic>x</italic>&#x0232A; is the same as (<italic>n</italic>&#x0002B;<italic>m</italic>)|<italic>x</italic>&#x0232A;; (3) the order and any round brackets in a sum do not matter. Thus, for instance, there is an equality of multisets:</p>
<disp-formula id="E7"><mml:math id="M13"><mml:mrow><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>a</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mn>5</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>b</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>0</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>c</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow><mml:mo>)</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>4</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>b</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mn>9</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>b</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>a</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow></mml:math></disp-formula>
<p>These conventions are especially relevant for functoriality in item 3, since multiple elements <italic>x, x</italic>&#x02032; may be mapped to the same outcome <italic>f</italic>(<italic>x</italic>) &#x0003D; <italic>f</italic>(<italic>x</italic>&#x02032;). We use this functoriality especially for projection functions &#x003C0;<sub><italic>i</italic></sub>:<italic>X</italic><sub>1</sub>&#x000D7;<italic>X</italic><sub>2</sub>&#x02192;<italic>X</italic><sub><italic>i</italic></sub>. It then yields marginalisaiton.</p>
<p>As briefly discussed in the introduction, it is illuminating to compare multisets to the datatypes of lists and subsets.</p>
<table-wrap position="float">
<table frame="box" rules="all">
<tbody>
<tr>
<td valign="top" align="left"><inline-graphic mimetype="image" mime-subtype="tiff" xlink:href="fcogn-04-1623227-i0003.tif"/></td>
</tr>
</tbody>
</table>
</table-wrap>
<p>If we write <inline-formula><mml:math id="M14"><mml:mrow><mml:mstyle mathvariant="script"><mml:mi>L</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>X</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> and <inline-formula><mml:math id="M15"><mml:mrow><mml:mstyle mathvariant="script"><mml:mi>P</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>X</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> for the sets of finite lists and of subsets over a set <italic>X</italic>, then we can form a diagram in which multisets sit inbetween lists and subsets:</p>
<disp-formula id="E8"><label>(5)</label><mml:math id="M16"><mml:mrow><mml:mstyle mathvariant="script"><mml:mi>L</mml:mi></mml:mstyle><mml:mo stretchy='false'>(</mml:mo><mml:mi>X</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mtext>&#x2009;</mml:mtext><mml:munder><mml:mrow><mml:mover><mml:mo>&#x2192;</mml:mo><mml:mrow><mml:mi>a</mml:mi><mml:mi>c</mml:mi><mml:mi>c</mml:mi></mml:mrow></mml:mover></mml:mrow><mml:mrow><mml:mtext>forget&#x00A0;order</mml:mtext></mml:mrow></mml:munder><mml:mstyle mathvariant="script"><mml:mi>M</mml:mi></mml:mstyle><mml:mo stretchy='false'>(</mml:mo><mml:mi>X</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mtext>&#x2009;</mml:mtext><mml:munder><mml:mrow><mml:mover><mml:mo>&#x2192;</mml:mo><mml:mrow><mml:mi>s</mml:mi><mml:mi>u</mml:mi><mml:mi>p</mml:mi><mml:mi>p</mml:mi></mml:mrow></mml:mover></mml:mrow><mml:mrow><mml:mtext>forget&#x00A0;multiplicity</mml:mtext></mml:mrow></mml:munder><mml:mstyle mathvariant="script"><mml:mi>P</mml:mi></mml:mstyle><mml:mo stretchy='false'>(</mml:mo><mml:mi>X</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow></mml:math></disp-formula>
<p>The function <italic>acc</italic> performs accumulation, via <italic>acc</italic>(<italic>x</italic><sub>1</sub>, &#x02026;, <italic>x</italic><sub><italic>n</italic></sub>): &#x0003D; 1|<italic>x</italic><sub>1</sub>&#x0232A;&#x0002B;&#x022EF;&#x0002B;1|<italic>x</italic><sub><italic>n</italic></sub>&#x0232A;. It counts the occurrences of elements in a list, for instance in:</p>
<disp-formula id="E9"><mml:math id="M17"><mml:mrow><mml:mi>a</mml:mi><mml:mi>c</mml:mi><mml:mi>c</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>c</mml:mi><mml:mo>,</mml:mo><mml:mi>b</mml:mi><mml:mo>,</mml:mo><mml:mi>a</mml:mi><mml:mo>,</mml:mo><mml:mi>a</mml:mi><mml:mo>,</mml:mo><mml:mi>a</mml:mi><mml:mo>,</mml:mo><mml:mi>b</mml:mi><mml:mo>,</mml:mo><mml:mi>c</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mn>3</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>a</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>b</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>c</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow></mml:math></disp-formula>
<p>Similarly, the accumulations of the lists of children in <xref ref-type="disp-formula" rid="E2">Equation 2</xref> yields the multisets of children in <xref ref-type="disp-formula" rid="E3">Equation 3</xref>.</p>
<p>The support map <italic>supp</italic> in <xref ref-type="disp-formula" rid="E8">Equation 5</xref> sends a multiset on a set to its subset of elements, see Definition 1 1. In the example, <italic>supp</italic>(3|<italic>a</italic>&#x0232A; &#x0002B; 2|<italic>b</italic>&#x0232A; &#x0002B; 2|<italic>c</italic>&#x0232A;) &#x0003D; {<italic>a, b, c</italic>}. As an aside for the mathematically oriented reader, both <italic>acc</italic> and <italic>supp</italic> preserve the monoid structures on these data types and they both are maps of monads. They are fundamental, well-behaved mappings.</p></sec>
<sec id="s3">
<title>3 Distributions, with probabilities of elements</title>
<p>This section introduces finite discrete probability distributions, using ket notation as for multisets. It is shown that multisets give rise to &#x0201C;fractional&#x0201D; distributions, via normalization, and that each distribution is in fact a limit of such fractional distributions.</p>
<p>A coin has two sides, namely head (<italic>H</italic>) and tails (<italic>T</italic>). A fair coin assigns a probability of <inline-formula><mml:math id="M18"><mml:mfrac><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:mfrac></mml:math></inline-formula> to both sides. In ket notation we write it as on the left below.</p>
<disp-formula id="E10"><mml:math id="M19"><mml:mrow><mml:mfrac><mml:mn>1</mml:mn><mml:mn>2</mml:mn></mml:mfrac><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>H</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>2</mml:mn></mml:mfrac><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>T</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mfrac><mml:mrow><mml:mn>51</mml:mn></mml:mrow><mml:mrow><mml:mn>100</mml:mn></mml:mrow></mml:mfrac><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>H</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mrow><mml:mn>49</mml:mn></mml:mrow><mml:mrow><mml:mn>100</mml:mn></mml:mrow></mml:mfrac><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>T</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>.</mml:mo></mml:mrow></mml:math></disp-formula>
<p>On the right, above, there is an almost fair coin, with a slight bias toward head. Characteristically for distributions, the numbers before the kets must be probabilties from the unit interval [0, 1] that add up to one.</p>
<p>Consider the urn multiset &#x003C5; &#x0003D; 4|<italic>R</italic>&#x0232A;&#x0002B;3|<italic>B</italic>&#x0232A;&#x0002B;2|<italic>G</italic>&#x0232A; from <xref ref-type="fig" rid="F1">Figure 1</xref>, with size ||&#x003C5;|| &#x0003D; 9. The probability of drawing a red ball is <inline-formula><mml:math id="M20"><mml:mfrac><mml:mrow><mml:mn>4</mml:mn></mml:mrow><mml:mrow><mml:mn>9</mml:mn></mml:mrow></mml:mfrac><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mi>&#x003C5;</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>R</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mo>|</mml:mo><mml:mo>|</mml:mo><mml:mi>&#x003C5;</mml:mi><mml:mo>|</mml:mo><mml:mo>|</mml:mo></mml:mrow></mml:mfrac></mml:math></inline-formula>. These draw probabilities arise via normalization of the urn-as-multiset, for which we use a function <italic>flrn</italic>, as in:</p>
<disp-formula id="E11"><mml:math id="M21"><mml:mtable columnalign='left'><mml:mtr><mml:mtd><mml:mi>f</mml:mi><mml:mi>l</mml:mi><mml:mi>r</mml:mi><mml:mi>n</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x003C5;</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mi>&#x003C5;</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>R</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mrow><mml:mo>&#x02016;</mml:mo><mml:mi>&#x003C5;</mml:mi><mml:mo>&#x02016;</mml:mo></mml:mrow></mml:mfrac><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>R</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mrow><mml:mi>&#x003C5;</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>B</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mrow><mml:mo>&#x02016;</mml:mo><mml:mi>&#x003C5;</mml:mi><mml:mo>&#x02016;</mml:mo></mml:mrow></mml:mfrac><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mrow><mml:mi>&#x003C5;</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>G</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mrow><mml:mo>&#x02016;</mml:mo><mml:mi>&#x003C5;</mml:mi><mml:mo>&#x02016;</mml:mo></mml:mrow></mml:mfrac><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mfrac><mml:mn>4</mml:mn><mml:mn>9</mml:mn></mml:mfrac><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>R</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>3</mml:mn><mml:mn>9</mml:mn></mml:mfrac><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>2</mml:mn><mml:mn>9</mml:mn></mml:mfrac><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>The latter distribution captures the probabilities of drawing a ball of a particular color from the urn &#x003C5;. The function <italic>flrn</italic> will be defined in general form below. It turns a (non-empty) multiset into a distribution. It learns this distribution by counting, so <italic>flrn</italic> is used as abbreviation of frequentist learning.</p>
<p>As argued in Gigerenzer and Hoffrage (<xref ref-type="bibr" rid="B11">1995</xref>), people are in general not very good at probabilistic (esp. Bayesian) reasoning, but they fare better at reasoning with what are called &#x0201C;frequency formats&#x0201D; but which are in fact multisets. The information that there is a 0.04 probability of getting a disease can be captured in a distribution <inline-formula><mml:math id="M22"><mml:mrow><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>25</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mi>D</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mrow><mml:mn>24</mml:mn></mml:mrow><mml:mrow><mml:mn>25</mml:mn></mml:mrow></mml:mfrac><mml:msup><mml:mi>D</mml:mi><mml:mo>&#x022A5;</mml:mo></mml:msup></mml:mrow></mml:math></inline-formula>, where <italic>D</italic><sup>&#x022A5;</sup> stands for no-disease. This appears more difficult to grasp than the information that 4 out of 100 people get the disease, as captured by the multiset 4|<italic>D</italic>&#x0232A;&#x0002B;96|<italic>D</italic>&#x022A5;&#x0232A;. Applying frequentist learning <italic>flrn</italic> to the latter multiset gives the disease distribution.</p>
<p>Definition 2. Let <italic>X, Y</italic> be arbitrary sets.</p>
<list list-type="simple">
<list-item><p>1. A <italic>distribution</italic> on <italic>X</italic> is a formal finite convex sum <italic>r</italic><sub>1</sub>|<italic>x</italic><sub>1</sub>&#x0232A;&#x0002B;&#x022EF;&#x0002B;<italic>r</italic><sub><italic>n</italic></sub>|<italic>x</italic><sub><italic>n</italic></sub>&#x0232A; with elements <italic>x</italic><sub><italic>i</italic></sub> &#x02208; <italic>X</italic> and associated probabilities <italic>r</italic><sub><italic>i</italic></sub> &#x02208; [0, 1] satisfying <inline-formula><mml:math id="M23"><mml:munder class="msub"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:munder><mml:msub><mml:mrow><mml:mi>r</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:math></inline-formula>.</p></list-item>
<list-item><p>Alternatively, a distribution is given by a &#x0201C;probability mass&#x0201D; function &#x003C9;:<italic>X</italic> &#x02192; [0, 1]&#x02286;&#x0211D; with finite support <italic>supp</italic>(&#x003C9;) &#x0003D; {<italic>x</italic> &#x02208; <italic>X</italic>|&#x003C9;(<italic>x</italic>) &#x02260; 0} and with <inline-formula><mml:math id="M24"><mml:munder class="msub"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>x</mml:mi></mml:mrow></mml:munder><mml:mi>&#x003C9;</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:math></inline-formula>.</p></list-item>
<list-item><p>Each non-empty multiset <inline-formula><mml:math id="M25"><mml:mrow><mml:mi>&#x003C6;</mml:mi><mml:mo>=</mml:mo><mml:mstyle displaystyle='true'><mml:msub><mml:mo>&#x02211;</mml:mo><mml:mi>i</mml:mi></mml:msub><mml:mrow><mml:msub><mml:mi>n</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:mstyle><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow><mml:mo>&#x02208;</mml:mo><mml:mstyle mathvariant="script"><mml:mi>M</mml:mi></mml:mstyle><mml:mo>(</mml:mo><mml:mi>X</mml:mi><mml:mo>)</mml:mo></mml:math></inline-formula> of size <inline-formula><mml:math id="M26"><mml:mi>n</mml:mi><mml:mo>=</mml:mo><mml:munder class="msub"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:munder><mml:msub><mml:mrow><mml:mi>n</mml:mi></mml:mrow><mml:mrow><mml:mi>i</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mo>|</mml:mo><mml:mo>|</mml:mo><mml:mi>&#x003C6;</mml:mi><mml:mo>|</mml:mo><mml:mo>|</mml:mo><mml:mo>&#x0003E;</mml:mo><mml:mn>0</mml:mn></mml:math></inline-formula> gives rise to a &#x0201C;fractional&#x0201D; distribution <inline-formula><mml:math id="M27"><mml:mi>fl</mml:mi><mml:mi>r</mml:mi><mml:mi>n</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>&#x003C6;</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x02208;</mml:mo><mml:mrow><mml:mstyle mathvariant="script"><mml:mi>D</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>X</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula>, given by:</p></list-item>
</list>
<disp-formula id="E12"><mml:math id="M28"><mml:mrow><mml:mi>f</mml:mi><mml:mi>l</mml:mi><mml:mi>r</mml:mi><mml:mi>n</mml:mi><mml:mo stretchy='true'>(</mml:mo><mml:mstyle displaystyle='true'><mml:msub><mml:mo>&#x02211;</mml:mo><mml:mi>i</mml:mi></mml:msub><mml:mrow><mml:msub><mml:mi>n</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow></mml:mstyle><mml:mo stretchy='true'>)</mml:mo><mml:mo>:</mml:mo><mml:mo>=</mml:mo><mml:mstyle displaystyle='true'><mml:msub><mml:mo>&#x02211;</mml:mo><mml:mi>i</mml:mi></mml:msub><mml:mrow><mml:mfrac><mml:mrow><mml:msub><mml:mi>n</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow><mml:mi>n</mml:mi></mml:mfrac><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow></mml:mstyle><mml:mtext>&#x000A0;&#x000A0;i.e.&#x000A0;</mml:mtext><mml:mi>f</mml:mi><mml:mi>l</mml:mi><mml:mi>r</mml:mi><mml:mi>n</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x003C6;</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>:</mml:mo><mml:mo>=</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>x</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mi>X</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:mfrac><mml:mrow><mml:mi>&#x003C6;</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mrow><mml:mo>&#x02016;</mml:mo><mml:mi>&#x003C6;</mml:mi><mml:mo>&#x02016;</mml:mo></mml:mrow></mml:mfrac></mml:mrow></mml:mstyle><mml:mrow><mml:mi>x</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>.</mml:mo><mml:mtext>&#x000A0;</mml:mtext></mml:mrow></mml:math></disp-formula>
<list list-type="simple">
<list-item><p>We write <inline-formula><mml:math id="M29"><mml:mrow><mml:mstyle mathvariant="script"><mml:mi>D</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>X</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> for the set of distributions on <italic>X</italic>. This <inline-formula><mml:math id="M30"><mml:mrow><mml:mstyle mathvariant="script"><mml:mi>D</mml:mi></mml:mstyle></mml:mrow></mml:math></inline-formula> is functorial, like <inline-formula><mml:math id="M31"><mml:mrow><mml:mstyle mathvariant="script"><mml:mi>M</mml:mi></mml:mstyle></mml:mrow></mml:math></inline-formula>, see Definition 1 (3): for a function <italic>f</italic>:<italic>X</italic>&#x02192;<italic>Y</italic> we write <inline-formula><mml:math id="M32"><mml:mrow><mml:mstyle mathvariant="script"><mml:mi>D</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>f</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>:</mml:mo><mml:mrow><mml:mstyle mathvariant="script"><mml:mi>D</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>X</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x02192;</mml:mo><mml:mrow><mml:mstyle mathvariant="script"><mml:mi>D</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>Y</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> for the function that produces the image distribution, as:</p></list-item>
</list>
<disp-formula id="E13"><label>(6)</label><mml:math id="M33"><mml:mrow><mml:mstyle mathvariant="script"><mml:mi>D</mml:mi></mml:mstyle><mml:mo stretchy='false'>(</mml:mo><mml:mi>f</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='true'>(</mml:mo><mml:mstyle displaystyle='true'><mml:msub><mml:mo>&#x02211;</mml:mo><mml:mi>i</mml:mi></mml:msub><mml:mrow><mml:msub><mml:mi>r</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow></mml:mstyle><mml:mo stretchy='true'>)</mml:mo><mml:mo>:</mml:mo><mml:mo>=</mml:mo><mml:mstyle displaystyle='true'><mml:msub><mml:mo>&#x02211;</mml:mo><mml:mi>i</mml:mi></mml:msub><mml:mrow><mml:msub><mml:mi>r</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x0007C;</mml:mo></mml:mrow></mml:mstyle><mml:mi>f</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo stretchy='false'>)</mml:mo><mml:mo>.</mml:mo></mml:mrow></mml:math></disp-formula>
<p>We apply the same conventions to distributions as formal sums of ket&#x00027;s as for multisets, so that terms 0|<italic>x</italic>&#x0232A; may be ommitted, <italic>etc</italic>.</p>
<p>Such fractional distributions <italic>flrn</italic>(&#x003C6;) have fractions as probabilities. We recall that the subset &#x0211A;&#x021AA;&#x0211D; is dense: each real number can be expressed as limit of fractions. An analogous situation applies to distributions. In order to formulate it we use the following <italic>total variation distance</italic> on distributions. It is a special case of the Kantorovich-Wasserstein distance (Kantorovich and Rubinshtein, <xref ref-type="bibr" rid="B26">1958</xref>). For two distributions <inline-formula><mml:math id="M35"><mml:mi>&#x003C9;</mml:mi><mml:mo>,</mml:mo><mml:msup><mml:mrow><mml:mi>&#x003C9;</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x02032;</mml:mi></mml:mrow></mml:msup><mml:mo>&#x02208;</mml:mo><mml:mrow><mml:mi mathvariant="script">D</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>X</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> one defines the distance <italic>d</italic>(&#x003C9;, &#x003C9;&#x02032;)&#x02208;[0, 1] as:</p>
<disp-formula id="E14"><label>(7)</label><mml:math id="M36"><mml:mrow><mml:mtable><mml:mtr><mml:mtd><mml:mrow><mml:mi>d</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x003C9;</mml:mi><mml:mo>,</mml:mo><mml:msup><mml:mi>&#x003C9;</mml:mi><mml:mo>&#x02032;</mml:mo></mml:msup><mml:mo stretchy='false'>)</mml:mo></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:mo>:</mml:mo><mml:mo>=</mml:mo></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:mfrac><mml:mn>1</mml:mn><mml:mn>2</mml:mn></mml:mfrac><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>x</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mi>X</mml:mi></mml:mrow></mml:munder></mml:mstyle><mml:mo>&#x0007C;</mml:mo><mml:mi>&#x003C9;</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x02212;</mml:mo><mml:msup><mml:mi>&#x003C9;</mml:mi><mml:mo>&#x02032;</mml:mo></mml:msup><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x0007C;</mml:mo><mml:mo>.</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:math></disp-formula>
<p>We can now formulate some basic topological properties of distributions. Stated informally, all distributions come from multisets.</p>
<p>Theorem 1. For a finite set <italic>X</italic>, the set <inline-formula><mml:math id="M37"><mml:mrow><mml:mi mathvariant="script">D</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>X</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> of distributions on <italic>X</italic>, with total variation distance (<xref ref-type="disp-formula" rid="E14">Equation 7</xref>), is a compact complete metric space, containing a countable dense subset of fractional distributions, given as image of frequentist learning <italic>flrn</italic>, from non-empty multisets to distributions.</p>
<p>There is another topic that we need to introduce, namely random variables, together with their expected value&#x02014;described here as validity.</p>
<p>Definition 3. Let <italic>X</italic> be an arbitrary set.</p>
<list list-type="simple">
<list-item><p>1. An <italic>observable</italic> on <italic>X</italic> is a function <italic>p</italic>:<italic>X</italic> &#x02192; &#x0211D;. These observables are closed under pointwise sum &#x0002B; and multiplication &#x00026;. There are the always-zero and always-one observables <bold>0</bold>, <bold>1</bold>:<italic>X</italic> &#x02192; &#x0211D;.</p></list-item>
<list-item><p>We write <italic>Obs</italic>(<italic>X</italic>): &#x0003D; &#x0211D;<sup><italic>X</italic></sup> for the vector space of observables on <italic>X</italic>.</p></list-item>
<list-item><p>2. A <italic>random variable</italic> is a pair (&#x003C9;, <italic>p</italic>) of a distribution <inline-formula><mml:math id="M38"><mml:mi>&#x003C9;</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mrow><mml:mi mathvariant="script">D</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>X</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> and an observable <italic>p</italic>:<italic>X</italic> &#x02192; &#x0211D;, on the same set <italic>X</italic>.</p></list-item>
<list-item><p>3. For a random variable <inline-formula><mml:math id="M39"><mml:mstyle mathsize="1.19em"><mml:mrow></mml:mrow></mml:mstyle><mml:mi>&#x003C9;</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mrow><mml:mi mathvariant="script">D</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>X</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>,</mml:mo><mml:mi>p</mml:mi><mml:mo>:</mml:mo><mml:mi>X</mml:mi><mml:mo>&#x02192;</mml:mo><mml:mi>&#x0211D;</mml:mi><mml:mstyle mathsize="1.19em"><mml:mrow></mml:mrow></mml:mstyle></mml:math></inline-formula> the <italic>expected value</italic> is written as validity &#x003C9;&#x022A7;<italic>p</italic> and defined as:</p></list-item>
</list>
<disp-formula id="E15"><label>(8)</label><mml:math id="M40"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mi>&#x003C9;</mml:mi><mml:mo>&#x022A7;</mml:mo><mml:mi>p</mml:mi></mml:mtd><mml:mtd><mml:mo>:</mml:mo><mml:mo>=</mml:mo></mml:mtd><mml:mtd><mml:mstyle displaystyle="true"><mml:munder class="msub"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>x</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mi>X</mml:mi></mml:mrow></mml:munder></mml:mstyle><mml:mi>&#x003C9;</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x000B7;</mml:mo><mml:mi>p</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>;</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>The expected value is commonly written as &#x1D53C;(<italic>p</italic>) with the distribution &#x003C9; left implicit. This may be inconvenient and confusing, especially when the distribution at hand may change, for instance in a computational setting. The notation &#x1D53C;<sub><italic>x</italic>&#x02190;&#x003C9;</sub><italic>p</italic>(<italic>x</italic>) fares better since it makes the distribution &#x003C9; explicit, but it introduces an additional bound variable, namely the <italic>x</italic> that is sampled from &#x003C9;. However, since the actual probability &#x003C9;(<italic>x</italic>) of the sampled element <italic>x</italic> does not occur in this expression &#x1D53C;<sub><italic>x</italic>&#x02190;&#x003C9;</sub><italic>p</italic>(<italic>x</italic>), it can not be used for calculations. Hence we prefer the (new) validity notation <inline-formula><mml:math id="M41"><mml:mi>&#x003C9;</mml:mi><mml:mo>&#x022A7;</mml:mo><mml:mi>p</mml:mi><mml:mo>=</mml:mo><mml:munder class="msub"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>x</mml:mi></mml:mrow></mml:munder><mml:mi>&#x003C9;</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x000B7;</mml:mo><mml:mi>p</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> for the expected value of observable <italic>p</italic> in distribution &#x003C9;.</p>
</sec>
<sec id="s4">
<title>4 Bayesian updating</title>
<p>There are two main schools in statistics, one of &#x0201C;frequentist&#x0201D; nature, assigning probabilities to data, and one of &#x0201C;Bayesian&#x0201D; kind, where probabilities are associated with hypotheses. The frequentistist approach is captured, in the discrete case, by distributions <inline-formula><mml:math id="M42"><mml:mrow><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mi>i</mml:mi></mml:munder><mml:mrow><mml:msub><mml:mi>r</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow></mml:mstyle><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:msub><mml:mi>x</mml:mi><mml:mi>i</mml:mi></mml:msub></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x02208;</mml:mo><mml:mi mathvariant="script">D</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>X</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow></mml:math></inline-formula>, with probabilities <italic>r</italic><sub><italic>i</italic></sub> associated with elements / objects / data item <italic>x</italic><sub><italic>i</italic></sub>&#x02208;<italic>X</italic>. The Bayesian approach may be formalized in terms of belief functions <italic>Obs</italic>(<italic>X</italic>) &#x02192; &#x0211D;, assigning values to observables / predicates / hypotheses / evidence. In Bayesian approaches these assignments are often called subjective, resulting from individual choices about how much value (like money) people wish to put on which possible outcomes.</p>
<p>The <xref ref-type="supplementary-material" rid="SM1">Appendix</xref> explores a mathematical perspective and describes an isomorphism (on the left in <xref ref-type="disp-formula" rid="E8">Equation 13</xref>, <xref ref-type="supplementary-material" rid="SM1">Appendix</xref>) that connects these Bayesian and frequentist approaches via a duality isomorphism <inline-formula><mml:math id="M43"><mml:mi>H</mml:mi><mml:mi>o</mml:mi><mml:mi>m</mml:mi><mml:mstyle mathsize="1.19em"><mml:mrow></mml:mrow></mml:mstyle><mml:mi>O</mml:mi><mml:mi>b</mml:mi><mml:mi>s</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>X</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>,</mml:mo><mml:mi>&#x0211D;</mml:mi><mml:mstyle mathsize="1.19em"><mml:mrow></mml:mrow></mml:mstyle><mml:mo>&#x02245;</mml:mo><mml:mrow><mml:mi mathvariant="script">D</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>X</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> between belief functions and distributions. Thus, one could say, the matter is solved, there is mathematically no difference between the two approaches&#x02014;up to isomorphism.</p>
<p>Conditional probabilities are typically developed on the Bayesian side, in terms of adapted belief functions. Using the approach of the <xref ref-type="supplementary-material" rid="SM1">Appendix</xref>, this approach can be pulled across the duality isomorphism, to the frequentist side. It leads to a form of updating in terms of adapted distributions &#x003C9;|<sub><italic>p</italic></sub>, see Section B in <xref ref-type="supplementary-material" rid="SM1">Appendix</xref> for the mathematical details. The next definition already formulates Bayesian updating of distributions with observables, in concrete form. It has been developed and used in a series of papers (Jacobs and Zanasi, <xref ref-type="bibr" rid="B22">2016</xref>, <xref ref-type="bibr" rid="B23">2017</xref>; Jacobs, <xref ref-type="bibr" rid="B13">2017a</xref>; Cho and Jacobs, <xref ref-type="bibr" rid="B6">2019</xref>; Jacobs, <xref ref-type="bibr" rid="B16">2019</xref>, <xref ref-type="bibr" rid="B17">2021</xref>, <xref ref-type="bibr" rid="B19">2024</xref>) aimed at systematizing probabilistic updating. It is with this formalization of Bayesian updating that we can clearly formulate and prove commutativity of updating, see Proposition 1 below.</p>
<p>Definition 4. Let <inline-formula><mml:math id="M44"><mml:mi>&#x003C9;</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mrow><mml:mi mathvariant="script">D</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>X</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> be a distribution with a non-negative observable <italic>p</italic>:<italic>X</italic> &#x02192; &#x0211D;&#x02265;0 such that the validity &#x003C9;&#x022A7;<italic>p</italic> is non-zero. In that case we define the <italic>Bayesian update</italic> <inline-formula><mml:math id="M45"><mml:mi>&#x003C9;</mml:mi><mml:mo>|</mml:mo><mml:mi>p</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mrow><mml:mi mathvariant="script">D</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>X</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> of &#x003C9; with &#x0201C;evidence&#x0201D; <italic>p</italic> as the normalized product:</p>
<disp-formula id="E16"><label>(9)</label><mml:math id="M46"><mml:mrow><mml:mtable><mml:mtr><mml:mtd><mml:mrow><mml:mi>&#x003C9;</mml:mi><mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mi>p</mml:mi></mml:msub></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:mo>:</mml:mo><mml:mo>=</mml:mo></mml:mrow></mml:mtd><mml:mtd><mml:mrow><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>x</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mi>X</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:mfrac><mml:mrow><mml:mi>&#x003C9;</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x000B7;</mml:mo><mml:mi>p</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mrow><mml:mi>&#x003C9;</mml:mi><mml:mo>&#x022A8;</mml:mo><mml:mi>p</mml:mi></mml:mrow></mml:mfrac></mml:mrow></mml:mstyle><mml:mrow><mml:mo>|</mml:mo><mml:mi>x</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>.</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:math></disp-formula>
<p>This formulation of Bayesian updating comes alive in illustrations. We present an example first and then show how the above formulation &#x003C9;|<italic>p</italic> generalizes the traditional formulation <italic>P</italic>(<italic>E</italic>&#x02223;<italic>D</italic>).</p>
<p>Example 1. We consider a study involving four common species of birds: robin (<italic>R</italic>), crow (<italic>C</italic>), sparrow (<italic>S</italic>), and woodpecker (<italic>W</italic>). We start from the following species distribution (in a particular area), on the set <italic>X</italic> &#x0003D; {<italic>R, C, S, W</italic>}.</p>
<disp-formula id="E17"><mml:math id="M47"><mml:mtable columnalign='left'><mml:mtr><mml:mtd><mml:mi>&#x003C3;</mml:mi><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mtext>&#x000A0;&#x000A0;</mml:mtext><mml:mfrac><mml:mn>1</mml:mn><mml:mn>4</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mi>R</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>3</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mi>C</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>4</mml:mn></mml:mfrac><mml:mi>S</mml:mi><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>6</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mi>W</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>&#x02248;</mml:mo><mml:mtext>&#x000A0;&#x000A0;</mml:mtext><mml:mn>0.25</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mi>R</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>0.333</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mi>C</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>0.25</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mi>S</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>0.167</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mi>W</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>This is our prior distribution. Then a day of bird counting happens, resulting in a count observable <italic>f</italic>:<italic>X</italic> &#x02192; &#x02115; with numbers:</p>
<disp-formula id="E18"><mml:math id="M48"><mml:mrow><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none none none none none none none none none none none none none none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mi>f</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>R</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mtd><mml:mtd><mml:mo>=</mml:mo></mml:mtd><mml:mtd><mml:mn>200</mml:mn></mml:mtd><mml:mtd><mml:mtext>&#x02003;</mml:mtext></mml:mtd><mml:mtd><mml:mi>f</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>C</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mtd><mml:mtd><mml:mo>=</mml:mo></mml:mtd><mml:mtd><mml:mn>150</mml:mn></mml:mtd><mml:mtd><mml:mtext>&#x02003;</mml:mtext></mml:mtd><mml:mtd><mml:mi>f</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>S</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mtd><mml:mtd><mml:mo>=</mml:mo></mml:mtd><mml:mtd><mml:mn>50</mml:mn></mml:mtd><mml:mtd><mml:mtext>&#x02003;</mml:mtext></mml:mtd><mml:mtd><mml:mi>f</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>W</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mtd><mml:mtd><mml:mo>=</mml:mo></mml:mtd><mml:mtd><mml:mn>100</mml:mn><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:math></disp-formula>
<p>We see that this observable does not really match the prior, for instance since the number of observed robins is higher than the number of crows, whereas the robin probability in &#x003C3; is lower than the crow probability. Also, the number of observed sparrows is low with respect to the woodpecker number. Hence we expect that updating &#x003C3; with <italic>f</italic> will lead to a considerable change of (relative) probabilities.</p>
<p>We first calculate the expected value, as validity:</p>
<disp-formula id="E19"><mml:math id="M49"><mml:mtable columnalign='left'><mml:mtr><mml:mtd><mml:mi>&#x003C3;</mml:mi><mml:mo>&#x022A8;</mml:mo><mml:mi>f</mml:mi><mml:mtext>&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mover><mml:mo>=</mml:mo><mml:mrow><mml:mo stretchy='false'>(</mml:mo><mml:mn>8</mml:mn><mml:mo stretchy='false'>)</mml:mo></mml:mrow></mml:mover><mml:mtext>&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>x</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mi>X</mml:mi></mml:mrow></mml:munder><mml:mi>&#x003C3;</mml:mi></mml:mstyle><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x000B7;</mml:mo><mml:mi>f</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mtext>&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mi>&#x003C3;</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>R</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x000B7;</mml:mo><mml:mi>f</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>R</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>+</mml:mo><mml:mi>&#x003C3;</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>C</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x000B7;</mml:mo><mml:mi>f</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>C</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>+</mml:mo><mml:mi>&#x003C3;</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>S</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x000B7;</mml:mo><mml:mi>f</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>S</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>+</mml:mo><mml:mi>&#x003C3;</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>W</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x000B7;</mml:mo><mml:mi>f</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>W</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mtext>&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mfrac><mml:mn>1</mml:mn><mml:mn>4</mml:mn></mml:mfrac><mml:mo>&#x000B7;</mml:mo><mml:mn>200</mml:mn><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>3</mml:mn></mml:mfrac><mml:mo>&#x000B7;</mml:mo><mml:mn>150</mml:mn><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>4</mml:mn></mml:mfrac><mml:mo>&#x000B7;</mml:mo><mml:mn>50</mml:mn><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>6</mml:mn></mml:mfrac><mml:mo>&#x000B7;</mml:mo><mml:mn>100</mml:mn></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mtext>&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mfrac><mml:mrow><mml:mn>775</mml:mn></mml:mrow><mml:mn>6</mml:mn></mml:mfrac><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>(<xref ref-type="disp-formula" rid="E17">8</xref>)</p>
<p>The updated, posterior distribution can now be computed, with this validity as normalization factor:</p>
<disp-formula id="E20"><mml:math id="M50"><mml:mtable columnalign='left'><mml:mtr><mml:mtd><mml:mi>&#x003C3;</mml:mi><mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mi>f</mml:mi></mml:msub><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mover><mml:mo>=</mml:mo><mml:mrow><mml:mo stretchy='false'>(</mml:mo><mml:mn>9</mml:mn><mml:mo stretchy='false'>)</mml:mo></mml:mrow></mml:mover><mml:mtext>&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>x</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mi>X</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:mfrac><mml:mrow><mml:mi>&#x003C3;</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x000B7;</mml:mo><mml:mi>f</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mrow><mml:mi>&#x003C3;</mml:mi><mml:mo>&#x022A8;</mml:mo><mml:mi>f</mml:mi></mml:mrow></mml:mfrac></mml:mrow></mml:mstyle><mml:mrow><mml:mo>|</mml:mo><mml:mi>x</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mtext>&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mfrac><mml:mrow><mml:mn>1</mml:mn><mml:mo>/</mml:mo><mml:mn>4</mml:mn><mml:mo>&#x000B7;</mml:mo><mml:mn>200</mml:mn></mml:mrow><mml:mrow><mml:mn>775</mml:mn><mml:mo>/</mml:mo><mml:mn>6</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mi>R</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mrow><mml:mn>1</mml:mn><mml:mo>/</mml:mo><mml:mn>3</mml:mn><mml:mo>&#x000B7;</mml:mo><mml:mn>150</mml:mn></mml:mrow><mml:mrow><mml:mn>775</mml:mn><mml:mo>/</mml:mo><mml:mn>6</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mi>C</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mrow><mml:mn>1</mml:mn><mml:mo>/</mml:mo><mml:mn>4</mml:mn><mml:mo>&#x000B7;</mml:mo><mml:mn>50</mml:mn></mml:mrow><mml:mrow><mml:mn>775</mml:mn><mml:mo>/</mml:mo><mml:mn>6</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mi>S</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mrow><mml:mn>1</mml:mn><mml:mo>/</mml:mo><mml:mn>6</mml:mn><mml:mo>&#x000B7;</mml:mo><mml:mn>100</mml:mn></mml:mrow><mml:mrow><mml:mn>775</mml:mn><mml:mo>/</mml:mo><mml:mn>6</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mi>W</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mtext>&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mfrac><mml:mrow><mml:mn>12</mml:mn></mml:mrow><mml:mrow><mml:mn>31</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mi>R</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mrow><mml:mn>12</mml:mn></mml:mrow><mml:mrow><mml:mn>31</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mi>C</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>3</mml:mn><mml:mrow><mml:mn>31</mml:mn></mml:mrow></mml:mfrac><mml:mi>S</mml:mi><mml:mo>+</mml:mo><mml:mfrac><mml:mn>4</mml:mn><mml:mrow><mml:mn>31</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mi>W</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>&#x02248;</mml:mo><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mn>0.387</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mi>R</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>0.387</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mi>C</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>0.0968</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mi>S</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>0.129</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mi>W</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>(<xref ref-type="disp-formula" rid="E18">9</xref>)</p>
<p>This posterior distribution &#x003C3;|<italic>f</italic> incorporates the evidence of the observable <italic>f</italic>. Its robin and crow probabilities are equal, and its woodpecker probability is higher than the sparrow probability, reflecting the count outcome.</p>
<p>A year later a new bird count happens, resulting in a new observable <italic>g</italic>:<italic>X</italic> &#x02192; &#x02115;, say with <italic>g</italic>(<italic>R</italic>) &#x0003D; 100, <italic>g</italic>(<italic>C</italic>) &#x0003D; 150, <italic>g</italic>(<italic>S</italic>) &#x0003D; 100, and <italic>g</italic>(<italic>W</italic>) &#x0003D; 50. This count is more in line with the prior. One can then update &#x003C3;|<italic>f</italic> once again, now with observable <italic>g</italic>&#x02014;last year&#x00027;s posterior is this year&#x00027;s prior. The resulting second update takes the form:</p>
<disp-formula id="E21"><mml:math id="M51"><mml:mtable columnalign='left'><mml:mtr><mml:mtd><mml:mi>&#x003C3;</mml:mi><mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mi>f</mml:mi></mml:msub><mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mi>g</mml:mi></mml:msub><mml:mtext>&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mtext>&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mfrac><mml:mrow><mml:mn>12</mml:mn></mml:mrow><mml:mrow><mml:mn>35</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mi>R</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mrow><mml:mn>18</mml:mn></mml:mrow><mml:mrow><mml:mn>35</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mi>C</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>3</mml:mn><mml:mrow><mml:mn>35</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mi>S</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>2</mml:mn><mml:mrow><mml:mn>35</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mi>W</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>&#x02248;</mml:mo><mml:mtext>&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mn>0.343</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mi>R</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>0.514</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mi>C</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>0.0857</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mi>S</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>0.0571</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mi>W</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>This second update brings us a bit closer to the original prior &#x003C3;.</p>
<p>Interestingly, this double update &#x003C3;|<italic>f</italic>|<italic>g</italic> is equal to the update &#x003C3;|<italic>g</italic>|<italic>f</italic> with swapped observables <italic>f, g</italic>. Thus, eventhough there is a clear order in the yearly bird counting, the mathematics of Bayesian updating ignores this order and produces the same outcome for both orders (<italic>f, g</italic>) and (<italic>g, f</italic>) of observables.</p>
<p>We now show how the conditional probability in traditional form fits into our form of Bayesian updating (9).</p>
<p>Lemma 1. Let <inline-formula><mml:math id="M52"><mml:mi>&#x003C9;</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mrow><mml:mi mathvariant="script">D</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>X</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> be a distribution, with a subset (event) <italic>E</italic>&#x02286;<italic>X</italic>. We write <bold>1</bold><italic>E</italic>:<italic>X</italic> &#x02192; &#x0211D; for the observable given by the indicator function of <italic>E</italic>, with associated validity:</p>
<disp-formula id="E22"><mml:math id="M53"><mml:mrow><mml:mstyle mathvariant="bold"><mml:mn>1</mml:mn></mml:mstyle><mml:mi>E</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>:</mml:mo><mml:mo>=</mml:mo><mml:mtext>&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mrow><mml:mo>{</mml:mo><mml:mrow><mml:mtable columnalign='left'><mml:mtr columnalign='left'><mml:mtd columnalign='left'><mml:mn>1</mml:mn></mml:mtd><mml:mtd columnalign='left'><mml:mrow><mml:mtext>if&#x000A0;</mml:mtext><mml:mi>x</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mi>E</mml:mi></mml:mrow></mml:mtd></mml:mtr><mml:mtr columnalign='left'><mml:mtd columnalign='left'><mml:mn>0</mml:mn></mml:mtd><mml:mtd columnalign='left'><mml:mrow><mml:mtext>if&#x000A0;</mml:mtext><mml:mi>x</mml:mi><mml:menclose notation='updiagonalstrike'><mml:mo>&#x02208;</mml:mo></mml:menclose><mml:mi>E</mml:mi></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:mrow><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;and&#x000A0;&#x000A0;</mml:mtext><mml:mi>P</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>E</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mtext>&#x000A0;</mml:mtext><mml:mo>:</mml:mo><mml:mo>=</mml:mo><mml:mtext>&#x000A0;&#x000A0;</mml:mtext><mml:mi>&#x003C9;</mml:mi><mml:mo>&#x022A8;</mml:mo><mml:msub><mml:mn>1</mml:mn><mml:mi>E</mml:mi></mml:msub><mml:mtext>&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mtext>&#x000A0;&#x000A0;</mml:mtext><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>x</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mi>E</mml:mi></mml:mrow></mml:munder><mml:mi>&#x003C9;</mml:mi></mml:mstyle><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>.</mml:mo></mml:mrow></mml:math></disp-formula>
<p>For another subset <italic>D</italic>&#x02286;<italic>X</italic> one has:</p>
<list list-type="simple">
<list-item><p>1. <bold>1</bold><italic>E&#x00026;</italic><bold>1</bold><italic>D</italic> &#x0003D; <bold>1</bold><italic>E</italic>&#x02229;<italic>D</italic>;</p></list-item>
<list-item><p>The conditional probability <italic>P</italic>(<italic>E</italic>&#x02223;<italic>D</italic>) is obtained as validity of <bold>1</bold><italic>E</italic> in the distribution &#x003C9; updated with <bold>1</bold><italic>D</italic>, that is, as:</p></list-item>
</list>
<disp-formula id="E23"><mml:math id="M54"><mml:mrow><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none none none none none none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mi>&#x003C9;</mml:mi><mml:mo>|</mml:mo><mml:mstyle mathvariant='bold-italic'><mml:mn>1</mml:mn></mml:mstyle><mml:mi>D</mml:mi><mml:mo>&#x022A7;</mml:mo><mml:mstyle mathvariant='bold-italic'><mml:mn>1</mml:mn></mml:mstyle><mml:mi>E</mml:mi></mml:mtd><mml:mtd><mml:mo>=</mml:mo></mml:mtd><mml:mtd><mml:mfrac><mml:mrow><mml:mi>&#x003C9;</mml:mi><mml:mo>&#x022A7;</mml:mo><mml:mstyle mathvariant='bold-italic'><mml:mn>1</mml:mn></mml:mstyle><mml:mi>E</mml:mi><mml:mi>&#x00026;</mml:mi><mml:mstyle mathvariant='bold-italic'><mml:mn>1</mml:mn></mml:mstyle><mml:mi>D</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003C9;</mml:mi><mml:mo>&#x022A7;</mml:mo><mml:mstyle mathvariant='bold-italic'><mml:mn>1</mml:mn></mml:mstyle><mml:mi>D</mml:mi></mml:mrow></mml:mfrac></mml:mtd><mml:mtd><mml:mo>=</mml:mo></mml:mtd><mml:mtd><mml:mfrac><mml:mrow><mml:mi>P</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>E</mml:mi><mml:mo>&#x02229;</mml:mo><mml:mi>D</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mi>P</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>D</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow></mml:mfrac></mml:mtd><mml:mtd><mml:mo>=</mml:mo><mml:mo>:</mml:mo></mml:mtd><mml:mtd><mml:mi>P</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>E</mml:mi><mml:mo>&#x02223;</mml:mo><mml:mi>D</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:math></disp-formula>
<p>This last equation defines the conditional probability <italic>P</italic>(<italic>E</italic>&#x02223;<italic>D</italic>).</p>
<p><bold>Proof</bold>. 1. For <italic>x</italic>&#x02208;<italic>X</italic>, using that &#x00026; is given by pointwise multiplication:</p>
<disp-formula id="E24"><mml:math id="M"><mml:mtable columnalign='left'><mml:mtr><mml:mtd><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mn>1</mml:mn><mml:mi>E</mml:mi></mml:msub><mml:mo>&#x00026;</mml:mo><mml:msub><mml:mn>1</mml:mn><mml:mi>D</mml:mi></mml:msub><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mn>1</mml:mn><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>&#x021D4;</mml:mo><mml:mtext>&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:msub><mml:mn>1</mml:mn><mml:mi>E</mml:mi></mml:msub><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x000B7;</mml:mo><mml:msub><mml:mn>1</mml:mn><mml:mi>D</mml:mi></mml:msub><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>&#x021D4;</mml:mo><mml:mtext>&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:msub><mml:mn>1</mml:mn><mml:mi>E</mml:mi></mml:msub><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mn>1</mml:mn><mml:mtext>&#x000A0;and&#x000A0;</mml:mtext><mml:msub><mml:mn>1</mml:mn><mml:mi>D</mml:mi></mml:msub><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>&#x021D4;</mml:mo><mml:mtext>&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mi>x</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mi>E</mml:mi><mml:mtext>&#x000A0;and&#x000A0;</mml:mtext><mml:mi>x</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mi>D</mml:mi></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>&#x021D4;</mml:mo><mml:mtext>&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mi>x</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mi>E</mml:mi><mml:mo>&#x02229;</mml:mo><mml:mi>D</mml:mi><mml:mo>&#x021D4;</mml:mo><mml:msub><mml:mn>1</mml:mn><mml:mrow><mml:mi>E</mml:mi><mml:mo>&#x02229;</mml:mo><mml:mi>D</mml:mi></mml:mrow></mml:msub><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mn>1.</mml:mn></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>(<xref ref-type="disp-formula" rid="E18">9</xref>)</p>
<list list-type="simple">
<list-item><p>2. We only have to prove the first equation, since the second one follows from the previous item. Thus:</p></list-item>
</list>
<disp-formula id="E25"><mml:math id="M56"><mml:mtable columnalign='left'><mml:mtr><mml:mtd><mml:mi>&#x003C9;</mml:mi><mml:mo>&#x0007C;</mml:mo><mml:msub><mml:mn>1</mml:mn><mml:mi>D</mml:mi></mml:msub><mml:mo>&#x022A8;</mml:mo><mml:msub><mml:mn>1</mml:mn><mml:mi>E</mml:mi></mml:msub><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mover><mml:mo>=</mml:mo><mml:mrow><mml:mo stretchy='false'>(</mml:mo><mml:mn>9</mml:mn><mml:mo stretchy='false'>)</mml:mo></mml:mrow></mml:mover><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>x</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mi>X</mml:mi></mml:mrow></mml:munder></mml:mstyle><mml:mi>&#x003C9;</mml:mi><mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:msub><mml:mn>1</mml:mn><mml:mi>D</mml:mi></mml:msub></mml:mrow></mml:msub><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x000B7;</mml:mo><mml:msub><mml:mn>1</mml:mn><mml:mi>E</mml:mi></mml:msub><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mover><mml:mo>=</mml:mo><mml:mrow><mml:mo stretchy='false'>(</mml:mo><mml:mn>9</mml:mn><mml:mo stretchy='false'>)</mml:mo></mml:mrow></mml:mover><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>x</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mi>X</mml:mi></mml:mrow></mml:munder></mml:mstyle><mml:mfrac><mml:mrow><mml:mi>&#x003C9;</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x000B7;</mml:mo><mml:msub><mml:mn>1</mml:mn><mml:mi>D</mml:mi></mml:msub><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mrow><mml:mi>&#x003C9;</mml:mi><mml:mo>&#x022A8;</mml:mo><mml:mi>D</mml:mi></mml:mrow></mml:mfrac><mml:mo>&#x000B7;</mml:mo><mml:msub><mml:mn>1</mml:mn><mml:mi>E</mml:mi></mml:msub><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mfrac><mml:mrow><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>x</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mi>X</mml:mi></mml:mrow></mml:munder></mml:mstyle><mml:mi>&#x003C9;</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x000B7;</mml:mo><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mn>1</mml:mn><mml:mi>D</mml:mi></mml:msub><mml:mo>&#x00026;</mml:mo><mml:msub><mml:mn>1</mml:mn><mml:mi>E</mml:mi></mml:msub><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mrow><mml:mi>P</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>D</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow></mml:mfrac></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>=</mml:mo><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mfrac><mml:mrow><mml:mi>P</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>E</mml:mi><mml:mo>&#x02229;</mml:mo><mml:mi>D</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mrow><mml:mo stretchy='false'>(</mml:mo><mml:mi>D</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow></mml:mfrac><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>The traditional <italic>P</italic>(&#x02212;) notation leaves the distribution implicit, which has many disadvantages. Most relevant in this context is that this <italic>P</italic>(&#x02212;) notation makes it impossible to express the commutativity of Bayesian updating, as formulated in Proposition 1 below.</p>
<p>We include another illustration that combines several of the topics that we discussed earlier: multisets, functoriality (for marginalization), and updating.</p>
<p>Example 2. Consider the following situation and questions, describing a typical update situation with observations about offspring.<xref ref-type="fn" rid="fn0001"><sup>1</sup></xref></p>
<p>A friend of mine has three children aged 4 and 5 with one twin.</p>
<list list-type="simple">
<list-item><p>(a) What is the probability that there are three girls, assuming that the probability of a girl is <inline-formula><mml:math id="M57"><mml:mfrac><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:mfrac></mml:math></inline-formula>? 2.</p></list-item>
<list-item><p>(b) I ring this friend&#x00027;s doorbell and I hear a girl&#x00027;s voice say that she will open the door soon. If I can assume that this is one of the three children, what is the probability that my friend has three daughters? 3.</p></list-item>
<list-item><p>(c) A 4-year-old boy opens the door. Still assuming this is one of the children, what is the probability that there are three boys?</p></list-item>
</list>
<p>Questions (b) and (c) are independent.</p>
<p>We address this situation in terms of multisets of children, as in <xref ref-type="disp-formula" rid="E3">Equation 3</xref>. The situation is more complicated now since we have to use a set <italic>C</italic> &#x0003D; {<italic>B, G</italic>} for the (sex of the) children but also a set <italic>A</italic> &#x0003D; {4, 5} for their ages. The possible offspring configurations are (certain) multisets of size 3 over the product set <italic>C</italic>&#x000D7;<italic>A</italic>. After a moment&#x00027;s thought we see that the prior &#x003C5; is a distribution of the following form.</p>
<disp-formula id="E26"><label>(10)</label><mml:math id="M58"><mml:mtable columnalign='left'><mml:mtr><mml:mtd><mml:mi>&#x003C5;</mml:mi><mml:mo>=</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>16</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>2</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>16</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x02009;&#x02009;&#x02009;&#x02009;</mml:mtext><mml:mo>+</mml:mo><mml:mtext>&#x02009;</mml:mtext><mml:mfrac><mml:mn>1</mml:mn><mml:mn>8</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>16</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>2</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x02009;&#x02009;&#x02009;&#x02009;</mml:mtext><mml:mo>+</mml:mo><mml:mtext>&#x02009;</mml:mtext><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>16</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>8</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x02009;&#x02009;&#x02009;&#x02009;</mml:mtext><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>16</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>2</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>8</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x02009;&#x02009;&#x02009;&#x02009;</mml:mtext><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>16</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>2</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>8</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x02009;&#x02009;&#x02009;&#x02009;</mml:mtext><mml:mo>+</mml:mo><mml:mtext>&#x02009;</mml:mtext><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>16</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>16</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>This &#x003C5; is a distribution over multisets, using nested kets. The outer, big kets are for the probabilities, with inside the different offspring configurations in the form of a multiset. For instance, the multisets 1|<italic>B</italic>, 4&#x0232A; &#x0002B; 2|<italic>G</italic>, 5&#x0232A; and 1|<italic>G</italic>, 4&#x0232A; &#x0002B; 2|<italic>G</italic>, 5&#x0232A; in the last line capture the situations with one boy (or girl) of 4 and two girls of 5 years old.<xref ref-type="fn" rid="fn0002"><sup>2</sup></xref></p>
<p>We shall write <italic>S</italic>: &#x0003D; <italic>supp</italic>(&#x003C5;) for the support of this distribution &#x003C5;. This set <italic>S</italic> contains all of the 12 different multisets &#x003C6; inside the big kets in <xref ref-type="disp-formula" rid="E26">Equation 10</xref>.</p>
<p>For the first question (a) we ask ourselves more generally what the children distribution is in this situation. It can be obtained by discarding the ages, via the first marginal of the multisets inside the big kets. This involves applying the marginalization function <inline-formula><mml:math id="M60"><mml:mrow><mml:mi mathvariant="script">M</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>&#x003C0;</mml:mi></mml:mrow><mml:mrow><mml:mn>1</mml:mn></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>:</mml:mo><mml:mrow><mml:mi mathvariant="script">M</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>C</mml:mi><mml:mo>&#x000D7;</mml:mo><mml:mi>A</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x02192;</mml:mo><mml:mrow><mml:mi mathvariant="script">M</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>C</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> to these multisets, see Definition 1 3. Since we wish to apply this marginalization function <inline-formula><mml:math id="M61"><mml:mrow><mml:mi mathvariant="script">M</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>&#x003C0;</mml:mi></mml:mrow><mml:mrow><mml:mn>1</mml:mn></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> inside the bigkets, we have to use functoriality of <inline-formula><mml:math id="M62"><mml:mrow><mml:mi mathvariant="script">D</mml:mi></mml:mrow></mml:math></inline-formula> as well, see Definition 2 3. Thus, the distribution of children marginals is obtained as:</p>
<disp-formula id="E27"><mml:math id="M63"><mml:mtable columnalign='left'><mml:mtr><mml:mtd><mml:mi mathvariant="script">D</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi mathvariant="script">M</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>&#x003C0;</mml:mi><mml:mn>1</mml:mn></mml:msub><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x003C5;</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>&#x003C6;</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mi>S</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:mi>&#x003C5;</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x003C6;</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mi mathvariant="script">M</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>&#x003C0;</mml:mi><mml:mn>1</mml:mn></mml:msub><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x003C6;</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>&#x003C6;</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mi>S</mml:mi></mml:mrow></mml:munder><mml:mrow><mml:mi>&#x003C5;</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x003C6;</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mstyle displaystyle='true'><mml:msub><mml:mo>&#x02211;</mml:mo><mml:mrow><mml:mi>x</mml:mi><mml:mo>,</mml:mo><mml:mi>y</mml:mi></mml:mrow></mml:msub><mml:mi>&#x003C6;</mml:mi></mml:mstyle><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo>,</mml:mo><mml:mi>y</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x0007C;</mml:mo><mml:mi>x</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo stretchy='true'>&#x0232A;</mml:mo></mml:mrow></mml:mrow></mml:mstyle></mml:mrow></mml:mstyle></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;</mml:mtext><mml:mo>=</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>16</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>3</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>16</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>3</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>8</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;</mml:mtext><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>16</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>16</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;</mml:mtext><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>8</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>16</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;</mml:mtext><mml:mo>+</mml:mo><mml:mtext>&#x02009;</mml:mtext><mml:mfrac><mml:mn>1</mml:mn><mml:mn>8</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>16</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>3</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>8</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;</mml:mtext><mml:mo>+</mml:mo><mml:mtext>&#x02009;</mml:mtext><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>16</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>16</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;</mml:mtext><mml:mo>=</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>8</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>3</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>3</mml:mn><mml:mn>8</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;</mml:mtext><mml:mo>+</mml:mo><mml:mfrac><mml:mn>3</mml:mn><mml:mn>8</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>8</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>3</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>(<xref ref-type="disp-formula" rid="E1">6</xref>)</p>
<p>(<xref ref-type="disp-formula" rid="E1">4</xref>)</p>
<p>The answer to question (a) is thus <inline-formula><mml:math id="M64"><mml:mfrac><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mn>8</mml:mn></mml:mrow></mml:mfrac></mml:math></inline-formula>, the probability associated in the last line with the three girls multiset 3|<italic>G</italic>&#x0232A;.</p>
<p>The interested reader may wish to check that taking the second marginals yields the (expected) age distribution of the form:</p>
<disp-formula id="E28"><mml:math id="M65"><mml:mrow><mml:mi mathvariant="script">D</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi mathvariant="script">M</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>&#x003C0;</mml:mi><mml:mn>2</mml:mn></mml:msub><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x003C5;</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>2</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mn>4</mml:mn><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mn>5</mml:mn><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>2</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mn>4</mml:mn><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mn>5</mml:mn><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>.</mml:mo></mml:mrow></mml:math></disp-formula>
<p>For question (b) we define an observable <italic>g</italic> : <italic>S</italic> &#x02192; {0, 1} which is 1 if and only if there is at least one girl:</p>
<disp-formula id="E29"><mml:math id="M66"><mml:mrow><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none none none none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mi>g</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>&#x003C6;</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:mtd><mml:mtd><mml:mo>&#x021D4;</mml:mo></mml:mtd><mml:mtd><mml:mrow><mml:mi mathvariant="script">M</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>&#x003C0;</mml:mi></mml:mrow><mml:mrow><mml:mn>1</mml:mn></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>&#x003C6;</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>g</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x02265;</mml:mo><mml:mn>1</mml:mn></mml:mtd><mml:mtd><mml:mo>&#x021D4;</mml:mo></mml:mtd><mml:mtd><mml:mi>&#x003C6;</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>g</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x0002B;</mml:mo><mml:mi>&#x003C6;</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>g</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x02265;</mml:mo><mml:mn>1</mml:mn><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:math></disp-formula>
<p>This observable <italic>g</italic> is {0, 1}-valued and may be identified with a subset of <italic>S</italic>, as in Lemma 1. Updating &#x003C5; with <italic>g</italic> involves removing the multisets &#x003C6; &#x02208; <italic>S</italic> with <italic>g</italic>(&#x003C6;) &#x0003D; 0, that is, with boys only, and then renormalising. The normalization factor is the validity:</p>
<disp-formula id="E30"><mml:math id="M67"><mml:mrow><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none none none none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mi>&#x003C5;</mml:mi><mml:mo>&#x022A7;</mml:mo><mml:mi>g</mml:mi></mml:mtd><mml:mtd><mml:mo>=</mml:mo></mml:mtd><mml:mtd><mml:mstyle displaystyle="true"><mml:munder class="msub"><mml:mrow><mml:mo>&#x02211;</mml:mo></mml:mrow><mml:mrow><mml:mi>&#x003C6;</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mi>S</mml:mi></mml:mrow></mml:munder></mml:mstyle><mml:mi>&#x003C5;</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>&#x003C6;</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x000B7;</mml:mo><mml:mi>g</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>&#x003C6;</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mtd><mml:mtd><mml:mo>=</mml:mo></mml:mtd><mml:mtd><mml:mfrac><mml:mrow><mml:mn>7</mml:mn></mml:mrow><mml:mrow><mml:mn>8</mml:mn></mml:mrow></mml:mfrac></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:math></disp-formula>
<p>The answer to question (b) is obtained by computing the update &#x003C5;|<sub><italic>g</italic></sub> and taking its children marginal, as before. This yields:</p>
<disp-formula id="E31"><mml:math id="M68"><mml:mtable columnalign='left'><mml:mtr><mml:mtd><mml:mi mathvariant="script">D</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi mathvariant="script">M</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>&#x003C0;</mml:mi><mml:mn>1</mml:mn></mml:msub><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x003C5;</mml:mi><mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mi>g</mml:mi></mml:msub><mml:mo stretchy='false'>)</mml:mo></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x02009;&#x02009;</mml:mtext><mml:mo>=</mml:mo><mml:mi mathvariant="script">D</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi mathvariant="script">M</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>&#x003C0;</mml:mi><mml:mn>1</mml:mn></mml:msub><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>(</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>7</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;</mml:mtext><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>14</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>14</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mtext>&#x02009;</mml:mtext></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;</mml:mtext><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>14</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>7</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mtext>&#x02009;</mml:mtext></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;</mml:mtext><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>7</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>7</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn><mml:mo>&#x0232A;</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;</mml:mtext><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>14</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mtext>&#x02009;</mml:mtext><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>14</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;</mml:mtext><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>14</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>4</mml:mn><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>,</mml:mo><mml:mn>5</mml:mn><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo stretchy='true'>)</mml:mo></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x02009;&#x02009;&#x02009;</mml:mtext><mml:mo>=</mml:mo><mml:mfrac><mml:mn>3</mml:mn><mml:mn>7</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>3</mml:mn><mml:mn>7</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>7</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>3</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>We can conclude that after seeing one girl the probability that there are three girls has risen from <inline-formula><mml:math id="M69"><mml:mfrac><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mn>8</mml:mn></mml:mrow></mml:mfrac></mml:math></inline-formula> to <inline-formula><mml:math id="M70"><mml:mfrac><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mn>7</mml:mn></mml:mrow></mml:mfrac></mml:math></inline-formula>. As an aside: the distribution of age marginals remains the same after this update.</p>
<p>What happens when we see a 4-year old boy? We capture this via an event / observable <italic>b</italic><sub>4</sub> : <italic>S</italic> &#x02192; {0, 1} with <italic>b</italic><sub>4</sub>(&#x003C6;) &#x0003D; 1 iff &#x003C6;(<italic>B</italic>, 4) &#x02265; 1. Its validity &#x003C5; &#x022A7; <italic>b</italic><sub>4</sub> in the prior distribution &#x003C5; is <inline-formula><mml:math id="M71"><mml:mfrac><mml:mrow><mml:mn>7</mml:mn></mml:mrow><mml:mrow><mml:mn>12</mml:mn></mml:mrow></mml:mfrac></mml:math></inline-formula>. We leave it to the interested reader to verify that the distributions of children / age marginals, after update with <italic>b</italic><sub>4</sub>, are:</p>
<disp-formula id="E32"><mml:math id="M72"><mml:mrow><mml:mtable><mml:mtr><mml:mtd><mml:mrow><mml:mi mathvariant="script">D</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi mathvariant="script">M</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>&#x003C0;</mml:mi><mml:mn>1</mml:mn></mml:msub><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x003C5;</mml:mi><mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:msub><mml:mi>b</mml:mi><mml:mn>4</mml:mn></mml:msub></mml:mrow></mml:msub><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mrow><mml:mfrac><mml:mn>1</mml:mn><mml:mn>5</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>3</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mrow><mml:mn>10</mml:mn></mml:mrow></mml:mfrac><mml:mo>&#x0007C;</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mrow><mml:mtext>&#x000A0;</mml:mtext></mml:mrow><mml:mrow><mml:mtext>&#x000A0;</mml:mtext></mml:mrow><mml:mrow><mml:mtext>&#x02003;&#x02003;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>+</mml:mo><mml:mfrac><mml:mn>3</mml:mn><mml:mrow><mml:mn>10</mml:mn></mml:mrow></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mrow><mml:mi mathvariant="script">D</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi mathvariant="script">M</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>&#x003C0;</mml:mi><mml:mn>2</mml:mn></mml:msub><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x003C5;</mml:mi><mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:msub><mml:mi>b</mml:mi><mml:mn>4</mml:mn></mml:msub></mml:mrow></mml:msub><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mrow><mml:mfrac><mml:mn>3</mml:mn><mml:mn>5</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mn>4</mml:mn><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mn>5</mml:mn><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>2</mml:mn><mml:mn>5</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mn>4</mml:mn><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mn>5</mml:mn><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>.</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:math></disp-formula>
<p>The first equation answers question (c): the probability of three boys is <inline-formula><mml:math id="M73"><mml:mfrac><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mn>5</mml:mn></mml:mrow></mml:mfrac></mml:math></inline-formula>, having seen one 4-year old boy. It is higher than the probability of seeing three girls, given that there is at least one girl? The boy-of-4 observation excludes more cases and the remaining cases thus get higher probability, after re-normalization.</p>
<p>The second equation about the age marginals shows that the configuration with two 4-year olds is more likely, after seeing at least one 4-year old (boy). This makes sense.</p>
<p>We can still ask what we can infer if we have seen both a girl and a boy-of-4. As before the order of updating is irrelevant: &#x003C5;|<italic>g</italic>|<italic>b</italic><sub>4</sub> &#x0003D; &#x003C5;|<italic>b</italic><sub>4</sub>|<italic>g</italic>. In that situation there are 5 multisets left, out of the original 12, in &#x003C5; in <xref ref-type="disp-formula" rid="E26">Equation 10</xref>. The distributions of marginals are:</p>
<disp-formula id="E33"><mml:math id="M74"><mml:mrow><mml:mtable><mml:mtr><mml:mtd><mml:mrow><mml:mi mathvariant="script">D</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi mathvariant="script">M</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>&#x003C0;</mml:mi><mml:mn>1</mml:mn></mml:msub><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x003C5;</mml:mi><mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mi>g</mml:mi></mml:msub><mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:msub><mml:mi>b</mml:mi><mml:mn>4</mml:mn></mml:msub></mml:mrow></mml:msub><mml:mo stretchy='false'>)</mml:mo></mml:mrow></mml:mtd><mml:mtd><mml:mo>=</mml:mo></mml:mtd><mml:mtd><mml:mrow><mml:mfrac><mml:mn>5</mml:mn><mml:mn>8</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>3</mml:mn><mml:mn>8</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>B</mml:mi><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mi>G</mml:mi><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mrow><mml:mi mathvariant="script">D</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi mathvariant="script">M</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>&#x003C0;</mml:mi><mml:mn>2</mml:mn></mml:msub><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>(</mml:mo><mml:mi>&#x003C5;</mml:mi><mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mi>g</mml:mi></mml:msub><mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mrow><mml:msub><mml:mi>b</mml:mi><mml:mn>4</mml:mn></mml:msub></mml:mrow></mml:msub><mml:mo stretchy='false'>)</mml:mo></mml:mrow></mml:mtd><mml:mtd><mml:mo>=</mml:mo></mml:mtd><mml:mtd><mml:mrow><mml:mfrac><mml:mn>5</mml:mn><mml:mn>8</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mn>4</mml:mn><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mn>5</mml:mn><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>+</mml:mo><mml:mfrac><mml:mn>3</mml:mn><mml:mn>8</mml:mn></mml:mfrac><mml:mrow><mml:mo>|</mml:mo><mml:mrow><mml:mn>1</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mn>4</mml:mn><mml:mo>&#x0232A;</mml:mo><mml:mo>+</mml:mo><mml:mn>2</mml:mn><mml:mo>&#x0007C;</mml:mo><mml:mn>5</mml:mn><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>&#x0232A;</mml:mo></mml:mrow><mml:mo>.</mml:mo></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:math></disp-formula>
<p>We conclude by proving in general what we have already seen several times, namely that multiple Bayesian updates commute (<xref ref-type="disp-formula" rid="E1">Equation 1</xref>). We do so by using the conjunction <italic>p &#x00026; q</italic> (pointwise multiplication) of observables, in order to emphasize the close connection between commutativity of conjunction and of updating. The result below already occurs in Jacobs (<xref ref-type="bibr" rid="B16">2019</xref>, Lem. 4.1), together with a generalized formulation of Bayes&#x00027; rule for observables. A proof is included for completeness.</p>
<p>Proposition 1. Let <inline-formula><mml:math id="M75"><mml:mi>&#x003C9;</mml:mi><mml:mo>&#x02208;</mml:mo><mml:mrow><mml:mi mathvariant="script">D</mml:mi></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>X</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:math></inline-formula> be a distribution with two non-negative observables <italic>p, q</italic> : <italic>X</italic> &#x02192; &#x0211D; &#x02265; 0. Then, assuming that the relevant validities are non-zero,</p>
<disp-formula id="E34"><mml:math id="M76"><mml:mrow><mml:mi>&#x003C9;</mml:mi><mml:msub><mml:mo stretchy="false">&#x0007C;</mml:mo><mml:mi>p</mml:mi></mml:msub><mml:msub><mml:mo stretchy="false">&#x0007C;</mml:mo><mml:mi>q</mml:mi></mml:msub><mml:mo>=</mml:mo><mml:mi>&#x003C9;</mml:mi><mml:msub><mml:mo stretchy="false">&#x0007C;</mml:mo><mml:mrow><mml:mi>p</mml:mi><mml:mo>&#x00026;</mml:mo><mml:mi>q</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mi>&#x003C9;</mml:mi><mml:msub><mml:mo stretchy="false">&#x0007C;</mml:mo><mml:mrow><mml:mi>q</mml:mi><mml:mo>&#x00026;</mml:mo><mml:mi>p</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mi>&#x003C9;</mml:mi><mml:msub><mml:mo stretchy="false">&#x0007C;</mml:mo><mml:mi>q</mml:mi></mml:msub><mml:msub><mml:mo stretchy="false">&#x0007C;</mml:mo><mml:mi>p</mml:mi></mml:msub><mml:mo>.</mml:mo></mml:mrow></mml:math></disp-formula>
<p><bold>Proof</bold>. We only have to prove the first equation, since the commutativity of &#x00026; is obvious (multiplication of numbers is commutative) and the last equation is an instance of the first (with <italic>p, q</italic> swapped). Using the functional description for distributions, we have for <italic>x</italic> &#x02208; <italic>X</italic>,</p>
<disp-formula id="E37"><mml:math id="M77"><mml:mtable columnalign='left'><mml:mtr><mml:mtd><mml:mi>&#x003C9;</mml:mi><mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mi>p</mml:mi></mml:msub><mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mi>q</mml:mi></mml:msub><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mi>&#x003C9;</mml:mi><mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mi>p</mml:mi></mml:msub><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x000B7;</mml:mo><mml:mi>q</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mrow><mml:mi>&#x003C9;</mml:mi><mml:msub><mml:mo>&#x0007C;</mml:mo><mml:mi>p</mml:mi></mml:msub><mml:mo>&#x022A8;</mml:mo><mml:mi>q</mml:mi></mml:mrow></mml:mfrac></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;</mml:mtext><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mfrac><mml:mrow><mml:mi>&#x003C9;</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x000B7;</mml:mo><mml:mi>p</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mrow><mml:mi>&#x003C9;</mml:mi><mml:mo>&#x022A8;</mml:mo><mml:mi>p</mml:mi></mml:mrow></mml:mfrac><mml:mo>&#x000B7;</mml:mo><mml:mi>q</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mrow><mml:mstyle displaystyle='true'><mml:msub><mml:mo>&#x02211;</mml:mo><mml:mi>y</mml:mi></mml:msub><mml:mrow><mml:mfrac><mml:mrow><mml:mi>&#x003C9;</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>y</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x000B7;</mml:mo><mml:mi>p</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>y</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mrow><mml:mi>&#x003C9;</mml:mi><mml:mo>&#x022A8;</mml:mo><mml:mi>p</mml:mi></mml:mrow></mml:mfrac></mml:mrow></mml:mstyle><mml:mo>&#x000B7;</mml:mo><mml:mi>q</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>y</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow></mml:mfrac></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;&#x02009;</mml:mtext><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mi>&#x003C9;</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>&#x022C5;</mml:mo><mml:mo stretchy='false'>(</mml:mo><mml:mi>p</mml:mi><mml:mo>&#x00026;</mml:mo><mml:mi>q</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mrow><mml:mi>&#x003C9;</mml:mi><mml:mo>&#x022A8;</mml:mo><mml:mi>p</mml:mi><mml:mo>&#x00026;</mml:mo><mml:mi>q</mml:mi></mml:mrow></mml:mfrac><mml:mo>=</mml:mo><mml:msub><mml:mi>&#x003C9;</mml:mi><mml:mrow><mml:mi>p</mml:mi><mml:mo>&#x00026;</mml:mo><mml:mi>q</mml:mi></mml:mrow></mml:msub><mml:mo stretchy='false'>(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy='false'>)</mml:mo><mml:mo>.</mml:mo><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>&#x025A1;</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>(<xref ref-type="disp-formula" rid="E1">9</xref>)</p>
</sec>
<sec id="s5">
<title>5 Concluding remarks</title>
<p>The Bayesian approach is popular in cognition theory, where the human mind is seen as a Bayesian prediction and inference engine, see for instance the recent books (Griffiths et al., <xref ref-type="bibr" rid="B12">2024</xref>; Parr et al., <xref ref-type="bibr" rid="B30">2022</xref>). In that line of work the mismatch caused by the commutativity of Bayesian updating does not get much attention. It is however known in the literature, see notably (Uzan, <xref ref-type="bibr" rid="B36">2023</xref>). One way out is to switch from classical to quantum probability, where conjunction and updating are non-commutative. This has led to a new line of &#x0201C;quantum&#x0201D; cognition theory (see e.g. Busemeyer and Bruza, <xref ref-type="bibr" rid="B3">2012</xref>; Yearsley and Busemeyer, <xref ref-type="bibr" rid="B37">2016</xref>; or Jacobs, <xref ref-type="bibr" rid="B14">2017b</xref> which is similar in style to this article).</p>
<p>When we take the commutativity of Bayesian updating seriously, the proper data structure to deal with multiple updates is: a multiset of observables. Indeed, as we have seen in Section 2, multisets abstract from lists by ignoring the order. This perspective is elaborated in Jacobs (<xref ref-type="bibr" rid="B19">2024</xref>), where the different update mechanisms of Pearl and Jeffrey (Jacobs, <xref ref-type="bibr" rid="B16">2019</xref>), and also the variational free update mechanism from predictive coding (Friston, <xref ref-type="bibr" rid="B8">2009</xref>; Tull et al., <xref ref-type="bibr" rid="B35">2023</xref>), are formulated in terms of such multisets of observables. Jeffrey&#x00027;s rule is non-commutative, but in a special way, namely for multiple such (non-singleton) multisets. All this suggests that the topic of commutativity may be a decisive element in further developing probabilistic perspectives in cognition and in AI.</p></sec>
</body>
<back>
<sec sec-type="data-availability" id="s6">
<title>Data availability statement</title>
<p>The original contributions presented in the study are included in the article/supplementary material, further inquiries can be directed to the corresponding author.</p>
</sec>
<sec sec-type="author-contributions" id="s7">
<title>Author contributions</title>
<p>BJ: Writing &#x02013; original draft, Methodology, Formal analysis, Investigation, Conceptualization, Writing &#x02013; review &#x00026; editing.</p>
</sec>
<sec sec-type="funding-information" id="s8">
<title>Funding</title>
<p>The author(s) declare that no financial support was received for the research and/or publication of this article.</p>
</sec>
<sec sec-type="COI-statement" id="conf1">
<title>Conflict of interest</title>
<p>The author declares that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec sec-type="ai-statement" id="s9">
<title>Generative AI statement</title>
<p>The author(s) declare that no Gen AI was used in the creation of this manuscript.</p></sec>
<sec sec-type="disclaimer" id="s10">
<title>Publisher&#x00027;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
<sec sec-type="supplementary-material" id="s11">
<title>Supplementary material</title>
<p>The Supplementary Material for this article can be found online at: <ext-link ext-link-type="uri" xlink:href="https://www.frontiersin.org/articles/10.3389/fcogn.2025.1623227/full#supplementary-material">https://www.frontiersin.org/articles/10.3389/fcogn.2025.1623227/full#supplementary-material</ext-link></p>
<supplementary-material xlink:href="Supplementary_file_1.pdf" id="SM1" mimetype="application/pdf" xmlns:xlink="http://www.w3.org/1999/xlink"/></sec>
<fn-group>
<fn id="fn0001"><p><sup>1</sup>See the special page <ext-link ext-link-type="uri" xlink:href="https://en.wikipedia.org/wiki/Boy_or_girl_paradox">https://en.wikipedia.org/wiki/Boy_or_girl_paradox</ext-link>.</p></fn>
<fn id="fn0002"><p><sup>2</sup>One can obtain the distribution &#x003C5; in <xref ref-type="disp-formula" rid="E26">Equation 10</xref> itself via conditioning, namely from the multinomial distribution, of draws of size three from the uniform distribution on <italic>C</italic>&#x000D7;<italic>A</italic>. One updates this multinomial distribution by keeping only those multiset &#x003C6; in which both ages occur, that is, for which the support <inline-formula><mml:math id="M59"><mml:mi>s</mml:mi><mml:mi>u</mml:mi><mml:mi>p</mml:mi><mml:mi>p</mml:mi><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mrow><mml:mstyle mathvariant="script"><mml:mi>M</mml:mi></mml:mstyle></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:msub><mml:mrow><mml:mi>&#x003C0;</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msub></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mi>&#x003C6;</mml:mi></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:mrow><mml:mo stretchy="false">)</mml:mo></mml:mrow><mml:mo>&#x02286;</mml:mo><mml:mi>A</mml:mi></mml:math></inline-formula> of the second marginal of &#x003C6; has two elements. This construction of &#x003C5; distracts from the main line, so we decided to simply present the relevant prior distribution in <xref ref-type="disp-formula" rid="E26">Equation 10</xref>.</p></fn>
</fn-group>
<ref-list>
<title>References</title>
<ref id="B1">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Alfsen</surname> <given-names>E.</given-names></name></person-group> (<year>1971</year>). <source>Compact Convex Sets and Boundary Integrals, Volume 57 of Ergebnisse der Mathematik und ihrer Grenzgebiete</source>. <publisher-loc>Cham</publisher-loc>: <publisher-name>Springer</publisher-name>.</citation>
</ref>
<ref id="B2">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Awodey</surname> <given-names>S.</given-names></name></person-group> (<year>2006</year>). <source>Category Theory. Oxford Logic Guides</source>. <publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>.<pub-id pub-id-type="pmid">34130967</pub-id></citation></ref>
<ref id="B3">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Busemeyer</surname> <given-names>J.</given-names></name> <name><surname>Bruza</surname> <given-names>P.</given-names></name></person-group> (<year>2012</year>). <source>Quantum Models of Cognition and Decision</source>. <publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>.</citation>
</ref>
<ref id="B4">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chaput</surname> <given-names>P.</given-names></name> <name><surname>Danos</surname> <given-names>V.</given-names></name> <name><surname>Panangaden</surname> <given-names>P.</given-names></name> <name><surname>Plotkin</surname> <given-names>G.</given-names></name></person-group> (<year>2014</year>). <article-title>Approximating Markov processes by averaging</article-title>. <source>J. ACM</source>, <volume>61</volume>, <fpage>1</fpage>&#x02013;<lpage>45</lpage>. <pub-id pub-id-type="doi">10.1145/2537948</pub-id></citation>
</ref>
<ref id="B5">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Cheng</surname> <given-names>E.</given-names></name></person-group> (<year>2022</year>). <source>The Joy of Abstraction. An Exploration of Math, Category Theory, and Life</source>. <publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>.</citation>
</ref>
<ref id="B6">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cho</surname> <given-names>K.</given-names></name> <name><surname>Jacobs</surname> <given-names>B.</given-names></name></person-group> (<year>2019</year>). <article-title>Disintegration and Bayesian inversion via string diagrams</article-title>. <source>Math. Struct. Comp. Sci</source>. <volume>29</volume>, <fpage>938</fpage>&#x02013;<lpage>971</lpage>. <pub-id pub-id-type="doi">10.1017/S0960129518000488</pub-id></citation>
</ref>
<ref id="B7">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Conway</surname> <given-names>J.</given-names></name></person-group> (<year>1990</year>). <source>A Course in Functional Analysis. Graduate Texts in Mathematics 96</source>, 2nd Edn. Cham: Springer.</citation>
</ref>
<ref id="B8">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Friston</surname> <given-names>K.</given-names></name></person-group> (<year>2009</year>). <article-title>The free-energy principle: a rough guide to the brain?</article-title> <source>Trends Cogn. Sci</source>. <volume>13</volume>, <fpage>293</fpage>&#x02013;<lpage>301</lpage>. <pub-id pub-id-type="doi">10.1016/j.tics.2009.04.005</pub-id><pub-id pub-id-type="pmid">19559644</pub-id></citation></ref>
<ref id="B9">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fritz</surname> <given-names>T.</given-names></name></person-group> (<year>2020</year>). <article-title>A synthetic approach to Markov kernels, conditional independence, and theorems on sufficient statistics</article-title>. <source>Adv. Math</source>. <volume>370</volume>:<fpage>107239</fpage>. <pub-id pub-id-type="doi">10.1016/j.aim.2020.107239</pub-id></citation>
</ref>
<ref id="B10">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gentili</surname> <given-names>P.</given-names></name></person-group> (<year>2021</year>). <article-title>Establishing a new link between fuzzy logic, neuroscience, and quantum mechanics through Bayesian probability: perspectives in artificial intelligence and unconventional computing</article-title>. <source>Molecules</source> <volume>26</volume>:<fpage>5987</fpage>. <pub-id pub-id-type="doi">10.3390/molecules26195987</pub-id><pub-id pub-id-type="pmid">34641530</pub-id></citation></ref>
<ref id="B11">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gigerenzer</surname> <given-names>G.</given-names></name> <name><surname>Hoffrage</surname> <given-names>U.</given-names></name></person-group> (<year>1995</year>). <article-title>How to improve Bayesian reasoning without instruction: frequency formats</article-title>. <source>Psychol. Rev</source>. <volume>102</volume>, <fpage>684</fpage>&#x02013;<lpage>704</lpage>.</citation>
</ref>
<ref id="B12">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Griffiths</surname> <given-names>T.</given-names></name> <name><surname>Chater</surname> <given-names>N.</given-names></name> <name><surname>Tenenbaum</surname> <given-names>J.</given-names></name></person-group> (<year>2024</year>). <source>Bayesian Models of Cognition. Reverse Engineering the Mind</source>. <publisher-loc>Cambridge, MA</publisher-loc>: <publisher-name>MIT Press</publisher-name>.<pub-id pub-id-type="pmid">25898807</pub-id></citation></ref>
<ref id="B13">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jacobs</surname> <given-names>B.</given-names></name></person-group> (<year>2017a</year>). <article-title>Hyper normalisation and conditioning for discrete probability distributions</article-title>. <source>Log. Methods Comp. Sci</source>. 13. <pub-id pub-id-type="doi">10.23638/LMCS-13(3:17)2017</pub-id></citation>
</ref>
<ref id="B14">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jacobs</surname> <given-names>B.</given-names></name></person-group> (<year>2017b</year>). <article-title>Quantum effect logic in cognition</article-title>. <source>J. Math. Psychol</source>. <volume>81</volume>, <fpage>1</fpage>&#x02013;<lpage>10</lpage>. <pub-id pub-id-type="doi">10.1016/j.jmp.2017.08.004</pub-id></citation>
</ref>
<ref id="B15">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jacobs</surname> <given-names>B.</given-names></name></person-group> (<year>2017c</year>). <article-title>A recipe for state and effect triangles</article-title>. <source>Log. Methods Comp. Sci</source>. <volume>13</volume>, <fpage>1</fpage>&#x02013;<lpage>26</lpage>. <pub-id pub-id-type="doi">10.23638/LMCS-13(2:6)2017</pub-id><pub-id pub-id-type="pmid">38430550</pub-id></citation></ref>
<ref id="B16">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jacobs</surname> <given-names>B.</given-names></name></person-group> (<year>2019</year>). <article-title>The mathematics of changing one&#x00027;s mind, via Jeffrey&#x00027;s or via Pearl&#x00027;s update rule</article-title>. <source>J. Artif. Intell. Res</source>. <volume>65</volume>, <fpage>783</fpage>&#x02013;<lpage>806</lpage>. <pub-id pub-id-type="doi">10.1613/jair.1.11349</pub-id></citation>
</ref>
<ref id="B17">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jacobs</surname> <given-names>B.</given-names></name></person-group> (<year>2021</year>). <article-title>&#x0201C;Learning from what&#x00027;s right and learning from what&#x00027;s wrong,&#x0201D;</article-title> in <source>Mathematical Foundations of Programming Semantics, number 351 in Elect. Proc. in Theor. Comp. Sci</source>., ed. A. <volume>Sokolova</volume>, <fpage>116</fpage>&#x02013;<lpage>133</lpage>. <pub-id pub-id-type="doi">10.4204/EPTCS.351.8</pub-id></citation>
</ref>
<ref id="B18">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jacobs</surname> <given-names>B.</given-names></name></person-group> (<year>2022</year>). <article-title>Urns &#x00026; tubes</article-title>. <source>Compositionality</source> 4. <pub-id pub-id-type="doi">10.32408/compositionality-4-4</pub-id></citation>
</ref>
<ref id="B19">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jacobs</surname> <given-names>B.</given-names></name></person-group> (<year>2024</year>). <article-title>Getting wiser from multiple data: probabilistic updating according to Jeffrey and Pearl</article-title>. <source>arXiv</source> [Preprint] arXiv:2405.12700. <pub-id pub-id-type="doi">10.48550/arXiv.2405.12700</pub-id></citation>
</ref>
<ref id="B20">
<citation citation-type="web"><person-group person-group-type="author"><name><surname>Jacobs</surname> <given-names>B.</given-names></name></person-group> (<year>2025</year>). <source>Structured Probabilitistic Reasoning</source>. Available online at: <ext-link ext-link-type="uri" xlink:href="http://www.cs.ru.nl/B.Jacobs/PAPERS/ProbabilisticReasoning.pdf">http://www.cs.ru.nl/B.Jacobs/PAPERS/ProbabilisticReasoning.pdf</ext-link> (Accessed June 2025).</citation>
</ref>
<ref id="B21">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jacobs</surname> <given-names>B.</given-names></name> <name><surname>Mandemaker</surname> <given-names>J.</given-names></name> <name><surname>Furber</surname> <given-names>R.</given-names></name></person-group> (<year>2016</year>). <article-title>The expectation monad in quantum foundations</article-title>. <source>Inf. Comput</source>. <volume>250</volume>, <fpage>87</fpage>&#x02013;<lpage>114</lpage>. <pub-id pub-id-type="doi">10.1016/j.ic.2016.02.009</pub-id></citation>
</ref>
<ref id="B22">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Jacobs</surname> <given-names>B.</given-names></name> <name><surname>Zanasi</surname> <given-names>F.</given-names></name></person-group> (<year>2016</year>). <article-title>&#x0201C;A predicate/state transformer semantics for Bayesian learning,&#x0201D;</article-title> in <source>Math. Found. of Programming Semantics, number 325 in Elect. Notes in Theor. Comp. Sci</source>., ed. L. Birkedal (<publisher-loc>Amsterdam</publisher-loc>: <publisher-name>Elsevier</publisher-name>), <fpage>185</fpage>&#x02013;<lpage>200</lpage>.</citation>
</ref>
<ref id="B23">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jacobs</surname> <given-names>B.</given-names></name> <name><surname>Zanasi</surname> <given-names>F.</given-names></name></person-group> (<year>2017</year>). <article-title>&#x0201C;A formal semantics of influence in Bayesian reasoning,&#x0201D;</article-title> in <source>Math. Found. of Computer Science, Volume 83 of LIPIcs</source>, eds. K. Larsen, H. Bodlaender, and J.-F. Raskin (Wadern: Schloss Dagstuhl), <volume>21</volume>:<fpage>1</fpage>&#x02013;<lpage>21</lpage>:14.<pub-id pub-id-type="pmid">23953959</pub-id></citation></ref>
<ref id="B24">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Johnson</surname> <given-names>N.</given-names></name> <name><surname>Kotz</surname> <given-names>S.</given-names></name></person-group> (<year>1977</year>). <source>Urn Models and Their Application: An Approach to Modern Discrete Probability Theory</source>. <publisher-loc>Hoboken, NJ</publisher-loc>: <publisher-name>John Wiley</publisher-name>.</citation>
</ref>
<ref id="B25">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Johnstone</surname> <given-names>P.</given-names></name></person-group> (<year>1982</year>). <source>Stone Spaces. Number 3 in Cambridge Studies in Advanced Mathematics</source>. <publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>.</citation>
</ref>
<ref id="B26">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kantorovich</surname> <given-names>L.</given-names></name> <name><surname>Rubinshtein</surname> <given-names>G.</given-names></name></person-group> (<year>1958</year>). <article-title>On a space of totally additive functions</article-title>. <source>Vestnik Leningrad Univ</source>. <volume>13</volume>, <fpage>52</fpage>&#x02013;<lpage>59</lpage>.</citation>
</ref>
<ref id="B27">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Leinster</surname> <given-names>T.</given-names></name></person-group> (<year>2014</year>). <source>Basic Category Theory. Cambridge Studies in Advanced Mathematics</source>. Cambridge: Cambridge University Press. <pub-id pub-id-type="doi">10.48550/arXiv.1612.09375</pub-id></citation>
</ref>
<ref id="B28">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Mahmoud</surname> <given-names>H.</given-names></name></person-group> (<year>2008</year>). <source>P&#x000F3;lya Urn Models</source>. <publisher-loc>Boca Raton, FL</publisher-loc>: <publisher-name>Chapman and Hall</publisher-name>.</citation>
</ref>
<ref id="B29">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Panangaden</surname> <given-names>P.</given-names></name></person-group> (<year>2009</year>). <source>Labelled Markov Processes</source>. <publisher-loc>London</publisher-loc>: <publisher-name>Imperial College Press</publisher-name>.</citation>
</ref>
<ref id="B30">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Parr</surname> <given-names>T.</given-names></name> <name><surname>Pezzulo</surname> <given-names>G.</given-names></name> <name><surname>Friston</surname> <given-names>K.</given-names></name></person-group> (<year>2022</year>). <source>Active Inference. The Free Energy Principle in Mind, Brain, and Behavior</source>. <publisher-loc>Cambridge, MA</publisher-loc>: <publisher-name>MIT Press</publisher-name>.</citation>
</ref>
<ref id="B31">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Perrone</surname> <given-names>P.</given-names></name></person-group> (<year>2024</year>). <source>Starting Category Theory</source>. <publisher-loc>Singapore</publisher-loc>: <publisher-name>World Scientific</publisher-name>.</citation>
</ref>
<ref id="B32">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Pierce</surname> <given-names>B.</given-names></name></person-group> (<year>1991</year>). <source>Basic Category Theory for Computer Scientists</source>. <publisher-loc>Cambridge, MA</publisher-loc>: <publisher-name>MIT Press</publisher-name>.</citation>
</ref>
<ref id="B33">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ross</surname> <given-names>S.</given-names></name></person-group> (<year>2018</year>). <source>A First Course in Probability</source>, 10th Edn. Upper Saddle River, NJ: Pearson Education.</citation>
</ref>
<ref id="B34">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Simons</surname> <given-names>H.</given-names></name></person-group> (<year>2011</year>). <source>An Introduction to Category Theory</source>. <publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>.</citation>
</ref>
<ref id="B35">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tull</surname> <given-names>S.</given-names></name> <name><surname>Kleiner</surname> <given-names>J.</given-names></name> <name><surname>Smithe</surname> <given-names>T. S. C.</given-names></name></person-group> (<year>2023</year>). <article-title>Active inference in string diagrams: a categorical account of predictive processing and free energy</article-title>. <source>arXiv</source> [Preprint] arXiv.2308.00861. <pub-id pub-id-type="doi">10.48550/arXiv.2308.00861</pub-id></citation>
</ref>
<ref id="B36">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Uzan</surname> <given-names>P.</given-names></name></person-group> (<year>2023</year>). <article-title>Bayesian rationality revisited: integrating order effects</article-title>. <source>Found. Sci</source>. <volume>28</volume>, <fpage>507</fpage>&#x02013;<lpage>528</lpage>. <pub-id pub-id-type="doi">10.1007/s10699-022-09838-0</pub-id></citation>
</ref>
<ref id="B37">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Yearsley</surname> <given-names>J.</given-names></name> <name><surname>Busemeyer</surname> <given-names>J.</given-names></name></person-group> (<year>2016</year>). <article-title>Quantum cognition and decision theories: a tutorial</article-title>. <source>J. Math. Psychol</source>. <volume>74</volume>, <fpage>99</fpage>&#x02013;<lpage>116</lpage>.</citation>
</ref>
</ref-list>
</back>
</article>