<?xml version="1.0" encoding="UTF-8" standalone="no"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xml:lang="EN" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Neurorobot.</journal-id>
<journal-title>Frontiers in Neurorobotics</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Neurorobot.</abbrev-journal-title>
<issn pub-type="epub">1662-5218</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fnbot.2022.896229</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Neuroscience</subject>
<subj-group>
<subject>Hypothesis and Theory</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Reclaiming saliency: Rhythmic precision-modulated action and perception</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name><surname>Anil Meera</surname> <given-names>Ajith</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="corresp" rid="c001"><sup>&#x0002A;</sup></xref>
<xref ref-type="author-notes" rid="fn001"><sup>&#x02020;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/1717011/overview"/>
</contrib>
<contrib contrib-type="author" corresp="yes">
<name><surname>Novicky</surname> <given-names>Filip</given-names></name>
<xref ref-type="aff" rid="aff2"><sup>2</sup></xref>
<xref ref-type="corresp" rid="c002"><sup>&#x0002A;</sup></xref>
<xref ref-type="author-notes" rid="fn001"><sup>&#x02020;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/1720876/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Parr</surname> <given-names>Thomas</given-names></name>
<xref ref-type="aff" rid="aff3"><sup>3</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/499268/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Friston</surname> <given-names>Karl</given-names></name>
<xref ref-type="aff" rid="aff3"><sup>3</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/20407/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Lanillos</surname> <given-names>Pablo</given-names></name>
<xref ref-type="aff" rid="aff4"><sup>4</sup></xref>
<xref ref-type="author-notes" rid="fn002"><sup>&#x02021;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/476110/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Sajid</surname> <given-names>Noor</given-names></name>
<xref ref-type="aff" rid="aff3"><sup>3</sup></xref>
<xref ref-type="author-notes" rid="fn002"><sup>&#x02021;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/1112823/overview"/>
</contrib>
</contrib-group>
<aff id="aff1"><sup>1</sup><institution>Department of Cognitive Robotics, Faculty of Mechanical, Maritime and Materials Engineering, Delft University of Technology</institution>, <addr-line>Delft</addr-line>, <country>Netherlands</country></aff>
<aff id="aff2"><sup>2</sup><institution>Department of Neurophysiology, Donders Institute for Brain Cognition and Behavior, Radboud University</institution>, <addr-line>Nijmegen</addr-line>, <country>Netherlands</country></aff>
<aff id="aff3"><sup>3</sup><institution>Wellcome Centre for Human Neuroimaging, University College London</institution>, <addr-line>London</addr-line>, <country>United Kingdom</country></aff>
<aff id="aff4"><sup>4</sup><institution>Department of Artificial Intelligence, Donders Institute for Brain Cognition and Behavior, Radboud University</institution>, <addr-line>Nijmegen</addr-line>, <country>Netherlands</country></aff>
<author-notes>
<fn fn-type="edited-by"><p>Edited by: Adam Safron, Johns Hopkins Medicine, United States</p></fn>
<fn fn-type="edited-by"><p>Reviewed by: Valerio Santangelo, University of Perugia, Italy; Nicholas Huang, Biofourmis Inc., United States</p></fn>
<corresp id="c001">&#x0002A;Correspondence: Ajith Anil Meera <email>a.anilmeera&#x00040;tudelft.nl</email></corresp>
<corresp id="c002">Filip Novicky <email>filip.novicky&#x00040;donders.ru.nl</email></corresp>
<fn fn-type="equal" id="fn001"><p>&#x02020;These authors have contributed equally to this work and share first authorship</p></fn>
<fn fn-type="equal" id="fn002"><p>&#x02021;These authors have contributed equally to this work</p></fn></author-notes>
<pub-date pub-type="epub">
<day>28</day>
<month>07</month>
<year>2022</year>
</pub-date>
<pub-date pub-type="collection">
<year>2022</year>
</pub-date>
<volume>16</volume>
<elocation-id>896229</elocation-id>
<history>
<date date-type="received">
<day>14</day>
<month>03</month>
<year>2022</year>
</date>
<date date-type="accepted">
<day>28</day>
<month>06</month>
<year>2022</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x000A9; 2022 Anil Meera, Novicky, Parr, Friston, Lanillos and Sajid.</copyright-statement>
<copyright-year>2022</copyright-year>
<copyright-holder>Anil Meera, Novicky, Parr, Friston, Lanillos and Sajid</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/"><p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p></license>
</permissions>
<abstract>
<p>Computational models of visual attention in artificial intelligence and robotics have been inspired by the concept of a saliency map. These models account for the mutual information between the (current) visual information and its estimated causes. However, they fail to consider the circular causality between perception and action. In other words, they do not consider where to sample next, given current beliefs. Here, we reclaim salience as an active inference process that relies on two basic principles: uncertainty minimization and rhythmic scheduling. For this, we make a distinction between attention and salience. Briefly, we associate attention with precision control, i.e., the confidence with which beliefs can be updated given sampled sensory data, and salience with uncertainty minimization that underwrites the selection of future sensory data. Using this, we propose a new account of attention based on rhythmic precision-modulation and discuss its potential in robotics, providing numerical experiments that showcase its advantages for state and noise estimation, system identification and action selection for informative path planning.</p></abstract>
<kwd-group>
<kwd>attention</kwd>
<kwd>saliency</kwd>
<kwd>free-energy principle</kwd>
<kwd>active inference</kwd>
<kwd>precision</kwd>
<kwd>brain-inspired robotics</kwd>
<kwd>cognitive robotics</kwd>
</kwd-group>
<counts>
<fig-count count="10"/>
<table-count count="2"/>
<equation-count count="7"/>
<ref-count count="134"/>
<page-count count="0"/>
<word-count count="14714"/>
</counts>
</article-meta>
</front>
<body>
<sec sec-type="intro" id="s1">
<title>1. Introduction</title>
<p>Attention is a fundamental cognitive ability that determines which events from the environment, and the body, are preferentially processed (Itti and Koch, <xref ref-type="bibr" rid="B54">2001</xref>). For example, the motor system directs the visual sensory stream by orienting the fovea centralis (i.e., the retinal region of highest visual acuity) toward points of interest within the visual scene. Thus, the confidence with which the causes of sampled visual information are inferred is constrained by the physical structure of the eye&#x02014;and eye movements are necessary to minimize uncertainty about visual percepts (Ahnelt, <xref ref-type="bibr" rid="B1">1998</xref>). In neuroscience, this can be attributed to two distinct, but highly interdependent attentional processes: (<italic>i</italic>) attentional gain mechanisms reliant on estimating the sensory precision of current data (Feldman and Friston, <xref ref-type="bibr" rid="B25">2010</xref>; Yang et al., <xref ref-type="bibr" rid="B133">2016a</xref>), and (<italic>ii</italic>) attentional salience that involves actively engaging with the sensorium to sample appropriate future data (Lengyel et al., <xref ref-type="bibr" rid="B73">2016</xref>; Parr and Friston, <xref ref-type="bibr" rid="B98">2019</xref>). Here we refer to perceptual-related salience, i.e., processing of low-level visual information (Santangelo, <xref ref-type="bibr" rid="B118">2015</xref>). Put simply, we formalize the fundamental difference between attention&#x02014;as optimizing perceptual processing&#x02014;and salience as optimizing the sampling of what is processed. This highlights the dynamic, circular nature with which biological agents acquire, and process, sensory information.</p>
<p>Understanding the computational mechanisms that undergird these two attentional phenomena is pertinent for deploying apt models of (visual) perception in artificial agents (Klink et al., <xref ref-type="bibr" rid="B60">2014</xref>; Mousavi et al., <xref ref-type="bibr" rid="B85">2016</xref>; Atrey et al., <xref ref-type="bibr" rid="B4">2019</xref>) and robots (Frintrop and Jensfelt, <xref ref-type="bibr" rid="B33">2008</xref>; Begum and Karray, <xref ref-type="bibr" rid="B8">2010</xref>; Ferreira and Dias, <xref ref-type="bibr" rid="B26">2014</xref>; Lanillos et al., <xref ref-type="bibr" rid="B67">2015a</xref>). Previous computational models of visual attention, used in artificial intelligence and robotics, have been inspired (and limited) by the feature integration theory proposed by Treisman and Gelade (<xref ref-type="bibr" rid="B126">1980</xref>) and the concept of a saliency map (Tsotsos et al., <xref ref-type="bibr" rid="B127">1995</xref>; Itti and Koch, <xref ref-type="bibr" rid="B54">2001</xref>; Borji and Itti, <xref ref-type="bibr" rid="B11">2012</xref>). Briefly, a saliency map is a static two-dimensional &#x02018;image&#x00027; that encodes stimulus relevance, e.g., the importance of particular region. These maps are then used to isolate relevant information for control (e.g., to direct foveation of the maximum valued region). Accordingly, computational models reliant on this formulation do not consider the circular-dependence between action selection and cue relevance&#x02014;and simply use these static saliency maps to guide action.</p>
<p>In this article, we adopt a first principles account to disambiguate the computational mechanisms that underpin attention and salience (Parr and Friston, <xref ref-type="bibr" rid="B98">2019</xref>) and provide a new account of attention. Specifically, our formulation can be effectively implemented for robotic systems and facilitates both state-estimation and action selection. For this, we associate attention with precision control, i.e., the confidence with which beliefs can be updated given (current) sampled sensory data. Salience is associated with uncertainty minimization that influences the selection of future sensory data. This formulation speaks to a computational distinction between action selection (i.e., where to look next) and visual sampling (i.e., what information is being processed). Importantly, recent evidence demonstrates the rhythmic nature of these processes <italic>via</italic> a theta-cycle coupling that fluctuates between high and low precision&#x02014;as unpacked in Section 2. From a robotics perspective, resolving uncertainty about states of affair speaks to a form of Bayesian optimality, in which decisions are made to maximize expected information gain (Lindley, <xref ref-type="bibr" rid="B74">1956</xref>; Friston et al., <xref ref-type="bibr" rid="B37">2021</xref>; Sajid et al., <xref ref-type="bibr" rid="B114">2021a</xref>). The duality between attention and salience is important for resolving uncertainty and enabling active perception. Significantly, it addresses an important challenge for defining autonomous robotics systems that can balance optimally between data assimilation (i.e., confidently perceiving current observations) and exploratory behavior to maximize information gain (Bajcsy et al., <xref ref-type="bibr" rid="B6">2018</xref>).</p>
<p>In what follows, we review the neuroscience of attention and salience (Section 2) to develop a novel (computational) account of attention based on precision-modulation that underwrites perception and action (Section 3). Next, we face-validate our formulation within a robotics context using numerical experiments (Section 4). The robotics implementation instantiates a free energy principle (FEP) approach to information processing (Friston, <xref ref-type="bibr" rid="B35">2010</xref>). This allows us to modulate the (appropriate) precision parameters to solve relevant robotics challenges in perception and control; namely, state-estimation (Section 4.2.2), system identification (Section 4.2.3), planning (Section 4.3), and active perception (Section 4.3.3). We conclude with a discussion of the requisite steps for instantiating a full-fledged computational model of precision-modulated attention&#x02014;and its implications in a robotics setting.</p></sec>
<sec id="s2">
<title>2. Attention and salience in neuroscience</title>
<p>Our interactions with the world are guided by efficient gathering and processing of sensory information. The quality of these acquired sensory data is reflected in attentional resources that select sensations which influence our beliefs about the (current and future) states of affairs (Lengyel et al., <xref ref-type="bibr" rid="B73">2016</xref>; Yang et al., <xref ref-type="bibr" rid="B134">2016b</xref>). This selection is often related to gain control, i.e., an increase of neural spikes when an object is attended to. However, gain control only accounts for half the story because we can only attend to those objects that are within our visual field. Accordingly, if a salient object is outside the center of our visual field, we orient the fovea to points of interest. This involves two separate, but often conflated, processes: attention and salience&#x02014;where the former relates to processing current visual data, and the latter to ensuring the agent samples salient data in the future (Parr and Friston, <xref ref-type="bibr" rid="B98">2019</xref>). That these two processes are strongly coupled is exemplified by the pre-motor theory of attention (Rizzolatti et al., <xref ref-type="bibr" rid="B110">1987</xref>), which highlights the close relationship between overt saccadic sampling of the visual field and the covert deployment of attention in the absence of eye movements. Specifically, it posits that covert attention<xref ref-type="fn" rid="fn0001"><sup>1</sup></xref> is realized <italic>via</italic> processes that are generated by particular eye movements but inhibits the action itself. In this sense, it does not distinguish between covert and overt<xref ref-type="fn" rid="fn0002"><sup>2</sup></xref> types of attention.</p>
<p>From a first principles (Bayesian) account, it is necessary to separate between attention and salience because they speak to different optimization processes. Explicitly, attention as a precision-dependent (neural) gain control mechanism that facilitates optimization of the <italic>current</italic> sampled sensory data (Desimone, <xref ref-type="bibr" rid="B21">1996</xref>; Feldman and Friston, <xref ref-type="bibr" rid="B25">2010</xref>). Conversely, salience is associated with selection of <italic>future</italic> data that reduces uncertainty (Friston et al., <xref ref-type="bibr" rid="B40">2015</xref>; Mirza et al., <xref ref-type="bibr" rid="B83">2016</xref>; Parr and Friston, <xref ref-type="bibr" rid="B98">2019</xref>). Put simply, it is possible to optimize attention in the absence of eye movements and active vision, whereas salience is necessary to optimize the deployment of eye movements. In what follows, we formalize this distinction with a particular focus on visual attention (Kanwisher and Wojciulik, <xref ref-type="bibr" rid="B56">2000</xref>), and discuss recent findings that speak to a rhythmic coupling that underwrites periodic deployment of gain control and saccades, <italic>via</italic> modulation of distinct precision parameters.</p>
<sec>
<title>2.1. Attention as neural gain control</title>
<p>Neural gain control can be regarded as an amplifier of neural communication during attention tasks (Reynolds et al., <xref ref-type="bibr" rid="B109">2000</xref>; Eldar et al., <xref ref-type="bibr" rid="B24">2013</xref>). Computationally, this is analogous to modulating a precision term, or the inverse temperature parameter (Feldman and Friston, <xref ref-type="bibr" rid="B25">2010</xref>; Parr and Friston, <xref ref-type="bibr" rid="B96">2017a</xref>). For this reason, we refer to precision and gain control interchangeably. An increase in gain amplifies the postsynaptic responses of neurons to their pre-synaptic input. Thus, gain control rests on synaptic modulation that can emphasize&#x02014;or preferentially select&#x02014;a particular type of sensory data. From a Bayesian perspective (Rao, <xref ref-type="bibr" rid="B105">2005</xref>; Spratling, <xref ref-type="bibr" rid="B124">2008</xref>; Parr et al., <xref ref-type="bibr" rid="B94">2018</xref>), this speaks to the confidence with which beliefs can be updated given sampled sensory data (i.e., optimal state estimation)&#x02014;under a generative model (Whiteley and Sahani, <xref ref-type="bibr" rid="B132">2008</xref>; Parr et al., <xref ref-type="bibr" rid="B94">2018</xref>). For example, affording high precision to certain sensory inputs would lead to confident Bayesian belief updating. However, low precision reduces the influence of sensory input by attenuating the precision of the likelihood, relative to a prior belief, and current observations would do little to resolve ensuing uncertainty. Thus, sampled visual data (from different areas) can be predicted with varying levels of precision, where attention accentuates sensory precision. The deployment of precision or attention is influenced by competition between stimuli (i.e., which sensory data to sample) and prior beliefs. Interestingly, casting attention as precision or, equivalently, synaptic gain offers a coherency between biased competition (Desimone, <xref ref-type="bibr" rid="B21">1996</xref>), predictive coding (Spratling, <xref ref-type="bibr" rid="B124">2008</xref>) and generic active inference schemes (Feldman and Friston, <xref ref-type="bibr" rid="B25">2010</xref>; Brown et al., <xref ref-type="bibr" rid="B13">2013</xref>; Kanai et al., <xref ref-type="bibr" rid="B55">2015</xref>; Parr et al., <xref ref-type="bibr" rid="B94">2018</xref>).</p>
<p>Naturally, gain control is accompanied by neuronal variability, i.e., sharpened neural responses for the same task over time. Consistent with gain control, these fluctuations in neural responses across trials can be explained by precision engineered message passing (Clark, <xref ref-type="bibr" rid="B18">2013</xref>) <italic>via</italic> (<italic>i</italic>) normalization models (Reynolds and Heeger, <xref ref-type="bibr" rid="B108">2009</xref>; Ruff and Cohen, <xref ref-type="bibr" rid="B113">2016</xref>), (<italic>ii</italic>) temperature parameter manipulation (Feldman and Friston, <xref ref-type="bibr" rid="B25">2010</xref>; Parr and Friston, <xref ref-type="bibr" rid="B96">2017a</xref>; Parr et al., <xref ref-type="bibr" rid="B94">2018</xref>, <xref ref-type="bibr" rid="B95">2019</xref>; Mirza et al., <xref ref-type="bibr" rid="B82">2019</xref>), or (<italic>iii</italic>) introduction of (conjugate hyper-)priors that are either pre-specified (Sajid et al., <xref ref-type="bibr" rid="B116">2020</xref>, <xref ref-type="bibr" rid="B115">2021b</xref>) or optimized using uninformed priors (Friston et al., <xref ref-type="bibr" rid="B42">2003</xref>; Anil Meera and Wisse, <xref ref-type="bibr" rid="B3">2021</xref>). Recently, these approaches have been used to simulate attention by accentuating predictions about a given visual stimulus (Reynolds and Heeger, <xref ref-type="bibr" rid="B108">2009</xref>; Feldman and Friston, <xref ref-type="bibr" rid="B25">2010</xref>; Ruff and Cohen, <xref ref-type="bibr" rid="B113">2016</xref>). For example, normalization models propose that every neuronal response is normalized within its neuronal ensemble (i.e., the surrounding neuronal responses) (Heeger, <xref ref-type="bibr" rid="B47">1992</xref>; Louie and Glimcher, <xref ref-type="bibr" rid="B75">2019</xref>). Thus, to amplify the neuronal response of particular neuron, the neuronal pool has to be inhibited such that particular neuron has a sharper evoked response (Schmitz and Duncan, <xref ref-type="bibr" rid="B121">2018</xref>). Importantly, these (superficially distinct) formulations simulate similar functions using different procedures to accentuate responses over a particular neuronal pool for a given neuron or a group of neurons. This introduces shifts in precision to produce attentional gain and the precision of neuronal encoding.</p>
</sec>
<sec>
<title>2.2. Salience as uncertainty minimization</title>
<p>In the neurosciences, (visual) salience refers to the &#x02018;significance&#x00027; of particular objects in the environment. Salience often implicates the superior colliculus, a region that encodes eye movements (White et al., <xref ref-type="bibr" rid="B131">2017</xref>). This makes intuitive sense, as the superior colliculus plays a role in generation of eye movements&#x02014;being an integral part of the brainstem oculomotor network (Raybourn and Keller, <xref ref-type="bibr" rid="B107">1977</xref>)&#x02014;and salient objects provide information that is best resolved in the center of the visual field, thus motivating eye movements to that location. For this reason, our understanding of salience is a quintessentially action-driving phenomenon (Parr and Friston, <xref ref-type="bibr" rid="B98">2019</xref>). Mathematically, salience has been defined as Bayesian surprise (Itti and Koch, <xref ref-type="bibr" rid="B54">2001</xref>; Itti and Baldi, <xref ref-type="bibr" rid="B53">2009</xref>), intrinsic motivation (Oudeyer and Kaplan, <xref ref-type="bibr" rid="B92">2009</xref>), and subsequently, epistemic value under active inference (Mirza et al., <xref ref-type="bibr" rid="B83">2016</xref>; Parr et al., <xref ref-type="bibr" rid="B94">2018</xref>). Active inference&#x02014;a Bayesian account of perception and action (Friston et al., <xref ref-type="bibr" rid="B38">2017a</xref>; Da Costa et al., <xref ref-type="bibr" rid="B20">2020</xref>)&#x02014;stipulates that action selection is determined by uncertainty minimization. Formally, uncertainty minimization speaks to minimization of an expected free energy functional over future trajectories (Da Costa et al., <xref ref-type="bibr" rid="B20">2020</xref>; Sajid et al., <xref ref-type="bibr" rid="B114">2021a</xref>). This action selection objective can be decomposed into epistemic and extrinsic value, where the former pertains to exploratory drives that encourage resolution of uncertainty by sampling salient observations, e.g., only checking one&#x00027;s watch when one does not know the time. However, after checking the watch there is little epistemic value in looking at it again. Generally, the tendency to seek out new locations&#x02014;once uncertainty has been resolved at the current fixation point&#x02014;is called inhibition of return (Klein, <xref ref-type="bibr" rid="B59">2000</xref>).</p>
<p>From an active inference perspective, this phenomenon is prevalent because a recent action has already resolved the uncertainty about the time and checking again would offer nothing more in terms of information gain (Parr and Friston, <xref ref-type="bibr" rid="B98">2019</xref>). Accordingly, salience involves seeking sensory data that have a predictable, uncertainty reducing, effect on current beliefs about states of affairs in the world (Mirza et al., <xref ref-type="bibr" rid="B83">2016</xref>; Parr et al., <xref ref-type="bibr" rid="B94">2018</xref>). Thus salience contends with beliefs about data that must be acquired and the precision of beliefs about policies (i.e., action trajectories) that dictate it. Formally, this emerges from the imperative to maximize the amount of information gained regarding beliefs, from observing the environment. Happily, prior studies have made the connection between eye movements, salience, and precision manipulation (Friston et al., <xref ref-type="bibr" rid="B39">2011</xref>; Brown et al., <xref ref-type="bibr" rid="B13">2013</xref>; Crevecoeur and Kording, <xref ref-type="bibr" rid="B19">2017</xref>). This connection emerges from planning strategies that allow the agent to minimize uncertainty by garnering the right kind of data.</p>
<p>Next, we consider recent findings on how the coupling of these two mechanisms, attention and salience, may be realized in the brain.</p>
</sec>
<sec>
<title>2.3. Rhythmic coupling of attention and salience</title>
<p>To illustrate the coupling between attention and salience, we turn to a recent rhythmic theory of attention. The theory proposes that coupling of saccades, during sampling of visual information, happens at neuronal and behavioral theta oscillations; a frequency of 3&#x02013;8 Hz (Fiebelkorn and Kastner, <xref ref-type="bibr" rid="B27">2019</xref>, <xref ref-type="bibr" rid="B29">2021</xref>). This frequency simultaneously allows for: (<italic>i</italic>) a systematic integration of visual samples with action, and (<italic>ii</italic>) a temporal schedule to disengage and search the environment for more relevant information.</p>
<p>Given that gain control is related to increased sensory precision, we can accordingly relate saccadic eye movements to the decreased precision. This introduces saccadic suppression, a phenomenon that decreases visual gain during eye movements (Crevecoeur and Kording, <xref ref-type="bibr" rid="B19">2017</xref>). This phenomenon was described by Helmholtz who observed that externally initiated eye movements (e.g., when oneself gently presses a side of an eye) eludes the saccadic suppression that accompanies normal eye movements&#x02014;and we see the world shift, because optic flow is not attenuated (Helmholtz, <xref ref-type="bibr" rid="B49">1925</xref>). An interesting consequence of this is that, as eye movements happen periodically (Rucci et al., <xref ref-type="bibr" rid="B112">2018</xref>; Benedetto et al., <xref ref-type="bibr" rid="B10">2020</xref>), there must be a periodic switch between high and low sensory precision, with high precision (or enhanced gain) during fixations and low precision (or suppressed gain) during saccades. Interestingly, it has been shown that rather than having action resetting the neural periodicity, it is better understood as something that aligns within an already existing rhythm (Hogendoorn, <xref ref-type="bibr" rid="B51">2016</xref>; Tomassini et al., <xref ref-type="bibr" rid="B125">2017</xref>). Additionally, the rhythmicity of higher and lower fidelity of sensory sampling has been shown to fluctuate rhythmically around 3 Hz (Benedetto and Morrone, <xref ref-type="bibr" rid="B9">2017</xref>), suggesting that action emerges rhythmically when visual precision is low (Hogendoorn, <xref ref-type="bibr" rid="B51">2016</xref>), triggering salience.</p>
<p>Building upon this, we hypothesize that theta rhythms generated in the fronto-parietal network (Fiebelkorn et al., <xref ref-type="bibr" rid="B30">2018</xref>; Helfrich et al., <xref ref-type="bibr" rid="B48">2018</xref>; Fiebelkorn and Kastner, <xref ref-type="bibr" rid="B28">2020</xref>) couples saccades with saccadic suppression causing the switches between visual sampling and saccadic shifting. This introduces a diachronic aspect to the belief updating process (Friston et al., <xref ref-type="bibr" rid="B44">2020</xref>; Parr and Pezzulo, <xref ref-type="bibr" rid="B99">2021</xref>; Sajid et al., <xref ref-type="bibr" rid="B117">2022</xref>); i.e., sequential fluctuations between attending to current data (perception) and seeking new data (action). This supports empirical findings that both eye movements (Sommer and Wurtz, <xref ref-type="bibr" rid="B123">2006</xref>) and filtering irrelevant information (Phillips et al., <xref ref-type="bibr" rid="B102">2016</xref>; Nakajima et al., <xref ref-type="bibr" rid="B87">2019</xref>; Fiebelkorn and Kastner, <xref ref-type="bibr" rid="B28">2020</xref>) are initiated in this cortical network. Interestingly, both eye movements and visual filtering then propagate to sub-cortical regions, i.e., the superior colliculus&#x02014;for saliency map composition (White et al., <xref ref-type="bibr" rid="B131">2017</xref>)&#x02014;and the thalamus&#x02014;for gain control (Kanai et al., <xref ref-type="bibr" rid="B55">2015</xref>; Fiebelkorn et al., <xref ref-type="bibr" rid="B31">2019</xref>), respectively. Furthermore, this is consistent with recent findings that the periodicity of neural responses are important for understanding the relation of motor responses and sensory information&#x02014;i.e., perception-action coupling (Benedetto et al., <xref ref-type="bibr" rid="B10">2020</xref>). Importantly, theta rhythms also speak to the speed (i.e., the temporal schedule) with which visual information is sampled from the environment (Busch and VanRullen, <xref ref-type="bibr" rid="B15">2010</xref>; Dugu&#x000E9; et al., <xref ref-type="bibr" rid="B22">2015</xref>, <xref ref-type="bibr" rid="B23">2016</xref>; Helfrich et al., <xref ref-type="bibr" rid="B48">2018</xref>). Meaning visual information is not sampled continuously, as our visual experiences would suggest, but rather it is made of successive discrete samples (VanRullen, <xref ref-type="bibr" rid="B129">2016</xref>; Parr et al., <xref ref-type="bibr" rid="B100">2021</xref>).</p>
<p>The prefrontal theta rhythm has been associated with working memory (WM), a process that holds compressed information about the previously observed stimuli, in the sense that measured power in this frequency range using electroencephalography increases during tasks that place demands on WM (Axmacher et al., <xref ref-type="bibr" rid="B5">2010</xref>; Hsieh and Ranganath, <xref ref-type="bibr" rid="B52">2014</xref>; K&#x000F6;ster et al., <xref ref-type="bibr" rid="B62">2018</xref>; Brzezicka et al., <xref ref-type="bibr" rid="B14">2019</xref>; Peters et al., <xref ref-type="bibr" rid="B101">2020</xref>; Balestrieri et al., <xref ref-type="bibr" rid="B7">2021</xref>; Pomper and Ansorge, <xref ref-type="bibr" rid="B103">2021</xref>). The implication is that the neural processes that underwrite WM may depend upon temporal cycles with periods similar to that of perceptual sampling. Importantly, this cognitive process is influenced by how salient a particular stimulus was (Fine and Minnery, <xref ref-type="bibr" rid="B32">2009</xref>; Santangelo and Macaluso, <xref ref-type="bibr" rid="B120">2013</xref>; Santangelo et al., <xref ref-type="bibr" rid="B119">2015</xref>). Moreover, WM has been implicated with attentional mechanisms (Knudsen, <xref ref-type="bibr" rid="B61">2007</xref>; Gazzaley and Nobre, <xref ref-type="bibr" rid="B46">2012</xref>; Oberauer, <xref ref-type="bibr" rid="B89">2019</xref>; Peters et al., <xref ref-type="bibr" rid="B101">2020</xref>; Panichello and Buschman, <xref ref-type="bibr" rid="B93">2021</xref>). This is aligned with our account where we illustrate a rhythmic coupling between salience and attention.</p>
<p>In summary, the computations that underwrite attention and active vision are coupled and exhibit circular causality. Briefly, selective attention and sensory attenuation optimize the processing of sensory samples and which particular visual percepts are inferred. In turn, this determines appropriateness of future eye movements (or actions) and shapes which prior stimuli are encoded into the agent&#x00027;s working memory. Interestingly, the close functional (and computational) link between the two mechanisms endorses the pre-motor theory of attention.</p>
</sec>
</sec>
<sec id="s3">
<title>3. Proposed precision-modulated account of attention and salience</title>
<p>Here, we introduce our precision-modulated account of perception and action. A graphical illustration is provided in <xref ref-type="fig" rid="F1">Figure 1</xref>. For this, we turn to attention and salient action selection which have their roots in biological processes relevant for acquiring task-relevant information. Under an active inference account, this attention influences (posterior) state estimation and can be associated with increased precision of belief updating and gain control&#x02014;described in Section 2.1. Furthermore, this is distinct from salience despite interdependent neuronal composition and computations.</p>
<fig id="F1" position="float">
<label>Figure 1</label>
<caption><p>A graphical illustration of the precision-modulated account of perception and action. Salience and attention are computed based upon beliefs (assumed to be) encoded in parts of the fronto-parietal network and realized in distinct brain regions: superior colliculus (SC) for perception as inference and thalamus for planning as inference, respectively. To deploy attentional processes efficiently, these two mechanisms have to be aligned, which is done rhythmically, hypothetically in theta frequency. This coupling enables the saccadic suppression phenomenon through fluctuations in precision (on an arbitrary scale). When precision is low (i.e., the trough of the theta rhythm), the saccade emerges. Note that there might be distinct processes inhibiting the action (e.g., covert attention), and (despite a decline in precision) saccades might not emerge in every theta cycle. On the other hand, high precision facilitates confident inferences about the causes of visual data. Under this account, thalamus is used for initiating gain control (or visual sampling in general) by providing stronger sensory input, while superior colliculus dictates next saccades, that lead to most informative fixation positions.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fnbot-16-896229-g0001.tif"/>
</fig>
<p>Further alignment between the two constructs can be revealed by considering the temporal scheduling between movement (i.e., action) and perception for uncertainty resolution (Parr and Friston, <xref ref-type="bibr" rid="B98">2019</xref>). We postulate that this perception-action coupling is best understood as a periodic fluctuation between minimizing uncertainty and precision control. Subsequently, action is deployed to reduce uncertainty. Such an alignment specifies what stimulus is selected and under what level of precision it is processed. Parr and Friston (<xref ref-type="bibr" rid="B98">2019</xref>) hypothesize that action alignment with precision is due to the eye structure that provides precise information in the fovea and requires the agent to foveate the most informative stimulus. We extend this by considering the periodic deployment of gain control with saccades (Hogendoorn, <xref ref-type="bibr" rid="B51">2016</xref>; Benedetto and Morrone, <xref ref-type="bibr" rid="B9">2017</xref>; Tomassini et al., <xref ref-type="bibr" rid="B125">2017</xref>; Fiebelkorn and Kastner, <xref ref-type="bibr" rid="B27">2019</xref>; Nakayama and Motoyoshi, <xref ref-type="bibr" rid="B88">2019</xref>).</p>
<p>Accordingly, our formulation defines attention as precision control and salience as uncertainty minimization supported by discrete sampling of visual information at a theta rhythm. This synchronizes perception and action together in an oscillatory fashion (Hogendoorn, <xref ref-type="bibr" rid="B51">2016</xref>). Importantly, a Bayesian formulation of this can be realized as precision manipulation over particular model parameters. We reserve further details for Section 4.</p>
<p><bold>Summary</bold> Based upon our review, we propose a precision-modulated account of attention and salience, emphasizing the diachronic realization of action and perception. In the following sections, we investigate a realization of this model for a robotic system.</p>
</sec>
<sec id="s4">
<title>4. Precision-based attention for Robotics</title>
<p>The previous section introduced a conceptual account to explain the computational mechanisms that undergird attention based on neuroscience findings. We focused on reclaiming saliency as an active process that relies on neural gain control, uncertainty minimization and structured scheduling. Here, we describe how we can mathematically realize some of these mechanisms in the context of well-known challenges in robotics. Enabling robots with this type of attention may be crucial to filter the sensory signals and internal variables that are relevant to estimate the robot/world state and complete any task. More importantly, the active component of salience (i.e., behavior) is essential to interact with the world&#x02014;as argued in active perception approaches (Bajcsy et al., <xref ref-type="bibr" rid="B6">2018</xref>).</p>
<p>We revisit the standard view of attention in robotics by introducing sensory precision (inverse variance) as the driving mechanism for modulating both perception and action (Friston et al., <xref ref-type="bibr" rid="B39">2011</xref>; Clark, <xref ref-type="bibr" rid="B18">2013</xref>). Although saliency was originally described to underwrite behavior, most models used in robotics, strongly biased by computer vision approaches, focus on computing the most relevant region of an image (Borji and Itti, <xref ref-type="bibr" rid="B11">2012</xref>)&#x02014;mainly computing human fixation maps&#x02014;relegating action to a secondary process. Illustratively, state-of-the-art deep learning saliency models&#x02014;as shown in the MIT saliency benchmark (Bylinskii et al., <xref ref-type="bibr" rid="B17">2019</xref>)&#x02014;do not have the action as an output. Conversely, the active perception approach properly defines the action as an essential process of active sensing to gather the relevant information. Our proposed model, based on precision modulated action and perception coupling (<italic>i</italic>) place attention as essential for state-estimation and system identification and (<italic>ii</italic>) and reclaims saliency as a driver for information-seeking behavior, as proposed in early works (Tsotsos et al., <xref ref-type="bibr" rid="B127">1995</xref>), but goes beyond human fixation maps for both improving the model of the environment (exploration) and solving the task (exploitation).</p>
<p>In what follows, we highlight the key role of precision by reviewing relevant brain-inspired attention models deployed in robotics (Section 4.1). We propose precision-modulated attentional mechanisms for robots in three contexts&#x02014;perception (Section 4.2), action (Section 4.3) and active perception (Section 4.3.3). The precision-modulated perception is formalized for a robotics setting; <italic>via</italic> (<italic>i</italic>) state estimation (i.e., estimating the hidden states of a dynamic system from sensory signals&#x02014;Section 4.2.2), and (<italic>ii</italic>) system identification (i.e., estimating the parameters of the dynamic system from sensory signals&#x02014;Section 4.2.3). Next, we show that precision-modulated action can be realized through precision optimization (planning future actions&#x02014;Section 4.3.2) and discuss practical considerations for coupling with precision-modulated perception (precision based active perception&#x02014;Section 4.3.3). <xref ref-type="table" rid="T1">Table 1</xref> summarizes our proposed precision manipulations to solve relevant problems in robot perception and action. <xref ref-type="table" rid="T2">Table 2</xref> provides the definitions of precision within our mechanism.</p>
<table-wrap position="float" id="T1">
<label>Table 1</label>
<caption><p>Robotics applications and their precision realizations.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th valign="top" align="left"><bold>Task</bold></th>
<th valign="top" align="left"><bold>Application</bold></th>
<th valign="top" align="left"><bold>Precision manipulation</bold></th>
<th valign="top" align="center"><bold>Sections</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Perception</td>
<td valign="top" align="left">State and input estimation</td>
<td valign="top" align="left">Noise precision modeling <inline-formula><mml:math id="M1"><mml:mover accent="true"><mml:mrow><mml:mi>&#x003A0;</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:math></inline-formula></td>
<td valign="top" align="center">4.2.2</td>
</tr>
<tr>
<td/>
<td valign="top" align="left">System Identification</td>
<td valign="top" align="left">Posterior parameter precision learning &#x003A0;<sup>&#x003B8;</sup></td>
<td valign="top" align="center">4.2.3</td>
</tr>
<tr>
<td/>
<td valign="top" align="left">Exploration-exploitation in learning</td>
<td valign="top" align="left">Prior parameter precision modeling <italic>P</italic><sup>&#x003B8;</sup></td>
<td valign="top" align="center">4.2.4</td>
</tr>
<tr>
<td/>
<td valign="top" align="left">Noise estimation</td>
<td valign="top" align="left">Noise precision learning <inline-formula><mml:math id="M2"><mml:mover accent="true"><mml:mrow><mml:mi>&#x003A0;</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:math></inline-formula></td>
<td valign="top" align="center">4.2.5</td>
</tr>
<tr>
<td valign="top" align="left">Action</td>
<td valign="top" align="left">Informative Path Planning (IPP)</td>
<td valign="top" align="left">Precision optimization (of map)</td>
<td valign="top" align="center">4.3.2</td>
</tr>
<tr>
<td valign="top" align="left">Active perception</td>
<td valign="top" align="left">IPP with action-perception cycle</td>
<td valign="top" align="left">Precision modulation</td>
<td valign="top" align="center">4.3.3</td>
</tr>
</tbody>
</table>
</table-wrap>
<table-wrap position="float" id="T2">
<label>Table 2</label>
<caption><p>Precision parameters that are manipulated in Section 4.2.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th valign="top" align="left"><bold>Term</bold></th>
<th valign="top" align="center"><bold>Symbol</bold></th>
<th valign="top" align="left"><bold>Definition</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Sensory precision</td>
<td valign="top" align="center">&#x003A0;<sup><italic>z</italic></sup></td>
<td valign="top" align="left">Inverse covariance of sensory noise <bold>z</bold> (Equation 1).</td>
</tr>
<tr>
<td valign="top" align="left">Prior parameter precision</td>
<td valign="top" align="center"><italic>P</italic><sup>&#x003B8;</sup></td>
<td valign="top" align="left">The robot&#x00027;s confidence on its prior parameters &#x003B7;<sup>&#x003B8;</sup>.</td>
</tr>
<tr>
<td valign="top" align="left">Noise precision</td>
<td valign="top" align="center"><inline-formula><mml:math id="M3"><mml:mover accent="true"><mml:mrow><mml:mi>&#x003A0;</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:math></inline-formula></td>
<td valign="top" align="left">The inverse covariance of all noises (Equation 5).</td>
</tr>
<tr>
<td valign="top" align="left">Posterior parameter precision</td>
<td valign="top" align="center">&#x003A0;<sup>&#x003B8;</sup></td>
<td valign="top" align="left">The robot&#x00027;s confidence on its parameter estimates.</td>
</tr>
</tbody>
</table>
</table-wrap>
<sec>
<title>4.1. Previous brain-inspired attention models in robotics</title>
<p>Brain-inspired attention has been mainly addressed in robotics from a &#x0201C;passive&#x0201D; visual saliency perspective, e.g., which pixels of the image are the most relevant. This saliency map is then generally used to foveate the most salient region. This approach was strongly influenced by early computational models of visual attention (Tsotsos et al., <xref ref-type="bibr" rid="B127">1995</xref>; Itti and Koch, <xref ref-type="bibr" rid="B54">2001</xref>). The first models deployed in robots were bottom-up, where the sensory input was transformed into an array of values that represents the importance (or salience) of each cue. Thus, the robot was able to identify which region of the scene has to look at, independently of the task performed&#x02014;see Borji and Itti (<xref ref-type="bibr" rid="B11">2012</xref>) for a review on visual saliency. These models have also been useful for acquiring meaningful visual features in applications, such as object recognition (Orabona et al., <xref ref-type="bibr" rid="B91">2005</xref>; Frintrop, <xref ref-type="bibr" rid="B34">2006</xref>), localization, mapping and navigation (Frintrop and Jensfelt, <xref ref-type="bibr" rid="B33">2008</xref>; Roberts et al., <xref ref-type="bibr" rid="B111">2012</xref>; Kim and Eustice, <xref ref-type="bibr" rid="B58">2013</xref>). Saliency computation was usually employed as a helper for the selection of the relevant characteristics of the environment to be encoded. Thus, reducing the information needed to process.</p>
<p>More refined methods of visual attention employed top-down modulation, where the context, task or goal bias the relevance of the visual input. These methods were used, for instance, to identify humans using motion patterns (Butko et al., <xref ref-type="bibr" rid="B16">2008</xref>; Mor&#x000E9;n et al., <xref ref-type="bibr" rid="B84">2008</xref>). A few works also focused on object/target search applications, where top-down and bottom-up saliency attention were used to find objects or people in a search and rescue scenario (Rasouli et al., <xref ref-type="bibr" rid="B106">2020</xref>).</p>
<p>Attention has also been considered in human-robot interaction and social robotics applications (Ferreira and Dias, <xref ref-type="bibr" rid="B26">2014</xref>), mainly for scene or task understanding (Kragic et al., <xref ref-type="bibr" rid="B63">2005</xref>; Ude et al., <xref ref-type="bibr" rid="B128">2005</xref>; Lanillos et al., <xref ref-type="bibr" rid="B66">2016</xref>), and gaze estimation (Shon et al., <xref ref-type="bibr" rid="B122">2005</xref>) and generation (Lanillos et al., <xref ref-type="bibr" rid="B67">2015a</xref>). For instance, computing where the human is looking at and where the robot should look at or which object should be grasped. Furthermore, multi-sensory and 3D saliency computation has also been investigated (Lanillos et al., <xref ref-type="bibr" rid="B68">2015b</xref>). Finally, more complex attention behaviors, particularly designed for social robotics and based on human non-verbal communication, such as joint attention, have also been addressed. Here the robot and the human share the attention of one object through meaningful saccades, i.e., head/eye movements (Nagai et al., <xref ref-type="bibr" rid="B86">2003</xref>; Kaplan and Hafner, <xref ref-type="bibr" rid="B57">2006</xref>; Lanillos et al., <xref ref-type="bibr" rid="B67">2015a</xref>).</p>
<p>Although attention mechanisms have been widely investigated in robotics, specially to model visual cognition (Kragic et al., <xref ref-type="bibr" rid="B63">2005</xref>; Begum and Karray, <xref ref-type="bibr" rid="B8">2010</xref>), the majority of the works have treated attention as an extra feature that can help the visual processing, instead of a crucial component needed for the proper functioning of the cognitive abilities of the robot (Lanillos and Cheng, <xref ref-type="bibr" rid="B71">2018a</xref>). Furthermore, these methods had the tendency to leave the action generation out of the attention process. One of the reasons for not including saliency computation, in robotic systems, is that the majority of the models only output &#x0201C;human-fixation map&#x0201D; predictions, given a static image. Saliency computation introduces extra computational complexity, which can be finessed by visual segmentation algorithms (e.g., line detectors in autonomous navigation). However, it does not resolve uncertainty nor select actions that maximize information gain in the future. In essence, the incomplete view of attention models that output human-fixation maps has arguably obscured the huge potential of neuroscience-inspired attentional mechanisms for robotics.</p>
<p>Our proposed model of attention, based on precision modulation, abandons the current robotics narrow view of attention and saliency by explicitly modeling attention within state estimation, learning and control. Thus, placing attentional processes at the core of the robot computation and not as an extra add-on. In the following sections, we describe the realization of our precision-based attention formulation in robotics using common practical applications as the backbone motif.</p>
</sec>
<sec>
<title>4.2. Precision-modulated perception</title>
<p>We formalize precision-modulated perception from a first principles Bayesian perspective&#x02014;explicitly the free energy principle approach proposed by Friston et al. (<xref ref-type="bibr" rid="B39">2011</xref>). Practically, this entails optimizing precision parameters over (particular) model parameters.</p>
<p>Through numerical examples show how our model is able to perform accurate state estimation (Bos et al., <xref ref-type="bibr" rid="B12">2021</xref>) and stable parameter learning (Meera and Wisse, <xref ref-type="bibr" rid="B79">2021a</xref>,<xref ref-type="bibr" rid="B80">b</xref>). To illustrate the approach, we first introduce a dynamic system modeled as a linear state space system in robotics (Section 4.2.1)&#x02014;we used this formulation in all our numerical experiments. We briefly review the formal terminologies for a robotics context to appropriately situate our precision-based mechanism for perception. Explicitly, we introduce: precision modeling (by adapting a known form of the precision matrix), precision learning (by learning the full precision matrix), and precision optimization (use precision as an objective function during learning). As a reminder, precision modeling is associated with (instantaneous) gain control and precision learning (at slower time scales) is associated with optimizing that control.</p>
<sec>
<title>4.2.1. Precision for state space models</title>
<p>A linear dynamic system can be modeled using the following state space equations (boldface notation denotes components of the real system and non-boldface notation its estimates):</p>
<disp-formula id="E1"><label>(1)</label><mml:math id="M4"><mml:mtable class="eqnarray" columnalign="right center left"><mml:mtr><mml:mtd><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mover accent="true"><mml:mrow><mml:mtext mathvariant='bold-italic'>x</mml:mtext></mml:mrow><mml:mo>&#x02219;</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:mtext mathvariant='bold-italic'>Ax</mml:mtext><mml:mo>&#x0002B;</mml:mo><mml:mtext mathvariant='bold-italic'>Bu</mml:mtext><mml:mo>&#x0002B;</mml:mo><mml:mtext mathvariant='bold-italic'>w</mml:mtext><mml:mo>,</mml:mo></mml:mtd></mml:mtr></mml:mtable><mml:mtext>&#x000A0;&#x000A0;</mml:mtext><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mtext mathvariant='bold-italic'>y</mml:mtext><mml:mo>=</mml:mo><mml:mtext mathvariant='bold-italic'>Cx</mml:mtext><mml:mo>&#x0002B;</mml:mo><mml:mtext mathvariant='bold-italic'>z</mml:mtext><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>where <bold>A</bold>, <bold>B</bold> and <bold>C</bold> are constant matrices defining the system parameters, <bold>x</bold> &#x02208; &#x0211D;<sup><italic>n</italic></sup> is the system state (usually an unobserved variable), <bold>u</bold> &#x02208; &#x0211D;<sup><italic>r</italic></sup> is the input or control actions, <bold>y</bold> &#x02208; &#x0211D;<sup><italic>m</italic></sup> is the output or the sensory measurements, <bold>w</bold> &#x02208; &#x0211D;<sup><italic>n</italic></sup> is the process noise with precision <bold>&#x003A0;<sup><italic>w</italic></sup></bold> (or inverse variance <bold>&#x003A3;</bold><sup><italic><bold>w</bold></italic>&#x02212;1</sup>), and <bold>z</bold> &#x02208; &#x0211D;<sup><italic>m</italic></sup> is the measurement noise with precision <bold>&#x003A0;<sup><italic>z</italic></sup></bold>.</p>
<p>For instance, we can describe a mass-spring damper system (depicted in <xref ref-type="fig" rid="F2">Figure 2B</xref>) using state space equations. A mass (<italic>m</italic> &#x0003D; 1.4<italic>kg</italic>) is attached to a spring with elasticity constant (<italic>k</italic> &#x0003D; 0.8<italic>N</italic>/<italic>m</italic>), and a damper with a damping coefficient (<italic>b</italic> &#x0003D; 0.4<italic>Ns</italic>/<italic>m</italic>). When a force (<italic>u</italic>(<italic>t</italic>) &#x0003D; <italic>e</italic><sup>&#x02212;0.25(<italic>t</italic>&#x02212;12)<sup>2</sup></sup>) is applied on the mass, it displaces <italic>x</italic> from its equilibrium point. The linear dynamics of this system is given by:</p>
<disp-formula id="E2"><label>(2)</label><mml:math id="M5"><mml:mtable class="eqnarray" columnalign="right center left"><mml:mtr><mml:mtd><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none none none none none none none none none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mi>&#x01E8B;</mml:mi></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mi>&#x01E8D;</mml:mi></mml:mtd></mml:mtr></mml:mtable></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none none none none none none none none none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mn>1</mml:mn></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mo>-</mml:mo><mml:mfrac><mml:mrow><mml:mi>k</mml:mi></mml:mrow><mml:mrow><mml:mi>m</mml:mi></mml:mrow></mml:mfrac></mml:mtd><mml:mtd><mml:mo>-</mml:mo><mml:mfrac><mml:mrow><mml:mi>b</mml:mi></mml:mrow><mml:mrow><mml:mi>m</mml:mi></mml:mrow></mml:mfrac></mml:mtd></mml:mtr></mml:mtable></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none none none none none none none none none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mi>x</mml:mi></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mi>&#x01E8B;</mml:mi></mml:mtd></mml:mtr></mml:mtable></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mo>&#x0002B;</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none none none none none none none none none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mn>0</mml:mn></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mfrac><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>m</mml:mi></mml:mrow></mml:mfrac></mml:mtd></mml:mtr></mml:mtable></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mi>u</mml:mi><mml:mo>,</mml:mo><mml:mtext>&#x000A0;&#x000A0;</mml:mtext><mml:mi>y</mml:mi><mml:mo>=</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none none none none none none none none none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mn>1</mml:mn></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd></mml:mtr></mml:mtable></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none none none none none none none none none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mi>x</mml:mi></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mi>&#x01E8B;</mml:mi></mml:mtd></mml:mtr></mml:mtable></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>Note that Equation (2) is equivalent to Equation (1) with parameters <bold>A</bold> <inline-formula><mml:math id="M6"><mml:mo>=</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none none none none none none none none none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mn>1</mml:mn></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mo>-</mml:mo><mml:mfrac><mml:mrow><mml:mi>k</mml:mi></mml:mrow><mml:mrow><mml:mi>m</mml:mi></mml:mrow></mml:mfrac></mml:mtd><mml:mtd><mml:mo>-</mml:mo><mml:mfrac><mml:mrow><mml:mi>b</mml:mi></mml:mrow><mml:mrow><mml:mi>m</mml:mi></mml:mrow></mml:mfrac></mml:mtd></mml:mtr></mml:mtable></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mo>,</mml:mo></mml:math></inline-formula><bold>B</bold> =<inline-formula><mml:math id="M7"><mml:msup><mml:mrow><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none none none none none none none none none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mn>0</mml:mn><mml:mo>,</mml:mo><mml:mfrac><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>m</mml:mi></mml:mrow></mml:mfrac></mml:mtd></mml:mtr></mml:mtable></mml:mrow><mml:mo>]</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mi>T</mml:mi></mml:mrow></mml:msup></mml:math></inline-formula> and <bold>C</bold> <inline-formula><mml:math id="M8"><mml:mo>=</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none none none none none none none none none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mn>1</mml:mn></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd></mml:mtr></mml:mtable></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mo>,</mml:mo></mml:math></inline-formula> and state x = <inline-formula><mml:math id="M9"><mml:msup><mml:mrow><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none none none none none none none none none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mi>x</mml:mi><mml:mo>,</mml:mo><mml:mi>&#x01E8B;</mml:mi></mml:mtd></mml:mtr></mml:mtable></mml:mrow><mml:mo>]</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mi>T</mml:mi></mml:mrow></mml:msup></mml:math></inline-formula>.</p>
<fig id="F2" position="float">
<label>Figure 2</label>
<caption><p>An illustration of an attention mechanism for state and input estimation of a system (shown in <bold>B</bold>). The quality of the estimation improves <bold>(C)</bold> as the embedding order (number of derivatives) of generalized coordinates are increased <bold>(A)</bold>. However, the imprecise information in the higher order derivatives of the sensory input <bold>y</bold> does not affect the final performance of the observer because of attentional selection, which selectively weighs the importance afforded to each derivative, in the free energy optimization scheme.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fnbot-16-896229-g0002.tif"/>
</fig>
<p>Now we introduce attention as precision modulation assuming that the robotic goal is to minimize the prediction error (Friston et al., <xref ref-type="bibr" rid="B39">2011</xref>; Lanillos and Cheng, <xref ref-type="bibr" rid="B65">2018b</xref>; Meera and Wisse, <xref ref-type="bibr" rid="B78">2020</xref>), i.e., to refine its model of the environment and perform accurate state estimation, given the information available. In other words, the robot has to estimate <bold>x</bold> and <bold>u</bold> from input prior &#x003B7;<sup><italic>u</italic></sup> with a prior precision of <italic>P</italic><sup><italic>u</italic></sup>, given the measurements <bold>y</bold>, parameters <bold>A</bold>, <bold>B</bold>, <bold>C</bold> and noise precision <bold>&#x003A0;<sup>w</sup></bold> and <bold>&#x003A0;<sup>z</sup></bold>. Formally, the prediction error <inline-formula><mml:math id="M10"><mml:mover accent="true"><mml:mrow><mml:mi>&#x003F5;</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:math></inline-formula> of the sensory measurements <inline-formula><mml:math id="M11"><mml:msup><mml:mrow><mml:mover accent="true"><mml:mrow><mml:mi>&#x003F5;</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>y</mml:mi></mml:mrow></mml:msup></mml:math></inline-formula>, control input reference <inline-formula><mml:math id="M12"><mml:msup><mml:mrow><mml:mover accent="true"><mml:mrow><mml:mi>&#x003F5;</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>u</mml:mi></mml:mrow></mml:msup></mml:math></inline-formula> and state <inline-formula><mml:math id="M13"><mml:msup><mml:mrow><mml:mover accent="true"><mml:mrow><mml:mi>&#x003F5;</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>x</mml:mi></mml:mrow></mml:msup></mml:math></inline-formula> are:</p>
<disp-formula id="E3"><label>(3)</label><mml:math id="M14"><mml:mtable class="eqnarray" columnalign="right center left"><mml:mtr><mml:mtd><mml:mover accent="true"><mml:mrow><mml:mi>&#x003F5;</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none none none none none none none none none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:msup><mml:mrow><mml:mover accent="true"><mml:mrow><mml:mi>&#x003F5;</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>y</mml:mi></mml:mrow></mml:msup></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:msup><mml:mrow><mml:mover accent="true"><mml:mrow><mml:mi>&#x003F5;</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>u</mml:mi></mml:mrow></mml:msup></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:msup><mml:mrow><mml:mover accent="true"><mml:mrow><mml:mi>&#x003F5;</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>x</mml:mi></mml:mrow></mml:msup></mml:mtd></mml:mtr></mml:mtable></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none none none none none none none none none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mover accent="true"><mml:mrow><mml:mtext mathvariant='bold-italic'>y</mml:mtext></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover><mml:mo>-</mml:mo><mml:mover accent="true"><mml:mrow><mml:mtext mathvariant='bold-italic'>C</mml:mtext></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover><mml:mover accent="true"><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mi>&#x00169;</mml:mi><mml:mo>-</mml:mo><mml:msup><mml:mrow><mml:mover accent="true"><mml:mrow><mml:mi>&#x003B7;</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>u</mml:mi></mml:mrow></mml:msup></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:msup><mml:mrow><mml:mi>D</mml:mi></mml:mrow><mml:mrow><mml:mi>x</mml:mi></mml:mrow></mml:msup><mml:mover accent="true"><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover><mml:mo>-</mml:mo><mml:mover accent="true"><mml:mrow><mml:mtext mathvariant='bold-italic'>A</mml:mtext></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover><mml:mover accent="true"><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover><mml:mo>-</mml:mo><mml:mover accent="true"><mml:mrow><mml:mtext mathvariant='bold-italic'>B</mml:mtext></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover><mml:mi>&#x00169;</mml:mi></mml:mtd></mml:mtr></mml:mtable></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mtext>&#x000A0;</mml:mtext><mml:mrow><mml:mo>{</mml:mo><mml:mrow><mml:mtable style="text-align:axis;" equalrows="false" columnlines="" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mtext class="textrm" mathvariant="normal">sensory prediction error</mml:mtext></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext class="textrm" mathvariant="normal">control input prediction</mml:mtext></mml:mtd></mml:mtr><mml:mtr><mml:mtd columnalign="left"><mml:mtext class="textrm" mathvariant="normal">error</mml:mtext></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext class="textrm" mathvariant="normal">state prediction error</mml:mtext></mml:mtd></mml:mtr></mml:mtable></mml:mrow></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>Note that <inline-formula><mml:math id="M16"><mml:msup><mml:mrow><mml:mover accent="true"><mml:mrow><mml:mi>&#x003F5;</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>y</mml:mi></mml:mrow></mml:msup><mml:mo>=</mml:mo><mml:mover accent="true"><mml:mrow><mml:mstyle class="text"><mml:mtext mathvariant="bold">y</mml:mtext></mml:mstyle></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover><mml:mo>-</mml:mo><mml:mover accent="true"><mml:mrow><mml:mstyle class="text"><mml:mtext mathvariant="bold">C</mml:mtext></mml:mstyle></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover><mml:mover accent="true"><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:math></inline-formula> is the difference between the observed measurement and the predicted sensory input given the state<xref ref-type="fn" rid="fn0003"><sup>3</sup></xref>. Here <italic>D</italic><sup><italic>x</italic></sup> performs the (block) derivative operation, which is equivalent to shifting up all the components in generalized coordinates by one block.</p>
<p>We can estimate the state and input using the Dynamic Expectation Maximization (DEM) algorithm (Friston et al., <xref ref-type="bibr" rid="B45">2008</xref>; Meera and Wisse, <xref ref-type="bibr" rid="B78">2020</xref>) that optimizes a free energy variational bound <inline-formula><mml:math id="M19"><mml:mrow><mml:mi mathvariant="-tex-caligraphic">F</mml:mi></mml:mrow></mml:math></inline-formula> to be tractable<xref ref-type="fn" rid="fn0004"><sup>4</sup></xref>. This is:</p>
<disp-formula id="E4"><label>(4)</label><mml:math id="M20"><mml:mtable class="eqnarray" columnalign="right center left"><mml:mtr><mml:mtd><mml:mi>X</mml:mi><mml:mo>=</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none none none none none none none none none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mover accent="true"><mml:mrow><mml:mi>x</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mi>&#x00169;</mml:mi></mml:mtd></mml:mtr></mml:mtable></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mo class="qopname">arg</mml:mo><mml:mstyle displaystyle="true"><mml:munder class="msub"><mml:mrow><mml:mo class="qopname">max</mml:mo></mml:mrow><mml:mrow><mml:mi>X</mml:mi></mml:mrow></mml:munder></mml:mstyle><mml:mrow><mml:mi mathvariant="-tex-caligraphic">F</mml:mi></mml:mrow><mml:mo>=</mml:mo><mml:mo class="qopname">arg</mml:mo><mml:mstyle displaystyle="true"><mml:munder class="msub"><mml:mrow><mml:mo class="qopname">max</mml:mo></mml:mrow><mml:mrow><mml:mi>X</mml:mi></mml:mrow></mml:munder></mml:mstyle><mml:mo>-</mml:mo><mml:mfrac><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:mfrac><mml:msup><mml:mrow><mml:mover accent="true"><mml:mrow><mml:mi>&#x003F5;</mml:mi></mml:mrow><mml:mo class="qopname">&#x0007E;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>T</mml:mi></mml:mrow></mml:msup><mml:mover accent="true"><mml:mrow><mml:mi>&#x003A0;</mml:mi></mml:mrow><mml:mo class="qopname">&#x0007E;</mml:mo></mml:mover><mml:mover accent="true"><mml:mrow><mml:mi>&#x003F5;</mml:mi></mml:mrow><mml:mo class="qopname">&#x0007E;</mml:mo></mml:mover></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>Crucially, <inline-formula><mml:math id="M21"><mml:mover accent="true"><mml:mrow><mml:mi>&#x003A0;</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:math></inline-formula> is the generalized noise precision that modulates the contribution of each prediction error to the estimation of the state and the computation of the action. Thus, <inline-formula><mml:math id="M22"><mml:mover accent="true"><mml:mrow><mml:mi>&#x003A0;</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:math></inline-formula> is equivalent to attentional gain. For instance, we can model the precision matrix to attend to the most informative signal derivatives in <inline-formula><mml:math id="M23"><mml:mover accent="true"><mml:mrow><mml:mstyle class="text"><mml:mtext mathvariant="bold">y</mml:mtext></mml:mstyle></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:math></inline-formula>. Concisely, the precision <inline-formula><mml:math id="M24"><mml:mover accent="true"><mml:mrow><mml:mi>&#x003A0;</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:math></inline-formula> has the following form:</p>
<disp-formula id="E5"><label>(5)</label><mml:math id="M25"><mml:mtable class="eqnarray" columnalign="right center left"><mml:mtr><mml:mtd><mml:mover accent="true"><mml:mrow><mml:mi>&#x003A0;</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none none none none none none none none none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mi>S</mml:mi><mml:mo>&#x02297;</mml:mo><mml:msup><mml:mrow><mml:mi>&#x003A0;</mml:mi></mml:mrow><mml:mrow><mml:mi>z</mml:mi></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mi>S</mml:mi><mml:mo>&#x02297;</mml:mo><mml:msup><mml:mrow><mml:mi>P</mml:mi></mml:mrow><mml:mrow><mml:mi>u</mml:mi></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mi>S</mml:mi><mml:mo>&#x02297;</mml:mo><mml:msup><mml:mrow><mml:mi>&#x003A0;</mml:mi></mml:mrow><mml:mrow><mml:mi>w</mml:mi></mml:mrow></mml:msup></mml:mtd></mml:mtr></mml:mtable></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mo>,</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>where <italic>S</italic> is the smoothness matrix. In Section 4.2.2, we show that modeling the precision matrix <inline-formula><mml:math id="M26"><mml:mover accent="true"><mml:mrow><mml:mi>&#x003A0;</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:math></inline-formula> using the <italic>S</italic> matrix improves the estimation quality.</p>
<p>The full free energy functional (time integral of free energy <inline-formula><mml:math id="M27"><mml:mover accent="true"><mml:mrow><mml:mrow><mml:mi mathvariant="-tex-caligraphic">F</mml:mi></mml:mrow></mml:mrow><mml:mo>&#x00304;</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:mo>&#x0222B;</mml:mo><mml:mrow><mml:mi mathvariant="-tex-caligraphic">F</mml:mi></mml:mrow><mml:mi>d</mml:mi><mml:mi>t</mml:mi></mml:math></inline-formula> at optimal precision) that the robot optimizes to perform state-estimation and system identification is described in Equation (6)&#x02014;for readability we omitted the details of the derivation of this cost function, and we refer to Anil Meera and Wisse (<xref ref-type="bibr" rid="B3">2021</xref>) for further details.</p>
<disp-formula id="E6"><label>(6)</label><mml:math id="M15"><mml:mtable class="eqnarray" columnalign="left"><mml:mtr><mml:mtd><mml:mi>&#x02131;</mml:mi><mml:mo>=</mml:mo><mml:mo>&#x02212;</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>2</mml:mn></mml:mfrac><mml:mstyle displaystyle='true'><mml:munder><mml:mo>&#x02211;</mml:mo><mml:mi>t</mml:mi></mml:munder><mml:mo stretchy='false'>[</mml:mo></mml:mstyle><mml:munder><mml:munder><mml:mrow><mml:msup><mml:mover accent='true'><mml:mi>&#x003F5;</mml:mi><mml:mo>&#x002DC;</mml:mo></mml:mover><mml:mrow><mml:mi>y</mml:mi><mml:mi>T</mml:mi></mml:mrow></mml:msup><mml:msup><mml:mover accent='true'><mml:mi>&#x003A0;</mml:mi><mml:mo>&#x002DC;</mml:mo></mml:mover><mml:mi>z</mml:mi></mml:msup><mml:msup><mml:mover accent='true'><mml:mi>&#x003F5;</mml:mi><mml:mo>&#x002DC;</mml:mo></mml:mover><mml:mi>y</mml:mi></mml:msup><mml:mo>+</mml:mo><mml:msup><mml:mover accent='true'><mml:mi>&#x003F5;</mml:mi><mml:mo>&#x002DC;</mml:mo></mml:mover><mml:mrow><mml:mi>u</mml:mi><mml:mi>T</mml:mi></mml:mrow></mml:msup><mml:msup><mml:mi>P</mml:mi><mml:mover accent='true'><mml:mi>u</mml:mi><mml:mo>&#x002DC;</mml:mo></mml:mover></mml:msup><mml:msup><mml:mover accent='true'><mml:mi>&#x003F5;</mml:mi><mml:mo>&#x002DC;</mml:mo></mml:mover><mml:mi>u</mml:mi></mml:msup><mml:mo>+</mml:mo><mml:msup><mml:mover accent='true'><mml:mi>&#x003F5;</mml:mi><mml:mo>&#x002DC;</mml:mo></mml:mover><mml:mrow><mml:mi>x</mml:mi><mml:mi>T</mml:mi></mml:mrow></mml:msup><mml:msup><mml:mover accent='true'><mml:mi>&#x003A0;</mml:mi><mml:mo>&#x002DC;</mml:mo></mml:mover><mml:mi>w</mml:mi></mml:msup><mml:msup><mml:mover accent='true'><mml:mi>&#x003F5;</mml:mi><mml:mo>&#x002DC;</mml:mo></mml:mover><mml:mi>x</mml:mi></mml:msup></mml:mrow><mml:mo stretchy='true'>&#x0FE38;</mml:mo></mml:munder><mml:mrow><mml:mtext>precision&#x000A0;weighed&#x000A0;prediction&#x000A0;error</mml:mtext></mml:mrow></mml:munder><mml:mo stretchy='false'>]</mml:mo><mml:mo>&#x02212;</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>2</mml:mn></mml:mfrac><mml:mo stretchy='false'>[</mml:mo><mml:munder><mml:munder><mml:mrow><mml:msup><mml:mi>&#x003F5;</mml:mi><mml:mrow><mml:mi>&#x003B8;</mml:mi><mml:mi>T</mml:mi></mml:mrow></mml:msup><mml:msup><mml:mi>P</mml:mi><mml:mi>&#x003B8;</mml:mi></mml:msup><mml:msup><mml:mi>&#x003F5;</mml:mi><mml:mi>&#x003B8;</mml:mi></mml:msup><mml:mo>+</mml:mo><mml:msup><mml:mi>&#x003F5;</mml:mi><mml:mrow><mml:mi>&#x003BB;</mml:mi><mml:mi>T</mml:mi></mml:mrow></mml:msup><mml:msup><mml:mi>P</mml:mi><mml:mi>&#x003BB;</mml:mi></mml:msup><mml:msup><mml:mi>&#x003F5;</mml:mi><mml:mi>&#x003BB;</mml:mi></mml:msup></mml:mrow><mml:mo stretchy='true'>&#x0FE38;</mml:mo></mml:munder><mml:mrow><mml:mtext>prior&#x000A0;precision&#x000A0;weighed&#x000A0;prediction&#x000A0;error&#x000A0;of&#x000A0;</mml:mtext><mml:mi>&#x003B8;</mml:mi><mml:mtext>&#x000A0;and&#x000A0;</mml:mtext><mml:mi>&#x003BB;</mml:mi></mml:mrow></mml:munder><mml:mo stretchy='false'>]</mml:mo></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mtext>&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;&#x000A0;</mml:mtext><mml:mo>+</mml:mo><mml:munder><mml:munder><mml:mrow><mml:mfrac><mml:mn>1</mml:mn><mml:mn>2</mml:mn></mml:mfrac><mml:msub><mml:mi>n</mml:mi><mml:mi>t</mml:mi></mml:msub><mml:mi>ln</mml:mi><mml:mo>&#x0007C;</mml:mo><mml:msup><mml:mi>&#x003A3;</mml:mi><mml:mi>X</mml:mi></mml:msup><mml:mo>&#x0007C;</mml:mo></mml:mrow><mml:mo stretchy='true'>&#x0FE38;</mml:mo></mml:munder><mml:mrow><mml:mtext>state&#x000A0;and&#x000A0;input&#x000A0;entropy</mml:mtext></mml:mrow></mml:munder><mml:mo>+</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mn>2</mml:mn></mml:mfrac><mml:msup><mml:mi>n</mml:mi><mml:mi>t</mml:mi></mml:msup><mml:munder><mml:munder><mml:mrow><mml:mo stretchy='false'>[</mml:mo><mml:mi>ln</mml:mi><mml:mo>&#x0007C;</mml:mo><mml:msup><mml:mover accent='true'><mml:mi>&#x003A0;</mml:mi><mml:mo>&#x002DC;</mml:mo></mml:mover><mml:mi>z</mml:mi></mml:msup><mml:mo>&#x0007C;</mml:mo><mml:mo>+</mml:mo><mml:mi>ln</mml:mi><mml:mo>&#x0007C;</mml:mo><mml:msup><mml:mi>P</mml:mi><mml:mover accent='true'><mml:mi>v</mml:mi><mml:mo>&#x002DC;</mml:mo></mml:mover></mml:msup><mml:mo>&#x0007C;</mml:mo><mml:mo>+</mml:mo><mml:mi>ln</mml:mi><mml:mo>&#x0007C;</mml:mo><mml:msup><mml:mover accent='true'><mml:mi>&#x003A0;</mml:mi><mml:mo>&#x002DC;</mml:mo></mml:mover><mml:mi>w</mml:mi></mml:msup><mml:mo>&#x0007C;</mml:mo><mml:mo stretchy='false'>]</mml:mo></mml:mrow><mml:mo stretchy='true'>&#x0FE38;</mml:mo></mml:munder><mml:mrow><mml:mtext>noise&#x000A0;entropy</mml:mtext></mml:mrow></mml:munder><mml:mo>+</mml:mo><mml:munder><mml:munder><mml:mrow><mml:mfrac><mml:mn>1</mml:mn><mml:mn>2</mml:mn></mml:mfrac><mml:mi>ln</mml:mi><mml:mo>&#x0007C;</mml:mo><mml:msup><mml:mi>&#x003A3;</mml:mi><mml:mi>&#x003B8;</mml:mi></mml:msup><mml:msup><mml:mi>P</mml:mi><mml:mi>&#x003B8;</mml:mi></mml:msup><mml:mo>&#x0007C;</mml:mo></mml:mrow><mml:mo stretchy='true'>&#x0FE38;</mml:mo></mml:munder><mml:mrow><mml:mtext>parameter&#x000A0;entropy</mml:mtext></mml:mrow></mml:munder><mml:mo>+</mml:mo><mml:munder><mml:munder><mml:mrow><mml:mfrac><mml:mn>1</mml:mn><mml:mn>2</mml:mn></mml:mfrac><mml:mi>ln</mml:mi><mml:mo>&#x0007C;</mml:mo><mml:msup><mml:mi>&#x003A3;</mml:mi><mml:mi>&#x003BB;</mml:mi></mml:msup><mml:msup><mml:mi>P</mml:mi><mml:mi>&#x003BB;</mml:mi></mml:msup><mml:mo>&#x0007C;</mml:mo></mml:mrow><mml:mo stretchy='true'>&#x0FE38;</mml:mo></mml:munder><mml:mrow><mml:mtext>hyperparameter&#x000A0;entropy</mml:mtext></mml:mrow></mml:munder></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>Here &#x003F5;<sup>&#x003B8;</sup> &#x0003D; &#x003B8;&#x02212;&#x003B7;<sup>&#x003B8;</sup>, &#x003F5;<sup>&#x003BB;</sup> &#x0003D; &#x003BB;&#x02212;&#x003B7;<sup>&#x003BB;</sup> are the prediction errors of parameters and hyper-parameters<xref ref-type="fn" rid="fn0005"><sup>5</sup></xref>. <inline-formula><mml:math id="M29"><mml:mover accent="true"><mml:mrow><mml:mrow><mml:mi mathvariant="-tex-caligraphic">F</mml:mi></mml:mrow></mml:mrow><mml:mo>&#x00304;</mml:mo></mml:mover></mml:math></inline-formula> consist of two main components: i) precision weighed prediction errors and ii) precision-based entropy. The dominant role of precision&#x02014;in the free energy objective&#x02014;is reflected in how modulating these precision parameters can have a profound influence perception and behavior. The theoretical guarantees for stable estimation (Meera and Wisse, <xref ref-type="bibr" rid="B80">2021b</xref>), and its application on real robots (Lanillos et al., <xref ref-type="bibr" rid="B70">2021</xref>) make this formulation very appealing to robotic systems.</p>
<p>Note that we can manipulate three kinds of precision within the state space formulation: (i) prior precision (<italic>P</italic><sup>&#x00169;</sup>, <italic>P</italic><sup>&#x003B8;</sup>, <italic>P</italic><sup>&#x003BB;</sup>), (ii) conditional precision on estimates (&#x003A0;<sup><italic>X</italic></sup>, &#x003A0;<sup>&#x003B8;</sup>, &#x003A0;<sup>&#x003BB;</sup>) and (iii) noise precision (&#x003A0;<sup><italic>z</italic></sup>, &#x003A0;<sup><italic>w</italic></sup>). Therefore, to learn the correct parameter values &#x003B8;, we (i) learn the parameter precision &#x003A0;<sup>&#x003B8;</sup>, (ii) model the prior parameter precision <italic>P</italic><sup>&#x003B8;</sup>, and (iii) learn the noise precision &#x003A0;<sup><italic>w</italic></sup> and &#x003A0;<sup><italic>z</italic></sup> (parameterised using &#x003BB;).</p>
</sec>
<sec>
<title>4.2.2. State and input estimation</title>
<p>State estimation is the process of estimating the unobserved states of a real system from (noisy) measurements. Here, we show how we can achieve accurate estimation through precision modulation in a linear time invariant system under the influence of colored noise (Meera and Wisse, <xref ref-type="bibr" rid="B78">2020</xref>). State estimation in the presence of colored noise is inherently challenging, owing to the non-white nature of the noise, which is often ignored in conventional approaches, such as the Kalman Filter (Welch and Bishop, <xref ref-type="bibr" rid="B130">2002</xref>).</p>
<p><xref ref-type="fig" rid="F2">Figure 2</xref> summarizes a numerical example that shows how one can use precision modulation to focus on the less noisy derivatives (lower derivatives) of measurements, relative to imprecise higher derivatives. Thus, enabling the robot to use the most informative data for state and input estimation, while discarding imprecise input. <xref ref-type="fig" rid="F2">Figure 2B</xref> depicts the mass-spring damper system used. The numerical results show that the quality of the estimation increases as the embedding ordering increases but the lack of information in the higher order derivatives of the sensory input do not affect the final performance due to the precision modulation. The higher order derivatives <xref ref-type="fig" rid="F2">(Figure 2A</xref>) are less precise than the lower derivatives, thereby reflecting the loss of information in higher derivatives. The state and input estimation was performed using the optimization framework described in the previous section. The quality of estimation is shown in <xref ref-type="fig" rid="F2">Figure 2C</xref>, where the input estimation using six derivatives (blue curve) is closer to the real input (yellow curve) than when compared to the estimation using only one derivative (red curve). The quality of the estimation reports the sum of squared error (SSE) in the estimation of states and inputs with respect to the embedding order (number of signal derivatives considered).</p>
<p>To obtain accurate state estimation by optimizing the precision parameters, we recall that the precision weights the prediction errors. From Equation (3), the structural form of <inline-formula><mml:math id="M30"><mml:mover accent="true"><mml:mrow><mml:mi>&#x003A0;</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:math></inline-formula> is mainly dictated by the smoothness matrix <italic>S</italic>, which establishes the interdependence between the components of the variable expressed in generalized coordinates (e.g., the dependence between <bold>y</bold>, <bold>y</bold>&#x02032; and <bold>y</bold>&#x02033; in <inline-formula><mml:math id="M31"><mml:mover accent="true"><mml:mrow><mml:mstyle class="text"><mml:mtext mathvariant="bold">y</mml:mtext></mml:mstyle></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:math></inline-formula>). For instance, the <italic>S</italic> matrix for a Gaussian kernel is as follows Meera and Wisse (<xref ref-type="bibr" rid="B81">2022</xref>):</p>
<disp-formula id="E7"><label>(7)</label><mml:math id="M32"><mml:mtable class="eqnarray" columnalign="right center left"><mml:mtr><mml:mtd><mml:mi>S</mml:mi><mml:mo>=</mml:mo><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mtable style="text-align:axis;" equalrows="false" columnlines="none none none none none none none none none" equalcolumns="false" class="array"><mml:mtr><mml:mtd><mml:mfrac><mml:mrow><mml:mn>35</mml:mn></mml:mrow><mml:mrow><mml:mn>16</mml:mn></mml:mrow></mml:mfrac></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mfrac><mml:mrow><mml:mn>35</mml:mn></mml:mrow><mml:mrow><mml:mn>8</mml:mn></mml:mrow></mml:mfrac><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mfrac><mml:mrow><mml:mn>7</mml:mn></mml:mrow><mml:mrow><mml:mn>4</mml:mn></mml:mrow></mml:mfrac><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>4</mml:mn></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mfrac><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mn>6</mml:mn></mml:mrow></mml:mfrac><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>6</mml:mn></mml:mrow></mml:msup></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mfrac><mml:mrow><mml:mn>35</mml:mn></mml:mrow><mml:mrow><mml:mn>4</mml:mn></mml:mrow></mml:mfrac><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mn>7</mml:mn><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>4</mml:mn></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>6</mml:mn></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mfrac><mml:mrow><mml:mn>35</mml:mn></mml:mrow><mml:mrow><mml:mn>8</mml:mn></mml:mrow></mml:mfrac><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mfrac><mml:mrow><mml:mn>77</mml:mn></mml:mrow><mml:mrow><mml:mn>4</mml:mn></mml:mrow></mml:mfrac><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>4</mml:mn></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mfrac><mml:mrow><mml:mn>19</mml:mn></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:mfrac><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>6</mml:mn></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>8</mml:mn></mml:mrow></mml:msup></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mn>7</mml:mn><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>4</mml:mn></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mn>8</mml:mn><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>6</mml:mn></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mfrac><mml:mrow><mml:mn>4</mml:mn></mml:mrow><mml:mrow><mml:mn>3</mml:mn></mml:mrow></mml:mfrac><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>8</mml:mn></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mfrac><mml:mrow><mml:mn>7</mml:mn></mml:mrow><mml:mrow><mml:mn>4</mml:mn></mml:mrow></mml:mfrac><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>4</mml:mn></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mfrac><mml:mrow><mml:mn>19</mml:mn></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:mfrac><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>6</mml:mn></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mfrac><mml:mrow><mml:mn>17</mml:mn></mml:mrow><mml:mrow><mml:mn>3</mml:mn></mml:mrow></mml:mfrac><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>8</mml:mn></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mfrac><mml:mrow><mml:mn>2</mml:mn></mml:mrow><mml:mrow><mml:mn>3</mml:mn></mml:mrow></mml:mfrac><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>10</mml:mn></mml:mrow></mml:msup></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>6</mml:mn></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mfrac><mml:mrow><mml:mn>4</mml:mn></mml:mrow><mml:mrow><mml:mn>3</mml:mn></mml:mrow></mml:mfrac><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>8</mml:mn></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mfrac><mml:mrow><mml:mn>4</mml:mn></mml:mrow><mml:mrow><mml:mn>15</mml:mn></mml:mrow></mml:mfrac><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>10</mml:mn></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mfrac><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mn>6</mml:mn></mml:mrow></mml:mfrac><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>6</mml:mn></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>8</mml:mn></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mfrac><mml:mrow><mml:mn>2</mml:mn></mml:mrow><mml:mrow><mml:mn>3</mml:mn></mml:mrow></mml:mfrac><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>10</mml:mn></mml:mrow></mml:msup></mml:mtd><mml:mtd><mml:mn>0</mml:mn></mml:mtd><mml:mtd><mml:mfrac><mml:mrow><mml:mn>4</mml:mn></mml:mrow><mml:mrow><mml:mn>45</mml:mn></mml:mrow></mml:mfrac><mml:msup><mml:mrow><mml:mi>s</mml:mi></mml:mrow><mml:mrow><mml:mn>12</mml:mn></mml:mrow></mml:msup></mml:mtd></mml:mtr><mml:mtr></mml:mtr></mml:mtable></mml:mrow><mml:mo>]</mml:mo></mml:mrow><mml:mo>,</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>
<p>where <italic>s</italic> is the kernel width of the Gaussian filter that is assumed to be responsible for serial correlations in measurement or state noise. Here, the order of generalized coordinates (number of derivatives under consideration) is taken as six (<italic>S</italic> &#x02208; &#x0211D;<sup>7 &#x000D7; 7</sup>). For practical robotics applications, the measurement frequency is high, resulting in 0 &#x0003C; <italic>s</italic> &#x0003C; 1. It can be observed that the diagonal elements of <italic>S</italic> decreases because <italic>s</italic> &#x0003C; 1, resulting in a higher attention (or weighting) on the prediction errors from the lower derivatives when compared to the higher derivatives. The higher the noise color (i.e., <italic>s</italic> increases), the higher the weight given to the higher state derivatives (last diagonal elements of <italic>S</italic> increases). This reflects the fact that smooth fluctuations have more information content in their higher derivatives. Having established the potential importance of precision weighting in state estimation, we now turn to the estimation (i.e., learning) of precision in any given context.</p>
</sec>
<sec>
<title>4.2.3. System identification</title>
<p>This section shows how to optimize system identification by means of precision learning (Anil Meera and Wisse, <xref ref-type="bibr" rid="B3">2021</xref>; Meera and Wisse, <xref ref-type="bibr" rid="B80">2021b</xref>). Specifically, we show how to fuse prior knowledge about the dynamic model with the data to recover unknown parameters of the system through an attention mechanism. This involves the learning of the (1) parameters and (2) noise precisions. Our model &#x0201C;turns&#x0201D; the attention to the least precise parameters and uses the data to update those parameters to increase their precision. Hence, allowing faster parameter learning.</p>
<p>For the sake of clarity, we use again the mass-spring-damper system as the driving example (Section 4.2.1). We formalize system identification as evaluating the unknown parameters <italic>k</italic>, <italic>m</italic> and <italic>b</italic>, given the input <bold>u</bold>, the output <bold>y</bold>, and the general form of the linear system in Equation (2).</p>
<p><xref ref-type="fig" rid="F3">Figure 3</xref> depicts the process of learning unknown parameters (dotted boxes denote the processes inside the robot brain). The robot measures its position <italic>x</italic>(<italic>t</italic>) using its sensors (e.g., vision or range sensor). We assume that the robot has observed the behavior of a mass-spring-damper system before or a model is provided by the expert designer. However, some of the parameters are unknown. The robot can reuse the prior learned model of the system to relearn the new system. This can be realized by setting a high prior precision on the known parameters and a low prior precision on the unknown parameters. By means of precision learning, the robot uses the sensory signals to learn the parameter precision &#x003A0;<sup>&#x003B8;</sup>, thereby improving the confidence in the parameter estimates &#x003B8;. This directs the robot&#x00027;s attention toward the refinement of the parameters with least precision as they are the most uncertain. The requisite parameter learning proceeds by the gradient ascent of the free energy functional given in Equation (6). The parameter precision learning proceeds by tracking the negative curvature of <inline-formula><mml:math id="M33"><mml:mover accent="true"><mml:mrow><mml:mrow><mml:mi mathvariant="-tex-caligraphic">F</mml:mi></mml:mrow></mml:mrow><mml:mo>&#x00304;</mml:mo></mml:mover></mml:math></inline-formula> as <inline-formula><mml:math id="M34"><mml:msup><mml:mrow><mml:mi>&#x003A0;</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x003B8;</mml:mi></mml:mrow></mml:msup><mml:mo>=</mml:mo><mml:mo>-</mml:mo><mml:mfrac><mml:mrow><mml:msup><mml:mrow><mml:mi>&#x02202;</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup><mml:mover accent="true"><mml:mrow><mml:mrow><mml:mi mathvariant="-tex-caligraphic">F</mml:mi></mml:mrow></mml:mrow><mml:mo>&#x00304;</mml:mo></mml:mover></mml:mrow><mml:mrow><mml:mi>&#x02202;</mml:mi><mml:msup><mml:mrow><mml:mi>&#x003B8;</mml:mi></mml:mrow><mml:mrow><mml:mn>2</mml:mn></mml:mrow></mml:msup></mml:mrow></mml:mfrac></mml:math></inline-formula> (Anil Meera and Wisse, <xref ref-type="bibr" rid="B3">2021</xref>).</p>
<fig id="F3" position="float">
<label>Figure 3</label>
<caption><p>The schematic of the robot&#x00027;s attention mechanism for learning the least precise parameters of a given generative model of a mass-spring-damper system (shown in <bold>D</bold>). <bold>(A)</bold> Learning the conditional precision on parameters and the noise precision. <bold>(B)</bold> The free energy optimization helping to identify the unknown system parameters. <bold>(C)</bold> The parameter learning.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fnbot-16-896229-g0003.tif"/>
</fig>
<p>The learning process&#x02014;by means of variational free energy optimization (maximization)&#x02014;is shown in <xref ref-type="fig" rid="F3">Figure 3B</xref>. The learning involves two parallel processes: precision learning <xref ref-type="fig" rid="F3">(Figure 3A</xref>), and parameter learning (<xref ref-type="fig" rid="F3">Figure 3C</xref>). Precision learning comprises of parameter precision learning (top graph)&#x02014;i.e., identifying the precision of an approximate posterior density for the parameters being estimated&#x02014;and noise precision learning (bottom graph). The high prior precision on the known system parameters (0 and 1), and low prior precision on the unknown system parameters (<inline-formula><mml:math id="M35"><mml:mo>-</mml:mo><mml:mfrac><mml:mrow><mml:mi>k</mml:mi></mml:mrow><mml:mrow><mml:mi>m</mml:mi></mml:mrow></mml:mfrac><mml:mo>,</mml:mo><mml:mo>-</mml:mo><mml:mfrac><mml:mrow><mml:mi>b</mml:mi></mml:mrow><mml:mrow><mml:mi>m</mml:mi></mml:mrow></mml:mfrac></mml:math></inline-formula> and <inline-formula><mml:math id="M36"><mml:mfrac><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>m</mml:mi></mml:mrow></mml:mfrac></mml:math></inline-formula>, highlighted in blue) directs attention toward learning the unknown parameters and their precision. Note that in <xref ref-type="fig" rid="F3">Figure 3A</xref>, the precision on the three unknown parameters start from a low prior precision of <italic>P</italic><sup>&#x003B8;</sup> &#x0003D; 1 and increase with each iteration, whereas the precision of known parameters (0 and 1) remains a constant (3.3 &#x000D7; 10<sup>6</sup>). The noise precisions are learned simultaneously, which starts from a low prior precision of <italic>P</italic><sup>&#x003BB;</sup><sup><italic>w</italic></sup> &#x0003D; <italic>P</italic><sup>&#x003BB;</sup><sup><italic>z</italic></sup> &#x0003D; 1 and finally converges to the true noise precision (dotted black line). Both precisions are used to learn the three parameters of the system (<xref ref-type="fig" rid="F3">Figure 3B</xref>), which starts from randomly selected values within the range [&#x02212;2,2] and finally converges to the true parameter values of the system (<inline-formula><mml:math id="M37"><mml:msub><mml:mrow><mml:mi>&#x003B8;</mml:mi></mml:mrow><mml:mrow><mml:mn>3</mml:mn></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mo>-</mml:mo><mml:mfrac><mml:mrow><mml:mi>k</mml:mi></mml:mrow><mml:mrow><mml:mi>m</mml:mi></mml:mrow></mml:mfrac><mml:mo>=</mml:mo><mml:mo>-</mml:mo><mml:mn>0</mml:mn><mml:mo>.</mml:mo><mml:mn>5714</mml:mn></mml:math></inline-formula>, <inline-formula><mml:math id="M38"><mml:msub><mml:mrow><mml:mi>&#x003B8;</mml:mi></mml:mrow><mml:mrow><mml:mn>4</mml:mn></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mo>-</mml:mo><mml:mfrac><mml:mrow><mml:mi>b</mml:mi></mml:mrow><mml:mrow><mml:mi>m</mml:mi></mml:mrow></mml:mfrac><mml:mo>=</mml:mo><mml:mo>-</mml:mo><mml:mn>0</mml:mn><mml:mo>.</mml:mo><mml:mn>2857</mml:mn></mml:math></inline-formula> and <inline-formula><mml:math id="M39"><mml:msub><mml:mrow><mml:mi>&#x003B8;</mml:mi></mml:mrow><mml:mrow><mml:mn>6</mml:mn></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mn>1</mml:mn></mml:mrow><mml:mrow><mml:mi>m</mml:mi></mml:mrow></mml:mfrac><mml:mo>=</mml:mo><mml:mn>0</mml:mn><mml:mo>.</mml:mo><mml:mn>7143</mml:mn></mml:math></inline-formula>), denoted by black dotted lines. From an attentional perspective, the lower plot in <xref ref-type="fig" rid="F3">Figure 3A</xref> is particularly significant here. This is because the robot discovers the data are more informative than initially assumed, thereby leading to an increase in its estimate of the precision of the data-generating process. This means that the robot is not only using the data to optimize its beliefs about states and parameters (system identification), it is also using these data to optimize the way in which it assimilates these data.</p>
<p>In summary, precision-based attention, in the form of precision learning, helps the robot to accurately learn unknown parameters by fusing prior knowledge with new incoming data (sensory measurements), and attending to the least precise parameters.</p>
</sec>
<sec>
<title>4.2.4. Precision-modulated exploration and exploitation in system identification</title>
<p>Exploration and exploitation in the parameter space can be advantageous to robots during system identification. Precision-based attention&#x02014;here the prior precision&#x02014;allows a graceful balance between the two, mediated by the prior precision<xref ref-type="fn" rid="fn0006"><sup>6</sup></xref>. A very high prior precision encourages exploitation and biases the robot toward believing its priors, while a low prior precision encourages exploration and makes the robot sensitive to new information.</p>
<p>We use again the mass-spring-damper system example but with a different prior parameter precision <italic>P</italic><sup>&#x003B8;</sup>. The prior parameters are initialized at random and learned using optimization. <xref ref-type="fig" rid="F4">Figure 4B</xref> shows the increase in parameter estimation error (SSE) as the prior parameter precision <italic>P</italic><sup>&#x003B8;</sup> increases until it finally saturates. The bottom left region (circled in red) indicates the region where the prior precision is low, encouraging exploration with high attention on the sensory signals for learning the model. This region over-exposes the robot to its sensory signals by neglecting the prior parameters. The top right region (circled in red) indicates the biased robot where the prior precision is high, encouraging the robot to exploit its prior beliefs by retaining high attention on prior parameters. This regime biases the robot into being confident about its priors and disregarding new information from the sensory signals. Between those extreme regimes (blue curve) the prior precision balances the exploration-exploitation trade-off. <xref ref-type="fig" rid="F4">Figure 4A</xref> describes how increased attention to sensory signals helped the robot to recover from poor initial estimates of parameter values and converge toward the correct values (dotted black line). Conversely, in <xref ref-type="fig" rid="F4">Figure 4C</xref>, high attention on prior parameters did not help the robot to learn the correct parameter values.</p>
<fig id="F4" position="float">
<label>Figure 4</label>
<caption><p><bold>(A)</bold> Lower <italic>P</italic><sup>&#x02227;</sup>&#x003B8; gives a high exploration strategy across the parameter space. <bold>(B)</bold> Precision-based attention allows exploration and exploitation balanced model learning mediated by the prior precisions on the parameters <italic>P</italic><sup>&#x02227;</sup>&#x003B8;. <bold>(C)</bold> The higher the <italic>P</italic><sup>&#x02227;</sup>&#x003B8;, the higher the attention on prior parameters &#x003B7;<sup>&#x02227;</sup>&#x003B8; and the lower the attention on the sensory signals while learning.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fnbot-16-896229-g0004.tif"/>
</fig>
<p>These results establish that prior precision modeling allows balanced exploration and exploitation of parameter space during system identification. Although the results show that an over-exposed robot provides better parameter learning, we show&#x02014;in the next section&#x02014;that this is not always be the case.</p>
</sec>
<sec>
<title>4.2.5. Noise estimation</title>
<p>In real-world applications, sensory measurements are often highly noisy and unpredictable. Furthermore, the robot does not have access to the noise levels. Thus, it needs to learn the noise precision (<bold>&#x003A0;<sup><italic>z</italic></sup></bold>) for accurate estimation and robust control. Precision-based attention enables this learning. In what follows, we show how one can estimate <bold>&#x003A0;<sup><italic>z</italic></sup></bold> using noise precision learning and that biasing the robot to prior beliefs can be advantageous in highly noisy environments.</p>
<p>Consider again the mass-spring-damper system in <xref ref-type="fig" rid="F5">Figure 5B</xref>, where heavy rainfall/snow corrupts visual sensory signals. We evaluate the parameter estimation error under different noise conditions, using different levels of noise variances (inverse precision). For an over-exposed robot (only attending to sensory measurements), left plot of <xref ref-type="fig" rid="F5">Figure 5A</xref>, the estimation error increases as the noise strength increases, to a point where the error surpasses the error from a prior-biased robot. This shows that a robot, confident in its prior model, assigns low attention to sensory signals and outperforms an over-exposed robot that assigns high attention to sensory signals, in a highly noisy environment. The right plot of <xref ref-type="fig" rid="F5">Figure 5A</xref> shows the quality of noise precision learning for an over-exposed robot. It can be seen that all the data points in red lie close to the blue line, indicating that the estimated noise precision is close to the real noise precision. Therefore, the robot is capable of recovering the correct sensory noise levels even when the environment is extremely noisy, where accurate parameter estimation is difficult.</p>
<fig id="F5" position="float">
<label>Figure 5</label>
<caption><p>Simulations demonstrating how a biased robot could be advantageous, especially while learning in a highly noisy environment (shown in <bold>B</bold>). <bold>(A i)</bold> As the sensor noise increases, the quality of parameter estimation deteriorates to a point where an explorative robot generates higher parameter estimation errors than when compared to the biased robot that relies on its prior parameters. <bold>(A ii)</bold> However, the sensor noise estimation is accurate even for high noise environments, demonstrating the success of the attention mechanism using the noise precision.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fnbot-16-896229-g0005.tif"/>
</fig>
<p>These numerical results show that attention mechanism&#x02014;by means of noise precision learning&#x02014;allows the estimation of the noise levels in the environment and thereby protects against over-fitting or overconfident parameter estimation.</p>
<p><bold>Summary</bold>. We have shown how precision-based attention&#x02013;through precision modeling and learning&#x02013; yields to accurate robot state estimation, parameter identification and sensory noise estimation. In the next section, we discuss how action is generated in this framework.</p>
</sec>
</sec>
<sec>
<title>4.3. Precision-modulated action</title>
<p>Selecting the optimal sequence of actions to fulfill a task is essential for robotics (LaValle, <xref ref-type="bibr" rid="B72">2006</xref>). One of the most prominent challenges is to ensure robust behavior given the uncertainty emerging from a highly complex and dynamic real world, where the robots have to operate on. A proper attention system should provide action plans that resolve uncertainty and maximize information gain. For instance, it may minimize the information entropy, thereby encouraging repeated sensory measurements (observations) on high uncertainty sensory information.</p>
<p>Salience, which in neuroscience is sometimes identified as Bayesian surprise (i.e., divergence between prior and posterior), describes which information is relevant to process. We go one step further by defining the saliency map as the epistemic value of a particular action (Friston et al., <xref ref-type="bibr" rid="B40">2015</xref>). Thus, the (expected) divergence now becomes the mutual information under a particular action or plan. This makes the saliency map more sophisticated because it is an explicit measure of the reduction in uncertainty or mutual information associated with a particular action (i.e., active sampling), and more pragmatic because it tells you where to sample data next, given current Bayesian beliefs.</p>
<p>We first describe a precision representation usually used in information gathering problems and then how to directly generate action plans through precision optimization. Afterwards, we discuss the realization of the full-fledged model presented in the neuroscience section for active perception. We use the informative path planning (IPP) problem, described in <xref ref-type="fig" rid="F6">Figure 6</xref>, as an illustrative example to drive intuitions.</p>
<fig id="F6" position="float">
<label>Figure 6</label>
<caption><p>IPP problem for localizing human victims in an urban search and rescue scenario (Meera et al., <xref ref-type="bibr" rid="B77">2019</xref>). <bold>(A)</bold> Action: a UAV, in a realistic simulation environment, plans a finite look-ahead path to minimize the uncertainty of its human occupancy map (e.g., modeled as a Gaussian process) of the world. The planned path is then executed, during which the UAV flies and captures images at a constant measurement frequency. <bold>(B)</bold> Perception: after the data acquisition is complete, a human detection algorithm is executed to detect all the humans on the images. These detections are then fused into the UAV&#x00027;s human location map. The cycle is repeated until the uncertainty of the map is completely resolved (this usually implies enough area coverage and repeated measurements on uncertain locations). The ground truth of the human occupancy map and the UAV belief is shown in <bold>(B,C)</bold> respectively. The final map approaches the ground truth and all the seven humans on the ground are correctly detected.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fnbot-16-896229-g0006.tif"/>
</fig>
<sec>
<title>4.3.1. Precision maps as saliency</title>
<p>One of the popular approaches in information gathering problems is to model the information map as a distribution [e.g., using Gaussian processes (Hitz et al., <xref ref-type="bibr" rid="B50">2017</xref>)]. This is widely used in applications, such as a target search, coverage and navigation. The robot keeps track of an occupancy map and the associated uncertainty map (covariance matrix or inverse precision). While the occupancy map records the presence of the target on the map, the uncertainty map records the quality of those observations. The goal of the robot is to learn the distribution using some learning algorithm (Marchant and Ramos, <xref ref-type="bibr" rid="B76">2014</xref>). A popular strategy is to plan the robot path such that it minimizes the uncertainty of the map in future (Popovi&#x00107; et al., <xref ref-type="bibr" rid="B104">2017</xref>). In Section 4.3.2, we will show how we can use the map precision to perform active perception, i.e., optimize the robot path for maximal information gain. Optimizing the map precision drives the robot toward an exploratory behavior.</p>
</sec>
<sec>
<title>4.3.2. Precision optimization for action planning</title>
<p>To introduce precision-based saliency we use an exemplary application of search and rescue. The goal is to find all humans using an unmanned air vehicle (UAV) (Lanillos, <xref ref-type="bibr" rid="B64">2013</xref>; Lanillos et al., <xref ref-type="bibr" rid="B69">2014</xref>; Meera et al., <xref ref-type="bibr" rid="B77">2019</xref>; Rasouli et al., <xref ref-type="bibr" rid="B106">2020</xref>). We use precision for two purposes: (i) precision optimization for action planning (plan flight path) and (ii) precision learning for map refinement. In contrast to previous models of action selection within active inference in robotics (Lanillos et al., <xref ref-type="bibr" rid="B70">2021</xref>; Oliver et al., <xref ref-type="bibr" rid="B90">2021</xref>) here precision explicitly drives the agent behavior. <xref ref-type="fig" rid="F7">Figure 7</xref> describes the scenario in simulation. The seven human targets on the ground are correctly identified by the UAV. We can formalize the solution as the UAV actions (next flight path) that minimize the future uncertainties of the human occupancy map. In our precision-based attention scheme, this objective is equivalent to maximizing the posterior precision of the map. <xref ref-type="fig" rid="F8">Figure 8</xref> shows the reduction in map uncertainty after subsequent assimilation of the measurements (camera images from the UAV, processed by a human detector). The map (and precision) is learned using a recursive Kalman Filter by fusing the human detector outcome onto the map (and precision). The algorithm drives the UAV toward the least explored regions in the environment, defined by the precision map.</p>
<fig id="F7" position="float">
<label>Figure 7</label>
<caption><p>Finding humans with unmanned air vehicles (UAVs): an informative path planning (IPP) approach (Anil Meera, <xref ref-type="bibr" rid="B2">2018</xref>). The simulation environment on the left consists of a tall building at the center, surrounded by seven humans lying on the floor. The goal of the UAV is to compute the action sequence that allows maximum information gathering, i.e., the humans location uncertainty is minimized. On the right is the final occupancy map colored with the probability of finding a human at that location. It can be observed that all humans on the simulation environment were correctly detected by the robot.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fnbot-16-896229-g0007.tif"/>
</fig>
<fig id="F8" position="float">
<label>Figure 8</label>
<caption><p>Variance map of the probability distribution of people location (<xref ref-type="fig" rid="F7">Figure 7</xref>)&#x02014;inverse precision of human occupancy map. The plot sequence shows the reduction of map uncertainty (inverse precision) after measurements (Anil Meera, <xref ref-type="bibr" rid="B2">2018</xref>).</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fnbot-16-896229-g0008.tif"/>
</fig>
<p>Furthermore, <xref ref-type="fig" rid="F9">Figure 9</xref> shows an example of uncertainty resolution under false positives. In this case, human targets are moved to the bottom half of the map. The first measurement provides a wrong human detection with high uncertainty. However, after repeated measurements at the same location in the map the algorithm was capable of resolving this ambiguity, to finally learn the correct ground truth map. Hence, the sought behavior is to take actions that encourage repeated measurements at uncertain locations for reducing uncertainty.</p>
<fig id="F9" position="float">
<label>Figure 9</label>
<caption><p>The human occupancy map (probability to find humans at every location of the environment) at four time instances during the UAV flight showing ambiguity resolution. The ambiguity arising from imprecise sensor measurements (false positive) is resolved through repeated measurements at the same location. The plot sequence shows how the assimilation of the measurements updates the probability of the people being in each location of the map (Meera et al., <xref ref-type="bibr" rid="B77">2019</xref>).</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fnbot-16-896229-g0009.tif"/>
</fig>
<p>Although the IPP example illustrates how to generate control actions through precision optimization, the task, by construction, is constrained to explicitly reduce uncertainty. This is similar to the description of visual search described in Friston et al. (<xref ref-type="bibr" rid="B36">2012</xref>), where the location was chosen maximize information gain. Information gain (i.e., the Bayesian surprise expected following an action) is a key part of the expected free energy functional that underwrite action selection in active inference. In brief, expected free energy can be decomposed into two parts the first corresponds to the information gain above (a.k.a., epistemic value or affordance). The second corresponds to the expected log evidence or marginal likelihood of sensory samples (a.k.a., pragmatic value). When this likelihood is read as a prior preference, it contextualizes the imperative to reduce uncertainty by including a goal-directed, imperative. For example, in the search paradigm above, we could have formulated the problem in terms of reducing uncertainty about whether each location was occupied by a human or not. We could have then equipped the agent with prior preferences for observing humans.</p>
<p>In principle, this would have produced searching behavior until uncertainty had been resolved about the scene; after which, the robot would seek out humans; simply because, these are its preferred outcomes. In thinking about how this kind of neuroscience inspired or biomimetic approach could be implemented in robotics, one has to consider carefully, the precision afforded sensory inputs (i.e., the likelihood of sensory data, given its latent causes)&#x02014;and how this changes during robotic flight and periods of data gathering. This brings us back to the precision modulation and the temporal scheduling of searching and securing data. In the final section, we conclude with a brief discussion of how this might be implemented in future applications.</p>
</sec>
<sec>
<title>4.3.3. Precision-based active perception</title>
<p>In this section, we discuss the realization of a biomimetic brain-inspired model in relation to existing solutions in robotics in the context of path-planning. <xref ref-type="fig" rid="F10">Figure 10</xref> compares our proposed precision-modulated attention model&#x02014;from <xref ref-type="fig" rid="F1">Figure 1</xref>&#x02014;with the action-perception loop widely used in robotics. By analogy with eye saccades to the next visual sample, the UAV flies (action) over the environment to assimilate sensory data for an informed scene construction (perception). Once the flight time of the UAV is exhausted (similar to saccade window of the eye), the action is complete, after which the map is updated, and the next flight path is planned.</p>
<fig id="F10" position="float">
<label>Figure 10</label>
<caption><p>Precision-modulated attention model adapted to the action-perception loop in robotics. Each cycle consists of two steps: (1) action (planning and execution of a finite-time look ahead of the robot path for data collection) and (2) perception (learning using the collected data). This scheduling, using a finite time look-ahead plan, is quite common in real applications and of particular importance when processing is computationally expensive, e.g., slow rate of classification, non-scalable data fusion algorithms, Exponential planners, etc. However, the benefits of incorporating &#x0201C;optimal&#x0201D; scheduled loop driven by precision should be further studied.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fnbot-16-896229-g0010.tif"/>
</fig>
<p>In standard applications of active inference, the information gain is supplemented with expected log preferences to provide a complete expected free energy functional (Sajid et al., <xref ref-type="bibr" rid="B114">2021a</xref>). This accommodates the two kinds of uncertainty that actions and choices typically reduce. The first kind of uncertainty is inherent in unknowns in the environment. This is the information gain we have focused on above. The second kind of uncertainty corresponds to expected surprise, where surprise rests upon a priori expected or preferred outcomes. As noted above, equipping robots with both epistemic and pragmatic aspects to their action selection or planning could produce realistic and useful behavior that automatically resolves the exploration-exploitation dilemma. This follows because the expected free energy contains the optical mixture of epistemic (information-seeking) and pragmatic (i.e., preference seeking) components. Usually, after a period of exploration, the preference seeking components predominate because uncertainty has been resolved. Although expected free energy provides a fairly universal objective function for sentient behavior, it does not specify how to deploy behavior and sensory processing optimally. This brings us to the precision modulation model, inspired by neuroscientific considerations of attention and salience.</p>
<p>Hence, there are key differences between biological and robotic implementations of the search behavior. First, the use of oscillatory precision to modulate visual sampling and movement cycles, as opposed to arbitrary discrete action and perception steps currently used in robotics. Second, precision modulation influences both state estimation and action following the same uncertainty reduction principle. Importantly, our salience formulation speaks to selecting future data that reduces this uncertainty. For instance, we have shown&#x02014;in the information gathering IPP example described in the previous subsection&#x02014;that by optimizing precision we also optimize behavior.</p>
<p>Hence, there are key differences between biological and robotic implementations of the search behavior. First, the use of oscillatory precision to modulate visual sampling and movement cycles, as opposed to arbitrary discrete action and perception steps currently used in robotics. Second, precision modulation influences both state estimation and action following the same uncertainty reduction principle. Importantly, our salience formulation speaks to selecting future data that reduces this uncertainty. For instance, we have shown&#x02014;in the information gathering IPP example described in the previous subsection&#x02014;that by optimizing precision we also optimize behavior.</p>
<p>We argue the potential need and the advantages of realizing precision based temporal scheduling, as described the our brain-inspired model, for two practically relevant test cases: (<italic>i</italic>) learning dynamic models and (<italic>ii</italic>) information seeking applications.</p>
<p>In Section 4.2.4, we have shown how the exploration-exploitation trade-off can be mediated by the prior parameter precision during learning. However, the accuracy-precision curve (<xref ref-type="fig" rid="F4">Figure 4B</xref>) is often practically unavailable due to unknown true parameters values, challenging the modeling of prior precision. An alternative would be to use a precision based temporal scheduling mechanism to alternate between exploration and exploitation by means of a varying <italic>P</italic><sup>&#x003B8;</sup> (similar to <xref ref-type="fig" rid="F10">Figure 10</xref>) during learning, such that system identification is neither biased nor over exposed to sensory measurements. In <xref ref-type="fig" rid="F5">Figure 5A</xref>, we showed how noise levels influence estimation accuracy, and how biasing the robot by modeling <italic>P</italic><sup>&#x003B8;</sup> can be beneficial for highly noisy environments. A precision based temporal scheduling mechanism by means of a varying <italic>P</italic><sup>&#x003B8;</sup> could provide a balanced solution between a biased robot (that exploits its model) and an exploratory one.</p>
<p>Furthermore, temporal scheduling, in the same way that eye saccades are generated, can be adapted for information gathering applications, such as target search, simultaneous localization and mapping, environment monitoring, etc. For instance, introducing precision-modulation scheduling for solving the IPP, and scheduling perception (map learning) and action (UAV flight). Precision modulation will switch between action and perception: when the precision is high, perception occurs (c.f., visual sampling), and when the precision is low, action occurs (c.f., eye movements). This switch, which is often implemented in the robotics literature using a budget for flight time, will be now dictated by precision dynamics.</p>
<p>In short, we have sketched the basis for a future realization of precision-based active perception, where the robot computes the actions to minimize the expected uncertainty. While most attentional mechanisms in robotics are limited to providing a &#x0201C;saliency&#x0201D; map highlighting the most relevant features, our attention mechanism proposes a general scheduling mechanism with action in the loop with perception, both driven by precision.</p>
</sec>
</sec>
</sec>
<sec id="s5">
<title>5. Concluding remarks</title>
<p>We have considered attention and salience as two distinct processes that rest upon oscillatory precision control processes. Accordingly, they require particular temporal considerations: attention to reliably estimate latent states from current sensory data and salience for uncertainty reduction regarding future data samples. This formulation addresses visual search from a first principles (Bayesian) account of how these mechanisms might manifest&#x02014;and the circular causality that undergirds them <italic>via</italic> a rhythmic theta-coupling. Crucially, we have revisited the definition of salience from the visual neurosciences; where it is read as Bayesian surprise (i.e., the Kullback Leibler divergence between prior and posterior beliefs). We took this one step further and defined salience as the expected Bayesian surprise (i.e., epistemic value) of a particular action (e.g., sampling this set of data) (Friston et al., <xref ref-type="bibr" rid="B43">2017b</xref>; Sajid et al., <xref ref-type="bibr" rid="B114">2021a</xref>). Formulating salience as the expected divergence renders it the mutual information under a particular action (or action trajectory) (Friston et al., <xref ref-type="bibr" rid="B37">2021</xref>),&#x02014;and highlights its role in encoding working memory (Parr and Friston, <xref ref-type="bibr" rid="B97">2017b</xref>). For brevity, our narrative was centered around visual attention and its realization <italic>via</italic> eye movements. However, this model does not strictly need to be limited to visual information processing, because it addresses sensorimotor and auditory processing in general. This means it explains how action and perception can be coupled in other sensory modalities. For instance, Tomassini et al. (<xref ref-type="bibr" rid="B125">2017</xref>) showed that visual information is coupled with finger movements at a theta rhythm.</p>
<p>The point of contact with the robotics use of salience emerges because the co-variation between a particular parameterisation and the inputs is a measure of the mutual information between the data and its estimated causes. In this sense, both definitions of salience reflect the mutual information&#x02014;or information about a particular representation of a (latent) cause&#x02014;afforded by an observation or consequence. However, our formulation is more sophisticated. Briefly, because it is an explicit measure of the reduction in uncertainty (i.e., mutual information) associated with a particular action (i.e., active sampling) and specifies where to sample data next, given current Bayesian beliefs. These processes (attention and salience) are a consequence of precision of beliefs over distinct model parameters. Explicitly, attention contends with precision over the causes of (current) outcomes and salience contends with beliefs about the data that has to be acquired and precision over beliefs about actions that dictate it. Since both processes can be linked <italic>via</italic> precision manipulation, the crucial thing is the precision that differentiates whether the agent acquires new information (under high precision) or resolves uncertainty by moving (low precision).</p>
<p>The focus of this work has been to illustrate the importance of optimizing precision at various places in generative models used for data assimilation, system identification and active sensing. A key point&#x02014;implicit in these demonstrations - rests upon the mean field approximation used in all applications. Crucially, this means that getting the precision right matters, because updating posterior estimates of states, parameters and precisions all depend upon each other. This may be particularly prescient for making the most sense of samples that maximizes information gain. In other words, although attention and salience are separable optimization processes, they depend upon each other during active sensing. This was the focus of our final numerical studies of action planning.</p>
<p>To face-validate our formulation, we evaluated precision-modulated attentional processes in the robotic domain. We presented numerical examples to show how precision manipulation underwrites accurate state and noise estimation (e.g., selecting relevant information), as well as allowing system identification (e.g., learning unknown parameters of the dynamics). We also showed how one can use precision-based optimization to solve interesting problems; like the informative path planning in search and rescue scenarios. Thus, in contrast to previous uses of attention in robotics, we placed attention and saliency as integral processes for efficient gathering and processing of sensory information. Accordingly, &#x02018;attention&#x00027; is not only about filtering the current flow of information from the sensors but performing those actions that minimize expected uncertainty. Still, the full potential of our proposal has yet to be realized, as the precision-based attention should be able to account for prior preferences beyond the IPP problem (e.g., localizing people using UAVs). Finally, we briefly considered the realization of temporal scheduling for information gathering tasks, opening up interesting lines of research to provide robots with biologically plausible attention.</p>
</sec>
<sec sec-type="data-availability" id="s6">
<title>Data availability statement</title>
<p>The original contributions presented in the study are included in the article/supplementary material, further inquiries can be directed to the corresponding author/s.</p>
</sec>
<sec id="s7">
<title>Author contributions</title>
<p>AA and FN are responsible for the novel account and its translation to robotics. AA, FN, PL, and NS wrote the manuscript. All authors contributed to conception and design of the work, manuscript revision, read, and approved the submitted version.</p>
</sec>
<sec sec-type="funding-information" id="s8">
<title>Funding</title>
<p>The open access publication of the manuscript was funded by the TU Delft Library. FN was funded by the Serotonin &#x00026; Beyond project (953327). PL was partially supported by Spikeference project, Human Brain Project Specific Grant Agreement 3 (ID: 945539). NS was funded by the Medical Research Council (MR/S502522/1) and 2021-2022 Microsoft Ph.D. Fellowship. KF is supported by funding for the Wellcome Centre for Human Neuroimaging (Ref: 205103/Z/16/Z) and a Canada UK Artificial Intelligence Initiative (Ref: ES/T01279X/1).</p>
</sec>
<sec id="s9">
<title>Conflict of interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec sec-type="disclaimer" id="s10">
<title>Publisher&#x00027;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
</body>
<back>
<ref-list>
<title>References</title>
<ref id="B1">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ahnelt</surname> <given-names>P.</given-names></name></person-group> (<year>1998</year>). <article-title>The photoreceptor mosaic</article-title>. <source>Eye</source> <volume>12</volume>, <fpage>531</fpage>&#x02013;<lpage>540</lpage>. <pub-id pub-id-type="doi">10.1038/eye.1998.142</pub-id><pub-id pub-id-type="pmid">9775214</pub-id></citation></ref>
<ref id="B2">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Anil Meera</surname> <given-names>A.</given-names></name></person-group> (<year>2018</year>). <source>Informative path planning for search and rescue using a uav</source> (<publisher-loc>Maters thesis</publisher-loc>). TU Delft.</citation>
</ref>
<ref id="B3">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Anil Meera</surname> <given-names>A.</given-names></name> <name><surname>Wisse</surname> <given-names>M.</given-names></name></person-group> (<year>2021</year>). <article-title>Dynamic expectation maximization algorithm for estimation of linear systems with colored noise</article-title>. <source>Entropy</source> <volume>23</volume>, <fpage>1306</fpage>. <pub-id pub-id-type="doi">10.3390/e23101306</pub-id><pub-id pub-id-type="pmid">34682030</pub-id></citation></ref>
<ref id="B4">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Atrey</surname> <given-names>A.</given-names></name> <name><surname>Clary</surname> <given-names>K.</given-names></name> <name><surname>Jensen</surname> <given-names>D.</given-names></name></person-group> (<year>2019</year>). <article-title>Exploratory not explanatory: counterfactual analysis of saliency maps for deep reinforcement learning</article-title>. <source>arXiv[preprint].arXiv:1912.05743</source>. <pub-id pub-id-type="doi">10.48550/arXiv.1912.05743</pub-id></citation>
</ref>
<ref id="B5">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Axmacher</surname> <given-names>N.</given-names></name> <name><surname>Henseler</surname> <given-names>M. M.</given-names></name> <name><surname>Jensen</surname> <given-names>O.</given-names></name> <name><surname>Weinreich</surname> <given-names>I.</given-names></name> <name><surname>Elger</surname> <given-names>C. E.</given-names></name> <name><surname>Fell</surname> <given-names>J.</given-names></name></person-group> (<year>2010</year>). <article-title>Cross-frequency coupling supports multi-item working memory in the human hippocampus</article-title>. <source>Proc. Natl. Acad. Sci. U.S.A</source>. <volume>107</volume>, <fpage>3228</fpage>&#x02013;<lpage>3233</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.0911531107</pub-id><pub-id pub-id-type="pmid">20133762</pub-id></citation></ref>
<ref id="B6">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bajcsy</surname> <given-names>R.</given-names></name> <name><surname>Aloimonos</surname> <given-names>Y.</given-names></name> <name><surname>Tsotsos</surname> <given-names>J. K.</given-names></name></person-group> (<year>2018</year>). <article-title>Revisiting active perception</article-title>. <source>Auton. Robots</source> <volume>42</volume>, <fpage>177</fpage>&#x02013;<lpage>196</lpage>. <pub-id pub-id-type="doi">10.1007/s10514-017-9615-3</pub-id><pub-id pub-id-type="pmid">31983809</pub-id></citation></ref>
<ref id="B7">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Balestrieri</surname> <given-names>E.</given-names></name> <name><surname>Ronconi</surname> <given-names>L.</given-names></name> <name><surname>Melcher</surname> <given-names>D.</given-names></name></person-group> (<year>2021</year>). <article-title>Shared resources between visual attention and visual working memory are allocated through rhythmic sampling</article-title>. <source>Eur. J. Neurosci</source>. <volume>55</volume>, <fpage>3040</fpage>&#x02013;<lpage>3053</lpage>. <pub-id pub-id-type="doi">10.1111/EJN.15264/v2/response1</pub-id><pub-id pub-id-type="pmid">33942394</pub-id></citation></ref>
<ref id="B8">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Begum</surname> <given-names>M.</given-names></name> <name><surname>Karray</surname> <given-names>F.</given-names></name></person-group> (<year>2010</year>). <article-title>Visual attention for robotic cognition: a survey</article-title>. <source>IEEE Trans. Auton. Ment. Dev</source>. <volume>3</volume>, <fpage>92</fpage>&#x02013;<lpage>105</lpage>. <pub-id pub-id-type="doi">10.1109/TAMD.2010.2096505</pub-id></citation>
</ref>
<ref id="B9">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Benedetto</surname> <given-names>A.</given-names></name> <name><surname>Morrone</surname> <given-names>M. C.</given-names></name></person-group> (<year>2017</year>). <article-title>Saccadic suppression is embedded within extended oscillatory modulation of sensitivity</article-title>. <source>J. Neurosci</source>. <volume>37</volume>, <fpage>3661</fpage>&#x02013;<lpage>3670</lpage>. <pub-id pub-id-type="doi">10.1523/JNEUROSCI.2390-16.2016</pub-id><pub-id pub-id-type="pmid">28270573</pub-id></citation></ref>
<ref id="B10">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Benedetto</surname> <given-names>A.</given-names></name> <name><surname>Morrone</surname> <given-names>M. C.</given-names></name> <name><surname>Tomassini</surname> <given-names>A.</given-names></name></person-group> (<year>2020</year>). <article-title>The common rhythm of action and perception</article-title>. <source>J. Cogn. Neurosci</source>. <volume>32</volume>, <fpage>187</fpage>&#x02013;<lpage>200</lpage>. <pub-id pub-id-type="doi">10.1162/jocn_a_01436</pub-id><pub-id pub-id-type="pmid">31210564</pub-id></citation></ref>
<ref id="B11">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Borji</surname> <given-names>A.</given-names></name> <name><surname>Itti</surname> <given-names>L.</given-names></name></person-group> (<year>2012</year>). <article-title>State-of-the-art in visual attention modeling</article-title>. <source>IEEE Trans. Pattern Anal. Mach. Intell</source>. <volume>35</volume>, <fpage>185</fpage>&#x02013;<lpage>207</lpage>. <pub-id pub-id-type="doi">10.1109/TPAMI.2012.89</pub-id><pub-id pub-id-type="pmid">22487985</pub-id></citation></ref>
<ref id="B12">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bos</surname> <given-names>F.</given-names></name> <name><surname>Meera</surname> <given-names>A. A.</given-names></name> <name><surname>Benders</surname> <given-names>D.</given-names></name> <name><surname>Wisse</surname> <given-names>M.</given-names></name></person-group> (<year>2021</year>). <article-title>Free energy principle for state and input estimation of a quadcopter flying in wind</article-title>. <source>arXiv[preprint].arXiv:2109.12052</source>. <pub-id pub-id-type="doi">10.48550/arXiv.2109.12052</pub-id></citation>
</ref>
<ref id="B13">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Brown</surname> <given-names>H.</given-names></name> <name><surname>Adams</surname> <given-names>R. A.</given-names></name> <name><surname>Parees</surname> <given-names>I.</given-names></name> <name><surname>Edwards</surname> <given-names>M.</given-names></name> <name><surname>Friston</surname> <given-names>K.</given-names></name></person-group> (<year>2013</year>). <article-title>Active inference, sensory attenuation and illusions</article-title>. <source>Cogn. Process</source>. <volume>14</volume>, <fpage>411</fpage>&#x02013;<lpage>427</lpage>. <pub-id pub-id-type="doi">10.1007/s10339-013-0571-3</pub-id><pub-id pub-id-type="pmid">23744445</pub-id></citation></ref>
<ref id="B14">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Brzezicka</surname> <given-names>A.</given-names></name> <name><surname>Kami&#x00144;ski</surname> <given-names>J.</given-names></name> <name><surname>Reed</surname> <given-names>C. M.</given-names></name> <name><surname>Chung</surname> <given-names>J. M.</given-names></name> <name><surname>Mamelak</surname> <given-names>A. N.</given-names></name> <name><surname>Rutishauser</surname> <given-names>U.</given-names></name></person-group> (<year>2019</year>). <article-title>Working memory load-related theta power decreases in dorsolateral prefrontal cortex predict individual differences in performance</article-title>. <source>J. Cogn. Neurosci</source>. <volume>31</volume>, <fpage>1290</fpage>&#x02013;<lpage>1307</lpage>. <pub-id pub-id-type="doi">10.1162/jocn_a_01417</pub-id><pub-id pub-id-type="pmid">31037988</pub-id></citation></ref>
<ref id="B15">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Busch</surname> <given-names>N. A.</given-names></name> <name><surname>VanRullen</surname> <given-names>R.</given-names></name></person-group> (<year>2010</year>). <article-title>Spontaneous eeg oscillations reveal periodic sampling of visual attention</article-title>. <source>Proc. Natl. Acad. Sci. U.S.A</source>. <volume>107</volume>, <fpage>16048</fpage>&#x02013;<lpage>16053</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.1004801107</pub-id><pub-id pub-id-type="pmid">20805482</pub-id></citation></ref>
<ref id="B16">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Butko</surname> <given-names>N. J.</given-names></name> <name><surname>Zhang</surname> <given-names>L.</given-names></name> <name><surname>Cottrell</surname> <given-names>G. W.</given-names></name> <name><surname>Movellan</surname> <given-names>J. R.</given-names></name></person-group> (<year>2008</year>). <article-title>Visual saliency model for robot cameras,</article-title> in <source>2008 IEEE International Conference on Robotics and Automation</source> (<publisher-loc>Pasadena, CA</publisher-loc>: <publisher-name>IEEE</publisher-name>), <fpage>2398</fpage>&#x02013;<lpage>2403</lpage>.<pub-id pub-id-type="pmid">35538154</pub-id></citation></ref>
<ref id="B17">
<citation citation-type="web"><person-group person-group-type="author"><name><surname>Bylinskii</surname> <given-names>Z.</given-names></name> <name><surname>Judd</surname> <given-names>T.</given-names></name> <name><surname>Borji</surname> <given-names>A.</given-names></name> <name><surname>Itti</surname> <given-names>L.</given-names></name> <name><surname>Durand</surname> <given-names>F.</given-names></name> <name><surname>Oliva</surname> <given-names>A.</given-names></name> <etal/></person-group>. (<year>2019</year>). <source>Mit Saliency Benchmark</source>. Available online at: <ext-link ext-link-type="uri" xlink:href="http://saliency.mit.edu/">http://saliency.mit.edu/</ext-link></citation>
</ref>
<ref id="B18">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Clark</surname> <given-names>A.</given-names></name></person-group> (<year>2013</year>). <article-title>The many faces of precision (replies to commentaries on &#x0201C;whatever next? neural prediction, situated agents, and the future of cognitive science&#x0201D;)</article-title>. <source>Front. Psychol</source>. <volume>4</volume>, <fpage>270</fpage>. <pub-id pub-id-type="doi">10.3389/fpsyg.2013.00270</pub-id><pub-id pub-id-type="pmid">23734133</pub-id></citation></ref>
<ref id="B19">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Crevecoeur</surname> <given-names>F.</given-names></name> <name><surname>Kording</surname> <given-names>K. P.</given-names></name></person-group> (<year>2017</year>). <article-title>Saccadic suppression as a perceptual consequence of efficient sensorimotor estimation</article-title>. <source>Elife</source> <volume>6</volume>, <fpage>e25073</fpage>. <pub-id pub-id-type="doi">10.7554/eLife.25073</pub-id><pub-id pub-id-type="pmid">28463113</pub-id></citation></ref>
<ref id="B20">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Da Costa</surname> <given-names>L.</given-names></name> <name><surname>Parr</surname> <given-names>T.</given-names></name> <name><surname>Sajid</surname> <given-names>N.</given-names></name> <name><surname>Veselic</surname> <given-names>S.</given-names></name> <name><surname>Neacsu</surname> <given-names>V.</given-names></name> <name><surname>Friston</surname> <given-names>K.</given-names></name></person-group> (<year>2020</year>). <article-title>Active inference on discrete state-spaces: a synthesis</article-title>. <source>J. Math. Psychol</source>. <volume>99</volume>, <fpage>102447</fpage>. <pub-id pub-id-type="doi">10.1016/j.jmp.2020.102447</pub-id><pub-id pub-id-type="pmid">33343039</pub-id></citation></ref>
<ref id="B21">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Desimone</surname> <given-names>R.</given-names></name></person-group> (<year>1996</year>). <article-title>Neural mechanisms for visual memory and their role in attention</article-title>. <source>Proc. Natl. Acad. Sci. U.S.A</source>. <volume>93</volume>, <fpage>13494</fpage>&#x02013;<lpage>13499</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.93.24.13494</pub-id><pub-id pub-id-type="pmid">8942962</pub-id></citation></ref>
<ref id="B22">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dugu&#x000E9;</surname> <given-names>L.</given-names></name> <name><surname>Marque</surname> <given-names>P.</given-names></name> <name><surname>VanRullen</surname> <given-names>R.</given-names></name></person-group> (<year>2015</year>). <article-title>Theta oscillations modulate attentional search performance periodically</article-title>. <source>J. Cogn. Neurosci</source>. <volume>27</volume>, <fpage>945</fpage>&#x02013;<lpage>958</lpage>. <pub-id pub-id-type="doi">10.1162/jocn_a_00755</pub-id><pub-id pub-id-type="pmid">25390199</pub-id></citation></ref>
<ref id="B23">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dugu&#x000E9;</surname> <given-names>L.</given-names></name> <name><surname>Roberts</surname> <given-names>M.</given-names></name> <name><surname>Carrasco</surname> <given-names>M.</given-names></name></person-group> (<year>2016</year>). <article-title>Attention reorients periodically</article-title>. <source>Curr. Biol</source>. <volume>26</volume>, <fpage>1595</fpage>&#x02013;<lpage>1601</lpage>. <pub-id pub-id-type="doi">10.1016/j.cub.2016.04.046</pub-id><pub-id pub-id-type="pmid">27265395</pub-id></citation></ref>
<ref id="B24">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Eldar</surname> <given-names>E.</given-names></name> <name><surname>Cohen</surname> <given-names>J. D.</given-names></name> <name><surname>Niv</surname> <given-names>Y.</given-names></name></person-group> (<year>2013</year>). <article-title>The effects of neural gain on attention and learning</article-title>. <source>Nat Neurosci</source>. <volume>16</volume>, <fpage>1146</fpage>&#x02013;<lpage>1153</lpage>. <pub-id pub-id-type="doi">10.1038/nn.3428</pub-id><pub-id pub-id-type="pmid">23770566</pub-id></citation></ref>
<ref id="B25">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Feldman</surname> <given-names>H.</given-names></name> <name><surname>Friston</surname> <given-names>K.</given-names></name></person-group> (<year>2010</year>). <article-title>Attention, uncertainty, and free-energy</article-title>. <source>Front. Hum. Neurosci</source>. <volume>4</volume>, <fpage>215</fpage>. <pub-id pub-id-type="doi">10.3389/fnhum.2010.00215</pub-id><pub-id pub-id-type="pmid">21160551</pub-id></citation></ref>
<ref id="B26">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ferreira</surname> <given-names>J. F.</given-names></name> <name><surname>Dias</surname> <given-names>J.</given-names></name></person-group> (<year>2014</year>). <article-title>Attentional mechanisms for socially interactive robots-a survey</article-title>. <source>IEEE Trans. Auton. Ment. Dev</source>. <volume>6</volume>, <fpage>110</fpage>&#x02013;<lpage>125</lpage>. <pub-id pub-id-type="doi">10.1109/TAMD.2014.2303072</pub-id></citation>
</ref>
<ref id="B27">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fiebelkorn</surname> <given-names>I. C.</given-names></name> <name><surname>Kastner</surname> <given-names>S.</given-names></name></person-group> (<year>2019</year>). <article-title>A rhythmic theory of attention</article-title>. <source>Trends Cogn. Sci</source>. <volume>23</volume>, <fpage>87</fpage>&#x02013;<lpage>101</lpage>. <pub-id pub-id-type="doi">10.1016/j.tics.2018.11.009</pub-id><pub-id pub-id-type="pmid">30591373</pub-id></citation></ref>
<ref id="B28">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fiebelkorn</surname> <given-names>I. C.</given-names></name> <name><surname>Kastner</surname> <given-names>S.</given-names></name></person-group> (<year>2020</year>). <article-title>Functional specialization in the attention network</article-title>. <source>Annu. Rev. Psychol</source>. <volume>71</volume>, <fpage>221</fpage>&#x02013;<lpage>249</lpage>. <pub-id pub-id-type="doi">10.1146/annurev-psych-010418-103429</pub-id><pub-id pub-id-type="pmid">31514578</pub-id></citation></ref>
<ref id="B29">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fiebelkorn</surname> <given-names>I. C.</given-names></name> <name><surname>Kastner</surname> <given-names>S.</given-names></name></person-group> (<year>2021</year>). <article-title>Spike timing in the attention network predicts behavioral outcome prior to target selection</article-title>. <source>Neuron</source> <volume>109</volume>, <fpage>177</fpage>&#x02013;<lpage>188</lpage>. <pub-id pub-id-type="doi">10.1016/j.neuron.2020.09.039</pub-id><pub-id pub-id-type="pmid">33098762</pub-id></citation></ref>
<ref id="B30">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fiebelkorn</surname> <given-names>I. C.</given-names></name> <name><surname>Pinsk</surname> <given-names>M. A.</given-names></name> <name><surname>Kastner</surname> <given-names>S.</given-names></name></person-group> (<year>2018</year>). <article-title>A dynamic interplay within the frontoparietal network underlies rhythmic spatial attention</article-title>. <source>Neuron</source> <volume>99</volume>, <fpage>842</fpage>&#x02013;<lpage>853</lpage>. <pub-id pub-id-type="doi">10.1016/j.neuron.2018.07.038</pub-id><pub-id pub-id-type="pmid">30138590</pub-id></citation></ref>
<ref id="B31">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fiebelkorn</surname> <given-names>I. C.</given-names></name> <name><surname>Pinsk</surname> <given-names>M. A.</given-names></name> <name><surname>Kastner</surname> <given-names>S.</given-names></name></person-group> (<year>2019</year>). <article-title>The mediodorsal pulvinar coordinates the macaque fronto-parietal network during rhythmic spatial attention</article-title>. <source>Nat. Commun</source>. <volume>10</volume>, <fpage>1</fpage>&#x02013;<lpage>15</lpage>. <pub-id pub-id-type="doi">10.1038/s41467-018-08151-4</pub-id><pub-id pub-id-type="pmid">30644391</pub-id></citation></ref>
<ref id="B32">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fine</surname> <given-names>M. S.</given-names></name> <name><surname>Minnery</surname> <given-names>B. S.</given-names></name></person-group> (<year>2009</year>). <article-title>Visual salience affects performance in a working memory task</article-title>. <source>J. Neurosci</source>. <volume>29</volume>, <fpage>8016</fpage>&#x02013;<lpage>8021</lpage>. <pub-id pub-id-type="doi">10.1523/JNEUROSCI.5503-08.2009</pub-id><pub-id pub-id-type="pmid">19553441</pub-id></citation></ref>
<ref id="B33">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Frintrop</surname> <given-names>S.</given-names></name> <name><surname>Jensfelt</surname> <given-names>P.</given-names></name></person-group> (<year>2008</year>). <article-title>Attentional landmarks and active gaze control for visual slam</article-title>. <source>IEEE Trans. Rob</source>. <volume>24</volume>, <fpage>1054</fpage>&#x02013;<lpage>1065</lpage>. <pub-id pub-id-type="doi">10.1109/TRO.2008.2004977</pub-id></citation>
</ref>
<ref id="B34">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Frintrop</surname> <given-names>S.</given-names></name></person-group> (<year>2006</year>). <source>VOCUS: A Visual Attention System for Object Detection and Goal-Directed Search, Vol. 3899</source>. <publisher-loc>Berlin</publisher-loc>: <publisher-name>Springer</publisher-name>.</citation>
</ref>
<ref id="B35">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Friston</surname> <given-names>K.</given-names></name></person-group> (<year>2010</year>). <article-title>The free-energy principle: a unified brain theory?</article-title> <source>Nat. Rev. Neurosci</source>. <volume>11</volume>, <fpage>127</fpage>&#x02013;<lpage>138</lpage>. <pub-id pub-id-type="doi">10.1038/nrn2787</pub-id><pub-id pub-id-type="pmid">20068583</pub-id></citation></ref>
<ref id="B36">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Friston</surname> <given-names>K.</given-names></name> <name><surname>Adams</surname> <given-names>R.</given-names></name> <name><surname>Perrinet</surname> <given-names>L.</given-names></name> <name><surname>Breakspear</surname> <given-names>M.</given-names></name></person-group> (<year>2012</year>). <article-title>Perceptions as hypotheses: saccades as experiments</article-title>. <source>Front. Psychol</source>. <volume>3</volume>, <fpage>151</fpage>. <pub-id pub-id-type="doi">10.3389/fpsyg.2012.00151</pub-id><pub-id pub-id-type="pmid">22654776</pub-id></citation></ref>
<ref id="B37">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Friston</surname> <given-names>K.</given-names></name> <name><surname>Da Costa</surname> <given-names>L.</given-names></name> <name><surname>Hafner</surname> <given-names>D.</given-names></name> <name><surname>Hesp</surname> <given-names>C.</given-names></name> <name><surname>Parr</surname> <given-names>T.</given-names></name></person-group> (<year>2021</year>). <article-title>Sophisticated inference</article-title>. <source>Neural Comput</source>. <volume>33</volume>, <fpage>713</fpage>&#x02013;<lpage>763</lpage>. <pub-id pub-id-type="doi">10.1162/neco_a_01351</pub-id><pub-id pub-id-type="pmid">33626312</pub-id></citation></ref>
<ref id="B38">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Friston</surname> <given-names>K.</given-names></name> <name><surname>FitzGerald</surname> <given-names>T.</given-names></name> <name><surname>Rigoli</surname> <given-names>F.</given-names></name> <name><surname>Schwartenbeck</surname> <given-names>P.</given-names></name> <name><surname>Pezzulo</surname> <given-names>G.</given-names></name></person-group> (<year>2017a</year>). <article-title>Active inference: a process theory</article-title>. <source>Neural Comput</source>. <volume>29</volume>, <fpage>1</fpage>&#x02013;<lpage>49</lpage>. <pub-id pub-id-type="doi">10.1162/NECO_a_00912</pub-id><pub-id pub-id-type="pmid">27870614</pub-id></citation></ref>
<ref id="B39">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Friston</surname> <given-names>K.</given-names></name> <name><surname>Mattout</surname> <given-names>J.</given-names></name> <name><surname>Kilner</surname> <given-names>J.</given-names></name></person-group> (<year>2011</year>). <article-title>Action understanding and active inference</article-title>. <source>Biol. Cybern</source>. <volume>104</volume>, <fpage>137</fpage>&#x02013;<lpage>160</lpage>. <pub-id pub-id-type="doi">10.1007/s00422-011-0424-z</pub-id><pub-id pub-id-type="pmid">21327826</pub-id></citation></ref>
<ref id="B40">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Friston</surname> <given-names>K.</given-names></name> <name><surname>Rigoli</surname> <given-names>F.</given-names></name> <name><surname>Ognibene</surname> <given-names>D.</given-names></name> <name><surname>Mathys</surname> <given-names>C.</given-names></name> <name><surname>Fitzgerald</surname> <given-names>T.</given-names></name> <name><surname>Pezzulo</surname> <given-names>G.</given-names></name></person-group> (<year>2015</year>). <article-title>Active inference and epistemic value</article-title>. <source>Cogn. Neurosci</source>. <volume>6</volume>, <fpage>187</fpage>&#x02013;<lpage>214</lpage>. <pub-id pub-id-type="doi">10.1080/17588928.2015.1020053</pub-id><pub-id pub-id-type="pmid">25689102</pub-id></citation></ref>
<ref id="B41">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Friston</surname> <given-names>K.</given-names></name> <name><surname>Stephan</surname> <given-names>K.</given-names></name> <name><surname>Li</surname> <given-names>B.</given-names></name> <name><surname>Daunizeau</surname> <given-names>J.</given-names></name></person-group> (<year>2010</year>). <article-title>Generalised filtering</article-title>. <source>Math. Problems Eng</source>. <volume>2010</volume>, <fpage>621670</fpage>. <pub-id pub-id-type="doi">10.1155/2010/621670</pub-id></citation>
</ref>
<ref id="B42">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Friston</surname> <given-names>K. J.</given-names></name> <name><surname>Harrison</surname> <given-names>L.</given-names></name> <name><surname>Penny</surname> <given-names>W.</given-names></name></person-group> (<year>2003</year>). <article-title>Dynamic causal modelling</article-title>. <source>Neuroimage</source> <volume>19</volume>, <fpage>1273</fpage>&#x02013;<lpage>1302</lpage>. <pub-id pub-id-type="doi">10.1016/S1053-8119(03)00202-7</pub-id><pub-id pub-id-type="pmid">12948688</pub-id></citation></ref>
<ref id="B43">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Friston</surname> <given-names>K. J.</given-names></name> <name><surname>Lin</surname> <given-names>M.</given-names></name> <name><surname>Frith</surname> <given-names>C. D.</given-names></name> <name><surname>Pezzulo</surname> <given-names>G.</given-names></name> <name><surname>Hobson</surname> <given-names>J. A.</given-names></name> <name><surname>Ondobaka</surname> <given-names>S.</given-names></name></person-group> (<year>2017b</year>). <article-title>Active inference, curiosity and insight</article-title>. <source>Neural Comput</source>. <volume>29</volume>, <fpage>2633</fpage>&#x02013;<lpage>2683</lpage>. <pub-id pub-id-type="doi">10.1162/neco_a_00999</pub-id><pub-id pub-id-type="pmid">28777724</pub-id></citation></ref>
<ref id="B44">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Friston</surname> <given-names>K. J.</given-names></name> <name><surname>Parr</surname> <given-names>T.</given-names></name> <name><surname>Yufik</surname> <given-names>Y.</given-names></name> <name><surname>Sajid</surname> <given-names>N.</given-names></name> <name><surname>Price</surname> <given-names>C. J.</given-names></name> <name><surname>Holmes</surname> <given-names>E.</given-names></name></person-group> (<year>2020</year>). <article-title>Generative models, linguistic communication and active inference</article-title>. <source>Neurosci. Biobehav. Rev</source>. <volume>118</volume>, <fpage>42</fpage>&#x02013;<lpage>64</lpage>. <pub-id pub-id-type="doi">10.1016/j.neubiorev.2020.07.005</pub-id><pub-id pub-id-type="pmid">32687883</pub-id></citation></ref>
<ref id="B45">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Friston</surname> <given-names>K. J.</given-names></name> <name><surname>Trujillo-Barreto</surname> <given-names>N.</given-names></name> <name><surname>Daunizeau</surname> <given-names>J.</given-names></name></person-group> (<year>2008</year>). <article-title>Dem: a variational treatment of dynamic systems</article-title>. <source>Neuroimage</source> <volume>41</volume>, <fpage>849</fpage>&#x02013;<lpage>885</lpage>. <pub-id pub-id-type="doi">10.1016/j.neuroimage.2008.02.054</pub-id><pub-id pub-id-type="pmid">18434205</pub-id></citation></ref>
<ref id="B46">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gazzaley</surname> <given-names>A.</given-names></name> <name><surname>Nobre</surname> <given-names>A. C.</given-names></name></person-group> (<year>2012</year>). <article-title>Top-down modulation: bridging selective attention and working memory</article-title>. <source>Trends Cogn. Sci</source>. <volume>16</volume>, <fpage>129</fpage>&#x02013;<lpage>135</lpage>. <pub-id pub-id-type="doi">10.1016/j.tics.2011.11.014</pub-id><pub-id pub-id-type="pmid">22209601</pub-id></citation></ref>
<ref id="B47">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Heeger</surname> <given-names>D. J.</given-names></name></person-group> (<year>1992</year>). <article-title>Normalization of cell responses in cat striate cortex</article-title>. <source>Vis. Neurosci</source>. <volume>9</volume>, <fpage>181</fpage>&#x02013;<lpage>197</lpage>. <pub-id pub-id-type="doi">10.1017/S0952523800009640</pub-id><pub-id pub-id-type="pmid">1504027</pub-id></citation></ref>
<ref id="B48">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Helfrich</surname> <given-names>R. F.</given-names></name> <name><surname>Fiebelkorn</surname> <given-names>I. C.</given-names></name> <name><surname>Szczepanski</surname> <given-names>S. M.</given-names></name> <name><surname>Lin</surname> <given-names>J. J.</given-names></name> <name><surname>Parvizi</surname> <given-names>J.</given-names></name> <name><surname>Knight</surname> <given-names>R. T.</given-names></name> <etal/></person-group>. (<year>2018</year>). <article-title>Neural mechanisms of sustained attention are rhythmic</article-title>. <source>Neuron</source> <volume>99</volume>, <fpage>854</fpage>&#x02013;<lpage>865</lpage>. <pub-id pub-id-type="doi">10.1016/j.neuron.2018.07.032</pub-id><pub-id pub-id-type="pmid">30138591</pub-id></citation></ref>
<ref id="B49">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Helmholtz</surname> <given-names>H. V.</given-names></name></person-group> (<year>1925</year>). <source>Treatise on Physiological Optics</source>. <publisher-loc>Rochester, NY</publisher-loc>: <publisher-name>Optical Society of America</publisher-name>.</citation>
</ref>
<ref id="B50">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hitz</surname> <given-names>G.</given-names></name> <name><surname>Galceran</surname> <given-names>E.</given-names></name> <name><surname>Garneau</surname> <given-names>M.-,&#x000C8;.</given-names></name> <name><surname>Pomerleau</surname> <given-names>F.</given-names></name> <name><surname>Siegwart</surname> <given-names>R.</given-names></name></person-group> (<year>2017</year>). <article-title>Adaptive continuous-space informative path planning for online environmental monitoring</article-title>. <source>J. Field Rob</source>. <volume>34</volume>, <fpage>1427</fpage>&#x02013;<lpage>1449</lpage>. <pub-id pub-id-type="doi">10.1002/rob.21722</pub-id></citation>
</ref>
<ref id="B51">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hogendoorn</surname> <given-names>H.</given-names></name></person-group> (<year>2016</year>). <article-title>Voluntary saccadic eye movements ride the attentional rhythm</article-title>. <source>J. Cogn. Neurosci</source>. <volume>28</volume>, <fpage>1625</fpage>&#x02013;<lpage>1635</lpage>. <pub-id pub-id-type="doi">10.1162/jocn_a_00986</pub-id><pub-id pub-id-type="pmid">27243615</pub-id></citation></ref>
<ref id="B52">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hsieh</surname> <given-names>L.-T.</given-names></name> <name><surname>Ranganath</surname> <given-names>C.</given-names></name></person-group> (<year>2014</year>). <article-title>Frontal midline theta oscillations during working memory maintenance and episodic encoding and retrieval</article-title>. <source>Neuroimage</source> <volume>85</volume>, <fpage>721</fpage>&#x02013;<lpage>729</lpage>. <pub-id pub-id-type="doi">10.1016/j.neuroimage.2013.08.003</pub-id><pub-id pub-id-type="pmid">23933041</pub-id></citation></ref>
<ref id="B53">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Itti</surname> <given-names>L.</given-names></name> <name><surname>Baldi</surname> <given-names>P.</given-names></name></person-group> (<year>2009</year>). <article-title>Bayesian surprise attracts human attention</article-title>. <source>Vis. Res</source>. <volume>49</volume>, <fpage>1295</fpage>&#x02013;<lpage>1306</lpage>. <pub-id pub-id-type="doi">10.1016/j.visres.2008.09.007</pub-id><pub-id pub-id-type="pmid">18834898</pub-id></citation></ref>
<ref id="B54">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Itti</surname> <given-names>L.</given-names></name> <name><surname>Koch</surname> <given-names>C.</given-names></name></person-group> (<year>2001</year>). <article-title>Computational modelling of visual attention</article-title>. <source>Nat. Rev. Neurosci</source>. <volume>2</volume>, <fpage>194</fpage>&#x02013;<lpage>203</lpage>. <pub-id pub-id-type="doi">10.1038/35058500</pub-id><pub-id pub-id-type="pmid">11256080</pub-id></citation></ref>
<ref id="B55">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kanai</surname> <given-names>R.</given-names></name> <name><surname>Komura</surname> <given-names>Y.</given-names></name> <name><surname>Shipp</surname> <given-names>S.</given-names></name> <name><surname>Friston</surname> <given-names>K.</given-names></name></person-group> (<year>2015</year>). <article-title>Cerebral hierarchies: predictive processing, precision and the pulvinar</article-title>. <source>Philos. Trans. R. Soc. B Biol. Sci</source>. <volume>370</volume>, <fpage>20140169</fpage>. <pub-id pub-id-type="doi">10.1098/rstb.2014.0169</pub-id><pub-id pub-id-type="pmid">25823866</pub-id></citation></ref>
<ref id="B56">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kanwisher</surname> <given-names>N.</given-names></name> <name><surname>Wojciulik</surname> <given-names>E.</given-names></name></person-group> (<year>2000</year>). <article-title>Visual attention: insights from brain imaging</article-title>. <source>Nat. Rev. Neurosci</source>. <volume>1</volume>, <fpage>91</fpage>&#x02013;<lpage>100</lpage>. <pub-id pub-id-type="doi">10.1038/35039043</pub-id><pub-id pub-id-type="pmid">11252779</pub-id></citation></ref>
<ref id="B57">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kaplan</surname> <given-names>F.</given-names></name> <name><surname>Hafner</surname> <given-names>V. V.</given-names></name></person-group> (<year>2006</year>). <article-title>The challenges of joint attention</article-title>. <source>Interact. Stud</source>. <volume>7</volume>, <fpage>135</fpage>&#x02013;<lpage>169</lpage>. <pub-id pub-id-type="doi">10.1075/is.7.2.04kap</pub-id></citation>
</ref>
<ref id="B58">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kim</surname> <given-names>A.</given-names></name> <name><surname>Eustice</surname> <given-names>R. M.</given-names></name></person-group> (<year>2013</year>). <article-title>Real-time visual slam for autonomous underwater hull inspection using visual saliency</article-title>. <source>IEEE Trans. Rob</source>. <volume>29</volume>, <fpage>719</fpage>&#x02013;<lpage>733</lpage>. <pub-id pub-id-type="doi">10.1109/TRO.2012.2235699</pub-id></citation>
</ref>
<ref id="B59">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Klein</surname> <given-names>R. M.</given-names></name></person-group> (<year>2000</year>). <article-title>Inhibition of return</article-title>. <source>Trends Cogn. Sci</source>. <volume>4</volume>, <fpage>138</fpage>&#x02013;<lpage>147</lpage>. <pub-id pub-id-type="doi">10.1016/S1364-6613(00)01452-2</pub-id><pub-id pub-id-type="pmid">10740278</pub-id></citation></ref>
<ref id="B60">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Klink</surname> <given-names>P. C.</given-names></name> <name><surname>Jentgens</surname> <given-names>P.</given-names></name> <name><surname>Lorteije</surname> <given-names>J. A.</given-names></name></person-group> (<year>2014</year>). <article-title>Priority maps explain the roles of value, attention, and salience in goal-oriented behavior</article-title>. <source>J. Neurosci</source>. <volume>34</volume>, <fpage>13867</fpage>&#x02013;<lpage>13869</lpage>. <pub-id pub-id-type="doi">10.1523/JNEUROSCI.3249-14.2014</pub-id><pub-id pub-id-type="pmid">25319682</pub-id></citation></ref>
<ref id="B61">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Knudsen</surname> <given-names>E. I.</given-names></name></person-group> (<year>2007</year>). <article-title>Fundamental components of attention</article-title>. <source>Annu. Rev. Neurosci</source>. <volume>30</volume>, <fpage>57</fpage>&#x02013;<lpage>78</lpage>. <pub-id pub-id-type="doi">10.1146/annurev.neuro.30.051606.094256</pub-id><pub-id pub-id-type="pmid">17417935</pub-id></citation></ref>
<ref id="B62">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>K&#x000F6;ster</surname> <given-names>M.</given-names></name> <name><surname>Finger</surname> <given-names>H.</given-names></name> <name><surname>Graetz</surname> <given-names>S.</given-names></name> <name><surname>Kater</surname> <given-names>M.</given-names></name> <name><surname>Gruber</surname> <given-names>T.</given-names></name></person-group> (<year>2018</year>). <article-title>Theta-gamma coupling binds visual perceptual features in an associative memory task</article-title>. <source>Sci. Rep</source>. <volume>8</volume>, <fpage>1</fpage>&#x02013;<lpage>9</lpage>. <pub-id pub-id-type="doi">10.1038/s41598-018-35812-7</pub-id><pub-id pub-id-type="pmid">30523336</pub-id></citation></ref>
<ref id="B63">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kragic</surname> <given-names>D.</given-names></name> <name><surname>Bj&#x000F6;rkman</surname> <given-names>M.</given-names></name> <name><surname>Christensen</surname> <given-names>H. I.</given-names></name> <name><surname>Eklundh</surname> <given-names>J.-O.</given-names></name></person-group> (<year>2005</year>). <article-title>Vision for robotic object manipulation in domestic settings</article-title>. <source>Rob. Auton. Syst</source>. <volume>52</volume>, <fpage>85</fpage>&#x02013;<lpage>100</lpage>. <pub-id pub-id-type="doi">10.1016/j.robot.2005.03.011</pub-id><pub-id pub-id-type="pmid">30959920</pub-id></citation></ref>
<ref id="B64">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Lanillos</surname> <given-names>P.</given-names></name></person-group> (<year>2013</year>). <source>Minimum time search of moving targets in uncertain environments</source> (<publisher-loc>Ph.D. thesis</publisher-loc>).</citation>
</ref>
<ref id="B65">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Lanillos</surname> <given-names>P.</given-names></name> <name><surname>Cheng</surname> <given-names>G.</given-names></name></person-group> (<year>2018b</year>). <article-title>Adaptive robot body learning and estimation through predictive coding,</article-title> in <source>2018 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)</source> (<publisher-loc>Madrid</publisher-loc>: <publisher-name>IEEE</publisher-name>), <fpage>4083</fpage>&#x02013;<lpage>4090</lpage>.</citation>
</ref>
<ref id="B66">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lanillos</surname> <given-names>P.</given-names></name> <name><surname>Dean-Leon</surname> <given-names>E.</given-names></name> <name><surname>Cheng</surname> <given-names>G.</given-names></name></person-group> (<year>2016</year>). <article-title>Yielding self-perception in robots through sensorimotor contingencies</article-title>. <source>IEEE Trans. Cogn. Dev. Syst</source>. <volume>9</volume>, <fpage>100</fpage>&#x02013;<lpage>112</lpage>. <pub-id pub-id-type="doi">10.1109/TCDS.2016.2627820</pub-id></citation>
</ref>
<ref id="B67">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Lanillos</surname> <given-names>P.</given-names></name> <name><surname>Ferreira</surname> <given-names>J. F.</given-names></name> <name><surname>Dias</surname> <given-names>J.</given-names></name></person-group> (<year>2015a</year>). <article-title>Designing an artificial attention system for social robots,</article-title> in <source>2015 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)</source> (<publisher-loc>Hamburg</publisher-loc>: <publisher-name>IEEE</publisher-name>), <fpage>4171</fpage>&#x02013;<lpage>4178</lpage>.</citation>
</ref>
<ref id="B68">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Lanillos</surname> <given-names>P.</given-names></name> <name><surname>Ferreira</surname> <given-names>J. F.</given-names></name> <name><surname>Dias</surname> <given-names>J.</given-names></name></person-group> (<year>2015b</year>). <article-title>Multisensory 3D saliency for artificial attention systems,</article-title> in <source>Proc. 3rd Workshop Recogn</source> (<publisher-loc>Action Scene Understanding</publisher-loc>), <fpage>1</fpage>&#x02013;<lpage>6</lpage>.</citation>
</ref>
<ref id="B69">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lanillos</surname> <given-names>P.</given-names></name> <name><surname>Gan</surname> <given-names>S. K.</given-names></name> <name><surname>Besada-Portas</surname> <given-names>E.</given-names></name> <name><surname>Pajares</surname> <given-names>G.</given-names></name> <name><surname>Sukkarieh</surname> <given-names>S.</given-names></name></person-group> (<year>2014</year>). <article-title>Multi-uav target search using decentralized gradient-based negotiation with expected observation</article-title>. <source>Inf. Sci</source>. <volume>282</volume>, <fpage>92</fpage>&#x02013;<lpage>110</lpage>. <pub-id pub-id-type="doi">10.1016/j.ins.2014.05.054</pub-id></citation>
</ref>
<ref id="B70">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lanillos</surname> <given-names>P.</given-names></name> <name><surname>Meo</surname> <given-names>C.</given-names></name> <name><surname>Pezzato</surname> <given-names>C.</given-names></name> <name><surname>Meera</surname> <given-names>A. A.</given-names></name> <name><surname>Baioumy</surname> <given-names>M.</given-names></name> <name><surname>Ohata</surname> <given-names>W.</given-names></name> <etal/></person-group>. (<year>2021</year>). <article-title>Active inference in robotics and artificial agents: Survey and challenges</article-title>. <source>arXiv[preprint].arXiv:2112.01871</source>. <pub-id pub-id-type="doi">10.48550/arXiv.2112.01871</pub-id></citation>
</ref>
<ref id="B71">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Lanillos</surname> <given-names>P.</given-names></name> <name><surname>Cheng</surname> <given-names>G.</given-names></name></person-group> (<year>2018a</year>). <article-title>Active attention applications in robotics,</article-title> in <source>International Workshop on Active Vision, Attention, and Learning</source> (<publisher-loc>Tokyo</publisher-loc>: <publisher-name>IEEE Developmental Learning and Epigenetic Robotics (ICDL-Epirob)</publisher-name>).</citation>
</ref>
<ref id="B72">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>LaValle</surname> <given-names>S. M.</given-names></name></person-group> (<year>2006</year>). <source>Planning Algorithms</source>. <publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>.</citation>
</ref>
<ref id="B73">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lengyel</surname> <given-names>M.</given-names></name> <name><surname>Yang</surname> <given-names>S. C.-H.</given-names></name> <name><surname>Wolpert</surname> <given-names>D. M.</given-names></name></person-group> (<year>2016</year>). <article-title>Active sensing in the categorization of visual patterns</article-title>. <source>eLife</source> <volume>5</volume>, <fpage>e12215</fpage>. <pub-id pub-id-type="doi">10.7554/eLife.12215</pub-id><pub-id pub-id-type="pmid">28165321</pub-id></citation></ref>
<ref id="B74">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lindley</surname> <given-names>D. V.</given-names></name></person-group> (<year>1956</year>). <article-title>On a measure of the information provided by an experiment</article-title>. <source>Ann. Math. Stat</source>. <volume>27</volume>, <fpage>986</fpage>&#x02013;<lpage>1005</lpage>. <pub-id pub-id-type="doi">10.1214/aoms/1177728069</pub-id></citation>
</ref>
<ref id="B75">
<citation citation-type="web"><person-group person-group-type="author"><name><surname>Louie</surname> <given-names>K.</given-names></name> <name><surname>Glimcher</surname> <given-names>P. W.</given-names></name></person-group> (<year>2019</year>). <article-title>Normalization principles in computational neuroscience,</article-title> in <source>Oxford Research Encyclopedia of Neuroscience</source>. Available online at: <ext-link ext-link-type="uri" xlink:href="https://oxfordre.com/neuroscience/view/10.1093/acrefore/9780190264086.001.0001/acrefore-9780190264086-e-43">https://oxfordre.com/neuroscience/view/10.1093/acrefore/9780190264086.001.0001/acrefore-9780190264086-e-43</ext-link> (accessed July 7, 2022).</citation>
</ref>
<ref id="B76">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Marchant</surname> <given-names>R.</given-names></name> <name><surname>Ramos</surname> <given-names>F.</given-names></name></person-group> (<year>2014</year>). <article-title>Bayesian optimisation for informative continuous path planning,</article-title> in <source>2014 IEEE International Conference on Robotics and Automation (ICRA)</source> (<publisher-loc>Hong Kong</publisher-loc>: <publisher-name>IEEE</publisher-name>), <fpage>6136</fpage>&#x02013;<lpage>6143</lpage>.</citation>
</ref>
<ref id="B77">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Meera</surname> <given-names>A. A.</given-names></name> <name><surname>Popovi&#x00107;</surname> <given-names>M.</given-names></name> <name><surname>Millane</surname> <given-names>A.</given-names></name> <name><surname>Siegwart</surname> <given-names>R.</given-names></name></person-group> (<year>2019</year>). <article-title>Obstacle-aware adaptive informative path planning for uav-based target search,</article-title> in <source>2019 International Conference on Robotics and Automation (ICRA)</source> (<publisher-loc>Montreal, QC</publisher-loc>: <publisher-name>IEEE</publisher-name>), <fpage>718</fpage>&#x02013;<lpage>724</lpage>.</citation>
</ref>
<ref id="B78">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Meera</surname> <given-names>A. A.</given-names></name> <name><surname>Wisse</surname> <given-names>M.</given-names></name></person-group> (<year>2020</year>). <article-title>Free energy principle based state and input observer design for linear systems with colored noise,</article-title> in <source>2020 American Control Conference (ACC)</source> (<publisher-loc>Denver, CO</publisher-loc>: <publisher-name>IEEE</publisher-name>), <fpage>5052</fpage>&#x02013;<lpage>5058</lpage>.</citation>
</ref>
<ref id="B79">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Meera</surname> <given-names>A. A.</given-names></name> <name><surname>Wisse</surname> <given-names>M.</given-names></name></person-group> (<year>2021a</year>). <article-title>A brain inspired learning algorithm for the perception of a quadrotor in wind</article-title>. <source>arXiv[preprint].arXiv:2109.11971</source>. <pub-id pub-id-type="doi">10.48550/arXiv.2109.11971</pub-id></citation>
</ref>
<ref id="B80">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Meera</surname> <given-names>A. A.</given-names></name> <name><surname>Wisse</surname> <given-names>M.</given-names></name></person-group> (<year>2021b</year>). <article-title>On the convergence of dem&#x00027;s linear parameter estimator,</article-title> in <source>Machine Learning and Principles and Practice of Knowledge Discovery in Databases</source> (<publisher-loc>Cham</publisher-loc>: <publisher-name>Springer International Publishing</publisher-name>), <fpage>692</fpage>&#x02013;<lpage>700</lpage>.</citation>
</ref>
<ref id="B81">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Meera</surname> <given-names>A. A.</given-names></name> <name><surname>Wisse</surname> <given-names>M.</given-names></name></person-group> (<year>2022</year>). <article-title>Free energy principle for the noise smoothness estimation of linear systems with colored noise</article-title>. <source>arXiv[preprint].arXiv:2204.01796</source>. <pub-id pub-id-type="doi">10.48550/arXiv.2204.01796</pub-id></citation>
</ref>
<ref id="B82">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mirza</surname> <given-names>M. B.</given-names></name> <name><surname>Adams</surname> <given-names>R. A.</given-names></name> <name><surname>Friston</surname> <given-names>K.</given-names></name> <name><surname>Parr</surname> <given-names>T.</given-names></name></person-group> (<year>2019</year>). <article-title>Introducing a bayesian model of selective attention based on active inference</article-title>. <source>Sci. Rep</source>. <volume>9</volume>, <fpage>1</fpage>&#x02013;<lpage>22</lpage>. <pub-id pub-id-type="doi">10.1038/s41598-019-50138-8</pub-id><pub-id pub-id-type="pmid">31558746</pub-id></citation></ref>
<ref id="B83">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mirza</surname> <given-names>M. B.</given-names></name> <name><surname>Adams</surname> <given-names>R. A.</given-names></name> <name><surname>Mathys</surname> <given-names>C. D.</given-names></name> <name><surname>Friston</surname> <given-names>K. J.</given-names></name></person-group> (<year>2016</year>). <article-title>Scene construction, visual foraging, and active inference</article-title>. <source>Front. Comput. Neurosci</source>. <volume>10</volume>, <fpage>56</fpage>. <pub-id pub-id-type="doi">10.3389/fncom.2016.00056</pub-id><pub-id pub-id-type="pmid">27378899</pub-id></citation></ref>
<ref id="B84">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mor&#x000E9;n</surname> <given-names>J.</given-names></name> <name><surname>Ude</surname> <given-names>A.</given-names></name> <name><surname>Koene</surname> <given-names>A.</given-names></name> <name><surname>Cheng</surname> <given-names>G.</given-names></name></person-group> (<year>2008</year>). <article-title>Biologically based top-down attention modulation for humanoid interactions</article-title>. <source>Int. J. Humanoid Rob</source>. <volume>5</volume>, <fpage>3</fpage>&#x02013;<lpage>24</lpage>. <pub-id pub-id-type="doi">10.1142/S0219843608001285</pub-id></citation>
</ref>
<ref id="B85">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mousavi</surname> <given-names>S.</given-names></name> <name><surname>Schukat</surname> <given-names>M.</given-names></name> <name><surname>Howley</surname> <given-names>E.</given-names></name> <name><surname>Borji</surname> <given-names>A.</given-names></name> <name><surname>Mozayani</surname> <given-names>N.</given-names></name></person-group> (<year>2016</year>). <article-title>Learning to predict where to look in interactive environments using deep recurrent q-learning</article-title>. <source>arXiv[preprint].arXiv:1612.05753</source>. <pub-id pub-id-type="doi">10.48550/arXiv.1612.05753</pub-id></citation>
</ref>
<ref id="B86">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Nagai</surname> <given-names>Y.</given-names></name> <name><surname>Hosoda</surname> <given-names>K.</given-names></name> <name><surname>Morita</surname> <given-names>A.</given-names></name> <name><surname>Asada</surname> <given-names>M.</given-names></name></person-group> (<year>2003</year>). <article-title>A constructive model for the development of joint attention</article-title>. <source>Conn. Sci</source>. <volume>15</volume>, <fpage>211</fpage>&#x02013;<lpage>229</lpage>. <pub-id pub-id-type="doi">10.1080/09540090310001655101</pub-id></citation>
</ref>
<ref id="B87">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Nakajima</surname> <given-names>M.</given-names></name> <name><surname>Schmitt</surname> <given-names>L. I.</given-names></name> <name><surname>Halassa</surname> <given-names>M. M.</given-names></name></person-group> (<year>2019</year>). <article-title>Prefrontal cortex regulates sensory filtering through a basal ganglia-to-thalamus pathway</article-title>. <source>Neuron</source> <volume>103</volume>, <fpage>445</fpage>&#x02013;<lpage>458</lpage>. <pub-id pub-id-type="doi">10.1016/j.neuron.2019.05.026</pub-id><pub-id pub-id-type="pmid">31202541</pub-id></citation></ref>
<ref id="B88">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Nakayama</surname> <given-names>R.</given-names></name> <name><surname>Motoyoshi</surname> <given-names>I.</given-names></name></person-group> (<year>2019</year>). <article-title>Attention periodically binds visual features as single events depending on neural oscillations phase-locked to action</article-title>. <source>J.Neurosci</source>. <volume>39</volume>, <fpage>4153</fpage>&#x02013;<lpage>4161</lpage>. <pub-id pub-id-type="doi">10.1523/JNEUROSCI.2494-18.2019</pub-id><pub-id pub-id-type="pmid">30886011</pub-id></citation></ref>
<ref id="B89">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Oberauer</surname> <given-names>K.</given-names></name></person-group> (<year>2019</year>). <article-title>Working memory and attention-a conceptual analysis and review</article-title>. <source>J. Cogn</source>. <volume>2</volume>, <fpage>58</fpage>. <pub-id pub-id-type="doi">10.5334/joc.58</pub-id><pub-id pub-id-type="pmid">31517246</pub-id></citation></ref>
<ref id="B90">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Oliver</surname> <given-names>G.</given-names></name> <name><surname>Lanillos</surname> <given-names>P.</given-names></name> <name><surname>Cheng</surname> <given-names>G.</given-names></name></person-group> (<year>2021</year>). <article-title>An empirical study of active inference on a humanoid robot</article-title>. <source>IEEE Trans. Cogn. Dev. Syst</source>. <volume>14</volume>, <fpage>462</fpage>&#x02013;<lpage>471</lpage>. <pub-id pub-id-type="doi">10.1109/TCDS.2021.3049907</pub-id></citation>
</ref>
<ref id="B91">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Orabona</surname> <given-names>F.</given-names></name> <name><surname>Metta</surname> <given-names>G.</given-names></name> <name><surname>Sandini</surname> <given-names>G.</given-names></name></person-group> (<year>2005</year>). <article-title>Object-based visual attention: a model for a behaving robot,</article-title> in <source>2005 IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR&#x00027;05)-Workshops</source> (<publisher-loc>San Diego, CA</publisher-loc>: <publisher-name>IEEE</publisher-name>), <fpage>89</fpage>&#x02013;<lpage>89</lpage>.</citation>
</ref>
<ref id="B92">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Oudeyer</surname> <given-names>P.-Y.</given-names></name> <name><surname>Kaplan</surname> <given-names>F.</given-names></name></person-group> (<year>2009</year>). <article-title>What is intrinsic motivation? a typology of computational approaches</article-title>. <source>Front. Neurorob</source>. <volume>1</volume>, <fpage>6</fpage>. <pub-id pub-id-type="doi">10.3389/neuro.12.006.2007</pub-id><pub-id pub-id-type="pmid">18958277</pub-id></citation></ref>
<ref id="B93">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Panichello</surname> <given-names>M. F.</given-names></name> <name><surname>Buschman</surname> <given-names>T. J.</given-names></name></person-group> (<year>2021</year>). <article-title>Shared mechanisms underlie the control of working memory and attention</article-title>. <source>Nature</source> <volume>592</volume>, <fpage>601</fpage>&#x02013;<lpage>605</lpage>. <pub-id pub-id-type="doi">10.1038/s41586-021-03390-w</pub-id><pub-id pub-id-type="pmid">33790467</pub-id></citation></ref>
<ref id="B94">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Parr</surname> <given-names>T.</given-names></name> <name><surname>Benrimoh</surname> <given-names>D. A.</given-names></name> <name><surname>Vincent</surname> <given-names>P.</given-names></name> <name><surname>Friston</surname> <given-names>K. J.</given-names></name></person-group> (<year>2018</year>). <article-title>Precision and false perceptual inference</article-title>. <source>Front. Integr. Neurosci</source>. <volume>12</volume>, <fpage>39</fpage>. <pub-id pub-id-type="doi">10.3389/fnint.2018.00039</pub-id><pub-id pub-id-type="pmid">30294264</pub-id></citation></ref>
<ref id="B95">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Parr</surname> <given-names>T.</given-names></name> <name><surname>Corcoran</surname> <given-names>A. W.</given-names></name> <name><surname>Friston</surname> <given-names>K. J.</given-names></name> <name><surname>Hohwy</surname> <given-names>J.</given-names></name></person-group> (<year>2019</year>). <article-title>Perceptual awareness and active inference</article-title>. <source>Neurosci. Consciousness</source> <volume>2019</volume>, <fpage>niz012</fpage>. <pub-id pub-id-type="doi">10.1093/nc/niz012</pub-id><pub-id pub-id-type="pmid">31528360</pub-id></citation></ref>
<ref id="B96">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Parr</surname> <given-names>T.</given-names></name> <name><surname>Friston</surname> <given-names>K. J.</given-names></name></person-group> (<year>2017a</year>). <article-title>Uncertainty, epistemics and active inference</article-title>. <source>J. R. Soc. Interface</source> <volume>14</volume>, <fpage>20170376</fpage>. <pub-id pub-id-type="doi">10.1098/rsif.2017.0376</pub-id><pub-id pub-id-type="pmid">29167370</pub-id></citation></ref>
<ref id="B97">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Parr</surname> <given-names>T.</given-names></name> <name><surname>Friston</surname> <given-names>K. J.</given-names></name></person-group> (<year>2017b</year>). <article-title>Working memory, attention, and salience in active inference</article-title>. <source>Sci. Rep</source>. <volume>7</volume>, <fpage>1</fpage>&#x02013;<lpage>21</lpage>. <pub-id pub-id-type="doi">10.1038/s41598-017-15249-0</pub-id><pub-id pub-id-type="pmid">29116142</pub-id></citation></ref>
<ref id="B98">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Parr</surname> <given-names>T.</given-names></name> <name><surname>Friston</surname> <given-names>K. J.</given-names></name></person-group> (<year>2019</year>). <article-title>Attention or salience?</article-title> <source>Curr. Opin. Psychol</source>. <volume>29</volume>, <fpage>1</fpage>&#x02013;<lpage>5</lpage>. <pub-id pub-id-type="doi">10.1016/j.copsyc.2018.10.006</pub-id><pub-id pub-id-type="pmid">30359960</pub-id></citation></ref>
<ref id="B99">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Parr</surname> <given-names>T.</given-names></name> <name><surname>Pezzulo</surname> <given-names>G.</given-names></name></person-group> (<year>2021</year>). <article-title>Understanding, explanation, and active inference</article-title>. <source>Front. Syst. Neurosci</source>. <volume>15</volume>, <fpage>772641</fpage>. <pub-id pub-id-type="doi">10.3389/fnsys.2021.772641</pub-id><pub-id pub-id-type="pmid">34803619</pub-id></citation></ref>
<ref id="B100">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Parr</surname> <given-names>T.</given-names></name> <name><surname>Sajid</surname> <given-names>N.</given-names></name> <name><surname>Da Costa</surname> <given-names>L.</given-names></name> <name><surname>Mirza</surname> <given-names>M. B.</given-names></name> <name><surname>Friston</surname> <given-names>K. J.</given-names></name></person-group> (<year>2021</year>). <article-title>Generative models for active vision</article-title>. <source>Front. Neurorobot</source>. <volume>15</volume>, <fpage>651432</fpage>. <pub-id pub-id-type="doi">10.3389/fnbot.2021.651432</pub-id><pub-id pub-id-type="pmid">33927605</pub-id></citation></ref>
<ref id="B101">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Peters</surname> <given-names>B.</given-names></name> <name><surname>Kaiser</surname> <given-names>J.</given-names></name> <name><surname>Rahm</surname> <given-names>B.</given-names></name> <name><surname>Bledowski</surname> <given-names>C.</given-names></name></person-group> (<year>2020</year>). <article-title>Object-based attention prioritizes working memory contents at a theta rhythm</article-title>. <source>J. Exp. Psychol. Gen</source>. <volume>150</volume>, <fpage>1250</fpage>&#x02013;<lpage>1256</lpage>. <pub-id pub-id-type="doi">10.1037/xge0000994</pub-id><pub-id pub-id-type="pmid">33211526</pub-id></citation></ref>
<ref id="B102">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Phillips</surname> <given-names>J. M.</given-names></name> <name><surname>Kambi</surname> <given-names>N. A.</given-names></name> <name><surname>Saalmann</surname> <given-names>Y. B.</given-names></name></person-group> (<year>2016</year>). <article-title>A subcortical pathway for rapid, goal-driven, attentional filtering</article-title>. <source>Trends Neurosci</source>. <volume>39</volume>, <fpage>49</fpage>&#x02013;<lpage>51</lpage>. <pub-id pub-id-type="doi">10.1016/j.tins.2015.12.003</pub-id><pub-id pub-id-type="pmid">26743499</pub-id></citation></ref>
<ref id="B103">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Pomper</surname> <given-names>U.</given-names></name> <name><surname>Ansorge</surname> <given-names>U.</given-names></name></person-group> (<year>2021</year>). <article-title>Theta-rhythmic oscillation of working memory performance</article-title>. <source>Psychol. Sci</source>. <volume>32</volume>, <fpage>1801</fpage>&#x02013;<lpage>1810</lpage>. <pub-id pub-id-type="doi">10.1177/09567976211013045</pub-id><pub-id pub-id-type="pmid">34592108</pub-id></citation></ref>
<ref id="B104">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Popovi&#x00107;</surname> <given-names>M.</given-names></name> <name><surname>Vidal-Calleja</surname> <given-names>T.</given-names></name> <name><surname>Hitz</surname> <given-names>G.</given-names></name> <name><surname>Sa</surname> <given-names>I.</given-names></name> <name><surname>Siegwart</surname> <given-names>R.</given-names></name> <name><surname>Nieto</surname> <given-names>J.</given-names></name></person-group> (<year>2017</year>). <article-title>Multiresolution mapping and informative path planning for uav-based terrain monitoring,</article-title> in <source>2017 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)</source> (<publisher-loc>Vancouver, BC</publisher-loc>: <publisher-name>IEEE</publisher-name>), <fpage>1382</fpage>&#x02013;<lpage>1388</lpage>.</citation>
</ref>
<ref id="B105">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Rao</surname> <given-names>R. P.</given-names></name></person-group> (<year>2005</year>). <article-title>Bayesian inference and attentional modulation in the visual cortex</article-title>. <source>Neuroreport</source> <volume>16</volume>, <fpage>1843</fpage>&#x02013;<lpage>1848</lpage>. <pub-id pub-id-type="doi">10.1097/01.wnr.0000183900.92901.fc</pub-id><pub-id pub-id-type="pmid">16237339</pub-id></citation></ref>
<ref id="B106">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Rasouli</surname> <given-names>A.</given-names></name> <name><surname>Lanillos</surname> <given-names>P.</given-names></name> <name><surname>Cheng</surname> <given-names>G.</given-names></name> <name><surname>Tsotsos</surname> <given-names>J. K.</given-names></name></person-group> (<year>2020</year>). <article-title>Attention-based active visual search for mobile robots</article-title>. <source>Auton. Robots</source> <volume>44</volume>, <fpage>131</fpage>&#x02013;<lpage>146</lpage>. <pub-id pub-id-type="doi">10.1007/s10514-019-09882-z</pub-id></citation>
</ref>
<ref id="B107">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Raybourn</surname> <given-names>M. S.</given-names></name> <name><surname>Keller</surname> <given-names>E. L.</given-names></name></person-group> (<year>1977</year>). <article-title>Colliculoreticular organization in primate oculomotor system</article-title>. <source>J. Neurophysiol</source>. <volume>40</volume>, <fpage>861</fpage>&#x02013;<lpage>878</lpage>. <pub-id pub-id-type="doi">10.1152/jn.1977.40.4.861</pub-id><pub-id pub-id-type="pmid">407334</pub-id></citation></ref>
<ref id="B108">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Reynolds</surname> <given-names>J. H.</given-names></name> <name><surname>Heeger</surname> <given-names>D. J.</given-names></name></person-group> (<year>2009</year>). <article-title>The normalization model of attention</article-title>. <source>Neuron</source> <volume>61</volume>, <fpage>168</fpage>&#x02013;<lpage>185</lpage>. <pub-id pub-id-type="doi">10.1016/j.neuron.2009.01.002</pub-id><pub-id pub-id-type="pmid">19186161</pub-id></citation></ref>
<ref id="B109">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Reynolds</surname> <given-names>J. H.</given-names></name> <name><surname>Pasternak</surname> <given-names>T.</given-names></name> <name><surname>Desimone</surname> <given-names>R.</given-names></name></person-group> (<year>2000</year>). <article-title>Attention increases sensitivity of v4 neurons</article-title>. <source>Neuron</source> <volume>26</volume>, <fpage>703</fpage>&#x02013;<lpage>714</lpage>. <pub-id pub-id-type="doi">10.1016/S0896-6273(00)81206-4</pub-id><pub-id pub-id-type="pmid">10896165</pub-id></citation></ref>
<ref id="B110">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Rizzolatti</surname> <given-names>G.</given-names></name> <name><surname>Riggio</surname> <given-names>L.</given-names></name> <name><surname>Dascola</surname> <given-names>I.</given-names></name> <name><surname>Umilt&#x000E1;</surname> <given-names>C.</given-names></name></person-group> (<year>1987</year>). <article-title>Reorienting attention across the horizontal and vertical meridians: evidence in favor of a premotor theory of attention</article-title>. <source>Neuropsychologia</source> <volume>25</volume>, <fpage>31</fpage>&#x02013;<lpage>40</lpage>. <pub-id pub-id-type="doi">10.1016/0028-3932(87)90041-8</pub-id><pub-id pub-id-type="pmid">3574648</pub-id></citation></ref>
<ref id="B111">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Roberts</surname> <given-names>R.</given-names></name> <name><surname>Ta</surname> <given-names>D.-N.</given-names></name> <name><surname>Straub</surname> <given-names>J.</given-names></name> <name><surname>Ok</surname> <given-names>K.</given-names></name> <name><surname>Dellaert</surname> <given-names>F.</given-names></name></person-group> (<year>2012</year>). <article-title>Saliency detection and model-based tracking: a two part vision system for small robot navigation in forested environment,</article-title> in <source>Unmanned Systems Technology XIV, Vol. 8387</source> (<publisher-loc>International Society for Optics and Photonics</publisher-loc>), <fpage>83870S</fpage>.</citation>
</ref>
<ref id="B112">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Rucci</surname> <given-names>M.</given-names></name> <name><surname>Ahissar</surname> <given-names>E.</given-names></name> <name><surname>Burr</surname> <given-names>D.</given-names></name></person-group> (<year>2018</year>). <article-title>Temporal coding of visual space</article-title>. <source>Trends Cogn. Sci</source>. <volume>22</volume>, <fpage>883</fpage>&#x02013;<lpage>895</lpage>. <pub-id pub-id-type="doi">10.1016/j.tics.2018.07.009</pub-id><pub-id pub-id-type="pmid">30266148</pub-id></citation></ref>
<ref id="B113">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ruff</surname> <given-names>D. A.</given-names></name> <name><surname>Cohen</surname> <given-names>M. R.</given-names></name></person-group> (<year>2016</year>). <article-title>Stimulus dependence of correlated variability across cortical areas</article-title>. <source>J. Neurosci</source>. <volume>36</volume>, <fpage>7546</fpage>&#x02013;<lpage>7556</lpage>. <pub-id pub-id-type="doi">10.1523/JNEUROSCI.0504-16.2016</pub-id><pub-id pub-id-type="pmid">27413163</pub-id></citation></ref>
<ref id="B114">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sajid</surname> <given-names>N.</given-names></name> <name><surname>Da Costa</surname> <given-names>L.</given-names></name> <name><surname>Parr</surname> <given-names>T.</given-names></name> <name><surname>Friston</surname> <given-names>K.</given-names></name></person-group> (<year>2021a</year>). <article-title>Active inference, bayesian optimal design, and expected utility</article-title>. <source>arXiv[preprint].arXiv:2110.04074</source>. <pub-id pub-id-type="doi">10.1017/9781009026949.007</pub-id></citation>
</ref>
<ref id="B115">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sajid</surname> <given-names>N.</given-names></name> <name><surname>Faccio</surname> <given-names>F.</given-names></name> <name><surname>Da Costa</surname> <given-names>L.</given-names></name> <name><surname>Parr</surname> <given-names>T.</given-names></name> <name><surname>Schmidhuber</surname> <given-names>J.</given-names></name> <name><surname>Friston</surname> <given-names>K.</given-names></name></person-group> (<year>2021b</year>). <article-title>Bayesian brains and the r\&#x00027;enyi divergence</article-title>. <source>arXiv[preprint].arXiv:2107.05438</source>. <pub-id pub-id-type="doi">10.48550/arXiv.2107.05438</pub-id></citation>
</ref>
<ref id="B116">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sajid</surname> <given-names>N.</given-names></name> <name><surname>Friston</surname> <given-names>K. J.</given-names></name> <name><surname>Ekert</surname> <given-names>J. O.</given-names></name> <name><surname>Price</surname> <given-names>C. J.</given-names></name> <name><surname>Green</surname> <given-names>D. W.</given-names></name></person-group> (<year>2020</year>). <article-title>Neuromodulatory control and language recovery in bilingual aphasia: An active inference approach</article-title>. <source>Behav. Sci</source>. <volume>10</volume>, <fpage>161</fpage>. <pub-id pub-id-type="doi">10.3390/bs10100161</pub-id><pub-id pub-id-type="pmid">33096824</pub-id></citation></ref>
<ref id="B117">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sajid</surname> <given-names>N.</given-names></name> <name><surname>Holmes</surname> <given-names>E.</given-names></name> <name><surname>Costa</surname> <given-names>L. D.</given-names></name> <name><surname>Price</surname> <given-names>C.</given-names></name> <name><surname>Friston</surname> <given-names>K.</given-names></name></person-group> (<year>2022</year>). <article-title>A mixed generative model of auditory word repetition</article-title>. <source>bioRxiv [preprint]</source>. <pub-id pub-id-type="doi">10.1101/2022.01.20.477138</pub-id></citation>
</ref>
<ref id="B118">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Santangelo</surname> <given-names>V.</given-names></name></person-group> (<year>2015</year>). <article-title>Forced to remember: when memory is biased by salient information</article-title>. <source>Behav. Brain Res</source>. <volume>283</volume>, <fpage>1</fpage>&#x02013;<lpage>10</lpage>. <pub-id pub-id-type="doi">10.1016/j.bbr.2015.01.013</pub-id><pub-id pub-id-type="pmid">25595422</pub-id></citation></ref>
<ref id="B119">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Santangelo</surname> <given-names>V.</given-names></name> <name><surname>Di Francesco</surname> <given-names>S. A.</given-names></name> <name><surname>Mastroberardino</surname> <given-names>S.</given-names></name> <name><surname>Macaluso</surname> <given-names>E.</given-names></name></person-group> (<year>2015</year>). <article-title>Parietal cortex integrates contextual and saliency signals during the encoding of natural scenes in working memory</article-title>. <source>Hum. Brain Mapp</source>. <volume>36</volume>, <fpage>5003</fpage>&#x02013;<lpage>5017</lpage>. <pub-id pub-id-type="doi">10.1002/hbm.22984</pub-id><pub-id pub-id-type="pmid">26333392</pub-id></citation></ref>
<ref id="B120">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Santangelo</surname> <given-names>V.</given-names></name> <name><surname>Macaluso</surname> <given-names>E.</given-names></name></person-group> (<year>2013</year>). <article-title>Visual salience improves spatial working memory via enhanced parieto-temporal functional connectivity</article-title>. <source>J. Neurosci</source>. <volume>33</volume>, <fpage>4110</fpage>&#x02013;<lpage>4117</lpage>. <pub-id pub-id-type="doi">10.1523/JNEUROSCI.4138-12.2013</pub-id><pub-id pub-id-type="pmid">23447619</pub-id></citation></ref>
<ref id="B121">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Schmitz</surname> <given-names>T. W.</given-names></name> <name><surname>Duncan</surname> <given-names>J.</given-names></name></person-group> (<year>2018</year>). <article-title>Normalization and the cholinergic microcircuit: a unified basis for attention</article-title>. <source>Trends Cogn. Sci</source>. <volume>22</volume>, <fpage>422</fpage>&#x02013;<lpage>437</lpage>. <pub-id pub-id-type="doi">10.1016/j.tics.2018.02.011</pub-id><pub-id pub-id-type="pmid">29576464</pub-id></citation></ref>
<ref id="B122">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Shon</surname> <given-names>A. P.</given-names></name> <name><surname>Grimes</surname> <given-names>D. B.</given-names></name> <name><surname>Baker</surname> <given-names>C. L.</given-names></name> <name><surname>Hoffman</surname> <given-names>M. W.</given-names></name> <name><surname>Zhou</surname> <given-names>S.</given-names></name> <name><surname>Rao</surname> <given-names>R. P.</given-names></name></person-group> (<year>2005</year>). <article-title>Probabilistic gaze imitation and saliency learning in a robotic head,</article-title> in <source>Proceedings of the 2005 IEEE International Conference on Robotics and Automation</source> (<publisher-loc>Barcelona</publisher-loc>: <publisher-name>IEEE</publisher-name>), <fpage>2865</fpage>&#x02013;<lpage>2870</lpage>.</citation>
</ref>
<ref id="B123">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sommer</surname> <given-names>M. A.</given-names></name> <name><surname>Wurtz</surname> <given-names>R. H.</given-names></name></person-group> (<year>2006</year>). <article-title>Influence of the thalamus on spatial visual processing in frontal cortex</article-title>. <source>Nature</source> <volume>444</volume>, <fpage>374</fpage>&#x02013;<lpage>377</lpage>. <pub-id pub-id-type="doi">10.1038/nature05279</pub-id><pub-id pub-id-type="pmid">17093408</pub-id></citation></ref>
<ref id="B124">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Spratling</surname> <given-names>M. W.</given-names></name></person-group> (<year>2008</year>). <article-title>Predictive coding as a model of biased competition in visual attention</article-title>. <source>Vis. Res</source>. <volume>48</volume>, <fpage>1391</fpage>&#x02013;<lpage>1408</lpage>. <pub-id pub-id-type="doi">10.1016/j.visres.2008.03.009</pub-id><pub-id pub-id-type="pmid">18442841</pub-id></citation></ref>
<ref id="B125">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tomassini</surname> <given-names>A.</given-names></name> <name><surname>Ambrogioni</surname> <given-names>L.</given-names></name> <name><surname>Medendorp</surname> <given-names>W. P.</given-names></name> <name><surname>Maris</surname> <given-names>E.</given-names></name></person-group> (<year>2017</year>). <article-title>Theta oscillations locked to intended actions rhythmically modulate perception</article-title>. <source>Elife</source> <volume>6</volume>, <fpage>e25618</fpage>. <pub-id pub-id-type="doi">10.7554/eLife.25618</pub-id><pub-id pub-id-type="pmid">28686161</pub-id></citation></ref>
<ref id="B126">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Treisman</surname> <given-names>A. M.</given-names></name> <name><surname>Gelade</surname> <given-names>G.</given-names></name></person-group> (<year>1980</year>). <article-title>A feature-integration theory of attention</article-title>. <source>Cogn. Psychol</source>. <volume>12</volume>, <fpage>97</fpage>&#x02013;<lpage>136</lpage>. <pub-id pub-id-type="doi">10.1016/0010-0285(80)90005-5</pub-id><pub-id pub-id-type="pmid">7351125</pub-id></citation></ref>
<ref id="B127">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tsotsos</surname> <given-names>J. K.</given-names></name> <name><surname>Culhane</surname> <given-names>S. M.</given-names></name> <name><surname>Wai</surname> <given-names>W. Y. K.</given-names></name> <name><surname>Lai</surname> <given-names>Y.</given-names></name> <name><surname>Davis</surname> <given-names>N.</given-names></name> <name><surname>Nuflo</surname> <given-names>F.</given-names></name></person-group> (<year>1995</year>). <article-title>Modeling visual attention via selective tuning</article-title>. <source>Artif. Intell</source>. <volume>78</volume>, <fpage>507</fpage>&#x02013;<lpage>545</lpage>. <pub-id pub-id-type="doi">10.1016/0004-3702(95)00025-9</pub-id></citation>
</ref>
<ref id="B128">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Ude</surname> <given-names>A.</given-names></name> <name><surname>Wyart</surname> <given-names>V.</given-names></name> <name><surname>Lin</surname> <given-names>L.-H.</given-names></name> <name><surname>Cheng</surname> <given-names>G.</given-names></name></person-group> (<year>2005</year>). <article-title>Distributed visual attention on a humanoid robot,</article-title> in <source>5th IEEE-RAS International Conference on Humanoid Robots, 2005</source> (<publisher-loc>Tsukuba</publisher-loc>: <publisher-name>IEEE</publisher-name>), <fpage>381</fpage>&#x02013;<lpage>386</lpage>.</citation>
</ref>
<ref id="B129">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>VanRullen</surname> <given-names>R.</given-names></name></person-group> (<year>2016</year>). <article-title>Perceptual cycles</article-title>. <source>Trends Cogn. Sci.</source> <volume>20</volume>, <fpage>723</fpage>&#x02013;<lpage>735</lpage>. <pub-id pub-id-type="doi">10.1016/j.tics.2016.07.006</pub-id><pub-id pub-id-type="pmid">27567317</pub-id></citation></ref>
<ref id="B130">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Welch</surname> <given-names>G.</given-names></name> <name><surname>Bishop</surname> <given-names>G.</given-names></name></person-group> (<year>2002</year>). <source>An Introduction to the Kalman Filter</source>. <publisher-loc>Chapel Hill, NC</publisher-loc>: <publisher-name>University of North Carolina</publisher-name>.</citation>
</ref>
<ref id="B131">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>White</surname> <given-names>B. J.</given-names></name> <name><surname>Berg</surname> <given-names>D. J.</given-names></name> <name><surname>Kan</surname> <given-names>J. Y.</given-names></name> <name><surname>Marino</surname> <given-names>R. A.</given-names></name> <name><surname>Itti</surname> <given-names>L.</given-names></name> <name><surname>Munoz</surname> <given-names>D. P.</given-names></name></person-group> (<year>2017</year>). <article-title>Superior colliculus neurons encode a visual saliency map during free viewing of natural dynamic video</article-title>. <source>Nat. Commun</source>. <volume>8</volume>, <fpage>1</fpage>&#x02013;<lpage>9</lpage>. <pub-id pub-id-type="doi">10.1038/ncomms14263</pub-id><pub-id pub-id-type="pmid">28117340</pub-id></citation></ref>
<ref id="B132">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Whiteley</surname> <given-names>L.</given-names></name> <name><surname>Sahani</surname> <given-names>M.</given-names></name></person-group> (<year>2008</year>). <article-title>Implicit knowledge of visual uncertainty guides decisions with asymmetric outcomes</article-title>. <source>J. Vis</source>. <volume>8</volume>, <fpage>2</fpage>&#x02013;<lpage>2</lpage>. <pub-id pub-id-type="doi">10.1167/8.3.2</pub-id><pub-id pub-id-type="pmid">18484808</pub-id></citation></ref>
<ref id="B133">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Yang</surname> <given-names>S. C.-H.</given-names></name> <name><surname>Lengyel</surname> <given-names>M.</given-names></name> <name><surname>Wolpert</surname> <given-names>D. M.</given-names></name></person-group> (<year>2016a</year>). <article-title>Active sensing in the categorization of visual patterns</article-title>. <source>Elife</source> <volume>5</volume>, <fpage>e12215</fpage>. <pub-id pub-id-type="doi">10.7554/eLife.12215</pub-id><pub-id pub-id-type="pmid">28165321</pub-id></citation></ref>
<ref id="B134">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Yang</surname> <given-names>S. C.-H.</given-names></name> <name><surname>Wolpert</surname> <given-names>D. M.</given-names></name> <name><surname>Lengyel</surname> <given-names>M.</given-names></name></person-group> (<year>2016b</year>). <article-title>Theoretical perspectives on active sensing</article-title>. <source>Curr. Opin. Behav. Sci</source>. <volume>11</volume>, <fpage>100</fpage>&#x02013;<lpage>108</lpage>. <pub-id pub-id-type="doi">10.1016/j.cobeha.2016.06.009</pub-id><pub-id pub-id-type="pmid">30175197</pub-id></citation></ref>
</ref-list>
<fn-group>
<fn id="fn0001"><p><sup>1</sup>Covert attention is where saccadic eye movements do not occur.</p></fn>
<fn id="fn0002"><p><sup>2</sup>Overt attention deals with how an agent tracks the object with eye movements.</p></fn>
<fn id="fn0003"><p><sup>3</sup>The tilde over the variable refers to the generalized coordinates, i.e., the variable includes all temporal derivatives. Thus, <inline-formula><mml:math id="M17"><mml:mover accent="true"><mml:mrow><mml:mi>&#x003F5;</mml:mi></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover></mml:math></inline-formula> is the combined prediction error of outputs, inputs and states. For example, the generalized output &#x01EF9; is given by <inline-formula><mml:math id="M18"><mml:mover accent="true"><mml:mrow><mml:mstyle class="text"><mml:mtext mathvariant="bold">y</mml:mtext></mml:mstyle></mml:mrow><mml:mo>&#x0007E;</mml:mo></mml:mover><mml:mo>=</mml:mo><mml:msup><mml:mrow><mml:mrow><mml:mo>[</mml:mo><mml:mrow><mml:mi>y</mml:mi><mml:mo>,</mml:mo><mml:msup><mml:mrow><mml:mi>y</mml:mi></mml:mrow><mml:mrow><mml:mi>&#x02032;</mml:mi></mml:mrow></mml:msup><mml:mo>,</mml:mo><mml:msup><mml:mrow><mml:mi>y</mml:mi></mml:mrow><mml:mrow><mml:mo>&#x02033;</mml:mo></mml:mrow></mml:msup><mml:mo>.</mml:mo><mml:mo>.</mml:mo><mml:mo>.</mml:mo></mml:mrow><mml:mo>]</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mi>T</mml:mi></mml:mrow></mml:msup></mml:math></inline-formula>, where the prime operator denotes the derivatives. We use generalized coordinates (Friston et al., <xref ref-type="bibr" rid="B41">2010</xref>) for achieving accurate state and input estimation during the presence of (colored) noise by modeling the time dependent quantities (<italic>x, v, y, w, z</italic>) in generalized coordinates. This involves keeping track of the evolution of the trajectory of the probability distributions of states, instead of just their point estimates. Here the colored noise <italic>w</italic> and <italic>z</italic> are modeled as a white noise convoluted with a Gaussian kernel. The use of generalized coordinates has recently shown to outperform classical approaches under colored noise on real quadrotor flight (Bos et al., <xref ref-type="bibr" rid="B12">2021</xref>).</p></fn>
<fn id="fn0004"><p><sup>4</sup>Note that this expression of the variational free energy is using the Laplace and mean-field approximations commonly used in the FEP literature.</p></fn>
<fn id="fn0005"><p><sup>5</sup>System identification involves the estimation of system parameters (denoted by &#x003B8;, e.g., vectorised <bold>A</bold>), given <bold>y</bold>, <bold>u</bold>, by starting from a parameter prior of &#x003B7;<sup>&#x003B8;</sup> with prior precision <italic>P</italic><sup>&#x003B8;</sup>, and a prior on noise hyper-parameter &#x003B7;<sup>&#x003BB;</sup> with a prior precision of <italic>P</italic><sup>&#x003BB;</sup>. Note that we parametrise noise precision (&#x003A0;<sup><italic>w</italic></sup> and &#x003A0;<sup><italic>z</italic></sup>) using <inline-formula><mml:math id="M28"><mml:mrow><mml:mi>&#x003BB;</mml:mi><mml:mo>&#x02208;</mml:mo><mml:msup><mml:mi>&#x0211D;</mml:mi><mml:mrow><mml:mn>2</mml:mn><mml:mo>&#x000D7;</mml:mo><mml:mn>1</mml:mn></mml:mrow></mml:msup><mml:mo>=</mml:mo><mml:mrow><mml:mo>[</mml:mo> <mml:mrow><mml:mtable><mml:mtr><mml:mtd><mml:mrow><mml:msup><mml:mi>&#x003BB;</mml:mi><mml:mi>z</mml:mi></mml:msup></mml:mrow></mml:mtd></mml:mtr><mml:mtr><mml:mtd><mml:mrow><mml:msup><mml:mi>&#x003BB;</mml:mi><mml:mi>w</mml:mi></mml:msup></mml:mrow></mml:mtd></mml:mtr></mml:mtable></mml:mrow> <mml:mo>]</mml:mo></mml:mrow></mml:mrow></mml:math></inline-formula> as an exponential relation (e.g., &#x003A0;<sup><italic>w</italic></sup>(&#x003BB;<sup><italic>w</italic></sup>) &#x0003D; exp(&#x003BB;<sup><italic>w</italic></sup>)<italic>I</italic><sup><italic>n</italic>&#x000D7;<italic>n</italic></sup>).</p></fn>
<fn id="fn0006"><p><sup>6</sup>Note that here we are using exploration and exploration not in terms of behavior but for parameter learning. Exploration means adapting the parameter to a different (unexplored) value and exploitation means keeping that value.</p></fn>
</fn-group>
</back>
</article>