<?xml version="1.0" encoding="UTF-8" standalone="no"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Hum. Neurosci.</journal-id>
<journal-title>Frontiers in Human Neuroscience</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Hum. Neurosci.</abbrev-journal-title>
<issn pub-type="epub">1662-5161</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fnhum.2017.00540</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Neuroscience</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Changes of Attention during Value-Based Reversal Learning Are Tracked by N2pc and Feedback-Related Negativity</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name><surname>Oemisch</surname> <given-names>Mariann</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="author-notes" rid="fn001"><sup>&#x0002A;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/464187/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Watson</surname> <given-names>Marcus R.</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/79586/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Womelsdorf</surname> <given-names>Thilo</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="aff" rid="aff2"><sup>2</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/12918/overview"/>
</contrib> 
<contrib contrib-type="author">
<name><surname>Schub&#x000F6;</surname> <given-names>Anna</given-names></name>
<xref ref-type="aff" rid="aff3"><sup>3</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/49296/overview"/>
</contrib>
</contrib-group>
<aff id="aff1"><sup>1</sup><institution>Department of Biology, Centre for Vision Research, York University</institution>, <addr-line>Toronto, ON</addr-line>, <country>Canada</country></aff>
<aff id="aff2"><sup>2</sup><institution>Department of Psychology, Vanderbilt University</institution>, <addr-line>Nashville, TN</addr-line>, <country>United States</country></aff>
<aff id="aff3"><sup>3</sup><institution>Department of Psychology, Philipps-University Marburg</institution>, <addr-line>Marburg</addr-line>, <country>Germany</country></aff>
<author-notes>
<fn fn-type="edited-by"><p>Edited by: Xiaolin Zhou, Peking University, China</p></fn>
<fn fn-type="edited-by"><p>Reviewed by: Lihui Wang, Medizinische Fakult&#x000E4;t, Universit&#x000E4;tsklinikum Magdeburg, Germany; Sirawaj Itthipuripat, King Mongkut&#x02019;s University of Technology Thonburi, Thailand</p></fn>
<fn fn-type="corresp" id="fn001"><p>&#x0002A;Correspondence: Mariann Oemisch <email>moemisch&#x00040;yorku.ca</email></p></fn>
</author-notes>
<pub-date pub-type="epub">
<day>07</day>
<month>11</month>
<year>2017</year>
</pub-date>
<pub-date pub-type="collection">
<year>2017</year>
</pub-date>
<volume>11</volume>
<elocation-id>540</elocation-id>
<history>
<date date-type="received">
<day>01</day>
<month>08</month>
<year>2017</year>
</date>
<date date-type="accepted">
<day>24</day>
<month>10</month>
<year>2017</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x000A9; 2017 Oemisch, Watson, Womelsdorf and Schub&#x000F6;.</copyright-statement>
<copyright-year>2017</copyright-year>
<copyright-holder>Oemisch, Watson, Womelsdorf and Schub&#x000F6;</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/"><p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) or licensor are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p>
</license>
</permissions>
<abstract><p>Previously learned reward values can have a pronounced impact, behaviorally and neurophysiologically, on the allocation of selective attention. All else constant, stimuli previously associated with a high value gain stronger attentional prioritization than stimuli previously associated with a low value. The N2pc, an ERP component indicative of attentional target selection, has been shown to reflect aspects of this prioritization, by changes of mean amplitudes closely corresponding to selective enhancement of high value target processing and suppression of high value distractor processing. What has remained unclear so far is whether the N2pc also reflects the flexible and repeated behavioral adjustments needed in a volatile task environment, in which the values of stimuli are reversed often and unannounced. Using a value-based reversal learning task, we found evidence that the N2pc amplitude flexibly and reversibly tracks value-based choices during the learning of reward associated stimulus colors. Specifically, successful learning of current value-contingencies was associated with reduced N2pc amplitudes, and this effect was more apparent for distractor processing, compared with target processing. In addition, following a value reversal the feedback related negativity(FRN), an ERP component that reflects feedback processing, was amplified and co-occurred with increased N2pc amplitudes in trials following low-value feedback. Importantly, participants that showed the greatest adjustment in N2pc amplitudes based on feedback were also the most efficient learners. These results allow further insight into how changes in attentional prioritization in an uncertain and volatile environment support flexible adjustments of behavior.</p></abstract>
<kwd-group>
<kwd>visual selective attention</kwd>
<kwd>attentional learning</kwd>
<kwd>feedback</kwd>
<kwd>N2pc</kwd>
<kwd>reversal learning</kwd>
<kwd>EEG</kwd>
<kwd>reward</kwd>
<kwd>value learning</kwd>
</kwd-group>
<contract-num rid="cn001">MOP-102482</contract-num>
<contract-num rid="cn002">CRC/TRR 135, TP B3</contract-num>
<contract-sponsor id="cn001">Canadian Institutes of Health Research<named-content content-type="fundref-id">10.13039/501100000024</named-content></contract-sponsor>
<contract-sponsor id="cn002">Deutsche Forschungsgemeinschaft<named-content content-type="fundref-id">10.13039/501100001659</named-content></contract-sponsor>
<counts>
<fig-count count="6"/>
<table-count count="0"/>
<equation-count count="0"/>
<ref-count count="73"/>
<page-count count="16"/>
<word-count count="10988"/>
</counts>
</article-meta>
</front>
<body>
<sec sec-type="introduction" id="s1">
<title>Introduction</title>
<p>Visual selective attention allows the prioritization of task-relevant over irrelevant stimuli in the visual field. Traditionally, selective attention has been dissociated into goal-directed &#x0201C;top-down&#x0201D; driven selective attention and salience-driven &#x0201C;bottom-up&#x0201D; selective attention (e.g., Posner and Petersen, <xref ref-type="bibr" rid="B53">1990</xref>; Kastner and Ungerleider, <xref ref-type="bibr" rid="B39">2000</xref>; Corbetta and Shulman, <xref ref-type="bibr" rid="B14">2002</xref>). However, in recent years it has become evident that this dichotomy does not suffice to explain all instances in which a stimulus, or a set of stimuli, become the target of attentional priority (Awh et al., <xref ref-type="bibr" rid="B6">2012</xref>; Anderson, <xref ref-type="bibr" rid="B1">2013</xref>; Womelsdorf and Everling, <xref ref-type="bibr" rid="B67">2015</xref>). A third source of attentional control, referred to as &#x0201C;experience-driven&#x0201D;, includes an individual&#x02019;s recent selection-history and reward learning (Della Libera and Chelazzi, <xref ref-type="bibr" rid="B16">2006</xref>; Anderson et al., <xref ref-type="bibr" rid="B2">2011a</xref>; Awh et al., <xref ref-type="bibr" rid="B6">2012</xref>; Irons and Leber, <xref ref-type="bibr" rid="B36">2016</xref>). In particular, previously learned reward value has been shown to be a strong modulator of attentional prioritization (e.g., Della Libera and Chelazzi, <xref ref-type="bibr" rid="B17">2009</xref>; Krebs et al., <xref ref-type="bibr" rid="B42">2010</xref>; Anderson et al., <xref ref-type="bibr" rid="B3">2011b</xref>, <xref ref-type="bibr" rid="B4">2013</xref>, <xref ref-type="bibr" rid="B5">2014</xref>; Della Libera et al., <xref ref-type="bibr" rid="B18">2011</xref>; Hickey et al., <xref ref-type="bibr" rid="B28">2011</xref>, <xref ref-type="bibr" rid="B30">2015</xref>; Sali et al., <xref ref-type="bibr" rid="B55">2014</xref>; Bucker and Theeuwes, <xref ref-type="bibr" rid="B8">2017</xref>). For example, non-salient and task-irrelevant distractors that have previously been associated with reward can capture attention involuntarily and cause slower reaction times (RTs) in classical visual search tasks, and this is modulated by reward level, such that the higher the previously-associated reward, the greater the subsequent capture (e.g., Anderson et al., <xref ref-type="bibr" rid="B3">2011b</xref>, <xref ref-type="bibr" rid="B4">2013</xref>; Munneke et al., <xref ref-type="bibr" rid="B90">2015</xref>).</p>
<p>Neurophysiological evidence for this increased attentional capture by high-valued stimuli has been found in the mean amplitude of the N2pc (e.g., Kiss et al., <xref ref-type="bibr" rid="B40">2009</xref>; Feldmann-W&#x000FC;stefeld et al., <xref ref-type="bibr" rid="B24">2015</xref>, <xref ref-type="bibr" rid="B22">2016</xref>; Sawaki et al., <xref ref-type="bibr" rid="B57">2015</xref>). The N2pc is an EEG component thought to reflect attentional target selection processes (Luck and Hillyard, <xref ref-type="bibr" rid="B91">1994a</xref>; Woodman and Luck, <xref ref-type="bibr" rid="B68">2003</xref>; Eimer and Grubert, <xref ref-type="bibr" rid="B21">2014</xref>; Eimer, <xref ref-type="bibr" rid="B20">2014</xref>), likely generated in intermediate and high levels of the ventral visual processing pathway (Hopf et al., <xref ref-type="bibr" rid="B35">2000</xref>, <xref ref-type="bibr" rid="B34">2006</xref>). It onsets earlier and is more pronounced when search targets are associated with a higher value (Kiss et al., <xref ref-type="bibr" rid="B40">2009</xref>), and the presence of higher value distractors causes a decrease in target enhancement and an increase in distractor suppression during visual search (Feldmann-W&#x000FC;stefeld et al., <xref ref-type="bibr" rid="B22">2016</xref>). Sawaki et al. (<xref ref-type="bibr" rid="B57">2015</xref>) found that prior to visual search, a high value cue elicited stronger distractor suppression than a low value cue, and thereafter a smaller N2pc component was observed during the visual search. The authors argue that increased suppression to a high value but non-informative (to target selection) cue may have allowed better performance in the following visual search, which was supported by decreased RTs as well as decreased alpha oscillation levels prior to the visual search that indicated heightened visual readiness (Sawaki et al., <xref ref-type="bibr" rid="B57">2015</xref>).</p>
<p>We have thus already gained substantial insight into the behavior and neurophysiological processes that underlie the selective processing of high- or low-valued target and distractor signals. However, the distinction between targets and distractors is often much less clear outside the classic experimental environment. Real life is substantially more volatile, and therefore the stimuli that surround us must continuously be reevaluated with regards to their relevance to our current goals. A critical goal of attentional prioritization is likely reward maximization and loss minimization (e.g., Navalpakkam et al., <xref ref-type="bibr" rid="B49">2010</xref>), meaning that in a dynamic world we have to continuously learn and update our choice criteria with regards to the stimuli we attend.</p>
<p>We do not yet know how flexibly attentional target selection, as tracked by the N2pc component, can change in response to changes in reward values. In this study, we therefore employed a value-based reversal learning task in which stimulus reward values changed repeatedly and without warning. Specifically, participants were asked to attend to and choose one of two target stimuli that differed in color and likelihood of leading to a high reward outcome. This color-value association changed often and unannounced, such that the previously high value stimulus became the low value stimulus and vice versa. Participants therefore had to continuously re-evaluate, based on trial and error, whether they were choosing the currently high valued stimulus. This allowed us to assess learning-related changes in behavior, and to compare neural processing when subjects were in the process of learning the current value contingency, with processing when they had already successfully learned the current value contingency. Using EEG recordings, we examined learning- and choice-related differences to the N2pc. Since participants used trial-by-trial feedback to evaluate choices, we also examined feedback-related differences to the N2pc and learning-related differences to frontal feedback related negativity (FRN), which has previously been shown to change during reversal learning and has been suggested to encode prediction error signals and behavioral adjustments (e.g., Cohen and Ranganath, <xref ref-type="bibr" rid="B12">2007</xref>; Chase et al., <xref ref-type="bibr" rid="B9">2011</xref>; Walsh and Anderson, <xref ref-type="bibr" rid="B64">2011</xref>).</p>
<p>Neural processing of the valued stimuli was isolated by always placing one valued stimulus on the vertical midline, thereby attributing the lateralized EEG activity to the second, lateralized stimulus (e.g., Hickey et al., <xref ref-type="bibr" rid="B29">2009</xref>; Feldmann-W&#x000FC;stefeld et al., <xref ref-type="bibr" rid="B22">2016</xref>). Importantly, our task design did not have a fixed dissociation into &#x0201C;target&#x0201D; and &#x0201C;distractor&#x0201D;, since either of the two stimuli could be selected for response and the identity of the high and low value stimuli changed frequently. Instead, we differentiated processing of the selected (target) and the non-selected (distractor) stimulus on a trial-by-trial basis dependent on participants&#x02019; choice performance. We expected to observe learning-dependent changes in attentional prioritization reflected in N2pc amplitudes, and in feedback-processing reflected in FRN amplitudes (Figure <xref ref-type="fig" rid="F1">1</xref>). When performing similar tasks (e.g., Cools et al., <xref ref-type="bibr" rid="B13">2002</xref>; Chase et al., <xref ref-type="bibr" rid="B9">2011</xref>) participants generally show a low probability of choosing the newly high-valued stimulus in the trials immediately following a value reversal and therefore <italic>during</italic> learning of the current value contingency, and a high probability of choosing the currently high valued stimulus <italic>after</italic> successful learning (Figure <xref ref-type="fig" rid="F1">1A</xref>). We hypothesized that changes in N2pc and FRN amplitudes would parallel these changes in learning behavior. Specifically, we hypothesized that with successful learning of the current value contingency, the N2pc elicited by the stimulus selected for response (and therefore presumably actively attended), should increase, potentially reflecting more precise attentional selection of the current target stimulus (Figure <xref ref-type="fig" rid="F1">1B</xref>, left). We furthermore hypothesized that the N2pc elicited by the stimulus that was <italic>not</italic> selected for response (therefore presumably not actively attended) should generally be smaller than that of the stimulus that was selected for response, and it should further decrease with successful learning, potentially reflecting more successful avoidance of attentional capture by the current distractor stimulus (Figure <xref ref-type="fig" rid="F1">1B</xref>, right). Alternatively, it is possible that relatively fast learning during frequent value reversals is not reflected in changes in early attentional stimulus selection as measured with the N2pc, or that attentional processing of only the current target or only the current distractor is affected. Finally, we expected feedback processing reflected in FRN amplitudes to be greater <italic>during</italic> learning of the current value contingency compared with <italic>after</italic> learning, potentially reflecting the greater propensity to keep track of accumulated feedback when behavior needs to potentially be adjusted following a value reversal (Figure <xref ref-type="fig" rid="F1">1C</xref>, Chase et al., <xref ref-type="bibr" rid="B9">2011</xref>).</p>
<fig id="F1" position="float">
<label>Figure 1</label>
<caption><p>Hypotheses for learning-related changes in behavior, N2pc and feedback related negativity (FRN) amplitudes. <bold>(A)</bold> Successful learning (<italic>after</italic> learning) in our value based reversal learning task is reflected in an increased probability of choosing the currently high valued target. <bold>(B,C)</bold> We expected N2pc and FRN amplitudes to change in parallel with learning behavior. We hypothesized that N2pc amplitudes for the stimulus that was chosen for response, and therefore presumably actively attended, would increase with learning, to potentially reflect more accurate attentional target selection with successful learning (<bold>B</bold>, left). In contrast, we hypothesized that N2pc amplitudes for the stimulus that was <italic>not</italic> chosen for response, and therefore presumably <italic>not</italic> actively attended, would be substantially smaller than that for the stimulus chosen for response, and would further decrease with learning, to potentially reflect more successful avoidance of attentional capture by the distracting stimulus (<bold>B</bold>, right). <bold>(C)</bold> We hypothesized that FRN amplitudes would be greater <italic>during</italic> learning of the current value contingency compared with <italic>after</italic> learning, potentially reflecting the greater propensity to actively assess feedback during periods that may require behavioral adjustment.</p></caption>
<graphic xlink:href="fnhum-11-00540-g0001.tif"/>
</fig>
</sec>
<sec sec-type="materials and methods" id="s2">
<title>Materials and Methods</title>
<sec id="s2-1">
<title>Participants</title>
<p>Twenty-six students of the Philipps-University Marburg participated in the experiment for payment (8&#x020AC;/h) or course credit. Contingent on performance, participants could collect an additional monetary bonus of up to 6&#x020AC;. Three participants were rejected from analysis because too many trials (&#x0003E;40%) were lost due to EEG artifacts and non-learning (see &#x0201C;Data Analysis&#x0201D; section). Two participants had to be rejected due to technical issues. Analyses are shown for the remaining 21 participants (15 females, 6 males, mean age = 21.4). Three out of those 21 participants were left-handed. All participants were na&#x000EF;ve to the purpose of the experiment and had normal or corrected-to-normal visual acuity and normal color vision. Visual acuity and color vision were tested with an OCULUS Binoptometer 3 (OCULUS Optikger&#x000E4;te GmbH, Wetzlar, Germany). This study was carried out in accordance with the recommendations of the Ethics Committee of the Department of Psychology at the Philipps University Marburg with written informed consent from all participants, in accordance with the Declaration of Helsinki.</p>
</sec>
<sec id="s2-2">
<title>Apparatus and Stimuli</title>
<p>Participants were seated in a comfortable chair in a dimly lit and electrically shielded room, facing a monitor placed at a distance of approximately 100 cm from their eyes. Stimuli were presented on a 22&#x02033; screen (1680 &#x000D7; 1050 px) using Unity3D 5.3.5 (Unity Technologies, San Francisco, CA, USA). The display showed eight stimuli (diameter of 2.7&#x000B0;) arranged equidistantly on an imaginary circle with an eccentricity of 5.5&#x000B0; of visual angle (Figure <xref ref-type="fig" rid="F2">2A</xref>). All stimuli were presented against a gray background. Six of the stimuli were dark gray circles (RGB 24, 24, 24); the remaining two stimuli were one of two different target colors. Half the participants were presented with one pink (RGB 237, 83, 255) and one green (RGB 29, 181, 13) circle, half with one orange (RGB 217, 148, 14) and one blue (RGB 44, 168, 255) circle. All four possible stimulus colors were approximately isoluminant with the gray background (45.8&#x02013;55.65 cd/m<sup>2</sup>, luminance background: 51.12 cd/m<sup>2</sup>). Each stimulus contained a black line. Lines inside black stimuli were tilted 30&#x000B0; to the left or the right, alternating around the circle. Lines inside colored target stimuli were always tilted in opposite directions, by 45&#x000B0; to the left or 45&#x000B0; to the right. The two colored stimuli were always separated by one dark gray stimulus, such that one stimulus was always presented on the vertical midline either below or above the fixation cross, while the other was presented laterally to the left or right of the fixation cross. This experimental design was chosen because it allows isolating the processing related to the color stimulus presented laterally from processing of the color stimulus presented vertically. Traditionally, this design is used to isolate target-related from distractor-related processing (e.g., Hickey et al., <xref ref-type="bibr" rid="B29">2009</xref>). However, we do not have a pre-defined target and distractor, rather the same color stimulus changes roles several times throughout the experiment and participants are free to decide which color stimulus they select in a trial (Irons and Leber, <xref ref-type="bibr" rid="B36">2016</xref>; see &#x0201C;Procedure&#x0201D; section below). Thus, we are interested in how observers process stimuli which are associated with different reward values that change roles throughout the experiment.</p>
<fig id="F2" position="float">
<label>Figure 2</label>
<caption><p>Value-based reversal learning task and example block performance computed with Expectation Maximization (EM) algorithm. <bold>(A)</bold> The task started with a central fixation cross, followed by the appearance of eight black circles. For the target display, lines with different orientation appeared in all circles and two circles changed color to pink and green (or blue and orange). Participants had 1500 ms to report the line orientation (+45&#x000B0; or &#x02212;45&#x000B0; tilt) inside one of the two colored stimuli. Selecting the high value stimulus would lead to a high value feedback (+8) in 70% of the trials and to a low value feedback (+2) in 30% of the trials. This was reversed for the low value stimulus. Feedback was presented for 600 ms. If an incorrect color-line orientation combination was reported, a +0 was shown in red font for 1200 ms. After 1000&#x02013;1300 ms, a new trial was initiated. The response pad used by participants is illustrated below the task. <bold>(B)</bold> Top: schematic illustration of the value reversals applied to the colored stimuli in consecutive blocks. Bottom: displayed is the probability of a high value choice across trials for 9 example blocks performed by one representative participant. The probability of a high value choice was computed using an EM algorithm (see &#x0201C;Materials and Methods&#x0201D; section and Smith et al., <xref ref-type="bibr" rid="B61">2004</xref>). Dotted gray lines represent 95% confidence intervals. Dotted blue lines indicate the trial at which learning has occurred (T<sub>Learn</sub>) according to the ideal observer confidence interval. Trials are split into <italic>during</italic> learning and <italic>after</italic> learning trials according to T<sub>Learn</sub>. Black and gray boxes above plots indicate high value and low value choice trials, respectively.</p></caption>
<graphic xlink:href="fnhum-11-00540-g0002.tif"/>
</fig>
<p>Four response buttons were arranged on an Ergodex DX1 response pad, such that participants could comfortably place the middle and index fingers of both hands on the buttons. Each participant was randomly assigned to respond to a given stimulus color with the left or right hand, e.g., participants with green and pink stimuli were randomly assigned to respond to pink stimuli with their right hand and green stimuli with their left hand, or vice versa. In either case, the left-most finger of each hand (middle finger left hand and index finger right hand) was used to respond to a 45&#x000B0; tilt to the left (assuming a vertical line) and the right-most finger of each hand (index finger left hand and middle finger right hand) was used to respond to a 45&#x000B0; tilt to the right.</p>
</sec>
<sec id="s2-3">
<title>Procedure</title>
<p>The task is illustrated in Figure <xref ref-type="fig" rid="F2">2</xref>. Participants were instructed to keep their eyes on the center of the screen throughout a trial. Each trial started with the appearance of a central fixation cross for 500 ms, followed by eight dark gray circles without lines, presented for 400&#x02013;700 ms. Two circles, one on the vertical midline and one on the horizontal midline, then changed colors, and the lines appeared in all stimuli. Since lines inside colored stimuli were always tilted in opposite directions and participants were free to respond to either color stimulus, two &#x0201C;correct&#x0201D; (high and low value) choices were possible in any given trial (e.g., pink&#x02014;rightward orientation and green&#x02014;leftward orientation; trial example from Figure <xref ref-type="fig" rid="F2">2A</xref>). An incorrect response was recorded when a color/line-orientation pair was reported that was not presented in the display (e.g., pink&#x02014;leftward orientation and green&#x02014;rightward orientation). This stimulus display was presented for 200 ms, after which participants had 1500 ms to respond, followed by a delay of 600 ms after which feedback was shown for another 600 ms. Feedback had three possible values: a high-value &#x0201C;+8&#x0201D;, a low-value &#x0201C;+2&#x0201D;, or a &#x0201C;+0&#x0201D; for incorrect responses, the latter shown in red font for 1200 ms. If no response was made within 1500 ms, participants were asked to respond faster in the next trial via a written visual display. The inter-trial interval was 1000&#x02013;1300 ms.</p>
<p>Each stimulus display always contained the two colored stimuli (e.g., pink and green) and participants freely chose to report the line orientation of either of the two stimuli. At a given time, one color stimulus (e.g., pink) was associated with a 70% probability (high value) of leading to the outcome &#x0201C;+8&#x0201D; and a 30% probability of leading to the outcome &#x0201C;+2&#x0201D;. The second color stimulus (e.g., green) was simultaneously associated with a 30% probability (low value) of leading to an outcome of &#x0201C;+8&#x0201D; and a 70% probability of leading to an outcome of &#x0201C;+2&#x0201D;. Across trials of a block, the color-outcome probability association remained constant for 25 to a maximum of 50 trials (e.g., color 1 high valued). After trial 25, a running average of 80% high value choices over the last 12 trials triggered a block change (e.g., now color 2 high valued), and if this did not occur by trial 50, the block change happened automatically. The reversal was unannounced, requiring the participant to use performance feedback to detect reversals.</p>
<p>Participants were instructed to collect as many points as possible and to respond as fast as possible without jeopardizing response accuracy. They were explicitly informed of the 70%&#x02013;30% reward outcome distribution and understood that they should optimally always try to choose the color stimulus with the 70% high value outcome probability. Participants were also informed that the color-value associations would change within the experiment. Participants performed a total of 1200 trials, where stimulus positions and target line orientations were pseudo-randomly chosen on each trial. The color that was first associated with a high value was randomly chosen in each experimental session. Each experimental session (1200 trials) lasted approximately 100 min including a 10-min break after 50&#x02013;60 min. Each participant took part in one experimental session. After the experiment, participants filled in a questionnaire to assess strategies and other factors that may have influenced performance.</p>
</sec>
<sec id="s2-4">
<title>EEG Recording</title>
<p>The EEG was recorded continuously using BrainAmp amplifiers (Brain Products, Munich, Germany) from 64 Ag/AgCl electrodes (actiCAP) positioned according to the international modified 10-20 system. Vertical (vEOG) and horizontal electrooculograms (hEOG) were recorded as voltage difference between electrodes positioned above and below the left eye, and to the left and right canthi of the eyes, respectively. All channels were initially referenced to FCz and re-referenced offline to the average of all electrodes. Electrode impedances were kept below 5 k&#x003A9;. The sampling rate was 1000 Hz with a high cut-off filter of 250 Hz (half-amplitude cut-off, 30 dB/oct) and a low cut-off filter of 0.016 Hz (half-amplitude cut-off, 6 dB/oct).</p>
</sec>
<sec id="s2-5">
<title>Data Analysis</title>
<p>Analysis was performed with custom MATLAB code (Mathworks, Natick, MA, USA), utilizing functions from the open-source Fieldtrip toolbox<xref ref-type="fn" rid="fn0001"><sup>1</sup></xref>.</p>
<sec id="s2-5-1">
<title>Behavioral Data</title>
<p>Incorrect choices, defined as the reporting of a color/line-orientation pair not present in the display (see &#x0201C;Procedure&#x0201D; section), were discarded from all further analyses (5.3 &#x000B1; 0.12%). To identify at which trial during a block a participant showed statistically reliable learning of the current value rule, we analyzed the trial-by-trial choice dynamics using the state-space framework introduced by Smith and Brown (<xref ref-type="bibr" rid="B60">2003</xref>), and implemented by Smith et al. (<xref ref-type="bibr" rid="B61">2004</xref>). This framework entails a state equation that describes the internal learning process as a hidden Markov or latent process and is updated with each trial. The learning state process estimates the probability of a high value choice in each trial and thus provides the learning curve of participants (Figures <xref ref-type="fig" rid="F2">2B</xref>, <xref ref-type="fig" rid="F3">3B</xref>). The algorithm estimates learning from the perspective of an ideal observer that takes into account all trial outcomes of participants&#x02019; choices in a block of trials to estimate the probability that the outcome in a single trial is a high value response or a low value response. This probability is then used to calculate the confidence range of observing a high value response. We defined the learning trial (T<sub>Learn</sub>) as the earliest trial in a block at which the lower confidence bound of the probability for a high value response exceeded the <italic>p</italic> = 0.5 chance level. This corresponds to a 0.95 confidence level for an ideal observer to identify learning. The very first block of an experimental session and blocks in which no learning was identified were removed from further analyses. For most analyses, trials were split based on their occurrence prior to T<sub>Learn</sub> and after, into <italic>during</italic> learning trials and <italic>after</italic> learning trials, respectively.</p>
<fig id="F3" position="float">
<label>Figure 3</label>
<caption><p>Behavioral performance of the reversal learning task. <bold>(A)</bold> shows performance across all participants with the mean proportion of high value choices across trials <bold>(Ai)</bold> and a histogram of the length of blocks <bold>(Aii)</bold> as well as numbers of blocks per session across all participants (<bold>Aii</bold>, inset). Length of blocks can occasionally be shorter than the minimum of 25 trials because incorrect or late responses were not counted towards block lengths. <bold>(B)</bold> shows performance as measured using the EM algorithm. The mean probability of a high value choice with the average trial at which learning has occurred across participants is shown in <bold>(Bi)</bold>, and the distribution of learning trials across all blocks of all participants is shown in <bold>(Bii)</bold> (mean learning trial: 12.51; median learning trial: 11). <bold>(C)</bold> displays the effect of learning on reaction time (RT) across participants. Two asterisks indicate <italic>p</italic> &#x0003C; 0.01 (<italic>t</italic>-test on RT difference (<italic>during</italic>-<italic>after</italic> learning) across participants).</p></caption>
<graphic xlink:href="fnhum-11-00540-g0003.tif"/>
</fig>
<p>RTs <italic>during</italic> and <italic>after</italic> learning were compared by computing the difference between <italic>during</italic> and <italic>after</italic> learning trials per participant and comparing these differences across participants against a distribution with a mean of zero (<italic>t</italic>-test, <italic>&#x003B1;</italic> = 0.05). We tested whether RT and the probability to switch between colors were dependent on the previous trials&#x02019; feedback. For the first analysis, we simply compared RT in trial n for trials in which trial n-1 ended with a high value feedback and trials in which trial n-1 ended with a low value feedback. This comparison was made within participants as well as between participants (<italic>t</italic>-test, <italic>&#x003B1;</italic> = 0.05). For the latter analysis, we extracted trials in which a switch of choice was made from stimulus 1 (color 1) to stimulus 2 (color 2) or vice versa. We then computed a ratio of switch trials that followed a low value feedback vs. switch trials that followed a high value feedback. This ratio was then compared across participants against a distribution with a mean of one (<italic>t</italic>-test, <italic>&#x003B1;</italic> = 0.05).</p>
</sec>
<sec id="s2-5-2">
<title>EEG Data</title>
<p>For the N2pc analysis the EEG data was segmented into epochs of 700 ms, starting 200 ms prior to stimulus display onset and ending 500 ms after stimulus display onset. The time period from &#x02212;200 ms to 0 ms was used as baseline. Trials with blinks (vEOG &#x0003E;80 &#x003BC;V) or horizontal eye movements (hEOG &#x0003E;40 &#x003BC;V) were excluded from all analyses (across participants: 78 &#x000B1; 17 trials). The total trial number available for analysis following artifact removal and removal of non-learned blocks (see &#x0201C;Materials and Methods&#x0201D; section above) was 1047 &#x000B1; 25 trials across participants. The N2pc was measured at parieto-occipital electrode sites (PO3/4, PO7/8) as lateralized response to the laterally presented colored stimulus. The choice of electrode locations was based on the previously shown topography of N2pc subcomponents (Hickey et al., <xref ref-type="bibr" rid="B29">2009</xref>) and equivalent to earlier studies (Feldmann-W&#x000FC;stefeld and Schub&#x000F6;, <xref ref-type="bibr" rid="B23">2013</xref>; Feldmann-W&#x000FC;stefeld et al., <xref ref-type="bibr" rid="B24">2015</xref>). Difference waves were calculated by subtracting activity ipsilateral from activity contralateral to the lateral stimulus, and averaged separately for chosen and non-chosen stimuli to isolate choice-related N2pc differences <italic>during</italic> and <italic>after</italic> learning. In line with previous studies, mean amplitudes for the N2pc were computed for the time interval from 200 ms to 300 ms after stimulus display onset (Luck and Hillyard, <xref ref-type="bibr" rid="B92">1994b</xref>; Woodman and Luck, <xref ref-type="bibr" rid="B94">1999</xref>; Kiss et al., <xref ref-type="bibr" rid="B93">2008</xref>; Hickey et al., <xref ref-type="bibr" rid="B27">2010</xref>). Initial comparisons were made using a two-way repeated-measures analysis of variance (ANOVA) with the factors selection (chosen vs. non-chosen) and learning (during vs. after), and followed up by one-way repeated-measures ANOVAs (<italic>&#x003B1;</italic> = 0.05) with the factor learning conducted separately for chosen and non-chosen stimuli.</p>
<p>To investigate whether feedback in trial n-1 had an impact on the mean amplitude of the N2pc component in trial n, we isolated trials in which a choice to a lateralized color target followed a choice to the same color target presented lateralized, to verify that any effect was solely due to feedback. Specifically, a trial combination (trial n and n-1) was only selected for analysis if, e.g., a response was made to color 1 in trial n-1 and in trial n, and stimulus color 1 was presented at a lateralized position (left or right) in both trials. Following this restriction, total trial numbers available for this analysis across participants were 212 &#x000B1; 7. These trials were then sorted based on the feedback (high or low) received in trial n-1 and the mean amplitude of the N2pc component in trial n was compared in these two groups of trials. Comparisons were made using two-way repeated-measures ANOVA with the factors feedback value in trial n-1 and learning.</p>
<p>For FRN component analyses we extracted the data into 800 ms epochs, lasting from &#x02212;200 ms to 600 ms around the feedback event. Similar to previous studies (e.g., Hajcak et al., <xref ref-type="bibr" rid="B26">2006</xref>; Cohen et al., <xref ref-type="bibr" rid="B11">2007</xref>), the FRN component was isolated at the Fz electrode, and as a control additionally at the FCz electrode (see &#x0201C;Results&#x0201D; section, data not shown). Difference waves were calculated by subtracting activity for high value feedback from activity for low value feedback in the 250&#x02013;325 ms following feedback onset, which generally fell within the time range investigated in previous studies (for review see Walsh and Anderson, <xref ref-type="bibr" rid="B65">2012</xref>). The comparison of FRN mean amplitude <italic>during</italic> vs. <italic>after</italic> learning was computed using two-way repeated measures ANOVA, with the factors feedback value and learning.</p>
<p>To assess a more general effect of learning on feedback processing independent of valence (FRN), we performed a three-way ANOVA with the factors learning, feedback value and time window (12 non-overlapping 50-ms windows ranging from 0 ms to 600 ms post feedback). Follow-up tests of simple main effects were done using one-way ANOVA&#x02019;s in each time window with <italic>p</italic>-values corrected for multiple comparisons using the Bonferroni-Holm method.</p>
<p>For visualization purposes only, the N2pc and FRN displayed in Figures <xref ref-type="fig" rid="F4">4</xref><xref ref-type="fig" rid="F5"></xref>&#x02013;<xref ref-type="fig" rid="F6">6</xref> were smoothed with a moving average filter of 25 ms (40 Hz).</p>
<fig id="F4" position="float">
<label>Figure 4</label>
<caption><p>Lateralized N2pc components. <bold>(A,B)</bold> Contra- and ipsi-lateral mean amplitudes from pooled left (PO3, PO7) and right (PO4, PO8) parieto-occipital electrodes are shown aligned to target display onset for lateralized chosen <bold>(A)</bold> and non-chosen <bold>(B)</bold> stimuli <italic>during</italic> <bold>(Ai,Bi)</bold> and <italic>after</italic> <bold>(Aii,Bii)</bold> learning. The gray bars highlight the N2pc time window of analysis (200&#x02013;300 ms). N2pc difference waves contrasted <italic>during</italic> <bold>(Ci)</bold> and <italic>after</italic> <bold>(Cii)</bold> learning. Example trials are illustrated in the top left corners. The topography of the N2pc (200&#x02013;300 ms) for non-chosen stimuli is shown below <bold>(Cii)</bold>.</p></caption>
<graphic xlink:href="fnhum-11-00540-g0004.tif"/>
</fig>
<fig id="F5" position="float">
<label>Figure 5</label>
<caption><p>Effect of previous trial feedback on N2pc amplitude. <bold>(A)</bold> Illustration of example trial sequences for &#x0201C;previous feedback high&#x0201D; and &#x0201C;previous feedback low&#x0201D; trials. Trials for analysis were chosen based on the previous trial feedback (n-1), the N2pc analysis was done on trial n. <bold>(B)</bold> N2pc difference wave in trial n following high value (blue line) or low value (red line) feedback in the previous trial. Gray shaded area highlight the N2pc analysis window (200&#x02013;300 ms after stimulus display onset). Asterisk indicates <italic>p</italic> &#x02264; 0.05 (one-way ANOVA). <bold>(C)</bold> Mean difference in N2pc amplitude following high vs. low value feedback <italic>during</italic> learning and <italic>after</italic> learning. <bold>(D)</bold> Correlation between individual participants&#x02019; T<sub>Learn</sub> and N2pc amplitude differences between feedback (previous high-value feedback &#x02212; previous low-value feedback) <italic>during</italic> learning (left) and <italic>after</italic> learning (right). Blue lines represent least square fit. Note that large positive differences indicate a large N2pc difference for high vs. low-value feedback in trial n-1.</p></caption>
<graphic xlink:href="fnhum-11-00540-g0005.tif"/>
</fig>
<fig id="F6" position="float">
<label>Figure 6</label>
<caption><p>FRN amplitudes. <bold>(A)</bold> Aligned to feedback presentation onset, mean activity at Fz electrode is shown for high and low value feedback <italic>during</italic> and <italic>after</italic> learning trials. <bold>(B)</bold> The mean difference (low value feedback &#x02212; high value feedback) for the FRN wave is shown <italic>during</italic> and <italic>after</italic> learning. The gray bar indicates the analysis time window for the FRN (250&#x02013;325 ms after feedback onset). For visualization purposes only, the FRN wave was smoothed with a moving average filter of 5 ms. <bold>(C)</bold> Topography of the FRN (difference elicited by low vs. high value feedback) <italic>during</italic> and <italic>after</italic> learning. Red circles identify Fz electrodes. <bold>(D)</bold> Mean activity from <bold>(A)</bold> is shown collapsed across feedback valence to illustrate the time span of the simple main effect of learning on feedback processing (gray bar).</p></caption>
<graphic xlink:href="fnhum-11-00540-g0006.tif"/>
</fig>
</sec>
<sec id="s2-5-3">
<title>Correlations</title>
<p>We compared mean differences in N2pc amplitudes following low vs. high value feedback with mean behavioral measures on an individual participant level using Pearson correlation (<italic>&#x003B1;</italic> = 0.05). The three behavioral measures tested included: (i) the proportion of blocks successfully learned; (ii) the mean trial at which learning was deemed successful across blocks (i.e., T<sub>Learn</sub>); and (iii) mean RT. These three behavioral measures were not correlated across participants (Pearson correlation, all <italic>p</italic> &#x0003E; 0.05). We compared correlation coefficients obtained e.g., <italic>during</italic> learning vs. <italic>after</italic> learning by using a <italic>z</italic>-test to assess differences between two dependent correlations (Steiger, <xref ref-type="bibr" rid="B62">1980</xref>). When the observed <italic>z</italic>-value was greater than |1.96|, we considered the correlation coefficients significantly different.</p>
</sec>
</sec>
</sec>
<sec sec-type="results" id="s3">
<title>Results</title>
<sec id="s3-1">
<title>Reversal Learning</title>
<p>Behavioral results are plotted in Figure <xref ref-type="fig" rid="F3">3</xref>. Participants performed the task very well and generally showed quick increases in the proportion of high value choices following a value reversal (Figures <xref ref-type="fig" rid="F2">2B</xref>, <xref ref-type="fig" rid="F3">3Ai,Bi</xref>), in line with the behavioral assumptions (Figure <xref ref-type="fig" rid="F1">1A</xref>). This is further shown in the distribution of block lengths observed across all participants, whereby the majority of blocks had a length of approximately 25 trials, indicating a performance of 80% high value choices around the time trial 25 was reached (Figure <xref ref-type="fig" rid="F3">3Aii</xref>, note that blocks shorter than 25 trials were possible following the rejection of incorrect responses, see &#x0201C;Materials and Methods&#x0201D; section). Participants performed a mean of 41.1 blocks per experimental session. 86% out of all blocks were successfully learned (across participants: 86 &#x000B1; 1.6%). Learning of the current value rule across blocks and participants occurred within a mean of 12.5 &#x000B1; 0.5 trials as identified using the Expectation Maximization algorithm by Smith et al. (<xref ref-type="bibr" rid="B61">2004</xref>; Figure <xref ref-type="fig" rid="F3">3B</xref>, median: 11).</p>
<p>RT significantly decreased with learning of the current value rule. Participants showed on average 13.2 ms shorter RTs in trials after acquiring the current rule (<italic>after</italic> learning, 609 &#x000B1; 10 ms) as opposed to trials beforehand (<italic>during</italic> learning, 622 &#x000B1; 12 ms; <italic>t</italic><sub>(20)</sub> = 3.31, <italic>p</italic> = 0.004, Figure <xref ref-type="fig" rid="F3">3C</xref>).</p>
</sec>
<sec id="s3-2">
<title>Attention Deployment Changes with Learning</title>
<p>Mean amplitudes for the N2pc were computed for the time interval from 200 ms to 300 ms after stimulus display onset (Figure <xref ref-type="fig" rid="F4">4</xref>). An initial two-way repeated-measures ANOVA tested the effects of the factors selection and learning on N2pc amplitudes. A main effect of selection (<italic>F</italic><sub>(1,20)</sub> = 44.04, <italic>p</italic> &#x0003C; 0.001) showed that a pronounced N2pc was elicited when the chosen stimulus was presented at a lateralized position (&#x00394;<sub>(contra-ipsi)</sub> = &#x02212;1.901 &#x000B1; 0.29 &#x003BC;V), which was substantially reduced when the non-chosen stimulus was presented laterally and the chosen stimulus on the vertical midline (&#x00394;<sub>(contra-ipsi)</sub> = &#x02212;0.271 &#x000B1; 0.16 &#x003BC;V). A main effect of learning additionally suggested that N2pc amplitudes differed <italic>during</italic> learning and <italic>after</italic> learning (<italic>F</italic><sub>(1,20)</sub> = 10.79, <italic>p</italic> = 0.004). The <italic>post hoc</italic> comparison showed that N2pc amplitudes were significantly larger (more negative) <italic>during</italic> learning (&#x00394;<sub>(contra-ipsi)</sub> = &#x02212;1.155 &#x000B1; 0.20 &#x003BC;V) than <italic>after</italic> learning (&#x00394;<sub>(contra-ipsi)</sub> = &#x02212;1.016 &#x000B1; 0.20 &#x003BC;V). Therefore, as initially predicted (Figure <xref ref-type="fig" rid="F1">1B</xref>) we found main effects of both learning and selection on the N2pc. However, contrary to our expectation (Figure <xref ref-type="fig" rid="F1">1B</xref>), we did not find a significant interaction between the two (<italic>F</italic><sub>(1,20)</sub> = 1.04, <italic>p</italic> = 0.319). While the absolute magnitude of the N2pc for the non-chosen stimulus was higher <italic>during</italic> learning than <italic>after</italic> learning (Figure <xref ref-type="fig" rid="F4">4Cii</xref>), and this was significant as a main effect of a one-way ANOVA (<italic>F</italic><sub>(1,20)</sub> = 5.36, <italic>p</italic> = 0.031) as predicted (Figure <xref ref-type="fig" rid="F1">1B</xref>, right), the magnitude for the chosen stimulus was virtually identical <italic>during</italic> and <italic>after</italic> learning (<italic>F</italic><sub>(1,20)</sub> = 1.41, <italic>p</italic> = 0.249, Figure <xref ref-type="fig" rid="F4">4Ci</xref>), which does not match our prediction (Figure <xref ref-type="fig" rid="F1">1B</xref>, left). Thus, our results suggest that the primary effect of learning on the N2pc amplitude in our task is to suppress attention to non-chosen distractors, rather than to enhance attention to chosen targets. The lack of apparent target enhancement might explain why the two-way interaction was not significant, despite the apparent effect of learning on the non-chosen stimulus. Given this lack of a significant interaction, the differential results should be treated as suggestive, rather than definitive.</p>
<p>In summary, N2pc results showed that attention was mainly deployed to the chosen stimulus compared with the non-chosen stimulus, and that attention deployment was generally more pronounced <italic>during</italic> learning compared with <italic>after</italic> learning. In contrast to our hypotheses (Figure <xref ref-type="fig" rid="F1">1B</xref>), the direction of the effects of learning were not opposing for chosen and non-chosen stimuli. Thus, even though we did not observe an interaction between learning and selection, we nevertheless found evidence suggestive of successful learning mainly leading to a decrease in processing of the non-chosen compared with the chosen stimulus (Figure <xref ref-type="fig" rid="F4">4C</xref>). We therefore found partial evidence in line with our hypothesis for the effect of learning on non-chosen stimuli (Figure <xref ref-type="fig" rid="F1">1B</xref>, right).</p>
</sec>
<sec id="s3-3">
<title>Low Value Feedback Is Followed by Increased Attentional Target Selection</title>
<p>To investigate the impact of low or high value feedback on behavioral or electrophysiological measures, we assessed whether RT, probability to switch color choices, and N2pc mean amplitudes were affected by the previous trial&#x02019;s feedback. Whether the feedback in trial n-1 was of high or low value had no impact on RT in the following trial n (across all trials: <italic>t</italic><sub>(12,470)</sub> = &#x02212;0.36, <italic>p</italic> = 0.719). It did, however, have an impact on the likelihood to switch choices from stimulus color 1 to stimulus color 2 or vice versa. Across participants, a choice switch was more likely to occur following a low value feedback compared with a high value feedback (<italic>t</italic><sub>(20)</sub> = 5.10, <italic>p</italic> = 5.4 &#x000D7; 10<sup>&#x02212;5</sup>).</p>
<p>Feedback in trial n-1 also had an impact on N2pc mean amplitude in trial n. Overall, if a choice was made to the same target in trials n-1 and n (Figure <xref ref-type="fig" rid="F5">5A</xref>), the N2pc amplitude in trial n was larger following a low value feedback in trial n-1 (&#x00394;<sub>(contra-ipsi)</sub> = &#x02212;2.041 &#x000B1; 0.30 &#x003BC;V), compared with following a high value feedback in trial n-1 (&#x00394;<sub>(contra-ipsi)</sub> = &#x02212;1.671 &#x000B1; 0.33 &#x003BC;V; <italic>F</italic><sub>(1,20)</sub> = 4.52, <italic>p</italic> = 0.046; Figure <xref ref-type="fig" rid="F5">5B</xref>). This was also the case when we did not explicitly control for the choice in trial n (i.e., a choice switch could occur from trial n-1 to trial n, <italic>F</italic><sub>(1,20)</sub> = 5.51, <italic>p</italic> = 0.029, data not shown). We were interested in whether feedback had a differential effect on N2pc amplitudes depending on the current state of learning, and therefore separated trials into <italic>during</italic> learning and <italic>after</italic> learning trials. We did not find a significant interaction between the factors feedback and learning in a two-way ANOVA (<italic>F</italic><sub>(1,20)</sub> = 0.82, <italic>p</italic> = 0.375). Nevertheless, the difference in N2pc amplitudes following high vs. low value feedback tended to be greater <italic>during</italic> learning than <italic>after</italic> learning (Figure <xref ref-type="fig" rid="F5">5C</xref>). As previously, this should be treated as suggestive rather than definitive.</p>
<p>Across individual participants, this difference in N2pc amplitude following high vs. low value feedback was significantly correlated with learning performance <italic>during</italic> learning (Figure <xref ref-type="fig" rid="F5">5D</xref>, left), but not <italic>after</italic> learning (Figure <xref ref-type="fig" rid="F5">5D</xref>, right). Specifically, the greater the individual difference in N2pc amplitude following high vs. low value feedback <italic>during</italic> learning, the faster the individual learned, i.e., T<sub>Learn</sub> was smaller (Pearson correlation, <italic>R</italic> = &#x02212;0.545, <italic>p</italic> = 0.011; Figure <xref ref-type="fig" rid="F5">5D</xref>, left). However, the difference in N2pc amplitudes following high vs. low value feedback <italic>after</italic> learning was not related to learning performance (Pearson correlation, <italic>R</italic> = 0.058, <italic>p</italic> = 0.803; Figure <xref ref-type="fig" rid="F5">5D</xref>, right). This difference in correlation coefficients between <italic>during</italic> learning and <italic>after</italic> learning was significant (<italic>z</italic>-test to compare <italic>R</italic>-values, <italic>z</italic> = 2.12, <italic>p</italic> = 0.034). The difference in N2pc amplitude following high vs. low value feedback was not correlated with average RT or the proportion of blocks learned (all <italic>p</italic> &#x0003E; 0.05).</p>
</sec>
<sec id="s3-4">
<title>Feedback Processing Is Increased <italic>during</italic> Learning</title>
<p>Considering the finding that feedback has a differential effect on N2pc amplitudes, and that this effect specifically <italic>during</italic> learning correlates with successful learning behavior, we asked whether feedback processing was affected by learning. We therefore computed the mean amplitude of the FRN as a proxy for negative feedback processing (Miltner et al., <xref ref-type="bibr" rid="B48">1997</xref>). The FRN, computed as the difference between low and high value feedback presentation, was measured at the Fz electrode, since the amplitude difference between low and high value feedback was largest at this electrode. The qualitative and quantitative pattern of results did not change when the FRN was measured at the FCz electrode instead (data not shown). The analysis was done within the 250&#x02013;325 ms after feedback presentation, since the difference between low and high value was largest in this window (see below) and it generally fell within the range used in the literature (for review see Walsh and Anderson, <xref ref-type="bibr" rid="B65">2012</xref>). We found that processing of feedback (low and high value) was generally increased <italic>during</italic> learning compared with <italic>after</italic> learning (Figure <xref ref-type="fig" rid="F6">6A</xref>), and the difference between low and high value feedback (FRN) was more pronounced <italic>during</italic> learning compared with <italic>after</italic> learning (<italic>during</italic>: &#x00394;<sub>(lowFB-highFB)</sub> = &#x02212;0.791 &#x000B1; 0.17 &#x003BC;V; <italic>after</italic>: &#x00394;<sub>(lowFB-highFB)</sub> = &#x02212;0.426 &#x000B1; 0.18 &#x003BC;V). The resulting FRN was therefore substantially larger <italic>during</italic> learning as compared with <italic>after</italic> learning (Figure <xref ref-type="fig" rid="F6">6B</xref>). This was confirmed with a two-way ANOVA that showed a significant main effect of feedback value (<italic>F</italic><sub>(1,20)</sub> = 14.70, <italic>p</italic> = 0.001), a significant main effect of learning (<italic>F</italic><sub>(1,20)</sub> = 37.18, <italic>p</italic> &#x0003C; 0.001), and a significant interaction between the two parameters (<italic>F</italic><sub>(1,20)</sub> = 6.04, <italic>p</italic> = 0.023). Topographical maps of the amplitude difference between low and high value feedback <italic>during</italic> and <italic>after</italic> learning are shown in Figure <xref ref-type="fig" rid="F6">6C</xref>.</p>
<p>In addition to the change in FRN amplitude with learning, visual inspection of the plots (Figure <xref ref-type="fig" rid="F6">6A</xref>), revealed a much longer effect of learning that was distinct from the effect of feedback value and the interaction of feedback value and learning. To tease apart these effects and to determine the time range of the effect of learning, we ran a three-way ANOVA with the factors learning (<italic>during</italic> learning, <italic>after</italic> learning), feedback value (high, low), and time window, where we defined 12 50 ms non-overlapping time windows running from 0 ms to 600 ms following feedback onset. We found interactions between the factors learning and time window (<italic>F</italic><sub>(11,220)</sub> = 7.13, <italic>p</italic> &#x0003C; 0.001), and feedback value and time window (<italic>F</italic><sub>(11,220)</sub> = 3.88, <italic>p</italic> &#x0003C; 001). Follow-up simple main effects across time windows showed that feedback processing <italic>per se</italic> differed with learning in all time windows from 150 ms to 400 ms following feedback onset (F-values between 18.07 and 47.29, all <italic>p</italic> &#x0003C; 0.005, <italic>p-values</italic> in all other time windows &#x0003E;0.05, Bonferroni-Holm multiple comparison corrected). A simple main effect of feedback was only found in the 250&#x02013;300 ms time window (<italic>F</italic><sub>(1,20)</sub> = 16.39, <italic>p</italic> = 0.008, all other <italic>p</italic> > <italic>0</italic>.05, Bonferroni-Holm multiple comparison corrected), confirming the initial FRN analysis above. The previous suggests that feedback processing independent of valence was increased <italic>during</italic> learning in the time window from 150 ms to 400 ms following feedback onset (Figure <xref ref-type="fig" rid="F6">6D</xref>).</p>
</sec>
</sec>
<sec sec-type="discussion" id="s4">
<title>Discussion</title>
<p>In this study, we implemented a value-based reversal learning task to explore in more detail how attentional target selection and feedback processing is realized in a highly volatile task environment. We measured the N2pc, an EEG component thought to reflect attentional target selection, and the FRN, an EEG component thought to reflect negative feedback processing or prediction error encoding, while participants performed a value-based reversal learning task with probabilistic feedback. Participants were required to frequently adjust their color-based stimulus choice using recent reward feedback. We found that: (i) participants used feedback efficiently to reverse back and forth between the two stimulus choices in accordance with the reversal of their respective values (Figures <xref ref-type="fig" rid="F1">1A</xref>, <xref ref-type="fig" rid="F2">2</xref>, <xref ref-type="fig" rid="F3">3</xref>); (ii) successful learning of the current value-contingency led to a decrease in N2pc amplitudes, which was particularly evident for non-chosen (distractor) stimuli compare with chosen (target) stimuli (Figure <xref ref-type="fig" rid="F4">4</xref>); (iii) negative feedback in the previous trial was associated with an increase in N2pc mean amplitude, which was selectively correlated with an enhanced learning rate <italic>during</italic> learning (Figure <xref ref-type="fig" rid="F5">5</xref>); and (iv) FRN amplitudes were increased <italic>during</italic> learning of the current value contingencies, which co-occurred with a more general increase of feedback processing <italic>during</italic> learning (Figure <xref ref-type="fig" rid="F6">6</xref>).</p>
<p>We live in an environment that is usually much more uncertain and volatile than a classical experimental setting, in which objects or actions need to be continuously evaluated for their relevance to our current goal. We attempted to imitate some of this volatility by employing a value-based reversal learning task, albeit one that is clearly more restrictive than the world outside the laboratory.</p>
<p>To the best of our knowledge this is the first study that investigated learned value-dependent changes of the N2pc amplitude elicited by chosen (i.e., selected) and non-chosen (non-selected) stimuli in a task in which participants were free to select a stimulus for response. Most studies that have investigated the effects of value or reward on attentional stimulus selection in behavior, have used tasks with a predefined, fixed target and distractor-assignment, and in which the trial-by-trial association between the specific stimulus features and reward were in fact irrelevant to solving the task (Della Libera and Chelazzi, <xref ref-type="bibr" rid="B17">2009</xref>; Hickey et al., <xref ref-type="bibr" rid="B27">2010</xref>, <xref ref-type="bibr" rid="B30">2015</xref>; Anderson et al., <xref ref-type="bibr" rid="B2">2011a</xref>, <xref ref-type="bibr" rid="B4">2013</xref>; Itthipuripat et al., <xref ref-type="bibr" rid="B38">2015</xref>; Sawaki et al., <xref ref-type="bibr" rid="B57">2015</xref>; Feldmann-W&#x000FC;stefeld et al., <xref ref-type="bibr" rid="B22">2016</xref>), often to dissociate reward-related processes from goal-related processes of attentional selection. Or they have used tasks in which value associations of targets or cues were kept constant (Kiss et al., <xref ref-type="bibr" rid="B40">2009</xref>; Raymond and O&#x02019;Brien, <xref ref-type="bibr" rid="B54">2009</xref>; Krebs et al., <xref ref-type="bibr" rid="B42">2010</xref>; Le Pelley et al., <xref ref-type="bibr" rid="B44">2013</xref>; San Mart&#x000ED;n et al., <xref ref-type="bibr" rid="B56">2016</xref>), or if they changed, were specifically trained (Navalpakkam et al., <xref ref-type="bibr" rid="B49">2010</xref>). None of these studies allowed insight into how attentional selection changes when the values of target stimuli change unannounced and require repeated adjustment of behavior.</p>
<p>We used probabilistic feedback so that subjects needed to keep track of feedback over multiple trials to determine the current high-value stimulus. Following an unannounced value reversal, participants tended to continue to choose the low value (previously high value) stimulus for response for multiple trials, before switching their choice behavior to the current high value stimulus. RTs were longer <italic>during</italic> learning than <italic>after</italic> learning (Figure <xref ref-type="fig" rid="F3">3C</xref>), suggesting that participants optimized their attention allocation to stimulus features with learning of the current value contingencies.</p>
<p>Since both stimuli were repeatedly associated with a high and a low value, both were frequently selected for a response. Thus, the dissociation between target and distractor is not as clear on a trial by trial basis as in previous literature (see above). For this reason, it was initially unclear to what extent processing of the chosen (that is, the selected target) stimulus would be enhanced throughout learning and how efficiently the brain could evade attentional capture by the non-chosen (non-selected distractor) stimulus, which always posed a distraction to solving the task quickly and accurately. We initially predicted that as attentional prioritization shifts towards the chosen stimulus with learning, this would concomitantly be reflected by an N2pc increase for the chosen stimulus (Figure <xref ref-type="fig" rid="F1">1B</xref>, left) and an N2pc decrease for the non-chosen stimulus (Figure <xref ref-type="fig" rid="F1">1B</xref>, right). Instead we found that N2pc amplitudes decreased with learning in general, which was true on average for both chosen and non-chosen stimuli. However, this decrease in amplitude with learning seemed more apparent for the non-chosen stimulus (Figure <xref ref-type="fig" rid="F4">4C</xref>). This suggests that the primary effect of learning in this task was a decrease in attention to the non-chosen lateralized stimulus, which potentially indicates suppression (Figure <xref ref-type="fig" rid="F4">4Cii</xref>). In contrast, the amplitude of the N2pc that reflected processing of the chosen lateralized stimulus did not seem to substantially change with learning (Figure <xref ref-type="fig" rid="F4">4Ci</xref>). Thus, we find some evidence for our initial hypothesis of how learning affects processing of non-chosen stimuli (Figure <xref ref-type="fig" rid="F1">1B</xref>, right), but not for our hypothesis of how learning affects processing of chosen stimuli (Figure <xref ref-type="fig" rid="F1">1B</xref>, left). These results suggest that following a value reversal, when participants needed to actively re-evaluate their current choices, distraction by the non-chosen stimulus was not as effectively evaded as <italic>after</italic> learning, when participants often showed plateau-performance with a high probability of choosing the currently high-valued stimulus (Figures <xref ref-type="fig" rid="F2">2</xref>, <xref ref-type="fig" rid="F3">3</xref>).</p>
<p>These results suggest that efficient attention allocation in this highly volatile task design was more likely observed for processing of the non-chosen stimulus in the form of a reduced or suppressed N2pc, and not as an N2pc enhancement of the chosen stimulus. We should note however that although we observed suppression for lateralized non-chosen stimuli, independent of learning, the amplitude elicited in this time interval was still negative, and not positive as has been observed previously (e.g., Hickey et al., <xref ref-type="bibr" rid="B29">2009</xref>; Feldmann-W&#x000FC;stefeld et al., <xref ref-type="bibr" rid="B24">2015</xref>, <xref ref-type="bibr" rid="B22">2016</xref>; Sawaki et al., <xref ref-type="bibr" rid="B57">2015</xref>). It is therefore difficult to decide whether value learning has led to a reduced capture by the non-chosen (i.e., non-selected) stimulus or an actual suppression, as both interpretations would account for a reduction in N2pc amplitude. Similarly, it is possible that in a less volatile task design in which learning takes place over a much longer time window (e.g., days), an effect of learning on attentional processing would predominantly be observed for the chosen (target) stimulus, as has been shown previously (Clark et al., <xref ref-type="bibr" rid="B10">2015</xref>; Itthipuripat et al., <xref ref-type="bibr" rid="B37">2017</xref>). We should therefore be careful of over-interpreting the absence of a strong effect of learning on N2pc amplitudes of the chosen stimulus in this task, as such an effect could have been revealed with a larger number of <italic>after</italic>-learning trials.</p>
<p>That attention and learning are closely intertwined concepts has been the subject of attentional learning theories for some time (Mackintosh, <xref ref-type="bibr" rid="B46">1975</xref>; Pearce and Hall, <xref ref-type="bibr" rid="B51">1980</xref>; Le Pelley, <xref ref-type="bibr" rid="B43">2004</xref>). The Mackintosh and the Pearce and Hall models of attentional learning predict contradicting relationships between attention and learning. According to Mackintosh (<xref ref-type="bibr" rid="B46">1975</xref>), attention is biased towards stimuli that have a higher predictive value, as they are more likely to yield a rewarding outcome (e.g., Mackintosh and Little, <xref ref-type="bibr" rid="B47">1969</xref>; Le Pelley et al., <xref ref-type="bibr" rid="B44">2013</xref>). The Pearce and Hall model on the contrary predicts that unexpected and surprising outcomes that lead to a prediction error are associated with an increase in attention (e.g., Wilson et al., <xref ref-type="bibr" rid="B66">1992</xref>; Anderson et al., <xref ref-type="bibr" rid="B4">2013</xref>). Both theories have received extensive empirical support (Pearce and Mackintosh, <xref ref-type="bibr" rid="B52">2010</xref>) and there have been efforts to reconcile their findings (Holland and Gallagher, <xref ref-type="bibr" rid="B32">1999</xref>; Dayan et al., <xref ref-type="bibr" rid="B15">2000</xref>; Le Pelley, <xref ref-type="bibr" rid="B43">2004</xref>; Hogarth et al., <xref ref-type="bibr" rid="B31">2010</xref>). One possible solution suggests that there is a distinction between two aspects of attention in associative learning: attention that is concerned with action, and attention that is concerned with learning. It suggests that one should attend to the most reward-predicting stimuli or features when making a choice, but should attend to the most uncertain stimuli or features when learning from prediction errors (Holland and Gallagher, <xref ref-type="bibr" rid="B32">1999</xref>; Dayan et al., <xref ref-type="bibr" rid="B15">2000</xref>; Hogarth et al., <xref ref-type="bibr" rid="B31">2010</xref>; Gottlieb, <xref ref-type="bibr" rid="B25">2012</xref>).</p>
<p>Even though our task was not designed to address this debate explicitly, we found evidence that in a highly volatile environment that encourages continuous learning, attention is increased during periods of uncertainty in line with Pearce and Hall (<xref ref-type="bibr" rid="B51">1980</xref>). In particular, we tested whether feedback itself could influence the allocation of attention to target stimuli and found that low-value feedback was followed by a larger N2pc amplitude on the next trial than was high-value feedback (Figure <xref ref-type="fig" rid="F5">5B</xref>), when the only difference between two consecutive trials was the feedback received after the first trial (Figure <xref ref-type="fig" rid="F5">5A</xref>). Importantly, only <italic>during</italic> learning did this difference in N2pc amplitude following low vs. high value feedback correlate with individual participants&#x02019; learning rates&#x02014;participants in which the difference in feedback-dependent N2pc amplitude was particularly large learned faster (Figure <xref ref-type="fig" rid="F5">5D</xref>, left). This suggests that <italic>during</italic> learning, when participants needed to actively reevaluate their stimulus choices and relearn the current value contingency, allocating more attention to the choice stimulus after experiencing a negative outcome compared with a positive outcome, led to faster and more successful adjustment of behavior according to the current value contingency. In addition, this suggests that in a highly volatile task environment that requires continuous learning and value updating, with sources of uncertainty found in the inherent reward probability distributions (70%&#x02013;30%) and in the sudden value reversals (Yu and Dayan, <xref ref-type="bibr" rid="B69">2005</xref>; Payzan-LeNestour and Bossaerts, <xref ref-type="bibr" rid="B50">2011</xref>), participants allocate more attention to an uncertain compared with a more certain choice stimulus, in line with Pearce and Hall (<xref ref-type="bibr" rid="B51">1980</xref>). It is possible that in a less volatile environment in which participants have much longer periods of consistent choices, i.e., periods in which participants know the current value contingency with high certainty and learning presumably no longer takes place, we could have observed an increase in attention during target selection, which would be in line with Mackintosh (<xref ref-type="bibr" rid="B46">1975</xref>). Although we are using the N2pc amplitude as a proxy for selective attention, which may limit the implications to be drawn, our results are consistent with the idea that when tasks demand states of active learning, attention is increased following uncertain choice outcomes or events, and this correlates with enhanced learning performance.</p>
<p>Another prominent EEG component that has been shown to change during reversal learning is the FRN (e.g., Chase et al., <xref ref-type="bibr" rid="B9">2011</xref>; von Borries et al., <xref ref-type="bibr" rid="B63">2013</xref>; Donaldson et al., <xref ref-type="bibr" rid="B19">2016</xref>). The FRN has been thought to encode negative feedback, prediction error signals, outcome valence and behavioral adjustment (Holroyd and Coles, <xref ref-type="bibr" rid="B33">2002</xref>; Cohen and Ranganath, <xref ref-type="bibr" rid="B12">2007</xref>; Bellebaum and Daum, <xref ref-type="bibr" rid="B7">2008</xref>; Chase et al., <xref ref-type="bibr" rid="B9">2011</xref>; Walsh and Anderson, <xref ref-type="bibr" rid="B64">2011</xref>; von Borries et al., <xref ref-type="bibr" rid="B63">2013</xref>; Donaldson et al., <xref ref-type="bibr" rid="B19">2016</xref>). Using a probabilistic reversal learning paradigm similar to ours, Chase et al. (<xref ref-type="bibr" rid="B9">2011</xref>) have shown that the FRN amplitude scales with a negative prediction error signal obtained with a reinforcement learning model, whereby the FRN amplitude was largest following a reversal and diminished as a behavioral adjustment approached. Recent evidence using a reversal learning task, in which positive as well as negative outcomes could signal a need for behavioral adjustment and could be equally unexpected, suggests that the FRN may be more related to outcome valence (positive vs. negative) than to expectancy or behavioral adjustment (von Borries et al., <xref ref-type="bibr" rid="B63">2013</xref>; Donaldson et al., <xref ref-type="bibr" rid="B19">2016</xref>). We computed the FRN as the difference wave between the presentation of low-value and high-value feedback, and then compared the difference <italic>during</italic> learning with <italic>after</italic> learning (Figure <xref ref-type="fig" rid="F6">6B</xref>). We found that the FRN was substantially larger <italic>during</italic> learning than <italic>after</italic> learning. Although we should be careful of over-interpreting, since our task confounds the accumulation of negative outcomes with the need for behavioral adjustment, these results suggest that low-value as well as high-value feedback are processed differently during periods of uncertainty (<italic>during</italic> learning) when stimulus-value associations need to be updated (Figure <xref ref-type="fig" rid="F6">6A</xref>), which cannot be solely explained by differences in outcome valence.</p>
<p>In addition to the effect of learning on the FRN amplitudes, we also observed a more general and longer lasting effect of learning on feedback processing that was independent of feedback valence (Figure <xref ref-type="fig" rid="F6">6D</xref>). Feedback processing differed between 150 ms and 400 ms following feedback onset d<italic>uring</italic> learning of current value contingencies compared with <italic>after</italic> learning, which could indicate an increase in feedback processing <italic>per se</italic>. The FRN is thought to originate in anterior cingulate cortex (ACC; Hickey et al., <xref ref-type="bibr" rid="B27">2010</xref>), and this prolonged window of differential feedback processing matches the time-resolved activity level of ACC during reward presentation observed previously (Hickey et al., <xref ref-type="bibr" rid="B27">2010</xref>). In our task, enhanced ACC activity, during periods in which participants experience increased levels of uncertainty that potentially require a behavioral adjustment, could potentially reflect the necessity of increased levels of cognitive control (Shenhav et al., <xref ref-type="bibr" rid="B58">2013</xref>, <xref ref-type="bibr" rid="B59">2016</xref>), or increased activity related to the decision of moving from a state of exploitation to a state of exploration (Kolling et al., <xref ref-type="bibr" rid="B41">2016</xref>). Together with increased attention <italic>during</italic> learning and following negative feedback, these signals may be part of the underlying neural network that drives behavioral adjustment during periods of increased uncertainty that concludes in the switch of stimulus choice.</p>
<p>Overall, we found evidence that during periods of active behavioral adjustment in a changing and volatile task environment, feedback processing of recent choices and attentional processing is amplified and co-occurs with increases in attentional allocation following low-value feedback compared with high-value feedback that possibly promotes increased learning speed. Following successful learning of current value contingencies, during periods of stable behavior, attentional allocation then becomes potentially more efficient by suppressing non-relevant distractor processing. These results provide insight into how changes in attentional prioritization and feedback processing may support flexible and repeated behavioral adjustments in humans.</p>
</sec>
<sec id="s5">
<title>Author Contributions</title>
<p>MO, MRW, TW and AS designed the experiment. MO collected the data, performed the analyses and drafted the manuscript. All authors edited the manuscript.</p>
</sec>
<sec id="s6">
<title>Conflict of Interest Statement</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
</body>
<back>
<ack>
<p>This research was supported by the International Research Training Group IRTG 1901 &#x0201C;The Brain in Action&#x0201D; from the Natural Sciences and Engineering Research Council (NSERC) CREATE program (MO, MRW), the Canadian Institute of Health Research MOP-102482 (TW) and by the Deutsche Forschungsgemeinschaft (German Research Foundation, CRC/TRR 135, TP B3) (AS).</p>
</ack>
<ref-list>
<title>References</title>
<ref id="B1"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Anderson</surname> <given-names>B. A.</given-names></name></person-group> (<year>2013</year>). <article-title>A value-driven mechanism of attentional selection</article-title>. <source>J. Vis.</source> <volume>13</volume>:<fpage>7</fpage>. <pub-id pub-id-type="doi">10.1167/13.3.7</pub-id><pub-id pub-id-type="pmid">23589803</pub-id></citation></ref>
<ref id="B2"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Anderson</surname> <given-names>B. A.</given-names></name> <name><surname>Laurent</surname> <given-names>P. A.</given-names></name> <name><surname>Yantis</surname> <given-names>S.</given-names></name></person-group> (<year>2011a</year>). <article-title>Learned value magnifies salience-based attentional capture</article-title>. <source>PLoS One</source> <volume>6</volume>:<fpage>e27926</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pone.0027926</pub-id><pub-id pub-id-type="pmid">22132170</pub-id></citation></ref>
<ref id="B3"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Anderson</surname> <given-names>B. A.</given-names></name> <name><surname>Laurent</surname> <given-names>P. A.</given-names></name> <name><surname>Yantis</surname> <given-names>S.</given-names></name></person-group> (<year>2011b</year>). <article-title>Value-driven attentional capture</article-title>. <source>Proc. Natl. Acad. Sci. U S A</source> <volume>108</volume>, <fpage>10367</fpage>&#x02013;<lpage>10371</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.1104047108</pub-id><pub-id pub-id-type="pmid">21646524</pub-id></citation></ref>
<ref id="B4"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Anderson</surname> <given-names>B. A.</given-names></name> <name><surname>Laurent</surname> <given-names>P. A.</given-names></name> <name><surname>Yantis</surname> <given-names>S.</given-names></name></person-group> (<year>2013</year>). <article-title>Reward predictions bias attentional selection</article-title>. <source>Front. Hum. Neurosci.</source> <volume>7</volume>:<fpage>262</fpage>. <pub-id pub-id-type="doi">10.3389/fnhum.2013.00262</pub-id><pub-id pub-id-type="pmid">23781185</pub-id></citation></ref>
<ref id="B5"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Anderson</surname> <given-names>B. A.</given-names></name> <name><surname>Laurent</surname> <given-names>P. A.</given-names></name> <name><surname>Yantis</surname> <given-names>S.</given-names></name></person-group> (<year>2014</year>). <article-title>Value-driven attentional priority signals in human basal ganglia and visual cortex</article-title>. <source>Brain Res.</source> <volume>1587</volume>, <fpage>88</fpage>&#x02013;<lpage>96</lpage>. <pub-id pub-id-type="doi">10.1016/j.brainres.2014.08.062</pub-id><pub-id pub-id-type="pmid">25171805</pub-id></citation></ref>
<ref id="B6"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Awh</surname> <given-names>E.</given-names></name> <name><surname>Belopolsky</surname> <given-names>A. V.</given-names></name> <name><surname>Theeuwes</surname> <given-names>J.</given-names></name></person-group> (<year>2012</year>). <article-title>Top-down versus bottom-up attentional control: a failed theoretical dichotomy</article-title>. <source>Trends Cogn. Sci.</source> <volume>16</volume>, <fpage>437</fpage>&#x02013;<lpage>443</lpage>. <pub-id pub-id-type="doi">10.1016/j.tics.2012.06.010</pub-id><pub-id pub-id-type="pmid">22795563</pub-id></citation></ref>
<ref id="B7"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bellebaum</surname> <given-names>C.</given-names></name> <name><surname>Daum</surname> <given-names>I.</given-names></name></person-group> (<year>2008</year>). <article-title>Learning-related changes in reward expectancy are reflected in the feedback-related negativity</article-title>. <source>Eur. J. Neurosci.</source> <volume>27</volume>, <fpage>1823</fpage>&#x02013;<lpage>1835</lpage>. <pub-id pub-id-type="doi">10.1111/j.1460-9568.2008.06138.x</pub-id><pub-id pub-id-type="pmid">18380674</pub-id></citation></ref>
<ref id="B8"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bucker</surname> <given-names>B.</given-names></name> <name><surname>Theeuwes</surname> <given-names>J.</given-names></name></person-group> (<year>2017</year>). <article-title>Pavlovian reward learning underlies value driven attentional capture</article-title>. <source>Atten. Percept. Psychophys.</source> <volume>79</volume>, <fpage>415</fpage>&#x02013;<lpage>428</lpage>. <pub-id pub-id-type="doi">10.3758/s13414-016-1241-1</pub-id><pub-id pub-id-type="pmid">27905069</pub-id></citation></ref>
<ref id="B9"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chase</surname> <given-names>H. W.</given-names></name> <name><surname>Swainson</surname> <given-names>R.</given-names></name> <name><surname>Durham</surname> <given-names>L.</given-names></name> <name><surname>Benham</surname> <given-names>L.</given-names></name> <name><surname>Cools</surname> <given-names>R.</given-names></name></person-group> (<year>2011</year>). <article-title>Feedback-related negativity codes prediction error but not behavioral adjustment during probabilistic reversal learning</article-title>. <source>J. Cogn. Neurosci.</source> <volume>23</volume>, <fpage>936</fpage>&#x02013;<lpage>946</lpage>. <pub-id pub-id-type="doi">10.1162/jocn.2010.21456</pub-id><pub-id pub-id-type="pmid">20146610</pub-id></citation></ref>
<ref id="B10"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Clark</surname> <given-names>K.</given-names></name> <name><surname>Appelbaum</surname> <given-names>L. G.</given-names></name> <name><surname>van den Berg</surname> <given-names>B.</given-names></name> <name><surname>Mitroff</surname> <given-names>S. R.</given-names></name> <name><surname>Woldorff</surname> <given-names>M. G.</given-names></name></person-group> (<year>2015</year>). <article-title>Improvement in visual search with practice: mapping learning-related changes in neurocognitive stages of processing</article-title>. <source>J. Neurosci.</source> <volume>35</volume>, <fpage>5351</fpage>&#x02013;<lpage>5359</lpage>. <pub-id pub-id-type="doi">10.1523/JNEUROSCI.1152-14.2015</pub-id><pub-id pub-id-type="pmid">25834059</pub-id></citation></ref>
<ref id="B11"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cohen</surname> <given-names>M. X.</given-names></name> <name><surname>Elger</surname> <given-names>C. E.</given-names></name> <name><surname>Ranganath</surname> <given-names>C.</given-names></name></person-group> (<year>2007</year>). <article-title>Reward expectation modulates feedback-related negativity and EEG spectra</article-title>. <source>Neuroimage</source> <volume>35</volume>, <fpage>968</fpage>&#x02013;<lpage>978</lpage>. <pub-id pub-id-type="doi">10.1016/j.neuroimage.2006.11.056</pub-id><pub-id pub-id-type="pmid">17257860</pub-id></citation></ref>
<ref id="B12"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cohen</surname> <given-names>M. X.</given-names></name> <name><surname>Ranganath</surname> <given-names>C.</given-names></name></person-group> (<year>2007</year>). <article-title>Reinforcement learning signals predict future decisions</article-title>. <source>J. Neurosci.</source> <volume>27</volume>, <fpage>371</fpage>&#x02013;<lpage>378</lpage>. <pub-id pub-id-type="doi">10.1523/JNEUROSCI.4421-06.2007</pub-id><pub-id pub-id-type="pmid">17215398</pub-id></citation></ref>
<ref id="B13"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cools</surname> <given-names>R.</given-names></name> <name><surname>Clark</surname> <given-names>L.</given-names></name> <name><surname>Owen</surname> <given-names>A. M.</given-names></name> <name><surname>Robbins</surname> <given-names>T. W.</given-names></name></person-group> (<year>2002</year>). <article-title>Defining the neural mechanisms of probabilistic reversal learning using event-related functional magnetic resonance imaging</article-title>. <source>J. Neurosci.</source> <volume>22</volume>, <fpage>4563</fpage>&#x02013;<lpage>4567</lpage>. <pub-id pub-id-type="pmid">12040063</pub-id></citation></ref>
<ref id="B14"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Corbetta</surname> <given-names>M.</given-names></name> <name><surname>Shulman</surname> <given-names>G. L.</given-names></name></person-group> (<year>2002</year>). <article-title>Control of goal-directed and stimulus-driven attention in the brain</article-title>. <source>Nat. Rev. Neurosci.</source> <volume>3</volume>, <fpage>215</fpage>&#x02013;<lpage>229</lpage>. <pub-id pub-id-type="doi">10.1038/nrn755</pub-id><pub-id pub-id-type="pmid">11994752</pub-id></citation></ref>
<ref id="B15"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dayan</surname> <given-names>P.</given-names></name> <name><surname>Kakade</surname> <given-names>S.</given-names></name> <name><surname>Montague</surname> <given-names>P. R.</given-names></name></person-group> (<year>2000</year>). <article-title>Learning and selective attention</article-title>. <source>Nat. Neurosci.</source> <volume>3</volume>, <fpage>1218</fpage>&#x02013;<lpage>1223</lpage>. <pub-id pub-id-type="doi">10.1038/81504</pub-id><pub-id pub-id-type="pmid">11127841</pub-id></citation></ref>
<ref id="B16"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Della Libera</surname> <given-names>C.</given-names></name> <name><surname>Chelazzi</surname> <given-names>L.</given-names></name></person-group> (<year>2006</year>). <article-title>Visual selective attention and the effects of monetary rewards</article-title>. <source>Psychol. Sci.</source> <volume>17</volume>, <fpage>222</fpage>&#x02013;<lpage>227</lpage>. <pub-id pub-id-type="doi">10.1111/j.1467-9280.2006.01689.x</pub-id><pub-id pub-id-type="pmid">16507062</pub-id></citation></ref>
<ref id="B17"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Della Libera</surname> <given-names>C.</given-names></name> <name><surname>Chelazzi</surname> <given-names>L.</given-names></name></person-group> (<year>2009</year>). <article-title>Learning to attend and to ignore is a matter of gains and losses</article-title>. <source>Psychol. Sci.</source> <volume>20</volume>, <fpage>778</fpage>&#x02013;<lpage>784</lpage>. <pub-id pub-id-type="doi">10.1111/j.1467-9280.2009.02360.x</pub-id><pub-id pub-id-type="pmid">19422618</pub-id></citation></ref>
<ref id="B18"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Della Libera</surname> <given-names>C.</given-names></name> <name><surname>Perlato</surname> <given-names>A.</given-names></name> <name><surname>Chelazzi</surname> <given-names>L.</given-names></name></person-group> (<year>2011</year>). <article-title>Dissociable effects of reward on attentional learning: from passive associations to active monitoring</article-title>. <source>PLoS One</source> <volume>6</volume>:<fpage>e19460</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pone.0019460</pub-id><pub-id pub-id-type="pmid">21559388</pub-id></citation></ref>
<ref id="B19"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Donaldson</surname> <given-names>K. R.</given-names></name> <name><surname>Ait Oumeziane</surname> <given-names>B.</given-names></name> <name><surname>H&#x000E9;lie</surname> <given-names>S.</given-names></name> <name><surname>Foti</surname> <given-names>D.</given-names></name></person-group> (<year>2016</year>). <article-title>The temporal dynamics of reversal learning: P3 amplitude predicts valence-specific behavioral adjustment</article-title>. <source>Physiol. Behav.</source> <volume>161</volume>, <fpage>24</fpage>&#x02013;<lpage>32</lpage>. <pub-id pub-id-type="doi">10.1016/j.physbeh.2016.03.034</pub-id><pub-id pub-id-type="pmid">27059320</pub-id></citation></ref>
<ref id="B20"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Eimer</surname> <given-names>M.</given-names></name></person-group> (<year>2014</year>). <article-title>The neural basis of attentional control in visual search</article-title>. <source>Trends Cogn. Sci.</source> <volume>18</volume>, <fpage>526</fpage>&#x02013;<lpage>535</lpage>. <pub-id pub-id-type="doi">10.1016/j.tics.2014.05.005</pub-id><pub-id pub-id-type="pmid">24930047</pub-id></citation></ref>
<ref id="B21"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Eimer</surname> <given-names>M.</given-names></name> <name><surname>Grubert</surname> <given-names>A.</given-names></name></person-group> (<year>2014</year>). <article-title>Spatial attention can be allocated rapidly and in parallel to new visual objects</article-title>. <source>Curr. Biol.</source> <volume>24</volume>, <fpage>193</fpage>&#x02013;<lpage>198</lpage>. <pub-id pub-id-type="doi">10.1016/j.cub.2013.12.001</pub-id><pub-id pub-id-type="pmid">24412208</pub-id></citation></ref>
<ref id="B22"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Feldmann-W&#x000FC;stefeld</surname> <given-names>T.</given-names></name> <name><surname>Brandhofer</surname> <given-names>R.</given-names></name> <name><surname>Schub&#x000F6;</surname> <given-names>A.</given-names></name></person-group> (<year>2016</year>). <article-title>Rewarded visual items capture attention only in heterogeneous contexts</article-title>. <source>Psychophysiology</source> <volume>53</volume>, <fpage>1063</fpage>&#x02013;<lpage>1073</lpage>. <pub-id pub-id-type="doi">10.1111/psyp.12641</pub-id><pub-id pub-id-type="pmid">26997364</pub-id></citation></ref>
<ref id="B23"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Feldmann-W&#x000FC;stefeld</surname> <given-names>T.</given-names></name> <name><surname>Schub&#x000F6;</surname> <given-names>A.</given-names></name></person-group> (<year>2013</year>). <article-title>Context homogeneity facilitates both distractor inhibition and target enhancement</article-title>. <source>J. Vis.</source> <volume>13</volume>:<fpage>11</fpage>. <pub-id pub-id-type="doi">10.1167/13.3.11</pub-id><pub-id pub-id-type="pmid">23650629</pub-id></citation></ref>
<ref id="B24"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Feldmann-W&#x000FC;stefeld</surname> <given-names>T.</given-names></name> <name><surname>Uengoer</surname> <given-names>M.</given-names></name> <name><surname>Schub&#x000F6;</surname> <given-names>A.</given-names></name></person-group> (<year>2015</year>). <article-title>You see what you have learned</article-title>. <source>Psychophysiology</source> <volume>52</volume>, <fpage>1483</fpage>&#x02013;<lpage>1497</lpage>. <pub-id pub-id-type="doi">10.1111/psyp.12514</pub-id><pub-id pub-id-type="pmid">26338030</pub-id></citation></ref>
<ref id="B25"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gottlieb</surname> <given-names>J.</given-names></name></person-group> (<year>2012</year>). <article-title>Attention, learning, and the value of information</article-title>. <source>Neuron</source> <volume>76</volume>, <fpage>281</fpage>&#x02013;<lpage>295</lpage>. <pub-id pub-id-type="doi">10.1016/j.neuron.2012.09.034</pub-id><pub-id pub-id-type="pmid">23083732</pub-id></citation></ref>
<ref id="B26"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hajcak</surname> <given-names>G.</given-names></name> <name><surname>Moser</surname> <given-names>J. S.</given-names></name> <name><surname>Holroyd</surname> <given-names>C. B.</given-names></name> <name><surname>Simons</surname> <given-names>R. F.</given-names></name></person-group> (<year>2006</year>). <article-title>The feedback-related negativity reflects the binary evaluation of good versus bad outcomes</article-title>. <source>Biol. Psychol.</source> <volume>71</volume>, <fpage>148</fpage>&#x02013;<lpage>154</lpage>. <pub-id pub-id-type="doi">10.1016/j.biopsycho.2005.04.001</pub-id><pub-id pub-id-type="pmid">16005561</pub-id></citation></ref>
<ref id="B27"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hickey</surname> <given-names>C.</given-names></name> <name><surname>Chelazzi</surname> <given-names>L.</given-names></name> <name><surname>Theeuwes</surname> <given-names>J.</given-names></name></person-group> (<year>2010</year>). <article-title>Reward changes salience in human vision via the anterior cingulate</article-title>. <source>J. Neurosci.</source> <volume>30</volume>, <fpage>11096</fpage>&#x02013;<lpage>11103</lpage>. <pub-id pub-id-type="doi">10.1523/JNEUROSCI.1026-10.2010</pub-id><pub-id pub-id-type="pmid">20720117</pub-id></citation></ref>
<ref id="B28"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hickey</surname> <given-names>C.</given-names></name> <name><surname>Chelazzi</surname> <given-names>L.</given-names></name> <name><surname>Theeuwes</surname> <given-names>J.</given-names></name></person-group> (<year>2011</year>). <article-title>Reward has a residual impact on target selection in visual search, but not on the suppression of distractors</article-title>. <source>Vis. Cogn.</source> <volume>19</volume>, <fpage>117</fpage>&#x02013;<lpage>128</lpage>. <pub-id pub-id-type="doi">10.1080/13506285.2010.503946</pub-id></citation></ref>
<ref id="B29"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hickey</surname> <given-names>C.</given-names></name> <name><surname>Di Lollo</surname> <given-names>V.</given-names></name> <name><surname>McDonald</surname> <given-names>J. J.</given-names></name></person-group> (<year>2009</year>). <article-title>Electrophysiological indices of target and distractor processing in visual search</article-title>. <source>J. Cogn. Neurosci.</source> <volume>21</volume>, <fpage>760</fpage>&#x02013;<lpage>775</lpage>. <pub-id pub-id-type="doi">10.1162/jocn.2009.21039</pub-id><pub-id pub-id-type="pmid">18564048</pub-id></citation></ref>
<ref id="B30"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hickey</surname> <given-names>C.</given-names></name> <name><surname>Kaiser</surname> <given-names>D.</given-names></name> <name><surname>Peelen</surname> <given-names>M. V.</given-names></name></person-group> (<year>2015</year>). <article-title>Reward guides attention to object categories in real-world scenes</article-title>. <source>J. Exp. Psychol. Gen.</source> <volume>144</volume>, <fpage>264</fpage>&#x02013;<lpage>273</lpage>. <pub-id pub-id-type="doi">10.1037/a0038627</pub-id><pub-id pub-id-type="pmid">25559653</pub-id></citation></ref>
<ref id="B31"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Hogarth</surname> <given-names>L.</given-names></name> <name><surname>Dickinson</surname> <given-names>A.</given-names></name> <name><surname>Duka</surname> <given-names>T.</given-names></name></person-group> (<year>2010</year>). &#x0201C;<article-title>Selective attention to conditioned stimuli in human discrimination learning: untangling the effects of outcome prediction, valence, arousal, and uncertainty</article-title>,&#x0201D; in <source>Attention and Associative Learning. From Brain to Behavior</source>, eds <person-group person-group-type="editor"><name><surname>Mitchell</surname> <given-names>C. J.</given-names></name> <name><surname>Le Pelley</surname> <given-names>M. E.</given-names></name></person-group> (<publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>), <fpage>71</fpage>&#x02013;<lpage>97</lpage>.</citation></ref>
<ref id="B32"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Holland</surname> <given-names>P. C.</given-names></name> <name><surname>Gallagher</surname> <given-names>M.</given-names></name></person-group> (<year>1999</year>). <article-title>Amygdala circuitry in attentional and representational processes</article-title>. <source>Trends Cogn. Sci.</source> <volume>3</volume>, <fpage>65</fpage>&#x02013;<lpage>73</lpage>. <pub-id pub-id-type="doi">10.1016/s1364-6613(98)01271-6</pub-id><pub-id pub-id-type="pmid">10234229</pub-id></citation></ref>
<ref id="B33"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Holroyd</surname> <given-names>C.</given-names></name> <name><surname>Coles</surname> <given-names>M.</given-names></name></person-group> (<year>2002</year>). <article-title>The neural basis of human error processing: reinforcement learning, dopamine, and the error-related negativity</article-title>. <source>Psychol. Rev.</source> <volume>109</volume>, <fpage>679</fpage>&#x02013;<lpage>709</lpage>. <pub-id pub-id-type="doi">10.1037/0033-295X.109.4.679</pub-id><pub-id pub-id-type="pmid">12374324</pub-id></citation></ref>
<ref id="B34"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hopf</surname> <given-names>J. M.</given-names></name> <name><surname>Luck</surname> <given-names>S. J.</given-names></name> <name><surname>Boelmans</surname> <given-names>K.</given-names></name> <name><surname>Schoenfeld</surname> <given-names>M. A.</given-names></name> <name><surname>Boehler</surname> <given-names>C. N.</given-names></name> <name><surname>Rieger</surname> <given-names>J.</given-names></name> <etal/></person-group>. (<year>2006</year>). <article-title>The neural site of attention matches the spatial scale of perception</article-title>. <source>J. Neurosci.</source> <volume>26</volume>, <fpage>3532</fpage>&#x02013;<lpage>3540</lpage>. <pub-id pub-id-type="doi">10.1523/JNEUROSCI.4510-05.2006</pub-id><pub-id pub-id-type="pmid">16571761</pub-id></citation></ref>
<ref id="B35"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hopf</surname> <given-names>J. M.</given-names></name> <name><surname>Luck</surname> <given-names>S. J.</given-names></name> <name><surname>Girelli</surname> <given-names>M.</given-names></name> <name><surname>Hagner</surname> <given-names>T.</given-names></name> <name><surname>Mangun</surname> <given-names>G. R.</given-names></name> <name><surname>Scheich</surname> <given-names>H.</given-names></name> <etal/></person-group>. (<year>2000</year>). <article-title>Neural sources of focused attention in visual search</article-title>. <source>Cereb. Cortex</source> <volume>10</volume>, <fpage>1233</fpage>&#x02013;<lpage>1241</lpage>. <pub-id pub-id-type="doi">10.1093/cercor/10.12.1233</pub-id><pub-id pub-id-type="pmid">11073872</pub-id></citation></ref>
<ref id="B36"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Irons</surname> <given-names>J. L.</given-names></name> <name><surname>Leber</surname> <given-names>A. B.</given-names></name></person-group> (<year>2016</year>). <article-title>Choosing attentional control settings in a dynamically changing environment</article-title>. <source>Atten. Percept. Psychophys.</source> <volume>78</volume>, <fpage>2031</fpage>&#x02013;<lpage>2048</lpage>. <pub-id pub-id-type="doi">10.3758/s13414-016-1125-4</pub-id><pub-id pub-id-type="pmid">27188652</pub-id></citation></ref>
<ref id="B37"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Itthipuripat</surname> <given-names>S.</given-names></name> <name><surname>Cha</surname> <given-names>K.</given-names></name> <name><surname>Byers</surname> <given-names>A.</given-names></name> <name><surname>Serences</surname> <given-names>J. T.</given-names></name></person-group> (<year>2017</year>). <article-title>Two different mechanisms support selective attention at different phases of training</article-title>. <source>PLoS Biol.</source> <volume>15</volume>:<fpage>e2001724</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pbio.2001724</pub-id><pub-id pub-id-type="pmid">28654635</pub-id></citation></ref>
<ref id="B38"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Itthipuripat</surname> <given-names>S.</given-names></name> <name><surname>Cha</surname> <given-names>K.</given-names></name> <name><surname>Rangsipat</surname> <given-names>N.</given-names></name> <name><surname>Serences</surname> <given-names>J. T.</given-names></name></person-group> (<year>2015</year>). <article-title>Value-based attentional capture influences context-dependent decision-making</article-title>. <source>J. Neurophysiol.</source> <volume>114</volume>, <fpage>560</fpage>&#x02013;<lpage>569</lpage>. <pub-id pub-id-type="doi">10.1152/jn.00343.2015</pub-id><pub-id pub-id-type="pmid">25995350</pub-id></citation></ref>
<ref id="B39"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kastner</surname> <given-names>S.</given-names></name> <name><surname>Ungerleider</surname> <given-names>L. G.</given-names></name></person-group> (<year>2000</year>). <article-title>Mechanisms of visual attention in the human cortex</article-title>. <source>Annu. Rev. Neurosci.</source> <volume>23</volume>, <fpage>315</fpage>&#x02013;<lpage>341</lpage>. <pub-id pub-id-type="doi">10.1146/annurev.neuro.23.1.315</pub-id><pub-id pub-id-type="pmid">10845067</pub-id></citation></ref>
<ref id="B40"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kiss</surname> <given-names>M.</given-names></name> <name><surname>Driver</surname> <given-names>J.</given-names></name> <name><surname>Eimer</surname> <given-names>M.</given-names></name></person-group> (<year>2009</year>). <article-title>Reward priority of visual target singletons modulates ERP signatures of attentional selection</article-title>. <source>Psychol. Sci.</source> <volume>20</volume>, <fpage>245</fpage>&#x02013;<lpage>251</lpage>. <pub-id pub-id-type="doi">10.1111/j.1467-9280.2009.02281.x</pub-id><pub-id pub-id-type="pmid">19175756</pub-id></citation></ref>
<ref id="B93"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kiss</surname> <given-names>M.</given-names></name> <name><surname>Van Velzen</surname> <given-names>J.</given-names></name> <name><surname>Eimer</surname> <given-names>M.</given-names></name></person-group> (<year>2008</year>). <article-title>The N2pc component and its links to attention shifts and spatially selective visual processing</article-title>. <source>Psychophysiology</source> <volume>45</volume>, <fpage>240</fpage>&#x02013;<lpage>249</lpage>. <pub-id pub-id-type="doi">10.1111/j.1469-8986.2007.00611.x</pub-id><pub-id pub-id-type="pmid">17971061 </pub-id></citation></ref>
<ref id="B41"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kolling</surname> <given-names>N.</given-names></name> <name><surname>Wittmann</surname> <given-names>M. K.</given-names></name> <name><surname>Behrens</surname> <given-names>T. E. J.</given-names></name> <name><surname>Boorman</surname> <given-names>E. D.</given-names></name> <name><surname>Mars</surname> <given-names>R. B.</given-names></name> <name><surname>Rushworth</surname> <given-names>M. F. S.</given-names></name></person-group> (<year>2016</year>). <article-title>Value, search, persistence and model updating in anterior cingulate cortex</article-title>. <source>Nat. Neurosci.</source> <volume>19</volume>, <fpage>1280</fpage>&#x02013;<lpage>1285</lpage>. <pub-id pub-id-type="doi">10.1038/nn.4382</pub-id><pub-id pub-id-type="pmid">27669988</pub-id></citation></ref>
<ref id="B42"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Krebs</surname> <given-names>R. M.</given-names></name> <name><surname>Boehler</surname> <given-names>C. N.</given-names></name> <name><surname>Woldorff</surname> <given-names>M. G.</given-names></name></person-group> (<year>2010</year>). <article-title>The influence of reward associations on conflict processing in the Stroop task</article-title>. <source>Cognition</source> <volume>117</volume>, <fpage>341</fpage>&#x02013;<lpage>347</lpage>. <pub-id pub-id-type="doi">10.1016/j.cognition.2010.08.018</pub-id><pub-id pub-id-type="pmid">20864094</pub-id></citation></ref>
<ref id="B43"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Le Pelley</surname> <given-names>M. E.</given-names></name></person-group> (<year>2004</year>). <article-title>The role of associative history in models of associative learning: a selective review and a hybrid model</article-title>. <source>Q. J. Exp. Psychol. B</source> <volume>57</volume>, <fpage>193</fpage>&#x02013;<lpage>243</lpage>. <pub-id pub-id-type="doi">10.1080/02724990344000141</pub-id><pub-id pub-id-type="pmid">15204108</pub-id></citation></ref>
<ref id="B44"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Le Pelley</surname> <given-names>M. E.</given-names></name> <name><surname>Vadillo</surname> <given-names>M.</given-names></name> <name><surname>Luque</surname> <given-names>D.</given-names></name></person-group> (<year>2013</year>). <article-title>Learned predictiveness influences rapid attentional capture: evidence from the dot probe task</article-title>. <source>J. Exp. Psychol. Learn. Mem. Cogn.</source> <volume>39</volume>, <fpage>1888</fpage>&#x02013;<lpage>1900</lpage>. <pub-id pub-id-type="doi">10.1037/a0033700</pub-id><pub-id pub-id-type="pmid">23855549</pub-id></citation></ref>
<ref id="B91"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Luck</surname> <given-names>S. J.</given-names></name> <name><surname>Hillyard</surname> <given-names>S. A.</given-names></name></person-group> (<year>1994a</year>). <article-title>Electrophysiological correlates of feature analysis during visual search</article-title>. <source>Psychophysiology</source> <volume>31</volume>, <fpage>291</fpage>&#x02013;<lpage>308</lpage>. <pub-id pub-id-type="doi">10.1111/j.1469-8986.1994.tb02218.x</pub-id><pub-id pub-id-type="pmid">8008793</pub-id></citation></ref>
<ref id="B92"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Luck</surname> <given-names>S. J.</given-names></name> <name><surname>Hillyard</surname> <given-names>S. A.</given-names></name></person-group> (<year>1994b</year>). <article-title>Spatial filtering during visual search: evidence from human electrophysiology</article-title>. <source>J. Exp. Psychol. Hum. Percept. Perform.</source> <volume>20</volume>, <fpage>1000</fpage>&#x02013;<lpage>1014</lpage>. <pub-id pub-id-type="doi">10.1037/0096-1523.20.5.1000</pub-id><pub-id pub-id-type="pmid">7964526</pub-id></citation></ref>
<ref id="B46"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mackintosh</surname> <given-names>N. J.</given-names></name></person-group> (<year>1975</year>). <article-title>A theory of attention: variations in the associability of stimuli with reinforcement</article-title>. <source>Psychol. Rev.</source> <volume>82</volume>, <fpage>276</fpage>&#x02013;<lpage>298</lpage>. <pub-id pub-id-type="doi">10.1037/h0076778</pub-id></citation></ref>
<ref id="B47"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mackintosh</surname> <given-names>N. J.</given-names></name> <name><surname>Little</surname> <given-names>L.</given-names></name></person-group> (<year>1969</year>). <article-title>Intradimensional and extradimensional shift learning by pigeons</article-title>. <source>Psychon. Sci.</source> <volume>14</volume>, <fpage>5</fpage>&#x02013;<lpage>6</lpage>. <pub-id pub-id-type="doi">10.3758/bf03336395</pub-id></citation></ref>
<ref id="B48"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Miltner</surname> <given-names>W. H. R.</given-names></name> <name><surname>Braun</surname> <given-names>C. H.</given-names></name> <name><surname>Coles</surname> <given-names>M. G. H.</given-names></name></person-group> (<year>1997</year>). <article-title>Event-related brain potentials following incorrect feedback in a time-estimation task: evidence for a &#x0201C;generic&#x0201D; neural system for error detection</article-title>. <source>J. Cogn. Neurosci.</source> <volume>9</volume>, <fpage>788</fpage>&#x02013;<lpage>798</lpage>. <pub-id pub-id-type="doi">10.1162/jocn.1997.9.6.788</pub-id><pub-id pub-id-type="pmid">23964600</pub-id></citation></ref>
<ref id="B90"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Munneke</surname> <given-names>J.</given-names></name> <name><surname>Hoppenbrouwers</surname> <given-names>S. S.</given-names></name> <name><surname>Theeuwes</surname> <given-names>J.</given-names></name></person-group> (<year>2015</year>). <article-title>Reward can modulate attentional capture, independent of top-down set</article-title>. <source>Atten. Percept. Psychophys.</source> <volume>77</volume>, <fpage>2540</fpage>&#x02013;<lpage>2548</lpage>. <pub-id pub-id-type="doi">10.3758/s13414-015-0958-6</pub-id><pub-id pub-id-type="pmid">26178858 </pub-id></citation></ref>
<ref id="B49"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Navalpakkam</surname> <given-names>V.</given-names></name> <name><surname>Koch</surname> <given-names>C.</given-names></name> <name><surname>Rangel</surname> <given-names>A.</given-names></name> <name><surname>Perona</surname> <given-names>P.</given-names></name></person-group> (<year>2010</year>). <article-title>Optimal reward harvesting in complex perceptual environments</article-title>. <source>Proc. Natl. Acad. Sci. U S A</source> <volume>107</volume>, <fpage>5232</fpage>&#x02013;<lpage>5237</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.0911972107</pub-id><pub-id pub-id-type="pmid">20194768</pub-id></citation></ref>
<ref id="B50"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Payzan-LeNestour</surname> <given-names>E.</given-names></name> <name><surname>Bossaerts</surname> <given-names>P.</given-names></name></person-group> (<year>2011</year>). <article-title>Risk, unexpected uncertainty and estimation uncertainty: bayesian learning in unstable settings</article-title>. <source>PLoS Comput. Biol.</source> <volume>7</volume>:<fpage>e1001048</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pcbi.1001048</pub-id><pub-id pub-id-type="pmid">21283774</pub-id></citation></ref>
<ref id="B51"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Pearce</surname> <given-names>J.</given-names></name> <name><surname>Hall</surname> <given-names>G.</given-names></name></person-group> (<year>1980</year>). <article-title>A model for Pavlovian learning: variation in the effectiveness of conditioned but not unconditioned stimuli</article-title>. <source>Psychol. Rev.</source> <volume>87</volume>, <fpage>532</fpage>&#x02013;<lpage>552</lpage>. <pub-id pub-id-type="doi">10.1037/0033-295x.87.6.532</pub-id><pub-id pub-id-type="pmid">7443916</pub-id></citation></ref>
<ref id="B52"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Pearce</surname> <given-names>J.</given-names></name> <name><surname>Mackintosh</surname> <given-names>N. J.</given-names></name></person-group> (<year>2010</year>). &#x0201C;<article-title>Two theories of attention: a review and possible integration</article-title>,&#x0201D; in <source>Attention and Associative Learning. From Brain to Behavior</source>, eds <person-group person-group-type="editor"><name><surname>Mitchell</surname> <given-names>C. J.</given-names></name> <name><surname>Le Pelley</surname> <given-names>M. E.</given-names></name></person-group> (<publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>), <fpage>11</fpage>&#x02013;<lpage>39</lpage>.</citation></ref>
<ref id="B53"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Posner</surname> <given-names>M.</given-names></name> <name><surname>Petersen</surname> <given-names>S. E.</given-names></name></person-group> (<year>1990</year>). <article-title>The attention system of the human brain</article-title>. <source>Annu. Rev. Neurosci.</source> <volume>13</volume>, <fpage>25</fpage>&#x02013;<lpage>42</lpage>. <pub-id pub-id-type="doi">10.1146/annurev.ne.13.030190.000325</pub-id><pub-id pub-id-type="pmid">2183676</pub-id></citation></ref>
<ref id="B54"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Raymond</surname> <given-names>J. E.</given-names></name> <name><surname>O&#x02019;Brien</surname> <given-names>J. L.</given-names></name></person-group> (<year>2009</year>). <article-title>Selective visual attention and motivation: the consequences of value learning in an attentional blink task</article-title>. <source>Psychol. Sci.</source> <volume>20</volume>, <fpage>981</fpage>&#x02013;<lpage>989</lpage>. <pub-id pub-id-type="doi">10.1111/j.1467-9280.2009.02391.x</pub-id><pub-id pub-id-type="pmid">19549080</pub-id></citation></ref>
<ref id="B55"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sali</surname> <given-names>A. W.</given-names></name> <name><surname>Anderson</surname> <given-names>B. A.</given-names></name> <name><surname>Yantis</surname> <given-names>S.</given-names></name></person-group> (<year>2014</year>). <article-title>The role of reward prediction in the control of attention</article-title>. <source>J. Exp. Psychol. Hum. Percept. Perform.</source> <volume>40</volume>, <fpage>1654</fpage>&#x02013;<lpage>1664</lpage>. <pub-id pub-id-type="doi">10.1037/a0037267</pub-id><pub-id pub-id-type="pmid">24955700</pub-id></citation></ref>
<ref id="B56"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>San Mart&#x000ED;n</surname> <given-names>R.</given-names></name> <name><surname>Appelbaum</surname> <given-names>L. G.</given-names></name> <name><surname>Huettel</surname> <given-names>S. A.</given-names></name> <name><surname>Woldorff</surname> <given-names>M. G.</given-names></name></person-group> (<year>2016</year>). <article-title>Cortical brain activity reflecting attentional biasing toward reward-predicting cues covaries with economic decision-making performance</article-title>. <source>Cereb. Cortex</source> <volume>26</volume>, <fpage>1</fpage>&#x02013;<lpage>11</lpage>. <pub-id pub-id-type="doi">10.1093/cercor/bhu160</pub-id><pub-id pub-id-type="pmid">25139941</pub-id></citation></ref>
<ref id="B57"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sawaki</surname> <given-names>R.</given-names></name> <name><surname>Luck</surname> <given-names>S. J.</given-names></name> <name><surname>Raymond</surname> <given-names>J. E.</given-names></name></person-group> (<year>2015</year>). <article-title>How attention changes in response to incentives</article-title>. <source>J. Cogn. Neurosci.</source> <volume>27</volume>, <fpage>2229</fpage>&#x02013;<lpage>2239</lpage>. <pub-id pub-id-type="doi">10.1162/jocn_a_00847</pub-id><pub-id pub-id-type="pmid">26151604</pub-id></citation></ref>
<ref id="B58"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Shenhav</surname> <given-names>A.</given-names></name> <name><surname>Botvinick</surname> <given-names>M. M.</given-names></name> <name><surname>Cohen</surname> <given-names>J. D.</given-names></name></person-group> (<year>2013</year>). <article-title>The expected value of control: an integrative theory of anterior cingulate cortex function</article-title>. <source>Neuron</source> <volume>79</volume>, <fpage>217</fpage>&#x02013;<lpage>240</lpage>. <pub-id pub-id-type="doi">10.1016/j.neuron.2013.07.007</pub-id><pub-id pub-id-type="pmid">23889930</pub-id></citation></ref>
<ref id="B59"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Shenhav</surname> <given-names>A.</given-names></name> <name><surname>Cohen</surname> <given-names>J. D.</given-names></name> <name><surname>Botvinick</surname> <given-names>M. M.</given-names></name></person-group> (<year>2016</year>). <article-title>Dorsal anterior cingulate cortex and the value of control</article-title>. <source>Nat. Neurosci.</source> <volume>19</volume>, <fpage>1286</fpage>&#x02013;<lpage>1291</lpage>. <pub-id pub-id-type="doi">10.1038/nn.4384</pub-id><pub-id pub-id-type="pmid">27669989</pub-id></citation></ref>
<ref id="B60"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Smith</surname> <given-names>A. C.</given-names></name> <name><surname>Brown</surname> <given-names>E. N.</given-names></name></person-group> (<year>2003</year>). <article-title>Estimating a state-space model from point process observations</article-title>. <source>Neural Comput.</source> <volume>15</volume>, <fpage>965</fpage>&#x02013;<lpage>991</lpage>. <pub-id pub-id-type="doi">10.1162/089976603765202622</pub-id><pub-id pub-id-type="pmid">12803953</pub-id></citation></ref>
<ref id="B61"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Smith</surname> <given-names>A. C.</given-names></name> <name><surname>Frank</surname> <given-names>L. M.</given-names></name> <name><surname>Wirth</surname> <given-names>S.</given-names></name> <name><surname>Yanike</surname> <given-names>M.</given-names></name> <name><surname>Hu</surname> <given-names>D.</given-names></name> <name><surname>Kubota</surname> <given-names>Y.</given-names></name> <etal/></person-group>. (<year>2004</year>). <article-title>Dynamic analysis of learning in behavioral experiments</article-title>. <source>J. Neurosci.</source> <volume>24</volume>, <fpage>447</fpage>&#x02013;<lpage>461</lpage>. <pub-id pub-id-type="doi">10.1523/JNEUROSCI.2908-03.2004</pub-id><pub-id pub-id-type="pmid">14724243</pub-id></citation></ref>
<ref id="B62"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Steiger</surname> <given-names>J. H.</given-names></name></person-group> (<year>1980</year>). <article-title>Tests for comparing elements of a correlation matrix</article-title>. <source>Psychol. Bull.</source> <volume>87</volume>, <fpage>245</fpage>&#x02013;<lpage>251</lpage>. <pub-id pub-id-type="doi">10.1037/0033-2909.87.2.245</pub-id></citation></ref>
<ref id="B63"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>von Borries</surname> <given-names>A. K. L.</given-names></name> <name><surname>Verkes</surname> <given-names>R. J.</given-names></name> <name><surname>Bulten</surname> <given-names>B. H.</given-names></name> <name><surname>Cools</surname> <given-names>R.</given-names></name> <name><surname>de Bruijn</surname> <given-names>E. R. A.</given-names></name></person-group> (<year>2013</year>). <article-title>Feedback-related negativity codes outcome valence, but not outcome expectancy, during reversal learning</article-title>. <source>Cogn. Affect. Behav. Neurosci.</source> <volume>13</volume>, <fpage>737</fpage>&#x02013;<lpage>746</lpage>. <pub-id pub-id-type="doi">10.3758/s13415-013-0150-1</pub-id><pub-id pub-id-type="pmid">24146314</pub-id></citation></ref>
<ref id="B64"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Walsh</surname> <given-names>M. M.</given-names></name> <name><surname>Anderson</surname> <given-names>J. R.</given-names></name></person-group> (<year>2011</year>). <article-title>Modulation of the feedback-related negativity by instruction and experience</article-title>. <source>Proc. Natl. Acad. Sci. U S A</source> <volume>208</volume>, <fpage>19048</fpage>&#x02013;<lpage>19053</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.1117189108</pub-id><pub-id pub-id-type="pmid">22065792</pub-id></citation></ref>
<ref id="B65"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Walsh</surname> <given-names>M. M.</given-names></name> <name><surname>Anderson</surname> <given-names>J. R.</given-names></name></person-group> (<year>2012</year>). <article-title>Learning from experience: event-related potential correlates of reward processing, neural adaptation, and behavioral choice</article-title>. <source>Neurosci. Biobehav. Rev.</source> <volume>36</volume>, <fpage>1870</fpage>&#x02013;<lpage>1884</lpage>. <pub-id pub-id-type="doi">10.1016/j.neubiorev.2012.05.008</pub-id><pub-id pub-id-type="pmid">22683741</pub-id></citation></ref>
<ref id="B66"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wilson</surname> <given-names>P. N.</given-names></name> <name><surname>Boumphrey</surname> <given-names>P.</given-names></name> <name><surname>Pearce</surname> <given-names>J. M.</given-names></name></person-group> (<year>1992</year>). <article-title>Restoration of the orienting response to a light by a change in its predictive accuracy</article-title>. <source>J. Exp. Psychol.</source> <volume>44</volume>, <fpage>17</fpage>&#x02013;<lpage>36</lpage>.</citation></ref>
<ref id="B67"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Womelsdorf</surname> <given-names>T.</given-names></name> <name><surname>Everling</surname> <given-names>S.</given-names></name></person-group> (<year>2015</year>). <article-title>Long-range attention networks: circuit motifs underlying endogenously controlled stimulus selection</article-title>. <source>Trends Neurosci.</source> <volume>38</volume>, <fpage>682</fpage>&#x02013;<lpage>700</lpage>. <pub-id pub-id-type="doi">10.1016/j.tins.2015.08.009</pub-id><pub-id pub-id-type="pmid">26549883</pub-id></citation></ref>
<ref id="B94"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Woodman</surname> <given-names>G. F.</given-names></name> <name><surname>Luck</surname> <given-names>S. J.</given-names></name></person-group> (<year>1999</year>). <article-title>Electrophysiological measurement of rapid shifts of attention during visual search</article-title>. <source>Nature</source> <volume>400</volume>, <fpage>867</fpage>&#x02013;<lpage>869</lpage>. <pub-id pub-id-type="doi">10.1038/23698</pub-id><pub-id pub-id-type="pmid">10476964</pub-id></citation></ref>
<ref id="B68"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Woodman</surname> <given-names>G. F.</given-names></name> <name><surname>Luck</surname> <given-names>S. J.</given-names></name></person-group> (<year>2003</year>). <article-title>Serial deployment of attention during visual search</article-title>. <source>J. Exp. Psychol. Hum. Percept. Perform.</source> <volume>29</volume>, <fpage>121</fpage>&#x02013;<lpage>138</lpage>. <pub-id pub-id-type="doi">10.1037/0096-1523.29.1.121</pub-id><pub-id pub-id-type="pmid">12669752</pub-id></citation></ref>
<ref id="B69"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Yu</surname> <given-names>A. J.</given-names></name> <name><surname>Dayan</surname> <given-names>P.</given-names></name></person-group> (<year>2005</year>). <article-title>Uncertainty, neuromodulation, and attention</article-title>. <source>Neuron</source> <volume>46</volume>, <fpage>681</fpage>&#x02013;<lpage>692</lpage>. <pub-id pub-id-type="doi">10.1016/j.neuron.2005.04.026</pub-id><pub-id pub-id-type="pmid">15944135</pub-id></citation></ref>
</ref-list>
<fn-group>
<fn id="fn0001"><p><sup>1</sup><ext-link ext-link-type="uri" xlink:href="http://www.ru.nl/fcdonders/fieldtrip/">http://www.ru.nl/fcdonders/fieldtrip/</ext-link></p></fn>
</fn-group>
</back>
</article>