<?xml version="1.0" encoding="UTF-8" standalone="no"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xml:lang="EN" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Psychol.</journal-id>
<journal-title>Frontiers in Psychology</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Psychol.</abbrev-journal-title>
<issn pub-type="epub">1664-1078</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fpsyg.2022.875744</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Psychology</subject>
<subj-group>
<subject>Hypothesis and Theory</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Linguistic Skill and Stimulus-Driven Attention: A Case for Linguistic Relativity</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name><surname>Ansorge</surname> <given-names>Ulrich</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="aff" rid="aff2"><sup>2</sup></xref>
<xref ref-type="aff" rid="aff3"><sup>3</sup></xref>
<xref ref-type="corresp" rid="c001"><sup>&#x002A;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/10564/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Baier</surname> <given-names>Diane</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="aff" rid="aff4"><sup>4</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/954496/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Choi</surname> <given-names>Soonja</given-names></name>
<xref ref-type="aff" rid="aff5"><sup>5</sup></xref>
<xref ref-type="aff" rid="aff6"><sup>6</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/616225/overview"/>
</contrib>
</contrib-group>
<aff id="aff1"><sup>1</sup><institution>Faculty of Psychology, University of Vienna</institution>, <addr-line>Vienna</addr-line>, <country>Austria</country></aff>
<aff id="aff2"><sup>2</sup><institution>Cognitive Science Hub, University of Vienna</institution>, <addr-line>Vienna</addr-line>, <country>Austria</country></aff>
<aff id="aff3"><sup>3</sup><institution>Research Platform Mediatised Lifeworlds, University of Vienna</institution>, <addr-line>Vienna</addr-line>, <country>Austria</country></aff>
<aff id="aff4"><sup>4</sup><institution>Acoustics Research Institute, Austrian Academy of Sciences</institution>, <addr-line>Vienna</addr-line>, <country>Austria</country></aff>
<aff id="aff5"><sup>5</sup><institution>Department of Linguistics and Asian/Middle Eastern Languages, San Diego State University</institution>, <addr-line>San Diego, CA</addr-line>, <country>United States</country></aff>
<aff id="aff6"><sup>6</sup><institution>Faculty of Philological and Cultural Studies, University of Vienna</institution>, <addr-line>Vienna</addr-line>, <country>Austria</country></aff>
<author-notes>
<fn fn-type="edited-by"><p>Edited by: Bernhard Hommel, University Hospital Carl Gustav Carus, Germany</p></fn>
<fn fn-type="edited-by"><p>Reviewed by: Klaus Blischke, Saarland University, Germany; Michele Feist, University of Louisiana at Lafayette, United States</p></fn>
<corresp id="c001">&#x002A;Correspondence: Ulrich Ansorge, <email>ulrich.ansorge@univie.ac.at</email></corresp>
<fn fn-type="other" id="fn004"><p>This article was submitted to Cognition, a section of the journal Frontiers in Psychology</p></fn>
</author-notes>
<pub-date pub-type="epub">
<day>20</day>
<month>05</month>
<year>2022</year>
</pub-date>
<pub-date pub-type="collection">
<year>2022</year>
</pub-date>
<volume>13</volume>
<elocation-id>875744</elocation-id>
<history>
<date date-type="received">
<day>14</day>
<month>02</month>
<year>2022</year>
</date>
<date date-type="accepted">
<day>02</day>
<month>05</month>
<year>2022</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x00A9; 2022 Ansorge, Baier and Choi.</copyright-statement>
<copyright-year>2022</copyright-year>
<copyright-holder>Ansorge, Baier and Choi</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/"><p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p></license>
</permissions>
<abstract>
<p>How does the language we speak affect our perception? Here, we argue for linguistic relativity and present an explanation through &#x201C;language-induced automatized stimulus-driven attention&#x201D; (LASA): Our respective mother tongue <italic>automatically</italic> influences our attention and, hence, perception, and in this sense determines what we see. As LASA is highly practiced throughout life, it is difficult to suppress, and even shows in language-independent non-linguistic tasks. We argue that attention is involved in language-dependent processing and point out that automatic or stimulus-driven forms of attention, albeit initially learned as serving a linguistic skill, account for linguistic relativity as they are automatized and generalize to non-linguistic tasks. In support of this possibility, we review evidence for such automatized stimulus-driven attention in language-independent non-linguistic tasks. We conclude that linguistic relativity is possible and in fact a reality, although it might not be as powerful as assumed by some of its strongest proponents.</p>
</abstract>
<kwd-group>
<kwd>language</kwd>
<kwd>attention</kwd>
<kwd>linguistic relativity</kwd>
<kwd>visual saccade</kwd>
<kwd>automatic processing</kwd>
</kwd-group>
<contract-sponsor id="cn001">Vienna Science and Technology Fund<named-content content-type="fundref-id">10.13039/501100001821</named-content></contract-sponsor>
<counts>
<fig-count count="2"/>
<table-count count="0"/>
<equation-count count="0"/>
<ref-count count="74"/>
<page-count count="9"/>
<word-count count="7833"/>
</counts>
</article-meta>
</front>
<body>
<sec id="S1" sec-type="intro">
<title>Introduction</title>
<p>According to the concept of linguistic relativity, the language a human speaks shapes her/his perception (for recent reviews of the evidence, see <xref ref-type="bibr" rid="B6">Athanasopoulos and Casaponsa, 2020</xref>; <xref ref-type="bibr" rid="B39">Lupyan et al., 2020</xref>). To put it in the words of Sapir, we humans &#x201C;see and hear and otherwise experience very largely as we do because the language habits of our community predispose certain choices of interpretation&#x201D; (<xref ref-type="bibr" rid="B53">Sapir, 1941/1964</xref>, p. 69; see also <xref ref-type="bibr" rid="B72">Whorf, 1956</xref>). We propose that stimulus-driven attention accounts for some of these profound influences of language on perception, specifically with our currently proposed model, which we call &#x201C;language-induced automatized stimulus-driven attention&#x201D; (LASA).</p>
<p>Within the framework supporting &#x201C;linguistic relativity,&#x201D; researchers have proposed different claims (<xref ref-type="bibr" rid="B60">Slobin, 1996</xref>; <xref ref-type="bibr" rid="B41">Majid et al., 2004</xref>) in the degree to which language directly influences our perception. As will be shown, our proposed model is closer to some recent views supporting linguistic relativity, but contrasts with others. Importantly, our model goes further than any current view by incorporating &#x201C;automatization&#x201D; &#x2013; from the stimulus-driven level &#x2013; in the mechanism of language shaping perception and explaining in detail how it works.</p>
<p>Our view aligns with a strong version of linguistic relativity, which drew on the concept of attention to argue for linguistic relativity, such as <xref ref-type="bibr" rid="B41">Majid et al. (2004)</xref> or <xref ref-type="bibr" rid="B66">Thierry et al. (2009)</xref>. <xref ref-type="bibr" rid="B41">Majid et al. (2004)</xref> for example, noted that languages show different preferred spatial frames of references for spatial object relations, such as absolute positions (e.g., &#x201C;to the North&#x201D; in Guugu Yimithirr, even for the locations of &#x201C;table-top&#x201D; objects, such as the position of a spoon relative to plate) versus relative positions (e.g., &#x201C;to the left&#x201D; in Dutch, for the same relations). <xref ref-type="bibr" rid="B41">Majid et al. (2004)</xref> observed that speakers of such different languages showed the consistent language-specific preferences for particular frames of references when solving non-linguistic object-arrangement tasks. Moreover, these authors speculated that attention &#x2013; the human ability to select some object, feature, location, or sense, while at the same time ignoring others &#x2013; could be responsible for these generalizations of &#x201C;linguistic habits&#x201D; to non-linguistic tasks. According to this idea, for example, speakers of Tzeltal, another language with a preference for absolute frames of references, would in general pay attention to update locations of objects in their surroundings in terms of [or &#x201C;following&#x201D;] the absolute frame of references.</p>
<p>However, our argument is in contrast to widely held but weak version of linguistic relativity, namely that the influences of linguistic specifics on attention only play out during verbal processing itself &#x2013; that is, while speaking, reading, writing, or verbally comprehending (e.g., <xref ref-type="bibr" rid="B60">Slobin, 1996</xref>, <xref ref-type="bibr" rid="B61">2000</xref>, <xref ref-type="bibr" rid="B62">2003</xref>). <xref ref-type="bibr" rid="B60">Slobin (1996</xref>, p. 89) admitted that during communication, attention can play a decisive role for what is represented: &#x201C;In brief, each native language has trained its speakers to pay different kinds of attention to events and experiences when talking about them.&#x201D; However, Slobin restricted these influences to linguistic processing (e.g., <xref ref-type="bibr" rid="B62">Slobin, 2003</xref>, p. 159): &#x201C;I wish to argue that serious study of <italic>language in use</italic> points to pervasive effects of language on selective attention and memory for particular event characteristics. As I&#x2019;ve argued in greater detail elsewhere (<xref ref-type="bibr" rid="B60">Slobin, 1996</xref>, <xref ref-type="bibr" rid="B61">2000</xref>), whatever effects language may have when people are not speaking or listening, the mental activity that goes <italic>on while formulating and interpreting utterances</italic> is not trivial or obvious, and it deserves attention&#x201D; [italics added by us]. In conclusion, <xref ref-type="bibr" rid="B60">Slobin (1996; 2000; 2003</xref>), clearly argues that <italic>use of language</italic> is involved in mental activity and it is such &#x201C;covert&#x201D; language use that influences our non-linguistic cognition during communication.</p>
<p>The question of whether <italic>attention that derives from language-specific grammar</italic> generalizes to non-linguistic tasks in a pervasive/direct way and, therefore, allows for linguistic relativity, is contested until today, as hitherto the evidence is not that clear. For example, critics could argue that observers in the studies on spatial frames of references reviewed by <xref ref-type="bibr" rid="B41">Majid et al. (2004)</xref> were free to use linguistic means to solve their spatial tasks. At the very least, language might be helpful when an observer has to decide where to put one object relative to another one as was required in the spatial puzzles reviewed in <xref ref-type="bibr" rid="B41">Majid et al. (2004)</xref>. Likewise, proof of linguistic influences on attention in linguistic tasks as provided by <xref ref-type="bibr" rid="B60">Slobin (1996; 2000; 2003</xref>), does not rule out that these linguistic influences on attention would not also be present in non-linguistic tasks. The current perspective, thus, seeks to explain how language-dependent attentional selection exerts its effects in non-linguistic tasks.</p>
</sec>
<sec id="S2">
<title>Language as a Skill</title>
<p>To start with, languages are symbolic systems for communicative purposes (cf. <xref ref-type="bibr" rid="B13">Clark and Clark, 1977</xref>). Its symbols are different from, but referring to such diverse things as ideas, objects, and events. Take the example of the word <italic>bread</italic>. It refers to a piece of bread as an object but is clearly different from the particular shape or material of the object. Language is also conventional. Its meaning, use, and structure are governed by the rules shared by the many people who speak it. For instance, only if people consent upon the meaning of the word <italic>bread</italic>, can this word be used to successfully refer to this object and can the word be correctly understood in a conversation. Language is also structured on many different levels from phonemes and letters via words, sentences, and texts, to turn-taking and discourses.</p>
<p>Critically, language execution &#x2013; whether in the form of speaking, reading, or auditory comprehension &#x2013; is based on acquired skills (e.g., <xref ref-type="bibr" rid="B64">Stark, 1980</xref>; <xref ref-type="bibr" rid="B1">Afflerbach et al., 2008</xref>; <xref ref-type="bibr" rid="B14">DeKeyser, 2020</xref>). Starting from an early age, humans practice their mother tongue for years. Skills consist of a number of different actions (e.g., with different effectors) and covert processes that are jointly executed in an integrated fashion, for example, with particular sequences and with switches back and forth between gross and fine motor actions or between overt motor actions and covert processing (e.g., <xref ref-type="bibr" rid="B43">Newell and Rosenbloom, 1981</xref>; <xref ref-type="bibr" rid="B4">Anderson et al., 2004</xref>). Think of reading as an example. The eyes move overtly across the text by jumping movements, so-called saccades, and intermittent fixations &#x2013; that is, times during which the eyes are relatively still. Without much effortful monitoring, these overt eye movement actions smoothly take turns with covert processes of lexical and phonological processing of inputs, mostly occurring during fixations (cf. <xref ref-type="bibr" rid="B51">Reichle et al., 2003</xref>; <xref ref-type="bibr" rid="B19">Engbert et al., 2005</xref>). There is evidence that, during fixations, covert processing steps such as lexical processing occur as indicated by both behavioral as well as electrophysiological evidence (e.g., <xref ref-type="bibr" rid="B58">Sereno and Rayner, 2003</xref>; <xref ref-type="bibr" rid="B16">Dimigen et al., 2011</xref>). For example, fixations tend to land on words carrying the bulk of the information rather than on mere functional words such as articles. Fixations are also typically longer for less than for more frequent words and for ambiguous than for unambiguous words (<xref ref-type="bibr" rid="B17">Duffy et al., 1988</xref>; <xref ref-type="bibr" rid="B49">Rayner et al., 1996</xref>). All of these findings indicate that, during fixations, information from the words is covertly processed, either about the currently fixated word or about the potential candidate words for the next fixation (or landing point).</p>
<p>Reading, speaking, or verbal comprehension are all skills. As a consequence of their regular practice, skills are represented in human long-term memory (e.g., <xref ref-type="bibr" rid="B30">Johnson, 2013</xref>). If sufficiently well-trained, they are executed automatically with high efficiency, meaning that they do not need much top-down control, active monitoring, or willing deliberation to start with (e.g., <xref ref-type="bibr" rid="B18">d&#x2018;Ydewalle et al., 1991</xref>; <xref ref-type="bibr" rid="B22">Feng et al., 2013</xref>). Instead, it is typical for skills to proceed automatically, at least in part, meaning that control is delegated to the fitting stimuli, and that the stimuli themselves can trigger the skill (cf. <xref ref-type="bibr" rid="B50">Reason, 1990</xref>). As an example, think of hearing someone calling your own name. Even if you are entirely occupied with doing something else, hearing your own name is often distracting (cf. <xref ref-type="bibr" rid="B45">Perrin et al., 1999</xref>; <xref ref-type="bibr" rid="B52">R&#x00F6;er et al., 2013</xref>). It can capture your attention and trigger verbal comprehension, though you might have been engaged in an entirely non-linguistic task, such as painting or running.</p>
</sec>
<sec id="S3">
<title>Language Skills and Visual Attention</title>
<p>Visual attention is critical part and parcel of most skills, including language skills (e.g., <xref ref-type="bibr" rid="B30">Johnson, 2013</xref>). In this context, <italic>attention</italic> denotes the selection of information for its use in action control (<italic>sic</italic>!) and covert processing (e.g., encoding information into memory or retrieving information from memory; <italic>sic</italic>!). <italic>Visual attention</italic> denotes the corresponding selections of information from the visual environment. That visual attention is involved in language skills is obvious in the case of reading. For example, for the successful fixation of the next word during reading, humans need to select the position of this word for the programming of a saccade with a fitting direction and amplitude (cf. <xref ref-type="bibr" rid="B51">Reichle et al., 2003</xref>; <xref ref-type="bibr" rid="B19">Engbert et al., 2005</xref>). In addition, in line with a high degree of automaticity, words seemingly attract visual attention automatically. For example, free-viewing studies of photographic pictures yields evidence for the influence of local visual feature contrasts on fixation directions: Humans prefer to look at locations characterized by local contrasts in terms of color, luminance, or orientation (cf. <xref ref-type="bibr" rid="B28">Itti et al., 1998</xref>). This is regarded as evidence for stimulus-driven capture of attention by visual salience, which is basically the summed local feature contrast. Critically, however, in such situations written words notoriously outperform saliency in their attraction of attention (cf. <xref ref-type="bibr" rid="B31">Judd et al., 2009</xref>; <xref ref-type="bibr" rid="B8">Borji, 2012</xref>). For example, they capture the eyes much more than would be expected based on their relative saliency alone and, thus, require amending the basic salience model (e.g., <xref ref-type="bibr" rid="B8">Borji, 2012</xref>). This finding is perfectly in line with a learned and skill-dependent stimulus-driven effect of visual words on attention: Even if currently no linguistic processing is required (as in a free-viewing task), a word that fits to the linguistic skill of reading would trigger an attention shift to the word as part of a stimulus-driven activation of this linguistic skill.</p>
<p>Importantly, the same type of stimulus-driven visual attention in the service of linguistic skills can be observed for other skills than reading. Language production and in particular naming of objects require that the current visual surroundings are taken into account for linguistic processing. For example, if I want to have a sip of milk that is in a jug out of my reach at a coffee table, I could ask a person closer to the milk if she could please hand me the jug. Though I would have some flexibility in how to express my wish, depending on the exact looks of an object, I could not just use any word for the object of my desire. For instance, asking to please hand me the milk container is probably also okay. However, asking for a cup or a spoon would not do the job. These labels would not fit the purpose, my table neighbor would not know that I want the milk jug. In this sense, visual attention to linguistically critical characteristics of my surroundings is necessary for choosing an appropriate label and, thus, visual attention is linked to language production.</p>
<p>Crucially, a tight attentional coupling between verbal labels and objects is not only a logical necessity. It is also the kind of connection observed during language acquisition. In an ingenious study, for example, <xref ref-type="bibr" rid="B63">Smith and Yu (2008)</xref>, presented pairs of objects to their 12- to 14-months old participants and consistently paired novel verbal labels with only one of the two objects per pair, while the second object was selected randomly from a set of additional novel or yet-to-be-learned labels. In this situation, children preferentially looked at consistently labeled objects, demonstrating that attention &#x2013; here, the spatial selection of fixated objects &#x2013; reflected statistical learning and picked up upon the cross-modal visual-verbal probabilities.</p>
</sec>
<sec id="S4">
<title>Language Differences and Their Effects on Stimulus-Driven Attention</title>
<p>Instances of language production can vary depending on the language that one speaks (<xref ref-type="bibr" rid="B10">Choi and Bowerman, 1991</xref>). While certain features of an object/event are habitually/regularly highlighted in the grammar of one language, they may not be in other languages. To the degree that habitual visual discriminations are required for linguistic productions in Language A but not in Language B, speakers of these languages could differ with respect to their vulnerability to stimulus-driven attention capture by a stimulus in question (cf. <xref ref-type="bibr" rid="B11">Choi et al., 1999</xref>). Most obviously, whenever I have to consider my visual environment for an appropriate linguistic expression in Language A, my attention would be directed to the corresponding visual information &#x2013; that is, the linguistically critical information would be selected as part of my linguistic production skills (cf. <xref ref-type="bibr" rid="B68">Tomasello and Farrar, 1986</xref>; <xref ref-type="bibr" rid="B69">Tomasello and Kruger, 1992</xref>). This means that attention is habitually shifted to particular features of objects as part of my linguistic practice (cf. <xref ref-type="bibr" rid="B34">Knott, 2012</xref>). However, if I, as a speaker of a particular language, now encounter a stimulus habitually fitting to my language skills but outside of the current linguistic/non-linguistic task &#x2013; that is, when I do not have to produce nor comprehend a fitting linguistic expression and am in an entirely non-linguistic situation &#x2013; by virtue of the fact that a fitting stimulus could automatically trigger a skill itself, this fitting stimulus could capture attention in a stimulus-driven way and lead to language-dependent visual selection of particular features (cf. <xref ref-type="bibr" rid="B26">Goller et al., 2017</xref>).</p>
<p><xref ref-type="bibr" rid="B25">Goller et al. (2020)</xref> recently tested and confirmed exactly this prediction in what may be called an instance of domain-centered research on linguistic relativity (<xref ref-type="bibr" rid="B36">Lucy, 2016</xref>). For their tests, they used an established benchmark of stimulus-driven capture of visual attention &#x2013; the distraction effect by a visual singleton (cf. <xref ref-type="bibr" rid="B65">Theeuwes, 1992</xref>). Here, a singleton denotes a visual stimulus that is salient and, thus, stands out by at least one of its features among several more feature-homogeneous non-singletons. For instance, a green apple among red apples would be a singleton among non-singletons. Some singletons can capture attention in a stimulus-driven way, even if entirely task-irrelevant (e.g., <xref ref-type="bibr" rid="B71">Weichselbaum and Ansorge, 2018</xref>). For instance, during visual search for a shape-defined target (e.g., for a circle among diamonds) presenting a singleton distractor in a different color than all other stimuli (e.g., a green diamond among red diamonds and one red circle as the target) and presenting this singleton distractor away from the target delays successful search for the target by attracting attention to the singleton distractor (<xref ref-type="bibr" rid="B65">Theeuwes, 1992</xref>).</p>
<p>Singleton capture, as we may call this kind of stimulus-driven attention, is not only a consequence of the type of visual information: It can also occur as a consequence of learning (<xref ref-type="bibr" rid="B3">Anderson et al., 2011</xref>; <xref ref-type="bibr" rid="B9">Bucker and Theeuwes, 2014</xref>; <xref ref-type="bibr" rid="B21">Failing and Theeuwes, 2018</xref>). For example, rewarding Color A (say red) on average more than Color B (e.g., green) during a training phase, leads to singleton capture and thus, interference by a color distractor with a previously rewarded color during visual search for a shape target in a subsequent test phase: Even though the color of the target and of just any stimulus in the shape search task is entirely task-irrelevant, that is, does no longer lead to any reward &#x2013; and participants know all that &#x2013; presenting the previously reward-associated color as a distractor away from the target delays search, as much as any other singleton would do (<xref ref-type="bibr" rid="B3">Anderson et al., 2011</xref>). The learned reward-associated color stands out by its acquired saliency so to say.</p>
<p>Such learning-dependent singleton capture also occurs for a different type of learning, namely linguistic skills. For their tests, <xref ref-type="bibr" rid="B25">Goller et al. (2020)</xref> made use of the highly practiced linguistic discrimination of the tightness of the spatial fit between objects in Korean language but not in English or German (<xref ref-type="bibr" rid="B26">Goller et al., 2017</xref>; <xref ref-type="bibr" rid="B74">Yun and Choi, 2018</xref>). In Korean, speakers ubiquitously use the word <italic>kkita</italic> for a tight fit, for example, when a cap fits on a pen tightly, and they use the words <italic>nehta</italic> or <italic>nohta</italic> for a loose fit, for example, when an olive is loosely surrounded by a bowl (<xref ref-type="bibr" rid="B74">Yun and Choi, 2018</xref>). In German, in contrast, such semantic distinctions are possible, but they are not obligatory, nor are they ubiquitously made. Neither is the choice of the suited verb determined by these distinctions, nor is a fit-discrimination an obligatory preposition/particle in German grammar. Thus, only Korean speakers but not speakers of German habitually and obligatorily have to discriminate linguistically between tight and loose spatial fits.</p>
<p><xref ref-type="bibr" rid="B25">Goller et al. (2020)</xref> reasoned that if such skilled linguistic discrimination influences which stimuli or features can capture attention in a stimulus-driven way in non-linguistic tasks, a visual distractor that is presented away from the target and that stands out by the linguistically discriminated feature should be salient and should capture attention in a stimulus-driven way &#x2013; even when the feature relates to distractors, and not to the target. If so, such attention would interfere with performance even in a non-linguistic color-target search task. Critically, this interference was expected in a group of language users that had to linguistically discriminate the corresponding visual input (here, Korean speakers that are obliged to linguistically discriminate between tight and loose fits for the choice of a correct verb) but not in a group of speakers of a different language without these linguistic obligations (here, German speakers).</p>
<p>This expectation was borne out by the results. During search for a color-defined (e.g., red) target among (e.g., green) distractors, presentation of a singleton distractor with a unique spatial fit (e.g., one loosely fitting object among several tight-fitting objects) spatially away from the target delayed visual search among Korean speakers but not among German speakers (Experiment 4 of <xref ref-type="bibr" rid="B25">Goller et al., 2020</xref>; see <xref ref-type="fig" rid="F1">Figure 1</xref>, below). In addition, control conditions showed that this stimulus-driven capture of attention by the singleton distractor was not due to a generally larger proneness of Korean speakers to stimulus-driven attention capture. In a control condition, with color-singleton distractors (i.e., a blue singleton-distractor among green non-singleton distractors) that should have been equally salient to both language groups, Korean and German speakers showed the same degree of singleton-distractor interference (Experiment 4 of <xref ref-type="bibr" rid="B25">Goller et al., 2020</xref>). In another control condition, Korean speakers showed evidence of stimulus-driven capture by fit-singleton distractors when the distractors were 2D depictions of 3D objects (i.e., pistons within tightly vs. loosely surrounding cylinders) but <underline>not</underline> when 2D images were used (i.e., disks within tightly or loosely surrounding rings). The latter finding supports the conclusion that stimulus-driven capture by the spatial-fit singletons was language-dependent, as Korean speakers would use tight-fit verbs (e.g., <italic>kkita</italic>) to discriminate the tightness of fit of 3D objects but not of 2D objects. At the same time, these data also demonstrated that Korean speakers were not simply more sensitive to just any type of contextual information (cf. <xref ref-type="bibr" rid="B44">Nisbett and Miyamoto, 2005</xref>) &#x2013; here: of the otherwise task-irrelevant spatial fits&#x2013;, but just the one that corresponded to highly practiced semantic discrimination in their language. In conclusion, <xref ref-type="bibr" rid="B25">Goller et al. (2020)</xref> provide strong evidence of language-specific semantics influencing perception at a stimulus-feature level outside linguistic tasks.</p>
<fig id="F1" position="float">
<label>FIGURE 1</label>
<caption><p>An illustration of the procedure of <xref ref-type="bibr" rid="B25">Goller et al. (2020)</xref>, Experiment 4. Korean- speaking and German-speaking participants had to search for a color-defined target (e.g., a red target as depicted), and a fit-singleton distractor (e.g., a loose-fit ring among tight-fit rings) was presented away from the target in half of the trials. Compared to a condition without spatial-fit distractor, Korean speakers but not German speakers showed slower search performance for the color targets. This is in line with linguistic relativity, as search for the targets was delayed by stimulus-driven capture of attention toward the spatial-fit singletons only among the Korean speakers that verbally discriminate between tightness of fit levels in an obligatory way, but not among the German speakers that do not have to verbally discriminate different fits.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-875744-g001.tif"/>
</fig>
<p>Importantly, in the color-search task, participants did not use a linguistic strategy to solve the task (<xref ref-type="bibr" rid="B7">Baier and Ansorge, 2019</xref>): During search for color-defined targets (e.g., a red target among green distractors), participants do not show any signs of verbal rehearsal &#x2013; that is, visual search performance is not disrupted by the concomitant task of having to rehearse a syllable &#x2013; and participants&#x2019; attention is not captured by visual color words, even if these words denote the searched-for color (e.g., the word red during search for red targets; <xref ref-type="bibr" rid="B7">Baier and Ansorge, 2019</xref>).</p>
</sec>
<sec id="S5">
<title>How Does Language-Dependent Stimulus-Driven Capture of Attention Account for Linguistic Relativity?</title>
<p>We have presented and argued for &#x201C;language-induced automatized stimulus-driven attention&#x201D; (LASA). According to this argument, attentional selection (of objects, features, or locations) has been repeatedly practiced as part of a linguistic skill to such an extent from early in life that the corresponding forms of selection do no longer require top-down control to trigger such a selection. Instead, with sufficient practice, the corresponding perceptual input can capture attention in a stimulus-driven way &#x2013; that is, stimuli can provoke their selection for processing by themselves, even outside linguistic tasks. To understand this, let us first look a little closer at skills and their automaticity or proceduralization (<xref ref-type="bibr" rid="B5">Anderson and Lebiere, 1998</xref>; <xref ref-type="bibr" rid="B4">Anderson et al., 2004</xref>). As <xref ref-type="bibr" rid="B5">Anderson and Lebiere (1998)</xref> described in their production model system ACT-R, skills in form of procedural knowledge exist as patterns connecting top-down goals with productions in procedural memory in recurrent feedback loops. Skills consist of sequences of procedures, with to-be-executed productions, where productions cover both overt actions (e.g., grasping an object) and covert processes (e.g., word comprehension) (cf. <xref ref-type="bibr" rid="B4">Anderson et al., 2004</xref>). These actions/processes are frequently and habitually practiced and, thus, are part of long-term memory (cf. <xref ref-type="bibr" rid="B20">Ericsson and Kintsch, 1995</xref>). As procedures, they follow a general form, consisting of sequences of processing steps that are executed conditionally on the fulfillment of specific eliciting conditions &#x2013; a process called &#x201C;pattern matching&#x201D; in <xref ref-type="fig" rid="F2">Figure 2 (Anderson and Lebiere, 1998</xref>). For example, in language production, a speaker would first look at the critical characteristics of an event to produce a sentence describing what she sees (cf. <xref ref-type="bibr" rid="B60">Slobin, 1996</xref>). Imagine a speaker of English who registers the durative nature of an illustrated event in a book and marks the progressive aspect on a fitting verb, for instance, &#x201C;the dog was running from the bees&#x201D; (cf. <xref ref-type="bibr" rid="B60">Slobin, 1996</xref>). Importantly, and in contrast to what <xref ref-type="bibr" rid="B60">Slobin (1996)</xref> believes, pattern matching and, hence, attention shifts to the corresponding characteristics of an image would not only run off in a top-down controlled fashion only, for example, when an intention to communicate requires this, but also, following repeated language practice, a stimulus as a fitting input could trigger the production. That is, the durative nature of an event would attract attention &#x2013; without intervening top-down control, therefore, practice facilitates stimulus-driven shifts or capture of attention. As skills are frequently or habitually practiced, they become automatized (or proceduralized): When a skill is originally acquired, it typically requires exerting top-down control to get started. However, practicing a skill means that control about what to do next in a sequence of overt motor responses or covert processing steps is delegated more and more to the stimuli that are used in the course of a skill&#x2019;s pattern matching process (<xref ref-type="bibr" rid="B42">Neisser and Becklen, 1975</xref>; <xref ref-type="bibr" rid="B5">Anderson and Lebiere, 1998</xref>; <xref ref-type="bibr" rid="B4">Anderson et al., 2004</xref>). Thus, with practice, stimuli used in pattern matching take over the role of initiating a procedure.</p>
<fig id="F2" position="float">
<label>FIGURE 2</label>
<caption><p>Mode of control &#x2013; top-down/by goals of the observer versus stimulus-driven &#x2013; of covert and overt productions in long-term skill memory as a function of skill practice. All productions (e.g., linguistic/attentional, motor) are triggered by pattern matching (see boxes on the left), comparing an input stimulus to a specific template or parameter predefined as a critical precondition for the execution of the production, and the execution of the production itself (see boxes on the right) (<xref ref-type="bibr" rid="B4">Anderson et al., 2004</xref>). With practice (downward pointing arrow on the far left), control shifts from top-down, goal-directed selection of the production (top row), to stimulus-driven elicitation of the same production (bottom row). See text for additional explanations.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-875744-g002.tif"/>
</fig>
<p>In the case of the sensation of a particular sensory pattern triggering an attention shift that originally served a correct linguistic production, this is really not such a wonder, as above we have already emphasized that such visual-verbal pattern-label associations form an important basis of language acquisition in the first place (<xref ref-type="bibr" rid="B63">Smith and Yu, 2008</xref>).</p>
<p>At this point, a related clarification is in order that might be responsible for reservations against our suggested hypotheses. We have emphasized that language comprehension and production are skills. Therefore, they are part of the (more implicit) procedural memory of humans (<xref ref-type="bibr" rid="B70">Ullman, 2001</xref>; <xref ref-type="bibr" rid="B37">Lum et al., 2012</xref>). This means that (1) the selectivity underlying perception implied by a particular language is not entirely under voluntary top-down control and (2) these forms of selective processing often go unnoticed by humans from an introspective, first-person perspective. In addition, this also means that language effects on the stimulus-driven forms of perceptual input selection are possible and even likely, as they are typical of well-practiced skills and procedures in general (<xref ref-type="bibr" rid="B2">Allport, 1987</xref>) and of language in particular (<xref ref-type="bibr" rid="B47">Pulverm&#x00FC;ller and Shtyrov, 2006</xref>; <xref ref-type="bibr" rid="B35">Kotz and Schwartze, 2010</xref>).</p>
<p>However, there might be more to language as a skill than the types of bidirectional associations between verbal labels and attention shifts that we described. Most critically, language is characterized by higher-order regularities besides the discussed item-adjacent visual-verbal dependencies. There are many more language-specific regularities concerning non-adjacent item dependencies that are reflective of a recursive or hierarchical language-inherent structure, and, importantly, these higher-order regularities could be more typical of natural languages than the simple visual-verbal dependencies which are at the focus of our argument (<xref ref-type="bibr" rid="B23">Fitch and Friederici, 2012</xref>; <xref ref-type="bibr" rid="B29">Jager and Rogers, 2012</xref>; <xref ref-type="bibr" rid="B24">Fitch and Martins, 2014</xref>). We do not claim that all of these language-inherent statistical characteristics need to be equally easily triggered by the mere presence of a matching pattern. It might well be that stimulus-driven language-specific skill elicitation is restricted to the somewhat simpler associations between sensory inputs and originally linguistic attention shifts that we described above.</p>
<p>This brings us to the interesting question about the &#x201C;complexity&#x201D; of these associations and the type of representational system that might be used to represent these types of skills. First of all, we want to emphasize that a degree of processing demand is implied by the cross-modal nature of the visual-verbal links that we described. A representational system sensitive to the statistical or temporal regularities present within a single modality would, thus, be insufficient to represent the corresponding information (cf. <xref ref-type="bibr" rid="B32">Keele et al., 2003</xref>). This makes it more likely that the corresponding knowledge is represented in what <xref ref-type="bibr" rid="B32">Keele et al. (2003)</xref> called the &#x201C;ventral system,&#x201D; possibly, with the Inferior Parietal Lobe as the area of multi-modal convergence in which the types of spatial tight- versus loose-fit relations that <xref ref-type="bibr" rid="B25">Goller et al. (2020)</xref> investigated could be represented (cf. <xref ref-type="bibr" rid="B33">Kemmerer, 2006</xref>). This system would be sensitive to the task relevance of the associated material &#x2013; here the visual-verbal pairs connecting fitting sensory input with a corresponding verbal label &#x2013; and, hence, its proper functioning could even sometimes be vulnerable to the characteristics of a non-linguistic test task. For example, it might simply be difficult to observe the corresponding language-dependent attention shifts in non-linguistic tasks if these test tasks would require or even only offer the acquisition of even simpler (e.g., within-modality) item associations. Here, we can see that there are reasons for the possible failure of tests of linguistic relativity beyond the ones discussed by <xref ref-type="bibr" rid="B60">Slobin (1996)</xref>, who speculated that such non-linguistic tests could fail because of the language-specific nature of linguistic discriminations that has little connection to other sensory and motor discriminations. On the contrary, we would argue that as the building blocks for linguistic skills are relatively similar to that of sensory and motor skills, a point that is particularly true of attention shifts to objects, their features or characteristics, there is plenty of opportunity for the alternative usage of the same non-linguistic part devices in non-linguistic tasks that could potentially block or allow their usage as according to a linguistic skill in these very same tasks.</p>
<p>Here, we want to conclude that the thereby assumed more gradual nature of the involved skills and representations as being neither perfectly sensory/motor, nor fully linguistic, is also better in line with the type of proto-conceptual representations that are obviously driving perception from infancy onward (<xref ref-type="bibr" rid="B73">Xu, 2019</xref>). In particular, <xref ref-type="bibr" rid="B73">Xu (2019)</xref> came to the conclusion that perceptual discriminations of infants already show characteristics of linguistic representations, such as (partial) categorization and (imperfect) enrichment of situations with organismic beliefs, that are not yet at an adult level but that would certainly also make it difficult to keep up the impermeable boundary between the linguistic and the sensory sphere that seems to have dominated the thinking of the great theoreticians for so long (cf. <xref ref-type="bibr" rid="B46">Piaget, 1954</xref>; <xref ref-type="bibr" rid="B12">Chomsky, 1987</xref>; or <xref ref-type="bibr" rid="B60">Slobin, 1996</xref>; <xref ref-type="bibr" rid="B48">Pylyshyn, 1999</xref>).</p>
<sec id="S5.SS1">
<title>Relationship Between Attention and Perception</title>
<p>Of interest for the theory of linguistic relativity is now the relation between attention and perception. The argument is basically that perception depends on attention and, thus, linguistic influences on attention can literally shape our view of the world. Inherent to this argument for linguistic relativity through language-dependent stimulus-driven attention is the conviction that attention precedes perception and, thus, shapes what humans perceive. Since long, it has been assumed that without attentional selection, perception is not possible (<xref ref-type="bibr" rid="B67">Titchener, 1908</xref>; <xref ref-type="bibr" rid="B15">Di Lollo et al., 2000</xref>). Although this position might be too extreme and there might be situations in which perception is possible without attention, attention as the selection of particular stimuli, features or locations is at least abundant. Attention can at least facilitate and, thus, modulate perception, when manipulated by the experimenter (<xref ref-type="bibr" rid="B54">Scharlau, 2002</xref>, <xref ref-type="bibr" rid="B55">2004</xref>; <xref ref-type="bibr" rid="B57">Scharlau and Neumann, 2003</xref>). For example, shifting attention to the position of a visual stimulus, in advance of this stimulus, can speed up this stimulus&#x2019; perception as reflected in its subjectively apparent temporal precedence relative to a concomitant second stimulus that does not benefit from a like shift of attention (<xref ref-type="bibr" rid="B54">Scharlau, 2002</xref>). In other words, of two simultaneous stimuli, the one that we attend to first is perceived earlier: It seems to precede the stimulus to which we do not attend. Thus, perception &#x2013; here of the time at which a stimulus is perceived &#x2013; is shaped by attention. Importantly, in line with a role of attention for perception, this selection of perceptual input can occur independently of and, thus, prior to the perceiver&#x2019;s awareness of a stimulus (<xref ref-type="bibr" rid="B54">Scharlau, 2002</xref>; <xref ref-type="bibr" rid="B56">Scharlau and Ansorge, 2003</xref>). This temporal sequence of attentional selection prior to perceptual awareness of a stimulus is in line with the modulating role of attention on perception. In addition, attentional selection of one stimulus, feature, or event can come at the expense of missing out on alternative stimuli, features, or events (<xref ref-type="bibr" rid="B40">Mack and Rock, 1998</xref>; <xref ref-type="bibr" rid="B59">Simons and Chabris, 1999</xref>; <xref ref-type="bibr" rid="B27">Horstmann and Ansorge, 2016</xref>). In this way, language-induced stimulus-driven attention (LASA) is also in a position to determine what exactly a human observer can perceive.</p>
</sec>
<sec id="S5.SS2">
<title>How Much Does Language Determine Humans&#x2019; Perception of the World?</title>
<p>As we have argued, as a skill, language is closely linked to attention in a way that the domains of language and perception are not such distinct unconnected spheres. However, we are not certain how much this extends to differences for human perception as a whole. The reasons for this are twofold. First, different languages share commonalities. For example, <xref ref-type="bibr" rid="B74">Yun and Choi (2018)</xref> proposed that languages share a common set of spatial features (e.g., containment, support, degree of fit) but that they differ in the degree to which they highlight those features in their semantic system. The varying degrees of commonalities also create a huge overlap in the way humans perceive objects, regardless of the particular language they speak. Second, the language-specific effects on stimulus-driven attention and perception are in general similar to other forms of practice-dependent long-term memory effects. Each skill that humans acquire also entails some forms of skill-dependent sensitivities and insensitivities for the selection of particular skill-implied perceptual inputs (cf. <xref ref-type="bibr" rid="B30">Johnson, 2013</xref>). Language is, thus, not the only way in which practice leads to a change of attention and perception. As a consequence, many other human skills, such as walking, grasping, driving, or eating, can all have an impact on how we humans attend to objects and, thus, perceive them. These non-linguistic skills provide a rich source for both language-independent commonalities and differences in the way humans allocate attention and perceive the world.</p>
<p>Nevertheless, in the present review, we highlighted that language alters the way humans&#x2019; attention is attracted by different stimuli and features. As each language has its own semantic system that systematically highlights a specific set of features or feature differences, which may differ &#x2013; and often do &#x2013; from other languages, attention to those language-specific features taken together can contribute to significant differences in the way speakers of different languages look at the world (cf. <xref ref-type="bibr" rid="B66">Thierry et al., 2009</xref>; <xref ref-type="bibr" rid="B38">Lupyan, 2012</xref>).</p>
</sec>
</sec>
<sec id="S6" sec-type="conclusion">
<title>Conclusion</title>
<p>Critics of linguistic relativity used task-dependent top-down attention to verbal concepts to explain away language-induced effects and, thus, would not accept that language can influence attention/perception in language-independent non-linguistic contexts. In this paper, we make a counter argument with a set of well-established processing mechanisms showing intimate interaction between language and perception/attention, and demonstrate that linguistic relativity is possible and even likely. In our proposed &#x201C;language-induced automatized stimulus-driven attention&#x201D; (LASA) explanation, attention &#x2013; indeed &#x2013; plays a very important role during language acquisition and processing, and in fact language and attentional skills are highly interconnected, which are then practiced jointly throughout life. These extensively rehearsed coevolved skills then subsequently lead to stimulus-driven, automatic attention capture by fitting stimuli. In this way, linguistic influences generalize, for example, to the visual selection and perception of specific objects or features, even in language-independent tasks.</p>
</sec>
<sec id="S7" sec-type="data-availability">
<title>Data Availability Statement</title>
<p>The original contributions presented in the study are included in the article/supplementary material, further inquiries can be directed to the corresponding author/s.</p>
</sec>
<sec id="S8">
<title>Ethics Statement</title>
<p>Ethical review and approval was not required for the study on human participants in accordance with the Local Legislation and Institutional Requirements. The patients/participants provided their written informed consent to participate in this study.</p>
</sec>
<sec id="S9">
<title>Author Contributions</title>
<p>UA drafted the manuscript. DB and SC revised it. All authors contributed to the article and approved the submitted version.</p>
</sec>
<sec id="conf1" sec-type="COI-statement">
<title>Conflict of Interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec id="pudiscl1" sec-type="disclaimer">
<title>Publisher&#x2019;s Note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
</body>
<back>
<sec id="S10" sec-type="funding-information">
<title>Funding</title>
<p>This research was supported by the Wiener Wissenschafts- und Technologiefonds (WWTF), Project CS-15-001, awarded to SC and UA.</p>
</sec>
<ref-list>
<title>References</title>
<ref id="B1"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Afflerbach</surname> <given-names>P.</given-names></name> <name><surname>Pearson</surname> <given-names>P. D.</given-names></name> <name><surname>Paris</surname> <given-names>S. G.</given-names></name></person-group> (<year>2008</year>). <article-title>Clarifying differences between reading skills and reading strategies.</article-title> <source><italic>Read. Teach.</italic></source> <volume>61</volume> <fpage>364</fpage>&#x2013;<lpage>373</lpage>. <pub-id pub-id-type="doi">10.1598/RT.61.5.1</pub-id></citation></ref>
<ref id="B2"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Allport</surname> <given-names>A.</given-names></name></person-group> (<year>1987</year>). &#x201C;<article-title>Selection-for-action: some behavioral and neurophysiological considerations of attention and action</article-title>,&#x201D; in <source><italic>Perspectives on Perception and Action</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>Heuer</surname> <given-names>H.</given-names></name> <name><surname>Sanders</surname> <given-names>A. F.</given-names></name></person-group> (<publisher-loc>Hillsdale, NJ</publisher-loc>: <publisher-name>Erlbaum</publisher-name>), <fpage>395</fpage>&#x2013;<lpage>419</lpage>.</citation></ref>
<ref id="B3"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Anderson</surname> <given-names>B. A.</given-names></name> <name><surname>Laurent</surname> <given-names>P. A.</given-names></name> <name><surname>Yantis</surname> <given-names>S.</given-names></name></person-group> (<year>2011</year>). <article-title>Value-driven attentional capture.</article-title> <source><italic>Proc. Natl. Sci. U.S.A.</italic></source> <volume>108</volume> <fpage>10367</fpage>&#x2013;<lpage>10371</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.1104047108</pub-id> <pub-id pub-id-type="pmid">21646524</pub-id></citation></ref>
<ref id="B4"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Anderson</surname> <given-names>J. R.</given-names></name> <name><surname>Bothell</surname> <given-names>D.</given-names></name> <name><surname>Byrne</surname> <given-names>M. D.</given-names></name> <name><surname>Douglass</surname> <given-names>S.</given-names></name> <name><surname>Lebiere</surname> <given-names>C.</given-names></name> <name><surname>Qin</surname> <given-names>Y.</given-names></name></person-group> (<year>2004</year>). <article-title>An integrated theory of the mind.</article-title> <source><italic>Psychol. Rev.</italic></source> <volume>111</volume> <fpage>1036</fpage>&#x2013;<lpage>1060</lpage>. <pub-id pub-id-type="doi">10.1037/0033-295X.111.4.1036</pub-id> <pub-id pub-id-type="pmid">15482072</pub-id></citation></ref>
<ref id="B5"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Anderson</surname> <given-names>J. R.</given-names></name> <name><surname>Lebiere</surname> <given-names>C.</given-names></name></person-group> (<year>1998</year>). <source><italic>The Atomic Components of Thought.</italic></source> <publisher-loc>Mawah, NJ</publisher-loc>: <publisher-name>Erlbaum</publisher-name>.</citation></ref>
<ref id="B6"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Athanasopoulos</surname> <given-names>P.</given-names></name> <name><surname>Casaponsa</surname> <given-names>A.</given-names></name></person-group> (<year>2020</year>). <article-title>The whorfian brain: neuroscientific approaches to linguistic relativity.</article-title> <source><italic>Cogn. Neuropsychol.</italic></source> <volume>37</volume> <fpage>393</fpage>&#x2013;<lpage>412</lpage>. <pub-id pub-id-type="doi">10.1080/02643294.2020.1769050</pub-id> <pub-id pub-id-type="pmid">32476559</pub-id></citation></ref>
<ref id="B7"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Baier</surname> <given-names>D.</given-names></name> <name><surname>Ansorge</surname> <given-names>U.</given-names></name></person-group> (<year>2019</year>). <article-title>Investigating the role of verbal templates in contingent capture by color.</article-title> <source><italic>Atten. Percept. Psycho.</italic></source> <volume>81</volume> <fpage>1846</fpage>&#x2013;<lpage>1879</lpage>. <pub-id pub-id-type="doi">10.3758/s13414-019-01701-y</pub-id> <pub-id pub-id-type="pmid">30924054</pub-id></citation></ref>
<ref id="B8"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Borji</surname> <given-names>A.</given-names></name></person-group> (<year>2012</year>). &#x201C;<article-title>Boosting bottom-up and top-down visual features for saliency estimation</article-title>,&#x201D; in <source><italic>Proceedings of 2012 IEEE Conference on Computer Vision and Pattern Recognition (IEEE)</italic></source>, <fpage>438</fpage>&#x2013;<lpage>445</lpage>. <publisher-loc>Providence, RI</publisher-loc> <pub-id pub-id-type="doi">10.1109/CVPR.2012.6247706</pub-id></citation></ref>
<ref id="B9"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bucker</surname> <given-names>B.</given-names></name> <name><surname>Theeuwes</surname> <given-names>J.</given-names></name></person-group> (<year>2014</year>). <article-title>The effect of reward on orienting and reorienting in exogenous cuing.</article-title> <source><italic>Cogn. Affect. Behav. Ne</italic></source> <volume>14</volume> <fpage>635</fpage>&#x2013;<lpage>646</lpage>. <pub-id pub-id-type="doi">10.3758/s13415-014-0278-7</pub-id> <pub-id pub-id-type="pmid">24671762</pub-id></citation></ref>
<ref id="B10"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Choi</surname> <given-names>S.</given-names></name> <name><surname>Bowerman</surname> <given-names>M.</given-names></name></person-group> (<year>1991</year>). <article-title>Learning to express motion events in English and Korean: the influence of language-specific lexicalization patterns.</article-title> <source><italic>Cognition</italic></source> <volume>41</volume> <fpage>83</fpage>&#x2013;<lpage>121</lpage>. <pub-id pub-id-type="doi">10.1016/0010-0277(91)90033-Z</pub-id></citation></ref>
<ref id="B11"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Choi</surname> <given-names>S.</given-names></name> <name><surname>McDonough</surname> <given-names>L.</given-names></name> <name><surname>Bowerman</surname> <given-names>M.</given-names></name> <name><surname>Mandler</surname> <given-names>J.</given-names></name></person-group> (<year>1999</year>). <article-title>Early sensitivity to language-specific spatial categories in English and Korean.</article-title> <source><italic>Cogn. Dev.</italic></source> <volume>14</volume> <fpage>241</fpage>&#x2013;<lpage>268</lpage>. <pub-id pub-id-type="doi">10.1016/S0885-2014(99)00004-0</pub-id></citation></ref>
<ref id="B12"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chomsky</surname> <given-names>N.</given-names></name></person-group> (<year>1987</year>). <source><italic>Language and Problems of Knowledge.</italic></source> <publisher-loc>Cambridge, MA</publisher-loc>: <publisher-name>MIT Press</publisher-name>.</citation></ref>
<ref id="B13"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Clark</surname> <given-names>H. H.</given-names></name> <name><surname>Clark</surname> <given-names>E. V.</given-names></name></person-group> (<year>1977</year>). <source><italic>Psychology and Language: An Introduction to Psycholinguistics.</italic></source> <publisher-loc>New York, NY</publisher-loc>: <publisher-name>Harcourt Brace Jovanovich</publisher-name>.</citation></ref>
<ref id="B14"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>DeKeyser</surname> <given-names>R.</given-names></name></person-group> (<year>2020</year>). &#x201C;<article-title>Skill acquisition theory</article-title>,&#x201D; in <source><italic>Theories in Second Language Acquisition</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>VanPatten</surname> <given-names>B.</given-names></name> <name><surname>Keating</surname> <given-names>G. D.</given-names></name> <name><surname>Wulff</surname> <given-names>S.</given-names></name></person-group> (<publisher-loc>London</publisher-loc>: <publisher-name>Routledge</publisher-name>), <fpage>83</fpage>&#x2013;<lpage>104</lpage>. <pub-id pub-id-type="doi">10.4324/9780429503986-5</pub-id></citation></ref>
<ref id="B15"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Di Lollo</surname> <given-names>V.</given-names></name> <name><surname>Enns</surname> <given-names>J. T.</given-names></name> <name><surname>Rensink</surname> <given-names>R. A.</given-names></name></person-group> (<year>2000</year>). <article-title>Competition for consciousness among visual events: the psychophysics of reentrant visual processes.</article-title> <source><italic>J. Exp. Psychol. G</italic></source> <volume>129</volume> <fpage>481</fpage>&#x2013;<lpage>507</lpage>. <pub-id pub-id-type="doi">10.1037/0096-3445.129.4.481</pub-id> <pub-id pub-id-type="pmid">11142864</pub-id></citation></ref>
<ref id="B16"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dimigen</surname> <given-names>O.</given-names></name> <name><surname>Sommer</surname> <given-names>W.</given-names></name> <name><surname>Hohlfeld</surname> <given-names>A.</given-names></name> <name><surname>Jacobs</surname> <given-names>A. M.</given-names></name> <name><surname>Kliegl</surname> <given-names>R.</given-names></name></person-group> (<year>2011</year>). <article-title>Coregistration of eye movements and EEG in natural reading: analyses and review.</article-title> <source><italic>J. Exp. Psychol. G</italic></source> <volume>140</volume> <fpage>552</fpage>&#x2013;<lpage>572</lpage>. <pub-id pub-id-type="doi">10.1037/a0023885</pub-id> <pub-id pub-id-type="pmid">21744985</pub-id></citation></ref>
<ref id="B17"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Duffy</surname> <given-names>S. A.</given-names></name> <name><surname>Morris</surname> <given-names>R. K.</given-names></name> <name><surname>Rayner</surname> <given-names>K.</given-names></name></person-group> (<year>1988</year>). <article-title>Lexical ambiguity and fixation times in reading.</article-title> <source><italic>J. Mem. Lang.</italic></source> <volume>27</volume> <fpage>429</fpage>&#x2013;<lpage>446</lpage>. <pub-id pub-id-type="doi">10.1016/0749-596X(88)90066-6d</pub-id></citation></ref>
<ref id="B18"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>d&#x2018;Ydewalle</surname> <given-names>G.</given-names></name> <name><surname>Praet</surname> <given-names>C.</given-names></name> <name><surname>Verfaillie</surname> <given-names>K.</given-names></name> <name><surname>Rensbergen</surname> <given-names>J. V.</given-names></name></person-group> (<year>1991</year>). <article-title>Watching subtitled television: automatic reading behavior.</article-title> <source><italic>Commun. Res.</italic></source> <volume>18</volume> <fpage>650</fpage>&#x2013;<lpage>666</lpage>. <pub-id pub-id-type="doi">10.1177/009365091018005005</pub-id></citation></ref>
<ref id="B19"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Engbert</surname> <given-names>R.</given-names></name> <name><surname>Nuthmann</surname> <given-names>A.</given-names></name> <name><surname>Richter</surname> <given-names>E. M.</given-names></name> <name><surname>Kliegl</surname> <given-names>R.</given-names></name></person-group> (<year>2005</year>). <article-title>SWIFT: a dynamical model of saccade generation during reading.</article-title> <source><italic>Psychol. Rev.</italic></source> <volume>112</volume> <fpage>777</fpage>&#x2013;<lpage>813</lpage>. <pub-id pub-id-type="doi">10.1037/0033-295X.112.4.777</pub-id> <pub-id pub-id-type="pmid">16262468</pub-id></citation></ref>
<ref id="B20"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ericsson</surname> <given-names>K. A.</given-names></name> <name><surname>Kintsch</surname> <given-names>W.</given-names></name></person-group> (<year>1995</year>). <article-title>Long-term working memory.</article-title> <source><italic>Psych. Rev.</italic></source> <volume>102</volume> <fpage>211</fpage>&#x2013;<lpage>245</lpage>. <pub-id pub-id-type="doi">10.1037/0033-295X.102.2.211</pub-id> <pub-id pub-id-type="pmid">7740089</pub-id></citation></ref>
<ref id="B21"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Failing</surname> <given-names>M.</given-names></name> <name><surname>Theeuwes</surname> <given-names>J.</given-names></name></person-group> (<year>2018</year>). <article-title>Selection history: how reward modulates selectivity of visual attention.</article-title> <source><italic>Psychon. B Rev.</italic></source> <volume>25</volume> <fpage>514</fpage>&#x2013;<lpage>538</lpage>. <pub-id pub-id-type="doi">10.3758/s13423-017-1380-y</pub-id> <pub-id pub-id-type="pmid">28986770</pub-id></citation></ref>
<ref id="B22"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Feng</surname> <given-names>S.</given-names></name> <name><surname>D&#x2019;Mello</surname> <given-names>S.</given-names></name> <name><surname>Graesser</surname> <given-names>A. C.</given-names></name></person-group> (<year>2013</year>). <article-title>Mind wandering while reading easy and difficult texts.</article-title> <source><italic>Psychon. B Rev.</italic></source> <volume>20</volume> <fpage>586</fpage>&#x2013;<lpage>592</lpage>. <pub-id pub-id-type="doi">10.3758/s13423-012-0367-y</pub-id> <pub-id pub-id-type="pmid">23288660</pub-id></citation></ref>
<ref id="B23"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fitch</surname> <given-names>W. T.</given-names></name> <name><surname>Friederici</surname> <given-names>A. D.</given-names></name></person-group> (<year>2012</year>). <article-title>Artificial grammar learning meets formal language theory: an overview.</article-title> <source><italic>Philos. Trans. Biol. Sci.</italic></source> <volume>367</volume> <fpage>1933</fpage>&#x2013;<lpage>1955</lpage>. <pub-id pub-id-type="doi">10.1098/rstb.2012.0103</pub-id> <pub-id pub-id-type="pmid">22688631</pub-id></citation></ref>
<ref id="B24"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fitch</surname> <given-names>T. W.</given-names></name> <name><surname>Martins</surname> <given-names>M. D.</given-names></name></person-group> (<year>2014</year>). <article-title>Hierarchical processing in music, language, and action: lashley revisited: music, language, and action hierarchical processing.</article-title> <source><italic>Ann. N Y Acad. Sci.</italic></source> <volume>1316</volume> <fpage>87</fpage>&#x2013;<lpage>104</lpage>. <pub-id pub-id-type="doi">10.1111/nyas.12406</pub-id> <pub-id pub-id-type="pmid">24697242</pub-id></citation></ref>
<ref id="B25"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Goller</surname> <given-names>F.</given-names></name> <name><surname>Choi</surname> <given-names>S.</given-names></name> <name><surname>Hong</surname> <given-names>U.</given-names></name> <name><surname>Ansorge</surname> <given-names>U.</given-names></name></person-group> (<year>2020</year>). <article-title>Whereof one cannot speak: how language and capture of visual attention interact.</article-title> <source><italic>Cognition</italic></source> <volume>194</volume>:<issue>Article104023</issue>. <pub-id pub-id-type="doi">10.1016/j.cognition.2019.104023</pub-id> <pub-id pub-id-type="pmid">31445296</pub-id></citation></ref>
<ref id="B26"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Goller</surname> <given-names>F.</given-names></name> <name><surname>Lee</surname> <given-names>D.</given-names></name> <name><surname>Ansorge</surname> <given-names>U.</given-names></name> <name><surname>Choi</surname> <given-names>S.</given-names></name></person-group> (<year>2017</year>). <article-title>Effects of language background on gaze behavior: a crosslinguistic comparison between Korean and German speakers.</article-title> <source><italic>Adv. Cogn. Psychol.</italic></source> <volume>13</volume> <fpage>267</fpage>&#x2013;<lpage>279</lpage>. <pub-id pub-id-type="doi">10.5709/acp-0227-z</pub-id> <pub-id pub-id-type="pmid">29362644</pub-id></citation></ref>
<ref id="B27"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Horstmann</surname> <given-names>G.</given-names></name> <name><surname>Ansorge</surname> <given-names>U.</given-names></name></person-group> (<year>2016</year>). <article-title>Surprise capture and inattentional blindness.</article-title> <source><italic>Cognition</italic></source> <volume>157</volume> <fpage>237</fpage>&#x2013;<lpage>249</lpage>. <pub-id pub-id-type="doi">10.1016/j.cognition.2016.09.005</pub-id> <pub-id pub-id-type="pmid">27665396</pub-id></citation></ref>
<ref id="B28"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Itti</surname> <given-names>L.</given-names></name> <name><surname>Koch</surname> <given-names>C.</given-names></name> <name><surname>Niebur</surname> <given-names>E.</given-names></name></person-group> (<year>1998</year>). <article-title>A model of saliency-based visual attention for rapid scene analysis.</article-title> <source><italic>IEEE Trans. Pattern Anal. Mach. Intell.</italic></source> <volume>20</volume> <fpage>1254</fpage>&#x2013;<lpage>1259</lpage>. <pub-id pub-id-type="doi">10.1109/34.730558</pub-id></citation></ref>
<ref id="B29"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jager</surname> <given-names>G.</given-names></name> <name><surname>Rogers</surname> <given-names>J.</given-names></name></person-group> (<year>2012</year>). <article-title>Formal language theory: refining the Chomsky hierarchy.</article-title> <source><italic>Philos. Trans. Biol. Sci.</italic></source> <volume>367</volume> <fpage>1956</fpage>&#x2013;<lpage>1970</lpage>. <pub-id pub-id-type="doi">10.1098/rstb.2012.0077</pub-id> <pub-id pub-id-type="pmid">22688632</pub-id></citation></ref>
<ref id="B30"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Johnson</surname> <given-names>A.</given-names></name></person-group> (<year>2013</year>). &#x201C;<article-title>Procedural memory and skill acquisition</article-title>,&#x201D; in <source><italic>Handbook of Psychology: Experimental Psychology</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>Healy</surname> <given-names>A. F.</given-names></name> <name><surname>Proctor</surname> <given-names>R. W.</given-names></name> <name><surname>Weiner</surname> <given-names>I. B.</given-names></name></person-group> (<publisher-loc>Hoboken, NJ</publisher-loc>: <publisher-name>John Wiley</publisher-name>), <fpage>495</fpage>&#x2013;<lpage>520</lpage>.</citation></ref>
<ref id="B31"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Judd</surname> <given-names>T.</given-names></name> <name><surname>Ehinger</surname> <given-names>K.</given-names></name> <name><surname>Durand</surname> <given-names>F.</given-names></name> <name><surname>Torralba</surname> <given-names>A.</given-names></name></person-group> (<year>2009</year>). &#x201C;<article-title>Learning to predict where humans look</article-title>,&#x201D; in <source><italic>Proceedings of 2009 IEEE 12th international Conference on Computer Vision</italic></source>, (<publisher-loc>Kyoto</publisher-loc>: <publisher-name>IEEE</publisher-name>), <fpage>2106</fpage>&#x2013;<lpage>2113</lpage>. <pub-id pub-id-type="doi">10.1109/ICCV.2009.5459462</pub-id></citation></ref>
<ref id="B32"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Keele</surname> <given-names>S. W.</given-names></name> <name><surname>Ivry</surname> <given-names>R.</given-names></name> <name><surname>Mayr</surname> <given-names>U.</given-names></name> <name><surname>Hazeltine</surname> <given-names>E.</given-names></name> <name><surname>Heuer</surname> <given-names>H.</given-names></name></person-group> (<year>2003</year>). <article-title>The cognitive and neural architecture of sequence representation.</article-title> <source><italic>Psychol. Rev.</italic></source> <volume>110</volume> <fpage>316</fpage>&#x2013;<lpage>339</lpage>. <pub-id pub-id-type="doi">10.1037/0033-295X.110.2.316</pub-id> <pub-id pub-id-type="pmid">12747526</pub-id></citation></ref>
<ref id="B33"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kemmerer</surname> <given-names>D.</given-names></name></person-group> (<year>2006</year>). <article-title>The semantics of space: integrating linguistic typology and cognitive neuroscience.</article-title> <source><italic>Neuropsychologia</italic></source> <volume>44</volume> <fpage>1607</fpage>&#x2013;<lpage>1621</lpage>. <pub-id pub-id-type="doi">10.1016/j.neuropsychologia.2006.01.025</pub-id> <pub-id pub-id-type="pmid">16516934</pub-id></citation></ref>
<ref id="B34"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Knott</surname> <given-names>A.</given-names></name></person-group> (<year>2012</year>). <source><italic>Sensorimotor Cognition and Natural Language Syntax.</italic></source> <publisher-loc>Cambridge, MA</publisher-loc>: <publisher-name>MIT Press</publisher-name>.</citation></ref>
<ref id="B35"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kotz</surname> <given-names>S. A.</given-names></name> <name><surname>Schwartze</surname> <given-names>M.</given-names></name></person-group> (<year>2010</year>). <article-title>Cortical speech processing unplugged: a timely subcortico-cortical framework.</article-title> <source><italic>Trends Cogn. Sci.</italic></source> <volume>14</volume> <fpage>392</fpage>&#x2013;<lpage>399</lpage>. <pub-id pub-id-type="doi">10.1016/j.tics.2010.06.005</pub-id> <pub-id pub-id-type="pmid">20655802</pub-id></citation></ref>
<ref id="B36"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lucy</surname> <given-names>J. A.</given-names></name></person-group> (<year>2016</year>). <article-title>Recent advances in the study of linguistic relativity in historical context: a critical assessment.</article-title> <source><italic>Lang. Learn.</italic></source> <volume>66</volume> <fpage>487</fpage>&#x2013;<lpage>515</lpage>. <pub-id pub-id-type="doi">10.1111/lang.12195</pub-id></citation></ref>
<ref id="B37"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lum</surname> <given-names>J. A. G.</given-names></name> <name><surname>Conti-Ramsden</surname> <given-names>G.</given-names></name> <name><surname>Page</surname> <given-names>D.</given-names></name> <name><surname>Ullman</surname> <given-names>M. T.</given-names></name></person-group> (<year>2012</year>). <article-title>Working, declarative and procedural memory in specific language impairment.</article-title> <source><italic>Cortex</italic></source> <volume>48</volume> <fpage>1138</fpage>&#x2013;<lpage>1154</lpage>. <pub-id pub-id-type="doi">10.1016/j.cortex.2011.06.001</pub-id> <pub-id pub-id-type="pmid">21774923</pub-id></citation></ref>
<ref id="B38"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lupyan</surname> <given-names>G.</given-names></name></person-group> (<year>2012</year>). <article-title>Linguistically modulated perception and cognition: the label-feedback hypothesis.</article-title> <source><italic>Front. Psychol.</italic></source> <volume>3</volume>:<issue>Article54</issue>. <pub-id pub-id-type="doi">10.3389/fpsyg.2012.00054</pub-id> <pub-id pub-id-type="pmid">22408629</pub-id></citation></ref>
<ref id="B39"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lupyan</surname> <given-names>G.</given-names></name> <name><surname>Rahman</surname> <given-names>R. A.</given-names></name> <name><surname>Boroditsky</surname> <given-names>L.</given-names></name> <name><surname>Clark</surname> <given-names>A.</given-names></name></person-group> (<year>2020</year>). <article-title>Effects of language on visual perception.</article-title> <source><italic>Trends Cogn. Sci.</italic></source> <volume>24</volume> <fpage>930</fpage>&#x2013;<lpage>944</lpage>. <pub-id pub-id-type="doi">10.1016/j.tics.2020.08.005</pub-id> <pub-id pub-id-type="pmid">33012687</pub-id></citation></ref>
<ref id="B40"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mack</surname> <given-names>A.</given-names></name> <name><surname>Rock</surname> <given-names>I.</given-names></name></person-group> (<year>1998</year>). &#x201C;<article-title>Inattentional blindness: perception without attention</article-title>,&#x201D; in <source><italic>Visual Attention</italic></source>, <role>ed.</role> <person-group person-group-type="editor"><name><surname>Wright</surname> <given-names>R. D.</given-names></name></person-group> (<publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>), <fpage>55</fpage>&#x2013;<lpage>76</lpage>.</citation></ref>
<ref id="B41"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Majid</surname> <given-names>A.</given-names></name> <name><surname>Bowerman</surname> <given-names>M.</given-names></name> <name><surname>Kita</surname> <given-names>S.</given-names></name> <name><surname>Haun</surname> <given-names>D. B.</given-names></name> <name><surname>Levinson</surname> <given-names>S. C.</given-names></name></person-group> (<year>2004</year>). <article-title>Can language restructure cognition? the case for space.</article-title> <source><italic>Trends Cogn. Sci.</italic></source> <volume>8</volume> <fpage>108</fpage>&#x2013;<lpage>114</lpage>. <pub-id pub-id-type="doi">10.1016/j.tics.2004.01.003</pub-id> <pub-id pub-id-type="pmid">15301750</pub-id></citation></ref>
<ref id="B42"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Neisser</surname> <given-names>U.</given-names></name> <name><surname>Becklen</surname> <given-names>R.</given-names></name></person-group> (<year>1975</year>). <article-title>Selective looking: attending to visually specified events.</article-title> <source><italic>Cogn. Psychol.</italic></source> <volume>7</volume> <fpage>480</fpage>&#x2013;<lpage>494</lpage>. <pub-id pub-id-type="doi">10.1016/0010-0285(75)90019-5</pub-id></citation></ref>
<ref id="B43"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Newell</surname> <given-names>A.</given-names></name> <name><surname>Rosenbloom</surname> <given-names>P. S.</given-names></name></person-group> (<year>1981</year>). &#x201C;<article-title>Mechanisms of skill acquisition and the law of practice</article-title>,&#x201D; in <source><italic>Cognitive Skills and their Acquisition</italic></source>, <role>ed.</role> <person-group person-group-type="editor"><name><surname>Anderson</surname> <given-names>J. R.</given-names></name></person-group> (<publisher-loc>Hillsdale, NJ</publisher-loc>: <publisher-name>Erlbaum</publisher-name>), <fpage>1</fpage>&#x2013;<lpage>55</lpage>.</citation></ref>
<ref id="B44"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Nisbett</surname> <given-names>R. E.</given-names></name> <name><surname>Miyamoto</surname> <given-names>Y.</given-names></name></person-group> (<year>2005</year>). <article-title>The influence of culture: holistic versus analytic perception.</article-title> <source><italic>Trends Cogn. Sci.</italic></source> <volume>9</volume> <fpage>467</fpage>&#x2013;<lpage>473</lpage>. <pub-id pub-id-type="doi">10.1016/j.tics.2005.08.004</pub-id> <pub-id pub-id-type="pmid">16129648</pub-id></citation></ref>
<ref id="B45"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Perrin</surname> <given-names>F.</given-names></name> <name><surname>Garc&#x0131;&#x00EC;a-Larrea</surname> <given-names>L.</given-names></name> <name><surname>Maugui&#x00E8;re</surname> <given-names>F.</given-names></name> <name><surname>Bastuji</surname> <given-names>H.</given-names></name></person-group> (<year>1999</year>). <article-title>A differential brain response to the subject&#x2019;s own name persists during sleep.</article-title> <source><italic>Clin. Neurophysiol.</italic></source> <volume>110</volume> <fpage>2153</fpage>&#x2013;<lpage>2164</lpage>. <pub-id pub-id-type="doi">10.1016/S1388-2457(99)00177-7</pub-id></citation></ref>
<ref id="B46"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Piaget</surname> <given-names>J.</given-names></name></person-group> (<year>1954</year>). <source><italic>The Construction of Reality in the Child.</italic></source> <publisher-loc>New York, NY</publisher-loc>: <publisher-name>Routledge</publisher-name>. <pub-id pub-id-type="doi">10.1037/11168-000</pub-id></citation></ref>
<ref id="B47"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Pulverm&#x00FC;ller</surname> <given-names>F.</given-names></name> <name><surname>Shtyrov</surname> <given-names>Y.</given-names></name></person-group> (<year>2006</year>). <article-title>Language outside the focus of attention: the mismatch negativity as a tool for studying higher cognitive processes.</article-title> <source><italic>Prog. Neurobiol.</italic></source> <volume>79</volume> <fpage>49</fpage>&#x2013;<lpage>71</lpage>. <pub-id pub-id-type="doi">10.1016/j.pneurobio.2006.04.004</pub-id> <pub-id pub-id-type="pmid">16814448</pub-id></citation></ref>
<ref id="B48"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Pylyshyn</surname> <given-names>Z.</given-names></name></person-group> (<year>1999</year>). <article-title>Is vision continuous with cognition?: the case for cognitive impenetrability of visual perception.</article-title> <source><italic>Behav. Brain Sci.</italic></source> <volume>22</volume> <fpage>341</fpage>&#x2013;<lpage>365</lpage>. <pub-id pub-id-type="doi">10.1017/S0140525X99002022</pub-id> <pub-id pub-id-type="pmid">11301517</pub-id></citation></ref>
<ref id="B49"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Rayner</surname> <given-names>K.</given-names></name> <name><surname>Sereno</surname> <given-names>S. C.</given-names></name> <name><surname>Raney</surname> <given-names>G. E.</given-names></name></person-group> (<year>1996</year>). <article-title>Eye movement control in reading: a comparison of two types of models.</article-title> <source><italic>J. Exp. Psychol. Hum.</italic></source> <volume>22</volume> <fpage>1188</fpage>&#x2013;<lpage>1200</lpage>. <pub-id pub-id-type="doi">10.1037/0096-1523.22.5.1188</pub-id> <pub-id pub-id-type="pmid">8865619</pub-id></citation></ref>
<ref id="B50"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Reason</surname> <given-names>J.</given-names></name></person-group> (<year>1990</year>). <source><italic>Human Error.</italic></source> <publisher-loc>Cambridge, UK</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>.</citation></ref>
<ref id="B51"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Reichle</surname> <given-names>E. D.</given-names></name> <name><surname>Rayner</surname> <given-names>K.</given-names></name> <name><surname>Pollatsek</surname> <given-names>A.</given-names></name></person-group> (<year>2003</year>). <article-title>The EZ Reader model of eye-movement control in reading: comparisons to other models.</article-title> <source><italic>Behav. Brain Sci.</italic></source> <volume>26</volume> <fpage>445</fpage>&#x2013;<lpage>526</lpage>. <pub-id pub-id-type="doi">10.1017/S0140525X03290105</pub-id></citation></ref>
<ref id="B52"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>R&#x00F6;er</surname> <given-names>J. P.</given-names></name> <name><surname>Bell</surname> <given-names>R.</given-names></name> <name><surname>Buchner</surname> <given-names>A.</given-names></name></person-group> (<year>2013</year>). <article-title>Self-relevance increases the irrelevant sound effect: attentional disruption by one&#x2019;s own name.</article-title> <source><italic>J. Cogn. Psychol.</italic></source> <volume>25</volume> <fpage>925</fpage>&#x2013;<lpage>931</lpage>. <pub-id pub-id-type="doi">10.1080/20445911.2013.828063</pub-id></citation></ref>
<ref id="B53"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sapir</surname> <given-names>E.</given-names></name></person-group> (<year>1941/1964</year>). <source><italic>Culture, Language and Personality.</italic></source> <publisher-loc>Berkley, CA</publisher-loc>: <publisher-name>University of California Press</publisher-name>.</citation></ref>
<ref id="B54"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Scharlau</surname> <given-names>I.</given-names></name></person-group> (<year>2002</year>). <article-title>Leading, but not trailing, primes influence temporal order perception: further evidence for an attentional account of perceptual latency priming.</article-title> <source><italic>Percept. Psychophys.</italic></source> <volume>64</volume> <fpage>1346</fpage>&#x2013;<lpage>1360</lpage>. <pub-id pub-id-type="doi">10.3758/BF03194777</pub-id> <pub-id pub-id-type="pmid">12519031</pub-id></citation></ref>
<ref id="B55"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Scharlau</surname> <given-names>I.</given-names></name></person-group> (<year>2004</year>). <article-title>Evidence against response bias in temporal order tasks with attention manipulation by masked primes.</article-title> <source><italic>Psychol. Res. Psych. Fo</italic></source> <volume>68</volume> <fpage>224</fpage>&#x2013;<lpage>236</lpage>. <pub-id pub-id-type="doi">10.1007/s00426-003-0135-8</pub-id> <pub-id pub-id-type="pmid">12827351</pub-id></citation></ref>
<ref id="B56"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Scharlau</surname> <given-names>I.</given-names></name> <name><surname>Ansorge</surname> <given-names>U.</given-names></name></person-group> (<year>2003</year>). <article-title>Direct parameter specification of an attention shift: evidence from perceptual latency priming.</article-title> <source><italic>Vision Res.</italic></source> <volume>43</volume> <fpage>1351</fpage>&#x2013;<lpage>1363</lpage>. <pub-id pub-id-type="doi">10.1016/s0042-6989(03)00141-x</pub-id> <pub-id pub-id-type="pmid">12742105</pub-id></citation></ref>
<ref id="B57"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Scharlau</surname> <given-names>I.</given-names></name> <name><surname>Neumann</surname> <given-names>O.</given-names></name></person-group> (<year>2003</year>). <article-title>Perceptual latency priming by masked and unmasked stimuli: evidence for an attentional interpretation.</article-title> <source><italic>Psychol. Res. Psych. Fo</italic></source> <volume>67</volume> <fpage>184</fpage>&#x2013;<lpage>196</lpage>. <pub-id pub-id-type="doi">10.1007/s00426-002-0116-3</pub-id> <pub-id pub-id-type="pmid">12955508</pub-id></citation></ref>
<ref id="B58"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sereno</surname> <given-names>S. C.</given-names></name> <name><surname>Rayner</surname> <given-names>K.</given-names></name></person-group> (<year>2003</year>). <article-title>Measuring word recognition in reading: eye movements and event-related potentials.</article-title> <source><italic>Trends Cogn. Sci.</italic></source> <volume>7</volume> <fpage>489</fpage>&#x2013;<lpage>493</lpage>. <pub-id pub-id-type="doi">10.1016/j.tics.2003.09.010</pub-id> <pub-id pub-id-type="pmid">14585445</pub-id></citation></ref>
<ref id="B59"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Simons</surname> <given-names>D. J.</given-names></name> <name><surname>Chabris</surname> <given-names>C. F.</given-names></name></person-group> (<year>1999</year>). <article-title>Gorillas in our midst: sustained inattentional blindness for dynamic events.</article-title> <source><italic>Perception</italic></source> <volume>28</volume> <fpage>1059</fpage>&#x2013;<lpage>1074</lpage>. <pub-id pub-id-type="doi">10.1068/p281059</pub-id> <pub-id pub-id-type="pmid">10694957</pub-id></citation></ref>
<ref id="B60"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Slobin</surname> <given-names>D. I.</given-names></name></person-group> (<year>1996</year>). &#x201C;<article-title>From &#x201C;thought and language&#x201D; to &#x201C;thinking for speaking</article-title>,&#x201D; in <source><italic>Rethinking Linguistic Relativity</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>Gumperz</surname> <given-names>J. J.</given-names></name> <name><surname>Levinson</surname> <given-names>S. C.</given-names></name></person-group> (<publisher-loc>Cambridge, MA</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>), <fpage>70</fpage>&#x2013;<lpage>96</lpage>.</citation></ref>
<ref id="B61"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Slobin</surname> <given-names>D. I.</given-names></name></person-group> (<year>2000</year>). &#x201C;<article-title>Verbalized events: a dynamic approach to linguistics relativity and determinism</article-title>,&#x201D; in <source><italic>Evidence for Linguistic Relativity</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>Niemeier</surname> <given-names>S.</given-names></name> <name><surname>Dirven</surname> <given-names>R.</given-names></name></person-group> (<publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins</publisher-name>), <fpage>107</fpage>&#x2013;<lpage>138</lpage>. <pub-id pub-id-type="doi">10.1075/cilt.198.10slo</pub-id> <pub-id pub-id-type="pmid">33486653</pub-id></citation></ref>
<ref id="B62"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Slobin</surname> <given-names>D. I.</given-names></name></person-group> (<year>2003</year>). &#x201C;<article-title>Language and thought online: cognitive consequences of linguistic relativity</article-title>,&#x201D; in <source><italic>Language in Mind</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>Gentner</surname> <given-names>D.</given-names></name> <name><surname>Goldin-Meadow</surname> <given-names>S.</given-names></name></person-group> (<publisher-loc>Cambridge, MA</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>), <fpage>157</fpage>&#x2013;<lpage>192</lpage>.</citation></ref>
<ref id="B63"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Smith</surname> <given-names>L.</given-names></name> <name><surname>Yu</surname> <given-names>C.</given-names></name></person-group> (<year>2008</year>). <article-title>Infants rapidly learn word-referent mappings via cross-situational statistics.</article-title> <source><italic>Cognition</italic></source> <volume>106</volume> <fpage>1558</fpage>&#x2013;<lpage>1568</lpage>. <pub-id pub-id-type="doi">10.1016/j.cognition.2007.06.010</pub-id> <pub-id pub-id-type="pmid">17692305</pub-id></citation></ref>
<ref id="B64"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Stark</surname> <given-names>R. E.</given-names></name></person-group> (<year>1980</year>). &#x201C;<article-title>Stages of speech development in the first year of life</article-title>,&#x201D; in <source><italic>Child Phonology</italic></source>, <role>ed.</role> <person-group person-group-type="editor"><name><surname>Stark</surname> <given-names>R. E.</given-names></name></person-group> (<publisher-loc>London, UK</publisher-loc>: <publisher-name>Academic Press</publisher-name>), <fpage>73</fpage>&#x2013;<lpage>92</lpage>. <pub-id pub-id-type="doi">10.1016/B978-0-12-770601-6.50010-3</pub-id></citation></ref>
<ref id="B65"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Theeuwes</surname> <given-names>J.</given-names></name></person-group> (<year>1992</year>). <article-title>Perceptual selectivity for color and form.</article-title> <source><italic>Percept. Psychophys.</italic></source> <volume>51</volume> <fpage>599</fpage>&#x2013;<lpage>606</lpage>. <pub-id pub-id-type="doi">10.3758/BF03211656</pub-id> <pub-id pub-id-type="pmid">1620571</pub-id></citation></ref>
<ref id="B66"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Thierry</surname> <given-names>G.</given-names></name> <name><surname>Athanasopoulos</surname> <given-names>P.</given-names></name> <name><surname>Wiggett</surname> <given-names>A.</given-names></name> <name><surname>Dering</surname> <given-names>B.</given-names></name> <name><surname>Kuipers</surname> <given-names>J.-R.</given-names></name></person-group> (<year>2009</year>). <article-title>Unconscious effects of language-specific terminology on preattentive color perception.</article-title> <source><italic>Proc. Natl. Sci. U.S.A.</italic></source> <volume>106</volume> <fpage>4567</fpage>&#x2013;<lpage>4570</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.0811155106</pub-id> <pub-id pub-id-type="pmid">19240215</pub-id></citation></ref>
<ref id="B67"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Titchener</surname> <given-names>E. M.</given-names></name></person-group> (<year>1908</year>). <source><italic>Lectures on the Elementary Psychology of Feeling and Attention.</italic></source> <publisher-loc>London</publisher-loc>: <publisher-name>MacMillan</publisher-name>.</citation></ref>
<ref id="B68"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tomasello</surname> <given-names>M.</given-names></name> <name><surname>Farrar</surname> <given-names>M. J.</given-names></name></person-group> (<year>1986</year>). <article-title>Joint attention and early language.</article-title> <source><italic>Child Dev.</italic></source> <volume>57</volume> <fpage>1454</fpage>&#x2013;<lpage>1463</lpage>. <pub-id pub-id-type="doi">10.2307/1130423</pub-id></citation></ref>
<ref id="B69"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tomasello</surname> <given-names>M.</given-names></name> <name><surname>Kruger</surname> <given-names>A. C.</given-names></name></person-group> (<year>1992</year>). <article-title>Joint attention on actions: acquiring verbs in ostensive and non-ostensive and non-ostensive contexts.</article-title> <source><italic>J. Child Lang.</italic></source> <volume>19</volume> <fpage>311</fpage>&#x2013;<lpage>333</lpage>. <pub-id pub-id-type="doi">10.1017/S0305000900011430</pub-id> <pub-id pub-id-type="pmid">1527205</pub-id></citation></ref>
<ref id="B70"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ullman</surname> <given-names>M. T.</given-names></name></person-group> (<year>2001</year>). <article-title>The neural basis of lexicon and grammar in first and second language: the declarative/procedural model.</article-title> <source><italic>Biling-Lang. Cogn.</italic></source> <volume>4</volume> <fpage>105</fpage>&#x2013;<lpage>122</lpage>. <pub-id pub-id-type="doi">10.1017/S1366728901000220</pub-id></citation></ref>
<ref id="B71"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Weichselbaum</surname> <given-names>H.</given-names></name> <name><surname>Ansorge</surname> <given-names>U.</given-names></name></person-group> (<year>2018</year>). <article-title>Bottom-up attention capture with distractor and target singletons defined in the same (color) dimension is not a matter of feature uncertainty.</article-title> <source><italic>Atten. Percept. Psycho.</italic></source> <volume>80</volume> <fpage>1350</fpage>&#x2013;<lpage>1361</lpage>. <pub-id pub-id-type="doi">10.3758/s13414-018-1538-3</pub-id> <pub-id pub-id-type="pmid">29777515</pub-id></citation></ref>
<ref id="B72"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Whorf</surname> <given-names>B. L.</given-names></name></person-group> (<year>1956</year>). <source><italic>Language, Thought, and Reality.</italic></source> <publisher-loc>Cambridge, MA</publisher-loc>: <publisher-name>MIT Press</publisher-name>.</citation></ref>
<ref id="B73"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Xu</surname> <given-names>F.</given-names></name></person-group> (<year>2019</year>). <article-title>Towards a rational constructivist theory of cognitive development.</article-title> <source><italic>Psychol. Rev.</italic></source> <volume>126</volume> <fpage>841</fpage>&#x2013;<lpage>864</lpage>. <pub-id pub-id-type="doi">10.1037/rev0000153</pub-id> <pub-id pub-id-type="pmid">31180701</pub-id></citation></ref>
<ref id="B74"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Yun</surname> <given-names>H.</given-names></name> <name><surname>Choi</surname> <given-names>S.</given-names></name></person-group> (<year>2018</year>). <article-title>Spatial semantics, cognition, and their interaction: a comparative study of spatial categorization in English and Korean.</article-title> <source><italic>Cogn. Sci.</italic></source> <volume>42</volume> <fpage>1736</fpage>&#x2013;<lpage>1776</lpage>. <pub-id pub-id-type="doi">10.1111/cogs.12622</pub-id> <pub-id pub-id-type="pmid">29790181</pub-id></citation></ref>
</ref-list>
</back>
</article>