<?xml version="1.0" encoding="UTF-8" standalone="no"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Psychol.</journal-id>
<journal-title>Frontiers in Psychology</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Psychol.</abbrev-journal-title>
<issn pub-type="epub">1664-1078</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fpsyg.2018.01211</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Psychology</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Adult Learning of Novel Words in a Non-native Language: Consonants, Vowels, and Tones</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name><surname>Poltrock</surname> <given-names>Silvana</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="aff" rid="aff2"><sup>2</sup></xref>
<xref ref-type="aff" rid="aff3"><sup>3</sup></xref>
<xref ref-type="corresp" rid="c001"><sup>&#x0002A;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/486710/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Chen</surname> <given-names>Hui</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="aff" rid="aff2"><sup>2</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/585876/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Kwok</surname> <given-names>Celia</given-names></name>
<xref ref-type="aff" rid="aff4"><sup>4</sup></xref>
</contrib>
<contrib contrib-type="author">
<name><surname>Cheung</surname> <given-names>Hintat</given-names></name>
<xref ref-type="aff" rid="aff4"><sup>4</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/585808/overview"/>
</contrib>
<contrib contrib-type="author" corresp="yes">
<name><surname>Nazzi</surname> <given-names>Thierry</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="aff" rid="aff2"><sup>2</sup></xref>
<xref ref-type="corresp" rid="c002"><sup>&#x0002A;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/10914/overview"/>
</contrib>
</contrib-group>
<aff id="aff1"><sup>1</sup><institution>Universit&#x000E9; Paris Descartes, Sorbonne Paris Cit&#x000E9;</institution>, <addr-line>Paris</addr-line>, <country>France</country></aff>
<aff id="aff2"><sup>2</sup><institution>CNRS, Laboratoire Psychologie de la Perception</institution>, <addr-line>Paris</addr-line>, <country>France</country></aff>
<aff id="aff3"><sup>3</sup><institution>Department Linguistik, Universit&#x000E4;t Potsdam</institution>, <addr-line>Potsdam</addr-line>, <country>Germany</country></aff>
<aff id="aff4"><sup>4</sup><institution>Department of Linguistics and Modern Language Studies, The Education University of Hong Kong</institution>, <addr-line>Tai Po</addr-line>, <country>Hong Kong</country></aff>
<author-notes>
<fn fn-type="edited-by"><p>Edited by: Jessica Hay, University of Tennessee, Knoxville, United States</p></fn>
<fn fn-type="edited-by"><p>Reviewed by: Mariapaola D&#x00027;Imperio, Aix-Marseille Universit&#x000E9;, France; Mireille Besson, Institut de Neurosciences Cognitives de la M&#x000E9;diterran&#x000E9;e (INCM), France; Aaron D. Mitchel, Bucknell University, United States</p></fn>
<corresp id="c001">&#x0002A;Correspondence: Silvana Poltrock <email>poltrocks&#x00040;gmail.com</email></corresp>
<corresp id="c002">Thierry Nazzi <email>thierry.nazzi&#x00040;parisdescartes.fr</email></corresp>
<fn fn-type="other" id="fn001"><p>This article was submitted to Language Sciences, a section of the journal Frontiers in Psychology</p></fn></author-notes>
<pub-date pub-type="epub">
<day>24</day>
<month>07</month>
<year>2018</year>
</pub-date>
<pub-date pub-type="collection">
<year>2018</year>
</pub-date>
<volume>9</volume>
<elocation-id>1211</elocation-id>
<history>
<date date-type="received">
<day>16</day>
<month>10</month>
<year>2017</year>
</date>
<date date-type="accepted">
<day>26</day>
<month>06</month>
<year>2018</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x000A9; 2018 Poltrock, Chen, Kwok, Cheung and Nazzi.</copyright-statement>
<copyright-year>2018</copyright-year>
<copyright-holder>Poltrock, Chen, Kwok, Cheung and Nazzi</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/"><p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p></license>
</permissions>
<abstract><p>While words are distinguished primarily by consonants and vowels in many languages, tones are also used in the majority of the world&#x00027;s languages to cue lexical contrasts. However, studies on novel word learning have largely concentrated on consonants and vowels. To shed more light on the use of tonal information in novel word learning and its relationship with the development of phonological categories, the present study explored how adults&#x00027; ability to learn minimal pair pseudowords in a tone language is modulated by their native phonological knowledge. Twenty-four adult speakers of three languages were tested: Cantonese, Mandarin, and French. Eye-tracking was used to record eye movements of these learners, while they were watching animated cartoons in Cantonese. On each trial, adults had to learn two new label-object associations, while the labels differed minimally by a consonant, a vowel, or a tone. Learning would therefore attest to participants&#x00027; ability to use phonological information to distinguish the paired words. Results first revealed that adult learners in each language group performed better than chance in all conditions. Moreover, compared to native Cantonese adults, both Mandarin- and French-speaking adults performed worse on all three contrasts. In addition, French adults were worse on tones when compared to Mandarin adults. Lastly, no advantage for consonantal information in native lexical processing was found for Cantonese-speaking adults as predicted by the &#x0201C;division of labor&#x0201D; proposal, thus confirming crosslinguistic differences in consonant/vowel weight between speakers of tonal vs. non-tonal languages. These findings establish rapid novel word learning in a non-native language (long-term learning will have to be further assessed), modulated by native phonological knowledge. The implications of the findings of this adult study for further infant word learning studies are discussed.</p></abstract>
<kwd-group>
<kwd>word learning</kwd>
<kwd>minimal pairs</kwd>
<kwd>non-native speech perception</kwd>
<kwd>tones</kwd>
<kwd>adults</kwd>
</kwd-group>
<counts>
<fig-count count="6"/>
<table-count count="3"/>
<equation-count count="0"/>
<ref-count count="80"/>
<page-count count="15"/>
<word-count count="12107"/>
</counts>
</article-meta>
</front>
<body>
<sec sec-type="intro" id="s1">
<title>Introduction</title>
<p>Learning words is a crucial step in learning a language, no matter whether it is one&#x00027;s initial, native language as an infant, or a new, non-native language as an adult. Importantly, learning new words requires the ability to process relevant phonetic information and represent it in proper phonological categories. This ability is largely based on which phonetic variations are relevant to word meaning and how phonological categories are established in one&#x00027;s native language. In all languages, lexical representations include segmental information related to the identity of consonants and vowels constituting the word forms. In many languages, lexical representations also include suprasegmental information, such as lexical stress, pitch accents or tones. Phonological repertoires vary across languages, and the same is true for lexically-relevant prosodic information. For instance, both Cantonese and Mandarin have tonal systems, but their systems differ in both number and identity of tones (Wang, <xref ref-type="bibr" rid="B75">1963</xref>; Mandarin: Cheng, <xref ref-type="bibr" rid="B16">1966</xref>; Hashimoto, <xref ref-type="bibr" rid="B39">1972</xref>; Howie, <xref ref-type="bibr" rid="B44">1976</xref>; Cantonese: Bauer and Benedict, <xref ref-type="bibr" rid="B4">1997</xref>; Duanmu, <xref ref-type="bibr" rid="B25">2000</xref>; Yip, <xref ref-type="bibr" rid="B80">2002</xref>), as illustrated in Figure <xref ref-type="fig" rid="F1">1</xref>. The present study will explore the interplay between word learning and phonological processing (of consonants, vowels and tones) by comparing three groups of adults when learning minimal pairs of Cantonese words in their native (Cantonese-speaking adults) vs. a non-native (Mandarin- and French-speaking adults) language.</p>
<fig id="F1" position="float">
<label>Figure 1</label>
<caption><p>F0 contours for tones in Hong Kong Cantonese (<bold>left</bold>, produced by a male speaker on the syllable &#x00027;ji&#x00027; /ji/) and in Beijing Mandarin (<bold>right</bold>, produced by a male speaker on the syllable&#x00027;da&#x00027; /ta/). Note that, because these Cantonese and Mandarin tones come from different speakers, absolute F0 levels cannot be compared across languages, only relative levels are comparable. Images come from a previously published article (Francis et al., <xref ref-type="bibr" rid="B29">2008</xref>), reproduced here with the authors&#x00027; permission.</p></caption>
<graphic xlink:href="fpsyg-09-01211-g0001.tif"/>
</fig>
<p>Decades of research have established that speech perception becomes language-specific during the first year of life, as attested by decreases in ability to discriminate many (though not all, see below) non-native phonological contrasts (that is, contrasts not used in one&#x00027;s native language), and increases in the ability to discriminate native contrasts. These changes have been found to happen later for consonants (by 10&#x02013;12 months of age, e.g., Werker and Tees, <xref ref-type="bibr" rid="B76">1984a</xref>; Best et al., <xref ref-type="bibr" rid="B7">1988</xref>; Rivera-Gaxiola et al., <xref ref-type="bibr" rid="B65">2005</xref>), than for vowels (by 6 months of age, Kuhl et al., <xref ref-type="bibr" rid="B46">1992</xref>; Polka and Werker, <xref ref-type="bibr" rid="B62">1994</xref>). The developmental timing for tones is less clear, as changes are usually reported around 10 months of age (Mattock and Burnham, <xref ref-type="bibr" rid="B50">2006</xref>; Mattock et al., <xref ref-type="bibr" rid="B51">2008</xref>; Liu and Kager, <xref ref-type="bibr" rid="B48">2014</xref>; Cabrera et al., <xref ref-type="bibr" rid="B13">2015</xref>) although evidence for changes as early as 4 months has been found in one study (Yeung et al., <xref ref-type="bibr" rid="B79">2013</xref>).</p>
<p>These developmental changes in speech perception, which attest to the early acquisition of the phonological repertoire of the native language, have continued effects in adulthood. Speech perception difficulties have been found for the processing of non-native consonants (e.g., Werker and Tees, <xref ref-type="bibr" rid="B77">1984b</xref>), non-native vowels (e.g., Polka, <xref ref-type="bibr" rid="B61">1995</xref>), and non-native tones (Gandour et al., <xref ref-type="bibr" rid="B31">2000</xref>; Hall&#x000E9; et al., <xref ref-type="bibr" rid="B35">2004</xref>; So and Best, <xref ref-type="bibr" rid="B69">2010</xref>). This is attested by the fact that adults will sometimes have difficulties identifying some non-native sounds, and/or discriminating between non-native sounds. For example, regarding consonant perception, the fact that the English /r/ consonant does not have an equivalent in both Japanese and German has been found to lead to differences in how Japanese- and German-speaking adults perceive this sound (in contrast to English /l/) when compared with English-speaking adults (e.g., Miyawaki et al., <xref ref-type="bibr" rid="B53">1975</xref>; Iverson et al., <xref ref-type="bibr" rid="B45">2003</xref>). Moreover, perception of this non-native contrast differs across the two language groups, with more difficulty observed for the Japanese-speaking adults, who appear to form only one sound category, compared to the German-speaking adults who perceive two sound categories (e.g., Iverson et al., <xref ref-type="bibr" rid="B45">2003</xref>). This further shows that these processing difficulties stem from interference with the native phonological system. With respect to tones, many studies have found that speakers of languages that do not use tone contrasts at the lexical level identify and discriminate non-native tones with more difficulty than speakers of tonal languages (Gandour et al., <xref ref-type="bibr" rid="B31">2000</xref>; Hall&#x000E9; et al., <xref ref-type="bibr" rid="B35">2004</xref>; So and Best, <xref ref-type="bibr" rid="B69">2010</xref>). Even though some discrimination ability is found in speakers of non-tonal languages, they appear to process tones differently. This is attested, for example, by the fact that (non-tonal) French-speaking adults, while being able to discriminate non-native Mandarin tones, perceive these tones less categorically than Mandarin-speaking adults. Some have proposed to link this to their lack of phonological categories for tones (Hall&#x000E9; et al., <xref ref-type="bibr" rid="B35">2004</xref>).</p>
<p>The first goal of the present study was thus to explore the effects of such perceptual changes on adults&#x00027; ability to learn new words in an unfamiliar language. Although this is a situation that adults have to cope with when learning a new language, it has received surprisingly little attention to this day (but see Chandrasekaran et al., <xref ref-type="bibr" rid="B15">2010</xref>; Cooper and Wang, <xref ref-type="bibr" rid="B17">2012</xref>, <xref ref-type="bibr" rid="B18">2013</xref>, for training studies on English-speaking adults&#x00027; processing of Cantonese or Mandarin tones in lexical contexts). Here, we evaluated monolingually-raised Mandarin- and French-speaking adults&#x00027; ability to learn new words in Cantonese, and compared their performance to baseline data from Cantonese-speaking adults. This was done in a laboratory setting, during which, on each trial, adults had to learn a pair of Cantonese pseudowords that differed by either a consonant, a vowel or a tone. Given the above language specialization findings at the perceptual level, evidenced by difficulties in low-level (discrimination or identification) non-native processing, we predict that adults (and infants from 6 to 12 months onward) should have more difficulty learning new words in a non-native language than in their native language, because they are made up of some sounds that do not belong to the native phonological repertoire. Hence, overall performance should be higher for Cantonese-speakers than for Mandarin- and French-speaking adults, and it might even be that the latter two groups fail at learning. Another possibility is that linguistic distance between the native language and the language of the stimuli affects performance. Since Cantonese and Mandarin, being both Sino-Tibetan languages, share many phonological, morphological, and syntactic properties (Li, <xref ref-type="bibr" rid="B47">1937</xref>; Gong, <xref ref-type="bibr" rid="B33">1980</xref>; DeLancey, <xref ref-type="bibr" rid="B21">2009</xref>), which is not the case with Cantonese and Indo-European French, Mandarin-speaking adults might have higher overall performance than French-speaking adults.</p>
<p>One important feature of our experimental design is the fact that on each trial adults had to learn a pair of new pseudowords. Therefore, for learning to take place, adults had to process the phonological contrast distinguishing the two paired words, which allowed us to explore in more detail the interplay between phonological and lexical processing in this process of acquiring new words. To begin with, the fact that the pseudowords contrasted in either consonant, vowel, or tone information allowed us to evaluate differential processing of these three phonological sound categories. This second goal of the present study was motivated by the proposal that consonants carry more information about the lexicon, whereas vowels play a more important role in syntactic and prosodic processing (Nespor et al., <xref ref-type="bibr" rid="B57">2003</xref>). For example, in word reconstruction studies in which English, Dutch, and Spanish listeners hear pseudowords and have to transform them into real words, they preserve consonantal over vocalic information, changing kebra into cobra rather than zebra (van Ooijen, <xref ref-type="bibr" rid="B74">1996</xref>; Cutler et al., <xref ref-type="bibr" rid="B20">2000</xref>). Evidence for this bias for consonantal information in lexical processing (often referred to as the C-bias in the literature) is supported by studies with adults across several non-tonal languages (Dutch, English, French, Italian, Spanish) and a variety of different methods such as word learning (e.g., Bonatti et al., <xref ref-type="bibr" rid="B11">2005</xref>; Creel et al., <xref ref-type="bibr" rid="B19">2006</xref>; Toro et al., <xref ref-type="bibr" rid="B72">2008</xref>; Havy et al., <xref ref-type="bibr" rid="B40">2014</xref>) and lexical access (e.g., New et al., <xref ref-type="bibr" rid="B58">2008</xref>; Carreiras et al., <xref ref-type="bibr" rid="B14">2009</xref>; Delle Luche et al., <xref ref-type="bibr" rid="B22">2014</xref>; New and Nazzi, <xref ref-type="bibr" rid="B59">2014</xref>).</p>
<p>Importantly though, when this project was started, little was known about the C-bias in non-European languages, and in particular in tone languages. Tone languages provide a particularly interesting test of the C-bias as lexical tones are mostly carried by vowels. This might affect performance, in two opposing ways. Indeed, the need for speakers of tone languages to attend to tones to identify words might increase their attention to the vowels (which carry them), and thus increase the weight given to vowels compared to consonants in tone languages compared to non-tone languages. This might result in a lack of bias or in a reversed advantage in processing vocalic information. In contrast, the fact that vowels carry tones might make the acoustic realization of vowels more variable in tone than in non-tone languages, making them more difficult to process and identify. If so, the consonant bias in lexical processing found in non-tone languages might be even more pronounced in tone languages.</p>
<p>To date, only two studies have explored this issue, but have focused on levels other than word learning: lexical access to known words, and word form segmentation in an artificial language. First, in a word reconstruction study based on van Ooijen (<xref ref-type="bibr" rid="B74">1996</xref>) and testing lexical access, Wiener and Turnbull (<xref ref-type="bibr" rid="B78">2016</xref>) asked participants to transform a pseudoword into a real word by changing either a consonant, a vowel (in fact, to conform to Chinese phonology, they were asked to change the final - in Chinese phonology, and in the stimuli used in that study, the final corresponds to V, VV, or VVN), a tone, or any of the three. Results show effects of condition, corresponding to the fact that Mandarin-speaking adults appear to preferentially change tones over both consonants and vowels/finals, with vowels appearing to be the less mutable sound category, contrary to what had been found in Dutch, English, and Spanish (van Ooijen, <xref ref-type="bibr" rid="B74">1996</xref>; Cutler et al., <xref ref-type="bibr" rid="B20">2000</xref>). These findings suggest a different balance in the weight given to consonants and vowels, with less weight given to consonants (or more weight given to vowels), by Mandarin-speaking adults. Second, an artificial language study exploring whether Cantonese-speaking adults use consonants or vowels (and tones) to segment a fluent speech stream revealed that they could not use consonantal information alone, but could rely either on vocalic information alone (although the difference between the consonant and vowel conditions was not significant), or more likely on a combination of vocalic and tonal information (G&#x000F3;mez et al., <xref ref-type="bibr" rid="B32">2017</xref>). This also suggests a different balance in the weight given to consonants and vowels by Cantonese-speaking adults as compared to French- or Italian-speaking adults (Bonatti et al., <xref ref-type="bibr" rid="B11">2005</xref>; Toro et al., <xref ref-type="bibr" rid="B72">2008</xref>). The present study will add to this literature by providing the first evaluation of this issue in a word learning task for tonal language speakers, for either the native language (Cantonese adults processing Cantonese stimuli) or a foreign language (Mandarin adults processing Cantonese stimuli). It will also provide the first evidence of whether the C-bias, found in French-speaking adults when processing native stimuli, would also extend to the processing of non-native stimuli (French adults processing Cantonese stimuli).</p>
<p>The results regarding the use of tonal information by French-speaking adults when learning words will also inform our understanding of the link between phonological and lexical processing. Previous studies on tone perception/identification have established that even though adult speakers of non-tonal languages have more difficulties at tone processing than speakers of tone languages (e.g., Gandour et al., <xref ref-type="bibr" rid="B31">2000</xref>; Hall&#x000E9; et al., <xref ref-type="bibr" rid="B35">2004</xref>; So and Best, <xref ref-type="bibr" rid="B69">2010</xref>), they tend to perform above chance levels. Recently, Liu and Kager (<xref ref-type="bibr" rid="B48">2014</xref>) found that tone discrimination in non-tonal Dutch-learning infants follows a U-shaped function: while 5&#x02013;6-month-olds can discriminate a Mandarin tone easily, their sensitivity declines between 8 and 15 months, but they regain sensitivity to it by 17&#x02013;18 months. They suggested that this rebound in sensitivity might be related to the acquisition of the intonation of the native language. If this increased sensitivity is not limited to low levels of processing, then the French speakers in our experiment might perform at above chance levels on the tone-contrasted trials. This prediction is supported by previous findings on English-speaking adults (Chandrasekaran et al., <xref ref-type="bibr" rid="B15">2010</xref>; Cooper and Wang, <xref ref-type="bibr" rid="B17">2012</xref>, <xref ref-type="bibr" rid="B18">2013</xref>) showing tone processing in lexical contexts, although in these studies, participants were subject to intense word training (several training sessions of about 30 min, in which some feedback was provided). It is unclear whether non-tone language speaking adults would show sensitivity to tone information in a less intensive task.</p>
<p>The present study used eyetracking to investigate the ability of Cantonese-, Mandarin-, and French-speaking adults to learn pairs of pseudowords in Cantonese, while processing fine phonetic information (consonant vs. vowel vs. tone information), whether used contrastively in the native language or not. On each of 24 trials, adults saw a pair of cartoons. In each cartoon, an unfamiliar object was presented visually while 6 sentences in Cantonese, each containing a pseudoword labeling that object, were heard. Between the two cartoons, the pseudowords differed by either a consonant (8 times), a vowel (8 times) or a tone (8 times). Adults were then tested on whether they had been able to learn the words following this short word learning phase, by presenting them with the two unfamiliar objects side-by-side, and observing their pattern of object looking before (prenaming phase) and after (postnaming phase) one of the objects was named. The current procedure was based on Experiment 1a of Havy et al. (<xref ref-type="bibr" rid="B40">2014</xref>) in which French-speaking adults were taught pairs of new pseudowords that differed either by a consonant or by a vowel. A comparison of performance in the two conditions to evaluate the consonant bias revealed that adults increased their looking times toward the target object (mean percentage of looking times at the target object in the postnaming&#x02014;prenaming phase) similarly in both the consonant and vowel conditions, but that latencies in shift from the distractor to the target at the time of naming were faster in the consonant than in the vowel condition, establishing a consonant bias. To explore the relative strength of the processing of consonant, vowel, and tone information in our three linguistic groups, we similarly analyzed changes in mean percentage of looking times at the target object between the prenaming and postnaming phases (which also evaluates whether the pseudowords were learned), and latencies to shift from the distractor to the target at the time of naming. We also performed cluster-based permutation analyses on the time course of looking times in the different conditions to determine when in processing looking times to the target differ among the three conditions.</p>
</sec>
<sec sec-type="materials and methods" id="s2">
<title>Materials and methods</title>
<sec>
<title>Participants</title>
<p>Seventy-two adults were tested in total: 24 Mandarin- (age range 21-27 years, mean age: 23.9 years, 21 female), 24 French-speaking adults (age range 21&#x02013;38 years, mean age: 25.2 years, 12 female), and 24 Cantonese-speaking adults (age range 20&#x02013;43 years, mean age: 27.4 years, 20 females) who served as the native speaker control group. French- and Mandarin-speaking participants had no knowledge of Cantonese (and French adults had no knowledge of other tone languages either). All participants had grown up monolingual, and no further background information (such as musical abilities&#x02026;) were collected. French-speaking adults were tested in Paris, Cantonese- and Mandarin-speaking adults were tested in Hong Kong, the latter group within a week of their arrival in Hong Kong. Before the experiment started, informed written consents were obtained from all participants. Both the experimental protocol and consent procedures were reviewed and approved by the CERES (Comit&#x000E9; d&#x00027;&#x000E9;valuation &#x000E9;thique des projets de recherche) of the Universit&#x000E9; Paris Descartes and the Human Research Ethics Committee of the Education University of Hong Kong. All data were obtained according to the principles expressed in the Declaration of Helsinki.</p>
</sec>
<sec>
<title>Stimuli</title>
<sec>
<title>Speech stimuli</title>
<p>The speech stimuli (presented in Table <xref ref-type="table" rid="T1">1</xref>) consisted of 24 pairs of disyllabic CVCV Cantonese pseudowords, differing by a minimal phonological contrast of 1 feature (except for two 2-feature contrasts for consonants, and two 2-feature contrasts for vowels, which could not be avoided due to phonological and lexical constraints in Cantonese). All contrasts were on the first syllable of the words, and the second syllable was always associated with the high level tone (T1). Eight pairs involved a consonant contrast (e.g., <bold>k</bold><sup>h</sup>&#x00254;2.l&#x003F5;1/&#x02212;/<bold>t</bold><sup>h</sup>&#x00254;2.l&#x003F5;1/), 8 involved a vowel contrast (e.g., /p<sup>h</sup><bold>u</bold>2.f&#x00254;1/&#x02212;/p<sup>h</sup><bold>y</bold>2.f&#x00254;1/), and 8 involved a tone contrast (e.g., /p<sup>h</sup>a<bold>5</bold>.mi1/&#x02212;/<italic>p</italic><sup>h</sup><italic>a</italic><bold>6</bold>.mi1/). The tones of the target syllables in the consonant and vowel pairs varied across trials.</p>
<table-wrap position="float" id="T1">
<label>Table 1</label>
<caption><p>Minimal pairs of pseudowords.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th valign="top" align="left"><bold>Condition</bold></th>
<th/>
<th valign="top" align="center"><bold>Pair</bold></th>
<th valign="top" align="left"><bold>Feature/Tone change</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Consonant contrasts</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">/tu3.la1/ - /nu3.la1/</td>
<td valign="top" align="left">Manner, voicing</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">2</td>
<td valign="top" align="center">/m&#x00254;4.h&#x00153;1/&#x02212;/p&#x00254;4.h&#x00153;1/</td>
<td valign="top" align="left">Manner, voicing</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">3</td>
<td valign="top" align="center">/si3.k<sup>h</sup>&#x00254;1/&#x02212;/ts<sup>h</sup>i3.k<sup>h</sup>&#x00254;1/</td>
<td valign="top" align="left">Manner</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">4</td>
<td valign="top" align="center">/k<sup>h</sup>&#x00254;2.l&#x003F5;1/&#x02212;/t<sup>h</sup>&#x00254;2.l&#x003F5;1/</td>
<td valign="top" align="left">Place</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">5</td>
<td valign="top" align="center">/sa1.k&#x00254;1/&#x02212;/fa1.k&#x00254;1/</td>
<td valign="top" align="left">Place</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">6</td>
<td valign="top" align="center">/k<sup>h</sup>i1.ka1/&#x02212;/kw<sup>h</sup>i1.ka1/</td>
<td valign="top" align="left">Place</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">7</td>
<td valign="top" align="center">/tsu2.m&#x003F5;1/&#x02212;/ts<sup>h</sup>u2.m&#x003F5;1/</td>
<td valign="top" align="left">Aspiration</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td/>
<td valign="top" align="center">8</td>
<td valign="top" align="center">/k<sup>h</sup>&#x003F5;4.t<sup>h</sup>&#x00153;1/&#x02212;/k&#x003F5;4.t<sup>h</sup>&#x00153;1/</td>
<td valign="top" align="left">Aspiration</td>
</tr> <tr>
<td valign="top" align="left">Vowel contrasts</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">/k<sup>h</sup>iu4.w&#x00254;1/&#x02212;/k<sup>h</sup>ui4.w&#x00254;1/</td>
<td valign="top" align="left">Place, roundedness</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">2</td>
<td valign="top" align="center">/s&#x00250;3.k&#x003F5;1/&#x02212;/s&#x00250;u3.k&#x003F5;1/</td>
<td valign="top" align="left">Place, roundedness</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">3</td>
<td valign="top" align="center">/fu1.ti1/ - /f&#x00254;1.ti1/</td>
<td valign="top" align="left">Height</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">4</td>
<td valign="top" align="center">/h&#x00153;2.t<sup>h</sup>i1/&#x02212;/hO2.t<sup>h</sup>i1/</td>
<td valign="top" align="left">Place</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">5</td>
<td valign="top" align="center">/p<sup>h</sup>u2.f&#x00254;1/&#x02212;/p<sup>h</sup>y2.f&#x00254;1/</td>
<td valign="top" align="left">Place</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">6</td>
<td valign="top" align="center">/li3.ts<sup>h</sup>a1/&#x02212;/ly3.ts<sup>h</sup>a1/</td>
<td valign="top" align="left">Roundedness</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">7</td>
<td valign="top" align="center">/k&#x003F5;1.ts&#x003F5;1/&#x02212;/k&#x00153;1.ts&#x003F5;1/</td>
<td valign="top" align="left">Roundedness</td>
</tr>
<tr style="border-bottom: thin solid #000000;">
<td/>
<td valign="top" align="center">8</td>
<td valign="top" align="center">/m&#x00250;4.t<sup>h</sup>u1/&#x02212;/mau4.t<sup>h</sup>u1/</td>
<td valign="top" align="left">Length</td>
</tr> <tr>
<td valign="top" align="left">Tone contrasts</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">/m&#x00254;2.li1/&#x02212;/m&#x00254;1.li1/</td>
<td valign="top" align="left">T1&#x02013;T2</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">2</td>
<td valign="top" align="center">/ki2.pa1 - /ki6.pa1/</td>
<td valign="top" align="left">T2&#x02013;T6</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">3</td>
<td valign="top" align="center">/pu1.fa1/ - /pu3.fa1/</td>
<td valign="top" align="left">T1&#x02013;T3</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">4</td>
<td valign="top" align="center">/ly1.khi1/ - /ly3.khi1/</td>
<td valign="top" align="left">T1&#x02013;T3</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">5</td>
<td valign="top" align="center">/tshu1.k&#x003F5;1/&#x02212;/tshu4.k&#x003F5;1/</td>
<td valign="top" align="left">T1&#x02013;T4</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">6</td>
<td valign="top" align="center">/pha5.mi1/ - /pha6.mi1/</td>
<td valign="top" align="left">T5&#x02013;T6</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">7</td>
<td valign="top" align="center">/f&#x00153;3.t&#x00254;1/&#x02212;/f4.t&#x00254;1/</td>
<td valign="top" align="left">T3&#x02013;T4</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">8</td>
<td valign="top" align="center">/t&#x003F5;4.s&#x00254;1/&#x02212;/t&#x003F5;6.s&#x00254;1/</td>
<td valign="top" align="left">T4&#x02013;T6</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<p><italic>T1 (High Level 55), T2 (High Rising 25), T3 (Mid-Level 33), T4 (Low Falling 21), T5 (Low Rising 23), T6 (Low Level 22)</italic>.</p>
</table-wrap-foot>
</table-wrap>
<p>While these pseudowords were all contrastive in Cantonese, some were not necessarily contrastive in French and/or Mandarin, as they were likely to assimilate to the same category in those languages, either equally well (such as /<bold>k</bold><sup>h</sup>&#x003F5;4.t<sup>h</sup>&#x00153;1/ - /<bold>k</bold>&#x003F5;4.t<sup>h</sup>&#x00153;1/ in French, where [k<sup>h</sup>] and [k] are both allophones of /k/), or with one of the sounds assimilating better than the other (such as /k&#x003F5;1.ts&#x003F5;1/ - /k<bold>&#x00153;</bold>1.ts&#x003F5;1/ in Mandarin, which has a front-mid-unrounded vowel [&#x003F5;] but not a front-mid-rounded vowel [&#x00153;]). See detailed explanations in the Appendix.</p>
<p>The words were presented in sentences in Cantonese. For the familiarization, they were embedded in a little passage, and appeared in six different sentences. In the test phase, one of the two words was designated twice, in two sentences (see details in &#x0201C;animated cartoons&#x0201D; section below). All speech stimuli were recorded in a quiet room by a female native adult speaker of Hong Kong Cantonese. One audio file of each condition can be find in the Supplementary Material (Consonant trial: /<bold>k</bold><sup>h</sup>&#x00254;2.l&#x003F5;1/&#x02212;/<bold>t</bold><sup>h</sup>&#x00254;2.l&#x003F5;1/; Vowel trial: /p<sup>h</sup><bold>u</bold>2.f&#x00254;1/&#x02212;/p<sup>h</sup><bold>y</bold>2.f&#x00254;1/; Tone trial: /p<sup>h</sup>a<bold>5</bold>.mi1/-/p<sup>h</sup>a<bold>6</bold>.mi1/). Note that while Mandarin and French adults did not speak Cantonese, the structure of the cartoon (with the moving object and the 6 sentences all embedding the target word) made it clear that each target word (which was thus the most frequent content word in each passage) was meant to name the object presented at the same time (which is confirmed by the results, see below).</p>
</sec>
<sec>
<title>Object stimuli</title>
<p>Images of eight pairs of objects differing in shape, color and texture (see Figure <xref ref-type="fig" rid="F2">2</xref>) were taken from a previous study by Gonzalez-Gomez et al. (<xref ref-type="bibr" rid="B34">2013</xref>). The reason for using clearly different objects was to facilitate learning of the word-object pairings. All objects were selected so that they would look novel to the participants. All 8 object pairs were used 3 times, once in each condition (consonant, vowel, tone). This was done in order to ensure that overall performance differences across conditions could not be due to the objects used.</p>
<fig id="F2" position="float">
<label>Figure 2</label>
<caption><p>The 8 pairs of novel objects used for word learning.</p></caption>
<graphic xlink:href="fpsyg-09-01211-g0002.tif"/>
</fig>
</sec>
<sec>
<title>Animated cartoons</title>
<p>The audio recordings were included in animated cartoons that have been successfully used in a computer-controlled word-learning task in toddlers by Gonzalez-Gomez et al. (<xref ref-type="bibr" rid="B34">2013</xref>). An example of a cartoon is illustrated in Figure <xref ref-type="fig" rid="F3">3</xref>.</p>
<fig id="F3" position="float">
<label>Figure 3</label>
<caption><p>Structure of a word-learning cartoon.</p></caption>
<graphic xlink:href="fpsyg-09-01211-g0003.tif"/>
</fig>
<p>On each trial, a female character behind a black board presented the two objects, one at a time (Figure <xref ref-type="fig" rid="F3">3</xref>, learning phase). The first object always appeared in the left upper corner of the screen. At the beginning, the object moved horizontally in the upper left part of the display, while it was labeled three times (&#x0201C;Look! A [label]! This is a [label]. Look at what I&#x00027;m doing with the [label]!&#x0201D;). Then, the object started shifting down, while it was labeled one more time (&#x0201C;I&#x00027;m putting the [label] here&#x0201D;). It started moving vertically in the lower left part of the screen and was labeled two more times (&#x0201C;Have you seen the [label]? Have a look at the [label]!&#x0201D;) before disappearing. The second object was always introduced in the upper right corner of the display and followed a trajectory analogous to that of the first object. The cartoon experimenter followed the objects&#x00027; movements with her eyes. Participants were successively trained on each label-object pairing for 30 s. The entire learning phase lasted 1 min and each label was repeated 6 times.</p>
<p>After the learning phase, participants were tested immediately on the given contrast. There was a close up on the face of the cartoon experimenter saying: &#x0201C;Look!&#x0201D; in order to direct the participants&#x00027; fixations to the center of the screen. After the face disappeared, the two objects appeared at the same time, each on the side where it had appeared during the learning phase, and started moving synchronously in a vertical way, for 5000 ms, while the out-of-sight speaker said: &#x0201C;Look at the [target]! Where&#x00027;s the [target]?&#x0201D; about half way through the presentation in order to divide the test phase into a prenaming and a postnaming phase of equal duration (Figure <xref ref-type="fig" rid="F3">3</xref>, test phase). Since the material was originally designed to test and compare performance in both adults and toddlers, and since it has been shown that it takes 367 ms for infants and toddlers to program eye movements (e.g., Swingley and Aslin, <xref ref-type="bibr" rid="B71">2000</xref>), the cartoons were constructed so that the onset of the <italic>postnaming phase</italic> corresponded to the onset of the first target word &#x0002B; 367 ms for consonant and tone trials; while it corresponded to the onset of the first vowel of the target word &#x0002B; 367 ms for vowel trials (hence, it corresponded to the onset of the contrasting phoneme in all trials). However, for the adult data analyses, we changed the timing by time-locking the onset of the postnaming phase 200 ms after the onset of the contrasting phoneme (as usually done in adult studies, e.g., Barr, <xref ref-type="bibr" rid="B2">2008</xref>), and reducing the size of the <italic>pre-</italic> and <italic>post-naming</italic> phases to 2,000 ms around this time point. Note that since there is debate whether tonal cues are already present in onset consonants, or whether they mostly become available with vowel onset, we conducted a preliminary analysis (see results section &#x0201C;Time course analysis for Cantonese speakers: onset of tonal information use&#x0201D;) to explore this issue in our data.</p>
<p>Every object pair was associated with one pseudoword pair in each experimental condition (e.g., object A and B were associated with /k<sup>h</sup>&#x00254;2.l&#x003F5;1/and/t<sup>h</sup>&#x00254;2.l&#x003F5;1/ in the consonant condition; with /p<sup>h</sup>u.2f&#x00254;1/&#x02212;/p<sup>h</sup>y2.f&#x00254;1/ in the vowel condition and /p<sup>h</sup>a5.mi1/&#x02212;/p<sup>h</sup>a6.mi1/ in the tone condition), for use in 24 different trials. Four versions of each cartoon were created so that in half of the trials, object/label A was the target (and consequently object/label B the distractor in those trials) and, in addition, the target was presented as first object in 50% of the trials and as second object in the other 50%. This yielded a total stimulus set of 96 movies, all having a resolution of 1280 &#x000D7; 930 pixel. Presentation of each of the four versions of each cartoon was counterbalanced across participants.</p>
</sec>
</sec>
<sec>
<title>Apparatus and procedure</title>
<p>In Paris, the movies were presented on a 17&#x02033; TFT monitor (1280 &#x000D7; 1024 pixel resolution) with an integrated Tobii T60 eyetracking system which was run by a Dell computer. The presentation of the stimuli and the storing of the data were performed with the Tobii Studio software. In Hong Kong, a Tobii TX300 was used, which was run by a Dell computer and with videos presented on a Tobii TX300 screen unit with a 1920 &#x000D7; 1080 pixel resolution.</p>
<p>Each participant was tested individually in a quiet, dimly lit laboratory room and watched 24 testing trials in total. As French- and Mandarin-speaking participants had no knowledge of Cantonese they received a warm-up trial in Cantonese, in which the two pseudowords used were phonetically different in every single segment (/ka/ - /su/) and in which subtitles were presented, in order to familiarize them with the task.</p>
<p>There were 12 pseudo-randomized orders, which were each presented to two participants in each of the three language groups. Four sub-blocks of 6 trials were presented. After every sub-block, the participant could take a break for as long as s/he wished. In each sub-block, there were 2 consonant trials (one with target on the left, one with target on the right), 2 vowel trials (left, right), and 2 tone trials (left, right). Consequently, within a subject, half of the time the target word was on the left, half of the time it was on the right. All 3 conditions were presented in the first 3 trials and there were never more than 2 target-left or 2 target-right trials in a row. None of the words was presented twice, but the objects occurred three times during the test. Note that the same object pairs were not presented within the same sub-block, in order to prevent learning interference. After creating order 1 with those constraints, it was mirrored to get order 2 (e.g., order 1: trial 1 - trial 24; order 2: trial 24 - trial 1). For order 3, we shuffled the trials of order 1 so that the ones that occurred in the first half in order 1 appeared in the second half (and the other way around). Order 4 was again a mirror of order 3. Orders 5-8 and 9-12 were exactly like orders 1-4 but the conditions were differently assigned following a Latin square design (e.g., order 1: V<sub>pair1</sub>, T<sub>pair8</sub>, C<sub>trial3</sub> &#x02026;; order 5: T<sub>pair1</sub>, C<sub>pair8</sub>, V<sub>trial3</sub> &#x02026;; order 9: C<sub>pair1</sub>, V<sub>pair8</sub>, T<sub>trial3</sub> &#x02026; Note that the number of each pair here means the specific object pair that was used). As a consequence, between-subjects counterbalancing ensured that each object-word pair was presented and tested on the right and left side equally often and occurred in all 3 conditions at the same serial position. The experiment lasted approximately 30 min.</p>
</sec>
<sec>
<title>Data analysis</title>
<p>The eye-tracking data used for the analysis consisted of the binocular gaze position (X and Y coordinates) at each timestamp, that is, every 16.6 ms for French-speaking adults and every 3.3 ms for Mandarin- and Cantonese-speaking adults. Trials in which no data was available for the postnaming phase were discarded from the analyses (27/1728 trials). The data was analyzed in R (version 3.4.3, R Core Team, <xref ref-type="bibr" rid="B64">2017</xref>, <ext-link ext-link-type="uri" xlink:href="http://www.r-project.org">http://www.r-project.org</ext-link>) using the <italic>eyetrackingR</italic> package (Dink and Ferguson, <xref ref-type="bibr" rid="B23">2015</xref>, <ext-link ext-link-type="uri" xlink:href="http://www.eyetrackingr.com">http://www.eyetrackingr.com</ext-link>) for the latency and the growth curve analysis as well as for the cluster-based permutation analysis.</p>
</sec>
</sec>
<sec sec-type="results" id="s3">
<title>Results</title>
<sec>
<title>Time course analysis for cantonese speakers: onset of tonal information use</title>
<p>To evaluate the issue of whether the onset of tones should be time-locked to the onset of the consonants or the vowels of the syllables in which they were embedded, we first plotted the time course of the Cantonese adults&#x00027; target looking behavior during the test phase based on two analyses. In the first one, as originally planned when preparing the videos, the postnaming phase was aligned with the beginning of the onset consonant of the target words (see Figure <xref ref-type="fig" rid="F4">4</xref>, top panel). In the second analysis, we corrected the time course aligning the postnaming phase to the onset of the vowel (see Figure <xref ref-type="fig" rid="F4">4</xref>, bottom panel). On average, we corrected for 127 ms (range 32&#x02013;235 ms). As can be seen from the comparison of the two figures, similar identification curves are found for the consonant and vowel conditions, with a very similar timing. While word recognition appears delayed in the tone condition compared to the other two conditions when recognition is time-locked to the onset of the consonant, this delay disappears when it is time-locked to the onset of the vowel. This suggests that tonal information is more likely available from vowel onset rather than consonant onset for the current set of pseudowords, and that the speed of use of tonal information in native processing is similar to that of consonantal and vocalic information.</p>
<fig id="F4" position="float">
<label>Figure 4</label>
<caption><p>Target looking behavior during the test phase, for Cantonese speakers only. Time point 0 refers to the onset of the first consonant for consonant trials, and the onset of the first vowel for vowel trials in both panels. For tones, time point 0 refers to the onset of the first consonant in the top panel, but the onset of the first vowel for the bottom panel. The dotted line represents the beginning of the postnaming phase.</p></caption>
<graphic xlink:href="fpsyg-09-01211-g0004.tif"/>
</fig>
<p>Given the above findings, all analyses presented in the following sections are based on the recalculation of the pre/postnaming phase for the tone-contrasted trials, taking <italic>vowel</italic> onset &#x0002B;200 ms as the beginning of the postnaming phase. Note however that equivalent analyses time-locked to consonant onset provided the same pattern of results.</p>
</sec>
<sec>
<title>Accuracy-overall analysis</title>
<p>We first calculated the mean proportion of target looking (PTL &#x0003D; total looking time to target/ total looking time to both objects) on each trial for both the pre- and postnaming phase. For this purpose, two areas of interest (AOI) were defined (575 &#x000D7; 895 Pixel), each including one object. Time stamps that were not in any of the AOIs were treated as missing data, so that the calculated proportion of looking to one AOI is always relative to both AOIs, resulting in values between zero and 1 (i.e., a proportion value of 0.5 means that each AOI was looked at equally long). Word learning is typically reflected by the <italic>naming effect</italic> which corresponds to an increase in the proportion of target looking between the pre- and the post-naming phases that is significantly above 0 (e.g., Singh et al., <xref ref-type="bibr" rid="B66">2015</xref>). The purpose of this prenaming correction is to control for looking preferences that are independent of the labeling. Difference scores between the pre- and postnaming phases were therefore calculated for each adult and each of the 24 contrast pairs, and then averaged for the 3 types of contrasts (see Figure <xref ref-type="fig" rid="F5">5</xref>). Zero corresponds to no increase in looking to target between the pre- and post-naming phases (chance performance). Positive difference scores mean an increase in target looking proportion.</p>
<fig id="F5" position="float">
<label>Figure 5</label>
<caption><p>Size of the naming effect, broken down by the language of the participants (Cantonese, Mandarin and French) and the type of the contrast (consonant, vowel, tone). Error bars indicate standard errors of the means.</p></caption>
<graphic xlink:href="fpsyg-09-01211-g0005.tif"/>
</fig>
<p>For each of the three types of contrasts, adults in each language group exhibit an above chance naming effect (all <italic>p</italic>s &#x0003C; 0.001; see Table <xref ref-type="table" rid="T2">2</xref> for details). This establishes that adults in all language groups could learn the words in all conditions, even though all stimuli were in Cantonese, a language not known by the Mandarin- and French-speaking adults.</p>
<table-wrap position="float" id="T2">
<label>Table 2</label>
<caption><p>Naming effect broken down by language and condition.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th/>
<th valign="top" align="center"><bold>Mean (SD)</bold></th>
<th valign="top" align="center"><bold>Comparison to 0 chance-level</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left" colspan="3" style="background-color:#bdbec1"><bold>CANTONESE-SPEAKING ADULTS:</bold></td>
</tr>
<tr>
<td valign="top" align="left">Consonant trials</td>
<td valign="top" align="center">0.40 (0.104)</td>
<td valign="top" align="center"><italic>t</italic><sub>(23)</sub> &#x0003D; 18.93; <italic>p</italic> &#x0003C; 0.001</td>
</tr>
<tr>
<td valign="top" align="left">Vowel trials</td>
<td valign="top" align="center">0.40 (0.091)</td>
<td valign="top" align="center"><italic>t</italic><sub>(23)</sub> &#x0003D; 21.56; <italic>p</italic> &#x0003C; 0.001</td>
</tr>
<tr>
<td valign="top" align="left">Tone trials</td>
<td valign="top" align="center">0.41 (0.094)</td>
<td valign="top" align="center"><italic>t</italic><sub>(23)</sub> &#x0003D; 21.31; <italic>p</italic> &#x0003C; 0.001</td>
</tr>
<tr>
<td valign="top" align="left" colspan="3" style="background-color:#bdbec1"><bold>MANDARIN-SPEAKING ADULTS:</bold></td>
</tr>
<tr>
<td valign="top" align="left">Consonant trials</td>
<td valign="top" align="center">0.32 (0.157)</td>
<td valign="top" align="center"><italic>t</italic><sub>(23)</sub> &#x0003D; 9.93; <italic>p</italic> &#x0003C; 0.001</td>
</tr>
<tr>
<td valign="top" align="left">Vowel trials</td>
<td valign="top" align="center">0.31 (0.156)</td>
<td valign="top" align="center"><italic>t</italic><sub>(23)</sub> &#x0003D; 9.87; <italic>p</italic> &#x0003C; 0.001</td>
</tr>
<tr>
<td valign="top" align="left">Tone trials</td>
<td valign="top" align="center">0.25 (0.171)</td>
<td valign="top" align="center"><italic>t</italic><sub>(23)</sub> &#x0003D; 7.05; <italic>p</italic> &#x0003C; 0.001</td>
</tr>
<tr>
<td valign="top" align="left" colspan="3" style="background-color:#bdbec1"><bold>FRENCH-SPEAKING ADULTS:</bold></td>
</tr>
<tr>
<td valign="top" align="left">Consonant trials</td>
<td valign="top" align="center">0.26 (0.145)</td>
<td valign="top" align="center"><italic>t</italic><sub>(23)</sub> &#x0003D; 8.83; <italic>p</italic> &#x0003C; 0.001</td>
</tr>
<tr>
<td valign="top" align="left">Vowel trials</td>
<td valign="top" align="center">0.30 (0.172)</td>
<td valign="top" align="center"><italic>t</italic><sub>(23)</sub> &#x0003D; 8.46; <italic>p</italic> &#x0003C; 0.001</td>
</tr>
<tr>
<td valign="top" align="left">Tone trials</td>
<td valign="top" align="center">0.12 (0.150)</td>
<td valign="top" align="center"><italic>t</italic><sub>(23)</sub> &#x0003D; 3.83; <italic>p</italic> &#x0003C; 0.001</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>To test for differences between language group and type of contrast, a 2-way ANOVA with the main factors of native language (Cantonese, Mandarin, French) and type of contrast (consonant, vowel, tone) was performed. A main effect of language [<italic>F</italic><sub>(2, 69)</sub> &#x0003D; 16.96; <italic>p</italic> &#x0003C; 0.001] was found. <italic>T</italic>-tests revealed that Cantonese-speaking adults had a larger naming effect (0.40) than both Mandarin- [0.29, <italic>t</italic><sub>(46)</sub> &#x0003D; 3.92; <italic>p</italic> &#x0003C; 0.001] and French-speaking adults [0.22, <italic>t</italic><sub>(46)</sub> &#x0003D; 6.19; <italic>p</italic> &#x0003C; 0.001], whose performance did marginally differ [<italic>t</italic><sub>(46)</sub> &#x0003D; 1.92; <italic>p</italic> &#x0003D; 0.06]. This indicates an advantage of learning in one&#x00027;s native language vs. in an unknown language, and furthermore points toward a linguistic distance effect as Mandarin and Cantonese are related languages while French is unrelated to Cantonese.</p>
<p>There was also a main effect of type of contrast [<italic>F</italic><sub>(2, 138)</sub> &#x0003D; 10.46; <italic>p</italic> &#x0003C; 0.001], naming effects being larger for both consonants (0.33) and vowels (0.34) than for tones [0.26; <italic>t</italic><sub>(71)</sub> &#x0003D; 3.34, <italic>p</italic> &#x0003D; 0.001, and <italic>t</italic><sub>(71)</sub> &#x0003D; 3.34; <italic>p</italic> &#x0003D; 0.001, respectively]. This indicates that tone contrasts were overall more difficult to process than consonant and vowel contrasts. Performance between the consonant and vowel condition did not differ [<italic>t</italic><sub>(71)</sub> &#x0003D; 0.65, <italic>p</italic> &#x0003D; 0.51]. In addition, the native language x type of contrast interaction was significant [<italic>F</italic><sub>(4, 138)</sub> &#x0003D; 4.86; <italic>p</italic> &#x0003D; 0.001]. This indicates that performance for the different types of contrasts was differently affected in the three language groups. Compared to native Cantonese-speaking adults, both Mandarin- and French-speaking adults performed significantly worse on all three contrasts [tone: 0.25 vs. 0.41, <italic>t</italic><sub>(46)</sub> &#x0003D; 4.09; <italic>p</italic> &#x0003C; 0.001; 0.12 vs. 0.41, <italic>t</italic><sub>(46)</sub> &#x0003D; 8.06; <italic>p</italic> &#x0003C; 0.001; vowel: 0.31 vs. 0.40, <italic>t</italic><sub>(46)</sub> &#x0003D; 2.34; <italic>p</italic> &#x0003D; 0.02; 0.30 vs. 0.40, <italic>t</italic><sub>(46)</sub> &#x0003D; 2.57; <italic>p</italic> &#x0003D; 0.01; consonant: 0.32 vs. 0.40, <italic>t</italic><sub>(46)</sub> &#x0003D; 2.16; <italic>p</italic> &#x0003D; 0.04; 0.26 vs. 0.40, <italic>t</italic><sub>(46)</sub> &#x0003D; 3.86; <italic>p</italic> &#x0003C; 0.001]. Additionally, French-speaking adults performed worse on tone contrasts than Mandarin-speaking adults [0.12 vs. 0.25, <italic>t</italic><sub>(46)</sub> &#x0003D; 2.77; <italic>p</italic> &#x0003D; 0.008].</p>
<p>Comparing conditions within each language, taking all 8 trials per condition into account, no difference in performance between the consonant and vowel conditions was found for the three language groups [Cantonese speakers: 0.40 vs. 0.40, <italic>t</italic><sub>(23)</sub> &#x0003D; 0.07, <italic>p</italic> &#x0003D; 0.94; Mandarin speakers: 0.32 vs. 0.31, <italic>t</italic><sub>(23)</sub> &#x0003D; 0.24, <italic>p</italic> &#x0003D; 0.82; French speakers: 0.26 vs. 0.30, <italic>t</italic><sub>(23)</sub> &#x0003D; 1.16, <italic>p</italic> &#x0003D; 0.26]. Performance on tone contrasts was lower than in the other two conditions for French speakers [0.12 vs. 0.28, <italic>t</italic><sub>(23)</sub> &#x0003D; 5.14, <italic>p</italic> &#x0003C; 0.001], but not for Mandarin [0.25 vs. 0.32, <italic>t</italic><sub>(23)</sub> &#x0003D; 1.61, <italic>p</italic> &#x0003D; 0.12] and Cantonese speakers [0.41 vs. 0.40, <italic>t</italic><sub>(23)</sub> &#x0003D; 0.49, <italic>p</italic> &#x0003D; 0.63]. Redoing these analyses removing the Single Category trials and the Category Goodness trials in each condition (see details in Appendix) confirmed the lack of difference in performance between the 8 consonant and 5 vowel native-like/Two Category pairs for Mandarin [0.32 vs. 0.35, <italic>t</italic><sub>(23)</sub> &#x0003D; 0.93, <italic>p</italic> &#x0003D; 0.36], and the 6 consonant and 7 vowel native-like/Two Category pairs for French [0.29 vs. 0.32, <italic>t</italic><sub>(23)</sub> &#x0003D; 0.58, <italic>p</italic> &#x0003D; 0.56].</p>
</sec>
<sec>
<title>Latency analysis</title>
<p>Second, following Havy et al. (<xref ref-type="bibr" rid="B40">2014</xref>), we examined the participants&#x00027; latency in shifting from the distractor to the target object, that is the time needed to orient from the initially fixated distractor object to the target object after labeling. Faster latencies to the target object in a condition would indicate a processing advantage compared to the other conditions. In a first step, distractor-initial trials were defined as those in which participants fixated the distractor object at the onset of the pivotal phoneme (first consonant of the target word for consonant trials; first vowel for tone and vowel trials). These distractor-initial trials corresponded to, on average, 46% of all the trials (Cantonese: 45%; Mandarin: 45%; French: 48%). From those trials, we excluded trials in which participants shifted before the postnaming phase began (i.e., within the next 200 ms) as these saccades were probably programmed before the name of the target was processed (Cantonese: 21%; Mandarin: 22%; French: 10%) or did not shift at all (Cantonese: 1%; Mandarin: 6%; French: 7%) as well as outliers, that is values greater or smaller than 2.5 standard deviations from the mean (Cantonese: 1%; Mandarin: 2%; French: 2%).</p>
<p>Mean latencies and standard deviations are shown in Table <xref ref-type="table" rid="T3">3</xref> for each language and condition. We used a linear mixed model using the function <italic>lmer</italic> of the R package <italic>lm4</italic>, with random effects for participants and items (Bates et al., <xref ref-type="bibr" rid="B3">2015</xref>), and the package <italic>languageR</italic> (Baayen, <xref ref-type="bibr" rid="B1">2015</xref>) to obtain <italic>p</italic>-values. The model included fixed effects of condition (compared in sliding contrasts: Consonants-Vowels; Vowels-Tones), language group (also compared in sliding contrasts: French vs. Cantonese; Cantonese vs. Mandarin), and the interaction between condition and language group. We decided to test the consonant-vowel contrast to be able to compare with previously reported results, and the tone-vowel comparison because of the same target phoneme onset. As for the language contrasts, we took Cantonese as the native speaker reference group with which to compare both non-native speaker groups. The output measure was mean shift latency. The only significant differences were found between Cantonese and Mandarin participants (&#x003B2; &#x0003D; 163.31, <italic>SE</italic> &#x0003D; 50.23, <italic>t</italic> &#x0003D; 3.25, <italic>p</italic> &#x0003D; 0.002) and between Cantonese and French participants (&#x003B2; &#x0003D; &#x02212;199.87, <italic>SE</italic> &#x0003D; 49.14, <italic>t</italic> &#x0003D; 4.07, <italic>p</italic> &#x0003C; 0.001), with the Cantonese participants having overall faster latencies then each of the two other language groups. This points, again, to a general native language advantage. Importantly, the conditions did not differ from each other or interact with language.</p>
<table-wrap position="float" id="T3">
<label>Table 3</label>
<caption><p>Mean shift latencies in ms and their SDs (in brackets), broken down by language and condition.</p></caption>
<table frame="hsides" rules="groups">
<thead><tr>
<th valign="top" align="left" colspan="5"><bold>LANGUAGE</bold></th>
</tr>
</thead>
<tbody>
<tr style="border-bottom: thin solid #000000;">
<td/>
<td valign="top" align="left"><bold>Cantonese</bold></td>
<td valign="top" align="left"><bold>Mandarin</bold></td>
<td valign="top" align="left"><bold>French</bold></td>
<td valign="top" align="left"><bold>Mean (condition)</bold></td>
</tr> <tr>
<td valign="top" align="left" colspan="5"><bold>CONDITION</bold></td>
</tr>
<tr>
<td valign="top" align="left">Consonants</td>
<td valign="top" align="left">369 (124)</td>
<td valign="top" align="left">489 (303)</td>
<td valign="top" align="left">548 (377)</td>
<td valign="top" align="left">469 (297)</td>
</tr>
<tr>
<td valign="top" align="left">Vowels</td>
<td valign="top" align="left">366 (260)</td>
<td valign="top" align="left">644 (514)</td>
<td valign="top" align="left">570 (386)</td>
<td valign="top" align="left">529 (408)</td>
</tr>
<tr>
<td valign="top" align="left">Tones</td>
<td valign="top" align="left">420 (281)</td>
<td valign="top" align="left">528 (325)</td>
<td valign="top" align="left">668 (395)</td>
<td valign="top" align="left">537 (350)</td>
</tr>
<tr>
<td valign="top" align="left"><bold>Mean (language)</bold></td>
<td valign="top" align="left">386 (227)</td>
<td valign="top" align="left">543 (380)</td>
<td valign="top" align="left">592 (387)</td>
<td/>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
<sec>
<title>Growth curve analysis</title>
<p>Third, we conducted a Growth Curve Analysis (GCA) which includes time as a predictor to estimate if differences between conditions emerged over time within each language group. As dependent measure we took the transformed proportion data during the postnaming phase using the empirical logit (<italic>elog</italic>, aggregated in 100 ms time bins) and analyzed it with a weighted mixed-effects linear regression model within the <italic>eyetrackingR</italic> package (modeled after Mirman et al., <xref ref-type="bibr" rid="B52">2008</xref>). For each language group separately, we entered condition (again compared in sliding contrasts: Consonants-Vowels; Vowels-Tones), orthogonal polynomials (linear, quadratic and cubic time component), and the interaction between each time term and condition as fixed effects. Participants and items were entered as random effects into the model.</p>
<p>For Cantonese-speaking adults (see Figure <xref ref-type="fig" rid="F6">6A</xref>), conditions (Consonant-Vowel; Vowel-Tone) did not differ in their mean target looking, but both contrasts interacted (marginally) significantly with time (linear parameter: &#x003B2; &#x0003D; &#x02212;0.93, <italic>SE</italic> &#x0003D; 0.21, <italic>p</italic> &#x0003C; 0.001; &#x003B2; &#x0003D; 0.52, <italic>SE</italic> &#x0003D; 0.21, <italic>p</italic> &#x0003D; 0.01; quadratic parameter: &#x003B2; &#x0003D; 0.45, <italic>SE</italic> &#x0003D; 0.21, <italic>p</italic> &#x0003C; 0.03; &#x003B2; &#x0003D; &#x02212;0.37, <italic>SE</italic> &#x0003D; 0.28, <italic>p</italic> &#x0003D; 0.08).</p>
<fig id="F6" position="float">
<label>Figure 6</label>
<caption><p>Time course during the postnaming phase of consonant, vowel and tone trials for Cantonese-<bold>(A)</bold>, Mandarin- <bold>(B)</bold>, and French-speaking adults <bold>(C)</bold>; shown as raw data (light) and fitted curves (bold).</p></caption>
<graphic xlink:href="fpsyg-09-01211-g0006.tif"/>
</fig>
<p>For Mandarin-speaking adults (see Figure <xref ref-type="fig" rid="F6">6B</xref>), there was no significant main effect of the Consonant-Vowel and the Vowel-Tone contrast (both <italic>ps</italic> &#x0003E; 0.22), indicating no differences in the overall target looking in the postnaming phase between those conditions. We found a significant interaction between the Consonant-Vowel contrast and time (specifically, the quadratic and cubic parameter: &#x003B2; &#x0003D; 0.94, <italic>SE</italic> &#x0003D; 0.28, <italic>p</italic> &#x0003C; 0.001; &#x003B2; &#x0003D; &#x02212;0.72, <italic>SE</italic> &#x0003D; 0.28, <italic>p</italic> &#x0003D; 0.009), and between the Vowel-Tone contrast and time (linear parameter: &#x003B2; &#x0003D; &#x02212;0.71, <italic>SE</italic> &#x0003D; 0.28, <italic>p</italic> &#x0003D; 0.01).</p>
<p>For French-speaking adults (see Figure <xref ref-type="fig" rid="F6">6C</xref>), the GCA revealed a significant main effect of the Vowel-Tone contrast on the intercept term, confirming the overall lower target fixations for the tone trials relative to the vowel trials (&#x003B2; &#x0003D; &#x02212;0.70, <italic>SE</italic> &#x0003D; 0.19, <italic>p</italic> &#x0003D; 0.001). In addition, the Vowel-Tone contrast interacted (marginally) significantly with time (linear time parameter: &#x003B2; &#x0003D; &#x02212;0.39, <italic>SE</italic> &#x0003D; 0.21, <italic>p</italic> &#x0003D; 0.06; quadratic time parameter: &#x003B2; &#x0003D; 0.43, <italic>SE</italic> &#x0003D; 0.21, <italic>p</italic> &#x0003D; 0.04), suggesting divergent linear and non-linear temporal trajectories for tone and vowel trials. While the Consonant-Vowel contrast on the intercept term was not significant (&#x003B2; &#x0003D; 0.05, <italic>SE</italic> &#x0003D; 0.19, <italic>p</italic> &#x0003D; 0.81), its interaction with time was marginally significant (linear time parameter: &#x003B2; &#x0003D; 0.38, <italic>SE</italic> &#x0003D; 0.21, <italic>p</italic> &#x0003D; 0.06; cubic time parameter: &#x003B2; &#x0003D; &#x02212;0.40, <italic>SE</italic> &#x0003D; 0.21, <italic>p</italic> &#x0003D; 0.06). This indicates that the temporal trajectory tends to differ between these conditions, although these differences are only trends, in line with the lack of mean target looking time differences during the postnaming phase between the consonant and vowel trials.</p>
<p>Note that <italic>eyetrackingR</italic> fits curves using orthogonal polynomials so that the estimated time parameters are independent from each other. As a consequence, the condition effect on the intercept corresponds to differences averaged across the entire postnaming phase. In a second model, we used natural polynomials in order to obtain so-called anticipatory effects, that is mean differences between conditions at the onset of postnaming phase (see Barr, <xref ref-type="bibr" rid="B2">2008</xref>). These analyses revealed no significant effect of condition on the intercept term (all <italic>p</italic>s &#x0003E; 0.35). Thus, it can be ruled out that differences between conditions were already present before the postnaming phase started, that is before the critical information in a trial was processed.</p>
</sec>
<sec>
<title>Cluster-based permutation analysis</title>
<p>To further explore the different temporal trajectories that the results of the CGAs indicated, we conducted a cluster-based permutation analysis (Maris and Oostenveld, <xref ref-type="bibr" rid="B49">2007</xref>) for each language group separately to identify the exact time periods where conditions differ significantly from each other. As dependent measure we took the proportion of target looking within each 100 ms bin across the postnaming phase (20 bins). In a first step, this analysis compares conditions at each time bin with a <italic>t</italic>-test and identifies any time period(s) of adjacent bins in which conditions significantly differ. As <italic>t</italic>-threshold we chose an &#x003B1;-level of 0.05 (two-tailed). This yields in cluster-level <italic>t</italic>-value(s) which correspond to the sum of all single sample <italic>t</italic>-values within the time period(s). In a second step, it generates a Monte-Carlo distribution to compare the cluster-level <italic>t</italic>-value(s) by randomly assigning the trials to conditions and repeating step 1 several times (for our data: 1000 times). This results in a Monte Carlo <italic>p</italic>-value for each observed time cluster which reflects the probability that this cluster could have occurred simply by chance.</p>
<p>This analysis revealed no differences between conditions for the Cantonese-speaking group. For Mandarin-speaking adults, Consonant and Vowel trials diverged from 300 to 900 ms during the postnaming phase (cluster <italic>t</italic> &#x0003D; 16.29, Monte Carlo <italic>p</italic> &#x0003D; 0.02) with Consonant trials having higher target looking proportions. While Tone trials did not differ from Vowel trials, they did from Consonant trials between 400 and 2000 ms during the postnaming phase (cluster <italic>t</italic> &#x0003D; 51.91, Monte Carlo <italic>p</italic> &#x0003C; 0.001), again Consonant trials having higher target looking proportions. Interestingly, redoing these analyses removing the Single Category and Category Goodness trials in each condition (see details in Appendix) there was no difference between conditions any more. For French-speaking adults, two significant clusters were found: both Tone and Vowel trials and Tone and Consonant trials diverged from 200 to 2000 ms during the postnaming phase (cluster <italic>t</italic> &#x0003D; 63.43, Monte Carlo <italic>p</italic> &#x0003C; 0.001; cluster <italic>t</italic> &#x0003D; 62.27, Monte Carlo <italic>p</italic> &#x0003C; 0.001, respectively), with Tone trials having lower target fixations. Consonant and Vowel trials did not differ. This was still the case after removing the Single Category and Category Goodness trials in the vowel and consonant conditions (see details in Appendix).</p>
</sec>
</sec>
<sec sec-type="discussion" id="s4">
<title>Discussion</title>
<p>In this study, we investigated whether and how adults can quickly learn new minimal pair words in a non-native tone language, Cantonese, and whether this ability is modulated by native phonological knowledge. We tested this learning ability in Mandarin- and French-speaking adults, using Cantonese pseudowords differing minimally in either a consonant, a vowel, or a tone, and compared their performance to those of native Cantonese-speaking adults. Overall, we found that all three groups of adults performed at above chance levels in learning the pseudowords, and this held for all three types of contrasts. Also, compared to native Cantonese-speaking adults, both Mandarin- and French-speaking adults performed worse on all three types of contrasts. Furthermore, French-speaking adults performed even worse on tones when compared to Mandarin-speaking adults.</p>
<p>The present findings first establish that adults in all three language groups could rapidly learn new words in a computer-based situation, after solely 6 repetitions of each word. Note that the present interpretation in terms of word-learning needs to be qualified by the fact that the present study does not establish long-term establishment of lexical items, and could result from simple associations between the pseudowords and either the objects (or the side of the screen on which the objects were presented). Future studies will have to further probe our word-learning interpretation, using designs testing for word learning independent of object localization, and in long term memory, for example adapting the word-learning design used in Dittinger et al. (<xref ref-type="bibr" rid="B24">2016</xref>). While this word-learning finding is in part trivial for the Cantonese-speaking adults (though see more discussion on this below), it holds even when the new words were presented to Mandarin- and French-speaking adults, for whom Cantonese was a non-native language, and who had no knowledge of Cantonese prior to taking part in the experiment. Our findings reveal a significant effect of nativeness status, as overall, Cantonese-speaking adults performed better than the other two groups (in overall performance and shift latency analyses).</p>
<p>The effect of linguistic distance is less clearcut. Indeed, although Cantonese is closer to Mandarin than to French (at many levels including phonology, morphology and syntax, Li, <xref ref-type="bibr" rid="B47">1937</xref>; Gong, <xref ref-type="bibr" rid="B33">1980</xref>; DeLancey, <xref ref-type="bibr" rid="B21">2009</xref>), this did not significantly impact overall performance and shift latencies, as French-speaking adults performed at the same overall level as Mandarin-speaking adults, in spite of a trend in the expected direction for overall performance (see further discussion in the Appendix for a more fine-grained approach). Our findings thus establish robust word learning abilities in a non-native language in adulthood, that contrast with the difficulties that adults have in learning some specific aspects of the phonology and syntax of non-native languages (e.g., Flege et al., <xref ref-type="bibr" rid="B27">1999</xref>; Birdsong and Molis, <xref ref-type="bibr" rid="B9">2001</xref>; Dupoux et al., <xref ref-type="bibr" rid="B26">2008</xref>; Boll-Avetisyan et al., <xref ref-type="bibr" rid="B10">2016</xref>). This difference might be due to the fact that while the acquisition of the phonology and syntax of one&#x00027;s native language is to a great extent completed in the first years of life, vocabulary acquisition is a lifelong, continuing process that allows for the acquisition of specialized vocabularies (as when, for example, becoming a -developmental- psychologist!) or learning the names of new objects and concepts (e.g., to &#x0201C;log into&#x0201D; a &#x0201C;googledoc&#x0201D; on one&#x00027;s &#x0201C;iphone&#x0201D;) in the native language.</p>
<p>Importantly, these word learning abilities were found in a specific learning context in which adults had to learn words presented in pairs, and in which the sound forms of the two words differed only by a consonant, vowel or tone. The fact that Mandarin- and French-speaking adults succeeded in learning the word pairs in all three conditions establishes that they could process fine segmental (consonantal and vocalic) and suprasegmental (tonal) information in doing so, and that they were establishing representations of the word forms that included specific segmental or tonal information. This finding is particularly striking for the French speakers&#x00027; performance with tone contrasts, given that tones are not used in French at the lexical level. It could be due to the fact that these contrasts were introduced to the adults in minimal pairs of words, where they had to pay attention to the fine phonetic detail in order to distinguish the objects and memorize the words. Further research will be needed to explore whether our French-speaking adults would have failed to use such precise phonetic information if they had not been presented with minimal pairs, leading to lower or at chance performance. Importantly though, the ability of the French-speaking adults to use tonal information when learning words suggest that the rebound in tone discrimination found in late infancy in Dutch, another non-tonal language (Liu and Kager, <xref ref-type="bibr" rid="B48">2014</xref>), interpreted in relation to the acquisition of the intonation of the native language, would not be limited to low levels of processing, but would extend to the lexical level.</p>
<p>Our findings also establish that adult performance is not solely based on the acoustic distance between the contrasted sounds, but is also dependent on their native phonological system. At this more fine-grained level, language distance appears to play a role, as our results clearly show that the Mandarin-speaking adults performed better than the French-speaking adults in learning words distinguished by Cantonese tonal contrasts. Since there was no difference in performance between the two language groups for consonants and vowels, this effect likely indicates that Mandarin-speaking adults, as experienced tone language users, exhibit greater ability in processing non-native tonal information at the lexical level, when compared to the non-tone user French speakers. In the Appendix, we present exploratory analyses, based on individual trials analyses, that allow some evaluation of the Perceptual Assimilation Model (PAM; for consonants: Best, <xref ref-type="bibr" rid="B6">1995</xref>; for vowels: Tyler et al., <xref ref-type="bibr" rid="B73">2014</xref>; for tones: Hall&#x000E9; et al., <xref ref-type="bibr" rid="B35">2004</xref>) applied here at the level of word learning rather than speech processing.</p>
<p>Besides providing data on the phonological/lexical interface in processing a new, non-native language, our results also provide an evaluation of the use of tonal information in word learning, and its impact on processing consonantal and vocalic information at the lexical level in native speakers of a tone language. Regarding the use of tonal contrasts, we found that tonal contrasts are as important as consonantal and vocalic contrasts in processing word meanings for native Cantonese-speaking adults. This is revealed by the overall accuracy analyses showing that Cantonese adults perform at the same level in all three contrast conditions. Our time course analysis further shows that all three kinds of contrasts are processed at the same speed from the onset of the contrasting phonemes. For the tones, the comparison of our two analyses time-locked to consonant vs. vowel onset suggests that tonal information became available from the onset of the vowel. This might be related to the fact that 6 of the 8 pairs we presented started with unvoiced consonants, so that tonal information was mostly carried by the vowels. Whether a similar pattern would be found for syllables starting with voiced consonants would need to be evaluated in an experimental design counterbalancing the two types of consonants.</p>
<p>Furthermore, our work bears on the issue of the relative weight given to consonantal and vocalic information in lexical processing. Previous studies on various Indo-European languages (English, Dutch, French, Italian, Spanish) have found that adults have a consonant bias in accessing or learning words (e.g., van Ooijen, <xref ref-type="bibr" rid="B74">1996</xref>; Cutler et al., <xref ref-type="bibr" rid="B20">2000</xref>; Bonatti et al., <xref ref-type="bibr" rid="B11">2005</xref>; Creel et al., <xref ref-type="bibr" rid="B19">2006</xref>; New et al., <xref ref-type="bibr" rid="B58">2008</xref>; Toro et al., <xref ref-type="bibr" rid="B72">2008</xref>; Carreiras et al., <xref ref-type="bibr" rid="B14">2009</xref>; Delle Luche et al., <xref ref-type="bibr" rid="B22">2014</xref>; Havy et al., <xref ref-type="bibr" rid="B40">2014</xref>; New and Nazzi, <xref ref-type="bibr" rid="B59">2014</xref>). This supports the &#x0201C;division of labor&#x0201D; proposal by Nespor et al. (<xref ref-type="bibr" rid="B57">2003</xref>) that consonants are given more weight than vowels in lexical processing (while vowels are given more weight than consonants at the prosodic/syntactic levels). Accordingly, in the present study, we investigated Cantonese, a tone language, where lexical meanings are also crucially cued by tones. Our interest came from the fact that since tones are essentially associated with the voiced portions of syllables, which mostly correspond to vowels (and nasal codas) in Cantonese, and since only a few onset consonants (/j, w, m, n, &#x003B7;/) are voiced in that language, the relative weight given to consonants and vowels might be different from what has been found for Indo-European languages. The effect of tones could either increase the weight given to vowels (since they carry both segmental and tonal information, compared to only segmental information in non-tone languages) or decrease their weight even further (due to additional acoustic variation related to tonal differences and the fact that each vowel in Cantonese can carry 6 different tones).</p>
<p>Our findings did not reveal any advantage for either consonantal or vocalic information in lexical processing for Cantonese-speaking adults, as shown by their overall similar performance in the consonant and vowel contrast conditions. This finding differs from all previous findings on adult speakers of non-tonal languages, and in particular with the findings of a clear C-bias in latency analyses found when French adults learn new words (Havy et al., <xref ref-type="bibr" rid="B40">2014</xref>). The present null effect (lack of difference between the C and V conditions), which thus needs to be interpreted with caution, might be taken as evidence that Cantonese-speaking adults pay more attention to vowels that carry tonal information than non-tone language users, resulting in a lack of C-bias. This interpretation needs to be considered cautiously given that a null effect was also found in the other two language groups, including in the French-speaking adults, which have been documented to have a C-bias in lexical processing when processing words in their native language (e.g., Bonatti et al., <xref ref-type="bibr" rid="B11">2005</xref>; New et al., <xref ref-type="bibr" rid="B58">2008</xref>; Havy et al., <xref ref-type="bibr" rid="B40">2014</xref>). This lack of effect in the French- and Mandarin-speaking adults could mean that there is something in the acoustics of the stimuli used in the present study that does not support a C-bias. Alternatively, it could mean that the C-bias only operates in the native language, or in languages in which adults have sufficient experience/knowledge, hence the null effect found here for the French-speaking (and Mandarin-speaking) adults who had no or limited knowledge of Cantonese. Importantly though, our interpretation in terms of lack of a C-bias in Cantonese is corroborated by two recent studies having explored similar issues in either Mandarin- (Wiener and Turnbull, <xref ref-type="bibr" rid="B78">2016</xref>) or Cantonese-speaking (G&#x000F3;mez et al., <xref ref-type="bibr" rid="B32">2017</xref>) adults, using a word reconstruction and word form segmentation task respectively. As discussed in the introduction, their findings differ from those previously found in Indo-European languages, failing to find a clear C-bias in both languages, thus suggesting a different balance in the weight given to consonants and vowels in these two tone languages.</p>
<p>The above findings that begin to establish crosslinguistic differences in consonant/vowel weight between adult listeners of tonal vs. non-tonal languages are to be considered in relation to infant studies on Indo-European languages that have shown that the C-bias is modulated in infancy both developmentally and crosslinguistically (see Nazzi et al., <xref ref-type="bibr" rid="B56">2016</xref>, for a complete review). Indeed, results from French and Italian show that while a C-bias is found from 8 months onward (Nazzi, <xref ref-type="bibr" rid="B54">2005</xref>; Hochmann et al., <xref ref-type="bibr" rid="B42">2011</xref>; Poltrock and Nazzi, <xref ref-type="bibr" rid="B63">2015</xref>; Nishibayashi and Nazzi, <xref ref-type="bibr" rid="B60">2016</xref>), it is not present up to 6 months of age (Benavides-Varela et al., <xref ref-type="bibr" rid="B5">2012</xref>; Bouchon et al., <xref ref-type="bibr" rid="B12">2015</xref>; Nishibayashi and Nazzi, <xref ref-type="bibr" rid="B60">2016</xref>; Hochmann et al., <xref ref-type="bibr" rid="B41">2018</xref>). Moreover, a C-bias could not be attested before 30 months in British English-learning infants (Nazzi et al., <xref ref-type="bibr" rid="B55">2009</xref>; Floccia et al., <xref ref-type="bibr" rid="B28">2014</xref>), and Danish-learning 20-month-olds demonstrate a V-bias (H&#x000F8;jen and Nazzi, <xref ref-type="bibr" rid="B43">2016</xref>). Taken together, these studies suggest that the C-bias is acquired and that its acquisition depends on the phonological and lexical properties of the native language.</p>
<p>Given that Cantonese- and Mandarin-speaking adults appear to have a reduced or reversed bias (Wiener and Turnbull, <xref ref-type="bibr" rid="B78">2016</xref>; G&#x000F3;mez et al., <xref ref-type="bibr" rid="B32">2017</xref>; present study), it is of great interest to expand research on the consonant bias to infants and toddlers learning a tone language, which was one of the original motivations for setting up the present study. At present, only one study has started to explore this issue in (Mandarin-dominant) Mandarin-English bilingual toddlers (aged 2.5&#x02013;3.5 years) and preschoolers (aged 4&#x02013;5 years). In a word recognition task exploring their sensitivity to mispronunciations of known words, the toddlers were found to be more sensitive to tone than consonant and vowel mispronunciations, while the reverse pattern was found in preschoolers (Singh et al., <xref ref-type="bibr" rid="B66">2015</xref>). However, at both ages, no differences in sensitivity were found between consonant and vowel mispronunciations. Future studies will have to expand on this first finding, exploring such effects in younger monolingual infants learning various tone languages, and exploring various aspects of lexical processing, including both word learning and lexical comprehension.</p>
<p>In conclusion, the present study establishes adults&#x00027; word learning abilities in an unknown language, and show that level of performance is modulated by how the phonologies of the native and non-native languages map onto each other. They also bring evidence suggesting that being a speaker of a tonal language reduces the consonant bias in lexical processing previously found in adults of several Indo-European languages, probably due to the fact that tones are carried by vowels more than by consonants. However, no clear bias could be found for either consonants or vowels, and future studies will have to further probe the link between phonological and lexical processing in tone languages. These findings nevertheless set up the foundations for equivalent developmental studies that will inform our understanding of what determines the phonological biases that are observed in lexical processing.</p>
</sec>
<sec id="s5">
<title>Author contributions</title>
<p>SP created the experimental stimuli and design, conducted data analyses and drafted the manuscript. CK assisted in stimuli creation, recruited participants and collected data in Hong Kong. HuiC interpretated the data and drafted the manuscript. HC and TN conceptualized the study and supervised all stages of the project.</p>
<sec>
<title>Conflict of interest statement</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
</sec>
</body>
<back>
<ack><p>This work was partly funded by ANR-13-BSH2-0004 to TN, and ANR-15-CE28-0011 to TN and HC. We would like to thank the participants for their time, and Sylvie Margules for help running the experiment.</p>
</ack>
<sec sec-type="supplementary-material" id="s6">
<title>Supplementary material</title>
<p>The Supplementary Material for this article can be found online at: <ext-link ext-link-type="uri" xlink:href="https://www.frontiersin.org/articles/10.3389/fpsyg.2018.01211/full#supplementary-material">https://www.frontiersin.org/articles/10.3389/fpsyg.2018.01211/full#supplementary-material</ext-link></p>
<supplementary-material xlink:href="Data_Sheet_1.docx" id="SM1" mimetype="application/vnd.openxmlformats-officedocument.wordprocessingml.document" xmlns:xlink="http://www.w3.org/1999/xlink"/>
<supplementary-material xlink:href="Audio_1.WAV" id="SM2" mimetype="audio/x-wav" xmlns:xlink="http://www.w3.org/1999/xlink"/>
<supplementary-material xlink:href="Audio_2.WAV" id="SM3" mimetype="audio/x-wav" xmlns:xlink="http://www.w3.org/1999/xlink"/>
<supplementary-material xlink:href="Audio_3.WAV" id="SM4" mimetype="audio/x-wav" xmlns:xlink="http://www.w3.org/1999/xlink"/>
</sec>
<ref-list>
<title>References</title>
<ref id="B1">
<citation citation-type="web"><person-group person-group-type="author"><name><surname>Baayen</surname> <given-names>R. H.</given-names></name></person-group> (<year>2015</year>). <source>languageR: Data sets and Functions with R. Analyzing Linguistic Data: A Practical Introduction to Statistics Using R. R package version 1.4.1</source>. Available online at: <ext-link ext-link-type="uri" xlink:href="https://CRAN.R-project.org/package=languageR">https://CRAN.R-project.org/package=languageR</ext-link></citation></ref>
<ref id="B2">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Barr</surname> <given-names>D. J.</given-names></name></person-group> (<year>2008</year>). <article-title>Analyzing &#x02018;visual world&#x02019; eyetracking data using multilevel logistic regression</article-title>. <source>J. Mem. Lang.</source> <volume>59</volume>, <fpage>457</fpage>&#x02013;<lpage>474</lpage>. <pub-id pub-id-type="doi">10.1016/j.jml.2007.09.002</pub-id></citation></ref>
<ref id="B3">
<citation citation-type="web"><person-group person-group-type="author"><name><surname>Bates</surname> <given-names>D.</given-names></name> <name><surname>Maechler</surname> <given-names>M.</given-names></name> <name><surname>Bolker</surname> <given-names>B.</given-names></name> <name><surname>Walker</surname> <given-names>S.</given-names></name></person-group> (<year>2015</year>). <source>lme4: Linear Mixed-Effects Models Using Eigen and S4. R package version 1.1-7</source>. Available online at: <ext-link ext-link-type="uri" xlink:href="http://CRAN.R-project.org/package=lme4">http://CRAN.R-project.org/package=lme4</ext-link></citation></ref>
<ref id="B4">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Bauer</surname> <given-names>R. S.</given-names></name> <name><surname>Benedict</surname> <given-names>P. K.</given-names></name></person-group> (<year>1997</year>). <source>Modern Cantonese Phonology.</source> <publisher-loc>Berlin; New York, NY</publisher-loc>: <publisher-name>Mouton de Gruyter</publisher-name>.</citation></ref>
<ref id="B5">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Benavides-Varela</surname> <given-names>S.</given-names></name> <name><surname>Hochmann</surname> <given-names>J. R.</given-names></name> <name><surname>Macagno</surname> <given-names>F.</given-names></name> <name><surname>Nespor</surname> <given-names>M.</given-names></name> <name><surname>Mehler</surname> <given-names>J.</given-names></name></person-group> (<year>2012</year>). <article-title>Newborn&#x00027;s brain activity signals the origin of word memories</article-title>. <source>Proc. Natl. Acad. Sci.</source> <volume>109</volume>, <fpage>17908</fpage>&#x02013;<lpage>17913</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.1205413109</pub-id><pub-id pub-id-type="pmid">23071325</pub-id></citation></ref>
<ref id="B6">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Best</surname> <given-names>C. T.</given-names></name></person-group> (<year>1995</year>). <article-title>A direct realist view of cross-language speech perception</article-title>, in <source>Speech Perception and Linguistic Experience: Issues in Cross-Language Research</source>, ed <person-group person-group-type="editor"><name><surname>Strange</surname> <given-names>W.</given-names></name></person-group> (<publisher-loc>Timonium, MD</publisher-loc>: <publisher-name>York Press</publisher-name>), <fpage>171</fpage>&#x02013;<lpage>204</lpage>.</citation></ref>
<ref id="B7">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Best</surname> <given-names>C. T.</given-names></name> <name><surname>McRoberts</surname> <given-names>G. W.</given-names></name> <name><surname>Sithole</surname> <given-names>N. M.</given-names></name></person-group> (<year>1988</year>). <article-title>Examination of perceptual reorganization for nonnative speech contrasts: zulu click discrimination by English-speaking adults and infants</article-title>. <source>J. Exp. Psychol.</source> <volume>14</volume>, <fpage>345</fpage>&#x02013;<lpage>360</lpage>. <pub-id pub-id-type="doi">10.1037/0096-1523.14.3.345</pub-id><pub-id pub-id-type="pmid">2971765</pub-id></citation></ref>
<ref id="B8">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Best</surname> <given-names>C. T.</given-names></name> <name><surname>Tyler</surname> <given-names>M.</given-names></name></person-group> (<year>2007</year>). <article-title>Nonnative and second-language speech learning: the role of language experience in speech perception and production</article-title>, in <source>Language Experience in Second Language Speech Learning: In Honor of James E. Flege</source>, eds <person-group person-group-type="editor"><name><surname>Bohn</surname> <given-names>O. D.</given-names></name> <name><surname>Munro</surname> <given-names>M.</given-names></name></person-group> (<publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins</publisher-name>), <fpage>13</fpage>&#x02013;<lpage>24</lpage>.</citation></ref>
<ref id="B9">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Birdsong</surname> <given-names>D.</given-names></name> <name><surname>Molis</surname> <given-names>M.</given-names></name></person-group> (<year>2001</year>). <article-title>On the evidence for maturational effects in second language acquisition</article-title>. <source>J. Memory Lang.</source> <volume>44</volume>, <fpage>235</fpage>&#x02013;<lpage>249</lpage>. <pub-id pub-id-type="doi">10.1006/jmla.2000.2750</pub-id></citation></ref>
<ref id="B10">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Boll-Avetisyan</surname> <given-names>N.</given-names></name> <name><surname>Bhatara</surname> <given-names>A.</given-names></name> <name><surname>Unger</surname> <given-names>A.</given-names></name> <name><surname>Nazzi</surname> <given-names>T.</given-names></name> <name><surname>H&#x000F6;hle</surname> <given-names>B.</given-names></name></person-group> (<year>2016</year>). <article-title>Effects of experience with L2 and music on rhythmic grouping by French listeners</article-title>. <source>Biling. Lang. Cogn.</source> <volume>19</volume>, <fpage>971</fpage>&#x02013;<lpage>986</lpage>. <pub-id pub-id-type="doi">10.1017/S1366728915000425</pub-id></citation></ref>
<ref id="B11">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bonatti</surname> <given-names>L. L.</given-names></name> <name><surname>Pe&#x000F1;a</surname> <given-names>M.</given-names></name> <name><surname>Nespor</surname> <given-names>M.</given-names></name> <name><surname>Mehler</surname> <given-names>J.</given-names></name></person-group> (<year>2005</year>). <article-title>Linguistic constraints on statistical computations: the role of consonants and vowels in continuous speech processing</article-title>. <source>Psychol. Sci.</source> <volume>16</volume>, <fpage>451</fpage>&#x02013;<lpage>459</lpage>. <pub-id pub-id-type="doi">10.1111/j.0956-7976.2005.01556.x</pub-id><pub-id pub-id-type="pmid">15943671</pub-id></citation></ref>
<ref id="B12">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bouchon</surname> <given-names>C.</given-names></name> <name><surname>Floccia</surname> <given-names>C.</given-names></name> <name><surname>Fux</surname> <given-names>T.</given-names></name> <name><surname>Adda-Decker</surname> <given-names>M.</given-names></name> <name><surname>Nazzi</surname> <given-names>T.</given-names></name></person-group> (<year>2015</year>). <article-title>Call me Alix, not Elix: vowels are more important than consonants in own-name recognition at 5 months</article-title>. <source>Dev. Sci.</source> <volume>18</volume>, <fpage>587</fpage>&#x02013;<lpage>598</lpage>. <pub-id pub-id-type="doi">10.1111/desc.12242</pub-id></citation></ref>
<ref id="B13">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cabrera</surname> <given-names>L.</given-names></name> <name><surname>Tsao</surname> <given-names>F. M.</given-names></name> <name><surname>Liu</surname> <given-names>H. M.</given-names></name> <name><surname>Li</surname> <given-names>L. Y.</given-names></name> <name><surname>Hu</surname> <given-names>Y. H.</given-names></name> <name><surname>Lorenzi</surname> <given-names>C.</given-names></name> <etal/></person-group>. (<year>2015</year>). <article-title>The perception of speech modulation cues in lexical tones is guided by early language-specific experience</article-title>. <source>Front. Psychol.</source> <volume>6</volume>:<fpage>1290</fpage>. <pub-id pub-id-type="doi">10.3389/fpsyg.2015.01290</pub-id><pub-id pub-id-type="pmid">26379605</pub-id></citation></ref>
<ref id="B14">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Carreiras</surname> <given-names>M.</given-names></name> <name><surname>Du&#x000F1;abeitia</surname> <given-names>J. A.</given-names></name> <name><surname>Molinaro</surname> <given-names>N.</given-names></name></person-group> (<year>2009</year>). <article-title>Consonants and vowels contribute differently to visual word recognition: ERPs of relative position priming</article-title>. <source>Cereb. Cortex</source> <volume>19</volume>, <fpage>2659</fpage>&#x02013;<lpage>2670</lpage>. <pub-id pub-id-type="doi">10.1093/cercor/bhp019</pub-id><pub-id pub-id-type="pmid">19273459</pub-id></citation></ref>
<ref id="B15">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chandrasekaran</surname> <given-names>B.</given-names></name> <name><surname>Sampath</surname> <given-names>P. D.</given-names></name> <name><surname>Wong</surname> <given-names>P. C.</given-names></name></person-group> (<year>2010</year>). <article-title>Individual variability in cue-weighting and lexical tone learning</article-title>. <source>J. Acoust. Soc. Am.</source> <volume>128</volume>, <fpage>456</fpage>&#x02013;<lpage>465</lpage>. <pub-id pub-id-type="doi">10.1121/1.3445785</pub-id><pub-id pub-id-type="pmid">20649239</pub-id></citation></ref>
<ref id="B16">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cheng</surname> <given-names>R. L.</given-names></name></person-group> (<year>1966</year>). <article-title>Mandarin phonological structure</article-title>. <source>J. Linguist.</source> <volume>2</volume>, <fpage>135</fpage>&#x02013;<lpage>158</lpage>. <pub-id pub-id-type="doi">10.1017/S0022226700001444</pub-id></citation></ref>
<ref id="B17">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cooper</surname> <given-names>A.</given-names></name> <name><surname>Wang</surname> <given-names>Y.</given-names></name></person-group> (<year>2012</year>). <article-title>The influence of linguistic and musical experience on Cantonese word learning</article-title>. <source>J. Acoust. Soc. Am.</source> <volume>131</volume>, <fpage>4756</fpage>&#x02013;<lpage>4769</lpage>. <pub-id pub-id-type="doi">10.1121/1.4714355</pub-id><pub-id pub-id-type="pmid">22712948</pub-id></citation></ref>
<ref id="B18">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cooper</surname> <given-names>A.</given-names></name> <name><surname>Wang</surname> <given-names>Y.</given-names></name></person-group> (<year>2013</year>). <article-title>Effects of tone training on Cantonese tone-word learning</article-title>. <source>J. Acoust. Soc. Am.</source> <volume>134</volume>, <fpage>EL133</fpage>&#x02013;<lpage>EL139</lpage>. <pub-id pub-id-type="doi">10.1121/1.4812435</pub-id><pub-id pub-id-type="pmid">23927215</pub-id></citation></ref>
<ref id="B19">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Creel</surname> <given-names>S. C.</given-names></name> <name><surname>Aslin</surname> <given-names>R. N.</given-names></name> <name><surname>Tanenhaus</surname> <given-names>M. K.</given-names></name></person-group> (<year>2006</year>). <article-title>Acquiring an artificial lexicon: segment type and order information in early lexical entries</article-title>. <source>J. Mem. Lang.</source> <volume>54</volume>, <fpage>1</fpage>&#x02013;<lpage>19</lpage>. <pub-id pub-id-type="doi">10.1016/j.jml.2005.09.003</pub-id></citation></ref>
<ref id="B20">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cutler</surname> <given-names>A.</given-names></name> <name><surname>Sebasti&#x000E1;n-Gall&#x000E9;s</surname> <given-names>N.</given-names></name> <name><surname>Soler-Vilageliu</surname> <given-names>O.</given-names></name> <name><surname>van Ooijen</surname> <given-names>B.</given-names></name></person-group> (<year>2000</year>). <article-title>Constraints of vowels and consonants on lexical selection: cross-linguistic comparisons</article-title>. <source>Mem. Cognit.</source> <volume>28</volume>, <fpage>746</fpage>&#x02013;<lpage>755</lpage>. <pub-id pub-id-type="doi">10.3758/BF03198409</pub-id><pub-id pub-id-type="pmid">10983448</pub-id></citation></ref>
<ref id="B21">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>DeLancey</surname> <given-names>S.</given-names></name></person-group> (<year>2009</year>). <article-title>Sino-Tibetan languages</article-title>, in <source>The World&#x00027;s Major Languages 2nd Edn</source>, ed <person-group person-group-type="editor"><name><surname>Comrie</surname> <given-names>B.</given-names></name></person-group> (<publisher-loc>London; New York, NY</publisher-loc>: <publisher-name>Routledge</publisher-name>), <fpage>693</fpage>&#x02013;<lpage>702</lpage>.</citation></ref>
<ref id="B22">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Delle Luche</surname> <given-names>C.</given-names></name> <name><surname>Poltrock</surname> <given-names>S.</given-names></name> <name><surname>Goslin</surname> <given-names>J.</given-names></name> <name><surname>New</surname> <given-names>B.</given-names></name> <name><surname>Floccia</surname> <given-names>C.</given-names></name> <name><surname>Nazzi</surname> <given-names>T.</given-names></name></person-group> (<year>2014</year>). <article-title>Differential processing of consonants and vowels in the auditory modality: a cross-linguistic study</article-title>. <source>J. Mem. Lang.</source> <volume>72</volume>, <fpage>1</fpage>&#x02013;<lpage>15</lpage>. <pub-id pub-id-type="doi">10.1016/j.jml.2013.12.001</pub-id></citation></ref>
<ref id="B23">
<citation citation-type="web"><person-group person-group-type="author"><name><surname>Dink</surname> <given-names>J. W.</given-names></name> <name><surname>Ferguson</surname> <given-names>B.</given-names></name></person-group> (<year>2015</year>). <source>eyetrackingR: An R Library for Eye-tracking Data Analysis</source>. Available online at <ext-link ext-link-type="uri" xlink:href="http://www.eyetrackingr.com">http://www.eyetrackingr.com</ext-link>.</citation></ref>
<ref id="B24">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dittinger</surname> <given-names>E.</given-names></name> <name><surname>Barbaroux</surname> <given-names>M.</given-names></name> <name><surname>D&#x00027;Imperio</surname> <given-names>M.</given-names></name> <name><surname>J&#x000E4;ncke</surname> <given-names>L.</given-names></name> <name><surname>Elmer</surname> <given-names>S.</given-names></name> <name><surname>Besson</surname> <given-names>M.</given-names></name></person-group> (<year>2016</year>). <article-title>Professional music training and novel word learning: from faster semantic encoding to longer-lasting word representations</article-title>. <source>J. Cogn. Neurosci.</source> <volume>28</volume>, <fpage>1584</fpage>&#x02013;<lpage>1602</lpage>. <pub-id pub-id-type="doi">10.1162/jocn_a_00997</pub-id><pub-id pub-id-type="pmid">27315272</pub-id></citation></ref>
<ref id="B25">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Duanmu</surname> <given-names>S.</given-names></name></person-group> (<year>2000</year>). <source>The Phonology of Standard Chinese.</source> <publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>.</citation></ref>
<ref id="B26">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dupoux</surname> <given-names>E.</given-names></name> <name><surname>Sebastian-Gall&#x000E9;s</surname> <given-names>N.</given-names></name> <name><surname>Navarrete</surname> <given-names>E.</given-names></name> <name><surname>Peperkamp</surname> <given-names>S.</given-names></name></person-group> (<year>2008</year>). <article-title>Persistent stress &#x0201C;deafness&#x0201D;: The case of French learners of Spanish</article-title>. <source>Cognition</source>, <volume>106</volume>, <fpage>682</fpage>&#x02013;<lpage>706</lpage>. <pub-id pub-id-type="doi">10.1016/j.cognition.2007.04.001</pub-id><pub-id pub-id-type="pmid">17592731</pub-id></citation></ref>
<ref id="B27">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Flege</surname> <given-names>J. E.</given-names></name> <name><surname>Yeni-Komshian</surname> <given-names>G. H.</given-names></name> <name><surname>Liu</surname> <given-names>S.</given-names></name></person-group> (<year>1999</year>). <article-title>Age constraints on second-language acquisition</article-title>. <source>J. Mem. Lang.</source> <volume>41</volume>, <fpage>78</fpage>&#x02013;<lpage>104</lpage>. <pub-id pub-id-type="doi">10.1006/jmla.1999.2638</pub-id></citation></ref>
<ref id="B28">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Floccia</surname> <given-names>C.</given-names></name> <name><surname>Nazzi</surname> <given-names>T.</given-names></name> <name><surname>Delle Luche</surname> <given-names>C.</given-names></name> <name><surname>Poltrock</surname> <given-names>S.</given-names></name> <name><surname>Goslin</surname> <given-names>J.</given-names></name></person-group> (<year>2014</year>). <article-title>English-learning one-to two-year-olds do not show a consonant bias in word learning</article-title>. <source>J. Child Lang.</source> <volume>41</volume>, <fpage>1085</fpage>&#x02013;<lpage>1114</lpage>. <pub-id pub-id-type="doi">10.1017/S0305000913000287</pub-id></citation></ref>
<ref id="B29">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Francis</surname> <given-names>A. L.</given-names></name> <name><surname>Ciocca</surname> <given-names>V.</given-names></name> <name><surname>Ma</surname> <given-names>L.</given-names></name> <name><surname>Fenn</surname> <given-names>K.</given-names></name></person-group> (<year>2008</year>). <article-title>Perceptual learning of Cantonese lexical tones by tone and non-tone language speakers</article-title>. <source>J. Phonet.</source> <volume>36</volume>, <fpage>268</fpage>&#x02013;<lpage>294</lpage>. <pub-id pub-id-type="doi">10.1016/j.wocn.2007.06.005</pub-id></citation></ref>
<ref id="B30">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Gandour</surname> <given-names>J.</given-names></name></person-group> (<year>1978</year>). <article-title>The perception of tone</article-title>, in <source>Tone: A Linguistic Survey</source>, ed <person-group person-group-type="editor"><name><surname>Fromkin</surname> <given-names>A. V.</given-names></name></person-group> (<publisher-loc>New York</publisher-loc>: <publisher-name>Academic Press</publisher-name>), <fpage>41</fpage>&#x02013;<lpage>76</lpage>.</citation></ref>
<ref id="B31">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gandour</surname> <given-names>J.</given-names></name> <name><surname>Wong</surname> <given-names>D.</given-names></name> <name><surname>Hsieh</surname> <given-names>L.</given-names></name> <name><surname>Weinzapfel</surname> <given-names>B.</given-names></name> <name><surname>Lancker</surname> <given-names>D.</given-names></name> <name><surname>Van Hutchins</surname> <given-names>G. D.</given-names></name></person-group> (<year>2000</year>). <article-title>A crosslinguistic PET study of tone perception</article-title>. <source>J. Cogn. Neurosci.</source> <volume>12</volume>, <fpage>207</fpage>&#x02013;<lpage>222</lpage>. <pub-id pub-id-type="doi">10.1162/089892900561841</pub-id><pub-id pub-id-type="pmid">10769317</pub-id></citation></ref>
<ref id="B32">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>G&#x000F3;mez</surname> <given-names>D. M.</given-names></name> <name><surname>Mok</surname> <given-names>P.</given-names></name> <name><surname>Ordin</surname> <given-names>M.</given-names></name> <name><surname>Mehler</surname> <given-names>J.</given-names></name> <name><surname>Nespor</surname> <given-names>M.</given-names></name></person-group> (<year>2017</year>). <article-title>Statistical Speech Segmentation in Tone languages: the role of lexical tones</article-title>. <source>Lang. Speech</source> <volume>61</volume>, <fpage>84</fpage>&#x02013;<lpage>96</lpage>. <pub-id pub-id-type="doi">10.1177/0023830917706529</pub-id><pub-id pub-id-type="pmid">28486862</pub-id></citation></ref>
<ref id="B33">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gong</surname> <given-names>H. C.</given-names></name></person-group> (<year>1980</year>). <article-title>A comparative study of the chinese, tibetan, and burmese vowel systems</article-title>. <source>Bull. Instit. Hist. Philol. Acad. Sin.</source> <volume>51</volume>, <fpage>455</fpage>&#x02013;<lpage>489</lpage>.</citation></ref>
<ref id="B34">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gonzalez-Gomez</surname> <given-names>N.</given-names></name> <name><surname>Poltrock</surname> <given-names>S.</given-names></name> <name><surname>Nazzi</surname> <given-names>T.</given-names></name></person-group> (<year>2013</year>). <article-title>A &#x0201C;bat&#x0201D; is easier to learn than a &#x0201C;tab&#x0201D;: Effects of relative phonotactic frequency on infant word learning</article-title>. <source>PLOS ONE</source> <volume>8</volume>:<fpage>e59601</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pone.0059601</pub-id><pub-id pub-id-type="pmid">23527227</pub-id></citation></ref>
<ref id="B35">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hall&#x000E9;</surname> <given-names>P. A.</given-names></name> <name><surname>Chang</surname> <given-names>Y.-C.</given-names></name> <name><surname>Best</surname> <given-names>C. T.</given-names></name></person-group> (<year>2004</year>). <article-title>Identification and discrimination of Mandarin Chinese tones by Mandarin Chinese vs</article-title>. <source>French listen. J. Phonet.</source> <volume>32</volume>, <fpage>395</fpage>&#x02013;<lpage>421</lpage>. <pub-id pub-id-type="doi">10.1016/S0095-4470(03)00016-0</pub-id></citation></ref>
<ref id="B36">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Harrison</surname> <given-names>P. A.</given-names></name></person-group> (<year>1998</year>). <article-title>Yoruba babies and unchained melody</article-title>. <source>UCL Working Papers in Phonet.</source> <volume>10</volume>, <fpage>33</fpage>&#x02013;<lpage>52</lpage>.</citation></ref>
<ref id="B37">
<citation citation-type="other"><person-group person-group-type="author"><name><surname>Harrison</surname> <given-names>P. A.</given-names></name></person-group> (<year>1999</year>). <source>The Acquisition of Phonology in the First Year of Life</source>. Ph.D. Dissertation, University College London.</citation></ref>
<ref id="B38">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Harrison</surname> <given-names>P. A.</given-names></name></person-group> (<year>2000</year>). <article-title>Acquiring the phonology of lexical tone in infancy</article-title>. <source>Lingua</source> <volume>110</volume>, <fpage>581</fpage>&#x02013;<lpage>616</lpage>. <pub-id pub-id-type="doi">10.1016/S0024-3841(00)00003-6</pub-id></citation></ref>
<ref id="B39">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Hashimoto</surname> <given-names>A. O.-K. Y.</given-names></name></person-group> (<year>1972</year>). <source>Phonology of Cantonese.</source> <publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>.</citation></ref>
<ref id="B40">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Havy</surname> <given-names>M.</given-names></name> <name><surname>Serres</surname> <given-names>J.</given-names></name> <name><surname>Nazzi</surname> <given-names>T.</given-names></name></person-group> (<year>2014</year>). <article-title>A consonant/vowel asymmetry in word-form processing: evidence in childhood and in adulthood</article-title>. <source>Lang. Speech</source> <volume>57</volume>, <fpage>254</fpage>&#x02013;<lpage>281</lpage>. <pub-id pub-id-type="doi">10.1177/0023830913507693</pub-id></citation></ref>
<ref id="B41">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hochmann</surname> <given-names>J. R.</given-names></name> <name><surname>Benavides-Varela</surname> <given-names>S.</given-names></name> <name><surname>Fl,&#x000F3;</surname> <given-names>A.</given-names></name> <name><surname>Nespor</surname> <given-names>M.</given-names></name> <name><surname>Mehler</surname> <given-names>J.</given-names></name></person-group> (<year>2018</year>). <article-title>Bias for vocalic over consonantal information in 6-month-olds</article-title>. <source>Infancy</source> <volume>23</volume>, <fpage>136</fpage>&#x02013;<lpage>151</lpage>. <pub-id pub-id-type="doi">10.1111/infa.12203</pub-id></citation></ref>
<ref id="B42">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hochmann</surname> <given-names>J. R.</given-names></name> <name><surname>Benavides-Varela</surname> <given-names>S.</given-names></name> <name><surname>Nespor</surname> <given-names>M.</given-names></name> <name><surname>Mehler</surname> <given-names>J.</given-names></name></person-group> (<year>2011</year>). <article-title>Consonants and vowels: different roles in early language acquisition</article-title>. <source>Dev. Sci.</source> <volume>14</volume>, <fpage>1445</fpage>&#x02013;<lpage>1458</lpage>. <pub-id pub-id-type="doi">10.1111/j.1467-7687.2011.01089.x</pub-id><pub-id pub-id-type="pmid">22010902</pub-id></citation></ref>
<ref id="B43">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>H&#x000F8;jen</surname> <given-names>A.</given-names></name> <name><surname>Nazzi</surname> <given-names>T.</given-names></name></person-group> (<year>2016</year>). <article-title>Vowel bias in danish word-learning: processing biases are language-specific</article-title>. <source>Dev. Sci.</source> <volume>19</volume>, <fpage>41</fpage>&#x02013;<lpage>49</lpage>. <pub-id pub-id-type="doi">10.1111/desc.12286</pub-id><pub-id pub-id-type="pmid">25660116</pub-id></citation></ref>
<ref id="B44">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Howie</surname> <given-names>J.</given-names></name></person-group> (<year>1976</year>). <source>Acoustical Studies of Mandarin Vowels and Tones.</source> <publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>.</citation></ref>
<ref id="B45">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Iverson</surname> <given-names>P.</given-names></name> <name><surname>Kuhl</surname> <given-names>P. K.</given-names></name> <name><surname>Akahane-yamada</surname> <given-names>R.</given-names></name> <name><surname>Diesch</surname> <given-names>E.</given-names></name></person-group> (<year>2003</year>). <article-title>A perceptual interference account of acquisition difficulties for non-native phonemes</article-title>. <source>Cognition</source> <volume>87</volume>, <fpage>B47</fpage>&#x02013;<lpage>B57</lpage>. <pub-id pub-id-type="doi">10.1016/S0010-0277(02)00198-1</pub-id><pub-id pub-id-type="pmid">12499111</pub-id></citation></ref>
<ref id="B46">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kuhl</surname> <given-names>P. K.</given-names></name> <name><surname>Williams</surname> <given-names>K. A.</given-names></name> <name><surname>Lacerda</surname> <given-names>F.</given-names></name> <name><surname>Stevens</surname> <given-names>K. N.</given-names></name> <name><surname>Lindblom</surname> <given-names>B.</given-names></name></person-group> (<year>1992</year>). <article-title>Linguistic experience alters phonetic perception in infants by 6 months of age</article-title>. <source>Science</source> <volume>255</volume>, <fpage>606</fpage>&#x02013;<lpage>608</lpage>. <pub-id pub-id-type="doi">10.1126/science.1736364</pub-id><pub-id pub-id-type="pmid">1736364</pub-id></citation></ref>
<ref id="B47">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Li</surname> <given-names>F.-K.</given-names></name></person-group> (<year>1937</year>). <article-title>Languages and Dialects, in Shih, Ch&#x00027;ao-ying; Chang, Ch&#x00027;i-hsien, The Chinese Year Book, Commercial Press, pp. 59&#x02013;65, reprinted as Li, Fang-Kuei. (1973), Languages and Dialects of China</article-title>, <source>J. Chin. Linguist.</source> <volume>1</volume>, <fpage>1</fpage>&#x02013;<lpage>13</lpage>.</citation></ref>
<ref id="B48">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Liu</surname> <given-names>L.</given-names></name> <name><surname>Kager</surname> <given-names>R.</given-names></name></person-group> (<year>2014</year>). <article-title>Perception of tones by infants learning a non-tone language</article-title>. <source>Cognition</source> <volume>133</volume>, <fpage>385</fpage>&#x02013;<lpage>394</lpage>. <pub-id pub-id-type="doi">10.1016/j.cognition.2014.06.004</pub-id><pub-id pub-id-type="pmid">25128796</pub-id></citation></ref>
<ref id="B49">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Maris</surname> <given-names>E.</given-names></name> <name><surname>Oostenveld</surname> <given-names>R.</given-names></name></person-group> (<year>2007</year>). <article-title>Nonparametric statistical testing of EEG-and MEG-data</article-title>. <source>J. Neurosci. Methods</source> <volume>164</volume>, <fpage>177</fpage>&#x02013;<lpage>190</lpage>. <pub-id pub-id-type="doi">10.1016/j.jneumeth.2007.03.024</pub-id><pub-id pub-id-type="pmid">17517438</pub-id></citation></ref>
<ref id="B50">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mattock</surname> <given-names>K.</given-names></name> <name><surname>Burnham</surname> <given-names>D.</given-names></name></person-group> (<year>2006</year>). <article-title>Chinese and English infants&#x00027; tone perception: evidence for perceptual reorganization</article-title>. <source>Infancy</source> <volume>10</volume>, <fpage>241</fpage>&#x02013;<lpage>265</lpage>. <pub-id pub-id-type="doi">10.1207/s15327078in1003_3</pub-id></citation></ref>
<ref id="B51">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mattock</surname> <given-names>K.</given-names></name> <name><surname>Molnar</surname> <given-names>M.</given-names></name> <name><surname>Polka</surname> <given-names>L.</given-names></name> <name><surname>Burnham</surname> <given-names>D.</given-names></name></person-group> (<year>2008</year>). <article-title>The developmental course of lexical tone perception in the first year of life</article-title>. <source>Cognition</source> <volume>106</volume>, <fpage>1367</fpage>&#x02013;<lpage>1381</lpage>. <pub-id pub-id-type="doi">10.1016/j.cognition.2007.07.002</pub-id><pub-id pub-id-type="pmid">17707789</pub-id></citation></ref>
<ref id="B52">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mirman</surname> <given-names>D.</given-names></name> <name><surname>Dixon</surname> <given-names>J. A.</given-names></name> <name><surname>Magnuson</surname> <given-names>J. S.</given-names></name></person-group> (<year>2008</year>). <article-title>Statistical and computational models of the visual world paradigm: growth curves and individual differences</article-title>. <source>J. Mem. Lang.</source> <volume>59</volume>, <fpage>475</fpage>&#x02013;<lpage>494</lpage>. <pub-id pub-id-type="doi">10.1016/j.jml.2007.11.006</pub-id><pub-id pub-id-type="pmid">19060958</pub-id></citation></ref>
<ref id="B53">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Miyawaki</surname> <given-names>K.</given-names></name> <name><surname>Strange</surname> <given-names>W.</given-names></name> <name><surname>Verbrugge</surname> <given-names>R.</given-names></name> <name><surname>Liberman</surname> <given-names>A.</given-names></name> <name><surname>Jenkins</surname> <given-names>J.</given-names></name> <name><surname>Fujimura</surname> <given-names>O.</given-names></name></person-group> (<year>1975</year>). <article-title>An effect of language experience: the discrimination of /r/ and /l/ by native speakers of Japanese and English</article-title>. <source>Percept. Psychophys.</source> <volume>18</volume>, <fpage>331</fpage>&#x02013;<lpage>340</lpage>. <pub-id pub-id-type="doi">10.3758/BF03211209</pub-id></citation></ref>
<ref id="B54">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Nazzi</surname> <given-names>T.</given-names></name></person-group> (<year>2005</year>). <article-title>Use of phonetic specificity during the acquisition of new words: differences between consonants and vowels</article-title>. <source>Cognition</source> <volume>98</volume>, <fpage>13</fpage>&#x02013;<lpage>30</lpage>. <pub-id pub-id-type="doi">10.1016/j.cognition.2004.10.005</pub-id><pub-id pub-id-type="pmid">16297674</pub-id></citation></ref>
<ref id="B55">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Nazzi</surname> <given-names>T.</given-names></name> <name><surname>Floccia</surname> <given-names>C.</given-names></name> <name><surname>Moquet</surname> <given-names>B.</given-names></name> <name><surname>Butler</surname> <given-names>J.</given-names></name></person-group> (<year>2009</year>). <article-title>Bias for consonantal information over vocalic information in 30-month-olds: Cross-linguistic evidence from French and English</article-title>. <source>J. Exp. Child Psychol.</source> <volume>102</volume>, <fpage>522</fpage>&#x02013;<lpage>537</lpage>. <pub-id pub-id-type="doi">10.1016/j.jecp.2008.05.003</pub-id><pub-id pub-id-type="pmid">18572185</pub-id></citation></ref>
<ref id="B56">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Nazzi</surname> <given-names>T.</given-names></name> <name><surname>Poltrock</surname> <given-names>S.</given-names></name> <name><surname>Von Holzen</surname> <given-names>K.</given-names></name></person-group> (<year>2016</year>). <article-title>The developmental origins of the consonant bias in lexical processing</article-title>. <source>Curr. Dir. Psychol. Sci.</source> <volume>25</volume>, <fpage>291</fpage>&#x02013;<lpage>296</lpage>. <pub-id pub-id-type="doi">10.1177/0963721416655786</pub-id></citation></ref>
<ref id="B57">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Nespor</surname> <given-names>M.</given-names></name> <name><surname>Pe&#x000F1;a</surname> <given-names>M.</given-names></name> <name><surname>Mehler</surname> <given-names>J.</given-names></name></person-group> (<year>2003</year>). <article-title>On the different roles of vowels and consonants in speech processing and language acquisition</article-title>. <source>Lingue E Linguaggio</source> <volume>2</volume>, <fpage>203</fpage>&#x02013;<lpage>230</lpage>. <pub-id pub-id-type="doi">10.1418/10879</pub-id></citation></ref>
<ref id="B58">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>New</surname> <given-names>B.</given-names></name> <name><surname>Ara&#x000FA;jo</surname> <given-names>V.</given-names></name> <name><surname>Nazzi</surname> <given-names>T.</given-names></name></person-group> (<year>2008</year>). <article-title>Differential processing of consonants and vowels in lexical access through reading</article-title>. <source>Psychol. Sci.</source> <volume>19</volume>, <fpage>1223</fpage>&#x02013;<lpage>1227</lpage>. <pub-id pub-id-type="doi">10.1111/j.1467-9280.2008.02228.x</pub-id><pub-id pub-id-type="pmid">19121127</pub-id></citation></ref>
<ref id="B59">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>New</surname> <given-names>B.</given-names></name> <name><surname>Nazzi</surname> <given-names>T.</given-names></name></person-group> (<year>2014</year>). <article-title>The time course of consonant and vowel processing during word recognition</article-title>. <source>Lang. Cogn. Neurosci.</source> <volume>29</volume>, <fpage>147</fpage>&#x02013;<lpage>157</lpage>. <pub-id pub-id-type="doi">10.1080/01690965.2012.735678</pub-id></citation></ref>
<ref id="B60">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Nishibayashi</surname> <given-names>L. L.</given-names></name> <name><surname>Nazzi</surname> <given-names>T.</given-names></name></person-group> (<year>2016</year>). <article-title>Vowels, then consonants: early bias switch in recognizing segmented word forms</article-title>. <source>Cognition</source> <volume>155</volume>, <fpage>188</fpage>&#x02013;<lpage>203</lpage>. <pub-id pub-id-type="doi">10.1016/j.cognition.2016.07.003</pub-id><pub-id pub-id-type="pmid">27428809</pub-id></citation></ref>
<ref id="B61">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Polka</surname> <given-names>L.</given-names></name></person-group> (<year>1995</year>). <article-title>Linguistic influences in adult perception of non-native vowel contrasts</article-title>. <source>J. Acoust. Soc. Am.</source> <volume>97</volume>, <fpage>1286</fpage>&#x02013;<lpage>1296</lpage>. <pub-id pub-id-type="doi">10.1121/1.412170</pub-id><pub-id pub-id-type="pmid">7876448</pub-id></citation></ref>
<ref id="B62">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Polka</surname> <given-names>L.</given-names></name> <name><surname>Werker</surname> <given-names>J. F.</given-names></name></person-group> (<year>1994</year>). <article-title>Developmental changes in perception of nonnative vowel contrasts</article-title>. <source>J. Exp. Psychol.</source> <volume>20</volume>, <fpage>421</fpage>&#x02013;<lpage>435</lpage>. <pub-id pub-id-type="doi">10.1037/0096-1523.20.2.421</pub-id><pub-id pub-id-type="pmid">8189202</pub-id></citation></ref>
<ref id="B63">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Poltrock</surname> <given-names>S.</given-names></name> <name><surname>Nazzi</surname> <given-names>T.</given-names></name></person-group> (<year>2015</year>). <article-title>Consonant/vowel asymmetry in early word form recognition</article-title>. <source>J. Exp. Child Psychol.</source> <volume>131</volume>, <fpage>135</fpage>&#x02013;<lpage>148</lpage>. <pub-id pub-id-type="doi">10.1016/j.jecp.2014.11.011</pub-id><pub-id pub-id-type="pmid">25544396</pub-id></citation></ref>
<ref id="B64">
<citation citation-type="book"><person-group person-group-type="author"><collab>R Core Team</collab></person-group> (<year>2017</year>). <source>R: A Language and Environment for Statistical Computing</source>. <publisher-loc>Vienna</publisher-loc>: <publisher-name>R Foundation for Statistical Computing</publisher-name>. Available online at: <ext-link ext-link-type="uri" xlink:href="https://www.R-project.org/">https://www.R-project.org/</ext-link></citation></ref>
<ref id="B65">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Rivera-Gaxiola</surname> <given-names>M.</given-names></name> <name><surname>Silva-Pereyra</surname> <given-names>J.</given-names></name> <name><surname>Kuhl</surname> <given-names>P. K.</given-names></name></person-group> (<year>2005</year>). <article-title>Brain potentials to native and non-native speech contrasts in 7- and 11-month-old American infants</article-title>. <source>Dev. Sci.</source> <volume>8</volume>, <fpage>162</fpage>&#x02013;<lpage>172</lpage>. <pub-id pub-id-type="doi">10.1111/j.1467-7687.2005.00403.x</pub-id><pub-id pub-id-type="pmid">15720374</pub-id></citation></ref>
<ref id="B66">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Singh</surname> <given-names>L.</given-names></name> <name><surname>Goh</surname> <given-names>H. H.</given-names></name> <name><surname>Wewalaarachchi</surname> <given-names>T. D.</given-names></name></person-group> (<year>2015</year>). <article-title>Spoken word recognition in early childhood: comparative effects of vowel, consonant and lexical tone variation</article-title>. <source>Cognition</source> <volume>142</volume>, <fpage>1</fpage>&#x02013;<lpage>11</lpage>. <pub-id pub-id-type="doi">10.1016/j.cognition.2015.05.010</pub-id><pub-id pub-id-type="pmid">26010558</pub-id></citation></ref>
<ref id="B67">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>So</surname> <given-names>C. K.</given-names></name></person-group> (<year>2010</year>). <article-title>Categorizing Mandarin tones into Japanese pitch-accent categories:The role of phonetic properties</article-title>, in <source>Proceedings of Interspeech 2010 Satellite Workshopon Second Language Studies</source> (<publisher-loc>Tokyo</publisher-loc>).</citation></ref>
<ref id="B68">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>So</surname> <given-names>C. K.</given-names></name> <name><surname>Best</surname> <given-names>C. T.</given-names></name></person-group> (<year>2008</year>). <article-title>Do English speakers assimilate Mandarin tones to Englishprosodic categories?</article-title>, in <source>Proceedings of Interspeech 2008</source> (<publisher-loc>Baixas</publisher-loc>), <fpage>1120</fpage>.</citation></ref>
<ref id="B69">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>So</surname> <given-names>C. K.</given-names></name> <name><surname>Best</surname> <given-names>C. T.</given-names></name></person-group> (<year>2010</year>). <article-title>Cross-language perception of non-native tonal contrasts: effects of native phonological and phonetic influences</article-title>. <source>Lang. Speech</source> <volume>53</volume>, <fpage>273</fpage>&#x02013;<lpage>293</lpage>. <pub-id pub-id-type="doi">10.1177/0023830909357156</pub-id><pub-id pub-id-type="pmid">20583732</pub-id></citation></ref>
<ref id="B70">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>So</surname> <given-names>C. K.</given-names></name> <name><surname>Best</surname> <given-names>C. T.</given-names></name></person-group> (<year>2014</year>). <article-title>Phonetic influences on English and French listeners&#x00027; assimilation of Mandarin tones to native prosodic categories</article-title>. <source>Stud. Second Lang. Acquisit.</source> <volume>36</volume>, <fpage>195</fpage>&#x02013;<lpage>221</lpage>. <pub-id pub-id-type="doi">10.1017/S0272263114000047</pub-id></citation></ref>
<ref id="B71">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Swingley</surname> <given-names>D.</given-names></name> <name><surname>Aslin</surname> <given-names>R. N.</given-names></name></person-group> (<year>2000</year>). <article-title>Spoken word recognition and lexical representation in very young children</article-title>. <source>Cognition</source> <volume>76</volume>, <fpage>147</fpage>&#x02013;<lpage>166</lpage>. <pub-id pub-id-type="doi">10.1016/S0010-0277(00)00081-0</pub-id><pub-id pub-id-type="pmid">10856741</pub-id></citation></ref>
<ref id="B72">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Toro</surname> <given-names>J. M.</given-names></name> <name><surname>Nespor</surname> <given-names>M.</given-names></name> <name><surname>Mehler</surname> <given-names>J.</given-names></name> <name><surname>Bonatti</surname> <given-names>L. L.</given-names></name></person-group> (<year>2008</year>). <article-title>Finding words and rules in a speech stream: functional differences between vowels and consonants</article-title>. <source>Psychol. Sci.</source> <volume>19</volume>, <fpage>137</fpage>&#x02013;<lpage>144</lpage>. <pub-id pub-id-type="doi">10.1111/j.1467-9280.2008.02059.x</pub-id><pub-id pub-id-type="pmid">18271861</pub-id></citation></ref>
<ref id="B73">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tyler</surname> <given-names>M. D.</given-names></name> <name><surname>Best</surname> <given-names>C. T.</given-names></name> <name><surname>Faber</surname> <given-names>A.</given-names></name> <name><surname>Levitt</surname> <given-names>A. G.</given-names></name></person-group> (<year>2014</year>). <article-title>Perceptual assimilation and discrimination of non-native vowel contrasts</article-title>. <source>Phonetica</source> <volume>71</volume>, <fpage>4</fpage>&#x02013;<lpage>21</lpage>. <pub-id pub-id-type="doi">10.1159/000356237</pub-id><pub-id pub-id-type="pmid">24923313</pub-id></citation></ref>
<ref id="B74">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>van Ooijen</surname> <given-names>B.</given-names></name></person-group> (<year>1996</year>). <article-title>Vowel mutability and lexical selection in English: evidence from a word reconstruction task</article-title>. <source>Mem. Cognit.</source> <volume>24</volume>, <fpage>573</fpage>&#x02013;<lpage>583</lpage>. <pub-id pub-id-type="doi">10.3758/BF03201084</pub-id><pub-id pub-id-type="pmid">8870528</pub-id></citation></ref>
<ref id="B75">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wang</surname> <given-names>W. S.-Y.</given-names></name></person-group> (<year>1963</year>). <article-title>Mandarin phonology</article-title>. <source>Project Linguist. Analy.</source> <volume>6</volume>, <fpage>1</fpage>&#x02013;<lpage>6</lpage>.</citation></ref>
<ref id="B76">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Werker</surname> <given-names>J. F.</given-names></name> <name><surname>Tees</surname> <given-names>R. C.</given-names></name></person-group> (<year>1984a</year>). <article-title>Cross-language speech perception: evidence for perceptual reorganization during the first year of life</article-title>. <source>Infant Behav. Dev.</source> <volume>7</volume>, <fpage>49</fpage>&#x02013;<lpage>63</lpage>. <pub-id pub-id-type="doi">10.1016/S0163-6383(84)80022-3</pub-id></citation></ref>
<ref id="B77">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Werker</surname> <given-names>J. F.</given-names></name> <name><surname>Tees</surname> <given-names>R. C.</given-names></name></person-group> (<year>1984b</year>). <article-title>Phonemic and phonetic actors in adult cross-language speech perception</article-title>. <source>J. Acoust. Soc. Am.</source> <volume>75</volume>, <fpage>1866</fpage>&#x02013;<lpage>1878</lpage>. <pub-id pub-id-type="doi">10.1121/1.390988</pub-id></citation></ref>
<ref id="B78">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wiener</surname> <given-names>S.</given-names></name> <name><surname>Turnbull</surname> <given-names>R.</given-names></name></person-group> (<year>2016</year>). <article-title>Constraints of tones, vowels and consonants on lexical selection in Mandarin Chinese</article-title>. <source>Lang. Speech</source> <volume>59</volume>, <fpage>59</fpage>&#x02013;<lpage>82</lpage>. <pub-id pub-id-type="doi">10.1177/0023830915578000</pub-id><pub-id pub-id-type="pmid">27089806</pub-id></citation></ref>
<ref id="B79">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Yeung</surname> <given-names>H. H.</given-names></name> <name><surname>Chen</surname> <given-names>K. H.</given-names></name> <name><surname>Werker</surname> <given-names>J. F.</given-names></name></person-group> (<year>2013</year>). <article-title>When does native language input reorganize phonetic perception? The precocious case of lexical tone</article-title>. <source>J. Memory Lang.</source> <volume>68</volume>, <fpage>123</fpage>&#x02013;<lpage>139</lpage>. <pub-id pub-id-type="doi">10.1016/j.jml.2012.09.004</pub-id></citation></ref>
<ref id="B80">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Yip</surname> <given-names>M.</given-names></name></person-group> (<year>2002</year>). <source>Tone</source>. <publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>.</citation></ref>
</ref-list> 
</back>
</article> 