<?xml version="1.0" encoding="utf-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" article-type="research-article" dtd-version="2.3" xml:lang="EN">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Psychol.</journal-id>
<journal-title>Frontiers in Psychology</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Psychol.</abbrev-journal-title>
<issn pub-type="epub">1664-1078</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fpsyg.2022.877684</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Psychology</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>The effects of alphabetic literacy, linguistic-processing demand and tone type on the dichotic listening of lexical tones</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author"><name><surname>Shao</surname><given-names>Jing</given-names></name>
<xref rid="aff1" ref-type="aff"><sup>1</sup></xref>
<xref rid="aff2" ref-type="aff"><sup>2</sup></xref>
<uri xlink:href="https://loop.frontiersin.org/people/927314/overview"/>
</contrib>
<contrib contrib-type="author" corresp="yes"><name><surname>Zhang</surname><given-names>Caicai</given-names></name>
<xref rid="aff3" ref-type="aff"><sup>3</sup></xref>
<xref rid="c001" ref-type="corresp"><sup>&#x002A;</sup></xref>
<uri xlink:href="https://loop.frontiersin.org/people/20448/overview"/>
</contrib>
<contrib contrib-type="author"><name><surname>Zhang</surname><given-names>Gaoyuan</given-names></name>
<xref rid="aff4" ref-type="aff"><sup>4</sup></xref>
</contrib>
<contrib contrib-type="author"><name><surname>Zhang</surname><given-names>Yubin</given-names></name>
<xref rid="aff5" ref-type="aff"><sup>5</sup></xref>
<uri xlink:href="https://loop.frontiersin.org/people/1221148/overview"/>
</contrib>
<contrib contrib-type="author"><name><surname>Pattamadilok</surname><given-names>Chotiga</given-names></name>
<xref rid="aff6" ref-type="aff"><sup>6</sup></xref>
<uri xlink:href="https://loop.frontiersin.org/people/28351/overview"/>
</contrib>
</contrib-group>
<aff id="aff1"><sup>1</sup><institution>Department of English Language and Literature, Hong Kong Baptist University</institution>, <addr-line>Kowloon Tong</addr-line>, <country>Hong Kong SAR, China</country></aff>
<aff id="aff2"><sup>2</sup><institution>Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences</institution>, <addr-line>Shenzhen</addr-line>, <country>China</country></aff>
<aff id="aff3"><sup>3</sup><institution>Research Centre for Language, Cognition, and Neuroscience, Department of Chinese and Bilingual Studies, The Hong Kong Polytechnic University</institution>, <addr-line>Hung Hom</addr-line>, <country>Hong Kong SAR, China</country></aff>
<aff id="aff4"><sup>4</sup><institution>Department of Chinese Language and Literature, Peking University</institution>, <addr-line>Beijing</addr-line>, <country>China</country></aff>
<aff id="aff5"><sup>5</sup><institution>Department of Linguistics, University of Southern California</institution>, <addr-line>Los Angeles, CA</addr-line>, <country>United States</country></aff>
<aff id="aff6"><sup>6</sup><institution>Aix Marseille Univ, CNRS, LPL, Laboratoire Parole et Langage</institution>, <addr-line>Aix-en-Provence</addr-line>, <country>France</country></aff>
<author-notes>
<fn id="fn0001" fn-type="edited-by">
<p>Edited by: William Choi, The University of Hong Kong, Hong Kong SAR, China</p>
</fn>
<fn id="fn0002" fn-type="edited-by">
<p>Reviewed by: Yu-Fu Chien, Fudan University, China; Yiu-Kei Tsang, Hong Kong Baptist University, Hong Kong SAR, China; Wentao Gu, Nanjing Normal University, China</p>
</fn>
<corresp id="c001">&#x002A;Correspondence: Caicai Zhang, <email>caicai.zhang@polyu.edu.hk</email></corresp>
<fn id="fn0003" fn-type="other">
<p>This article was submitted to Language Sciences, a section of the journal Frontiers in Psychology</p>
</fn>
</author-notes>
<pub-date pub-type="epub">
<day>26</day>
<month>07</month>
<year>2022</year>
</pub-date>
<pub-date pub-type="collection">
<year>2022</year>
</pub-date>
<volume>13</volume>
<elocation-id>877684</elocation-id>
<history>
<date date-type="received">
<day>17</day>
<month>02</month>
<year>2022</year>
</date>
<date date-type="accepted">
<day>28</day>
<month>06</month>
<year>2022</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x00A9; 2022 Shao, Zhang, Zhang, Zhang and Pattamadilok.</copyright-statement>
<copyright-year>2022</copyright-year>
<copyright-holder>Shao, Zhang, Zhang, Zhang and Pattamadilok</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/">
<p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p>
</license>
</permissions>
<abstract>
<p>Brain lateralization of lexical tone processing remains a matter of debate. In this study we used a dichotic listening paradigm to examine the influences of the knowledge of <italic>Jyutping</italic> (a romanization writing system which provides explicit Cantonese tone markers), linguistic-processing demand and tone type on the ear preference pattern of native tone processing in Hong Kong Cantonese speakers. While participants with little knowledge of <italic>Jyutping</italic> showed a previously reported left-ear advantage (LEA), those with a good level of <italic>Jyutping</italic> expertise exhibited either a right-ear advantage or bilateral processing during lexical tone identification and contour tone discrimination, respectively. As for the effect of linguistic-processing demand, while an LEA was found in acoustic/phonetic perception situations, this advantage disappeared and was replaced by a bilateral pattern in conditions that involved a greater extent of linguistic processing, suggesting an increased involvement of the left hemisphere. Regarding the effect of tone type, both groups showed an LEA in level tone discrimination, but only the <italic>Jyutping</italic> group demonstrated a bilateral pattern in contour tone discrimination. Overall, knowledge of written codes of tones, greater degree of linguistic processing and contour tone processing seem to influence the brain lateralization of lexical tone processing in native listeners of Cantonese by increasing the recruitment of the left-hemisphere language network.</p>
</abstract>
<kwd-group>
<kwd>dichotic listening</kwd>
<kwd>alphabetic literacy</kwd>
<kwd>linguistic-processing demand</kwd>
<kwd>ear preference</kwd>
<kwd>lexical tone perception</kwd>
<kwd>Cantonese</kwd>
</kwd-group>
<contract-num rid="cn1">P0008738</contract-num>
<contract-num rid="cn2">ECS: 25603916</contract-num>
<contract-num rid="cn3">NSFC: 11504400</contract-num>
<contract-sponsor id="cn1">Departmental General Research Funds</contract-sponsor>
<contract-sponsor id="cn2">Research Grants Council of Hong Kong</contract-sponsor>
<contract-sponsor id="cn3">National Natural Science Foundation of China<named-content content-type="fundref-id">10.13039/501100001809</named-content>
</contract-sponsor>
<counts>
<fig-count count="5"/>
<table-count count="1"/>
<equation-count count="0"/>
<ref-count count="79"/>
<page-count count="16"/>
<word-count count="12820"/>
</counts>
</article-meta>
</front>
<body>
<sec id="sec1" sec-type="intro">
<title>Introduction</title>
<sec id="sec2">
<title>Ear preference in lexical tone perception revealed by dichotic listening studies</title>
<p>Over the past few decades, numerous studies have shown that the human brain is functionally specialized. A left-hemisphere (LH) specialization has been found for verbal material processing (<xref ref-type="bibr" rid="ref35">Kimura, 1967</xref>; <xref ref-type="bibr" rid="ref30">Hugdahl et al., 1999</xref>) and a right-hemisphere (RH) specialization for nonverbal material processing (<xref ref-type="bibr" rid="ref8">Boucher and Bryden, 1997</xref>; <xref ref-type="bibr" rid="ref30">Hugdahl et al., 1999</xref>). Dichotic listening, in which different auditory stimuli are presented simultaneously to left and right ears, is a neuropsychological technique for studying perceptual laterality. Findings have shown a right-ear advantage (REA) for the processing of spoken syllables and digits, indicating an LH dominance (<xref ref-type="bibr" rid="ref34">Kimura, 1961</xref>, <xref ref-type="bibr" rid="ref35">1967</xref>; <xref ref-type="bibr" rid="ref63">Studdert-Kennedy and Shankweiler, 1970</xref>; <xref ref-type="bibr" rid="ref30">Hugdahl et al., 1999</xref>), and a left-ear advantage (LEA) for music and pitch processing (<xref ref-type="bibr" rid="ref75">Wioland et al., 1999</xref>; <xref ref-type="bibr" rid="ref9">Brancucci et al., 2005</xref>; <xref ref-type="bibr" rid="ref29">Hoch and Tillmann, 2010</xref>), indicating an RH dominance.</p>
<p>Since the 19th century, the REA has been thought to be dominant in perception of segmental speech input such as consonants and vowels (<xref ref-type="bibr" rid="ref17">Cutting, 1974</xref>; <xref ref-type="bibr" rid="ref21">Dwyer et al., 1982</xref>; <xref ref-type="bibr" rid="ref12">Bryden and Murray, 1985</xref>). However, the ear preference and its underlying brain lateralization of suprasegmental elements like lexical tones are still an issue of debate. Two hypotheses have been put forward to explain the brain lateralization patterns in lexical tone processing, i.e., the functional hypothesis and the acoustic hypothesis. The functional hypothesis (<xref ref-type="bibr" rid="ref67">Van Lancker, 1980</xref>; <xref ref-type="bibr" rid="ref74">Whalen and Liberman, 1987</xref>; <xref ref-type="bibr" rid="ref25">Gandour Wong and Hutchins, 1998</xref>; <xref ref-type="bibr" rid="ref24">Gandour et al., 2000</xref>; <xref ref-type="bibr" rid="ref42">Liberman and Whalen, 2000</xref>) assumes that brain lateralization is dependent on the functional role of the auditory signal. This view predicts that speech stimuli are primarily processed in the LH, typically considered as the dominant hemisphere for language processing in right-handed individuals, whereas non-speech signals are processed primarily in the RH. On the other hand, the acoustic hypothesis (<xref ref-type="bibr" rid="ref77">Zatorre and Belin, 2001</xref>; <xref ref-type="bibr" rid="ref78">Zatorre et al., 2002</xref>; <xref ref-type="bibr" rid="ref53">Poeppel, 2003</xref>) claims that the acoustic structures of auditory inputs determine the brain lateralization: spectral processing, including pitch-related and suprasegmental information, is lateralized to the RH, whereas fast temporal processing, such as segmental information, induces more LH activation. Given the unique nature of lexical tones, the functional and acoustic hypotheses make diverging predictions. The functional hypothesis predicts an LH dominance in native speakers of tonal languages, based on their linguistic functions; however, the acoustic hypothesis predicts an RH dominance for processing lexical tones, based on their acoustic features.</p>
<p>While dichotic listening studies on native tone perception have generated empirical support for both hypotheses (<xref ref-type="bibr" rid="ref71">Wang et al., 2001</xref>; <xref ref-type="bibr" rid="ref44">Luo et al., 2006</xref>; <xref ref-type="bibr" rid="ref31">Jia et al., 2013</xref>), the actual lateralization patterns are more complex than those predicted by the two hypotheses, and appear to vary across languages. In summary, three distinct patterns of brain lateralization of lexical tones have been revealed by dichotic listening studies: (1) an REA in processing lexical tones by native Mandarin Chinese, Thai and Norwegian speakers (<xref ref-type="bibr" rid="ref68">Van Lancker and Fromkin, 1973</xref>, <xref ref-type="bibr" rid="ref69">1978</xref>; <xref ref-type="bibr" rid="ref46">Moen, 1993</xref>; <xref ref-type="bibr" rid="ref71">Wang et al., 2001</xref>); (2) bilateral processing by native Mandarin Chinese speakers (<xref ref-type="bibr" rid="ref3">Baudoin-Chial, 1986</xref>); (3) an LEA in the perception of lexical tones by native Hong Kong Cantonese speakers (<xref ref-type="bibr" rid="ref31">Jia et al., 2013</xref>), regardless of the tone type (level tones and contour tones), stimulus type (hums, real syllables and pseudosyllables) and task (discrimination task and identification task). Of particular note is the last pattern reported in Hong Kong Cantonese, which diverges from those observed in Thai, Norwegian and Mandarin Chinese speakers, and further deviates from the convergent finding of LH activation in tone processing in native tonal language listeners as revealed by a meta-analysis on neuroimaging studies (<xref ref-type="bibr" rid="ref40">Liang and Du, 2018</xref>). This discrepancy across studies indicates that additional factors may influence the hemispheric laterality of native tone processing other than the functional and acoustic explanations.</p>
<p>One possible factor is the lack of training in a native alphabetic script of spoken Cantonese in Hong Kong Cantonese speakers, as also argued by <xref ref-type="bibr" rid="ref31">Jia et al. (2013)</xref>. In contrast to Mandarin Chinese and Thai in which tones are marked as written labels with various degrees of precision in the respective alphabet (<italic>Pinyin</italic> or the Thai script), most Hong Kong Cantonese speakers learned logographic Chinese without exposure to a native alphabetic script of spoken Cantonese. Based on the well-established link between alphabetic literacy and phonological awareness, including tone awareness (<xref ref-type="bibr" rid="ref48">Morais et al., 1979</xref>; <xref ref-type="bibr" rid="ref16">Cheung et al., 2001</xref>; <xref ref-type="bibr" rid="ref60">Shu et al., 2008</xref>; <xref ref-type="bibr" rid="ref39">Li and Suk-Han Ho, 2011</xref>), lack of exposure to written codes of the native tones may account for the LEA observed in Hong Kong Cantonese listeners. Other factors that may explain the discrepancy are the differences in stimulus type (e.g., nonspeech and speech tones) and tone type (e.g., level tones and contour tones). Thus, the first and primary aim of this study is to examine whether knowledge of a native alphabetic script with codes for lexical tones would influence the ear preference pattern of lexical tone processing in Cantonese listeners. The second aim is to further investigate the influences of linguistic-processing demand and tone type on dichotic listening of lexical tones, and the potential impact of the knowledge of tone markers on these two factors.</p>
<sec id="sec3">
<title>The effect of alphabetic literacy</title>
<p>It has been well established that learning an alphabetic script boosts phonological awareness at the behavioral level (<xref ref-type="bibr" rid="ref48">Morais et al., 1979</xref>; <xref ref-type="bibr" rid="ref16">Cheung et al., 2001</xref>; <xref ref-type="bibr" rid="ref60">Shu et al., 2008</xref>; <xref ref-type="bibr" rid="ref82">Zhang Y. et al., 2021</xref>). Phonological awareness usually refers to the ability to analyze the spoken language into smaller units such as phonemes (<xref ref-type="bibr" rid="ref41">Liberman et al., 1974</xref>). Since lexical tone is a suprasegmental unit that distinguishes word meanings in tonal languages, tone awareness can be considered as a component of phonological awareness, which involves the ability to recognize and extract the lexical tone from a speech unit (<xref ref-type="bibr" rid="ref60">Shu et al., 2008</xref>; <xref ref-type="bibr" rid="ref39">Li and Suk-Han Ho, 2011</xref>). For example, <xref ref-type="bibr" rid="ref16">Cheung et al. (2001)</xref> examined the development of phonological awareness in Hong Kong and Guangzhou children who both spoke Cantonese. Guangzhou children usually had early experience with alphabetic Chinese reading (<italic>Pinyin</italic>) in addition to logographic Chinese reading, whereas Hong Kong children read only logographic Chinese. The results revealed that at the pre-reading stage, the Hong Kong and Guangzhou children showed similar performance on phonological awareness. However, after learning to read, the Guangzhou children outperformed their Hong Kong counterparts on several phonological awareness tasks, indicating that learning <italic>Pinyin</italic> boosts phonological awareness. With regard to tone awareness, <xref ref-type="bibr" rid="ref60">Shu et al. (2008)</xref> reported that phonological coding (<italic>Pinyin</italic>) instruction, which provides explicit markers for individual lexical tones, significantly improved this ability in Mandarin-speaking children: Tone awareness arose from chance level in preschoolers to over 74% accuracy in first graders after receiving <italic>Pinyin</italic> instruction. Although tones are marked with non-alphabetic symbols (e.g., diacritics in <italic>Pinyin</italic>), the development of tone awareness is presumably governed by the same principle that supports the development of phonemic awareness at the segmental level when one learns to read in an alphabetic system (<xref ref-type="bibr" rid="ref48">Morais et al., 1979</xref>; <xref ref-type="bibr" rid="ref55">Read et al., 1986</xref>; <xref ref-type="bibr" rid="ref47">Morais, 2021</xref>).</p>
<p>At the neural level, learning an alphabetic script has been reported to induce functional reorganization of the LH language network (<xref ref-type="bibr" rid="ref18">Dehaene et al., 2010</xref>; <xref ref-type="bibr" rid="ref11">Brennan et al., 2013</xref>). <xref ref-type="bibr" rid="ref11">Brennan et al. (2013)</xref> conducted a study that examined the influence of learning to read on brain activities associated with spoken language processing in English and Chinese speakers. The authors reported that, compared to Chinese speakers who had much less experience with an alphabetic writing system, English speakers showed developmental increases of brain activity in the LH phonological network, including the superior temporal gyrus, inferior parietal lobule and inferior frontal gyrus. These findings led the authors to conclude that learning to read in an alphabetic writing system reorganizes the phonological awareness network, and that such reorganization might lead to better phonological awareness skills.</p>
<p>In light of the aforementioned impacts of alphabetic literacy on boosting tone awareness and reorganizing the LH phonological network, we hypothesized that Cantonese listeners with knowledge of written codes that allow clear identification of the individual lexical tones present in the spoken language would demonstrate greater REA in native tone processing compared to those without such knowledge. Studying tone processing in Cantonese allowed us to further investigate the RH dominance of tone processing in Hong Kong Cantonese speakers, which has so far been reported only by <xref ref-type="bibr" rid="ref31">Jia et al. (2013)</xref>. Additionally, it also allowed us to avoid possible confounds related to age, education level or the maturation of the spoken language system, that may occur when the same question is addressed, for instance, by comparing performance in young native Mandarin Chinese speakers with different levels of <italic>Pinyin</italic> skills. Indeed, in contrast to the widespread <italic>Pinyin</italic> instruction that typically starts in the primary school in the mainland, the majority of Hong Kong Cantonese speakers learned logographic Chinese without being exposed to a native alphabetic script of spoken Cantonese. Although many Cantonese speakers in Hong Kong also learned English and its alphabetic script from childhood, tones are not coded in the English script, and it is highly unlikely that proficiency in the English script would boost tone awareness in Cantonese speakers (<xref ref-type="bibr" rid="ref5">Bialystok et al., 2005</xref>; <xref ref-type="bibr" rid="ref20">Dodd et al., 2008</xref>; <xref ref-type="bibr" rid="ref19">Deng et al., 2019</xref>). Furthermore, as an intonation language, English uses large (coarse-grained) pitch modulations to index intonation differences (statement/question) and lexical stress, whereas smaller (more refined) pitch modulations are used to differentiate lexical tones in Cantonese (<xref ref-type="bibr" rid="ref43">Liu et al., 2010</xref>), which would make any transfer of English phonological knowledge to tone awareness in Cantonese difficult. Despite the lack of an official Cantonese alphabetic script, several non-standard Cantonese romanization systems are concurrently in use in Hong Kong, primarily as online Chinese input methods. Among these systems, <italic>Jyutping</italic> is one of the few systems that provide a precise coding of lexical tones. <italic>Jyutping</italic> is a romanization system of spoken Cantonese devised by the Linguistic Society of Hong Kong in 1993. Tones are transcribed in <italic>Jyutping</italic> as numbers 1&#x2013;6 (e.g., &#x2018;&#x8B8A;&#x5316;&#x2019; bin3 faa3 &#x2018;change&#x2019;, with &#x2018;3&#x2019; indicating the third tone in Cantonese), which correspond to the six tones in Cantonese: T1 /55/ high level tone, T2 /25/ high rising tone, T3 /33/ mid-level tone, T4 /21/ low falling/extra low level tone, T5 /23/ low rising tone, and T6 /22/ low level tone (<xref ref-type="bibr" rid="ref4">Bauer and Benedict, 1997</xref>). Thus, <italic>Jyutping</italic> is particularly suitable for examining the ear preference pattern of lexical tone processing in Cantonese listeners.</p>
</sec>
<sec id="sec4">
<title>The effect of linguistic-processing demand</title>
<p>The second issue addressed in this study is the role of linguistic-processing demand on the modulation of ear preference patterns in dichotic listening of lexical tones. Acoustically, lexical tones result from a modulation of pitch contour. At the same time, it also serves as a distinctive feature for distinguishing word meanings in tonal languages, comparable to the role of phonemes. As mentioned above, these unique characteristics of lexical tone might have led to the mixed findings regarding its ear preference and brain lateralization: At the acoustic level, it may be mainly processed based on its acoustic features and leads to an ear preference pattern similar to that of prosody. At the phonological level, it may elicit an ear preference pattern resembling that of phonemes. Indeed, as revealed by <xref ref-type="bibr" rid="ref40">Liang and Du (2018)</xref> using a meta-analysis approach, lexical tones, like prosody, showed more extensive activations in the right than the left auditory cortex in both tonal and non-tonal language speakers, whereas the LH was recruited during lexical tone processing exclusively by native tonal language speakers, consistent with the activation pattern of phonemes. In other words, it is likely that lexical tones induce RH activation in both tonal and non-tonal language speakers due to their low-level acoustic feature, whereas tonal language speakers additionally engage the LH because lexical tones are further processed as a phonological unit.</p>
<p>In addition to the role of tonal vs. non-tonal language experience, previous studies have revealed that even within native tonal language speakers, different amounts of linguistic information contained in the stimuli or different degrees of linguistic processing elicited different ear preference patterns (<xref ref-type="bibr" rid="ref62">Shuai and Gong, 2014</xref>; <xref ref-type="bibr" rid="ref45">Mei et al., 2020</xref>). According to an EEG study using the dichotic listening paradigm, auditory processing of pitch variations elicited greater activation of the RH as a bottom-up effect; linguistic processing of lexical tones, on the other hand, evoked greater activation of the LH as a top-down effect (<xref ref-type="bibr" rid="ref61">Shuai, 2009</xref>; <xref ref-type="bibr" rid="ref62">Shuai and Gong, 2014</xref>). <xref ref-type="bibr" rid="ref45">Mei et al. (2020)</xref> reported that when the stimuli contained slow frequency modulation of tones only, such as in hums, an LEA was observed in native Mandarin listeners. However, when the stimuli contained linguistic cues such as phoneme and lexical information (e.g., in the condition where lexical tones were carried by a single vowel), the LEA was likely to disappear. Furthermore, bilateral processing was found when more phonological and lexical-semantic attributes were included, as in the consonant-vowel (CV), pseudo-word, and word conditions. These findings suggest that linguistic complexity of the stimuli and linguistic processing demand contribute to the brain specialization of lexical tones.</p>
<p>Based on the above existing observations, the second aim of this study is to further test the hypothesis that different degrees of linguistic processing result in different ear preference patterns, by employing three types of stimuli to vary the degree of linguistic processing&#x2014;non-speech stimuli, speech stimuli with low syllable variation and speech stimuli with high syllable variation (see descriptions in <italic>The Current Study</italic> below).</p>
</sec>
<sec id="sec5">
<title>The effect of tone type</title>
<p>The last aim of the current study is to revisit the influence of tone type on ear preference. The Cantonese tonal system can be classified as comprising three static level tones that primarily differ in pitch height (T1 &#x2013; high level tone, T3 &#x2013; mid level tone, and T6 &#x2013; low level tone) and three dynamic contour tones that primarily differ in the direction of pitch change (T2 &#x2013; high rising tone, T4 &#x2013; extra low level/low falling tone, and T5 &#x2013; low rising tone; <xref ref-type="bibr" rid="ref4">Bauer and Benedict, 1997</xref>). Whereas pitch height is the primary cue for distinguishing the three level tones with discernible pitch differences from the pitch onset, both pitch height and direction are involved in contour tone distinction and pitch cues in the later portion of the pitch curve might be more critical for contour tone perception (<xref ref-type="bibr" rid="ref33">Khouw and Ciocca, 2007</xref>). It has been found that native Cantonese listeners placed more weight on or were more sensitive to pitch height than pitch contour cues (<xref ref-type="bibr" rid="ref23">Gandour, 1983</xref>; <xref ref-type="bibr" rid="ref33">Khouw and Ciocca, 2007</xref>; <xref ref-type="bibr" rid="ref31">Jia et al., 2013</xref>; <xref ref-type="bibr" rid="ref81">Zhang C. et al., 2021</xref>). Accordingly, a study showed that the accuracy in the perception of three level tones was higher compared to the more complex contour tones (<xref ref-type="bibr" rid="ref31">Jia et al., 2013</xref>). Early processing of these two types of tones were also found to be different, in that they elicited different ERP components, namely, a prominent MMN in level tones but a P3a in contour tones in Cantonese listeners (<xref ref-type="bibr" rid="ref66">Tsang et al., 2011</xref>). Altogether, these observations point out processing differences between level and contour tones.</p>
<p>According to the temporal integration hypothesis (<xref ref-type="bibr" rid="ref53">Poeppel, 2003</xref>; <xref ref-type="bibr" rid="ref6">Boemio et al., 2005</xref>; <xref ref-type="bibr" rid="ref57">Sanders and Poeppel, 2007</xref>; <xref ref-type="bibr" rid="ref65">Teng et al., 2016</xref>; <xref ref-type="bibr" rid="ref22">Flinker et al., 2019</xref>), the left and right auditory cortices have differential sensitivity towards acoustic information over varied time-scales: whereas the left auditory cortex (AC) preferentially processes information from short temporal integration windows (25&#x2013;50&#x2009;ms), the right AC preferentially processes information from long temporal integration windows (200&#x2013;300&#x2009;ms). In light of this hypothesis, it is likely that the fast-changing and dynamic contour tones would require the extraction of pitch information over short temporal windows (<xref ref-type="bibr" rid="ref36">Krishnan and Gandour, 2009</xref>), increasing the LH participation, in contrast to the slowly-changing and static level tones. In addition, as mentioned above, the dynamic contour tones that are perceptually more challenging might require deeper processing, which may also increase the LH participation.</p>
<p>However, mixed findings have been reported regarding the ear preference pattern of contour vs. level tone processing (<xref ref-type="bibr" rid="ref28">Ho, 2010</xref>; <xref ref-type="bibr" rid="ref31">Jia et al., 2013</xref>). Although <xref ref-type="bibr" rid="ref28">Ho (2010)</xref> found that pitch height and contour changes induced different hemispheric advantages, both level and contour tones were reported to elicit a greater RH advantage in <xref ref-type="bibr" rid="ref31">Jia et al. (2013)</xref>. We aimed to revisit the influence of tone type on ear preference and further examine whether the effect of tone type interacted with that of <italic>Jyutping</italic> expertise on ear preference in the current study.</p>
</sec>
</sec>
<sec id="sec6">
<title>The current study</title>
<p>In the present study, we examined the impacts of three factors &#x2013; <italic>Jyutping</italic> expertise (<italic>Jyutping</italic> vs. non-<italic>Jyutping</italic> group), linguistic-processing demand (nonspeech vs. low syllable variation vs. high syllable variation) and tone type (level vs. contour tones), as well as their interactions on ear preference of lexical tone processing in native Cantonese speakers using the dichotic listening paradigm. As with the previous study on Cantonese (<xref ref-type="bibr" rid="ref31">Jia et al., 2013</xref>), we employed an identification task and a discrimination task to examine the dichotic listening of lexical tones. In the current study, the effect of <italic>Jyutping</italic> expertise was investigated <italic>via</italic> a comparison of two matched groups of native Cantonese speakers without <italic>Jyutping</italic> expertise (henceforth, non-<italic>Jyutping</italic> participants) and with <italic>Jyutping</italic> expertise (henceforth, <italic>Jyutping</italic> participants), in order to elucidate the influence of knowledge of Cantonese tonal codes on the ear preference of lexical tone processing. The impact of linguistic-processing demand was examined by three stimulus types&#x2014;nonspeech tones, speech materials with low syllable variation and speech materials with high syllable variation. The nonspeech tone condition, which only contained pitch information extracted from the speech materials, was included to induce primarily acoustic processing of lexical tones (<xref ref-type="bibr" rid="ref68">Van Lancker and Fromkin, 1973</xref>; <xref ref-type="bibr" rid="ref62">Shuai and Gong, 2014</xref>; <xref ref-type="bibr" rid="ref45">Mei et al., 2020</xref>). The low and high variation conditions both used meaningful Cantonese words and would engage more linguistic (e.g., phonological and lexical) processing relative to the nonspeech condition. The critical difference between these two conditions concerned the carrying syllables in a dichotic pair, which remained constant in the low variation condition (e.g., /ji55/ &#x2018;doctor&#x2019; &#x2013; /ji22/ &#x2018;second&#x2019;), but varied in the high variation condition (e.g., /ji55/ &#x2018;doctor&#x2019; &#x2013; /f&#x0250;n22/ &#x2018;part&#x2019;). Note that the two stimuli in a dichotic pair involved a meaning change in both low and high variation conditions, and thus the two conditions may be deemed to be largely matched in this regard, constraining the primary difference between them to the changing of carrying syllables. Previous studies have indicated that different degrees of syllable variability may tap into different levels of linguistic processing (<xref ref-type="bibr" rid="ref37">Lee et al., 2008</xref>; <xref ref-type="bibr" rid="ref60">Shu et al., 2008</xref>; <xref ref-type="bibr" rid="ref58">Shao et al., 2019</xref>). In the low variation condition, tone perception could be carried out primarily by comparing the acoustic forms of the stimuli without having to segregate the tone from the segmental units (<xref ref-type="bibr" rid="ref13">Burton et al., 2000</xref>), thus requiring relatively less phonological processing. In contrast, in the high variation condition, initial separation between segmental and suprasegmental units seems necessary before conducting the comparison of tone categories (<xref ref-type="bibr" rid="ref13">Burton et al., 2000</xref>), demanding greater efforts in phonological processing. Moreover, this meta-phonological ability has been reported to develop with language experience and reading ability (<xref ref-type="bibr" rid="ref60">Shu et al., 2008</xref>; <xref ref-type="bibr" rid="ref82">Zhang Y. et al., 2021</xref>). Therefore, we employed these three types of stimuli to further test the hypothesis that different degrees of linguistic processing, especially phonological segmentation demand, influence the patterns of ear preference. Lastly, we investigated the effects of tone type by comparing the processing of three level tones (T1, T3 and T6) versus three contour tones (T2, T4 and T5) in Cantonese.</p>
<p>With regard to the effect of <italic>Jyutping</italic> expertise, based on previous observations that alphabetic literacy boosts phonological awareness (including tone awareness) and induces the activation of the LH phonological network (<xref ref-type="bibr" rid="ref16">Cheung et al., 2001</xref>; <xref ref-type="bibr" rid="ref60">Shu et al., 2008</xref>; <xref ref-type="bibr" rid="ref18">Dehaene et al., 2010</xref>; <xref ref-type="bibr" rid="ref11">Brennan et al., 2013</xref>; <xref ref-type="bibr" rid="ref82">Zhang Y. et al., 2021</xref>), we predicted that participants with <italic>Jyutping</italic> knowledge may show greater REA (LH dominance) in the dichotic listening of lexical tones compared to their non-<italic>Jyutping</italic> peers. Regarding the effects of linguistic-processing demand, we predicted that situations that place a higher demand on linguistic processing, especially phonological segmentation (e.g., the high variation condition), would lead to greater engagement of the LH, yielding bilateral processing or even REA. On the contrary, an LEA was expected in the conditions that required minimal linguistic processing (e.g., the nonspeech condition). In terms of the effect of tone type, level tones are expected to elicit an LEA, in contrast to contour tones that may exhibit more bilateral processing or even REA. Finally, we also explored whether there would be an interaction between <italic>Jyutping</italic> expertise and the effects of linguistic-processing demand and tone type. It is possible that participants with knowledge of <italic>Jyutping</italic> are likely to show an REA especially in the processing conditions that require more LH engagement.</p>
</sec>
</sec>
<sec id="sec7" sec-type="materials|methods">
<title>Materials and methods</title>
<sec id="sec8">
<title>Participants</title>
<p>Eighteen non-<italic>Jyutping</italic> participants (8&#x2009;M, 10F) and 16 <italic>Jyutping</italic> participants (9&#x2009;M, 7F) were recruited based on their self-report of <italic>Jyutping</italic> knowledge which was confirmed by a <italic>Jyutping</italic> transcription test (see below). All the participants were native speakers of Hong Kong Cantonese. The participants were pre-screened based on the criteria of being right-handed as assessed by the Edinburgh Handedness Inventory (<xref ref-type="bibr" rid="ref50">Oldfield, 1971</xref>), having no hearing impairment, and having no formal musical training. Participants with linguistics background were deliberately excluded. The two groups were largely matched in age (<italic>Jyutping</italic>: mean&#x2009;=&#x2009;22.2, age range&#x2009;=&#x2009;20&#x2013;24; non-<italic>Jyutping</italic>: mean&#x2009;=&#x2009;22.3, age range&#x2009;=&#x2009;18&#x2013;28) and education level. The <italic>Jyutping</italic> proficiency (or lack of it) of the two groups of participants was confirmed using a timed <italic>Jyutping</italic> transcription test, which contained 20 disyllabic words that covered the full Cantonese phonetic inventory. The 20 words were presented in Chinese characters on a piece of paper, and the participants were instructed to write down the <italic>Jyutping</italic> transcriptions of these words (including tones) as fast as possible. The participants without <italic>Jyutping</italic> knowledge were instructed to skip the trials or guess the transcriptions. Only the participants who scored above 50% accuracy in tone transcription were included in the <italic>Jyutping</italic> group. Accuracy of tone transcription in the <italic>Jyutping</italic> group was significantly higher than the non-<italic>Jyutping</italic> group (<italic>t</italic>(33)&#x2009;=&#x2009;&#x2212;9.539, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001; <italic>Jyutping</italic>: <italic>M</italic>&#x2009;=&#x2009;72.94%, <italic>SD</italic>&#x2009;=&#x2009;13.8%; non-<italic>Jyutping</italic>: <italic>M</italic>&#x2009;=&#x2009;29.13%, <italic>SD</italic>&#x2009;=&#x2009;15.1%).</p>
<p>The experimental procedure was approved by the Human Subjects Ethics Sub-committee of The Hong Kong Polytechnic University (Application number: HSEARS20190502004). Informed written consent was obtained from the participants in compliance with the experiment protocols.</p>
</sec>
<sec id="sec9">
<title>Stimuli</title>
<p>There were three stimulus conditions: nonspeech tone, low variation and high variation conditions. The stimuli used in the low variation condition were six words contrasting six Cantonese tones on the syllable /ji/. The stimuli used in the high variation condition were 18 words contrasting six Cantonese tones on the syllables /f&#x0250;n/, /j&#x0250;u/, and /w&#x0250;i/ (see <xref rid="tab1" ref-type="table">Table 1</xref>). In addition to these critical stimuli that were used in both identification and discrimination task, six tones carried by the base syllable /&#x014B;a/ were employed as mask items in the discrimination task (see further details in Procedure). These base syllables were selected because they can yield meaningful morphemes in combination with every tone, which enables us to have a full tonal coverage while controlling for the base syllable variability (5 syllables &#x00D7; 6 tones). All the syllables are free or bound morphemes. While syllables carrying T1, T3, and T6 were grouped into the level tone condition, syllables carrying T2, T4, and T5 were grouped into the contour tone condition.</p>
<table-wrap position="float" id="tab1">
<label>Table 1</label>
<caption>
<p>The five sets of syllables used in the experiment.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th/>
<th align="left" valign="top">T1 high level /55/</th>
<th align="left" valign="top">T2 high rising /25/</th>
<th align="left" valign="top">T3 mid-level /33/</th>
<th align="left" valign="top">T4 low falling /21/</th>
<th align="left" valign="top">T5 low rising /23/</th>
<th align="left" valign="top">T6 low level /22/</th>
</tr>
</thead>
<tbody>
<tr>
<td align="char" valign="top" char=".">/ji/</td>
<td align="left" valign="top">&#x91AB; &#x201C;doctor&#x201D;</td>
<td align="left" valign="top">&#x6905;<break/>&#x201C;chair&#x201D;</td>
<td align="left" valign="top">&#x610F;<break/>&#x201C;meaning&#x201D;</td>
<td align="left" valign="top">&#x5152;<break/>&#x201C;son&#x201D;</td>
<td align="left" valign="top">&#x8033;<break/>&#x201C;ear&#x201D;</td>
<td align="left" valign="top">&#x4E8C;<break/>&#x201C;two&#x201D;</td>
</tr>
<tr>
<td align="char" valign="top" char=".">/f&#x0250;n/</td>
<td align="left" valign="top">&#x5A5A; &#x201C;marriage&#x201D;</td>
<td align="left" valign="top">&#x7C89;<break/>&#x201C;pink&#x201D;</td>
<td align="left" valign="top">&#x8A13;<break/>&#x201C;train&#x201D;</td>
<td align="left" valign="top">&#x711A;<break/>&#x201C;burn&#x201D;</td>
<td align="left" valign="top">&#x596E;<break/>&#x201C;strive&#x201D;</td>
<td align="left" valign="top">&#x4EFD;<break/>&#x201C;part&#x201D;</td>
</tr>
<tr>
<td align="char" valign="top" char=".">/w&#x0250;i/</td>
<td align="left" valign="top">&#x5A01; &#x201C;power&#x201D;</td>
<td align="left" valign="top">&#x59D4; &#x201C;council&#x201D;</td>
<td align="left" valign="top">&#x9935;<break/>&#x201C;feed&#x201D;</td>
<td align="left" valign="top">&#x570D; &#x201C;surround&#x201D;</td>
<td align="left" valign="top">&#x5049;<break/>&#x201C;grand&#x201D;</td>
<td align="left" valign="top">&#x80C3;<break/>&#x201C;stomach&#x201D;</td>
</tr>
<tr>
<td align="char" valign="top" char=".">/j&#x0250;u/</td>
<td align="left" valign="top">&#x4F11;<break/>&#x201C;rest&#x201D;</td>
<td align="left" valign="top">&#x9EDD;<break/>&#x201C;dark&#x201D;</td>
<td align="left" valign="top">&#x5E7C;<break/>&#x201C;young&#x201D;</td>
<td align="left" valign="top">&#x6CB9;<break/>&#x201C;oil&#x201D;</td>
<td align="left" valign="top">&#x53CB;<break/>&#x201C;friend&#x201D;</td>
<td align="left" valign="top">&#x53F3;<break/>&#x201C;right&#x201D;</td>
</tr>
<tr>
<td align="char" valign="top" char=".">/&#x014B;a/</td>
<td align="left" valign="top">&#x9D09;<break/>&#x201C;crow&#x201D;</td>
<td align="left" valign="top">&#x555E;<break/>&#x201C;mute&#x201D;</td>
<td align="left" valign="top">&#x4E9E;<break/>&#x201C;Asia&#x201D;</td>
<td align="left" valign="top">&#x7259;<break/>&#x201C;teeth&#x201D;</td>
<td align="left" valign="top">&#x96C5;<break/>&#x201C;proper&#x201D;</td>
<td align="left" valign="top">&#x8A1D; &#x201C;astonished&#x201D;</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>One female native Cantonese speaker was recorded reading aloud these words in a carrier sentence, &#x5462;&#x500B;&#x5B57;&#x4FC2; /li55 ko33 tsi22 h&#x0250;i22/ (&#x2018;This word is&#x2019;) in a clear and deliberate manner. Each sentence was recorded six times. Then, the most clearly produced token was selected and the word was segmented out of the carrier sentence. All selected words were normalized such that they had the same acoustic intensity (60&#x2009;dB) and duration (620&#x2009;ms, which corresponded to the mean duration of all selected words; Praat: <xref ref-type="bibr" rid="ref7">Boersma and Weenink, 2014</xref>). The first author checked the naturalness of the stimuli after normalization.</p>
<p>The nonspeech tone stimuli were nonspeech analogues of the stimuli that were used in the low variation condition. A 620-ms pure tone sound was first generated using Praat, and then a total of 12 F0 contours of the syllables /ji/ and /&#x014B;a/ were extracted and superimposed on the pure tone sound, generating 12 pure tone stimuli. The mean acoustic intensity of the pure tone stimuli was set to 75&#x2009;dB, which was 15&#x2009;dB louder than the speech stimuli. This adjustment allowed us to match the subjective intensity of the two types of stimuli (<xref ref-type="bibr" rid="ref83">Zhang et al., 2017</xref>; <xref ref-type="bibr" rid="ref59">Shao and Zhang, 2020</xref>).</p>
</sec>
<sec id="sec10">
<title>Procedure</title>
<p>Both identification and discrimination tasks were employed in a total of six conditions as defined by the stimulus type and tone type (3&#x2009;&#x00D7;&#x2009;2). The design of the identification task was adopted from <xref ref-type="bibr" rid="ref31">Jia et al. (2013)</xref>. In each trial, there was a dichotic pair presented to the two ears simultaneously. Level tones and contour tones were presented in separate blocks. For both tone types, there were three same-tone pairs and six different-tone pairs. For the level tones, the same-tone pairs included T1-T1, T3-T3, and T6-T6, while the different-tone pairs included T1-T3, T1-T6, T3-T6, T3-T1, T6-T1, and T6-T3. For the contour tones, the same-tone pairs included T2-T2, T4-T4, and T5-T5, while the different-tone pairs included T2-T5, T5-T2, T4-T5, T5-T4, T4-T2, and T2-T4. In the low variation condition, the two items within each trial had the same base syllable (e.g., /ji55/&#x2212;/ji33/). In the high variation condition, the base syllables differed (e.g., /f&#x0250;n55/&#x2212;/j&#x0250;u33/). Note that the three syllables formed three pairs of syllables (f&#x0250;n/&#x2212;/j&#x0250;u/, /f&#x0250;n/&#x2212;/w&#x0250;i/ and /j&#x0250;u/&#x2212;/w&#x0250;i/) with equal probabilities to occur. In the nonspeech tone condition, all the stimuli were pure tones with the same F0 trajectories as the stimuli used in the low variation condition. For all stimulus types, the task was to identify the most clearly heard tone (presented to either the left or right ear) by pressing the button on the keyboard as soon as possible, i.e., 1&#x2013;6 (corresponding to T1 to T6) within 5&#x2009;s. The identification task contained six blocks in total, which corresponded to the combination of the two tone types and the three stimulus types. Within each block, the three same pairs were repeated six times and the six different pairs were also repeated six times, generating a total of 54 trials. The presentation of stimulus pairs was randomized. In this task, we did not balance the number of same and different pairs, because identical tones were presented to both ears in the same pairs, which did not probe dichotic listening and was not our primary interest. Each block lasted about 3&#x2013;3.5&#x2009;min and the whole identification task took about 20&#x2009;min. Participants were asked to take a five-minute break every three blocks.</p>
<p>The design of the discrimination task followed that of <xref ref-type="bibr" rid="ref10">Brancucci et al. (2008)</xref> and <xref ref-type="bibr" rid="ref31">Jia et al. (2013)</xref>. Within each trial, two dichotic pairs were consecutively presented. The participants were instructed to direct their attention to a designated ear (i.e., the testing ear). The first pair composed of a target and a mask. The second pair composed of a probe and a mask. The target and probe were always presented in the testing ear; the mask was always presented in the ear to be ignored. The task was to judge whether the target and the probe were, or not, pronounced with the same tone as soon as possible within 3&#x2009;s, by pressing the button on the keyboard (&#x201C;left arrow&#x201D; if same, and &#x201C;right arrow&#x201D; if different).</p>
<p>In the nonspeech tone condition, the masks were pure tones with the same F0 trajectories as the syllable /&#x014B;a/, and the targets and probes were pure tones carrying the same F0 trajectories as the syllable /ji/. In the low and high variation conditions, the masks were words with the syllable /&#x014B;a/. The carrying syllables for the targets and probes for the low variation condition was /ji/. For the high variation condition, they were /f&#x0250;n/, /j&#x0250;u/, and /w&#x0250;i/. These syllables were grouped into three pairs (/f&#x0250;n/&#x2212;/j&#x0250;u/, /f&#x0250;n/&#x2212;/w&#x0250;i/ and /j&#x0250;u/&#x2212;/w&#x0250;i/), which had equal chance to occur. The discrimination task contained 12 blocks in total, which corresponded to the combination of the two tone types, the three stimulus types and the two testing ears. Within each block, the sequences of stimulus presentation were randomized. The same pairs were repeated six times (3&#x2009;&#x00D7;&#x2009;6), and different pairs were repeated three times (6&#x2009;&#x00D7;&#x2009;3), creating equal numbers of same and different pairs in each block. Each block lasted about 2&#x2013;2.5&#x2009;min and the whole discrimination task took about 25&#x2009;min. Participants were asked to take a five-minute break every three blocks.</p>
<p>The tasks were programmed with E-prime 1.0 (Psychology Software Tools, Pittsburgh, PA). All the tasks were conducted in a soundproof booth in the Speech and Language Sciences Lab at the Hong Kong Polytechnic University. The stimuli were presented <italic>via</italic> headphones to the participants at a comfortable listening level. The volume level was kept constant across all the tasks within each participant. No participants reported any difficulty hearing the stimuli. The discrimination task required the participants to direct their attention to a specific ear in each block, whereas the identification task did not. To avoid the effect of direction of attention transferring from the discrimination task to the identification task, all participants completed the identification task before the discrimination task. Before each task, a practice session was provided to familiarize the participants with the procedure and to ensure that they fully understood the instruction. No feedback was given during the practice sessions. The presentation order of blocks within each task was counterbalanced across participants.</p>
</sec>
<sec id="sec11">
<title>Data analysis</title>
<p>We measured both accuracy and reaction time (RT) to examine the hemispheric lateralization pattern, following previous studies (<xref ref-type="bibr" rid="ref31">Jia et al., 2013</xref>; <xref ref-type="bibr" rid="ref56">Reilly et al., 2015</xref>). Linear mixed-effects (LME) analyses were performed on the R platform using the <italic>lme4</italic> (<xref ref-type="bibr" rid="ref2">Bates et al., 2014</xref>), <italic>lmerTest</italic> (<xref ref-type="bibr" rid="ref001">Kuznetsova et al., 2017</xref>), and <italic>emmeans</italic> packages (<xref ref-type="bibr" rid="ref002">Lenth and Lenth, 2018</xref>). The <italic>anova</italic> function of the R package was used to obtain the <italic>p</italic> values of the main effects and the interactions in the models. The <italic>emmeans</italic> package was used to conduct pairwise comparisons with Tukey&#x2019;s correction.</p>
<p>Identification accuracy was computed as the relative portion of the correct responses in each ear. The maximal model was first fitted using the following variables and their interactions: <italic>group</italic> (<italic>Jyutping</italic> vs. non-<italic>Jyutping</italic>), <italic>stimulus type</italic> (nonspeech tone vs. low syllable variation vs. high syllable variation), <italic>tone type</italic> (level tone vs. contour tone) and <italic>ear</italic> (left ear vs. right ear). The random intercepts of subjects, along with the random slope of the interaction between categories, stimulus type and ear per subject were treated as the random factors.<xref rid="fn0004" ref-type="fn"><sup>1</sup></xref> To reach a simpler model, the random intercepts and slopes were removed one by one using the likelihood ratio test (LRT) in R. The same method was used to remove the least contributing predictors in terms of fixed factors.</p>
<p>Regarding the discrimination accuracy, generalized mixed-effects models were fitted on the responses to each trial (correct response was coded as &#x201C;1&#x201D; and incorrect response was coded as &#x201C;0&#x201D;). The fixed effects were <italic>group</italic>, <italic>stimulus type, tone type, ear</italic> and <italic>their interactions</italic>. The random intercepts of subjects and stimulus, along with the random slope of the interaction between categories, stimulus type and ear per subject and random slope of group per stimulus were treated as the random factors. The same model comparison procedure as in the identification task was applied.</p>
<p>In the RT analyses, identification RT was measured from the offset of the stimuli to the time that a response was made. RT in the discrimination task was measured from the offset of the second dichotic pair to the time that a response was made. In both tasks, trials with null and incorrect responses were excluded from the analysis and RT data of the remaining trials were log-transformed. Linear mixed-effects models were fitted with the fixed effects including <italic>group</italic>, <italic>stimulus type, tone type, ear</italic> and <italic>their interactions</italic>. The random intercepts of subjects and stimulus, along with the random slope of the interaction between categories, stimulus type and ear per subject and random slope of group per stimulus were treated as the random factors. The same model comparison procedure described was applied. Given that these analyses yielded a number of main effects and interactions, for the sake of simplicity, the results are presented in different sub-sections that correspond to the research questions posed in the Introduction. Unless stated otherwise, only the significant effects are reported. The full results of the models are reported in the supplemental materials.</p>
</sec>
</sec>
<sec id="sec12" sec-type="results">
<title>Results</title>
<sec id="sec13">
<title>Does jyutping expertise have an influence on ear preference of lexical tone processing?</title>
<p>This research question was addressed by examining the presence or absence of the interaction between <italic>ear</italic> and <italic>group</italic>.<xref rid="fn0005" ref-type="fn"><sup>2</sup></xref>,<xref rid="fn0006" ref-type="fn"><sup>3</sup></xref> <xref rid="fig1" ref-type="fig">Figures 1A</xref>&#x2013;<xref rid="fig1" ref-type="fig">D</xref> displays the identification RT, identification accuracy, discrimination RT and discrimination accuracy in terms of <italic>ear</italic> and <italic>group</italic>. Among the analyses on the RTs and accuracy scores of the identification and discrimination tasks, the interaction was significant only in the RT analysis of the identification task [<italic>F</italic>(2, 5,897)&#x2009;=&#x2009;4.39, <italic>p&#x2009;=&#x2009;0</italic>.01]. Pairwise comparisons examining the <italic>ear</italic> effect in each group of participants showed that, in the non-<italic>Jyutping</italic> group, RTs obtained on the stimuli presented in the left ear were significantly shorter than RTs obtained on stimuli presented in the right ear (Estimate&#x2009;=&#x2009;&#x2212;0.0160, Std. Error&#x2009;=&#x2009;0.007, <italic>t</italic>&#x2009;=&#x2009;&#x2212;2.198, <italic>p</italic>&#x2009;=&#x2009;0.028), suggesting an LEA. The <italic>Jyutping</italic> group showed the opposite pattern: The RTs obtained on stimuli presented in the right ear were significantly shorter (Estimate&#x2009;=&#x2009;0.0203, Std. Error&#x2009;=&#x2009;0.007, <italic>t</italic>&#x2009;=&#x2009;2.710, <italic>p</italic>&#x2009;=&#x2009;0.006), suggesting an REA in the <italic>Jyutping</italic> group (<xref rid="fig1" ref-type="fig">Figure 1A</xref>).</p>
<fig position="float" id="fig1">
<label>Figure 1</label>
<caption>
<p>Plots displaying the interaction of ear and group. <bold>(A)</bold> Identification RT, <bold>(B)</bold> identification accuracy, <bold>(C)</bold> discrimination RT, and <bold>(D)</bold> discrimination accuracy of the stimuli presented in the left and right ear in the <italic>Jyutping</italic> and non-<italic>Jyutping</italic> participants. The error bars indicate 95% confidence interval.</p>
</caption>
<graphic xlink:href="fpsyg-13-877684-g001.tif"/>
</fig>
</sec>
<sec id="sec14">
<title>Do linguistic-processing demand and tone type have an influence on ear preference of lexical tone processing?</title>
<p>These two research questions were addressed by examining the presence or absence of the interaction between <italic>ear</italic> and linguistic-processing demand (<italic>stimulus type</italic>) or <italic>tone type</italic>. Since there were not many significant effects in relation to these two questions, they are reported together in this section. <xref rid="fig2" ref-type="fig">Figures 2A</xref>&#x2013;<xref rid="fig2" ref-type="fig">D</xref>, <xref rid="fig3" ref-type="fig">3A</xref>&#x2013;<xref rid="fig3" ref-type="fig">D</xref> displays the identification RT, identification accuracy, discrimination RT and discrimination accuracy in terms of the interaction of <italic>stimulus type</italic> and <italic>ear</italic>, and the interaction of <italic>tone type</italic> and <italic>ear</italic>, respectively.</p>
<fig position="float" id="fig2">
<label>Figure 2</label>
<caption>
<p>Plots displaying the interaction of stimulus type and ear. <bold>(A)</bold> Identification RT, <bold>(B)</bold> identification accuracy, <bold>(C)</bold> discrimination RT, and <bold>(D)</bold> discrimination accuracy obtained on the stimuli presented in the left and right ear in the nonspeech tone (NS), low variation (LV) and high variation (HV) conditions. The error bars indicate 95% confidence interval.</p>
</caption>
<graphic xlink:href="fpsyg-13-877684-g002.tif"/>
</fig>
<fig position="float" id="fig3">
<label>Figure 3</label>
<caption>
<p>Plots displaying the interaction of tone type and ear. <bold>(A)</bold> Identification RT, <bold>(B)</bold> identification accuracy, <bold>(C)</bold> discrimination RT, and <bold>(D)</bold> discrimination accuracy obtained on the level and contour tones presented in the left and right ear. The error bars indicate 95% confidence interval.</p>
</caption>
<graphic xlink:href="fpsyg-13-877684-g003.tif"/>
</fig>
<p>Regarding the influence of <italic>stimulus type</italic>, analysis on discrimination accuracy revealed a significant two-way interaction between <italic>stimulus type</italic> and <italic>ear</italic> [<italic>&#x03C7;</italic><sup>2</sup>(4)&#x2009;=&#x2009;92.91, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001; <xref rid="fig2" ref-type="fig">Figure 2D</xref>]. Pairwise comparisons showed that accuracy scores on stimuli presented in the left ear were higher than scores on stimuli presented in the right ear in the nonspeech tone condition (Estimate&#x2009;=&#x2009;0.325, Std. Error&#x2009;=&#x2009;0.1, <italic>z</italic>&#x2009;=&#x2009;3.251, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.01) and low variation condition (Estimate&#x2009;=&#x2009;0.428, Std. Error&#x2009;=&#x2009;0.139, <italic>z</italic>&#x2009;=&#x2009;3.080, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001), but not in the high variation condition (Estimate&#x2009;=&#x2009;0.063, Std. Error&#x2009;=&#x2009;0.07, <italic>z</italic>&#x2009;=&#x2009;0.724, <italic>p</italic>&#x2009;=&#x2009;0.47), indicating that a significant LEA was observed in the nonspeech and low variation conditions, whereas the high variation condition showed a bilateral pattern. No two-way interaction between <italic>ear</italic> and <italic>stimulus type</italic> was found in the other analyses (<italic>p</italic>s&#x2009;&#x003E;&#x2009;0.05, see <xref rid="fig2" ref-type="fig">Figure 2</xref>).</p>
<p>Regarding the influence of <italic>tone type</italic>, no two-way interaction between <italic>ear</italic> and <italic>tone type</italic> was found (<italic>p</italic>s&#x2009;&#x003E;&#x2009;0.05, see <xref rid="fig3" ref-type="fig">Figure 3</xref>).</p>
</sec>
<sec id="sec15">
<title>Does the interaction between Jyutping expertise, linguistic-processing demand and tone type influence ear preference on lexical tone processing?</title>
<p>This section examined the existence of complex interaction patterns between <italic>ear</italic>, <italic>group</italic> and linguistic-processing demand (<italic>stimulus type</italic>) and <italic>tone type</italic>.</p>
<p>We observed that the impact of <italic>Jyutping</italic> expertise on ear preference further interacted with <italic>tone type</italic> on discrimination RT, as shown in a significant three-way interaction among <italic>group</italic>, <italic>ear</italic> and <italic>tone type</italic> [<italic>F</italic> (7, 41.146)&#x2009;=&#x2009;5.56, <italic>p</italic>&#x2009;=&#x2009;0.001]. The interaction was plotted in <xref rid="fig4" ref-type="fig">Figure 4</xref>. Linear mixed-effect models were fitted within each <italic>tone type</italic> to explore this three-way interaction. For the slowly-changing level tones, there was only a main effect of <italic>ear</italic> [<italic>&#x03C7;</italic><sup>2</sup>(1)&#x2009;=&#x2009;13.101, <italic>p</italic>&#x2009;=&#x2009;0.002], where the mean RT on the stimuli presented in the left ear was shorter than that on the stimuli presented to the right ear. For the fast-changing contour tones, there was a main effect of <italic>ear</italic> [<italic>&#x03C7;</italic><sup>2</sup>(1)&#x2009;=&#x2009;38.089, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001] and a significant two-way interaction between <italic>ear</italic> and <italic>group</italic> [<italic>&#x03C7;</italic><sup>2</sup>(1)&#x2009;=&#x2009;12.153, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001]. Pairwise comparisons showed that the non-<italic>Jyutping</italic> group exhibited significantly shorter RT in the left ear than the right ear (Estimate&#x2009;=&#x2009;&#x2212;0.0246, Std. Error&#x2009;=&#x2009;0.003, <italic>t</italic>&#x2009;=&#x2009;&#x2212;6.896, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001). This LEA was no longer significant in the <italic>Jyutping</italic> group (Estimate&#x2009;=&#x2009;&#x2212;0.006, Std. Error&#x2009;=&#x2009;0.003, <italic>t</italic>&#x2009;=&#x2009;&#x2212;1.754, <italic>p</italic>&#x2009;&#x003E;&#x2009;0.05).</p>
<fig position="float" id="fig4">
<label>Figure 4</label>
<caption>
<p>A plot displaying the three-way interaction among <italic>group</italic>, <italic>ear</italic> and <italic>tone type</italic> in the discrimination RT. Discrimination RT were shown for the level and contour tones presented in the left and the right ear in the <italic>Jyutping</italic> and non-<italic>Jyutping</italic> participants. The error bars indicate 95% confidence interval.</p>
</caption>
<graphic xlink:href="fpsyg-13-877684-g004.tif"/>
</fig>
<p>No three-way interaction involving <italic>ear</italic>, <italic>group</italic> and the other two factors was observed in the other analyses (<italic>p</italic>s&#x2009;&#x003E;&#x2009;0.05).</p>
</sec>
<sec id="sec16">
<title>Task difference in the ear preference pattern and general impacts of Jyutping expertise, linguistic-processing demand and tone type</title>
<p>This last section presents the remaining significant effects. Although not a central aim of this study, we observed different ear preference patterns in the identification and discrimination tasks. In addition, we found main and interaction effects that were unrelated to the ear preference issue, but they allowed us to verify whether lexical tone processing performance is influenced by <italic>Jyutping</italic> expertise, linguistic-processing demand (<italic>stimulus type</italic>) and <italic>tone type</italic> as we initially assumed.</p>
<p>Regarding the task difference in the ear preference pattern, the analyses on discrimination accuracy (<xref rid="fig1" ref-type="fig">Figure 1D</xref>) revealed a main effect of <italic>ear</italic> [<italic>&#x03C7;</italic><sup>2</sup>(1)&#x2009;=&#x2009;8.762, <italic>p</italic>&#x2009;=&#x2009;0.003]. Discrimination accuracy on stimuli presented in the left ear was significantly higher than that on stimuli presented in the right ear, suggesting an overall LEA in this task. In contrast, there was no significant main effect of <italic>ear</italic> in the identification accuracy (<italic>p</italic>s&#x2009;&#x003E;&#x2009;0.05, <xref rid="fig1" ref-type="fig">Figures 1A</xref>,<xref rid="fig1" ref-type="fig">B</xref>).</p>
<p>As for the remaining effects, there was a significant main effect of <italic>group</italic> [<italic>&#x03C7;</italic><sup>2</sup>(1)&#x2009;=&#x2009;6.865, <italic>p</italic>&#x2009;=&#x2009;0.008] in the discrimination accuracy (<xref rid="fig5" ref-type="fig">Figure 5E</xref>), with the <italic>Jyutping</italic> group showing higher discrimination accuracy. As for the effects of <italic>tone type</italic>, there was a significant main effect of <italic>tone type</italic> [<italic>F</italic> (1, 406)&#x2009;=&#x2009;5.497, <italic>p</italic>&#x2009;=&#x2009;0.01] in the identification accuracy (<xref rid="fig5" ref-type="fig">Figure 5A</xref>), with level tones showing higher accuracy scores than contour tones. We also observed a significant main effect of <italic>tone type</italic> [<italic>F</italic> (1, 14.9)&#x2009;=&#x2009;5.987, <italic>p</italic>&#x2009;=&#x2009;0.02] in the identification RT (<xref rid="fig5" ref-type="fig">Figure 5B</xref>), where level tones elicited shorter RT than contour tones. These effects concerning the tone type are consistent with what was previously reported in the literature (<xref ref-type="bibr" rid="ref31">Jia et al., 2013</xref>). Finally, with regard to the effects of <italic>stimulus type</italic>, we observed a significant two-way interaction between <italic>stimulus type</italic> and [<italic>F</italic> (1, 35)&#x2009;=&#x2009;31.46, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001] in the identification RT. Despite this significant interaction, pairwise comparisons with Bonferroni correction found that the group difference was not significant in any stimulus type (<italic>p</italic>s&#x2009;&#x003E;&#x2009;0.05). As illustrated in <xref rid="fig5" ref-type="fig">Figure 5C</xref>, in both groups, the RT elicited in the high variation condition was significantly longer than that in the low variation and nonspeech tone conditions (<italic>p</italic>s&#x2009;&#x003C;&#x2009;0.001), whereas the difference between the low variation and nonspeech tone conditions was not significant (<italic>p</italic>&#x2009;&#x003E;&#x2009;0.05). In addition, there was a main effect of <italic>stimulus type</italic> [<italic>F</italic> (2, 32)&#x2009;=&#x2009;50.88, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001] in the discrimination RT. As shown in <xref rid="fig5" ref-type="fig">Figure 5D</xref>, RT elicited in the low variation condition was the shortest, followed by the nonspeech tone condition and then the high variation condition (<italic>p</italic>s&#x2009;&#x003C;&#x2009;0.01).</p>
<fig position="float" id="fig5">
<label>Figure 5</label>
<caption>
<p>Plots displaying the general effects of <italic>Jyutping</italic> expertise, processing demand and tone type without interaction with <italic>ear</italic>. <bold>(A)</bold> Identification accuracy and <bold>(B)</bold> Identification RT in the contour and level tone conditions; <bold>(C)</bold> Identification RT in the <italic>Jyutping</italic> and non-<italic>Jyutping</italic> participants in the nonspeech tone (NS), low variation (LV) and high variation (HV) conditions; <bold>(D)</bold> Discrimination RT in the nonspeech tone (NS), low variation (LV) and high variation (HV) conditions; <bold>(E)</bold> Discrimination accuracy in <italic>Jyutping</italic> and non-<italic>Jyutping</italic> participants. The error bars indicate 95% confidence interval.</p>
</caption>
<graphic xlink:href="fpsyg-13-877684-g005.tif"/>
</fig>
</sec>
<sec id="sec17">
<title>Summary</title>
<p>Regardless of ear preference, the global performance was in accordance with our expectations: <italic>Jyutping</italic> expertise was associated with higher accuracy in the tone discrimination task; processing speech tones in a high variation context was more challenging than processing speech tones in a low variation context or processing non-speech tones, lengthening the RT in both identification and discrimination tasks; identifying level tones elicited higher accuracy and shorter RT than identifying contour tones.</p>
<p>Regarding the effect of <italic>Jyutping</italic> expertise on ear preference, the non-<italic>Jyutping</italic> group showed an LEA whereas the <italic>Jyutping</italic> group showed an REA in the identification RT, implying different lateralization patterns between the two groups (<xref rid="fig1" ref-type="fig">Figure 1A</xref>). Moreover, the discrimination RT results showed an LEA in both groups when they processed the level tones. When they processed the contour tones, an LEA was observed in the non-<italic>Jyutping</italic> group but disappeared in the <italic>Jyutping</italic> group (<xref rid="fig4" ref-type="fig">Figure 4</xref>). Nevertheless, both groups showed an LEA in the discrimination accuracy scores (<xref rid="fig1" ref-type="fig">Figure 1D</xref>).</p>
<p>The analysis of the effect of linguistic-processing demand on ear preference showed no ear preference in the identification task on the accuracy data. In the discrimination task, the LEA was observed on the accuracy scores, although it was restricted to the nonspeech tone and low variation conditions, while the high variation condition showed a bilateral pattern (<xref rid="fig3" ref-type="fig">Figure 3D</xref>).</p>
</sec>
</sec>
<sec id="sec18" sec-type="discussions">
<title>Discussion</title>
<p>Ear preference and the underling brain lateralization for lexical tone processing remain an issue of debate. As mentioned in the introduction, in addition to functional and acoustic explanations, the complex patterns of ear preference could be driven by the experience with written codes of lexical tones and the demand on linguistic processing. Additionally, acoustic differences between level and contour tones may also have an impact on brain literalization. To examine these questions, we used a dichotic listening paradigm to investigate the effects of <italic>Jyutping</italic> expertise (<italic>Jyutping</italic> vs. non-<italic>Jyutping</italic> group), linguistic-processing demand (nonspeech vs. low syllable variation vs. high syllable variation), and tone type (level vs. contour tones) on the ear preference pattern in lexical tone processing in Hong Kong Cantonese speakers. In the text below we first discussed the results regarding the effects of these three factors as well as their interactions, followed by a general discussion in the end.</p>
<sec id="sec19">
<title>The effect of Jyutping expertise on ear preference in lexical tone processing</title>
<p>We found that <italic>Jyutping</italic> expertise contributes to some extent to ear preference of lexical tone processing. In the discrimination task, there was a shift from the LEA in the non-<italic>Jyutping</italic> group to a bilateral pattern in <italic>Jyutping</italic> group. However, this shift was observed only in the RT data and during the discrimination of contour tones, which was hypothesized to preferentially rely on the LH auditory cortex based on the temporal integration window hypothesis (<xref ref-type="bibr" rid="ref53">Poeppel, 2003</xref>; <xref ref-type="bibr" rid="ref6">Boemio et al., 2005</xref>; <xref ref-type="bibr" rid="ref57">Sanders and Poeppel, 2007</xref>; <xref ref-type="bibr" rid="ref65">Teng et al., 2016</xref>; <xref ref-type="bibr" rid="ref22">Flinker et al., 2019</xref>), but not in level tone discrimination (<xref rid="fig4" ref-type="fig">Figure 4</xref>; see the text below for further discussion of the interaction of <italic>Jyutping</italic> expertise and tone type). In the identification task, a more pronounced REA, which suggests a stronger engagement of the LH in lexical tone processing, clearly emerged in the <italic>Jyutping</italic> group. This pattern was observed in the RT but not accuracy, in that the <italic>Jyutping</italic> group showed shorter RTs to stimuli presented in the right ear than the left ear, whereas non-<italic>Jyutping</italic> participants showed shorter RTs in response to stimuli presented in the left ear (<xref rid="fig1" ref-type="fig">Figure 1A</xref>). Interestingly, in the identification task, the REA observed in the <italic>Jyutping</italic> group was generalized to all stimulus types and tone types, implying that individuals with <italic>Jyutping</italic> knowledge might have recruited the LH more systematically during lexical tone processing in the identification task.</p>
</sec>
<sec id="sec20">
<title>The effect of linguistic-processing demand on ear preference in lexical tone processing</title>
<p>A piece of evidence for the influence of linguistic-processing demand on ear advantage is the observation that the LEA found in discrimination accuracy was restricted to the nonspeech tone and low variation conditions, whereas no ear advantage was found in the high variation condition (<xref rid="fig3" ref-type="fig">Figure 3D</xref>). In both nonspeech tone and low variation conditions, tone discrimination could be performed based on acoustic processing. Pitch or tone processing in such contexts may mainly recruit the RH, which is thought to predominantly process pitch information (<xref ref-type="bibr" rid="ref77">Zatorre and Belin, 2001</xref>; <xref ref-type="bibr" rid="ref78">Zatorre et al., 2002</xref>; <xref ref-type="bibr" rid="ref53">Poeppel, 2003</xref>). However, in the high variation condition, tones were carried by different syllables. This increase in syllable variability would make the comparison of tones at the purely acoustic/phonetic level difficult. As a result, more effort in linguistic analysis of the speech signal was required, including the separation of tonal information from the segmental elements before performing a tonal comparison. Indeed, high syllable variability led to a significant increase of RTs compared to the nonspeech and low variation conditions (<xref rid="fig5" ref-type="fig">Figures 5C</xref>,<xref rid="fig5" ref-type="fig">D</xref>). The segmentation process that puts more demand on linguistic processing might entail an increase of activity in the LH spoken language network (<xref ref-type="bibr" rid="ref13">Burton et al., 2000</xref>). The greater involvement of the LH presumably resulted in bilateral processing of lexical tones in the high variability condition found in the current study, in contrast to the LEA observed in the situations where no segmentation was required. However, we might have to be cautious when interpreting the patterns obtained in the high syllable variation condition. Since segmental variation is present in this condition but not in others, it might have had some contribution to the ear preference pattern in this condition. Future studies that tease apart the influence of segmental variation can shed more light on the effect of linguistic-processing demand on the brain lateralization of lexical tone processing.</p>
</sec>
<sec id="sec21">
<title>The impact of the interaction between Jyutping expertise and tone type on ear preference in lexical tone processing</title>
<p>Consistent with the extant literature (<xref ref-type="bibr" rid="ref31">Jia et al., 2013</xref>), we found that level tones were identified more accurately than contour tones (<xref rid="fig5" ref-type="fig">Figures 5A</xref>,<xref rid="fig5" ref-type="fig">B</xref>). As mentioned in the introduction, lexical tones are characterized by multiple acoustic features and vary along more than one dimension (e.g., pitch height and contour; <xref ref-type="bibr" rid="ref23">Gandour, 1983</xref>; <xref ref-type="bibr" rid="ref14">Chandrasekaran et al., 2007a</xref>). Whereas pitch height is the primary cue for distinguishing the three level tones with discernible pitch differences from the beginning of the F0 curve, both pitch height and direction are involved in contour tone distinction and F0 cues in the later portion of the F0 curve might be more critical for contour tone perception (<xref ref-type="bibr" rid="ref33">Khouw and Ciocca, 2007</xref>). These differences may explain why the participants showed lower accuracy and longer RT when processing Cantonese contour tones compared to level tones.</p>
<p>We also found a complex interaction effect between <italic>Jyutping</italic> expertise and tone type on the ear preference pattern in the discrimination RT (<xref rid="fig4" ref-type="fig">Figure 4</xref>). For the level tones, the discrimination RT exhibited an LEA for both groups of participants, whereas for the contour tones, the LEA remained in the non-<italic>Jyutping</italic> group but disappeared in the <italic>Jyutping</italic> group. In other words, only the <italic>Jyutping</italic> group exhibited bilateral processing in the discrimination RT of contour tones. We hypothesized in the introduction that the fast-changing and dynamic contour tones would require the extraction of pitch information over short temporal windows. Moreover, contour tone perception does not only rely on pitch height at the onset, but also on pitch changes in later portions of the pitch curve. These two factors may lead to deeper processing of contour tones and therefore more LH processing, compared to the slowly-changing and static level tones (<xref ref-type="bibr" rid="ref53">Poeppel, 2003</xref>; <xref ref-type="bibr" rid="ref6">Boemio et al., 2005</xref>; <xref ref-type="bibr" rid="ref57">Sanders and Poeppel, 2007</xref>; <xref ref-type="bibr" rid="ref65">Teng et al., 2016</xref>; <xref ref-type="bibr" rid="ref22">Flinker et al., 2019</xref>). In line with the discussion here, the group difference we found in the discrimination RT of contour tone processing may suggest that the effect of <italic>Jyutping</italic> expertise was more prominent in the processing of tonal features that preferentially rely on the LH auditory cortex.</p>
</sec>
<sec id="sec22">
<title>General discussion</title>
<p>As mentioned in the introduction, previous dichotic listening studies on the brain specialization of lexical tones have found conflicting results, reporting three distinct patterns: (1) an REA in processing lexical tones by native Mandarin Chinese, Thai and Norwegian speakers (<xref ref-type="bibr" rid="ref68">Van Lancker and Fromkin, 1973</xref>, <xref ref-type="bibr" rid="ref69">1978</xref>; <xref ref-type="bibr" rid="ref46">Moen, 1993</xref>; <xref ref-type="bibr" rid="ref71">Wang et al., 2001</xref>); (2) bilateral processing by native Mandarin Chinese speakers (<xref ref-type="bibr" rid="ref3">Baudoin-Chial, 1986</xref>); (3) an LEA in the perception of lexical tones by native Hong Kong Cantonese speakers (<xref ref-type="bibr" rid="ref31">Jia et al., 2013</xref>). The last pattern in Hong Kong Cantonese not only deviates from those observed in other tonal language speakers, but also from an EEG study on Cantonese that revealed left hemispheric lateralization of lexical pitch and acoustic pitch processing, as indexed by the mismatch negativity (MMN; <xref ref-type="bibr" rid="ref26">Gu et al., 2013</xref>). It also differs from another MMN study which suggests an absence of brain specialization in the processing of Cantonese lexical tones (<xref ref-type="bibr" rid="ref32">Jia et al., 2015</xref>). These discrepancies warrant more empirical studies.</p>
<p>In order to explain the aforementioned complex results of ear preference in native tone processing, we postulated that a native alphabetic script with codes for lexical tones might play a role. The current study is a first attempt to empirically test this hypothesis and provided some crucial evidence in this direction. Even though the Cantonese speakers in the <italic>Jyutping</italic> group in the current study learned <italic>Jyutping</italic> after childhood, they exhibited either greater REA or a bilateral pattern in the identification task and contour tone discrimination, respectively. The finding suggests that even late acquisition of codes of lexical tones can shape the ear preference, and presumably, underlying brain lateralization, to some extent. Our observation was indeed consistent with other studies reporting that alphabetic literacy enhanced the LH participation in speech processing even in individuals who became literate in adulthood (<xref ref-type="bibr" rid="ref18">Dehaene et al., 2010</xref>).</p>
<p>It is worth noting that the contribution of <italic>Jyutping</italic> knowledge to ear preference reported here was found on the processing speed but not on response accuracy. This result may be attributable to the fact that RT is a more sensitive measure in capturing the processing advantage of <italic>Jyutping</italic> participants in lexical tone processing ability. Another explanation is related to the relatively late learning of <italic>Jyutping</italic> in Cantonese speakers, unlike in Mandarin Chinese speakers who learn <italic>Pinyin</italic> at the beginning of primary school or even in kindergartens. It is possible that the late acquisition of tonal codes might have a weaker impact on the hemispheric lateralization of native tone processing compared to early acquisition when children&#x2019;s phonological system is still developing. Lastly, the fact that <italic>Jyutping</italic> uses abstract numbers (1&#x2013;6) to represent Cantonese lexical tones may also have played some role in the relatively weak impact of <italic>Jyutping</italic>. In contrast, <italic>Pinyin</italic> employs diacritics that contains explicit visual&#x2013;spatial representations of the pitch information, which has been reported to facilitate Mandarin lexical tone learning (<xref ref-type="bibr" rid="ref49">Morett and Chang, 2015</xref>). These explanations are not mutually exclusive.</p>
<p>Note that even though dichotic listening paradigm has been long used to investigate brain lateralization in auditory processing, several factors should be taken into consideration in the experimental design to ensure the reliability and validity (<xref ref-type="bibr" rid="ref70">Voyer, 1998</xref>; <xref ref-type="bibr" rid="ref72">Westerhausen, 2019</xref>; <xref ref-type="bibr" rid="ref73">Westerhausen and Samuelsen, 2020</xref>). These factors include stimulus characteristics, stimulus-presentation features, response collection and instruction, participants variables and generalizability (<xref ref-type="bibr" rid="ref72">Westerhausen, 2019</xref>). In terms of stimulus characteristics, the selection of stimulus materials (e.g., numeric words, non-numeric words, and non-word syllables) should be based on the consideration of processing stages that are of interest (<xref ref-type="bibr" rid="ref72">Westerhausen, 2019</xref>). In our study, the aim is to investigate the influence of linguistic processing on ear preference. We used nonspeech stimuli and real words which have been widely used to investigate brain specialization and are suitable to investigate our hypothesis. In terms of stimulus-presentation features, the number of 90 to 120 trials seems to be ideal for the reliability and the intensity level of 70 to 80&#x2009;dB is most commonly used (<xref ref-type="bibr" rid="ref72">Westerhausen, 2019</xref>). In the current study, the sound level is within the suggested range. In order control the experimental length and avoid fatigue, 54 trials were presented in each block in identification task and 72 trials in each condition in discrimination task. Future studies should increase trial numbers to achieve the reliability. Regarding response collection and instruction, it was found that both verbal and manual responses seem suitable to collect accuracy data (<xref ref-type="bibr" rid="ref72">Westerhausen, 2019</xref>). The current study required the participants to respond by pressing keys on the keyboard. Regarding the participants, one most important factor is hearing ability, especially the absence of significantly difference in hearing acuity between the two ears should be ensured in the aging population (<xref ref-type="bibr" rid="ref72">Westerhausen, 2019</xref>). The participants in our study are all young college students and none of them reported hearing problem. However, future studies need to measure the hearing threshold carefully to rule out the possible influence of hearing acuity.</p>
<p>In conclusion, our findings further expanded the understanding of the brain lateralization of lexical tones and addressed some discrepancies in the literature. They suggest that, even among native tonal language speakers, the hemispheric lateralization pattern is not a fixed process but could be influenced by listeners&#x2019; phonological skills that are induced or boosted by alphabetic literacy, and by the processing demand inherent to different degrees of linguistic processing as well as acoustic features of lexical tones. The current study also left open several questions to be addressed in future studies. First, it remains to be investigated whether the <italic>Jyutping</italic> participants resorted to the abstract labels of lexical tones while they performed the tasks (e.g., 1&#x2013;6 in the <italic>Jyutping</italic> transcription), that is the observed impact of Jyutping knowledge would reflect the surface connection between the written and spoken codes of the lexical tone, or whether tone awareness boosted by <italic>Jyutping</italic> skills has a profound influence by progressively restructuring and fine-tuning the phonological representations of tones in Cantonese speakers (<xref ref-type="bibr" rid="ref52">Perre et al., 2009</xref>; <xref ref-type="bibr" rid="ref51">Pattamadilok et al., 2010</xref>; <xref ref-type="bibr" rid="ref11">Brennan et al., 2013</xref>). In the latter case, group differences may be found in tasks that probe phonological representations, such as categorical perception. Although both mechanisms might have contributed to promoting the REA and bilateral pattern observed here, their relative role can be further investigated. Second, it is unknown how much experience with an alphabetic script with tonal codes is necessary to alter the hemispheric laterality of lexical tone processing. In other words, future studies should examine when a shift from the RH dominance to bilateral processing or LH dominance takes place in native speakers when they learn the codes of tones. A training study conducted on Cantonese-speaking adults where their hemispheric laterality will be measured before and after learning tonal codes could provide an approach to address this issue. In relation to this point, a training study like this can also provide more evidence for the causal effect of learning tonal codes on the hemispheric lateralization of native tone processing, because the current study, which is correlational in nature, cannot exclude pre-existing differences between <italic>Jyutping</italic> and non-<italic>Jyutping</italic> participants. In line with this interindividual variation issue, as all the participants (in the <italic>Jyutping</italic> as well as non-<italic>Jyutping</italic> group) are college students in Hong Kong, they are bi-literate in both logographic Chinese and English. It begs the question of whether the participants&#x2019; alphabetic English knowledge has any influence on their lexical tone processing and ear preference patterns. However, since the focus of this study is on lexical tone processing and its ear preference pattern, it is highly unlikely that knowledge of the English alphabetic script would affect lexical tone perception. Indeed, previous research suggested that the influence of knowledge of the English alphabetic script on phonological awareness skills in native spoken language processing is limited (<xref ref-type="bibr" rid="ref11">Brennan et al., 2013</xref>; <xref ref-type="bibr" rid="ref82">Zhang Y. et al., 2021</xref>). This argument is also partially corroborated by a large amount of cross-linguistic studies that demonstrated the challenge faced by native English speakers when processing or learning lexical tones (<xref ref-type="bibr" rid="ref23">Gandour, 1983</xref>; <xref ref-type="bibr" rid="ref15">Chandrasekaran et al., 2007b</xref>; <xref ref-type="bibr" rid="ref76">Wong et al., 2007</xref>; <xref ref-type="bibr" rid="ref54">Qin and Jongman, 2016</xref>). A similar issue could be raised regarding the level of proficiency in Chinese. In the present study, we did not measure the participants&#x2019; level of Chinese proficiency, which was expected to be relatively high since all participants were native speakers of Chinese. Although unlikely, we cannot objectively rule out the possibility that the <italic>Jyutping</italic> group might somehow have a higher level of Chinese proficiency than the control group, and this potential difference could contribute to their ear preference patterns. Overall, while the hypotheses regarding linguistic-processing demand and tone type are well informed by the literature, a caveat is that these hypothesized operations may not be what actually happened in the listeners&#x2019; brain. There may be other sources of individual variance in the processing strategies or mechanisms that future studies should look into. In the present dataset, the analyses were conducted on a relatively small sample size (<italic>N</italic>&#x2009;=&#x2009;16 in the <italic>Jyutping</italic> group; <italic>N</italic>&#x2009;=&#x2009;18 in the non-<italic>Jyutping</italic> group). The COVID-19 pandemic has created a difficult environment for recruiting a large sample of participants, especially those well matched on demographic characteristics. Future studies with a larger sample size should try to replicate the current results and obtain more clear-cut ear preference patterns.</p>
</sec>
</sec>
<sec id="sec23" sec-type="data-availability">
<title>Data availability statement</title>
<p>The datasets presented in this study can be found in online repositories. The names of the repository/repositories and accession number(s) can be found at: https://osf.io/g79sm/.</p>
</sec>
<sec id="sec24">
<title>Ethics statement</title>
<p>The studies involving human participants were reviewed and approved by The Human Subjects Ethics Sub-committee of The Hong Kong Polytechnic University. The patients/participants provided their written informed consent to participate in this study.</p>
</sec>
<sec id="sec25">
<title>Author contributions</title>
<p>JS and CZ contributed to conception of the study. JS designed the study and performed the statistical analysis. JS, GZ, and YZ contributed to the data collection. JS wrote the first draft of the manuscript. CZ wrote sections of the manuscript. CZ and CP revised the manuscript. All authors contributed to the article and approved the submitted version.</p>
</sec>
<sec id="sec26" sec-type="funding-information">
<title>Funding</title>
<p>This work was supported by grants from the Departmental General Research Funds (P0008738; <ext-link xlink:href="https://www.polyu.edu.hk/cbs/web/en/" ext-link-type="uri">https://www.polyu.edu.hk/cbs/web/en/</ext-link>), the Research Grants Council of Hong Kong (ECS: 25603916; <ext-link xlink:href="https://www.ugc.edu.hk/eng/rgc/" ext-link-type="uri">https://www.ugc.edu.hk/eng/rgc/</ext-link>), the National Natural Science Foundation of China (NSFC: 11504400; <ext-link xlink:href="http://www.nsfc.gov.cn/" ext-link-type="uri">http://www.nsfc.gov.cn/</ext-link>), the Departmental Reward Scheme for Research Publications in Indexed Journals (<ext-link xlink:href="https://www.polyu.edu.hk/cbs/web/en/" ext-link-type="uri">https://www.polyu.edu.hk/cbs/web/en/</ext-link>) and the Hong Kong Polytechnic Univerity Project of Strategic Importance Scheme (<ext-link xlink:href="https://www.polyu.edu.hk/en/rio/about-rio/committees/areas-of-excellence-committee/" ext-link-type="uri">https://www.polyu.edu.hk/en/rio/about-rio/committees/areas-of-excellence-committee/</ext-link>) to CZ. The funders had no role in study design, data collection and analysis, decision to publish, or preparation of the manuscript.</p>
</sec>
<sec id="conf1" sec-type="COI-statement">
<title>Conflict of interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest. The reviewer Y-KT declared a shared affiliation, with no collaboration, with the authors to the handling editor at the time of the review.</p>
</sec>
<sec id="sec100" sec-type="disclaimer">
<title>Publisher&#x2019;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
</body>
<back>
<sec id="sec28" sec-type="supplementary-material">
<title>Supplementary material</title>
<p>The Supplementary Material for this article can be found online at: <ext-link xlink:href="https://www.frontiersin.org/articles/10.3389/fpsyg.2022.877684/full#supplementary-material" ext-link-type="uri">https://www.frontiersin.org/articles/10.3389/fpsyg.2022.877684/full#supplementary-material</ext-link></p>
<supplementary-material xlink:href="Data_Sheet_1.docx" id="SM1" mimetype="application/vnd.openxmlformats-officedocument.wordprocessingml.document" xmlns:xlink="http://www.w3.org/1999/xlink"/>
</sec>
<ref-list>
<title>References</title>
<ref id="ref2"><citation citation-type="other"><person-group person-group-type="author"><name><surname>Bates</surname> <given-names>D.</given-names></name> <name><surname>M&#x00E4;chler</surname> <given-names>M.</given-names></name> <name><surname>Bolker</surname> <given-names>B.</given-names></name> <name><surname>Walker</surname> <given-names>S.</given-names></name></person-group> (<year>2014</year>). <article-title>Fitting linear mixed-effects models using lme4</article-title>. <source>arXiv preprint arXiv:1406.5823</source>.</citation></ref>
<ref id="ref3"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Baudoin-Chial</surname> <given-names>S.</given-names></name></person-group> (<year>1986</year>). <article-title>Hemispheric lateralization of modern standard Chinese tone processing</article-title>. <source>J. Neurolinguistics</source> <volume>2</volume>, <fpage>189</fpage>&#x2013;<lpage>199</lpage>. doi: <pub-id pub-id-type="doi">10.1016/S0911-6044(86)80012-4</pub-id></citation></ref>
<ref id="ref4"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Bauer</surname> <given-names>R.</given-names></name> <name><surname>Benedict</surname> <given-names>P. K.</given-names></name></person-group> (<year>1997</year>). <source>Modern Cantonese Phonology</source>. <publisher-loc>Berlin</publisher-loc>: <publisher-name>Mouton de Gruyter</publisher-name>.</citation></ref>
<ref id="ref5"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bialystok</surname> <given-names>E.</given-names></name> <name><surname>Luk</surname> <given-names>G.</given-names></name> <name><surname>Kwan</surname> <given-names>E.</given-names></name></person-group> (<year>2005</year>). <article-title>Bilingualism, biliteracy, and learning to read: interactions among languages and writing systems</article-title>. <source>Sci. Stud. Read.</source> <volume>9</volume>, <fpage>43</fpage>&#x2013;<lpage>61</lpage>. doi: <pub-id pub-id-type="doi">10.1207/s1532799xssr0901_4</pub-id></citation></ref>
<ref id="ref6"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Boemio</surname> <given-names>A.</given-names></name> <name><surname>Fromm</surname> <given-names>S.</given-names></name> <name><surname>Braun</surname> <given-names>A.</given-names></name> <name><surname>Poeppel</surname> <given-names>D.</given-names></name></person-group> (<year>2005</year>). <article-title>Hierarchical and asymmetric temporal sensitivity in human auditory cortices</article-title>. <source>Nat. Neurosci.</source> <volume>8</volume>, <fpage>389</fpage>&#x2013;<lpage>395</lpage>. doi: <pub-id pub-id-type="doi">10.1038/nn1409</pub-id>, PMID: <pub-id pub-id-type="pmid">15723061</pub-id></citation></ref>
<ref id="ref7"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Boersma</surname> <given-names>P.</given-names></name> <name><surname>Weenink</surname> <given-names>D.</given-names></name></person-group> (<year>2014</year>). <source>Praat: Doing Phonetics by Computer.</source> <publisher-loc>Amsterdam</publisher-loc>: <publisher-name>Institute of Phonetic Sciences, University of Amsterdam</publisher-name>.</citation></ref>
<ref id="ref8"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Boucher</surname> <given-names>R.</given-names></name> <name><surname>Bryden</surname> <given-names>M. P.</given-names></name></person-group> (<year>1997</year>). <article-title>Laterality effects in the processing of melody and timbre</article-title>. <source>Neuropsychologia</source> <volume>35</volume>, <fpage>1467</fpage>&#x2013;<lpage>1473</lpage>. doi: <pub-id pub-id-type="doi">10.1016/S0028-3932(97)00066-3</pub-id>, PMID: <pub-id pub-id-type="pmid">9352524</pub-id></citation></ref>
<ref id="ref9"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Brancucci</surname> <given-names>A.</given-names></name> <name><surname>Babiloni</surname> <given-names>C.</given-names></name> <name><surname>Rossini</surname> <given-names>P. M.</given-names></name> <name><surname>Romani</surname> <given-names>G. L.</given-names></name></person-group> (<year>2005</year>). <article-title>Right hemisphere specialization for intensity discrimination of musical and speech sounds</article-title>. <source>Neuropsychologia</source> <volume>43</volume>, <fpage>1916</fpage>&#x2013;<lpage>1923</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.neuropsychologia.2005.03.005</pub-id>, PMID: <pub-id pub-id-type="pmid">16168732</pub-id></citation></ref>
<ref id="ref10"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Brancucci</surname> <given-names>A.</given-names></name> <name><surname>D&#x2019;Anselmo</surname> <given-names>A.</given-names></name> <name><surname>Martello</surname> <given-names>F.</given-names></name> <name><surname>Tommasi</surname> <given-names>L.</given-names></name></person-group> (<year>2008</year>). <article-title>Left hemisphere specialization for duration discrimination of musical and speech sounds</article-title>. <source>Neuropsychologia</source> <volume>46</volume>, <fpage>2013</fpage>&#x2013;<lpage>2019</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.neuropsychologia.2008.01.019</pub-id>, PMID: <pub-id pub-id-type="pmid">18329056</pub-id></citation></ref>
<ref id="ref11"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Brennan</surname> <given-names>C.</given-names></name> <name><surname>Cao</surname> <given-names>F.</given-names></name> <name><surname>Pedroarena-leal</surname> <given-names>N.</given-names></name> <name><surname>Mcnorgan</surname> <given-names>C.</given-names></name> <name><surname>Booth</surname> <given-names>J. R.</given-names></name></person-group> (<year>2013</year>). <article-title>Reading acquisition reorganized the phonological awareness network only in alphabetic writing systems</article-title>. <source>Hum. Brain Mapp.</source> <volume>34</volume>, <fpage>3354</fpage>&#x2013;<lpage>3368</lpage>. doi: <pub-id pub-id-type="doi">10.1002/hbm.22147.Reading</pub-id></citation></ref>
<ref id="ref12"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bryden</surname> <given-names>M. P.</given-names></name> <name><surname>Murray</surname> <given-names>J. E.</given-names></name></person-group> (<year>1985</year>). <article-title>Toward a model of dichotic listening performance</article-title>. <source>Brain Cogn.</source> <volume>4</volume>, <fpage>241</fpage>&#x2013;<lpage>257</lpage>. doi: <pub-id pub-id-type="doi">10.1016/0278-2626(85)90019-3</pub-id>, PMID: <pub-id pub-id-type="pmid">4027059</pub-id></citation></ref>
<ref id="ref13"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Burton</surname> <given-names>M. W.</given-names></name> <name><surname>Small</surname> <given-names>S. L.</given-names></name> <name><surname>Blumstein</surname> <given-names>S. E.</given-names></name></person-group> (<year>2000</year>). <article-title>The role of segmentation in phonological processing: an fMRI investigation</article-title>. <source>J. Cogn. Neurosci.</source> <volume>12</volume>, <fpage>679</fpage>&#x2013;<lpage>690</lpage>. doi: <pub-id pub-id-type="doi">10.1162/089892900562309</pub-id>, PMID: <pub-id pub-id-type="pmid">10936919</pub-id></citation></ref>
<ref id="ref14"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chandrasekaran</surname> <given-names>B.</given-names></name> <name><surname>Gandour</surname> <given-names>J. T.</given-names></name> <name><surname>Krishnan</surname> <given-names>A.</given-names></name></person-group> (<year>2007a</year>). <article-title>Neuroplasticity in the processing of pitch dimensions: a multidimensional scaling analysis of the mismatch negativity</article-title>. <source>Restor. Neurol. Neurosci.</source> <volume>25</volume>, <fpage>195</fpage>&#x2013;<lpage>210</lpage>.</citation></ref>
<ref id="ref15"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chandrasekaran</surname> <given-names>B.</given-names></name> <name><surname>Krishnan</surname> <given-names>A.</given-names></name> <name><surname>Gandour</surname> <given-names>J. T.</given-names></name></person-group> (<year>2007b</year>). <article-title>Experience-dependent neural plasticity is sensitive to shape of pitch contours</article-title>. <source>Neuroreport</source> <volume>18</volume>, <fpage>1963</fpage>&#x2013;<lpage>1967</lpage>. doi: <pub-id pub-id-type="doi">10.1097/WNR.0b013e3282f213c5</pub-id>, PMID: <pub-id pub-id-type="pmid">18007195</pub-id></citation></ref>
<ref id="ref16"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cheung</surname> <given-names>H.</given-names></name> <name><surname>Chen</surname> <given-names>H. C.</given-names></name> <name><surname>Lai</surname> <given-names>C. Y.</given-names></name> <name><surname>Wong</surname> <given-names>O. C.</given-names></name> <name><surname>Hills</surname> <given-names>M.</given-names></name></person-group> (<year>2001</year>). <article-title>The development of phonological awareness: effects of spoken language experience and orthography</article-title>. <source>Cognition</source> <volume>81</volume>, <fpage>227</fpage>&#x2013;<lpage>241</lpage>. doi: <pub-id pub-id-type="doi">10.1016/S0010-0277(01)00136-6</pub-id>, PMID: <pub-id pub-id-type="pmid">11483171</pub-id></citation></ref>
<ref id="ref17"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cutting</surname> <given-names>J. E.</given-names></name></person-group> (<year>1974</year>). <article-title>Two left-hemisphere mechanisms in speech perception</article-title>. <source>Percept. Psychophys.</source> <volume>16</volume>, <fpage>601</fpage>&#x2013;<lpage>612</lpage>. doi: <pub-id pub-id-type="doi">10.3758/BF03198592</pub-id></citation></ref>
<ref id="ref18"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dehaene</surname> <given-names>S.</given-names></name> <name><surname>Pegado</surname> <given-names>F.</given-names></name> <name><surname>Braga</surname> <given-names>L. W.</given-names></name> <name><surname>Ventura</surname> <given-names>P.</given-names></name> <name><surname>Nunes Filho</surname> <given-names>G.</given-names></name> <name><surname>Jobert</surname> <given-names>A.</given-names></name> <etal/></person-group>. (<year>2010</year>). <article-title>How learning to read changes the cortical networks for vision and language</article-title>. <source>Science</source> <volume>330</volume>, <fpage>1359</fpage>&#x2013;<lpage>1364</lpage>. doi: <pub-id pub-id-type="doi">10.1126/science.1194140</pub-id>, PMID: <pub-id pub-id-type="pmid">21071632</pub-id></citation></ref>
<ref id="ref19"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Deng</surname> <given-names>Q.</given-names></name> <name><surname>Choi</surname> <given-names>W.</given-names></name> <name><surname>Tong</surname> <given-names>X.</given-names></name></person-group> (<year>2019</year>). <article-title>Bidirectional cross-linguistic association of phonological skills and reading comprehension: evidence from Hong Kong Chinese-English bilingual readers</article-title>. <source>J. Learn. Disabil.</source> <volume>52</volume>, <fpage>299</fpage>&#x2013;<lpage>311</lpage>. doi: <pub-id pub-id-type="doi">10.1177/0022219419842914</pub-id>, PMID: <pub-id pub-id-type="pmid">31046555</pub-id></citation></ref>
<ref id="ref20"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dodd</surname> <given-names>B. J.</given-names></name> <name><surname>So</surname> <given-names>L. K. H.</given-names></name> <name><surname>Lam</surname> <given-names>K. K. C.</given-names></name></person-group> (<year>2008</year>). <article-title>Bilingualism and learning: the effect of language pair on phonological awareness abilities</article-title>. <source>Aust. J. Learn. Difficulties</source> <volume>13</volume>, <fpage>99</fpage>&#x2013;<lpage>113</lpage>. doi: <pub-id pub-id-type="doi">10.1080/19404150802380514</pub-id></citation></ref>
<ref id="ref21"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dwyer</surname> <given-names>J.</given-names></name> <name><surname>Blumstein</surname> <given-names>S. E.</given-names></name> <name><surname>Ryalls</surname> <given-names>J.</given-names></name></person-group> (<year>1982</year>). <article-title>The role of duration and rapid temporal processing on the lateral perception of consonants and vowels</article-title>. <source>Brain Lang.</source> <volume>17</volume>, <fpage>272</fpage>&#x2013;<lpage>286</lpage>. doi: <pub-id pub-id-type="doi">10.1016/0093-934X(82)90021-9</pub-id>, PMID: <pub-id pub-id-type="pmid">7159836</pub-id></citation></ref>
<ref id="ref22"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Flinker</surname> <given-names>A.</given-names></name> <name><surname>Doyle</surname> <given-names>W. K.</given-names></name> <name><surname>Mehta</surname> <given-names>A. D.</given-names></name> <name><surname>Devinsky</surname> <given-names>O.</given-names></name> <name><surname>Poeppel</surname> <given-names>D.</given-names></name></person-group> (<year>2019</year>). <article-title>Spectrotemporal modulation provides a unifying framework for auditory cortical asymmetries</article-title>. <source>Nat. Hum. Behav.</source> <volume>3</volume>, <fpage>393</fpage>&#x2013;<lpage>405</lpage>. doi: <pub-id pub-id-type="doi">10.1038/s41562-019-0548-z</pub-id>, PMID: <pub-id pub-id-type="pmid">30971792</pub-id></citation></ref>
<ref id="ref23"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gandour</surname> <given-names>J.</given-names></name></person-group> (<year>1983</year>). <article-title>Tone perception in far eastern languages</article-title>. <source>J. Phon.</source> <volume>11</volume>, <fpage>149</fpage>&#x2013;<lpage>175</lpage>. doi: <pub-id pub-id-type="doi">10.1016/S0095-4470(19)30813-7</pub-id></citation></ref>
<ref id="ref24"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gandour</surname> <given-names>J.</given-names></name> <name><surname>Wong</surname> <given-names>D.</given-names></name> <name><surname>Hsieh</surname> <given-names>L.</given-names></name> <name><surname>Weinzapfel</surname> <given-names>B.</given-names></name> <name><surname>Van Lancker</surname> <given-names>D.</given-names></name> <name><surname>Hutchins</surname> <given-names>G. D.</given-names></name></person-group> (<year>2000</year>). <article-title>A crosslinguistic PET study of tone perception</article-title>. <source>J. Cogn. Neurosci.</source> <volume>12</volume>, <fpage>207</fpage>&#x2013;<lpage>222</lpage>. doi: <pub-id pub-id-type="doi">10.1162/089892900561841</pub-id>, PMID: <pub-id pub-id-type="pmid">10769317</pub-id></citation></ref>
<ref id="ref25"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gandour Wong</surname> <given-names>D.</given-names></name> <name><surname>Hutchins</surname> <given-names>G. J.</given-names></name></person-group> (<year>1998</year>). <article-title>Pitch processing in the human brain is influenced by language experience</article-title>. <source>Neuroreport</source> <volume>9</volume>, <fpage>2115</fpage>&#x2013;<lpage>2119</lpage>. doi: <pub-id pub-id-type="doi">10.1097/00001756-199806220-00038</pub-id>, PMID: <pub-id pub-id-type="pmid">9674604</pub-id></citation></ref>
<ref id="ref26"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gu</surname> <given-names>F.</given-names></name> <name><surname>Zhang</surname> <given-names>C.</given-names></name> <name><surname>Hu</surname> <given-names>A.</given-names></name> <name><surname>Zhao</surname> <given-names>G.</given-names></name></person-group> (<year>2013</year>). <article-title>Left hemisphere lateralization for lexical and acoustic pitch processing in Cantonese speakers as revealed by mismatch negativity</article-title>. <source>NeuroImage</source> <volume>83</volume>, <fpage>637</fpage>&#x2013;<lpage>645</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.neuroimage.2013.02.080</pub-id></citation></ref>
<ref id="ref28"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Ho</surname> <given-names>J. P.-K.</given-names></name></person-group> (<year>2010</year>). <source>An ERP Study on the Effect of Tone Features on Lexical Tone Lateralization in Cantonese</source>. <publisher-loc>Ma Liu Shui</publisher-loc>: <publisher-name>The Chinese University of Hong Kong</publisher-name>.</citation></ref>
<ref id="ref29"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hoch</surname> <given-names>L.</given-names></name> <name><surname>Tillmann</surname> <given-names>B.</given-names></name></person-group> (<year>2010</year>). <article-title>Laterality effects for musical structure processing: a dichotic listening study</article-title>. <source>Neuropsychology</source> <volume>24</volume>, <fpage>661</fpage>&#x2013;<lpage>666</lpage>. doi: <pub-id pub-id-type="doi">10.1037/a0019653</pub-id>, PMID: <pub-id pub-id-type="pmid">20804254</pub-id></citation></ref>
<ref id="ref30"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hugdahl</surname> <given-names>K.</given-names></name> <name><surname>Br&#x00F8;nnick</surname> <given-names>K.</given-names></name> <name><surname>Kyllingsbrk</surname> <given-names>S.</given-names></name> <name><surname>Law</surname> <given-names>I.</given-names></name> <name><surname>Gade</surname> <given-names>A.</given-names></name> <name><surname>Paulson</surname> <given-names>O. B.</given-names></name></person-group> (<year>1999</year>). <article-title>Brain activation during dichotic presentations of consonant-vowel and musical instrument stimuli: a 15O-PET study</article-title>. <source>Neuropsychologia</source> <volume>37</volume>, <fpage>431</fpage>&#x2013;<lpage>440</lpage>. doi: <pub-id pub-id-type="doi">10.1016/S0028-3932(98)00101-8</pub-id>, PMID: <pub-id pub-id-type="pmid">10215090</pub-id></citation></ref>
<ref id="ref31"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jia</surname> <given-names>S.</given-names></name> <name><surname>Tsang</surname> <given-names>Y.-K.</given-names></name> <name><surname>Huang</surname> <given-names>J.</given-names></name> <name><surname>Chen</surname> <given-names>H.-C.</given-names></name></person-group> (<year>2013</year>). <article-title>Right hemisphere advantage in processing Cantonese level and contour tones: evidence from dichotic listening</article-title>. <source>Neurosci. Lett.</source> <volume>556</volume>, <fpage>135</fpage>&#x2013;<lpage>139</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.neulet.2013.10.014</pub-id>, PMID: <pub-id pub-id-type="pmid">24140004</pub-id></citation></ref>
<ref id="ref32"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jia</surname> <given-names>S.</given-names></name> <name><surname>Tsang</surname> <given-names>Y.-K.</given-names></name> <name><surname>Huang</surname> <given-names>J.</given-names></name> <name><surname>Chen</surname> <given-names>H.-C.</given-names></name></person-group> (<year>2015</year>). <article-title>Processing Cantonese lexical tones: evidence from oddball paradigms</article-title>. <source>Neuroscience</source> <volume>305</volume>, <fpage>351</fpage>&#x2013;<lpage>360</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.neuroscience.2015.08.009</pub-id>, PMID: <pub-id pub-id-type="pmid">26265553</pub-id></citation></ref>
<ref id="ref33"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Khouw</surname> <given-names>E.</given-names></name> <name><surname>Ciocca</surname> <given-names>V.</given-names></name></person-group> (<year>2007</year>). <article-title>Perceptual correlates of Cantonese tones</article-title>. <source>J. Phon.</source> <volume>35</volume>, <fpage>104</fpage>&#x2013;<lpage>117</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.wocn.2005.10.003</pub-id></citation></ref>
<ref id="ref34"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kimura</surname> <given-names>D.</given-names></name></person-group> (<year>1961</year>). <article-title>Cerebral dominance and the perception of verbal stimuli</article-title>. <source>Can. J. Psychol./Revue Canadienne de Psychologie</source> <volume>15</volume>, <fpage>166</fpage>&#x2013;<lpage>171</lpage>. doi: <pub-id pub-id-type="doi">10.1037/h0083219</pub-id></citation></ref>
<ref id="ref35"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kimura</surname> <given-names>D.</given-names></name></person-group> (<year>1967</year>). <article-title>Functional asymmetry of the brain in dichotic listening</article-title>. <source>Cortex</source> <volume>3</volume>, <fpage>163</fpage>&#x2013;<lpage>178</lpage>. doi: <pub-id pub-id-type="doi">10.1016/S0010-9452(67)80010-8</pub-id></citation></ref>
<ref id="ref36"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Krishnan</surname> <given-names>A.</given-names></name> <name><surname>Gandour</surname> <given-names>J. T.</given-names></name></person-group> (<year>2009</year>). <article-title>The role of the auditory brainstem in processing linguistically relevant pitch patterns</article-title>. <source>Brain Lang.</source> <volume>110</volume>, <fpage>135</fpage>&#x2013;<lpage>148</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.bandl.2009.03.005</pub-id>, PMID: <pub-id pub-id-type="pmid">19366639</pub-id></citation></ref>
<ref id="ref001"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kuznetsova</surname> <given-names>A.</given-names></name> <name><surname>Brockhoff</surname> <given-names>P. B.</given-names></name> <name><surname>Christensen</surname> <given-names>R. H.</given-names></name></person-group> (<year>2017</year>). <article-title>lmerTest package: tests in linear mixed effects models</article-title>. <source>J. Stat. Softw.</source> <volume>82</volume>, <fpage>1</fpage>&#x2013;<lpage>23</lpage>.</citation></ref>
<ref id="ref37"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lee</surname> <given-names>C.-Y.</given-names></name> <name><surname>Tao</surname> <given-names>L.</given-names></name> <name><surname>Bond</surname> <given-names>Z. S.</given-names></name></person-group> (<year>2008</year>). <article-title>Identification of acoustically modified mandarin tones by native listeners</article-title>. <source>J. Phon.</source> <volume>36</volume>, <fpage>537</fpage>&#x2013;<lpage>563</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.wocn.2008.01.002</pub-id></citation></ref>
<ref id="ref002"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lenth</surname> <given-names>R.</given-names></name> <name><surname>Lenth</surname> <given-names>M. R.</given-names></name></person-group> (<year>2018</year>). <article-title>Package &#x2018;lsmeans&#x2019;</article-title>. <source>Am. Stat.</source> <volume>34</volume>, <fpage>216</fpage>&#x2013;<lpage>221</lpage>., PMID: <pub-id pub-id-type="pmid">21092370</pub-id></citation></ref>
<ref id="ref39"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Li</surname> <given-names>W. S.</given-names></name> <name><surname>Suk-Han Ho</surname> <given-names>C.</given-names></name></person-group> (<year>2011</year>). <article-title>Lexical tone awareness among Chinese children with developmental dyslexia</article-title>. <source>J. Child Lang.</source> <volume>38</volume>, <fpage>793</fpage>&#x2013;<lpage>808</lpage>. doi: <pub-id pub-id-type="doi">10.1017/S0305000910000346</pub-id>, PMID: <pub-id pub-id-type="pmid">21092370</pub-id></citation></ref>
<ref id="ref40"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Liang</surname> <given-names>B.</given-names></name> <name><surname>Du</surname> <given-names>Y.</given-names></name></person-group> (<year>2018</year>). <article-title>The functional neuroanatomy of lexical tone perception: an activation likelihood estimation meta-analysis</article-title>. <source>Front. Neurosci.</source> <volume>12</volume>, <fpage>1</fpage>&#x2013;<lpage>17</lpage>. doi: <pub-id pub-id-type="doi">10.3389/fnins.2018.00495</pub-id>, PMID: <pub-id pub-id-type="pmid">30087589</pub-id></citation></ref>
<ref id="ref41"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Liberman</surname> <given-names>I. Y.</given-names></name> <name><surname>Shankweiler</surname> <given-names>D.</given-names></name> <name><surname>Fischer</surname> <given-names>F. W.</given-names></name> <name><surname>Carter</surname> <given-names>B.</given-names></name></person-group> (<year>1974</year>). <article-title>Explicit syllable and phoneme segmentation in the young child</article-title>. <source>J. Exp. Child Psychol.</source> <volume>18</volume>, <fpage>201</fpage>&#x2013;<lpage>212</lpage>. doi: <pub-id pub-id-type="doi">10.1016/0022-0965(74)90101-5</pub-id></citation></ref>
<ref id="ref42"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Liberman</surname> <given-names>A. M.</given-names></name> <name><surname>Whalen</surname> <given-names>D. H.</given-names></name></person-group> (<year>2000</year>). <article-title>On the relation of speech to language</article-title>. <source>Trends Cogn. Sci.</source> <volume>4</volume>, <fpage>187</fpage>&#x2013;<lpage>196</lpage>. doi: <pub-id pub-id-type="doi">10.1016/S1364-6613(00)01471-6</pub-id></citation></ref>
<ref id="ref43"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Liu</surname> <given-names>F.</given-names></name> <name><surname>Patel</surname> <given-names>A. D.</given-names></name> <name><surname>Fourcin</surname> <given-names>A.</given-names></name> <name><surname>Stewart</surname> <given-names>L.</given-names></name></person-group> (<year>2010</year>). <article-title>Intonation processing in congenital amusia: discrimination, identification and imitation</article-title>. <source>Brain</source> <volume>133</volume>, <fpage>1682</fpage>&#x2013;<lpage>1693</lpage>. doi: <pub-id pub-id-type="doi">10.1093/brain/awq089</pub-id>, PMID: <pub-id pub-id-type="pmid">20418275</pub-id></citation></ref>
<ref id="ref44"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Luo</surname> <given-names>H.</given-names></name> <name><surname>Ni</surname> <given-names>J.-T.</given-names></name> <name><surname>Li</surname> <given-names>Z.-H.</given-names></name> <name><surname>Li</surname> <given-names>X.-O.</given-names></name> <name><surname>Zhang</surname> <given-names>D.-R.</given-names></name> <name><surname>Zeng</surname> <given-names>F.-G.</given-names></name> <etal/></person-group>. (<year>2006</year>). <article-title>Opposite patterns of hemisphere dominance for early auditory processing of lexical tones and consonants</article-title>. <source>Proc. Natl. Acad. Sci.</source> <volume>103</volume>, <fpage>19558</fpage>&#x2013;<lpage>19563</lpage>. doi: <pub-id pub-id-type="doi">10.1073/pnas.0607065104</pub-id>, PMID: <pub-id pub-id-type="pmid">17159136</pub-id></citation></ref>
<ref id="ref45"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mei</surname> <given-names>N.</given-names></name> <name><surname>Flinker</surname> <given-names>A.</given-names></name> <name><surname>Zhu</surname> <given-names>M.</given-names></name> <name><surname>Cai</surname> <given-names>Q.</given-names></name> <name><surname>Tian</surname> <given-names>X.</given-names></name></person-group> (<year>2020</year>). <article-title>Lateralization in the dichotic listening of tones is influenced by the content of speech</article-title>. <source>Neuropsychologia</source> <volume>140</volume>:<fpage>107389</fpage>. doi: <pub-id pub-id-type="doi">10.1016/j.neuropsychologia.2020.107389</pub-id>, PMID: <pub-id pub-id-type="pmid">32057939</pub-id></citation></ref>
<ref id="ref46"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Moen</surname> <given-names>I.</given-names></name></person-group> (<year>1993</year>). <article-title>Functional lateralization of the perception of Norwegian word tones-evidence from a dichotic listening experiment</article-title>. <source>Brain Lang.</source> <volume>44</volume>, <fpage>400</fpage>&#x2013;<lpage>413</lpage>. doi: <pub-id pub-id-type="doi">10.1006/brln.1993.1024</pub-id>, PMID: <pub-id pub-id-type="pmid">8319080</pub-id></citation></ref>
<ref id="ref47"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Morais</surname> <given-names>J.</given-names></name></person-group> (<year>2021</year>). <article-title>The phoneme: a conceptual heritage from alphabetic literacy</article-title>. <source>Cognition</source> <volume>213</volume>:<fpage>104740</fpage>. doi: <pub-id pub-id-type="doi">10.1016/j.cognition.2021.104740</pub-id>, PMID: <pub-id pub-id-type="pmid">33895003</pub-id></citation></ref>
<ref id="ref48"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Morais</surname> <given-names>J.</given-names></name> <name><surname>Cary</surname> <given-names>L.</given-names></name> <name><surname>Alegria</surname> <given-names>J.</given-names></name> <name><surname>Bertelson</surname> <given-names>P.</given-names></name></person-group> (<year>1979</year>). <article-title>Does awareness of speech as a sequence of phones arise spontaneously?</article-title> <source>Cognition</source> <volume>7</volume>, <fpage>323</fpage>&#x2013;<lpage>331</lpage>. doi: <pub-id pub-id-type="doi">10.1016/0010-0277(79)90020-9</pub-id></citation></ref>
<ref id="ref49"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Morett</surname> <given-names>L. M.</given-names></name> <name><surname>Chang</surname> <given-names>L.-Y.</given-names></name></person-group> (<year>2015</year>). <article-title>Emphasising sound and meaning: pitch gestures enhance mandarin lexical tone acquisition</article-title>. <source>Lang. Cognit. Neurosci.</source> <volume>30</volume>, <fpage>347</fpage>&#x2013;<lpage>353</lpage>. doi: <pub-id pub-id-type="doi">10.1080/23273798.2014.923105</pub-id></citation></ref>
<ref id="ref50"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Oldfield</surname> <given-names>R. C.</given-names></name></person-group> (<year>1971</year>). <article-title>The assessment and analysis of handedness: the Edinburgh inventory</article-title>. <source>Neuropsychologia</source> <volume>9</volume>, <fpage>97</fpage>&#x2013;<lpage>113</lpage>. doi: <pub-id pub-id-type="doi">10.1016/0028-3932(71)90067-4</pub-id>, PMID: <pub-id pub-id-type="pmid">5146491</pub-id></citation></ref>
<ref id="ref51"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Pattamadilok</surname> <given-names>C.</given-names></name> <name><surname>Knierim</surname> <given-names>I. N.</given-names></name> <name><surname>Kawabata Duncan</surname> <given-names>K. J.</given-names></name> <name><surname>Devlin</surname> <given-names>J. T.</given-names></name></person-group> (<year>2010</year>). <article-title>How does learning to read affect speech perception?</article-title> <source>J. Neurosci.</source> <volume>30</volume>, <fpage>8435</fpage>&#x2013;<lpage>8444</lpage>. doi: <pub-id pub-id-type="doi">10.1523/JNEUROSCI.5791-09.2010</pub-id>, PMID: <pub-id pub-id-type="pmid">20573891</pub-id></citation></ref>
<ref id="ref52"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Perre</surname> <given-names>L.</given-names></name> <name><surname>Pattamadilok</surname> <given-names>C.</given-names></name> <name><surname>Montant</surname> <given-names>M.</given-names></name> <name><surname>Ziegler</surname> <given-names>J. C.</given-names></name></person-group> (<year>2009</year>). <article-title>Orthographic effects in spoken language: on-line activation or phonological restructuring?</article-title> <source>Brain Res.</source> <volume>1275</volume>, <fpage>73</fpage>&#x2013;<lpage>80</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.brainres.2009.04.018</pub-id>, PMID: <pub-id pub-id-type="pmid">19376099</pub-id></citation></ref>
<ref id="ref53"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Poeppel</surname> <given-names>D.</given-names></name></person-group> (<year>2003</year>). <article-title>The analysis of speech in different temporal integration windows: cerebral lateralization as &#x2018;asymmetric sampling in time&#x2019;</article-title>. <source>Speech Comm.</source> <volume>41</volume>, <fpage>245</fpage>&#x2013;<lpage>255</lpage>. doi: <pub-id pub-id-type="doi">10.1016/S0167-6393(02)00107-3</pub-id></citation></ref>
<ref id="ref54"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Qin</surname> <given-names>Z.</given-names></name> <name><surname>Jongman</surname> <given-names>A.</given-names></name></person-group> (<year>2016</year>). <article-title>Does second language experience modulate perception of tones in a third language?</article-title> <source>Lang. Speech</source> <volume>59</volume>, <fpage>318</fpage>&#x2013;<lpage>338</lpage>. doi: <pub-id pub-id-type="doi">10.1177/0023830915590191</pub-id></citation></ref>
<ref id="ref55"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Read</surname> <given-names>C.</given-names></name> <name><surname>Zhang</surname> <given-names>Y.-F.</given-names></name> <name><surname>Nie</surname> <given-names>H.-Y.</given-names></name> <name><surname>Ding</surname> <given-names>B.-Q.</given-names></name></person-group> (<year>1986</year>). <article-title>The ability to manipulate speech sounds depends on knowing alphabetic reading</article-title>. <source>Cognition</source> <volume>24</volume>, <fpage>31</fpage>&#x2013;<lpage>44</lpage>. doi: <pub-id pub-id-type="doi">10.1016/0010-0277(86)90003-X</pub-id>, PMID: <pub-id pub-id-type="pmid">3791920</pub-id></citation></ref>
<ref id="ref56"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Reilly</surname> <given-names>M.</given-names></name> <name><surname>Machado</surname> <given-names>N.</given-names></name> <name><surname>Blumstein</surname> <given-names>S. E.</given-names></name></person-group> (<year>2015</year>). <article-title>Hemispheric lateralization of semantic feature distinctiveness</article-title>. <source>Neuropsychologia</source> <volume>75</volume>, <fpage>99</fpage>&#x2013;<lpage>108</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.neuropsychologia.2015.05.025</pub-id>, PMID: <pub-id pub-id-type="pmid">26022059</pub-id></citation></ref>
<ref id="ref57"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sanders</surname> <given-names>L. D.</given-names></name> <name><surname>Poeppel</surname> <given-names>D.</given-names></name></person-group> (<year>2007</year>). <article-title>Local and global auditory processing: behavioral and ERP evidence</article-title>. <source>Neuropsychologia</source> <volume>45</volume>, <fpage>1172</fpage>&#x2013;<lpage>1186</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.neuropsychologia.2006.10.010</pub-id>, PMID: <pub-id pub-id-type="pmid">17113115</pub-id></citation></ref>
<ref id="ref58"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Shao</surname> <given-names>J.</given-names></name> <name><surname>Lau</surname> <given-names>R. Y. M.</given-names></name> <name><surname>Tang</surname> <given-names>P. O. C.</given-names></name> <name><surname>Zhang</surname> <given-names>C.</given-names></name></person-group> (<year>2019</year>). <article-title>The effects of acoustic variation on the perception of lexical tone in Cantonese-speaking congenital amusics</article-title>. <source>J. Speech Lang. Hear. Res.</source> <volume>62</volume>, <fpage>190</fpage>&#x2013;<lpage>205</lpage>. doi: <pub-id pub-id-type="doi">10.1044/2018_JSLHR-H-17-0483</pub-id>, PMID: <pub-id pub-id-type="pmid">30950752</pub-id></citation></ref>
<ref id="ref59"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Shao</surname> <given-names>J.</given-names></name> <name><surname>Zhang</surname> <given-names>C.</given-names></name></person-group> (<year>2020</year>). <article-title>Dichotic perception of lexical tones in Cantonese-speaking congenital amusics</article-title>. <source>Front. Psychol.</source> <volume>11</volume>, <fpage>1</fpage>&#x2013;<lpage>12</lpage>. doi: <pub-id pub-id-type="doi">10.3389/fpsyg.2020.01411</pub-id>, PMID: <pub-id pub-id-type="pmid">32733321</pub-id></citation></ref>
<ref id="ref60"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Shu</surname> <given-names>H.</given-names></name> <name><surname>Peng</surname> <given-names>H.</given-names></name> <name><surname>McBride-Chang</surname> <given-names>C.</given-names></name></person-group> (<year>2008</year>). <article-title>Phonological awareness in young Chinese children</article-title>. <source>Dev. Sci.</source> <volume>11</volume>, <fpage>171</fpage>&#x2013;<lpage>181</lpage>. doi: <pub-id pub-id-type="doi">10.1111/j.1467-7687.2007.00654.x</pub-id>, PMID: <pub-id pub-id-type="pmid">18171377</pub-id></citation></ref>
<ref id="ref61"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Shuai</surname> <given-names>L.</given-names></name></person-group> (<year>2009</year>). <source>ERP Studies of tone Lateralization</source>. <publisher-loc>Ma Liu Shui</publisher-loc>: <publisher-name>Chinese University of Hong Kong</publisher-name>.</citation></ref>
<ref id="ref62"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Shuai</surname> <given-names>L.</given-names></name> <name><surname>Gong</surname> <given-names>T.</given-names></name></person-group> (<year>2014</year>). <article-title>Temporal relation between top-down and bottom-up processing in lexical tone perception</article-title>. <source>Front. Behav. Neurosci.</source> <volume>8</volume>, <fpage>1</fpage>&#x2013;<lpage>16</lpage>. doi: <pub-id pub-id-type="doi">10.3389/fnbeh.2014.00097</pub-id>, PMID: <pub-id pub-id-type="pmid">24723863</pub-id></citation></ref>
<ref id="ref63"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Studdert-Kennedy</surname> <given-names>M.</given-names></name> <name><surname>Shankweiler</surname> <given-names>D.</given-names></name></person-group> (<year>1970</year>). <article-title>Hemispheric specialization for speech perception</article-title>. <source>J. Acoust. Soc. Am.</source> <volume>48</volume>, <fpage>579</fpage>&#x2013;<lpage>594</lpage>. doi: <pub-id pub-id-type="doi">10.1121/1.1912174</pub-id></citation></ref>
<ref id="ref65"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Teng</surname> <given-names>X.</given-names></name> <name><surname>Tian</surname> <given-names>X.</given-names></name> <name><surname>Poeppel</surname> <given-names>D.</given-names></name></person-group> (<year>2016</year>). <article-title>Testing multi-scale processing in the auditory system</article-title>. <source>Sci. Rep.</source> <volume>6</volume>:<fpage>34390</fpage>. doi: <pub-id pub-id-type="doi">10.1038/srep34390</pub-id>, PMID: <pub-id pub-id-type="pmid">27713546</pub-id></citation></ref>
<ref id="ref66"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tsang</surname> <given-names>Y.-K.</given-names></name> <name><surname>Jia</surname> <given-names>S.</given-names></name> <name><surname>Huang</surname> <given-names>J.</given-names></name> <name><surname>Chen</surname> <given-names>H.-C.</given-names></name></person-group> (<year>2011</year>). <article-title>ERP correlates of pre-attentive processing of Cantonese lexical tones: the effects of pitch contour and pitch height</article-title>. <source>Neurosci. Lett.</source> <volume>487</volume>, <fpage>268</fpage>&#x2013;<lpage>272</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.neulet.2010.10.035</pub-id>, PMID: <pub-id pub-id-type="pmid">20970477</pub-id></citation></ref>
<ref id="ref67"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Van Lancker</surname> <given-names>D.</given-names></name></person-group> (<year>1980</year>). <article-title>Cerebral lateralization of pitch cues in the linguistic signal</article-title>. <source>Res. Lang. Soc. Interact.</source> <volume>13</volume>, <fpage>201</fpage>&#x2013;<lpage>277</lpage>. doi: <pub-id pub-id-type="doi">10.1080/08351818009370498</pub-id></citation></ref>
<ref id="ref68"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Van Lancker</surname> <given-names>D.</given-names></name> <name><surname>Fromkin</surname> <given-names>V. A.</given-names></name></person-group> (<year>1973</year>). <article-title>Hemispheric specialization for pitch and &#x201C;tone&#x201D;: evidence from Thai</article-title>. <source>J. Phon.</source> <volume>1</volume>, <fpage>101</fpage>&#x2013;<lpage>109</lpage>. doi: <pub-id pub-id-type="doi">10.1016/S0095-4470(19)31414-7</pub-id></citation></ref>
<ref id="ref69"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Van Lancker</surname> <given-names>D.</given-names></name> <name><surname>Fromkin</surname> <given-names>V. A.</given-names></name></person-group> (<year>1978</year>). <article-title>Cerebral dominance for pitch contrasts in tone language speakers and in musically untrained and trained English speakers</article-title>. <source>J. Phon.</source> <volume>6</volume>, <fpage>19</fpage>&#x2013;<lpage>23</lpage>. doi: <pub-id pub-id-type="doi">10.1016/S0095-4470(19)31082-4</pub-id></citation></ref>
<ref id="ref70"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Voyer</surname> <given-names>D.</given-names></name></person-group> (<year>1998</year>). <article-title>On the reliability and validity of noninvasive laterality measures</article-title>. <source>Brain Cogn.</source> <volume>36</volume>, <fpage>209</fpage>&#x2013;<lpage>236</lpage>. doi: <pub-id pub-id-type="doi">10.1006/brcg.1997.0953</pub-id>, PMID: <pub-id pub-id-type="pmid">9520314</pub-id></citation></ref>
<ref id="ref71"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wang</surname> <given-names>Y.</given-names></name> <name><surname>Jongman</surname> <given-names>A.</given-names></name> <name><surname>Sereno</surname> <given-names>J. A.</given-names></name></person-group> (<year>2001</year>). <article-title>Dichotic perception of mandarin tones by Chinese and American listeners</article-title>. <source>Brain Lang.</source> <volume>78</volume>, <fpage>332</fpage>&#x2013;<lpage>348</lpage>. doi: <pub-id pub-id-type="doi">10.1006/brln.2001.2474</pub-id>, PMID: <pub-id pub-id-type="pmid">11703061</pub-id></citation></ref>
<ref id="ref72"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Westerhausen</surname> <given-names>R.</given-names></name></person-group> (<year>2019</year>). <article-title>A primer on dichotic listening as a paradigm for the assessment of hemispheric asymmetry</article-title>. <source>Laterality</source> <volume>24</volume>, <fpage>740</fpage>&#x2013;<lpage>771</lpage>. doi: <pub-id pub-id-type="doi">10.1080/1357650X.2019.1598426</pub-id>, PMID: <pub-id pub-id-type="pmid">30922169</pub-id></citation></ref>
<ref id="ref73"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Westerhausen</surname> <given-names>R.</given-names></name> <name><surname>Samuelsen</surname> <given-names>F.</given-names></name></person-group> (<year>2020</year>). <article-title>An optimal dichotic-listening paradigm for the assessment of hemispheric dominance for speech processing</article-title>. <source>PLoS One</source> <volume>15</volume>, <fpage>e0234611</fpage>&#x2013;<lpage>e0234665</lpage>. doi: <pub-id pub-id-type="doi">10.1371/journal.pone.0234665</pub-id>, PMID: <pub-id pub-id-type="pmid">32544204</pub-id></citation></ref>
<ref id="ref74"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Whalen</surname> <given-names>D. H.</given-names></name> <name><surname>Liberman</surname> <given-names>A. M.</given-names></name></person-group> (<year>1987</year>). <article-title>Speech perception takes precedence over nonspeech perception</article-title>. <source>Science</source> <volume>237</volume>, <fpage>169</fpage>&#x2013;<lpage>171</lpage>. doi: <pub-id pub-id-type="doi">10.1126/science.3603014</pub-id>, PMID: <pub-id pub-id-type="pmid">3603014</pub-id></citation></ref>
<ref id="ref75"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wioland</surname> <given-names>N.</given-names></name> <name><surname>Rudolf</surname> <given-names>G.</given-names></name> <name><surname>Metz-Lutz</surname> <given-names>M. N.</given-names></name> <name><surname>Mutschler</surname> <given-names>V.</given-names></name> <name><surname>Marescaux</surname> <given-names>C.</given-names></name></person-group> (<year>1999</year>). <article-title>Cerebral correlates of hemispheric lateralization during a pitch discrimination task: an ERP study in dichotic situation</article-title>. <source>Clin. Neurophysiol.</source> <volume>110</volume>, <fpage>516</fpage>&#x2013;<lpage>523</lpage>. doi: <pub-id pub-id-type="doi">10.1016/S1388-2457(98)00051-0</pub-id>, PMID: <pub-id pub-id-type="pmid">10363775</pub-id></citation></ref>
<ref id="ref76"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wong</surname> <given-names>P. C. M.</given-names></name> <name><surname>Skoe</surname> <given-names>E.</given-names></name> <name><surname>Russo</surname> <given-names>N. M.</given-names></name> <name><surname>Dees</surname> <given-names>T.</given-names></name> <name><surname>Kraus</surname> <given-names>N.</given-names></name></person-group> (<year>2007</year>). <article-title>Musical experience shapes human brainstem encoding of linguistic pitch patterns</article-title>. <source>Nat. Neurosci.</source> <volume>10</volume>, <fpage>420</fpage>&#x2013;<lpage>422</lpage>. doi: <pub-id pub-id-type="doi">10.1038/nn1872</pub-id>, PMID: <pub-id pub-id-type="pmid">17351633</pub-id></citation></ref>
<ref id="ref77"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zatorre</surname> <given-names>R. J.</given-names></name> <name><surname>Belin</surname> <given-names>P.</given-names></name></person-group> (<year>2001</year>). <article-title>Spectral and temporal processing in human auditory cortex</article-title>. <source>Cereb. Cortex</source> <volume>11</volume>, <fpage>946</fpage>&#x2013;<lpage>953</lpage>. doi: <pub-id pub-id-type="doi">10.1093/cercor/11.10.946</pub-id></citation></ref>
<ref id="ref78"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zatorre</surname> <given-names>R. J.</given-names></name> <name><surname>Belin</surname> <given-names>P.</given-names></name> <name><surname>Penhune</surname> <given-names>V. B.</given-names></name></person-group> (<year>2002</year>). <article-title>Structure and function of auditory cortex: music and speech</article-title>. <source>Trends Cogn. Sci.</source> <volume>6</volume>, <fpage>37</fpage>&#x2013;<lpage>46</lpage>. doi: <pub-id pub-id-type="doi">10.1016/S1364-6613(00)01816-7</pub-id></citation></ref>
<ref id="ref81"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zhang</surname> <given-names>C.</given-names></name> <name><surname>Ho</surname> <given-names>O. Y.</given-names></name> <name><surname>Shao</surname> <given-names>J.</given-names></name> <name><surname>Ou</surname> <given-names>J.</given-names></name> <name><surname>Law</surname> <given-names>S. P.</given-names></name></person-group> (<year>2021</year>). <article-title>Dissociation of tone merger and congenital amusia in Hong Kong Cantonese</article-title>. <source>PLoS One</source> <volume>16</volume>:<fpage>e0253982</fpage>. doi: <pub-id pub-id-type="doi">10.1371/journal.pone.0253982</pub-id>, PMID: <pub-id pub-id-type="pmid">34197546</pub-id></citation></ref>
<ref id="ref82"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zhang</surname> <given-names>Y.</given-names></name> <name><surname>Pattamadilok</surname> <given-names>C.</given-names></name> <name><surname>Lau</surname> <given-names>D. K. Y.</given-names></name> <name><surname>Bakhtiar</surname> <given-names>M.</given-names></name> <name><surname>Yim</surname> <given-names>L. Y.</given-names></name> <name><surname>Leung</surname> <given-names>K. Y.</given-names></name> <etal/></person-group>. (<year>2021</year>). <article-title>Early auditory event-related potentials are modulated by alphabetic literacy skills in logographic Chinese readers</article-title>. <source>Front. Psychol.</source> <volume>12</volume>, <fpage>1</fpage>&#x2013;<lpage>14</lpage>. doi: <pub-id pub-id-type="doi">10.3389/fpsyg.2021.663166</pub-id>, PMID: <pub-id pub-id-type="pmid">34393900</pub-id></citation></ref>
<ref id="ref83"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zhang</surname> <given-names>C.</given-names></name> <name><surname>Shao</surname> <given-names>J.</given-names></name> <name><surname>Huang</surname> <given-names>X.</given-names></name></person-group> (<year>2017</year>). <article-title>Deficits of congenital amusia beyond pitch: evidence from impaired categorical perception of vowels in Cantonese-speaking congenital amusics</article-title>. <source>PLoS One</source> <volume>12</volume>:<fpage>e0183151</fpage>. doi: <pub-id pub-id-type="doi">10.1371/journal.pone.0183151</pub-id>, PMID: <pub-id pub-id-type="pmid">28829808</pub-id></citation></ref>
</ref-list>
<fn-group>
<fn id="fn0004">
<p><sup>1</sup>Item was not treated as a random factor in the model because identification accuracy was a set of calculated data which was averaged across items (tone pairs). First, the response to each trial was categorized as either accurate in the left ear or the right ear, or inaccurate. For example, in one trial, the right ear was presented with T1 and the left ear was presented with T3. If a participant&#x2019;s response was T1, this trial was considered as correct in the right ear, and if the response was T3, it was considered as correct in the left ear; if the response was neither T1 nor T3, it was deemed as incorrect. Identification accuracy was then computed as the relative portion of correct responses in each ear; during this process, item (tone pair) info was not preserved.</p>
</fn>
<fn id="fn0005">
<p><sup>2</sup>As the stimulus set contained tones T3-T6/T6-T3 and T2-T5/T5-T2, which are currently undergoing tone merger in some native speakers of Hong Kong Cantonese, a separate data analysis was conducted excluding these tone pairs. Another separate data analysis was conducted excluding the syllable /f&#x0250;n/, the only syllable in the stimulus set (/f&#x0250;n/, /j&#x0250;u/, and /w&#x0250;i/) that does not carry F0 at the syllable onset. Both sets of additional data analyses yield patterns that are qualitatively identical to those reported in the paper.</p>
</fn>
<fn id="fn0006">
<p><sup>3</sup>Additional Bayesian analyses: For the four sets of analyses reported above, the mixed-effects models revealed significant interactions between <italic>group</italic> and <italic>ear</italic> on identification and discrimination RT, but not on identification and discrimination accuracy. To further test the null effects (H0), we conducted Bayesian two-way ANOVA (group by ear) with JASP (JASP Team, 2019). The dependent variable was identification and discrimination accuracy, respectively. For the identification accuracy, Bayesian analysis revealed a BF<sub>01</sub> value of 25.381, which means that the data were approximately 25 times more likely to occur under the H0 (null hypothesis) than under the H1 (the alternative hypothesis). The error percentage was 1.397%, which reflects the stability of the numerical algorithm that was used to obtain the results. As for the discrimination accuracy, BF<sub>01</sub> value was 5.005, meaning that the data were approximately 5 times more likely to occur under the H0 than under the H1, with an error percentage of 1.8%. Altogether, these Bayes factors provided additional support for the possibility that the absence of interaction effects in the analysis on identification and discrimination accuracy might not be due to insufficient power.</p>
</fn>
</fn-group>
</back>
</article>