<?xml version="1.0" encoding="UTF-8" standalone="no"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xml:lang="EN" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Lang. Sci.</journal-id>
<journal-title>Frontiers in Language Sciences</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Lang. Sci.</abbrev-journal-title>
<issn pub-type="epub">2813-4605</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/flang.2024.1243678</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Language Sciences</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Exploring effects of brief daily exposure to unfamiliar accent on listening performance and cognitive load</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name><surname>McLaughlin</surname> <given-names>Drew J.</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="aff" rid="aff2"><sup>2</sup></xref>
<xref ref-type="corresp" rid="c001"><sup>&#x0002A;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/2038053/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Baese-Berk</surname> <given-names>Melissa M.</given-names></name>
<xref ref-type="aff" rid="aff3"><sup>3</sup></xref>
<xref ref-type="aff" rid="aff4"><sup>4</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/96929/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Van Engen</surname> <given-names>Kristin J.</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/115374/overview"/>
</contrib>
</contrib-group>
<aff id="aff1"><sup>1</sup><institution>Department of Psychological and Brain Sciences, Washington University in St. Louis</institution>, <addr-line>St. Louis, MO</addr-line>, <country>United States</country></aff>
<aff id="aff2"><sup>2</sup><institution>Basque Center on Cognition, Brain and Language</institution>, <addr-line>San Sebastian</addr-line>, <country>Spain</country></aff>
<aff id="aff3"><sup>3</sup><institution>Department of Linguistics, University of Oregon</institution>, <addr-line>Eugene, OR</addr-line>, <country>United States</country></aff>
<aff id="aff4"><sup>4</sup><institution>Department of Linguistics, University of Chicago</institution>, <addr-line>Chicago, IL</addr-line>, <country>United States</country></aff>
<author-notes>
<fn fn-type="edited-by"><p>Edited by: Juhani J&#x000E4;rvikivi, University of Alberta, Canada</p></fn>
<fn fn-type="edited-by"><p>Reviewed by: Melissa Paquette-Smith, University of California, Los Angeles, United States</p>
<p>Xin Xie, University of California, Irvine, United States</p></fn>
<corresp id="c001">&#x0002A;Correspondence: Drew J. McLaughlin <email>d.mclaughlin&#x00040;bcbl.eu</email></corresp>
</author-notes>
<pub-date pub-type="epub">
<day>27</day>
<month>05</month>
<year>2024</year>
</pub-date>
<pub-date pub-type="collection">
<year>2024</year>
</pub-date>
<volume>3</volume>
<elocation-id>1243678</elocation-id>
<history>
<date date-type="received">
<day>21</day>
<month>06</month>
<year>2023</year>
</date>
<date date-type="accepted">
<day>06</day>
<month>05</month>
<year>2024</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x000A9; 2024 McLaughlin, Baese-Berk and Van Engen.</copyright-statement>
<copyright-year>2024</copyright-year>
<copyright-holder>McLaughlin, Baese-Berk and Van Engen</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/"><p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p></license>
</permissions>
<abstract>
<sec>
<title>Introduction</title>
<p>Listeners rapidly &#x0201C;tune&#x0201D; to unfamiliar accented speech, and some evidence also suggests that they may improve over multiple days of exposure. The present study aimed to measure accommodation of unfamiliar second language- (L2-) accented speech over a consecutive 5-day period using both a measure of listening performance (speech recognition accuracy) and a measure of cognitive load (a dual-task paradigm).</p></sec>
<sec>
<title>Methods</title>
<p>All subjects completed a dual-task paradigm with L1 and L2 accent on Days 1 and 5, and were given brief exposure to either L1 (control group) or unfamiliar L2 (training groups) accent on Days 2&#x02013;4. One training group was exposed to the L2 accent via a standard speech transcription task while the other was exposed to the L2 accent via a transcription task that included implicit feedback (i.e., showing the correct answer after each trial).</p></sec>
<sec>
<title>Results</title>
<p>Although overall improvement in listening performance and reduction in cognitive load were observed from Days 1 to 5, our results indicated neither a larger benefit for the L2 accent training groups compared to the control group nor a difference based on the implicit feedback manipulation.</p></sec>
<sec>
<title>Discussion</title>
<p>We conclude that the L2 accent trainings implemented in the present study did not successfully promote long-term learning benefits of a statistically meaningful magnitude, presenting our findings as a methodologically informative starting point for future research on this topic.</p></sec></abstract>
<kwd-group>
<kwd>speech perception</kwd>
<kwd>accent</kwd>
<kwd>perceptual training</kwd>
<kwd>listening effort</kwd>
<kwd>cognitive load</kwd>
</kwd-group>
<contract-num rid="cn001">2146993</contract-num>
<contract-num rid="cn001">BCS-2020805</contract-num>
<contract-num rid="cn001">DGE-1745038</contract-num>
<contract-num rid="cn002">BERC 2022-2025</contract-num>
<contract-num rid="cn003">CEX2020-001010-S</contract-num>
<contract-num rid="cn004">101103964</contract-num>
<contract-sponsor id="cn001">National Science Foundation<named-content content-type="fundref-id">10.13039/100000001</named-content></contract-sponsor>
<contract-sponsor id="cn002">Eusko Jaurlaritza<named-content content-type="fundref-id">10.13039/501100003086</named-content></contract-sponsor>
<contract-sponsor id="cn003">Fundaci&#x000F3;n Carmen y Severo Ochoa<named-content content-type="fundref-id">10.13039/100017199</named-content></contract-sponsor>
<contract-sponsor id="cn004">Horizon 2020<named-content content-type="fundref-id">10.13039/501100007601</named-content></contract-sponsor>
<counts>
<fig-count count="3"/>
<table-count count="3"/>
<equation-count count="0"/>
<ref-count count="34"/>
<page-count count="12"/>
<word-count count="9406"/>
</counts>
<custom-meta-wrap>
<custom-meta>
<meta-name>section-at-acceptance</meta-name>
<meta-value>Bilingualism</meta-value>
</custom-meta>
</custom-meta-wrap>
</article-meta>
</front>
<body>
<sec sec-type="intro" id="s1">
<title>Introduction</title>
<p>Evidence suggests that listeners can rapidly &#x0201C;tune&#x0201D; to unfamiliar accented speech, thereby improving their ability to understand a given speaker over time. For second language- (L2-) accented speech, improvements to listening performance (often measured with transcription/repetition accuracy, or &#x0201C;intelligibility&#x0201D;) can be facilitated by exposure to a single accented speaker, to multiple speakers with the same accent, or even to a variety of speakers with different accents (Bradlow and Bent, <xref ref-type="bibr" rid="B6">2008</xref>; Sidaras et al., <xref ref-type="bibr" rid="B29">2009</xref>; Baese-Berk et al., <xref ref-type="bibr" rid="B2">2013</xref>). Similarly, the cognitive demands of speech processing have been shown to rapidly decrease following exposure to L2-accented speech in single-session experiments (Clarke and Garrett, <xref ref-type="bibr" rid="B10">2004</xref>; Brown et al., <xref ref-type="bibr" rid="B7">2020</xref>). Based on correlational evidence, it also appears that the efficiency and accuracy of L2 accent processing depends on a listener&#x00027;s prior (real world) experience: More experienced listeners typically process L2 accent faster and more accurately (Kennedy and Trofimovich, <xref ref-type="bibr" rid="B15">2008</xref>; Porretta et al., <xref ref-type="bibr" rid="B26">2020</xref>). However, empirical evidence connecting these two literatures is lacking, with few studies that have examined perceptual accommodation of L2 accent across multiple days (or weeks, etc.). In the present study, we take a first step toward filling this empirical gap. Across five consecutive daily sessions, we sought to document changes in listening performance and cognitive load for (previously unfamiliar) L2 accent.</p>
<sec>
<title>Accent experience: a theoretical framework</title>
<p>Because spoken language varies from talker to talker, due to both idiosyncratic differences and accent, listeners have to be adaptable when mapping complex acoustic input onto linguistic representations in the mental lexicon (Bent and Baese-Berk, <xref ref-type="bibr" rid="B4">2021</xref>). Changes to representations (and/or decision processes, see Xie et al., <xref ref-type="bibr" rid="B33">2023</xref><xref ref-type="fn" rid="fn0001"><sup>1</sup></xref>) based on listeners&#x00027; global and recent exposure are supported by multiple leading language models, including exemplar (Johnson, <xref ref-type="bibr" rid="B14">1997</xref>; Pierrehumbert, <xref ref-type="bibr" rid="B25">2001</xref>), non-analytic episodes (Goldinger, <xref ref-type="bibr" rid="B13">1998</xref>; Nygaard and Pisoni, <xref ref-type="bibr" rid="B24">1998</xref>), and Bayesian inference (i.e., &#x0201C;the ideal adaptor&#x0201D;; Kleinschmidt and Jaeger, <xref ref-type="bibr" rid="B16">2015</xref>) models. Under these frameworks, it is posited that listeners create categories systematically linking social groupings and phonetic patterns, including accent-specific representations. On this view, listeners&#x00027; prior experience with a given accent ought to determine their ability to efficiently and accurately process speech produced with that accent. In the same way that processing speech produced by a familiar talker is faster (Newman and Evers, <xref ref-type="bibr" rid="B22">2007</xref>; Magnuson et al., <xref ref-type="bibr" rid="B20">2021</xref>), processing speech produced in a familiar accent ought to be faster.</p>
<p>Correlational evidence aligns with the supposition that the ability to process an L2 accent accurately and efficiently can be developed over time with sufficient (real world) exposure. For example, Kennedy and Trofimovich (<xref ref-type="bibr" rid="B15">2008</xref>) found that both semantically meaningful and semantically anomalous Mandarin Chinese-accented sentences were transcribed with higher accuracy by L1 English listeners who had greater experience with Mandarin Chinese accent. Psychophysiological evidence from pupillometry and eye-tracking has also indicated a benefit of experience: Task-evoked pupil response in Porretta and Tucker (<xref ref-type="bibr" rid="B27">2019</xref>) indicated that L1 English listeners who had more experience with Mandarin Chinese accent processed Mandarin Chinese-accented English words (presented in noise) more easily, and gaze behavior in Porretta et al. (<xref ref-type="bibr" rid="B26">2020</xref>) demonstrated that greater experience with Mandarin Chinese accent also resulted in faster speech processing for Mandarin Chinese-accented English. Further behavioral evidence from L1 Dutch listeners also suggests that activation of L2-accented words depends on listener experience with the target accent. Witteman et al. (<xref ref-type="bibr" rid="B32">2013</xref>) used a cross-modal lexical decision task in which each trial participants were presented auditorily with a German-accented word or non-word in Dutch followed by an orthographic probe word or non-word in Dutch. The listeners&#x00027; task was to make lexical decisions for the visual probe items. Results indicated that listeners with less experience with German accent were less primed by auditory items that were strongly accented than participants with greater experience.</p>
<p>Altogether, these studies suggest a critical role of prior experience in L2 accent processing, aligning with predictions from episodic models of speech processing. What remains to be empirically determined, however, is the amount and rate of exposure that is necessary to observe a benefit of prior experience. Under all of the theoretical frameworks mentioned above (exemplar, non-analytic episodes, and Bayesian inference), listeners with more experience with a given accent ought to be more adept at processing speech produced with that accent. The present study aims to test these theoretical models of speech processing.</p>
</sec>
<sec>
<title>Accent experience: empirical evidence</title>
<p>Only a small number of studies have investigated the benefits of prolonged exposure to L2 accent using causal experimental design. Within a single experimental session, rapid improvements to listening performance and reductions of listening effort are well-documented (Clarke and Garrett, <xref ref-type="bibr" rid="B10">2004</xref>; Bradlow and Bent, <xref ref-type="bibr" rid="B6">2008</xref>; Sidaras et al., <xref ref-type="bibr" rid="B29">2009</xref>; Baese-Berk et al., <xref ref-type="bibr" rid="B2">2013</xref>; Brown et al., <xref ref-type="bibr" rid="B7">2020</xref>). Beyond a single experimental session, however, investigations of sustained benefits of L2 accent exposure are limited, and those that exist have produced mixed results.</p>
<p>Examining both transcription accuracy and comprehension of Korean-accented English materials, Lindemann et al. (<xref ref-type="bibr" rid="B18">2016</xref>) found a sustained benefit of L2 accent exposure at 1&#x02013;2 days post-training. The authors presented L1 English listeners with either L2 accent exposure or an explicit linguistic training (i.e., teaching phonemic differences between Korean and English, etc.). Notable for the present study, participants in the L2 accent exposure group completed a speech transcription task in which the correct sentence was presented each trial after submitting the typed response. At test, both training groups had better sentence transcription accuracy than a control group; comprehension scores were the same across groups. Thus, the results of Lindemann et al. demonstrate that a (brief) training session with L2 accent can lead to improvements in perceptual accuracy that last into subsequent day(s).</p>
<p>Sustained adaptation to L2 accent over a half-day (12-h) period&#x02014;as well as generalization&#x02014;was also demonstrated in a sleep consolidation study conducted by Xie et al. (<xref ref-type="bibr" rid="B34">2017</xref>). L1 English listeners in the study were trained with word-length stimuli from a Mandarin Chinese-accented talker, focusing on a key accented phoneme (/d/). All subjects completed a test with a novel Mandarin Chinese-accented talker immediately after training as well as a second iteration of this test 12-h later, but for half of the subjects this 12-h period spanned the day (e.g., 8 a.m. to 8 p.m.) and for the other half is spanned the night (e.g., 8 p.m. to 8 a.m.). In both groups, retention of the training benefit was observed. Critically, however, the overnight group showed unique generalization of learning to another Mandarin Chinese-accented speaker and phonemic category (/t/), suggesting that sleep consolidation promoted generalization of learning.</p>
<p>Whether benefits of L2 accent exposure are retained over intervals longer than 1 day, however, remains unclear. Bieber and Gordon-Salant (<xref ref-type="bibr" rid="B5">2021</xref>), for example, failed to find evidence of a training benefit in test sessions administered 1 week after training. In their study, the authors examined accent-generalizable learning (i.e., performance on a novel/untrained accent following training with multiple other accents) for speech presented in six-talker babble. They employed a dual-task paradigm (similar to the one used in the present study), which combines a speech transcription task with a simultaneous reaction time-based visual task (it is assumed that with finite cognitive resources, reaction times will slow for the secondary visual task as the demands of the primary speech task increase). L1 English-speaking young adults and older adults with and without hearing loss completed three experimental sessions across approximately 3 weeks, where the beginning portion of Week 2&#x00027;s session and Week 3&#x00027;s session each served as measures of retention. Results indicated that within an experimental session listeners rapidly improved, as in prior work (Bradlow and Bent, <xref ref-type="bibr" rid="B6">2008</xref>; Baese-Berk et al., <xref ref-type="bibr" rid="B2">2013</xref>). However, the benefit of each prior week&#x00027;s training session on transcription accuracy was not retained (i.e., the beginning of the Week 2 and 3 sessions did not demonstrate improvement). Reaction times from the secondary measure, on the other hand, were significantly improved for Week 3&#x00027;s session, in particular, indicating that the cognitive load associated with L2 accent processing may have been reduced.<xref ref-type="fn" rid="fn0002"><sup>2</sup></xref></p>
<p>Focusing on both listening performance and attitudes toward L2 speakers, Derwing et al. (<xref ref-type="bibr" rid="B11">2002</xref>) implemented what appears to be the longest L2 accent training protocol to date, occurring over an 8-week period. The authors sought to train L1 English listeners to better understand L2 (specifically, Vietnamese) accent, comparing the effects of a training with explicit phonetic lectures and a training with cross-cultural awareness lectures. Unfortunately, results of the study indicated no significant benefits of either training for speech transcription or comprehension. Attitude questionnaires, however, did reveal that both training groups showed increased empathy toward immigrants, and participants given explicit phonetic training reported increased confidence in their ability to understand L2 accent.</p>
<p>Altogether, the current body of empirical evidence suggests that benefits of L2 accent training sessions may persist into subsequent days but diminish over longer (week-long) intervals. Additionally, benefits observed for cognitive load may diverge from those observed for listening performance (i.e., recognition/transcription accuracy). Based on these observations, in the present study we sought to examine the benefits of a training protocol administered over multiple consecutive days. From Pre-Test to Post-Test, we also incorporated a measure of cognitive load (similar to Bieber and Gordon-Salant, <xref ref-type="bibr" rid="B5">2021</xref>) to determine whether different benefits may be observed for measures of listening performance vs. cognitive load.</p>
</sec>
<sec>
<title>The present study</title>
<p>In the present study, we implemented a combination of dual-task paradigms and a speech transcription tasks over a 5-day period. On Days 1 and 5, participants completed a Pre-Test and Post-Test (dual-task paradigm), and on Days 2, 3, and 4 participants completed exposure-based training sessions (speech transcription). Participants were randomly assigned to one of three groups for the training days: Control (no exposure to L2 accent or feedback during training), Exposure (exposure to L2 accent but no feedback during training), and Feedback (exposure to L2 accent and feedback during training). All groups had the exact same Pre- and Post-Test with both L1- and L2-accented speech stimuli.</p>
<p>We predicted that response times to the secondary task in the dual-task paradigm (an index of cognitive load) would be shorter on Day 5 than Day 1 for all groups, indicating improvement on the task itself. Critically, we expected that this improvement would be greater for the Exposure and Feedback training groups&#x02014;particularly in the L2-accented speech condition&#x02014;than it would be for the Control group. Additionally, we predicted that the Feedback group would show greater reduction in cognitive load than the Exposure group, given that the feedback manipulation provided lexical context to guide perceptual adaptation.</p>
<p>For listening performance (speech recognition accuracy) in the Pre-Test and Post-Test data, we had similar predictions, although we also anticipated the possibility that subjects may demonstrate reduced cognitive load without gains in listening performance (as in Bieber and Gordon-Salant, <xref ref-type="bibr" rid="B5">2021</xref>). We expected that speech recognition scores from the primary task would be larger on Day 5 than Day 1 for all groups (indicating improvement on the task itself), and that the Exposure and Feedback training groups would improve more than the Control group in the L2-accented speech condition, in particular. We also predicted that the Feedback group would show greater improvement than the Exposure group.</p>
<p>Lastly, we planned to examine listening performance (speech transcription<xref ref-type="fn" rid="fn0003"><sup>3</sup></xref> accuracy) data from the training sessions on Days 2, 3, and 4. We predicted that, if any differences existed, they would be as follows: Higher scores for the Feedback group than the Exposure group overall, and an interaction with days reflecting greater improvement for the Feedback group over time.</p></sec>
</sec>
<sec sec-type="methods" id="s2">
<title>Methods</title>
<p>The current study was approved by Washington University&#x00027;s Institutional Review Board. Due to the COVID-19 pandemic, the final version of the study deviated substantially from the pre-registered version (full details can be found in the <xref ref-type="supplementary-material" rid="SM1">Supplementary material</xref>).</p>
<sec>
<title>Participants</title>
<p>Young adult subjects (age mean = 19.5; age range = 18&#x02013;30) were recruited from Washington University in St. Louis&#x00027;s Psychology Participants Pool. Inclusion criteria (set via demographic filters in SONA Systems) selected for L1 English speakers with normal hearing and vision (or corrected-to-normal vision). Additional criteria on the SONA listing indicated that subjects should not sign up for the study if they had extensive exposure to Mandarin Chinese (for example, they should not speak Mandarin Chinese, have studied Mandarin Chinese, or have parents or roommates who are fluent in Mandarin Chinese). Subjects who did not complete all 5 days of the study were excluded from analyses.</p>
<p>Due to COVID-19-related recruitment issues, we decided to combine a pilot version of the experiment with the main dataset to reach more desirable sample size (<italic>N</italic> = 160). We report full details regarding the minor differences between the pilot and primary subject groups below. In brief, the two differences were: (1) During the dual-task sessions for the primary group but not the pilot group, the practice session provided feedback instructing participants to &#x0201C;speed up&#x0201D; if they took longer than 3 s to respond; and, (2) Additional measures of cognitive ability (not analyzed in the present manuscript) were not collected from the pilot participants. Results of all analyses remained the same when accounting for time of participation (i.e., when including a two-level fixed effect denoting &#x0201C;pilot&#x0201D; vs. &#x0201C;main&#x0201D; experiment status); because time of participation did not improve model fits or impact the outcomes for the effects of interest, this factor was dropped from all models.</p>
<p>After combining the two datasets, the sample size by group was as follows: Control <italic>n</italic> = 54, Exposure <italic>n</italic> = 53, Feedback <italic>n</italic> = 53. In the pilot version of the experiment, a total of 43 subjects participated. Two subjects were excluded from this sample for failing to complete all days of the study, one for reporting exposure to Chinese, and three for having an average reaction time in the dual-task paradigm &#x0003E; 3,000 ms (the significance of this cut-off is discussed further in the Procedures section). After exclusions, 37 valid subjects remained (by group: Control <italic>n</italic> = 11, Exposure <italic>n</italic> = 14, Feedback <italic>n</italic> = 12). For the main experiment, 152 subjects participated in total. Twenty-nine of these subjects were excluded for one the following reasons: Failing to complete all 5 days of the study (24), self-reporting too much prior exposure to Mandarin Chinese (four), and, in one case, self-reporting a receptive and productive language disorder. After exclusions, 123 valid subjects remained (by group: Control <italic>n</italic> = 43, Exposure <italic>n</italic> = 39, Feedback <italic>n</italic> = 41).</p>
<p>We report information about participants&#x00027; language experience by random assignment group in <xref ref-type="table" rid="T1">Table 1</xref>. All participants reported English as their primary language, and as a language learned from birth. All participants reported at least one additional language (which is to be expected given high school language requirements in the U.S.). As can be seen in <xref ref-type="table" rid="T1">Table 1</xref>, a fairly large proportion of the sample can be classified as simultaneous (21%) or early (8%) bilinguals; these trends are unsurprising when considering current estimates of bilingualism in the United States (&#x0007E;1 in 5 speak a language other than English at home; Dietrich and Hernandez, <xref ref-type="bibr" rid="B12">2022</xref>). Including bilingual status as an effect in the response time and accuracy analyses did not change the results or improve model fits, and was thus not included as a factor in the final models.</p>
<table-wrap position="float" id="T1">
<label>Table 1</label>
<caption><p>Summary of participants&#x00027; language experiences.</p></caption>
<table frame="box" rules="all">
<thead>
<tr style="background-color:#919498;color:#ffffff">
<th/>
<th/>
<th/>
<th valign="top" align="center" colspan="3"><bold>Proportion of participants with one or more L2(s) acquired from:</bold></th>
</tr>
<tr style="background-color:#919498;color:#ffffff">
<th/>
<th valign="top" align="center"><bold>Count of participants</bold></th>
<th valign="top" align="center"><bold>Count of languages</bold></th>
<th valign="top" align="center"><bold>Birth</bold></th>
<th valign="top" align="center"><bold> &#x02264; age 5</bold></th>
<th valign="top" align="center"><bold> &#x02264; age 10</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left"><bold>All groups</bold></td>
<td valign="top" align="center"><bold>160</bold></td>
<td valign="top" align="center"><bold>2.66 (0.69)</bold></td>
<td valign="top" align="center"><bold>0.21</bold></td>
<td valign="top" align="center"><bold>0.29</bold></td>
<td valign="top" align="center"><bold>0.54</bold></td>
</tr> <tr>
<td valign="top" align="left">Control</td>
<td valign="top" align="center">54</td>
<td valign="top" align="center">2.65 (0.69)</td>
<td valign="top" align="center">0.22</td>
<td valign="top" align="center">0.31</td>
<td valign="top" align="center">0.57</td>
</tr> <tr>
<td valign="top" align="left">Exposure</td>
<td valign="top" align="center">53</td>
<td valign="top" align="center">2.67 (0.71)</td>
<td valign="top" align="center">0.17</td>
<td valign="top" align="center">0.21</td>
<td valign="top" align="center">0.56</td>
</tr> <tr>
<td valign="top" align="left">Feedback</td>
<td valign="top" align="center">53</td>
<td valign="top" align="center">2.66 (0.69)</td>
<td valign="top" align="center">0.26</td>
<td valign="top" align="center">0.36</td>
<td valign="top" align="center">0.48</td>
</tr></tbody>
</table>
<table-wrap-foot>
<p>Standard deviations shown in parentheses. The bold value indicate that its summary of the below rows.</p>
</table-wrap-foot>
</table-wrap>
</sec>
<sec>
<title>Materials</title>
<sec>
<title>Auditory stimuli</title>
<p>Semantically anomalous sentences from the Semantically Normal Sentence Test (SNST; Nye and Gaitenby, <xref ref-type="bibr" rid="B23">1974</xref>) were adapted for use in the present experiment. The SNST includes items with four keywords each (defined as any adjectives, verbs, or nouns) such as &#x0201C;the <italic>wrong shot led</italic> the <italic>farm</italic>.&#x0201D; This sentence set was selected with the aim of examining perception of L2 accent in quiet listening conditions while preventing ceiling effects for transcription accuracy. The original SNST set contains 200 items, and for the present study we created an additional 110 items. This provided enough unique items to avoid repeating auditory stimuli at any point during the study (i.e., more than 297 unique items total).</p>
<p>Recordings of these sentences were created in a sound-reduction booth using MOTU UltraLite-mk3 Hybrid microphone hardware and Audacity (version 2.4.2) run on iMac (version 10.15.7). For the L2 accent condition, we selected Mandarin Chinese-accented English. Six young adult, female speakers were recorded reading all of the semantically-anomalous items. To select three speakers for the present study, we piloted the stimuli with 253 participants, for a total of &#x0007E;10 transcriptions per item (i.e., each participant listened to only 90 items). Across all items and responses, transcription performance for three of the speakers was fairly well-matched and met the experiment needs. These speakers were estimated to be 51.8, 53.1, and 55.8% intelligible. All talkers were proficient English speakers who began learning English in China as children (at ages 11, 8, and 5, respectively) but had only been in the United States for &#x0007E;1 year of graduate-level studies.</p>
<p>For the L1-accented condition, three female L1 speakers of English from the Midwestern United States were recorded. Given that timing of responses in the dual-task paradigm was of critical interest, we decided to match speaking rate across the L1 and L2 speakers. Toward this goal, the L1 speakers were instructed to produce items at a typical speed, a slightly slower than normal speed, and a slower than normal speed. Items were selected based on their total duration in order to match the speaking rate across the L2 and L1 speakers. The final average stimuli length for the L2 speakers was 1,790 ms, and the final average stimuli length for the L1 speakers was 1,784 ms.</p></sec>
<sec>
<title>Questionnaire</title>
<p>Participants completed a questionnaire on Days 2, 3, and 4 of the study that assessed motivation (composite of three questions), self-perceived performance, and effort. The questions for the motivation composite score included the following: (1) How motivated were you to perform well-during the listening task? (1 = very unmotivated, 7 = very motivated); (2) How much did you like performing trials in the listening task? (1 = strongly dislike, 7 = strongly like); (3) How much did you desire to challenge yourself during the listening task? (1 = not at all, 7 = very much). The self-perceived performance question (&#x0201C;Please rate your performance on the listening task&#x0201D;) included a scale from &#x0201C;1 = absolute worst&#x0201D; to &#x0201C;7 = absolute best&#x0201D;, and the effort question (&#x0201C;How effortful did you find the listening task?&#x0201D;) included a scale from &#x0201C;1 = not at all effortful&#x0201D; to &#x0201C;7 = very effortful.&#x0201D;</p>
</sec>
</sec>
<sec>
<title>Procedures</title>
<sec>
<title>Overview</title>
<p>Subjects completed five experimental sessions lasting 30 min each over the course of five consecutive days (typically Monday through Friday). On Days 1 and 5, the primary speech perception task involved a dual-task paradigm, and on Days 2, 3, and 4, the primary speech perception task involved self-paced speech transcription only. The speech perception tasks were administered on a 21.5 inch iMac (version 10.15.7, &#x0201C;Catalina&#x0201D;) and programmed with SuperLab (Cedrus, version 5). Audio was presented via circumaural Beyerdynamic DT 100 headphones.</p>
<p>Additional measures occurred after the primary experimental tasks on specific days of the week as follows: On Day 1, participants completed a demographic and language background questionnaire; on Day 2, they completed the Trail-Making Task (Arbuthnott and Frank, <xref ref-type="bibr" rid="B1">2000</xref>); on Day 3, they completed a Stroop task (MacLeod, <xref ref-type="bibr" rid="B19">1991</xref>); on Day 4, they completed the Word Auditory Recognition and Recall Measure (WARRM; a measure of working memory capacity; Smith et al., <xref ref-type="bibr" rid="B30">2016</xref>); and for Days 2, 3, and 4, they completed a questionnaire each day to assess their motivation, self-perceived performance, and effort (method and results reported in <xref ref-type="supplementary-material" rid="SM1">Supplementary material</xref>). Note that in the pilot version of the experiment, the Trail-Making, Stroop, and WARRM tasks were not included. We do not report on these individual difference measures in the present study.</p></sec>
<sec>
<title>Dual-task paradigm (Days 1 and 5)</title>
<p>The dual-task paradigm included a speech perception primary task and a non-linguistic visual categorization secondary task (used previously in Strand et al., <xref ref-type="bibr" rid="B31">2018</xref>; Brown et al., <xref ref-type="bibr" rid="B7">2020</xref>). Participants were instructed that they would be completing both tasks simultaneously, but that the speech perception task was the primary and more important task.</p>
<p>Each trial, subjects were presented with a single auditory sentence. Their goal was to repeat the sentence at the end of the trial as accurately as possible. At the onset of the soundfile, two empty squares appeared on the screen. After an interstimulus interval (ISI) of 600&#x02013;800 ms (in 100 ms intervals), a number between 1 and 8 appeared in either the left or the right box. Using a button box, participants were instructed to make either a left response or a right response depending on the following: If an odd number appeared (1, 3, 5, or 7), they were supposed to press the button on the opposite side as the box on the screen; if, however, an even number appears (2, 4, 6, or 8) they were supposed to press the button on the same side as the box on the screen. For example, the correct response for a 1 appearing in the left box on the screen was pressing the right button, and the correct response for a 2 appearing in the left box on the screen was pressing the left button. Participants were instructed to respond as quickly as possible while prioritizing accuracy. The timing of the ISI ensured that the demands of the secondary task occurred approximately midway through the presentation of the target sentences for the primary task. Thus, trials in which the demands of the primary task were greater should result in longer response times to the secondary task.</p>
<p>For the primary task, participants repeated the target sentence aloud after both their key press was made for the secondary task and the auditory stimulus was completed for the primary task. Verbal responses were recorded and scored for accuracy offline. Between trials, an ISI of 5,000, 5,500, or 6,000 ms occurred before automatic presentation of the next trial.</p>
<p>The combination of items for the primary and secondary tasks was randomized across participants. For the secondary task, the occurrence of each number at each of the two locations occurred randomly. For the primary task, auditory files were presented in a random order within a list used for practice trials (12 total) and a list used for regular trials (78 total). An equal number of trials for each accent condition and each speaker were included. For the regular trials, this resulted in 39 trials per accent, and 13 trials per speaker. Four counterbalanced orders were used to rotate which target sentences appeared on Day 1 vs. Day 5, and whether these targets were presented in the L1 vs. the L2 accent condition on a given day.</p>
<p>During the practice trials, a researcher remained in the room to observe the participant and confirm they were making responses in the correct order (i.e., button press and then verbal repetition). After the pilot version of the experiment was complete, we decided to add feedback to the practice session. If subjects took longer than 3,000 ms to respond with a button press after presentation of the number target, &#x0201C;Too slow!&#x0201D; appeared onscreen. Data from practice trials was excluded from analyses.</p>
<p>A 72-trial block of the secondary task (i.e., presented in isolation) was completed after the critical dual-task session was complete. In this block, subjects were only presented with numbers to sort, and no auditory input. Pilot subjects were not presented with this block, and, thus, we do not report on this data in the present paper.</p></sec>
<sec>
<title>Speech transcription task (Days 2, 3, and 4)</title>
<p>A speech transcription task was administered on training days (Days 2, 3, and 4) instead of the dual-task paradigm. For the training sessions, subjects were randomly assigned to one of three conditions: Control, Exposure, and Feedback. Subjects assigned to the Control group heard only the three L1-accented talkers on training days, while those assigned to the Exposure and Feedback groups heard only the three L2-accented talkers on training days. The key difference between the Exposure and Feedback groups was that subjects in the Feedback group were shown the correct target sentence after submitting their transcription each trial (thus providing implicit feedback on performance).</p>
<p>Participants completed 39 trials each session (13 trials per talker) presented in a randomized order. None of the target sentences repeated across training sessions or overlapped with target sentences from the dual-task sessions. Transcriptions were completed with a keyboard and self-paced. Subjects were instructed to do their best to spell accurately. After entering their transcriptions, participants were shown either a series of eight hashtags (Control and Exposure groups) or the target sentence (Feedback group) for 5,000 ms. An inter-stimulus interval of 3,000 ms occurred before presentation of the next trial.</p></sec></sec>
</sec>
<sec id="s3">
<title>Analysis</title>
<sec>
<title>Model specifications: recognition accuracy data</title>
<p>Generalized linear mixed-effects regression was used to model the recognition accuracy data in R (version 4.0.4; R Core Team, <xref ref-type="bibr" rid="B28">2021</xref>) with the <italic>glmer()</italic> function from the <italic>lme4</italic> package (Bates et al., <xref ref-type="bibr" rid="B3">2015</xref>). Likelihood ratio tests were conducted to determine the significance of effects of interest, and <italic>p-</italic>values for model parameters were estimated using the <italic>lmerTest</italic> package (Kuznetsova et al., <xref ref-type="bibr" rid="B17">2017</xref>). Recognition accuracy was treated as a grouped binomial, meaning that models predicted performance using two columns of data (number of correct words, number of incorrect/missed words) for each sentence. A logit link function was specified. Fixed effects included: Condition (dummy-coded levels: L1 accent, L2 accent), Session (dummy-coded levels: Pre-Test, Post-Test), Group (dummy-coded levels: Control, Exposure, Feedback), as well as all possible two- and three-way interactions between Condition, Session, and Group. Random intercepts were included by item and by subject. Random slopes of Day and Condition were attempted but ultimately removed from all models due to issues with model singularity. Model syntax is provided in <xref ref-type="supplementary-material" rid="SM1">Supplemental materials</xref>.</p>
</sec>
<sec>
<title>Model specifications: response time data</title>
<p>For the response time data, linear mixed-effects regression was implemented with the <italic>lmer()</italic> function. Fixed effects included: Condition (dummy-coded levels: L1 accent, L2 accent), Session (dummy-coded levels: Pre-Test, Post-Test), Group (dummy-coded levels: Control, Exposure, Feedback), and all two- and three-way interactions between Condition, Session, and Group. In all models, random effects included random intercepts by subject and by item, and random slopes of Condition and Session by subject. Model syntax is provided in <xref ref-type="supplementary-material" rid="SM1">Supplemental materials</xref>.</p></sec>
</sec>
<sec sec-type="results" id="s4">
<title>Results</title>
<sec>
<title>Pre-Test and Post-Test (dual-task paradigm) data</title>
<sec>
<title>Recognition accuracy data from speech perception task</title>
<p>Recognition accuracy data from the dual-task paradigm is presented in <xref ref-type="fig" rid="F1">Figure 1</xref>. We report all log-likelihood model comparisons in <xref ref-type="table" rid="T2">Table 2</xref> and provide full model summaries in <xref ref-type="supplementary-material" rid="SM1">Supplemental materials</xref>. In brief, results indicated improved accuracy from Pre-Test to Post-Test (&#x000DF; = 0.13, <italic>p</italic> &#x0003C; 0.001), but this improvement was similar for all participant groups (non-significant three-way interaction of Condition, Session, and Group: &#x003C7;<sup>2</sup> = 4.56, <italic>p</italic> = 0.10).</p>
<fig id="F1" position="float">
<label>Figure 1</label>
<caption><p>Recognition accuracy data from the dual-task paradigm at Pre-Test and Post-Test, for each group and accent condition, is presented with violin density distributions, mean points, and standard error bars.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="flang-03-1243678-g0001.tif"/>
</fig>
<table-wrap position="float" id="T2">
<label>Table 2</label>
<caption><p>Log-likelihood model comparisons from analyses of dual-task recognition accuracy data.</p></caption>
<table frame="box" rules="all">
<thead>
<tr style="background-color:#919498;color:#ffffff">
<th valign="top" align="left"><bold>Effect</bold></th>
<th valign="top" align="center"><bold>&#x003C7;<sup>2</sup></bold></th>
<th valign="top" align="center"><bold>df</bold></th>
<th valign="top" align="center"><bold><italic>p</italic></bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Condition</td>
<td valign="top" align="center">16,709</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">&#x0003C; 0.001</td>
</tr> <tr>
<td valign="top" align="left">Session</td>
<td valign="top" align="center">51.91</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">&#x0003C; 0.001</td>
</tr> <tr>
<td valign="top" align="left">Group</td>
<td valign="top" align="center">1.20</td>
<td valign="top" align="center">2</td>
<td valign="top" align="center">0.55</td>
</tr> <tr>
<td valign="top" align="left">Condition: session</td>
<td valign="top" align="center">1.52</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">0.22</td>
</tr> <tr>
<td valign="top" align="left">Condition: group</td>
<td valign="top" align="center">8.52</td>
<td valign="top" align="center">2</td>
<td valign="top" align="center">0.01</td>
</tr> <tr>
<td valign="top" align="left">Session: group</td>
<td valign="top" align="center">4.02</td>
<td valign="top" align="center">2</td>
<td valign="top" align="center">0.13</td>
</tr> <tr>
<td valign="top" align="left">Condition: session: group</td>
<td valign="top" align="center">4.56</td>
<td valign="top" align="center">2</td>
<td valign="top" align="center">0.10</td>
</tr></tbody>
</table>
<table-wrap-foot>
<p><italic>df</italic> , degrees of freedom.</p>
</table-wrap-foot>
</table-wrap>
<p>As expected, overall performance for the L2 accent was significantly poorer than performance for the L1 accent (&#x000DF; = &#x02212;2.44). The fixed effect of Group indicated that all groups had similar recognition accuracy, overall. Of the two-way interactions, only the interaction of Condition and Group significantly improved model fit. Model estimates indicated an overall smaller difference in performance between the L1 and L2 accent conditions for the Exposure group (&#x000DF; = 0.15, <italic>p</italic> = 0.006) and the Feedback group (&#x000DF; = 0.12, <italic>p</italic> = 0.006) compared to the Control group. As noted above, the log-likelihood model comparison (i.e., omnibus test) of the critical three-way interaction was non-significant (&#x003C7;<sup>2</sup> = 4.56, <italic>p</italic> = 0.10). However, the model estimates within the full model indicated a significant difference between the Exposure and Control groups (&#x000DF; = 0.23, <italic>p</italic> = 0.04); the difference between the Feedback and Control groups trended in the same direction but was non-significant (&#x000DF; = 0.15, <italic>p</italic> = 0.16). To better understand the three-way interaction, we created <italic>post-hoc</italic> models to directly compare performance by the Control and Exposure groups for the L2 accent condition at Pre-Test and then (in a separate model) at Post-Test. Model estimates indicated that at Pre-Test the Exposure group had (non-significantly) poorer performance (&#x000DF; = &#x02212;0.06, <italic>p</italic> = 0.19) than the Control group, and (non-significantly) better performance (&#x000DF; = 0.01, <italic>p</italic> = 0.77) at Post-Test; critically, the size of these trends suggest that the three-way interaction was driven by a difference at Pre-Test, not Post-Test. Given that the omnibus test was non-significant, and the significant model estimate appears to have been driven by a Pre-Test difference, we conclude that no meaningful (training-related) differences emerged between the Control and training groups in the recognition accuracy dataset.</p></sec>
<sec>
<title>Response time data from visual categorization task</title>
<p>Response time data for all conditions is presented in <xref ref-type="fig" rid="F2">Figure 2</xref>. We report all log-likelihood model comparisons in <xref ref-type="table" rid="T3">Table 3</xref> and provide full model summaries in <xref ref-type="supplementary-material" rid="SM1">Supplemental materials</xref>. Matching the results of the recognition accuracy analysis, results of the response time analysis indicated improvement (i.e., reduction in response times) from Pre-Test to Post-Test (&#x000DF; = &#x02212;99.94, <italic>p</italic> &#x0003C; 0.001). Improvements were also largest for the L2 accent condition (significant interaction of Condition and Session: &#x003C7;<sup>2</sup> = 50.05, <italic>p</italic> &#x0003C; 0.001). However, no differences in improvement emerged based on Group (all <italic>p</italic>s &#x0003E; 0.05).</p>
<fig id="F2" position="float">
<label>Figure 2</label>
<caption><p>Response time data from the dual-task paradigm at Pre-Test and Post-Test, for each group and accent condition, is presented with violin density distributions, mean points, and standard error bars.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="flang-03-1243678-g0002.tif"/>
</fig>
<table-wrap position="float" id="T3">
<label>Table 3</label>
<caption><p>Log-likelihood model comparisons from analyses of dual-task response time data.</p></caption>
<table frame="box" rules="all">
<thead>
<tr style="background-color:#919498;color:#ffffff">
<th valign="top" align="left"><bold>Effect</bold></th>
<th valign="top" align="center"><bold>&#x003C7;<sup>2</sup></bold></th>
<th valign="top" align="center"><bold>df</bold></th>
<th valign="top" align="center"><bold><italic>p</italic></bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Condition</td>
<td valign="top" align="center">34.00</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">&#x0003C; 0.001</td>
</tr> <tr>
<td valign="top" align="left">Session</td>
<td valign="top" align="center">33.35</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">&#x0003C; 0.001</td>
</tr> <tr>
<td valign="top" align="left">Group</td>
<td valign="top" align="center">1.38</td>
<td valign="top" align="center">2</td>
<td valign="top" align="center">0.50</td>
</tr> <tr>
<td valign="top" align="left">Condition: session</td>
<td valign="top" align="center">50.05</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">&#x0003C; 0.001</td>
</tr> <tr>
<td valign="top" align="left">Condition: group</td>
<td valign="top" align="center">0.01</td>
<td valign="top" align="center">2</td>
<td valign="top" align="center">0.99</td>
</tr> <tr>
<td valign="top" align="left">Session: group</td>
<td valign="top" align="center">3.03</td>
<td valign="top" align="center">2</td>
<td valign="top" align="center">0.22</td>
</tr> <tr>
<td valign="top" align="left">Condition: session: group</td>
<td valign="top" align="center">2.79</td>
<td valign="top" align="center">2</td>
<td valign="top" align="center">0.25</td>
</tr></tbody>
</table>
<table-wrap-foot>
<p><italic>df</italic> , degrees of freedom.</p>
</table-wrap-foot>
</table-wrap>
<p>Overall, participants had significantly slower response times on the secondary task when presented with an L2 accent in the primary task (as compared to an L1 accent; &#x000DF; = 39.45, <italic>p</italic> &#x0003C; 0.001). The response times did not differ overall by Group (&#x003C7;<sup>2</sup> = 1.38, <italic>p</italic> = 0.50), nor did the effect of Group interact with Condition (&#x003C7;<sup>2</sup> = 0.01, <italic>p</italic> = 0.99), Session (&#x003C7;<sup>2</sup> = 3.03, <italic>p</italic> = 0.22), or a combination of Condition and Session (&#x003C7;<sup>2</sup> = 2.79, <italic>p</italic> = 0.25). Model estimates of the three-way interactions were also non-significant, although the direction of the trends was as predicted: The difference in response times for the L1 and L2 accent conditions was reduced at Post-Test to a (non-significantly) larger degree for the Exposure (&#x000DF; = &#x02212;21.11, <italic>p</italic> = 0.16) and Feedback (&#x000DF; = &#x02212;21.91, <italic>p</italic> = 0.14) groups.</p>
</sec>
</sec>
<sec>
<title>Training data</title>
<sec>
<title>Transcription accuracy</title>
<p>Recognition accuracy data from the training sessions is presented in <xref ref-type="fig" rid="F3">Figure 3</xref>. Fixed effects of the model included: Day (dummy-coded levels: Days 2, 3, 4), Group (dummy-coded levels: Control, Exposure, Feedback), as well as the interaction between Day and Group.</p>
<fig id="F3" position="float">
<label>Figure 3</label>
<caption><p>Mean transcription accuracy and 95% confidence intervals are presented as a function of Day and Group for the training data. Participants in the Control group were presented with L1-accented trials on training days, while participants in the Exposure and Feedback groups were presented with L2-accented trials.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="flang-03-1243678-g0003.tif"/>
</fig>
<p>Log-likelihood model comparisons indicated that the effects of Day (&#x003C7;<sup>2</sup> = 25.76, <italic>p</italic> &#x0003C; 0.001) and Group (&#x003C7;<sup>2</sup> = 496.50, <italic>p</italic> &#x0003C; 0.001) both improved model fit. Model estimates revealed improvement across the training days, such that performance on Day 3 (&#x000DF; = 0.05, <italic>p</italic> &#x0003C; 0.05) and Day 4 (&#x000DF; = 0.12, <italic>p</italic> &#x0003C; 0.001) were each better than Day 2. Releveling of the fixed effect of Day in the model confirmed that the difference in performance between Days 3 and 4 was also significant (&#x000DF; = 0.07, <italic>p</italic> = 0.002). For the effect of Group, performance of subjects assigned to the Exposure and Feedback groups (both of which received entirely Mandarin-accented stimuli) was poorer than that of subjects assigned to the Control group (which received entirely American-accented stimuli; <italic>p</italic>s &#x0003C; 0.001). The fixed effect of Group was releveled in the model to directly compare the Exposure and Feedback groups but revealed no significant difference in overall performance (&#x000DF; = 0.05, <italic>p</italic> = 0.22). The interaction of Day and Group was non-significant (&#x003C7;<sup>2</sup> = 3.20, <italic>p</italic> = 0.53), indicating consistent improvement across days regardless of the assigned training.</p></sec></sec>
</sec>
<sec sec-type="discussion" id="s5">
<title>Discussion</title>
<p>In the present study, we investigated whether brief daily exposure to unfamiliar L2 accent improves listeners&#x00027; ability to accurately understand speech and, simultaneously, whether it reduces the cognitive load associated with speech processing. At Pre-Test and Post-Test, participants were presented with multiple L1- and L2-accented speakers while completing a dual-task paradigm. We predicted that response times (an index of cognitive load) during the L2 accent trials would be shortened (improved) for the subjects assigned to L2 accent training groups as compared to a Control group. Additionally, we predicted that speech recognition accuracy would improve in the L2 accent condition for the L2 accent training groups. Overall, our results indicated similar improvements for all groups. Critically, Post-Test performance for the L2 accent condition for the Control and L2 accent training groups did not differ significantly (although all trends were in the predicted directions). We conclude that the L2 accent trainings implemented in the present study did not successfully promote long-term learning benefits of a statistically meaningful magnitude. However, we also emphasize that the present effort is a methodologically informative starting point for future research on this topic.</p>
<p>Our examination of the data from the dual-task paradigm on Days 1 and 5 consistently revealed the following across all three random assignments: (1) Participants improved at the task from Pre- to Post-Test, making it easier both in terms of cognitive load (faster response times) and perceptual processing (better recognition accuracy); (2) Listening performance was poorer and cognitive load was greater for the L2 accent condition as compared the L1 accent condition. These outcomes were to be expected given the design of the experiment, and general participant learning effects. Of particular interest to the present study&#x00027;s aims was the interaction of these two elements with the training manipulation.</p>
<p>In the analysis of recognition accuracy, we found some evidence that the Exposure group (i.e., L2 accent training without implicit feedback), in particular, may have improved from Pre-Test to Post-Test to a larger degree than the Control group. However, <italic>post-hoc</italic> analyses following the critical three-way interaction revealed that the Exposure group likely demonstrated larger improvement because of a difference at Pre-Test, not Post-Test. In other words, it is impossible to determine whether they improved to a larger degree than the Control group because they had more &#x0201C;room to improve&#x0201D; or because of they received training with the L2 accent. Given that the omnibus test of the three-way interaction (i.e., including the Feedback group) was non-significant, and the corresponding model estimate for the Feedback group was non-significant, we conclude that the present study did <italic>not</italic> find sufficient evidence to indicate a benefit of the L2 accent trainings on listening performance. In this same vein, we also did not find evidence that indicated any benefit of the L2 accent training with implicit feedback over the L2 accent training without implicit feedback. Given prior evidence that the presentation of subtitles can promote adaptation to L2 accent (Chan et al., <xref ref-type="bibr" rid="B9">2020</xref>), we had predicted that listeners in the Feedback group would show a larger benefit than listeners in the Exposure group. Our results suggest that presenting target sentences after listening (instead of in tandem with listening) may not provide similar training benefits (cf., Burchill et al., <xref ref-type="bibr" rid="B8">2018</xref>). In future research, a manipulation that presents lexical items in tandem with auditory targets may prove to have a larger effect.</p>
<p>For the response time data (i.e., cognitive load), analyses indicated similar trends: The difference in cognitive load for the L1 and L2 accent conditions was reduced on Day 5 compared to Day 1, but to a similar degree for all participant groups. One limitation of the present dataset may be the use of a reaction time task to measure cognitive load. Indeed, the amount of variance in subject response times may have reduced our power to detect the critical interaction. Measures of cognitive load (or &#x0201C;listening effort&#x0201D;) can differ markedly in their sensitivity to detect differences; for example, examining the cognitive load associated with speech-in-noise perception, Strand et al. (<xref ref-type="bibr" rid="B31">2018</xref>) found that effect sizes were larger when using a semantic dual-task paradigm than a complex dual-task paradigm (the latter of which is most similar to the present study&#x00027;s paradigm). Pupillometry, a psychophysiological measure of cognitive load, was even more sensitive than these dual-task measures. In future work, using a psychophysiological measure, such as pupillometry or eye-tracking, may provide greater precision and ability to detect changes in cognitive load as well as processing speed.</p>
<p>As expected, participants presented with L1 accent (the Control group) on Days 2, 3, and 4 of the study had higher overall transcription accuracy on those days of the study than participants presented with L2 accent. Matching this outcome, participants in the L2 accent training groups also self-reported that the task was more effortful than participants in the Control group. Across days, all groups showed steady improvement in listening performance, matching prior work that has demonstrated sustained benefits of L2 accent trainings over brief periods (Lindemann et al., <xref ref-type="bibr" rid="B18">2016</xref>; Xie et al., <xref ref-type="bibr" rid="B34">2017</xref>). There was no difference, however, in the rate of improvement between the Control and L2 accent training groups. Additionally, we had predicted that the Feedback group may show more rapid improvement than the Exposure group, but this was not the case. Participants in the Exposure and Feedback groups did, however, perceive their performance as improving across days, whereas the Control group perceived their performance as declining. The Feedback group also perceived their performance as marginally more positive than the Exposure group, which may reflect their superior ability to self-assess performance with the implicit feedback available to them.</p>
<p>Self-reported motivation also varied by group. Participants in the Feedback group reported greater motivation to do well at the task than participants in the Control or Exposure groups. The Exposure and Control groups did not significantly differ, although the trends in the data suggested that both of the L2 accent training groups reported higher motivation than the Control group. It may be the case that the L2 accent stimuli were more engaging to listeners, albeit more challenging.</p>
<sec>
<title>Limitations and future directions</title>
<p>We acknowledge the possibility that limited statistical power may have encumbered our ability to detect significant training benefits in the present study. In the response time data, in particular, the degree of variance may have reduced our ability to detect effects. We suggest increasing the number of trials per condition in future work when comparing dual-task data across sessions or using a cross-modal matching task in place of a dual-task paradigm (also referred to as a &#x0201C;semantic&#x0201D; or &#x0201C;linguistic&#x0201D; dual-task paradigm; Strand et al., <xref ref-type="bibr" rid="B31">2018</xref>). Although we created 100 novel stimuli in addition to the 200 SNST items, across 5 days this resulted in only 39 items per accent condition per task. In lieu of a larger set of sentence recordings, one solution would be to repeat items, particularly on training days, in future research (see Bradlow and Bent, <xref ref-type="bibr" rid="B6">2008</xref>; Baese-Berk et al., <xref ref-type="bibr" rid="B2">2013</xref>). In Pre- and Post-Test measures, it is critical to include novel stimuli in order to prevent item-specific learning effects from artificially inflating performance; however, analyses of training sessions are typically less crucial, and items could be repeated. There may even be a benefit to repeating items across training sessions, although to our knowledge this has yet to be examined directly. It is our hope that the present study can serve as a benchmark when selecting paradigms and estimating power in future investigations of multi-day accent trainings. We recommend that future studies maximize potential effect sizes via a combination of the following methods: (1) Using more sensitive measures of cognitive load, (2) Increasing the number of trials in test sessions, (3) Increasing the number of trials in training sessions, and (4) Increasing the number of participants.</p>
<p>With regard to the first recommendation, we predict based on prior evidence from speech-in-noise perception (Strand et al., <xref ref-type="bibr" rid="B31">2018</xref>) that cross-modal matching tasks may produce larger effects than dual-task paradigms that involve a non-linguistic secondary task&#x02014;although direct comparisons of the sensitivity of these paradigms for L2 accent perception/adaptation have yet to be conducted. Comparing the results of Clarke and Garrett (<xref ref-type="bibr" rid="B10">2004</xref>) with Brown et al. (<xref ref-type="bibr" rid="B7">2020</xref>), however, provides some indication of what types tasks may be most sensitive in the context of L2 accent adaptation: In Clarke and Garrett (<xref ref-type="bibr" rid="B10">2004</xref>), rapid (single session) adaptation to L2 accent was robustly demonstrated both across four experimental blocks (each containing 16 trials) and within the early trials. In that study, a cross-modal matching task was used; specifically, participants completed a task where they responded &#x0201C;yes&#x0201D; or &#x0201C;no&#x0201D; to visually-presented sentence-final probe words. In contrast, Brown et al. (<xref ref-type="bibr" rid="B7">2020</xref>) used the same non-linguistic dual-task paradigm as the present study and only found evidence of rapid adaptation within the first 20 trials, not across the full 50-trial session. Other differences between the two studies (type of L2 accent, presence of background noise, etc.) may account for the deviating outcomes, but we (cautiously) recommend based on outcomes of these prior studies that researchers may be best served with cross-modal matching tasks in future work. We can also (more confidently) recommend pupillometry as a measure of cognitive load, which proved to be more sensitive than the dual-task paradigms in both Strand et al. (<xref ref-type="bibr" rid="B31">2018</xref>) and Brown et al. (<xref ref-type="bibr" rid="B7">2020</xref>).</p>
<p>From Pre- to Post-Test, (non-significant) trends in the data indicated larger benefits for the measure of listening performance than the measure of cognitive effort. This outcome runs counter to the findings of Bieber and Gordon-Salant (<xref ref-type="bibr" rid="B5">2021</xref>), in which benefits were observed at a test session 1-week after training for cognitive load but not listening performance. One possible explanation for the contrary outcome in the present study may be the design of the middle (training) days, which utilized a speech transcription task rather than the dual-task paradigm from the Pre- and Post-Test sessions. Thus, participants received more extensive training with the linguistic task, but not the non-linguistic (visual) task, from the dual-task paradigm. It may be the case that these training sessions were better situated to promote near-transfer to the more similar (linguistic) task. In future work, matching the designs of training and test sessions may be ideal to remove any potential differential transfer effects by task type.</p>
<p>One strength of the present study was the inclusion of L2-accented stimuli presented in quiet, as opposed to in noise. Although adding noise to stimuli can make it easier to match intelligibilities across conditions (e.g., matching L1 to L2 speakers), prior evidence also indicates that the cognitive and/or perceptual resources recruited to support noisy vs. accented listening conditions may differ (McLaughlin et al., <xref ref-type="bibr" rid="B21">2018</xref>). Thus, when examining questions pertaining to the perception of L2 accent, using L2-accented stimuli presented in noise may not always be suitable. To prevent ceiling effects in the present study, we decided to use semantically-anomalous sentences (e.g., &#x0201C;the wrong shot led the farm&#x0201D;), which pose a different potential issue: Namely, anomalous sentences reduce a listener&#x00027;s ability to use top-down information during speech processing, and are therefore less ecologically valid. Studies that use these types of items thus give a more direct assessment of bottom-up processing at the cost of limited generalizability of the findings. In future work, focusing solely on measures of cognitive load (as opposed to a combination of cognitive load and intelligibility measures) can remove these types of obstacles and allow for more ecological examinations of accent accommodation.</p></sec>
</sec>
<sec sec-type="conclusions" id="s6">
<title>Conclusion</title>
<p>Although L2 accent can pose a challenge during speech processing, listeners are able to rapidly accommodate L2 speakers&#x00027; unique productions, thereby reducing cognitive load (Clarke and Garrett, <xref ref-type="bibr" rid="B10">2004</xref>; Brown et al., <xref ref-type="bibr" rid="B7">2020</xref>). Additionally, correlational evidence suggests that the efficiency and accuracy of L2 accent processing depends on a listener&#x00027;s prior (real world) experience, with more experienced listeners typically processing L2 accent faster and more accurately (Kennedy and Trofimovich, <xref ref-type="bibr" rid="B15">2008</xref>; Porretta et al., <xref ref-type="bibr" rid="B26">2020</xref>). Empirical evidence connecting these two literatures, however, is lacking: Few studies to date have examined perceptual accommodation of L2 accent across multiple days (or weeks, etc.). In the present study, we took a first step toward filling this empirical gap, implementing a dual-task paradigm to measure changes in cognitive load and listening performance for perception of L2 accent across a 5-day period. Participants were either exposed to the L1 (Control) or L2 accent in the interim days, and half of the subjects exposed to L2 accent were provided with implicit feedback. Our results did not show a benefit of the L2 accent trainings, despite a larger sample size (<italic>n</italic> &#x0003E; 50 per group) than prior work (although all trends were in the predicted directions). We conclude that the L2 accent trainings implemented in the present study did not successfully promote long-term learning benefits of a statistically meaningful magnitude, but also emphasize that the present effort is a methodologically informative starting point for future research on this topic.</p></sec>
<sec sec-type="data-availability" id="s7">
<title>Data availability statement</title>
<p>The datasets presented in this study can be found in online repositories. The names of the repository/repositories and accession number(s) can be found below: <ext-link ext-link-type="uri" xlink:href="https://osf.io/2qz5n/files/osfstorage">https://osf.io/2qz5n/files/osfstorage</ext-link>.</p></sec>
<sec sec-type="ethics-statement" id="s8">
<title>Ethics statement</title>
<p>The studies involving humans were approved by Washington University in St. Louis IRB. The studies were conducted in accordance with the local legislation and institutional requirements. The participants provided their written informed consent to participate in this study.</p></sec>
<sec sec-type="author-contributions" id="s9">
<title>Author contributions</title>
<p>DM, MB-B, and KV contributed to conception and design of the study. DM performed the statistical analysis and wrote the first draft of the manuscript. All authors contributed to manuscript revision, read, and approved the submitted version.</p></sec>
</body>
<back>
<sec sec-type="funding-information" id="s10">
<title>Funding</title>
<p>This work was supported by National Science Foundation (NSF) Graduate Research Fellowship DGE-1745038, NSF BCS-2020805, NSF 2146993, the European Union&#x00027;s Horizon 2020 research and innovation programme under the Marie Sk&#x00142;odowska-Curie grant agreement no. 101103964, the Basque Government through the BERC 2022-2025 program, and the Spanish State Research Agency through BCBL Severo Ochoa excellence accreditation (CEX2020-001010-S).</p>
</sec>
<sec sec-type="COI-statement" id="conf1">
<title>Conflict of interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec sec-type="disclaimer" id="s11">
<title>Publisher&#x00027;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
<sec sec-type="supplementary-material" id="s12">
<title>Supplementary material</title>
<p>The Supplementary Material for this article can be found online at: <ext-link ext-link-type="uri" xlink:href="https://www.frontiersin.org/articles/10.3389/flang.2024.1243678/full#supplementary-material">https://www.frontiersin.org/articles/10.3389/flang.2024.1243678/full#supplementary-material</ext-link></p>
<supplementary-material xlink:href="Data_Sheet_1.docx" id="SM1" mimetype="application/vnd.openxmlformats-officedocument.wordprocessingml.document" xmlns:xlink="http://www.w3.org/1999/xlink"/>
<supplementary-material xlink:href="Data_Sheet_2.docx" id="SM2" mimetype="application/vnd.openxmlformats-officedocument.wordprocessingml.document" xmlns:xlink="http://www.w3.org/1999/xlink"/></sec>
<fn-group>
<fn id="fn0001"><p><sup>1</sup>Computational evidence from Xie et al. (<xref ref-type="bibr" rid="B33">2023</xref>) suggests that many of the benefits observed in adaptive speech perception experiments may be explained by multiple mechanisms, including: (1) changes to phonemic representations, (2) pre-linguistic signal normalization, and (3) changes in post-perceptual decision-making criteria.</p></fn>
<fn id="fn0002"><p><sup>2</sup>One limitation of this finding is that, without a control condition, the reductions in cognitive demands cannot be solely attributed to the training. It is possible that familiarization with the secondary task led to improved reaction times.</p></fn>
<fn id="fn0003"><p><sup>3</sup>The difference in terminology for the Pre-Test and Post-Test sessions (speech <italic>recognition</italic>) vs. the training sessions (speech <italic>transcription</italic>) corresponds to the different task demands. In the dual-task paradigm, participants heard target sentences and then repeated them aloud (because the secondary task required use of their hands to make responses). In the training sessions, however, there was no secondary (dual) task, so participants listened to the target sentences and then typed what they heard into a response box. Both measures are used to index <italic>listening performance</italic>.</p></fn>
</fn-group>
<ref-list>
<title>References</title>
<ref id="B1">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Arbuthnott</surname> <given-names>K.</given-names></name> <name><surname>Frank</surname> <given-names>J.</given-names></name></person-group> (<year>2000</year>). <article-title>Trail making test, part B as a measure of executive control: validation using a set-switching paradigm</article-title>. <source>J. Clin. Exp. Neuropsychol.</source> <volume>22</volume>, <fpage>518</fpage>&#x02013;<lpage>528</lpage>. <pub-id pub-id-type="doi">10.1076/1380-3395(200008)22:4</pub-id><pub-id pub-id-type="pmid">10923061</pub-id></citation></ref>
<ref id="B2">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Baese-Berk</surname> <given-names>M. M.</given-names></name> <name><surname>Bradlow</surname> <given-names>A. R.</given-names></name> <name><surname>Wright</surname> <given-names>B. A.</given-names></name></person-group> (<year>2013</year>). <article-title>Accent-independent adaptation to foreign accented speech</article-title>. <source>J. Acoust. Soc. Am.</source> 133, EL174&#x02013;EL180. <pub-id pub-id-type="doi">10.1121/1.4789864</pub-id><pub-id pub-id-type="pmid">23464125</pub-id></citation></ref>
<ref id="B3">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bates</surname> <given-names>D. M. M.</given-names></name> <name><surname>Bolker</surname> <given-names>B.</given-names></name> <name><surname>Walker</surname> <given-names>S.</given-names></name></person-group> (<year>2015</year>). <article-title>Fitting linear mixed-effects models using lme4</article-title>. <source>J. Stat. Softw.</source> <volume>67</volume>, <fpage>1</fpage>&#x02013;<lpage>48</lpage>. <pub-id pub-id-type="doi">10.18637/jss.v067.i01</pub-id></citation>
</ref>
<ref id="B4">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bent</surname> <given-names>T.</given-names></name> <name><surname>Baese-Berk</surname> <given-names>M.</given-names></name></person-group> (<year>2021</year>). <article-title>&#x0201C;Perceptual learning of accented speech,&#x0201D;</article-title> in <source>The Handbook of Speech Perception</source>, eds. J. S. Pardo, L. C. Nygaard, R. E. Remez and D. B. Pisoni. <pub-id pub-id-type="doi">10.1002/9781119184096.ch16</pub-id></citation>
</ref>
<ref id="B5">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bieber</surname> <given-names>R. E.</given-names></name> <name><surname>Gordon-Salant</surname> <given-names>S.</given-names></name></person-group> (<year>2021</year>). <article-title>Improving older adults&#x00027; understanding of challenging speech: auditory training, rapid adaptation and perceptual learning</article-title>. <source>Hear. Res</source>. <volume>402</volume>:<fpage>108054</fpage>. <pub-id pub-id-type="doi">10.1016/j.heares.2020.108054</pub-id><pub-id pub-id-type="pmid">32826108</pub-id></citation></ref>
<ref id="B6">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bradlow</surname> <given-names>A. R.</given-names></name> <name><surname>Bent</surname> <given-names>T.</given-names></name></person-group> (<year>2008</year>). <article-title>Perceptual adaptation to non-native speech</article-title>. <source>Cognition</source> <volume>106</volume>, <fpage>707</fpage>&#x02013;<lpage>729</lpage>. <pub-id pub-id-type="doi">10.1016/j.cognition.2007.04.005</pub-id><pub-id pub-id-type="pmid">17532315</pub-id></citation></ref>
<ref id="B7">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Brown</surname> <given-names>V. A.</given-names></name> <name><surname>McLaughlin</surname> <given-names>D. J.</given-names></name> <name><surname>Strand</surname> <given-names>J. F.</given-names></name> <name><surname>Van Engen</surname> <given-names>K. J.</given-names></name></person-group> (<year>2020</year>). <article-title>Rapid adaptation to fully intelligible nonnative-accented speech reduces listening effort</article-title>. <source>Q. J. Exp. Psychol.</source> <volume>73</volume>, <fpage>1431</fpage>&#x02013;<lpage>1443</lpage>. <pub-id pub-id-type="doi">10.1177/1747021820916726</pub-id><pub-id pub-id-type="pmid">32192390</pub-id></citation></ref>
<ref id="B8">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Burchill</surname> <given-names>Z.</given-names></name> <name><surname>Liu</surname> <given-names>L.</given-names></name> <name><surname>Jaeger</surname> <given-names>T. F.</given-names></name></person-group> (<year>2018</year>). <article-title>Maintaining information about speech input during accent adaptation</article-title>. <source>PLoS ONE</source> <volume>13</volume>:<fpage>e0199358</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pone.0199358</pub-id><pub-id pub-id-type="pmid">30086140</pub-id></citation></ref>
<ref id="B9">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chan</surname> <given-names>K. Y.</given-names></name> <name><surname>Lyons</surname> <given-names>C.</given-names></name> <name><surname>Kon</surname> <given-names>L. L.</given-names></name> <name><surname>Stine</surname> <given-names>K.</given-names></name> <name><surname>Manley</surname> <given-names>M.</given-names></name> <name><surname>Crossley</surname> <given-names>A.</given-names></name></person-group> (<year>2020</year>). <article-title>Effect of on-screen text on multimedia learning with native and foreign-accented narration</article-title>. <source>Learn. Instruct.</source> <volume>67</volume>:<fpage>101305</fpage>. <pub-id pub-id-type="doi">10.1016/j.learninstruc.2020.101305</pub-id></citation>
</ref>
<ref id="B10">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Clarke</surname> <given-names>C. M.</given-names></name> <name><surname>Garrett</surname> <given-names>M. F.</given-names></name></person-group> (<year>2004</year>). <article-title>Rapid adaptation to foreign-accented English</article-title>. <source>J. Acoust. Soc. Am.</source> <volume>116</volume>, <fpage>3647</fpage>&#x02013;<lpage>3658</lpage>. <pub-id pub-id-type="doi">10.1121/1.1815131</pub-id><pub-id pub-id-type="pmid">15658715</pub-id></citation></ref>
<ref id="B11">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Derwing</surname> <given-names>T. M.</given-names></name> <name><surname>Rossiter</surname> <given-names>M. J.</given-names></name> <name><surname>Munro</surname> <given-names>M. J.</given-names></name></person-group> (<year>2002</year>). <article-title>Teaching native speakers to listen to foreign-accented speech</article-title>. <source>J. Multiling. Multicult. Dev.</source> <volume>23</volume>, <fpage>245</fpage>&#x02013;<lpage>259</lpage>. <pub-id pub-id-type="doi">10.1080/01434630208666468</pub-id></citation>
</ref>
<ref id="B12">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Dietrich</surname> <given-names>S.</given-names></name> <name><surname>Hernandez</surname> <given-names>E.</given-names></name></person-group> (<year>2022</year>). <source>Language Use in the United States: 2019</source>. <publisher-loc>Suitland, MD</publisher-loc>: <publisher-name>American Community Survey Reports</publisher-name>.</citation>
</ref>
<ref id="B13">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Goldinger</surname> <given-names>S. D.</given-names></name></person-group> (<year>1998</year>). <article-title>Echoes of echoes? An episodic theory of lexical access</article-title>. <source>Psychol. Rev.</source> <volume>105</volume>, <fpage>251</fpage>&#x02013;<lpage>279</lpage>. <pub-id pub-id-type="doi">10.1037/0033-295x.105.2.251</pub-id><pub-id pub-id-type="pmid">9577239</pub-id></citation></ref>
<ref id="B14">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Johnson</surname> <given-names>K.</given-names></name></person-group> (<year>1997</year>). <article-title>&#x0201C;Speech perception without speaker normalization: an exemplar model,&#x0201D;</article-title> in <source>Talker Variability in Speech Processing</source>, eds. K. Johnson and J. Mullenix (Academic Press), <fpage>145</fpage>&#x02013;<lpage>166</lpage>.</citation>
</ref>
<ref id="B15">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kennedy</surname> <given-names>S.</given-names></name> <name><surname>Trofimovich</surname> <given-names>P.</given-names></name></person-group> (<year>2008</year>). <article-title>Intelligibility, comprehensibility, and accentedness of L2 speech: the role of listener experience and semantic context</article-title>. <source>Can. Modern Lang. Rev.</source> <volume>64</volume>, <fpage>459</fpage>&#x02013;<lpage>489</lpage>. <pub-id pub-id-type="doi">10.3138/cmlr.64.3.459</pub-id><pub-id pub-id-type="pmid">38144408</pub-id></citation></ref>
<ref id="B16">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kleinschmidt</surname> <given-names>D. F.</given-names></name> <name><surname>Jaeger</surname> <given-names>T. F.</given-names></name></person-group> (<year>2015</year>). <article-title>Robust speech perception: recognize the familiar, generalize to the similar, and adapt to the novel</article-title>. <source>Psychol. Rev.</source> <volume>122</volume>, <fpage>148</fpage>&#x02013;<lpage>203</lpage>. <pub-id pub-id-type="doi">10.1037/a0038695</pub-id><pub-id pub-id-type="pmid">25844873</pub-id></citation></ref>
<ref id="B17">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kuznetsova</surname> <given-names>A.</given-names></name> <name><surname>Brockhoff</surname> <given-names>P. B.</given-names></name> <name><surname>Christensen</surname> <given-names>R. H. B.</given-names></name></person-group> (<year>2017</year>). <article-title>lmerTest package: tests in linear mixed effects models</article-title>. <source>J. Stat. Softw.</source> <volume>82</volume>, <fpage>1</fpage>&#x02013;<lpage>26</lpage>. <pub-id pub-id-type="doi">10.18637/jss.v082.i13</pub-id></citation>
</ref>
<ref id="B18">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lindemann</surname> <given-names>S.</given-names></name> <name><surname>Campbell</surname> <given-names>M. A.</given-names></name> <name><surname>Litzenberg</surname> <given-names>J.</given-names></name> <name><surname>Close Subtirelu</surname> <given-names>N.</given-names></name></person-group> (<year>2016</year>). <article-title>Explicit and implicit training methods for improving native English speakers? comprehension of nonnative speech</article-title>. <source>J. Second Lang. Pronunc.</source> <volume>2</volume>, <fpage>93</fpage>&#x02013;<lpage>108</lpage>. <pub-id pub-id-type="doi">10.1075/jslp.2.1.04lin</pub-id><pub-id pub-id-type="pmid">33486653</pub-id></citation></ref>
<ref id="B19">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>MacLeod</surname> <given-names>C. M.</given-names></name></person-group> (<year>1991</year>). <article-title>Half a century of research on the Stroop effect: an integrative review</article-title>. <source>Psychol. Bull.</source> <volume>109</volume>:<fpage>163</fpage>. <pub-id pub-id-type="doi">10.1037/0033-2909.109.2.163</pub-id><pub-id pub-id-type="pmid">2034749</pub-id></citation></ref>
<ref id="B20">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Magnuson</surname> <given-names>J. S.</given-names></name> <name><surname>Nusbaum</surname> <given-names>H. C.</given-names></name> <name><surname>Akahane-Yamada</surname> <given-names>R.</given-names></name> <name><surname>Saltzman</surname> <given-names>D.</given-names></name></person-group> (<year>2021</year>). <article-title>Talker familiarity and the accommodation of talker variability</article-title>. <source>Attent. Percept. Psychophys.</source> <volume>83</volume>, <fpage>1842</fpage>&#x02013;<lpage>1860</lpage>. <pub-id pub-id-type="doi">10.3758/s13414-020-02203-y</pub-id><pub-id pub-id-type="pmid">33398658</pub-id></citation></ref>
<ref id="B21">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>McLaughlin</surname> <given-names>D. J.</given-names></name> <name><surname>Baese-Berk</surname> <given-names>M. M.</given-names></name> <name><surname>Bent</surname> <given-names>T.</given-names></name> <name><surname>Borrie</surname> <given-names>S. A.</given-names></name> <name><surname>Van Engen</surname> <given-names>K. J.</given-names></name></person-group> (<year>2018</year>). <article-title>Coping with adversity: individual differences in the perception of noisy and accented speech</article-title>. <source>Attent. Percept. Psychophys.</source> <volume>80</volume>, <fpage>1559</fpage>&#x02013;<lpage>1570</lpage>. <pub-id pub-id-type="doi">10.3758/s13414-018-1537-4</pub-id><pub-id pub-id-type="pmid">29740795</pub-id></citation></ref>
<ref id="B22">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Newman</surname> <given-names>R. S.</given-names></name> <name><surname>Evers</surname> <given-names>S.</given-names></name></person-group> (<year>2007</year>). <article-title>The effect of talker familiarity on stream segregation</article-title>. <source>J. Phonet.</source> <volume>35</volume>, <fpage>85</fpage>&#x02013;<lpage>103</lpage>. <pub-id pub-id-type="doi">10.1016/j.wocn.2005.10.004</pub-id></citation>
</ref>
<ref id="B23">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Nye</surname> <given-names>P. W.</given-names></name> <name><surname>Gaitenby</surname> <given-names>J. H.</given-names></name></person-group> (<year>1974</year>). <article-title>The intelligibility of synthetic monosyllabic words in short, syntactically normal sentences</article-title>. <source>Haskins Lab. Status Rep. Speech Res.</source> <volume>38</volume>:<fpage>43</fpage>.</citation>
</ref>
<ref id="B24">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Nygaard</surname> <given-names>L. C.</given-names></name> <name><surname>Pisoni</surname> <given-names>D. B.</given-names></name></person-group> (<year>1998</year>). <article-title>Talker-specific learning in speech perception</article-title>. <source>Percept. Psychophys.</source> <volume>60</volume>, <fpage>355</fpage>&#x02013;<lpage>376</lpage>. <pub-id pub-id-type="doi">10.3758/BF03206860</pub-id><pub-id pub-id-type="pmid">9599989</pub-id></citation></ref>
<ref id="B25">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Pierrehumbert</surname> <given-names>J.</given-names></name></person-group> (<year>2001</year>). <article-title>&#x0201C;Lenition and contrast,&#x0201D;</article-title> in <source>Frequency and the Emergence of Linguistic Structure</source>, 137.</citation>
</ref>
<ref id="B26">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Porretta</surname> <given-names>V.</given-names></name> <name><surname>Buchanan</surname> <given-names>L.</given-names></name> <name><surname>J&#x000E4;rvikivi</surname> <given-names>J.</given-names></name></person-group> (<year>2020</year>). <article-title>When processing costs impact predictive processing: the case of foreign-accented speech and accent experience</article-title>. <source>Attent. Percept. Psychophys.</source> <volume>82</volume>, <fpage>1558</fpage>&#x02013;<lpage>1565</lpage>. <pub-id pub-id-type="doi">10.3758/s13414-019-01946-7</pub-id><pub-id pub-id-type="pmid">31970710</pub-id></citation></ref>
<ref id="B27">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Porretta</surname> <given-names>V.</given-names></name> <name><surname>Tucker</surname> <given-names>B. V.</given-names></name></person-group> (<year>2019</year>). <article-title>Eyes wide open: pupillary response to a foreign accent varying in intelligibility</article-title>. <source>Front. Commun.</source> <volume>4</volume>:<fpage>8</fpage>. <pub-id pub-id-type="doi">10.3389/fcomm.2019.00008</pub-id></citation>
</ref>
<ref id="B28">
<citation citation-type="web"><person-group person-group-type="author"><collab>R Core Team</collab></person-group> (<year>2021</year>). <source>R: A Language and Environment for Statistical Computing</source>. Vienna: R Foundation for Statistical Computing. Available online at: <ext-link ext-link-type="uri" xlink:href="https://www.R-project.org/">https://www.R-project.org/</ext-link> (accessed January 1, 2022).</citation>
</ref>
<ref id="B29">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sidaras</surname> <given-names>S. K.</given-names></name> <name><surname>Alexander</surname> <given-names>J. E.</given-names></name> <name><surname>Nygaard</surname> <given-names>L. C.</given-names></name></person-group> (<year>2009</year>). <article-title>Perceptual learning of systematic variation in Spanish-accented speech</article-title>. <source>J. Acoust. Soc. Am.</source> <volume>125</volume>, <fpage>3306</fpage>&#x02013;<lpage>3316</lpage>. <pub-id pub-id-type="doi">10.1121/1.3101452</pub-id><pub-id pub-id-type="pmid">19425672</pub-id></citation></ref>
<ref id="B30">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Smith</surname> <given-names>S. L.</given-names></name> <name><surname>Pichora-Fuller</surname> <given-names>M. K.</given-names></name> <name><surname>Alexander</surname> <given-names>G.</given-names></name></person-group> (<year>2016</year>). <article-title>Development of the word auditory recognition and recall measure: a working memory test for use in rehabilitative audiology</article-title>. <source>Ear Hear.</source> <volume>37</volume>, <fpage>e360</fpage>&#x02013;<lpage>e376</lpage>. <pub-id pub-id-type="doi">10.1097/AUD.0000000000000329</pub-id><pub-id pub-id-type="pmid">27438869</pub-id></citation></ref>
<ref id="B31">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Strand</surname> <given-names>J. F.</given-names></name> <name><surname>Brown</surname> <given-names>V. A.</given-names></name> <name><surname>Merchant</surname> <given-names>M. B.</given-names></name> <name><surname>Brown</surname> <given-names>H. E.</given-names></name> <name><surname>Smith</surname> <given-names>J.</given-names></name></person-group> (<year>2018</year>). <article-title>Measuring listening effort: convergent validity, sensitivity, and links with cognitive and personality measures</article-title>. <source>J. Speech Lang. Hear. Res.</source> <volume>61</volume>, <fpage>1463</fpage>&#x02013;<lpage>1486</lpage>. <pub-id pub-id-type="doi">10.1044/2018_JSLHR-H-17-0257</pub-id><pub-id pub-id-type="pmid">29800081</pub-id></citation></ref>
<ref id="B32">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Witteman</surname> <given-names>M. J.</given-names></name> <name><surname>Weber</surname> <given-names>A.</given-names></name> <name><surname>McQueen</surname> <given-names>J. M.</given-names></name></person-group> (<year>2013</year>). <article-title>Foreign accent strength and listener familiarity with an accent codetermine speed of perceptual adaptation</article-title>. <source>Attent. Percept. Psychophys.</source> <volume>75</volume>, <fpage>537</fpage>&#x02013;<lpage>556</lpage>. <pub-id pub-id-type="doi">10.3758/s13414-012-0404-y</pub-id><pub-id pub-id-type="pmid">23456266</pub-id></citation></ref>
<ref id="B33">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Xie</surname> <given-names>X.</given-names></name> <name><surname>Jaeger</surname> <given-names>T. F.</given-names></name> <name><surname>Kurumada</surname> <given-names>C.</given-names></name></person-group> (<year>2023</year>). <article-title>What we do (not) know about the mechanisms underlying adaptive speech perception: a computational framework and review</article-title>. <source>Cortex</source> <volume>66</volume>, <fpage>377</fpage>&#x02013;<lpage>424</lpage>. <pub-id pub-id-type="doi">10.1016/j.cortex.2023.05.003</pub-id><pub-id pub-id-type="pmid">37506665</pub-id></citation></ref>
<ref id="B34">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Xie</surname> <given-names>X.</given-names></name> <name><surname>Theodore</surname> <given-names>R. M.</given-names></name> <name><surname>Myers</surname> <given-names>E. B.</given-names></name></person-group> (<year>2017</year>). <article-title>More than a boundary shift: Perceptual adaptation to foreign-accented speech reshapes the internal structure of phonetic categories</article-title>. <source>J. Exp. Psychol.</source> <volume>43</volume>:<fpage>206</fpage>. <pub-id pub-id-type="doi">10.1037/xhp0000285</pub-id><pub-id pub-id-type="pmid">27819457</pub-id></citation></ref>
</ref-list>
</back>
</article>