<?xml version="1.0" encoding="UTF-8" standalone="no"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<?covid-19-tdm?>
<article xml:lang="EN" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Psychol.</journal-id>
<journal-title>Frontiers in Psychology</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Psychol.</abbrev-journal-title>
<issn pub-type="epub">1664-1078</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fpsyg.2022.874345</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Psychology</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Face Masks Impact Auditory and Audiovisual Consonant Recognition in Children With and Without Hearing Loss</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name><surname>Lalonde</surname> <given-names>Kaylah</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="corresp" rid="c001"><sup>&#x002A;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/1279092/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Buss</surname> <given-names>Emily</given-names></name>
<xref ref-type="aff" rid="aff2"><sup>2</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/759845/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Miller</surname> <given-names>Margaret K.</given-names></name>
<xref ref-type="aff" rid="aff3"><sup>3</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/1699939/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Leibold</surname> <given-names>Lori J.</given-names></name>
<xref ref-type="aff" rid="aff3"><sup>3</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/609119/overview"/>
</contrib>
</contrib-group>
<aff id="aff1"><sup>1</sup><institution>Audiovisual Speech Processing Laboratory, Boys Town National Research Hospital, Center for Hearing Research</institution>, <addr-line>Omaha, NE</addr-line>, <country>United States</country></aff>
<aff id="aff2"><sup>2</sup><institution>Speech Perception and Auditory Research at Carolina Laboratory, Department of Otolaryngology Head and Neck Surgery, University of North Carolina School of Medicine</institution>, <addr-line>Chapel Hill, NC</addr-line>, <country>United States</country></aff>
<aff id="aff3"><sup>3</sup><institution>Human Auditory Development Laboratory, Boys Town National Research Hospital, Center for Hearing Research</institution>, <addr-line>Omaha, NE</addr-line>, <country>United States</country></aff>
<author-notes>
<fn fn-type="edited-by"><p>Edited by: Monika Molnar, University of Toronto, Canada</p></fn>
<fn fn-type="edited-by"><p>Reviewed by: Yi Yuan, The Ohio State University, United States; Yang Zhang, University of Minnesota, Twin Cities, United States</p></fn>
<corresp id="c001">&#x002A;Correspondence: Kaylah Lalonde, <email>kaylah.lalonde@boystown.org</email></corresp>
<fn fn-type="other" id="fn004"><p>This article was submitted to Perception Science, a section of the journal Frontiers in Psychology</p></fn>
</author-notes>
<pub-date pub-type="epub">
<day>13</day>
<month>05</month>
<year>2022</year>
</pub-date>
<pub-date pub-type="collection">
<year>2022</year>
</pub-date>
<volume>13</volume>
<elocation-id>874345</elocation-id>
<history>
<date date-type="received">
<day>12</day>
<month>02</month>
<year>2022</year>
</date>
<date date-type="accepted">
<day>22</day>
<month>03</month>
<year>2022</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x00A9; 2022 Lalonde, Buss, Miller and Leibold.</copyright-statement>
<copyright-year>2022</copyright-year>
<copyright-holder>Lalonde, Buss, Miller and Leibold</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/"><p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p></license>
</permissions>
<abstract>
<p>Teachers and students are wearing face masks in many classrooms to limit the spread of the coronavirus. Face masks disrupt speech understanding by concealing lip-reading cues and reducing transmission of high-frequency acoustic speech content. Transparent masks provide greater access to visual speech cues than opaque masks but tend to cause greater acoustic attenuation. This study examined the effects of four types of face masks on auditory-only and audiovisual speech recognition in 18 children with bilateral hearing loss, 16 children with normal hearing, and 38 adults with normal hearing tested in their homes, as well as 15 adults with normal hearing tested in the laboratory. Stimuli simulated the acoustic attenuation and visual obstruction caused by four different face masks: hospital, fabric, and two transparent masks. Participants tested in their homes completed auditory-only and audiovisual consonant recognition tests with speech-spectrum noise at 0 dB SNR. Adults tested in the lab completed the same tests at 0 and/or &#x2212;10 dB SNR. A subset of participants from each group completed a visual-only consonant recognition test with no mask. Consonant recognition accuracy and transmission of three phonetic features (place of articulation, manner of articulation, and voicing) were analyzed using linear mixed-effects models. Children with hearing loss identified consonants less accurately than children with normal hearing and adults with normal hearing tested at 0 dB SNR. However, all the groups were similarly impacted by face masks. Under auditory-only conditions, results were consistent with the pattern of high-frequency acoustic attenuation; hospital masks had the least impact on performance. Under audiovisual conditions, transparent masks had less impact on performance than opaque masks. High-frequency attenuation and visual obstruction had the greatest impact on place perception. The latter finding was consistent with the visual-only feature transmission data. These results suggest that the combination of noise and face masks negatively impacts speech understanding in children. The best mask for promoting speech understanding in noisy environments depend on whether visual cues will be accessible: hospital masks are best under auditory-only conditions, but well-fit transparent masks are best when listeners have a clear, consistent view of the talker&#x2019;s face.</p>
</abstract>
<kwd-group>
<kwd>COVID-19</kwd>
<kwd>face mask</kwd>
<kwd>hearing loss</kwd>
<kwd>children</kwd>
<kwd>speech intelligibility</kwd>
<kwd>audiovisual perception</kwd>
</kwd-group>
<contract-num rid="cn001">5R01 DC011038</contract-num>
<contract-num rid="cn001">P20GM109023</contract-num>
<contract-sponsor id="cn001">National Institutes of Health<named-content content-type="fundref-id">10.13039/100000002</named-content></contract-sponsor>
<counts>
<fig-count count="7"/>
<table-count count="1"/>
<equation-count count="0"/>
<ref-count count="70"/>
<page-count count="17"/>
<word-count count="14758"/>
</counts>
</article-meta>
</front>
<body>
<sec id="S1" sec-type="intro">
<title>Introduction</title>
<p>Face masks are important for limiting the spread of the coronavirus disease 2019 (COVID-19), because they decrease aerosol transmission of the virus during speaking and coughing (<xref ref-type="bibr" rid="B7">Centers for Disease Control and Prevention, 2021a</xref>). To limit the spread of the coronavirus, the Centers for Disease Control and Prevention (CDC) stated in the summer of 2020 that individuals aged 2 years and older should wear masks in public and when around people who do not live in their household (<xref ref-type="bibr" rid="B8">Centers for Disease Control and Prevention, 2021b</xref>). Many United States schools returned to in-person or hybrid learning for the 2020&#x2013;2021 school year, with teachers and students wearing face masks in most classrooms. Even after resolution of the current pandemic, face masks may be required to control outbreaks of COVID-19 or other contagious diseases.</p>
<p>While face masks are critically important for preventing the spread of COVID-19, behavioral studies have demonstrated a detrimental effect of face masks on adults&#x2019; speech recognition and recall (<xref ref-type="bibr" rid="B69">Wittum et al., 2013</xref>; <xref ref-type="bibr" rid="B2">Atcherson et al., 2017</xref>; <xref ref-type="bibr" rid="B5">Bottalico et al., 2020</xref>; <xref ref-type="bibr" rid="B6">Brown et al., 2021</xref>; <xref ref-type="bibr" rid="B37">Magee et al., 2021</xref>; <xref ref-type="bibr" rid="B44">Muzzi et al., 2021</xref>; <xref ref-type="bibr" rid="B56">Smiljanic et al., 2021</xref>; <xref ref-type="bibr" rid="B62">Toscano and Toscano, 2021</xref>; <xref ref-type="bibr" rid="B63">Truong et al., 2021</xref>; <xref ref-type="bibr" rid="B67">Vos et al., 2021</xref>; <xref ref-type="bibr" rid="B70">Yi et al., 2021</xref>). Specifically, the use of face masks can lead to difficulties with speech understanding by concealing lip-reading cues and reducing the transmission of high-frequency speech content (<xref ref-type="bibr" rid="B47">Palmiero et al., 2016</xref>; <xref ref-type="bibr" rid="B1">Atcherson et al., 2020</xref>; <xref ref-type="bibr" rid="B14">Corey et al., 2020</xref>; <xref ref-type="bibr" rid="B18">Goldin et al., 2020</xref>; <xref ref-type="bibr" rid="B21">Jeong et al., 2020</xref>; <xref ref-type="bibr" rid="B48">P&#x00F6;rschmann et al., 2020</xref>; <xref ref-type="bibr" rid="B37">Magee et al., 2021</xref>; <xref ref-type="bibr" rid="B44">Muzzi et al., 2021</xref>; <xref ref-type="bibr" rid="B56">Smiljanic et al., 2021</xref>; <xref ref-type="bibr" rid="B62">Toscano and Toscano, 2021</xref>; <xref ref-type="bibr" rid="B63">Truong et al., 2021</xref>; <xref ref-type="bibr" rid="B67">Vos et al., 2021</xref>; <xref ref-type="bibr" rid="B70">Yi et al., 2021</xref>). Adults with and without hearing loss report that both factors degrade speech recognition (<xref ref-type="bibr" rid="B45">Naylor et al., 2020</xref>; <xref ref-type="bibr" rid="B53">Saunders et al., 2020</xref>). When a conversation partner wears a mask, it affects the listener&#x2019;s hearing and feeling of engagement with the talker (<xref ref-type="bibr" rid="B53">Saunders et al., 2020</xref>); among adults, those with hearing loss are more greatly impacted (<xref ref-type="bibr" rid="B53">Saunders et al., 2020</xref>).</p>
<p>Most prior studies investigating the effects of face masks on speech understanding are limited to adult subjects. Therefore, effects of face masks on children&#x2019;s auditory and audiovisual speech understanding in environments, such as classrooms, are not yet understood. Additionally, there are many types of face masks in use, with different effects on acoustic and visual speech cues (e.g., <xref ref-type="bibr" rid="B14">Corey et al., 2020</xref>; <xref ref-type="bibr" rid="B70">Yi et al., 2021</xref>). The purpose of this study is to examine the effects of various types of face masks on auditory and audiovisual speech recognition in children with and without hearing loss, in hopes of providing recommendations regarding types of face masks that best support children&#x2019;s speech understanding.</p>
<p>For the general public, including teachers, the CDC recommended non-medical disposable masks, breathable cloth masks made of multiple layers of tightly woven fabrics (such as cotton and cotton blends), or respirators without vents (<xref ref-type="bibr" rid="B10">Centers for Disease Control and Prevention, 2022a</xref>). However, teachers and others who interact with individuals who are deaf and hard of hearing were encouraged to wear transparent masks (<xref ref-type="bibr" rid="B11">Centers for Disease Control and Prevention, 2022b</xref>). In earlier guidelines, teachers were also encouraged to wear a transparent mask if they interact with students with special education or healthcare needs, teach young students who are learning to read, teach English as a second language, or teach students with disabilities, including hearing loss (<xref ref-type="bibr" rid="B9">Centers for Disease Control and Prevention, 2020</xref>). Acoustic attenuation is related to material attributes, such as thickness, weight, weave density, and porosity (<xref ref-type="bibr" rid="B35">Llamas et al., 2008</xref>). Among the common alternatives, hospital masks cause the least transmission loss (<xref ref-type="bibr" rid="B1">Atcherson et al., 2020</xref>; <xref ref-type="bibr" rid="B5">Bottalico et al., 2020</xref>; <xref ref-type="bibr" rid="B14">Corey et al., 2020</xref>; <xref ref-type="bibr" rid="B18">Goldin et al., 2020</xref>; <xref ref-type="bibr" rid="B67">Vos et al., 2021</xref>). Cloth masks vary considerably depending on fabric attributes and number of layers, with one study showing between 0.4 to 9 dB greater attenuation between 2 and 16 kHz for cloth masks than for a disposable hospital mask (<xref ref-type="bibr" rid="B14">Corey et al., 2020</xref>).</p>
<p>Opaque face masks conceal a substantial portion of visual cues used for lip-reading and audiovisual speech enhancement. In one study, occluding the lower half of the face decreased adults&#x2019; lip-reading of consonant-vowel (CV) syllables from an 80 to 20% accuracy, and their audiovisual enhancement from a &#x223C;21 to 4% benefit (<xref ref-type="bibr" rid="B22">Jordan and Thomas, 2011</xref>). Transparent masks (made of plastic or vinyl) provide access to more visual speech cues compared to opaque masks. Adults with moderate to profound hearing loss benefit from visual cues available when the talker is wearing a mask with a transparent window (<xref ref-type="bibr" rid="B2">Atcherson et al., 2017</xref>). When listening to speech in noisy backgrounds, adults with normal hearing (ANH) also benefit from visual cues available when the talker is wearing a transparent mask (<xref ref-type="bibr" rid="B70">Yi et al., 2021</xref>). However, transparent masks also cause 7&#x2013;16 dB greater high-frequency attenuation than opaque disposable masks (<xref ref-type="bibr" rid="B1">Atcherson et al., 2020</xref>; <xref ref-type="bibr" rid="B14">Corey et al., 2020</xref>).</p>
<p>Different types of transparent masks vary with respect to visual cues they provide access to. Some transparent masks, such as ClearMask&#x2122; (Clear Mask, LLC), conceal very little of the face (<xref ref-type="fig" rid="F1">Figure 1E</xref>). Others, such as The Communicator&#x2122; (Safe&#x2019;N&#x2019;Clear, Inc.), are primarily opaque with a transparent window in the region of the mouth (<xref ref-type="fig" rid="F1">Figure 1D</xref>). Although not previously examined, differences in visual cues available with different transparent masks might affect audiovisual speech perception. Some non-oral regions of the face contain subtle visual cues that could contribute to audiovisual enhancement (<xref ref-type="bibr" rid="B54">Scheinberg, 1980</xref>; <xref ref-type="bibr" rid="B49">Preminger et al., 1998</xref>). However, the impact of these non-oral visual cues may be negligible when oral cues are available, because the shape and movement of the oral area are highly correlated with movements of the jaw and cheeks (<xref ref-type="bibr" rid="B43">Munhall and Vatikiotis-Bateson, 1998</xref>).</p>
<fig id="F1" position="float">
<label>FIGURE 1</label>
<caption><p>Illustration of the five mask conditions. The top row contains photographs of the five conditions: <bold>(A)</bold> no mask, <bold>(B)</bold> hospital mask, <bold>(C)</bold> fabric mask, <bold>(D)</bold> Communicator&#x2122;, and <bold>(E)</bold> ClearMask&#x2122;. The bottom row shows the associated visual face mask simulations: <bold>(F)</bold> no mask, <bold>(G)</bold> hospital mask, <bold>(H)</bold> fabric mask, <bold>(I)</bold> Communicator&#x2122;, and <bold>(J)</bold> ClearMask&#x2122;.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-874345-g001.tif"/>
</fig>
<p>Several recent studies have presented data on the effects of face masks on speech recognition and recall in adults. Among ANH, researchers have reported 12&#x2013;23% poorer auditory-only (<xref ref-type="bibr" rid="B69">Wittum et al., 2013</xref>; <xref ref-type="bibr" rid="B5">Bottalico et al., 2020</xref>; <xref ref-type="bibr" rid="B44">Muzzi et al., 2021</xref>; <xref ref-type="bibr" rid="B70">Yi et al., 2021</xref>) and audiovisual (<xref ref-type="bibr" rid="B6">Brown et al., 2021</xref>; <xref ref-type="bibr" rid="B56">Smiljanic et al., 2021</xref>; <xref ref-type="bibr" rid="B70">Yi et al., 2021</xref>) word and sentence recognition in noise with face masks than without face masks. ANH are worse at recalling sentences presented audiovisually with a cloth or a hospital face mask than without a face mask (<xref ref-type="bibr" rid="B56">Smiljanic et al., 2021</xref>; <xref ref-type="bibr" rid="B63">Truong et al., 2021</xref>), and they report greater listening effort for speech produced with a face mask than without one (<xref ref-type="bibr" rid="B5">Bottalico et al., 2020</xref>; <xref ref-type="bibr" rid="B6">Brown et al., 2021</xref>). Among adult hearing aid users with moderate hearing loss, <xref ref-type="bibr" rid="B2">Atcherson et al. (2017)</xref> observed 7&#x2013;8% worse auditory sentence recognition in multi-talker babble with the talker wearing a hospital mask than without one. Among adult cochlear implant users, <xref ref-type="bibr" rid="B67">Vos et al. (2021)</xref> observed 26% lower auditory sentence recognition accuracy in quiet with an N95 respirator in combination with a face shield but no effect of the N95 respirator alone. Deficits depend on the amount of background noise in the listening environment (<xref ref-type="bibr" rid="B6">Brown et al., 2021</xref>; <xref ref-type="bibr" rid="B56">Smiljanic et al., 2021</xref>), with limited effects among ANH tested in quiet or at high SNRs (+5 to +13 dB) (<xref ref-type="bibr" rid="B35">Llamas et al., 2008</xref>; <xref ref-type="bibr" rid="B40">Mendel et al., 2008</xref>; <xref ref-type="bibr" rid="B2">Atcherson et al., 2017</xref>; <xref ref-type="bibr" rid="B6">Brown et al., 2021</xref>; <xref ref-type="bibr" rid="B56">Smiljanic et al., 2021</xref>; <xref ref-type="bibr" rid="B62">Toscano and Toscano, 2021</xref>). Additionally, auditory-only performance may not be affected by face masks in individuals with severe to profound hearing loss (<xref ref-type="bibr" rid="B2">Atcherson et al., 2017</xref>; <xref ref-type="bibr" rid="B67">Vos et al., 2021</xref>), potentially because their hearing loss restricts the speech bandwidth with and without a face mask.</p>
<p>Recent studies have demonstrated that acoustic effects of face masks significantly impact auditory-only speech recognition in children with normal hearing (CNH; <xref ref-type="bibr" rid="B17">Flaherty, 2021</xref>; <xref ref-type="bibr" rid="B55">Sfakianaki et al., 2021</xref>), with a similar impact on children and adults. <xref ref-type="bibr" rid="B17">Flaherty, 2021</xref> tested auditory-only speech recognition in a two-talker speech masker among 24 adults and 30 children (8&#x2013;12 years of age), comparing masked speech recognition thresholds at baseline to a primarily transparent mask (ClearMask&#x2122;), a face shield, an N95, and a hospital mask. Subjects required a more favorable target-to-masker ratio to understand speech produced with the ClearMask&#x2122; and face shield than with no mask, with similar effect sizes in children and adults. <xref ref-type="bibr" rid="B55">Sfakianaki et al. (2021)</xref> tested 10 adults and 10 children (6&#x2013;8 years of age) in quiet, classroom noise, and a two-talker masker, comparing accuracy with no mask to accuracy with a hospital mask. Subjects accurately repeated back fewer words when listening to speech with the hospital mask than without it. There was a non-significant trend for greater impact of the mask in children than in adults.</p>
<p>Although the CDC recommended using a transparent mask when communicating with individuals who are deaf or hard of hearing (<xref ref-type="bibr" rid="B11">Centers for Disease Control and Prevention, 2022b</xref>), only one study has examined the effects of face masks on speech perception in children with hearing loss (CHL). <xref ref-type="bibr" rid="B34">Lipps et al. (2021)</xref> examined audiovisual word identification in quiet in 13 CHL (3&#x2013;7 years of age), comparing accuracy at baseline to accuracy with a hospital mask, ClearMask&#x2122;, and a transparent apron mask. The hospital mask and the transparent apron mask negatively impacted children&#x2019;s audiovisual word identification accuracy, but the ClearMask&#x2122; did not. This finding indicates that a transparent mask is likely the best choice for young CHL listening to speech under quiet audiovisual conditions. We have only begun to scratch the surface of understanding how face masks affect communication in children with and without hearing loss, which limits our ability to make evidence-based recommendations about what type of face mask best supports communication in children with and without hearing loss.</p>
<p>Measurements of the acoustic transmission loss associated with the use of opaque and transparent face masks indicate a trade-off between concealment of visual cues and the degree of high-frequency acoustic attenuation (<xref ref-type="bibr" rid="B1">Atcherson et al., 2020</xref>; <xref ref-type="bibr" rid="B14">Corey et al., 2020</xref>; <xref ref-type="bibr" rid="B67">Vos et al., 2021</xref>). When listening to speech in moderate levels of background noise (&#x2212;5 dB SNR), the availability of visual cues outweighs the additional high-frequency acoustic attenuation caused by the transparent mask in ANH (<xref ref-type="bibr" rid="B70">Yi et al., 2021</xref>). We expect that age- and hearing-related differences in susceptibility to both high-frequency acoustic attenuation and reliance on visual speech cues will affect where children fall in this trade-off. For instance, children may be particularly affected by the acoustic attenuation caused by face masks, because they require greater bandwidths than adults for masked speech understanding (<xref ref-type="bibr" rid="B57">Stelmachowicz et al., 2000</xref>, <xref ref-type="bibr" rid="B58">2001</xref>; <xref ref-type="bibr" rid="B41">Mlot et al., 2010</xref>; <xref ref-type="bibr" rid="B39">McCreery and Stelmachowicz, 2011</xref>). However, the acoustic attenuation of face masks may have a smaller effect on CHL who have poor high-frequency aided audibility, because the hearing loss restricts the speech bandwidth both with and without a face mask. The loss of visual cues resulting from the use of opaque face masks is likely to affect both CNH and CHL, as both groups benefit significantly from visual speech cues (see reviews by <xref ref-type="bibr" rid="B27">Lalonde and McCreery, 2020</xref>; <xref ref-type="bibr" rid="B28">Lalonde and Werner, 2021</xref>). However, CNH are likely to be less affected than adults by the loss of these visual cues, because they benefit less from visual speech (<xref ref-type="bibr" rid="B68">Wightman et al., 2006</xref>; <xref ref-type="bibr" rid="B51">Ross et al., 2011</xref>; <xref ref-type="bibr" rid="B26">Lalonde and Holt, 2016</xref>). Finally, CHL (and particularly those with more severe hearing loss) may be more impacted by the loss of visual cues, because they benefit more from visual speech than CNH and CHL with less severe hearing loss (<xref ref-type="bibr" rid="B27">Lalonde and McCreery, 2020</xref>).</p>
<p>Previous studies examining the impact of face masks on speech understanding have measured effects at the word or sentence level. Face masks likely have a greater detrimental impact on some speech features than others. Reduced transmission of high-frequency speech content is likely to degrade the perception of speech features based primarily on high-frequency acoustic cues, such as consonants and specifically their place of articulation (<xref ref-type="bibr" rid="B59">Stevens, 2000</xref>). Similarly, reduced transmission of visual speech cues is likely to degrade the perception of speech features that are easy to speech-read, such as consonant place of articulation (<xref ref-type="bibr" rid="B4">Binnie et al., 1974</xref>; <xref ref-type="bibr" rid="B46">Owens and Blazek, 1985</xref>).</p>
<p>The purpose of this study was twofold. First, we aimed to determine the impact of face masks on auditory and audiovisual speech recognition in noise among CNH and CHL. Second, we aimed to determine what type of face mask may best support communication in these groups. We used four types of face masks: a hospital mask, a fabric mask, a primarily transparent mask (ClearMask&#x2122;), and a primarily opaque mask with a transparent window (The Communicator&#x2122;). We compared performance with these face masks to baseline conditions with no mask. Audiovisual conditions were included to investigate the impact of face masks on communication under ideal conditions. Auditory-only conditions were included to evaluate performance under less-than-ideal conditions, based on limited evidence that young children do not consistently orient to the target talker the way adults do (<xref ref-type="bibr" rid="B50">Ricketts and Galster, 2008</xref>).</p>
<p>Based on previous studies and children&#x2019;s susceptibility to the detrimental effects of decreased auditory bandwidth, we expected all the masks to affect auditory-only performance in children with good high-frequency aided audibility, including CNH and some CHL. More specifically, we expected that reduced transmission of high-frequency speech content would affect the discrimination of consonants, especially consonant place of articulation (<xref ref-type="bibr" rid="B59">Stevens, 2000</xref>). Based on acoustic transmission data in previous studies, we expected hospital masks to affect auditory-only performance the least and transparent masks to affect auditory-only performance the most. In children with poor high-frequency aided audibility, we expected a smaller difference between scores in the auditory-only conditions. In audiovisual conditions, we expected opaque masks to reduce the transmission of visual speech cues, which would affect the discrimination of consonants, especially consonant place of articulation (<xref ref-type="bibr" rid="B4">Binnie et al., 1974</xref>; <xref ref-type="bibr" rid="B46">Owens and Blazek, 1985</xref>). As such, under audiovisual conditions, we expected children with poor high-frequency aided audibility to perform best with the transparent masks, because they benefit from visual speech cues (<xref ref-type="bibr" rid="B27">Lalonde and McCreery, 2020</xref>) and may be less impacted by the high-frequency attenuation caused by the masks. It is uncertain what to predict for CNH and CHL with good high-frequency aided audibility with respect to the trade-off between loss of visual cues and high-frequency attenuation. The results from this study will provide an evidence base for advising educators and others who work with children as to the best face masks for promoting speech understanding.</p>
</sec>
<sec id="S2">
<title>Overview, General Materials, and General Methods</title>
<p>This study was designed to test the effects of four types of face masks on auditory and audiovisual speech recognition in noise by ANH, CNH, and CHL. The four types of masks (<xref ref-type="fig" rid="F1">Figures 1B&#x2013;E</xref>) include a disposable hospital mask, a homemade pleated cloth mask with two layers of cotton blend, the Communicator&#x2122;, and the ClearMask&#x2122;. The Communicator&#x2122; is an FDA-registered single-use, disposable device that meets ASTM Level 1 hospital mask standards and includes a fog-resistant transparent window. The ClearMask&#x2122; is an FDA-approved, class II single-use transparent face mask that meets ASTM Level 3 standards. It serves the same function as traditional masks and provides a full, anti-fog plastic barrier.</p>
<p>We simulated the effects of face masks on audiovisual stimuli by filtering the acoustic speech from an unmasked talker to match the long-term average spectrum of each mask and by overlaying mask shapes onto videos of the unmasked talker. This method does not capture differences in production that talkers may adopt when wearing a face mask, but it has the advantage of ensuring that idiosyncrasies in production across conditions do not affect results.</p>
<p>Experiment 1 was conducted by delivering research equipment to the homes of CHL. CHL and members of their household completed auditory-only and audiovisual speech recognition tests in noise at 0 dB SNR. In Experiment 2, young ANH completed the same test in a laboratory sound booth. The effects of face masks vary depending on baseline difficulty; mask effects differ in subjects as a function of SNR (<xref ref-type="bibr" rid="B62">Toscano and Toscano, 2021</xref>). Therefore, the ANH in Experiment 2 were also tested at &#x2212;10 dB SNR to better match the performance level of the CHL tested in Experiment 1. Both experiments were reviewed and approved by the Boys Town National Research Hospital Institutional Review Board. Adult participants provided their written informed consent to participate in the study; child participants and their caregivers provided written assent and permission, respectively.</p>
<sec id="S2.SS1">
<title>Stimuli</title>
<p>Target stimuli included 36 audiovisual recordings of a 32-year-old female native speaker of mainstream American English (author KL) repeating CV words with the vowel /i/ and the format, &#x201C;Choose /CV/&#x201D;. The 12 CVs (/bi/, /si/, /di/, /hi/, /ki/, /mi/, /ni/, /pi/, /&#x222B;i/, /ti/, /vi/, and /zi/) were the same CVs used in a previous study evaluating auditory-only speech perception in children (<xref ref-type="bibr" rid="B31">Leibold and Buss, 2013</xref>), based on the Audiovisual Feature Test for Young Children (<xref ref-type="bibr" rid="B65">Tyler et al., 1991</xref>). Three tokens of each CV were used so that idiosyncratic differences between the videos (e.g., blinking) could not be used to discriminate the syllables. Including the carrier, these recordings had a mean duration of 856 ms (728&#x2013;941 ms) and a mean F0 of 238 Hz (212&#x2013;276 Hz).</p>
<p>The original unmasked target stimuli were professionally recorded in a large sound-attenuating booth with a Sennheiser EW 112-P G3-G lapel microphone and a JVC GY-GM710U video camera. Final Cut Pro video editing software was used to splice the recordings into individual videos, which began 333 ms (10 frames) before the onset of acoustic speech and ended 333 ms after the end of the acoustic speech, with the exception that additional frames may have been added to avoid beginning or ending the video mid-blink. Using Adobe Audition, the intensity of each sound file was adjusted so that the total RMS of the speech portion of all the files matched.</p>
<sec id="S2.SS1.SSS1">
<title>Acoustic Mask Simulation</title>
<p>Target stimuli in the four mask conditions were generated by filtering the original (no mask) target recordings. Filters were constructed based on recordings of the rainbow passage (<xref ref-type="bibr" rid="B16">Fairbanks, 1969</xref>), produced by the target talker, each lasting 1.7&#x2013;1.9 min. Two recordings were made using a Shure-KSM42 (condenser) microphone in each of five conditions: no mask, hospital mask, fabric mask, ClearMask&#x2122;, and Communicator&#x2122;. The two recordings for each condition were concatenated, and the result was transformed into the frequency domain using the <italic>pwelch</italic> function in MATLAB (Mathworks), with a 512-point window and a 256-point overlap between sequential windows. Attenuation as a function of frequency for each mask was determined as the difference in amplitude spectrum relative to the no-mask condition. These attenuation functions were used to generate 128-point FIR filters using the <italic>fir2</italic> function in MATLAB. An all-pass filter was generated for use with the no-mask stimuli; this was done so the stimulus preparation steps were identical across stimuli. The amplitude spectrum of the no-mask recordings was also used to construct a filter for generating speech-shaped noise, following the same procedures. Stimuli were filtered using the <italic>filter</italic> function in MATLAB and saved to a disk. Long-term average magnitude spectra for targets in each condition are shown in <xref ref-type="fig" rid="F2">Figure 2A</xref>. The speech-shaped noise was 30 s in duration, and its long-term average magnitude spectrum (not shown) was similar to that of the no-mask targets.</p>
<fig id="F2" position="float">
<label>FIGURE 2</label>
<caption><p><bold>(A)</bold> Long-term magnitude spectra for stimuli in each of the five conditions. <bold>(B)</bold> Example target stimulus waveform after filtering for the no mask (blue) and fabric mask (green) conditions. Boxes indicate consonant regions.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-874345-g002.tif"/>
</fig>
<p>As illustrated in <xref ref-type="fig" rid="F2">Figure 2A</xref>, power spectra in the four mask conditions were similar to the power spectrum of the no-mask to within approximately &#x00B1;6 dB up to 2 kHz. Mean attenuation between 2 and 16 kHz was 2.4 dB for the hospital mask, 5.5 dB for the Communicator&#x2122;, 5.9 dB for the ClearMask&#x2122;, and 8.2 dB for the fabric mask. Peak attenuation within a 1/3-octave-wide band by mask type was 5.5 dB at 8 kHz for the hospital mask, 14.9 dB at 3.6 kHz for the Communicator&#x2122;, 11.9 dB at 16 kHz for the ClearMask&#x2122;, and 12.7 dB at 3.2 kHz for the fabric mask. <xref ref-type="fig" rid="F2">Figure 2B</xref> shows an example stimulus waveform after filtering for no mask (blue) and fabric mask (green) conditions. The mask causes greater attenuation of the consonants than the vowels. We observed similar acoustic differences between the mask and no mask conditions in recordings of the talker wearing a mask.</p>
</sec>
<sec id="S2.SS1.SSS2">
<title>Visual Mask Simulation</title>
<p>To visually simulate face masks, video files were processed using specialized software that automatically determines the position of 66 points on the face in each frame of the videos (<xref ref-type="bibr" rid="B52">Saragih et al., 2011</xref>; available at <ext-link ext-link-type="uri" xlink:href="http://osf.io/g9jtr/">osf.io/g9jtr/</ext-link>). These included 17 points on the perimeter of the lower half of the face, 18 points at the inner and outer rim of the lips, 6 points for each eye, 5 points for each eyebrow, 4 points on the bridge of the noise, and 5 points at the bottom of the nostrils.</p>
<p>The <italic>x</italic> and <italic>y</italic> coordinates of these points were imported into MATLAB. The Computer Vision Toolbox was used to superimpose filled white polygons onto each video frame. For opaque face masks (hospital, cloth), the size and shape of the polygon were created based on the location of the 17 points at the perimeter of the lower half of the face and the marker at the upper middle of the bridge of the nose. An example of the simulated opaque face mask is shown in <xref ref-type="fig" rid="F1">Figure 1G</xref>. The simulated Communicator&#x2122; mask was created in a similar fashion, except that two filled polygons were used to create the effect of a cutout in the middle of the simulated opaque mask shape. Adjustments to the simulated masks were made to correct for problems noted by independent observers (see <xref ref-type="supplementary-material" rid="TS1">Supplementary Material</xref>). An example of the simulated Communicator&#x2122; mask is shown in <xref ref-type="fig" rid="F1">Figure 1I</xref>. For the ClearMask&#x2122;, no mask was superimposed. The stimuli are available at <ext-link ext-link-type="uri" xlink:href="https://osf.io/5wapg">https://osf.io/5wapg</ext-link>.</p>
</sec>
</sec>
</sec>
<sec id="S3">
<title>Experiment 1</title>
<p>The first experiment was conducted remotely in participants&#x2019; homes. The goal was to evaluate auditory-only and audiovisual speech recognition in the five conditions. Remote data collection made it possible to collect data without bringing the subjects to the laboratory. Collecting data from members of the CHL&#x2019;s household increased the efficiency of the protocol and reduced the variance of test conditions across participant groups.</p>
<sec id="S3.SS1">
<title>Materials and Methods</title>
<sec id="S3.SS1.SSS1">
<title>Participants</title>
<p>Eighteen children with bilateral hearing loss (8 males) and 39 members of their households completed a remote study. CHL varied in age, between 7.4 and 18.9 years (mean = 12.7 years, SD = 3 years). Sixteen had sensorineural hearing loss, and two had mixed hearing loss. Audiograms and aided Speech Intelligibility Index (SII) scores at 65 and 75 dB SPL were available from 17 children&#x2019;s previous research visits, 3&#x2013;24 months before data were collected for this study (median 7.5 months). An audiogram (but not aided SII scores) was acquired from the other child&#x2019;s previous clinical visit 7.5 months before data were collected for this study. Mean better-ear aided SII was 72.1 at 55 dB SPL, 80.8 at 65 dB SPL, and 81.2 at 75 dB SPL. The distribution of better-ear aided SII scores and audiograms for each of the CHL are shown in <xref ref-type="fig" rid="F3">Figure 3</xref>. Medical records indicate that 11 children had congenital hearing loss. Five passed their newborn hearing screening and had hearing loss identified between age 4 and 7 years. Two others had hearing loss identified at age 2 or 3, but no newborn hearing screening results were reported. Family members of the CHL included 16 children with parent-reported normal hearing (9 males) between 7.5 and 19.8 years of age (mean = 12.1 years, SD = 3.5 years) and 23 adults with self-reported normal hearing (8 male) between 33.6 and 50.7 years of age (mean = 43 years, SD = 4.3 years). No audiometric data were available for family members of the CHL. Sample sizes are consistent with previous studies involving CHL (e.g., <xref ref-type="bibr" rid="B29">Leibold et al., 2013</xref>, <xref ref-type="bibr" rid="B30">2019</xref>; <xref ref-type="bibr" rid="B27">Lalonde and McCreery, 2020</xref>).</p>
<fig id="F3" position="float">
<label>FIGURE 3</label>
<caption><p>Values of <bold>(A)</bold> SII and <bold>(B)</bold> pure-tone thresholds for children with hearing loss (CHL). <bold>(A)</bold> The distribution of SII is plotted as a function of stimulus level, with results shown separately for the better and worse ears, as indicated at the top of the panel. Horizontal lines indicate the median, boxes span the 25th to 75th percentiles, and vertical lines span the 10th to 90th percentiles. <bold>(B)</bold> Pure-tone thresholds for the better and worse ears are shown. Better and worse ear were identified for each subject based on SII55. Better-ear thresholds are shown in the left panel, and worse-ear thresholds are shown in the right panel. Individual data are shown in gray, and means are shown with thick black lines.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-874345-g003.tif"/>
</fig>
</sec>
<sec id="S3.SS1.SSS2">
<title>Apparatus</title>
<p>The experiments were conducted using custom software on a 13&#x2033; MacBook Air laptop (MacOS Catalina 10.15.6). Auditory stimuli were routed directly from the laptop to two loudspeakers (QSC CP8 compact powered 90&#x00B0;) <italic>via</italic> a breakout cable with dual XLR inserts. The equipment was placed on a mat that was marked to show the intended location of the laptop and speakers (see <xref ref-type="supplementary-material" rid="TS1">Supplementary Figure 3</xref>). Speakers were placed 23 cm away from the edge of the laptop and oriented toward the participants&#x2019; ears. A manual explaining how to set up the equipment and run the experiment was also provided.</p>
</sec>
<sec id="S3.SS1.SSS3">
<title>Task</title>
<p>The participants completed a closed-set consonant identification task in speech-shaped noise using a picture selection response. On each trial, the noise was presented at 70 dB SPL, and the target was presented at 0 dB SNR. The noise had 20-ms ramps at onset and offset. The videos began 186 ms after the initiation of noise onset and &#x2265;333 ms before the onset of speech-related movements, ensuring that each video began with a neutral face. On average, acoustic speech began 756 ms (SD = 11 ms) after the offset of the noise ramp. The participants identified the CV token from a matrix of 12 alphabetically ordered illustrations that appeared immediately after the stimulus (<xref ref-type="supplementary-material" rid="TS1">Supplementary Figure 4</xref>). The same illustrations and similar methods were used in a previous experiment with children as young as 5 years old (<xref ref-type="bibr" rid="B31">Leibold and Buss, 2013</xref>). In each trial, the stimulus was presented once; there was no option to repeat the stimulus. Participants were instructed to take their best guess if they were unsure.</p>
<p>The participants completed testing in 11 randomly ordered conditions (four mask conditions conducted in auditory-only and audiovisual modalities and the no mask condition in auditory-only, audiovisual, and visual-only modalities). The five conditions included a no-mask baseline (<xref ref-type="fig" rid="F1">Figures 1A,F</xref>) and simulations of speech produced with a hospital mask (<xref ref-type="fig" rid="F1">Figures 1B,G</xref>), a fabric mask (<xref ref-type="fig" rid="F1">Figures 1C,H</xref>), and two transparent masks [ClearMask&#x2122; (<xref ref-type="fig" rid="F1">Figures 1E,J</xref>) and Communicator&#x2122; (<xref ref-type="fig" rid="F1">Figures 1D,I</xref>)]. In audiovisual conditions, the participants saw a synchronous and congruent video of the speaker, with or without a simulated face mask. In auditory-only conditions, the participants saw a blank gray screen during auditory stimulus presentation. The 36 consonant stimuli (12 CVs, 3 samples each) were presented in random order in a block of trials.</p>
</sec>
<sec id="S3.SS1.SSS4">
<title>Procedure</title>
<p>Informed consent was obtained by a research audiologist by videoconferencing <italic>via</italic> the HIPAA-protected WebEx Internet-based application and electronic consent forms developed in the Internet-based software called Research Electronic Data Capture (REDCap). All the participating family members were consented together, and each individually signed their forms electronically. Following completion of all consent paperwork, the audiologist provided an overview of the instructions. The participants were instructed to complete the testing in a quiet space free from distraction and to listen carefully to the woman who appeared on the screen. The audiologist arranged a time to deliver test equipment to the participants&#x2019; homes and provided her contact information so that she could be available for troubleshooting.</p>
<p>The participants received a manual with the test equipment. In the manual, parents were instructed to choose the most technologically savvy and most available participant in the home to participate first, to ensure the equipment was set up properly, and that the person could assist all other participants at home if needed. Furthermore, the participants were instructed to set up the equipment in a quiet space, away from all other participants, to avoid premature exposure to the stimuli. The recommended equipment configuration was to place the mat on a long table (at least 1.52 m &#x00D7; 0.76 m). There were cases in which a family member did not have a table large enough on which to set the equipment, so they sat with the equipment on the floor. Once the equipment was in place, parent participants were asked to test each speaker with a feature on the user interface. To ensure appropriate calibration, the participants were instructed to keep the laptop set at full volume. The speakers were also fitted with a custom 3D printed plastic barrier that prevented the participants from adjusting the speaker volume. It was recommended that the equipment stay in place once it was set up.</p>
<p>When they were ready for the test, the participants were instructed to sit directly in front of the laptop. Instructions specified that the CHL should wear their hearing aids in their typical configuration while completing the study. With the help of the manual, the participants were instructed to test themselves in the randomized order indicated on their individualized data collection sheet. Despite these instructions, 15 participants tested conditions in alphabetical order [some participants ran both sets of each condition in alphabetical order (AA, BB, CC, &#x2026;), while others ran through each condition in alphabetical order twice (ABC&#x2026;, ABC&#x2026;)]. This resulted in the following order: no mask, ClearMask&#x2122;, Communicator&#x2122;, fabric mask, and hospital mask, with each auditory condition preceding the corresponding audiovisual condition; the visual-only condition was last. There were also instances in which participants began testing and then stopped in the middle of a condition once they realized they were not following the order on their individualized datasheet. These incomplete conditions were excluded from the analysis. The parents were asked to assist their children during testing. Total test time was 1.25 to 2 h. Breaks were suggested for all the participants, but strongly encouraged for younger participants. Most of the participants completed the testing in one sitting; others split up the testing into two sessions. The participants were paid with an electronic gift card for their participation. A research audiologist was on stand-by remotely to answer questions or help with the protocol.</p>
</sec>
<sec id="S3.SS1.SSS5">
<title>Analyses</title>
<p>Analyses were performed in RStudio (version 1.2.1335). Linear mixed-effects models were fitted using the <italic>lmer</italic> and <italic>anova</italic> functions in the <italic>lmerTest</italic> package (<xref ref-type="bibr" rid="B23">Kuznetsova et al., 2017</xref>). The <italic>anova</italic> function provided <italic>F</italic>-statistics for the model generated using the <italic>lmer</italic> function. Non-significant interactions were systematically eliminated to arrive at the final model. Reference conditions were systematically varied as needed for <italic>post hoc</italic> comparisons. Auditory-only and audiovisual proportion correct data were transformed into rationalized arcsine units (RAUs) for statistical analyses.</p>
</sec>
</sec>
<sec id="S3.SS2">
<title>Results</title>
<p>Most subjects in Experiment 1 completed two runs of data collection in each condition, but there were exceptions. Because of a programming error, two subjects in each group heard the no-mask audiovisual condition at 20 dB SNR rather than 0 dB SNR. These data were omitted. The visual-only condition was added after the data collection commenced; as a result, data in the visual-only condition were only obtained for 16/23 ANH, 11/16 CNH, and 12/18 CHL. Of the remaining 603 cases (subjects &#x00D7; conditions, minus missing data), 27% included data from a single run, 72% from two runs, and only 1% from more than two runs. <xref ref-type="fig" rid="F4">Figure 4</xref> shows the distribution of mean performance plotted by stimulus condition for the three groups of subjects tested in Experiment 1.</p>
<fig id="F4" position="float">
<label>FIGURE 4</label>
<caption><p>Distribution of scores for each of the auditory-only and audiovisual stimulus conditions, plotted separately for each of the three groups of subjects tested in Experiment 1, as indicated at the top of each panel. Mask type is indicated on the horizontal axis, and cue condition is indicated with box fill, as defined in the legend. Vertical lines indicate the median, boxes span the 25th to 75th percentiles, and vertical lines span the 10th to 90th percentiles.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-874345-g004.tif"/>
</fig>
<p>Overall, auditory-only and audiovisual accuracies were poorer and more variable for CHL than for the two groups with normal hearing. These group differences were similar for auditory-only and audiovisual conditions. Audiovisual accuracy was higher than auditory-only accuracy, especially in the no mask, ClearMask&#x2122;, and Communicator&#x2122; conditions. The masks degrade performance relative to the no mask condition, but the impact varied across groups, modalities, and mask conditions.</p>
<p>The initial linear mixed-effects model examining RAU-transformed accuracy included fixed effects and interactions of mask condition, modality, and group, as well as a random intercept per subject. The baseline no mask condition, auditory-only modality, and CNH group served as the reference condition for the model. The final model included significant main effects of modality (<italic>F</italic><sub>1</sub>,<sub>916</sub> = 981.9, <italic>p</italic> &#x003C; 0.0001), mask condition (<italic>F</italic><sub>4</sub>,<sub>916</sub> = 176.6, <italic>p</italic> &#x003C; 0.0001), and group (<italic>F</italic><sub>2</sub>,<sub>55</sub> = 37.4, <italic>p</italic> &#x003C; 0.0001), as well as interactions of modality and mask condition (<italic>F</italic><sub>4</sub>,<sub>916</sub> = 72.1, <italic>p</italic> &#x003C; 0.0001) and modality and group (<italic>F</italic><sub>2</sub>,<sub>916</sub> = 18.9, <italic>p</italic> &#x003C; 0.0001). Model estimates and <italic>post hoc</italic> comparisons are shown in <xref ref-type="supplementary-material" rid="TS1">Supplementary Table 1</xref>. The CHL had lower speech recognition accuracy than the groups with normal hearing, in both the auditory-only and audiovisual modalities. The ANH and the CNH performed similarly in the auditory-only modality, but the ANH were more accurate than the CNH in the audiovisual modality. All the groups demonstrated significantly better overall performance in the audiovisual modality than in the auditory-only modality. The effect of modality differed across all the groups. The ANH exhibited greater audiovisual benefit (greater difference between the auditory-only and audiovisual modalities) than the CNH and the CHL. The CHL exhibited greater audiovisual benefit than the CNH.</p>
<p>The effect of mask type differed across modalities but not across groups. In the auditory-only modality, the ClearMask&#x2122;, Communicator&#x2122;, and fabric masks significantly degraded performance relative to the no-mask condition (<italic>B</italic> &#x2264; &#x2212;8.15, <italic>t</italic><sub>916</sub> &#x2264; &#x2212;6.337, <italic>p</italic> &#x003C; 0.0001), but the hospital mask did not (<italic>B</italic> = &#x2212;2.17, <italic>t</italic><sub>916</sub> = &#x2212;1.7, <italic>p</italic> = 0.0895). The masks differed significantly in the severity of degradation. The fabric mask had the greatest detrimental impact on performance, and the hospital mask had the least impact; the Communicator&#x2122; mask had a greater detrimental impact than the ClearMask&#x2122;. This pattern of results is consistent with the degree of high-frequency acoustic attenuation shown in <xref ref-type="fig" rid="F2">Figure 2</xref>. In the audiovisual modality, all four of the masks degraded performance relative to the no-mask condition (<italic>B</italic> &#x2264; &#x2212;7.23, <italic>t</italic><sub>916</sub> &#x2264; &#x2212;5.391, <italic>p</italic> &#x003C; 0.0001). The opaque masks (hospital and fabric) had a greater detrimental impact than the transparent masks (ClearMask&#x2122;, Communicator&#x2122;) due to loss of visual cues. The fabric mask resulted in the greatest degradation in audiovisual conditions. Significant audiovisual benefit was observed for the no mask, Communicator&#x2122;, and ClearMask&#x2122; conditions, and did not differ among these conditions. No significant audiovisual benefit was observed for the hospital and fabric mask conditions with the CNH as the reference group. Although there was no significant three-way interaction, audiovisual benefit was observed for the hospital and fabric mask conditions with the CHL or the ANH as the reference group.</p>
<p>Although this study was not designed to examine individual differences, we explored the impact of child age on the results reported above. Overall, older children recognized speech in noise with greater accuracy than younger children. An initial linear mixed-effects model examining RAU-transformed accuracy included fixed effects of mask condition, modality, group (CNH and CHL), and age; three-way interactions of modality, age, and group and modality, age, and mask condition; and a random intercept per subject. In addition to the effects and interactions of condition, modality, and group noted above, the model indicated a significant effect of age (<italic>F</italic><sub>1,32</sub> = 13.5, <italic>p</italic> = 0.0009), an interaction of age and modality (<italic>F</italic><sub>1,517</sub> = 5.5, <italic>p</italic> = 0.0193), and a marginal three-way interaction of age, modality, and mask condition (<italic>F</italic><sub>4,517</sub> = 2.1, <italic>p</italic> = 0.0802). Age effects did not significantly differ between the CNH and CHL groups. The <italic>post hoc</italic> comparisons demonstrated that age effects were present in every condition (<italic>B</italic> &#x2265; 1.6, <italic>t</italic><sub>86</sub> &#x2265; 2.648, <italic>p</italic> &#x2264; 0.0097) except for the auditory-only no mask and auditory-only hospital mask conditions (<italic>B</italic> &#x2264; 1, <italic>t</italic><sub>86</sub> &#x2264; 1.795, <italic>p</italic> &#x2264; 0.0762). In auditory-only conditions, younger children were more negatively impacted by the ClearMask&#x2122;, Communicator&#x2122;, and fabric mask than older children (<italic>B</italic> &#x2265; 1.1, <italic>t</italic><sub>517</sub> &#x2264; 2.065, <italic>p</italic> &#x2264; 0.0395). This could indicate greater bandwidth requirements in younger children, but ceiling effects could produce this pattern of results. These ceiling effects, along with limited representation of older children in our sample (<italic>n</italic> = 6 age 15 and higher), prohibit strong conclusions about how the negative impact of masks varies from 7 to 19 years of age.</p>
</sec>
</sec>
<sec id="S4">
<title>Experiment 2</title>
<p>The second experiment was conducted in a sound booth in the laboratory. The goal of Experiment 2 was twofold. One goal was to replicate the results obtained remotely from the ANH under more controlled conditions. Another goal was to obtain data from adults at a more challenging SNR (&#x2212;10 dB) to approximately match the overall performance level of CHL tested at 0 dB SNR.</p>
<sec id="S4.SS1">
<title>Materials and Methods</title>
<sec id="S4.SS1.SSS1">
<title>Participants</title>
<p>Fifteen ANH (3 males) between 19 and 28 years of age (mean = 22.3 years, SD = 2.4 years) participated. All the participants passed a pure-tone hearing screening bilaterally at 20 dB HL at octave intervals from 0.25 to 8 kHz. They also demonstrated at least 20/30 vision bilaterally, with or without corrective lenses, based on a Snellen eye chart screening. Ten subjects provided data at 0 dB SNR and ten at &#x2212;10 dB SNR. The first five subjects tested at 0 dB SNR provided pilot data at &#x2212;8 dB SNR. Five subjects were tested at both 0 dB SNR and &#x2212;10 dB SNR, and another five subjects were tested only at &#x2212;10 dB SNR. These sample sizes were determined based on previous studies with similar comparisons to CNH and/or CHL (<xref ref-type="bibr" rid="B31">Leibold and Buss, 2013</xref>; <xref ref-type="bibr" rid="B27">Lalonde and McCreery, 2020</xref>).</p>
</sec>
<sec id="S4.SS1.SSS2">
<title>Apparatus, Task, and Procedures</title>
<p>The task, apparatus, and procedures were the same as in Experiment 1 but with a few exceptions. The experiment was completed in a laboratory setting. The equipment was set up in a sound booth by an experimenter in the same configuration as in Experiment 1, except that two JBL Professional IRX108BT portable powered loudspeakers (James B. Lansing Sound, Inc.) were used. Written consent was obtained in person by a research assistant who also provided an oral description of the study at the start of the session. The participants were tested in one or two sessions, depending on the number of SNRs in which they completed testing. During each session, the participants completed each of the four masks and no mask conditions in auditory-only and audiovisual conditions twice. In two cases, a participant mistakenly completed a condition three times. During one session (randomized across subjects), the participants also completed the visual-only condition twice. Two participants completed the visual-only condition during both sessions. The order of SNRs and conditions was randomized across the participants. Two subjects completed the conditions in alphabetical order when tested at 0 dB SNR.</p>
</sec>
</sec>
<sec id="S4.SS2">
<title>Results</title>
<p><xref ref-type="fig" rid="F5">Figure 5</xref> shows the distribution of performance for ANH tested in the lab, plotted by stimulus condition. The results for subjects tested at &#x2212;10 dB and 0 dB SNR are shown in the right and left panels, respectively. A linear mixed-effects model with a random intercept per subject was used to examine the effects of SNR, mask condition, modality, and their interactions on consonant recognition. Model estimates and <italic>post hoc</italic> comparisons are shown in <xref ref-type="supplementary-material" rid="TS1">Supplementary Table 2</xref>. The final model included significant interactions of SNR and modality (<italic>F</italic><sub>1</sub>,<sub>372</sub> = 46.4, <italic>p</italic> &#x003C; 0.0001), SNR, and mask condition (<italic>F</italic><sub>4</sub>,<sub>372</sub> = 13.1, <italic>p</italic> &#x003C; 0.0001), and modality and mask condition (<italic>F</italic><sub>4</sub>,<sub>372</sub> = 25.8, <italic>p</italic> &#x003C; 0.0001).</p>
<fig id="F5" position="float">
<label>FIGURE 5</label>
<caption><p>Distribution of scores for each of the auditory-only and audiovisual stimulus conditions, plotted separately for each of the two SNRs tested in Experiment 2, as indicated at the top of each panel. Mask type is indicated on the horizontal axis, and cue condition is indicated with box fill, as defined in the legend. Vertical lines indicate the median, boxes span the 25th to 75th percentiles, and vertical lines span the 10th to 90th percentiles.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-874345-g005.tif"/>
</fig>
<p>The interactions of SNR with modality and SNR with mask condition reflect increases in the effects of mask and modality at &#x2212;10 dB SNR compared to 0 dB SNR, likely because ceiling effects were eliminated at the more difficult SNR. At &#x2212;10 dB SNR, auditory-only performance was degraded by all the mask types, with hospital masks having the smallest effect and fabric masks having the largest. At 0 dB SNR, auditory-only performance was degraded by the ClearMask&#x2122;, Communicator&#x2122;, and fabric mask, and was poorer for the fabric mask than for the ClearMask&#x2122;. At both SNRs, audiovisual performance was degraded by all the mask types, with the same pattern as in Experiment 1. At &#x2212;10 dB SNR, there was a significant audiovisual benefit in all the mask conditions. At 0 dB SNR, audiovisual benefit was significant in all mask conditions except the hospital mask. Audiovisual benefit was greater for the transparent masks and no mask conditions as compared to the opaque masks. The significant interaction of mask and modality reflects reduced audiovisual benefit for the two opaque masks, as observed in Experiment 1.</p>
<p>Comparing the ANH tested in the lab at 0 dB SNR (<xref ref-type="fig" rid="F5">Figure 5</xref>, right) to the ANH tested remotely at the same SNR (<xref ref-type="fig" rid="F4">Figure 4</xref>, left), there was a small but significant overall effect of group (<italic>F</italic><sub>1</sub>,<sub>31</sub> = 5.9, <italic>p</italic> = 0.0215) and an interaction between group and modality (<italic>F</italic><sub>1</sub>,<sub>572</sub> = 31.9, <italic>p</italic> &#x003E; 0.0001). The group effect was significant in the auditory-only condition (<italic>B</italic> = 10.74, <italic>t</italic><sub>35</sub> = 3.781, <italic>p</italic> = 0.0005), but not in the audiovisual condition (<italic>B</italic> = 2.58, <italic>t</italic><sub>35</sub> = 0.909, <italic>p</italic> = 0.3696). No other variables interacted with group. The poorer performance of participants tested in their homes may reflect a limited control of the test environment. It may also reflect the difference in age of the participants, as the adults tested at home had a mean age of 43 years (SD = 4.3 years), and the adults tested in the lab had a mean age of 22.3 years (SD = 2.4 years). Additionally, we tested hearing detection thresholds in the lab but not at home, so it is possible that the difference is due to inaccurate self-reports of normal hearing among the adults tested at home.</p>
<p>The lower overall performance of the ANH tested at &#x2212;10 dB SNR allows for a performance-matched comparison to the CHL. A mixed-effects linear model was used with fixed effects and interactions of group, mask, and modality, and a random intercept per subject. From this model, we examined the effects of group and interactions between group and each within-subjects variable. Model estimates and <italic>post hoc</italic> comparisons are shown in <xref ref-type="supplementary-material" rid="TS1">Supplementary Table 3</xref>. There was a marginal overall group effect (<italic>F</italic><sub>1,26</sub> = 3.5, <italic>p</italic> = 0.0723), with the ANH at &#x2212;10 dB SNR performing slightly worse overall than the CHL at 0 dB SNR. The group effect was not significant in the auditory-only baseline condition (<italic>B</italic> = &#x2212;6.23, <italic>t</italic><sub>41</sub> = &#x2212;1.38, <italic>p</italic> = 0.1724), but there were significant interactions of modality and mask condition (<italic>F</italic><sub>4,471</sub> = 39.2, <italic>p</italic> &#x003C; 0.0001), modality and group (<italic>F</italic><sub>1,471</sub> = 13.4, <italic>p</italic> = 0.0003), and mask and group (<italic>F</italic><sub>4,471</sub> = 8.1, <italic>p</italic> &#x003C; 0.0001). The interaction of modality and group was driven by greater audiovisual benefit in the ANH than in the CHL (<italic>B</italic> = 6.62, <italic>t</italic><sub>471</sub> = 3.66, <italic>p</italic> = 0.0003). The interaction of group and mask condition was driven by the fact that the fabric mask degraded performance more in the ANH than in the CHL (<italic>B</italic> = 13.91, <italic>t</italic><sub>471</sub> = 4.833, <italic>p</italic> &#x003C; 0.0001). The interaction of fabric mask and group is consistent with our hypothesis that the CHL may be less impacted by the acoustic attenuation of the face mask, because their hearing loss reduces access to high-frequency cues even in the no-mask condition. However, there is no relationship between individual differences in relative degradation caused by the fabric mask and either CHL&#x2019;s aided SII or 4 kHz thresholds. This could be due to the delay between audiometric testing and the study tasks. It is also possible that this relationship may emerge with a larger sample of CHL.</p>
</sec>
</sec>
<sec id="S5">
<title>Auditory-Only and Audiovisual Phonetic Feature Transmission</title>
<p>We conducted additional analyses to determine which speech features were impacted by the face masks. Using data from the CHL and ANH tested at &#x2212;10 dB SNR, we examined the effects of face masks on the perception of three speech features: voicing, place of articulation, and manner of articulation. <xref ref-type="table" rid="T1">Table 1</xref> shows feature classification for each consonant. Other data sets were excluded from this analysis because of the common occurrence of ceiling performance. <xref ref-type="fig" rid="F6">Figure 6</xref> demonstrates the mean and distribution of voicing, place, and manner feature transmission accuracy in each mask condition in auditory-only (top) and audiovisual (bottom) conditions for the CHL (left) and the ANH (right). The full consonant accuracy data are provided in gray for reference.</p>
<table-wrap position="float" id="T1">
<label>TABLE 1</label>
<caption><p>Phonetic feature assignment.</p></caption>
<table cellspacing="5" cellpadding="5" frame="hsides" rules="groups">
<thead>
<tr>
<td valign="top" align="left" colspan="3">Voicing</td>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Voiced</td>
<td valign="top" align="left" colspan="2">Unvoiced</td>
</tr>
<tr>
<td valign="top" align="left">b, d, m, n, v, z</td>
<td valign="top" align="left" colspan="2">s, h, k, t, p, &#x222B;</td>
</tr>
<tr>
<td valign="top" align="left" colspan="3"><hr/></td>
</tr>
<tr>
<td valign="top" align="left" colspan="3"><bold>Manner</bold></td>
</tr>
<tr>
<td valign="top" align="left">Stop</td>
<td valign="top" align="left">Fricative</td>
<td valign="top" align="left">Nasal</td>
</tr>
<tr>
<td valign="top" align="left">b, d, p, t, k</td>
<td valign="top" align="left">v, z, s, h, &#x222B;</td>
<td valign="top" align="left">m, n</td>
</tr>
<tr>
<td valign="top" align="left" colspan="3"><hr/></td>
</tr>
<tr>
<td valign="top" align="left" colspan="3"><bold>Place</bold></td>
</tr>
<tr>
<td valign="top" align="left">Front</td>
<td valign="top" align="left">Middle</td>
<td valign="top" align="left">Back</td>
</tr>
<tr>
<td valign="top" align="left">p, b, m, v</td>
<td valign="top" align="left">s, z, t, d, n</td>
<td valign="top" align="left">&#x222B;, k, h</td>
</tr>
</tbody>
</table>
</table-wrap>
<fig id="F6" position="float">
<label>FIGURE 6</label>
<caption><p>Mean and standard deviation of scores for each of the auditory-only and audiovisual stimulus conditions, plotted separately for each group and modality, as indicated at the top and right of each panel, respectively. Mask type is indicated on the horizontal axis, and articulatory feature is indicated with color, as defined in the legend.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-874345-g006.tif"/>
</fig>
<p>For each modality and group, a linear mixed-effects model with a random intercept per subject was used to analyze the effects and interactions of feature and mask conditions on RAU-transformed feature transmission data. The no mask condition and place feature served as the reference condition. Reference conditions were systematically varied as needed for <italic>post hoc</italic> comparisons.</p>
<p>In both groups, the auditory-only transmission was best for the voicing feature and worst for the place feature. Rank differences between the face mask conditions for each phonetic feature were similar to the differences in consonant recognition accuracy. These findings were confirmed by two statistical models. The final model of the feature transmission data in the CHL included main effects of mask (<italic>F</italic><sub>4</sub>,<sub>453</sub> = 20.6, <italic>p</italic> &#x003C; 0.0001) and feature (<italic>F</italic><sub>2</sub>,<sub>453</sub> = 423.6, <italic>p</italic> &#x003C; 0.0001) and no significant interaction. Accuracy was lower for place than manner (<italic>t</italic><sub>453</sub> = 9.989, <italic>p</italic> &#x003C; 0.0001) and lower for manner than voicing (<italic>t</italic><sub>453</sub> = 18.682, <italic>p</italic> &#x003C; 0.001).</p>
<p>The final model for the ANH tested at &#x2212;10 dB SNR in auditory-only conditions included main effects of mask (<italic>F</italic><sub>4</sub>,<sub>279</sub> = 47.5, <italic>p</italic> &#x003C; 0.0001) and feature (<italic>F</italic><sub>2</sub>,<sub>279</sub> = 190.1, <italic>p</italic> &#x003C; 0.0001), and an interaction of mask and feature (<italic>F</italic><sub>8</sub>,<sub>279</sub> = 2.5, <italic>p</italic> = 0.0122). Model estimates and <italic>post hoc</italic> comparisons are shown in <xref ref-type="supplementary-material" rid="TS1">Supplementary Table 4</xref>. With no mask, accuracy was lower for place than manner (<italic>t</italic><sub>453</sub> = 9.989, <italic>p</italic> &#x003C; 0.0001) and lower for manner than voicing (<italic>t</italic><sub>453</sub> = 18.682, <italic>p</italic> &#x003C; 0.001). The interaction of mask condition and feature reflects the fact that face masks impacted perception of place more than manner and voicing. More specifically, the hospital mask only impacted perception of the place feature. The other masks impacted perception of all features, but the ClearMask&#x2122; and fabric masks impacted place perception more than voicing (<italic>B</italic> = 10.27, <italic>t</italic><sub>279</sub> = 2.648, <italic>p</italic> = 0.0086; <italic>B</italic> = 14.82, <italic>t</italic><sub>279</sub> = 3.822, <italic>p</italic> = 0.0002) and manner (<italic>B</italic> = 9.94, <italic>t</italic><sub>279</sub> = 2.563, <italic>p</italic> = 0.0109; <italic>B</italic> = 8.33, <italic>t</italic><sub>279</sub> = 2.149, <italic>p</italic> = 0.0325). The impact of the Communicator&#x2122; was also marginally greater for place than voicing (<italic>B</italic> = 7.06, <italic>t</italic><sub>279</sub> = 1.842, <italic>p</italic> = 0.0666). Finally, the impact of the fabric mask was marginally greater for manner than voicing (<italic>B</italic> = 6.49, <italic>t</italic><sub>279</sub> = 1.673, <italic>p</italic> = 0.0954).</p>
<p>Audiovisual feature transmission data from the CHL tested at 0 dB SNR and the ANH tested at &#x2212;10 dB SNR are shown in the bottom half of <xref ref-type="fig" rid="F6">Figure 6</xref>. Rank differences between the face mask conditions for each phonetic feature were similar to the differences in consonant recognition accuracy. However, there was a notable difference in the patterns of feature transmission between conditions in which the participants could see the talker&#x2019;s mouth region (no mask, ClearMask&#x2122;, and Communicator&#x2122; conditions) and conditions in which the mouth region was obscured (fabric and hospital mask conditions). These findings were confirmed with two statistical models.</p>
<p>The final audiovisual model for the CHL included effects of mask (<italic>F</italic><sub>4,427</sub> = 72.8, <italic>p</italic> &#x003C; 0.0001) and feature (<italic>F</italic><sub>2,427</sub> = 101.2, <italic>p</italic> &#x003C; 0.0001), and an interaction of mask condition and feature (<italic>F</italic><sub>8,427</sub> = 8.1, <italic>p</italic> &#x003C; 0.0001). Model estimates and <italic>post hoc</italic> comparisons are shown in <xref ref-type="supplementary-material" rid="TS1">Supplementary Table 5</xref>. The interaction reflects the fact that the difference in transmission between conditions in which the participants can see the talker&#x2019;s mouth region and conditions in which the mouth region is obscured was larger for place than voicing and manner (<italic>B</italic> &#x2265; 13.65, <italic>t</italic><sub>427</sub> &#x2265; 3.389, <italic>p</italic> &#x2264; 0.0007).</p>
<p>The final audiovisual model for the ANH tested at &#x2212;10 dB SNR included the effects of mask (<italic>F</italic><sub>4,276</sub> = 135.3, <italic>p</italic> &#x003C; 0.0001) and feature (<italic>F</italic><sub>2,276</sub> = 50.2, <italic>p</italic> &#x003C; 0.0001). The effect of feature varied significantly across mask conditions (<italic>F</italic><sub>8,276</sub> = 10.7, <italic>p</italic> &#x003C; 0.0001). Model estimates and <italic>post hoc</italic> comparisons are shown in <xref ref-type="supplementary-material" rid="TS1">Supplementary Table 6</xref>. As in the CHL, the difference in transmission between conditions in which the participants could see the talker&#x2019;s mouth region and conditions in which the mouth region was obscured was larger for place than voicing and manner (<italic>B</italic> &#x2265; 12.76, <italic>t</italic><sub>276</sub> &#x2265; 3.135, <italic>p</italic> &#x2264; 0.0019). The difference between the fabric mask and the two transparent masks was also greater for manner than voicing (<italic>B</italic> &#x2265; 11.65, <italic>t</italic><sub>276</sub> &#x2265; 2.861, <italic>p</italic> &#x2264; 0.0045).</p>
</sec>
<sec id="S6">
<title>Visual-Only Phonetic Feature Transmission</title>
<p>Visual-only data were obtained for 11/16 CNH, 12/18 CHL, and 16/23 ANH in Experiment 1, as well as 15/15 ANH in Experiment 2. Visual-only consonant recognition accuracy data are plotted in gray in <xref ref-type="fig" rid="F7">Figure 7</xref>. Mean accuracy was 31% for the group of CNH, 33% for the group of CHL, and 42% for both ANH groups. Welch&#x2019;s <italic>t</italic>-tests indicated greater visual-only consonant recognition accuracy in the ANH than in the CNH (<italic>t</italic><sub>24</sub> = 4.695, <italic>p</italic> &#x003C; 0.0001) and the CHL (<italic>t</italic><sub>26</sub> = 3.257, <italic>p</italic> = 0.0031) but no difference between the CNH and the CHL (<italic>t</italic><sub>41</sub> = 0.286, <italic>p</italic> = 0.7761).</p>
<fig id="F7" position="float">
<label>FIGURE 7</label>
<caption><p>Mean and standard deviation of scores for the visual-only condition. Group is indicated on the horizontal axis, and phonetic feature is indicated with color, as defined in the legend.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-874345-g007.tif"/>
</fig>
<p>Visual-only feature transmission data are plotted in color in <xref ref-type="fig" rid="F7">Figure 7</xref>. This figure shows higher transmission of the place feature than the voicing and manner features and higher transmission in the ANH than in the CNH and CHL. Additionally, there was large variability in place transmission among children.</p>
<p>A linear mixed-effects model with a random intercept per subject was used to analyze the effects and interactions of feature and group on RAU-transformed visual-only feature transmission data. We combined the adults tested at home and the adults tested in the lab into one group. The CHL group and place feature served as the reference condition. Reference conditions were systematically varied as needed for <italic>post hoc</italic> comparisons. Model estimates and <italic>post hoc</italic> comparisons are shown in <xref ref-type="supplementary-material" rid="TS1">Supplementary Table 7</xref>. There were significant effects of group (<italic>F</italic><sub>2,51</sub> = 12.4, <italic>p</italic> &#x003C; 0.0001) and feature (<italic>F</italic><sub>2,269</sub> = 175, <italic>p</italic> &#x003C; 0.0001), as well as an interaction of group and feature (<italic>F</italic><sub>4,269</sub> = 6.7, <italic>p</italic> &#x003C; 0.0001). In all the groups, visual-only feature transmission was higher for place than for voicing (<italic>B</italic> &#x2264; &#x2212;11.42, <italic>t</italic><sub>269</sub> &#x2265; &#x2212;4.321, <italic>p</italic> &#x003C; 0.0001) and manner (<italic>B</italic> &#x2264; &#x2212;19.18, <italic>t</italic><sub>269</sub> &#x2264; &#x2212;7.254, <italic>p</italic> &#x003C; 0.0001), and higher for voicing than for manner (<italic>B</italic> &#x2265; 4.98, <italic>t</italic><sub>269</sub> &#x2265; 2.933, <italic>p</italic> &#x2264; 0.0036). However, it is important to note the lack of independence between features (for example, all nasal sounds are voiced), and that the chance of correct transmission of voicing is greater than the chance of correct transmission of place and manner.</p>
<p>Feature transmission accuracy was higher in the ANH than in the CNH for place (<italic>B</italic> = 19.54, <italic>t</italic><sub>103</sub> = 6.155, <italic>p</italic> &#x003C; 0.0001), manner (<italic>B</italic> = 9.78, <italic>t</italic><sub>103</sub> = 3.081, <italic>p</italic> = 0.0026), and voicing (<italic>B</italic> = 7.01, <italic>t</italic><sub>103</sub> = 2.207, <italic>p</italic> = 0.0295). Feature transmission accuracy was higher in the adults than in the CHL for place (<italic>B</italic> = 13.63, <italic>t</italic><sub>97</sub> = 4.519, <italic>p</italic> &#x003C; 0.0001) and manner (<italic>B</italic> = 7.148, <italic>t</italic><sub>97</sub> = 2.369, <italic>p</italic> = 0.0198). The difference between age groups was greater for place than for voicing (<italic>B</italic> &#x2265; 10.45, <italic>t</italic><sub>269</sub> &#x2265; 3.707, <italic>p</italic> &#x2264; 0.0003) and manner (<italic>B</italic> &#x2265; 6.49, <italic>t</italic><sub>269</sub> &#x2265; 2.302, <italic>p</italic> &#x2264; 0.0221). There were no differences in visual-only feature transmission between the CNH and the CHL.</p>
</sec>
<sec id="S7" sec-type="discussion">
<title>Discussion</title>
<p>The purpose of this study was to determine the impact of four types of face masks on auditory and audiovisual speech recognition in noise among CHL, CNH, and ANH. We compared performance with a hospital mask, a fabric mask, a primarily transparent mask (ClearMask&#x2122;), and a primarily opaque disposable mask with a transparent window (The Communicator&#x2122;) to baseline conditions with no mask. We expected the face masks to reduce transmission of high-frequency speech content and, thus, impact discrimination of consonants in noise, especially consonant features that rely on high-frequency acoustic cues, such as place of articulation. We hypothesized that the impact of high-frequency acoustic attenuation depends on audibility, such that there is little impact on the masks in CHL with poor high-frequency audibility. In audiovisual conditions, we also hypothesized that the loss of visual cues associated with the opaque masks would impact discrimination of consonants in noise, especially consonant features that are easier to speech-read, such as place of articulation. Thus, we expected the children with poor high-frequency audibility to do best with the transparent masks. However, given that previous studies have shown greater acoustic attenuation in transparent masks than opaque (hospital and fabric) masks (<xref ref-type="bibr" rid="B1">Atcherson et al., 2020</xref>; <xref ref-type="bibr" rid="B14">Corey et al., 2020</xref>; <xref ref-type="bibr" rid="B67">Vos et al., 2021</xref>), it was uncertain how the participants with good high-frequency audibility would fare in the trade-off between loss of visual cues and acoustic attenuation.</p>
<sec id="S7.SS1">
<title>Effects of Face Masks on Speech Acoustics</title>
<p>We measured the acoustic power spectra of speech produced by the talker at baseline and while wearing each mask. These acoustic measurements demonstrated that the masks caused high-frequency attenuation. As in previous studies, the hospital mask caused the least attenuation (e.g., <xref ref-type="bibr" rid="B1">Atcherson et al., 2020</xref>; <xref ref-type="bibr" rid="B5">Bottalico et al., 2020</xref>; <xref ref-type="bibr" rid="B14">Corey et al., 2020</xref>; <xref ref-type="bibr" rid="B18">Goldin et al., 2020</xref>; <xref ref-type="bibr" rid="B67">Vos et al., 2021</xref>). Although previous studies have shown greater high-frequency acoustic attenuation from transparent masks than from fabric masks (<xref ref-type="bibr" rid="B14">Corey et al., 2020</xref>; <xref ref-type="bibr" rid="B6">Brown et al., 2021</xref>), the fabric mask used in our study caused the greatest attenuation. The acoustic transmission properties of fabric masks vary considerably depending on type of fabric and number of fabric layers. One study showed 0.4- to 9-dB greater attenuation between 2 and 16 kHz for fabric masks than for a disposable hospital mask (<xref ref-type="bibr" rid="B14">Corey et al., 2020</xref>). The double-layered mask used in this study resulted in 7.4-dB greater attenuation between 2 and 16 kHz than the hospital mask, suggesting that our fabric mask is on the higher end of acoustic attenuation range.</p>
</sec>
<sec id="S7.SS2">
<title>Effects of Face Masks on Adults&#x2019; Speech Perception in Noise</title>
<p>In the perception experiments, we found that face masks degraded consonant recognition in noise. The relative impact of each mask was consistent across the groups of adults tested at home and in the lab, as well as across SNRs. The impact of the masks on auditory-only consonant recognition in noise varied in accordance with the high-frequency acoustic attenuation caused by the masks. The hospital mask had the least high-frequency acoustic attenuation and least impact on performance; the fabric mask had the greatest high-frequency acoustic attenuation and greatest impact on performance. The differences in the relative impact of the masks are consistent with other studies comparing between effects of fabric and hospital masks (<xref ref-type="bibr" rid="B5">Bottalico et al., 2020</xref>) or between effects of a hospital mask and the ClearMask&#x2122; (<xref ref-type="bibr" rid="B70">Yi et al., 2021</xref>) on auditory-only perception in noise or babble.</p>
<p>In audiovisual conditions, the two transparent masks (ClearMask&#x2122; and Communicator&#x2122;) had the least impact on performance, even though acoustic attenuation is greater for the transparent masks than for the hospital mask. Thus, in the tradeoff between loss of visual cues and high-frequency acoustic attenuation, visual cues seem to be more important for consonant recognition. One caveat is that the test configuration, a close, frontal view of the talker&#x2019;s face, short test sessions, and CV stimuli, may not be representative of listening in daily life. These results are consistent with a previous study that compared the effects of a hospital mask and the ClearMask&#x2122; on the perception of audiovisual speech in noise and babble (<xref ref-type="bibr" rid="B70">Yi et al., 2021</xref>). In that study, the results depended on the masker and speaking style. The ClearMask&#x2122; impacted the perception of clear speech in babble, conversational speech in noise, and conversational speech in babble but not the perception of clear speech in noise. The hospital mask impacted perception more than the ClearMask&#x2122; when speech was spoken in a clear style in babble, in a clear style in noise, or in a conversational style in noise but not in a conversational style in babble.</p>
<p>The finding that visual cues available with a transparent mask outweigh the additional acoustic attenuation diverges from published results (<xref ref-type="bibr" rid="B6">Brown et al., 2021</xref>). <xref ref-type="bibr" rid="B6">Brown et al. (2021)</xref> compared audiovisual sentence recognition accuracy among a hospital mask, a cloth mask with a paper filter, a cloth mask without a paper filter, and a cloth mask with a transparent window. In moderate (&#x2212;5 dB SNR) and high (&#x2212;9 dB SNR) levels of noise, the hospital mask impacted performance the least, followed by the cloth mask with no paper filter. The cloth mask with a paper filter and the cloth mask with a transparent window had the greatest impact. The cloth mask with a filter seemed to have caused slightly less attenuation than the transparent mask. Therefore, the fact that the transparent mask and the cloth mask with a filter similarly impacted performance may represent a small audiovisual benefit from the transparent window. However, unlike in our study, the benefit of the transparent window was not large enough to overcome the detrimental effects of high-frequency attenuation relative to the hospital mask. In other words, the availability of visual cues from the transparent window did not outweigh the added acoustic attenuation.</p>
<p>There are several differences between the <xref ref-type="bibr" rid="B6">Brown et al. (2021)</xref> study and this study that could account for the difference in results. First, the type of transparent mask differed between the two studies. Second, we simulated the effects of masks on speech acoustics and visual speech cues, whereas <xref ref-type="bibr" rid="B6">Brown et al. (2021)</xref> audio-visually recorded speech stimuli in each mask. Therefore, the participants in our study had consistent access to a full view of the mouth, whereas those in the previous study (and in natural conversations) may have been affected by improper mask placement and/or fogging. Furthermore, our stimuli did not capture phoneme-specific effects of the masks on articulation, including any targeted adjustments to articulation that talkers might make in response to wearing a mask (<xref ref-type="bibr" rid="B12">Cohn et al., 2021</xref>). This is important, because 60% of survey respondents report communicating differently when wearing a mask, including changing their manner of speaking, minimizing linguistic content, and using gestures, facial expressions, and eye contact more often and more purposefully than when communicating without a face mask (<xref ref-type="bibr" rid="B53">Saunders et al., 2020</xref>). Finally, we examined perception of consonants in CV syllables, whereas <xref ref-type="bibr" rid="B6">Brown et al. (2021)</xref> examined perception of words in sentences. Unlike this study, the stimuli used by <xref ref-type="bibr" rid="B6">Brown et al. (2021)</xref> include linguistic context and may include additional visual cues to prosody, such as rigid head movements (<xref ref-type="bibr" rid="B42">Munhall et al., 2004</xref>; <xref ref-type="bibr" rid="B15">Davis and Kim, 2006</xref>).</p>
</sec>
<sec id="S7.SS3">
<title>Effect of Face Masks on Children&#x2019;s Speech Perception in Noise</title>
<p>The effects of face masks on speech perception in noise for the CNH did not differ from those for the ANH. High-frequency acoustic attenuation impacted children&#x2019;s auditory speech sound recognition, consistent with previous studies on children&#x2019;s susceptibility to the detrimental effects of decreased bandwidth (<xref ref-type="bibr" rid="B57">Stelmachowicz et al., 2000</xref>, <xref ref-type="bibr" rid="B58">2001</xref>; <xref ref-type="bibr" rid="B41">Mlot et al., 2010</xref>; <xref ref-type="bibr" rid="B39">McCreery and Stelmachowicz, 2011</xref>) and with previous data on the impact of face masks on auditory-only speech recognition in CNH (<xref ref-type="bibr" rid="B17">Flaherty, 2021</xref>; <xref ref-type="bibr" rid="B55">Sfakianaki et al., 2021</xref>). Although we observed no effects of age group, it is possible that these effects would emerge in younger children. Our exploration of the impact of child age suggested that younger children may be more impacted by face masks in auditory-only conditions. However, it was impossible to separate age effects from ceiling effects in the current child data. Additional data are needed to determine how the negative impact of face masks varies as a function of age.</p>
<p>We expected that CHL with poor high-frequency audibility would be less impacted by high-frequency acoustic attenuation caused by the masks than children with good high-frequency audibility. However, we did not observe a difference in the impact of high-frequency acoustic attenuation between the CHL and the CNH. Overall, the CHL in this study had relatively good aided audibility, with a mean aided SII score of 80.8 at 65 dB SPL. It is possible that a more heterogeneous sample of CHL, including those with poorer high-frequency aided audibility, would have a smaller detrimental effect of face masks in the auditory-only condition.</p>
</sec>
<sec id="S7.SS4">
<title>Audiovisual Benefit With Face Masks</title>
<p>When the mouth region was visible, there was significant audiovisual benefit in all the groups, with the greatest benefit in the ANH and the least benefit in the CNH. The finding that CNH benefited less than ANH is consistent with previous research (<xref ref-type="bibr" rid="B68">Wightman et al., 2006</xref>; <xref ref-type="bibr" rid="B26">Lalonde and Holt, 2016</xref>; <xref ref-type="bibr" rid="B27">Lalonde and McCreery, 2020</xref>). However, the finding that CHL benefited less than ANH conflicts with our previous study, which showed similar benefits in ANH and CHL (<xref ref-type="bibr" rid="B27">Lalonde and McCreery, 2020</xref>).</p>
<p>In this study, we found that audiovisual benefit was greatest when the mouth was visible, consistent with research on the importance of the mouth region for lip-reading. The difference in benefit between the opaque and transparent masks is consistent with previous findings from adults indicating that there is a minimal difference between lip-reading accuracy when viewing a whole face and viewing only the lips or the lower half of the face (<xref ref-type="bibr" rid="B19">Greenberg and Bode, 1968</xref>; <xref ref-type="bibr" rid="B20">Ijsseldijk, 1992</xref>; <xref ref-type="bibr" rid="B38">Marassa and Lansing, 1995</xref>).</p>
<p>Most of the groups in this study also showed a small yet significant audiovisual benefit with the opaque face masks. This benefit was observed with the hospital and fabric masks in the CHL and ANH tested at home, and the ANH tested in the lab at &#x2212;10 dB SNR, as well as with the fabric mask in the ANH tested in the lab at 0 dB SNR. These findings are consistent with those of <xref ref-type="bibr" rid="B70">Yi et al. (2021)</xref> who observed a significant audiovisual benefit with the ClearMask&#x2122; in ANH tested in noise and a significant audiovisual benefit with both the ClearMask&#x2122; and a hospital mask in ANH tested in babble. The finding of audiovisual benefit with opaque masks is consistent with studies suggesting that non-oral regions of the face contain subtle visual cues that contribute to audiovisual speech recognition in noise (<xref ref-type="bibr" rid="B60">Summerfield, 1979</xref>; <xref ref-type="bibr" rid="B54">Scheinberg, 1980</xref>; <xref ref-type="bibr" rid="B3">Beno&#x00EE;t et al., 1996</xref>; <xref ref-type="bibr" rid="B49">Preminger et al., 1998</xref>; <xref ref-type="bibr" rid="B61">Thomas and Jordan, 2004</xref>; <xref ref-type="bibr" rid="B15">Davis and Kim, 2006</xref>; <xref ref-type="bibr" rid="B22">Jordan and Thomas, 2011</xref>). For example, although occluding the bottom half of the face decreases visual-only recognition and audiovisual enhancement of CVs, visual-only recognition accuracy remains above chance, and a small audiovisual enhancement remains for some CVs (<xref ref-type="bibr" rid="B22">Jordan and Thomas, 2011</xref>). <xref ref-type="bibr" rid="B61">Thomas and Jordan (2004)</xref> suggested that occluding the lower half of the face might affect viewing and attention strategies. Non-oral face and head motions become more important when oral cues are not available; therefore, these cues could play a substantial role in understanding audiovisual speech with face masks (<xref ref-type="bibr" rid="B61">Thomas and Jordan, 2004</xref>; <xref ref-type="bibr" rid="B22">Jordan and Thomas, 2011</xref>). The significant audiovisual benefit we observed, despite the presence of an opaque face mask, suggests that talkers should try to keep their face visible to listeners in noisy conditions, even when wearing an opaque mask, as many listeners will benefit from limited visual cues available.</p>
</sec>
<sec id="S7.SS5">
<title>Effects of Face Masks on Phonetic Feature Transmission</title>
<p>We measured the effects of face masks on phonetic feature transmission to determine the type of perceptual errors face masks are likely to cause. Rank differences between performance with each mask were the same for all three consonant features as for overall consonant recognition accuracy. However, the face masks impacted the perception of some consonant features more than others.</p>
<p>We expected consonant features based primarily on high-frequency acoustic cues, such as the place of articulation of voiceless fricatives and stops (<xref ref-type="bibr" rid="B59">Stevens, 2000</xref>), to be most impacted by face masks. In the adults tested at &#x2212;10 dB SNR, this expectation was confirmed. The masks affected adults&#x2019; perception of auditory-only place more than voicing and manner. In CHL, the effect of the masks on auditory perception did not differ across features. However, more complex models are required to confirm the significance of this difference in patterns. In audiovisual conditions, we also expected consonant features that are easier to lip-read, such as place of articulation (<xref ref-type="bibr" rid="B4">Binnie et al., 1974</xref>; <xref ref-type="bibr" rid="B46">Owens and Blazek, 1985</xref>), to be most impacted by the opaque mask. This expectation was confirmed in both the CHL and the ANH; there was a greater difference between the opaque masks and transparent masks for place than for voicing and manner.</p>
<p>Although baseline performance in the auditory-only condition did not differ between the CHL tested at 0 dB SNR and the ANH tested at &#x2212;10 dB SNR, the feature transmission patterns differed. The CHL tested at 0 dB SNR showed greater differences between the perception of voicing and other features than the ANH tested at &#x2212;10 dB SNR. In other words, the increased masking noise used to decrease ANH performance to the level of CHL likely resulted in an error pattern different than that resulting from the combination of noise and children&#x2019;s hearing loss. Given that voicing is poorly transmitted visually, the visual signal may have less potential benefit to consonant perception in the ANH tested at &#x2013;10 dB SNR than the CHL tested at 0 dB SNR. In other words, differences in audiovisual benefit across groups who are tested at different SNRs may reflect differences in patterns of acoustic errors, in addition to differences in the ability to use visual speech information.</p>
</sec>
<sec id="S7.SS6">
<title>Visual-Only Consonant Perception in Children With Hearing Loss, Children With Normal Hearing, and Adults With Normal Hearing</title>
<p>Our results demonstrated that ANH are better at visual-only consonant recognition than CNH and CHL. There was large variability in visual-only consonant perception in the children, especially in the transmission of the place feature, but there was no difference in visual-only consonant perception between the children with and without hearing loss. Previous studies have provided conflicting results regarding whether CHL are better at lip-reading than CNH. Although some studies have shown no effect of hearing status on lip-reading (<xref ref-type="bibr" rid="B13">Conrad, 1977</xref>; <xref ref-type="bibr" rid="B24">Kyle et al., 2013</xref>), others have shown that a subset of CHL has better lip-reading than CNH (<xref ref-type="bibr" rid="B36">Lyxell and Holmberg, 2000</xref>; <xref ref-type="bibr" rid="B25">Kyle and Harris, 2006</xref>; <xref ref-type="bibr" rid="B64">Tye-Murray et al., 2014</xref>). Additional research is needed to understand the factors underlying these conflicting findings.</p>
</sec>
</sec>
<sec id="S8">
<title>Limitations and Future Directions</title>
<p>This study examined the perceptual consequences of the acoustic attenuation and visual obstruction caused by a variety of face masks. However, the test configuration in this study, a close frontal view of the talker&#x2019;s face, short test sessions, CV stimuli, and simulated masks, may not be representative of listening in daily life. The mask simulations do not capture any phoneme-specific effects of masks on articulation, including targeted adjustments to articulation that talkers might adopt when wearing a face mask. The mask simulations also do not account for the effects of improper mask placement and/or fogging. In future studies, it would be helpful to compare the acoustic consequences and perception of speech produced while wearing a mask to speech with mask simulations. Future studies could also test the perception of consonants, words, and sentences using the same participants and target talker to see how well results for consonant perception might generalize to the type of speech we encounter in everyday life.</p>
<p>We tested children from a broad age range, but ceiling effects precluded an examination of developmental differences in the negative impact of face masks. In future studies, a larger sample of children and performance-matching techniques would allow us to examine how the negative impact of face masks varies over development, and a more heterogeneous sample of CHL would allow us to probe whether the impact of acoustic attenuation caused by face masks depends on high-frequency aided audibility. A better understanding of children&#x2019;s orienting behaviors in classrooms would allow us to make more concrete recommendations. The combination of noise and face masks negatively impacts speech understanding in children. Future studies could examine the impact of masks on children in other listening conditions, such as conversational (rather than clear) speech (<xref ref-type="bibr" rid="B6">Brown et al., 2021</xref>; <xref ref-type="bibr" rid="B56">Smiljanic et al., 2021</xref>; <xref ref-type="bibr" rid="B70">Yi et al., 2021</xref>), listening to speech produced by a non-native talker (<xref ref-type="bibr" rid="B56">Smiljanic et al., 2021</xref>), or listening to speech in one&#x2019;s non-native language. Future studies could include children who use cochlear implants, as their deficits in acoustic-phonetic access differ qualitatively from children who use hearing aids.</p>
</sec>
<sec id="S9" sec-type="conclusion">
<title>Conclusion</title>
<list list-type="simple">
<list-item>
<label>&#x2022;</label>
<p>The best face masks for promoting speech understanding depend on whether visual cues will be available. When listeners will not be able to view the talker, a hospital mask is the best option for ANH, CNH, and CHL. Fabric masks that cause less acoustic attenuation than the one used in the current study (those made with fewer layers and less dense fabric) could be another good option. From a communication perspective, when working one-on-one and face-to-face with a listener at a short distance, a well-fit ClearMask&#x2122; is the best option, as it provides visual speech cues that more than compensate for higher levels of acoustic attenuation.</p>
</list-item>
<list-item>
<label>&#x2022;</label>
<p>In a classroom setting, the best mask for promoting speech understanding will depend on the degree to which children orient to the talker and whether the teacher is positioned to provide visual cues for all students. Previous studies have shown that young CNH do not consistently orient to the target talker the way adults do (<xref ref-type="bibr" rid="B50">Ricketts and Galster, 2008</xref>; <xref ref-type="bibr" rid="B66">Valente et al., 2012</xref>; <xref ref-type="bibr" rid="B32">Lewis et al., 2015</xref>, <xref ref-type="bibr" rid="B33">Lewis et al., 2018</xref>). The hospital mask may be better when communicating with children who do not orient to the talker.</p>
</list-item>
<list-item>
<label>&#x2022;</label>
<p>Even when a talker wears an opaque mask, many listeners benefit from seeing the head and facial movements that are not obscured by the mask. Therefore, talkers wearing opaque masks should try to keep their faces visible.</p>
</list-item>
<list-item>
<label>&#x2022;</label>
<p>One caveat to these conclusions is that the test configuration in this study, a close frontal view of the talker&#x2019;s face, short test sessions, CV stimuli, and simulated masks, may not be representative of listening in daily life.</p>
</list-item>
</list>
</sec>
<sec id="S10" sec-type="data-availability">
<title>Data Availability Statement</title>
<p>The datasets and stimuli presented in this study can be found at <ext-link ext-link-type="uri" xlink:href="https://osf.io/5wapg/">https://osf.io/5wapg/</ext-link>.</p>
</sec>
<sec id="S11">
<title>Ethics Statement</title>
<p>The studies involving human participants were reviewed and approved by the Institutional Review Board at Boys Town National Research Hospital. Written informed consent to participate in this study was provided by the participants or their legal guardians. Written informed consent was obtained from the individual(s) for the publication of any identifiable images or data included in this article.</p>
</sec>
<sec id="S12">
<title>Author Contributions</title>
<p>All authors contributed to conception and design of the study. LL secured the funding and assembled the research team. KL and EB created the stimuli and performed the data analyses. MM recruited and collected the data for Experiment 1. KL led the recruitment and data collection for Experiment 2. KL, EB, and MM wrote sections of the manuscript. All authors contributed to the manuscript revision, and read and approved the submitted version.</p>
</sec>
<sec id="conf1" sec-type="COI-statement">
<title>Conflict of Interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec id="pudiscl1" sec-type="disclaimer">
<title>Publisher&#x2019;s Note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
</body>
<back>
<sec id="S13" sec-type="funding-information">
<title>Funding</title>
<p>This project was funded by the National Institutes of Health (5R01DC011038 and 5P20GM109023).</p>
</sec>
<ack>
<p>Seth Bashford of the Boys Town National Research Hospital Technology Core developed the experimental software. Heather Porter, Taylor Corbaley, Megan Klinginsmith, and Randi S. Knox contributed to stimulus development. Taylor Corbaley contributed to data collection and recruitment for Experiment 2. Grace Dywer contributed to manuscript formatting and proofreading.</p>
</ack>
<sec id="S15" sec-type="supplementary-material">
<title>Supplementary Material</title>
<p>The Supplementary Material for this article can be found online at: <ext-link ext-link-type="uri" xlink:href="https://www.frontiersin.org/articles/10.3389/fpsyg.2022.874345/full#supplementary-material">https://www.frontiersin.org/articles/10.3389/fpsyg.2022.874345/full#supplementary-material</ext-link></p>
<supplementary-material xlink:href="Table_1.docx" id="TS1" mimetype="application/vnd.openxmlformats-officedocument.wordprocessingml.document" xmlns:xlink="http://www.w3.org/1999/xlink"/>
</sec>
<ref-list>
<title>References</title>
<ref id="B1"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Atcherson</surname> <given-names>S. R.</given-names></name> <name><surname>Finley</surname> <given-names>E. T.</given-names></name> <name><surname>McDowell</surname> <given-names>B. R.</given-names></name> <name><surname>Watson</surname> <given-names>C.</given-names></name></person-group> (<year>2020</year>). <source><italic>More Speech Degradations and Considerations in the Search for Transparent Face Coverings during the COVID-19 Pandemic. Audiology Today</italic>.</source> Available online at: <ext-link ext-link-type="uri" xlink:href="https://www.audiology.org/audiology-today-julyaugust-2020/online-feature-more-speech-degradations-and-considerations-search">https://www.audiology.org/audiology-today-julyaugust-2020/online-feature-more-speech-degradations-and-considerations-search</ext-link> <comment>(accessed September 30, 2020)</comment>.</citation></ref>
<ref id="B2"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Atcherson</surname> <given-names>S. R.</given-names></name> <name><surname>Poussonf</surname> <given-names>M.</given-names></name> <name><surname>Spann</surname> <given-names>M. J.</given-names></name></person-group> (<year>2017</year>). <article-title>The effect of conventional and transparent surgical masks on speech understanding in individuals with and without hearing loss.</article-title> <source><italic>J. Am. Acad. Audiol.</italic></source> <volume>28</volume> <fpage>58</fpage>&#x2013;<lpage>67</lpage>.</citation></ref>
<ref id="B3"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Beno&#x00EE;t</surname> <given-names>C.</given-names></name> <name><surname>Guiard-Marigny</surname> <given-names>T.</given-names></name> <name><surname>Le Goff</surname> <given-names>B.</given-names></name> <name><surname>Adjoudani</surname> <given-names>A.</given-names></name></person-group> (<year>1996</year>). &#x201C;<article-title>Which components of the face do humans and machines best speechread?</article-title>,&#x201D; in <source><italic>Speechreading by Humans and Machines: Models, Systems and Applications</italic></source>, <edition>150th Edn</edition>, <role>eds</role> <person-group person-group-type="editor"><name><surname>Stork</surname> <given-names>D. G.</given-names></name> <name><surname>Hennecke</surname> <given-names>M. E.</given-names></name></person-group> (<publisher-loc>Berlin</publisher-loc>: <publisher-name>Springer</publisher-name>), <fpage>315</fpage>&#x2013;<lpage>328</lpage>. <pub-id pub-id-type="doi">10.1007/978-3-662-13015-5_24</pub-id></citation></ref>
<ref id="B4"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Binnie</surname> <given-names>C. A.</given-names></name> <name><surname>Montgomery</surname> <given-names>A. A.</given-names></name> <name><surname>Jackson</surname> <given-names>P. L.</given-names></name></person-group> (<year>1974</year>). <article-title>Auditory and visual contributions to the perception of consonants.</article-title> <source><italic>J. Speech Hear. Res.</italic></source> <volume>17</volume> <fpage>619</fpage>&#x2013;<lpage>630</lpage>. <pub-id pub-id-type="doi">10.1044/jshr.1704.619</pub-id> <pub-id pub-id-type="pmid">4444283</pub-id></citation></ref>
<ref id="B5"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bottalico</surname> <given-names>P.</given-names></name> <name><surname>Murgia</surname> <given-names>S.</given-names></name> <name><surname>Puglisi</surname> <given-names>G. E.</given-names></name> <name><surname>Astolfi</surname> <given-names>A.</given-names></name> <name><surname>Kirk</surname> <given-names>K. I.</given-names></name></person-group> (<year>2020</year>). <article-title>Effect of masks on speech intelligibility in auralized classrooms.</article-title> <source><italic>J. Acoust. Soc. Am</italic>.</source> <volume>148</volume> <fpage>2878</fpage>&#x2013;<lpage>2884</lpage>. <pub-id pub-id-type="doi">10.1121/10.0002450</pub-id></citation></ref>
<ref id="B6"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Brown</surname> <given-names>V. A.</given-names></name> <name><surname>Van Engen</surname> <given-names>K.</given-names></name> <name><surname>Peelle</surname> <given-names>J. E.</given-names></name></person-group> (<year>2021</year>). <article-title>Face mask type affects audiovisual speech intelligibility and subjective listening effort in young and older adults.</article-title> <source><italic>Cogn. Res.</italic></source> <volume>6</volume>:<fpage>49</fpage>. <pub-id pub-id-type="doi">10.1186/s41235-021-00314-0</pub-id> <pub-id pub-id-type="pmid">34275022</pub-id></citation></ref>
<ref id="B7"><citation citation-type="journal"><collab>Centers for Disease Control and Prevention</collab> (<year>2021a</year>). <source><italic>Scientific Brief: Community Use of Cloth Masks to Control the Spread of SARS-CoV-2.</italic></source> Available online at: <ext-link ext-link-type="uri" xlink:href="https://www.cdc.gov/coronavirus/2019-ncov/more/masking-science-sars-cov2.html">https://www.cdc.gov/coronavirus/2019-ncov/more/masking-science-sars-cov2.html</ext-link> <comment>(accessed February 8, 2022)</comment>.</citation></ref>
<ref id="B8"><citation citation-type="journal"><collab>Centers for Disease Control and Prevention</collab> (<year>2021b</year>). <source><italic>Use Masks to Slow the Spread of COVID-19</italic>.</source> Available online at: <ext-link ext-link-type="uri" xlink:href="https://www.cdc.gov/coronavirus/2019-ncov/prevent-getting-sick/diy-cloth-face-coverings.html">https://www.cdc.gov/coronavirus/2019-ncov/prevent-getting-sick/diy-cloth-face-coverings.html</ext-link> <comment>(accessed February 8, 2022)</comment>.</citation></ref>
<ref id="B9"><citation citation-type="journal"><collab>Centers for Disease Control and Prevention</collab> (<year>2020</year>). <source><italic>Masks in Schools</italic>.</source> Available online at: <ext-link ext-link-type="uri" xlink:href="https://www.cdc.gov/coronavirus/2019-ncov/community/schools-childcare/cloth-face-cover.html">https://www.cdc.gov/coronavirus/2019-ncov/community/schools-childcare/cloth-face-cover.html</ext-link> <comment>(accessed January 22, 2021)</comment>.</citation></ref>
<ref id="B10"><citation citation-type="journal"><collab>Centers for Disease Control and Prevention</collab> (<year>2022a</year>). <source><italic>COVID-19 Types of Masks and Respirators Key Messages?: Choosing a Mask or Respirator for Different Situations.</italic></source> Available online at: <ext-link ext-link-type="uri" xlink:href="https://www.cdc.gov/coronavirus/2019-ncov/prevent-getting-sick/types-of-masks.html">https://www.cdc.gov/coronavirus/2019-ncov/prevent-getting-sick/types-of-masks.html</ext-link> <comment>(accessed February 8, 2022)</comment>.</citation></ref>
<ref id="B11"><citation citation-type="journal"><collab>Centers for Disease Control and Prevention</collab> (<year>2022b</year>). <source><italic>Your Guide to Masks.</italic></source> Available online at: <ext-link ext-link-type="uri" xlink:href="https://www.cdc.gov/coronavirus/2019-ncov/prevent-getting-sick/about-face-coverings.html">https://www.cdc.gov/coronavirus/2019-ncov/prevent-getting-sick/about-face-coverings.html</ext-link> <comment>(accessed February 8, 2022)</comment>.</citation></ref>
<ref id="B12"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cohn</surname> <given-names>M.</given-names></name> <name><surname>Pycha</surname> <given-names>A.</given-names></name> <name><surname>Zellou</surname> <given-names>G.</given-names></name></person-group> (<year>2021</year>). <article-title>Intelligibility of face-masked speech depends on speaking style: comparing casual, clear, and emotional speech.</article-title> <source><italic>Cognition</italic></source> <volume>210</volume>:<fpage>104570</fpage>. <pub-id pub-id-type="doi">10.1016/j.cognition.2020.104570</pub-id> <pub-id pub-id-type="pmid">33450446</pub-id></citation></ref>
<ref id="B13"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Conrad</surname> <given-names>R.</given-names></name></person-group> (<year>1977</year>). <article-title>Lip-reading by deaf and hearing children.</article-title> <source><italic>Br. J. Educ. Psychol.</italic></source> <volume>47</volume> <fpage>60</fpage>&#x2013;<lpage>65</lpage>. <pub-id pub-id-type="doi">10.1111/j.2044-8279.1977.tb03001.x</pub-id> <pub-id pub-id-type="pmid">843431</pub-id></citation></ref>
<ref id="B14"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Corey</surname> <given-names>R. M.</given-names></name> <name><surname>Jones</surname> <given-names>U.</given-names></name> <name><surname>Singer</surname> <given-names>A. C.</given-names></name></person-group> (<year>2020</year>). <article-title>Acoustic effects of medical, cloth, and transparent face masks on speech signals.</article-title> <source><italic>J. Acoust. Soc. Am.</italic></source> <volume>148</volume> <fpage>2371</fpage>&#x2013;<lpage>2375</lpage>. <pub-id pub-id-type="doi">10.1121/10.0002279</pub-id></citation></ref>
<ref id="B15"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Davis</surname> <given-names>C.</given-names></name> <name><surname>Kim</surname> <given-names>J.</given-names></name></person-group> (<year>2006</year>). <article-title>Audio-visual speech perception off the top of the head.</article-title> <source><italic>Cognition</italic></source> <volume>100</volume> <fpage>21</fpage>&#x2013;<lpage>31</lpage>. <pub-id pub-id-type="doi">10.1016/j.cognition.2005.09.002</pub-id> <pub-id pub-id-type="pmid">16289070</pub-id></citation></ref>
<ref id="B16"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fairbanks</surname> <given-names>G.</given-names></name></person-group> (<year>1969</year>). <source><italic>Voice and Articulation Drillbook</italic></source>, <edition>2nd Edn</edition>. <publisher-loc>New York, NY</publisher-loc>: <publisher-name>Harper &#x0026; Row</publisher-name>.</citation></ref>
<ref id="B17"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Flaherty</surname> <given-names>M. M.</given-names></name></person-group> (<year>2021</year>). <article-title>The effects of face masks on speech-in-speech recognition for children and adults.</article-title> <source><italic>J. Acoust. Soc. Am.</italic></source> <volume>1500</volume>:<fpage>A153</fpage>.</citation></ref>
<ref id="B18"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Goldin</surname> <given-names>A.</given-names></name> <name><surname>Weinstein</surname> <given-names>B.</given-names></name> <name><surname>Shiman</surname> <given-names>N.</given-names></name></person-group> (<year>2020</year>). <article-title>Speech blocked by surgical masks becomes a more important issue in the era of COVID-19.</article-title> <source><italic>Hear. Rev.</italic></source> <volume>27</volume> <fpage>8</fpage>&#x2013;<lpage>9</lpage>.</citation></ref>
<ref id="B19"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Greenberg</surname> <given-names>H. J.</given-names></name> <name><surname>Bode</surname> <given-names>D. L.</given-names></name></person-group> (<year>1968</year>). <article-title>Visual discrimination of consonants.</article-title> <source><italic>J. Speech Hear. Res.</italic></source> <volume>11</volume> <fpage>869</fpage>&#x2013;<lpage>874</lpage>. <pub-id pub-id-type="doi">10.1044/jshr.1104.869</pub-id> <pub-id pub-id-type="pmid">5719244</pub-id></citation></ref>
<ref id="B20"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ijsseldijk</surname> <given-names>F. J.</given-names></name></person-group> (<year>1992</year>). <article-title>Speechreading performance under different conditions of video image, repetition, and speech rate.</article-title> <source><italic>J. Speech Hear. Res.</italic></source> <volume>35</volume> <fpage>466</fpage>&#x2013;<lpage>471</lpage>. <pub-id pub-id-type="doi">10.1044/jshr.3502.466</pub-id> <pub-id pub-id-type="pmid">1573883</pub-id></citation></ref>
<ref id="B21"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jeong</surname> <given-names>J.</given-names></name> <name><surname>Kim</surname> <given-names>M.</given-names></name> <name><surname>Kim</surname> <given-names>Y.</given-names></name></person-group> (<year>2020</year>). <article-title>Changes on speech transmission characteristics by types of mask.</article-title> <source><italic>Audiol. Speech Res.</italic></source> <volume>16</volume> <fpage>295</fpage>&#x2013;<lpage>304</lpage>. <pub-id pub-id-type="doi">10.21848/asr.200053</pub-id></citation></ref>
<ref id="B22"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jordan</surname> <given-names>T. R.</given-names></name> <name><surname>Thomas</surname> <given-names>S. M.</given-names></name></person-group> (<year>2011</year>). <article-title>When half a face is as good as a whole: effects of simple substantial occlusion on visual and audiovisual speech perception.</article-title> <source><italic>Atten. Percept. Psychophys.</italic></source> <volume>73</volume> <fpage>2270</fpage>&#x2013;<lpage>2285</lpage>. <pub-id pub-id-type="doi">10.3758/s13414-011-0152-4</pub-id> <pub-id pub-id-type="pmid">21842332</pub-id></citation></ref>
<ref id="B23"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kuznetsova</surname> <given-names>A.</given-names></name> <name><surname>Brockhoff</surname> <given-names>P. B.</given-names></name> <name><surname>Christensen</surname> <given-names>R. H. B.</given-names></name></person-group> (<year>2017</year>). <source><italic>&#x2018;lmerTest: Test for Random and Fixed Effects for Linear Mixed Effects Models (R package Version, 2.0-2.5)&#x2019;, Computer Software.</italic></source> Available online at: <ext-link ext-link-type="uri" xlink:href="https://cran.r-project.org/web/packages/lmerTest/lmerTest.pdf">https://cran.r-project.org/web/packages/lmerTest/lmerTest.pdf</ext-link> <comment>(accessed January 22, 2018)</comment>.</citation></ref>
<ref id="B24"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kyle</surname> <given-names>F. E.</given-names></name> <name><surname>Campbell</surname> <given-names>R.</given-names></name> <name><surname>Mohammed</surname> <given-names>T.</given-names></name> <name><surname>Coleman</surname> <given-names>M.</given-names></name> <name><surname>Macsweeney</surname> <given-names>M.</given-names></name></person-group> (<year>2013</year>). <article-title>Speechreading development in deaf and hearing children: introducing the test of child speechreading.</article-title> <source><italic>J. Speech Lang. Hear. Res.</italic></source> <volume>56</volume> <fpage>416</fpage>&#x2013;<lpage>426</lpage>. <pub-id pub-id-type="doi">10.1044/1092-4388(2012/12-0039)</pub-id></citation></ref>
<ref id="B25"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kyle</surname> <given-names>F. E.</given-names></name> <name><surname>Harris</surname> <given-names>M.</given-names></name></person-group> (<year>2006</year>). <article-title>Concurrent correlates and predictors of reading and spelling achievement in deaf and hearing school children.</article-title> <source><italic>J. Deaf Stud. Deaf Educ.</italic></source> <volume>11</volume> <fpage>273</fpage>&#x2013;<lpage>288</lpage>. <pub-id pub-id-type="doi">10.1093/deafed/enj037</pub-id> <pub-id pub-id-type="pmid">16556897</pub-id></citation></ref>
<ref id="B26"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lalonde</surname> <given-names>K.</given-names></name> <name><surname>Holt</surname> <given-names>R. F.</given-names></name></person-group> (<year>2016</year>). <article-title>Audiovisual speech perception development at varying levels of perceptual processing.</article-title> <source><italic>J. Acoust. Soc. Am.</italic></source> <volume>139</volume> <fpage>1713</fpage>&#x2013;<lpage>1723</lpage>. <pub-id pub-id-type="doi">10.1121/1.4945590</pub-id></citation></ref>
<ref id="B27"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lalonde</surname> <given-names>K.</given-names></name> <name><surname>McCreery</surname> <given-names>R. W.</given-names></name></person-group> (<year>2020</year>). <article-title>Audiovisual enhancement of speech perception in noise by school-age children who are hard of hearing.</article-title> <source><italic>Ear Hear.</italic></source> <volume>41</volume> <fpage>705</fpage>&#x2013;<lpage>719</lpage>. <pub-id pub-id-type="doi">10.1097/AUD.0000000000000830</pub-id> <pub-id pub-id-type="pmid">32032226</pub-id></citation></ref>
<ref id="B28"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lalonde</surname> <given-names>K.</given-names></name> <name><surname>Werner</surname> <given-names>L. A.</given-names></name></person-group> (<year>2021</year>). <article-title>Development of the mechanisms underlying audiovisual speech perception benefit.</article-title> <source><italic>Brain Sci.</italic></source> <volume>11</volume>:<fpage>49</fpage>. <pub-id pub-id-type="doi">10.3390/brainsci11010049</pub-id> <pub-id pub-id-type="pmid">33466253</pub-id></citation></ref>
<ref id="B29"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Leibold</surname> <given-names>J. J.</given-names></name> <name><surname>Hillock-Dunn</surname> <given-names>A.</given-names></name> <name><surname>Duncan</surname> <given-names>N.</given-names></name> <name><surname>Roush</surname> <given-names>P. A.</given-names></name> <name><surname>Buss</surname> <given-names>E.</given-names></name></person-group> (<year>2013</year>). <article-title>Influence of hearing loss on children&#x2019;s identification of spondee words in speech-shaped noise or a two-talker masker.</article-title> <source><italic>Ear Hear.</italic></source> <volume>34</volume> <fpage>575</fpage>&#x2013;<lpage>584</lpage>. <pub-id pub-id-type="doi">10.1097/aud.0b013e3182857742</pub-id> <pub-id pub-id-type="pmid">23492919</pub-id></citation></ref>
<ref id="B30"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Leibold</surname> <given-names>L. J.</given-names></name> <name><surname>Browning</surname> <given-names>J. M.</given-names></name> <name><surname>Buss</surname> <given-names>E.</given-names></name></person-group> (<year>2019</year>). <article-title>Masking release for speech-in-speech recognition due to target/masker sex mismatch in children with heairng loss.</article-title> <source><italic>Ear Hear.</italic></source> <volume>41</volume> <fpage>259</fpage>&#x2013;<lpage>267</lpage>. <pub-id pub-id-type="doi">10.1097/AUD.0000000000000752</pub-id> <pub-id pub-id-type="pmid">31365355</pub-id></citation></ref>
<ref id="B31"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Leibold</surname> <given-names>L. J.</given-names></name> <name><surname>Buss</surname> <given-names>E.</given-names></name></person-group> (<year>2013</year>). <article-title>Children&#x2019;s identification of consonants in a speech-shaped noise or a two-talker masker.</article-title> <source><italic>J. Speech Lang. Hear. Res.</italic></source> <volume>56</volume> <fpage>1144</fpage>&#x2013;<lpage>1155</lpage>. <pub-id pub-id-type="doi">10.1044/1092-4388(2012/12-0011)</pub-id></citation></ref>
<ref id="B32"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lewis</surname> <given-names>D.</given-names></name> <name><surname>Valente</surname> <given-names>D. L.</given-names></name> <name><surname>Spalding</surname> <given-names>J. L.</given-names></name></person-group> (<year>2015</year>). <article-title>Effect of minial/mild hearing loss on children&#x2019;s speech understanding in a simulated classroom.</article-title> <source><italic>Ear Hear.</italic></source> <volume>36</volume> <fpage>136</fpage>&#x2013;<lpage>144</lpage>. <pub-id pub-id-type="doi">10.1097/AUD.0000000000000092</pub-id> <pub-id pub-id-type="pmid">25170780</pub-id></citation></ref>
<ref id="B33"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lewis</surname> <given-names>D. E.</given-names></name> <name><surname>Smith</surname> <given-names>N. A.</given-names></name> <name><surname>Spalding</surname> <given-names>J. L.</given-names></name> <name><surname>Valente</surname> <given-names>D. L.</given-names></name></person-group> (<year>2018</year>). <article-title>Looking behavior and audiovisual speech understanding in children with normal hearing and children with mild bilateral or unilateral hearing loss.</article-title> <source><italic>Ear Hear.</italic></source> <volume>39</volume> <fpage>783</fpage>&#x2013;<lpage>794</lpage>. <pub-id pub-id-type="doi">10.1097/AUD.0000000000000534</pub-id> <pub-id pub-id-type="pmid">29252979</pub-id></citation></ref>
<ref id="B34"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lipps</surname> <given-names>E.</given-names></name> <name><surname>Caldwell-Kurtzman</surname> <given-names>J.</given-names></name> <name><surname>Motlagh-Zadeh</surname> <given-names>L.</given-names></name> <name><surname>Blankenship</surname> <given-names>C. M.</given-names></name> <name><surname>Moore</surname> <given-names>D. R.</given-names></name> <name><surname>Hunter</surname> <given-names>L. L.</given-names></name><etal/></person-group> (<year>2021</year>). <article-title>Impact of face masks on audiovisual word recognition in young children with hearing loss during the COVID-19 pandemic.</article-title> <source><italic>J. Early Hear. Detect. Interv.</italic></source> <volume>6</volume> <fpage>70</fpage>&#x2013;<lpage>78</lpage>.</citation></ref>
<ref id="B35"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Llamas</surname> <given-names>C.</given-names></name> <name><surname>Harrison</surname> <given-names>P.</given-names></name> <name><surname>Donnelly</surname> <given-names>D.</given-names></name> <name><surname>Watt</surname> <given-names>D.</given-names></name></person-group> (<year>2008</year>). <article-title>Effects of different types of face coverings on speech acoustics and intelligibility.</article-title> <source><italic>York Pap. Linguist. Ser.</italic></source> <volume>2</volume> <fpage>80</fpage>&#x2013;<lpage>104</lpage>.</citation></ref>
<ref id="B36"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lyxell</surname> <given-names>B.</given-names></name> <name><surname>Holmberg</surname> <given-names>I.</given-names></name></person-group> (<year>2000</year>). <article-title>Visual speechreading and cognitive performance in hearing-impaired and normal hearing children (11-14 years).</article-title> <source><italic>Br. J. Educ. Psychol.</italic></source> <volume>70</volume> <fpage>505</fpage>&#x2013;<lpage>518</lpage>. <pub-id pub-id-type="doi">10.1348/000709900158272</pub-id> <pub-id pub-id-type="pmid">11191184</pub-id></citation></ref>
<ref id="B37"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Magee</surname> <given-names>M.</given-names></name> <name><surname>Lewis</surname> <given-names>C.</given-names></name> <name><surname>Noffs</surname> <given-names>G.</given-names></name> <name><surname>Reece</surname> <given-names>H.</given-names></name> <name><surname>Chan</surname> <given-names>J. C. S.</given-names></name> <name><surname>Zaga</surname> <given-names>C. J.</given-names></name><etal/></person-group> (<year>2021</year>). <article-title>Effects of face masks on acoustic analysis and speech perception?: implications for peri-pandemic protocols.</article-title> <source><italic>J. Acoust. Soc. Am.</italic></source> <volume>148</volume> <fpage>3562</fpage>&#x2013;<lpage>3568</lpage>. <pub-id pub-id-type="doi">10.1121/10.0002873</pub-id> <pub-id pub-id-type="pmid">27367518</pub-id></citation></ref>
<ref id="B38"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Marassa</surname> <given-names>L. K.</given-names></name> <name><surname>Lansing</surname> <given-names>C. R.</given-names></name></person-group> (<year>1995</year>). <article-title>Visual word recognition in two facial motion conditions?: full face versus lips-plus-mandible.</article-title> <source><italic>J. Speech Hear. Res.</italic></source> <volume>38</volume> <fpage>1387</fpage>&#x2013;<lpage>1394</lpage>. <pub-id pub-id-type="doi">10.1044/jshr.3806.1387</pub-id> <pub-id pub-id-type="pmid">8747830</pub-id></citation></ref>
<ref id="B39"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>McCreery</surname> <given-names>R. W.</given-names></name> <name><surname>Stelmachowicz</surname> <given-names>P. G.</given-names></name></person-group> (<year>2011</year>). <article-title>Audibility-based predictions of speech recognition for children and adults with normal hearing.</article-title> <source><italic>J. Acoust. Soc. Am.</italic></source> <volume>130</volume> <fpage>4070</fpage>&#x2013;<lpage>4081</lpage>. <pub-id pub-id-type="doi">10.1121/1.3658476</pub-id></citation></ref>
<ref id="B40"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mendel</surname> <given-names>L. L.</given-names></name> <name><surname>Gardino</surname> <given-names>J. A.</given-names></name> <name><surname>Atcherson</surname> <given-names>S. R.</given-names></name></person-group> (<year>2008</year>). <article-title>Speech understanding using surgical masks: A problem in health care?</article-title> <source><italic>J. Am. Acad. Audiol.</italic></source> <volume>19</volume> <fpage>686</fpage>&#x2013;<lpage>695</lpage>. <pub-id pub-id-type="doi">10.3766/jaaa.19.9.4</pub-id> <pub-id pub-id-type="pmid">19418708</pub-id></citation></ref>
<ref id="B41"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mlot</surname> <given-names>S.</given-names></name> <name><surname>Buss</surname> <given-names>E.</given-names></name> <name><surname>Hall</surname> <given-names>J. W.</given-names></name></person-group> (<year>2010</year>). <article-title>Spectral integration and bandwidth effects on speech recognition in school-aged children and adults.</article-title> <source><italic>Ear Hear.</italic></source> <volume>31</volume> <fpage>56</fpage>&#x2013;<lpage>62</lpage>. <pub-id pub-id-type="doi">10.1097/AUD.0b013e3181ba746b</pub-id> <pub-id pub-id-type="pmid">19816182</pub-id></citation></ref>
<ref id="B42"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Munhall</surname> <given-names>K. G.</given-names></name> <name><surname>Jones</surname> <given-names>J. A.</given-names></name> <name><surname>Callan</surname> <given-names>D. E.</given-names></name> <name><surname>Kuratate</surname> <given-names>T.</given-names></name> <name><surname>Vatikiotis-Bateson</surname> <given-names>E.</given-names></name></person-group> (<year>2004</year>). &#x201C;<article-title>Visual Prosody and speech intelligibility: Head movement improves auditory speech perception.</article-title>&#x201D; <source><italic>Psychol. Sci</italic></source>. <volume>15</volume>, <fpage>133</fpage>&#x2013;<lpage>137</lpage>.</citation></ref>
<ref id="B43"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Munhall</surname> <given-names>K. G.</given-names></name> <name><surname>Vatikiotis-Bateson</surname> <given-names>E.</given-names></name></person-group> (<year>1998</year>). &#x201C;<article-title>The moving face during speech communication</article-title>,&#x201D; in <source><italic>Hearing by Eye II: Advances in the Psychology of Speechreading and Auditory&#x2013;Visual Speech</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>Campbell</surname> <given-names>R.</given-names></name> <name><surname>Dodd</surname> <given-names>B.</given-names></name> <name><surname>Burnham</surname> <given-names>D.</given-names></name></person-group> (<publisher-loc>London</publisher-loc>: <publisher-name>Psychology Press</publisher-name>), <fpage>123</fpage>&#x2013;<lpage>139</lpage>.</citation></ref>
<ref id="B44"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Muzzi</surname> <given-names>E.</given-names></name> <name><surname>Chermaz</surname> <given-names>C.</given-names></name> <name><surname>Castro</surname> <given-names>V.</given-names></name> <name><surname>Zaninoni</surname> <given-names>M.</given-names></name> <name><surname>Saksida</surname> <given-names>A.</given-names></name> <name><surname>Orzan</surname> <given-names>E.</given-names></name></person-group> (<year>2021</year>). <article-title>Short report on the effects of SARS-CoV-2 face protective equipment on verbal communication.</article-title> <source><italic>Eur. Arch. Oto rhino Laryngol.</italic></source> <volume>278</volume> <fpage>3565</fpage>&#x2013;<lpage>3570</lpage>. <pub-id pub-id-type="doi">10.1007/s00405-020-06535-1</pub-id> <pub-id pub-id-type="pmid">33389012</pub-id></citation></ref>
<ref id="B45"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Naylor</surname> <given-names>G.</given-names></name> <name><surname>Burke</surname> <given-names>L. A.</given-names></name> <name><surname>Holman</surname> <given-names>J. A.</given-names></name></person-group> (<year>2020</year>). <article-title>COVID-19 lockdown affects hearing disability and handicap in diverse ways?: a rapid online survey study.</article-title> <source><italic>Ear Hear.</italic></source> <volume>41</volume> <fpage>1442</fpage>&#x2013;<lpage>1449</lpage>. <pub-id pub-id-type="doi">10.1097/AUD.0000000000000948</pub-id> <pub-id pub-id-type="pmid">33136621</pub-id></citation></ref>
<ref id="B46"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Owens</surname> <given-names>E.</given-names></name> <name><surname>Blazek</surname> <given-names>B.</given-names></name></person-group> (<year>1985</year>). <article-title>Visemes observed by hearing-impaired and normal-hearing adult viewers.</article-title> <source><italic>J. Speech Hear. Res.</italic></source> <volume>28</volume> <fpage>381</fpage>&#x2013;<lpage>393</lpage>. <pub-id pub-id-type="doi">10.1044/jshr.2803.381</pub-id> <pub-id pub-id-type="pmid">4046579</pub-id></citation></ref>
<ref id="B47"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Palmiero</surname> <given-names>A. J.</given-names></name> <name><surname>Symons</surname> <given-names>D.</given-names></name> <name><surname>Morgan</surname> <given-names>J. W.</given-names> <suffix>III</suffix></name> <name><surname>Shaffer</surname> <given-names>R. E.</given-names></name></person-group> (<year>2016</year>). <article-title>Speech intelligibility assessment of protective facemasks and air purifying respirators.</article-title> <source><italic>J. Occup. Environ. Hyg.</italic></source> <volume>13</volume> <fpage>960</fpage>&#x2013;<lpage>968</lpage>. <pub-id pub-id-type="doi">10.1080/15459624.2016.1200723</pub-id> <pub-id pub-id-type="pmid">27362358</pub-id></citation></ref>
<ref id="B48"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>P&#x00F6;rschmann</surname> <given-names>C.</given-names></name> <name><surname>L&#x00FC;beck</surname> <given-names>T.</given-names></name> <name><surname>Arend</surname> <given-names>J. M.</given-names></name></person-group> (<year>2020</year>). <article-title>Impact of face masks on voice radiation.</article-title> <source><italic>J. Acoust. Soc. Am.</italic></source> <volume>148</volume> <fpage>3663</fpage>&#x2013;<lpage>3670</lpage>. <pub-id pub-id-type="doi">10.1121/10.0002853</pub-id></citation></ref>
<ref id="B49"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Preminger</surname> <given-names>J. E.</given-names></name> <name><surname>Lin</surname> <given-names>H. B.</given-names></name> <name><surname>Payen</surname> <given-names>M.</given-names></name> <name><surname>Levitt</surname> <given-names>H.</given-names></name></person-group> (<year>1998</year>). <article-title>Selective visual masking in speechreading.</article-title> <source><italic>J. Speech Lang. Hear. Res.</italic></source> <volume>41</volume> <fpage>564</fpage>&#x2013;<lpage>575</lpage>. <pub-id pub-id-type="doi">10.1044/jslhr.4103.564</pub-id> <pub-id pub-id-type="pmid">9638922</pub-id></citation></ref>
<ref id="B50"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ricketts</surname> <given-names>T. A.</given-names></name> <name><surname>Galster</surname> <given-names>J.</given-names></name></person-group> (<year>2008</year>). <article-title>Head angle and elevation in classroom environments: implications for amplification.</article-title> <source><italic>J. Speech Lang. Hear. Res.</italic></source> <volume>51</volume> <fpage>516</fpage>&#x2013;<lpage>525</lpage>. <pub-id pub-id-type="doi">10.1044/1092-4388(2008/037)</pub-id></citation></ref>
<ref id="B51"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ross</surname> <given-names>L. A.</given-names></name> <name><surname>Molholm</surname> <given-names>S.</given-names></name> <name><surname>Blanco</surname> <given-names>D.</given-names></name> <name><surname>Gomez-Ramirez</surname> <given-names>M.</given-names></name> <name><surname>Saint-Amour</surname> <given-names>D.</given-names></name> <name><surname>Foxe</surname> <given-names>J. J.</given-names></name><etal/></person-group> (<year>2011</year>). <article-title>The development of multisensory speech perception continues into the late childhood years.</article-title> <source><italic>Eur. J. Neurosci.</italic></source> <volume>33</volume> <fpage>2329</fpage>&#x2013;<lpage>2337</lpage>. <pub-id pub-id-type="doi">10.1111/j.1460-9568.2011.07685.x</pub-id> <pub-id pub-id-type="pmid">21615556</pub-id></citation></ref>
<ref id="B52"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Saragih</surname> <given-names>J. M.</given-names></name> <name><surname>Lucey</surname> <given-names>S.</given-names></name> <name><surname>Cohn</surname> <given-names>J. E.</given-names></name></person-group> (<year>2011</year>). <article-title>Deformable model fitting by regularized landmark mean-shift.</article-title> <source><italic>Int. J. Comput. Vis.</italic></source> <volume>91</volume> <fpage>200</fpage>&#x2013;<lpage>215</lpage>. <pub-id pub-id-type="doi">10.1007/s11263-010-0380-4</pub-id></citation></ref>
<ref id="B53"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Saunders</surname> <given-names>G. H.</given-names></name> <name><surname>Jackson</surname> <given-names>I. R.</given-names></name> <name><surname>Visram</surname> <given-names>A. S.</given-names></name></person-group> (<year>2020</year>). <article-title>Impacts of face coverings on communication: an indirect impact of COVID-19.</article-title> <source><italic>Int. J. Audiol.</italic></source> <volume>60</volume> <fpage>495</fpage>&#x2013;<lpage>506</lpage>. <pub-id pub-id-type="doi">10.1080/14992027.2020.1851401</pub-id> <pub-id pub-id-type="pmid">33246380</pub-id></citation></ref>
<ref id="B54"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Scheinberg</surname> <given-names>J. S.</given-names></name></person-group> (<year>1980</year>). <article-title>Analysis of speecreading cues using an interleaved technique.</article-title> <source><italic>J. Commun. Disord.</italic></source> <volume>13</volume> <fpage>489</fpage>&#x2013;<lpage>492</lpage>. <pub-id pub-id-type="doi">10.1016/0021-9924(80)90048-9</pub-id></citation></ref>
<ref id="B55"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sfakianaki</surname> <given-names>A.</given-names></name> <name><surname>Kafentzis</surname> <given-names>G. P.</given-names></name> <name><surname>Kiagiadaki</surname> <given-names>D.</given-names></name> <name><surname>Vlahavas</surname> <given-names>G.</given-names></name></person-group> (<year>2021</year>). &#x201C;<article-title>Effect of face mask and noise on word recognition by children and adults</article-title>,&#x201D; in <source><italic>Proceedings of 12th International Conference of Experimental Linguistics</italic></source>, <role>ed.</role> <person-group person-group-type="editor"><name><surname>Botinis</surname> <given-names>A.</given-names></name></person-group> (<publisher-loc>Athens</publisher-loc>: <publisher-name>ExLing Society</publisher-name>), <fpage>207</fpage>&#x2013;<lpage>210</lpage>.</citation></ref>
<ref id="B56"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Smiljanic</surname> <given-names>R.</given-names></name> <name><surname>Keerstock</surname> <given-names>S.</given-names></name> <name><surname>Meemann</surname> <given-names>K.</given-names></name> <name><surname>Ransom</surname> <given-names>S. M.</given-names></name></person-group> (<year>2021</year>). <article-title>Face masks and speaking style affect audio-visual word recognition and memory of native and non-native speech.</article-title> <source><italic>J. Acoust. Soc. Am.</italic></source> <volume>149</volume> <fpage>4013</fpage>&#x2013;<lpage>4023</lpage>. <pub-id pub-id-type="doi">10.1121/10.0005191</pub-id></citation></ref>
<ref id="B57"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Stelmachowicz</surname> <given-names>P. G.</given-names></name> <name><surname>Hoover</surname> <given-names>B. M.</given-names></name> <name><surname>Lewis</surname> <given-names>D. E.</given-names></name> <name><surname>Kortekaas</surname> <given-names>R. W.</given-names></name> <name><surname>Pittman</surname> <given-names>A. L.</given-names></name></person-group> (<year>2000</year>). <article-title>The relation between stimulus context, speech audibility, and perception for normal-hearing and hearing-impaired children.</article-title> <source><italic>J. Speech Hear. Res.</italic></source> <volume>43</volume> <fpage>902</fpage>&#x2013;<lpage>914</lpage>. <pub-id pub-id-type="doi">10.1044/jslhr.4304.902</pub-id> <pub-id pub-id-type="pmid">11386477</pub-id></citation></ref>
<ref id="B58"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Stelmachowicz</surname> <given-names>P. G.</given-names></name> <name><surname>Pittman</surname> <given-names>A. L.</given-names></name> <name><surname>Hoover</surname> <given-names>B. M.</given-names></name> <name><surname>Lewis</surname> <given-names>D. E.</given-names></name></person-group> (<year>2001</year>). <article-title>Effect of stimulus bandwidth on the perception of &#x201C;s&#x201D; in normal- and hearing-impaired children and adults.</article-title> <source><italic>J. Acoust. Soc. Am.</italic></source> <volume>110</volume> <fpage>2183</fpage>&#x2013;<lpage>2190</lpage>. <pub-id pub-id-type="doi">10.1121/1.1400757</pub-id></citation></ref>
<ref id="B59"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Stevens</surname> <given-names>K. N.</given-names></name></person-group> (<year>2000</year>). <source><italic>Acoustic Phonetics.</italic></source> <publisher-loc>Cambridge, MA</publisher-loc>: <publisher-name>MIT Press</publisher-name>.</citation></ref>
<ref id="B60"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Summerfield</surname> <given-names>Q.</given-names></name></person-group> (<year>1979</year>). <article-title>Use of visual information for phonetic perception.</article-title> <source><italic>Phonetica</italic></source> <volume>36</volume> <fpage>314</fpage>&#x2013;<lpage>331</lpage>. <pub-id pub-id-type="doi">10.1159/000259969</pub-id> <pub-id pub-id-type="pmid">523520</pub-id></citation></ref>
<ref id="B61"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Thomas</surname> <given-names>S. M.</given-names></name> <name><surname>Jordan</surname> <given-names>T. R.</given-names></name></person-group> (<year>2004</year>). <article-title>Contributions of oral and extraoral facial movement to visual and audiovisual speech perception.</article-title> <source><italic>J. Exp. Psychol.</italic></source> <volume>30</volume> <fpage>873</fpage>&#x2013;<lpage>888</lpage>. <pub-id pub-id-type="doi">10.1037/0096-1523.30.5.873</pub-id> <pub-id pub-id-type="pmid">15462626</pub-id></citation></ref>
<ref id="B62"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Toscano</surname> <given-names>J. C.</given-names></name> <name><surname>Toscano</surname> <given-names>C. M.</given-names></name></person-group> (<year>2021</year>). <article-title>Effects of face masks on speech recognition in multi-talker babble noise.</article-title> <source><italic>PLoS One</italic></source> <volume>16</volume>:<fpage>e0246842</fpage>. <pub-id pub-id-type="doi">10.1371/journal.pone.0246842</pub-id> <pub-id pub-id-type="pmid">33626073</pub-id></citation></ref>
<ref id="B63"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Truong</surname> <given-names>T. L.</given-names></name> <name><surname>Beck</surname> <given-names>S. D.</given-names></name> <name><surname>Weber</surname> <given-names>A.</given-names></name></person-group> (<year>2021</year>). <article-title>The impact of face masks on the recall of spoken sentences.</article-title> <source><italic>J. Acoust. Soc. Am.</italic></source> <volume>149</volume> <fpage>142</fpage>&#x2013;<lpage>144</lpage>. <pub-id pub-id-type="doi">10.1121/10.0002951</pub-id> <pub-id pub-id-type="pmid">34093817</pub-id></citation></ref>
<ref id="B64"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tye-Murray</surname> <given-names>N.</given-names></name> <name><surname>Hale</surname> <given-names>S.</given-names></name> <name><surname>Spehar</surname> <given-names>B.</given-names></name> <name><surname>Myerson</surname> <given-names>J.</given-names></name> <name><surname>Sommers</surname> <given-names>M. S.</given-names></name></person-group> (<year>2014</year>). <article-title>Lipreading in school-age children: the roles of age, hearing status, and cognitive ability.</article-title> <source><italic>J. Speech Lang. Hear. Res.</italic></source> <volume>57</volume> <fpage>556</fpage>&#x2013;<lpage>565</lpage>. <pub-id pub-id-type="doi">10.1044/2013_JSLHR-H-12-0273</pub-id> <pub-id pub-id-type="pmid">33249997</pub-id></citation></ref>
<ref id="B65"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tyler</surname> <given-names>R. S.</given-names></name> <name><surname>Fryauf-Bertschy</surname> <given-names>H.</given-names></name> <name><surname>Kelsay</surname> <given-names>D.</given-names></name></person-group> (<year>1991</year>). <source><italic>Audiovisual Feature Test for Young Children.</italic></source> <publisher-loc>Iowa City, IA</publisher-loc>: <publisher-name>University of Iowa</publisher-name>.</citation></ref>
<ref id="B66"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Valente</surname> <given-names>D. L.</given-names></name> <name><surname>Plevinsky</surname> <given-names>H. M.</given-names></name> <name><surname>Franco</surname> <given-names>J. M.</given-names></name> <name><surname>Heinrichs-Graham</surname> <given-names>E. C.</given-names></name> <name><surname>Lewis</surname> <given-names>D. E.</given-names></name></person-group> (<year>2012</year>). <article-title>Experimental investigation of the effects of the acoustical conditions in a simulated classroom on speech recognition and learning in children.</article-title> <source><italic>J. Acoust. Soc. Am.</italic></source> <volume>131</volume> <fpage>232</fpage>&#x2013;<lpage>246</lpage>. <pub-id pub-id-type="doi">10.1121/1.3662059</pub-id></citation></ref>
<ref id="B67"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Vos</surname> <given-names>T. G.</given-names></name> <name><surname>Dillon</surname> <given-names>M. T.</given-names></name> <name><surname>Buss</surname> <given-names>E.</given-names></name> <name><surname>Rooth</surname> <given-names>M. A.</given-names></name> <name><surname>Bucker</surname> <given-names>A. L.</given-names></name> <name><surname>Dillon</surname> <given-names>S.</given-names></name><etal/></person-group> (<year>2021</year>). <article-title>Influence of protective face coverings on the speech recognition of cochlear implant patients.</article-title> <source><italic>Laryngoscope</italic></source> <volume>131</volume> <fpage>E2038</fpage>&#x2013;<lpage>E2043</lpage>. <pub-id pub-id-type="doi">10.1002/lary.29447</pub-id> <pub-id pub-id-type="pmid">33590898</pub-id></citation></ref>
<ref id="B68"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wightman</surname> <given-names>F.</given-names></name> <name><surname>Kistler</surname> <given-names>D.</given-names></name> <name><surname>Brungart</surname> <given-names>D.</given-names></name></person-group> (<year>2006</year>). <article-title>Informational masking of speech in children: auditory-visual integration.</article-title> <source><italic>J. Acoust. Soc. Am.</italic></source> <volume>119</volume> <fpage>3939</fpage>&#x2013;<lpage>3949</lpage>. <pub-id pub-id-type="doi">10.1121/1.2195121</pub-id></citation></ref>
<ref id="B69"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wittum</surname> <given-names>K. J.</given-names></name> <name><surname>Feth</surname> <given-names>L.</given-names></name> <name><surname>Hoglund</surname> <given-names>E.</given-names></name></person-group> (<year>2013</year>). <article-title>The effects of surgical masks on speech perception in noise research.</article-title> <source><italic>Proc. Mtgs. Acoust.</italic></source> <volume>19</volume>:<fpage>060125</fpage>.</citation></ref>
<ref id="B70"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Yi</surname> <given-names>H.</given-names></name> <name><surname>Pingsterhaus</surname> <given-names>A.</given-names></name> <name><surname>Song</surname> <given-names>W.</given-names></name></person-group> (<year>2021</year>). <article-title>Effects of wearing face masks while using different speaking styles in noise on speech intelligibility during the COVID-19 pandemic.</article-title> <source><italic>Front. Psychol.</italic></source> <volume>12</volume>:<fpage>682677</fpage>. <pub-id pub-id-type="doi">10.3389/fpsyg.2021.682677</pub-id> <pub-id pub-id-type="pmid">34295288</pub-id></citation></ref>
</ref-list>
<glossary>
<title>Abbreviations</title>
<def-list id="DL1">
<def-item><term>CHL</term><def><p>children with hearing loss</p></def></def-item>
<def-item><term>CNH</term><def><p>children with normal hearing</p></def></def-item>
<def-item><term>ANH</term><def><p>adults with normal hearing</p></def></def-item>
<def-item><term>SNR</term><def><p>signal-to-noise ratio</p></def></def-item>
<def-item><term>CDC</term><def><p>Centers for Disease Control and Prevention.</p></def></def-item>
</def-list>
</glossary>
</back>
</article>