<?xml version="1.0" encoding="UTF-8" standalone="no"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xml:lang="EN" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="editorial">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Neurosci.</journal-id>
<journal-title>Frontiers in Neuroscience</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Neurosci.</abbrev-journal-title>
<issn pub-type="epub">1662-453X</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fnins.2023.1192459</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Neuroscience</subject>
<subj-group>
<subject>Editorial</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Editorial: Insights in auditory cognitive neuroscience: 2021</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name><surname>Sch&#x000F6;nwiesner</surname> <given-names>Marc</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="aff" rid="aff2"><sup>2</sup></xref>
<xref ref-type="corresp" rid="c001"><sup>&#x0002A;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/92470/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Alain</surname> <given-names>Claude</given-names></name>
<xref ref-type="aff" rid="aff3"><sup>3</sup></xref>
<xref ref-type="aff" rid="aff4"><sup>4</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/2708/overview"/>
</contrib>
</contrib-group>
<aff id="aff1"><sup>1</sup><institution>Institute of Biology, Faculty of Life Sciences, Leipzig University</institution>, <addr-line>Leipzig</addr-line>, <country>Germany</country></aff>
<aff id="aff2"><sup>2</sup><institution>Department of Psychology, Facult&#x000E9; des Arts et des Sciences, Universit&#x000E9; de Montr&#x000E9;al</institution>, <addr-line>Montreal, QC</addr-line>, <country>Canada</country></aff>
<aff id="aff3"><sup>3</sup><institution>Department of Psychology, University of Toronto</institution>, <addr-line>Toronto, ON</addr-line>, <country>Canada</country></aff>
<aff id="aff4"><sup>4</sup><institution>Rotman Research Institute, Baycrest Hospital</institution>, <addr-line>Toronto, ON</addr-line>, <country>Canada</country></aff>
<author-notes>
<fn fn-type="edited-by"><p>Edited and reviewed by: Robert J. Zatorre, McGill University, Canada</p></fn>
<corresp id="c001">&#x0002A;Correspondence: Marc Sch&#x000F6;nwiesner <email>marcs&#x00040;uni-leipzig.de</email></corresp>
<fn fn-type="other" id="fn001"><p>This article was submitted to Auditory Cognitive Neuroscience, a section of the journal Frontiers in Neuroscience</p></fn></author-notes>
<pub-date pub-type="epub">
<day>11</day>
<month>04</month>
<year>2023</year>
</pub-date>
<pub-date pub-type="collection">
<year>2023</year>
</pub-date>
<volume>17</volume>
<elocation-id>1192459</elocation-id>
<history>
<date date-type="received">
<day>23</day>
<month>03</month>
<year>2023</year>
</date>
<date date-type="accepted">
<day>24</day>
<month>03</month>
<year>2023</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x000A9; 2023 Sch&#x000F6;nwiesner and Alain.</copyright-statement>
<copyright-year>2023</copyright-year>
<copyright-holder>Sch&#x000F6;nwiesner and Alain</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/"><p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p></license>
</permissions>
<related-article id="RA1" related-article-type="commentary-article" xlink:href="https://www.frontiersin.org/research-topics/26912/insights-in-auditory-cognitive-neuroscience-2021" ext-link-type="uri">Editorial on the Research Topic <article-title>Insights in auditory cognitive neuroscience: 2021</article-title>
</related-article>
<kwd-group>
<kwd>auditory</kwd>
<kwd>pitch</kwd>
<kwd>MMN (mismatch negativity)</kwd>
<kwd>speech-in-noise</kwd>
<kwd>hemispheric asymmetries</kwd>
<kwd>voice</kwd>
<kwd>what/where system</kwd>
<kwd>hearing disorders</kwd>
</kwd-group>
<counts>
<fig-count count="0"/>
<table-count count="0"/>
<equation-count count="0"/>
<ref-count count="12"/>
<page-count count="3"/>
<word-count count="1844"/>
</counts>
</article-meta>
</front>
<body>
<p>Imagine an expensive research and development meeting at a large company. The presenter: &#x0201C;We have our top people working on this. Our top people!&#x0201D; This is how we feel about the many recent breakthroughs in auditory cognitive neuroscience research. Researchers like Tim Griffiths, Robert Zatorre, Andrew Oxenham, and the other contributors to this Frontiers&#x00027; Research Topic have shaped and advanced the field for years. This collection of ten short perspective papers aims to provide a readable overview of several current (and, in many cases, timeless) topics in auditory cognitive neuroscience through the vantage point of some of the main actors. The papers are best enjoyed as a collection rather than independently because of the many interconnections between the topics they discuss, some of which we will point out here.</p>
<p>We start with topic of processing and representation of critical auditory features. The mechanism of pitch perception is among the oldest such topics in hearing science, going back to Strutt (<xref ref-type="bibr" rid="B12">1907</xref>). The brain encoding of time-based pitch cues has seen strong empirical support using delay-and-add noise in brain imaging studies (Griffiths et al., <xref ref-type="bibr" rid="B4">1998</xref>). A classical study by Oxenham et al. (<xref ref-type="bibr" rid="B8">2004</xref>) demonstrated that time-based cues are not sufficient and that pitch perception also requires correct cochlear frequency-to-place mapping of the spectral components of the stimulus. After over a 100 years of research, the relationship between these two cues in pitch perception and representation is still under debate. The perspective by <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/fnins.2022.1074752">Oxenham</ext-link> discusses recent developments and directions in the study of pitch coding and perception.</p>
<p>From pitch extraction is the extraction of voice features: Pascal Belin&#x00027;s discovery of the temporal voice area in 2001 (Belin et al., <xref ref-type="bibr" rid="B1">2000</xref>) opened up new research into the cortical processing of voices and non-speech vocal sounds. This area around the middle of the superior temporal sulcus responds more strongly to voices than other sounds. There is some discussion of whether this area is processing speech rather than voice information, which is reminiscent of the debate around whether the fusiform face area genuinely represents faces or any stimuli that observers have acquired expertise with (Gauthier et al., <xref ref-type="bibr" rid="B2">1999</xref>). Here <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/fnins.2022.1075288">Trapeau et al.</ext-link> present evidence-based arguments to support the role of the temporal voice area in genuine voice processing.</p>
<p>The mismatch negativity is one of the most popular neural metrics to study preattentive processing, predictive coding mechanisms, auditory memory, and many other phenomena. Its discovery in late 1978 by Finish psychologist Risto N&#x000E4;t&#x000E4;&#x000E4;nen created a paradigm shift in auditory neuroscience. <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/fnins.2022.1025763">Tervaniemi</ext-link> discusses the development of stimulation paradigms from simple sine tones to complex multi-feature sounds and paradigms, including recent efforts to achieve ecological validity in experiments with such tightly controlled and repetitive stimuli. These new developments will ensure that the mismatch negativity remains among the most significant and versatile tools in auditory cognitive neuroscience for years to come.</p>
<p>Our understanding of the function and organization of the human primary (core) auditory cortex needs to catch up to that of the visual cortex. The auditory core is much smaller than V1 and is divided into subfields, nested on the superior temporal gyrus. Several functional and anatomical markers have been discovered and allow some non-invasive access, for example, increased myelination (Sigalovsky et al., <xref ref-type="bibr" rid="B11">2006</xref>), the 40-Hz auditory steady-state response (Gutschalk et al., <xref ref-type="bibr" rid="B5">1999</xref>), or a peak in the slope of the magneto-encephalographic response at about 20 ms (L&#x000FC;tkenh&#x000F6;ner et al., <xref ref-type="bibr" rid="B6">2003</xref>). <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/fnins.2022.1075369">Simon et al.</ext-link> argues that early time-locked high gamma band responses to natural speech can track primary cortical activity, adding a robust and ecologically valid method to study primary auditory cortex function non-invasively.</p>
<p>We now turn to the organization of the auditory system. <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/fnins.2022.1075511">Zatorre</ext-link> provides a perspective of hemispherical asymmetries in music and speech processing, in which his group has contributed significant theoretical and empirical advances. This is a topic with deep historical roots going back to the recognition of lateralized language areas by Broca and Wernicke in the late 19th century. Zatorre unifies recent results on the processing of musical pitch patterns in auditory networks of the right hemisphere (and complementary lateralization of speech sounds) in the framework of spectrotemporal modulation processing. The paper discusses the importance of low-level differential sensitivity to acoustical features of communication sounds (bottom-up) and high-level modulation of asymmetries by learning, attention, or other top-down factors.</p>
<p>A central concept of sensory processing in the cortex is that of partially segregated streams with different functions. This idea was initially conceived to explain different sensitivities, and latencies in cortical fields along the visual pathway (Schneider, <xref ref-type="bibr" rid="B10">1969</xref>; Mishkin and Ungerleider, <xref ref-type="bibr" rid="B7">1982</xref>; Goodale and Milner, <xref ref-type="bibr" rid="B3">1992</xref>) and later applied to audition by Rauschecker and Tian (<xref ref-type="bibr" rid="B9">2000</xref>) with the proposal of &#x0201C;what&#x0201D; and &#x0201C;where&#x0201D; pathways. This idea was reconceptualized several times, and the dual pathways have lost their initial clear functional separation and are now often referred to by location. These ventral and dorsal processing streams originate in the secondary (belt) auditory cortex in rostral and caudal fields, which then connect to different downstream areas in the frontal and parietal cortex. A recurrent functional distinction that has held up since the original studies in non-human primates is that rostral fields tend to be more involved in sound recognition and caudal areas more in sound localization. <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/fnins.2022.1076374">Scott and Jasmin</ext-link> discuss the origins and recent developments of the dual stream concept and its interaction with speech and voice processing of simultaneous talkers.</p>
<p>The feedback or top-down auditory projections is another principle of brain organization with powerful implications. The cortico-fugal pathway, the thickest efferent projection in the human brain after the pyramidal tract, instructively illustrates this. <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/fnins.2023.1081295">McAlpine and de Hoz</ext-link> discuss how adaptation in such feedback pathways of the auditory system aids in adaptive en- and decoding of complex sounds by building a representation of their statistical structure at different time scales. Exploring these feedback loops at different granularities, from <italic>in vivo</italic> recording to human neuroimaging, may reveal the fundamental listening processes.</p>
<p>Finally, we turn to topics in more applied auditory neuroscience. <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/fnins.2023.1077344">Griffiths</ext-link> provides an overview of recent work in the lab on predicting speech-in-noise ability based on performance with non-speech material in basic auditory cognitive tests. Speech in noise perception is the most important human auditory capacity and a consistent problem for persons with hearing disorders. Such tests may reveal the basic auditory factors that determine speech-in-noise understanding and enable more robust, language-independent clinical diagnosis. <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3389/fpsyg.2022.967260">R&#x000F6;nnberg et al.</ext-link> discusses the ongoing trend of including more cognitive factors in this effort to add to the classical models based on system identification approaches to peripheral hearing mechanisms. He proposes the Ease of Language Understanding model, which models complex interactions of cognitive modules, such as the different memory systems, lexical access, and predictive and postdictive processes. Such models help to understand the perceptual consequences of hearing disorders and mirror the trend to include cognitive factors in hearing aids and rehabilitation.</p>
<sec sec-type="author-contributions" id="s1">
<title>Author contributions</title>
<p>All authors listed have made a substantial, direct, and intellectual contribution to the work and approved it for publication.</p>
</sec>
</body>
<back>
<ack><p>We would like to thank all authors and reviewers who participated in the special issue.</p>
</ack>
<sec sec-type="COI-statement" id="conf1">
<title>Conflict of interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec sec-type="disclaimer" id="s2">
<title>Publisher&#x00027;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
<ref-list>
<title>References</title>
<ref id="B1">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Belin</surname> <given-names>P.</given-names></name> <name><surname>Zatorre</surname> <given-names>R. J.</given-names></name> <name><surname>Lafaille</surname> <given-names>P.</given-names></name> <name><surname>Ahad</surname> <given-names>P.</given-names></name> <name><surname>Pike</surname> <given-names>B.</given-names></name></person-group> (<year>2000</year>). <article-title>Voice-selective areas in human auditory cortex</article-title>. <source>Nature</source> <volume>403</volume>, <fpage>309</fpage>&#x02013;<lpage>312</lpage> <pub-id pub-id-type="doi">10.1038/35002078</pub-id><pub-id pub-id-type="pmid">10659849</pub-id></citation></ref>
<ref id="B2">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gauthier</surname> <given-names>I.</given-names></name> <name><surname>Tarr</surname> <given-names>M. J.</given-names></name> <name><surname>Anderson</surname> <given-names>A. W.</given-names></name> <name><surname>Skudlarski</surname> <given-names>P.</given-names></name> <name><surname>Gore</surname> <given-names>J. C.</given-names></name></person-group> (<year>1999</year>). <article-title>Activation of the middle fusiform &#x02018;face area&#x00027; increases with expertise in recognizing novel objects</article-title>. <source>Nat. Neurosci</source>. <volume>2</volume>, <fpage>568</fpage>&#x02013;<lpage>573</lpage> <pub-id pub-id-type="doi">10.1038/9224</pub-id><pub-id pub-id-type="pmid">10448223</pub-id></citation></ref>
<ref id="B3">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Goodale</surname> <given-names>M. A.</given-names></name> <name><surname>Milner</surname> <given-names>A. D.</given-names></name></person-group> (<year>1992</year>). <article-title>Separate visual pathways for perception and action</article-title>. <source>Trends Neurosci</source>. <volume>15</volume>, <fpage>20</fpage>&#x02013;<lpage>25</lpage>. <pub-id pub-id-type="doi">10.1016/0166-2236(92)90344-8</pub-id><pub-id pub-id-type="pmid">1374953</pub-id></citation></ref>
<ref id="B4">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Griffiths</surname> <given-names>T. D.</given-names></name> <name><surname>B&#x000FC;chel</surname> <given-names>C.</given-names></name> <name><surname>Frackowiak</surname> <given-names>R. S.</given-names></name> <name><surname>Patterson</surname> <given-names>R. D.</given-names></name></person-group> (<year>1998</year>). <article-title>Analysis of temporal structure in sound by the human brain</article-title>. <source>Nat. Neurosci</source>. <volume>1</volume>, <fpage>422</fpage>&#x02013;<lpage>427</lpage> <pub-id pub-id-type="doi">10.1038/1637</pub-id><pub-id pub-id-type="pmid">10196534</pub-id></citation></ref>
<ref id="B5">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gutschalk</surname> <given-names>A.</given-names></name> <name><surname>Mase</surname> <given-names>R.</given-names></name> <name><surname>Roth</surname> <given-names>R.</given-names></name> <name><surname>Ille</surname> <given-names>N.</given-names></name> <name><surname>Rupp</surname> <given-names>A.</given-names></name> <name><surname>H&#x000E4;hnel</surname> <given-names>S.</given-names></name> <etal/></person-group>. (<year>1999</year>). <article-title>Deconvolution of 40 Hz steady-state fields reveals two overlapping source activities of the human auditory cortex</article-title>. <source>Clin. Neurophysiol.</source> <volume>110</volume>, <fpage>856</fpage>&#x02013;<lpage>868</lpage> <pub-id pub-id-type="doi">10.1016/S1388-2457(99)00019-X</pub-id><pub-id pub-id-type="pmid">10400199</pub-id></citation></ref>
<ref id="B6">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>L&#x000FC;tkenh&#x000F6;ner</surname> <given-names>B.</given-names></name> <name><surname>Krumbholz</surname> <given-names>K.</given-names></name> <name><surname>Lammertmann</surname> <given-names>C.</given-names></name> <name><surname>Seither-Preisler</surname> <given-names>A.</given-names></name> <name><surname>Steinstr&#x000E4;ter</surname> <given-names>O.</given-names></name> <name><surname>Patterson</surname> <given-names>R. D.</given-names></name></person-group> (<year>2003</year>). <article-title>Localization of primary auditory cortex in humans by magnetoencephalography</article-title>. <source>Neuroimage</source> <volume>18</volume>, <fpage>58</fpage>&#x02013;<lpage>66</lpage> <pub-id pub-id-type="doi">10.1006/nimg.2002.1325</pub-id><pub-id pub-id-type="pmid">12507443</pub-id></citation></ref>
<ref id="B7">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mishkin</surname> <given-names>M.</given-names></name> <name><surname>Ungerleider</surname> <given-names>L. G.</given-names></name></person-group> (<year>1982</year>). <article-title>Contribution of striate inputs to the visuospatial functions of parieto-preoccipital cortex in monkeys</article-title>. <source>Behav. Brain Res</source>. <volume>6</volume>, <fpage>57</fpage>&#x02013;<lpage>77</lpage>. <pub-id pub-id-type="doi">10.1016/0166-4328(82)90081-X</pub-id><pub-id pub-id-type="pmid">7126325</pub-id></citation></ref>
<ref id="B8">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Oxenham</surname> <given-names>A. J.</given-names></name> <name><surname>Bernstein</surname> <given-names>J. G.</given-names></name> <name><surname>Penagos</surname> <given-names>H.</given-names></name></person-group> (<year>2004</year>). <article-title>Correct tonotopic representation is necessary for complex pitch perception</article-title>. <source>Proc. Natl. Acad. Sci. U. S. A</source>. <volume>101</volume>, <fpage>1421</fpage>&#x02013;<lpage>1425</lpage> <pub-id pub-id-type="doi">10.1073/pnas.0306958101</pub-id><pub-id pub-id-type="pmid">14718671</pub-id></citation></ref>
<ref id="B9">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Rauschecker</surname> <given-names>J. P.</given-names></name> <name><surname>Tian</surname> <given-names>B.</given-names></name></person-group> (<year>2000</year>). <article-title>Mechanisms and streams for processing of &#x0201C;what&#x0201D; and &#x0201C;where&#x0201D; in auditory cortex</article-title>. <source>Proc. Natl. Acad. Sci. U. S. A</source>. <volume>97</volume>, <fpage>11800</fpage>&#x02013;<lpage>11806</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.97.22.11800</pub-id><pub-id pub-id-type="pmid">11050212</pub-id></citation></ref>
<ref id="B10">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Schneider</surname> <given-names>G. E.</given-names></name></person-group> (<year>1969</year>). <article-title>Two visual systems</article-title>. <source>Science</source> <volume>163</volume>, <fpage>895</fpage>&#x02013;<lpage>902</lpage> <pub-id pub-id-type="doi">10.1126/science.163.3870.895</pub-id><pub-id pub-id-type="pmid">5763873</pub-id></citation></ref>
<ref id="B11">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sigalovsky</surname> <given-names>I. S.</given-names></name> <name><surname>Fischl</surname> <given-names>B.</given-names></name> <name><surname>Melcher</surname> <given-names>J. R.</given-names></name></person-group> (<year>2006</year>). <article-title>Mapping an intrinsic MR property of gray matter in auditory cortex of living humans: a possible marker for primary cortex and hemispheric differences</article-title>. <source>Neuroimage</source> <volume>32</volume>, <fpage>1524</fpage>&#x02013;<lpage>1537</lpage> <pub-id pub-id-type="doi">10.1016/j.neuroimage.2006.05.023</pub-id><pub-id pub-id-type="pmid">16806989</pub-id></citation></ref>
<ref id="B12">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Strutt</surname> <given-names>J. W.</given-names></name></person-group> (<year>1907</year>). <article-title>On our perception of sound direction</article-title>. <source>Philos. Mag.</source> <volume>13</volume>, <fpage>214</fpage>&#x02013;<lpage>232</lpage>.</citation>
</ref>
</ref-list> 
</back>
</article>