<?xml version="1.0" encoding="utf-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" article-type="research-article" dtd-version="2.3" xml:lang="EN">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Psychol.</journal-id>
<journal-title>Frontiers in Psychology</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Psychol.</abbrev-journal-title>
<issn pub-type="epub">1664-1078</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fpsyg.2023.1233176</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Psychology</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Markers of schizophrenia at the prosody/pragmatics interface. Evidence from corpora of spontaneous speech interactions</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes"><name><surname>Saccone</surname> <given-names>Valentina</given-names></name><xref rid="c001" ref-type="corresp"><sup>&#x002A;</sup></xref>
<uri xlink:href="https://loop.frontiersin.org/people/2253261/overview"/>
</contrib>
<contrib contrib-type="author"><name><surname>Trillocco</surname> <given-names>Simona</given-names></name>
</contrib>
<contrib contrib-type="author"><name><surname>Moneglia</surname> <given-names>Massimo</given-names></name>
<uri xlink:href="https://loop.frontiersin.org/people/1916323/overview"/>
</contrib>
</contrib-group>
<aff><institution>LABLITA Laboratory, Department of &#x201C;Lettere e Filosofia&#x201D;, University of Florence</institution>, <addr-line>Florence</addr-line>, <country>Italy</country></aff>
<author-notes>
<fn fn-type="edited-by" id="fn0029">
<p>Edited by: Gloria Gagliardi, University of Bologna, Italy</p>
</fn>
<fn fn-type="edited-by" id="fn0030">
<p>Reviewed by: Przemys&#x0142;aw Zakowicz, Poznan University of Medical Sciences, Poland; Luca Bischetti, University Institute of Higher Studies in Pavia, Italy</p>
</fn>
<corresp id="c001">&#x002A;Correspondence: Valentina Saccone, <email>valentina.saccone@unifi.it</email></corresp>
</author-notes>
<pub-date pub-type="epub">
<day>12</day>
<month>10</month>
<year>2023</year>
</pub-date>
<pub-date pub-type="collection">
<year>2023</year>
</pub-date>
<volume>14</volume>
<elocation-id>1233176</elocation-id>
<history>
<date date-type="received">
<day>01</day>
<month>06</month>
<year>2023</year>
</date>
<date date-type="accepted">
<day>27</day>
<month>09</month>
<year>2023</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x00A9; 2023 Saccone, Trillocco and Moneglia.</copyright-statement>
<copyright-year>2023</copyright-year>
<copyright-holder>Saccone, Trillocco and Moneglia</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/">
<p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p>
</license>
</permissions>
<abstract>
<p>The speech of individuals with schizophrenia exhibits atypical prosody and pragmatic dysfunctions, producing monotony. The paper presents the outcomes of corpus-based research on the prosodic features of the pathology as they manifest in real-life spontaneous interactions. The research relies on a corpus of schizophrenic speech recorded during psychiatric interviews (CIPPS) compared to a sampling of non-pathological speech derived from the LABLITA corpus of spoken Italian, which has been selected according to comparability requirements. Corpora has been intensively analyzed in the Language into Act Theory (L-AcT) frame, which links prosodic cues and pragmatic values. A cluster of linguistic parameters marked by prosody has been considered: utterance boundaries, information structure, speech disfluency, and prosodic prominence. The speech flow of patients turns out to be organized into small chunks of information that are shorter and scarcely structured, with an atypical proportion of post-nuclear information units (Appendix). It is pervasively scattered with silences, especially with long pauses between utterances and long silences at turn-taking. Fluency is hindered by retracing phenomena that characterize complex information structures. The acoustic parameters that give rise to prosodic prominence (f0 mean, f0 standard deviation, spectral emphasis, and intensity variation) have been measured considering the pragmatic roles of the prosodic units, distinguishing prominences within the illocutionary units (Comment) from those characterizing Topic units. Patients show a flattening of the Comment-prominence, reflecting impairments in performing the illocutionary activity. Reduced values of spectral emphasis and intensity variation also suggest a lack of engagement in communication. Conversely, Topic-prominence shows higher values for f0 standard deviation and spectral emphasis, suggesting effort when defining the domain of relevance of the illocutionary force. When comparing Topic and Comment-prominences of patients, the former consistently exhibit higher values across all parameters. In contrast, the non-pathological group displays the opposite pattern.</p>
</abstract>
<kwd-group>
<kwd>schizophrenic speech</kwd>
<kwd>prosodic-pragmatic correlates</kwd>
<kwd>information structure</kwd>
<kwd>prominence</kwd>
<kwd>disfluency</kwd>
</kwd-group>
<counts>
<fig-count count="6"/>
<table-count count="9"/>
<equation-count count="0"/>
<ref-count count="78"/>
<page-count count="17"/>
<word-count count="11668"/>
</counts>
<custom-meta-wrap>
<custom-meta>
<meta-name>section-at-acceptance</meta-name>
<meta-value>Psychology of Language</meta-value>
</custom-meta>
</custom-meta-wrap>
</article-meta>
</front>
<body>
<sec sec-type="intro" id="sec1">
<label>1.</label>
<title>Introduction</title>
<p>Language and communication dysfunction characterize all the symptoms of schizophrenia. Verbal communication impairments appear among the symptoms as positive/negative thought disorder (<xref ref-type="bibr" rid="ref42">Liddle et al., 2002</xref>; <xref ref-type="bibr" rid="ref40">Kuperberg, 2010</xref>; <xref ref-type="bibr" rid="ref29">DSM, 2013</xref>). The literature widely describes patients&#x2019; &#x201C;thought disorders,&#x201D; including poverty of speech, disorganization in the discourse, which is hard to follow, derailment and tangentiality with a loosening of associations (<xref ref-type="bibr" rid="ref10">Bleuler, 1950</xref>; <xref ref-type="bibr" rid="ref4">Andreasen, 1986</xref>; <xref ref-type="bibr" rid="ref29">DSM, 2013</xref>). The impairments lead to difficulties in interpersonal communication for patients (<xref ref-type="bibr" rid="ref32">Elvev&#x00E5;g et al., 2010</xref>) and damage pragmatic abilities, contributing to social dysfunction (<xref ref-type="bibr" rid="ref13">Bowie and Harvey, 2008</xref>); moreover, it is possible to underline correlations between types of schizophrenic pathology and linguistic functioning (<xref ref-type="bibr" rid="ref6">Bambini et al., 2022</xref>), the damage of which is associated with a reduced brain specialization (<xref ref-type="bibr" rid="ref16">Cavelti et al., 2018</xref>; <xref ref-type="bibr" rid="ref11">Boer et al., 2020</xref>). These phenomena depict an overall monotony in schizophrenic speech (<xref ref-type="bibr" rid="ref26">Dovetto et al., 2015</xref>; <xref ref-type="bibr" rid="ref21">Cresti and Moneglia, 2017</xref>).</p>
<p>To assess the psychopathology of schizophrenia, numerous evaluation scales have been employed since the 1960s.<xref rid="fn0001" ref-type="fn">
<sup>1</sup></xref> However, it has become evident that these scales rely on human judgment, necessitating fresh approaches or analyses to interpret the symptomatic heterogeneity of the disease accurately (<xref ref-type="bibr" rid="ref6">Bambini et al., 2022</xref>), characterized by variations from one individual to another and within the same individual at different disease stages.</p>
<p>The present research focuses on the qualitative evaluation of linguistic profiles within schizophrenia. It deals with prosodic and pragmatic features that characterize speech productions in spontaneous interactions and takes a corpus-based approach. We will search for markers of schizophrenic speech at three levels of the prosody/pragmatic interface, which in principle may be responsible for the monotony effect: (a) the informational complexity of the utterance; (b) the disfluencies of the speech flow; and (c) the prosodic prominence of the information units.</p>
<p>The research exploits an existing dataset of spontaneous speech of a small number of patients (4 schizophrenic subjects) compared with a control group (23 speakers), which is not sex-aged matched. The validity of the quantitative difference between the number of schizophrenic patients (<italic>n</italic> =&#x2009;4) and the control group (<italic>n</italic> =&#x2009;23) lies in the corpus-linguistics method. For qualitative analyses, the comparison group is restricted to 4 speakers to guarantee the relevance of the comparison. The analysis should be considered as a preliminary proof of concept study.</p>
<p>The Language into Act Theory (L-AcT) is the theoretical framework adopted for the research. L-AcT focuses on the pragmatic role played by prosody in speech organization and is specifically designed for spontaneous speech corpora analysis (<xref ref-type="bibr" rid="ref19">Cresti, 2000</xref>; <xref ref-type="bibr" rid="ref22">Cresti and Moneglia, 2018</xref>). The framework provides explicit methods for speech segmentation into utterances (<xref ref-type="bibr" rid="ref51">Moneglia, 2005</xref>) and for the annotation of information structure that are based on the hypothesis of a systematic correspondence between prosodic units and information functions (<xref ref-type="bibr" rid="ref19">Cresti, 2000</xref>; <xref ref-type="bibr" rid="ref53">Moneglia and Raso, 2014</xref>). L-AcT has been extensively applied to spoken Romance languages and tested on English, Japanese, and Chinese (<xref ref-type="bibr" rid="ref22">Cresti and Moneglia, 2018</xref>; <xref ref-type="bibr" rid="ref23">Cresti et al., forthcoming</xref>). Among the main achievements, the C-ORAL-ROM &#x2013; C-ORAL-BRASIL collections of comparable spoken romance corpora (Italian, French Spanish; European Portuguese; Basilian Portuguese (<xref ref-type="bibr" rid="ref96">Cresti and Moneglia, 2005</xref>; <xref ref-type="bibr" rid="ref150">Raso and Mello, 2013</xref>), the DBIPIC crosslinguistic Information structure Data Base (<xref ref-type="bibr" rid="ref130">Panunzi and Gregori, 2012</xref>), which allows comparative studies of speech organization in Italian, Spanish, English, and Brazilian Portuguese, and a Corpus-based Taxonomy of Illocution Acts based on the prosodic performance (<xref ref-type="bibr" rid="ref95">Cresti, 2020</xref>). Praat (<xref ref-type="bibr" rid="ref12">Boersma and Weenink, 2021</xref>) and Winpitch (<xref ref-type="bibr" rid="ref49">Martin, 2004</xref>) voice analysis software are the analysis tools.</p>
<p>L-AcT has already generated studies focusing on schizophrenia in Italian and Brazilian Portuguese. It has been made the hypothesis that patients have a specific difficulty in building up utterances presenting a Topic (<xref ref-type="bibr" rid="ref57">Rocha et al., 2022</xref>, <xref ref-type="bibr" rid="ref56">forthcoming</xref>) while they show an atypical preference for post-nuclear units (Appendix; <xref ref-type="bibr" rid="ref26">Dovetto et al., 2015</xref>; <xref ref-type="bibr" rid="ref22">Cresti and Moneglia, 2018</xref>). This difficulty seems to emerge in complex discourse contexts where patients do less structured speech productions, with a statistically significant decrease in Topic and a relevant increase in Appendix (<xref ref-type="bibr" rid="ref18">Costa, 2022</xref>). In addition, for what concerns Italian, it has been highlighted that schizophrenic speech records an abnormal quantity of pauses and retracing phenomena (<xref ref-type="bibr" rid="ref59">Saccone and Trillocco, 2022</xref>), and that pauses characterize schizophrenic speech, specifically in turn-taking position (in line with <xref ref-type="bibr" rid="ref45">Lucarini et al., 2022</xref>).</p>
<p>The paper is organized as follows. In 3.1, the complexity in schizophrenic speech is studied compared to controls by observing the amount of information in the utterance in terms of its length (MLU) and from the point of view of its informational complexity. Results, which only partially fit the expectations, give a measure of the atypical profile of schizophrenic speech considering the individual variability of patients. Values scored by patients will be compared to the controls and the general measures available for Italian (<xref ref-type="bibr" rid="ref20">Cresti, 2005</xref>, p. 227; <xref ref-type="bibr" rid="ref58">Saccone, 2022</xref>).</p>
<p>In 3.2, based on the segmentation of the speech flow into utterances and information units, a fine-grained analysis of disfluencies will be presented. Disfluencies, which strongly characterize schizophrenic speech, refer to hesitation phenomena and indicate the speaker&#x2019;s effort in planning, production, and post-articulatory evaluation (<xref ref-type="bibr" rid="ref35">Ginzburg et al., 2014</xref>). Disfluencies are dysfunctional (<xref ref-type="bibr" rid="ref1">Allwood, 2017</xref>), &#x201C;disturb&#x201D; the flow of communication (<xref ref-type="bibr" rid="ref31">Eklund, 2004</xref>), and are also pervasive in everyday language performance (<xref ref-type="bibr" rid="ref19">Cresti, 2000</xref>). Pauses and retracing phenomena have been investigated face to their possible positions inside the turn and considering their qualitative characteristics.</p>
<p>Finally, in 3.3 prosodic analysis of pathological speech has been carried out, in line with the most recent research (<xref ref-type="bibr" rid="ref24">Dickey et al., 2012</xref>; <xref ref-type="bibr" rid="ref17">Compton et al., 2018</xref>; <xref ref-type="bibr" rid="ref46">Lucarini et al., 2020</xref>, <xref ref-type="bibr" rid="ref44">2021</xref>, <xref ref-type="bibr" rid="ref45">2022</xref>). The focus is on prosodic prominences, a perceptual phenomenon that emphasizes linguistic segments compared to the surrounding context (<xref ref-type="bibr" rid="ref34">Gagliardi et al., 2012</xref>; <xref ref-type="bibr" rid="ref43">Lombardi Vallauri, 2014</xref>; <xref ref-type="bibr" rid="ref8">Barbosa, 2019</xref>). Prominence is determined by a complex interaction of prosodic and phonetic/acoustic parameters, essentially pitch and force accents. Pitch accent refers to fundamental frequency values, while force accent refers to intensity and duration.</p>
<p>The relevance of the prosodic prominence parameter in schizophrenia is highlighted in <xref ref-type="bibr" rid="ref50">Mart&#x00ED;nez-S&#x00E1;nchez et al. (2015)</xref>: at the nucleus&#x2019; syllabic level, slowness in the movement of the f0 and different realization of risings (peaks) and fallings (valleys) emerge with lower values in patients. In particular, the greater the number of years since diagnosis, the lower the intrasyllabic trajectories of f0, and the greater the amount of time since the last relapse, the less intrasyllabic trajectories of f0.</p>
<p>Further studies underline a direct correlation between a lowering of f0 and negative symptoms of schizophrenia (see aprosody in <xref ref-type="bibr" rid="ref17">Compton et al., 2018</xref>) as well as different pathologies such as depression (<xref ref-type="bibr" rid="ref160">Silva et al., 2021</xref>), mutational falsetto, laryngeal carcinoma, and vocal cord polyps (<xref ref-type="bibr" rid="ref100">Li et al., 2021</xref>).</p>
<p>Following the L-AcT approach, we will analyze acoustic indices specifically in the nucleus of the illocutionary unit of Comment and in the nucleus of the Topic Information Units whose prosodic profile presents prominence. To this end, we used the automatic script of <xref ref-type="bibr" rid="ref9">Barbosa et al. (2019)</xref>, which provides parameters to measure the movements of f0 and its variation. Spectral emphasis and intensity variation have also been calculated, correlating with a lack of engagement in communicative events (cf. <xref ref-type="bibr" rid="ref140">Pellet-Rostaing et al., 2023</xref>).</p>
<p>The paper aims to highlight distinctive properties of the speech flow in patients with schizophrenia through empirical research and data retrieved specifically from spontaneous speech corpora. Spontaneous spoken language is the field of communication in which idea processing needs to be synchronized with the interaction; thus, observing patients&#x2019; speech in a spontaneous interactive environment enables us to examine the actual context in which the linguistic outcomes of the pathology manifest.</p>
</sec>
<sec sec-type="materials|methods" id="sec2">
<label>2.</label>
<title>Materials and methods</title>
<sec id="sec3">
<label>2.1.</label>
<title>Data collections</title>
<p>The research relies on a case study of schizophrenic speech recorded during psychiatric interviews (Corpus of Italian Spoken Pathological/Schizophrenic CIPPS, Dovetto and Gemelli,<xref rid="fn0002" ref-type="fn">
<sup>2</sup></xref> 2013; <xref ref-type="bibr" rid="ref28">Dovetto et al., 2021</xref>), which has been intensively analyzed from the perspective of pragmatic and acoustic studies (<xref ref-type="bibr" rid="ref21">Cresti and Moneglia, 2017</xref>; <xref ref-type="bibr" rid="ref59">Saccone and Trillocco, 2022</xref>; <xref ref-type="bibr" rid="ref23">Cresti et al., forthcoming</xref>) in comparison with a control-group of non-pathological spontaneous speech derived from the LABLITA corpus of spoken Italian<xref rid="fn0003" ref-type="fn">
<sup>3</sup></xref> (<xref ref-type="bibr" rid="ref23">Cresti et al., forthcoming</xref>).</p>
<p>CIPPS collects about 9&#x2009;h of recordings (44.270 tokens; 6.707 utterances) of 4 male speakers with Schizophrenia aged 35&#x2013;45. Patients originate from Naples and metropolitan areas and are conventionally identified as A, B, C, and D.</p>
<p>The recording sessions are in the form of medical interviews between each patient and the psychiatrist and mainly consist of monologic excerpts due to the low presence of the doctor&#x2019;s turns. The interviews are about daily habits or topics the patient wants to discuss. They have been originally manually transcribed with orthographic criteria based on <xref ref-type="bibr" rid="ref60">Savy (2005)</xref>. Transcripts have been adapted to the CHAT-LABLITA format (<xref ref-type="bibr" rid="ref52">Moneglia and Cresti, 1997</xref>; <xref ref-type="bibr" rid="ref47">MacWhinney, 2000</xref>, <xref ref-type="bibr" rid="ref110">2012</xref>), comprehending prosodic and pragmatic annotations.</p>
<p>The four patients differ in the severity of the pathology and are characterized by different subtypes of schizophrenia (no longer considered in the DSM5), reflected in the speech flow.<xref rid="fn0004" ref-type="fn">
<sup>4</sup></xref></p>
<p>The clinical characterization of the patients in CIPPS follows the approach of phenomenological psychiatry,<xref rid="fn0005" ref-type="fn">
<sup>5</sup></xref> which was strongly influenced by Husserl&#x2019;s philosophy (<xref ref-type="bibr" rid="ref99">Jaspers, 1963</xref>) and Heidegger&#x2019;s existentialism (<xref ref-type="bibr" rid="ref92">Binswanger, 1942</xref>). This perspective considers that, in the realm of the human, the explanation of behavior through the observation of regularity and patterns (<italic>Erkl&#x00E4;rende Psychologie</italic>) must be supplemented by an understanding of the &#x201C;meaning-relations&#x201D; experienced by human beings (<italic>Verstehende Psychologie</italic>). Patients&#x2019; experience is accessed through the clinician&#x2019;s ability to &#x201C;identify&#x201D; with his psychic states (Jaspers). The clinical interviews collected in CIPPS are part of this attempt and are characterized by the maximum possible spontaneity and empathy.</p>
<p>In short, the diagnoses joint to the original data collection are as follows:</p>
<list list-type="alpha-upper">
<list-item>
<p>Pre-delusional condition of <italic>Wahnstimmung</italic> without hallucinations.</p>
</list-item>
<list-item>
<p>Paranoid schizophrenia with unstructured delirium without hallucinations.</p>
</list-item>
<list-item>
<p>Paranoid schizophrenia with structured delirium and hallucinations.</p>
</list-item>
<list-item>
<p>Paranoid schizophrenia with delirium.</p>
</list-item>
</list>
<p><xref rid="tab1" ref-type="table">Table 1</xref> gives a summary of the corpus.</p>
<table-wrap position="float" id="tab1">
<label>Table 1</label>
<caption>
<p>Summary of CIPPS data.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top">Patient</th>
<th align="center" valign="top">Recording duration</th>
<th align="center" valign="top">Stretch of speech</th>
<th align="center" valign="top">Tokens</th>
<th align="center" valign="top">Utterances</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="top">A</td>
<td align="center" valign="top">2&#x2009;h 30&#x2009;m</td>
<td align="center" valign="top">1&#x2009;h 3&#x2009;m</td>
<td align="center" valign="top">2,563</td>
<td align="center" valign="top">619</td>
</tr>
<tr>
<td align="left" valign="top">B</td>
<td align="center" valign="top">3&#x2009;h 58&#x2009;m</td>
<td align="center" valign="top">3&#x2009;h 43&#x2009;m</td>
<td align="center" valign="top">30,021</td>
<td align="center" valign="top">4,204</td>
</tr>
<tr>
<td align="left" valign="top">C</td>
<td align="center" valign="top">2&#x2009;h 8&#x2009;m</td>
<td align="center" valign="top">1&#x2009;h 26&#x2009;m</td>
<td align="center" valign="top">10,409</td>
<td align="center" valign="top">1,552</td>
</tr>
<tr>
<td align="left" valign="top">D</td>
<td align="center" valign="top">28&#x2009;m</td>
<td align="center" valign="top">17&#x2009;m</td>
<td align="center" valign="top">1,277</td>
<td align="center" valign="top">332</td>
</tr>
<tr>
<td align="left" valign="top">Tot.</td>
<td align="center" valign="top">9&#x2009;h 4&#x2009;m</td>
<td align="center" valign="top">6&#x2009;h 29&#x2009;m</td>
<td align="center" valign="top">44,270</td>
<td align="center" valign="top">6,707</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>The context of the clinical interview of CIPPS is not replicable in a non-pathological population. For instance, the therapeutic goal influences the relationship; the doctor tries not to interrupt the patient and stimulates his language activities. The control group corpus (CORCON) collects 3&#x2009;h and 57&#x2009;min of spontaneous speech of 23 healthy controls recorded during interviews in a friendly and motivating environment on various subjects, such as the speaker&#x2019;s life, work, habits, and family. For each recording, the interviewer is a friend or a well-known person by the main speaker. Most speakers are from Central Italy. Since this control group is not balanced in terms of age, gender, diatopic, diaphasic, and diastratic characteristics (see <xref rid="tab2" ref-type="table">Table 2</xref>), two subsets have been selected for specific analyses (SAMP and SAMP(100)). The main control group only compares the mean length of terminated sequences (MLU) and silences within the speech flow.</p>
<table-wrap position="float" id="tab2">
<label>Table 2</label>
<caption>
<p>Summary of groups&#x2019; demographic data.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th/>
<th align="left" valign="top">Gender</th>
<th align="left" valign="top">Age</th>
<th align="left" valign="top">Geographic origin</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="top">CIPPS</td>
<td align="left" valign="top">4 men</td>
<td align="center" valign="top">35&#x2013;45</td>
<td align="left" valign="top">Naples</td>
</tr>
<tr>
<td align="left" valign="top">CORCON</td>
<td align="left" valign="top">17 men, 6 women</td>
<td align="center" valign="top">26&#x2013;40 (6), 41&#x2013;60 (10), &#x003E;60 (7)</td>
<td align="left" valign="top">Florence (15), Siena (3), Arezzo (2), Milano (1), L&#x2019;Aquila (1), Terni (1)</td>
</tr>
<tr>
<td align="left" valign="top">SAMP</td>
<td align="left" valign="top">4 men</td>
<td align="center" valign="top">35&#x2013;45</td>
<td align="left" valign="top">Florence</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>SAMP was used for fine-grained analyses, such as information structure and the retracing phenomena, for which we need a more precise comparison selection concerning gender, age, and qualitative features of the interaction. To reduce the differences with the communicative context of CIPPS, SAMP selects four interviews, three about the work experience in life and one on the psychological problems experienced in family life, thus maintaining the presence of a main speaker and a solid motivation to interact in the intersubjective relation.<xref rid="fn0006" ref-type="fn">
<sup>6</sup></xref></p>
<p>SAMP(100) is a balanced subset of SAMP consisting of each speaker&#x2019;s first 100 terminated sequences; it was used for fine-grained acoustic research on prosodic prominence.</p>
<p><xref rid="tab3" ref-type="table">Table 3</xref> gives a summary of CORCON and the two subsets.</p>
<table-wrap position="float" id="tab3">
<label>Table 3</label>
<caption>
<p>Summary of control groups speech data.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top">Control group</th>
<th align="left" valign="top">Speakers</th>
<th align="center" valign="top">Stretch of speech</th>
<th align="left" valign="top">Tokens</th>
<th align="left" valign="top">Utterances</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="top">CORCON</td>
<td align="center" valign="top">23</td>
<td align="center" valign="top">3&#x2009;h 57&#x2009;m</td>
<td align="center" valign="top">34,398</td>
<td align="center" valign="top">4,016</td>
</tr>
<tr>
<td align="left" valign="top">SAMP</td>
<td align="center" valign="top">4</td>
<td align="center" valign="top">1&#x2009;h 8&#x2009;m</td>
<td align="center" valign="top">9,108</td>
<td align="center" valign="top">966</td>
</tr>
<tr>
<td align="left" valign="top">SAMP(100)</td>
<td align="center" valign="top">4</td>
<td align="center" valign="top">33&#x2009;m</td>
<td align="center" valign="top">4,639</td>
<td align="center" valign="top">436</td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
<sec id="sec4">
<label>2.2.</label>
<title>Methods and theoretical framework</title>
<p>The research is carried out within the Language into Act Theory (L-AcT, <xref ref-type="bibr" rid="ref19">Cresti, 2000</xref>; <xref ref-type="bibr" rid="ref53">Moneglia and Raso, 2014</xref>; <xref ref-type="bibr" rid="ref22">Cresti and Moneglia, 2018</xref>). According to L-AcT, the utterance is the primary referring unit for the analysis of spoken language, which results from pragmatic activities by the speaker; it is autonomous and conveys an illocutionary act. The segmentation of the speech flow into utterances is achieved through perceptual judgments into terminated sequences (TS) identified through their prosodic profile (<xref ref-type="bibr" rid="ref39">Izre'el et al., 2020</xref>). Subsequently, TSs are segmented into prosodic/information units, showing their information structure independently from their syntactic form. Thus, prosodic boundaries recognized in the speech flow provide its segmentation into utterances (terminal prosodic boundary, &#x2018;//&#x2019;) and smaller chunks, i.e., prosodic-information units (non-terminal prosodic boundary, &#x2018;/&#x2019;).</p>
<p>Through prosody, it is also possible to define which unit inside the utterance bears the illocution and, therefore, carries the pragmatic and prosodic autonomy of the sequence; this unit is named Comment (COM) and is necessary and sufficient to form an utterance. The prosodic contour of the COM can be described as a <italic>root</italic> unit (<xref ref-type="bibr" rid="ref61">&#x2018;t Hart et al., 1990</xref>); it widely varies as a function of its illocutionary value.</p>
<p>According to L-AcT, utterances can be simple or complex regarding their information structure: a simple utterance consists of only one prosodic/information unit, necessarily a COM bearing an illocutionary value (see example 1); conversely, a complex utterance consists of more than one prosodic/information unit, one of which is always the COM (see example 2, in which the COM is underlined).</p>
<list list-type="order">
<list-item>
<p>faccio un po&#x2019; di tutto // [LABLITA: prvmnl01-cami]</p>
</list-item>
</list>
<disp-quote>
<p>I do a bit of everything//</p>
</disp-quote>
<list list-type="order">
<list-item>
<p>e poi/niente// [LABLITA: prvmnl01-cami]</p>
</list-item>
</list>
<disp-quote>
<p>and then/nothing//</p>
</disp-quote>
<p>When an utterance is complex, the COM is supported by other units. Therefore, apart from the units that bear the illocution, for our goals, it is relevant to introduce two units identified within the L-AcT theoretical framework: Topic and Appendix.</p>
<p>Following <xref ref-type="bibr" rid="ref53">Moneglia and Raso (2014)</xref>, the Topic (TOP) provides the field of application for the illocutionary force of the Comment; it supplies the semantic representation of the domain of facts to which the illocutionary act refers (&#x201C;pragmatic aboutness&#x201D;). That is, utterances without a TOP necessarily refer to the context. Regarding its distribution, TOP units always precede the COM and have a <italic>prefix</italic> prosodic contour (<xref ref-type="bibr" rid="ref61">&#x2018;t Hart et al., 1990</xref>; <xref ref-type="bibr" rid="ref15">Cavalcante, 2016</xref>). On the other hand, the Appendix (APC) integrates the text of the COM and necessarily follows it. APC is performed with a <italic>suffix</italic> prosodic contour (in &#x2018;t Hart&#x2019;s terms) and does not have functional prosodic prominence (<xref ref-type="bibr" rid="ref23">Cresti et al., forthcoming</xref>).</p>
<p>Identifying these units leans on recognizing and perceiving relevant prosodic movements &#x2013; <italic>root</italic> for COM; <italic>prefix</italic> for TOP; <italic>suffix</italic> for APC. Both <italic>prefix</italic> and <italic>root</italic> prosodic contours can comprise a preparation and a nucleus. The nucleus corresponds to the minimal prosodic contour sufficient to perform the information unit; its contour can be composed of a simple movement (rising/falling/holding) or several movements aligned to the syllables participating in the contour (<xref ref-type="bibr" rid="ref97">Cresti and Moneglia, 2023</xref>); thus it is possible to identify a prosodically prominent part in both units of Topic and Comment whose relevance is connected to their functional value.</p>
<p>See in (3) an example of a complex utterance with the information structure of TOP/COM/APC; <xref rid="fig1" ref-type="fig">Figure 1</xref> shows the prosodic contour and the text labeled following the information tags.</p>
<list list-type="order">
<list-item>
<p>allora / i&#x2019; camionista /<sup>TOP</sup> ho iniziato a venti / tre anni /<sup>COM</sup> a farlo //<sup>APC</sup> [LABLITA: prvmnl01-cami]</p>
</list-item>
</list>
<fig position="float" id="fig1">
<label>Figure 1</label>
<caption>
<p>Annotation of a complex utterance.</p>
</caption>
<graphic xlink:href="fpsyg-14-1233176-g001.tif"/>
</fig>
<disp-quote>
<p>so/ the trucker/I started at twenty/three/doing it//</p>
</disp-quote>
<p>In <xref rid="fig1" ref-type="fig">Figure 1</xref>, the prominences of TOP and COM are circled in red. They include the rising movement, the peak, and the falling movement.</p>
<p>The previous examples (1), (2), and (3) show utterances in which only one unit bears the illocutionary force (&#x2018;faccio un po&#x2019; di tutto&#x2019; in 1; &#x2018;niente&#x2019; in 2; &#x2018;ho iniziato a venti / tre anni&#x2019; in 3); however, empirical studies in spontaneous speech, in particular in monologs, led to the identification of a different kind of terminated sequences in which more than one unit bears an illocution. It is usually the case of long excerpts of speech flow, in which the speaker develops a thought through a chain of semantic <italic>foci</italic>, and the illocution tends to remain unchanged (usually assertive). See an example in (4):</p>
<list list-type="order">
<list-item>
<p>l&#x2019; ho fatto per diversi anni / poi mi sono messo in proprio / s&#x2019; &#x00E8; creato una piccola azienda / da una piccola azienda viene poi / quell&#x2019; altra / e via // [LABLITA: prvmnl01-cami]</p>
</list-item>
</list>
<disp-quote>
<p>I did it for several years /then I branched out on my own /we set up a small company/from a small company then comes/ another/and so on//</p>
</disp-quote>
<p>Each unit in (4) bears a weak illocution. This type of TS, named <italic>stanza</italic>, has specific characteristics such as a monotonous prosodic trend and a &#x201C;step-by-step&#x201D; adjunctive structure. They are usually present where the implementation of speech is less interactive, as in monologs, and the speaker focuses on the semantic elaboration of the text (<xref ref-type="bibr" rid="ref20">Cresti, 2005</xref>; <xref ref-type="bibr" rid="ref55">Panunzi and Scarano, 2009</xref>; <xref ref-type="bibr" rid="ref58">Saccone, 2022</xref>). Inside a stanza, the units bearing an illocutionary value are named Bound Comments (COB) since they are linked together (bound) through prosodic and pragmatic features.</p>
<p>Assuming the L-AcT framework, automatic temporal and acoustic measurements of the signal are linked to the perceptual processing of linguistic data. The sound is aligned with the transcription and segmented both at the utterance level and, more specifically, at the information unit level.</p>
<p>Based on this multilayer annotation process, the analysis will explore (i) the structure and length of the utterance; (ii) speech disfluencies such as pauses and retracing phenomena (false starts, repetitions, corrections); (iii) a chosen set of acoustic parameters that highlight perceptual prosodic correlates of the schizophrenic atypia (mainly based on f0 and intensity).</p>
<p>On the first point, according to L-AcT, the audio files are segmented into TS (utterances and stanzas), and subsequently in information units. The segmentation in TS allows the quantitative measurement of their length in word numbers, while the segmentation in information units allows the qualitative measure of the information strategies adopted by each speaker.</p>
<p>Regarding pauses, as already stated in <xref ref-type="bibr" rid="ref4">Andreasen (1986)</xref> and cf. <xref ref-type="bibr" rid="ref42">Liddle et al. (2002)</xref>, one of the symptoms of schizophrenia is <italic>blocking</italic>, i.e., the interruption of thought followed by a phase of silence that can last from a few seconds to a few minutes. In <xref ref-type="bibr" rid="ref36">Goldman-Eisler (1961)</xref> and <xref ref-type="bibr" rid="ref7">Banfi (1999)</xref>, the length of the pauses is a clear distinction between pathological and non-pathological speech, and in <xref ref-type="bibr" rid="ref14">Cannizzaro et al. (2005)</xref> the abnormal quantity of silence is highlighted as a clear marker of patients&#x2019; speech. The most recent linguistic studies, albeit with different approaches, confirm these results (<xref ref-type="bibr" rid="ref37">Heldner and Edlund, 2010</xref>; <xref ref-type="bibr" rid="ref33">Fors, 2011</xref>; <xref ref-type="bibr" rid="ref25">Dodane and Hirsch, 2018</xref>; <xref ref-type="bibr" rid="ref6">Bambini et al., 2022</xref>). <xref ref-type="bibr" rid="ref44">Lucarini et al. (2021)</xref> do a conversation analysis of schizophrenic speech and observe a specific correlation between pause duration and negative symptoms.<xref rid="fn0007" ref-type="fn">
<sup>7</sup></xref></p>
<p>CIPPS and CORCON audio files are segmented into &#x201C;sounding&#x201D; and &#x201C;silent&#x201D; based on Praat&#x2019;s script. All silences over 150&#x2009;ms are considered and grouped quantitatively by duration thresholds and qualitatively by their position. Exploiting the L-AcT approach, position labeling distinguishes pauses between utterances of the same turn and between information units within the utterance. Moreover, considering the latest generation typological approach (cf. <italic>inter-tours</italic> and <italic>intra-tours</italic> in <xref ref-type="bibr" rid="ref25">Dodane and Hirsch, 2018</xref>; <italic>gaps</italic>/<italic>lapses</italic> and <italic>pauses</italic> in <xref ref-type="bibr" rid="ref37">Heldner and Edlund, 2010</xref>; <xref ref-type="bibr" rid="ref33">Fors, 2011</xref>), each silent is labeled according to the following types:</p>
<list list-type="bullet">
<list-item>
<p>T (&#x003C;turns): When the pause occurs between the turns of the two different speakers, it is, in principle, an index of the interviewed responsiveness in the intersubjective interaction. Therefore, the count of pauses T is limited only to pauses &#x201C;before&#x201D; the turn because they are an index of the patient&#x2019;s reaction time to the interlocutor&#x2019;s questions<xref rid="fn0008" ref-type="fn">
<sup>8</sup></xref></p>
</list-item>
<list-item>
<p>UT (&#x003C;utterances): When the pause occurs between two utterances of the same turn by the same speaker, it refers in principle to the difficulty of maintaining the turn programming a new speech act.</p>
</list-item>
<list-item>
<p>IU (&#x003C;informational units): When the pause occurs between two information units of the same utterance, it deals with the problems in conceiving the locutionary content of the information unit.</p>
</list-item>
</list>
<p>One added value of the CHAT/LABLITA transcription is the annotation of retracing phenomena such as hesitations, repeated words or fragments of words, false starts, and repairs. Often considered an <italic>error</italic> (<xref ref-type="bibr" rid="ref38">Hieke, 1981</xref>) or, more generally, an <italic>alteration</italic> (<xref ref-type="bibr" rid="ref35">Ginzburg et al., 2014</xref>), retracing is a fragmentation of the locutionary program, which is widely present in spontaneous speech performance (<xref ref-type="bibr" rid="ref19">Cresti, 2000</xref>). In our transcription format, the symbols &#x002A; and [/] respectively mark a retracted unit&#x2019;s beginning and end. The system allows accuracy in identifying the retracing events and the number of retracted tokens. Data were analyzed based on the different positions in the terminated sequences (at the very beginning of a TS -Start of TS-; inside a TS -Inside TS-; and at the beginning of an information unit -Start of IU-inside TS), distinguishing between isolated episodes and successions of retracing, called <italic>chains</italic>.</p>
<p>Lastly, to highlight perceptual prosodic correlates of the schizophrenic atypia, prominences are manually identified on Praat for each COM- and TOP unit. Four acoustic parameters are selected for each prominence: (i) f0 mean, the mean of the average number of oscillations of the vocal folds per second, starting parameters for the voice description; (ii) f0 standard deviation, which measures the variability of the f0 (connected to the neuromuscular control and the regularity of laryngeal vibration of the vocal folds in <xref ref-type="bibr" rid="ref101">Lopes et al., 2017</xref>); (iii) Spectral emphasis, which measures the vocal effort (<xref ref-type="bibr" rid="ref62">Traunm&#x00FC;ller and Eriksson, 2000</xref>) and correlates with the energy expended during the speech flow; and (vi) the coefficient of intensity variation, which reports the ratio between the mean and the intensity standard deviation.</p>
</sec>
</sec>
<sec sec-type="results" id="sec5">
<label>3.</label>
<title>Results</title>
<sec id="sec6">
<label>3.1.</label>
<title>The structure of the utterance</title>
<p>The direct relation between prosody and pragmatics foreseen by the L-AcT theoretical framework allows for outlining a first sketch of the linguistic complexity and productivity in the 4 patients compared to the control groups based on the annotation of the terminated sequences and their division into prosodic units. We will first observe the measurements for the Mean Length of Utterance (MLU); subsequently, we will report data about the inner structure of the terminated sequences (information structure).</p>
<sec id="sec7">
<label>3.1.1.</label>
<title>Mean length of utterance</title>
<p>The MLU reflects the complexity of the spoken structures in terms of the number of words contributing to the semantic content of a TS.<xref rid="fn0009" ref-type="fn">
<sup>9</sup></xref> The analysis has been carried out on the whole set of corpora under consideration (CIPPS and CORCON). <xref rid="fig2" ref-type="fig">Figure 2</xref> and <xref rid="tab4" ref-type="table">Table 4</xref> show the measurements of length for each utterances, the mean values per patient (colored box plots), and the collected measurements for the control group (distribution in the gray box plot).</p>
<fig position="float" id="fig2">
<label>Figure 2</label>
<caption>
<p>Length of utterances.</p>
</caption>
<graphic xlink:href="fpsyg-14-1233176-g002.tif"/>
</fig>
<table-wrap position="float" id="tab4">
<label>Table 4</label>
<caption>
<p>MLU values.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th/>
<th align="left" valign="top">MLU</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="top">A</td>
<td align="left" valign="top">4.3</td>
</tr>
<tr>
<td align="left" valign="top">B</td>
<td align="left" valign="top">6.5</td>
</tr>
<tr>
<td align="left" valign="top">C</td>
<td align="left" valign="top">6.9</td>
</tr>
<tr>
<td align="left" valign="top">D</td>
<td align="left" valign="top">5</td>
</tr>
<tr>
<td align="left" valign="top">CORCON</td>
<td align="left" valign="top">mean per speaker: 5.2&#x2013;15.7 (average value: 8.8)</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>The 4 boxes of CIPPS extend behind the CORCON mean (indicated with an &#x2018;x&#x2019; inside the gray box), and when considering whiskers, the CIPPS extension never exceeds that of CORCON. Patients B (blue box) and C (green box) show, on average, closer proximity to the controls (B: 6.5; C: 6.9; CORCON: 8.8 words/utterance), while A (red box) and D (pink box) exhibit lower MLU values (A: 4.3; D: 5). For patients A and D, more than a quarter of their utterances consist of a single word, whereas this applies to only 1 out of 20 of the CORCON&#x2019;s utterances. The high peaks in the control group variation (mean maximum rate: 15.7 words/utterance) correlate with the monologic context of the recordings.<xref rid="fn0010" ref-type="fn">
<sup>10</sup></xref> Schizophrenic speech is characterized by qualitatively shorter utterances, where the discourse is structured in smaller chunks.</p>
<p>To assess statistical significance, the Kruskal-Wallis test for not normally distributed data has been conducted, but it did not yield <italic>p</italic>-values &#x003C;0.05 (A: <italic>p</italic>-value&#x2009;=&#x2009;0.1409; B: <italic>p</italic>-value&#x2009;=&#x2009;0.578; C: <italic>p</italic>-value&#x2009;=&#x2009;0.2688; D: <italic>p</italic>-value&#x2009;=&#x2009;0.1029).</p>
</sec>
<sec id="sec8">
<label>3.1.2.</label>
<title>Information structure and complexity</title>
<p>Further analysis has been performed to examine how TSs are structured and evaluate their complexity, considering whether TSs give rise to utterances or stanzas and whether utterances consist of a single COM unit or are structured from an informational point of view. The analysis has been processed on a CIPPS Sample of 4,892 tokens (755 terminated sequences); the chosen excerpts are the first 15&#x2009;min of each patient.<xref rid="fn0011" ref-type="fn">
<sup>11</sup></xref> TSs have been segmented into units and labeled following their prosodic form and information function; hence, data concerning the information structure were extracted.<xref rid="fn0012" ref-type="fn">
<sup>12</sup></xref> Schizophrenic data are compared with the control group SAMP. <xref rid="tab5" ref-type="table">Table 5</xref> presents the comparison. For these parameters, applying a statistical significance test was impossible as the initial samples were not calibrated for statistical comparison.</p>
<table-wrap position="float" id="tab5">
<label>Table 5</label>
<caption>
<p>Information structure: classification of terminated sequences.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th/>
<th align="center" valign="top">Terminated sequences</th>
<th align="center" valign="top">Simple utterances</th>
<th align="center" valign="top">Complex utterances</th>
<th align="center" valign="top">Stanzas</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="top">A</td>
<td align="center" valign="top">77</td>
<td align="center" valign="top">39 (50.6%)</td>
<td align="center" valign="top">32 (41.6%)</td>
<td align="center" valign="top">6 (7.8%)</td>
</tr>
<tr>
<td align="left" valign="top">B</td>
<td align="center" valign="top">274</td>
<td align="center" valign="top">125 (45.6%)</td>
<td align="center" valign="top">109 (39.8%)</td>
<td align="center" valign="top">40 (14.6%)</td>
</tr>
<tr>
<td align="left" valign="top">C</td>
<td align="center" valign="top">202</td>
<td align="center" valign="top">86 (42.6%)</td>
<td align="center" valign="top">84 (41.6%)</td>
<td align="center" valign="top">32 (15.8%)</td>
</tr>
<tr>
<td align="left" valign="top">D</td>
<td align="center" valign="top">202</td>
<td align="center" valign="top">93 (46.0%)</td>
<td align="center" valign="top">84 (41.6%)</td>
<td align="center" valign="top">25 (12.4%)</td>
</tr>
<tr>
<td align="left" valign="top">CIPPS</td>
<td align="center" valign="top">755</td>
<td align="center" valign="top">343 (45.4%)</td>
<td align="center" valign="top">309 (40.9%)</td>
<td align="center" valign="top">103 (13.7%)</td>
</tr>
<tr>
<td align="left" valign="top">SAMP</td>
<td align="center" valign="top">966</td>
<td align="center" valign="top">304 (14.2&#x2013;42.6%; mean: 31.4%)</td>
<td align="center" valign="top">448 (38.7&#x2013;58.2%; mean: 46.4%)</td>
<td align="center" valign="top">214 (15.6&#x2013;35.0%; mean 22.2%)</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>Regarding the frequency of simple utterances, the average value for the control group in SAMP (31.4%) aligns with the trend of Italian monologic informal speech observed in previous studies (<xref ref-type="bibr" rid="ref20">Cresti, 2005</xref>, p. 227), i.e., 30.5%. However, the variation among the four speakers is high (14.2&#x2013;42.6%); two speakers produce nearly 15% of simple utterances, while the others are close to 43%. Despite individual differences, complex TSs (complex utterances and stanzas) overtake simple ones in non-pathological speech. In contrast, the trend of schizophrenic patients is less heterogeneous and shows a reduced gap between simple and complex TSs. CIPPS simple utterances always outnumber the control percentage (&#x2265;42.6%): For patient A, simple utterances go slightly beyond half of the total (50.6% simple), while in the other three (B, C, and D), the percentage of complex TSs increases moving closer to the non-pathological distribution.</p>
<p>Beyond the relation between simple utterances and complex TSs, <xref rid="tab5" ref-type="table">Table 5</xref> shows the frequency of stanzas. As pointed out in the method section, a high presence of stanzas is expected in monologs. Previous corpus-based studies (<xref ref-type="bibr" rid="ref58">Saccone, 2022</xref>) reveal that in Italian speech, the number of stanzas increases from 6.3% of TSs in dialogs/conversations to 19.8% in monologs.<xref rid="fn0013" ref-type="fn">
<sup>13</sup></xref> The recurrence of stanzas allows the speaker to extend his turn, performing his thought chunk by chunk, using small pieces of information, each with a weak illocutionary value. Using these macrostructures requires the speaker to have an overall idea of what should be said, even if the content can be progressively planned during the production of the discourse. Given these premises, we might expect a low presence of stanzas in schizophrenic speech where thoughts are, in principle, less organized. Again, the variation of the percentage of stanzas among the four controls is high (15.6&#x2013;35.0%) with an average of 22.2% of TSs. CIPPS&#x2019; rates are approximately under the minimum of the controls (15.6%), and, again, the value decreases to 7.8% for patient A.</p>
<p>Data, therefore, indicate a tendency of CIPPS patients to reduce the informational complexity of the speech flow, both about the information structure of the utterance (as expected in <xref ref-type="bibr" rid="ref26">Dovetto et al., 2015</xref>; <xref ref-type="bibr" rid="ref22">Cresti and Moneglia, 2018</xref>) and also about the stanzas.</p>
</sec>
<sec id="sec9">
<label>3.1.3.</label>
<title>Information units</title>
<p>Lastly, the inner composition of TS (complex utterances and stanzas) has been analyzed by looking at the frequency of Topic (TOP) and Appendix (APC) units. Data show relevant intersubjective variation for non-pathological and schizophrenic speech, as summarized in <xref rid="tab6" ref-type="table">Table 6</xref>. The reported values indicate the percentages of TSs with TOP/APC. Also, applying a statistical significance test was impossible for these parameters as the initial samples were not calibrated for statistical comparison.</p>
<table-wrap position="float" id="tab6">
<label>Table 6</label>
<caption>
<p>Information structure: presence of topic and appendix.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th/>
<th align="center" valign="top">Terminated sequences</th>
<th align="center" valign="top">TOP</th>
<th align="center" valign="top">APC</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="top">A</td>
<td align="center" valign="top">77</td>
<td align="center" valign="top">6 (7.8%)</td>
<td align="center" valign="top">6 (7.8%)</td>
</tr>
<tr>
<td align="left" valign="top">B</td>
<td align="center" valign="top">274</td>
<td align="center" valign="top">71 (25.9%)</td>
<td align="center" valign="top">17 (6.2%)</td>
</tr>
<tr>
<td align="left" valign="top">C</td>
<td align="center" valign="top">202</td>
<td align="center" valign="top">36 (17.8%)</td>
<td align="center" valign="top">15 (7.4%)</td>
</tr>
<tr>
<td align="left" valign="top">D</td>
<td align="center" valign="top">202</td>
<td align="center" valign="top">24 (12.4%)</td>
<td align="center" valign="top">22 (10.9%)</td>
</tr>
<tr>
<td align="left" valign="top">CIPPS</td>
<td align="center" valign="top">786</td>
<td align="center" valign="top">137 (17.4%)</td>
<td align="center" valign="top">60 (7.6%)</td>
</tr>
<tr>
<td align="left" valign="top">SAMP</td>
<td align="center" valign="top">966</td>
<td align="center" valign="top">315 (4.9&#x2013;64.1%; mean: 32.6%)</td>
<td align="center" valign="top">57 (3.7&#x2013;8.8%; mean: 5.9%)</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>Regarding TOP, both groups show a variable behavior, especially the control one. CIPPS values are always beyond the control&#x2019;s mean (&#x003C;32.6); while staying in the lower part of the distribution, they are still included in the range of variation of SAMP. On the other hand, the presence of APC shows the opposite trend: all four patients&#x2019; values are distributed above the controls&#x2019; mean (&#x003E;5.9%), and in one case (D), APC frequency overcomes the controls&#x2019; maximum (10.9%).</p>
<p>The TOP is more frequent than the APC in every speaker (except A, which shows the same number of both). Still, the CIPPS trend is remarkably different from the non-pathological ones since the percentages for the two units in schizophrenic patients are much closer, which leads to a higher relative frequency of APC. Indeed, the reported number of APCs is noteworthy, showing a marked preference for delocalizing and defocusing information in the right periphery of the utterance.</p>
</sec>
</sec>
<sec id="sec10">
<label>3.2.</label>
<title>Disfluencies</title>
<p>Based on the segmentation of the speech flow into TSs and information units, a fine-grained analysis of disfluencies is presented here, focusing on pauses and retracing phenomena, which have been investigated for their distribution inside the turn and their qualitative characteristics.<xref rid="fn0014" ref-type="fn">
<sup>14</sup></xref></p>
<sec id="sec11">
<label>3.2.1.</label>
<title>Pauses</title>
<p>For the analysis of the pauses, the Control Group is CORCON. Pauses have been automatically identified in the signal and manually classified in terms of inside/between utterances and turn-taking pauses and length (for a detailed description of the data processing, see <xref ref-type="bibr" rid="ref59">Saccone and Trillocco, 2022</xref>; <xref ref-type="bibr" rid="ref63">Trillocco, forthcoming</xref>). Related to the automatic identification, the sounding/silent script on Praat was used and manually checked by two revisors.<xref rid="fn0015" ref-type="fn">
<sup>15</sup></xref> In <xref rid="fig3" ref-type="fig">Figure 3</xref>, an excerpt from CIPPS (patient A) shows the abnormal length of pauses in schizophrenic speech: pink parts are pauses, and white parts are speech.</p>
<fig position="float" id="fig3">
<label>Figure 3</label>
<caption>
<p>Silent and sounding in schizophrenic speech.</p>
</caption>
<graphic xlink:href="fpsyg-14-1233176-g003.tif"/>
</fig>
<p>Silences so identified have been studied considering their position and duration.</p>
<p>The minimum threshold established (150&#x2009;ms) corresponds to the average duration of the stop consonants.<xref rid="fn0016" ref-type="fn">
<sup>16</sup></xref> According to the literature (<xref ref-type="bibr" rid="ref30">Duez, 1985</xref>; <xref ref-type="bibr" rid="ref27">Dovetto and Gemelli, 2013</xref>), only four duration thresholds have been considered: 150&#x2013;250&#x2009;ms, 251&#x2013;500&#x2009;ms, 501&#x2013;1,000&#x2009;ms, and&#x2009;&#x003E;&#x2009;1,001&#x2009;ms. The percentages of pauses of each type were then calculated in relation to their position. <xref rid="fig4" ref-type="fig">Figures 4</xref>, <xref rid="fig5" ref-type="fig">5</xref> present the results.</p>
<fig position="float" id="fig4">
<label>Figure 4</label>
<caption>
<p>Duration of pauses.</p>
</caption>
<graphic xlink:href="fpsyg-14-1233176-g004.tif"/>
</fig>
<fig position="float" id="fig5">
<label>Figure 5</label>
<caption>
<p>T-pauses in CIPPS.</p>
</caption>
<graphic xlink:href="fpsyg-14-1233176-g005.tif"/>
</fig>
<p>Firstly, we observe in <xref rid="fig4" ref-type="fig">Figure 4</xref> the comparison between CIPPS and CORCON: the length of transparent bars is shorter in the CIPPS for T pauses (26.15% vs. 71.55%) and UT pauses (58.41% vs. 70.43%), revealing the greater pervasiveness of silences in relation to turn-taking and between utterances. At the IU level, the two groups show a lower difference (68.61% vs. 79.13% without pauses).</p>
<p>The difference of the yellow bars (pauses &#x003E;1&#x2009;s) is the most evident data: their length is more extensive for the CIPPS regardless of the type of pause considered (IU: 9.39% vs. 1.40%; UT: 17.98% vs. 8.81%; T: 33.94% vs. 6.00%). In controls, these only sporadically exceed 2&#x2009;s; in the pathological, they can even exceed the 20s.</p>
<p>Moreover, the same trend is observed for the green bars UT and T (500&#x2013;1,000&#x2009;ms), longer than those of non-pathological speech (UT: 15.10% vs. 11.26; T: 22.16% vs. 10.73%), in line with expectations (<xref ref-type="bibr" rid="ref7">Banfi, 1999</xref>; <xref ref-type="bibr" rid="ref37">Heldner and Edlund, 2010</xref>).</p>
<p>The trend is markedly different regarding T pauses: while in most cases, there is no pause at the start of the turn in the non-pathological (71.55%), in the pathological, only 26.15% of turns do not present silences before. This difference does not regard short pauses (almost 5% in both corpora) but mainly pauses longer than 500&#x2009;ms. In short, pauses do not characterize locutionary programming but mainly occur between utterances and in a marked manner at turn-taking.</p>
<p>Observing the various patients confirms the peculiarity of long pauses at the turn&#x2019;s start. <xref rid="fig5" ref-type="fig">Figure 5</xref> reports individual differences: Patient A very rarely (11.11%) starts a turn without silence, while B, C, and D slightly more often (25.12, 37.68, and 30.68%). The turn-taking delay, recently observed by <xref ref-type="bibr" rid="ref45">Lucarini et al. (2022)</xref>, is confirmed.</p>
</sec>
<sec id="sec12">
<label>3.2.2.</label>
<title>Retracing phenomena</title>
<p>Retracing can be associated with both repetition (see examples 5, 6, and 7a) or modification (7b) of words; when the locutionary content is repeated, it can be total (5, 6) or partial (7a).</p>
<p>Retracing can occur in different positions of the terminated sequences: at the very beginning of a TS (Start of TS), otherwise inside a TS (Inside TS); the second case can be further split into two classes to isolate the retracing phenomena occurring inside a complex TS at the beginning of an information unit (Start of IU-inside TS). Retracing phenomena can occur in isolated episodes (5, 7a, 7b) or successions (6), called <italic>chains</italic>.</p>
<p>See the following examples:</p>
<list list-type="order">
<list-item>
<p>Start of TS</p>
</list-item>
</list>
<disp-quote>
<p>&#x002A;pe' [/] pe' dargli un colore pi&#x00F9; uniforme // [LABLITA: fammnl02-fale]</p>
</disp-quote>
<disp-quote>
<p>&#x002A;<italic>to</italic> [/] <italic>to give it a more uniform color</italic>//</p>
</disp-quote>
<list list-type="order">
<list-item>
<p>Start of IU-inside TS</p>
</list-item>
</list>
<disp-quote>
<p>i' ramo / &#x002A;&#x0026;d [/] &#x002A;&#x0026;d [/] d' un noce / gl' &#x00E8; pi&#x00F9; chiaro d' i' fusto // [LABLITA: fammnl02-fale]</p>
</disp-quote>
<disp-quote>
<p>the branch/ &#x002A;of [/] &#x002A;of [/] of a walnut/is lighter than the trunk//</p>
</disp-quote>
<list list-type="order">
<list-item>
<p>Inside TS</p>
</list-item>
</list>
<list list-type="alpha-lower">
<list-item>
<p>a casa mia &#x002A;s&#x2019; era [/] gl&#x2019; eran poveri / e quindi &#x2018;un c&#x2019; era / tanto da mangiare // [LABLITA: fammnl02-fale].</p>
</list-item>
</list>
<disp-quote>
<p>at home &#x002A;we were [/] they were poor / and so there wasn&#x2019;t / so much to eat //.</p>
</disp-quote>
<list list-type="alpha-lower">
<list-item>
<p>ha fatto &#x002A;le [/] il tecnico industriale // [LABLITA: pubdlr12-vefa].</p>
</list-item>
</list>
<disp-quote>
<p><italic>he went to &#x002A;the-PL-F</italic> [/] <italic>the-SN-M technical industrial institute</italic> //.</p>
</disp-quote>
<p>Once all the retracing phenomena of CIPPS and control groups had been labeled, data were analyzed to verify possible differences.</p>
<sec id="sec13">
<label>3.2.2.1.</label>
<title>Retracted tokens and units</title>
<p>We analyzed the phenomenon concerning the number of tokens produced (retracted tokens vs. total tokens) and the number of information units in which speech is articulated (retracing phenomena on information units) in both corpora.<xref rid="fn0017" ref-type="fn">
<sup>17</sup></xref> The results, summarized in <xref rid="fig6" ref-type="fig">Figure 6</xref>, show the tendency to produce retracing phenomena in schizophrenic speech.<xref rid="fn0018" ref-type="fn">
<sup>18</sup></xref></p>
<fig position="float" id="fig6">
<label>Figure 6</label>
<caption>
<p>Retracted/total tokens.</p>
</caption>
<graphic xlink:href="fpsyg-14-1233176-g006.tif"/>
</fig>
<p>While the box represents the distribution of values in the control group, the colored dots indicate the 4 CIPPS patients, showing the incidence of the retracing phenomena on the number of tokens. All the patients (A: 13.19%; B: 11.44%; C: 6.72%; D: 10.02%) outnumber the mean distribution of the controls (1.43&#x2013;4.76%).</p>
</sec>
<sec id="sec14">
<label>3.2.2.2.</label>
<title>Single episodes and chains</title>
<p>A fine-grained analysis is carried out to display the quantity and typology (single episodes/chains) of retracing phenomena related to the different types of terminated sequences. CIPPS data refer to the extract of <xref rid="tab6" ref-type="table">Table 6</xref>, while those for the control group refer to SAMP.</p>
<p>In <xref rid="tab7" ref-type="table">Table 7</xref>, the frequency of single retracting episodes is reported. The values are calculated by dividing the number of single episodes by the number of not retracted information units, according to the different types of terminated sequences in which they appear:</p>
<table-wrap position="float" id="tab7">
<label>Table 7</label>
<caption>
<p>Single episodes and chains: CIPPS and control group.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th/>
<th/>
<th align="center" valign="top">Terminated sequences</th>
<th align="center" valign="top">Simple utterances</th>
<th align="center" valign="top">Complex utterances</th>
<th align="center" valign="top">Stanzas</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="top" rowspan="2">Single episodes</td>
<td align="left" valign="middle">CIPPS</td>
<td align="center" valign="middle">9.02%</td>
<td align="center" valign="middle">7.63%</td>
<td align="center" valign="middle">11.26%</td>
<td align="center" valign="middle">14.53%</td>
</tr>
<tr>
<td align="left" valign="middle">SAMP</td>
<td align="center" valign="middle">5.30%</td>
<td align="center" valign="middle">3.38%</td>
<td align="center" valign="middle">5.70%</td>
<td align="center" valign="middle">5.47%</td>
</tr>
<tr>
<td align="left" valign="top" rowspan="2">Chains</td>
<td align="left" valign="top">CIPPS</td>
<td align="center" valign="top">3.31%</td>
<td align="center" valign="top">2.93%</td>
<td align="center" valign="top">3.45%</td>
<td align="center" valign="top">3.33%</td>
</tr>
<tr>
<td align="left" valign="top">SAMP</td>
<td align="center" valign="top">1.15%</td>
<td align="center" valign="top">0.82%</td>
<td align="center" valign="top">1.66%</td>
<td align="center" valign="top">0.95%</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>The difference between the two groups manifests in the frequency of single retracing episodes for TS, which almost doubled in CIPPS (9.02% vs. 5.30%).</p>
<p>The trend remains roughly the same in the different types of terminated sequences: the percentage of single retracing episodes is more than double in simple (7.63% vs. 3.38%) and complex utterances (11.26% vs. 5.70%), while the greatest atypia is found in stanzas (14.53% vs. 5.47%), which is the type of TS less frequent in schizophrenic speech (13.7% vs. 22.2%, see <xref rid="tab5" ref-type="table">Table 5</xref>). This highlights the difficulty in CIPPS to produce a more complex structure (complex utterances and stanzas).</p>
<p>The difference between single episodes and chains concerns the &#x201C;intensity&#x201D; of the disfluency phenomenon: a retracing chain indicates greater difficulty processing a single information unit. Retracing chains generally appear to be a typical trait of stuttering but are also present in the non-pathological, albeit with very low percentages.<xref rid="fn0019" ref-type="fn">
<sup>19</sup></xref> <xref rid="tab7" ref-type="table">Table 7</xref> shows the frequency of retracing chains in the two corpora.</p>
<p>These data show that schizophrenic patients produce roughly three times more chains in simple utterances (2.93% vs. 0.82%), complex utterances (3.45% vs. 1.66%), and stanzas (3.33% vs. 0.95%). Therefore, this type of disfluency seems associated with the disease in a more substantial way than single episodes.</p>
</sec>
<sec id="sec15">
<label>3.2.2.3.</label>
<title>Distribution</title>
<p>Regarding the distribution of retracing inside the terminated sequence, it might be relevant to observe if a speaker retracts the first words of a terminated sequence (Start of TS) or retracts words when the unit and the TS are ongoing (Start of IU-inside TS or Inside TS). The two cases seem to respond to different causes of the retracing; the first is most likely related to uncertainty in building the locutionary content in its connection to the illocutionary programming, while the second concerns the locutionary level only since the illocutionary activity has already been conceived and planned.</p>
<p><xref rid="tab8" ref-type="table">Table 8</xref> summarizes the comparison between pathological and non-pathological speakers.<xref rid="fn0020" ref-type="fn">
<sup>20</sup></xref></p>
<table-wrap position="float" id="tab8">
<label>Table 8</label>
<caption>
<p>Distribution of retracing phenomena.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th/>
<th align="center" valign="top">Start of TS</th>
<th align="center" valign="top">Start of IU-inside TS</th>
<th align="center" valign="top">Inside TS</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="top">A</td>
<td align="center" valign="top">17.6%</td>
<td align="center" valign="top">41.0%</td>
<td align="center" valign="top">41.4%</td>
</tr>
<tr>
<td align="left" valign="top">B</td>
<td align="center" valign="top">10.6%</td>
<td align="center" valign="top">59.5%</td>
<td align="center" valign="top">29.9%</td>
</tr>
<tr>
<td align="left" valign="top">C</td>
<td align="center" valign="top">21.9%</td>
<td align="center" valign="top">35.5%</td>
<td align="center" valign="top">42.6%</td>
</tr>
<tr>
<td align="left" valign="top">D</td>
<td align="center" valign="top">16.1%</td>
<td align="center" valign="top">49.4%</td>
<td align="center" valign="top">34.5%</td>
</tr>
<tr>
<td align="left" valign="top">CIPPS</td>
<td align="center" valign="top">12.6%</td>
<td align="center" valign="top">54.9%</td>
<td align="center" valign="top">32.4%</td>
</tr>
<tr>
<td align="left" valign="top">CORCON</td>
<td align="center" valign="top">13.1%</td>
<td align="center" valign="top">46.1%</td>
<td align="center" valign="top">40.8%</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>In both corpora, retracing phenomena occur above all when the utterance is ongoing (Start of IU-inside TS and Inside TS), while the position Start of TS rarely holds retracted words (CIPPS: 12.6%; CORCON: 13.1%). This trend is emphasized in B, who reports the lowest percentage at the Start of TS (10.6%) and the highest uncertainty in the locutionary processing at the Start of IU, i.e., after the first information unit of a complex TS (59.5%). No specific pathological trend emerges from this study. According to this data set, retracing characterizes most as a disfluency at the locutionary level, and no particular influence by the pragmatic level can be noted in patients.</p>
</sec>
</sec>
</sec>
<sec id="sec16">
<label>3.3.</label>
<title>Prosodic prominence</title>
<p>To study the prosodic prominence as a possible marker of the pathology, we have independently analyzed acoustic indices in the illocutionary unit of Comment and in the unit of Topic.<xref rid="fn0021" ref-type="fn">
<sup>21</sup></xref></p>
<p>The nucleus of the <italic>root</italic> and <italic>prefix</italic> prosodic units, corresponding to the minimal prosodic contour sufficient to perform the information function, is perceptively identified and is selected as the prominence, as highlighted in <xref rid="fig1" ref-type="fig">Figure 1</xref>.</p>
<p>The perceptive choice is validated using the values of f0, intensity, and duration observable in the spectrogram. For the replicability of the procedure, the prominence is segmented on the speech wave following a specific workflow:</p>
<list list-type="bullet">
<list-item>
<p>The movement starts with a rise of the f0, reaches a peak, and ends with a fall.</p>
</list-item>
<list-item>
<p>Since the prominence can often concern only portions of words, in order not to break the semantic unity, it is arbitrarily established to include up to a maximum of two syllables before and two after the entire movement considered.</p>
</list-item>
<list-item>
<p>The segment thus identified is labeled with the number of syllables of which it is composed<xref rid="fn0022" ref-type="fn">
<sup>22</sup></xref>;</p>
</list-item>
</list>
<p>The analyses are conducted on the first 100 TSs for each patient and control of SAMP(100) using an automatic script (<xref ref-type="bibr" rid="ref9">Barbosa et al., 2019</xref>).<xref rid="fn0023" ref-type="fn">
<sup>23</sup></xref></p>
<p>The measured acoustic parameters for each prominence are f0 mean and f0 standard deviation<xref rid="fn0024" ref-type="fn">
<sup>24</sup></xref>; spectral emphasis (<italic>emph</italic>) that measures the vocal effort (<xref ref-type="bibr" rid="ref62">Traunm&#x00FC;ller and Eriksson, 2000</xref>); intensity variation coefficient (<italic>cvint</italic>) that reports the ratio between the mean and the intensity standard deviation.</p>
<p>In this case as well, to assess statistical significance, the Kruskal-Wallis test for not normally distributed data has been conducted, but it did not yield significant results (find in the footnotes below the report of the values per each parameter).</p>
<p><xref rid="tab9" ref-type="table">Table 9</xref> summarizes the results obtained for COM and TOP in the two corpora. Data are reported as a whole and for each speaker.</p>
<table-wrap position="float" id="tab9">
<label>Table 9</label>
<caption>
<p>Acoustic parameters of Comment and Topic prominences.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="center" valign="top" colspan="8">COM-prominences</th>
</tr>
<tr>
<th/>
<th align="center" valign="top">n COM</th>
<th align="center" valign="top">f0mean Hz</th>
<th align="center" valign="top">f0mean st</th>
<th align="center" valign="top">f0sd Hz</th>
<th align="center" valign="top">f0sd st</th>
<th align="center" valign="top">emph</th>
<th align="center" valign="top">cvint</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="top">A</td>
<td align="center" valign="top">89</td>
<td align="center" valign="top">59.29</td>
<td align="center" valign="top">69.30</td>
<td align="center" valign="top">18.82</td>
<td align="center" valign="top">1.96</td>
<td align="center" valign="top">0.58</td>
<td align="center" valign="top">1.04</td>
</tr>
<tr>
<td align="left" valign="top">B</td>
<td align="center" valign="top">83</td>
<td align="center" valign="top">101.43</td>
<td align="center" valign="top">75.64</td>
<td align="center" valign="top">63.10</td>
<td align="center" valign="top">7.60</td>
<td align="center" valign="top">3.56</td>
<td align="center" valign="top">4.14</td>
</tr>
<tr>
<td align="left" valign="top">C</td>
<td align="center" valign="top">71</td>
<td align="center" valign="top">100.44</td>
<td align="center" valign="top">75.82</td>
<td align="center" valign="top">44.39</td>
<td align="center" valign="top">6.01</td>
<td align="center" valign="top">2.35</td>
<td align="center" valign="top">3.95</td>
</tr>
<tr>
<td align="left" valign="top">D</td>
<td align="center" valign="top">69</td>
<td align="center" valign="top">83.39</td>
<td align="center" valign="top">72.33</td>
<td align="center" valign="top">44.17</td>
<td align="center" valign="top">4.81</td>
<td align="center" valign="top">1.69</td>
<td align="center" valign="top">2.1</td>
</tr>
<tr>
<td align="left" valign="top">
<bold>CIPPS</bold>
</td>
<td align="center" valign="top">
<bold>312</bold>
</td>
<td align="center" valign="top">
<bold>85.2</bold>
</td>
<td align="center" valign="top">
<bold>73.14</bold>
</td>
<td align="center" valign="top">
<bold>42.02</bold>
</td>
<td align="center" valign="top">
<bold>5.01</bold>
</td>
<td align="center" valign="top">
<bold>2.02</bold>
</td>
<td align="center" valign="top">
<bold>2.78</bold>
</td>
</tr>
<tr>
<td align="left" valign="top">cami</td>
<td align="center" valign="top">56</td>
<td align="center" valign="top">110.34</td>
<td align="center" valign="top">80.75</td>
<td align="center" valign="top">35.19</td>
<td align="center" valign="top">3.99</td>
<td align="center" valign="top">2.11</td>
<td align="center" valign="top">7.84</td>
</tr>
<tr>
<td align="left" valign="top">fale</td>
<td align="center" valign="top">46</td>
<td align="center" valign="top">108.57</td>
<td align="center" valign="top">80.61</td>
<td align="center" valign="top">26.37</td>
<td align="center" valign="top">3.37</td>
<td align="center" valign="top">3.56</td>
<td align="center" valign="top">9.57</td>
</tr>
<tr>
<td align="left" valign="top">pell</td>
<td align="center" valign="top">62</td>
<td align="center" valign="top">135.65</td>
<td align="center" valign="top">84.31</td>
<td align="center" valign="top">19.39</td>
<td align="center" valign="top">2.31</td>
<td align="center" valign="top">2.91</td>
<td align="center" valign="top">4.28</td>
</tr>
<tr>
<td align="left" valign="top">vefa</td>
<td align="center" valign="top">76</td>
<td align="center" valign="top">140.28</td>
<td align="center" valign="top">84.37</td>
<td align="center" valign="top">31.16</td>
<td align="center" valign="top">3.18</td>
<td align="center" valign="top">4.52</td>
<td align="center" valign="top">6.13</td>
</tr>
<tr>
<td align="left" valign="top">
<bold>SAMP(100)</bold>
</td>
<td align="center" valign="top">
<bold>240</bold>
</td>
<td align="center" valign="top">
<bold>126.02</bold>
</td>
<td align="center" valign="top">
<bold>82.79</bold>
</td>
<td align="center" valign="top">
<bold>28.14</bold>
</td>
<td align="center" valign="top">
<bold>3.18</bold>
</td>
<td align="center" valign="top">
<bold>3.36</bold>
</td>
<td align="center" valign="top">
<bold>6.72</bold>
</td>
</tr>
</tbody>
</table>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="center" valign="top" colspan="8">TOP-prominences</th>
</tr>
<tr>
<th/>
<th align="center" valign="top">n TOP</th>
<th align="center" valign="top">f0mean Hz</th>
<th align="center" valign="top">f0mean st</th>
<th align="center" valign="top">f0sd Hz</th>
<th align="center" valign="top">f0sd st</th>
<th align="center" valign="top">emph</th>
<th align="center" valign="top">cvint</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="top">A</td>
<td align="center" valign="top">7</td>
<td align="center" valign="top">107.43</td>
<td align="center" valign="top">73.29</td>
<td align="center" valign="top">37.67</td>
<td align="center" valign="top">3.23</td>
<td align="center" valign="top">1.76</td>
<td align="center" valign="top">1.8</td>
</tr>
<tr>
<td align="left" valign="top">B</td>
<td align="center" valign="top">23</td>
<td align="center" valign="top">142.48</td>
<td align="center" valign="top">82.39</td>
<td align="center" valign="top">66.80</td>
<td align="center" valign="top">6.99</td>
<td align="center" valign="top">4.83</td>
<td align="center" valign="top">3.38</td>
</tr>
<tr>
<td align="left" valign="top">C</td>
<td align="center" valign="top">16</td>
<td align="center" valign="top">155.38</td>
<td align="center" valign="top">84.88</td>
<td align="center" valign="top">53.67</td>
<td align="center" valign="top">6.81</td>
<td align="center" valign="top">3.86</td>
<td align="center" valign="top">4.44</td>
</tr>
<tr>
<td align="left" valign="top">D</td>
<td align="center" valign="top">3</td>
<td align="center" valign="top">109.33</td>
<td align="center" valign="top">79</td>
<td align="center" valign="top">33.25</td>
<td align="center" valign="top">5.69</td>
<td align="center" valign="top">2.93</td>
<td align="center" valign="top">2.33</td>
</tr>
<tr>
<td align="left" valign="top">
<bold>CIPPS</bold>
</td>
<td align="center" valign="top">
<bold>49</bold>
</td>
<td align="center" valign="top">
<bold>139.65</bold>
</td>
<td align="center" valign="top">
<bold>81.69</bold>
</td>
<td align="center" valign="top">
<bold>56.29</bold>
</td>
<td align="center" valign="top">
<bold>6.31</bold>
</td>
<td align="center" valign="top">
<bold>3.97</bold>
</td>
<td align="center" valign="top">
<bold>3.44</bold>
</td>
</tr>
<tr>
<td align="left" valign="top">cami</td>
<td align="center" valign="top">38</td>
<td align="center" valign="top">154.29</td>
<td align="center" valign="top">86.47</td>
<td align="center" valign="top">27.43</td>
<td align="center" valign="top">2.87</td>
<td align="center" valign="top">2.84</td>
<td align="center" valign="top">5.02</td>
</tr>
<tr>
<td align="left" valign="top">fale</td>
<td align="center" valign="top">29</td>
<td align="center" valign="top">117.52</td>
<td align="center" valign="top">82.45</td>
<td align="center" valign="top">13.10</td>
<td align="center" valign="top">1.60</td>
<td align="center" valign="top">3.38</td>
<td align="center" valign="top">6.59</td>
</tr>
<tr>
<td align="left" valign="top">pell</td>
<td align="center" valign="top">11</td>
<td align="center" valign="top">134.64</td>
<td align="center" valign="top">84.36</td>
<td align="center" valign="top">7.43</td>
<td align="center" valign="top">0.84</td>
<td align="center" valign="top">1.82</td>
<td align="center" valign="top">3</td>
</tr>
<tr>
<td align="left" valign="top">vefa</td>
<td align="center" valign="top">18</td>
<td align="center" valign="top">133.44</td>
<td align="center" valign="top">84.56</td>
<td align="center" valign="top">22.45</td>
<td align="center" valign="top">2.49</td>
<td align="center" valign="top">3.58</td>
<td align="center" valign="top">5.22</td>
</tr>
<tr>
<td align="left" valign="top">
<bold>SAMP (100)</bold>
</td>
<td align="center" valign="top">
<bold>96</bold>
</td>
<td align="center" valign="top">
<bold>137.02</bold>
</td>
<td align="center" valign="top">
<bold>84.66</bold>
</td>
<td align="center" valign="top">
<bold>19.77</bold>
</td>
<td align="center" valign="top">
<bold>2.17</bold>
</td>
<td align="center" valign="top">
<bold>3.03</bold>
</td>
<td align="center" valign="top">
<bold>5.32</bold>
</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<p>In bold, corpus average values.</p>
</table-wrap-foot>
</table-wrap>
<sec id="sec17">
<label>3.3.1.</label>
<title>f0mean</title>
<p>Comparing <italic>f&#x0030;mean</italic> in the two groups, we can observe that the values for Comment are lower in CIPPS (85.20&#x2009;Hz and 73.14st) than in SAMP(100) (126.02&#x2009;Hz and 82.79st). On the contrary, for the Topic, the values of the <italic>f0mean</italic> in CIPPS (139.65&#x2009;Hz and 81.69st) are similar to those of the non-pathological group (137.02&#x2009;Hz and 84.66 st).</p>
<p>In schizophrenic speech, there is a higher <italic>f0mean</italic> for TOP-prominences (81.69st) compared to the COM-prominences (73.141st). For the control group, the two values are roughly equivalent (82.79st for the COM-prominences and 84.65st for the TOP-prominences).<xref rid="fn0025" ref-type="fn">
<sup>25</sup></xref></p>
</sec>
<sec id="sec18">
<label>3.3.2.</label>
<title>f0sd</title>
<p>In principle, <italic>f0sd</italic> in COM-prominences might correlate with the variability of illocutions, so the initial hypothesis is that pathological speech, which is perceived as monotonous, might show low values of <italic>f0sd</italic>.</p>
<p>Nevertheless, although the <italic>f0mean</italic> values are lower for schizophrenic speech, the COM-prominences have higher <italic>f0sd</italic> values in CIPPS (42.02&#x2009;Hz and 5.01st) than in SAMP(100) (28.14&#x2009;Hz and 3.18st). The higher <italic>f0sd</italic> in pathological speech is even more evident for the TOP-prominences (56.29&#x2009;Hz and 6.31st), almost three times those of the control group (19.77&#x2009;Hz and 2.17st).</p>
<p>Further observations rely on the different features of the two information units. While there is a great variety of illocutions, we only know three prosodic profiles for the Topic (<xref ref-type="bibr" rid="ref15">Cavalcante, 2016</xref>); hence, a higher <italic>f0sd</italic> in the Comment might be expected. This hypothesis is confirmed by the data of non-pathological speech (COM-prominences: 48.01&#x2009;Hz and 5.42st vs. TOP-prominences: 36.14&#x2009;Hz and 4.16st); instead, in CIPPS, the <italic>f0sd</italic> values are lower for the Comment (42.02&#x2009;Hz and 5.01st) than for the Topic (56.29&#x2009;Hz and 6.31st).<xref rid="fn0026" ref-type="fn">
<sup>26</sup></xref></p>
<p>Although the reason for this finding in schizophrenic patients must still be investigated, it is worth noticing that the recorded qualitatively higher <italic>f0sd</italic> is consistent with previous studies on other pathologies (depression in <xref ref-type="bibr" rid="ref160">Silva et al., 2021</xref>; mutational falsetto, laryngeal carcinoma, and vocal cord polyps in <xref ref-type="bibr" rid="ref100">Li et al., 2021</xref>).</p>
</sec>
<sec id="sec19">
<label>3.3.3.</label>
<title>emph</title>
<p>Regarding COM-prominences, the <italic>emph</italic> is 2.02&#x2009;dB in CIPPS and 3.36&#x2009;dB in SAMP(100). Thus, the schizophrenic speakers put less vocal effort than the non-pathological in producing the prominences bearing the illocution, in correlation with lower values of <italic>f0mean</italic>.</p>
<p>No particular differences, instead, are identified for TOP-prominences between the two groups: values in CIPPS (3.97&#x2009;dB) are similar to those in SAMP(100) (3.03&#x2009;dB).<xref rid="fn0027" ref-type="fn">
<sup>27</sup></xref></p>
<p>In other words, the performance of the illocution results in an attitude of acoustic &#x201C;weakness,&#x201D; flattening, and less effort is recorded. The datum is even more relevant, considering that this does not regard TOP-prominence. Therefore, a possible correlation with the monotony effect seems to be relative specifically to COM-prominence (<xref ref-type="bibr" rid="ref17">Compton et al., 2018</xref>).</p>
</sec>
<sec id="sec20">
<label>3.3.4.</label>
<title>cvint</title>
<p>For our goal, the coefficient of intensity variation (<italic>cvint</italic>) is more reliable than the direct intensity measurement since, in our corpora, neither the distance from the microphone nor the angle between the microphone and the speaker&#x2019;s mouth was fixed, so altering the recorded intensity.</p>
<p>Again, given the monotony perceived in pathological speech, the starting hypothesis is that in CIPPS, <italic>cvint</italic> values are lower than those of the control group.</p>
<p>The data confirms expectations: for COM-prominences, the <italic>cvint</italic> is three times lower in CIPPS (2.78) than in non-pathological speech (6.72), while for TOP-prominences the difference is reduced (3.44 vs. 5.32).<xref rid="fn0028" ref-type="fn">
<sup>28</sup></xref></p>
</sec>
</sec>
</sec>
<sec sec-type="discussions" id="sec21">
<label>4.</label>
<title>Discussion</title>
<p>The analysis conducted on CIPPS and its comparison with the control group highlights the peculiarity of schizophrenic speech compared to the threshold values recorded in non-pathological trends of spontaneous dialogs in the various linguistic domains considered in this research. Results can be summarized as follows.</p>
<p>Regarding the structure of the TS, from a qualitative point of view, utterances are shorter in terms of MLU and less articulated in schizophrenic patients, but the intersubjective variability is high. All patients, however, prefer delocalized post-nuclear information units (Appendix) and, as expected, a low number of stanzas; thus, the speech is structured in smaller chunks and less organized from an informational point of view.</p>
<p>Moreover, the fluency is interrupted by an atypical number of very long pauses (1&#x2013;20&#x2009;s). Pauses do not occur in connection to the locutionary programming inside the utterance but mostly regard its pragmatic conception with a substantial turn-taking delay (cf. <xref ref-type="bibr" rid="ref2">Alpert et al., 2000</xref>; <xref ref-type="bibr" rid="ref45">Lucarini et al., 2022</xref>).</p>
<p>The quantity of retracing phenomena highlights patients&#x2019; difficulty in programming the locution; according to our findings, the incidence of retracing rises specifically when the discourse is structured in complex utterances and stanzas. Retracing chains turn out to be associated with the disease in a more substantial way than single episodes.</p>
<p>Lastly, the analysis of prominences brought about the following findings:</p>
<list list-type="bullet">
<list-item>
<p>In CIPPS, the nuclear part of the COM unit is characterized by lower values of <italic>f0mean</italic>, <italic>emph</italic>, and <italic>cvint,</italic> while the <italic>f0sd</italic> is higher. The prosodic parameters reflect an attitude of acoustic &#x201C;weakness&#x201D; of the performance of the illocution, which can be one of the causes of the perceived monotony. The lowering of the above values suggests an impairment in dealing with the variability of the illocutions and a lack of engagement in the communicative events (cf. <xref ref-type="bibr" rid="ref140">Pellet-Rostaing et al., 2023</xref>).</p>
</list-item>
<list-item>
<p>On the other hand, the measured values of the nuclear part of the TOP are lower for <italic>cvint</italic>, similar for <italic>f0mean</italic> but higher for <italic>f0sd</italic> and <italic>emph</italic> concerning the controls. This suggests that schizophrenic speech is characterized by greater effort when defining the Topic, i.e., the domain of illocutionary force.</p>
</list-item>
<list-item>
<p>The differences between COM- and TOP-prominences highlight the relevance of dividing the analysis for the two information units. Beyond the previous differences, COM- and TOP-prominences record a high variation between them for <italic>f0mean</italic>, <italic>f0sd</italic>, and <italic>emph</italic> in CIPPS, which is not found in SAMP(100). Moreover, in CIPPS, TOP-prominences record higher values than the COM according to all the detected parameters. In contrast, the control group follows the opposite trend, except for the <italic>f0mean,</italic> which varies in a limited manner. The different attitudes toward the performance of the two units could be an index of schizophrenic atypia.</p>
</list-item>
</list>
<p>All the findings have been processed to investigate whether the results have statistical relevance. The Kruskal-Wallis test for not normally distributed data has been used, and data for each patient have been compared to the control groups, although without reporting significant differences. Our sample sizes are not conducive to inferential statistics due to the preference for a corpus-based methodology, which represents spontaneous speech variability rather than verifying the behavior of two populations facing the same task.</p>
<p>The results discussed here shall be understood as a qualitative description and shed light on the specificity of schizophrenic linguistic profiles, which still need more extensive studies. Moreover, one implication of our analyses is to suggest future directions of investigation where the tests above highlight differences between the datasets. For this purpose, designing larger and statistically sound samplings will be useful.</p>
<p>In sum, the terminated sequences of CIPPS appear generally short, lacking in informative articulation, often interrupted by disfluency phenomena, and prosodically flat when performing the illocutionary pragmatic activity.</p>
<p>Thanks to the L-AcT approach, it has been possible to divide the linguistic analysis into distinct levels, allowing the highlighting of the specific features for each level responsible for the perceived monotony of schizophrenic speech.</p>
</sec>
<sec sec-type="data-availability" id="sec22">
<title>Data availability statement</title>
<p>Publicly available datasets were analyzed in this study. This data can be found here: <ext-link xlink:href="http://corpus.lablita.it/?locale=en" ext-link-type="uri">http://corpus.lablita.it/?locale=en</ext-link>.</p>
</sec>
<sec sec-type="ethics-statement" id="sec23">
<title>Ethics statement</title>
<p>Ethical approval was not required for the study involving humans in accordance with the local legislation and institutional requirements. Written informed consent to participate in this study was not required from the participants or the participants&#x2019; legal guardians/next of kin in accordance with the national legislation and the institutional requirements.</p>
</sec>
<sec sec-type="author-contributions" id="sec24">
<title>Author contributions</title>
<p>VS wrote sections 1, 2.2, and 3.1. ST wrote sections 2.1, 3.2, and 3.3. MM supervised the research. VS, ST, and MM wrote and conceived the discussion section together. All the authors contributed to the conception and design of the study and have approved the final version of the manuscript.</p>
</sec>
</body>
<back>
<sec sec-type="COI-statement" id="sec25">
<title>Conflict of interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec id="sec100" sec-type="disclaimer">
<title>Publisher&#x2019;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
<fn-group>
<fn id="fn0001">
<p><sup>1</sup>Cf. <xref ref-type="bibr" rid="ref54">Overall and Gorham (1962)</xref> for the Brief Psychiatric Rating Scale &#x201C;BPRS&#x201D; with 16 items; <xref ref-type="bibr" rid="ref90">Andreasen (1979)</xref> for the Scale for the Assessment of Thought, Language, and Communication Disorders &#x201C;TLC&#x201D;; <xref ref-type="bibr" rid="ref3">Andreasen (1982)</xref> for the Scale for the Assessment of Negative Symptoms &#x201C;SANS&#x201D;; <xref ref-type="bibr" rid="ref4">Andreasen (1986)</xref> for the Scale for the Assessment of Positive Symptoms &#x201C;SAPS.&#x201D;</p>
</fn>
<fn id="fn0002">
<p><sup>2</sup>Patients have been recruited in collaboration with Doctor Pastore at &#x201C;Scuola Sperimentale per la Formazione alla Psicoterapia e alla Ricerca nel Campo delle Scienze Umane Applicate&#x201D; of ASL NA1 of Naples and Prof Albano Leoni at CIRASS in 2005. All the participants are recorded with informed written consent. The source audio files are publicly available on CD.</p>
</fn>
<fn id="fn0003">
<p><sup>3</sup>The source audio files are available on <ext-link xlink:href="http://corpus.lablita.it" ext-link-type="uri">http://corpus.lablita.it</ext-link>.</p>
</fn>
<fn id="fn0004">
<p><sup>4</sup>The psychopathological description of each patient does not comprise standard testing, which is not available from published materials.</p>
</fn>
<fn id="fn0005">
<p><sup>5</sup>See <xref ref-type="bibr" rid="ref91">Berrios (1996)</xref> for the background of the terminology. In particular, the term Wahnstimmung, also called &#x201C;Delusional mood&#x201D; (<xref ref-type="bibr" rid="ref94">Conrad, 1958</xref>; <xref ref-type="bibr" rid="ref120">Mishara, 2010</xref>), is a prodromal feature of an impending psychotic illness in which the patient has the feeling that &#x201C;something is in the air,&#x201D; a delusion of catastrophe in the world. The prevalence of Wahnstimmung in schizophrenia spectrum disorder was recently described as between 1 and 8% (<xref ref-type="bibr" rid="ref93">Blom, 2015</xref>).</p>
</fn>
<fn id="fn0006">
<p><sup>6</sup>The speakers of SAMP are named &#x201C;cami,&#x201D; &#x201C;fale,&#x201D; &#x201C;pell,&#x201D; and &#x201C;vefa.&#x201D;</p>
</fn>
<fn id="fn0007">
<p><sup>7</sup>They find a negative association between average pause duration and &#x201C;existential reorientation,&#x201D; which refers to a fundamental rearrangement concerning patients&#x2019; general metaphysical worldview and/or hierarchy of values, projects, and interests.</p>
</fn>
<fn id="fn0008">
<p><sup>8</sup>Pauses &#x201C;after&#x201D; the turn are excluded because they are influenced by the attitude of the doctor in the communicative context.</p>
</fn>
<fn id="fn0009">
<p><sup>9</sup>Excluding the retracted tokens (see below).</p>
</fn>
<fn id="fn0010">
<p><sup>10</sup>Cf. <xref ref-type="bibr" rid="ref51">Moneglia (2005</xref>, p. 58-59) for a description of MLU variations across language contexts in Italian non-pathological speech, which is consistent with these data for what regards monologs.</p>
</fn>
<fn id="fn0011">
<p><sup>11</sup>Cutting the Sample following the duration parameter highlights the peculiarity of A&#x2019;s behavior in the communicative exchange with the doctor. For him, the number of terminated sequences in 15&#x2032; of recordings is the lowest of the Sample, so he covers only 1/10 of the CIPPS excerpts here commented. This is reflected in the massive presence of pauses (see 3.2.1).</p>
</fn>
<fn id="fn0012">
<p><sup>12</sup>Similar analyses have been conducted on schizophrenic speech in Brazilian Portuguese, see <xref ref-type="bibr" rid="ref57">Rocha et al. (2022)</xref>.</p>
</fn>
<fn id="fn0013">
<p><sup>13</sup>It should be noted that in the work mentioned above, data is measured in relation to the number of terminated sequences per communicative event (dialog/conversation/monolog); since our data here are measured speaker by speaker, the numbers do not consider the interlocutor&#x2019;s turns. Hence, we expect the percentage of stanzas to be higher than the reference value reported in <xref ref-type="bibr" rid="ref58">Saccone (2022)</xref>.</p>
</fn>
<fn id="fn0014">
<p><sup>14</sup>For pauses and retracing phenomena, statistical tests were not applied because of data aggregation strategy.</p>
</fn>
<fn id="fn0015">
<p><sup>15</sup>We carried out an agreement test between annotators, resulting in a rate of 0.85. The test agreement has been made on a sample of D. On the basis of the silent/sounding detection, we observed the manually verified boundaries comparing starting (t-min) and ending (t-max) times of silences. We adopted a fluctuation range of 150&#x2009;ms, based on the minimum chosen threshold.</p>
</fn>
<fn id="fn0016">
<p><sup>16</sup>Cf. <xref ref-type="bibr" rid="ref30">Duez (1985)</xref> and <xref ref-type="bibr" rid="ref98">Giannini (2008)</xref> for silences &#x003E;180&#x2009;ms.</p>
</fn>
<fn id="fn0017">
<p><sup>17</sup>For these first preliminary analyses, the comparison group is the CORCON.</p>
</fn>
<fn id="fn0018">
<p><sup>18</sup>See <xref ref-type="bibr" rid="ref56">Cresti et al. (forthcoming)</xref> for a more detailed description.</p>
</fn>
<fn id="fn0019">
<p><sup>19</sup>Only for one speaker of SAMP the values do reach the average of CIPPS.</p>
</fn>
<fn id="fn0020">
<p><sup>20</sup>For this analysis, the comparison group is the CORCON.</p>
</fn>
<fn id="fn0021">
<p><sup>21</sup>Parallel works are in progress at the LEEL lab of UFMG of Belo Horizonte, under the supervision of Tommaso Raso and Bruno Rocha.</p>
</fn>
<fn id="fn0022">
<p><sup>22</sup>The part that precedes (or, rarely, follows) the prominence (<italic>preparation</italic>/<italic>tail</italic>) is also isolated for future works.</p>
</fn>
<fn id="fn0023">
<p><sup>23</sup>Parameters of the script have been settled for each audio file according to the f0 range of the speaker.</p>
</fn>
<fn id="fn0024">
<p><sup>24</sup>The parameters related to the f0 are calculated both in Hertz and in semitones. The values in Hertz show the absolute number of vibrations of the vocal cords in one second, while the ones in semitones, being a logarithmic transformation, indicate how the auditory system processes the vibrations (<xref ref-type="bibr" rid="ref8">Barbosa, 2019</xref>) and therefore better reflects perceptual differences between frequencies.</p>
</fn>
<fn id="fn0025">
<p><sup>25</sup>COM-prominence, values in Hz: A: value of <italic>p</italic>&#x2009;=&#x2009;0.477; B: value of <italic>p</italic>&#x2009;=&#x2009;0.282; C: value of <italic>p</italic>&#x2009;=&#x2009;0.449; D: value of <italic>p</italic>&#x2009;=&#x2009;0.545. COM-prominence, values in st: A: value of <italic>p</italic>&#x2009;=&#x2009;0.09; B: value of <italic>p</italic>&#x2009;=&#x2009;0.404; C: value of <italic>p</italic>&#x2009;=&#x2009;0.901; D: value of <italic>p</italic>&#x2009;=&#x2009;0.412. TOP-prominence, values in Hz: A: value of <italic>p</italic>&#x2009;=&#x2009;0.33; B: value of <italic>p</italic>&#x2009;=&#x2009;0.821; C: value of <italic>p</italic>&#x2009;=&#x2009;0.391; D: value of <italic>p</italic>&#x2009;=&#x2009;0.367. TOP-prominence, values in st: A: value of <italic>p</italic>&#x2009;=&#x2009;0.306; B: value of <italic>p</italic>&#x2009;=&#x2009;0.578; C: value of <italic>p</italic>&#x2009;=&#x2009;0.375; D: value of <italic>p</italic>&#x2009;=&#x2009;0.367.</p>
</fn>
<fn id="fn0026">
<p><sup>26</sup>COM-prominence, values in Hz: A: value of <italic>p</italic>&#x2009;=&#x2009;0.479; B: value of <italic>p</italic>&#x2009;=&#x2009;0.479; C: value of <italic>p</italic>&#x2009;=&#x2009;0.477; D: value of <italic>p</italic>&#x2009;=&#x2009;0.477. COM-prominence, values in st: A: value of <italic>p</italic>&#x2009;=&#x2009;0.451; B: value of <italic>p</italic>&#x2009;=&#x2009;0.431; C: value of <italic>p</italic>&#x2009;=&#x2009;0.366; D: value of <italic>p</italic>&#x2009;=&#x2009;0.575. TOP-prominence, values in Hz: A: value of <italic>p</italic>&#x2009;=&#x2009;0.33; B: value of <italic>p</italic>&#x2009;=&#x2009;0.821; C: value of <italic>p</italic>&#x2009;=&#x2009;0.391; D: value of <italic>p</italic>&#x2009;=&#x2009;0.367. TOP-prominence, values in st: A: value of p&#x2009;=&#x2009;0.; B: value of p&#x2009;=&#x2009;0.; C: value of p&#x2009;=&#x2009;0.; D: value of <italic>p</italic>&#x2009;=&#x2009;0.</p>
</fn>
<fn id="fn0027">
<p><sup>27</sup>COM-prominence: A: value of <italic>p</italic>&#x2009;=&#x2009;0.988; B: value of <italic>p</italic>&#x2009;=&#x2009;0.118; C: value of <italic>p</italic>&#x2009;=&#x2009;0.752; D: value of <italic>p</italic>&#x2009;=&#x2009;0.741. TOP-prominence: A: value of <italic>p</italic>&#x2009;=&#x2009;0.423; B: value of p&#x2009;=&#x2009;0.4; C: value of <italic>p</italic>&#x2009;=&#x2009;0.481; D: value of p&#x2009;=&#x2009;0.367.</p>
</fn>
<fn id="fn0028">
<p><sup>28</sup>COM-prominence: A: value of <italic>p</italic>&#x2009;=&#x2009;0.787; B: value of <italic>p</italic>&#x2009;=&#x2009;0.368; C: value of <italic>p</italic>&#x2009;=&#x2009;0.145; D: value of <italic>p</italic>&#x2009;=&#x2009;0.866. TOP-prominence: A: value of <italic>p</italic>&#x2009;=&#x2009;0.306; B: value of <italic>p</italic>&#x2009;=&#x2009;0.966; C: value of <italic>p</italic>&#x2009;=&#x2009;0.795; D: value of <italic>p</italic>&#x2009;=&#x2009;0.479.</p>
</fn>
</fn-group>
<ref-list>
<title>References</title>
<ref id="ref1">
<citation citation-type="book"><person-group person-group-type="editor"><name><surname>Allwood</surname> <given-names>J.</given-names></name></person-group> (<year>2017</year>). &#x201C;<article-title>Fluency or disfluency?</article-title>&#x201D;,  in <source>Proceedings of DiSS 2017. TMH-QPSR Volume, Eds. R. Eklund and R. Rose 58</source> (<publisher-loc>Stockholm Sweden</publisher-loc>: <publisher-name>Royal Institute of Technology</publisher-name>), <fpage>18</fpage>&#x2013;<lpage>19</lpage>.</citation>
</ref>
<ref id="ref2">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Alpert</surname> <given-names>M.</given-names></name> <name><surname>Rosenberg</surname> <given-names>S. D.</given-names></name> <name><surname>Pouget</surname> <given-names>E. R.</given-names></name> <name><surname>Shaw</surname> <given-names>R. J.</given-names></name></person-group> (<year>2000</year>). <article-title>Prosody and lexical accuracy in flat affect schizophrenia</article-title>. <source>Psychiatry Res.</source> <volume>97</volume>, <fpage>107</fpage>&#x2013;<lpage>118</lpage>. doi: <pub-id pub-id-type="doi">10.1016/s0165-1781(00)00231-6</pub-id>, PMID: <pub-id pub-id-type="pmid">11166083</pub-id></citation>
</ref>
<ref id="ref90">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Andreasen</surname> <given-names>N. C.</given-names></name>
</person-group> (<year>1979</year>). <article-title>Thought, language, and communication disorders. I. Clinical assessment, definition of terms, and evaluation of their reliability</article-title>. <source>Arch Gen Psychiatry</source> <volume>36</volume>, <fpage>1315</fpage>&#x2013;<lpage>1321</lpage>.</citation>
</ref>
<ref id="ref3">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Andreasen</surname> <given-names>N. C.</given-names></name>
</person-group> (<year>1982</year>). <article-title>Negative symptoms in schizophrenia. Definition and reliability</article-title>. <source>Arch. Gen. Psychiatry</source> <volume>39</volume>, <fpage>784</fpage>&#x2013;<lpage>788</lpage>. doi: <pub-id pub-id-type="doi">10.1001/archpsyc.1982.04290070020005</pub-id></citation>
</ref>
<ref id="ref4">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Andreasen</surname> <given-names>N. C.</given-names></name>
</person-group> (<year>1986</year>). <article-title>The scale for assessment of thought, language and communication (TLC)</article-title>. <source>Schizophr. Bull.</source> <volume>12</volume>, <fpage>473</fpage>&#x2013;<lpage>482</lpage>. doi: <pub-id pub-id-type="doi">10.1093/schbul/12.3.473</pub-id></citation>
</ref>
<ref id="ref6">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bambini</surname> <given-names>V.</given-names></name> <name><surname>Frau</surname> <given-names>F.</given-names></name> <name><surname>Bischetti</surname> <given-names>L.</given-names></name> <name><surname>Cuoco</surname> <given-names>F.</given-names></name> <name><surname>Bechi</surname> <given-names>M.</given-names></name> <name><surname>Buonocore</surname> <given-names>M.</given-names></name> <etal/></person-group>. (<year>2022</year>). <article-title>Deconstructing heterogeneity in schizophrenia through language: a semi-automated linguistic analysis and data-driven clustering approach</article-title>. <source>Schizophrenia</source> <volume>8</volume>, <fpage>1</fpage>&#x2013;<lpage>12</lpage>. doi: <pub-id pub-id-type="doi">10.1038/s41537-022-00306-z</pub-id></citation>
</ref>
<ref id="ref7">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Banfi</surname> <given-names>E.</given-names></name>
</person-group> (<year>1999</year>). <source>Pause, interruzioni, silenzi. Un percorso interdisciplinare</source>. <publisher-loc>Trento</publisher-loc>, <publisher-name>Dipartimento di Scienze Filologiche e Storiche: Labirinti 36</publisher-name>.</citation>
</ref>
<ref id="ref8">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Barbosa</surname> <given-names>P.A.</given-names></name>
</person-group> (<year>2019</year>). <source>Pros&#x00F3;dia</source>. <publisher-loc>S&#x00E3;o Paulo</publisher-loc>: <publisher-name>Par&#x00E1;bola Editorial</publisher-name>.</citation>
</ref>
<ref id="ref9">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Barbosa</surname> <given-names>P. A.</given-names></name> <name><surname>Camargo</surname> <given-names>Z. A.</given-names></name> <name><surname>Madureira</surname> <given-names>S.</given-names></name></person-group> (<year>2019</year>). &#x201C;<article-title>Acoustic-based tools and scripts for the automatic analysis of speech in clinical and non-clinical settings</article-title>&#x201D; in <source>Signal and acoustic modeling for speech and communication disorders</source>. eds. <person-group person-group-type="editor"><name><surname>Patil</surname> <given-names>H. A.</given-names></name> <name><surname>Kulshreshtha</surname> <given-names>M.</given-names></name> <name><surname>Neustein</surname> <given-names>A.</given-names></name></person-group> (<publisher-loc>Berlin, Boston</publisher-loc>: <publisher-name>De Gruyter</publisher-name>), <fpage>69</fpage>&#x2013;<lpage>86</lpage>.</citation>
</ref>
<ref id="ref91">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Berrios</surname> <given-names>G. E.</given-names></name>
</person-group> (<year>1996</year>). <source>The History of Mental Symptoms. Descriptive Psychopathology since the Nineteenth Century (Cambridge University Press)</source>.</citation>
</ref>
<ref id="ref92">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Binswanger</surname> <given-names>L.</given-names></name>
</person-group> (<year>1942</year>). <source>Grundformen und Erkenntnis menschlichen Daseins (Z&#x00FC;rich: Cambridge University Press &#x0026; Assessment)</source>.</citation>
</ref>
<ref id="ref93">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Blom</surname> <given-names>J. D.</given-names></name>
</person-group> (<year>2015</year>). <article-title>The delusion of world catastrophe. Is this classic symptom still relevant today?</article-title>, <source>Tijdschr Psychiatr</source>. <volume>57</volume>, <fpage>730</fpage>&#x2013;<lpage>8</lpage>.</citation>
</ref>
<ref id="ref10">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Bleuler</surname> <given-names>E.</given-names></name>
</person-group> (<year>1950</year>). <source><italic>Dementia praecox, or the Group of Schizophrenia</italic>s</source>. <publisher-loc>New York</publisher-loc>: <publisher-name>International Universities Pre ss</publisher-name>.</citation>
</ref>
<ref id="ref11">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Boer</surname> <given-names>J. N.</given-names></name> <name><surname>van Hoogdalem</surname> <given-names>M.</given-names></name> <name><surname>Mandl</surname> <given-names>R. C. W.</given-names></name> <name><surname>Brummelman</surname> <given-names>J.</given-names></name> <name><surname>Voppel</surname> <given-names>A. E.</given-names></name> <name><surname>Begemann</surname> <given-names>M. J. H.</given-names></name> <etal/></person-group>. (<year>2020</year>). <article-title>Language in schizophrenia: relation with diagnosis, symptomatology and white matter tracts</article-title>. <source>NJP Schizophr.</source> <volume>6</volume>, <fpage>1</fpage>&#x2013;<lpage>10</lpage>. doi: <pub-id pub-id-type="doi">10.1038/s41537-020-0099-3</pub-id></citation>
</ref>
<ref id="ref12">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Boersma</surname> <given-names>P.</given-names></name> <name><surname>Weenink</surname> <given-names>D.</given-names></name></person-group> (<year>2021</year>). <source>Praat: Doing phonetics by computer [computer program]</source>. <publisher-name>University of Amsterdam</publisher-name>. <publisher-loc>Amsterdam, The Netherlands</publisher-loc>.</citation>
</ref>
<ref id="ref13">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bowie</surname> <given-names>C. R.</given-names></name> <name><surname>Harvey</surname> <given-names>P. D.</given-names></name></person-group> (<year>2008</year>). <article-title>Communication abnormalities predict functional outcomes in chronic schizophrenia: differential associations with social and adaptive functions</article-title>. <source>Schizophr. Res.</source> <volume>103</volume>, <fpage>240</fpage>&#x2013;<lpage>247</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.schres.2008.05.006</pub-id>, PMID: <pub-id pub-id-type="pmid">18571378</pub-id></citation>
</ref>
<ref id="ref14">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cannizzaro</surname> <given-names>M. S.</given-names></name> <name><surname>Cohen</surname> <given-names>H.</given-names></name> <name><surname>Rappard</surname> <given-names>F.</given-names></name> <name><surname>Snyder</surname> <given-names>P. J.</given-names></name></person-group> (<year>2005</year>). <article-title>Bradyphrenia and bradykinesia both contribute to altered speech in schizophrenia: a quantitative acoustic study</article-title>. <source>Cog. Behav. Neurol.</source> <volume>18</volume>, <fpage>206</fpage>&#x2013;<lpage>210</lpage>. doi: <pub-id pub-id-type="doi">10.1097/01.wnn.0000185278.21352.e5</pub-id>, PMID: <pub-id pub-id-type="pmid">16340393</pub-id></citation>
</ref>
<ref id="ref15">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Cavalcante</surname> <given-names>F. A.</given-names></name>
</person-group> (<year>2016</year>). <source>The topic unit in spontaneous American English: A corpus-based study</source>. <publisher-loc>Belo Horizonte</publisher-loc>: <publisher-name>Federal University of Minas Gerais</publisher-name>.</citation>
</ref>
<ref id="ref16">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cavelti</surname> <given-names>M.</given-names></name> <name><surname>Winkelbeiner</surname> <given-names>S.</given-names></name> <name><surname>Federspiel</surname> <given-names>A.</given-names></name> <name><surname>Walther</surname> <given-names>S.</given-names></name> <name><surname>Stegmayer</surname> <given-names>K.</given-names></name> <name><surname>Giezendanner</surname> <given-names>S.</given-names></name> <etal/></person-group>. (<year>2018</year>). <article-title>Formal thought disorder is related to aberrations in language-related white matter tracts in patients with schizophrenia</article-title>. <source>Psychiatry Res. Neuroimaging</source> <volume>279</volume>, <fpage>40</fpage>&#x2013;<lpage>50</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.pscychresns.2018.05.011</pub-id>, PMID: <pub-id pub-id-type="pmid">29861197</pub-id></citation>
</ref>
<ref id="ref17">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Compton</surname> <given-names>M. T.</given-names></name> <name><surname>Lunden</surname> <given-names>A.</given-names></name> <name><surname>Cleary</surname> <given-names>S. D.</given-names></name> <name><surname>Pauselli</surname> <given-names>L.</given-names></name> <name><surname>Alolayan</surname> <given-names>Y.</given-names></name> <name><surname>Halpern</surname> <given-names>B.</given-names></name> <etal/></person-group>. (<year>2018</year>). <article-title>The aprosody of schizophrenia: computationally derived acoustic phonetic underpinnings of monotone speech</article-title>. <source>Schizophr. Res.</source> <volume>197</volume>, <fpage>392</fpage>&#x2013;<lpage>399</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.schres.2018.01.007</pub-id>, PMID: <pub-id pub-id-type="pmid">29449060</pub-id></citation>
</ref>
<ref id="ref94">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Conrad</surname> <given-names>K.</given-names></name>
</person-group> (<year>1958</year>). <source>Die beginnende Schizophrenie. (Stuttgart: Thieme Verlag)</source>.</citation>
</ref>
<ref id="ref95">
<citation citation-type="book"><person-group person-group-type="editor"><name><surname>Cresti</surname> <given-names>E.</given-names></name></person-group> (<year>2020</year>). <source>&#x201C;The Pragmatic Analysis of Speech and Its Illocutionary Classification According to the Language into Act Theory&#x201D;, in Search of basic units of spoken language: a Corpus-driven approach, Eds. S. Izre&#x2019;el, H. Mello, A. Panunzi, and T. Raso (Amsterdam: John Benjamins Publishing Company)</source>. <fpage>181</fpage>&#x2013;<lpage>219</lpage>.</citation>
</ref>
<ref id="ref96">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cresti</surname> <given-names>E.</given-names></name> <name><surname>Moneglia</surname> <given-names>M.</given-names></name></person-group> (<year>2005</year>). <article-title>The role of prosody for the expression of illocutionary types. The prosodic system of questions in spoken Italian and French according to Language into Act Theory, Front</article-title>. <source>Psychology of Language</source>. <fpage>1</fpage>&#x2013;<lpage>28</lpage>.</citation>
</ref>
<ref id="ref97">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cresti</surname> <given-names>E.</given-names></name> <name><surname>Moneglia</surname> <given-names>M.</given-names></name></person-group> (<year>2023</year>). <source>C-ORAL-ROM. Integrated Reference Corpora for Spoken Romance Languages (Amsterdam: John Benjamins Publishing Company)</source>.</citation>
</ref>
<ref id="ref18">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Costa</surname> <given-names>J. C.</given-names><suffix>Jr</suffix></name>
</person-group>. (<year>2022</year>). <source>Padr&#x00E3;o informacional de stanzas de pacientes com esquizofrenia</source>. <publisher-loc>Brasil</publisher-loc>. <publisher-name>Universidade Federal</publisher-name>.</citation>
</ref>
<ref id="ref19">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Cresti</surname> <given-names>E.</given-names></name>
</person-group> (<year>2000</year>). &#x201C;<article-title>Corpus di italiano parlato</article-title>&#x201D; in <source>Studi di grammatica italiana pubblicati dall'Accademia della Crusca</source>. ed. <person-group person-group-type="editor">
<name><surname>Cresti</surname> <given-names>E.</given-names></name>
</person-group> (<publisher-loc>Firenze</publisher-loc>: <publisher-name>Accademia della Crusca</publisher-name>)</citation>
</ref>
<ref id="ref20">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Cresti</surname> <given-names>E.</given-names></name>
</person-group> (<year>2005</year>). &#x201C;<article-title>Notes on lexical strategy, structural strategies and surface clause indexes in the C-ORAL-ROM spoken corpora</article-title>&#x201D; in <source>C-ORAL-ROM. Integrated reference corpora for spoken romance languages</source>. eds. <person-group person-group-type="editor"><name><surname>Cresti</surname> <given-names>E.</given-names></name> <name><surname>Moneglia</surname> <given-names>M.</given-names></name></person-group> (<publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins</publisher-name>), <fpage>209</fpage>&#x2013;<lpage>256</lpage>.</citation>
</ref>
<ref id="ref21">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Cresti</surname> <given-names>E.</given-names></name> <name><surname>Moneglia</surname> <given-names>M.</given-names></name></person-group> (<year>2017</year>). &#x201C;<article-title>Prosodic monotony and schizophrenia</article-title>&#x201D; in <source>Lingua e patologia</source>. ed. <person-group person-group-type="editor">
<name><surname>Dovetto</surname> <given-names>F. M.</given-names></name>
</person-group> (<publisher-loc>Aracne</publisher-loc>: <publisher-name>Napoli</publisher-name>), <fpage>147</fpage>&#x2013;<lpage>197</lpage>.</citation>
</ref>
<ref id="ref22">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Cresti</surname> <given-names>E.</given-names></name> <name><surname>Moneglia</surname> <given-names>M.</given-names></name></person-group> (<year>2018</year>). &#x201C;<article-title>The illocutionary basis of information structure. Language into act theory</article-title>&#x201D; in <source>Information structure in lesser-described languages: Studies in prosody and syntax</source>. eds. <person-group person-group-type="editor"><name><surname>Adamou</surname> <given-names>E.</given-names></name> <name><surname>Haude</surname> <given-names>K.</given-names></name> <name><surname>Vanhove</surname> <given-names>M.</given-names></name></person-group> (<publisher-loc>Amsterdam</publisher-loc>: <publisher-name>Benjamins</publisher-name>), <fpage>359</fpage>&#x2013;<lpage>401</lpage>.</citation>
</ref>
<ref id="ref23">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Cresti</surname> <given-names>E.</given-names></name> <name><surname>Moneglia</surname> <given-names>M.</given-names></name> <name><surname>Gregori</surname> <given-names>L.</given-names></name> <name><surname>Saccone</surname> <given-names>V.</given-names></name> <name><surname>Trillocco</surname> <given-names>S.</given-names></name></person-group> (<year>forthcoming</year>). <source>Segmentazione in enunciati del parlato schizofrenico e correlati della patologia nel parlato spontaneo. 4 casi di studio. Tra medici e linguisti 4: parole dentro, parole fuori, 2021</source>. <publisher-name>Aracne</publisher-name>: <publisher-loc>Napoli</publisher-loc>.</citation>
</ref>
<ref id="ref24">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dickey</surname> <given-names>C. C.</given-names></name> <name><surname>Vu</surname> <given-names>M. T.</given-names></name> <name><surname>Voglmaier</surname> <given-names>M.</given-names></name> <name><surname>Niznikiewicz</surname> <given-names>M. A.</given-names></name> <name><surname>McCarley</surname> <given-names>R. W.</given-names></name> <name><surname>Panych</surname> <given-names>L. P.</given-names></name></person-group> (<year>2012</year>). <article-title>Prosodic abnormalities in schizotypal personality disorder</article-title>. <source>Schizophr. Res.</source> <volume>142</volume>, <fpage>20</fpage>&#x2013;<lpage>30</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.schres.2012.09.006</pub-id></citation>
</ref>
<ref id="ref25">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dodane</surname> <given-names>C.</given-names></name> <name><surname>Hirsch</surname> <given-names>F.</given-names></name></person-group> (<year>2018</year>). <article-title>L&#x2019;organisation spatiale et temporelle de la pause en parole et en discours</article-title>. <source>Langages</source> <volume>211</volume>, <fpage>5</fpage>&#x2013;<lpage>12</lpage>. doi: <pub-id pub-id-type="doi">10.3917/lang.211.0005</pub-id></citation>
</ref>
<ref id="ref26">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Dovetto</surname> <given-names>F. M.</given-names></name> <name><surname>Cresti</surname> <given-names>E.</given-names></name> <name><surname>Rocha</surname> <given-names>B.</given-names></name></person-group> (<year>2015</year>). &#x201C;<article-title>Schizofrenia tra prosodia e lessico. Prime analisi</article-title>&#x201D; in <source>Studi italiani di linguistica teorica e applicata</source>. eds. <person-group person-group-type="editor"><name><surname>Orletti</surname> <given-names>F.</given-names></name> <name><surname>Cardinaletti</surname> <given-names>A.</given-names></name> <name><surname>Dovetto</surname> <given-names>F. M.</given-names></name></person-group>, vol. <volume>3</volume> (<publisher-loc>Italy</publisher-loc>: <publisher-name>Pacini Editore</publisher-name>), <fpage>486</fpage>&#x2013;<lpage>507</lpage>.</citation>
</ref>
<ref id="ref27">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Dovetto</surname> <given-names>F. M.</given-names></name> <name><surname>Gemelli</surname> <given-names>M.</given-names></name></person-group> (<year>2013</year>). <source><italic>Il parlar matto. Schizofrenia tra fenomenologia e linguistica. Il corpus CIPPS</italic>, Prefazione di Federico Albano Leoni</source>. <publisher-loc>Napoli</publisher-loc>: <publisher-name>Aracne</publisher-name>.</citation>
</ref>
<ref id="ref28">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Dovetto</surname> <given-names>F. M.</given-names></name> <name><surname>Guida</surname> <given-names>A.</given-names></name> <name><surname>Pagliaro</surname> <given-names>A. C.</given-names></name> <name><surname>Guarasci</surname> <given-names>R.</given-names></name> <name><surname>Raggio</surname> <given-names>L.</given-names></name> <name><surname>Sorrentino</surname> <given-names>A.</given-names></name> <etal/></person-group>. (<year>2021</year>). &#x201C;<article-title>Corpora di italiano parlato patologico dell'et&#x00E0; adulta e senile</article-title>&#x201D; in <source>Corpora e Studi Linguistici. Atti del LIV Congresso Internazionale di Studi della Societ&#x00E0; di Linguistica Italiana (online, 8&#x2013;10 settembre)</source>. eds. <person-group person-group-type="editor"><name><surname>Cresti</surname> <given-names>E.</given-names></name> <name><surname>Moneglia</surname> <given-names>M.</given-names></name></person-group> (<publisher-loc>Milano</publisher-loc>: <publisher-name>Officinaventuno</publisher-name>), <fpage>165</fpage>&#x2013;<lpage>177</lpage>.</citation>
</ref>
<ref id="ref29">
<citation citation-type="book"><person-group person-group-type="author">
<collab id="coll1">DSM</collab>
</person-group> (<year>2013</year>). <source>DSM 5 the diagnostic and statistical manual of mental disorders</source>. <edition>5th</edition> Edn, <publisher-name>Arlington: American Psychiatric Association</publisher-name>.</citation>
</ref>
<ref id="ref30">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Duez</surname> <given-names>D.</given-names></name>
</person-group> (<year>1985</year>). <article-title>Perception of silent pauses in continuous speech</article-title>. <source>Lang Speech</source> <volume>28</volume>, <fpage>377</fpage>&#x2013;<lpage>389</lpage>. doi: <pub-id pub-id-type="doi">10.1177/002383098502800403</pub-id></citation>
</ref>
<ref id="ref31">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Eklund</surname> <given-names>R.</given-names></name>
</person-group> (<year>2004</year>). <source>Disfluency in Swedish human-human and human machine travel booking dialogues</source>. <publisher-loc>Link&#x00F6;ping</publisher-loc>: <publisher-name>Link&#x00F6;ping University</publisher-name>.</citation>
</ref>
<ref id="ref32">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Elvev&#x00E5;g</surname> <given-names>B.</given-names></name> <name><surname>Foltz</surname> <given-names>P. W.</given-names></name> <name><surname>Rosenstein</surname> <given-names>M.</given-names></name> <name><surname>DeLisi</surname> <given-names>L. E.</given-names></name></person-group> (<year>2010</year>). <article-title>An automated method to analyze language use in patients with schizophrenia and their first-degree relatives</article-title>. <source>J. Neurolinguistics</source> <volume>23</volume>, <fpage>270</fpage>&#x2013;<lpage>284</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.jneuroling.2009.05.002</pub-id>, PMID: <pub-id pub-id-type="pmid">20383310</pub-id></citation>
</ref>
<ref id="ref33">
<citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Fors</surname> <given-names>K. L.</given-names></name>
</person-group>, (<year>2011</year>). &#x201C;<article-title>Pause length variations within and between speakers over time</article-title>&#x201D;,in <conf-name>Proceedings of the 15th Workshop on the Semantics and Pragmatics of Dialogue</conf-name>, <conf-loc>Los Angeles</conf-loc>, <fpage>198</fpage>&#x2013;<lpage>199</lpage>.</citation>
</ref>
<ref id="ref34">
<citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Gagliardi</surname> <given-names>G.</given-names></name> <name><surname>Tamburini</surname> <given-names>F.</given-names></name> <name><surname>Lombardi Vallauri</surname> <given-names>E.</given-names></name></person-group> (<year>2012</year>). <article-title>La prominenza in italiano: demarcazione pi&#x00F9; che culminazione?</article-title>. <conf-name>Atti del VIII&#x00B0; Convegno dell&#x2019;Associazione Italiana Scienze della Voce</conf-name>. <publisher-name>Bulzoni</publisher-name>. <publisher-loc>Roma</publisher-loc></citation>
</ref>
<ref id="ref98">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Giannini</surname> <given-names>A.</given-names></name>
</person-group> (<year>2008</year>). &#x201C;<article-title>I silenzi del telegiornale</article-title>&#x201D;, in <source>La comunicazione parlata (I), Atti del Congresso Internazionale (Napoli 23-25 febbraio 2006), Ed. M. Pettorino, A. Giannini, M. Vallone, and R. Savy (Napoli: Liguori)</source>. <fpage>97</fpage>&#x2013;<lpage>108</lpage>.</citation>
</ref>
<ref id="ref35">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ginzburg</surname> <given-names>J.</given-names></name> <name><surname>Fern&#x00E0;ndez</surname> <given-names>R.</given-names></name> <name><surname>Schlangen</surname> <given-names>D.</given-names></name></person-group> (<year>2014</year>). <article-title>Disfluences as intra-utterance dialogue moves</article-title>. <source>Semantics Pragmat.</source> <volume>7</volume>, <fpage>1</fpage>&#x2013;<lpage>64</lpage>. doi: <pub-id pub-id-type="doi">10.3765/sp.7.9</pub-id></citation>
</ref>
<ref id="ref36">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Goldman-Eisler</surname> <given-names>F.</given-names></name>
</person-group> (<year>1961</year>). <article-title>The significance of changes in the rate of articulation</article-title>. <source>Lang. Speech</source> <volume>4</volume>, <fpage>171</fpage>&#x2013;<lpage>174</lpage>. doi: <pub-id pub-id-type="doi">10.1177/002383096100400305</pub-id></citation>
</ref>
<ref id="ref37">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Heldner</surname> <given-names>M.</given-names></name> <name><surname>Edlund</surname> <given-names>J.</given-names></name></person-group> (<year>2010</year>). <article-title>Pauses, gaps and overlaps in conversations</article-title>. <source>Phonetics</source> <volume>38</volume>, <fpage>555</fpage>&#x2013;<lpage>568</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.wocn.2010.08.002</pub-id></citation>
</ref>
<ref id="ref38">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hieke</surname> <given-names>A. E.</given-names></name>
</person-group> (<year>1981</year>). <article-title>A content-processing view of hesitation phenomena</article-title>. <source>Lang. Speech</source> <volume>24</volume>, <fpage>147</fpage>&#x2013;<lpage>160</lpage>. doi: <pub-id pub-id-type="doi">10.1177/002383098102400203</pub-id></citation>
</ref>
<ref id="ref39">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Izre'el</surname> <given-names>Sh.</given-names></name> <name><surname>Mello</surname> <given-names>H.</given-names></name> <name><surname>Panunzi</surname> <given-names>A.</given-names></name> <name><surname>Raso</surname> <given-names>T.</given-names></name></person-group> (<year>2020</year>). <source>In search of basic units of spoken language: A Corpus-driven approach</source>. <publisher-loc>Amsterdam</publisher-loc>: <publisher-name>Benjamins</publisher-name>.</citation>
</ref>
<ref id="ref99">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jaspers</surname> <given-names>K.</given-names></name>
</person-group> (<year>1963</year>). <source>General Psychopathology. Eds. J. Hoenig, and M. W. Hamilton (Chicago: The University of Chicago Press)</source>.</citation>
</ref>
<ref id="ref40">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kuperberg</surname> <given-names>G. R.</given-names></name>
</person-group> (<year>2010</year>). <article-title>Language in schizophrenia part 1: an introduction</article-title>. <source>Lang. Linguist. Compass</source> <volume>4</volume>, <fpage>576</fpage>&#x2013;<lpage>589</lpage>. doi: <pub-id pub-id-type="doi">10.1111/j.1749-818X.2010.00216.x</pub-id>, PMID: <pub-id pub-id-type="pmid">20936080</pub-id></citation>
</ref>
<ref id="ref100">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Li</surname> <given-names>G.</given-names></name> <name><surname>Hou</surname> <given-names>Q.</given-names></name> <name><surname>Zhang</surname> <given-names>C.</given-names></name> <name><surname>Jiang</surname> <given-names>Z.</given-names></name> <name><surname>Gong</surname> <given-names>S</given-names></name></person-group>. (<year>2021</year>). <article-title>Acoustic parameters for the evaluation of voice quality in patients with voice disorders</article-title>. <source>Annals of Palliative Medicine</source>, <volume>10</volume>, <fpage>1</fpage>&#x2013;<lpage>7</lpage>.</citation>
</ref>
<ref id="ref42">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Liddle</surname> <given-names>P. F.</given-names></name> <name><surname>Ngan</surname> <given-names>E. T. C.</given-names></name> <name><surname>Caissie</surname> <given-names>S. L.</given-names></name> <name><surname>Anderson</surname> <given-names>C. M.</given-names></name> <name><surname>Bates</surname> <given-names>A. T.</given-names></name> <name><surname>Quested</surname> <given-names>D. J.</given-names></name> <etal/></person-group>. (<year>2002</year>). <article-title>Thought and language index: an instrument for assessing thought and language in schizophrenia</article-title>. <source>Br. J. Psychiatry</source> <volume>181</volume>, <fpage>326</fpage>&#x2013;<lpage>330</lpage>. doi: <pub-id pub-id-type="doi">10.1192/bjp.181.4.326</pub-id></citation>
</ref>
<ref id="ref43">
<citation citation-type="other"><person-group person-group-type="author"><name><surname>Lombardi Vallauri</surname> <given-names>E.</given-names></name>
</person-group> (<year>2014</year>). &#x201C;<article-title>Le topologic hypothesis of prominence as a cue to information structure in Italian</article-title>&#x201D; in <source>Discourse segmentation in romance languages</source>. ed. <person-group person-group-type="editor">
<name><surname>Border&#x00ED;a</surname> <given-names>S. P.</given-names></name>
</person-group> <source>(Amsterdam, Philadelphia: John Benjamins)</source> <fpage>219</fpage>&#x2013;<lpage>242</lpage>.</citation>
</ref>
<ref id="ref101">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lopes</surname> <given-names>L. W.</given-names></name> <name><surname>Sim&#x00F5;es</surname> <given-names>L. W.</given-names></name> <name><surname>da Silva</surname> <given-names>J. D.</given-names></name> <name><surname>da Silva Evangelista</surname> <given-names>D.</given-names></name> <name><surname>da N&#x00F3;brega e Ugulino</surname> <given-names>A. C.</given-names></name> <name><surname>Costa Silva</surname> <given-names>P. O.</given-names></name> <etal/></person-group>. (<year>2017</year>). <article-title>Accuracy of acoustic analysis measurements in the evaluation of patients with different laryngeal diagnoses</article-title>. <source>Journal of voice</source> <volume>31</volume>, <fpage>e15</fpage>&#x2013;<lpage>e26</lpage>.</citation>
</ref>
<ref id="ref44">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lucarini</surname> <given-names>V.</given-names></name> <name><surname>Cangemi</surname> <given-names>F.</given-names></name> <name><surname>Daniel</surname> <given-names>B. D.</given-names></name> <name><surname>Lucchese</surname> <given-names>J.</given-names></name> <name><surname>Paraboschi</surname> <given-names>F.</given-names></name> <name><surname>Cattani</surname> <given-names>C.</given-names></name> <etal/></person-group>. (<year>2021</year>). <article-title>Conversational metrics, psychopathological dimensions and self-disturbances in patients with schizophrenia</article-title>. <source>Eur. Arch. Psychiatry Clin. Neurosci.</source> <volume>272</volume>, <fpage>997</fpage>&#x2013;<lpage>1005</lpage>. doi: <pub-id pub-id-type="doi">10.1007/s00406-021-01329-w</pub-id></citation>
</ref>
<ref id="ref45">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Lucarini</surname> <given-names>V.</given-names></name> <name><surname>Cangemi</surname> <given-names>F.</given-names></name> <name><surname>Tonna</surname> <given-names>M.</given-names></name> <name><surname>Grice</surname> <given-names>M.</given-names></name></person-group> (<year>2022</year>). <source>Turn-taking analysis in patients with schizophrenia, poster presented at Exling 2022</source>, <publisher-loc>Paris</publisher-loc>.</citation>
</ref>
<ref id="ref46">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lucarini</surname> <given-names>V.</given-names></name> <name><surname>Grice</surname> <given-names>M.</given-names></name> <name><surname>Cangemi</surname> <given-names>F.</given-names></name> <name><surname>Zimmermann</surname> <given-names>J. T.</given-names></name> <name><surname>Marchesi</surname> <given-names>C.</given-names></name> <name><surname>Vogeley</surname> <given-names>K.</given-names></name> <etal/></person-group>. (<year>2020</year>). <article-title>Speech prosody as a bridge between psychopathology and linguistics: the case of the schizophrenia Spectrum</article-title>. <source>Front. Psych.</source> <volume>11</volume>:<fpage>531863</fpage>. doi: <pub-id pub-id-type="doi">10.3389/fpsyt.2020.531863</pub-id>, PMID: <pub-id pub-id-type="pmid">33101074</pub-id></citation>
</ref>
<ref id="ref110">
<citation citation-type="book"><person-group person-group-type="editor"><name><surname>MacWhinney</surname> <given-names>B.</given-names></name>
</person-group> (<year>2012</year>). &#x201C;<article-title>The Logic of the Unified Model</article-title>&#x201D;, in <source>Handbook of Second Language Acquisition, Eds. S. Gass, and A. Mackey (London: Routledge)</source>. <fpage>211</fpage>&#x2013;<lpage>227</lpage>. </citation>
</ref>
<ref id="ref47">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>MacWhinney</surname> <given-names>B.</given-names></name>
</person-group> (<year>2000</year>). <source>The CHILDES project: Tools for analyzing talk</source>. <edition>3</edition>. <publisher-loc>Mahwah</publisher-loc>: <publisher-name>Lawrence Erlbaum Associates</publisher-name>.</citation>
</ref>
<ref id="ref49">
<citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Martin</surname> <given-names>P.</given-names></name>
</person-group> (<year>2004</year>). <article-title>WinPitch Corpus: a text to speech alignment tool for multimodal corpora</article-title>. <conf-name>In Proceedings of the Fourth International Conference on Language Resources and Evaluation (LREC&#x2019;04)</conf-name>, <publisher-loc>Lisboa, Portugal</publisher-loc>, <publisher-name>European Language Resources Association (ELRA)</publisher-name>, <fpage>537</fpage>&#x2013;<lpage>540</lpage>.</citation>
</ref>
<ref id="ref50">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mart&#x00ED;nez-S&#x00E1;nchez</surname> <given-names>F.</given-names></name> <name><surname>Muela-Mart&#x00ED;nez</surname> <given-names>J. A.</given-names></name> <name><surname>Cort&#x00E9;s-Soto</surname> <given-names>P.</given-names></name> <name><surname>Garc&#x00ED;a-Meil&#x00E1;n</surname> <given-names>J. J.</given-names></name> <name><surname>Vera Ferr&#x00E1;ndiz</surname> <given-names>J. A.</given-names></name> <name><surname>Egea-Caparr&#x00F3;s</surname> <given-names>A.</given-names></name> <etal/></person-group>. (<year>2015</year>). <article-title>Can the acoustic analysis of expressive prosody discriminate schizophrenia?</article-title> <source>Span. J. Psychol.</source> <volume>18</volume>, <fpage>E86</fpage>&#x2013;<lpage>E89</lpage>. doi: <pub-id pub-id-type="doi">10.1017/sjp.2015.85</pub-id>, PMID: <pub-id pub-id-type="pmid">26522128</pub-id></citation>
</ref>
<ref id="ref120">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mishara</surname> <given-names>A. L.</given-names></name>
</person-group> (<year>2010</year>). <article-title>Klaus Conrad (1905-1961): delusional mood, psychosis, and beginning schizophrenia</article-title>. <source>Schizophr Bull</source> <volume>36</volume>, <fpage>9</fpage>&#x2013;<lpage>13</lpage>.</citation>
</ref>
<ref id="ref51">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Moneglia</surname> <given-names>M.</given-names></name>
</person-group> (<year>2005</year>). &#x201C;<article-title>The C-ORAL-ROM resource</article-title>&#x201D; in <source>C-ORAL-ROM. Integrated reference corpora for spoken romance languages</source>. eds. <person-group person-group-type="editor"><name><surname>Cresti</surname> <given-names>E.</given-names></name> <name><surname>Moneglia</surname> <given-names>M.</given-names></name></person-group> (<publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins</publisher-name>), <fpage>209</fpage>&#x2013;<lpage>256</lpage>.</citation>
</ref>
<ref id="ref52">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Moneglia</surname> <given-names>M.</given-names></name> <name><surname>Cresti</surname> <given-names>E.</given-names></name></person-group> (<year>1997</year>). <source>Il progetto CHILDES: strumenti per l&#x2019;analisi del linguaggio parlato, vol II</source>. <publisher-name>Pisa</publisher-name>: <publisher-loc>Pacini Ediroe</publisher-loc>, <fpage>57</fpage>&#x2013;<lpage>90</lpage>.</citation>
</ref>
<ref id="ref53">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Moneglia</surname> <given-names>M.</given-names></name> <name><surname>Raso</surname> <given-names>T.</given-names></name></person-group> (<year>2014</year>). &#x201C;<article-title>Notes on language into act theory (L-AcT)</article-title>&#x201D; in <source>Spoken corpora and linguistic studies</source>. eds. <person-group person-group-type="editor"><name><surname>Raso</surname> <given-names>T.</given-names></name> <name><surname>Mello</surname> <given-names>H.</given-names></name></person-group> (<publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins</publisher-name>), <fpage>468</fpage>&#x2013;<lpage>495</lpage>.</citation>
</ref>
<ref id="ref54">
<citation citation-type="other"><person-group person-group-type="author"><name><surname>Overall</surname> <given-names>J. E.</given-names></name> <name><surname>Gorham</surname> <given-names>D. R.</given-names></name></person-group> (<year>1962</year>). <source>Brief psychiatric rating scale (BPRS)</source></citation>
</ref>
<ref id="ref130">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Panunzi</surname> <given-names>A.</given-names></name> <name><surname>Gregori</surname> <given-names>L.</given-names></name></person-group> (<year>2012</year>). &#x201C;<article-title>DB-IPIC. An XML database for the representation of information structure in spoken language</article-title>&#x201D;, in <source>Pragmatics and Prosody- Illocution, Modality, Attitude, Information Patterning and Speech Annotation</source>.Eeds. <person-group person-group-type="editor"><name><surname>Mello</surname> <given-names>H.</given-names></name> <name><surname>Panunzi</surname> <given-names>A.</given-names></name> <name><surname>Raso</surname> <given-names>T.</given-names></name></person-group> (<publisher-loc>Firenze</publisher-loc>: <publisher-name>University Press</publisher-name>), <fpage>133</fpage>&#x2013;<lpage>150</lpage>.</citation>
</ref>
<ref id="ref140">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Pellet-Rostaing</surname> <given-names>A.</given-names></name> <name><surname>Bertrand</surname> <given-names>R.</given-names></name> <name><surname>Boudin</surname> <given-names>A.</given-names></name> <name><surname>Rauzy</surname> <given-names>S.</given-names></name> <name><surname>Blache</surname> <given-names>P.</given-names></name></person-group> (<year>2023</year>). <article-title>A multimodal approach for modeling engagement in conversation</article-title>. <source>Front. Comput. Sci</source>. <fpage>1</fpage>&#x2013;<lpage>14</lpage>.</citation>
</ref>
<ref id="ref55">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Panunzi</surname> <given-names>A.</given-names></name> <name><surname>Scarano</surname>
</name></person-group> (<year>2009</year>). &#x201C;<article-title>Parlato spontaneo e testo: Analisi del racconto di vita</article-title>&#x201D; in <source>I parlanti e le loro storie. Competenze linguistiche, strategie comunicative, livelli di analisi. Atti Del Convegno Carini-Valderice</source>. eds. <person-group person-group-type="editor"><name><surname>Amenta</surname> <given-names>L.</given-names></name> <name><surname>Paternostro</surname> <given-names>G.</given-names></name></person-group> (<publisher-loc>Palermo</publisher-loc>: <publisher-name>Centro di studi filologici e linguistici siciliani</publisher-name>), <fpage>121</fpage>&#x2013;<lpage>132</lpage>.</citation>
</ref>
<ref id="ref56">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Rocha</surname> <given-names>B.</given-names></name> <name><surname>Raso</surname> <given-names>T.</given-names></name> <name><surname>Bicalho</surname> <given-names>M.</given-names></name></person-group> (<year>forthcoming</year>). &#x201C;<article-title>Il corpus schizofrenico del GdL M.G</article-title>&#x201D; in <source>Tra medici e linguisti 4: Parole dentro, parole fuori, 2021</source> (<publisher-loc>Napoli</publisher-loc>: <publisher-name>Aracne</publisher-name>).</citation>
</ref>
<ref id="ref150">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Raso</surname> <given-names>T.</given-names></name> <name><surname>Mello</surname> <given-names>H.</given-names></name></person-group> (<year>2013</year>). <source>Frames e fala espont&#x00E2;nea, Cadernos de Estudos Ling&#x00FC;&#x00ED;sticos (55.1), (Campinas, SP)</source>. <fpage>99</fpage>&#x2013;<lpage>108</lpage>.</citation>
</ref>
<ref id="ref57">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Rocha</surname> <given-names>B.</given-names></name> <name><surname>Raso</surname> <given-names>T.</given-names></name> <name><surname>Mello</surname> <given-names>H.</given-names></name> <name><surname>Ferrari</surname> <given-names>L.</given-names></name></person-group> (<year>2022</year>). <article-title>Information structure in the speech of individuals with schizophrenia</article-title>. <source>Methodology and first analyses from corpus-based data. CHIMERA: Romance Corpora and Linguistic Studies</source>. <volume>9</volume>, <fpage>217</fpage>&#x2013;<lpage>242</lpage>.</citation>
</ref>
<ref id="ref160">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Silva</surname> <given-names>W. J.</given-names></name> <name><surname>Lopes</surname> <given-names>L.</given-names></name> <name><surname>Cavalcanti Galdino</surname> <given-names>M. K.</given-names></name> <name><surname>Almeida</surname> <given-names>A. A.</given-names></name></person-group>, (<year>2021</year>). <source>Voice Acoustic Parameters as Predictors of Depression, Journal of Voice</source>, <fpage>1</fpage>&#x2013;<lpage>9</lpage>.</citation>
</ref>
<ref id="ref58">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Saccone</surname> <given-names>V.</given-names></name>
</person-group>, <source>Le unit&#x00E0; del parlato e dello scritto mediato dal computer a confronto: La dimensione testuale della comunicazione spontanea</source>. <publisher-name>Edizioni dell&#x2019;Orso</publisher-name>, <publisher-loc>Alessandria</publisher-loc>; (<year>2022</year>).</citation>
</ref>
<ref id="ref59">
<citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Saccone</surname> <given-names>V.</given-names></name> <name><surname>Trillocco</surname> <given-names>S.</given-names></name></person-group> (<year>2022</year>). <article-title>Segmentation of the speech flow for the evaluation of spontaneous productions in pathologies affecting the language capacity. A case study of schizophrenia</article-title>. <conf-name>In Proceedings of the RaPID-4 @LREC &#x00A9; European Language Resources Association (ELRA)</conf-name>, <publisher-name>European Language Resources Association</publisher-name>. <publisher-loc>France</publisher-loc>. <fpage>94</fpage>&#x2013;<lpage>99</lpage>.</citation>
</ref>
<ref id="ref60">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Savy</surname> <given-names>R.</given-names></name>
</person-group> (<year>2005</year>). &#x201C;<article-title>Specifiche per la trascrizione ortografica annotata dei testi</article-title>&#x201D; in <source>Italiano Parlato. Analisi di un dialogo</source>. eds. <person-group person-group-type="editor"><name><surname>Leoni</surname> <given-names>F. A.</given-names></name> <name><surname>Giordano</surname> <given-names>R.</given-names></name></person-group> (<publisher-loc>Napoli</publisher-loc>: <publisher-name>Liguori</publisher-name>), <fpage>1</fpage>&#x2013;<lpage>37</lpage>.</citation>
</ref>
<ref id="ref61">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>&#x2018;t Hart</surname> <given-names>J.</given-names></name> <name><surname>Collier</surname> <given-names>R.</given-names></name> <name><surname>Cohen</surname> <given-names>A.</given-names></name></person-group> (<year>1990</year>). <source>A perceptual study of intonation an experimental-phonetic approach to speech melody</source>. <publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>.</citation>
</ref>
<ref id="ref62">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Traunm&#x00FC;ller</surname> <given-names>H.</given-names></name> <name><surname>Eriksson</surname> <given-names>A.</given-names></name></person-group> (<year>2000</year>). <article-title>Acoustic effects of variation in vocal effort by men, women, and children</article-title>. <source>J. Acoust. Soc. Am.</source> <volume>107</volume>, <fpage>3438</fpage>&#x2013;<lpage>3451</lpage>. doi: <pub-id pub-id-type="doi">10.1121/1.429414</pub-id></citation>
</ref>
<ref id="ref63">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Trillocco</surname> <given-names>S.</given-names></name>
</person-group> (<year>forthcoming</year>). <article-title>Il dato silente in un corpus di parlato schizofrenico. DILEF</article-title>. <source>Rivista digitale del dipartimento di lettere e filosofia.</source></citation>
</ref>
</ref-list>
</back>
</article>