<?xml version="1.0" encoding="UTF-8" standalone="no"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xml:lang="EN" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Psychol.</journal-id>
<journal-title>Frontiers in Psychology</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Psychol.</abbrev-journal-title>
<issn pub-type="epub">1664-1078</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fpsyg.2022.1063158</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Psychology</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Discourse markers in TV interviews: A corpus-based comparative study of Chinese and the western media</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name><surname>Fu</surname> <given-names>Yanli</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="corresp" rid="c001"><sup>&#x002A;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/1995926/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Ho</surname> <given-names>Victor</given-names></name>
<xref ref-type="aff" rid="aff2"><sup>2</sup></xref>
</contrib>
</contrib-group>
<aff id="aff1"><sup>1</sup><institution>Faculty of Humanities, The Hong Kong Polytechnic University</institution>, <addr-line>Hong Kong</addr-line>, <country>Hong Kong SAR, China</country></aff>
<aff id="aff2"><sup>2</sup><institution>Department of English and Communication, The Hong Kong Polytechnic University</institution>, <addr-line>Hong Kong</addr-line>, <country>Hong Kong SAR, China</country></aff>
<author-notes>
<fn fn-type="edited-by"><p>Edited by: Muhammad Afzaal, Shanghai International Studies University, China</p></fn>
<fn fn-type="edited-by"><p>Reviewed by: Shrouq Almaghlouth, King Faisal University, Saudi Arabia; Yating Yu, The Hong Kong Polytechnic University, Hong Kong SAR, China; Xujun Tian, China Three Gorges University, China</p></fn>
<corresp id="c001">&#x002A;Correspondence: Yanli Fu, <email>yan-li.fu@connect.polyu.hk</email></corresp>
<fn fn-type="other" id="fn004"><p>This article was submitted to Language Sciences, a section of the journal Frontiers in Psychology</p></fn>
</author-notes>
<pub-date pub-type="epub">
<day>01</day>
<month>12</month>
<year>2022</year>
</pub-date>
<pub-date pub-type="collection">
<year>2022</year>
</pub-date>
<volume>13</volume>
<elocation-id>1063158</elocation-id>
<history>
<date date-type="received">
<day>06</day>
<month>10</month>
<year>2022</year>
</date>
<date date-type="accepted">
<day>02</day>
<month>11</month>
<year>2022</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x00A9; 2022 Fu and Ho.</copyright-statement>
<copyright-year>2022</copyright-year>
<copyright-holder>Fu and Ho</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/"><p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p></license>
</permissions>
<abstract>
<p>This article, which is part of an on-going large-scale study, quantitatively explores and compares the frequency, patterns, and positions of the three most frequently used discourse markers (DMs): <italic>so, and, but</italic> in TV interviews. The data comprise three corpora consisting of three media programs from China, the US, and the UK. Results show that there is a statistically significant difference in the frequency of the DM <italic>so</italic> and the DM <italic>and</italic>, with each DM having the highest frequency in a specific corpus. Four co-occurring strings (<italic>&#x201C;and so,&#x201D; &#x201C;and but,&#x201D; &#x201C;so but,&#x201D; &#x201C;but so&#x201D;</italic>) are identified in the three corpora with the DM co-occurrence <italic>&#x201C;and so&#x201D;</italic> having the highest frequency in the American program, supporting the claim that this combination is a typical use in American English. The general positional distribution of the three DMs is similar with the highest tendency in the initial position, which can be attributed to the program&#x2019;s interactivity. The findings will enhance our understanding of the three DMs used in media discourse and should be of practical significance to media hosts and guests in achieving better bilateral communication.</p>
</abstract>
<kwd-group>
<kwd>discourse marker (DM)</kwd>
<kwd>DM frequency</kwd>
<kwd>DM pattern</kwd>
<kwd>DM position</kwd>
<kwd>TV interviews</kwd>
<kwd>corpus-based comparative study</kwd>
</kwd-group>
<counts>
<fig-count count="6"/>
<table-count count="6"/>
<equation-count count="0"/>
<ref-count count="94"/>
<page-count count="13"/>
<word-count count="8991"/>
</counts>
</article-meta>
</front>
<body>
<sec id="S1" sec-type="intro">
<title>Introduction</title>
<p>Signals, such as discourse markers (DMs), are frequently utilized by speakers in utterances to direct the hearer through the process of interpretation (<xref ref-type="bibr" rid="B37">Foolen, 2011</xref>). Lexical expressions that are predominantly produced from conjunctions, adverbials, and prepositional phrases are referred to as DMs within the subclass of pragmatic markers (<xref ref-type="bibr" rid="B38">Fraser, 1996</xref>, <xref ref-type="bibr" rid="B40">2009</xref>). <xref ref-type="bibr" rid="B83">Schiffrin (1987)</xref> defines DMs as sequentially dependent elements that bracket units of talk. The three monosyllabic DMs&#x2014;<italic>so, and, but</italic>&#x2014;have been chosen as the central focus in this study given their high frequency and keyness, as evidenced by spoken corpora (<xref ref-type="bibr" rid="B81">R&#x00FC;hlemann, 2019</xref>) such as British National Corpus 64 (BNC64) and British National Corpus 2014 (BNC2014), in which the three selected DMs rank first on the frequency list. Moreover, according to <xref ref-type="bibr" rid="B40">Fraser&#x2019;s (2009)</xref> taxonomy, the three DMs belong to distinct groups; <italic>so</italic> is an inferential discourse marker (IDM); <italic>and</italic> is an elaborative discourse marker (EDM); <italic>but</italic> is a contrastive discourse marker (CDM).</p>
<p>Against this backdrop, the current study aims to explore media talk, specifically TV interviews. Media talk, as a particular genre, provides insights into the nature of mass communication and serves as a bridge between the media, public opinion, and public knowledge. Recently, media talk has begun to be studied as a phenomenon in its own right (<xref ref-type="bibr" rid="B51">Hutchby, 2006</xref>). Studies on media talk have been carried out focusing on the spoken discourse, such as radio talk shows (<xref ref-type="bibr" rid="B51">Hutchby, 2006</xref>; <xref ref-type="bibr" rid="B92">Tolson, 2006</xref>), television talk shows (<xref ref-type="bibr" rid="B52">Ilie, 2001</xref>; <xref ref-type="bibr" rid="B64">Lauerbach and Aijmer, 2007</xref>), quiz shows (<xref ref-type="bibr" rid="B33">Culpeper, 2005</xref>), and web page talk (<xref ref-type="bibr" rid="B62">Kopf, 2022</xref>). Created in the 20th century, TV interview, as a semi-institutionalized socio-cultural practice, has grown more popular and received consistently high ratings over the years (<xref ref-type="bibr" rid="B52">Ilie, 2001</xref>). This type of program frequently demonstrates stringent host-initiated queries, typically including face-threatening activities such as direct and unpleasant questions (<xref ref-type="bibr" rid="B45">Furk&#x00F3; and Abuczki, 2014</xref>). These features are attributable to the program&#x2019;s discursive and linguistic qualities.</p>
<p>The characteristics of the program reveal the use of a set of pragmatic language realizations, such as the use of discourse markers (DMs) (<xref ref-type="bibr" rid="B45">Furk&#x00F3; and Abuczki, 2014</xref>). DMs can process pragmatic inferences by reducing the hearer&#x2019;s processing effort (<xref ref-type="bibr" rid="B4">Aijmer and Simon-Vandenbergen, 2004</xref>; <xref ref-type="bibr" rid="B45">Furk&#x00F3; and Abuczki, 2014</xref>). DMs can function on the politeness level (<xref ref-type="bibr" rid="B73">&#x00D6;stman, 1981</xref>), phatic level (<xref ref-type="bibr" rid="B1">Aijmer, 2002</xref>), as a face mitigator (<xref ref-type="bibr" rid="B29">Crible, 2018</xref>), and for weakening the illocutionary force (<xref ref-type="bibr" rid="B66">Leech, 2014</xref>). In addition, the high frequency of DMs appearing in spoken genres makes their use a distinctive feature and a pivotal role in spoken English (<xref ref-type="bibr" rid="B25">Carter and McCarthy, 2006</xref>; <xref ref-type="bibr" rid="B35">Farahani and Ghane, 2022</xref>). TV interview is a type of oral interaction between the host and the guest that provides a good opportunity to examine DMs in spoken discourse (<xref ref-type="bibr" rid="B74">Oyeleye and Olutayo, 2012</xref>). As a result, an increasing number of studies have been dedicated to the investigation of DM in media discourse with DM function being the most explored area, such as the mapping of the DM functional spectrum in media discourse (<xref ref-type="bibr" rid="B44">Furk&#x00F3;, 2015</xref>), the examination of DM types and functions in mediatized interviews (<xref ref-type="bibr" rid="B45">Furk&#x00F3; and Abuczki, 2014</xref>), talk shows (<xref ref-type="bibr" rid="B55">Kang, 2018</xref>, <xref ref-type="bibr" rid="B56">2019</xref>), and interview videos (<xref ref-type="bibr" rid="B93">Tsoy, 2022</xref>). Despite the widespread interest in DMs, other aspects such as frequencies, patterns, and positions have received less attention in the literature, particularly in the context of TV interviews. The current study, which is a part of an on-going large-scale comparative project examining DMs in media discourse, attempts to contribute to the existing literature by exploring from a quantitative perspective.</p>
</sec>
<sec id="S2">
<title>Previous studies</title>
<p>In the past, there has been a great deal of interest in the theoretical study of DMs, with much of that interest focusing on their definition, meaning, and functions. For instance, the majority of the studies tend to advocate for different interpretations of DMs, such as discourse connectives (<xref ref-type="bibr" rid="B13">Blakemore, 2002</xref>), discourse particles (<xref ref-type="bibr" rid="B85">Schourup, 1999</xref>), and connect a variety of theoretical models, like <xref ref-type="bibr" rid="B77">Redeker&#x2019;s (1990)</xref> model and <xref ref-type="bibr" rid="B83">Schiffrin&#x2019;s (1987)</xref> five distinct planes. The last few years have witnessed an increase in the number of empirical studies examining the role of DMs in various circumstances, such as mediatized institutional political interviews (<xref ref-type="bibr" rid="B45">Furk&#x00F3; and Abuczki, 2014</xref>), scientific papers (<xref ref-type="bibr" rid="B80">Rezanova and Kogut, 2015</xref>), Asian presidents&#x2019; addresses (<xref ref-type="bibr" rid="B10">Banguis-Bantawig, 2019</xref>), therapeutic interviews (<xref ref-type="bibr" rid="B26">Cepeda and Poblete, 2006</xref>), academic spoken English (<xref ref-type="bibr" rid="B35">Farahani and Ghane, 2022</xref>), and laboratory experiments (<xref ref-type="bibr" rid="B49">Holtgraves and Bonnefon, 2017</xref>). Another emerging trend in DM research is a growing interest in comparing the use of DMs in terms of functions and frequency between English native speakers (NSs) and non-native speakers (NNSs) from a variety of L1 backgrounds (<xref ref-type="bibr" rid="B70">M&#x00FC;ller, 2005</xref>; <xref ref-type="bibr" rid="B2">Aijmer, 2011</xref>; <xref ref-type="bibr" rid="B8">Asik and Cephe, 2013</xref>; <xref ref-type="bibr" rid="B6">Al-khazraji, 2019</xref>; <xref ref-type="bibr" rid="B82">&#x015E;ahin K&#x0131;z&#x0131;l, 2021</xref>).</p>
<p>Previous research on the three selected DMs has also examined their functions in structural relations, cohesive relations, and interactional relations, particularly their role in achieving discourse coherence. These studies use either natural data, such as sociolinguistic group interviews in which several people are invited to prompt one another to speak (<xref ref-type="bibr" rid="B83">Schiffrin, 1987</xref>), a film description experiment in which discourse markers are elicited from native American English speakers&#x2019; descriptions of films that they have seen to others without having watched (<xref ref-type="bibr" rid="B77">Redeker, 1990</xref>), or constructed examples created by the scholar for the purpose of examining DMs (<xref ref-type="bibr" rid="B39">Fraser, 1999</xref>, <xref ref-type="bibr" rid="B40">2009</xref>). According to <xref ref-type="bibr" rid="B50">Huang (2019)</xref>, corpus methodology is frequently used in DM research, which is supported by <xref ref-type="bibr" rid="B83">Schiffrin&#x2019;s (1987)</xref> interview data, who also proposes that corpus is useful for analyzing discourse markers. Moreover, with the employment of corpus, <xref ref-type="bibr" rid="B84">Schirm (2012)</xref> examines Hungarian DM <italic>h&#x00E1;t</italic> in semi-guided informal conversations and job interview dialogues; <xref ref-type="bibr" rid="B45">Furk&#x00F3; and Abuczki (2014)</xref> compare six DMs (<italic>I mean, of course, oh, well, I think, you know</italic>) in BBC and CNN political interviews; <xref ref-type="bibr" rid="B50">Huang (2019)</xref> and <xref ref-type="bibr" rid="B24">Buysse (2020)</xref> compare the use of DM <italic>well</italic> in a spoken learner corpus between NSs and NNSs.</p>
<p>Studies on DM <italic>so</italic> can be classified into two types. One is the investigation of the multifunctionality of the DM <italic>so</italic> in various contexts, and the other is primarily employing the comparative approach. The exploration of the function of the DM <italic>so</italic> has been conducted in learner corpora (<xref ref-type="bibr" rid="B21">Buysse, 2007</xref>; <xref ref-type="bibr" rid="B5">Algouzi, 2021</xref>), in naturally occurring face-to-face and telephone interactions (<xref ref-type="bibr" rid="B15">Bolden, 2008</xref>, <xref ref-type="bibr" rid="B16">2009</xref>; <xref ref-type="bibr" rid="B11">Barske and Golato, 2010</xref>), in seminar talks (<xref ref-type="bibr" rid="B78">Rendle-Short, 2003</xref>), in video-mediated communication (<xref ref-type="bibr" rid="B27">Collet et al., 2021</xref>), and in English TV programs (<xref ref-type="bibr" rid="B67">Li and Xiang, 2020</xref>). On the other hand, comparative studies on the DM <italic>so</italic> have focused exclusively on comparing <italic>so</italic> with other DMs, while others have been particularly interested in comparing how the DM <italic>so</italic> is used by NSs and NNSs. <xref ref-type="bibr" rid="B14">Bolden (2006)</xref> compares the interactional role of the DM <italic>so</italic> and <italic>oh</italic> in a corpus of everyday face-to-face and telephone conversations. <xref ref-type="bibr" rid="B71">Nneka (2022)</xref> analyzes and compares the function and frequency of DM <italic>so</italic> and DM <italic>well</italic> in a small corpus of three presidential chats, showing that the two DMs can be used to effectively manage the discourse flow, with the DM <italic>so</italic> occurring more frequently than the DM <italic>well</italic>. <xref ref-type="bibr" rid="B63">Lam (2007)</xref> compares the frequency, functions, and positions of the DM <italic>so</italic> and the DM <italic>well</italic> between the Hong Kong Corpus of Spoken English (HKCSE) and British National Corpus (BNC) and discovers that the function and frequency of <italic>so</italic> vary according to text genre. <xref ref-type="bibr" rid="B5">Algouzi (2021)</xref> analyzes the frequency of the functional distribution of the DM <italic>so</italic> in the LINDSEI-AR sub-corpus. Some comparative studies on the DM <italic>so</italic> in learner corpora are also reported in <xref ref-type="bibr" rid="B70">M&#x00FC;ller&#x2019;s (2005)</xref>, <xref ref-type="bibr" rid="B22">Buysse&#x2019;s (2012</xref>, <xref ref-type="bibr" rid="B23">2014)</xref>, and <xref ref-type="bibr" rid="B69">Liu&#x2019;s (2017)</xref> studies. Studies examine the function of the DM <italic>and</italic> tend to rely on loosely extracted examples from multiple sources, such as examples retrieved from the spoken learner corpora (<xref ref-type="bibr" rid="B48">Heng, 2005</xref>), showing that the DM <italic>and</italic> can serve as an addition, a comparison, and a delaying device. Research on the DM <italic>but</italic> focuses on how it functions across a range of genres and literary forms, such as oral narratives (<xref ref-type="bibr" rid="B72">Norrick, 2001</xref>), a diachronic corpus of Northern English conversations (<xref ref-type="bibr" rid="B46">Hancil, 2018</xref>), and on comparisons between the DM <italic>but</italic> and its counterparts in other languages (<xref ref-type="bibr" rid="B7">Alsager et al., 2020</xref>; <xref ref-type="bibr" rid="B58">Khammee, 2022</xref>; <xref ref-type="bibr" rid="B89">Shirzadi et al., 2022</xref>).</p>
<p>TV interviews have been studied from the im/politeness, language, and ideological perspectives, such as how im/politeness models and strategies are used (<xref ref-type="bibr" rid="B33">Culpeper, 2005</xref>; <xref ref-type="bibr" rid="B28">Cook, 2014</xref>; <xref ref-type="bibr" rid="B36">Fedyna, 2016</xref>; <xref ref-type="bibr" rid="B75">Rabab&#x2019;Ah et al., 2019</xref>; <xref ref-type="bibr" rid="B34">Damayanti and Mubarak, 2021</xref>; <xref ref-type="bibr" rid="B90">Sitorus et al., 2022</xref>) linguistic features (<xref ref-type="bibr" rid="B52">Ilie, 2001</xref>), the representation of ideologies and power relations (<xref ref-type="bibr" rid="B12">Bilal et al., 2012</xref>; <xref ref-type="bibr" rid="B86">Sharifi et al., 2017</xref>), the host&#x2019;s role in managing discourse (<xref ref-type="bibr" rid="B74">Oyeleye and Olutayo, 2012</xref>), structural units (<xref ref-type="bibr" rid="B54">Kamil Ali, 2018</xref>), and guests&#x2019; non-serious responses (<xref ref-type="bibr" rid="B87">Sheikhan and Haugh, 2022</xref>), to name but a few. However, as previously stated, few studies have focused specifically on DMs in the context of TV interviews. Notable exceptions are <xref ref-type="bibr" rid="B28">Cook (2014)</xref>, <xref ref-type="bibr" rid="B36">Fedyna (2016)</xref>, and <xref ref-type="bibr" rid="B54">Kamil Ali (2018)</xref>, who indirectly show the DM role and functions. There are a number of additional studies, which, however, focus on the qualitative analysis of the DM functions, seldom do they explore the frequency, patterns, and position from quantitative and comparative perspectives, exceptions can be found in the investigation of Korean DM position (<xref ref-type="bibr" rid="B59">Kim et al., 2021a</xref>,<xref ref-type="bibr" rid="B60">b</xref>), the examination of co-occurrences of Persian DM <italic>vae</italic>, equivalent to English DM <italic>and</italic> (<xref ref-type="bibr" rid="B57">Kazemian and Amouzadeh, 2022</xref>), the DM combination &#x201C;<italic>and now</italic>&#x201D; (<xref ref-type="bibr" rid="B88">Shirtz, 2021</xref>), the DM sequence &#x201C;<italic>and so</italic>&#x201D; and &#x201C;<italic>so and</italic>&#x201D; (<xref ref-type="bibr" rid="B61">Koops and Lohmann, 2022</xref>), and the frequency of &#x201C;<italic>so</italic>&#x201D; (<xref ref-type="bibr" rid="B5">Algouzi, 2021</xref>) and <italic>&#x201C;just so&#x201D;</italic> (<xref ref-type="bibr" rid="B53">Kaltenb&#x00F6;ck and Ten Wolde, 2022</xref>).</p>
<p>Discourse marker co-occurrence is pervasive and relatively frequent, but little work has been done on their ability to combine (<xref ref-type="bibr" rid="B43">Fraser, 2015</xref>), and it has been somewhat overlooked until recently (<xref ref-type="bibr" rid="B32">Cuenca and Crible, 2019</xref>). For example, <xref ref-type="bibr" rid="B41">Fraser (2010</xref>, <xref ref-type="bibr" rid="B42">2013)</xref> examines the acceptability of CDM in examples drawn from COCA (Corpus of Contemporary American English) and BNC (British National Corpus), as well as discussing general functions of the DM <italic>but</italic>. In addition, <xref ref-type="bibr" rid="B43">Fraser (2015)</xref> extends the scope by investigating the combination of CDM and IDM, showing acceptable cases for such combinations. Although Fraser&#x2019;s studies on DM cluster are insightful, he did not provide satisfactory explanations for such co-occurrences, and he also failed to explore the combination of EDM with the other two types. In view of this, quite a number of underlying motivations for such combinations are explored, such as syntactic and functional criteria (<xref ref-type="bibr" rid="B32">Cuenca and Crible, 2019</xref>), multifunctionality for certain DM clusters (<xref ref-type="bibr" rid="B31">Crible and Degand, 2021</xref>; <xref ref-type="bibr" rid="B88">Shirtz, 2021</xref>; <xref ref-type="bibr" rid="B61">Koops and Lohmann, 2022</xref>). However, the co-occurrence of the three types of DMs is still overlooked. As a result, the current study intends to embark on this perspective and investigate the possibility of combing EDM, CDM, and IDM, but the investigation of reasons is beyond the scope of this study and will not be discussed further here. Apart from this, many comparative studies rely extensively on existing corpora (<xref ref-type="bibr" rid="B70">M&#x00FC;ller, 2005</xref>; <xref ref-type="bibr" rid="B63">Lam, 2007</xref>; <xref ref-type="bibr" rid="B69">Liu, 2017</xref>; <xref ref-type="bibr" rid="B46">Hancil, 2018</xref>; <xref ref-type="bibr" rid="B50">Huang, 2019</xref>; <xref ref-type="bibr" rid="B24">Buysse, 2020</xref>; <xref ref-type="bibr" rid="B5">Algouzi, 2021</xref>; <xref ref-type="bibr" rid="B61">Koops and Lohmann, 2022</xref>) and rarely build their own. This study, guided by the two research questions below, will build three corpora based on TV interviews from China, the US, and the UK to conduct a quantitative analysis of the frequency, patterns, and positions of the three DMs. The examination and comparison of the use of the three DMs in the three corpora will shed light on DMs in greater detail.</p>
<list list-type="simple">
<list-item>
<label>1:</label>
<p>What are the frequencies, patterns, and positional distributions of the three DMs in the three corpora?</p>
</list-item>
<list-item>
<label>2:</label>
<p>What are the similarities and differences (if any) of the three DMs across the three corpora in terms of the above-mentioned aspects?</p>
</list-item>
</list>
</sec>
<sec id="S3" sec-type="materials|methods">
<title>Materials and methods</title>
<sec id="S3.SS1">
<title>Corpora of the study</title>
<p>The study uses three corpora of TV interviews from China, the US, and the UK. One representative program is chosen from each of the three countries. The three programs are highly representative with a combination of global vision and unique local characteristics, in which celebrities from various fields are interviewed. Each of the three programs begins with a concise introduction with some background material, and then they move on to the conversation with challenging questions and discussions; the total running time of each episode is no more than 30 mins. <italic>The Point with Liu Xin</italic> has been selected as the Chinese TV interview. The data of this interview consist of two episodes, which are together referred to as the Chinese Corpus. <italic>Amanpour and Company</italic> is a global-news interview program on public broadcasting service (PBS). Two episodes are selected as the sample data, termed the US Corpus. <italic>HARDtalk</italic> is a BBC television and radio program that airs on BBC News Channel. Likewise, the data from this interview also comprise two episodes and is coded as the UK Corpus. The data for this study are randomly extracted from an on-going large-scale project of 120 episodes (almost 3,000 mins). The composition of the three corpora is shown in <xref ref-type="table" rid="T1">Table 1</xref> in terms of text code, number of tokens, and proportion of each episode. The data in the present study consist of 20,517 tokens. Due to the balanced sample size and interviewed guests, the three corpora are quite comparable despite their small scale. Each corpus, for example, comprises two interviewees, one of whom is a politician and the other a researcher, resulting in unbiased topics. Furthermore, it is possible to conduct media discourse analysis with small sample size. <xref ref-type="bibr" rid="B28">Cook&#x2019;s (2014)</xref> analysis of politeness and DMs in one episode of a talk show, and <xref ref-type="bibr" rid="B30">Crible and Degand&#x2019;s (2019)</xref> investigation of DM functions in 7,545 words, are two typical illustrations. Therefore, the sample size in this study is acceptable and manageable.</p>
<table-wrap position="float" id="T1">
<label>TABLE 1</label>
<caption><p>The composition of the three corpora.</p></caption>
<table cellspacing="5" cellpadding="5" frame="hsides" rules="groups">
<thead>
<tr>
<td valign="top" align="left"></td>
<td valign="top" align="center">Text code</td>
<td valign="top" align="center">Number of tokens</td>
<td valign="top" align="center">Proportion (%)</td>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Chinese Corpus</td>
<td valign="top" align="center">C_1</td>
<td valign="top" align="center">3437</td>
<td valign="top" align="center">52.56</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">C_2</td>
<td valign="top" align="center">3102</td>
<td valign="top" align="center">47.44</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">Total</td>
<td valign="top" align="center">6539</td>
<td valign="top" align="center">100.00</td>
</tr>
<tr>
<td valign="top" align="left">US Corpus</td>
<td valign="top" align="center">A_1</td>
<td valign="top" align="center">2955</td>
<td valign="top" align="center">51.12</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">A_2</td>
<td valign="top" align="center">2826</td>
<td valign="top" align="center">48.44</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">Total</td>
<td valign="top" align="center">5781</td>
<td valign="top" align="center">100.00</td>
</tr>
<tr>
<td valign="top" align="left">UK Corpus</td>
<td valign="top" align="center">B_1</td>
<td valign="top" align="center">4238</td>
<td valign="top" align="center">49.88</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">B_2</td>
<td valign="top" align="center">4259</td>
<td valign="top" align="center">50.12</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">Total</td>
<td valign="top" align="center">8497</td>
<td valign="top" align="center">100.00</td>
</tr>
</tbody>
</table></table-wrap>
</sec>
<sec id="S3.SS2">
<title>Methodology</title>
<p>The corpus method complements quantitative and qualitative approaches and often works well and effectively in conjunction with them (<xref ref-type="bibr" rid="B47">Handford, 2015</xref>). The application of the corpus technique in the field of pragmatics is both productive and potent due to the automatic search functions (<xref ref-type="bibr" rid="B17">Bonelli, 2010</xref>). The availability of corpora has been of great help to recent research on DMs, which has benefited considerably from it. A case in point is the investigation of politeness, hedges, boosters, DMs, deixis, and speech acts (<xref ref-type="bibr" rid="B4">Aijmer and Simon-Vandenbergen, 2004</xref>). The three corpora are built individually and make use of analytical tools, such as LancsBox (<xref ref-type="bibr" rid="B20">Brezina et al., 2021</xref>) and iFLYTEK&#x2019;s Hearing App. The first one is for corpus construction, while the second one is for data transcription. LancsBox is a user-friendly, new-generation software package with multiple functions for analyzing corpus and language data. The iFLYTEK&#x2019;s Hearing App enables multi-terminal, multi-language, multi-scenario, and multi-form voice-to-text transcription. As stated in the last section, six episodes are collected for corpus building, two for each corpus, as the first step. The transcription system, in line with <xref ref-type="bibr" rid="B70">M&#x00FC;ller&#x2019;s (2005)</xref>, is implemented thoroughly to ensure consistency. After the transcription work is complete, the text needs to be cleaned up because the manually entered text may have some non-standard symbols and formats (<xref ref-type="bibr" rid="B68">Liang et al., 2019</xref>). Due to the computer&#x2019;s inability to detect errors, manual checking is required for verifying each transcription, including spelling, enclitic form, punctuation, anonyms, and proper nouns (<xref ref-type="bibr" rid="B65">Leech, 2005</xref>). The following step is to add markup and annotations. Although LancsBox can perform the majority of automatic annotations, some cannot be performed accurately due to the complexity and ambiguity of language (<xref ref-type="bibr" rid="B65">Leech, 2005</xref>). For example, syntactic annotation (segmentation) and prosodic annotation (pauses) are conducted. To clearly define the category and identify the corpus, descriptive metadata is presented in a separate file, including the file name, setting, speakers, and length (<xref ref-type="bibr" rid="B79">Reppen, 2010</xref>). The names of both the host and the guest are documented so that they may be identified easily. Then the following step is to save the content in a format known as plain text (<xref ref-type="bibr" rid="B94">Wynne, 2005</xref>). In the end, each set of texts is uploaded to LancsBox on its own, resulting in a total of three corpora: the Chinese Corpus (6,539 tokens), the US Corpus (5,781 tokens), and the UK Corpus (8,497 tokens). The detailed procedures are outlined in <xref ref-type="fig" rid="F1">Figure 1</xref>. As shown in the subsequent section, a quantitative method is used to compare the three DMs across the three corpora in terms of their frequencies, patterns, and positions. The study uses normalization (<xref ref-type="bibr" rid="B19">Brezina, 2018</xref>) and the UCREL log-likelihood test (<xref ref-type="bibr" rid="B76">Rayson, 2016</xref>) for the quantitative analysis of the corpora.</p>
<fig id="F1" position="float">
<label>FIGURE 1</label>
<caption><p>Stages in the development of corpora.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-1063158-g001.tif"/>
</fig>
</sec>
<sec id="S3.SS3">
<title>Procedures</title>
<p>The corpora were built for the examination of the frequency, patterns, and position of the three selected DMs, due to the importance of the spoken corpus in DM research (<xref ref-type="bibr" rid="B3">Aijmer, 2015</xref>). First, using the KWIC tool in LancsBox (<xref ref-type="bibr" rid="B20">Brezina et al., 2021</xref>), a list of all instances of <italic>so, and, but</italic> in the form of concordance lines in the three corpora can be drawn. The corresponding generated concordance lines are saved for subsequent analysis. The frequency of detected DMs is calculated through statistical procedures such as the calculation of the absolute frequency of the three DMs, which include both DMs and non-DM uses, and then the DM frequency is calculated in line with <xref ref-type="bibr" rid="B40">Fraser&#x2019;s (2009)</xref> DM definition and criteria. During this process, cases are excluded if <italic>so</italic> is a pro-form (I guess <italic>so</italic>), degree adverb (<italic>so</italic> good), or in fixed patterns (<italic>so</italic>&#x2026;<italic>that</italic>); if <italic>and</italic> connects elements below the clause level; if <italic>but</italic> in fixed expressions (<italic>all but</italic>).</p>
<p>Second, all the concordance lines registering with the co-occurrence of the three DMs are extracted <italic>via</italic> the filter function. The co-occurrences can also be visualized using the GraphColl tool, through which collocations are identified and displayed in a collocation graph, which visualizes the collocates&#x2019; strength, frequency, and position. The combination of the three DMs can be generated by excluding the non-DM clusters. Using the filter and GraphColl function, the collocation frequency is identified, enabling the investigation of the co-occurrence of the three DMs.</p>
<p>Third, the positional distribution of the three DMs can be observed and counted using the search results for DM frequency. The operational definition of DM position is in line with <xref ref-type="bibr" rid="B40">Fraser&#x2019;s (2009)</xref> and <xref ref-type="bibr" rid="B61">Koops and Lohmann&#x2019;s (2022)</xref> criteria: a complete utterance is a linguistic unit expressing a complete proposition. For example, <xref ref-type="bibr" rid="B40">Fraser (2009)</xref> shows that a DM can appear in the initial position (<italic>But</italic>, we arrived on time.), medial position (We, <italic>however</italic>, arrived on time.), and final position (We arrived on time, <italic>however</italic>). Using the KWIC function, the positional distribution of the three DMs in particular concordances in the three corpora was extracted, and their frequency was calculated.</p>
<p>Finally, regarding the comparison of the frequency of the three DMs across the three corpora, normalization is used to allocate the frequency of the specific word to a common basis (<xref ref-type="bibr" rid="B68">Liang et al., 2019</xref>). Counts correspond to linear distributions, and normalization is an appropriate methodological choice (<xref ref-type="bibr" rid="B9">Baker, 2006</xref>). To enable comparison across corpora of different sizes, the normalized frequency is calculated and 1,000 was used as the common basis for normalization (<xref ref-type="bibr" rid="B19">Brezina, 2018</xref>).</p>
<p>The frequency ranking can be calculated by comparing the normalized frequency. The log-likelihood test was used to determine whether there is a significant difference in the frequency of the DM <italic>so, and, but</italic> across the three corpora. The log-likelihood statistic is preferred in frequency comparisons between corpora, as demonstrated in <xref ref-type="bibr" rid="B2">Aijmer&#x2019;s (2011)</xref> study. Using the UCREL log-likelihood wizard (<xref ref-type="bibr" rid="B76">Rayson, 2016</xref>), the significance test in frequency between two corpora was performed. Based on the results, the statistical significance can be calculated.</p>
</sec>
</sec>
<sec id="S4" sec-type="results">
<title>Results</title>
<sec id="S4.SS1">
<title>Frequency of <italic>so, and, but</italic> in the three corpora</title>
<p>There are 60 instances of <italic>so</italic> in the Chinese Corpus. Forty-five instances are used as DMs, while the other 15 instances are non-DM. The DM <italic>and</italic> occurs 162 times in the corpus, of which 82 instances are excluded; hence, <italic>and</italic> as DM occurs 80 times. There are 31 instances of <italic>but</italic> in total, of which four are excluded; thus, <italic>but</italic> as DM occurs 27 times. The frequency of the three DMs in the Chinese Corpus is shown in <xref ref-type="table" rid="T2">Table 2</xref>.</p>
<table-wrap position="float" id="T2">
<label>TABLE 2</label>
<caption><p>The frequency of the DMs <italic>so, and, but</italic> in the three corpora.</p></caption>
<table cellspacing="5" cellpadding="5" frame="hsides" rules="groups">
<thead>
<tr>
<td valign="top" align="left"></td>
<td valign="top" align="center">Total</td>
<td valign="top" align="center">DM</td>
<td valign="top" align="center">Non-DM</td>
<td valign="top" align="center">DM (%)</td>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Chinese Corpus</td>
<td/>
<td/>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left">so</td>
<td valign="top" align="center">60</td>
<td valign="top" align="center">45</td>
<td valign="top" align="center">15</td>
<td valign="top" align="center">75.00</td>
</tr>
<tr>
<td valign="top" align="left">and</td>
<td valign="top" align="center">162</td>
<td valign="top" align="center">80</td>
<td valign="top" align="center">82</td>
<td valign="top" align="center">49.38</td>
</tr>
<tr>
<td valign="top" align="left">but</td>
<td valign="top" align="center">31</td>
<td valign="top" align="center">27</td>
<td valign="top" align="center">4</td>
<td valign="top" align="center">87.10</td>
</tr>
<tr>
<td valign="top" align="left">US Corpus</td>
<td/>
<td/>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left">so</td>
<td valign="top" align="center">55</td>
<td valign="top" align="center">37</td>
<td valign="top" align="center">18</td>
<td valign="top" align="center">67.27</td>
</tr>
<tr>
<td valign="top" align="left">and</td>
<td valign="top" align="center">253</td>
<td valign="top" align="center">155</td>
<td valign="top" align="center">98</td>
<td valign="top" align="center">61.26</td>
</tr>
<tr>
<td valign="top" align="left">but</td>
<td valign="top" align="center">29</td>
<td valign="top" align="center">23</td>
<td valign="top" align="center">6</td>
<td valign="top" align="center">79.31</td>
</tr>
<tr>
<td valign="top" align="left">UK Corpus</td>
<td/>
<td/>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left">so</td>
<td valign="top" align="center">31</td>
<td valign="top" align="center">22</td>
<td valign="top" align="center">9</td>
<td valign="top" align="center">70.97</td>
</tr>
<tr>
<td valign="top" align="left">and</td>
<td valign="top" align="center">182</td>
<td valign="top" align="center">87</td>
<td valign="top" align="center">95</td>
<td valign="top" align="center">47.80</td>
</tr>
<tr>
<td valign="top" align="left">but</td>
<td valign="top" align="center">54</td>
<td valign="top" align="center">47</td>
<td valign="top" align="center">7</td>
<td valign="top" align="center">87.04</td>
</tr>
</tbody>
</table></table-wrap>
<p>In the US Corpus, the DM <italic>so</italic> occurs 55 times in total, of which 37 are used as a DM. The DM <italic>and</italic> occurs 253 times in total, with 155 of those instances being used as a DM. There are 29 instances of <italic>but</italic> in the corpus, except for six cases; thus, <italic>but</italic> as DM appears 23 times.</p>
<p>Altogether, <xref ref-type="table" rid="T2">Table 2</xref> shows that there are 31 instances of <italic>so</italic> in the UK Corpus, 22 are used as DMs when the remaining nine instances are excluded. There are 182 instances of <italic>and</italic>, of which 87 instances are used as a DM after excluding 95 cases. There are 54 cases of <italic>but</italic>, with 47 cases used as DM when the other seven non-DM uses are excluded.</p>
</sec>
<sec id="S4.SS2">
<title>Patterns of <italic>so, and, but</italic> in the three corpora</title>
<p>The patterns discussed in the present study are the co-occurrence/combination of DMs or DM clusters. The frequency of the combination of the three selected DMs in the Chinese Corpus can be seen in <xref ref-type="table" rid="T3">Table 3</xref>, which demonstrates that there are two co-occurrences (<italic>&#x201C;so but,&#x201D; &#x201C;and so&#x201D;</italic>) that each occurs only once. In other words, the DM cluster <italic>&#x201C;and so&#x201D;</italic> is an example of EDM-IDM, while <italic>&#x201C;so but&#x201D;</italic> is a DM cluster of IDM-CDM. The visualization of the combination of <italic>&#x201C;so but&#x201D;</italic> is shown in <xref ref-type="fig" rid="F2">Figure 2</xref>.</p>
<table-wrap position="float" id="T3">
<label>TABLE 3</label>
<caption><p>Co-occurrence of the DMs <italic>so, and, but</italic> in the three corpora.</p></caption>
<table cellspacing="5" cellpadding="5" frame="hsides" rules="groups">
<thead>
<tr>
<td valign="top" align="left"></td>
<td valign="top" align="center">DM clusters</td>
<td valign="top" align="center">Node</td>
<td valign="top" align="center">Frequency</td>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Chinese Corpus</td>
<td valign="top" align="center">so but</td>
<td valign="top" align="center">so/but</td>
<td valign="top" align="center">1</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">and so</td>
<td valign="top" align="center">so/and</td>
<td valign="top" align="center">1</td>
</tr>
<tr>
<td valign="top" align="left">US Corpus</td>
<td valign="top" align="center">and so</td>
<td valign="top" align="center">so/and</td>
<td valign="top" align="center">12</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">but so</td>
<td valign="top" align="center">so/but</td>
<td valign="top" align="center">1</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">and but</td>
<td valign="top" align="center">but/and</td>
<td valign="top" align="center">1</td>
</tr>
<tr>
<td valign="top" align="left">UK Corpus</td>
<td valign="top" align="center">and but</td>
<td valign="top" align="center">and/but</td>
<td valign="top" align="center">1</td>
</tr>
</tbody>
</table></table-wrap>
<fig id="F2" position="float">
<label>FIGURE 2</label>
<caption><p>The GraphColl of the co-occurrence of &#x201C;so but&#x201D; in the Chinese Corpus.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-1063158-g002.tif"/>
</fig>
<p>Similarly, the collocation in the US Corpus can be obtained by employing the filter and GraphColl function. By analyzing the concordance lines with the three DMs as the nodes, 12 instances of EDM-IDM co-occurrence (<italic>&#x201C;and so&#x201D;</italic>), one example of CDM-IDM co-occurrence of (<italic>&#x201C;but so&#x201D;</italic>), and one instance of the combination of EDM-CDM (<italic>&#x201C;and but&#x201D;</italic>) are identified. The number of DM clusters and the collocation of the searched node are shown in <xref ref-type="table" rid="T3">Table 3</xref> and are visualized in <xref ref-type="fig" rid="F3">Figure 3</xref>.</p>
<fig id="F3" position="float">
<label>FIGURE 3</label>
<caption><p>The GraphColl of the co-occurrence of &#x201C;and so&#x201D; in the US Corpus.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-1063158-g003.tif"/>
</fig>
<p>The co-occurrence of the DMs <italic>so, and, but</italic> in the UK Corpus can also be obtained by using a similar method. However, no instances of DM clusters were generated when the DM <italic>so</italic> was searched as the keyword. There is only one co-occurrence of the DM cluster <italic>&#x201C;and but&#x201D;</italic> in the extracted concordance line. This collocation can be visualized through the GraphColl function in <xref ref-type="fig" rid="F4">Figure 4</xref>.</p>
<fig id="F4" position="float">
<label>FIGURE 4</label>
<caption><p>The GraphColl of the co-occurrence of &#x201C;and but&#x201D; in the UK Corpus.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-1063158-g004.tif"/>
</fig>
</sec>
<sec id="S4.SS3">
<title>Positions of <italic>so, and, but</italic> in the three corpora</title>
<p><xref ref-type="table" rid="T4">Table 4</xref> displays the frequency of the positional distribution of the DMs <italic>so, and, but</italic>. We can see that the three DMs have a similar distribution in terms of the overall position. According to <xref ref-type="bibr" rid="B40">Fraser&#x2019;s (2009)</xref> study, the frequency of appearing in the initial position accounts for a larger proportion, followed by the medial position and the final position. The distribution of the three DMs in the UK Corpus serves as a good illustration.</p>
<table-wrap position="float" id="T4">
<label>TABLE 4</label>
<caption><p>The frequency of DM position in the three corpora.</p></caption>
<table cellspacing="5" cellpadding="5" frame="hsides" rules="groups">
<thead>
<tr>
<td valign="top" align="left"></td>
<td valign="top" align="center">DM</td>
<td valign="top" align="center">Initial (%)</td>
<td valign="top" align="center">Medial (%)</td>
<td valign="top" align="center">Final (%)</td>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Chinese Corpus</td>
<td valign="top" align="center">so</td>
<td valign="top" align="center">36 (80%)</td>
<td valign="top" align="center">9 (20%)</td>
<td valign="top" align="center">0</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">and</td>
<td valign="top" align="center">39 (48.75%)</td>
<td valign="top" align="center">41 (51.25%)</td>
<td valign="top" align="center">0</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">but</td>
<td valign="top" align="center">14 (51.85%)</td>
<td valign="top" align="center">13 (48.15%)</td>
<td valign="top" align="center">0</td>
</tr>
<tr>
<td valign="top" align="left">US Corpus</td>
<td valign="top" align="center">so</td>
<td valign="top" align="center">29 (78.38%)</td>
<td valign="top" align="center">8 (21.62%)</td>
<td valign="top" align="center">0</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">and</td>
<td valign="top" align="center">39 (25.16%)</td>
<td valign="top" align="center">116 (74.84%)</td>
<td valign="top" align="center">0</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">but</td>
<td valign="top" align="center">8 (34.78%)</td>
<td valign="top" align="center">15 (65.22%)</td>
<td valign="top" align="center">0</td>
</tr>
<tr>
<td valign="top" align="left">UK Corpus</td>
<td valign="top" align="center">so</td>
<td valign="top" align="center">20 (90.91%)</td>
<td valign="top" align="center">2 (9.09%)</td>
<td valign="top" align="center">0</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">and</td>
<td valign="top" align="center">46 (52.87%)</td>
<td valign="top" align="center">37 (42.53%)</td>
<td valign="top" align="center">4 (4.60%)</td>
</tr>
<tr>
<td/>
<td valign="top" align="center">but</td>
<td valign="top" align="center">26 (55.32%)</td>
<td valign="top" align="center">21 (44.68%)</td>
<td valign="top" align="center">0</td>
</tr>
</tbody>
</table></table-wrap>
</sec>
<sec id="S4.SS4">
<title>Comparison of <italic>so, and, but</italic> in the three corpora</title>
<p>A comparison of the frequency of the three DMs in each corpus reveals two common features. One is that the DM <italic>and</italic> occurs more frequently than the other two DMs (<italic>so, but</italic>). The other one is that the frequency ranking of the three DMs is the same in the Chinese Corpus and the US Corpus. For instance, <italic>and</italic> appears 80 times, followed by <italic>so</italic> (45 times) and <italic>but</italic> (27 times) in the Chinese Corpus (<xref ref-type="table" rid="T2">Table 2</xref>). Similarly, there are 155 instances of the DM <italic>and</italic>, followed by the DM <italic>so</italic> with 37 instances and the DM <italic>but</italic> with 23 instances in the US Corpus (<xref ref-type="table" rid="T2">Table 2</xref>). A minor distinction in the UK Corpus is the ranking order of the DM <italic>but</italic> and the DM <italic>so</italic>: <italic>but</italic> has a higher frequency (47 times) than <italic>so</italic> (22 times; <xref ref-type="table" rid="T2">Table 2</xref>). The frequency ranking can be calculated by comparing the normalized frequency (<xref ref-type="table" rid="T5">Table 5</xref>).</p>
<table-wrap position="float" id="T5">
<label>TABLE 5</label>
<caption><p>The normalized frequency of the three DMs in the three corpora.</p></caption>
<table cellspacing="5" cellpadding="5" frame="hsides" rules="groups">
<thead>
<tr>
<td valign="top" align="left">Corpus</td>
<td valign="top" align="center">Absolute frequency of DM <italic>so</italic></td>
<td valign="top" align="center">Number of tokens in corpus</td>
<td valign="top" align="center">Normalized frequency</td>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Chinese Corpus</td>
<td valign="top" align="center">45</td>
<td valign="top" align="center">6,539</td>
<td valign="top" align="center">6.88</td>
</tr>
<tr>
<td valign="top" align="left">US Corpus</td>
<td valign="top" align="center">37</td>
<td valign="top" align="center">5,781</td>
<td valign="top" align="center">6.40</td>
</tr>
<tr>
<td valign="top" align="left">UK Corpus</td>
<td valign="top" align="center">22</td>
<td valign="top" align="center">8,497</td>
<td valign="top" align="center">2.59</td>
</tr>
<tr>
<td valign="top" align="left" colspan="4"><hr/></td>
</tr>
<tr>
<td valign="top" align="left"><bold>Corpus</bold></td>
<td valign="top" align="center"><bold>Absolute frequency of DM <italic>and</italic></bold></td>
<td valign="top" align="center"><bold>Number of tokens in corpus</bold></td>
<td valign="top" align="center"><bold>Normalized frequency</bold></td>
</tr>
<tr>
<td valign="top" align="left" colspan="4"><hr/></td>
</tr>
<tr>
<td valign="top" align="left">Chinese Corpus</td>
<td valign="top" align="center">80</td>
<td valign="top" align="center">6,539</td>
<td valign="top" align="center">12.23</td>
</tr>
<tr>
<td valign="top" align="left">US Corpus</td>
<td valign="top" align="center">155</td>
<td valign="top" align="center">5,781</td>
<td valign="top" align="center">26.81</td>
</tr>
<tr>
<td valign="top" align="left">UK Corpus</td>
<td valign="top" align="center">87</td>
<td valign="top" align="center">8,497</td>
<td valign="top" align="center">10.24</td>
</tr>
<tr>
<td valign="top" align="left" colspan="4"><hr/></td>
</tr>
<tr>
<td valign="top" align="left"><bold>Corpus</bold></td>
<td valign="top" align="center"><bold>Absolute frequency of DM <italic>but</italic></bold></td>
<td valign="top" align="center"><bold>Number of tokens in the corpus</bold></td>
<td valign="top" align="center"><bold>Normalized frequency</bold></td>
</tr>
<tr>
<td valign="top" align="left" colspan="4"><hr/></td>
</tr>
<tr>
<td valign="top" align="left">Chinese Corpus</td>
<td valign="top" align="center">27</td>
<td valign="top" align="center">6,539</td>
<td valign="top" align="center">4.13</td>
</tr>
<tr>
<td valign="top" align="left">US Corpus</td>
<td valign="top" align="center">23</td>
<td valign="top" align="center">5,781</td>
<td valign="top" align="center">3.98</td>
</tr>
<tr>
<td valign="top" align="left">UK Corpus</td>
<td valign="top" align="center">47</td>
<td valign="top" align="center">8,497</td>
<td valign="top" align="center">5.53</td>
</tr>
</tbody>
</table></table-wrap>
<p><xref ref-type="table" rid="T5">Table 5</xref> above indicates that the DM <italic>so</italic> occurs the most frequently in the Chinese Corpus, followed by the US Corpus and the UK Corpus; the DM <italic>and</italic> has the highest frequency in the US Corpus, followed by the Chinese Corpus and the UK Corpus; and the DM <italic>but</italic> ranks the first in the UK Corpus, followed by the Chinese Corpus and the US Corpus.</p>
<p>The log-likelihood formula shows that the LL (log-likelihood) must be above 3.84 for the difference to be significant at the <italic>p</italic> &#x003C; 0.05 level. The greater the LL score, the more statistically significant the result. <xref ref-type="table" rid="T6">Table 6</xref> shows that there is a statistically significant difference between the Chinese Corpus and the UK Corpus (LL = 15.23) and between the US Corpus and the UK Corpus 3 (LL = 11.81) regarding the frequency of the DM <italic>so</italic>. Regarding the frequency of the DM <italic>and</italic>, there is a statistically significant difference between the Chinese Corpus and the US Corpus (LL = 34.49) and between the US Corpus and the UK Corpus (LL = 54.48). <xref ref-type="table" rid="T6">Table 6</xref> shows the results of the significance test of the DM <italic>and</italic> between corpora. However, no statistically significant difference was observed in the frequency of the DM <italic>but</italic> across the three corpora.</p>
<table-wrap position="float" id="T6">
<label>TABLE 6</label>
<caption><p>The frequency comparison of the DM <italic>so, and</italic> between the corpora.</p></caption>
<table cellspacing="5" cellpadding="5" frame="hsides" rules="groups">
<thead>
<tr>
<td valign="top" align="left">Corpus</td>
<td valign="top" align="center">Frequency of the DM <italic>so</italic></td>
<td valign="top" align="center">LL</td>
<td valign="top" align="center">Significance</td>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Chinese Corpus</td>
<td valign="top" align="center">45</td>
<td valign="top" align="center">15.23</td>
<td valign="top" align="center"><italic>p</italic> &#x003C; 0.05</td>
</tr>
<tr>
<td valign="top" align="left">UK Corpus</td>
<td valign="top" align="center">22</td>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left">US Corpus</td>
<td valign="top" align="center">37</td>
<td valign="top" align="center">11.81</td>
<td valign="top" align="center"><italic>p</italic> &#x003C; 0.05</td>
</tr>
<tr>
<td valign="top" align="left">UK Corpus</td>
<td valign="top" align="center">22</td>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left" colspan="4"><hr/></td>
</tr>
<tr>
<td valign="top" align="left"><bold>Corpus</bold></td>
<td valign="top" align="center"><bold>Frequency of the DM <italic>and</italic></bold></td>
<td valign="top" align="center"><bold>LL</bold></td>
<td valign="top" align="center"><bold>Significance</bold></td>
</tr>
<tr>
<td valign="top" align="left" colspan="4"><hr/></td>
</tr>
<tr>
<td valign="top" align="left">Chinese Corpus</td>
<td valign="top" align="center">80</td>
<td valign="top" align="center">34.49</td>
<td valign="top" align="center"><italic>p</italic> &#x003C; 0.05</td>
</tr>
<tr>
<td valign="top" align="left">US Corpus</td>
<td valign="top" align="center">155</td>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left">US Corpus</td>
<td valign="top" align="center">155</td>
<td valign="top" align="center">54.48</td>
<td valign="top" align="center"><italic>p</italic> &#x003C; 0.05</td>
</tr>
<tr>
<td valign="top" align="left">UK Corpus</td>
<td valign="top" align="center">87</td>
<td/>
<td/>
</tr>
</tbody>
</table></table-wrap>
<p>The above pairwise comparison displays that there is a statistically significant difference between corpora regarding the use of DM <italic>so</italic>, DM <italic>and</italic>. However, no statistically significant difference is found between corpora in terms of the frequency of DM <italic>but</italic>.</p>
<p>As shown in <xref ref-type="table" rid="T3">Table 3</xref>, the most frequently occurring combination among the emerging patterns is <italic>&#x201C;and so,&#x201D;</italic> which occurs once in the Chinese Corpus and 12 times in the US Corpus. However, it never appears in the UK Corpus. As for the other identified sequencing patterns, <italic>&#x201C;so but&#x201D;</italic> only appears once in the Chinese Corpus, <italic>&#x201C;but so&#x201D;</italic> only occurs once in the US Corpus, <italic>&#x201C;and but&#x201D;</italic> occurs only once in both the US Corpus and the UK Corpus. The comparison of the frequency of the combinations of the three DMs can be visualized in <xref ref-type="fig" rid="F5">Figure 5</xref>.</p>
<fig id="F5" position="float">
<label>FIGURE 5</label>
<caption><p>The frequency of the co-occurrences of the three DMs in the three corpora. (1-the Chinese Corpus, 2-the US Corpus, 3-the UK Corpus).</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-1063158-g005.tif"/>
</fig>
<p><xref ref-type="fig" rid="F6">Figure 6</xref> shows that the positional distribution of the DM <italic>so</italic> is quite similar across the three corpora since it has a higher frequency of appearing in the initial position followed by the medial position in all three corpora. As for the DM <italic>and</italic>, the positional distribution in the Chinese Corpus and the US Corpus is quite similar because the DM <italic>and</italic> has a higher frequency in the medial position followed by the initial position, notably in the US Corpus. However, the number of occurrences of the initial position of the DM <italic>and</italic> in the UK Corpus is, however, higher than that in the medial position. Another intriguing feature in the UK Corpus is that only the DM <italic>and</italic> appears four times in the final position. In the US Corpus, the DM <italic>but</italic> has a higher frequency in the medial position followed by the initial position, as seen in <xref ref-type="fig" rid="F6">Figure 6</xref>. This is in contrast to the positional distribution in the other two corpora, where the DM <italic>but</italic> occurs more in the initial position followed by the medial position.</p>
<fig id="F6" position="float">
<label>FIGURE 6</label>
<caption><p>The positional distribution of the three DMs in the three corpora. (1-the Chinese Corpus, 2-the US Corpus, 3-the UK Corpus).</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-1063158-g006.tif"/>
</fig>
</sec>
</sec>
<sec id="S5" sec-type="discussion">
<title>Discussion</title>
<p>The present study compares the frequency of the three selected DMs, their co-occurrences, and their positions in the three corpora by adopting a corpus-based approach and a quantitative method. The results, to some extent, offer some evidence for the comparability of the three corpora.</p>
<p>First, three DMs are the top three in each of the three corpora in terms of frequency, confirming previous research that they are frequently used in the spoken genre (<xref ref-type="bibr" rid="B25">Carter and McCarthy, 2006</xref>). The resemblance can be found between the present study and previous studies (<xref ref-type="bibr" rid="B77">Redeker, 1990</xref>; <xref ref-type="bibr" rid="B8">Asik and Cephe, 2013</xref>; <xref ref-type="bibr" rid="B29">Crible, 2018</xref>) and BNC64 and BNC2014 is that the DM <italic>and</italic> ranks first. However, there are slight differences in the ranking order. For instance, the ranking order of the Chinese Corpus (<xref ref-type="table" rid="T2">Table 2</xref>) and the US Corpus (<xref ref-type="table" rid="T2">Table 2</xref>) corresponds to Asik and Cephe&#x2019;s and Redeker&#x2019;s studies, as well as spoken BNC64&#x2014;that is, the DM <italic>and</italic> ranks first, followed by the DM <italic>so</italic> and the DM <italic>but</italic>. Whereas the ranking order of the UK Corpus (<xref ref-type="table" rid="T2">Table 2</xref>) is consistent with <xref ref-type="bibr" rid="B29">Crible&#x2019;s (2018)</xref> study and BNC64&#x2014;which shows that the DM <italic>and</italic> ranks first, followed by the DM <italic>but</italic> and the DM <italic>so</italic>. When the frequency of the three DMs is compared across the three corpora, the DM <italic>so</italic> appears the most frequently in the Chinese Corpus (<xref ref-type="table" rid="T5">Table 5</xref>), the DM <italic>and</italic> occurs the most frequently in the US Corpus (<xref ref-type="table" rid="T5">Table 5</xref>), and the DM <italic>but</italic> has the highest frequency in the UK Corpus (<xref ref-type="table" rid="T5">Table 5</xref>). This indicates that inferential expressions associated with the DM <italic>so</italic> are most frequently used in the Chinese interview, elaborative expressions embedded with the DM <italic>and</italic> occur relatively more frequently in the American interview, and contrastive expressions with the DM <italic>but</italic> are used the most in the British interview. In line with previous studies (<xref ref-type="bibr" rid="B63">Lam, 2007</xref>; <xref ref-type="bibr" rid="B69">Liu, 2017</xref>), some DMs&#x2019; observed differences are statistically significant. For example, there is a statistically significant difference between the Chinese Corpus and the UK Corpus (<xref ref-type="table" rid="T6">Table 6</xref>) and between the US Corpus and the UK Corpus (<xref ref-type="table" rid="T6">Table 6</xref>) regarding the frequency of the DM <italic>so</italic>. The difference in the rate of the DM <italic>and</italic> between the Chinese Corpus and the US Corpus (<xref ref-type="table" rid="T6">Table 6</xref>) and between the US Corpus and the UK Corpus (<xref ref-type="table" rid="T6">Table 6</xref>) also achieves a high statistical significance. However, no statistically significant difference is observed in the frequency of DM <italic>but</italic>. Gender is one of the variables accounting for the frequency of DM use. For example, it is typically women&#x2019;s language (<xref ref-type="bibr" rid="B73">&#x00D6;stman, 1981</xref>). However, the current study cannot draw such firm conclusions because the presented DM frequency is used by both men and women. The Chinese Corpus, for example, has one female host and two male guests; the US Corpus includes one female host, one male guest, and one female guest; and the UK Corpus consists of one male host and two male guests. As a result, further research is required to interpret this phenomenon.</p>
<p>Previous studies have reported that the DM cluster is a frequent phenomenon that is not random and has some discourse-functional motivation (<xref ref-type="bibr" rid="B29">Crible, 2018</xref>). <xref ref-type="bibr" rid="B31">Crible and Degand (2021)</xref> disentangle some linguistic features that constrain DM clusters and propose a reasonable rule that governs this integration. The three DMs in the present study can be combined to form six co-occurring strings: <italic>&#x201C;and so,&#x201D; &#x201C;and but,&#x201D; &#x201C;so but,&#x201D; &#x201C;so and,&#x201D; &#x201C;but and,&#x201D; &#x201C;but so.&#x201D;</italic> The results (<xref ref-type="fig" rid="F5">Figure 5</xref>) show that four out of the six patterns occur among the three corpora. The combination <italic>&#x201C;and so&#x201D;</italic> occurs the most frequently in the US Corpus. <xref ref-type="bibr" rid="B29">Crible&#x2019;s (2018)</xref> study on the frequency of English DM clusters echoes the same finding, revealing the possibility of the combination of EDM and IDM. In contrast to the current study, <xref ref-type="bibr" rid="B63">Lam&#x2019;s (2007)</xref> research does not appear to demonstrate a clear preference for this combination in terms of DM collocations, which is contradictory. Another intriguing aspect is that <xref ref-type="bibr" rid="B43">Fraser&#x2019;s (2015)</xref> study does not find evidence of the possibility of IDM and CDM working together (referred to here as <italic>&#x201C;so but&#x201D;</italic>), which is demonstrated in the present study. The results of the patterns reveal a solid tendency for the combination of different types of DMs. <xref ref-type="bibr" rid="B61">Koops and Lohmann&#x2019;s (2022)</xref> general order principle can be used to explain this integration&#x2014;different DM patterns achieve different effects. In other words, the earlier DM constrains the interpretation of the later one. The DM co-occurrence <italic>&#x201C;and so&#x201D;</italic> indicates that the upcoming utterance is a result or conclusion, while <italic>&#x201C;so and</italic>&#x201D; marks a topic shift. The one that should be placed first or second is determined by DM functions. The DM sequencing <italic>&#x201C;and so&#x201D;</italic> indicates that it is frequently used to start a new topic or turn, particularly in the US corpus. In addition, <italic>&#x201C;and so&#x201D;</italic> is regular use in American English (<xref ref-type="bibr" rid="B61">Koops and Lohmann, 2022</xref>), which explains the higher frequency of its appearance in American interview. These may be preliminary explanations for speakers&#x2019; preference in choosing the DM pattern. Overall, given the frequency of DM clusters in the present study, it does not show a high consistency with previous studies that DM co-occurrence is a frequent phenomenon, or at least it is not as frequent as claimed in previous studies, which may be attributed to the small data set.</p>
<p>Finally, the general positional distribution of the three DMs across the three corpora is quite similar&#x2014;with the initial position having the highest proportion, followed by the medial position and the final position (<xref ref-type="table" rid="T4">Table 4</xref>). This is in line with <xref ref-type="bibr" rid="B29">Crible&#x2019;s (2018)</xref> finding that utterance initial is the most typical use of DM followed by utterance medial and final. Moreover, a higher rate of their occurrence in discourse initial position also echoes previous studies (<xref ref-type="bibr" rid="B40">Fraser, 2009</xref>; <xref ref-type="bibr" rid="B7">Alsager et al., 2020</xref>). Despite the general consistency in the distribution pattern, there is a certain degree of disparity in the proportions of the DM <italic>and</italic> in the Chinese Corpus (48.75 vs. 51.25%) and the US Corpus (25.16 vs. 74.84%) and the DM <italic>but</italic> in the US Corpus (34.78 vs. 65.25%), in which the medial proportion is greater than that of the initial position. As shown in <xref ref-type="fig" rid="F6">Figure 6</xref>, the positional distribution of the DM <italic>so</italic> is less flexible than that of the other two DMs, whose initial position is always the first. This finding, however, contradicts <xref ref-type="bibr" rid="B63">Lam&#x2019;s (2007)</xref> study, which shows that the DM <italic>so</italic> in the utterance medial position constitutes a greater proportion. Moreover, they rarely occur in the utterance final position. Only the DM <italic>and</italic> appears four times (4.6%) in the final position in the UK Corpus (<xref ref-type="fig" rid="F6">Figure 6</xref>). Thus, we may draw a preliminary conclusion that the DM positional distribution does not exhibit a high degree of positional freedom as reported in previous studies (<xref ref-type="bibr" rid="B91">Tanghe, 2016</xref>; <xref ref-type="bibr" rid="B18">Border&#x00ED;a and Fischer, 2021</xref>). The DM positional distribution can be attributed to different variables. Register plays a crucial role, specifically in the degree of interactivity: the more interactive the register, the more frequently it occurs in the initial position (<xref ref-type="bibr" rid="B29">Crible, 2018</xref>). Three interviews are all interactive, interpreting their higher proportion of the occurrence in the initial position. In addition, their rare appearance in the final position is also due to the limited number of tag questions in the three corpora (<xref ref-type="bibr" rid="B29">Crible, 2018</xref>). More thorough larger-scale research is required to better understand this phenomenon.</p>
</sec>
<sec id="S6" sec-type="conclusion">
<title>Conclusion</title>
<p>The present study examined and compared the frequency, patterns, and positional distribution of the three DMs in the three corpora. The study concluded that there are statistically significant differences and marginal variations within and between corpora in different aspects. The study differed most significantly from previous ones in terms of the data type and objectives. While previous studies concentrated more on qualitative analysis of DM functions in academic settings or only one of the above-mentioned aspects, this study attempted to investigate all three aspects in one single goal in media discourse from a quantitative perspective. The study revealed that DM <italic>so</italic> occurs most frequently in the Chinese Corpus; the US Corpus has the highest frequency of the DM <italic>and</italic>; and the UK Corpus has the highest frequency of DM <italic>but</italic>, which provides a basis for a more productive analysis of the types of DMs used in media talk. Do all IDMs, for example, appear frequently in Chinese media discourse? Or are there more EDMs in American interviews? Or are CDMs more likely to occur in British media discourse? The article found that the DM cluster <italic>&#x201C;and so&#x201D;</italic> is most frequently used in American interviews, confirming that this combination is typical use in American English. The higher frequency of DMs appearing in the initial position also confirmed that the more interactive the genre, the more likely DMs appear in the initial position.</p>
<p>Hence, the study contributes to the existing literature by filling the gap and adding more insights into the field of discourse markers. Apart from academe, the study also sheds light on the role of DMs in interviews, which may inspire hosts and guests to consider how to achieve the best bilateral communication by using certain DMs. Despite the foregoing insights, this study falls short of expectations and does have some limitations. The findings in the present study cannot be generalized due to the small data set. Although this study provides some preliminary explanations for the differences, it fails to account for more specific variables due to the small sample size and the scope of this study. This opens new avenues for the future research in encompassing both the large corpus and another genre in order to address this problem. It is intended that the results presented in this study would be of benefit not only to individuals who conduct interviews for the media and those who are interviewed for the media, but also to those who aspire to be successful doing DM research in media settings.</p>
</sec>
<sec id="S7" sec-type="data-availability">
<title>Data availability statement</title>
<p>The raw data supporting the conclusions of this article will be made available by the authors, without undue reservation.</p>
</sec>
<sec id="S8">
<title>Author contributions</title>
<p>Both authors listed have made a substantial, direct, and intellectual contribution to the work, and approved it for publication.</p>
</sec>
</body>
<back>
<sec id="S9" sec-type="COI-statement">
<title>Conflict of interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest. The reviewer YY and handling editor declared their shared affiliation.</p>
</sec>
<sec id="S10" sec-type="disclaimer">
<title>Publisher&#x2019;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
<ref-list>
<title>References</title>
<ref id="B1"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Aijmer</surname> <given-names>K.</given-names></name></person-group> (<year>2002</year>). <source><italic>English discourse particles: Evidence from a corpus.</italic></source> <publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins Pub. Co</publisher-name>. <pub-id pub-id-type="doi">10.1075/scl.10</pub-id></citation></ref>
<ref id="B2"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Aijmer</surname> <given-names>K.</given-names></name></person-group> (<year>2011</year>). <article-title>Well I&#x2019;m not sure I think. The use of well by non-native speakers.</article-title> <source><italic>Int. J. Corpus Linguist.</italic></source> <volume>16</volume> <fpage>231</fpage>&#x2013;<lpage>254</lpage>. <pub-id pub-id-type="doi">10.1075/ijcl.16.2.04aij</pub-id> <pub-id pub-id-type="pmid">33486653</pub-id></citation></ref>
<ref id="B3"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Aijmer</surname> <given-names>K.</given-names></name></person-group> (<year>2015</year>). &#x201C;<article-title>Analyzing discourse markers in spoken corpora: Actually as a case study</article-title>,&#x201D; in <source><italic>Corpora and discourse studies</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>Paul</surname> <given-names>B.</given-names></name> <name><surname>McEnery</surname> <given-names>T.</given-names></name></person-group> (<publisher-loc>London</publisher-loc>: <publisher-name>Palgrave Macmillan</publisher-name>), <fpage>88</fpage>&#x2013;<lpage>109</lpage>. <pub-id pub-id-type="doi">10.1057/9781137431738_5</pub-id></citation></ref>
<ref id="B4"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Aijmer</surname> <given-names>K.</given-names></name> <name><surname>Simon-Vandenbergen</surname> <given-names>A. M.</given-names></name></person-group> (<year>2004</year>). <article-title>A model and a methodology for the study of pragmatic markers: The semantic field of expectation.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>36</volume> <fpage>1781</fpage>&#x2013;<lpage>1805</lpage>. <pub-id pub-id-type="doi">10.1016/j.pragma.2004.05.005</pub-id></citation></ref>
<ref id="B5"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Algouzi</surname> <given-names>S.</given-names></name></person-group> (<year>2021</year>). <article-title>Functions of the discourse marker So in the LINDSEI-AR corpus.</article-title> <source><italic>Cogent Arts Hum.</italic></source> <volume>8</volume> <fpage>1</fpage>&#x2013;<lpage>13</lpage>. <pub-id pub-id-type="doi">10.1080/23311983.2021.1872166</pub-id></citation></ref>
<ref id="B6"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Al-khazraji</surname> <given-names>A.</given-names></name></person-group> (<year>2019</year>). <article-title>Analysis of Discourse Markers in Essays Writing in ESL Classroom.</article-title> <source><italic>Int. J. Instruct.</italic></source> <volume>12</volume> <fpage>559</fpage>&#x2013;<lpage>572</lpage>. <pub-id pub-id-type="doi">10.29333/iji.2019.12235a</pub-id></citation></ref>
<ref id="B7"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Alsager</surname> <given-names>H. N.</given-names></name> <name><surname>Afzal</surname> <given-names>N.</given-names></name> <name><surname>Aldawood</surname> <given-names>A.</given-names></name></person-group> (<year>2020</year>). <article-title>Discourse Markers in Arabic and English Newspaper Articles: The Case of the Arabic Lakin and its English equivalent But.</article-title> <source><italic>Arab World Engl. J.</italic></source> <volume>11</volume> <fpage>154</fpage>&#x2013;<lpage>165</lpage>. <pub-id pub-id-type="doi">10.24093/awej/vol11no1.13</pub-id></citation></ref>
<ref id="B8"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Asik</surname> <given-names>A.</given-names></name> <name><surname>Cephe</surname> <given-names>P. T.</given-names></name></person-group> (<year>2013</year>). <article-title>Discourse Markers and Spoken English: Nonnative Use in the Turkish EFL Setting.</article-title> <source><italic>Engl. Lang. Teach.</italic></source> <volume>6</volume>:<issue>144</issue>. <pub-id pub-id-type="doi">10.5539/elt.v6n12p144</pub-id></citation></ref>
<ref id="B9"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Baker</surname> <given-names>P.</given-names></name></person-group> (<year>2006</year>). <source><italic>Using corpora in discourse analysis.</italic></source> <publisher-loc>London</publisher-loc>: <publisher-name>Continuum</publisher-name>. <pub-id pub-id-type="doi">10.5040/9781350933996</pub-id></citation></ref>
<ref id="B10"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Banguis-Bantawig</surname> <given-names>R.</given-names></name></person-group> (<year>2019</year>). <article-title>The role of discourse markers in the speeches of selected Asian Presidents.</article-title> <source><italic>Heliyon</italic></source> <volume>5</volume>:<issue>e01298</issue>. <pub-id pub-id-type="doi">10.1016/j.heliyon.2019.e01298</pub-id> <pub-id pub-id-type="pmid">30976666</pub-id></citation></ref>
<ref id="B11"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Barske</surname> <given-names>T.</given-names></name> <name><surname>Golato</surname> <given-names>A.</given-names></name></person-group> (<year>2010</year>). <article-title>German so: Managing sequence and action.</article-title> <source><italic>Text Talk</italic></source> <volume>30</volume> <fpage>245</fpage>&#x2013;<lpage>266</lpage>. <pub-id pub-id-type="doi">10.1515/text.2010.013</pub-id></citation></ref>
<ref id="B12"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bilal</surname> <given-names>H. A.</given-names></name> <name><surname>Ahsan</surname> <given-names>H. M.</given-names></name> <name><surname>Gohar</surname> <given-names>S.</given-names></name> <name><surname>Younis</surname> <given-names>S.</given-names></name> <name><surname>Awan</surname> <given-names>S. J.</given-names></name></person-group> (<year>2012</year>). <article-title>Critical Discourse Analysis of Political TV Talk Shows of Pakistani Media.</article-title> <source><italic>Int. J. Linguist.</italic></source> <volume>4</volume> <fpage>203</fpage>&#x2013;<lpage>219</lpage>. <pub-id pub-id-type="doi">10.5296/ijl.v4i1.1425</pub-id></citation></ref>
<ref id="B13"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Blakemore</surname> <given-names>D.</given-names></name></person-group> (<year>2002</year>). <source><italic>Relevance and linguistic meaning: The semantics and pragmatics of discourse markers.</italic></source> <publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>. <pub-id pub-id-type="doi">10.1017/CBO9780511486456</pub-id></citation></ref>
<ref id="B14"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bolden</surname> <given-names>G.</given-names></name></person-group> (<year>2006</year>). <article-title>Little Words That Matter: Discourse Markers So and Oh and the. Doing of Other-Attentiveness in Social Interaction.</article-title> <source><italic>J. Commun.</italic></source> <volume>56</volume> <fpage>661</fpage>&#x2013;<lpage>688</lpage>. <pub-id pub-id-type="doi">10.1111/j.1460-2466.2006.00314.x</pub-id></citation></ref>
<ref id="B15"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bolden</surname> <given-names>G.</given-names></name></person-group> (<year>2008</year>). <article-title>So What&#x2019;s Up? Using the Discourse Marker So to Launch. Conversational Business.</article-title> <source><italic>Res. Lang. Soc. Interact.</italic></source> <volume>41</volume> <fpage>302</fpage>&#x2013;<lpage>337</lpage>. <pub-id pub-id-type="doi">10.1080/08351810802237909</pub-id></citation></ref>
<ref id="B16"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bolden</surname> <given-names>G.</given-names></name></person-group> (<year>2009</year>). <article-title>Implementing incipient actions: The discourse marker &#x201C;so&#x201D; in English conversation.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>41</volume> <fpage>974</fpage>&#x2013;<lpage>998</lpage>. <pub-id pub-id-type="doi">10.1016/j.pragma.2008.10.004</pub-id></citation></ref>
<ref id="B17"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bonelli</surname> <given-names>E.</given-names></name></person-group> (<year>2010</year>). &#x201C;<article-title>Theoretical overview of the evolution of corpus linguistics</article-title>,&#x201D; in <source><italic>The Routledge Handbook of Corpus Linguistics</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>O&#x2019;Keeffe</surname> <given-names>A.</given-names></name> <name><surname>McCarthy</surname> <given-names>M.</given-names></name></person-group> (<publisher-loc>London</publisher-loc>: <publisher-name>Routledge</publisher-name>), <fpage>14</fpage>&#x2013;<lpage>28</lpage>. <pub-id pub-id-type="doi">10.4324/9780203856949-2</pub-id> <pub-id pub-id-type="pmid">36153787</pub-id></citation></ref>
<ref id="B18"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Border&#x00ED;a</surname> <given-names>S. P.</given-names></name> <name><surname>Fischer</surname> <given-names>K.</given-names></name></person-group> (<year>2021</year>). <article-title>Using discourse segmentation to account for the polyfunctionality of discourse markers: The case of well.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>173</volume> <fpage>101</fpage>&#x2013;<lpage>118</lpage>. <pub-id pub-id-type="doi">10.1016/j.pragma.2020.11.021</pub-id></citation></ref>
<ref id="B19"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Brezina</surname> <given-names>V.</given-names></name></person-group> (<year>2018</year>). <source><italic>Statistics in Corpus Linguistics.</italic></source> <publisher-loc>Cambridge, MA</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>, <pub-id pub-id-type="doi">10.1017/9781316410899</pub-id></citation></ref>
<ref id="B20"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Brezina</surname> <given-names>V.</given-names></name> <name><surname>Weill-Tessier</surname> <given-names>P.</given-names></name> <name><surname>McEnery</surname> <given-names>A.</given-names></name></person-group> (<year>2021</year>). <source><italic>#LancsBox v. 6.x. [software. package].</italic></source> Available online at: <ext-link ext-link-type="uri" xlink:href="http://corpora.lancs.ac.uk/lancsbox">http://corpora.lancs.ac.uk/lancsbox</ext-link> <comment>(accessed on 26 Feb, 2021)</comment>.</citation></ref>
<ref id="B21"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Buysse</surname> <given-names>L.</given-names></name></person-group> (<year>2007</year>). <article-title>Discourse marker so in the English of Flemish university students.</article-title> <source><italic>Belgian J. Engl. Lang. Lit.</italic></source> <volume>5</volume> <fpage>79</fpage>&#x2013;<lpage>95</lpage>.</citation></ref>
<ref id="B22"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Buysse</surname> <given-names>L.</given-names></name></person-group> (<year>2012</year>). <article-title>So as a multifunctional discourse marker in native and learner speech.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>44</volume> <fpage>1764</fpage>&#x2013;<lpage>1782</lpage>. <pub-id pub-id-type="doi">10.1016/j.pragma.2012.08.012</pub-id></citation></ref>
<ref id="B23"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Buysse</surname> <given-names>L.</given-names></name></person-group> (<year>2014</year>). <article-title>So what&#x2019;s a year in a lifetime so&#x201D;. Non-prefatory use of so in native and learner English.</article-title> <source><italic>Text Talk</italic></source> <volume>34</volume> <fpage>23</fpage>&#x2013;<lpage>47</lpage>. <pub-id pub-id-type="doi">10.1515/text-2013-0036</pub-id></citation></ref>
<ref id="B24"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Buysse</surname> <given-names>L.</given-names></name></person-group> (<year>2020</year>). <article-title>&#x2018;It was a bit stressy as well actually&#x2019;. The pragmatic markers actually and in fact in spoken learner English.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>156</volume> <fpage>28</fpage>&#x2013;<lpage>40</lpage>. <pub-id pub-id-type="doi">10.1016/j.pragma.2018.11.004</pub-id></citation></ref>
<ref id="B25"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Carter</surname> <given-names>R.</given-names></name> <name><surname>McCarthy</surname> <given-names>M.</given-names></name></person-group> (<year>2006</year>). <source><italic>Cambridge grammar of English: A comprehensive guide: Spoken and written English grammar and usage.</italic></source> <publisher-loc>Cambridge, MA</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>.</citation></ref>
<ref id="B26"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cepeda</surname> <given-names>G.</given-names></name> <name><surname>Poblete</surname> <given-names>M. T.</given-names></name></person-group> (<year>2006</year>). <article-title>Politeness and modality: Discourse markers.</article-title> <source><italic>Rev. Signos</italic></source> <volume>39</volume> <fpage>357</fpage>&#x2013;<lpage>377</lpage>. <pub-id pub-id-type="doi">10.4067/S0718-09342006000300002</pub-id> <pub-id pub-id-type="pmid">27315006</pub-id></citation></ref>
<ref id="B27"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Collet</surname> <given-names>C.</given-names></name> <name><surname>Diemer</surname> <given-names>S.</given-names></name> <name><surname>Brunner</surname> <given-names>M.</given-names></name></person-group> (<year>2021</year>). <article-title>So in video-mediated communication in the Expanding Circle.</article-title> <source><italic>World Engl.</italic></source> <volume>40</volume> <fpage>594</fpage>&#x2013;<lpage>610</lpage>. <pub-id pub-id-type="doi">10.1111/weng.12543</pub-id></citation></ref>
<ref id="B28"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cook</surname> <given-names>J.</given-names></name></person-group> (<year>2014</year>). <article-title>Interaction of face and rapport in an American TV talk show.</article-title> <source><italic>Lang. Res.</italic></source> <volume>50</volume> <fpage>311</fpage>&#x2013;<lpage>332</lpage>.</citation></ref>
<ref id="B29"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Crible</surname> <given-names>L.</given-names></name></person-group> (<year>2018</year>). <source><italic>Discourse markers and (dis)fluency: Forms and functions across languages and registers.</italic></source> <publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins Publishing Company</publisher-name>. <pub-id pub-id-type="doi">10.1075/pbns.286</pub-id> <pub-id pub-id-type="pmid">33486653</pub-id></citation></ref>
<ref id="B30"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Crible</surname> <given-names>L.</given-names></name> <name><surname>Degand</surname> <given-names>L.</given-names></name></person-group> (<year>2019</year>). <article-title>Domains and Functions: A Two-Dimensional Account of Discourse Markers.</article-title> <source><italic>Discours</italic></source> <volume>24</volume>:<issue>33</issue>. <pub-id pub-id-type="doi">10.4000/discours.9997</pub-id></citation></ref>
<ref id="B31"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Crible</surname> <given-names>L.</given-names></name> <name><surname>Degand</surname> <given-names>L.</given-names></name></person-group> (<year>2021</year>). <article-title>Co-occurrence and ordering of discourse markers in sequences: A multifactorial study in spoken French.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>177</volume> <fpage>18</fpage>&#x2013;<lpage>28</lpage>. <pub-id pub-id-type="doi">10.1016/j.pragma.2021.02.006</pub-id></citation></ref>
<ref id="B32"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cuenca</surname> <given-names>M. J.</given-names></name> <name><surname>Crible</surname> <given-names>L.</given-names></name></person-group> (<year>2019</year>). <article-title>Co-occurrence of discourse markers in English: From juxtaposition to composition.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>140</volume> <fpage>171</fpage>&#x2013;<lpage>184</lpage>. <pub-id pub-id-type="doi">10.1016/j.pragma.2018.12.001</pub-id></citation></ref>
<ref id="B33"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Culpeper</surname> <given-names>J.</given-names></name></person-group> (<year>2005</year>). <article-title>Impoliteness and entertainment in the television quiz show: The Weakest Link.</article-title> <source><italic>J. Politeness Res.</italic></source> <volume>1</volume> <fpage>35</fpage>&#x2013;<lpage>72</lpage>. <pub-id pub-id-type="doi">10.1515/jplr.2005.1.1.35</pub-id></citation></ref>
<ref id="B34"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Damayanti</surname> <given-names>A.</given-names></name> <name><surname>Mubarak</surname> <given-names>Z.</given-names></name></person-group> (<year>2021</year>). <article-title>Positive politeness in &#x201C;Oprah&#x2019;s 2020 vision tour&#x201D; how reasons and factors influenced the choice of strategy.</article-title> <source><italic>J. Basis Upb</italic></source> <volume>8</volume> <fpage>13</fpage>&#x2013;<lpage>22</lpage>. <pub-id pub-id-type="doi">10.33884/basisupb.v8i1.2790</pub-id></citation></ref>
<ref id="B35"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Farahani</surname> <given-names>M. V.</given-names></name> <name><surname>Ghane</surname> <given-names>Z.</given-names></name></person-group> (<year>2022</year>). <article-title>Unpacking the function(s) of discourse markers in academic spoken English: A corpus-based study.</article-title> <source><italic>Aust. J. Lang. Lit.</italic></source> <volume>45</volume> <fpage>49</fpage>&#x2013;<lpage>70</lpage>. <pub-id pub-id-type="doi">10.1007/s44020-022-00005-3</pub-id></citation></ref>
<ref id="B36"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fedyna</surname> <given-names>M.</given-names></name></person-group> (<year>2016</year>). <article-title>The Pragmatics of Politeness in the American TV Talk Show Piers Morgan Live.</article-title> <source><italic>Inozenma Philol.</italic></source> <volume>129</volume> <fpage>81</fpage>&#x2013;<lpage>90</lpage>. <pub-id pub-id-type="doi">10.30970/fpl.2016.129.600</pub-id></citation></ref>
<ref id="B37"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Foolen</surname> <given-names>A.</given-names></name></person-group> (<year>2011</year>). &#x201C;<article-title>Pragmatic markers in a sociopragmatic perspective</article-title>,&#x201D; in <source><italic>Pragmatics of Society</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>Andersen</surname> <given-names>G.</given-names></name> <name><surname>Aijmer</surname> <given-names>K.</given-names></name></person-group> (<publisher-loc>Berlin</publisher-loc>: <publisher-name>De Gruyter Mouton</publisher-name>), <fpage>217</fpage>&#x2013;<lpage>242</lpage>. <pub-id pub-id-type="doi">10.1515/9783110214420.217</pub-id></citation></ref>
<ref id="B38"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fraser</surname> <given-names>B.</given-names></name></person-group> (<year>1996</year>). <article-title>Pragmatic Markers.</article-title> <source><italic>Pragmatics</italic></source> <volume>6</volume> <fpage>167</fpage>&#x2013;<lpage>190</lpage>. <pub-id pub-id-type="doi">10.1075/prag.6.2.03fra</pub-id> <pub-id pub-id-type="pmid">33486653</pub-id></citation></ref>
<ref id="B39"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fraser</surname> <given-names>B.</given-names></name></person-group> (<year>1999</year>). <article-title>What are discourse markers?.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>31</volume> <fpage>931</fpage>&#x2013;<lpage>952</lpage>. <pub-id pub-id-type="doi">10.1016/S0378-2166(98)00101-5</pub-id></citation></ref>
<ref id="B40"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fraser</surname> <given-names>B.</given-names></name></person-group> (<year>2009</year>). <article-title>An Account of Discourse Markers.</article-title> <source><italic>Int. Rev. Pragmat.</italic></source> <volume>1</volume> <fpage>293</fpage>&#x2013;<lpage>320</lpage>. <pub-id pub-id-type="doi">10.1163/187730909X12538045489818</pub-id></citation></ref>
<ref id="B41"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fraser</surname> <given-names>B.</given-names></name></person-group> (<year>2010</year>). <article-title>The Sequencing of Contrastive Discourse Markers in English.</article-title> <source><italic>Baltic J. Engl. Lang. Lit. Cult.</italic></source> <volume>1</volume>, <fpage>1</fpage>&#x2013;<lpage>7</lpage>.</citation></ref>
<ref id="B42"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fraser</surname> <given-names>B.</given-names></name></person-group> (<year>2013</year>). <article-title>Combinations of Contrastive Discourse Markers in English.</article-title> <source><italic>Int. Rev. Pragmat.</italic></source> <volume>5</volume> <fpage>318</fpage>&#x2013;<lpage>340</lpage>. <pub-id pub-id-type="doi">10.1163/18773109-13050209</pub-id></citation></ref>
<ref id="B43"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fraser</surname> <given-names>B.</given-names></name></person-group> (<year>2015</year>). <article-title>The Combining of Discourse Markers-A beginning.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>86</volume> <fpage>48</fpage>&#x2013;<lpage>53</lpage>.</citation></ref>
<ref id="B44"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Furk&#x00F3;</surname> <given-names>B. P.</given-names></name></person-group> (<year>2015</year>). <article-title>From mediatized political discourse to The Hobbit: The role of pragmatic markers in the construction of dialogues, stereotypes and literary style.</article-title> <source><italic>Lang. Dialogue</italic></source> <volume>5</volume> <fpage>264</fpage>&#x2013;<lpage>282</lpage>.</citation></ref>
<ref id="B45"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Furk&#x00F3;</surname> <given-names>P.</given-names></name> <name><surname>Abuczki</surname> <given-names>A.</given-names></name></person-group> (<year>2014</year>). <article-title>English Discourse Markers in Mediatised Political Interviews.</article-title> <source><italic>Brno Stud. Engl.</italic></source> <volume>40</volume> <fpage>45</fpage>&#x2013;<lpage>64</lpage>. <pub-id pub-id-type="doi">10.5817/BSE2014-1-3</pub-id></citation></ref>
<ref id="B46"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hancil</surname> <given-names>S.</given-names></name></person-group> (<year>2018</year>). <article-title>Discourse coherence and intersubjectivity: The development of final but in dialogues.</article-title> <source><italic>Lang. Sci.</italic></source> <volume>68</volume> <fpage>78</fpage>&#x2013;<lpage>93</lpage>. <pub-id pub-id-type="doi">10.1016/j.langsci.2017.12.002</pub-id></citation></ref>
<ref id="B47"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Handford</surname> <given-names>M.</given-names></name></person-group> (<year>2015</year>). &#x201C;<article-title>Corpus analysis</article-title>,&#x201D; in <source><italic>Research methods in intercultural communication: A practical guide</italic></source>, <role>ed.</role> <person-group person-group-type="editor"><name><surname>Zhu</surname> <given-names>H.</given-names></name></person-group> (<publisher-loc>Hoboken, NJ</publisher-loc>: <publisher-name>John Wiley &#x0026; Sons</publisher-name>), <fpage>311</fpage>&#x2013;<lpage>326</lpage>. <pub-id pub-id-type="doi">10.1002/9781119166283.ch21</pub-id></citation></ref>
<ref id="B48"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Heng</surname> <given-names>R. Q.</given-names></name></person-group> (<year>2005</year>). <article-title>A Pragmatic Account of the Discourse Marker &#x201C;And&#x201D; in Conversational Interaction.</article-title> <source><italic>Shandong Foreign Lang. Teach. J.</italic></source> <volume>107</volume> <fpage>23</fpage>&#x2013;<lpage>25</lpage>.</citation></ref>
<ref id="B49"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Holtgraves</surname> <given-names>T.</given-names></name> <name><surname>Bonnefon</surname> <given-names>J. F.</given-names></name></person-group> (<year>2017</year>). &#x201C;<article-title>Experimental Approaches to Linguistic. (Im)politeness</article-title>,&#x201D; in <source><italic>The Palgrave. Handbook of Linguistic (Im)politeness</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>Culpeper</surname> <given-names>J.</given-names></name> <name><surname>Haugh</surname> <given-names>M.</given-names></name> <name><surname>K&#x00E1;d&#x00E1;r</surname> <given-names>D.</given-names></name></person-group> (<publisher-loc>London</publisher-loc>: <publisher-name>Palgrave Macmillan</publisher-name>), <fpage>381</fpage>&#x2013;<lpage>401</lpage>. <pub-id pub-id-type="doi">10.1057/978-1-137-37508-7_15</pub-id> <pub-id pub-id-type="pmid">27193501</pub-id></citation></ref>
<ref id="B50"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Huang</surname> <given-names>L.</given-names></name></person-group> (<year>2019</year>). <article-title>A Corpus-Based Exploration of the Discourse Marker Well in Spoken Interlanguage.</article-title> <source><italic>Lang. Speech</italic></source> <volume>62</volume> <fpage>570</fpage>&#x2013;<lpage>593</lpage>. <pub-id pub-id-type="doi">10.1177/0023830918798863</pub-id> <pub-id pub-id-type="pmid">30232924</pub-id></citation></ref>
<ref id="B51"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hutchby</surname> <given-names>I.</given-names></name></person-group> (<year>2006</year>). <source><italic>Media talk: Conversation analysis and the study of broadcasting.</italic></source> <publisher-loc>Berkshire</publisher-loc>: <publisher-name>Open University Press</publisher-name>.</citation></ref>
<ref id="B52"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ilie</surname> <given-names>C.</given-names></name></person-group> (<year>2001</year>). <article-title>Semi-institutional discourse: The case of talk shows.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>33</volume> <fpage>209</fpage>&#x2013;<lpage>254</lpage>. <pub-id pub-id-type="doi">10.1016/S0378-2166(99)00133-2</pub-id></citation></ref>
<ref id="B53"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kaltenb&#x00F6;,ck</surname> <given-names>G.</given-names></name> <name><surname>Ten Wolde</surname> <given-names>E.</given-names></name></person-group> (<year>2022</year>). <article-title>A Just So Story: On the recent emergence of the purpose subordinator just so.</article-title> <source><italic>Engl. Lang. Linguist.</italic></source> <fpage>1</fpage>&#x2013;<lpage>27</lpage>. <pub-id pub-id-type="doi">10.1017/S136067432200020X</pub-id></citation></ref>
<ref id="B54"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kamil Ali</surname> <given-names>H.</given-names></name></person-group> (<year>2018</year>). <article-title>Conversation analysis of the structural units of interaction in american and iraqi TV talk shows: The doctors and shabab wbanat.</article-title> <source><italic>Int. J. Lang. Acad.</italic></source> <volume>6</volume> <fpage>311</fpage>&#x2013;<lpage>333</lpage>. <pub-id pub-id-type="doi">10.18033/ijla.3870</pub-id></citation></ref>
<ref id="B55"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kang</surname> <given-names>M.</given-names></name></person-group> (<year>2018</year>). <article-title>Functions of I mean in American Talk Shows.</article-title> <source><italic>Sociolinguist. J. Korea</italic></source> <volume>26</volume> <fpage>63</fpage>&#x2013;<lpage>84</lpage>. <pub-id pub-id-type="doi">10.14353/sjk.2018.26.2.03</pub-id></citation></ref>
<ref id="B56"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kang</surname> <given-names>M.</given-names></name></person-group> (<year>2019</year>). <article-title>Analysis of Discourse Marker Ani: Focusing on Korean Talk Shows.</article-title> <source><italic>Eoneohag</italic></source> <volume>85</volume> <fpage>81</fpage>&#x2013;<lpage>98</lpage>.</citation></ref>
<ref id="B57"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kazemian</surname> <given-names>R.</given-names></name> <name><surname>Amouzadeh</surname> <given-names>M.</given-names></name></person-group> (<year>2022</year>). <article-title>Aspects of v&#x00E6; (&#x201C;and&#x201D;) as a discourse marker in Persian.</article-title> <source><italic>Pragmatics</italic></source> <volume>32</volume> <fpage>588</fpage>&#x2013;<lpage>619</lpage>. <pub-id pub-id-type="doi">10.1075/prag.21011.amo</pub-id> <pub-id pub-id-type="pmid">33486653</pub-id></citation></ref>
<ref id="B58"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Khammee</surname> <given-names>K.</given-names></name></person-group> (<year>2022</year>). <article-title>On Pragmatics of Contrastiveness: The Discourse Marker But in English, Thai, and Korean.</article-title> <source><italic>J. Linguist. Sci.</italic></source> <volume>100</volume> <fpage>1</fpage>&#x2013;<lpage>18</lpage>. <pub-id pub-id-type="doi">10.21296/jls.2022.3.100.1</pub-id></citation></ref>
<ref id="B59"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kim</surname> <given-names>M.</given-names></name> <name><surname>Kim</surname> <given-names>S.</given-names></name> <name><surname>Sohn</surname> <given-names>S.</given-names></name></person-group> (<year>2021a</year>). <article-title>The Korean discourse particle ya across multiple turn positions: An interactional resource for turn-taking and stance-taking.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>186</volume> <fpage>251</fpage>&#x2013;<lpage>276</lpage>. <pub-id pub-id-type="doi">10.1016/j.pragma.2021.10.012</pub-id></citation></ref>
<ref id="B60"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kim</surname> <given-names>M.</given-names></name> <name><surname>Rhee</surname> <given-names>S.</given-names></name> <name><surname>Smith</surname> <given-names>H.</given-names></name></person-group> (<year>2021b</year>). <article-title>The Korean discourse particle kulssey across discrete positions and contexts in talk-in-interaction.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>182</volume> <fpage>16</fpage>&#x2013;<lpage>41</lpage>. <pub-id pub-id-type="doi">10.1016/j.pragma.2021.06.002</pub-id></citation></ref>
<ref id="B61"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Koops</surname> <given-names>C.</given-names></name> <name><surname>Lohmann</surname> <given-names>A.</given-names></name></person-group> (<year>2022</year>). <article-title>Explaining reversible discourse marker sequences: A case study of and and so.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>191</volume> <fpage>156</fpage>&#x2013;<lpage>171</lpage>. <pub-id pub-id-type="doi">10.1016/j.pragma.2022.01.014</pub-id></citation></ref>
<ref id="B62"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kopf</surname> <given-names>S.</given-names></name></person-group> (<year>2022</year>). <article-title>Participation and deliberative discourse on social media-Wikipedia talk pages as transnational public spheres?.</article-title> <source><italic>Crit. Discourse Stud.</italic></source> <volume>19</volume> <fpage>196</fpage>&#x2013;<lpage>211</lpage>. <pub-id pub-id-type="doi">10.1080/17405904.2020.1822896</pub-id></citation></ref>
<ref id="B63"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lam</surname> <given-names>P.</given-names></name></person-group> (<year>2007</year>). <source><italic>Discourse Particles in an Intercultural Corpus of Spoken English.</italic></source> Ph.D. thesis. <publisher-loc>Hong Kong</publisher-loc>: <publisher-name>The Hong Kong Polytechnic University</publisher-name>.</citation></ref>
<ref id="B64"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lauerbach</surname> <given-names>G.</given-names></name> <name><surname>Aijmer</surname> <given-names>K.</given-names></name></person-group> (<year>2007</year>). <article-title>Argumentation in dialogic media genres&#x2014;Talk shows and interviews.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>39</volume> <fpage>1333</fpage>&#x2013;<lpage>1341</lpage>. <pub-id pub-id-type="doi">10.1016/j.pragma.2007.04.007</pub-id></citation></ref>
<ref id="B65"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Leech</surname> <given-names>G.</given-names></name></person-group> (<year>2005</year>). &#x201C;<article-title>Adding Linguistic Annotation</article-title>,&#x201D; in <source><italic>Developing. Linguistic Corpora: A Guide to Good Practice</italic></source>, <role>ed.</role> <person-group person-group-type="editor"><name><surname>Wynne</surname> <given-names>M.</given-names></name></person-group> (<publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxbow Books</publisher-name>), <fpage>17</fpage>&#x2013;<lpage>29</lpage>.</citation></ref>
<ref id="B66"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Leech</surname> <given-names>G.</given-names></name></person-group> (<year>2014</year>). <source><italic>The pragmatics of politeness.</italic></source> <publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>. <pub-id pub-id-type="doi">10.1093/acprof:oso/9780195341386.001.0001</pub-id> <pub-id pub-id-type="pmid">36389024</pub-id></citation></ref>
<ref id="B67"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Li</surname> <given-names>X.</given-names></name> <name><surname>Xiang</surname> <given-names>Y.</given-names></name></person-group> (<year>2020</year>). <article-title>Analysis of the Pragmatic Functions of the Discourse Marker So in Breaking Bad.</article-title> <source><italic>J. Hubei Univ. Technol.</italic></source> <volume>35</volume> <fpage>89</fpage>&#x2013;<lpage>93</lpage>.</citation></ref>
<ref id="B68"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Liang</surname> <given-names>M. C.</given-names></name> <name><surname>Li</surname> <given-names>W. Z.</given-names></name> <name><surname>Xu</surname> <given-names>J. J.</given-names></name></person-group> (<year>2019</year>). <source><italic>Using Corpora: A Practical Coursebook.</italic></source> <publisher-loc>Beijing</publisher-loc>: <publisher-name>Foreign Language Teaching and Research Press</publisher-name>.</citation></ref>
<ref id="B69"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Liu</surname> <given-names>B. M.</given-names></name></person-group> (<year>2017</year>). <article-title>The use of discourse markers but and so by native English speakers and Chinese speakers of English.</article-title> <source><italic>Pragmatics</italic></source> <volume>27</volume> <fpage>479</fpage>&#x2013;<lpage>506</lpage>.</citation></ref>
<ref id="B70"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>M&#x00FC;ller</surname> <given-names>S.</given-names></name></person-group> (<year>2005</year>). <source><italic>Discourse markers in native and non-native English discourse.</italic></source> <publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins Pub.</publisher-name> <pub-id pub-id-type="doi">10.1075/pbns.138</pub-id> <pub-id pub-id-type="pmid">33486653</pub-id></citation></ref>
<ref id="B71"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Nneka</surname> <given-names>O.</given-names></name></person-group> (<year>2022</year>). <article-title>&#x201C;Well&#x201D; and &#x201C;So&#x201D; as Discourse Marker in President Goodluck Ebele Jonathan&#x2019;s Media Chat.</article-title> <source><italic>Ansu J. Lang. Lit.</italic></source> <volume>2</volume> <fpage>82</fpage>&#x2013;<lpage>91</lpage>.</citation></ref>
<ref id="B72"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Norrick</surname> <given-names>N. R.</given-names></name></person-group> (<year>2001</year>). <article-title>Discourse markers in oral narrative.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>33</volume> <fpage>849</fpage>&#x2013;<lpage>878</lpage>. <pub-id pub-id-type="doi">10.1016/S0378-2166(01)80032-1</pub-id></citation></ref>
<ref id="B73"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>&#x00D6;stman</surname> <given-names>J. O.</given-names></name></person-group> (<year>1981</year>). <source><italic>You know: A discourse functional approach.</italic></source> <publisher-loc>Amsterdam</publisher-loc>: <publisher-name>Benjamins</publisher-name>. <pub-id pub-id-type="doi">10.1075/pb.ii.7</pub-id> <pub-id pub-id-type="pmid">33486653</pub-id></citation></ref>
<ref id="B74"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Oyeleye</surname> <given-names>A. K.</given-names></name> <name><surname>Olutayo</surname> <given-names>O. G.</given-names></name></person-group> (<year>2012</year>). <article-title>Interaction Management in Nigerian Television Talk Shows.</article-title> <source><italic>Int. J. Engl. Linguist.</italic></source> <volume>2</volume> <fpage>149</fpage>&#x2013;<lpage>161</lpage>. <pub-id pub-id-type="doi">10.5539/ijel.v2n1p149</pub-id></citation></ref>
<ref id="B75"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Rabab&#x2019;Ah</surname> <given-names>B.</given-names></name> <name><surname>Rabab&#x2019;Ah</surname> <given-names>G.</given-names></name> <name><surname>Naimi</surname> <given-names>T.</given-names></name></person-group> (<year>2019</year>). <article-title>Oprah Winfrey Talk show: An analysis of the relationship between positive politeness strategies and speaker&#x2019;s ethnic background.</article-title> <source><italic>Kemanusiaan</italic></source> <volume>26</volume> <fpage>25</fpage>&#x2013;<lpage>50</lpage>. <pub-id pub-id-type="doi">10.21315/kajh2019.26.1.2</pub-id> <pub-id pub-id-type="pmid">33679231</pub-id></citation></ref>
<ref id="B76"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Rayson</surname> <given-names>P.</given-names></name></person-group> (<year>2016</year>). <source><italic>Log-likelihood and effect size calculator.</italic></source> Available online at: <ext-link ext-link-type="uri" xlink:href="https://ucrel.lancs.ac.uk/llwizard.html">https://ucrel.lancs.ac.uk/llwizard.html</ext-link></citation></ref>
<ref id="B77"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Redeker</surname> <given-names>G.</given-names></name></person-group> (<year>1990</year>). <article-title>Ideational and pragmatic markers of discourse structure.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>14</volume> <fpage>367</fpage>&#x2013;<lpage>381</lpage>. <pub-id pub-id-type="doi">10.1016/0378-2166(90)90095-U</pub-id></citation></ref>
<ref id="B78"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Rendle-Short</surname> <given-names>J.</given-names></name></person-group> (<year>2003</year>). <article-title>So what does this show us?</article-title> <source><italic>Aust. Rev. Appl. Linguist.</italic></source> <volume>26</volume> <fpage>46</fpage>&#x2013;<lpage>62</lpage>. <pub-id pub-id-type="doi">10.1075/aral.26.2.04ren</pub-id> <pub-id pub-id-type="pmid">33486653</pub-id></citation></ref>
<ref id="B79"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Reppen</surname> <given-names>R.</given-names></name></person-group> (<year>2010</year>). &#x201C;<article-title>Building a corpus. What are key considerations?</article-title>,&#x201D; in <source><italic>The Routledge Handbook of Corpus Linguistics</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>O&#x2019;Keeffe</surname> <given-names>A.</given-names></name> <name><surname>McCarthy</surname> <given-names>M. J.</given-names></name></person-group> (<publisher-loc>London</publisher-loc>: <publisher-name>Routledge</publisher-name>), <fpage>31</fpage>&#x2013;<lpage>37</lpage>. <pub-id pub-id-type="doi">10.4324/9780203856949.ch3</pub-id> <pub-id pub-id-type="pmid">36153787</pub-id></citation></ref>
<ref id="B80"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Rezanova</surname> <given-names>Z. I.</given-names></name> <name><surname>Kogut</surname> <given-names>S. V.</given-names></name></person-group> (<year>2015</year>). <article-title>Types of Discourse Markers: Their Ethnocultural. Diversity in Scientific Text.</article-title> <source><italic>Proc. Soc. Behav. Sci.</italic></source> <volume>215</volume> <fpage>266</fpage>&#x2013;<lpage>272</lpage>. <pub-id pub-id-type="doi">10.1016/j.sbspro.2015.11.633</pub-id></citation></ref>
<ref id="B81"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>R&#x00FC;hlemann</surname> <given-names>C.</given-names></name></person-group> (<year>2019</year>). <source><italic>Corpus linguistics for pragmatics: A guide for research.</italic></source> <publisher-loc>London</publisher-loc>: <publisher-name>Routledge</publisher-name>.</citation></ref>
<ref id="B82"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>&#x015E;ahin K&#x0131;z&#x0131;l</surname> <given-names>A.</given-names></name></person-group> (<year>2021</year>). <article-title>Discourse Markers in Learner Speech: A Corpus-Based Comparative Study.</article-title> <source><italic>J. Lang. Educ. Res.</italic></source> <volume>7</volume> <fpage>1</fpage>&#x2013;<lpage>16</lpage>. <pub-id pub-id-type="doi">10.31464/jlere.769613</pub-id></citation></ref>
<ref id="B83"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Schiffrin</surname> <given-names>D.</given-names></name></person-group> (<year>1987</year>). <source><italic>Discourse markers.</italic></source> <publisher-loc>Cambridge, MA</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>.</citation></ref>
<ref id="B84"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Schirm</surname> <given-names>A.</given-names></name></person-group> (<year>2012</year>). <source><italic>Discourse markers in talk shows: The examples of Hungarian h&#x00E1;t.</italic></source> <publisher-loc>Lisla, IL</publisher-loc>: <publisher-name>Triera communications</publisher-name>, <fpage>124</fpage>&#x2013;<lpage>133</lpage>.</citation></ref>
<ref id="B85"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Schourup</surname> <given-names>L.</given-names></name></person-group> (<year>1999</year>). <article-title>Discourse markers.</article-title> <source><italic>Lingua</italic></source> <volume>107</volume> <fpage>227</fpage>&#x2013;<lpage>265</lpage>. <pub-id pub-id-type="doi">10.1016/S0024-3841(96)90026-1</pub-id></citation></ref>
<ref id="B86"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sharifi</surname> <given-names>M.</given-names></name> <name><surname>Ansari</surname> <given-names>N.</given-names></name> <name><surname>Asadollahzadeh</surname> <given-names>M.</given-names></name></person-group> (<year>2017</year>). <article-title>A critical discourse analytic approach to discursive construction of Islam in Western talk shows: The case of CNN talk shows.</article-title> <source><italic>Int. Commun. Gazette</italic></source> <volume>79</volume> <fpage>45</fpage>&#x2013;<lpage>63</lpage>. <pub-id pub-id-type="doi">10.1177/1748048516656301</pub-id></citation></ref>
<ref id="B87"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sheikhan</surname> <given-names>A.</given-names></name> <name><surname>Haugh</surname> <given-names>M.</given-names></name></person-group> (<year>2022</year>). <article-title>Non-serious answers to (improper) questions in talk. shows.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>191</volume> <fpage>32</fpage>&#x2013;<lpage>45</lpage>. <pub-id pub-id-type="doi">10.1016/j.pragma.2022.01.020</pub-id></citation></ref>
<ref id="B88"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Shirtz</surname> <given-names>S.</given-names></name></person-group> (<year>2021</year>). <article-title>And now, co-occurrence and functionality of discourse markers on the Oregon Coast.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>185</volume> <fpage>164</fpage>&#x2013;<lpage>175</lpage>. <pub-id pub-id-type="doi">10.1016/j.pragma.2021.09.001</pub-id></citation></ref>
<ref id="B89"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Shirzadi</surname> <given-names>S.</given-names></name> <name><surname>Amoozadeh</surname> <given-names>M.</given-names></name> <name><surname>Kalantari</surname> <given-names>S.</given-names></name></person-group> (<year>2022</year>). <article-title>Pragmatic Aspects of M&#x00E6;g&#x00E6;r (&#x2018;unless&#x2019;/&#x2018;but&#x2019;) as a Discourse Marker in Persian.</article-title> <source><italic>J. Lang. Res.</italic></source> <volume>2</volume> <fpage>123</fpage>&#x2013;<lpage>146</lpage>. <pub-id pub-id-type="doi">10.22059/JOLR.2021.323183.666712</pub-id></citation></ref>
<ref id="B90"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sitorus</surname> <given-names>M. L.</given-names></name> <name><surname>Mono</surname> <given-names>U.</given-names></name> <name><surname>Setia</surname> <given-names>E.</given-names></name></person-group> (<year>2022</year>). <article-title>Language Politeness on Mata Najwa&#x2019;s Talk Show with Covid-19 Theme: Sociopragmatics.</article-title> <source><italic>Budapest Int. Res. Crit. Inst.</italic></source> <volume>5</volume> <fpage>3749</fpage>&#x2013;<lpage>3760</lpage>.</citation></ref>
<ref id="B91"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tanghe</surname> <given-names>S.</given-names></name></person-group> (<year>2016</year>). <article-title>Position and polyfunctionality of discourse markers: The case of Spanish markers derived from motion verbs.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>93</volume> <fpage>16</fpage>&#x2013;<lpage>31</lpage>. <pub-id pub-id-type="doi">10.1016/j.pragma.2015.12.002</pub-id></citation></ref>
<ref id="B92"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tolson</surname> <given-names>A.</given-names></name></person-group> (<year>2006</year>). <source><italic>Media talk: Spoken discourse on TV and radio.</italic></source> <publisher-loc>Edinburgh</publisher-loc>: <publisher-name>Edinburgh University Press</publisher-name>, <pub-id pub-id-type="doi">10.3366/j.ctt1g09vbv</pub-id></citation></ref>
<ref id="B93"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tsoy</surname> <given-names>E.</given-names></name></person-group> (<year>2022</year>). <article-title>Functions of Russian verb-derived discourse markers slu&#x0161;aj and smotri from the perspective of information management and interpersonal regulation.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>193</volume> <fpage>105</fpage>&#x2013;<lpage>121</lpage>. <pub-id pub-id-type="doi">10.1016/j.pragma.2022.03.010</pub-id></citation></ref>
<ref id="B94"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wynne</surname> <given-names>M.</given-names></name></person-group> (<year>2005</year>). <source><italic>Developing linguistic corpora: A guide to good practice.</italic></source> <publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxbow Books</publisher-name>.</citation></ref>
</ref-list>
</back>
</article>
