<?xml version="1.0" encoding="UTF-8" standalone="no"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xml:lang="EN" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Psychol.</journal-id>
<journal-title>Frontiers in Psychology</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Psychol.</abbrev-journal-title>
<issn pub-type="epub">1664-1078</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fpsyg.2022.788438</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Psychology</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Multimodal Irregular Self-Selection in Chinese Postgraduate English as a Foreign Language Learners&#x2019; Conversation: When, How, and Why</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author">
<name><surname>Ji</surname> <given-names>Mengmeng</given-names></name>
<uri xlink:href="http://loop.frontiersin.org/people/1484803/overview"/>
</contrib>
<contrib contrib-type="author" corresp="yes">
<name><surname>Zhang</surname> <given-names>Huiping</given-names></name>
<xref ref-type="corresp" rid="c001"><sup>&#x002A;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/1199073/overview"/>
</contrib>
</contrib-group>
<aff><institution>School of Foreign Languages, Northeast Normal University</institution>, <addr-line>Changchun</addr-line>, <country>China</country></aff>
<author-notes>
<fn fn-type="edited-by"><p>Edited by: Xiaowei Zhao, Emmanuel College, United States</p></fn>
<fn fn-type="edited-by"><p>Reviewed by: Hang Su, Sichuan International Studies University, China; Ping Yang, Western Sydney University, Australia</p></fn>
<corresp id="c001">&#x002A;Correspondence: Huiping Zhang, <email>zhanghp387@nenu.edu.cn</email></corresp>
<fn fn-type="other" id="fn004"><p>This article was submitted to Language Sciences, a section of the journal Frontiers in Psychology</p></fn>
</author-notes>
<pub-date pub-type="epub">
<day>25</day>
<month>03</month>
<year>2022</year>
</pub-date>
<pub-date pub-type="collection">
<year>2022</year>
</pub-date>
<volume>13</volume>
<elocation-id>788438</elocation-id>
<history>
<date date-type="received">
<day>04</day>
<month>10</month>
<year>2021</year>
</date>
<date date-type="accepted">
<day>23</day>
<month>02</month>
<year>2022</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x00A9; 2022 Ji and Zhang.</copyright-statement>
<copyright-year>2022</copyright-year>
<copyright-holder>Ji and Zhang</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/"><p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p></license>
</permissions>
<abstract>
<p>Irregular self-selection is a demonstration of active involvement in interaction. English as a foreign language (EFL) learners&#x2019; talk-in-interaction is one of such cases. Yet, little research has explored when, how, and why learners implement this action. The aim of this article is to address these issues in Chinese postgraduate EFL learners&#x2019; conversations from the perspective of multimodal interaction. To this end, we provide descriptive statistics and use multimodal conversation analysis to investigate the detailed process of irregular self-selection. The results show the interactional sensitivity of learners, and all the successful irregular self-selection can be divided into the following three types: turn interruption (TI), turn competition TC), and turn holding abortion (THA). Learners implement this action by using multimodal resources, including lexis, syntax, pitch reset, intensity enhancement, gaze, and so forth. However, their body movements lack diversity, causing behaviors to be constrained and inactive. The main purpose of irregular self-selection is to provide knowledge that contributes to topical development. This study reveals that Chinese postgraduate EFL learners are interactionally competent members. They are able to achieve communicative goals but in a low diversity of body movements. The findings help to understand the detailed process of speakership claiming in EFL learners&#x2019; conversations.</p>
</abstract>
<kwd-group>
<kwd>irregular self-selection</kwd>
<kwd>Chinese postgraduate EFL learners</kwd>
<kwd>descriptive statistics</kwd>
<kwd>multimodal conversation analysis</kwd>
<kwd>detailed process</kwd>
</kwd-group>
<contract-sponsor id="cn001">National Office for Philosophy and Social Sciences<named-content content-type="fundref-id">10.13039/501100012325</named-content></contract-sponsor>
<counts>
<fig-count count="9"/>
<table-count count="5"/>
<equation-count count="0"/>
<ref-count count="61"/>
<page-count count="17"/>
<word-count count="10730"/>
</counts>
</article-meta>
</front>
<body>
<sec id="S1" sec-type="intro">
<title>Introduction</title>
<p>Interaction is one of the core matrixes for human social life. A mechanism for coordinating interaction is a turn-taking system that regulates who is to speak and when (<xref ref-type="bibr" rid="B46">Sacks et al., 1974</xref>). Until now, a substantial amount of research has been conducted concerning how the system works either in ordinary or institutional conversation (e.g., <xref ref-type="bibr" rid="B34">Mondada, 2007</xref>, <xref ref-type="bibr" rid="B35">2013</xref>; <xref ref-type="bibr" rid="B50">Stivers and Rossano, 2010</xref>; <xref ref-type="bibr" rid="B57">Wei&#x00DF;, 2018</xref>; <xref ref-type="bibr" rid="B61">Yang, 2019</xref>; <xref ref-type="bibr" rid="B2">Auer, 2021</xref>). In studies of turn-taking practices, conversation analysis (CA), which adopts a participant-relevant emic perspective, offers a valid tool for observing and analyzing the dynamic details of turn and sequence organization in talk-in-interaction. Recently, how multimodal resources work together in the construction and organization of turns has received much attention (<xref ref-type="bibr" rid="B34">Mondada, 2007</xref>, <xref ref-type="bibr" rid="B35">2013</xref>; <xref ref-type="bibr" rid="B51">Streeck, 2009</xref>; <xref ref-type="bibr" rid="B60">Yang, 2011</xref>; <xref ref-type="bibr" rid="B8">Dressel, 2020</xref>). In line with this research, multimodal CA has largely remained a framework used by and shared with scholars in the social sciences.</p>
<p>However, mainstream multimodal CA research predominantly investigates first language interactions, and thus research on foreign language conversation is still marginalized in the literature (<xref ref-type="bibr" rid="B42">Pietikainen, 2018</xref>). In a few studies of non-native discourse, attention often turns to language classrooms since this is one of the main locations where foreign language learners have the opportunity and need to use the language they are learning (<xref ref-type="bibr" rid="B56">Waring, 2011</xref>). Yet, as noted by <xref ref-type="bibr" rid="B21">Kasper and Wagner (2011)</xref>, language classrooms offer a limited range of interaction. Obviously, turn-taking in the classroom is under the guidance of the teacher, who has a high knowledge status and dominates the class interaction. Because of the low knowledge status and unequal relationship between teachers and students, self-selection is quite challenging and unusual in the teacher-student interaction context (<xref ref-type="bibr" rid="B54">Takahashi, 2018</xref>).</p>
<p>In contrast to it, peer interaction involves students who have equal identity status and similar knowledge status. In such a context, there is a tendency of English as a foreign language (EFL) learners to implement self-selection in an irregular turn-taking way, that is, irregular self-selection (<xref ref-type="bibr" rid="B1">Almoaily, 2020</xref>). However, by referring to the previous literature, relatively little is known about this action in EFL learners&#x2019; conversation from the perspective of multimodal interaction. To fill this gap, the present article aims to examine irregular self-selection in Chinese postgraduate EFL learners&#x2019; conversations by mainly using a multimodal conversation analytic approach. It tries to reveal the nature and essence of this action from when, how, and why to some extent. This article is expected to gain a deeper and more detailed understanding of Chinese postgraduate EFL learners&#x2019; irregular self-selection and add to the research on multimodal turn-taking practices of EFL learners. In all, we seek to address the following research questions:</p>
<list list-type="simple">
<list-item>
<label>(1)</label>
<p>When does irregular self-selection occur (i.e., type)?</p>
</list-item>
<list-item>
<label>(2)</label>
<p>How does the hearer mobilize multimodal resources (in verbal, vocal, and non-verbal aspects) to achieve irregular self-selection?</p>
</list-item>
<list-item>
<label>(3)</label>
<p>Why do they implement irregular self-selection (i.e., purpose)?</p>
</list-item>
</list>
</sec>
<sec id="S2">
<title>Literature Review</title>
<sec id="S2.SS1">
<title>Turn-Taking System and Irregular Self-Selection</title>
<p>A major feature of conversation is that people overwhelmingly talk in turns. The conversation analytic account of how turn-taking is managed has been noted by <xref ref-type="bibr" rid="B46">Sacks et al. (1974)</xref>. They claim that the speech-exchange system consists of turn-constructional components and turn-taking rules. The characteristics of the turn-taking system can be summarized in three points as follows:</p>
<list list-type="simple">
<list-item>
<label>(1)</label>
<p>Turn is made up of turn-constructional units (TCUs). These units can be lexical, phrasal, clausal, or sentential, used to complete a communicative act. The possible completion point (PCP) at the end of each TCU may become a place for speaker transition, called a transition relevance place (TRP), which is the appropriate position for turn-taking.</p>
</list-item>
<list-item>
<label>(2)</label>
<p>An essential feature of TCU is its projectability, which allows participants to project where a turn will reach possible completion. A range of resources is used to project the possible completion of a TCU, including syntax, intonation, and pragmatics (<xref ref-type="bibr" rid="B10">Ford and Thompson, 1996</xref>). That is, a TCU has &#x201C;syntactic, intonational, semantic, and/or pragmatic status as potentially complete&#x201D; (<xref ref-type="bibr" rid="B26">Lazaraton, 2002</xref>, p. 32). To be specific, an utterance is grammatically complete if it could be interpreted as a complete clause in its discourse context. This utterance can be a word, a phrase, or a sentence. Intonation completion refers to a point at which a rising or falling intonation can be clearly heard as a final intonation. An utterance is pragmatically complete when it can be heard as a complete conversational action within its discourse context. When grammatical, intonational, and pragmatic completions of a TCU converge, a complex transition relevance place (CTRP) occurs (<xref ref-type="bibr" rid="B10">Ford and Thompson, 1996</xref>). CTRP is also where actual turn transitions are most likely to occur. Moreover, a substantial body of scholars notes that non-verbal conducts also figure in projecting TCU completion, such as eye gaze, open hand palm up (OHPU), pointing gesture, and so on (<xref ref-type="bibr" rid="B22">Kendon, 1967</xref>; <xref ref-type="bibr" rid="B12">Goodwin, 1980</xref>; <xref ref-type="bibr" rid="B34">Mondada, 2007</xref>; <xref ref-type="bibr" rid="B51">Streeck, 2009</xref>; <xref ref-type="bibr" rid="B30">Li, 2014</xref>).</p>
</list-item>
<list-item>
<label>(3)</label>
<p>The organization of turn-taking obeys three rules. Rule 1: Current speaker selects the next; Rule 2: Other participants self-select for the next speaker; Rule 3: If no hearer takes the turn, the current speaker should continue the turn until the next speaker takes it. These &#x201C;gross observations&#x201D; are critical in understanding how turn-taking is managed and &#x201C;when a turn is complete from the participants&#x2019; perspective&#x201D; (<xref ref-type="bibr" rid="B14">Greer and Potter, 2008</xref>, p. 299). By adhering to turn-taking rules, the ideal condition of turn-taking is featured as only one party speaking at a time, avoidance of overlapping talk, and the minimization of gaps and silences between turns.</p>
</list-item>
</list>
<p>Self-selection is one of the efficient ways to participate in a conversation and gain the speakership actively. The basic principle for the next speaker&#x2019;s self-selection is to &#x201C;start as early as possible at the earliest transition relevance place&#x201D; (<xref ref-type="bibr" rid="B46">Sacks et al., 1974</xref>, p. 719). However, because of the dynamic and unpredictable features of interaction, participants do not always adhere to the turn-taking system to participate in a conversation. Thus, the ideal condition of turn-taking cannot be guaranteed, &#x201C;turn-taking irregularities&#x201D; (<xref ref-type="bibr" rid="B1">Almoaily, 2020</xref>, p. 188) are of frequent occurrence. Violating the turn-taking system, irregular self-selection is mainly characterized as a form of interruption or overlap, which occurs in the following two circumstances: C1. The hearer may self-select at non-TRPs to achieve a certain communicative goal, especially when the speaker is still speaking or has no intention to give up the speaking turn (i.e., not project the turn completion). This kind of behavior results in an interruption in a conversation. C2. The hearer may self-select at a possible TRP (rule 2), which simultaneously occurs with the current speaker&#x2019;s application of rule 3 (i.e., the current speaker&#x2019;s continuation) (<xref ref-type="bibr" rid="B24">Konakahara, 2020</xref>). This kind of behavior results in overlap in a conversation. All of them reflect the dynamic and unpredictable nature of interactions and are vital to the understanding and study of irregular self-selection.</p>
</sec>
<sec id="S2.SS2">
<title>Multimodality of Self-Selection</title>
<p>In earlier CA research, embodied actions are considered to have a subsidiary role in interaction as &#x201C;a hand-maiden to speech&#x201D; (<xref ref-type="bibr" rid="B51">Streeck, 2009</xref>, p. 26), but more recent work has started to consider bodily actions to be as important as talk (<xref ref-type="bibr" rid="B52">Streeck et al., 2011</xref>). A substantial body of scholars have confirmed that a large range of interactional resources is relevant to the organization of turn-taking, encompassing linguistic resources such as lexis, grammar, prosody (<xref ref-type="bibr" rid="B10">Ford and Thompson, 1996</xref>), as well as embodied multimodal resources such as eye gaze (e.g., <xref ref-type="bibr" rid="B22">Kendon, 1967</xref>; <xref ref-type="bibr" rid="B12">Goodwin, 1980</xref>; <xref ref-type="bibr" rid="B13">Goodwin and Goodwin, 1986</xref>; <xref ref-type="bibr" rid="B44">Rossano, 2012</xref>), gesture (<xref ref-type="bibr" rid="B34">Mondada, 2007</xref>; <xref ref-type="bibr" rid="B51">Streeck, 2009</xref>; <xref ref-type="bibr" rid="B59">Yang, 2010</xref>; <xref ref-type="bibr" rid="B30">Li, 2014</xref>), head movement (<xref ref-type="bibr" rid="B33">Markaki and Mondada, 2012</xref>; <xref ref-type="bibr" rid="B31">Li, 2019</xref>), and body posture (<xref ref-type="bibr" rid="B35">Mondada, 2013</xref>; <xref ref-type="bibr" rid="B30">Li, 2014</xref>).</p>
<p>With regard to self-selection, &#x201C;a possible next speaker may start gearing up for his or her turn before the current speaker&#x2019;s turn completion&#x201D; (<xref ref-type="bibr" rid="B27">Lee, 2017</xref>, p. 672), and the participants use multimodal resources to implement this action in various contexts. For example, in a multi-party conversation, the linguistic resources &#x201C;<italic>I&#x2019;m sorry (to interrupt)</italic>&#x201D; are used as self-selection devices to obtain the speakership (<xref ref-type="bibr" rid="B41">Park and Duey, 2020</xref>). The pointing gesture of the hearer also severs the action-projecting function to &#x201C;self-selection for would be next speakers&#x201D; (<xref ref-type="bibr" rid="B34">Mondada, 2007</xref>, p. 207). In a teacher-fronted classroom (<xref ref-type="bibr" rid="B47">Sahlstr&#x00F6;m, 2002</xref>; <xref ref-type="bibr" rid="B25">Lauzon and Berger, 2015</xref>; <xref ref-type="bibr" rid="B54">Takahashi, 2018</xref>), <xref ref-type="bibr" rid="B47">Sahlstr&#x00F6;m (2002)</xref> reported that students used hand raising to self-select as the next speaker. In ordinary conversation (<xref ref-type="bibr" rid="B53">Streeck and Hartge, 1992</xref>; <xref ref-type="bibr" rid="B19">Iwasaki, 2009</xref>), <xref ref-type="bibr" rid="B53">Streeck and Hartge (1992)</xref> observed that facial configurations display the speakers&#x2019; intent. For example, facial expression (a) was used as a self-selection device among Ilokano speakers in their interactions.</p>
<p>Although most of the studies focus on self-selection actions in first language conversations, the remaining non-native speakers&#x2019; interactions are somewhat marginalized. It must be emphasized that EFL learners, &#x201C;despite their limited proficiency in the target language, are interactionally competent members who manage to participate in discussions&#x201D; (<xref ref-type="bibr" rid="B27">Lee, 2017</xref>, p. 673). For instance, <xref ref-type="bibr" rid="B5">Carroll (2004)</xref> observed that Japanese novice speakers of English used recycled turn beginnings (words) in ways similar to those of native speakers of English as a self-selection device. Moreover, non-verbal resources such as gestures, gaze orientation, and posture are also used by learners to show participating interests in conversation (<xref ref-type="bibr" rid="B39">Olsher, 2004</xref>; <xref ref-type="bibr" rid="B23">Konakahara, 2015</xref>, <xref ref-type="bibr" rid="B24">2020</xref>; <xref ref-type="bibr" rid="B55">Taleghani-Nikazm, 2015</xref>; <xref ref-type="bibr" rid="B27">Lee, 2017</xref>; <xref ref-type="bibr" rid="B32">Majlesi and Markee, 2018</xref>). For instance, <xref ref-type="bibr" rid="B27">Lee (2017)</xref> found that learners used to gaze and gesture to prepare for self-selection.</p>
<p>In line with the multimodality of self-selection, irregular self-selection has the same nature. Moreover, this action has been reported as a demonstration of active involvement in interactions, with EFL learners&#x2019; talk-in-interaction being one of such cases (<xref ref-type="bibr" rid="B6">Cogo and Dewey, 2012</xref>; <xref ref-type="bibr" rid="B23">Konakahara, 2015</xref>). Explorations of irregular self-selection could enrich multimodal CA-based turn-taking studies. However, relevant research is still scarce in this field, especially in EFL learners&#x2019; conversations. Thus, more investigations are needed.</p>
</sec>
<sec id="S2.SS3">
<title>Irregular Self-Selection: When, How, and Why</title>
<p>To begin with, scholars have conducted a few studies on when, how, and why learners implement self-selection. They are referential to the study of irregular self-selection. <xref ref-type="bibr" rid="B40">Orletti (1981)</xref> reported that self-selection occurs during a pause or after another speaker has completed the previous turn. Referring to how, <xref ref-type="bibr" rid="B43">Richard and Nunan (1990)</xref> demonstrated that self-selection could be achieved linguistically, non-verbally, pragmatically, and tactically. As for why, <xref ref-type="bibr" rid="B56">Waring (2011)</xref> reported three types of self-selections: to initiate a sequence, to volunteer response, and to proceed with the agenda. Furthermore, <xref ref-type="bibr" rid="B11">Garton (2012)</xref> found that confirmation checks, clarification requests, and information requests were the three most common uses of self-selection. The existing findings are beneficial to understanding the nature of participants&#x2019; self-selection.</p>
<p>Recently, scholars have been interested in how learners engage embodied resources in self-selection. For instance, <xref ref-type="bibr" rid="B27">Lee (2017)</xref> investigated the multimodal resources used by EFL learners to gain primary speakership within their peer group discussions. This study showed that the hearer actively moved into the primary speaker position by utilizing an ensemble of talk, gaze, gesture, and bodily orientation. For instance, learners used to gaze and gesture to claim for the speakership and to prepare for self-selection and used touch to interrupt the ongoing talk to join the conversation. It must be emphasized that both regular and irregular self-selection are a crucial part of the learning process because they both allow learners to claim speakership for the exchange of views, analyses, and opinions. However, the analysis of irregular self-selection is scarce, and only a few studies have explored it in learners&#x2019; conversations (e.g., <xref ref-type="bibr" rid="B15">Guillot, 2009</xref>, <xref ref-type="bibr" rid="B16">2012</xref>; <xref ref-type="bibr" rid="B23">Konakahara, 2015</xref>, <xref ref-type="bibr" rid="B24">2020</xref>; <xref ref-type="bibr" rid="B27">Lee, 2017</xref>). For example, <xref ref-type="bibr" rid="B23">Konakahara (2015)</xref> examined the interactional environment in which overlapping questions occur (i.e., when) and the interactional functions they serve (i.e., why). He found that this kind of irregular self-selection results from the simultaneous application of a next speaker&#x2019;s self-selects, and the current speaker continues turn-taking. Moreover, without clinging to the overlap, participants cooperatively moved the talk forward. <xref ref-type="bibr" rid="B24">Konakahara (2020)</xref> further reported two kinds of irregular self-selections (i.e., floor-taking overlap and floor-attempting overlap) from when and how, but this study did not consider why. <xref ref-type="bibr" rid="B27">Lee (2017)</xref> investigated how learners used touch to interrupt the ongoing talk to join the multi-party interaction. However, when and why has not been concluded in the study.</p>
<p>Previous literature reveals that irregular self-selection research is still insufficient. Thus, the aim of the present study is to enrich the research of this action. We focus on relatively naturally occurring peer conversations among Chinese postgraduate EFL learners, to illustrate their irregular self-selection. This study provides overall descriptive statistics of the number, type, and purpose of this action, and then uses single-case analysis and a multimodal conversation analytic approach. It exemplifies the process of irregular self-selection from when, how, and why in detail by analyzing their turn construction and sequence organization. The study contributes to the growing body of knowledge of the multimodal nature of EFL learners&#x2019; interaction. It also helps to understand what they actually do to achieve successful outcomes in different interactional contexts.</p>
</sec>
</sec>
<sec id="S3" sec-type="materials|methods">
<title>Materials and Methods</title>
<sec id="S3.SS1">
<title>Participants</title>
<p>This study involved 40 Chinese postgraduate EFL learners (4 men and 36 women), who were in the first year of their master&#x2019;s program in September 2020. Their average age was 23 years (<italic>SD</italic> = 1.48; range 21&#x2013;27). They generally shared the same first language background (Chinese) and had studied English for about 13.6 years on average (<italic>SD</italic> = 2.23; range 10&#x2013;18). Their overall English language proficiency can be characterized as high, because they had passed the Test for English Major-8 (TEM-8) with 70.5 points out of 100 on average (<italic>SD</italic> = 4.92; range 65&#x2013;80), and those who can reach 60 points are identified as advanced EFL learners in China. Before recording, informed consent was obtained from the participants at the time of the recruitment, and they volunteered to participate with great zeal.</p>
</sec>
<sec id="S3.SS2">
<title>Data Collection</title>
<p>Before collecting data, participants were not informed of the general study purpose. We only informed them of the video recording and required them to carry on an ordinary, casual, and natural conversation as much as possible. The data are at best characterized as &#x201C;non-pedagogic casual talk&#x201D; (<xref ref-type="bibr" rid="B5">Carroll, 2004</xref>, p. 203), because it was collected after class and was chatted among friends who are familiar with each other. In dyadic dialogue, the participants finally formed 20 peer-to-peer conversation groups by adopting a free combination at their own will. Each group chose one of the 10 topics to discuss, such as &#x201C;friends,&#x201D; &#x201C;travel,&#x201D; and &#x201C;traditional Chinese festival,&#x201D; which had been delivered to them in advance for preparation. The topics were slightly general to give the participants enough &#x201C;chat space&#x201D; to show their interactional ability.</p>
<p>The data was collected in a quiet room which is commonly used by these participants. They were requested to sit comfortably close to each other, with the video camera placed about 1 m away on a tripod in front of them, and a voice recorder placed behind them. The conversational interactions among participants were recorded by the current first researcher utilizing &#x201C;non-participant observations&#x201D; (<xref ref-type="bibr" rid="B7">Davies, 2007</xref>, p. 174). That is, the researcher does not participate in the discussion, but instead sits in a corner hidden from the participants&#x2019; view, to observe their behaviors without interference. In the study of multimodal interaction, facial expressions, gestures, head movements, and body movements of the participants are all important information. Therefore, their upper bodies were mainly captured by the closed-set-up video camera. Furthermore, the sound was also recorded by the high-quality voice recorder used to conduct prosody analysis. For each recording, the researcher started the video recording, checked the audio, and then sat. The participants could freely begin and end their conversations without a time limit. In all, the duration time of each conversation varied from 8 to 23 min, and the total communication time was 295 min, with 31,759 words.</p>
</sec>
<sec id="S3.SS3">
<title>Single-Case Analysis and Multimodal Conversation Analysis</title>
<p>A single-case analysis means &#x201C;the techniques of seeing significant interactional detail in the ongoing production of singular sequences of talk-in-interaction&#x201D; (<xref ref-type="bibr" rid="B18">Hutchby and Wooffitt, 2008</xref>, p. 113). Its goal is to explain a single complex phenomenon of interest. This approach can be enhanced further by combining with multimodal CA (<xref ref-type="bibr" rid="B37">Mondada, 2018</xref>). As <xref ref-type="bibr" rid="B37">Mondada (2018</xref>, p. 86) puts it, multimodal CA pays &#x201C;careful and precise attention [&#x2026;] to temporally and sequentially organized details of actions that account for how co-participants orient to each other&#x2019;s conduct and assemble it in meaningful ways, moment by moment.&#x201D; Multimodal CA allows analysts to detailly identify a range of interactional resources that interactants utilize and organize to achieve communicative goals in the extended sequences of talk, from participant-relevant emic and multimodal perspectives. Overall, the combined approach can help us to enrich our understanding of the learners&#x2019; irregular self-selection deeply and detailly to some extent, especially the interplay of verbal and non-verbal resources in specific interactional contexts.</p>
</sec>
<sec id="S3.SS4">
<title>Data Analysis</title>
<p>To address the research questions, four analytical procedures were followed:</p>
<list list-type="simple">
<list-item>
<label>(1)</label>
<p>The simplified Jeffersonian convention (<xref ref-type="bibr" rid="B20">Jefferson, 2004</xref>) was used to transcribe verbal behavior in the data by the first author with help of Transcriber software (<xref ref-type="bibr" rid="B4">Boudahmane et al., 2022</xref>) (see <xref ref-type="app" rid="A1">Appendix</xref>). To improve the accuracy and reliability of the transcribed data, the two researchers worked together to check it over. Then, we conducted a line-by-line analysis, in a larger sequence closely examining what and when the participants said. We narrowed the focus down to sequences in which the participants implemented irregular self-selection to speak next and gathered a collection of all such turns (i.e., number).</p>
</list-item>
<list-item>
<label>(2)</label>
<p>With reference to the two aforementioned circumstances of irregular self-selection, we carefully analyzed and coded each case within the collection. This round of analysis yielded findings of the types. We counted the frequency of each type and recorded it in an Excel sheet. Adhering to the purposes of self-selection classified by <xref ref-type="bibr" rid="B56">Waring (2011)</xref> and <xref ref-type="bibr" rid="B11">Garton (2012)</xref>, we conducted another round of analysis and coding of the successful irregular self-selection, guided by the question: &#x201C;Why that now?&#x201D; (<xref ref-type="bibr" rid="B49">Schegloff and Sacks, 1973</xref>, p. 299). Through classification, comparison, and modification, this round of analysis yielded findings of the purposes. We counted the frequency of each purpose and recorded it in the Excel sheet. During the process of analysis, the inter-rater agreement of the two researchers was over 80%.</p>
</list-item>
<list-item>
<label>(3)</label>
<p>To further explore the issues of when, how, and why, the single-case analysis combined with multimodal CA was used to carefully examine three excerpts of irregular self-selection representing the types, respectively. A slightly modified version of <xref ref-type="bibr" rid="B36">Mondada&#x2019;s (2016)</xref> annotation was used to describe the embodied actions within interactions (see <xref ref-type="app" rid="A1">Appendix</xref>). Details about participants&#x2019; body movements were noted on a separate line above the verbal line in the transcript. Screenshots were also used to show the participants&#x2019; body movements capturing the moment of when and how, and their specific occurrences were noted with a &#x201C;#&#x201D; in the transcript.</p>
</list-item>
<list-item>
<label>(4)</label>
<p>To analyze the prosodic features of irregular self-selection, particularly focused on pitch and intensity, we used <italic>Praat</italic> software, a combination of auditory and acoustic analysis (<xref ref-type="bibr" rid="B3">Boersma and Weenink, 2021</xref>). The form of spectrogram, waveforms, pitch traces, and intensity traces were all analyzed by this software. They are the compelling evidence that self-selector claims for turn space (<xref ref-type="bibr" rid="B48">Schegloff, 2000</xref>).</p>
</list-item>
</list>
</sec>
</sec>
<sec id="S4" sec-type="results">
<title>Results</title>
<p>In this part, we first presented the overall descriptive statistics of number, type, and purpose of irregular self-selection. Then, three representative excerpts were analyzed in detail using multimodal CA to further address the issues of when, how, and why.</p>
<sec id="S4.SS1">
<title>Descriptive Statistics</title>
<sec id="S4.SS1.SSS1">
<title>Number</title>
<p>Through repeated line-by-line observation and analysis of the transcribed data, this study yielded 152 cases of irregular self-selection. As <xref ref-type="table" rid="T1">Table 1</xref> showed, in a total of 20 groups, seventeen groups contained irregular self-selection, ranging from 1 to 38 cases in each conversation. According to the number of cases, the participation model of seventeen groups could be characterized as conventional, active, and highly active. Among them, five groups contained irregular self-selection in less than 5 cases, indicating that turn-taking devices in these groups were more conventional because they tended to obey turn-taking rules and used less irregular self-selection devices to take turns. The groups containing cases in 5&#x2013;9 were the most, including eight groups. These groups were active because more interruptions or overlaps occurred in their conversations, indicating that the participants more actively joined in the discussion. Finally, four groups had cases over 10, and the most were 38 cases, indicating the highly active participation of the learners in the interaction. This means that hearers in these groups were more eager to obtain the speakership. The uneven distribution of irregular self-selection among different groups reflected the group or individual discrepancy of participation in interaction. In summary, seeing from the overall data, the Chinese postgraduate EFL learners were relatively active in participating in peer conversations.</p>
<table-wrap position="float" id="T1">
<label>TABLE 1</label>
<caption><p>Number of irregular self-selection in each group.</p></caption>
<table cellspacing="5" cellpadding="5" frame="hsides" rules="groups">
<thead>
<tr>
<td valign="top" align="left">Group (G)</td>
<td valign="top" align="center">G1</td>
<td valign="top" align="center">G2</td>
<td valign="top" align="center">G3</td>
<td valign="top" align="center">G4</td>
<td valign="top" align="center">G5</td>
<td valign="top" align="center">G6</td>
<td valign="top" align="center">G7</td>
<td valign="top" align="center">G8</td>
<td valign="top" align="center">G9</td>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Number</td>
<td valign="top" align="center">9</td>
<td valign="top" align="center">18</td>
<td valign="top" align="center">38</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">6</td>
<td valign="top" align="center">2</td>
<td valign="top" align="center">12</td>
<td valign="top" align="center">1</td>
<td valign="top" align="center">7</td>
</tr>
<tr>
<td valign="top" colspan="10"><hr/></td>
</tr>
<tr>
<td valign="top" align="left"><bold>Group (G)</bold></td>
<td valign="top" align="center"><bold>G10</bold></td>
<td valign="top" align="center"><bold>G11</bold></td>
<td valign="top" align="center"><bold>G12</bold></td>
<td valign="top" align="center"><bold>G13</bold></td>
<td valign="top" align="center"><bold>G14</bold></td>
<td valign="top" align="center"><bold>G15</bold></td>
<td valign="top" align="center"><bold>G16</bold></td>
<td valign="top" align="center"><bold>G17</bold></td>
<td/>
</tr>
<tr>
<td valign="top" colspan="10"><hr/></td>
</tr>
<tr>
<td valign="top" align="left">Number</td>
<td valign="top" align="center">5</td>
<td valign="top" align="center">9</td>
<td valign="top" align="center">8</td>
<td valign="top" align="center">3</td>
<td valign="top" align="center">5</td>
<td valign="top" align="center">6</td>
<td valign="top" align="center">4</td>
<td valign="top" align="center">18</td>
<td/>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
<sec id="S4.SS1.SSS2">
<title>Type</title>
<p>According to the two aforementioned circumstances of irregular self-selection in the literature review, we divided 152 cases into four types (see <xref ref-type="table" rid="T2">Table 2</xref>). The frequency of each type was reported in <xref ref-type="table" rid="T3">Table 3</xref>.</p>
<table-wrap position="float" id="T2">
<label>TABLE 2</label>
<caption><p>Types of irregular self-selection.</p></caption>
<table cellspacing="5" cellpadding="5" frame="hsides" rules="groups">
<thead>
<tr>
<td valign="top" align="left">Type</td>
<td valign="top" align="left">Meaning</td>
<td valign="top" align="left">Example</td>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Turn interruption (TI)</td>
<td valign="top" align="left">The current speaker&#x2019;s turn is at a non-TRP, signified by a uncomplete TCU, the hearer interrupts to obtain the speakership. (C1).</td>
<td valign="top" align="left">A: &#x201C;Where are you come from, that&#x2019;s&#x201D;<break/>B: &#x201C;Or what&#x2019;s your name&#x201D;</td>
</tr>
<tr>
<td valign="top" align="left">Turn competition (TC)</td>
<td valign="top" align="left">The current speaker&#x2019;s turn reaches the TRP. Then, the two participants speak simultaneously to compete for speakership. It is the hearer who wins the competition for the turn space here. This kind of situation is a result of application of rule 2 and rule 3 as described in C2.</td>
<td valign="top" align="left">A: &#x201C;So have you some uh did you have some maybe some hum example teacher [in your]&#x201D;<break/>B: &#x201C;[uh]You know yes here is one.&#x201D;</td>
</tr>
<tr>
<td valign="top" align="left">Turn holding abortion (THA)</td>
<td valign="top" align="left">The current speaker&#x2019;s turn reaches TRP. However, the speaker has no intention of giving up the speakership by using non-lexical words, such as hum/uh/mm. At this time, the hearer chooses to speak to abort this turn holding process to obtain speakership. (C1).</td>
<td valign="top" align="left">A: &#x201C;I can hold parties many times hum&#x201D;<break/>B: &#x201C;There must be a garden in your house&#x201D;</td>
</tr>
<tr>
<td valign="top" align="left">Self-selection failed (TF)</td>
<td valign="top" align="left">The hearer fails to gain the speakership when he/she self-selects. (C1 or C2).</td>
<td valign="top" align="left">A: &#x201C;I hope all of us can&#x201D;<break/>B: &#x201C;Can&#x201D;<break/>A: &#x201C;Find a Mr. Right.&#x201D;</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<fn><p><italic>All the examples in <xref ref-type="table" rid="T2">Tables 2</xref>, <xref ref-type="table" rid="T4">4</xref> are real cases that occurred in the participants&#x2019; conversations.</italic></p></fn>
</table-wrap-foot>
</table-wrap>
<table-wrap position="float" id="T3">
<label>TABLE 3</label>
<caption><p>Frequency of types.</p></caption>
<table cellspacing="5" cellpadding="5" frame="hsides" rules="groups">
<thead>
<tr>
<td valign="top" align="left">Type</td>
<td valign="top" align="center">TI</td>
<td valign="top" align="center">TC</td>
<td valign="top" align="center">THA</td>
<td valign="top" align="center">TF</td>
<td valign="top" align="center">Total</td>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Number</td>
<td valign="top" align="center">96</td>
<td valign="top" align="center">37</td>
<td valign="top" align="center">10</td>
<td valign="top" align="center">9</td>
<td valign="top" align="center">152</td>
</tr>
<tr>
<td valign="top" align="left">Percentage</td>
<td valign="top" align="center">63.4%</td>
<td valign="top" align="center">24.2%</td>
<td valign="top" align="center">6.5%</td>
<td valign="top" align="center">5.9%</td>
<td valign="top" align="center">100%</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>As shown in <xref ref-type="table" rid="T2">Table 2</xref>, successful irregular self-selections mainly occurred in three interactional contexts: (1) when the speaker&#x2019;s turn was at a non-TRP, the hearer interrupted to obtain the speakership; (2) the two participants competed for the speakership; and (3) when the speaker was holding the current turn, the hearer chose to speak to abort this turn holding process, and then gain the speakership. The detailed multimodal process of those irregular self-selections will be analyzed in the subsequent section.</p>
<p><xref ref-type="table" rid="T3">Table 3</xref> revealed various types of irregular self-selection. The most frequent type was TI, consisting of 96 cases (63.4%) of all the irregular self-selection. TC came next, with 37 cases (24.2%). THA came third, with only 10 cases (6.5%). Because of the small percentage of the last type, TF, where hearer failed to gain the speakership, we excluded them from the analysis.</p>
</sec>
<sec id="S4.SS1.SSS3">
<title>Purpose</title>
<p>With reference to the classifications of <xref ref-type="bibr" rid="B56">Waring (2011)</xref> and <xref ref-type="bibr" rid="B11">Garton (2012)</xref>, and combining them with the actual cases that occurred in the present data, we divided the purposes of irregular self-selection into six types, as shown in <xref ref-type="table" rid="T4">Table 4</xref>. The frequency of each type was reported in <xref ref-type="table" rid="T5">Table 5</xref>.</p>
<table-wrap position="float" id="T4">
<label>TABLE 4</label>
<caption><p>Purposes of irregular self-selection.</p></caption>
<table cellspacing="5" cellpadding="5" frame="hsides" rules="groups">
<thead>
<tr>
<td valign="top" align="left">Purpose</td>
<td valign="top" align="left">Meaning</td>
<td valign="top" align="left">Example</td>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Aiding</td>
<td valign="top" align="left">To help the current speaker when he/she faces some expression difficulties (e.g., disfluency or pause)</td>
<td valign="top" align="left">A: &#x201C;Hum, I think the challenges means more. (2.3)&#x201D;<break/>B: &#x201C;Means more chances&#x201D;</td>
</tr>
<tr>
<td valign="top" align="left">Cooperative completion</td>
<td valign="top" align="left">To cooperatively complete the following turn content with the current speaker based on contextual information</td>
<td valign="top" align="left">A: &#x201C;The character acted by the by [Zhou Dong Yu]&#x201D;<break/>B: &#x201C;[Zhou Dong Yu]&#x201D;</td>
</tr>
<tr>
<td valign="top" align="left">Displaying knowledge</td>
<td valign="top" align="left">To display own thoughts and views or providing comments and new information contributing to the topical development</td>
<td valign="top" align="left">A: &#x201C;You can also listen some informal materials such as the Allen Show or Friends [this]&#x201D;<break/>B: &#x201C;[Yeah] they are popular.&#x201D;</td>
</tr>
<tr>
<td valign="top" align="left">Agreement</td>
<td valign="top" align="left">To express agreement and support of the current speaker&#x2019;s speech</td>
<td valign="top" align="left">A: &#x201C;He usually she usually do some small punishment to us uh then&#x201D;<break/>B: &#x201C;Yes I agree with you&#x201D;</td>
</tr>
<tr>
<td valign="top" align="left">Clarification</td>
<td valign="top" align="left">To request the current speaker to clarify some vague information in the previous turn in order to reach a mutual understanding, usually by using some lexical bundles like <italic>you mean X</italic></td>
<td valign="top" align="left">A: &#x201C;So he (0.5) [pro-]&#x201D;<break/>B: &#x201C;[You mean] leave her family a big fortune?&#x201D;</td>
</tr>
<tr>
<td valign="top" align="left">Information request</td>
<td valign="top" align="left">To elicit further information that relates to the ongoing topic based on the previous utterance/sequence, usually by using interrogative sentence</td>
<td valign="top" align="left">A: &#x201C;We waited a very long time, very very very long, so [in that]&#x201D;<break/>B: &#x201C;[Is in] midnight?&#x201D;</td>
</tr>
</tbody>
</table>
</table-wrap>
<table-wrap position="float" id="T5">
<label>TABLE 5</label>
<caption><p>Frequency of purposes.</p></caption>
<table cellspacing="5" cellpadding="5" frame="hsides" rules="groups">
<thead>
<tr>
<td valign="top" align="left">Purpose</td>
<td valign="top" align="center">Aiding</td>
<td valign="top" align="center">Cooperative completion</td>
<td valign="top" align="center">Displaying knowledge</td>
<td valign="top" align="center">Agreement</td>
<td valign="top" align="center">Clarification</td>
<td valign="top" align="center">Information request</td>
<td valign="top" align="center">Total</td>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Number</td>
<td valign="top" align="center">18</td>
<td valign="top" align="center">11</td>
<td valign="top" align="center">76</td>
<td valign="top" align="center">15</td>
<td valign="top" align="center">7</td>
<td valign="top" align="center">16</td>
<td valign="top" align="center">143</td>
</tr>
<tr>
<td valign="top" align="left">Percentage</td>
<td valign="top" align="center">12.6%</td>
<td valign="top" align="center">7.70%</td>
<td valign="top" align="center">53.1%</td>
<td valign="top" align="center">10.5%</td>
<td valign="top" align="center">4.90%</td>
<td valign="top" align="center">11.2%</td>
<td valign="top" align="center">100%</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>In accordance with <xref ref-type="table" rid="T4">Table 4</xref>, we calculated the frequency of each type in participants&#x2019; conversations, and the result was reported in <xref ref-type="table" rid="T5">Table 5</xref>.</p>
<p>As can be seen in <xref ref-type="table" rid="T5">Table 5</xref>, the most important purpose of irregular self-selection was to display knowledge, which had 76 cases (53.1%). It indicated that irregular self-selection was used by the hearer to show understanding of and interest in what had been said, and then displayed their own thoughts and views or provided comments and new information. All of them can contribute to topical development. The purposes of aiding and cooperative completion had 18 cases (12.6%) and 11 cases (7.7%), respectively. They both indicated high participation of the hearer in the conversation to help or cooperate with the speaker to complete their current turn. The information request purpose had 16 cases (11.2%). By asking a question, the hearer attempted to show interest in what has been said and tried to elicit further information relating to the ongoing topic from the speaker. The agreement purpose had 15 cases (10.5%), which showed the hearer&#x2019;s attentive listening and understanding of what had been said. The last one, clarification purpose, only had seven cases (4.9%), indicating the negotiation of information between participants to achieve mutual understanding.</p>
<p>Having shown the overall number, type, and purpose of irregular self-selection, the next section will analyze three representative excerpts of successful irregular self-selection in detail to further illustrate when, how, and why the self-selectors implement them.</p>
</sec>
</sec>
<sec id="S4.SS2">
<title>Case Analysis of Irregular Self-Selection Using Multimodal Conversation Analysis Method</title>
<p>The focal episodes of the analysis centered on the following three types: turn interruption (TI), turn competition (TC), and turn holding abortion (THA). In what follows, by focusing on three representative episodes in as much detailed as possible, we showcased when, how, and why self-selectors accomplished self-selection in the peer-to-peer conversation by using the multimodal CA method.</p>
<sec id="S4.SS2.SSS1">
<title>Turn Interruption</title>
<p>This subsection shows the analysis of the first type of irregular self-selection, that is, TI, one of the practices frequently used by self-selectors. TI means that the hearer implements self-selection when the speaker&#x2019;s turn is at non-TRP. More specifically, when the speaker is still in the state of event narrating or storytelling, the obvious sign is that the TCU is incomplete. However, at this time, the hearer self-selects to speak, resulting in a TCU being interrupted before it has reached a point of possible completion, as shown in excerpt 1.</p>
<list list-type="simple">
<list-item><p>Excerpt 1. Rubbish Sorting</p>
</list-item>
<list-item><p>01 W U:m I have heard that like Beijing and Shanghai</p>
</list-item>
<list-item><p>02 um they have put forward the (.) project</p>
</list-item>
<list-item><p>Hand<sub>L</sub> -.-.-.-.-.</p>
</list-item>
<list-item><p>Torso<sub>L</sub> F&#x2014;-</p>
</list-item>
<list-item><p>03 like #hum</p>
</list-item>
</list>
<list list-type="simple">
<list-item><p>Gaze <underline>mutual gaze</underline></p>
</list-item>
<list-item><p>Hand<sub>L</sub> <sup>&#x002A;&#x002A;&#x002A;&#x002A;</sup></p>
</list-item>
<list-item><p>Torso<sub>L</sub> H&#x2014;-</p>
</list-item>
<list-item><p>04 L #Ah</p>
</list-item>
<list-item><p>05 W [rubbish classification]</p>
</list-item>
<list-item><p>06 L [rubbish classification]</p>
</list-item>
<list-item><p>07 W yeah</p>
</list-item>
<list-item><p>08 L yeah</p>
</list-item>
</list>
<list list-type="simple">
<list-item><p>a. When</p>
</list-item>
</list>
<p>In lines 01&#x2013;02, W describes the garbage management scheme proposed by the environmental institutions in Beijing and Shanghai. She makes a concrete elaboration of the program in line 03. Grammatically, W uses preposition &#x201C;<italic>like</italic>,&#x201D; but due to the lack of object, it does not constitute the complete preposition phrase; Prosodically (<xref ref-type="fig" rid="F1">Figure 1</xref>), as shown in pitch trace, the word &#x201C;<italic>like</italic>&#x201D; is in flat intonation, which generally indicates the maintaining of the turn and shows that this turn has not ended (<xref ref-type="bibr" rid="B10">Ford and Thompson, 1996</xref>); Pragmatically, W does not elaborate the specific scheme, and thus the statement is incomplete. Non-verbally, W&#x2019;s gesture is still away from the &#x201C;home position&#x201D; (<xref ref-type="bibr" rid="B45">Sacks and Schegloff, 2002</xref>), and she does not gaze at the recipient L at the end of her turn (<xref ref-type="fig" rid="F2">Figure 2</xref>). The above multimodal resources show that W&#x2019;s current turn is not complete in terms of syntax, intonation, pragmatic behavior, and body movement. So, the turn does not reach TRP at the current moment, implying that turn-taking does not occur here.</p>
<fig id="F1" position="float">
<label>FIGURE 1</label>
<caption><p>Spectrogram, waveform, pitch trace (dotted line), and intensity trace (solid line) of lines 02&#x2013;08 in excerpt 1.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-788438-g001.tif"/>
</fig>
<fig id="F2" position="float">
<label>FIGURE 2</label>
<caption><p>Bodily movement of W in line 03.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-788438-g002.tif"/>
</fig>
<p>However, hearer L gets the thought W wants to express based on the contextual information, and then she makes her own choice to interrupt L&#x2019;s speech and self-select to gain the speakership (lines 04&#x2013;06).</p>
<list list-type="simple">
<list-item><p>b. How</p>
</list-item>
</list>
<p>(1) Grammatically, L uses an exclamatory word plus noun phrase, which is &#x201C;<italic>Ah</italic>&#x201D; and &#x201C;<italic>rubbish classification</italic>&#x201D; as self-selection words. First, she produces the non-lexical token &#x201C;<italic>Ah.</italic>&#x201D; On the one hand, &#x201C;<italic>Ah</italic>&#x201D; is used to show a change of state indicating she has known what W wants to say. On the other hand, it is used to attract W&#x2019;s attention and show her interest and attentive listening to the current conversation. Moreover, <italic>Ah</italic> can be seen as a kind of self-selection signal showing her follow-up participation. Then, the noun phrase &#x201C;<italic>rubbish classification</italic>&#x201D; overlaps with W&#x2019;s words, that is, they speak it simultaneously; (2) Prosodically (<xref ref-type="fig" rid="F1">Figure 1</xref>), as analyzed by Praat, L makes a pitch reset and an intensity enhancement. First, through pitch reset, the pitch of self-selection words becomes higher. As can be seen in pitch trace that the peak pitch at &#x201C;<italic>Ah</italic>&#x201D; is 270 Hz, which is higher than the pitch at the previous word &#x201C;<italic>hum</italic>&#x201D; (207 Hz). Second, through the intensity enhancement, the speech loudness increases; namely, the sound becomes louder. Intensity trace suggests that the peak intensity at &#x201C;<italic>Ah</italic>&#x201D; is 75 dB, which is louder than that at <italic>hum</italic> (50 dB). Finally, the color of the spectrogram at &#x201C;<italic>Ah</italic>&#x201D; becomes darker and acoustic amplitude becomes larger. They indicate that the energy value at this place becomes larger, which supports the above findings. (3) Non-verbally, L uses gaze, gesture, body posture, and head movement when she prepares for self-selection. Her constructions of complex multimodal &#x201C;gestalt&#x201D; (<xref ref-type="bibr" rid="B30">Li, 2014</xref>, p. 7) are assembled simultaneously. Specifically, L and W at mutual gaze status, her gesture gradually deviates from home position (<xref ref-type="fig" rid="F3">Figure 3</xref>), and upper body posture and head position change from leaning forward to relaxed position (<xref ref-type="fig" rid="F2">Figures 2</xref>, <xref ref-type="fig" rid="F3">3</xref>). The changes in body posture and head position, simultaneously with the self-selection signal &#x201C;<italic>Ah</italic>,&#x201D; indicate that she has recognized the thought W wants to express. Overall, L uses lexical phrases, pitch reset, intensity enhancement, and non-verbal resources including gaze, gesture, body posture, and head movement to project and achieve self-selection.</p>
<fig id="F3" position="float">
<label>FIGURE 3</label>
<caption><p>Bodily movement of L in self-selection in line 04.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-788438-g003.tif"/>
</fig>
<list list-type="simple">
<list-item><p>c. Why</p>
</list-item>
</list>
<p>As can be seen in conversation sequences (lines 01&#x2013;06), L&#x2019;s self-selection purpose is to cooperate with W to complete the specific rubbish management scheme mentioned in the previous turn, namely rubbish classification (lines 05&#x2013;06). After this, L&#x2019;s self-selection words get W&#x2019;s approval by using the acknowledgment token &#x201C;<italic>yeah</italic>&#x201D; (line 07). In all, the self-selection does not cause trouble to the current conversation, rather than show the interests and active participation of L. It also demonstrates her interactional sensitivity, turn-monitoring awareness, and participation ability in the conversation. Moreover, the overlapped noun phrase &#x201C;<italic>rubbish classification</italic>&#x201D; is recognitional overlap, which is a type of overlap that occurs when a potential next speaker recognizes the &#x201C;thrust or upshot&#x201D; of the prior talk (<xref ref-type="bibr" rid="B23">Konakahara, 2015</xref>, p. 39). It is usually considered legitimate or non-intrusive within the turn-taking system.</p>
</sec>
<sec id="S4.SS2.SSS2">
<title>Turn Competition</title>
<p>The second type of self-selection is TC, causing overlap. It occurs as a result of the hearer&#x2019;s application of rule 2 at a possible TRP (i.e., self-selection), simultaneously occurring with the speaker&#x2019;s application of rule 3 (i.e., the current speaker&#x2019;s continuation). Moreover, sometimes it occurs with the speaker&#x2019;s multimodal resources &#x201C;divergence&#x201D; from each other (<xref ref-type="bibr" rid="B30">Li, 2014</xref>, p. 205). That syntax, prosody, pragmatic behavior, and gaze resources project the end of the turn, indicating the occurrence of TRP and turn-taking. However, divergence from turn-ending projection, gesture projects the turn holding, implying that the current speaker is not ready to transfer the turn, as shown in excerpt 2.</p>
<list list-type="simple">
<list-item><p>Excerpt 2. School Bullying</p>
</list-item>
<list-item><p>01 J I also think hum our country should carry out some relevant laws</p>
</list-item>
<list-item><p>02 to about this kind hum event</p>
</list-item>
<list-item><p>03 J so that our teenagers and even (.)</p>
</list-item>
<list-item><p>Gaze<sub>J</sub> <underline><sub>away | at</sub></underline></p>
</list-item>
<list-item><p>Hand<sub>J</sub> <sup>&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;</sup></p>
</list-item>
<list-item><p>04 college students can be protected hum #well(0.3)</p>
</list-item>
</list>
<list list-type="simple">
<list-item><p>Gaze <underline>mutual gaze</underline></p>
</list-item>
<list-item><p>Hand<sub>J</sub> <sup>&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;</sup>| -.-.-.</p>
</list-item>
<list-item><p>05 Y [yes I think #so]</p>
</list-item>
<list-item><p>Hand<sub>J</sub> <sup>&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;&#x002A;</sup>| -.-.-.-.&#x2013;.-.-.-.-.-.&#x2013;.-.-.-.-.-.-.</p>
</list-item>
<list-item><p>06 J [that&#x2019;s my #per&#x00B0;sonal&#x00B0;] understanding</p>
</list-item>
<list-item><p>07 Y and I think the students should stand up</p>
</list-item>
<list-item><p>08 to stop this phenomenon</p>
</list-item>
</list>
<list list-type="simple">
<list-item><p>a. When</p>
</list-item>
</list>
<p>In lines 01&#x2013;04, J states that the Chinese government must promulgate laws and policies to protect teenagers from school bullying, including college students. Grammatically, it is an adverbial clause directed by &#x201C;<italic>so that</italic>,&#x201D; with complete syntactic structure; Prosodically, the sentence can be judged to be in falling intonation by combining with listening and discrimination. In addition, a pause of 0.3 s following &#x201C;<italic>well</italic>&#x201D; suggests that turn-taking may occur (<xref ref-type="bibr" rid="B10">Ford and Thompson, 1996</xref>); Pragmatically, J&#x2019;s declarative statement is complete and expresses his point of view; Non-verbally, J looks at Y at the end of the turn (the word &#x201C;well&#x201D;), forming a mutual gaze with Yang at the same time (<xref ref-type="fig" rid="F4">Figure 4</xref>), which projects possible end of the turn (<xref ref-type="bibr" rid="B22">Kendon, 1967</xref>). In all, the above four multimodal resources indicate that J&#x2019;s turn may end and arrive at TRP at this moment and turn-taking may occur.</p>
<fig id="F4" position="float">
<label>FIGURE 4</label>
<caption><p>Bodily movements of Y and J in lines 04&#x2013;05.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-788438-g004.tif"/>
</fig>
<p>However, J&#x2019;s gesture does not return to the home position but still keeps the &#x201C;open hand palm up (OHPU)&#x201D; (<xref ref-type="bibr" rid="B30">Li, 2014</xref>, p. 219; <xref ref-type="fig" rid="F4">Figure 4</xref>), implying turn holding. It can be seen that the gesture resource of J is in conflict with the aforementioned four kinds of multimodal resources. All the other resources indicate turn completion and occurrence of turn-taking, while gesture indicates turn holding; that is, J still wants to talk and has no intention of giving up the speaking turn. A noteworthy observation is that J&#x2019;s OHPU gesture in the current turn lasts about 15 s in the video, almost throughout his whole turn. It means from the perspective of J&#x2019;s gesture using habit, he prefers it. Previous studies have shown that the OHPU gesture usually appears at the possible end of a turn, indicating the yielding of that turn (<xref ref-type="bibr" rid="B51">Streeck, 2009</xref>). However, in this case, the OHPU gesture is J&#x2019;s habitual practice; hence, it does not indicate the end of the turn but implies the holding of this turn. The intention can also be seen in J&#x2019;s following actions. In lines 05&#x2013;06, J&#x2019;s turn overlaps with Y&#x2019;s. They compete for the speakership, but Y wins the competition for the turning space here. In addition, J returns his gesture to the home position at the word &#x201C;<italic>so</italic>&#x201D; and <italic>&#x201C;personal</italic>&#x201D; in lines 05 and 06 (<xref ref-type="fig" rid="F5">Figure 5</xref>) and completes the turn of himself after overlapping resolution (line 06). Thus, it shows that J wants to continue his turn, and only the return of gesture indicates the possible end of the turn (<xref ref-type="bibr" rid="B30">Li, 2014</xref>). This excerpt reflects whether noticing the individual discrepancy is an essential factor in deciding self-selection time. It is more unpredictable and needs the hearer to monitor the turn momentarily. Just as in this case, Y ignores the diverging of gesture and thus implements self-selection to show her understanding of and interests in what has been said by J.</p>
<list list-type="simple">
<list-item><p>b. How</p>
</list-item>
</list>
<fig id="F5" position="float">
<label>FIGURE 5</label>
<caption><p>Gesture of J after vertical bar in lines 05&#x2013;06.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-788438-g005.tif"/>
</fig>
<p>(1) Grammatically, Y uses the acknowledgment token &#x201C;<italic>yes</italic>&#x201D; to express her approval of J&#x2019;s statement, followed by the sentence &#x201C;<italic>I think so.</italic>&#x201D; (2) Prosodically (<xref ref-type="fig" rid="F6">Figure 6</xref>), the software Praat shows that Y makes a pitch reset but hardly makes an intensity enhancement. First, pitch trace shows that the pitch at the end of J&#x2019;s turn is about 124 Hz, while the pitch at <italic>yes</italic> is 176 Hz (300 min 124 Hz), which is 52 Hz higher than that at the end of J&#x2019;s turn. Second, as can be seen in the intensity trace, the intensity at <italic>yes</italic> is about 27 dB (72 min 45 dB), which is lower than 45 dB at the end of J&#x2019;s turn. Last, the color of the spectrogram at <italic>yes</italic> becomes darker, and the amplitude of the sound wave in the acoustic map becomes larger. They indicate that the energy value at <italic>yes</italic> becomes larger, which supports the occurrence of the overlapped speech of Y and J here. (3) Non-verbally, Y implements self-selection while at a mutual gaze state with J (<xref ref-type="fig" rid="F4">Figure 4</xref>). In all, Y uses lexis, syntax, pitch reset, and gaze to achieve self-selection.</p>
<list list-type="simple">
<list-item><p>c. Why</p>
</list-item>
</list>
<fig id="F6" position="float">
<label>FIGURE 6</label>
<caption><p>Spectrogram, waveform, pitch trace (dotted line), and intensity trace (solid line) in lines 03&#x2013;05 in excerpt 2.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-788438-g006.tif"/>
</fig>
<p>The conversation sequence (lines 03&#x2013;04) shows that the purpose of Y&#x2019;s self-selection is to express her agreement with J&#x2019;s statement by using the word &#x201C;<italic>yes</italic>&#x201D; and the sentence &#x201C;<italic>I think so.</italic>&#x201D; It shows her active participation in the current topic.</p>
</sec>
<sec id="S4.SS2.SSS3">
<title>Turn Holding Abortion</title>
<p>The usage frequency of the third type of irregular self-selection is less than the aforementioned two ones, but it is still a way for the hearer to gain the speakership. This type is THA, wherein when a speaker uses some multimodal resources to indicate the continuity of speakership, the hearer implements self-selection to obtain the speakership. The continuation of talk is represented by the use of the non-lexical word &#x201C;<italic>hum</italic>,&#x201D; gaze shift, and the holding of &#x201C;thinking face&#x201D; (<xref ref-type="bibr" rid="B13">Goodwin and Goodwin, 1986</xref>, p. 57). As illustrated in excerpt 3.</p>
<list list-type="simple">
<list-item><p>Excerpt 3. House</p>
</list-item>
<list-item><p>01 M I will buy hum a very very big house</p>
</list-item>
<list-item><p>02 and I can hold (.) many parties.</p>
</list-item>
<list-item><p>Gaze<sub>M</sub> <underline>away</underline></p>
</list-item>
<list-item><p>03 #hum</p>
</list-item>
</list>
<list list-type="simple">
<list-item><p>Gaze <underline>mutual gaze</underline></p>
</list-item>
<list-item><p>04 N #there must be a garden in your (.) home(.)</p>
</list-item>
<list-item><p>05 M yes</p>
</list-item>
</list>
<list list-type="simple">
<list-item><p>a. When</p>
</list-item>
</list>
<p>In lines 01&#x2013;02, M states that she wants to buy a big house in the future so that many parties can be held in it. Grammatically, this sentence is a compound sentence combined by a connective <italic>and</italic>, with complete syntactic structure. Prosodically, the point number shows that the second part of the sentence is in falling intonation, indicating the possible end of the sentence (<xref ref-type="bibr" rid="B10">Ford and Thompson, 1996</xref>). Pragmatically, M&#x2019;s declarative statement is complete, and she finishes her opinion of the house. At this time, her turn is complete in grammar, prosodic, and pragmatic behavior. The above multimodal resources indicate a high possibility of turn completion and turn-taking. However, instead of ending the turn here, M uses several turn holding strategies to indicate the continuity of the current turn, including non-lexical word &#x201C;<italic>hum</italic>,&#x201D; gaze shift, and holding of a thinking face (<xref ref-type="fig" rid="F7">Figure 7</xref>). To facilitate the back-and-forth flow of a natural conversation, N participates in the conversation actively to self-select and to gain the speakership, which shows her turn controlling awareness (line 04). Moreover, since the decision-making of self-selection lies with the hearer, the timing of it is related to participation state, knowledge, or emotional status (<xref ref-type="bibr" rid="B17">Heritage, 2013</xref>).</p>
<list list-type="simple">
<list-item><p>b. How</p>
</list-item>
</list>
<fig id="F7" position="float">
<label>FIGURE 7</label>
<caption><p>Gaze of M in line 03.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-788438-g007.tif"/>
</fig>
<p>(1) Grammatically, N uses the existential sentence guided by <italic>&#x201C;there be.</italic>&#x201D; (2) Prosodically, in <xref ref-type="fig" rid="F8">Figure 8</xref>, the pitch trace shows that N does not reset the pitch but maintains the same pitch range just as M has (about 250 Hz), so the pitch changes slightly. However, N enhances the intensity of the words &#x201C;<italic>there must</italic>&#x201D; to draw the attention of speaker M. In <xref ref-type="fig" rid="F8">Figure 8</xref>, the peak value of the intensity at &#x201C;<italic>there must</italic>&#x201D; is 71 dB, higher than that in <italic>hum</italic> at the end of the M&#x2019;s turn (60 dB). In addition, the color of the spectrum at &#x201C;<italic>there must</italic>&#x201D; becomes deeper and the amplitude of the sound wave in the acoustic map becomes larger. They indicate the increment of the energy value, which are proofs of the above findings; (3) Non-verbally, N and M form mutual gaze when she implements self-selection (<xref ref-type="fig" rid="F9">Figure 9</xref>). In all, N uses syntax, intensity reset, and mutual gaze to achieve self-selection.</p>
<list list-type="simple">
<list-item><p>c. Why</p>
</list-item>
</list>
<fig id="F8" position="float">
<label>FIGURE 8</label>
<caption><p>Spectrogram, waveform, pitch trace (dotted line), and intensity trace (solid line) in lines 03&#x2013;05 in excerpt 3.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-788438-g008.tif"/>
</fig>
<fig id="F9" position="float">
<label>FIGURE 9</label>
<caption><p>Bodily movement of N in self-selection in line 04.</p></caption>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fpsyg-13-788438-g009.tif"/>
</fig>
<p>Conversation sequence (lines 01&#x2013;04) reveals that the self-selection sentence &#x201C;<italic>There must be a garden in your house</italic>&#x201D; is used for displaying knowledge and functions as a supplement of new information to M&#x2019;s speaking content. It can promote topical development (line 04). M then acknowledges this with acknowledgment token <italic>yes</italic> (line 05).</p>
</sec>
</sec>
</sec>
<sec id="S5" sec-type="discussion">
<title>Discussion</title>
<p>As an important turn-taking method and conversation monitoring strategy, the implementation of irregular self-selection reflects certain interactional features of Chinese postgraduate EFL learners. We explore it by providing overall descriptive statistics of this action and then conducting a detailed analysis (i.e., multimodal CA) of three representative examples of irregular self-selection from when, how, and why. The results are of great significance to enrich the existing research on turn-taking practices of EFL learners from the perspective of multimodal interaction.</p>
<p>First of all, regarding when to self-selection, as shown in the number of irregular self-selections of each group, in a total of twenty groups, seventeen groups contain irregular self-selections, varying from 1 to 38 cases in each conversation. The participation mode of seventeen groups can be characterized as conventional (five groups), active (eight groups), and highly active (four groups). To summarize, 85% of groups contain irregular self-selections, and 60% of groups are active in implementing such actions to participate in the conversation. The results are in line with the evidence that EFL learners are interactionally competent members to participate in a conversation (see <xref ref-type="bibr" rid="B5">Carroll, 2004</xref>; <xref ref-type="bibr" rid="B9">Firth and Wagner, 2007</xref>; <xref ref-type="bibr" rid="B27">Lee, 2017</xref>; <xref ref-type="bibr" rid="B24">Konakahara, 2020</xref>). For example, <xref ref-type="bibr" rid="B27">Lee (2017)</xref> found that learners actively interrupted the ongoing talk to move into the primary speaker position. They are able to achieve certain communicative goals despite their limited proficiency in the target language. Just as <xref ref-type="bibr" rid="B5">Carroll (2004)</xref> observed that Japanese novice speakers of English used similar ways to self-select as those of native speakers of English.</p>
<p>Based on an analysis of the 152 cases, we find three types of successful irregular self-selection in learners&#x2019; interactions: TI, TC, and THA. The TI occupies the most of them (63.4%), reflecting that Chinese postgraduate EFL learners are used to interrupt to obtain the speakership in conversations. <xref ref-type="bibr" rid="B41">Park and Duey (2020)</xref> also found that in multi-party workplace meetings, the self-selector interrupted to gain the speakership. This indicates that interruption is an important device for both native and non-native speakers to participate in conversations. The TC occupies 24.2%, which is the second most frequent way of irregular self-selection mentioned as &#x201C;floor-taking overlap&#x201D; and studied by <xref ref-type="bibr" rid="B23">Konakahara (2015</xref>, <xref ref-type="bibr" rid="B24">2020)</xref> in casual ELF conversations. The result reveals that turn competition is also a way of active involvement in non-native speakers&#x2019; conversation. The third one is THA, which only occupies 6.5%, but it also implies that learners are eager to participate in interaction by aborting the holding of speakers&#x2019; turn. They aim to claim the speakership and boost conversational development.</p>
<p>Referring to the underlying reasons for the initiation of irregular self-selection, we think it may be because learners in peer interaction are naturally in an equal position; thus, they will initiate and participate in a conversation more actively. Moreover, when hearers have relevant conversational knowledge, they will initiate irregular self-selection in different interactional contexts to show their &#x201C;knowledge status&#x201D; (<xref ref-type="bibr" rid="B17">Heritage, 2013</xref>, p. 376) and willingness to express opinions concerning a certain domain of knowledge. The findings of this study show that after a long period of English learning, about 10 years, Chinese postgraduate EFL learners have got certain self-selection capabilities and turn controlling awareness to obtain speakership, as shown in the aforementioned three cases. However, the initiation of irregular self-selection is also based on the flow of conversation, being context-dependent and unpredictable. On the one hand, the initiation of it is context-dependent (<xref ref-type="bibr" rid="B29">Lerner, 2003</xref>), which requires both participants to construct the context together. They work to promote the flow of conversation and to provide a self-selection context at the same time. For example, in excerpt 1, L can implement irregular self-selection because of the existence of a &#x201C;rubbish sorting&#x201D; context. On the other hand, the initiation of this kind of action is unpredictable, and the decision-making is in the hands of the hearer. For example, hearers&#x2019; participation state, knowledge state, emotional state, and individual differences will all affect their implementation of self-selection actions. Thus, only when learners have a certain awareness of turn monitoring can they initiate this action in a conversation.</p>
<p>Second, regarding how to self-select, by detailed analysis of the three cases, we find that (1) in lexical and syntactic dimension (i.e., verbal aspect), learners can provide appropriate language resources to participate in conversation according to the flow of conversation. (2) in prosodic dimension (i.e., vocal aspect), pitch reset and intensity enhancement are used by learners to different degrees. Some of them use both ways to increase the volume of their self-selection words, for example in excerpt 1, while others only use pitch reset or intensity enhancement to achieve self-selection, as shown in excerpts 2 and 3. (3) in non-verbal dimension or aspect, except in excerpt 1, L uses four kinds of non-verbal resources in her irregular self-selection, including gaze, gesture, head movement, and body posture. The other two hearers only use gaze to predict and achieve self-selection. From the use of multimodal resources, it can be seen that the learners in these cases use at least three kinds of resources to implement irregular self-selection, which shows their ability to use multimodal resources to some extent.</p>
<p>However, through the overall investigation of irregular self-selection that occurred in our data, we find that about 80% of the self-selectors only use one or two kinds of body movements to project or implement irregular self-selection, such as gaze, gesture, or head movement. However, other body movement resources rarely occur, such as facial expression and body posture. <xref ref-type="bibr" rid="B27">Lee (2017)</xref> found that learners utilized an ensemble of talk, gaze, gesture, and bodily orientation to gain the speakership. <xref ref-type="bibr" rid="B24">Konakahara (2020)</xref> also found that in overlap sequences, interactants collaboratively exploited multiple non-verbal resources, such as gaze, posture, and gesture, for organizing turn-taking and conveying meaning. Compared with these two findings, the overall modal complexity and diversity of Chinese postgraduate EFL learners are low, causing their behaviors to be restrained and inactive. This phenomenon may be related to the Chinese culture emphasizing introversion and restraint of conversation participation. It reflects that culture has a profound influence on one&#x2019;s behavior, even when they use other languages to communicate and participate in the interaction. However, some studies show that in Mandarin Chinese talk-in-interaction, participants will use plenty of multimodal resources to take turns or manage their affiliation (<xref ref-type="bibr" rid="B58">Yang, 2007</xref>, <xref ref-type="bibr" rid="B60">2011</xref>). For example, <xref ref-type="bibr" rid="B58">Yang (2007)</xref> found that Chinese speakers used non-verbal resources to manage turns, such as hand drop, gaze, non-gaze, touch, thinking face, and finger count. The result was not the same as found in the present study. Maybe another possible reason for the low diversity of body movements in this study is that learners are aware that they and their conversations are being recorded. Thus, they cannot behave naturally when using the English language to talk and tend to control and restrain their behaviors to some extent.</p>
<p>Last, regarding why to self-selection, irregular self-selection can be divided into six types: displaying knowledge (53.1%), aiding (12.6%), information request (11.2%), agreement (10.5%), cooperative completion (7.7%), and clarification (4.9%). It can be seen that the main purpose of irregular self-selection is to display knowledge, also mentioned as one of the self-selection purposes by <xref ref-type="bibr" rid="B56">Waring (2011)</xref>. The result indicates that the hearer contributes to topical development by displaying his/her own thoughts and views or providing comments and new information (for example, in excerpt 3). Then, the purposes of aiding, agreement, and cooperative completion occupy 29.4%, used to support the current speaker in the meaning-making process. They also help to maintain the rhythm or pace of the conversation by showing listenership, understanding, active participation, and agreement (see <xref ref-type="bibr" rid="B38">Murata, 1994</xref>; <xref ref-type="bibr" rid="B28">Lerner, 2002</xref>) (for example in excerpts 1 and 2). Finally, the purposes of information request and clarification occupy 16.1%, used to interact with the speaker of vague information and to elicit further information. They also serve to show high interactional sensitivity and active participation (<xref ref-type="bibr" rid="B24">Konakahara, 2020</xref>).</p>
<p>Although the purposes of irregular self-selection are various, showing different communication intentions of the learner, the common characteristic of them is that they reveal the learners&#x2019; active involvement in interaction (<xref ref-type="bibr" rid="B6">Cogo and Dewey, 2012</xref>), topical development, and interactive sensitivity of conversation. Moreover, with irregular self-selection, the participants cooperatively move the talk forward, reflecting their cooperative communication intention. <xref ref-type="bibr" rid="B23">Konakahara (2015)</xref> obtained the same finding in casual ELF conversations of the overlapping questions. Consequently, non-native speakers are successful in &#x201C;achieving mutual understanding and developing interpersonal relationships&#x201D; (<xref ref-type="bibr" rid="B23">Konakahara, 2015</xref>, p. 37).</p>
</sec>
<sec id="S6" sec-type="conclusion">
<title>Conclusion</title>
<p>This study investigated when, how, and why Chinese postgraduate EFL learners implement irregular self-selection from the multimodal interaction perspective. By providing descriptive statistics and using a multimodal conversation analytic approach to examine three excerpts in detail, the results show that learners are interactionally competent members to participate in the conversation. They are able to achieve communicative goals, but their body movements lack diversity as compared with other non-native English speakers, causing behaviors to be constrained and inactive.</p>
<p>Based on the findings, this study provides some implications for EFL learners, especially other East Asian EFL learners who are commonly characterized as silent, reserved, and inactive during discussions, particularly in the classroom. This study shows that EFL learners with high language proficiency will benefit from peer interaction to develop their interactional competence, as evidenced by the initiation of irregular self-selection and active involvement in participation. Thus, in oral English learning and teaching, more high-level peer interaction without teacher involvement should be carried out. Although irregular self-selection violates the turn-taking system, it is harmless to the topical development. Therefore, learners should be encouraged to use this kind of turn-taking way to participate in the conversation, making their interaction more natural and vivid. However, when participating in interactions, EFL learners need to pay much attention to the use of multimodal resources, especially a variety of body movements, such as facial expression, gesture, head movement, and body posture. The use of these resources can improve the diversity of body movements and enhance interactional ability with native or other non-native English speakers.</p>
<p>This study adds to the scarce research on EFL learners&#x2019; irregular turn-taking practices and the growing literature on the use of multimodal resources in their interactions. At the same time, it verifies the applicability of the multimodal CA approach to the studies of learners&#x2019; conversation again. It is of significance in the detailed investigation of the learners&#x2019; turn management and their embodied participation in the conversation. It helps to understand the visible processes through which learners positively claim the speakership to participate in the conversation and build a cooperative relationship. It has also provided new empirical evidence to confirm the fact that EFL learners are interactionally competent members to successfully participate in the interaction, although with limited proficiency in the target language.</p>
<p>Despite its significance, the potential limitation of a single-case analysis is that it can only be representative of the analyzed phenomenon. To gain a richer and more comprehensive understanding of the phenomenon of interest, more investigations are needed. Our study also suggests directions for future research. Although topic discussion is one of the most efficient and natural ways to collect participants&#x2019; interactional data, it would be beneficial for future research to investigate irregular self-selection in varied tasks, for instance, role-play games, jigsaw puzzles, quiz games, and so on. Moreover, as the conversations of the present study were collected between friends, it is worth exploring whether the observation also applies to conversations between participants who are not familiar with each other or participants of unequal power relations.</p>
</sec>
<sec id="S7" sec-type="data-availability">
<title>Data Availability Statement</title>
<p>The original contributions presented in the study are included in the article/<xref ref-type="supplementary-material" rid="DS1">Supplementary Material</xref>, further inquiries can be directed to the corresponding author/s.</p>
</sec>
<sec id="S8">
<title>Ethics Statement</title>
<p>The studies involving human participants were reviewed and approved by the Professor Committee of School of Foreign Languages, Northeast Normal University. The patients/participants provided their written informed consent to participate in this study. Written informed consent was obtained from the individual(s) for the publication of any potentially identifiable images or data included in this article.</p>
</sec>
<sec id="S9">
<title>Author Contributions</title>
<p>MMJ contributed to the conception and design of the study, data collection, data analysis and interpretation, writing, and developing the manuscript. HPZ was responsible for data analysis and interpretation, manuscript development, writing, and editing. Both authors contributed to the article and approved the submitted version.</p>
</sec>
<sec id="conf1" sec-type="COI-statement">
<title>Conflict of Interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec id="pudiscl1" sec-type="disclaimer">
<title>Publisher&#x2019;s Note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
</body>
<back>
<sec id="S10" sec-type="funding-information">
<title>Funding</title>
<p>This research was funded by the National Office for Philosophy and Social Sciences of China (Fund No. 20BYY209), the Office for Philosophy and Social Sciences of Jilin Province (Fund No. 2020B214), and the Youth Team Foundation of Northeast Normal University (Fund No. 2021QT004).</p>
</sec>
<sec id="S11" sec-type="supplementary-material">
<title>Supplementary Material</title>
<p>The Supplementary Material for this article can be found online at: <ext-link ext-link-type="uri" xlink:href="https://www.frontiersin.org/articles/10.3389/fpsyg.2022.788438/full#supplementary-material">https://www.frontiersin.org/articles/10.3389/fpsyg.2022.788438/full#supplementary-material</ext-link></p>
<supplementary-material xlink:href="Table_1.xlsx" id="TS1" mimetype="application/vnd.openxmlformats-officedocument.spreadsheetml.sheet" xmlns:xlink="http://www.w3.org/1999/xlink"/>
<supplementary-material xlink:href="Table_2.xlsx" id="TS2" mimetype="application/vnd.openxmlformats-officedocument.spreadsheetml.sheet" xmlns:xlink="http://www.w3.org/1999/xlink"/>
<supplementary-material xlink:href="Data_Sheet_1.zip" id="DS1" mimetype="application/zip" xmlns:xlink="http://www.w3.org/1999/xlink"/>
<supplementary-material xlink:href="Data_Sheet_2.docx" id="DS2" mimetype="application/vnd.openxmlformats-officedocument.wordprocessingml.document" xmlns:xlink="http://www.w3.org/1999/xlink"/>
</sec>
<app-group>
<app id="A1">
<title>Appendix</title>
<sec id="S12.SS1">
<title>Transcription Conventions</title>
<list list-type="simple">
<list-item><p>(.) micro pause (0.3) pauses of 0.3 s [] overlap</p>
</list-item>
<list-item><p>: prolongation or stretching of the sound = latching</p>
</list-item>
<list-item><p>&#x00B0;&#x00B0; the word is markedly quiet or soft . falling intonation</p>
</list-item>
<list-item><p><underline>away</underline> gaze away <underline>at</underline> gaze at</p>
</list-item>
<list-item><p>&#x002A; stroke of gesticulation -. recovery of gesticulation</p>
</list-item>
<list-item><p>F forward movement H home position</p>
</list-item>
<list-item><p>----- close dashes indicate the holding of the body movements</p>
</list-item>
</list>
</sec>
</app>
</app-group>
<ref-list>
<title>References</title>
<ref id="B1"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Almoaily</surname> <given-names>M.</given-names></name></person-group> (<year>2020</year>). <article-title>Impact of age and gender on frequency of interruption in dyadic interviews.</article-title> <source><italic>Interact. Stud.</italic></source> <volume>21</volume> <fpage>187</fpage>&#x2013;<lpage>199</lpage>. <pub-id pub-id-type="doi">10.1075/is.17011.alm</pub-id> <pub-id pub-id-type="pmid">33486653</pub-id></citation></ref>
<ref id="B2"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Auer</surname> <given-names>P.</given-names></name></person-group> (<year>2021</year>). <article-title>Turn-allocation and gaze: a multimodal revision of the &#x201C;current-speaker selects-next&#x201D; rule of the turn-taking system of conversation analysis.</article-title> <source><italic>Discourse Stud.</italic></source> <volume>23</volume> <fpage>117</fpage>&#x2013;<lpage>140</lpage>. <pub-id pub-id-type="doi">10.1177/1461445620966922</pub-id></citation></ref>
<ref id="B3"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Boersma</surname> <given-names>P.</given-names></name> <name><surname>Weenink</surname> <given-names>D.</given-names></name></person-group> (<year>2021</year>). <source><italic>Praat: Doing Phonetics by Computer.</italic></source> Available online at: <ext-link ext-link-type="uri" xlink:href="http://www.praat.org/">http://www.praat.org/</ext-link> <comment>(accessed September 14, 2021)</comment>.</citation></ref>
<ref id="B4"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Boudahmane</surname> <given-names>K.</given-names></name> <name><surname>Manta</surname> <given-names>M.</given-names></name> <name><surname>Antoine</surname> <given-names>F.</given-names></name> <name><surname>Galliano</surname> <given-names>S.</given-names></name> <name><surname>Barras</surname> <given-names>C.</given-names></name></person-group> (<year>2022</year>). <source><italic>Transcriber: a Tool for Segmenting, Labeling and Transcribing Speech.</italic></source> Available online at: <ext-link ext-link-type="uri" xlink:href="http://trans.sourceforge.net/">http://trans.sourceforge.net/</ext-link> <comment>(accessed March 11, 2022)</comment>.</citation></ref>
<ref id="B5"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Carroll</surname> <given-names>D.</given-names></name></person-group> (<year>2004</year>). &#x201C;<article-title>Restarts in novice turn beginnings: disfluencies or interactional achievements?</article-title>,&#x201D; in <source><italic>Second Language Conversations</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>Gardner</surname> <given-names>R.</given-names></name> <name><surname>Wagner</surname> <given-names>J.</given-names></name></person-group> (<publisher-loc>New York, NY</publisher-loc>: <publisher-name>Continuum</publisher-name>), <fpage>201</fpage>&#x2013;<lpage>220</lpage>.</citation></ref>
<ref id="B6"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cogo</surname> <given-names>A.</given-names></name> <name><surname>Dewey</surname> <given-names>M.</given-names></name></person-group> (<year>2012</year>). <source><italic>Analysing English as a Lingua Franca: Corpus-Driven Investigation.</italic></source> <publisher-loc>London</publisher-loc>: <publisher-name>Continuum</publisher-name>.</citation></ref>
<ref id="B7"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Davies</surname> <given-names>M. B.</given-names></name></person-group> (<year>2007</year>). <source><italic>Doing a Successful Research Project: Using Qualitative or Quantitative Methods.</italic></source> <publisher-loc>New York, NY</publisher-loc>: <publisher-name>Palgrave Macmillan</publisher-name>.</citation></ref>
<ref id="B8"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dressel</surname> <given-names>D.</given-names></name></person-group> (<year>2020</year>). <article-title>Multimodal word searches in collaborative storytelling: on the local mobilization and negotiation of coparticipation.</article-title> <source><italic>J. Pragmat</italic></source> <volume>170</volume> <fpage>37</fpage>&#x2013;<lpage>54</lpage>. <pub-id pub-id-type="doi">10.1016/j.pragma.2020.08.010</pub-id></citation></ref>
<ref id="B9"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Firth</surname> <given-names>A.</given-names></name> <name><surname>Wagner</surname> <given-names>J.</given-names></name></person-group> (<year>2007</year>). <article-title>Second/foreign language learning as a social accomplishment: elaborations on a reconceptualized SLA.</article-title> <source><italic>Mod. Lang. J.</italic></source> <volume>91</volume> <fpage>800</fpage>&#x2013;<lpage>819</lpage>. <pub-id pub-id-type="doi">10.1111/j.1540-4781.2007.00670.x</pub-id></citation></ref>
<ref id="B10"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ford</surname> <given-names>C.</given-names></name> <name><surname>Thompson</surname> <given-names>S.</given-names></name></person-group> (<year>1996</year>). &#x201C;<article-title>Interactional units in conversation: syntactic, intonational, and pragmatic resources for the management of turns</article-title>,&#x201D; in <source><italic>Interaction and Grammar</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>Ochs</surname> <given-names>E.</given-names></name> <name><surname>Schegloff</surname> <given-names>E.</given-names></name> <name><surname>Thompson</surname> <given-names>S.</given-names></name></person-group> (<publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>), <fpage>134</fpage>&#x2013;<lpage>184</lpage>.</citation></ref>
<ref id="B11"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Garton</surname> <given-names>S.</given-names></name></person-group> (<year>2012</year>). <article-title>Speaking out of turn? Taking the initiative in teacher-fronted classroom interaction.</article-title> <source><italic>Classroom Discourse</italic></source> <volume>3</volume> <fpage>29</fpage>&#x2013;<lpage>45</lpage>. <pub-id pub-id-type="doi">10.1080/19463014.2012.666022</pub-id></citation></ref>
<ref id="B12"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Goodwin</surname> <given-names>C.</given-names></name></person-group> (<year>1980</year>). <article-title>Restarts, pauses, and the achievement of a state of mutual gaze at turn-beginning.</article-title> <source><italic>Socio. Inq.</italic></source> <volume>50</volume> <fpage>272</fpage>&#x2013;<lpage>302</lpage>. <pub-id pub-id-type="doi">10.1111/j.1475-682x.1980.tb00023.x</pub-id></citation></ref>
<ref id="B13"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Goodwin</surname> <given-names>M. H.</given-names></name> <name><surname>Goodwin</surname> <given-names>C.</given-names></name></person-group> (<year>1986</year>). <article-title>Gesture and co-participation in the activity of searching for a word.</article-title> <source><italic>Semiotica</italic></source> <volume>62</volume> <fpage>51</fpage>&#x2013;<lpage>75</lpage>.</citation></ref>
<ref id="B14"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Greer</surname> <given-names>T.</given-names></name> <name><surname>Potter</surname> <given-names>H.</given-names></name></person-group> (<year>2008</year>). <article-title>Turn-taking practices in multi-party EFL oral proficiency tests.</article-title> <source><italic>J. Appl. Linguist.</italic></source> <volume>5</volume> <fpage>297</fpage>&#x2013;<lpage>320</lpage>. <pub-id pub-id-type="doi">10.1558/japl.v5i2.297</pub-id></citation></ref>
<ref id="B15"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Guillot</surname> <given-names>M. N.</given-names></name></person-group> (<year>2009</year>). <article-title>Interruption in advanced learner French: issues of pragmatic discrimination.</article-title> <source><italic>Lang. Contrast.</italic></source> <volume>9</volume> <fpage>98</fpage>&#x2013;<lpage>123</lpage>. <pub-id pub-id-type="doi">10.1075/lic.9.1.06gui</pub-id> <pub-id pub-id-type="pmid">33486653</pub-id></citation></ref>
<ref id="B16"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Guillot</surname> <given-names>M. N.</given-names></name></person-group> (<year>2012</year>). <article-title>Conversational management and pragmatic discrimination in foreign talk: overlap in advanced L2 French.</article-title> <source><italic>Intercult. Pragmat.</italic></source> <volume>9</volume> <fpage>307</fpage>&#x2013;<lpage>333</lpage>. <pub-id pub-id-type="doi">10.1515/ip-2012-0019</pub-id></citation></ref>
<ref id="B17"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Heritage</surname> <given-names>J.</given-names></name></person-group> (<year>2013</year>). &#x201C;<article-title>Epistemics in conversation</article-title>,&#x201D; in <source><italic>The Handbook of Conversation Analysis</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>Sidnell</surname> <given-names>J.</given-names></name> <name><surname>Stivers</surname> <given-names>T.</given-names></name></person-group> (<publisher-loc>London</publisher-loc>: <publisher-name>Wiley-Blackwell</publisher-name>), <fpage>370</fpage>&#x2013;<lpage>394</lpage>.</citation></ref>
<ref id="B18"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hutchby</surname> <given-names>I.</given-names></name> <name><surname>Wooffitt</surname> <given-names>R.</given-names></name></person-group> (<year>2008</year>). <source><italic>Conversation Analysis</italic></source>, <edition>2nd Edn</edition>. <publisher-loc>Cambridge</publisher-loc>: <publisher-name>Polity</publisher-name>.</citation></ref>
<ref id="B19"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Iwasaki</surname> <given-names>S.</given-names></name></person-group> (<year>2009</year>). <article-title>Initiating interactive turn spaces in Japanese conversation: local projection and collaborative action.</article-title> <source><italic>Discourse Process.</italic></source> <volume>46</volume> <fpage>226</fpage>&#x2013;<lpage>246</lpage>. <pub-id pub-id-type="doi">10.1080/01638530902728918</pub-id></citation></ref>
<ref id="B20"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jefferson</surname> <given-names>G.</given-names></name></person-group> (<year>2004</year>). &#x201C;<article-title>Glossary of transcript symbols with an introduction</article-title>,&#x201D; in <source><italic>Conversation Analysis: Studies from the First Generation</italic></source>, <role>ed.</role> <person-group person-group-type="editor"><name><surname>Lerner</surname> <given-names>G.</given-names></name></person-group> (<publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins</publisher-name>), <fpage>13</fpage>&#x2013;<lpage>25</lpage>.</citation></ref>
<ref id="B21"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kasper</surname> <given-names>G.</given-names></name> <name><surname>Wagner</surname> <given-names>J.</given-names></name></person-group> (<year>2011</year>). &#x201C;<article-title>A conversation-analytic approach to second language acquisition</article-title>,&#x201D; in <source><italic>Alternative Approaches in Second Language Acquisition</italic></source>, <role>ed.</role> <person-group person-group-type="editor"><name><surname>Atkinson</surname> <given-names>D.</given-names></name></person-group> (<publisher-loc>New York, NY</publisher-loc>: <publisher-name>Routledge</publisher-name>), <fpage>117</fpage>&#x2013;<lpage>142</lpage>.</citation></ref>
<ref id="B22"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kendon</surname> <given-names>A.</given-names></name></person-group> (<year>1967</year>). <article-title>Some functions of gaze direction in social interaction.</article-title> <source><italic>Acta Psychol.</italic></source> <volume>26</volume> <fpage>22</fpage>&#x2013;<lpage>63</lpage>. <pub-id pub-id-type="doi">10.1016/0001-6918(67)90005-4</pub-id> <pub-id pub-id-type="pmid">6043092</pub-id></citation></ref>
<ref id="B23"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Konakahara</surname> <given-names>M.</given-names></name></person-group> (<year>2015</year>). <article-title>An analysis overlapping questions in casual ELF conversation: cooperative or competitive contribution.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>84</volume> <fpage>37</fpage>&#x2013;<lpage>53</lpage>. <pub-id pub-id-type="doi">10.1016/j.pragma.2015.04.014</pub-id></citation></ref>
<ref id="B24"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Konakahara</surname> <given-names>M.</given-names></name></person-group> (<year>2020</year>). <article-title>Single case analyses of two overlap sequences in casual ELF conversations from a multimodal perspective: toward the consideration of mutual benefits of ELF and CA.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>170</volume> <fpage>301</fpage>&#x2013;<lpage>316</lpage>. <pub-id pub-id-type="doi">10.1016/j.pragma.2020.09.024</pub-id></citation></ref>
<ref id="B25"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lauzon</surname> <given-names>V. F.</given-names></name> <name><surname>Berger</surname> <given-names>E.</given-names></name></person-group> (<year>2015</year>). <article-title>The multimodal organization of speaker selection in classroom interaction.</article-title> <source><italic>Linguist. Educ.</italic></source> <volume>31</volume> <fpage>14</fpage>&#x2013;<lpage>29</lpage>. <pub-id pub-id-type="doi">10.1016/j.linged.2015.05.001</pub-id></citation></ref>
<ref id="B26"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lazaraton</surname> <given-names>A.</given-names></name></person-group> (<year>2002</year>). <source><italic>A Qualitative Approach to the Validation of Oral Proficiency Tests.</italic></source> <publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>.</citation></ref>
<ref id="B27"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lee</surname> <given-names>J.</given-names></name></person-group> (<year>2017</year>). <article-title>Multimodal turn allocation in ESL peer group discussions.</article-title> <source><italic>Social Semiotics</italic></source> <volume>27</volume> <fpage>671</fpage>&#x2013;<lpage>692</lpage>. <pub-id pub-id-type="doi">10.1080/10350330.2016.1207353</pub-id></citation></ref>
<ref id="B28"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lerner</surname> <given-names>G. H.</given-names></name></person-group> (<year>2002</year>). &#x201C;<article-title>Turn-sharing: the choral co-production of talk-in-interaction</article-title>,&#x201D; in <source><italic>The Language of Turn and Sequence</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>Ford</surname> <given-names>C. E.</given-names></name> <name><surname>Fox</surname> <given-names>B. A.</given-names></name> <name><surname>Thompson</surname> <given-names>S. A.</given-names></name></person-group> (<publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>), <fpage>225</fpage>&#x2013;<lpage>256</lpage>.</citation></ref>
<ref id="B29"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lerner</surname> <given-names>G. H.</given-names></name></person-group> (<year>2003</year>). <article-title>Selecting next speaker: the context-sensitive operation of a context-free organization.</article-title> <source><italic>Lang. Soc.</italic></source> <volume>32</volume> <fpage>177</fpage>&#x2013;<lpage>201</lpage>. <pub-id pub-id-type="doi">10.1017/S004740450332202X</pub-id></citation></ref>
<ref id="B30"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Li</surname> <given-names>X. T.</given-names></name></person-group> (<year>2014</year>). <source><italic>Multimodality, Interaction and Turn-taking in Mandarin Conversation.</italic></source> <publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins</publisher-name>.</citation></ref>
<ref id="B31"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Li</surname> <given-names>X. T.</given-names></name></person-group> (<year>2019</year>). &#x201C;<article-title>Negotiating activity closings with reciprocal head nods in mandarin conversation</article-title>,&#x201D; in <source><italic>Embodied Activities in Face-to-face and Mediated Settings</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>Reber</surname> <given-names>E.</given-names></name> <name><surname>Gerhardt</surname> <given-names>C.</given-names></name></person-group> (<publisher-loc>London</publisher-loc>: <publisher-name>Palgrave Macmillan</publisher-name>), <fpage>369</fpage>&#x2013;<lpage>396</lpage>. <pub-id pub-id-type="doi">10.1007/978-3-319-97325-8_11</pub-id></citation></ref>
<ref id="B32"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Majlesi</surname> <given-names>A. R.</given-names></name> <name><surname>Markee</surname> <given-names>N.</given-names></name></person-group> (<year>2018</year>). <article-title>Multimodality in second language talk: the impact of video analysis on SLA research.</article-title> <source><italic>Tartu Semiotics Library</italic></source> <volume>19</volume> <fpage>247</fpage>&#x2013;<lpage>260</lpage>.</citation></ref>
<ref id="B33"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Markaki</surname> <given-names>V.</given-names></name> <name><surname>Mondada</surname> <given-names>L.</given-names></name></person-group> (<year>2012</year>). <article-title>Embodied orientations towards Co-participants in multinational meetings.</article-title> <source><italic>Discourse Stud.</italic></source> <volume>14</volume> <fpage>31</fpage>&#x2013;<lpage>52</lpage>. <pub-id pub-id-type="doi">10.1177/1461445611427210</pub-id></citation></ref>
<ref id="B34"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mondada</surname> <given-names>L.</given-names></name></person-group> (<year>2007</year>). <article-title>Multimodal resources or turn-taking: pointing and the emergence of possible next speakers.</article-title> <source><italic>Discourse Stud.</italic></source> <volume>9</volume> <fpage>194</fpage>&#x2013;<lpage>225</lpage>. <pub-id pub-id-type="doi">10.1177/1461445607075346</pub-id></citation></ref>
<ref id="B35"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mondada</surname> <given-names>L.</given-names></name></person-group> (<year>2013</year>). <article-title>Embodied and spatial resources for turn-taking in institutional multiparty interactions: participatory democracy debates.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>46</volume> <fpage>39</fpage>&#x2013;<lpage>68</lpage>. <pub-id pub-id-type="doi">10.1016/j.pragma.2012.03.010</pub-id></citation></ref>
<ref id="B36"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mondada</surname> <given-names>L.</given-names></name></person-group> (<year>2016</year>). <article-title>Challenges of multimodality: language and the body in social interaction.</article-title> <source><italic>J. Sociolinguist.</italic></source> <volume>20</volume> <fpage>336</fpage>&#x2013;<lpage>366</lpage>. <pub-id pub-id-type="doi">10.1111/josl.1_12177</pub-id></citation></ref>
<ref id="B37"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mondada</surname> <given-names>L.</given-names></name></person-group> (<year>2018</year>). <article-title>Multiple temporalities of language and body in interaction: challenges for transcribing multimodality.</article-title> <source><italic>Res. Lang. Soc. Interact.</italic></source> <volume>51</volume> <fpage>85</fpage>&#x2013;<lpage>106</lpage>. <pub-id pub-id-type="doi">10.1080/08351813.2018.1413878</pub-id></citation></ref>
<ref id="B38"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Murata</surname> <given-names>K.</given-names></name></person-group> (<year>1994</year>). <article-title>Intrusive or co-operative? A cross-cultural study of interruption.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>21</volume> <fpage>385</fpage>&#x2013;<lpage>400</lpage>. <pub-id pub-id-type="doi">10.1016/0378-2166(94)90011-6</pub-id></citation></ref>
<ref id="B39"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Olsher</surname> <given-names>D.</given-names></name></person-group> (<year>2004</year>). &#x201C;<article-title>Talk and gesture: the embodied completion of sequential actions in spoken interaction</article-title>,&#x201D; in <source><italic>Second Language Conversations</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>Gardner</surname> <given-names>R.</given-names></name> <name><surname>Wagner</surname> <given-names>J.</given-names></name></person-group> (<publisher-loc>New York, NY</publisher-loc>: <publisher-name>Continuum</publisher-name>), <fpage>221</fpage>&#x2013;<lpage>245</lpage>.</citation></ref>
<ref id="B40"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Orletti</surname> <given-names>F.</given-names></name></person-group> (<year>1981</year>). &#x201C;<article-title>Classroom verbal interaction: a conversational analysis</article-title>,&#x201D; in <source><italic>Possibilities and Limitations of Pragmatics</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>Par-ret</surname> <given-names>H.</given-names></name> <name><surname>Sbisa</surname> <given-names>M.</given-names></name> <name><surname>Verschueren</surname> <given-names>J.</given-names></name></person-group> (<publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins</publisher-name>), <fpage>531</fpage>&#x2013;<lpage>549</lpage>.</citation></ref>
<ref id="B41"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Park</surname> <given-names>I.</given-names></name> <name><surname>Duey</surname> <given-names>M.</given-names></name></person-group> (<year>2020</year>). <article-title>I&#x2019;m sorry (to interrupt): the use of explicit apology in turn-taking.</article-title> <source><italic>App. Linguist. Rev.</italic></source> <volume>11</volume> <fpage>377</fpage>&#x2013;<lpage>401</lpage>. <pub-id pub-id-type="doi">10.1515/applirev-2018-0017</pub-id></citation></ref>
<ref id="B42"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Pietikainen</surname> <given-names>K. S.</given-names></name></person-group> (<year>2018</year>). <article-title>Misunderstandings and ensuring understanding in private ELF talk.</article-title> <source><italic>Appl. Linguist.</italic></source> <volume>39</volume> <fpage>188</fpage>&#x2013;<lpage>212</lpage>. <pub-id pub-id-type="doi">10.1093/applin/amw005</pub-id></citation></ref>
<ref id="B43"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Richard</surname> <given-names>J. C.</given-names></name> <name><surname>Nunan</surname> <given-names>D.</given-names></name></person-group> (<year>1990</year>). <source><italic>Second Language Teacher Education.</italic></source> <publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>.</citation></ref>
<ref id="B44"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Rossano</surname> <given-names>F.</given-names></name></person-group> (<year>2012</year>). &#x201C;<article-title>Gaze in conversation</article-title>,&#x201D; in <source><italic>Handbook of Conversation Analysis</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>Sidnell</surname> <given-names>J.</given-names></name> <name><surname>Stivers</surname> <given-names>T.</given-names></name></person-group> (<publisher-loc>West Sussex</publisher-loc>: <publisher-name>Wiley-Blackwell</publisher-name>), <fpage>308</fpage>&#x2013;<lpage>329</lpage>. <pub-id pub-id-type="doi">10.16910/jemr.14.1.1</pub-id> <pub-id pub-id-type="pmid">34122746</pub-id></citation></ref>
<ref id="B45"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sacks</surname> <given-names>H.</given-names></name> <name><surname>Schegloff</surname> <given-names>E. A.</given-names></name></person-group> (<year>2002</year>). <article-title>Home position.</article-title> <source><italic>Gesture.</italic></source> <volume>2</volume> <fpage>133</fpage>&#x2013;<lpage>146</lpage>. <pub-id pub-id-type="doi">10.1075/gest.2.2.02sac</pub-id> <pub-id pub-id-type="pmid">33486653</pub-id></citation></ref>
<ref id="B46"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sacks</surname> <given-names>H.</given-names></name> <name><surname>Schegloff</surname> <given-names>E. A.</given-names></name> <name><surname>Jefferson</surname> <given-names>G.</given-names></name></person-group> (<year>1974</year>). <article-title>A simplest systematics for the organization of turn-taking for conversation.</article-title> <source><italic>Lang. Soc.</italic></source> <volume>50</volume> <fpage>696</fpage>&#x2013;<lpage>735</lpage>.</citation></ref>
<ref id="B47"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sahlstr&#x00F6;m</surname> <given-names>F.</given-names></name></person-group> (<year>2002</year>). <article-title>The interactional organization of hand raising in classroom interaction.</article-title> <source><italic>J. Classroom Interact.</italic></source> <volume>37</volume> <fpage>47</fpage>&#x2013;<lpage>57</lpage>.</citation></ref>
<ref id="B48"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Schegloff</surname> <given-names>E. A.</given-names></name></person-group> (<year>2000</year>). <article-title>Overlapping talk and the organization of turn-taking for conversation.</article-title> <source><italic>Lang. Soc.</italic></source> <volume>29</volume> <fpage>1</fpage>&#x2013;<lpage>63</lpage>. <pub-id pub-id-type="doi">10.1017/S0047404500001019</pub-id></citation></ref>
<ref id="B49"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Schegloff</surname> <given-names>E. A.</given-names></name> <name><surname>Sacks</surname> <given-names>H.</given-names></name></person-group> (<year>1973</year>). <article-title>Opening up closings.</article-title> <source><italic>Semiotica</italic></source> <volume>8</volume> <fpage>289</fpage>&#x2013;<lpage>327</lpage>. <pub-id pub-id-type="doi">10.1515/semi.1973.8.4.289</pub-id></citation></ref>
<ref id="B50"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Stivers</surname> <given-names>T.</given-names></name> <name><surname>Rossano</surname> <given-names>F.</given-names></name></person-group> (<year>2010</year>). <article-title>Mobilizing response.</article-title> <source><italic>Res. Lang. Soc. Interact.</italic></source> <volume>43</volume> <fpage>3</fpage>&#x2013;<lpage>31</lpage>. <pub-id pub-id-type="doi">10.1080/08351810903471258</pub-id></citation></ref>
<ref id="B51"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Streeck</surname> <given-names>J.</given-names></name></person-group> (<year>2009</year>). <source><italic>Gesturecraft: The Manu-Facture of Meaning.</italic></source> <publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins</publisher-name>.</citation></ref>
<ref id="B52"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Streeck</surname> <given-names>J.</given-names></name> <name><surname>Goodwin</surname> <given-names>C.</given-names></name> <name><surname>Lebaron</surname> <given-names>C.</given-names></name></person-group> (<year>2011</year>). &#x201C;<article-title>Embodied Interaction in the Material World: an Introduction</article-title>,&#x201D; in <source><italic>Embodied Interaction: Language and Body in the Material World</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>Streeck</surname> <given-names>J.</given-names></name> <name><surname>Goodwin</surname> <given-names>C.</given-names></name> <name><surname>Lebaron</surname> <given-names>C.</given-names></name></person-group> (<publisher-loc>New York, NY</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>), <fpage>1</fpage>&#x2013;<lpage>26</lpage>.</citation></ref>
<ref id="B53"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Streeck</surname> <given-names>J.</given-names></name> <name><surname>Hartge</surname> <given-names>U.</given-names></name></person-group> (<year>1992</year>). &#x201C;<article-title>Previews: gestures at the transition place</article-title>,&#x201D; in <source><italic>The contextualization of language</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>Auer</surname> <given-names>P.</given-names></name> <name><surname>Luzio</surname> <given-names>A. D.</given-names></name></person-group> (<publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins</publisher-name>), <fpage>135</fpage>&#x2013;<lpage>157</lpage>.</citation></ref>
<ref id="B54"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Takahashi</surname> <given-names>J.</given-names></name></person-group> (<year>2018</year>). <article-title>Practices of self-selection in the graduate classroom: extension, redirection, and disjunction.</article-title> <source><italic>Linguist. Educ.</italic></source> <volume>46</volume> <fpage>70</fpage>&#x2013;<lpage>81</lpage>. <pub-id pub-id-type="doi">10.1016/j.linged.2018.06.002</pub-id></citation></ref>
<ref id="B55"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Taleghani-Nikazm</surname> <given-names>C.</given-names></name></person-group> (<year>2015</year>). &#x201C;<article-title>On multimodality and coordinated participation in second language interaction: a conversation-analytic perspective</article-title>,&#x201D; in <source><italic>Dialogue in Multilingual and Multimodal Communities</italic></source>, <role>eds</role> <person-group person-group-type="editor"><name><surname>Koike</surname> <given-names>D. A.</given-names></name> <name><surname>Blyth</surname> <given-names>S. C.</given-names></name></person-group> (<publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins</publisher-name>), <fpage>79</fpage>&#x2013;<lpage>103</lpage>.</citation></ref>
<ref id="B56"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Waring</surname> <given-names>H. Z.</given-names></name></person-group> (<year>2011</year>). <article-title>Learner initiatives and learning opportunities in the language classroom.</article-title> <source><italic>Classroom Discourse.</italic></source> <volume>2</volume> <fpage>201</fpage>&#x2013;<lpage>218</lpage>. <pub-id pub-id-type="doi">10.1080/19463014.2011.614053</pub-id></citation></ref>
<ref id="B57"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wei&#x00DF;</surname> <given-names>C.</given-names></name></person-group> (<year>2018</year>). <article-title>When gaze-selected next speakers do not take the turn.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>133</volume> <fpage>28</fpage>&#x2013;<lpage>44</lpage>. <pub-id pub-id-type="doi">10.1016/j.pragma.2018.05.016</pub-id></citation></ref>
<ref id="B58"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Yang</surname> <given-names>P.</given-names></name></person-group> (<year>2007</year>). <article-title>Nonverbal affiliative phenomena in Mandarin conversation.</article-title> <source><italic>J. Intercult. Communication.</italic></source> <volume>15</volume> <fpage>1</fpage>&#x2013;<lpage>42</lpage>.</citation></ref>
<ref id="B59"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Yang</surname> <given-names>P.</given-names></name></person-group> (<year>2010</year>). <article-title>Nonverbal gender differences: examining gestures of university-educated Mandarin Chinese speakers.</article-title> <source><italic>Text Talk.</italic></source> <volume>33</volume> <fpage>333</fpage>&#x2013;<lpage>357</lpage>. <pub-id pub-id-type="doi">10.1515/text.2010.017</pub-id></citation></ref>
<ref id="B60"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Yang</surname> <given-names>P.</given-names></name></person-group> (<year>2011</year>). <article-title>Nonverbal aspects of turn taking in Mandarin Chinese interaction.</article-title> <source><italic>Chin. Lang. Discourse</italic></source> <volume>2</volume> <fpage>99</fpage>&#x2013;<lpage>130</lpage>. <pub-id pub-id-type="doi">10.1075/cld.2.1.09yan</pub-id> <pub-id pub-id-type="pmid">33486653</pub-id></citation></ref>
<ref id="B61"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Yang</surname> <given-names>Z.</given-names></name></person-group> (<year>2019</year>). <article-title>Turn allocation within the medical-service-seeking party in Chinese accompanied medical consultations.</article-title> <source><italic>J. Pragmat.</italic></source> <volume>143</volume> <fpage>135</fpage>&#x2013;<lpage>155</lpage>. <pub-id pub-id-type="doi">10.1016/j.pragma.2019.02.005</pub-id></citation></ref>
</ref-list>
</back>
</article>
