<?xml version="1.0" encoding="utf-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" article-type="research-article" dtd-version="2.3" xml:lang="EN">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Psychol.</journal-id>
<journal-title>Frontiers in Psychology</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Psychol.</abbrev-journal-title>
<issn pub-type="epub">1664-1078</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fpsyg.2023.1202455</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Psychology</subject>
<subj-group>
<subject>Hypothesis and Theory</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Joint attention and its linguistic representation in dialogue: <italic>embodiment</italic> revisited</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes"><name><surname>Zeng</surname><given-names>Guocai</given-names></name><xref rid="c001" ref-type="corresp"><sup>&#x002A;</sup></xref>
<uri xlink:href="https://loop.frontiersin.org/people/2275146/overview"/>
</contrib>
</contrib-group>
<aff><institution>College of Foreign Languages and Cultures, Sichuan University</institution>, <addr-line>Chengdu</addr-line>, <country>China</country></aff>
<author-notes>
<fn fn-type="edited-by" id="fn0013"><p>Edited by: Ana Paula Soares, University of Minho, Portugal</p></fn>
<fn fn-type="edited-by" id="fn0014"><p>Reviewed by: Keshu Xiang, Guangxi Medical University, China; Jose Teixeira, University of Minho, Portugal</p></fn>
<corresp id="c001">&#x002A;Correspondence: Guocai Zeng, <email>1020785310@qq.com</email></corresp>
</author-notes>
<pub-date pub-type="epub">
<day>26</day>
<month>07</month>
<year>2023</year>
</pub-date>
<pub-date pub-type="collection">
<year>2023</year>
</pub-date>
<volume>14</volume>
<elocation-id>1202455</elocation-id>
<history>
<date date-type="received">
<day>17</day>
<month>04</month>
<year>2023</year>
</date>
<date date-type="accepted">
<day>30</day>
<month>06</month>
<year>2023</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x00A9; 2023 Zeng.</copyright-statement>
<copyright-year>2023</copyright-year>
<copyright-holder>Zeng</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/"><p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p>
</license>
</permissions>
<abstract>
<p>Lakoff and Johnson, among many others, have discussed the role of the human body in structuring meaning in communication, aiming to reveal the interrelation between the human body, language, and cognition. This study revisits the concept of <italic>embodiment</italic> and investigates its interactive nature functioning in speakers constructing repeated structures in conversation, based on the hypothesis made in this work that the joint attention of interlocutors essentially indicates the interaction of their embodied experience of the language used in the situated context, where speakers not only share their propositional commitments but also make individual contributions to establishing common ground in dialogue. Viewed in this way, at the linguistic level, the implicitly and/or explicitly repeated language resources displayed between utterances are in fact the encoding of speakers&#x2019; co-construction of joint attention and demonstrate the interplay of speakers&#x2019; syntactic and pragmatic knowledge in producing utterances in the talk turns. This research hopefully sheds some light on studies concerning the relationship between language and cognition as well as how language is constructed in dialogue from the interactive view of the syntax&#x2013;pragmatics interface.</p>
</abstract>
<kwd-group>
<kwd>embodiment</kwd>
<kwd>joint attention</kwd>
<kwd>commitment</kwd>
<kwd>common ground</kwd>
<kwd>repetition</kwd>
</kwd-group>
<contract-num rid="cn1">18BYY076</contract-num>
<contract-sponsor id="cn1">National Social Science Fund of China<named-content content-type="fundref-id">10.13039/501100012325</named-content></contract-sponsor>
<counts>
<fig-count count="1"/>
<table-count count="3"/>
<equation-count count="0"/>
<ref-count count="67"/>
<page-count count="8"/>
<word-count count="6424"/>
</counts>
<custom-meta-wrap>
<custom-meta>
<meta-name>section-at-acceptance</meta-name>
<meta-value>Psychology of Language</meta-value>
</custom-meta>
</custom-meta-wrap>
</article-meta>
</front>
<body>
<sec id="sec1">
<label>1.</label>
<title>Structuring language and meaning: an <italic>embodiment</italic> view</title>
<p>In line with the theoretical view of the cognitive-functional approach to natural language, meaning constructed in language communication is fundamentally rooted in the speaker&#x2019;s experience of the objective world,<xref rid="fn0001" ref-type="fn"><sup>1</sup></xref> thus suggesting the epistemic stance that a speaker holds towards the entities s/he physically or mentally experiences in the reality. In this sense, language used to encode a speaker&#x2019;s knowledge is embodied. For this study, the term <italic>embodiment</italic>, basically denoting that the human body matters when a speaker constructs his/her language, is construed mainly from two aspects, one of which is that in a broad sense a speaker understands the reality by dint of his/her bodily interaction with the objective world, while the other, in a more narrow sense and much more important, is that in daily conversations the <italic>language</italic> itself is the most crucial object speakers experience more frequently than other entities when they construct their dialogues. The latter point significantly indicates the embodied experience-based strategy that interlocutors employ when they take language to structure language in dialogue (<italic>cf.</italic> <xref ref-type="bibr" rid="ref16">Du Bois, 2014</xref>; <xref ref-type="bibr" rid="ref66">Zeng, 2021</xref>). The experience-based view of meaning construction, distinct from the rule-based structuring of meaning in the generative tradition of language studies, entails that meaning is personalized but coordinated between speakers in the communication. However, findings on how the human body functions in structuring language from a dialogic view are still rarely seen.</p>
<p>This study is supposed to bridge this gap to a certain degree by investigating the interactive nature of <italic>embodiment</italic> that virtually functions as the basis for interlocutors to establish joint attention, through which they produce their utterances in language communication. For this account, this research closely looks at the repeated structures in paired utterances that are produced by different speakers, proposing that explicitly and/or implicitly repeated grammatical structures are in essence the linguistic representation of the joint attention of interlocutors in dialogue, significantly indicating the speaker&#x2019;s perception-based strategy of taking language to make language in dialogue.</p>
</sec>
<sec id="sec2">
<label>2.</label>
<title>Extant views on <italic>embodiment</italic></title>
<p>According to <xref ref-type="bibr" rid="ref4">Bergen (2015</xref>, p.11), in general, <italic>embodiment</italic> seems to be used to mean something about how the mind relates to the body. And in the view of <xref ref-type="bibr" rid="ref54">Smith (2017</xref>, p.1), embodiment&#x2014;having, being in, or being associated with a body&#x2014;is a feature of the existence of many entities, perhaps even of all entities. For <xref ref-type="bibr" rid="ref49">Rohrer (2007</xref>, p.27), in its broadest definition, the embodiment hypothesis is that human physical, cognitive, and social embodiment grounds our conceptual and linguistic system, while in the view of <xref ref-type="bibr" rid="ref62">Walsh (2020)</xref>, embodiment is a bi-directional link between the body and body language, where the body both demonstrates and creates our being. For cognitive linguists, language, as part of humans&#x2019; cognition, is fundamentally motivated by embodied experience (<xref ref-type="bibr" rid="ref64">Wen and Jiang, 2021</xref>).</p>
<p>Remarkably, <xref ref-type="bibr" rid="ref49">Rohrer (2007</xref>, p. 28&#x2013;31) proposes that, with respect to human&#x2019;s cognition, the term <italic>embodiment</italic> can be used in at least 12 different important senses, including <xref ref-type="bibr" rid="ref33">Lakoff&#x2019;s (1987</xref>, p. xiv) view that the core of our conceptual systems is directly grounded in perception, body movement, and experience of a physical and social nature. In line with <xref ref-type="bibr" rid="ref34">Lakoff and Johnson&#x2019;s (1999</xref>, p. 37) interpretation, the very properties of concepts are created as a result of the way the brain and body are structured and the way they function in interpersonal relations and in the physical world. <xref ref-type="bibr" rid="ref34">Lakoff and Johnson (1999)</xref> also suggest that there are three levels of embodiment which together shape the embodied mind, namely the neural level, the level of phenomenological conscious experience, and that of the cognitive unconscious. In addition, <xref ref-type="bibr" rid="ref51">Shapiro (2014)</xref>, taking both empirical and philosophical views, investigates the properties of embodied cognition, especially focusing on such themes as embedded, extended, and enactive cognition, with the finding that there are strong interrelations between language and perception, reasoning, social and moral cognition, emotion, and consciousness, as well as human memory.</p>
<p>These findings have undoubtedly expanded our views about <italic>embodiment</italic> from different theoretical standpoints. But closer scrutiny of these studies suggests that most of the discussion on embodiment is not conducted from a dialogic view. This research will hopefully make a certain contribution to shortening this gap.</p>
</sec>
<sec id="sec3">
<label>3.</label>
<title>Interactive <italic>embodiment</italic></title>
<p>Utterances in dialogue are interactive in nature. Thinking in this pattern, it is natural to describe and explain how language is produced from an interactive embodiment view. To be specific, language is constructed based on a speaker&#x2019;s sharing of individually embodied experience of how language is used.</p>
<sec id="sec4">
<label>3.1.</label>
<title>Language: the object speakers experience in dialogue</title>
<p>In the studies on the relationship between language and cognition, it is assumed that the human body plays a key role in language production and comprehension (e.g., <xref ref-type="bibr" rid="ref33">Lakoff, 1987</xref>; <xref ref-type="bibr" rid="ref37">Langacker, 1987</xref>, <xref ref-type="bibr" rid="ref38">1991</xref>; <xref ref-type="bibr" rid="ref34">Lakoff and Johnson, 1999</xref>; <xref ref-type="bibr" rid="ref49">Rohrer, 2007</xref>), with the most concern about a single speaker&#x2019;s embodied experience of the physical world. But from the view that language is used dialogically (<xref ref-type="bibr" rid="ref3">Bakhtin, 1981</xref>; <xref ref-type="bibr" rid="ref47">Pickering and Garrod, 2021</xref>), how the experiences of speakers interact and what the role of such interaction could be in language production in conversation are in fact not paid much attention.</p>
<p>Unarguably, it is interlocutors<xref rid="fn0002" ref-type="fn"><sup>2</sup></xref> who participate in the <italic>embodiment</italic> process to structure utterances used in communication. In this process, typically a speaker first perceives the object(s) based on his/her body interacting with the physical world, then narrows down his/her attention to certain aspect(s) of the given entity, which is also the grounding process of abstract conceptual content in the speaker&#x2019;s mind (cf. <xref ref-type="bibr" rid="ref39">Langacker, 2008</xref>). The consequence of this grounding is eventually mapped onto the grammatical structures of the utterances, showing the linguistic encoding of one&#x2019;s experiencing of the reality.</p>
<p>Strikingly, also in this process, the language resources, which can be lexical items, sentence structures, functions, or prosodies of utterances used previously by a speaker, are the objects another speaker experiences physically or mentally. To put it another way, not only the human body but the language used to depict human&#x2019;s embodied experiencing of the world is exactly the object humans interact with in daily conversations, and should be highlighted when the <italic>embodiment</italic> view of language is examined. Convincing evidence for this observation is the language phenomena of repetition, which fundamentally refer to a speaker&#x2019;s imitation of his/her dialogic partner&#x2019;s speech acts (<xref ref-type="bibr" rid="ref11">Clark, 1977</xref>; <xref ref-type="bibr" rid="ref9">Bybee, 2006</xref>; <xref ref-type="bibr" rid="ref25">Huang, 2010</xref>) or a form of structural priming effect (<xref ref-type="bibr" rid="ref47">Pickering and Garrod, 2021</xref>) in dialogue. Grounded on such imitations or structural priming, interlocutors verify their own comprehension of the reality in the interaction of embodied and individualized experience.</p>
</sec>
<sec id="sec5">
<label>3.2.</label>
<title>Interaction: the core of the concept <italic>embodiment</italic></title>
<p>As <xref ref-type="bibr" rid="ref67">Zlatev (2017)</xref> proclaims, the significance of examining embodied intersubjectivity can never be overestimated. <xref ref-type="bibr" rid="ref64">Wen and Jiang (2021</xref>, p.150) hold a similar view that linguistic conceptualization is embodied in social interaction. To put it more simply, language is interactively embodied. The embodiment view of language is not only structured on the interaction of a single speaker&#x2019;s body with the physical world, but also built on the interaction between interlocutors&#x2019; experiencing their dialogic partners&#x2019; language uses. To consolidate this view, how the joint attention of speakers in conversation is framed and then linguistically encoded is especially surveyed in the following sections.</p>
</sec>
<sec id="sec6">
<label>3.3.</label>
<title>Joint attention: how interactive <italic>embodiment</italic> works</title>
<p>Language communication is driven by the speaker&#x2019;s attention (<xref ref-type="bibr" rid="ref36">Langacker, 1984</xref>; <xref ref-type="bibr" rid="ref40">Levelt, 1993</xref>; <xref ref-type="bibr" rid="ref21">Giora, 2003</xref>; <xref ref-type="bibr" rid="ref46">Myachykov and Posner, 2005</xref>; <xref ref-type="bibr" rid="ref57">Talmy, 2007</xref>, <xref ref-type="bibr" rid="ref58">2017</xref>; <xref ref-type="bibr" rid="ref8">Breyer, 2009</xref>; <xref ref-type="bibr" rid="ref52">Shtyrov et al., 2010</xref>; <xref ref-type="bibr" rid="ref35">Lampert, 2015</xref>; <xref ref-type="bibr" rid="ref65">Yliniemi, 2021</xref>; <xref ref-type="bibr" rid="ref13">Dash et al., 2022</xref>). This view might work to instantiate the principle of &#x2018;<italic>what you see is what you get</italic> (e.g., <xref ref-type="bibr" rid="ref5">Boas, 2021</xref>, p.64)&#x2019; followed in cognitive linguistic studies on natural language. With this thinking in mind, dialogic partners&#x2019; experiencing of each other&#x2019;s language use is in fact the basis for interlocutors to structure joint attention in conversation.</p>
<sec id="sec7">
<label>3.3.1.</label>
<title>Attention in dialogue</title>
<p>Attention is the window through which a speaker perceives the world based on his/her body. A speaker&#x2019;s attention in dialogue might be visual attention or mental attention. Both of them are structured on human organs, particularly the former on eyes and the latter on the mind. The scope a speaker&#x2019;s attention can reach is roughly classified into the maximal scope, immediate scope, and the focused area (<xref ref-type="bibr" rid="ref39">Langacker, 2008</xref>, p. 260&#x2013;263).</p>
<p>When it comes to a speaker&#x2019;s talking about something verbally or non-verbally, s/he directs her or his own visual or mental attention to the candidate instance(s) within a category of entities first, and then zooms the attention in to particular one(s), focusing on the feature(s) of the targeted instance. To attract the speaker&#x2019;s attention, the candidate instance in a category is prototypically salient to a certain degree in the speaker&#x2019;s physical or mental world. As <xref ref-type="bibr" rid="ref59">Talmy (2021)</xref> argues in his analyses of attention phenomena, entities with more salience in terms of their locations and/or shapes, etc., are foregrounded while those which are less salient will be backgrounded in the conversational settings; those closely related to or participating in the ongoing dialogue process are more likely to be the center of the speakers&#x2019; attention; moreover, entities with a higher degree of complexity in the internal organization demand more cognitive processing efforts from the speaker, thus they are more likely to be attended to by a speaker; what is more, moveable objects rather than static ones in the attended scope are supposed to attract more attention from the interlocutors. Put briefly, entities prominent in certain features in the speaker&#x2019;s attention scope are more likely to be qualified as the dialogic focus in communication.</p>
</sec>
<sec id="sec8">
<label>3.3.2.</label>
<title>Joint attention</title>
<p><xref ref-type="bibr" rid="ref67">Zlatev (2017)</xref> proposes that meaning is sourced from humans&#x2019; interaction and is especially associated with speakers&#x2019; interactive and enactive perception, when he expounds embodied intersubjectivity in general and mimetic acts in particular, based on the analyses of body schema, body language, body memory, bodily movement, and perception. This study goes further, assuming that interactive perception is the speaker&#x2019;s bodily ground to co-construct joint attention in dialogue.</p>
<p>Joint attention, by definition, is the mental window shared by speakers (<italic>cf.</italic> <xref ref-type="bibr" rid="ref60">Tomasello, 1995</xref>; <xref ref-type="bibr" rid="ref12">Clark and Josie, 2008</xref>; <xref ref-type="bibr" rid="ref29">Kecskes, 2008</xref>; <xref ref-type="bibr" rid="ref44">Mondada, 2009</xref>) and it is interactively embodied in nature. To construct the joint attention, the speaker grounds the salient entity in his/her visual or mental world, and then directs the hearer&#x2019;s attention to the same entity via language or non-language cues. In doing so, an overlapping of attention from both the speaker and the hearer is displayed in the conversation. Before joint attention is formed, interlocutors do have their own focuses of attention. For this reason, what is attentively aimed at in joint attention might differ from the one that is the individual speaker&#x2019;s concern. The process where an object is attentively targeted by both speaker and hearer reveals the interlocutors&#x2019; cognitive coordination to establish the ground for ongoing dialogic interactions. The frame of joint attention can be illustrated as in <xref rid="fig1" ref-type="fig">Figure 1</xref>.<xref rid="fn0003" ref-type="fn"><sup>3</sup></xref></p>
<fig position="float" id="fig1">
<label>Figure 1</label>
<caption>
<p>The frame of joint attention in dialogue.</p>
</caption>
<graphic xlink:href="fpsyg-14-1202455-g001.tif"/>
</fig>
<p><xref rid="fig1" ref-type="fig">Figure 1</xref> shows that in conversational settings two speakers<xref rid="fn0004" ref-type="fn"><sup>4</sup></xref> interact to shift their attentions to the same entity, namely the dialogic focus marked by the rectangle with solid bold black lines. This jointly attended entity is the one located in the immediate attention scope (short for IAS), where other entities related to the focused one but with less salient features in shape, location, or color are backgrounded in the maximal attention scope (short for MAS), which is the largest size of joint attention area. The MAS and IAS suggest the speakers&#x2019; different allocations of their attentions. That is, a speaker devotes least attention to the MAS and most to the focused entity. Rooted in the interactive perception-based joint attention scope, speakers make and share their propositional commitments towards the dialogic focus and progressively make contributions to enlarging the common ground, based on which the interlocutors coordinate their stance-takings in the taking of turns.</p>
</sec>
</sec>
<sec id="sec9">
<label>3.4.</label>
<title>Beyond joint attention</title>
<p>In the process of building joint attention, a speaker makes propositional commitments about the targeted entity to his/her dialogic partner on the one hand (<xref ref-type="bibr" rid="ref28">Katriel and Dascal, 1989</xref>; <xref ref-type="bibr" rid="ref14">De Brabanter and Dendale, 2008</xref>; <xref ref-type="bibr" rid="ref20">Gilbert, 2013</xref>; <xref ref-type="bibr" rid="ref23">Heinonen, 2015</xref>; <xref ref-type="bibr" rid="ref6">Bonalumi et al., 2020</xref>; <xref ref-type="bibr" rid="ref18">Elder, 2021</xref>), and contributes to constructing the common ground shared by the interlocutors in conversation on the other hand.</p>
<sec id="sec10">
<label>3.4.1.</label>
<title>Propositional commitment shared</title>
<p>According to <xref ref-type="bibr" rid="ref19">Geurts (2019</xref>, p.1), human communication is first and foremost a matter of negotiating commitments and every speech act causes the speaker to become committed to the hearer to act on a propositional content. On his account, commitment is a three-place relation between two individuals, <italic>a</italic> and <italic>b</italic>, and a propositional content, <italic>p.</italic> That is, <italic>a</italic> is committed to <italic>b</italic> to act on <italic>p</italic> (<italic>ibid</italic>:3). Commitment is therefore understood as a social relationship subserving action coordination between individuals. Following this view, in the zone of joint attention, the speaker&#x2019;s language coding of events <italic>de facto</italic> makes an epistemic commitment (<italic>cf.</italic> <xref ref-type="bibr" rid="ref24">Hoff, 2019</xref>) concerning specific propositional contents to the hearer, who in turn makes a propositional commitment towards the speaker by producing a responsive utterance (<italic>cf.</italic> <xref ref-type="bibr" rid="ref7">Boulat, 2014</xref>).</p>
<p>From a dialogic view, the interlocutor&#x2019;s mutual commitments are joint attention-based in that the proposition contents in commitments are related to entities that are mentally contacted by both speaker and hearer, whereas a single speaker&#x2019;s propositional commitment is not insofar as there is no engagement of speakers&#x2019; interaction. In the joint attention zone, a speaker might have strong or weak commitment toward the shared target, suggesting the different allocations of attention of the interlocutors in conversation. The speaker&#x2019;s epistemic stance towards the dialogic focus could be weakened, reaffirmed, or even completely undermined because of the dialogic partner&#x2019;s propositional commitment to it, denoting the degrees of speakers&#x2019; subjectivity in construing the object in the attention scope and reflecting the different consequences of the speakers&#x2019; experiencing of the reality.</p>
<p>Therefore, in the exchange of talk turns, speakers actually share their propositional commitments to the jointly attended objects. The commitment interaction, which is speaker-centered (<xref ref-type="bibr" rid="ref43">Moeschler, 2013</xref>) or hearer-based (<xref ref-type="bibr" rid="ref45">Morency et al., 2008</xref>), is then justified as the engine driving the dialogue process to go forward. Since joint attention is structured in humans&#x2019; perception interaction, the shared commitment made in the dialogue is substantially embodied.</p>
</sec>
<sec id="sec11">
<label>3.4.2.</label>
<title>Common ground established</title>
<p>In the joint attention, the commitment made by a speaker suggests his/her knowledge about the world. Such knowledge contributes to the structuring of the common ground<xref rid="fn0005" ref-type="fn"><sup>5</sup></xref> in developing the size of local dialogue (<xref ref-type="bibr" rid="ref55">Stalnaker, 2002</xref>; <xref ref-type="bibr" rid="ref1">Abbott, 2008</xref>; <xref ref-type="bibr" rid="ref30">Kecskes and Zhang, 2009</xref>, <xref ref-type="bibr" rid="ref31">2013</xref>; <xref ref-type="bibr" rid="ref2">Allan, 2013</xref>; <xref ref-type="bibr" rid="ref22">Green, 2017</xref>; <xref ref-type="bibr" rid="ref50">Semeijn, 2017</xref>; <xref ref-type="bibr" rid="ref56">Swanson, 2020</xref>; <xref ref-type="bibr" rid="ref42">Marsili, 2021</xref>). Particularly, according to <xref ref-type="bibr" rid="ref30">Kecskes and Zhang (2009</xref>, p.347), there are two sides to common ground: <italic>core common ground</italic> and <italic>emergent common ground</italic>. In their view, the former refers to the relatively static, generalized, shared knowledge that belongs to a particular speech community, while the latter designates the relatively dynamic, specific, private knowledge created in the progress of communication that belongs to the individual(s).</p>
<p>In a broad sense, common ground encapsulates conventional social-cultural information that is by default understood by interlocutors and their personally embodied experience in the world (<xref ref-type="bibr" rid="ref26">Jaszczolt, 2005</xref>, <xref ref-type="bibr" rid="ref27">2016</xref>). Narrowly speaking, what is emergent in the ongoing speakers&#x2019; interaction could be the newly built common ground knowledge, some of which might be much more salient in the focused and/or the immediate attention area than that in the maximal attention scope.</p>
<p><xref ref-type="bibr" rid="ref30">Kecskes and Zhang (2009)</xref> also propose that the individual attention, which is the cause of the interlocutor&#x2019;s egocentrism, and the speaker&#x2019;s intention, which through relevance is expressed in cooperation, are equally important in constructing common ground. From an interactive embodiment view of the speaker&#x2019;s attention, common ground is naturally constructed through the process where speakers with egocentric behaviors interact with each other to build interpersonal cooperation. The more commonalities speakers construct with collaborative efforts in the joint attention zone, the more opportunities interlocutors have for achieving agreement in the negotiation of stances.</p>
</sec>
</sec>
</sec>
<sec id="sec12">
<label>4.</label>
<title>Repetition<xref rid="fn0006" ref-type="fn"><sup>6</sup></xref>: the linguistic representation of joint attention</title>
<p><xref ref-type="bibr" rid="ref49">Rohrer (2007</xref>, p. 26) mentions <italic>using language to establish joint attention</italic>. <xref ref-type="bibr" rid="ref61">Tomasello and Farrar (1986)</xref>, <xref ref-type="bibr" rid="ref32">Krause (1997)</xref>, <xref ref-type="bibr" rid="ref10">Charman (2003)</xref>, <xref ref-type="bibr" rid="ref48">Robins et al. (2004)</xref>, <xref ref-type="bibr" rid="ref17">Eilan et al. (2005)</xref>, <xref ref-type="bibr" rid="ref15">Diessel (2006)</xref>, and <xref ref-type="bibr" rid="ref53">Skarabela (2007)</xref>, among others, have also investigated the relationship between language and speaker&#x2019;s attention, but how grammatical structures linguistically encode joint attention in conversation has not been examined with great detail.</p>
<p>As <xref rid="fig1" ref-type="fig">Figure 1</xref> implies, language emerges in dynamic conversation, in which the interaction of the speakers&#x2019; embodied experience occurs. Linguistically speaking, grammatical structures that are shared by speakers in such interaction basically encode the mutual commitments made and the common ground shared by speakers. Or more precisely, repetitions of language resource in dialogue function to highlight the embodied interaction of individual attention, displaying how joint attention is co-constructed by different speakers in talk turns, which can be elaborated in dialogue (1).</p>
<list list-type="simple">
<list-item>
<p>Dialogue (1)<xref rid="fn0007" ref-type="fn"><sup>7</sup></xref></p>
<p><inline-graphic xlink:href="fpsyg-14-1202455-igr0001.tif"/></p>
<p>Diagraph for dialogue (1)<xref rid="fn0009" ref-type="fn"><sup>9</sup></xref></p>
</list-item>
</list>
<table-wrap position="anchor" id="tab1">
<table frame="hsides" rules="groups">
<tbody>
<tr>
<td align="left" valign="top">Speaker 1</td>
<td align="left" valign="top" rowspan="5">&#x2193;</td>
<td align="left" valign="top">...</td>
<td align="left" valign="top">What</td>
<td align="left" valign="top">was</td>
<td align="left" valign="top">it</td>
<td align="left" valign="top">?</td>
</tr>
<tr>
<td align="left" valign="top" rowspan="3">Speaker 2</td>
<td rowspan="4"/>
<td/>
<td/>
<td align="left" valign="top">It</td>
<td/>
</tr>
<tr>
<td/>
<td align="left" valign="top">was</td>
<td/>
<td/>
</tr>
<tr>
<td align="left" valign="top">a canceled check</td>
<td/>
<td/>
<td align="left" valign="top">.</td>
</tr>
<tr>
<td align="left" valign="top">Speaker 3</td>
<td align="left" valign="top">A canceled check from 1992</td>
<td/>
<td/>
<td align="left" valign="top">.</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>We can see from the diagraph for this local dialogue that there are three talk turns. At the end of Speaker 1&#x2019;s talk turn is a wh-question &#x2018;<italic>what was it?</italic>&#x2019;. Cognitively speaking, in contrast with the previous utterance that works as the background for speaker 1 to produce this wh-question, the <italic>what</italic> is more saliently positioned in the conversation because of its unspecified semantic content in this question. The utterance backgrounded for the wh-question in the first talk turn is in the maximal attention scope of speakers, and works as the common ground for speakers 2 and 3 to construe the instantiation of <italic>what,</italic> which is in the immediate attention scope of all speakers.</p>
<p>From an interactive view of embodiment, the heading position and the schematic content of <italic>what</italic> in the question are the motivation for speakers 1&#x2013;3 to construct joint attention, whose focus is exactly the <italic>what</italic>. In speaker 2&#x2019;s talk turn, <italic>what</italic> is specified as <italic>a canceled check,</italic> while it is <italic>a canceled check from 1992</italic> in speaker 3&#x2019;s talk turn, a slightly more detailed instance of <italic>what</italic>, demonstrating the process in which speakers 2 and 3 make their own propositional commitments to the schematic <italic>what</italic>. By doing so, the semantic content of <italic>what</italic>, the jointly attended target by the interlocutors in the ongoing conversation, is specified step by step with finer details.</p>
<p>Furthermore, the explicitly paralleled structures <italic>&#x2018;it:it; was:was&#x2019;</italic> within the question-answer 1(QA1) and &#x2018;<italic>what</italic>: <italic>a canceled check&#x2019;</italic> in QA 2<italic>,</italic> as well as the symmetrical structures &#x2018;<italic>a canceled check: a canceled check&#x2019;</italic> between speakers 2 and 3&#x2019;s talk turns, work together to indicate the common grounds for the speakers to further interpret <italic>what.</italic> The syntactic parallelism in this sense significantly reveals the linguistic evidence of the speakers&#x2019; joint attention on <italic>what</italic>. The alignment of speakers&#x2019; attention at the same time signifies that three speakers successfully build cooperation to specify the schematic content of <italic>what</italic> as <italic>a canceled check</italic> and <italic>a canceled check from 1992</italic>.</p>
<p>The way that joint attention is constructed in dialogue is also observed in child-to-child interaction, as shown in dialogue (2), a case of Mandarin-speaking children&#x2019;s conversation.</p>
<list list-type="simple">
<list-item>
<p>Dialogue (2): (Loc:Chinese/Mandarin/LiZhou/3/06.cha)<xref rid="fn0010" ref-type="fn"><sup>10</sup></xref></p>
<p>@ID: zho|LiZhou|CH1|3;00.|female|||Target_Child|||</p>
<p>@ID: zho|LiZhou|CH2|3;00.|male|||Target_Child|||</p>
<p>342 CH1<xref rid="fn0011" ref-type="fn"><sup>11</sup></xref>: &#x4F60; &#x8FD9; &#x4E2A; &#x4F1A; &#x5531; &#x6B4C; &#x5427;?</p>
<p>Ni zhe ge hui chang ge ba?</p>
<p>You this can sing song PARTICLE.</p>
<p>&#x2018;Can your this one sing songs?&#x2019;</p>
<p>343 CH2: &#x8FD9; &#x4E0D; &#x4F1A; &#x5531; &#x6B4C;&#x3002;</p>
<p>Zhe bu hui chang ge.</p>
<p>This cannot sing song.</p>
<p>&#x2018;This one cannot sing songs.&#x2019;</p>
<p>344 CH1: &#x8FD9; &#x4E2A; &#x4F1A; &#x5531; &#x6B4C;&#x3002;</p>
<p>Zhe ge hui chang ge.</p>
<p>This one can sing song.</p>
<p>&#x2018;This one can sing songs.&#x2019;</p>
<p>345 CH2: &#x8001;&#x864E; &#x7684; &#x6B4C;&#x3002;</p>
<p>Laohu de ge.</p>
<p>Tiger PARTICLE song</p>
<p>&#x2018;Songs about tigers.&#x2019;</p>
<p>346 CH1: &#x8001;&#x864E; &#x7684; &#x6B4C;&#x3002;</p>
<p>Laohu de ge.</p>
<p>Tiger PARTICLE song</p>
<p>&#x2018;Songs about tigers.&#x2019;</p>
<p>347 CH2: &#x54C8;&#x54C8;&#x3002;</p>
<p>Ha ha.</p>
<p>PARTICLE.</p>
<p>&#x2018;Ha-ha.&#x2019;</p>
<p>Diagraph for dialogue (2)</p>
</list-item>
</list>
<table-wrap position="anchor" id="tab2">
<table frame="hsides" rules="groups">
<thead>
<tr>
<th/>
<th/>
<th/>
<th align="center" valign="middle">particle</th>
<th align="center" valign="middle">personal pronoun</th>
<th align="center" valign="middle">demonstrative pronoun</th>
<th align="center" valign="middle">negation</th>
<th align="center" valign="middle">modal word</th>
<th align="center" valign="middle">verb</th>
<th align="center" valign="middle">noun</th>
<th align="center" valign="middle">auxiliary word</th>
<th align="center" valign="middle">noun</th>
<th align="center" valign="middle">article</th>
<th/>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="top">342</td>
<td align="center" valign="top">CH1:</td>
<td align="center" valign="top" rowspan="6">&#x2193;</td>
<td rowspan="5"/>
<td align="center" valign="top">Ni</td>
<td align="center" valign="top"><italic>zhe ge</italic></td>
<td/>
<td align="center" valign="top"><italic>hui</italic></td>
<td align="center" valign="top"><italic>chang</italic></td>
<td/>
<td/>
<td align="center" valign="top"><italic>ge</italic></td>
<td align="center" valign="top">ba</td>
<td align="center" valign="top">?</td>
</tr>
<tr>
<td align="left" valign="top">343</td>
<td align="center" valign="top">CH2:</td>
<td/>
<td align="center" valign="top"><italic>zhe ge</italic></td>
<td align="center" valign="top">bu</td>
<td align="center" valign="top"><italic>hui</italic></td>
<td align="center" valign="top"><italic>chang</italic></td>
<td/>
<td/>
<td align="center" valign="top"><italic>ge</italic></td>
<td/>
<td align="center" valign="top">.</td>
</tr>
<tr>
<td align="left" valign="top">344</td>
<td align="center" valign="top">CH1:</td>
<td/>
<td align="center" valign="top"><italic>zhe ge</italic></td>
<td/>
<td align="center" valign="top"><italic>hui</italic></td>
<td align="center" valign="top"><italic>chang</italic></td>
<td/>
<td/>
<td align="center" valign="top"><italic>ge</italic></td>
<td/>
<td align="center" valign="top">.</td>
</tr>
<tr>
<td align="left" valign="top">345</td>
<td align="center" valign="top">CH2:</td>
<td/>
<td/>
<td/>
<td/>
<td/>
<td align="center" valign="top"><italic>Laohu</italic></td>
<td align="center" valign="top"><italic>de</italic></td>
<td align="center" valign="top"><italic>ge</italic></td>
<td/>
<td align="center" valign="top">.</td>
</tr>
<tr>
<td align="left" valign="top">346</td>
<td align="center" valign="top">CH1:</td>
<td/>
<td/>
<td/>
<td/>
<td/>
<td align="center" valign="top"><italic>Laohu</italic></td>
<td align="center" valign="top"><italic>de</italic></td>
<td align="center" valign="top"><italic>ge</italic></td>
<td/>
<td align="center" valign="top">.</td>
</tr>
<tr>
<td align="left" valign="top">347</td>
<td align="center" valign="top">CH2:</td>
<td align="center" valign="top">Haha</td>
<td colspan="9"/>
<td align="center" valign="top">.</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>As demonstrated in this diagraph, repetitions of words (marked as italics) are obviously produced along with the ongoing dialogue process, simultaneously displaying the emergent shared template structures &#x2018;<italic>pronoun&#x2009;+&#x2009;modal word&#x2009;+&#x2009;verb&#x2009;+&#x2009;noun</italic>&#x2019; and &#x2018;<italic>noun&#x2009;+&#x2009;auxiliary word&#x2009;+&#x2009;noun</italic>&#x2019;, with &#x2018;<italic>zhe ge hui chang ge (this one can sing songs)</italic>&#x2019; and &#x2018;<italic>laohu de ge (songs about tigers)&#x2019;</italic> as the instances. These patterned structures, jointly attended to by these two child speakers, are also the common grounds for them to develop the size of the local discourse.</p>
<p>To be more specific, within talk turns 342&#x2013;344, both of them focus on the schematic event &#x2018;<italic>X hui Y (X can do Y)</italic>&#x2019; and its instance &#x2018;<italic>zhe ge hui chang ge (this one can sing songs)</italic>&#x2019;, while in talk turns 345&#x2013;346, child 1 and child 2 attend the same and more specific object, which is &#x2018;<italic>laohu de ge (songs about tigers)</italic>,&#x2019; revealing the cognitive coordination and interpersonal cooperation between the child interlocutors by making commitment to each other. In the last talk turn (347), the particle &#x2018;<italic>haha</italic>&#x2019; indicates that, at the end of this episode of a short conversation, child 2&#x2019;s attention is successfully directed by child 1 to the instance of &#x2018;<italic>ge (songs)</italic>&#x2019;, namely <italic>laohu de ge (songs about tigers)</italic>.</p>
<p>In this sense, repeated structures are viewed as the linguistic encoding of the joint attention of speakers in dialogues (1) and (2), which at the meantime implies the interplay of interlocutors&#x2019; syntactic-pragmatic knowledge in structuring utterances, thus bringing forth dialogic resonance (<italic>cf.</italic> <xref ref-type="bibr" rid="ref16">Du Bois, 2014</xref>) in communication, as can be further observed in dialogue (3).</p>
<list list-type="simple">
<list-item>
<p>Dialogue (3) (<italic>Tastes Very Special</italic> SBC031: 533.430&#x2013;541.201)<xref rid="fn0012" ref-type="fn"><sup>12</sup></xref></p>
<p>1 SHERRY; @^I @don&#x2019;t even like <bold>ice</bold> ^<bold>tea</bold>.</p>
<p>2 BETH; (H) (0.7) Do you like &#x00BF;^<bold>hot tea</bold>?</p>
<p>3 (0.6)</p>
<p>4 SHERRY;&#x2009;^Yeah,</p>
<p>5&#x2009;I ^love <bold>hot tea</bold>.</p>
<p>Diagraph for dialogue (3)</p>
</list-item>
</list>
<table-wrap position="anchor" id="tab3">
<table frame="hsides" rules="groups">
<thead>
<tr>
<th/>
<th/>
<th/>
<th align="left" valign="top">pronoun</th>
<th align="left" valign="top">auxiliary verb</th>
<th align="left" valign="top">verb</th>
<th align="left" valign="top">noun</th>
<th align="left" valign="top">punctuation</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="top">SHERRY</td>
<td align="left" valign="top" rowspan="5">&#x2193;</td>
<td/>
<td align="left" valign="top"><italic>I</italic></td>
<td align="left" valign="top"><italic>do not</italic></td>
<td align="left" valign="top"><italic>like</italic></td>
<td align="left" valign="top"><italic>ice tea</italic></td>
<td align="left" valign="top">.</td>
</tr>
<tr>
<td align="left" valign="middle" rowspan="2">BETH</td>
<td/>
<td/>
<td align="left" valign="top"><italic>do</italic></td>
<td/>
<td/>
<td/>
</tr>
<tr>
<td/>
<td align="left" valign="top"><italic>you</italic></td>
<td/>
<td align="left" valign="top"><italic>like</italic></td>
<td align="left" valign="top"><italic>hot tea</italic></td>
<td align="left" valign="top">?</td>
</tr>
<tr>
<td align="left" valign="middle" rowspan="2">SHERRY</td>
<td align="left" valign="top">Yeah</td>
<td/>
<td/>
<td/>
<td/>
<td align="left" valign="top">,</td>
</tr>
<tr>
<td/>
<td align="left" valign="top"><italic>I</italic></td>
<td/>
<td align="left" valign="top"><italic>love</italic></td>
<td align="left" valign="top"><italic>hot tea</italic></td>
<td align="left" valign="top">.</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>As seen in this diagraph, SHERRY first made a commitment with a negative statement introducing &#x2018;<italic>ice tea</italic>&#x2019;. SHERRY&#x2019;s negative attitude towards <italic>ice tea</italic> primarily functions as the partial common ground for BETH to structure her expression containing &#x2018;<italic>hot tea</italic>&#x2019;. The partially repeated structures &#x2018;<italic>ice tea: hot tea</italic>&#x2019; based on &#x2018;<italic>tea</italic>&#x2019; shows that both speakers have mental contact with &#x2018;<italic>tea</italic>&#x2019;. That is to say, the categorical entity encoded by <italic>tea</italic> is jointly attended by the two speakers when they make speech acts concerning SHERRY&#x2019;S preference of the instance of <italic>tea</italic>. Meanwhile, this parallelism suggests SHERRY and BETH have their own allocation of attention in the joint attention scope, with the former on <italic>hot tea</italic> but the latter on <italic>ice tea</italic>.</p>
<p>Notably, the shared structural pattern &#x2018;<italic>X like Y</italic>&#x2019; abstracted in talk turns 1&#x2013;2 based on the joint attended &#x2018;<italic>tea</italic>&#x2019; is the background for SHERRY to structure her utterance &#x2018;<italic>I love hot tea.</italic>&#x2019; The partial symmetry in the semantic structure between <italic>love</italic> and <italic>like,</italic> which in this case indicates a certain degree of preference for <italic>tea</italic>, also signifies the interpersonal interaction founded through the structural alignment &#x2018;<italic>I: you: I.</italic>&#x2019; In addition to that, grounded on this grammatical pattern, the two speakers&#x2019; embodied interactive stance-takings are entirely presented via the negative tone of talk turn 1, the interrogative tone in talk turn 2, and the assertion in the last talk turn, altogether displaying the pragmatic function of joint attention in the dialogue.</p>
</sec>
<sec id="sec13">
<label>5.</label>
<title>Concluding remarks</title>
<p>To sum up, as <xref rid="fig1" ref-type="fig">Figure 1</xref> and dialogues (1)&#x2013;(3) suggest, utterance interaction in conversation entails the speakers&#x2019; co-construction of joint attention, which is rooted in the speaker&#x2019;s general ability to perceive the world through the human body. Language production is hence driven by the formation of the speaker&#x2019;s joint attention, that is, the interaction of individual attention in the communication. More precisely, the interlocutors in dialogue take language to structure language through setting up joint attention. In doing so, the speakers at the same time make commitment to each other and establish common ground for the ongoing dialogue, based on their embodied experience of their partner&#x2019;s language use in situated context. At the linguistic level, repeated language structures are in essence the encoding of speakers&#x2019; joint attention. In this sense, language is interactively embodied in nature. Human&#x2019;s interactively embodied experience of other persons&#x2019; language use essentially reveals interlocutors&#x2019; cognitive coordination and interpersonal cooperation in communication. According to <xref ref-type="bibr" rid="ref63">Wang (2019)</xref>, the <italic>embodiment</italic> view of natural language, one of the fundamental claims in cognitive linguistic studies, cannot be overemphasized and the theoretical assumptions in <italic>Cognitive Linguistics</italic> should be reinterpreted or redefined within the framework of <italic>Embodied Cognitive Linguistics</italic>. This study will, hopefully, widen the views of the sense of <italic>embodiment</italic> from a dialogic perspective on language and shed some light on the research concerning the relationship between language and cognition as well as how language is constructed in dialogue from the interactive view of the syntax&#x2013;pragmatics interface.</p>
</sec>
<sec sec-type="data-availability" id="sec14">
<title>Data availability statement</title>
<p>The original contributions presented in the study are included in the article/supplementary material, further inquiries can be directed to the corresponding author.</p>
</sec>
<sec id="sec15">
<title>Ethics statement</title>
<p>Ethical review and approval was not required for the study on human participants in accordance with the local legislation and institutional requirements. Written informed consent from the patients/ participants or patients/participants' legal guardian/next of kin was not required to participate in this study in accordance with the national legislation and the institutional requirements.</p>
</sec>
<sec id="sec16">
<title>Author contributions</title>
<p>The author confirms being the sole contributor of this work and has approved it for publication.</p>
</sec>
<sec sec-type="funding-information" id="sec17">
<title>Funding</title>
<p>This work was supported by National Social Science Fund of China (Grant number: 18BYY076).</p>
</sec>
<sec sec-type="COI-statement" id="sec18">
<title>Conflict of interest</title>
<p>The author declares that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec id="sec100">
<title>Publisher&#x2019;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
</body>
<back>
<ack>
<p>I am grateful to the reviewers for their insightful comments and would like to express my greatest appreciation to professor Kasia Jaszczolt at the University of Cambridge for her valuable comments on an earlier version of this article when I worked as a visiting scholar at the University of Cambridge from Sep. 2021 to Sep. 2022. This work was supported by National Social Science Fund of China (Grant number: 18BYY076).</p>
</ack>
<ref-list>
<title>References</title>
<ref id="ref1"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Abbott</surname> <given-names>B.</given-names></name></person-group> (<year>2008</year>). <article-title>Presuppositions and common ground</article-title>. <source>Linguist. Philosophy</source> <volume>31</volume>, <fpage>523</fpage>&#x2013;<lpage>538</lpage>. doi: <pub-id pub-id-type="doi">10.1007/s10988-008-9048-8</pub-id></citation></ref>
<ref id="ref2"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Allan</surname> <given-names>K.</given-names></name></person-group> (<year>2013</year>). &#x201C;<article-title>What is common ground?</article-title>&#x201D; in <source>Perspectives on linguistic pragmatics</source>. eds. <person-group person-group-type="editor"><name><surname>Capone</surname> <given-names>A.</given-names></name> <name><surname>Piparo</surname> <given-names>F. L.</given-names></name> <name><surname>Carapezza</surname> <given-names>M.</given-names></name></person-group> (<publisher-loc>Springer Cham</publisher-loc>: <publisher-name>New York/London</publisher-name>), <fpage>285</fpage>&#x2013;<lpage>310</lpage>.</citation></ref>
<ref id="ref3"><citation citation-type="other"><person-group person-group-type="author"><name><surname>Bakhtin</surname> <given-names>M.</given-names></name></person-group> (<year>1981</year>). <source>The dialogic imagination: Four essays</source> (Trans. by <person-group person-group-type="translator"><name><surname>Emerson</surname><given-names>C.</given-names></name><name><surname>Holquist</surname><given-names>M.</given-names></name></person-group>). <publisher-loc>Austin</publisher-loc>: <publisher-name>University of Texas Press</publisher-name>.</citation></ref>
<ref id="ref4"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Bergen</surname> <given-names>B.</given-names></name></person-group> (<year>2015</year>). &#x201C;<article-title>Embodiment</article-title>&#x201D; in <source>Handbook of cognitive linguistics</source>. eds. <person-group person-group-type="editor"><name><surname>Dabrowska</surname> <given-names>E.</given-names></name> <name><surname>Divjak</surname> <given-names>D.</given-names></name></person-group> (<publisher-loc>Berlin/Boston</publisher-loc>: <publisher-name>De Gruyter Mouton</publisher-name>), <fpage>10</fpage>&#x2013;<lpage>30</lpage>.</citation></ref>
<ref id="ref5"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Boas</surname> <given-names>H.</given-names></name></person-group> (<year>2021</year>). &#x201C;<article-title>Construction grammar and frame semantics</article-title>&#x201D; in <source>The Routledge handbook of cognitive linguistics</source>. eds. <person-group person-group-type="editor"><name><surname>Wen</surname> <given-names>X.</given-names></name> <name><surname>Taylor</surname> <given-names>J. R.</given-names></name></person-group> (<publisher-loc>New York</publisher-loc>: <publisher-name>Taylor &#x0026; Francis Group</publisher-name>), <fpage>43</fpage>&#x2013;<lpage>77</lpage>.</citation></ref>
<ref id="ref6"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bonalumi</surname> <given-names>F.</given-names></name> <name><surname>Scott-Phillips</surname> <given-names>T.</given-names></name> <name><surname>Tacha</surname> <given-names>J.</given-names></name> <name><surname>Heintz</surname> <given-names>C.</given-names></name></person-group> (<year>2020</year>). <article-title>Commitment and communication: are we committed to what we mean, or what we say?</article-title> <source>Lang. Cogn.</source> <volume>12</volume>, <fpage>360</fpage>&#x2013;<lpage>384</lpage>. doi: <pub-id pub-id-type="doi">10.1017/langcog.2020.2</pub-id></citation></ref>
<ref id="ref7"><citation citation-type="other"><person-group person-group-type="author"><name><surname>Boulat</surname> <given-names>K.</given-names></name></person-group> (<year>2014</year>). Are you committed? A pragmatic model of commitment. Presented at the 8th Days of Swiss Linguistics, 19&#x2013;21 June 2014, Zurich.</citation></ref>
<ref id="ref8"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Breyer</surname> <given-names>T.</given-names></name></person-group> (<year>2009</year>). <article-title>Attention and language: preliminary remarks on a philosophically important connection</article-title>. <source>Intuitio</source> <volume>2</volume>, <fpage>245</fpage>&#x2013;<lpage>256</lpage>.</citation></ref>
<ref id="ref9"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bybee</surname> <given-names>J.</given-names></name></person-group> (<year>2006</year>). <article-title>From usage to grammar: the mind&#x2019;s response to repetition</article-title>. <source>Language</source> <volume>82</volume>, <fpage>711</fpage>&#x2013;<lpage>733</lpage>. doi: <pub-id pub-id-type="doi">10.1353/lan.2006.0186</pub-id></citation></ref>
<ref id="ref10"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Charman</surname> <given-names>T.</given-names></name></person-group> (<year>2003</year>). <article-title>Why is joint attention a pivotal skill in autism?</article-title> <source>Philos. Trans. R. Soc. Lond. Ser. B Biol. Sci.</source> <volume>358</volume>, <fpage>315</fpage>&#x2013;<lpage>324</lpage>. doi: <pub-id pub-id-type="doi">10.1098/rstb.2002.1199</pub-id>, PMID: <pub-id pub-id-type="pmid">12639329</pub-id></citation></ref>
<ref id="ref11"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Clark</surname> <given-names>R.</given-names></name></person-group> (<year>1977</year>). <article-title>What's the use of imitation?</article-title> <source>J. Child Lang.</source> <volume>4</volume>, <fpage>341</fpage>&#x2013;<lpage>358</lpage>. doi: <pub-id pub-id-type="doi">10.1017/S0305000900001732</pub-id></citation></ref>
<ref id="ref12"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Clark</surname> <given-names>E. V.</given-names></name> <name><surname>Josie</surname> <given-names>B.</given-names></name></person-group> (<year>2008</year>). <article-title>Repetition as ratification: how parents and children place information in common ground</article-title>. <source>J. Child Lang.</source> <volume>35</volume>, <fpage>349</fpage>&#x2013;<lpage>371</lpage>. doi: <pub-id pub-id-type="doi">10.1017/S0305000907008537</pub-id>, PMID: <pub-id pub-id-type="pmid">18416863</pub-id></citation></ref>
<ref id="ref13"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dash</surname> <given-names>T.</given-names></name> <name><surname>Joanette</surname> <given-names>Y.</given-names></name> <name><surname>Ansaldo</surname> <given-names>A. I.</given-names></name></person-group> (<year>2022</year>). <article-title>Exploring attention in the bilingualism continuum: a resting-state functional connectivity study</article-title>. <source>Brain Lang.</source> <volume>224</volume>:<fpage>105048</fpage>. doi: <pub-id pub-id-type="doi">10.1016/j.bandl.2021.105048</pub-id>, PMID: <pub-id pub-id-type="pmid">34781212</pub-id></citation></ref>
<ref id="ref14"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>De Brabanter</surname> <given-names>P.</given-names></name> <name><surname>Dendale</surname> <given-names>P.</given-names></name></person-group> (<year>2008</year>). <article-title>Commitment: the term and the notions</article-title>. <source>Belgian J. Linguist.</source> <volume>22</volume>, <fpage>1</fpage>&#x2013;<lpage>14</lpage>. doi: <pub-id pub-id-type="doi">10.1075/bjl.22.01de</pub-id></citation></ref>
<ref id="ref15"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Diessel</surname> <given-names>H.</given-names></name></person-group> (<year>2006</year>). <article-title>Demonstratives, joint attention, and the emergence of grammar. Cognitive</article-title>. <source>Linguistics</source> <volume>17</volume>, <fpage>463</fpage>&#x2013;<lpage>489</lpage>. doi: <pub-id pub-id-type="doi">10.1515/COG.2006.015</pub-id></citation></ref>
<ref id="ref16"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Du Bois</surname> <given-names>J. W.</given-names></name></person-group> (<year>2014</year>). <article-title>Towards a dialogic syntax</article-title>. <source>Cogn. Linguist.</source> <volume>25</volume>, <fpage>359</fpage>&#x2013;<lpage>410</lpage>. doi: <pub-id pub-id-type="doi">10.1515/cog-2014-0024</pub-id></citation></ref>
<ref id="ref17"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Eilan</surname> <given-names>N.</given-names></name> <name><surname>Hoerl</surname> <given-names>C.</given-names></name> <name><surname>McCormack</surname> <given-names>T.</given-names></name> <name><surname>Roessler</surname> <given-names>J.</given-names></name></person-group> (<year>2005</year>). <source>Joint attention: Communication and other minds: Issues in philosophy and psychology</source>. <publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>.</citation></ref>
<ref id="ref18"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Elder</surname> <given-names>C.</given-names></name></person-group> (<year>2021</year>). &#x201C;<article-title>Speaker meaning, commitment and accountability</article-title>&#x201D; in <source>The Cambridge handbook of Sociopragmatics Cambridge handbooks in language and linguistics</source>. eds. <person-group person-group-type="editor"><name><surname>Haugh</surname> <given-names>M.</given-names></name> <name><surname>K&#x00E1;d&#x00E1;r</surname> <given-names>D.</given-names></name> <name><surname>Terkourafi</surname> <given-names>M.</given-names></name></person-group> (<publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>), <fpage>48</fpage>&#x2013;<lpage>68</lpage>.</citation></ref>
<ref id="ref19"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Geurts</surname> <given-names>B.</given-names></name></person-group> (<year>2019</year>). <article-title>Communication as commitment sharing: speech acts, implicatures, common ground</article-title>. <source>Theoret. Linguist.</source> <volume>45</volume>, <fpage>1</fpage>&#x2013;<lpage>30</lpage>. doi: <pub-id pub-id-type="doi">10.1515/tl-2019-0001</pub-id></citation></ref>
<ref id="ref20"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Gilbert</surname> <given-names>M.</given-names></name></person-group> (<year>2013</year>). <source>Joint commitment: How we make the social world</source>. <publisher-loc>New York</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>.</citation></ref>
<ref id="ref21"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Giora</surname> <given-names>R.</given-names></name></person-group> (<year>2003</year>). <source>On our mind: Salience, context, and figurative language</source>. <publisher-loc>New York</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>.</citation></ref>
<ref id="ref22"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Green</surname> <given-names>M.</given-names></name></person-group> (<year>2017</year>). <article-title>Conversation and common ground</article-title>. <source>Philos. Stud.</source> <volume>174</volume>, <fpage>1587</fpage>&#x2013;<lpage>1604</lpage>. doi: <pub-id pub-id-type="doi">10.1007/s11098-016-0779-z</pub-id></citation></ref>
<ref id="ref23"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Heinonen</surname> <given-names>M.</given-names></name></person-group> (<year>2015</year>). <article-title>Joint commitment: how we make the social world</article-title>. <source>J. Soc. Ontol.</source> <volume>1</volume>, <fpage>175</fpage>&#x2013;<lpage>178</lpage>. doi: <pub-id pub-id-type="doi">10.1515/jso-2014-0032</pub-id></citation></ref>
<ref id="ref24"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hoff</surname> <given-names>M.</given-names></name></person-group> (<year>2019</year>). <article-title>Epistemic commitment and mood alternation: a semantic-pragmatic analysis of Spanish future-framed adverbials</article-title>. <source>J. Pragmat.</source> <volume>139</volume>, <fpage>97</fpage>&#x2013;<lpage>108</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.pragma.2018.10.016</pub-id></citation></ref>
<ref id="ref25"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Huang</surname> <given-names>C. C.</given-names></name></person-group> (<year>2010</year>). <article-title>Other-repetition in mandarin child language: a discourse-pragmatic perspective</article-title>. <source>J. Pragmat.</source> <volume>42</volume>, <fpage>825</fpage>&#x2013;<lpage>839</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.pragma.2009.08.005</pub-id></citation></ref>
<ref id="ref26"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Jaszczolt</surname> <given-names>K. M.</given-names></name></person-group> (<year>2005</year>). <source>Default semantics: Foundations of a compositional theory of acts of communication</source>. <publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>.</citation></ref>
<ref id="ref27"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Jaszczolt</surname> <given-names>K.M.</given-names></name></person-group> (<year>2016</year>). <source>Meaning in linguistic interaction: Semantics, metasemantics, philosophy of language</source>. <publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>.</citation></ref>
<ref id="ref28"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Katriel</surname> <given-names>T.</given-names></name> <name><surname>Dascal</surname> <given-names>M.</given-names></name></person-group> (<year>1989</year>). &#x201C;<article-title>Speaker&#x2019;s commitment and involvement in discourse</article-title>&#x201D; in <source>From sign to text/Ed.Tobin, Y</source> (<publisher-loc>Amsterdam (Philadelphia)</publisher-loc>: <publisher-name>John Benjamins Publishing Company</publisher-name>), <fpage>275</fpage>&#x2013;<lpage>295</lpage>.</citation></ref>
<ref id="ref29"><citation citation-type="other"><person-group person-group-type="author"><name><surname>Kecskes</surname> <given-names>I.</given-names></name> <name><surname>Mey</surname> <given-names>J.</given-names></name></person-group> (<year>2008</year>). <source>Intention, common ground and the egocentric speaker-hearer</source>. <publisher-loc>Berlin/New York</publisher-loc>: <publisher-name>Walter de Gruyter</publisher-name>.</citation></ref>
<ref id="ref30"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kecskes</surname> <given-names>I.</given-names></name> <name><surname>Zhang</surname> <given-names>F.</given-names></name></person-group> (<year>2009</year>). <article-title>Activating, seeking, and creating common ground: a socio-cognitive approach</article-title>. <source>Pragmat. Cogn.</source> <volume>17</volume>, <fpage>331</fpage>&#x2013;<lpage>355</lpage>. doi: <pub-id pub-id-type="doi">10.1075/pc.17.2.06kec</pub-id></citation></ref>
<ref id="ref31"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Kecskes</surname> <given-names>I.</given-names></name> <name><surname>Zhang</surname> <given-names>F.</given-names></name></person-group> (<year>2013</year>). &#x201C;<article-title>On the dynamic relations between common ground and presupposition</article-title>&#x201D; in <source>Perspectives on linguistic pragmatics</source>. eds. <person-group person-group-type="editor"><name><surname>Capone</surname> <given-names>A.</given-names></name> <name><surname>Piparo</surname> <given-names>F. L.</given-names></name> <name><surname>Carapezza</surname> <given-names>M.</given-names></name></person-group> (<publisher-loc>New York/London</publisher-loc>: <publisher-name>Springer Cham</publisher-name>), <fpage>375</fpage>&#x2013;<lpage>395</lpage>.</citation></ref>
<ref id="ref32"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Krause</surname> <given-names>M. A.</given-names></name></person-group> (<year>1997</year>). <article-title>Comparative perspectives on pointing and joint attention in children and apes</article-title>. <source>Int. J. Comp. Psychol.</source> <volume>10</volume>, <fpage>137</fpage>&#x2013;<lpage>157</lpage>. doi: <pub-id pub-id-type="doi">10.46867/C44K5H</pub-id></citation></ref>
<ref id="ref33"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Lakoff</surname> <given-names>G.</given-names></name></person-group> (<year>1987</year>). <source>Women, fire, and dangerous things</source>. <publisher-loc>Chicago</publisher-loc>: <publisher-name>University of Chicago Press</publisher-name>.</citation></ref>
<ref id="ref34"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Lakoff</surname> <given-names>G.</given-names></name> <name><surname>Johnson</surname> <given-names>M.</given-names></name></person-group> (<year>1999</year>). <source>Philosophy in the flesh --- the embodied mind and its challenge to Western thought</source>. <publisher-loc>New York</publisher-loc>: <publisher-name>Basic Books</publisher-name>.</citation></ref>
<ref id="ref35"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Lampert</surname> <given-names>M.</given-names></name></person-group> (<year>2015</year>). &#x201C;<article-title>How attention determines meaning: a cognitive-semantic study of the steady-state causatives remain, stay, continue, keep, still, on</article-title>&#x201D; in <source>Attention and meaning: The attentional basis of meaning</source>. eds. <person-group person-group-type="editor"><name><surname>Marchetti</surname> <given-names>G.</given-names></name> <name><surname>Benedetti</surname> <given-names>G.</given-names></name> <name><surname>Alharbi</surname> <given-names>A.</given-names></name></person-group> (<publisher-loc>New York</publisher-loc>: <publisher-name>Nova Science Publishers</publisher-name>), <fpage>207</fpage>&#x2013;<lpage>238</lpage>.</citation></ref>
<ref id="ref36"><citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Langacker</surname> <given-names>R.W.</given-names></name></person-group> (<year>1984</year>). <article-title>Active zones</article-title>. <conf-name>Proceedings of the Annual Meeting of the Berkeley Linguistics Society</conf-name> <volume>10</volume>, <fpage>172</fpage>&#x2013;<lpage>188</lpage>, doi: <pub-id pub-id-type="doi">10.3765/bls.v10i0.3175</pub-id></citation></ref>
<ref id="ref37"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Langacker</surname> <given-names>R.W.</given-names></name></person-group> (<year>1987</year>). <source>Foundations of cognitive grammar vol. I: Theoretical prerequisites</source>. <publisher-loc>Stanford, California</publisher-loc>: <publisher-name>Stanford University Press</publisher-name></citation></ref>
<ref id="ref38"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Langacker</surname> <given-names>R.W.</given-names></name></person-group> (<year>1991</year>). <source>Foundations of cognitive grammar vol. II: Descriptive application</source>. <publisher-loc>Stanford, California</publisher-loc>: <publisher-name>Stanford University Press</publisher-name>.</citation></ref>
<ref id="ref39"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Langacker</surname> <given-names>R.W.</given-names></name></person-group> (<year>2008</year>). <source>Cognitive grammar: A basic introduction</source>. <publisher-loc>Oxford</publisher-loc>: <publisher-name>OUP</publisher-name>.</citation></ref>
<ref id="ref40"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Levelt</surname> <given-names>W. J.</given-names></name></person-group> (<year>1993</year>). <source>Speaking: From intention to articulation</source>, vol. <volume>1</volume>. <publisher-loc>Cambridge, Massachusetts(MA)</publisher-loc>: <publisher-name>MIT press</publisher-name>.</citation></ref>
<ref id="ref41"><citation citation-type="book"><person-group person-group-type="author"><name><surname>MacWhinney</surname> <given-names>B.</given-names></name></person-group> (<year>2000</year>). <source>The CHILDES project: Tools for analyzing talk</source>, <edition>3rd edn</edition>. <publisher-loc>Mahwah, NJ</publisher-loc>: <publisher-name>Lawrence Erlbaum</publisher-name>.</citation></ref>
<ref id="ref42"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Marsili</surname> <given-names>N.</given-names></name></person-group> (<year>2021</year>). <article-title>Lies, common ground and performative utterances</article-title>. <source>Erkenntnis</source>, <fpage>1</fpage>&#x2013;<lpage>12</lpage>. doi: <pub-id pub-id-type="doi">10.1007/s10670-020-00368-4</pub-id></citation></ref>
<ref id="ref43"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Moeschler</surname> <given-names>J.</given-names></name></person-group> (<year>2013</year>). <article-title>Is a speaker-based pragmatics possible? Or how can a hearer infer a speaker's commitment?</article-title> <source>J. Pragmat.</source> <volume>48</volume>, <fpage>84</fpage>&#x2013;<lpage>97</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.pragma.2012.11.019</pub-id></citation></ref>
<ref id="ref44"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mondada</surname> <given-names>L.</given-names></name></person-group> (<year>2009</year>). <article-title>Emergent focused interactions in public places: a systematic analysis of the multimodal achievement of a common interactional space</article-title>. <source>J. Pragmat.</source> <volume>41</volume>, <fpage>1977</fpage>&#x2013;<lpage>1997</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.pragma.2008.09.019</pub-id></citation></ref>
<ref id="ref45"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Morency</surname> <given-names>P.</given-names></name> <name><surname>Oswald</surname> <given-names>S.</given-names></name> <name><surname>de Saussure</surname> <given-names>L.</given-names></name></person-group> (<year>2008</year>). <article-title>Explicitness, implicitness and commitment attribution: a cognitive pragmatic approach</article-title>. In <person-group person-group-type="editor"><name><surname>Brabanter</surname> <given-names>P.</given-names><prefix>de</prefix></name> <name><surname>Dendale</surname> <given-names>P.</given-names></name></person-group> (eds.), <source>Commitment (Belgian journal of linguistics 22)</source>.<publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins</publisher-name>, <fpage>197</fpage>&#x2013;<lpage>220</lpage>.</citation></ref>
<ref id="ref46"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Myachykov</surname> <given-names>A.</given-names></name> <name><surname>Posner</surname> <given-names>M. I.</given-names></name></person-group> (<year>2005</year>). &#x201C;<article-title>Attention in language</article-title>&#x201D; in <source>Neurobiology of attention</source>. eds. <person-group person-group-type="editor"><name><surname>Itti</surname> <given-names>L.</given-names></name> <name><surname>Rees</surname> <given-names>G.</given-names></name> <name><surname>Tsotsos</surname> <given-names>J. K.</given-names></name></person-group> (<publisher-loc>London/Burlington</publisher-loc>: <publisher-name>Elsevier Academic Press</publisher-name>), <fpage>324</fpage>&#x2013;<lpage>329</lpage>.</citation></ref>
<ref id="ref47"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Pickering</surname> <given-names>M. J.</given-names></name> <name><surname>Garrod</surname> <given-names>S</given-names></name></person-group>. (<year>2021</year>). <source>Understanding dialogue:Language use and social interaction</source>. <publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>.</citation></ref>
<ref id="ref48"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Robins</surname> <given-names>B.</given-names></name> <name><surname>Dickerson</surname> <given-names>P.</given-names></name> <name><surname>Stribling</surname> <given-names>P.</given-names></name> <name><surname>Dautenhahn</surname> <given-names>K.</given-names></name></person-group> (<year>2004</year>). <article-title>Robot-mediated joint attention in children with autism: a case study in robot-human interaction</article-title>. <source>Interact. Stud.</source> <volume>5</volume>, <fpage>161</fpage>&#x2013;<lpage>198</lpage>. doi: <pub-id pub-id-type="doi">10.1075/is.5.2.02rob</pub-id></citation></ref>
<ref id="ref49"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Rohrer</surname> <given-names>T.</given-names></name></person-group> (<year>2007</year>). &#x201C;<article-title>Embodiment and experientialism</article-title>&#x201D; in <source>The Oxford handbook of cognitive linguistics</source>. eds. <person-group person-group-type="editor"><name><surname>Geeraerts</surname> <given-names>D.</given-names></name> <name><surname>Cuyckens</surname> <given-names>H.</given-names></name></person-group> (<publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>), <fpage>25</fpage>&#x2013;<lpage>47</lpage>.</citation></ref>
<ref id="ref50"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Semeijn</surname> <given-names>M.</given-names></name></person-group> (<year>2017</year>). <article-title>A Stalnakerian analysis of Metafictive statements</article-title>. In <person-group person-group-type="editor"><name><surname>Cremers</surname> <given-names>A.</given-names></name> <name><surname>Gessel</surname> <given-names>T.</given-names><prefix>van</prefix></name> <name><surname>Roelofsen</surname> <given-names>F.</given-names></name></person-group> (Eds.), <source>Proceedings of the 21st Amsterdam colloquium Amsterdam.ILLC/Department of Philosophy</source>. (<publisher-loc>Amsterdam</publisher-loc>: <publisher-name>University of Amsterdam</publisher-name>), <fpage>415</fpage>&#x2013;<lpage>425</lpage>.</citation></ref>
<ref id="ref51"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Shapiro</surname> <given-names>L. A.</given-names></name></person-group> (<year>2014</year>). <source>The Routledge handbook of embodied cognition</source>. <publisher-loc>London/New York</publisher-loc>: <publisher-name>Taylor &#x0026; Francis Group</publisher-name>.</citation></ref>
<ref id="ref52"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Shtyrov</surname> <given-names>Y.</given-names></name> <name><surname>Kujala</surname> <given-names>T.</given-names></name> <name><surname>Pulverm&#x00FC;ller</surname> <given-names>F.</given-names></name></person-group> (<year>2010</year>). <article-title>Interactions between language and attention systems: early automatic lexical processing?</article-title> <source>J. Cogn. Neurosci.</source> <volume>22</volume>, <fpage>1465</fpage>&#x2013;<lpage>1478</lpage>. doi: <pub-id pub-id-type="doi">10.1162/jocn.2009.21292</pub-id>, PMID: <pub-id pub-id-type="pmid">19580394</pub-id></citation></ref>
<ref id="ref53"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Skarabela</surname> <given-names>B.</given-names></name></person-group> (<year>2007</year>). <article-title>Signs of early social cognition in children's syntax: the case of joint attention in argument realization in child Inuktitut</article-title>. <source>Lingua</source> <volume>117</volume>, <fpage>1837</fpage>&#x2013;<lpage>1857</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.lingua.2006.11.010</pub-id></citation></ref>
<ref id="ref54"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Smith</surname> <given-names>J.E.</given-names></name></person-group> (<year>2017</year>). <source>Embodiment: A history</source>. <publisher-loc>New York</publisher-loc>:<publisher-name>Oxford University Press</publisher-name>.</citation></ref>
<ref id="ref55"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Stalnaker</surname> <given-names>R.</given-names></name></person-group> (<year>2002</year>). <article-title>Common ground</article-title>. <source>Linguist. Philosophy</source> <volume>25</volume>, <fpage>701</fpage>&#x2013;<lpage>721</lpage>. doi: <pub-id pub-id-type="doi">10.1023/A:1020867916902</pub-id></citation></ref>
<ref id="ref56"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Swanson</surname> <given-names>E.</given-names></name></person-group> (<year>2020</year>). <article-title>Channels for common ground</article-title>. <source>Philos. Phenomenol. Res.</source> <volume>00</volume>, <fpage>1</fpage>&#x2013;<lpage>15</lpage>. doi: <pub-id pub-id-type="doi">10.1111/phpr.12741</pub-id></citation></ref>
<ref id="ref57"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Talmy</surname> <given-names>L.</given-names></name></person-group> (<year>2007</year>). &#x201C;<article-title>Attention phenomena</article-title>&#x201D; in <source>The Oxford handbook of cognitive linguistics</source>. eds. <person-group person-group-type="editor"><name><surname>Geeraerts</surname> <given-names>D.</given-names></name> <name><surname>Cuyckens</surname> <given-names>H.</given-names></name></person-group> (<publisher-loc>Oxford, New York</publisher-loc>: <publisher-name>University Press</publisher-name>), <fpage>264</fpage>&#x2013;<lpage>293</lpage>.</citation></ref>
<ref id="ref58"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Talmy</surname> <given-names>L.</given-names></name></person-group> (<year>2017</year>). <source>The targeting system of language</source>.<publisher-loc>Cambridge, Massachusetts,London, England</publisher-loc>: <publisher-name>The MIT Press</publisher-name>.</citation></ref>
<ref id="ref59"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Talmy</surname> <given-names>L.</given-names></name></person-group> (<year>2021</year>). <article-title>Structure within morphemic meaning</article-title>. <source>Cogn. Semantics</source> <volume>7</volume>, <fpage>155</fpage>&#x2013;<lpage>231</lpage>. doi: <pub-id pub-id-type="doi">10.1163/23526416-07020003</pub-id></citation></ref>
<ref id="ref60"><citation citation-type="other"><person-group person-group-type="author"><name><surname>Tomasello</surname> <given-names>M.</given-names></name></person-group> (<year>1995</year>). <source>Joint attention as social cognition. In joint attention: Its origins and role in development</source>. Moore,C., and <person-group person-group-type="editor"><name><surname>Dunham</surname> <given-names>P.</given-names></name></person-group> (ed). <publisher-loc>Hillsdale, NJ</publisher-loc>: <publisher-name>Erlbaum</publisher-name>, <fpage>103</fpage>&#x2013;<lpage>120</lpage>.</citation></ref>
<ref id="ref61"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tomasello</surname> <given-names>M.</given-names></name> <name><surname>Farrar</surname> <given-names>M. J.</given-names></name></person-group> (<year>1986</year>). <article-title>Joint attention and early language</article-title>. <source>Child Dev.</source> <volume>57</volume>, <fpage>1454</fpage>&#x2013;<lpage>1463</lpage>. doi: <pub-id pub-id-type="doi">10.2307/1130423</pub-id></citation></ref>
<ref id="ref62"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Walsh</surname> <given-names>M.</given-names></name></person-group> (<year>2020</year>). <source>Embodiment - moving beyond mindfulness</source>. <publisher-loc>London</publisher-loc>: <publisher-name>Unicorn Slayer Press</publisher-name>.</citation></ref>
<ref id="ref63"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wang</surname> <given-names>Y.</given-names></name></person-group> (<year>2019</year>). <article-title>Essential thoughts on embodied cognitive linguistics</article-title>. <source>Foreign Lang. China</source> <volume>16</volume>, <fpage>18</fpage>&#x2013;<lpage>25</lpage>. doi: <pub-id pub-id-type="doi">10.13564/j.cnki.issn.1672-9382.2019.06.004</pub-id></citation></ref>
<ref id="ref64"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Wen</surname> <given-names>X.</given-names></name> <name><surname>Jiang</surname> <given-names>C.</given-names></name></person-group> (<year>2021</year>). &#x201C;<article-title>Embodiment</article-title>&#x201D; in <source>The Routledge handbook of cognitive linguistics</source>. eds. <person-group person-group-type="editor"><name><surname>Wen</surname> <given-names>X.</given-names></name> <name><surname>Taylor</surname> <given-names>J. R.</given-names></name></person-group> (<publisher-loc>New York</publisher-loc>: <publisher-name>Taylor &#x0026; Francis Group</publisher-name>), <fpage>145</fpage>&#x2013;<lpage>160</lpage>.</citation></ref>
<ref id="ref65"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Yliniemi</surname> <given-names>J.</given-names></name></person-group> (<year>2021</year>). <article-title>Similarity of mirative and contrastive focus: three parameters for describing attention markers</article-title>. <source>Linguist. Typol.</source> <volume>27</volume>, <fpage>77</fpage>&#x2013;<lpage>111</lpage>. doi: <pub-id pub-id-type="doi">10.1515/lingty-2020-0134</pub-id></citation></ref>
<ref id="ref66"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zeng</surname> <given-names>G.</given-names></name></person-group> (<year>2021</year>). <article-title>Repetition in mandarin-speaking children&#x2019;s dialogs: its distribution and structural dimensions</article-title>. <source>Linguist. Vanguard</source> <volume>7</volume>:<fpage>20200059</fpage>. doi: <pub-id pub-id-type="doi">10.1515/lingvan-2020-0059</pub-id></citation></ref>
<ref id="ref67"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Zlatev</surname> <given-names>J.</given-names></name></person-group> (<year>2017</year>). &#x201C;<article-title>Embodied Intersubjectivity</article-title>&#x201D; in <source>The Cambridge handbook of cognitive linguistics</source>. ed. <person-group person-group-type="editor"><name><surname>Dancygier</surname> <given-names>B.</given-names></name></person-group> (<publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>), <fpage>172</fpage>&#x2013;<lpage>187</lpage>.</citation></ref>
</ref-list>
<fn-group>
<fn id="fn0001"><p><sup>1</sup>The &#x2018;objective world (<italic>cf.</italic> <xref ref-type="bibr" rid="ref34">Lakoff and Johnson, 1999</xref>)&#x2019; in this work is applied particularly in the sense of the physical world where human beings live.</p></fn>
<fn id="fn0002"><p><sup>2</sup>This study concerns typical conversations where the speaker and the hearer are different persons. A monologue, where the speaker and hearer is the same person, can be analyzed in the same way.</p></fn>
<fn id="fn0003"><p><sup>3</sup>This figure is based on the modifications of Figures present in <xref ref-type="bibr" rid="ref39">Langacker (2008</xref>, p. 260-263).</p></fn>
<fn id="fn0004"><p><sup>4</sup>Only two speakers are mentioned here and they are two different persons. A trialogue or monologue can be analyzed in the same way.</p></fn>
<fn id="fn0005"><p><sup>5</sup>Common ground here is defined to include the emergent knowledge in the dialogue, which goes beyond the discussions by <xref ref-type="bibr" rid="ref55">Stalnaker (2002)</xref> and <xref ref-type="bibr" rid="ref1">Abbott (2008)</xref>.</p></fn>
<fn id="fn0006"><p><sup>6</sup>Since the joint attention is constructed with different speakers&#x2019; efforts, other-repetitions rather than self-repetitions (<italic>cf.</italic> <xref ref-type="bibr" rid="ref25">Huang, 2010</xref>) in dialogue are concerned in this study.</p></fn>
<fn id="fn0007"><p><sup>7</sup>This dialogue is retrieved from COCA corpus (source-SPOK:CNN-Chung; Date 2002-11-12; Title: Latest bin Laden Tape Stirs Debate, Fear of New Attacks).</p></fn>
<fn id="fn00090"><p><sup>8</sup>Three dots here indicate the omitted utterances that are not directly relevant to the analyses at hand.</p></fn>
<fn id="fn0009"><p><sup>9</sup>A diagraph (<italic>cf.</italic> <xref ref-type="bibr" rid="ref16">Du Bois, 2014</xref>) is used here to indicate the structural mapping and parallelism between utterances and the symbol &#x2018;&#x2193;&#x2019; in the diagraph indicates the direction of turn-construction.</p></fn>
<fn id="fn0010"><p><sup>10</sup>This dialogue is retrieved from CHILDES corpus (<xref ref-type="bibr" rid="ref41">MacWhinney, 2000</xref>).</p></fn>
<fn id="fn0011"><p><sup>11</sup>CH1 and CH2 here refer to child-speaker 1 and child-speaker 2, respectively.</p></fn>
<fn id="fn0012"><p><sup>12</sup>This dialogue is quoted from <xref ref-type="bibr" rid="ref16">Du Bois (2014</xref>, p. 381).</p></fn>
</fn-group>
</back>
</article>