<?xml version="1.0" encoding="UTF-8" standalone="no"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Hum. Neurosci.</journal-id>
<journal-title>Frontiers in Human Neuroscience</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Hum. Neurosci.</abbrev-journal-title>
<issn pub-type="epub">1662-5161</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fnhum.2017.00073</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Neuroscience</subject>
<subj-group>
<subject>Perspective</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Child-Robot Interactions for Second Language Tutoring to Preschool Children</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name><surname>Vogt</surname> <given-names>Paul</given-names></name>
<xref ref-type="author-notes" rid="fn001"><sup>&#x0002A;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/134335/overview"/>
<xref ref-type="aff" rid="aff1"/>
</contrib>
<contrib contrib-type="author">
<name><surname>de Haas</surname> <given-names>Mirjam</given-names></name>
<uri xlink:href="http://loop.frontiersin.org/people/404540/overview"/>
<xref ref-type="aff" rid="aff1"/>
</contrib>
<contrib contrib-type="author">
<name><surname>de Jong</surname> <given-names>Chiara</given-names></name>
<uri xlink:href="http://loop.frontiersin.org/people/414941/overview"/>
<xref ref-type="aff" rid="aff1"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Baxter</surname> <given-names>Peta</given-names></name>
<uri xlink:href="http://loop.frontiersin.org/people/413498/overview"/>
<xref ref-type="aff" rid="aff1"/>
</contrib> 
<contrib contrib-type="author">
<name><surname>Krahmer</surname> <given-names>Emiel</given-names></name>
<uri xlink:href="http://loop.frontiersin.org/people/82107/overview"/>
<xref ref-type="aff" rid="aff1"/>
</contrib>
</contrib-group>
<aff id="aff1"><institution>Tilburg Center for Cognition and Communication, Tilburg University</institution> <country>Tilburg, Netherlands</country></aff>
<author-notes>
<fn fn-type="edited-by"><p>Edited by: Mila Vulchanova, Norwegian University of Science and Technology, Norway</p></fn>
<fn fn-type="edited-by"><p>Reviewed by: Ramesh Kumar Mishra, University of Hyderabad, India; Vera Kempe, Abertay University, UK</p></fn>
<fn fn-type="corresp" id="fn001"><p>&#x0002A;Correspondence: Paul Vogt <email>p.a.vogt&#x00040;uvt.nl</email></p></fn>
</author-notes>
<pub-date pub-type="epub">
<day>02</day>
<month>03</month>
<year>2017</year>
</pub-date>
<pub-date pub-type="collection">
<year>2017</year>
</pub-date>
<volume>11</volume>
<elocation-id>73</elocation-id>
<history>
<date date-type="received">
<day>26</day>
<month>10</month>
<year>2016</year>
</date>
<date date-type="accepted">
<day>06</day>
<month>02</month>
<year>2017</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x000A9; 2017 Vogt, de Haas, de Jong, Baxter and Krahmer.</copyright-statement>
<copyright-year>2017</copyright-year>
<copyright-holder>Vogt, de Haas, de Jong, Baxter and Krahmer</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/"><p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution and reproduction in other forums is permitted, provided the original author(s) or licensor are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p>
</license>
</permissions>
<abstract><p>In this digital age social robots will increasingly be used for educational purposes, such as second language tutoring. In this perspective article, we propose a number of design features to develop a child-friendly social robot that can effectively support children in second language learning, and we discuss some technical challenges for developing these. The features we propose include choices to develop the robot such that it can act as a peer to motivate the child during second language learning and build trust at the same time, while still being more knowledgeable than the child and scaffolding that knowledge in adult-like manner. We also believe that the first impressions children have about robots are crucial for them to build trust and common ground, which would support child-robot interactions in the long term. We therefore propose a strategy to introduce the robot in a safe way to toddlers. Other features relate to the ability to adapt to individual children&#x02019;s language proficiency, respond contingently, both temporally and semantically, establish joint attention, use meaningful gestures, provide effective feedback and monitor children&#x02019;s learning progress. Technical challenges we observe include automatic speech recognition (ASR) for children, reliable object recognition to facilitate semantic contingency and establishing joint attention, and developing human-like gestures with a robot that does not have the same morphology humans have. We briefly discuss an experiment in which we investigate how children respond to different forms of feedback the robot can give.</p></abstract>
<kwd-group>
<kwd>social robots</kwd>
<kwd>second language tutoring</kwd>
<kwd>education</kwd>
<kwd>child-robot interaction</kwd>
<kwd>robot assisted language learning</kwd>
</kwd-group>
<contract-num rid="cn001">688014</contract-num>
<contract-sponsor id="cn001">European Commission<named-content content-type="fundref-id">10.13039/501100000780</named-content></contract-sponsor>
<counts>
<fig-count count="1"/>
<table-count count="0"/>
<equation-count count="0"/>
<ref-count count="41"/>
<page-count count="7"/>
<word-count count="5271"/>
</counts>
</article-meta>
</front>
<body>
<sec id="s1">
<title>Social Robots for Second Language Tutoring</title>
<p>Given the globalization of our society, it is becoming increasingly important for people to speak multiple languages. For instance, the ability to speak foreign languages fosters people&#x02019;s mobility and increases their chances for employment. Moreover, immigrants to a country need to learn the official host language. Since young children are most flexible at learning languages, starting second language (L2) learning in preschool would provide them a good opportunity to acquire the second language more fluently at a later age (Hoff, <xref ref-type="bibr" rid="B18">2013</xref>).</p>
<p>One trend in the digital age of the 21st century is that technologies are being developed for educational purposes, including technologies to support L2 tutoring. There exist many forms of digital technologies for PCs, laptops or tablet computers that support second language learning, although there is little evidence about their efficacy (Golonka et al., <xref ref-type="bibr" rid="B15">2014</xref>; Hsin et al., <xref ref-type="bibr" rid="B20">2014</xref>). While children can benefit from playing with such technologies, these systems lack the situated and embodied interactions that young children naturally engage in and learn from (Glenberg, <xref ref-type="bibr" rid="B14">2010</xref>; Leyzberg et al., <xref ref-type="bibr" rid="B25">2012</xref>). Social robots represent an emerging technology that provides situatedness and embodiment, and thus have potential benefits for educational purposes. In essence, social robots are autonomous physical agents, often with human-like feature, that can interact socially with humans in a semi-natural way for prolonged periods of time (Dautenhahn, <xref ref-type="bibr" rid="B7">2007</xref>). The use of social robots, in comparison to more traditional digital technologies, allows for the development of tutoring systems more akin to human tutors, especially with respect to the situated and embodied social interactions between child and robot. Thus, this offers the opportunity to design robots such that they interact in a way that optimizes the child&#x02019;s language learning.</p>
<p>Recently, an increasing interest has emerged to develop social robots to support children with learning a second language (Kanda et al., <xref ref-type="bibr" rid="B21">2004</xref>; Belpaeme et al., <xref ref-type="bibr" rid="B3">2015</xref>; Kennedy et al., <xref ref-type="bibr" rid="B22">2016</xref>). While a social robot cannot provide tutoring to the level humans can, recent studies suggest that using social robots can result in an increased learning gain compared to digital learning environments for tablets or computers (Han et al., <xref ref-type="bibr" rid="B16">2008</xref>; Leyzberg et al., <xref ref-type="bibr" rid="B25">2012</xref>). It is, however, unclear why this is the case. Perhaps the physical presence of the robot draws the attention of children for longer periods of time, but the embodiment and situatedness of the learning environment perhaps also helps the children to ground the language more strongly than interactions with virtual objects do.</p>
<p>While there is a fair body of research on robot tutors, a comprehensive description of the design features for a second language robot tutor based on what is known about children&#x02019;s language acquisition is lacking. What are the design features of child-robot interactions that would support second language learning? And, to what extent can these interactions be implemented in today&#x02019;s social robot technologies? In this perspective article, we try to answer these questions based on theoretical accounts from the literature on children&#x02019;s language acquisition in combination with our own experiences in designing a tutor robot.</p>
</sec>
<sec id="s2">
<title>Designing Child-Robot Interactions</title>
<p>In our project, we aim to design a digital learning environment in which preschool children interact one-on-one with a social robot that supports either their learning of English as a foreign language, or the school language for those children who have a different native language (Belpaeme et al., <xref ref-type="bibr" rid="B3">2015</xref>). In particular, the project aims to develop a series of tutoring sessions revolving around three increasingly complex domains (numbers, spatial relations and mental vocabulary). In each session, the child will engage with the robot (a Softbank Robotics NAO robot) in a game-like scenario focusing on learning a small number of target words. The contextual setting is generally displayed on a tablet computer that occasionally also provides some verbal support, however, the robot acts as the interactive tutor. Below we discuss the design features and considerations that we believe are crucial to design a successful tutoring system.</p>
<sec id="s2-1">
<title>Peer-Like Tutoring</title>
<p>One of the first questions that comes up when designing a robot tutor is whether the robot should take the role of a teacher or a peer. Research on children&#x02019;s language acquisition has demonstrated that children learn more effectively from an adult who can use well-defined pedagogical methods for teaching children using clear directions, explanations and positive feedback methods (Matthews et al., <xref ref-type="bibr" rid="B30">2007</xref>). However, designing and framing the robot as an adult tutor has the disadvantage that children will form expectations about the robot&#x02019;s behavior and proficiency that cannot be met with current technology (Kennedy et al., <xref ref-type="bibr" rid="B23">2015</xref>). Due to technological limitations of the robot and underlying software, communication breakdowns are more likely to occur than with a human. For a peer robot introduced as a fellow language learner, breakdowns in communication are more acceptable. Moreover, interacting with robots acting as peers is conceived as more fun (Kanda et al., <xref ref-type="bibr" rid="B21">2004</xref>), allows for learning-by-teaching (Tanaka and Matsuzoe, <xref ref-type="bibr" rid="B34">2012</xref>) and has a proven to be efficient in teaching children how to write (Hood et al., <xref ref-type="bibr" rid="B19">2015</xref>). Furthermore, there is some evidence that children&#x02019;s learning can benefit from interacting with peers (Mashburn et al., <xref ref-type="bibr" rid="B29">2009</xref>). Given these considerations, we believe it is desirable to frame or introduce the robot as a peer and friend, yet design its interactions insofar possible based on pedagogically well-established strategies to scaffold language learning.</p>
</sec>
<sec id="s2-2">
<title>First Impressions</title>
<p>To implement effective tutoring, the robot needs to interact with children in multiple sessions, so they have to be motivated to engage in long-term interactions with the robot. Establishing common ground between child and robot can contribute to this (Kanda et al., <xref ref-type="bibr" rid="B21">2004</xref>), but first impressions to establish trust and rapport are also crucial (Hancock et al., <xref ref-type="bibr" rid="B17">2011</xref>).</p>
<p>Despite the wealth of studies regarding the introduction of entertainment robots as toys to children (e.g., Lund, <xref ref-type="bibr" rid="B28">2003</xref>), surprisingly little research has been conducted on designing protocols on how to introduce a robot tutor to a group of preschool children. Fridin (<xref ref-type="bibr" rid="B11">2014</xref>) presents one exception, and found that introducing a robot tutor to children in group sessions improved subsequent interactions compared to introducing the robot to children in individual sessions. Another study by Westlund et al. (<xref ref-type="bibr" rid="B38">2016</xref>) found that the way a robot is framed, either as a machine or a social entity, affected the way children later engaged with the robot. They concluded that introducing the robot as a machine could create a more distant relation between child and robot, thus reducing acceptance. We therefore decided to frame the robot in our project as a social playmate for the children and introduced the robot in a group session. However, the NAO robot is slightly taller and more rigid than the fluffy huggable Tega robot, which Westlund et al. (<xref ref-type="bibr" rid="B38">2016</xref>) used, and we observed that some 3-year-old children were somewhat intimidated by the NAO robot on their first encounter. Such a first impression of the robot could reduce the trust that the child had for the robot, which could negatively affect their willingness to interact with the robot in the short-term, but also in the long-term. To develop a successful first encounter and to build trust between the child and robot, we designed the following strategy for introducing the robot to 3-year-old children at their preschool.</p>
<p>Pilot studies revealed that some children got anxious when the robot was introduced and then suddenly started to move. To familiarize children prior to their first encounter with the robot, it is therefore advisable to prepare them well. For our study, we sent coloring pages of the robot to the preschools during recruitment and asked the pedagogical assistants to talk a little bit about the robots to the children. About 1 week before the experimental trials, the experimenters introduced the robot in class during their daily &#x0201C;circle time&#x0201D;, as this provided a safe and familiar environment with the whole group in which the pedagogical assistants usually introduce new topics or new activities. One experimenter first introduced the robot by telling a story about Robin, the name of our robot, using a makeshift picture book. In this story we explained the similarities and dissimilarities between the robot and children to construct the type of common ground considered to have a positive effect on the learning outcome (Kanda et al., <xref ref-type="bibr" rid="B21">2004</xref>). For example, we told that Robin enjoys dancing and wants to meet new friends, and even though he does not have a mouth and because of that cannot smile, he can smile using his eye LEDs.</p>
<p>After this story, another experimenter entered the room with the robot while it was actively looking at faces to provide an animate feeling. The robot introduced itself with a small story about itself and by performing a dance in which the children were encouraged to participate. The end of the circle time consisted of getting a blanket for the robot so it could &#x0201C;sleep&#x0201D;. This introduction was repeated later on the days we conducted the experiment in one-on-one sessions. While by then most children were comfortable interacting with the robot, some were still timid and anxious. To encourage these children to feel comfortable, one of the experiment leaders would sit next to the child during the warm-up phase of the experiment and motivate the child to respond to the robot when necessary until the child was sufficiently comfortable to interact with the robot by herself/himself. We found that the younger 3-year olds required more support from the experimenters than the older 3-year olds (Baxter et al., <xref ref-type="bibr" rid="B2">2017</xref>). Although we are still analyzing the experiments, preliminary findings suggest that our introduction helped children to build trust and common ground with the robot effectively.</p>
</sec>
<sec id="s2-3">
<title>Temporal Contingency</title>
<p>Research has shown that it is crucial for children&#x02019;s language development that their communication bids are responded to in a temporally contingent manner (Bornstein et al., <xref ref-type="bibr" rid="B4">2008</xref>; McGillion et al., <xref ref-type="bibr" rid="B31">2013</xref>). This, however, faces a technological challenge. While adults tend to take over turns very rapidly, robots require relatively long processing time to produce a response. Nevertheless, in our first experiment (de Haas et al., <xref ref-type="bibr" rid="B9">2016</xref>), we observed that children were at first surprised by the delayed responses, but quickly adapted to the robot and waited patiently for a response. Perhaps this is because children also require longer than adults to take turns (Garvey and Berninger, <xref ref-type="bibr" rid="B13">1981</xref>) and having framed the robot as a peer children made the delays more plausible or expected. Nevertheless, while a lag in temporal contingency may not harm the interaction with children, it may harm learning. One way to remedy this may be to have the robot start responding by providing a backchannel signal, such as &#x0201C;uhm&#x0201D; to indicate the robot is (still) taking his turn, but requires more time to process (Clark, <xref ref-type="bibr" rid="B6">1996</xref>).</p>
</sec>
<sec id="s2-4">
<title>Semantic Contingency</title>
<p>Robots should not only respond to children in a timely fashion, but also in a semantically contingent fashion (i.e., consistent with the child&#x02019;s focus of attention), as this too has a positive effect on children&#x02019;s language acquisition (Bornstein et al., <xref ref-type="bibr" rid="B4">2008</xref>; McGillion et al., <xref ref-type="bibr" rid="B31">2013</xref>). For instance, research has shown that by responding in a semantically contingent manner, either verbally or by following children&#x02019;s gaze, (joint) attention is sustained for a longer duration (Yu and Smith, <xref ref-type="bibr" rid="B39">2016</xref>), allowing children to learn more about a situation. To achieve semantically contingent responses, the robot should be able to understand the child&#x02019;s communication bids, construct joint attention with the child, or at least identify what the child is attending to. Monitoring children&#x02019;s behavior and establishing joint attention are therefore considered crucial for designing a successful robot tutor.</p>
</sec>
<sec id="s2-5">
<title>Monitoring Children&#x02019;s Behavior</title>
<p>To understand children&#x02019;s communication bids, as well as to test their pronunciation of the L2, it is important that the robot be equipped with well-functioning automatic speech recognition (ASR). However, the performance of state-of-the-art ASR for children is still suboptimal, especially for preschool-aged children (Fringi et al., <xref ref-type="bibr" rid="B12">2015</xref>; Kennedy et al., <xref ref-type="bibr" rid="B24">2017</xref>). Reasons for this include that children&#x02019;s pronunciation is often flawed and that their speech has a different pitch than adults. Moreover, relatively little research has been carried out in this domain and not much data exist to train ASR on. While it can be expected that the performance of ASR for children will improve in the not too distant future (Liao et al., <xref ref-type="bibr" rid="B26">2015</xref>), until then alternative strategies need to be developed that do not (exclusively) rely on ASR.</p>
<p>In our project, we explore various strategies to achieve this, both based on monitoring non-verbal behaviors of the children and focusing on comprehending rather than producing L2. The first strategy relies on providing children tasks they have to perform in the learning environment, such as placing &#x0201C;a toy cow behind a tree&#x0201D; when teaching spatial language. This, however, requires the visual object recognition on the robot to work well, which is only the case when the scene contains a limited set of distinctively recognizable objects, such as distinctly colored objects (Nguyen et al., <xref ref-type="bibr" rid="B32">2015</xref>). A potential solution explored in our project is to use objects with build-in RFID sensors that can be tracked automatically. The second solution we explore is to use a touch screen tablet that displays scenes the child can manipulate, which not only has the advantage of avoiding the problem of object recognition, but also allows us to control the robot&#x02019;s responses and vary the scenes in real time. A downside, however, is that it takes away the 3-dimensional physical aspect of embodied cognition that would help the children to better entrench what they learn (Glenberg, <xref ref-type="bibr" rid="B14">2010</xref>). Currently, experiments are underway to investigate the effect of using real vs. virtual objects. These solutions not only aid in understanding the child&#x02019;s communication bids, it also helps in identifying their attention and can thus contribute to establishing joint attention.</p>
</sec>
<sec id="s2-6">
<title>Joint Attention and Gestures</title>
<p>Joint attention, where interlocutors attend on the same referent, is a form of social interaction that has been shown to support children&#x02019;s language learning (Tomasello and Farrar, <xref ref-type="bibr" rid="B36">1986</xref>). One way to establish joint attention with a child is to guide their attention to a referent using gestures, such as pointing or iconic gestures. The ability to produce gestures in the real world is potentially one of the main advantages of using physical robots as opposed to virtual agents, who may have a harder time to establish joint attention. However, many robots&#x02019; physical morphologies do not correspond one-to-one to the human body. Hence, many human gestures cannot be translated directly to robot gestures. For instance, the NAO robot that we use in our research has a hand with three fingers that cannot be controlled independently, so index finger pointing cannot be achieved (see Figure <xref ref-type="fig" rid="F1">1</xref>). Will children still recognize NAO&#x02019;s arm extension as a pointing gesture? And if so, will they be able to identify the object the robot refers to? We are currently running an experiment to investigate how NAO&#x02019;s pointing gestures are perceived, and preliminary findings show that participants have difficulty identifying the referred object on a small tablet screen. Similar issues arise when developing other gestures. One of the other non-verbal behaviors we are using is the coloring of NAO&#x02019;s eye LEDSs to indicate the robot&#x02019;s happiness as a form of positive feedback, since the robot cannot smile with its mouth.</p>
<fig id="F1" position="float">
<label>Figure 1</label>
<caption><p><bold>NAO pointing to a block with three fingers</bold>. (Note that written, informed consent was obtained from the parents of the child for the publication of this image).</p></caption>
<graphic xlink:href="fnhum-11-00073-g0001.tif"/>
</fig>
</sec>
<sec id="s2-7">
<title>Feedback</title>
<p>Feedback, too, is an interactional feature known to help language learning (Matthews et al., <xref ref-type="bibr" rid="B30">2007</xref>; Ate&#x0015F;-&#x0015E;en and K&#x000FC;ntay, <xref ref-type="bibr" rid="B1">2015</xref>). The question is how should the robot provide feedback, such that it is both pleasant and effective for learning? While adults provide positive feedback explicitly, they usually provide negative feedback implicitly by reformulating children&#x02019;s errors in the correct form. In child-child interactions, however, Long (<xref ref-type="bibr" rid="B27">2006</xref>) found that there was a clear advantage in learning from explicit negative feedback (e.g., by saying &#x0201C;no, that&#x02019;s wrong, you need to say &#x02018;he ran&#x0201D;&#x02019;) when compared to reformulating feedback (the learner says &#x0201C;he runned&#x0201D; and the teacher reacts with &#x0201C;he ran&#x0201D;).</p>
<p>To investigate how children experience feedback from a peer robot, we carried out an experiment among 85 3-year-old Dutch-speaking children at preschools in Netherlands (de Haas et al., <xref ref-type="bibr" rid="B9">2016</xref>, <xref ref-type="bibr" rid="B8">2017</xref>). In this experiment, the children interacted with a NAO robot during which they received a short lesson on how to count from 1 to 4 in English. After a short training phase, in which the children were presented with the four counting words twice in relation to body parts and wooden blocks, they were given instructions by the robot to pick up a given number of blocks. While the instructions were given in their native language, the numbers were uttered in English. In response to the child&#x02019;s ability to achieve the task, the robot provided feedback. The experiment followed a between-subjects design with three conditions: adult-like feedback (explicit positive and implicit negative), peer-like feedback (no positive and explicit negative) and no feedback. We did not find significant differences in learning gain between the conditions, probably because the target words were insufficiently often repeated. However, we explored the way in which the children engaged with the robot after they received feedback and we found that children looked less often at the experimenter in the feedback conditions than in the no feedback condition. Further analyses are carried out to evaluate how the children responded to the various forms of feedback to find out what type of feedback would be most effective for achieving both acceptable and effective tutoring interactions.</p>
</sec>
<sec id="s2-8">
<title>Zone of Proximity and Adaptivity</title>
<p>Finally, from a pedagogical point of view it is desirable that the interactions between child and robot be sufficiently challenging and varied so that the child has a target to learn from, but at the same time interactions should not be too difficult, because that may frustrate the child causing it to lose interest in the robot (Charisi et al., <xref ref-type="bibr" rid="B5">2016</xref>). In other words, the robot should remain in Vygotsky&#x02019;s Zone of Proximity that supports an effective learning environment (Vygotsky, <xref ref-type="bibr" rid="B37">1978</xref>). In order to achieve this, the robot should be able to keep track of the children&#x02019;s advancements in language learning and perhaps their emotional states during the tutoring sessions, and adapt to these. While the former can be monitored as discussed previously, it may be possible to detect emotional states known to influence learning (e.g., concentration, confusion, frustration and boredom) using methods from affective computing (D&#x02019;Mello and Graesser, <xref ref-type="bibr" rid="B10">2012</xref>). Using this type of information, it is possible to adapt the tutoring sessions by either reducing or increasing the number of repetitions, and/or change the subject (Schodde et al., <xref ref-type="bibr" rid="B33">2017</xref>).</p>
</sec>
</sec>
<sec sec-type="conclusion" id="s3">
<title>Conclusion</title>
<p>This perspective article presented some design features that we consider crucial for developing a social robot as an effective second language tutor. We believe the robot is most effective when it is framed as a peer, i.e., as a fellow language learner and playmate, but that is designed to use adult-like interaction strategies to optimize learning efficacy. In order to establish common ground and trust to facilitate long-term interactions, we consider it essential that the robot be introduced with appropriate care on the first encounter. As an example, we outlined our strategy for introducing a robot to preschool children. Interactions between child and robot should be contingent and multimodal, and provide appropriate forms of feedback. We argued that the robot should remain within Vygotsky (<xref ref-type="bibr" rid="B37">1978</xref>) Zone of Proximal Development and thus should adapt to the individual level of the child.</p>
<p>We also discussed some technical challenges that need to be solved in order to implement contingent interactions; the most important of which we believe is ASR, which presently does not work well for children&#x02019;s speech. While various technical challenges still remain, we expect that social robots will provide effective digital technologies to support second language development in the years to come.</p>
<p>The present list of design features covers many aspects that need to be considered when developing a tutor robot, but it is not yet comprehensive. One aspect that has not been covered, for instance, concerns the design of robots for children from different cultures, which could require different design choices (Shahid et al., <xref ref-type="bibr" rid="B330">2014</xref>). For example, in some cultures education is more teaching-centered (Hofstede, <xref ref-type="bibr" rid="B180">1986</xref>) and thus designing the tutor as a peer robot may be less effective or acceptable (Tazhigaliyeva et al., <xref ref-type="bibr" rid="B35">2016</xref>). Concluding, this perspective article offers only a first step towards a comprehensive list of design features for tutor robots and additional research is needed to complete and optimize the list.</p>
</sec>
<sec id="s4">
<title>Ethics Statement</title>
<p>The Research Ethics Committee of Tilburg School of Humanities approved this study, and the parents of all participating children gave written informed consent in accordance with the Declaration of Helsinki.</p>
</sec>
<sec id="s5">
<title>Author Contributions</title>
<p>PV, MH and EK designed the conceptual aspects of the article; PV, MH, CJ and PB carried out the literature review; PV, EK and MH designed the feedback study; MH, CJ and PB designed the introduction study; MH, CJ and PB carried out the studies; PV and MH wrote the article; CJ, PB and EK revised the article critically.</p>
</sec>
<sec id="s6">
<title>Funding</title>
<p>This work has been supported by the EU H2020 L2TOR project (grant 688014). CJ and PB thank the research trainee program of the Tilburg School of Humanities for their support.</p>
</sec>
<sec id="s7">
<title>Conflict of Interest Statement</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
</body>
<back>
<ack>
<p>The authors wish to thank all members of the L2TOR project for their support and advice regarding this research. We also thank Kinderopvanggroep Tilburg and all participating daycare centers and preschools for their assistance in this research. Finally, a big thank you to all the children and their parents for participating in our research.</p>
</ack>
<ref-list>
<title>References</title>
<ref id="B1"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Ate&#x0015F;-&#x0015E;en</surname> <given-names>B. A.</given-names></name> <name><surname>K&#x000FC;ntay</surname> <given-names>A. C.</given-names></name></person-group> (<year>2015</year>). <article-title>Children&#x02019;s sensitivity to caregiver cues and the role of adult feedback in the development of referential communication</article-title>. <source>The Acquisition of Reference</source>, eds <person-group person-group-type="editor"><name><surname>Serratrice</surname> <given-names>L.</given-names></name> <name><surname>Allen</surname> <given-names>S. E. M.</given-names></name></person-group> (<publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins</publisher-name>), <fpage>241</fpage>&#x02013;<lpage>262</lpage>.</citation></ref>
<ref id="B2"><citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Baxter</surname> <given-names>P.</given-names></name> <name><surname>De Jong</surname> <given-names>C.</given-names></name> <name><surname>Aarts</surname> <given-names>A.</given-names></name> <name><surname>de Haas</surname> <given-names>M.</given-names></name> <name><surname>Vogt</surname> <given-names>P.</given-names></name></person-group> (<year>2017</year>). &#x0201C;<article-title>The effect of age on engagement in preschoolers&#x02019; child-robot interactions</article-title>,&#x0201D; in <source>Companion proceedings of the 12th Annual ACM International Conference on Human-Robot Interaction</source> (HRI&#x02019;17), (Vienna, Austria).</citation></ref>
<ref id="B3"><citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Belpaeme</surname> <given-names>T.</given-names></name> <name><surname>Kennedy</surname> <given-names>J.</given-names></name> <name><surname>Baxter</surname> <given-names>P.</given-names></name> <name><surname>Vogt</surname> <given-names>P.</given-names></name> <name><surname>Krahmer</surname> <given-names>E. E. J.</given-names></name> <name><surname>Kopp</surname> <given-names>S.</given-names></name> <etal/></person-group>. (<year>2015</year>). &#x0201C;<article-title>L2TOR-second language tutoring using social robots</article-title>,&#x0201D; in <source>First Workshop on Educational Robots</source> (WONDER), (Paris, France).</citation></ref>
<ref id="B4"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bornstein</surname> <given-names>M. H.</given-names></name> <name><surname>Tamis-LeMonda</surname> <given-names>C. S.</given-names></name> <name><surname>Hahn</surname> <given-names>C. S.</given-names></name> <name><surname>Haynes</surname> <given-names>O. M.</given-names></name></person-group> (<year>2008</year>). <article-title>Maternal responsiveness to young children at three ages: longitudinal analysis of a multidimensional, modular and specific parenting construct</article-title>. <source>Dev. Psychol.</source> <volume>44</volume>, <fpage>867</fpage>&#x02013;<lpage>874</lpage>. <pub-id pub-id-type="doi">10.1037/0012-1649.44.3.867</pub-id><pub-id pub-id-type="pmid">18473650</pub-id></citation></ref>
<ref id="B5"><citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Charisi</surname> <given-names>V.</given-names></name> <name><surname>Davison</surname> <given-names>D.</given-names></name> <name><surname>Reidsma</surname> <given-names>D.</given-names></name> <name><surname>Evers</surname> <given-names>V.</given-names></name></person-group> (<year>2016</year>). &#x0201C;<article-title>Children and robots: a preliminary review of methodological approaches in learning settings</article-title>,&#x0201D; in <source>2nd Workshop on Evaluating Child Robot Interaction - HRI</source>, (Christchurch, New Zealand).</citation></ref>
<ref id="B6"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Clark</surname> <given-names>H. H.</given-names></name></person-group> (<year>1996</year>). <source>Using Language.</source> <publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>.</citation></ref>
<ref id="B7"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dautenhahn</surname> <given-names>K.</given-names></name></person-group> (<year>2007</year>). <article-title>Socially intelligent robots: dimensions of human-robot interaction</article-title>. <source>Philos. Trans. R. Soc. B Biol. Sci.</source> <volume>1480</volume>, <fpage>679</fpage>&#x02013;<lpage>704</lpage>. <pub-id pub-id-type="doi">10.1098/rstb.2006.2004</pub-id><pub-id pub-id-type="pmid">17301026</pub-id></citation></ref>
<ref id="B8"><citation citation-type="confproc"><person-group person-group-type="author"><name><surname>de Haas</surname> <given-names>M.</given-names></name> <name><surname>Baxter</surname> <given-names>P.</given-names></name> <name><surname>de Jong</surname> <given-names>C.</given-names></name> <name><surname>Vogt</surname> <given-names>P.</given-names></name> <name><surname>Krahmer</surname> <given-names>E.</given-names></name></person-group> (<year>2017</year>). &#x0201C;<article-title>Exploring different types of feedback in preschooler and robot interaction</article-title>,&#x0201D; in <source>Companion proceedings of the 12th Annual ACM International Conference on Human-Robot Interaction</source> (HRI&#x02019;17), (Vienna, Austria).</citation></ref>
<ref id="B9"><citation citation-type="confproc"><person-group person-group-type="author"><name><surname>de Haas</surname> <given-names>M.</given-names></name> <name><surname>Vogt</surname> <given-names>P.</given-names></name> <name><surname>Krahmer</surname> <given-names>E. J.</given-names></name></person-group> (<year>2016</year>). &#x0201C;<article-title>Enhancing child-robot tutoring interactions with appropriate feedback</article-title>,&#x0201D; in <source>Proceedings of First Workshop on Long-Term Child-Robot Interaction. IEEE Ro-Man</source>, (<conf-loc>New York, NY</conf-loc>).</citation></ref>
<ref id="B10"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>D&#x02019;Mello</surname> <given-names>S.</given-names></name> <name><surname>Graesser</surname> <given-names>A.</given-names></name></person-group> (<year>2012</year>). <article-title>Dynamics of affective states during complex learning</article-title>. <source>Learn. Instr.</source> <volume>22</volume>, <fpage>145</fpage>&#x02013;<lpage>157</lpage>. <pub-id pub-id-type="doi">10.1016/j.learninstruc.2011.10.001</pub-id></citation></ref>
<ref id="B11"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fridin</surname> <given-names>M.</given-names></name></person-group> (<year>2014</year>). <article-title>Kindergarten social assistive robot: first meeting and ethical issues</article-title>. <source>Comput. Hum. Behav.</source> <volume>30</volume>, <fpage>262</fpage>&#x02013;<lpage>272</lpage>. <pub-id pub-id-type="doi">10.1016/j.chb.2013.09.005</pub-id></citation></ref>
<ref id="B12"><citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Fringi</surname> <given-names>E.</given-names></name> <name><surname>Lehman</surname> <given-names>J.</given-names></name> <name><surname>Russell</surname> <given-names>M. J.</given-names></name></person-group> (<year>2015</year>). &#x0201C;<article-title>Evidence of phonological processes in automatic recognition of children&#x02019;s speech</article-title>,&#x0201D; in <source>16th Annual Conference of the International Speech Communication Association</source>, (Dresden, Germany), <fpage>1621</fpage>&#x02013;<lpage>1624</lpage>. </citation></ref>
<ref id="B13"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Garvey</surname> <given-names>C.</given-names></name> <name><surname>Berninger</surname> <given-names>G.</given-names></name></person-group> (<year>1981</year>). <article-title>Timing and turn taking in children&#x02019;s conversations</article-title>. <source>Discourse Process.</source> <volume>4</volume>, <fpage>27</fpage>&#x02013;<lpage>57</lpage>. <pub-id pub-id-type="doi">10.1080/01638538109544505</pub-id></citation></ref>
<ref id="B14"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Glenberg</surname> <given-names>A. M.</given-names></name></person-group> (<year>2010</year>). <article-title>Embodiment as a unifying perspective for psychology</article-title>. <source>Wiley Interdiscip. Rev. Cogn. Sci.</source> <volume>4</volume>, <fpage>586</fpage>&#x02013;<lpage>596</lpage>. <pub-id pub-id-type="doi">10.1002/wcs.55</pub-id><pub-id pub-id-type="pmid">26271505</pub-id></citation></ref>
<ref id="B15"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Golonka</surname> <given-names>E. M.</given-names></name> <name><surname>Bowles</surname> <given-names>A. R.</given-names></name> <name><surname>Frank</surname> <given-names>V. M.</given-names></name> <name><surname>Richardson</surname> <given-names>D. L.</given-names></name> <name><surname>Freynik</surname> <given-names>S.</given-names></name></person-group> (<year>2014</year>). <article-title>Technologies for foreign language learning: a review of technology types and their effectiveness</article-title>. <source>Comput. Assist. Lang. Learn.</source> <volume>27</volume>, <fpage>70</fpage>&#x02013;<lpage>105</lpage>. <pub-id pub-id-type="doi">10.1080/09588221.2012.700315</pub-id></citation></ref>
<ref id="B16"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Han</surname> <given-names>J. H.</given-names></name> <name><surname>Jo</surname> <given-names>M. H.</given-names></name> <name><surname>Jones</surname> <given-names>V.</given-names></name> <name><surname>Jo</surname> <given-names>J. H.</given-names></name></person-group> (<year>2008</year>). <article-title>Comparative study on the educational use of home robots for children</article-title>. <source>J. Inf. Process. Syst.</source> <volume>4</volume>, <fpage>159</fpage>&#x02013;<lpage>168</lpage>. <pub-id pub-id-type="doi">10.3745/jips.2008.4.4.159</pub-id></citation></ref>
<ref id="B17"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hancock</surname> <given-names>P. A.</given-names></name> <name><surname>Billings</surname> <given-names>D. R.</given-names></name> <name><surname>Schaefer</surname> <given-names>K. E.</given-names></name> <name><surname>Chen</surname> <given-names>J. Y. C.</given-names></name> <name><surname>de Visser</surname> <given-names>E. J.</given-names></name> <name><surname>Parasuraman</surname> <given-names>R.</given-names></name></person-group> (<year>2011</year>). <article-title>A meta-analysis of factors affecting trust in human-robot interaction</article-title>. <source>Hum. Factors</source> <volume>53</volume>, <fpage>517</fpage>&#x02013;<lpage>527</lpage>. <pub-id pub-id-type="doi">10.1177/0018720811417254</pub-id><pub-id pub-id-type="pmid">22046724</pub-id></citation></ref>
<ref id="B18"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hoff</surname> <given-names>E.</given-names></name></person-group> (<year>2013</year>). <article-title>Interpreting the early language trajectories of children from low-SES and language minority homes: implications for closing achievement gaps</article-title>. <source>Dev. Psychol.</source> <volume>49</volume>, <fpage>4</fpage>&#x02013;<lpage>14</lpage>. <pub-id pub-id-type="doi">10.1037/a0027238</pub-id><pub-id pub-id-type="pmid">22329382</pub-id></citation></ref>
<ref id="B180"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hofstede</surname> <given-names>G.</given-names></name></person-group> (<year>1986</year>). <article-title>Cultural differences in teaching and learning</article-title>. <source>Int. J. Intercult. Relat.</source> <volume>10</volume>, <fpage>301</fpage>&#x02013;<lpage>320</lpage>. <pub-id pub-id-type="doi">10.1016/0147-1767(86)90015-5</pub-id><pub-id pub-id-type="pmid">22329382</pub-id></citation></ref>
<ref id="B19"><citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Hood</surname> <given-names>D.</given-names></name> <name><surname>Lemaignan</surname> <given-names>S.</given-names></name> <name><surname>Dillenbourg</surname> <given-names>P.</given-names></name></person-group> (<year>2015</year>). &#x0201C;<article-title>When children teach a robot to write: an autonomous teachable humanoid which uses simulated handwriting</article-title>,&#x0201D; in <source>Proceedings of the Tenth Annual ACM/IEEE International Conference on Human-Robot Interaction</source>, (<conf-loc>New York, NY</conf-loc>), <fpage>83</fpage>&#x02013;<lpage>90</lpage>.</citation></ref>
<ref id="B20"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hsin</surname> <given-names>C.-T.</given-names></name> <name><surname>Li</surname> <given-names>M.-C.</given-names></name> <name><surname>Tsai</surname> <given-names>C.-C.</given-names></name></person-group> (<year>2014</year>). <article-title>The influence of young children&#x02019;s use of technology on their learning: A</article-title>. <source>Educ. Technol. Soc.</source> <volume>17</volume>, <fpage>85</fpage>&#x02013;<lpage>99</lpage>. Available online at: <ext-link ext-link-type="uri" xlink:href="https://eric.ed.gov/?id=EJ1045554">https://eric.ed.gov/?id=EJ1045554</ext-link></citation></ref>
<ref id="B21"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kanda</surname> <given-names>T.</given-names></name> <name><surname>Hirano</surname> <given-names>T.</given-names></name> <name><surname>Eaton</surname> <given-names>D.</given-names></name> <name><surname>Ishiguro</surname> <given-names>H.</given-names></name></person-group> (<year>2004</year>). <article-title>Interactive robots as social partners and peer tutors for children: a field trial</article-title>. <source>Hum. Comput. Interact.</source> <volume>19</volume>, <fpage>61</fpage>&#x02013;<lpage>84</lpage>. <pub-id pub-id-type="doi">10.1207/s15327051hci1901&#x00026;2_4</pub-id></citation></ref>
<ref id="B23"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Kennedy</surname> <given-names>J.</given-names></name> <name><surname>Baxter</surname> <given-names>P.</given-names></name> <name><surname>Senft</surname> <given-names>E.</given-names></name> <name><surname>Belpaeme</surname> <given-names>T.</given-names></name></person-group> (<year>2015</year>). &#x0201C;<article-title>Higher nonverbal immediacy leads to greater learning gains in child-robot tutoring interactions</article-title>,&#x0201D; in <source>International Conference on Social Robotics</source>, eds <person-group person-group-type="editor"><name><surname>Tapus</surname> <given-names>A.</given-names></name> <name><surname>Andr&#x000E9;</surname> <given-names>E.</given-names></name> <name><surname>Martin</surname> <given-names>J.-C.</given-names></name> <name><surname>Ferland</surname> <given-names>F.</given-names></name> <name><surname>Ammi</surname> <given-names>M.</given-names></name></person-group> (<publisher-loc>New York, NY</publisher-loc>: <publisher-name>Springer International Publishing</publisher-name>), <fpage>327</fpage>&#x02013;<lpage>336</lpage>.</citation></ref>
<ref id="B22"><citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Kennedy</surname> <given-names>J.</given-names></name> <name><surname>Baxter</surname> <given-names>P.</given-names></name> <name><surname>Senft</surname> <given-names>E.</given-names></name> <name><surname>Belpaeme</surname> <given-names>T.</given-names></name></person-group> (<year>2016</year>). &#x0201C;<article-title>Social robot tutoring for child second language learning</article-title>,&#x0201D; in <source>Proceedings of the 11th Annual ACM/IEEE International Conference on Human-Robot Interaction</source> (HRI&#x02019;16), (<conf-loc>Christchurch, New Zealand</conf-loc>), <fpage>231</fpage>&#x02013;<lpage>238</lpage>.</citation></ref>
<ref id="B24"><citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Kennedy</surname> <given-names>J.</given-names></name> <name><surname>Lemaignan</surname> <given-names>S.</given-names></name> <name><surname>Montassier</surname> <given-names>C.</given-names></name> <name><surname>Lavalade</surname> <given-names>P.</given-names></name> <name><surname>Irfan</surname> <given-names>B.</given-names></name> <name><surname>Papadopoulos</surname> <given-names>F.</given-names></name> <etal/></person-group>. (<year>2017</year>). &#x0201C;<article-title>Child speech recognition in human-robot interaction: evaluations and recommendations</article-title>,&#x0201D; in <source>Proceedings of the 12th Annual ACM International Conference on Human-Robot Interaction</source> (HRI&#x02019;17), (<conf-loc>Vienna, Austria</conf-loc>).</citation></ref>
<ref id="B25"><citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Leyzberg</surname> <given-names>D.</given-names></name> <name><surname>Spaulding</surname> <given-names>S.</given-names></name> <name><surname>Toneva</surname> <given-names>M.</given-names></name> <name><surname>Scassellati</surname> <given-names>B.</given-names></name></person-group> (<year>2012</year>). &#x0201C;<article-title>The physical presence of a robot tutor increases cognitive learning gains</article-title>,&#x0201D; in <source>Proceedings of the 34th Annual Conference of the Cognitive Science Society</source>, (Sapporo, Japan), <fpage>1882</fpage>&#x02013;<lpage>1887</lpage>.</citation></ref>
<ref id="B26"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Liao</surname> <given-names>H.</given-names></name> <name><surname>Pundak</surname> <given-names>G.</given-names></name> <name><surname>Siohan</surname> <given-names>O.</given-names></name> <name><surname>Carroll</surname> <given-names>M. K.</given-names></name> <name><surname>Coccaro</surname> <given-names>N.</given-names></name> <name><surname>Jiang</surname> <given-names>Q. M.</given-names></name> <etal/></person-group>. (<year>2015</year>). &#x0201C;<article-title>Large vocabulary automatic speech recognition for children</article-title>,&#x0201D; in <source>Interspeech</source>, (Dresden, Germany), <fpage>1611</fpage>&#x02013;<lpage>1615</lpage>.</citation></ref>
<ref id="B27"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Long</surname> <given-names>M. H.</given-names></name></person-group> Ed. (<year>2006</year>). &#x0201C;<article-title>Recasts in SLA: the story so far</article-title>,&#x0201D; in <source>Problems in SLA. Second Language Acquisition Research Series</source>, (<publisher-loc>Mahwah, NJ</publisher-loc>: <publisher-name>Lawrence Erlbaum Associates</publisher-name>), <fpage>75</fpage>&#x02013;<lpage>116</lpage>.</citation></ref>
<ref id="B28"><citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Lund</surname> <given-names>H. H.</given-names></name></person-group> (<year>2003</year>). &#x0201C;<article-title>Adaptive robotics in the entertainment industry</article-title>,&#x0201D; in <source>Proceedings of IEEE International Symposium on Computational Intelligence in Robotics and Automation (cira2003)</source>, (Vol. 2), (Kobe, Japan), <fpage>595</fpage>&#x02013;<lpage>602</lpage>.</citation></ref>
<ref id="B29"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mashburn</surname> <given-names>A. J.</given-names></name> <name><surname>Justice</surname> <given-names>L. M.</given-names></name> <name><surname>Downer</surname> <given-names>J. T.</given-names></name> <name><surname>Pianta</surname> <given-names>R. C.</given-names></name></person-group> (<year>2009</year>). <article-title>Peer effects on children&#x02019;s language achievement during pre-kindergarten,</article-title>. <source>Child Dev.</source> <volume>80</volume>, <fpage>686</fpage>&#x02013;<lpage>702</lpage>. <pub-id pub-id-type="doi">10.1111/j.1467-8624.2009.01291.x</pub-id><pub-id pub-id-type="pmid">19489897</pub-id></citation></ref>
<ref id="B30"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Matthews</surname> <given-names>D.</given-names></name> <name><surname>Lieven</surname> <given-names>E.</given-names></name> <name><surname>Tomasello</surname> <given-names>M.</given-names></name></person-group> (<year>2007</year>). <article-title>How toddlers and preschoolers learn to uniquely identify referents for others: a training study</article-title>. <source>Child Dev.</source> <volume>6</volume>, <fpage>1744</fpage>&#x02013;<lpage>1759</lpage>. <pub-id pub-id-type="doi">10.1111/j.1467-8624.2007.01098.x</pub-id><pub-id pub-id-type="pmid">17988318</pub-id></citation></ref>
<ref id="B31"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>McGillion</surname> <given-names>M.</given-names></name> <name><surname>Herbert</surname> <given-names>J.</given-names></name> <name><surname>Pine</surname> <given-names>J.</given-names></name> <name><surname>Keren-Portnoy</surname> <given-names>T.</given-names></name> <name><surname>Vihman</surname> <given-names>M.</given-names></name> <name><surname>Matthews</surname> <given-names>D.</given-names></name></person-group> (<year>2013</year>). <article-title>Supporting early vocabulary development: what sort of responsiveness matters?</article-title> <source>IEEE Trans. Auton. Ment. Dev.</source> <volume>5</volume>, <fpage>240</fpage>&#x02013;<lpage>248</lpage>. <pub-id pub-id-type="doi">10.1109/tamd.2013.2275949</pub-id></citation></ref>
<ref id="B32"><citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Nguyen</surname> <given-names>T. L.</given-names></name> <name><surname>Boukezzoula</surname> <given-names>R.</given-names></name> <name><surname>Coquin</surname> <given-names>D.</given-names></name> <name><surname>Benoit</surname> <given-names>E.</given-names></name> <name><surname>Perrin</surname> <given-names>S.</given-names></name></person-group> (<year>2015</year>). &#x0201C;<article-title>Interaction between humans, NAO robot and multiple cameras for colored objects recognition using information fusion</article-title>,&#x0201D; in <source>8th International Conference on Human System Interaction</source> (HSI), (<conf-loc>Warsaw, Poland</conf-loc>), <fpage>322</fpage>&#x02013;<lpage>328</lpage>.</citation></ref>
<ref id="B33"><citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Schodde</surname> <given-names>T.</given-names></name> <name><surname>Bergmann</surname> <given-names>K.</given-names></name> <name><surname>Kopp</surname> <given-names>S.</given-names></name></person-group> (<year>2017</year>). &#x0201C;<article-title>Adaptive robot language tutoring based on bayesian knowledge tracing and predictive decision-making</article-title>,&#x0201D; in <source>Proceedings of the 12th ACM/IEEE International Conference on Human-Robot Interaction</source> (HRI 2017), (<conf-loc>Vienna, Austria</conf-loc>).</citation></ref>
<ref id="B330"><citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Shahid</surname> <given-names>S.</given-names></name> <name><surname>Krahmer</surname> <given-names>E.</given-names></name> <name><surname>Swerts</surname> <given-names>M.</given-names></name></person-group> (<year>2014</year>). &#x0201C;<article-title>Child&#x02013;robot interaction across cultures: how does playing a game with a social robot compare to playing a game alone or with a friend?</article-title> <source>Comput. Human Behav.</source> <volume>40</volume>, <fpage>86</fpage>&#x02013;<lpage>100</lpage>. <pub-id pub-id-type="doi">10.1016/j.chb.2014.07.043</pub-id></citation></ref>
<ref id="B34"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tanaka</surname> <given-names>F.</given-names></name> <name><surname>Matsuzoe</surname> <given-names>S.</given-names></name></person-group> (<year>2012</year>). <article-title>Children teach a care-receiving robot to promote their learning: field experiments in a classroom for vocabulary learning</article-title>. <source>J. Hum. Robot Interact.</source> <volume>1</volume>, <fpage>78</fpage>&#x02013;<lpage>95</lpage>. <pub-id pub-id-type="doi">10.5898/jhri.1.1.tanaka</pub-id></citation></ref>
<ref id="B35"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Tazhigaliyeva</surname> <given-names>N.</given-names></name> <name><surname>Diyas</surname> <given-names>Y.</given-names></name> <name><surname>Brakk</surname> <given-names>D.</given-names></name> <name><surname>Aimambetov</surname> <given-names>Y.</given-names></name> <name><surname>Sandygulova</surname> <given-names>A.</given-names></name></person-group> (<year>2016</year>). &#x0201C;<article-title>Learning with or from the robot: exploring robot roles in educational context with children</article-title>,&#x0201D; in <source>International Conference on Social Robotics</source>, eds <person-group person-group-type="editor"><name><surname>Agah</surname> <given-names>A.</given-names></name> <name><surname>Cabibihan</surname> <given-names>J.-J.</given-names></name> <name><surname>Howard</surname> <given-names>A. M.</given-names></name> <name><surname>Salichs</surname> <given-names>M. A.</given-names></name> <name><surname>He</surname> <given-names>H.</given-names></name></person-group> (<publisher-loc>New York, NY</publisher-loc>: <publisher-name>Springer International Publishing</publisher-name>), <fpage>327</fpage>&#x02013;<lpage>336</lpage>.</citation></ref>
<ref id="B36"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tomasello</surname> <given-names>M.</given-names></name> <name><surname>Farrar</surname> <given-names>M. J.</given-names></name></person-group> (<year>1986</year>). <article-title>Joint attention and early language</article-title>. <source>Child Dev.</source> <volume>57</volume>, <fpage>1454</fpage>&#x02013;<lpage>1463</lpage>. <pub-id pub-id-type="doi">10.2307/1130423</pub-id><pub-id pub-id-type="pmid">3802971</pub-id></citation></ref>
<ref id="B37"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Vygotsky</surname> <given-names>L.</given-names></name></person-group> (<year>1978</year>). <source>Mind in Society.</source> <publisher-loc>Harvard</publisher-loc>: <publisher-name>Harvard University Press</publisher-name>.</citation></ref>
<ref id="B38"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Westlund</surname> <given-names>J. M. K.</given-names></name> <name><surname>Martinez</surname> <given-names>M.</given-names></name> <name><surname>Archie</surname> <given-names>M.</given-names></name> <name><surname>Das</surname> <given-names>M.</given-names></name> <name><surname>Breazeal</surname> <given-names>C.</given-names></name></person-group> (<year>2016</year>). &#x0201C;<article-title>Effects of framing a robot as a social agent or as a machine on children&#x02019;s social behavior</article-title>,&#x0201D; in <source>Proceedings of the 25th IEEE International Symposium on Robot and Human Interactive Communication (RO-MAN)</source>, eds <person-group person-group-type="editor"><name><surname>Okita</surname> <given-names>S. Y.</given-names></name> <name><surname>Shibata</surname> <given-names>T.</given-names></name> <name><surname>Mutlu</surname> <given-names>B.</given-names></name></person-group> (<publisher-loc>Washington, DC</publisher-loc>: <publisher-name>IEEE</publisher-name>), <fpage>688</fpage>&#x02013;<lpage>693</lpage>.</citation></ref>
<ref id="B39"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Yu</surname> <given-names>C.</given-names></name> <name><surname>Smith</surname> <given-names>L. B.</given-names></name></person-group> (<year>2016</year>). <article-title>The social origins of sustained attention in one-year-old human infants</article-title>. <source>Curr. Biol.</source> <volume>26</volume>, <fpage>1235</fpage>&#x02013;<lpage>1240</lpage>. <pub-id pub-id-type="doi">10.1016/j.cub.2016.03.026</pub-id><pub-id pub-id-type="pmid">27133869</pub-id></citation></ref>
</ref-list>
</back>
</article> 