<?xml version="1.0" encoding="utf-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.3 20210610//EN" "JATS-journalpublishing1-3-mathml3.dtd">
<article article-type="research-article" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:ali="http://www.niso.org/schemas/ali/1.0/" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" dtd-version="1.3" xml:lang="EN">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Comput. Sci.</journal-id>
<journal-title-group>
<journal-title>Frontiers in Computer Science</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Comput. Sci.</abbrev-journal-title>
</journal-title-group>
<issn pub-type="epub">2624-9898</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fcomp.2025.1659594</article-id><article-version article-version-type="Version of Record" vocab="NISO-RP-8-2008"/>
<article-categories>
<subj-group subj-group-type="heading"><subject>Original Research</subject></subj-group>
</article-categories>
<title-group>
<article-title>The impact of usage experience and input modality on trust experience and cognitive load in older adults</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes"><name><surname>Huang</surname> <given-names>Hui</given-names></name><xref ref-type="aff" rid="aff1"><sup>1</sup></xref><xref ref-type="corresp" rid="c001"><sup>&#x002A;</sup></xref>
<uri xlink:href="https://loop.frontiersin.org/people/1974622"/>
<role vocab="credit" vocab-identifier="https://credit.niso.org/" vocab-term="Data curation" vocab-term-identifier="https://credit.niso.org/contributor-roles/data-curation/">Data curation</role>
<role vocab="credit" vocab-identifier="https://credit.niso.org/" vocab-term="Formal analysis" vocab-term-identifier="https://credit.niso.org/contributor-roles/formal-analysis/">Formal analysis</role>
<role vocab="credit" vocab-identifier="https://credit.niso.org/" vocab-term="Writing &#x2013; original draft" vocab-term-identifier="https://credit.niso.org/contributor-roles/writing-original-draft/">Writing &#x2013; original draft</role>
</contrib>
<contrib contrib-type="author"><name><surname>Hou</surname> <given-names>Guanhua</given-names></name><xref ref-type="aff" rid="aff2"><sup>2</sup></xref>
<uri xlink:href="https://loop.frontiersin.org/people/1226044"/>
<role vocab="credit" vocab-identifier="https://credit.niso.org/" vocab-term="supervision" vocab-term-identifier="https://credit.niso.org/contributor-roles/supervision/">Supervision</role>
<role vocab="credit" vocab-identifier="https://credit.niso.org/" vocab-term="Writing &#x2013; review &amp; editing" vocab-term-identifier="https://credit.niso.org/contributor-roles/writing-review-editing/">Writing &#x2013; review &#x0026; editing</role>
</contrib>
</contrib-group>
<aff id="aff1"><label>1</label><institution>Yibin Hospital Affiliated to Children&#x2019;s Hospital of Chongqing Medical University</institution>, <city>Yibin</city>, <country country="cn">China</country></aff>
<aff id="aff2"><label>2</label><institution>Southeast University</institution>, <city>Nanjing</city>, <country country="cn">China</country></aff>
<author-notes><corresp id="c001"><label>&#x002A;</label>Correspondence: Hui Huang, <email xlink:href="mailto:1132419014@qq.com">1132419014@qq.com</email></corresp></author-notes>
<pub-date publication-format="electronic" date-type="pub" iso-8601-date="2025-11-13">
<day>13</day>
<month>11</month>
<year>2025</year>
</pub-date>
<pub-date publication-format="electronic" date-type="collection">
<year>2025</year>
</pub-date>
<volume>7</volume>
<elocation-id>1659594</elocation-id>
<history>
<date date-type="received">
<day>07</day>
<month>07</month>
<year>2025</year>
</date>
<date date-type="accepted">
<day>06</day>
<month>10</month>
<year>2025</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x00A9; 2025 Huang and Hou.</copyright-statement>
<copyright-year>2025</copyright-year>
<copyright-holder>Huang and Hou</copyright-holder>
<license><ali:license_ref start_date="2025-11-13">https://creativecommons.org/licenses/by/4.0/</ali:license_ref>
<license-p>This is an open-access article distributed under the terms of the <ext-link ext-link-type="uri" xlink:href="https://creativecommons.org/licenses/by/4.0/">Creative Commons Attribution License (CC BY)</ext-link>. The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</license-p>
</license>
</permissions>
<abstract>
<p>Trust experience plays a pivotal role in human&#x2013;computer interaction, particularly for older adults, where it serves as a critical psychological threshold for technology adoption and sustained usage. Against the backdrop of increasingly diverse intelligent interaction modalities, trust directly influences older adults&#x2019; initial acceptance and long-term reliance on technological systems. This study focuses on the interactive effects of users&#x2019; experience and input modality on trust experience and cognitive load in the elderly. Employing a 2 (prior experience: experienced vs. inexperienced)&#x202F;&#x00D7;&#x202F;3 (input modality: touch, speech, eye control) mixed experimental design. Following each task, participants completed NASA-TLX scales and trust perception questionnaires, supplemented by eye-tracking data to quantify cognitive load and behavioral patterns. The results showed that (1) Experience-dependent divergence in trust perception: Experienced older adults exhibited higher trust in touch input, attributable to established press-response mental models from prior device usage, while inexperienced users preferred speech input due to its alignment with natural conversational paradigms. (2) Cognitive load mediation effect: Although voice input reduces the learning cost of user interfaces for inexperienced elderly users (NASA-TLX is 24% lower than touch), recognition errors can cause a sharp drop in trust; This study reveals that the trust experience of elderly users is influenced by both usage experience and input methods, with cognitive load being a key mediating factor. In terms of design, the touch physical metaphor should be retained for experienced elderly users, and the voice fault tolerance mechanism should be strengthened for inexperienced elderly users, while reducing technical anxiety by enhancing operational visibility.</p>
</abstract>
<kwd-group>
<kwd>input modality</kwd>
<kwd>human&#x2013;computer interaction</kwd>
<kwd>older adult interaction</kwd>
<kwd>user trust experience</kwd>
<kwd>cognitive load</kwd>
</kwd-group><funding-group><funding-statement>The author(s) declare that no financial support was received for the research and/or publication of this article.</funding-statement></funding-group>
<counts>
<fig-count count="9"/>
<table-count count="5"/>
<equation-count count="0"/>
<ref-count count="57"/>
<page-count count="14"/>
<word-count count="8318"/>
</counts>
<custom-meta-group>
<custom-meta>
<meta-name>section-at-acceptance</meta-name>
<meta-value>Human-Media Interaction</meta-value>
</custom-meta>
</custom-meta-group>
</article-meta>
</front>
<body>
<sec sec-type="intro" id="sec1">
<label>1</label>
<title>Introduction</title>
<p>The advancement of AI and computer vision technologies has provided strong support for the creation of more natural and efficient human&#x2013;computer interaction scenarios, as well as multimodal forms of input, but it has also raised many concerns about the trust experience. The trust experience is the most profound feeling that users have when using a product, and trust is critical in regulating the relationship between humans and automated systems, as well as the acceptance, adoption, and continued use of interactive systems (<xref ref-type="bibr" rid="ref13">Gefen et al., 2003</xref>; <xref ref-type="bibr" rid="ref28">Lee and See, 2004</xref>). The input modality (speech, gesture, eye gaze, and touch) that a user uses to communicate information to a mobile self-service terminal is particularly significant because it allows the user to access personal data and perform tasks like viewing personal information and conducting online transactions (<xref ref-type="bibr" rid="ref53">Vildjiounaite et al., 2006</xref>; <xref ref-type="bibr" rid="ref28">Lee and See, 2004</xref>). However, because different input modes require different ways of operating, this is especially challenging for older users who are already dealing with the challenges of the digital divide. At the moment, older adults have relatively poor trust experiences and low acceptance of smart technologies due to lower transparency and interpretability in human-computer interaction. Given the widespread use of smart products in mobile self-service terminals, it is critical to investigate the effect of various input modalities on the trust experience of older users.</p>
<p>Elderly people&#x2019;s cognitive modes are heavily influenced by their prior experiences and knowledge; thus, interaction systems that correspond to their experiences are more likely to be understood and accepted. Research has shown that an individual experiential knowledge has a significant impact on the trust experience in human&#x2013;computer interaction (HCI), influencing not only the user&#x2019;s trust in the system but also their understanding and satisfaction with the system. A complete human-computer interaction process consists of two major components: user input mode and system output feedback. The user implements control behaviors using input commands and interprets their responses using system feedback. Different input modalities have a significant impact on the user experience.</p>
<p>Based on this, this study focuses on the various input modes used by elderly users when interacting with smart bodies, as well as how differences in usage experience influence elderly user trust experience. In addition, this study will investigate whether usage experience modifies older trust experience. Furthermore, the study aims to determine whether using different input modes increases the cognitive load of older adults and whether usage experience plays a moderating role in the process.</p>
<p>Therefore, this study aims to answer the following research questions:</p>
<list list-type="order">
<list-item>
<p>What are the differences in task performance and cognitive load of older adults when using different input modalities (touch, speech, and eye control)?</p>
</list-item>
<list-item>
<p>Do older adults with different usage experiences (experienced vs. inexperienced) differ in input modalities (touch, voice, eye control), trust experience, and cognitive load?</p>
</list-item>
<list-item>
<p>Is there an interaction between user trust experience, cognitive load, and task performance?</p>
</list-item>
</list>
</sec>
<sec id="sec2">
<label>2</label>
<title>Related work</title>
<sec id="sec3">
<label>2.1</label>
<title>User trust experience</title>
<p>The critical role of trust in technology adoption has been well-documented across various domains, particularly in the acceptance of novel interactive systems (<xref ref-type="bibr" rid="ref19">Hoff and Bashir, 2015</xref>). Studies on human-automation interaction consistently identify trust as a pivotal determinant of user acceptance (<xref ref-type="bibr" rid="ref13">Gefen et al., 2003</xref>), serving as a psychological mediator between users and technological systems (<xref ref-type="bibr" rid="ref14">Ghazizadeh et al., 2012</xref>). Trust significantly influences users&#x2019; willingness to engage with technology, especially in contexts requiring risk mitigation (<xref ref-type="bibr" rid="ref18">Hengstler et al., 2016</xref>). In the realm of self-service terminals, trust formation varies substantially across input modalities (e.g., speech, touch, eye control) and is further modulated by age-related differences. Prior research suggests that older adults&#x2019; trust in touch interfaces often stems from schema transfer from legacy devices (e.g., feature phones, ATMs), whereas their trust in speech interfaces reflects alignment with natural communication paradigms (<xref ref-type="bibr" rid="ref38">Oviatt, 2022</xref>). However, excessive cognitive load&#x2014;common in complex modality interactions&#x2014;can erode trust by depleting metacognitive resources necessary for confidence calibration (<xref ref-type="bibr" rid="ref40">Parasuraman, 2000</xref>).</p>
<p>Trust is a multidimensional construct defined as&#x201D; the attitude that a system will fulfill user goals amid uncertainty&#x201D; (<xref ref-type="bibr" rid="ref28">Lee and See, 2004</xref>). Its calibration depends on three factors: (1) modality familiarity (e.g., older users&#x2019; predisposition toward tactile interfaces) (<xref ref-type="bibr" rid="ref9">Claypoole et al., 2016</xref>), (2) transparency of system feedback (<xref ref-type="bibr" rid="ref34">Mikulski, 2014</xref>), and (3) error tolerance (e.g., voice recognition errors disproportionately reduce trust in novice users) (<xref ref-type="bibr" rid="ref24">Kaye et al., 2018</xref>). Mismatches in trust calibration&#x2014;such as overreliance on flawed touch systems or distrust of efficient speech interfaces&#x2014;can lead to either misuse or disuse of self-service technologies (<xref ref-type="bibr" rid="ref43">Pop et al., 2015</xref>). While <xref ref-type="bibr" rid="ref28">Lee and See (2004)</xref> emphasize that trust builds through cognitive and affective pathways, its dynamic nature means it fluctuates with user experience.</p>
<p>Despite its centrality, trust in multimodal self-service systems remains understudied, particularly for older adults. Existing work seldom addresses how age-specific factors (e.g., cognitive decline, technophobia) interact with modality-dependent trust formation <xref ref-type="bibr" rid="ref8">Choi and Ji (2015)</xref>. This gap is critical because older users&#x2014;facing steeper learning curves with touch or eye control&#x2014;may default to speech despite its potential cognitive overhead (<xref ref-type="bibr" rid="ref3">Basu, 2021</xref>). A systematic integration of trust dynamics into self-service design could mitigate adoption barriers across age groups.</p>
</sec>
<sec id="sec4">
<label>2.2</label>
<title>Input modality</title>
<sec id="sec5">
<label>2.2.1</label>
<title>Utility of touch input modality in self-service terminals</title>
<p>The low cost of hardware means that touch input devices are a popular input modality for self-service terminals installed in public places, offering services such as hospitals (<xref ref-type="bibr" rid="ref51">Shang et al., 2020</xref>), internet access (<xref ref-type="bibr" rid="ref15">Guo et al., 2007</xref>), information (<xref ref-type="bibr" rid="ref52">Slay et al., 2006</xref>), city guides (<xref ref-type="bibr" rid="ref21">Johnston and Bangalore, 2004</xref>), banks (<xref ref-type="bibr" rid="ref39">Paradi and Ghazarian-Rock, 1998</xref>) and many more. These self-service terminals reduce the need for costly staff, and contents can be changed in real time. Compared to traditional physical-button-based interaction, touch input has the advantage of being concise and convenient, and was reported as having high usability and being more preferred by young users (<xref ref-type="bibr" rid="ref27">Lee et al., 2020</xref>). Most touch-based self-service terminals are based on absolute positioned virtual buttons which are difficult to locate without any tactile, audible or visual cues (<xref ref-type="bibr" rid="ref48">Sandnes et al., 2012</xref>). Older adults in particular may have struggle searching and clicking on targets due to reduced information processing, precise movement and timely control (<xref ref-type="bibr" rid="ref12">Fisk et al., 2020</xref>; <xref ref-type="bibr" rid="ref29">Leonardi et al., 2010</xref>). <xref ref-type="bibr" rid="ref6">Ch&#x00EA;ne et al. (2016)</xref> also found that they have a harder time using clicks than younger adults do because of invalid touches. In conclusion, touch input is the dominant input modality for self-service terminals and is popular with young people for its simplicity of operation. However, it may not be as user-friendly for older adults when performing actions such as searching and manipulating. Therefore, it is necessary to explore the accessibility of touch input for different age groups when using medical self-service terminals.</p>
</sec>
<sec id="sec6">
<label>2.2.2</label>
<title>Utility of touchless interaction for older adults</title>
<p>The user&#x2019;s touchless input modalities such as voice, gesture, or eye control are commonly utilized as a means to transmit information to the system. Recently, the input modalities that gain a lot of attention are voice and eye-control input (<xref ref-type="bibr" rid="ref25">Kim et al., 2021</xref>). As voice input device become a mature technology, it has emerged as a popular touchless input modality. Voice input allows users to utter a command that can be recognized by the system to execute an operation (e.g., return) or enter certain information (e.g., outpatient charges) (<xref ref-type="bibr" rid="ref55">Zhang et al., 2023</xref>). The use of voice input technology can go some way to improving the usability and user experience of self-service terminals (<xref ref-type="bibr" rid="ref47">Saji&#x0107; et al., 2021</xref>). It increases the confidence of older adults and improves their acceptance of self-service terminals (<xref ref-type="bibr" rid="ref7">Chi et al., 2020</xref>). <xref ref-type="bibr" rid="ref23">Kaufman et al. (1993)</xref> showed that voice input provided users with introductory interaction to increase usability. <xref ref-type="bibr" rid="ref33">Manzke et al. (1998)</xref> investigated the use of self-service terminals by visually impaired users, in which the use of voice input was evaluated. For visually impaired users, voice input showed significant performance advantages. This also benefits older adults with reduced vision. Voice input is one of the touchless input modalities, which improves the accessibility of self-service terminals. However, does voice input benefit older adults in using self-service terminals and improve their task performance and usability of operations? These questions are yet to be thoroughly researched and explored.</p>
<p>Eye-control input relies on gaze behavior for computer interaction. Users can use eye gaze to select products (<xref ref-type="bibr" rid="ref26">Kim et al., 2015</xref>). Determining the appropriate dwell time for the eye-control input becomes critical when the eye-control input is used to simulate a click and confirm operation (<xref ref-type="bibr" rid="ref16">Hansen et al., 2001</xref>). Researchers have tried to improve user experience by setting different gaze times to complete information selection (<xref ref-type="bibr" rid="ref41">Pfeuffer et al., 2021</xref>). Regarding eye-control input, the best dwell time is 600 which is recorded in milliseconds (ms) in the trigger system (<xref ref-type="bibr" rid="ref54">Ya-feng et al., 2022</xref>). <xref ref-type="bibr" rid="ref35">Niu et al. (2019)</xref> also showed that when the gaze dwell time is set to 600&#x202F;ms, the efficiency of the interaction is the highest, and the task load of users is minimal as well. However, prolonged gaze dwells time and eye fatigue can result in an excessive cognitive load (<xref ref-type="bibr" rid="ref49">Sato et al., 2018</xref>). In addition, eye-control research has focused on young individuals (<xref ref-type="bibr" rid="ref46">Rozado et al., 2012</xref>). It is questionable whether it is suitable for older people to use eye-control input to operate self-service terminals and cause the high cognitive load.</p>
</sec>
</sec>
</sec>
<sec sec-type="methods" id="sec7">
<label>3</label>
<title>Methods</title>
<sec id="sec8">
<label>3.1</label>
<title>Participants</title>
<p>This study recruited 40 participants and a vision and listening ability test was conducted among the participants by using the Chinese version of the Functional Visual Screening Questionnaire (FVSQ) (1991). Each participant was able to walk easily and able to complete tasks independently. <xref ref-type="table" rid="tab1">Table 1</xref> shows a detailed breakdown of participant demographics.</p>
<table-wrap position="float" id="tab1">
<label>Table 1</label>
<caption>
<p>Information for participants.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top">Information</th>
<th align="center" valign="top">Participants data</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="top">Gender</td>
<td align="center" valign="top">20 (male), 20 (female)</td>
</tr>
<tr>
<td align="left" valign="top">Age</td>
<td align="center" valign="top">60&#x202F;years and older (Mean&#x202F;&#x00B1;&#x202F;SD: 64&#x202F;&#x00B1;&#x202F;3.127)</td>
</tr>
<tr>
<td align="left" valign="top">Experience</td>
<td align="center" valign="top">20 (experienced), 20 (inexperienced)</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<p>All the participants had a visual acuity of 1.0 LogMAR. and a hearing ability of 25&#x202F;dB or less (<xref ref-type="bibr" rid="ref45">Roth et al., 2011</xref>; <xref ref-type="bibr" rid="ref10">De Raedemaeker et al., 2022</xref>).</p>
</table-wrap-foot>
</table-wrap>
<p>Experienced users were defined as those who reported using smart devices &#x2265;3 times per week over the past year and could independently perform complex operations (e.g., online payments, app installations). Inexperienced users were defined as those who used smart devices &#x003C;1 time per week or required assistance with basic operations (e.g., download software, register account). This dichotomous approach helped minimize within-group variance and ensured clearer detection of experimental effects. The older adults were active or retired school employees. Informed consent was obtained from all the participants before the experiment was conducted, and each participant was paid 50 RMB for their participation.</p>
</sec>
<sec id="sec9">
<label>3.2</label>
<title>Experimental material</title>
<p>Using both web-based and field research, this study analyzed the interfaces of the medical self-service systems in China and used a standard interface of a medical self-service system as the experimental material. Diagram of experimental material consisting of the home page and six other pages for each step in the task to complete a registration (see <xref ref-type="fig" rid="fig1">Figure 1</xref>; English version in <xref ref-type="fig" rid="fig2">Figure 2</xref>).</p>
<fig position="float" id="fig1">
<label>Figure 1</label>
<caption>
<p>Experimental material diagram (corresponding to the experimental task).</p>
</caption>
<graphic xlink:href="fcomp-07-1659594-g001.tif" mimetype="image" mime-subtype="tiff">
<alt-text content-type="machine-generated">A series of interface screenshots for a hospital appointment system. The first image displays the home page with registration options. The second shows appointment date selections from May thirteenth to twenty-first, 2022. The third image enables choosing a hospital department. The fourth depicts more specific department options. The fifth shows selection of hospital doctors with names and available slots. The sixth provides consultation number options with specific times. The final image offers confirmation information, including patient details and an option to confirm the appointment.</alt-text>
</graphic>
</fig>
<fig position="float" id="fig2">
<label>Figure 2</label>
<caption>
<p>Experimental Task Diagram (English version of interface content).</p>
</caption>
<graphic xlink:href="fcomp-07-1659594-g002.tif" mimetype="image" mime-subtype="tiff">
<alt-text content-type="machine-generated">The image displays a sequence of screenshots from a hospital&#x2019;s online appointment booking system. It includes seven steps: registration selection, appointment time selection, hospital department selection, detailed department selection, hospital doctor selection, consultation number selection, and confirmation information. Each screenshot shows a user interface with buttons and options relevant to each step, such as date and department choices, available doctors, and confirmation of appointment details.</alt-text>
</graphic>
</fig>
</sec>
<sec id="sec10">
<label>3.3</label>
<title>Experimental design</title>
<p>This study employed a mixed-design experiment with two factors: a three-level within-subjects factor of Input Modality (Touch, Speech, Eye Control) and a two-level between-subjects factor of User Type (Experienced, Inexperienced). The design was implemented to systematically evaluate the impact of different interaction modes and prior experience on key dimensions of user experience. Specifically, each participant interacted with all three input modalities, while being assigned to one of the two user type groups. The dependent variables encompassed a multi-dimensional set of metrics, primarily including subjective perceptions (e.g., user trust and cognitive load) and objective behavioral performance (e.g., task completion time), as detailed in the table below. The study recruited a total of 40 participants, comprising an experienced group (<italic>n</italic>&#x202F;=&#x202F;20) and an inexperienced group (<italic>n</italic>&#x202F;=&#x202F;20). Task completion time was measured in seconds (s) and was recorded for every trial starting from the start cue up until the participants submitted their responses to end each trial. The system cannot tolerate skipping each step, as shown in <xref ref-type="table" rid="tab2">Table 2</xref>.</p>
<table-wrap position="float" id="tab2">
<label>Table 2</label>
<caption>
<p>Experimental design.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top">Variable type</th>
<th align="left" valign="top">Variable name</th>
<th align="left" valign="top">Levels/measurement method</th>
<th align="left" valign="top">Design type (within-/between-subjects)</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="middle">Independent variable 1</td>
<td align="left" valign="middle">Input modality</td>
<td align="left" valign="middle">Touch, speech, eye control (3 levels)</td>
<td align="left" valign="middle">Within-subjects</td>
</tr>
<tr>
<td align="left" valign="middle">Independent variable 2</td>
<td align="left" valign="middle">User type</td>
<td align="left" valign="middle">Experienced, inexperienced (2 levels)</td>
<td align="left" valign="middle">Between-subjects</td>
</tr>
<tr>
<td align="left" valign="middle" rowspan="5">Dependent variable</td>
<td align="left" valign="middle">Trust experience questionnaire</td>
<td align="left" valign="middle">Subjective scale measurement</td>
<td align="left" valign="middle">&#x2013;</td>
</tr>
<tr>
<td align="left" valign="middle">Cognitive load</td>
<td align="left" valign="middle">Subjective scale measurement</td>
<td align="left" valign="middle">&#x2013;</td>
</tr>
<tr>
<td align="left" valign="middle">Task completion time</td>
<td align="left" valign="middle">Objective recording (seconds)</td>
<td align="left" valign="middle">&#x2013;</td>
</tr>
<tr>
<td align="left" valign="middle" rowspan="2">Eye-tracking data</td>
<td align="left" valign="middle">The total fixation duration</td>
<td align="left" valign="middle">&#x2013;</td>
</tr>
<tr>
<td align="left" valign="middle">The number of fixations</td>
<td align="left" valign="middle">&#x2013;</td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
<sec id="sec11">
<label>3.4</label>
<title>Measurement method</title>
<sec id="sec12">
<label>3.4.1</label>
<title>Trust experience questionnaire</title>
<p>Research has demonstrated that trust scales can be a straightforward and efficient way to assess the level of trust between humans and computers. These scales are based on the perception of the individual who is placing their trust and are both user-friendly and adaptable to large-scale applications (<xref ref-type="bibr" rid="ref32">Ma et al., 2025</xref>). <xref ref-type="bibr" rid="ref20">Jian et al. (2000)</xref> classified user trust into three dimensions: motivation, operation, and utility, based on the order of interaction. They developed and validated a user trust scale that is highly usable (Cronbach&#x2019;s <italic>&#x03B1;</italic>&#x202F;=&#x202F;0.92). This experiment is divided into three trust dimensions: motivation, which is the initial trust generated by the user&#x2019;s use of the input modality; operation, which is the real-time trust in the user&#x2019;s behavior when carrying out the input modality; and utility, which is the ex-post trust formed by the final control effect of such an input modality. The questionnaire for this study contains three trust factors and twelve measurement items on a seven-point Likert scale (least agree, strongly disagree, disagree, neutral, agree, strongly agree, most agree) (<xref ref-type="bibr" rid="ref30">Likert, 2017</xref>).</p>
</sec>
<sec id="sec13">
<label>3.4.2</label>
<title>NASA-TLX scales</title>
<p>Cognitive load was measured using the NASA-TLX scale developed by <xref ref-type="bibr" rid="ref17">Hart and Staveland (1988)</xref>. The NASA-TLX scale items were rated on a 20-point scale (0&#x202F;=&#x202F;low, 20&#x202F;=&#x202F;high). The mental demand, physical demand, temporal demand, performance, effort, and frustration subscales were combined to create a composite NASA-TLX workload score (scaled to 0&#x202F;=&#x202F;low, 100&#x202F;=&#x202F;high) (<xref ref-type="bibr" rid="ref31">Lowndes et al., 2020</xref>).</p>
</sec>
<sec id="sec14">
<label>3.4.3</label>
<title>Eye-tracking data measurement</title>
<p>Eye-tracking fixation data is a measure of visual perceptual engagement. For this experiment, visual workload was evaluated using two metrics: (1) the total fixation duration on the areas of interest (AOIs), and (2) the number of fixations. These metrics were collected and preprocessed for subsequent analysis. Eye-tracking studies used predefined areas of interest (AOIs) based on relevant regions and targets. Fixation duration is an attention distribution indicator that measures how long the eye stays in the area of interest (AOI) (<xref ref-type="bibr" rid="ref11">Eckstein et al., 2017</xref>). A longer fixation duration indicates difficulty in extracting information (<xref ref-type="bibr" rid="ref22">Just and Carpenter, 1976</xref>). The number of fixations is the number of user gaze points located within the target AOIs. A longer fixation duration indicates difficulty in extracting information (<xref ref-type="bibr" rid="ref22">Just and Carpenter, 1976</xref>). The number of fixations is the number of user gaze points located within the target AOIs. The higher the number of fixations, the more difficult it is to identify the target in the search task and the greater the cognitive load (<xref ref-type="bibr" rid="ref42">Poole and Ball, 2006</xref>). Fixation is the relatively static state of the eye in a certain period. This makes the foveal vision stable in a certain location so that the visual system can obtain the details of an object. Measurable fixation must have a minimum duration of 60 which is recorded in milliseconds (ms), and the gaze velocity should not exceed 30&#x00B0;/s (<xref ref-type="bibr" rid="ref37">Olsen, 2012</xref>). Among them, &#x201C;ms&#x201D; is the abbreviation of &#x201C;millisecond,&#x201D; which means &#x201C;millisecond.&#x201D; It is a unit of time, and 1&#x202F;ms is equal to 10<sup>&#x2212;3</sup> s (i.e., one-thousandth of a second). &#x201C;&#x00B0;/s&#x201D; is the abbreviation of &#x201C;degree per second,&#x201D; which means &#x201C;degree per second.&#x201D; It is a unit of angular velocity and is used to express the angle that an object rotates through per second.</p>
</sec>
</sec>
</sec>
<sec id="sec15">
<label>4</label>
<title>Experimental equipment</title>
<p>The operation of the three input methods involves the utilization of hardware devices, software programming, and operating methods, as demonstrated and elucidated in <xref ref-type="table" rid="tab3">Table 3</xref>.</p>
<table-wrap position="float" id="tab3">
<label>Table 3</label>
<caption>
<p>Input modality operation.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top">Input modality</th>
<th align="left" valign="top">Hardware equipment</th>
<th align="left" valign="top">Operating method</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="top">Touch</td>
<td align="left" valign="top">Touch screen PC</td>
<td align="left" valign="top">Touch screen</td>
</tr>
<tr>
<td align="left" valign="top">Speech</td>
<td align="left" valign="top">Touch screen PC</td>
<td align="left" valign="top">Voice Calls, e.g., &#x2018;Select Dr. XX&#x2019;.</td>
</tr>
<tr>
<td align="left" valign="top">Eye-control</td>
<td align="left" valign="top">The Tobii Pro spectrum</td>
<td align="left" valign="top">Gaze trigger time of 600&#x202F;ms</td>
</tr>
</tbody>
</table>
</table-wrap>
<sec id="sec16">
<label>4.1</label>
<title>Hardware environment of the system construction</title>
<p>The Tobii Pro Spectrum eye-tracking device was used in this study for eye-control input, as shown in <xref ref-type="fig" rid="fig3">Figure 3A</xref>. The eye-control input device is particularly suited for scientific research involving the observation of eye movements across various experimental settings. This device accommodates a wide range of head movements, enabling participants to record data with high accuracy and precision. It can be mounted on a monitor, laptop, or other compatible devices to facilitate eye-controlled interactions. As depicted in <xref ref-type="fig" rid="fig3">Figure 3B</xref>, a 24-inch IPS bezel-less touchscreen monitor is utilized for touch input. This monitor features a Full HD 1080p display with a maximum resolution of 1920&#x202F;&#x00D7;&#x202F;1080 and an aspect ratio of 16:9, providing clear and detailed visuals. Aliyun&#x2019;s intelligent voice interaction technology can be integrated with a touchscreen computer, a spatial audio speaker, and a two-channel microphone array to support audio input and output. This combination enables seamless voice-controlled interactions and enhances the overall user experience.</p>
<fig position="float" id="fig3">
<label>Figure 3</label>
<caption>
<p>Hardware device of the system. <bold>(A)</bold> shows the Tobii Pro Spectrum eye-tracking device, <bold>(B)</bold> shows the 24&#x201D; IPS bezel-less touchscreen display and Tobii Pro Spectrum.</p>
</caption>
<graphic xlink:href="fcomp-07-1659594-g003.tif" mimetype="image" mime-subtype="tiff">
<alt-text content-type="machine-generated">Image showing two parts: (a) a slim, black speaker bar with a stand; (b) a computer monitor with dimensions labeled as 509 millimeters wide and 306 millimeters tall. The side view shows a thickness of 200 millimeters and a height including the stand of 379 millimeters. The monitor displays a Windows 7 desktop screen.</alt-text>
</graphic>
</fig>
</sec>
<sec id="sec17">
<label>4.2</label>
<title>The software environment of the system construction</title>
<p>This study employed the following software tools for the development of the voice interaction system:</p>
<list list-type="bullet">
<list-item>
<p>Visual Studio 2022: Used for C# programming, compilation, and debugging.</p>
</list-item>
<list-item>
<p>Unity3D: Utilized for the development of the input system.</p>
</list-item>
<list-item>
<p>Adobe Icon Design Tools (e.g., Adobe Photoshop, Adobe Illustrator): Employed for icon creation and interface design.</p>
</list-item>
</list>
<p>All software tools were deployed on a Windows 10 operating system environment. Voice materials were synthesized using a text-to-speech mini program, which was based on the Aliyun Text-to-Speech Platform API.<xref ref-type="fn" rid="fn0001"><sup>1</sup></xref> A neutral female synthetic voice (system identifier: Zhitian) was configured with a speech rate of approximately 5 words per second while maintaining default prosodic parameters (fundamental frequency&#x202F;=&#x202F;0&#x202F;Hz; amplitude modulation&#x202F;=&#x202F;0&#x202F;dB).</p>
</sec>
</sec>
<sec id="sec18">
<label>5</label>
<title>Procedure</title>
<p>The participants were asked to answer a questionnaire regarding their age, educational background, and experience using self-service system before the start of the experiment. In addition, a visual and auditory ability test was used to screen participants. Prior to the formal experiment, a practice session was implemented to familiarize participants with the experimental procedure. This session also served to verify that their actual operational ability aligned with the experience level identified during pre-screening. Each participant was required to practice three modalities by completing hospital department search tasks, which were different from the experimental task. Training continued until the participants were familiar with and correctly completed the task for each input modality. Then, the Tobii Pro Spectrum device started eye calibration and recorded data. Participants followed the instructions on the display and completed seven medical registration tasks using the input modality, as shown in <xref ref-type="table" rid="tab4">Table 4</xref>. At the end of each level of testing, participants would complete the trust experience questionnaire and NASA to record feedback on their experience. The time it took for each individual to complete all the tests under uniform screen brightness and ambient light ranged from 20 to 30&#x202F;min, with a 2&#x2013;5&#x202F;min delay between each of the three variants of the experiment.</p>
<table-wrap position="float" id="tab4">
<label>Table 4</label>
<caption>
<p>Experimental task content arrangement.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top">Task</th>
<th align="left" valign="top">Touch</th>
<th align="left" valign="top">Speech</th>
<th align="left" valign="top">Eye control</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="top">Registration selection</td>
<td align="left" valign="top">Appointment booking</td>
<td align="left" valign="top">Appointment booking</td>
<td align="left" valign="top">Appointment booking</td>
</tr>
<tr>
<td align="left" valign="top">Appointment time selection</td>
<td align="left" valign="top">May-17</td>
<td align="left" valign="top">May-13</td>
<td align="left" valign="top">May-21</td>
</tr>
<tr>
<td align="left" valign="top">Hospital department selection</td>
<td align="left" valign="top">Orthopedics</td>
<td align="left" valign="top">Dermatology</td>
<td align="left" valign="top">Gastrointestinal surgery</td>
</tr>
<tr>
<td align="left" valign="top">Detailed department selection</td>
<td align="left" valign="top">Orthopedic care</td>
<td align="left" valign="top">Skin laser</td>
<td align="left" valign="top">Gastrointestinal surgery care</td>
</tr>
<tr>
<td align="left" valign="top">Hospital doctor selection</td>
<td align="left" valign="top">Dr. Zheyang Wang</td>
<td align="left" valign="top">Dr. Hongming Zhu</td>
<td align="left" valign="top">Dr. Jianming Xie</td>
</tr>
<tr>
<td align="left" valign="top">Consultation number selection</td>
<td align="left" valign="top">No. 36 in the morning</td>
<td align="left" valign="top">No. 48 in the morning</td>
<td align="left" valign="top">No. 31 in the morning</td>
</tr>
<tr>
<td align="left" valign="top">Verifying information</td>
<td align="left" valign="top">Confirmation</td>
<td align="left" valign="top">Confirmation</td>
<td align="left" valign="top">Confirmation</td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
<sec id="sec19">
<label>6</label>
<title>Data collection and analysis</title>
<p>A total of 120 video data and scales were collected for this experiment. Before conducting further analysis, the AOI for each experimental video was defined according to the tasks (see <xref ref-type="fig" rid="fig4">Figure 4</xref>). In <xref ref-type="fig" rid="fig4">Figure 4</xref>, A refers to the touch input interface, where the yellow area is the task AOI; B represents the speech input interface, where the purple area is the task AOI; and C represents the eye-control input interface, where the green area is the task AOI. The individuals&#x2019; processing levels with respect to various AOIs were investigated by comparing their total duration of fixation and total number of fixations for these AOIs. All data were analyzed through repeated-measures analysis of variance. Repeated measures ANOVA examines data collected from the same subjects across different time points or conditions by partitioning variance components and employing an <italic>F</italic>-statistic to evaluate treatment effects against error, thereby controlling for individual differences and assessing significance among measurements. IBM SPSS Statistics 19 software was used to analyze the aforementioned results, with <italic>p</italic>&#x202F;&#x003C;&#x202F;0.05 set as the significance level. Before analysis of variance (ANOVA) was performed, the normal data distribution was examined for each condition. In addition, Mauchly&#x2019;s spherical test was conducted to correct the results of the repeated-measures ANOVA for different input modalities and different user types.</p>
<fig position="float" id="fig4">
<label>Figure 4</label>
<caption>
<p>AOI divisions for different input modes. <bold>(A)</bold> shows the areas of interest for the touch input interface, <bold>(B)</bold> shows the areas of interest for the voice input interface, and <bold>(C)</bold> shows the areas of interest for the eye-control input interface.</p>
</caption>
<graphic xlink:href="fcomp-07-1659594-g004.tif" mimetype="image" mime-subtype="tiff">
<alt-text content-type="machine-generated">A grid of medical appointment booking screens demonstrating different input methods: touch, voice, and eye control. Each section shows steps for booking, selecting dates, and confirming details for various doctors and departments like orthopedics, dermatology, and gastrointestinal surgery, with specific dates and appointment details.</alt-text>
</graphic>
</fig>
</sec>
<sec sec-type="results" id="sec20">
<label>7</label>
<title>Results</title>
<p>The six experimental conditions (2 user experiences and 3 input modalities) were assessed using five measures: the User Trust Scale, the NASA-TLX Scale, the Total Duration of Fixation, the Total number of Fixations, and the Task Completion Time. <xref ref-type="table" rid="tab5">Table 5</xref> summarizes the main effects from the analysis of variance (ANOVA) for the six experimental conditions. To facilitate the presentation, we abbreviate the names of the five measures as follows:</p>
<list list-type="bullet">
<list-item>
<p>TS: Mean score on the user trust experience questionnaire.</p>
</list-item>
<list-item>
<p>NASA: NASA-TLX Scale aka cognitive load scale mean score.</p>
</list-item>
<list-item>
<p>TDOF: Total Duration of Fixation of interest in the AOI for the user.</p>
</list-item>
<list-item>
<p>NOF: Total number of Fixations of interest in the AOI for the user.</p>
</list-item>
<list-item>
<p>TCT: Time for user to complete tasks.</p>
</list-item>
</list>
<table-wrap position="float" id="tab5">
<label>Table 5</label>
<caption>
<p>Analysis of variance (ANOVA) for six experimental conditions.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top">Dependent variables</th>
<th align="left" valign="top">Factors</th>
<th align="center" valign="top">df</th>
<th align="center" valign="top">
<italic>F</italic>
</th>
<th align="center" valign="top">
<italic>p</italic>
</th>
<th align="center" valign="top">
<italic>&#x03B7;</italic>
<sub>p</sub>
<sup>2</sup>
</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="top" rowspan="2">TS</td>
<td align="left" valign="top">Experience</td>
<td align="center" valign="top">1</td>
<td align="center" valign="top">0.100</td>
<td align="center" valign="top">0.760</td>
<td align="center" valign="top">0.012</td>
</tr>
<tr>
<td align="left" valign="top">Input modality</td>
<td align="center" valign="top">2</td>
<td align="center" valign="top">13,295</td>
<td align="center" valign="top">0.000<sup>&#x002A;&#x002A;&#x002A;</sup></td>
<td align="center" valign="top">0.624</td>
</tr>
<tr>
<td align="left" valign="top" rowspan="2">NASA</td>
<td align="left" valign="top">Experience</td>
<td align="center" valign="top">1</td>
<td align="center" valign="top">33.172</td>
<td align="center" valign="top">0.000<sup>&#x002A;&#x002A;&#x002A;</sup></td>
<td align="center" valign="top">0.806</td>
</tr>
<tr>
<td align="left" valign="top">Input modality</td>
<td align="center" valign="top">2</td>
<td align="center" valign="top">3.848</td>
<td align="center" valign="top">0.043&#x002A;</td>
<td align="center" valign="top">0.325</td>
</tr>
<tr>
<td align="left" valign="top" rowspan="2">TDOF</td>
<td align="left" valign="top">Experience</td>
<td align="center" valign="top">1</td>
<td align="center" valign="top">6.611</td>
<td align="center" valign="top">0.030<sup>&#x002A;</sup></td>
<td align="center" valign="top">0.424</td>
</tr>
<tr>
<td align="left" valign="top">Input modality</td>
<td align="center" valign="top">2</td>
<td align="center" valign="top">47.478</td>
<td align="center" valign="top">0.000<sup>&#x002A;&#x002A;&#x002A;</sup></td>
<td align="center" valign="top">0.841</td>
</tr>
<tr>
<td align="left" valign="top" rowspan="2">NOF</td>
<td align="left" valign="top">Experience</td>
<td align="center" valign="top">1</td>
<td align="center" valign="top">3.227</td>
<td align="center" valign="top">0.106</td>
<td align="center" valign="top">0.264</td>
</tr>
<tr>
<td align="left" valign="top">Input modality</td>
<td align="center" valign="top">2</td>
<td align="center" valign="top">67.865</td>
<td align="center" valign="top">0.000<sup>&#x002A;&#x002A;&#x002A;</sup></td>
<td align="center" valign="top">0.883</td>
</tr>
<tr>
<td align="left" valign="top" rowspan="2">TCT</td>
<td align="left" valign="top">Experience</td>
<td align="center" valign="top">1</td>
<td align="center" valign="top">19.980</td>
<td align="center" valign="top">0.002<sup>&#x002A;&#x002A;</sup></td>
<td align="center" valign="top">0.698</td>
</tr>
<tr>
<td align="left" valign="top">Input modality</td>
<td align="center" valign="top">2</td>
<td align="center" valign="top">55.489</td>
<td align="center" valign="top">0.000<sup>&#x002A;&#x002A;&#x002A;</sup></td>
<td align="center" valign="top">0.860</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<p>Asterisks denote statistical significance: &#x002A;<italic>p</italic> &#x003C; 0.05, &#x002A;&#x002A;<italic>p</italic> &#x003C; 0.01, &#x002A;&#x002A;&#x002A;<italic>p</italic> &#x003C; 0.001.</p>
</table-wrap-foot>
</table-wrap>
<sec id="sec21">
<label>7.1</label>
<title>Trust experience questionnaire</title>
<p>The results of the ANOVA showed that there was no significant difference between the older adults&#x2019; experience of use (<italic>F</italic>&#x202F;=&#x202F;0.10, <italic>p</italic>&#x202F;=&#x202F;0.760, <italic>&#x03B7;<sup>2</sup></italic>&#x202F;=&#x202F;0.012) on the user trust experience. Input modality made a significant difference in user trust experience (<italic>F</italic>&#x202F;=&#x202F;13,295, <italic>p</italic>&#x202F;=&#x202F;0.000, <italic>&#x03B7;<sup>2</sup></italic>&#x202F;=&#x202F;0.624) (see <xref ref-type="table" rid="tab1">Table 1</xref>). Touch input (M&#x202F;=&#x202F;60.889 SD&#x202F;=&#x202F;3.683) had a higher user trust experience than speech input (M&#x202F;=&#x202F;58.667, SD&#x202F;=&#x202F;3.037) and eye-control input (M&#x202F;=&#x202F;27.333, SD&#x202F;=&#x202F;5.637). Multiple comparisons (see <xref ref-type="fig" rid="fig5">Figure 5</xref>) revealed that experienced older adults had a higher trust experience with touch input (M&#x202F;=&#x202F;72, SD&#x202F;=&#x202F;4.42) than speech input (M&#x202F;=&#x202F;46.22, SD&#x202F;=&#x202F;4.9) and eye-control input (M&#x202F;=&#x202F;28, SD&#x202F;=&#x202F;5.77); inexperienced older adults had the highest trust experience with speech input (M&#x202F;=&#x202F;71.11, SD&#x202F;=&#x202F;4.04), which was significantly higher than touch input (M&#x202F;=&#x202F;49.78, SD&#x202F;=&#x202F;6.5) and eye-control input (M&#x202F;=&#x202F;26.67, SD&#x202F;=&#x202F;6.03).</p>
<fig position="float" id="fig5">
<label>Figure 5</label>
<caption>
<p>Multiple comparisons of trust experience questionnaire of experience and input modalities.</p>
</caption>
<graphic xlink:href="fcomp-07-1659594-g005.tif" mimetype="image" mime-subtype="tiff">
<alt-text content-type="machine-generated">Box plot comparing "Trust Experience Questionnaire" scores for "Used Experience" and "Unused Experience" across three methods: speech (gray), touch (red), and eye-control (blue). Touch shows higher median scores, with significant differences marked by asterisks.</alt-text>
</graphic>
</fig>
</sec>
<sec id="sec22">
<label>7.2</label>
<title>The NASA-TLX scale</title>
<p>Experience of use (<italic>F</italic>&#x202F;=&#x202F;33.172, <italic>p</italic>&#x202F;=&#x202F;0.000, <italic>&#x03B7;<sup>2</sup></italic>&#x202F;=&#x202F;0.806) and input modality (<italic>F</italic>&#x202F;=&#x202F;3.848, <italic>p</italic>&#x202F;=&#x202F;0.043, <italic>&#x03B7;<sup>2</sup></italic>&#x202F;=&#x202F;0.325) had a differentially significant effect on cognitive load (see <xref ref-type="table" rid="tab1">Table 1</xref>). Older adults with use experience had a lower cognitive load (M&#x202F;=&#x202F;5.241, SD&#x202F;=&#x202F;0.380) compared to those without use experience (M&#x202F;=&#x202F;9.119, SD&#x202F;=&#x202F;0.518). Speech input (M&#x202F;=&#x202F;5.794, SD&#x202F;=&#x202F;0.373) had the lowest cognitive load, significantly less than touch input (M&#x202F;=&#x202F;7.522, SD&#x202F;=&#x202F;0.751) and eye control input (M&#x202F;=&#x202F;8.222, SD&#x202F;=&#x202F;0.623). All older adults had the lowest cognitive load using speech input; experienced older adults had a significantly lower cognitive load using speech input (M&#x202F;=&#x202F;3.911, SD&#x202F;=&#x202F;0.495) than eye-control input (M&#x202F;=&#x202F;6.93, SD&#x202F;=&#x202F;0.775); and inexperienced older adults had a significantly lower cognitive load using speech input (M&#x202F;=&#x202F;7.678, SD&#x202F;=&#x202F;0.425) than touch input (M&#x202F;=&#x202F;10.167, SD&#x202F;=&#x202F;0.954) (see <xref ref-type="fig" rid="fig6">Figure 6</xref>).</p>
<fig position="float" id="fig6">
<label>Figure 6</label>
<caption>
<p>Multiple comparisons of the subjective cognitive loads of users and input modalities.</p>
</caption>
<graphic xlink:href="fcomp-07-1659594-g006.tif" mimetype="image" mime-subtype="tiff">
<alt-text content-type="machine-generated">Box plot comparing "Used Experience" and "Unused Experience" on the NASA-TLX scale for speech, touch, and eye-control interfaces. Scores range from 0 to 16. Notable differences are marked with asterisks, with the "Touch" method showing the highest variability and median score in both categories.</alt-text>
</graphic>
</fig>
</sec>
<sec id="sec23">
<label>7.3</label>
<title>Eye track data</title>
<p>Experience of use (<italic>F</italic>&#x202F;=&#x202F;6.611, <italic>p</italic>&#x202F;=&#x202F;0.03, <italic>&#x03B7;<sup>2</sup></italic>&#x202F;=&#x202F;0.424) and input modality (<italic>F</italic>&#x202F;=&#x202F;47.478, <italic>p</italic>&#x202F;=&#x202F;0.000, <italic>&#x03B7;<sup>2</sup></italic>&#x202F;=&#x202F;0.841) had a differentially significant effect on total duration of fixation. There was no significant difference in the number of fixations for older adults&#x2019; experience (<italic>F</italic>&#x202F;=&#x202F;3.227, <italic>p</italic>&#x202F;=&#x202F;0.106, <italic>&#x03B7;<sup>2</sup></italic>&#x202F;=&#x202F;0.264). Input modality significantly impacted the number of fixations (<italic>F</italic>&#x202F;=&#x202F;67.865, <italic>p</italic>&#x202F;=&#x202F;0.000, <italic>&#x03B7;<sup>2</sup></italic>&#x202F;=&#x202F;0.883) (<xref ref-type="table" rid="tab1">Table 1</xref>). Older adults with usage experience had significantly shorter fixation times and fewer fixations for speech input (M&#x202F;=&#x202F;4.247, SD&#x202F;=&#x202F;0.404) (M&#x202F;=&#x202F;13.786, SD&#x202F;=&#x202F;5.487) compared to touch input (M&#x202F;=&#x202F;6.553, SD&#x202F;=&#x202F;0.783) (M&#x202F;=&#x202F;26.819, SD&#x202F;=&#x202F;2.212) and eye-control input (M&#x202F;=&#x202F;17.426, SD&#x202F;=&#x202F;1.088) (M&#x202F;=&#x202F;47.069, SD&#x202F;=&#x202F;5.225). Older adults with no usage experience had significantly shorter fixation times and fewer fixations for speech input (M&#x202F;=&#x202F;9.205, SD&#x202F;=&#x202F;1.138) (M&#x202F;=&#x202F;24, SD&#x202F;=&#x202F;1.972) and touch input (M&#x202F;=&#x202F;11.514, SD&#x202F;=&#x202F;1.852) (M&#x202F;=&#x202F;30.2, SD&#x202F;=&#x202F;3.999) compared to eye-control input (M&#x202F;=&#x202F;28.817, SD&#x202F;=&#x202F;5.59) (M&#x202F;=&#x202F;60.4, SD&#x202F;=&#x202F;5.884) (see <xref ref-type="fig" rid="fig7">Figure 7</xref>).</p>
<fig position="float" id="fig7">
<label>Figure 7</label>
<caption>
<p>Multiple comparisons of the total number of fixations according to users and input modalities.</p>
</caption>
<graphic xlink:href="fcomp-07-1659594-g007.tif" mimetype="image" mime-subtype="tiff">
<alt-text content-type="machine-generated">Bar chart comparing eye track data, with total duration of fixation on the left and number of fixations on the right. It is segmented by used and unused experiences, and categorized by eye-control, touch, and speech, as shown in the legend. Eye-control has the highest values in both categories.</alt-text>
</graphic>
</fig>
</sec>
<sec id="sec24">
<label>7.4</label>
<title>Task completion time</title>
<p>Older adults used experience (<italic>F</italic>&#x202F;=&#x202F;19.98, <italic>p</italic>&#x202F;=&#x202F;0.002, <italic>&#x03B7;<sup>2</sup></italic>&#x202F;=&#x202F;0.689) and input modality (<italic>F</italic>&#x202F;=&#x202F;55.489, <italic>p</italic>&#x202F;=&#x202F;0.000, <italic>&#x03B7;<sup>2</sup></italic>&#x202F;=&#x202F;0.86) to have a significant effect on task completion time (see <xref ref-type="table" rid="tab1">Table 1</xref>). All older adults took significantly longer to complete the task using speech input (M&#x202F;=&#x202F;84.088, SD&#x202F;=&#x202F;3.279) (M&#x202F;=&#x202F;126.302, SD&#x202F;=&#x202F;7.933) than touch (M&#x202F;=&#x202F;53.553, SD&#x202F;=&#x202F;2.868) (M&#x202F;=&#x202F;67.140, SD&#x202F;=&#x202F;3.879) and eye-control input (M&#x202F;=&#x202F;57.514, SD&#x202F;=&#x202F;3.879) (M&#x202F;=&#x202F;73.454, SD&#x202F;=&#x202F;5.702); there was no significant difference between touch and eye-control input in terms of task completion (see <xref ref-type="fig" rid="fig8">Figure 8</xref>).</p>
<fig position="float" id="fig8">
<label>Figure 8</label>
<caption>
<p>Multiple comparisons of task completion time according to users and input modalities.</p>
</caption>
<graphic xlink:href="fcomp-07-1659594-g008.tif" mimetype="image" mime-subtype="tiff">
<alt-text content-type="machine-generated">Bar chart comparing task completion time for used vs. unused experience across three input methods: speech, eye-control, and touch. Speech has the highest time, while touch is the quickest for both experiences.</alt-text>
</graphic>
</fig>
</sec>
<sec id="sec25">
<label>7.5</label>
<title>Correlation analysis of questionnaire data, eyetracking data and task completion time data</title>
<p>As shown in <xref ref-type="fig" rid="fig9">Figure 9</xref>, a Pearson correlation analysis (<xref ref-type="bibr" rid="ref7010">Pearson, 1895</xref>) has been performed on five sets of data, with the first two sets of data being questionnaire data (subjective), the third and fourth sets of data being eye-tracking data (objective), and the fifth set of data being task completion time data (objective). From these three types of data, it can be concluded that NASA is positively correlated with NOF (0.33&#x002A;&#x002A;&#x002A;, <italic>p</italic>&#x202F;&#x003C;&#x202F;0.001) and TDOF (0.39&#x002A;&#x002A;, <italic>p</italic>&#x202F;&#x003C;&#x202F;0.01); TS is negatively correlated with NOF (&#x2212;0.48&#x002A;, <italic>p</italic>&#x202F;&#x003C;&#x202F;0.05) and TDOF (&#x2212;0.64&#x002A;, <italic>p</italic>&#x202F;&#x003C;&#x202F;0.05); NOF is positively correlated with TDOF (0.68&#x002A;, <italic>p</italic>&#x202F;&#x003C;&#x202F;0.05) was positively correlated; TDOF was negatively correlated with TCT (&#x2212;0.28&#x002A;&#x002A;&#x002A;, <italic>p</italic>&#x202F;&#x003C;&#x202F;0.001). Correlation analyses further demonstrated the influence of older adults on the experience of trust with cognitive load. As the cognitive load increased, the eye movement data also increased. Higher eye-movement data indicated lower user trust.</p>
<fig position="float" id="fig9">
<label>Figure 9</label>
<caption>
<p>Pearson&#x2019;s correlation and significance for questionnaire data, eye tracking data, and task completion time data.</p>
</caption>
<graphic xlink:href="fcomp-07-1659594-g009.tif" mimetype="image" mime-subtype="tiff">
<alt-text content-type="machine-generated">Correlation matrix showing relationships between variables NASA, TS, NOF, TDOF, and TCT. Color intensity represents correlation strength: red for positive, blue for negative. Significant values are marked with asterisks: &#x002A; for p&#x2264;0.05, &#x002A;&#x002A; for p&#x2264;0.01, &#x002A;&#x002A;&#x002A; for p&#x2264;0.001. The color bar ranges from -1 to 1.</alt-text>
</graphic>
</fig>
</sec>
</sec>
<sec sec-type="discussion" id="sec26">
<label>8</label>
<title>Discussion</title>
<sec id="sec27">
<label>8.1</label>
<title>The difference in trust experience, cognitive load, and task performance</title>
<p>Significant differences were observed in trust perception, cognitive load, and task performance among older adults when using different input modalities (touch, speech, and eye control). Touch input elicited significantly higher user trust perception compared to speech and eye control inputs. This preference stems from touch interaction being a more prevalent and familiar input method for elderly users. As demonstrated by <xref ref-type="bibr" rid="ref4">Bong et al. (2018)</xref>, touch interfaces incorporate widely recognized graphical UI elements that enhance system accessibility and senior user acceptance. The long-term usage patterns have established a dependency effect, consequently elevating users&#x2019; trust perception scores. Empirical evidence further confirms superior performance outcomes with touch input among elderly populations (<xref ref-type="bibr" rid="ref2001">Sultana and Moffatt, 2019</xref>).</p>
<p>Compared to touch and eye control, older adults have the lowest cognitive load when using speech. Older adults have the shortest total duration of fixation and the fewest number of fixations when using speech input; the duration of fixation and the number of fixations are the longest when using eye control input. As established by <xref ref-type="bibr" rid="ref50">Seaborn et al. (2023)</xref>, speech interaction leverages natural speech-based communication patterns, mirroring daily conversational flows. <xref ref-type="bibr" rid="ref44">Portet et al. (2013)</xref> further confirmed speech input as a comfortable and low-stress modality, eliminating visual-motor coordination demands inherent to touch/eye control input. The latter modalities impose additional physical constraints&#x2014;requiring precise finger movements or sustained visual attention&#x2014;thereby reducing overall comfort and usability. However, existing literature notes that speech input necessitates continuous maintenance of command context (consuming 2.5 working memory chunks on average), whereas graphical interfaces provide visual persistence (e.g., button states) that offloads memory demands (<xref ref-type="bibr" rid="ref2">Baddeley, 2012</xref>). The NASA-TLX scale shows that the cognitive load of speech input is 23% lower than that of touch input. For the elderly, with the decline of limb operation ability, the situations of accidental clicking and no feedback after clicking often occur, which increases the cognitive load of the elderly. Instead, voice can help the elderly better interact.</p>
<p>Among older adults, touch input yielded the shortest task completion time, while speech input resulted in the longest task completion time. This performance disparity stems from two fundamental factors: First, touch represents the most familiar interaction paradigm for older users, leveraging established mental models that enhance operational efficiency (<xref ref-type="bibr" rid="ref4">Bong et al., 2018</xref>). Second, while speech enforces a linear dialog pattern (speak&#x2013;listen&#x2013;speak cycle) that creates sequential processing bottlenecks, touch permits parallel operations&#x2014;enabling simultaneous finger gestures and visual scanning of interface elements (<xref ref-type="bibr" rid="ref38">Oviatt, 2022</xref>).</p>
</sec>
<sec id="sec28">
<label>8.2</label>
<title>Trust experience and cognitive load varied significantly among older adults with different levels of prior experience</title>
<p>Older adults with technological experience demonstrate greater trust in touch interfaces, cognitively mapping smartphone touch interactions to conventional button operations (<xref ref-type="bibr" rid="ref56">Zhou et al., 2017</xref>). Conversely, the inexperienced people exhibit strongest trust in speech input. This disparity stems from experience-mediated schema differentiation&#x2014;seasoned users develop ingrained press-response mental models through prolonged use of tactile devices (e.g., feature phones, ATMs) (<xref ref-type="bibr" rid="ref36">Norman, 2013</xref>), while novices adopt speech to bypass the cognitive demands of hierarchical UI navigation (typically 5&#x2013;7 menu levels in touch systems), instead employing natural language expressions that mirror familiar conversational patterns.</p>
<p>The observed divergence further reflects generational asymmetries in technology adoption. Experienced users, having established digital self-efficacy through successful touch interactions, perceive speech as more cognitively demanding and question its reliability. Inexperienced users, while judging speech and touch as equally taxing cognitively, face a compound learning barrier with touch&#x2014;simultaneously acquiring gesture vocabulary and interface logic&#x2014;making speech comparatively more accessible despite equivalent perceived effort.</p>
</sec>
<sec id="sec29">
<label>8.3</label>
<title>Trust experience, cognitive load, and task efficiency interacted significantly</title>
<p>The significant positive correlation among NASA-TLX scores, the duration of fixation, and the number of fixations demonstrates robust convergent validity in cognitive load measurement. This tripartite alignment confirms cross-modal consistency between subjective scales (NASA-TLX) and objective physiological metrics (eye-tracking data), satisfying the multitrait-multimethod matrix (MTMM) validation framework proposed by <xref ref-type="bibr" rid="ref5">Campbell and Fiske (1959)</xref>.</p>
<p>Further analysis reveals an inverse relationship between user trust perception and cognitive load&#x2014;as cognitive demands increase, trust formation becomes progressively compromised. Neurocognitive evidence suggests that excessive task load consumption diminishes metacognitive monitoring capacity essential for trust establishment, indicating that trust development requires sufficient &#x201C;cognitive slack.&#x201D; Empirical thresholds show that when NASA-TLX scores exceed 60 points, trust assessment accuracy declines by approximately 42%, as demonstrated in controlled experiments (CHI Conference 2023).</p>
</sec>
<sec id="sec30">
<label>8.4</label>
<title>From empirical findings to adaptive design strategies</title>
<p>To address the impracticality of direct user experience inquiries in real-world settings such as hospitals, future intelligent systems could infer user proficiency through continuous and low-intrusion analysis of micro-behavioral indicators. These metrics include interaction fluency (e.g., accuracy, hesitation time, task efficiency), error patterns, and exploration of advanced features. The system could initially operate in a default &#x201C;guided mode&#x201D; and automatically transition to an &#x201C;advanced mode&#x201D; upon detecting proficient behaviors. Furthermore, the system should be capable of continuous learning, dynamically adjusting interface complexity in response to evolving user proficiency. This approach directly accommodates the continuous nature of user experience, moving beyond reliance on static binary classifications. To improve accessibility, public terminal interfaces (e.g., in hospital lobbies) should be preset with novice-friendly modes such as &#x201C;voice-first&#x201D; or &#x201C;large-button touch,&#x201D; while ensuring clear and readily available alternatives for mode switching.</p>
<p>We observed that trust levels among inexperienced users dropped sharply following speech recognition errors. This indicates that robust error tolerance is a more critical system requirement than achieving perfectly accurate user classification. Consequently, system design must prioritize mechanisms that effectively mitigate the negative consequences of interaction failures. An enhanced fault-tolerance approach should be implemented, including: providing clear command examples, offering contextual prompts after recognition failures, and ensuring seamless switching to alternative input modalities (e.g., one-touch fallback to a touch interface). Such a &#x201C;safety net&#x201D; design bridges experience gaps through inherent system qualities, ensuring interface resilience and robustness even without perfect user state awareness.</p>
</sec>
</sec>
<sec id="sec31">
<label>9</label>
<title>Conclusions and limitations</title>
<p>This study highlights the important role of input method and experience of use in influencing trust and cognitive load when using smart devices in older adults. It was found that experienced users preferred touch input because it fitted their existing mental models, while inexperienced users preferred voice input because of its natural interactive approach. Cognitive load plays a mediating role in this, and while voice input reduces initial learning costs, it also poses reliability issues. Design recommendations state that the touch input habits of experienced users should be preserved, whilst increasing voice input tolerance and system visibility for novices to increase trust and reduce usage anxiety.</p>
<p>The sample in this study may not be fully representative of all older populations, especially those with significant cognitive decline or limited access to technology. In addition, the controlled experimental environment may not reflect the complexities encountered when using speech recognition in the real world, such as the effects of environmental noise. Future research should explore longitudinal trust development and multimodal interaction design to better support diverse older users.</p>
</sec>
</body>
<back>
<sec sec-type="data-availability" id="sec32">
<title>Data availability statement</title>
<p>The original contributions presented in the study are included in the article/supplementary material, further inquiries can be directed to the corresponding author.</p>
</sec>
<sec sec-type="ethics-statement" id="sec33">
<title>Ethics statement</title>
<p>This study was reviewed and approved by the Inclusive User Experience Design Center of Ningbo University. The studies were conducted in accordance with the local legislation and institutional requirements. The participants provided their written informed consent to participate in this study.</p>
</sec>
<sec sec-type="author-contributions" id="sec34">
<title>Author contributions</title>
<p>HH: Data curation, Formal analysis, Writing &#x2013; original draft. GH: Supervision, Writing &#x2013; review &#x0026; editing.</p>
</sec>

<sec sec-type="COI-statement" id="sec36">
<title>Conflict of interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec sec-type="ai-statement" id="sec37">
<title>Generative AI statement</title>
<p>The authors declare that no Gen AI was used in the creation of this manuscript.</p>
<p>Any alternative text (alt text) provided alongside figures in this article has been generated by Frontiers with the support of artificial intelligence and reasonable efforts have been made to ensure accuracy, including review by the authors wherever possible. If you identify any issues, please contact us.</p>
</sec>
<sec sec-type="disclaimer" id="sec38">
<title>Publisher&#x2019;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec><ref-list>
<title>References</title>
<ref id="ref2"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Baddeley</surname> <given-names>A.</given-names></name></person-group> (<year>2012</year>). <article-title>Working memory: theories, models, and controversies</article-title>. <source>Annu. Rev. Psychol.</source> <volume>63</volume>, <fpage>1</fpage>&#x2013;<lpage>29</lpage>. doi: <pub-id pub-id-type="doi">10.1146/annurev-psych-120710-100422</pub-id>, PMID: <pub-id pub-id-type="pmid">21961947</pub-id></mixed-citation></ref>
<ref id="ref3"><mixed-citation publication-type="other"><person-group person-group-type="author"><name><surname>Basu</surname> <given-names>R.</given-names></name></person-group> (<year>2021</year>). Age and Interface equipping older adults with technological tools. (doctoral dissertation). OCAD University.</mixed-citation></ref>
<ref id="ref4"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Bong</surname> <given-names>W. K.</given-names></name> <name><surname>Chen</surname> <given-names>W.</given-names></name> <name><surname>Bergland</surname> <given-names>A.</given-names></name></person-group> (<year>2018</year>). <article-title>Tangible user interface for social interactions for the elderly: a review of literature</article-title>. <source>Advances in Human-Computer Interaction</source> <volume>2018</volume>:<fpage>7249378</fpage>. doi: <pub-id pub-id-type="doi">10.1155/2018/7249378</pub-id></mixed-citation></ref>
<ref id="ref5"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Campbell</surname> <given-names>D. T.</given-names></name> <name><surname>Fiske</surname> <given-names>D. W.</given-names></name></person-group> (<year>1959</year>). <article-title>Convergent and discriminant validation by the multitrait-multimethod matrix</article-title>. <source>Psychol. Bull.</source> <volume>56</volume>, <fpage>81</fpage>&#x2013;<lpage>105</lpage>. doi: <pub-id pub-id-type="doi">10.1037/h0046016</pub-id></mixed-citation></ref>
<ref id="ref6"><mixed-citation publication-type="confproc"><person-group person-group-type="author"><name><surname>Ch&#x00EA;ne</surname> <given-names>D.</given-names></name> <name><surname>Pillot</surname> <given-names>V.</given-names></name> <name><surname>Bobillier Chaumon</surname> <given-names>M. &#x00C9;.</given-names></name></person-group> (<year>2016</year>). <article-title>Tactile interaction for novice user</article-title>. In <conf-name>International conference on human aspects of IT for the aged population</conf-name> (pp. <fpage>412</fpage>&#x2013;<lpage>423</lpage>). <publisher-name>Springer</publisher-name>, <publisher-loc>Cham</publisher-loc>.</mixed-citation></ref>
<ref id="ref7"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Chi</surname> <given-names>O. H.</given-names></name> <name><surname>Denton</surname> <given-names>G.</given-names></name> <name><surname>Gursoy</surname> <given-names>D.</given-names></name></person-group> (<year>2020</year>). <article-title>Artificially intelligent device use in service delivery: a systematic review, synthesis, and research agenda</article-title>. <source>J. Hosp. Mark. Manag.</source> <volume>29</volume>, <fpage>757</fpage>&#x2013;<lpage>786</lpage>. doi: <pub-id pub-id-type="doi">10.1080/19368623.2020.1721394</pub-id></mixed-citation></ref>
<ref id="ref8"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Choi</surname> <given-names>J. K.</given-names></name> <name><surname>Ji</surname> <given-names>Y. G.</given-names></name></person-group> (<year>2015</year>). <article-title>Investigating the importance of trust on adopting an autonomous vehicle</article-title>. <source>Int. J. Hum. Comput. Interact.</source> <volume>31</volume>, <fpage>692</fpage>&#x2013;<lpage>702</lpage>. doi: <pub-id pub-id-type="doi">10.1080/10447318.2015.1070549</pub-id></mixed-citation></ref>
<ref id="ref9"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Claypoole</surname> <given-names>V. L.</given-names></name> <name><surname>Schroeder</surname> <given-names>B. L.</given-names></name> <name><surname>Mishler</surname> <given-names>A. D.</given-names></name></person-group> (<year>2016</year>). <article-title>Keeping in touch: Tactile interface design for older users</article-title>. <source>Ergonomics in Design</source> <volume>24</volume>, <fpage>18</fpage>&#x2013;<lpage>24</lpage>. doi: <pub-id pub-id-type="doi">10.1177/1064804615611271</pub-id></mixed-citation></ref>
<ref id="ref10"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>De Raedemaeker</surname> <given-names>K.</given-names></name> <name><surname>Foulon</surname> <given-names>I.</given-names></name> <name><surname>Azzopardi</surname> <given-names>R. V.</given-names></name> <name><surname>Lichtert</surname> <given-names>E.</given-names></name> <name><surname>Buyl</surname> <given-names>R.</given-names></name> <name><surname>Topsakal</surname> <given-names>V.</given-names></name> <etal/></person-group>. (<year>2022</year>). <article-title>Audiometric findings in senior adults of 80 years and older</article-title>. <source>Front. Psychol.</source> <volume>13</volume>:<fpage>861555</fpage>. doi: <pub-id pub-id-type="doi">10.3389/fpsyg.2022.861555</pub-id>, PMID: <pub-id pub-id-type="pmid">35936317</pub-id></mixed-citation></ref>
<ref id="ref11"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Eckstein</surname> <given-names>M. K.</given-names></name> <name><surname>Guerra-Carrillo</surname> <given-names>B.</given-names></name> <name><surname>Miller Singley</surname> <given-names>A. T.</given-names></name> <name><surname>Bunge</surname> <given-names>S. A.</given-names></name></person-group> (<year>2017</year>). <article-title>Beyond eye gaze: what else can eye tracking reveal about cognition and cognitive development?</article-title> <source>Dev. Cogn. Neurosci.</source> <volume>25</volume>, <fpage>69</fpage>&#x2013;<lpage>91</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.dcn.2016.11.001</pub-id>, PMID: <pub-id pub-id-type="pmid">27908561</pub-id></mixed-citation></ref>
<ref id="ref12"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Fisk</surname> <given-names>A. D.</given-names></name> <name><surname>Czaja</surname> <given-names>S. J.</given-names></name> <name><surname>Rogers</surname> <given-names>W. A.</given-names></name> <name><surname>Charness</surname> <given-names>N.</given-names></name> <name><surname>Sharit</surname> <given-names>J.</given-names></name></person-group> (<year>2020</year>). <article-title>Designing for older adults: Principles and creative human factors approaches</article-title>. <publisher-name>CRC Press</publisher-name>. doi: <pub-id pub-id-type="doi">10.1201/9781420080681</pub-id></mixed-citation></ref>
<ref id="ref13"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Gefen</surname> <given-names>D.</given-names></name> <name><surname>Karahanna</surname> <given-names>E.</given-names></name> <name><surname>Straub</surname> <given-names>D. W.</given-names></name></person-group> (<year>2003</year>). <article-title>Trust and TAM in online shopping: An integrated model</article-title>. <source>MIS Q.</source>, <volume>27</volume>, <fpage>51</fpage>&#x2013;<lpage>90</lpage>. doi: <pub-id pub-id-type="doi">10.2307/30036519</pub-id></mixed-citation></ref>
<ref id="ref14"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Ghazizadeh</surname> <given-names>M.</given-names></name> <name><surname>Lee</surname> <given-names>J. D.</given-names></name> <name><surname>Boyle</surname> <given-names>L. N.</given-names></name></person-group> (<year>2012</year>). <article-title>Extending the technology acceptance model to assess automation</article-title>. <source>Cogn. Tech. Work</source> <volume>14</volume>, <fpage>39</fpage>&#x2013;<lpage>49</lpage>. doi: <pub-id pub-id-type="doi">10.1007/s10111-011-0194-3</pub-id></mixed-citation></ref>
<ref id="ref15"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Guo</surname> <given-names>S.</given-names></name> <name><surname>Falaki</surname> <given-names>M. H.</given-names></name> <name><surname>Oliver</surname> <given-names>E. A.</given-names></name> <name><surname>Ur Rahman</surname> <given-names>S.</given-names></name> <name><surname>Seth</surname> <given-names>A.</given-names></name> <name><surname>Zaharia</surname> <given-names>M. A.</given-names></name> <etal/></person-group>. (<year>2007</year>). <article-title>Very low-cost internet access using KioskNet</article-title>. <source>ACM SIGCOMM Comput. Commun. Rev.</source> <volume>37</volume>, <fpage>95</fpage>&#x2013;<lpage>100</lpage>. doi: <pub-id pub-id-type="doi">10.1145/1290168.1290181</pub-id></mixed-citation></ref>
<ref id="ref16"><mixed-citation publication-type="other"><person-group person-group-type="author"><name><surname>Hansen</surname> <given-names>J. P.</given-names></name> <name><surname>Hansen</surname> <given-names>D. W.</given-names></name> <name><surname>Johansen</surname> <given-names>A. S.</given-names></name></person-group> (<year>2001</year>). &#x201C;<article-title>Bringing gaze-based interaction back to basics</article-title>&#x201D; in <source>HCI</source>, <fpage>325</fpage>&#x2013;<lpage>329</lpage>.</mixed-citation></ref>
<ref id="ref17"><mixed-citation publication-type="book"><person-group person-group-type="author"><name><surname>Hart</surname> <given-names>S. G.</given-names></name> <name><surname>Staveland</surname> <given-names>L. E.</given-names></name></person-group> (<year>1988</year>). &#x201C;<article-title>Development of NASA-TLX (Task Load Index): Results of empirical and theoretical research</article-title>&#x201D; in <source>Human mental workload</source>. eds. <person-group person-group-type="editor"><name><surname>Hancock</surname> <given-names>P. A.</given-names></name> <name><surname>Meshkati</surname> <given-names>N.</given-names></name></person-group>, vol. <volume>5</volume> (<publisher-loc>Amsterdam</publisher-loc>: <publisher-name>North-Holland</publisher-name>), <fpage>139</fpage>&#x2013;<lpage>183</lpage>.</mixed-citation></ref>
<ref id="ref18"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Hengstler</surname> <given-names>M.</given-names></name> <name><surname>Enkel</surname> <given-names>E.</given-names></name> <name><surname>Duelli</surname> <given-names>S.</given-names></name></person-group> (<year>2016</year>). <article-title>Applied artificial intelligence and trust&#x2014;the case of autonomous vehicles and medical assistance devices</article-title>. <source>Technol. Forecast. Soc. Change</source> <volume>105</volume>, <fpage>105</fpage>&#x2013;<lpage>120</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.techfore.2015.12.014</pub-id></mixed-citation></ref>
<ref id="ref19"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Hoff</surname> <given-names>K. A.</given-names></name> <name><surname>Bashir</surname> <given-names>M.</given-names></name></person-group> (<year>2015</year>). <article-title>Trust in automation: integrating empirical evidence on factors that influence trust</article-title>. <source>Hum. Factors</source> <volume>57</volume>, <fpage>407</fpage>&#x2013;<lpage>434</lpage>. doi: <pub-id pub-id-type="doi">10.1177/0018720814547570</pub-id>, PMID: <pub-id pub-id-type="pmid">25875432</pub-id></mixed-citation></ref>
<ref id="ref20"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Jian</surname> <given-names>J. Y.</given-names></name> <name><surname>Bisantz</surname> <given-names>A. M.</given-names></name> <name><surname>Drury</surname> <given-names>C. G.</given-names></name></person-group> (<year>2000</year>). <article-title>Foundations for an empirically determined scale of trust in automated systems</article-title>. <source>Int. J. Cogn. Ergon.</source> <volume>4</volume>, <fpage>53</fpage>&#x2013;<lpage>71</lpage>. doi: <pub-id pub-id-type="doi">10.1207/S15327566IJCE0401_04</pub-id></mixed-citation></ref>
<ref id="ref21"><mixed-citation publication-type="confproc"><person-group person-group-type="author"><name><surname>Johnston</surname> <given-names>M.</given-names></name> <name><surname>Bangalore</surname> <given-names>S.</given-names></name></person-group> (<year>2004</year>). <article-title>MATCHKiosk: a multimodal interactive city guide</article-title>. In <conf-name>Proceedings of the ACL interactive poster and demonstration sessions</conf-name> (pp. <fpage>222</fpage>&#x2013;<lpage>225</lpage>).</mixed-citation></ref>
<ref id="ref22"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Just</surname> <given-names>M. A.</given-names></name> <name><surname>Carpenter</surname> <given-names>P. A.</given-names></name></person-group> (<year>1976</year>). <article-title>Eye fixations and cognitive processes</article-title>. <source>Cogn. Psychol.</source> <volume>8</volume>, <fpage>441</fpage>&#x2013;<lpage>480</lpage>. doi: <pub-id pub-id-type="doi">10.1016/0010-0285(76)90015-3</pub-id></mixed-citation></ref>
<ref id="ref23"><mixed-citation publication-type="confproc"><person-group person-group-type="author"><name><surname>Kaufman</surname> <given-names>A. E.</given-names></name> <name><surname>Bandopadhay</surname> <given-names>A.</given-names></name> <name><surname>Shaviv</surname> <given-names>B. D.</given-names></name></person-group> (<year>1993</year>). <article-title>An eye tracking computer user interface</article-title>. In <conf-name>Proceedings of 1993 IEEE research properties in virtual reality symposium</conf-name> (pp. <fpage>120</fpage>&#x2013;<lpage>121</lpage>). <publisher-name>IEEE</publisher-name>.</mixed-citation></ref>
<ref id="ref24"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Kaye</surname> <given-names>S. A.</given-names></name> <name><surname>Lewis</surname> <given-names>I.</given-names></name> <name><surname>Freeman</surname> <given-names>J.</given-names></name></person-group> (<year>2018</year>). <article-title>Comparison of self-report and objective measures of driving behavior and road safety: a systematic review</article-title>. <source>J. Saf. Res.</source> <volume>65</volume>, <fpage>141</fpage>&#x2013;<lpage>151</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.jsr.2018.02.012</pub-id>, PMID: <pub-id pub-id-type="pmid">29776523</pub-id></mixed-citation></ref>
<ref id="ref25"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Kim</surname> <given-names>J. C.</given-names></name> <name><surname>Laine</surname> <given-names>T. H.</given-names></name> <name><surname>&#x00C5;hlund</surname> <given-names>C.</given-names></name></person-group> (<year>2021</year>). <article-title>Multimodal interaction systems based on internet of things and augmented reality: a systematic literature review</article-title>. <source>Appl. Sci.</source> <volume>11</volume>:<fpage>1738</fpage>. doi: <pub-id pub-id-type="doi">10.3390/app11041738</pub-id></mixed-citation></ref>
<ref id="ref26"><mixed-citation publication-type="confproc"><person-group person-group-type="author"><name><surname>Kim</surname> <given-names>M.</given-names></name> <name><surname>Lee</surname> <given-names>M. K.</given-names></name> <name><surname>Dabbish</surname> <given-names>L.</given-names></name></person-group> (<year>2015</year>). <article-title>Shop-i: gaze based interaction in the physical world for in-store social shopping experience</article-title>. In <conf-name>Proceedings of the 33rd Annual ACM Conference Extended Abstracts on Human Factors in Computing Systems</conf-name> (pp. <fpage>1253</fpage>&#x2013;<lpage>1258</lpage>).</mixed-citation></ref>
<ref id="ref27"><mixed-citation publication-type="confproc"><person-group person-group-type="author"><name><surname>Lee</surname> <given-names>Y.</given-names></name> <name><surname>Jeon</surname> <given-names>H.</given-names></name> <name><surname>Kim</surname> <given-names>H. K.</given-names></name> <name><surname>Park</surname> <given-names>S.</given-names></name></person-group> (<year>2020</year>). <article-title>Literature review on accessibility guidelines for self-service terminals</article-title>. <conf-name>Proceedings of the ACHI</conf-name>.</mixed-citation></ref>
<ref id="ref28"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Lee</surname> <given-names>J. D.</given-names></name> <name><surname>See</surname> <given-names>K. A.</given-names></name></person-group> (<year>2004</year>). <article-title>Trust in automation: designing for appropriate reliance</article-title>. <source>Hum. Factors</source> <volume>46</volume>, <fpage>50</fpage>&#x2013;<lpage>80</lpage>. doi: <pub-id pub-id-type="doi">10.1518/hfes.46.1.50.30392</pub-id>, PMID: <pub-id pub-id-type="pmid">15151155</pub-id></mixed-citation></ref>
<ref id="ref29"><mixed-citation publication-type="confproc"><person-group person-group-type="author"><name><surname>Leonardi</surname> <given-names>C.</given-names></name> <name><surname>Albertini</surname> <given-names>A.</given-names></name> <name><surname>Pianesi</surname> <given-names>F.</given-names></name> <name><surname>Zancanaro</surname> <given-names>M.</given-names></name></person-group> (<year>2010</year>). <article-title>An exploratory study of a touch-based gestural interface for elderly</article-title>. In <conf-name>Proceedings of the 6th nordic conference on human-computer interaction: Extending boundaries</conf-name> (pp. <fpage>845</fpage>&#x2013;<lpage>850</lpage>).</mixed-citation></ref>
<ref id="ref30"><mixed-citation publication-type="book"><person-group person-group-type="author"><name><surname>Likert</surname> <given-names>R.</given-names></name></person-group> (<year>2017</year>). &#x201C;<article-title>The method of constructing an attitude scale</article-title>&#x201D; in <source>Scaling</source> (<publisher-loc>London and New York</publisher-loc>: <publisher-name>Routledge</publisher-name>), <fpage>233</fpage>&#x2013;<lpage>242</lpage>.</mixed-citation></ref>
<ref id="ref31"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Lowndes</surname> <given-names>B. R.</given-names></name> <name><surname>Forsyth</surname> <given-names>K. L.</given-names></name> <name><surname>Blocker</surname> <given-names>R. C.</given-names></name> <name><surname>Dean</surname> <given-names>P. G.</given-names></name> <name><surname>Truty</surname> <given-names>M. J.</given-names></name> <name><surname>Heller</surname> <given-names>S. F.</given-names></name> <etal/></person-group>. (<year>2020</year>). <article-title>NASA-TLX assessment of surgeon workload variation across specialties</article-title>. <source>Ann. Surg.</source> <volume>271</volume>, <fpage>686</fpage>&#x2013;<lpage>692</lpage>. doi: <pub-id pub-id-type="doi">10.1097/SLA.0000000000003058</pub-id>, PMID: <pub-id pub-id-type="pmid">30247331</pub-id></mixed-citation></ref>
<ref id="ref32"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Ma</surname> <given-names>J.</given-names></name> <name><surname>Zuo</surname> <given-names>Y.</given-names></name> <name><surname>Du</surname> <given-names>H.</given-names></name> <name><surname>Wang</surname> <given-names>Y.</given-names></name> <name><surname>Tan</surname> <given-names>M.</given-names></name> <name><surname>Li</surname> <given-names>J.</given-names></name></person-group> (<year>2025</year>). <article-title>Interactive output modalities Design for Enhancement of user trust experience in highly autonomous driving</article-title>. <source>Int. J. Hum. Comput. Interact.</source> <volume>41</volume>, <fpage>6172</fpage>&#x2013;<lpage>6190</lpage>. doi: <pub-id pub-id-type="doi">10.1080/10447318.2024.2375697</pub-id></mixed-citation></ref>
<ref id="ref33"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Manzke</surname> <given-names>J. M.</given-names></name> <name><surname>Egan</surname> <given-names>D. H.</given-names></name> <name><surname>Felix</surname> <given-names>D.</given-names></name> <name><surname>Krueger</surname> <given-names>H.</given-names></name></person-group> (<year>1998</year>). <article-title>What makes an automated teller machine usable by blind users?</article-title> <source>Ergonomics</source> <volume>41</volume>, <fpage>982</fpage>&#x2013;<lpage>999</lpage>. doi: <pub-id pub-id-type="doi">10.1080/001401398186540</pub-id>, PMID: <pub-id pub-id-type="pmid">9674373</pub-id></mixed-citation></ref>
<ref id="ref34"><mixed-citation publication-type="confproc"><person-group person-group-type="author"><name><surname>Mikulski</surname> <given-names>D. G.</given-names></name></person-group> (<year>2014</year>). <article-title>Trust-based controller for convoy string stability</article-title>. <conf-name>In 2014 IEEE symposium on computational intelligence in vehicles and transportation systems (CIVTS)</conf-name> (pp. <fpage>69</fpage>&#x2013;<lpage>75</lpage>). <publisher-name>IEEE</publisher-name>.</mixed-citation></ref>
<ref id="ref35"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Niu</surname> <given-names>Y. F.</given-names></name> <name><surname>Gao</surname> <given-names>Y.</given-names></name> <name><surname>Zhang</surname> <given-names>Y. T.</given-names></name> <name><surname>Xue</surname> <given-names>C. Q.</given-names></name> <name><surname>Yang</surname> <given-names>L. X.</given-names></name></person-group> (<year>2019</year>). <article-title>Improving eye&#x2013;computer interaction interface design: Ergonomic investigations of the optimum target size and gaze-triggering dwell time</article-title>. <source>J. Eye Mov. Res.</source> <volume>12</volume>. doi: <pub-id pub-id-type="doi">10.16910/jemr.12.3.8</pub-id></mixed-citation></ref>
<ref id="ref36"><mixed-citation publication-type="book"><person-group person-group-type="author"><name><surname>Norman</surname> <given-names>D.</given-names></name></person-group> (<year>2013</year>). <source>The Design of EverydayThings: Revised and Expanded Edition</source>. <publisher-loc>New York</publisher-loc>:<publisher-name>Basic Books</publisher-name>.</mixed-citation></ref>
<ref id="ref37"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Olsen</surname> <given-names>A.</given-names></name></person-group> (<year>2012</year>). <article-title>The Tobii I-VT fixation filter: algorithm description</article-title>. <source>Tobii Technol.</source> <volume>21</volume>:<fpage>15</fpage>.</mixed-citation></ref>
<ref id="ref38"><mixed-citation publication-type="book"><person-group person-group-type="author"><name><surname>Oviatt</surname> <given-names>S.</given-names></name></person-group> (<year>2022</year>). &#x201C;<article-title>Multimodal interaction, interfaces, and analytics</article-title>&#x201D; in <source>Handbook of human computer interaction</source> (<publisher-loc>Cham</publisher-loc>: <publisher-name>Springer International Publishing</publisher-name>), <fpage>1</fpage>&#x2013;<lpage>29</lpage>.</mixed-citation></ref>
<ref id="ref39"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Paradi</surname> <given-names>J. C.</given-names></name> <name><surname>Ghazarian-Rock</surname> <given-names>A.</given-names></name></person-group> (<year>1998</year>). <article-title>A framework to evaluate video banking kiosks</article-title>. <source>Omega</source> <volume>26</volume>, <fpage>523</fpage>&#x2013;<lpage>539</lpage>. doi: <pub-id pub-id-type="doi">10.1016/S0305-0483(97)00080-7</pub-id></mixed-citation></ref>
<ref id="ref40"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Parasuraman</surname> <given-names>R.</given-names></name></person-group> (<year>2000</year>). <article-title>Designing automation for human use: empirical studies and quantitative models</article-title>. <source>Ergonomics</source> <volume>43</volume>, <fpage>931</fpage>&#x2013;<lpage>951</lpage>. doi: <pub-id pub-id-type="doi">10.1080/001401300409125</pub-id>, PMID: <pub-id pub-id-type="pmid">10929828</pub-id></mixed-citation></ref>
<ref id="ref7010"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Pearson</surname> <given-names>K.</given-names></name></person-group> (<year>1895</year>). <article-title>VII. Note on regression and inheritance in the case of two parents</article-title>. <source>Proceedings of the Royal Society of London</source> <volume>58</volume>, <fpage>240</fpage>&#x2013;<lpage>242</lpage>. doi: <pub-id pub-id-type="doi">10.1098/rspl.1895.0041</pub-id></mixed-citation></ref>
<ref id="ref41"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Pfeuffer</surname> <given-names>K.</given-names></name> <name><surname>Abdrabou</surname> <given-names>Y.</given-names></name> <name><surname>Esteves</surname> <given-names>A.</given-names></name> <name><surname>Rivu</surname> <given-names>R.</given-names></name> <name><surname>Abdelrahman</surname> <given-names>Y.</given-names></name> <name><surname>Meitner</surname> <given-names>S.</given-names></name> <etal/></person-group>. (<year>2021</year>). <article-title>ARtention: a design space for gaze-adaptive user interfaces in augmented reality</article-title>. <source>Comput. Graph.</source> <volume>95</volume>, <fpage>1</fpage>&#x2013;<lpage>12</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.cag.2021.01.001</pub-id></mixed-citation></ref>
<ref id="ref42"><mixed-citation publication-type="book"><person-group person-group-type="author"><name><surname>Poole</surname> <given-names>A.</given-names></name> <name><surname>Ball</surname> <given-names>L. J.</given-names></name></person-group> (<year>2006</year>). &#x201C;<article-title>Eye tracking in HCI and usability research</article-title>&#x201D; in <source>Encyclopedia of human computer interaction</source> (<publisher-loc>Hershey, PA</publisher-loc>: <publisher-name>IGI Global Scientific Publishing</publisher-name>), <fpage>211</fpage>&#x2013;<lpage>219</lpage>.</mixed-citation></ref>
<ref id="ref43"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Pop</surname> <given-names>V. L.</given-names></name> <name><surname>Shrewsbury</surname> <given-names>A.</given-names></name> <name><surname>Durso</surname> <given-names>F. T.</given-names></name></person-group> (<year>2015</year>). <article-title>Individual differences in the calibration of trust in automation</article-title>. <source>Hum. Factors</source> <volume>57</volume>, <fpage>545</fpage>&#x2013;<lpage>556</lpage>. doi: <pub-id pub-id-type="doi">10.1177/0018720814564422</pub-id>, PMID: <pub-id pub-id-type="pmid">25977317</pub-id></mixed-citation></ref>
<ref id="ref44"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Portet</surname> <given-names>F.</given-names></name> <name><surname>Vacher</surname> <given-names>M.</given-names></name> <name><surname>Golanski</surname> <given-names>C.</given-names></name> <name><surname>Roux</surname> <given-names>C.</given-names></name> <name><surname>Meillon</surname> <given-names>B.</given-names></name></person-group> (<year>2013</year>). <article-title>Design and evaluation of a smart home voice interface for the elderly: acceptability and objection aspects</article-title>. <source>Pers. Ubiquit. Comput.</source> <volume>17</volume>, <fpage>127</fpage>&#x2013;<lpage>144</lpage>. doi: <pub-id pub-id-type="doi">10.1007/s00779-011-0470-5</pub-id></mixed-citation></ref>
<ref id="ref45"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Roth</surname> <given-names>T. N.</given-names></name> <name><surname>Hanebuth</surname> <given-names>D.</given-names></name> <name><surname>Probst</surname> <given-names>R.</given-names></name></person-group> (<year>2011</year>). <article-title>Prevalence of age-related hearing loss in Europe: a review</article-title>. <source>Eur. Arch. Otorrinolaringol.</source> <volume>268</volume>, <fpage>1101</fpage>&#x2013;<lpage>1107</lpage>. doi: <pub-id pub-id-type="doi">10.1007/s00405-011-1597-8</pub-id>, PMID: <pub-id pub-id-type="pmid">21499871</pub-id></mixed-citation></ref>
<ref id="ref46"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Rozado</surname> <given-names>D.</given-names></name> <name><surname>Agustin</surname> <given-names>J. S.</given-names></name> <name><surname>Rodriguez</surname> <given-names>F. B.</given-names></name> <name><surname>Varona</surname> <given-names>P.</given-names></name></person-group> (<year>2012</year>). <article-title>Gliding and saccadic gaze gesture recognition in real time</article-title>. <source>ACM Transactions on Interactive Intelligent Systems (TiiS)</source> <volume>1</volume>, <fpage>1</fpage>&#x2013;<lpage>27</lpage>. doi: <pub-id pub-id-type="doi">10.1145/2070719.2070723</pub-id></mixed-citation></ref>
<ref id="ref47"><mixed-citation publication-type="confproc"><person-group person-group-type="author"><name><surname>Saji&#x0107;</surname> <given-names>M.</given-names></name> <name><surname>Bundalo</surname> <given-names>D.</given-names></name> <name><surname>Vidovi&#x0107;</surname> <given-names>&#x017D;.</given-names></name> <name><surname>Bundalo</surname> <given-names>Z.</given-names></name> <name><surname>Lalic</surname> <given-names>D.</given-names></name></person-group> (<year>2021</year>). <article-title>Smart digital terminal devices with speech recognition and speech control</article-title>. In <conf-name>2021 10th Mediterranean conference on embedded computing (MECO)</conf-name> (pp. <fpage>1</fpage>&#x2013;<lpage>5</lpage>). <publisher-name>IEEE</publisher-name>.</mixed-citation></ref>
<ref id="ref48"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Sandnes</surname> <given-names>F. E.</given-names></name> <name><surname>Tan</surname> <given-names>T. B.</given-names></name> <name><surname>Johansen</surname> <given-names>A.</given-names></name> <name><surname>Sulic</surname> <given-names>E.</given-names></name> <name><surname>Vesterhus</surname> <given-names>E.</given-names></name> <name><surname>Iversen</surname> <given-names>E. R.</given-names></name></person-group> (<year>2012</year>). <article-title>Making touch-based kiosks accessible to blind users through simple gestures</article-title>. <source>Universal Access Inf. Soc.</source> <volume>11</volume>, <fpage>421</fpage>&#x2013;<lpage>431</lpage>. doi: <pub-id pub-id-type="doi">10.1007/s10209-011-0258-4</pub-id></mixed-citation></ref>
<ref id="ref49"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Sato</surname> <given-names>H.</given-names></name> <name><surname>Abe</surname> <given-names>K.</given-names></name> <name><surname>Ohi</surname> <given-names>S.</given-names></name> <name><surname>Ohyama</surname> <given-names>M.</given-names></name></person-group> (<year>2018</year>). <article-title>A text input system based on information of voluntary blink and eye-gaze using an image analysis</article-title>. <source>Electr. Commun. Japan</source> <volume>101</volume>, <fpage>9</fpage>&#x2013;<lpage>22</lpage>. doi: <pub-id pub-id-type="doi">10.1002/ecj.12025</pub-id></mixed-citation></ref>
<ref id="ref50"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Seaborn</surname> <given-names>K.</given-names></name> <name><surname>Sekiguchi</surname> <given-names>T.</given-names></name> <name><surname>Tokunaga</surname> <given-names>S.</given-names></name> <name><surname>Miyake</surname> <given-names>N. P.</given-names></name> <name><surname>Otake-Matsuura</surname> <given-names>M.</given-names></name></person-group> (<year>2023</year>). <article-title>Voice over body? Older adults&#x2019; reactions to robot and voice assistant facilitators of group conversation</article-title>. <source>Int. J. Soc. Robot.</source> <volume>15</volume>, <fpage>143</fpage>&#x2013;<lpage>163</lpage>. doi: <pub-id pub-id-type="doi">10.1007/s12369-022-00925-7</pub-id>, PMID: <pub-id pub-id-type="pmid">36406778</pub-id></mixed-citation></ref>
<ref id="ref51"><mixed-citation publication-type="confproc"><person-group person-group-type="author"><name><surname>Shang</surname> <given-names>H.</given-names></name> <name><surname>Shi</surname> <given-names>X.</given-names></name> <name><surname>Wang</surname> <given-names>C.</given-names></name></person-group> (<year>2020</year>). <article-title>Research on the design of medical self-service terminal for the elderly</article-title>. In <conf-name>Advances in human factors and ergonomics in healthcare and medical devices: proceedings of the AHFE 2020 virtual conference on human factors and ergonomics in healthcare and medical devices, July 16&#x2013;20, 2020, USA</conf-name> (pp. <fpage>335</fpage>&#x2013;<lpage>349</lpage>). <publisher-name>Springer International Publishing</publisher-name>.</mixed-citation></ref>
<ref id="ref52"><mixed-citation publication-type="confproc"><person-group person-group-type="author"><name><surname>Slay</surname> <given-names>H.</given-names></name> <name><surname>Wentworth</surname> <given-names>P.</given-names></name> <name><surname>Locke</surname> <given-names>J.</given-names></name></person-group> (<year>2006</year>). <article-title>BingBee, an information kiosk for social enablement in marginalized communities</article-title>. In <conf-name>Proceedings of the 2006 annual research conference of the South African institute of computer scientists and information technologists on IT research in developing countries</conf-name> (pp. <fpage>107</fpage>&#x2013;<lpage>116</lpage>).</mixed-citation></ref>
<ref id="ref2001"><mixed-citation publication-type="book"><person-group person-group-type="author"><name><surname>Sultana</surname> <given-names>A.</given-names></name> <name><surname>Moffatt</surname> <given-names>K.</given-names></name></person-group> (<year>2019</year>). <article-title>Effects of aging on small target selection with touch input</article-title>. <source>ACM Transactions on Accessible Computing (TACCESS)</source> <volume>12</volume>, <fpage>1</fpage>&#x2013;<lpage>35</lpage>.</mixed-citation></ref>
<ref id="ref53"><mixed-citation publication-type="confproc"><person-group person-group-type="author"><name><surname>Vildjiounaite</surname> <given-names>E.</given-names></name> <name><surname>M&#x00E4;kel&#x00E4;</surname> <given-names>S. M.</given-names></name> <name><surname>Lindholm</surname> <given-names>M.</given-names></name> <name><surname>Riihim&#x00E4;ki</surname> <given-names>R.</given-names></name> <name><surname>Kyll&#x00F6;nen</surname> <given-names>V.</given-names></name> <name><surname>M&#x00E4;ntyj&#x00E4;rvi</surname> <given-names>J.</given-names></name> <etal/></person-group>. (<year>2006</year>). <article-title>Unobtrusive multimodal biometrics for ensuring privacy and information security with personal devices</article-title>. In <conf-name>International Conference on Pervasive Computing</conf-name>(pp. <fpage>187</fpage>&#x2013;<lpage>201</lpage>). <publisher-loc>Berlin, Heidelberg</publisher-loc>: <publisher-name>Springer Berlin Heidelberg</publisher-name>.</mixed-citation></ref>
<ref id="ref54"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Ya-feng</surname> <given-names>N.</given-names></name> <name><surname>Jin</surname> <given-names>L.</given-names></name> <name><surname>Jia-qi</surname> <given-names>C.</given-names></name> <name><surname>Wen-jun</surname> <given-names>Y.</given-names></name> <name><surname>Hong-rui</surname> <given-names>Z.</given-names></name> <name><surname>Jia-xin</surname> <given-names>H.</given-names></name> <etal/></person-group>. (<year>2022</year>). <article-title>Research on visual representation of icon colour in eye-controlled systems</article-title>. <source>Adv. Eng. Inform.</source> <volume>52</volume>:<fpage>101570</fpage>. doi: <pub-id pub-id-type="doi">10.1016/j.aei.2022.101570</pub-id></mixed-citation></ref>
<ref id="ref55"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Zhang</surname> <given-names>T.</given-names></name> <name><surname>Liu</surname> <given-names>X.</given-names></name> <name><surname>Zeng</surname> <given-names>W.</given-names></name> <name><surname>Tao</surname> <given-names>D.</given-names></name> <name><surname>Li</surname> <given-names>G.</given-names></name> <name><surname>Qu</surname> <given-names>X.</given-names></name></person-group> (<year>2023</year>). <article-title>Input modality matters: a comparison of touch, speech, and gesture based in-vehicle interaction</article-title>. <source>Appl. Ergon.</source> <volume>108</volume>:<fpage>103958</fpage>. doi: <pub-id pub-id-type="doi">10.1016/j.apergo.2022.103958</pub-id>, PMID: <pub-id pub-id-type="pmid">36587503</pub-id></mixed-citation></ref>
<ref id="ref56"><mixed-citation publication-type="journal"><person-group person-group-type="author"><name><surname>Zhou</surname> <given-names>J.</given-names></name> <name><surname>Chourasia</surname> <given-names>A.</given-names></name> <name><surname>Vanderheiden</surname> <given-names>G.</given-names></name></person-group> (<year>2017</year>). <article-title>Interface adaptation to novice older adults&#x2019; mental models through concrete metaphors</article-title>. <source>Int. J. Hum. Comput. Interact.</source> <volume>33</volume>, <fpage>592</fpage>&#x2013;<lpage>606</lpage>. doi: <pub-id pub-id-type="doi">10.1080/10447318.2016.1265827</pub-id></mixed-citation></ref>
</ref-list>
<fn-group><fn id="fn0002" fn-type="custom" custom-type="edited-by"><p>Edited by: <ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/183771/overview">Carlos Duarte</ext-link>, University of Lisbon, Portugal</p></fn>
<fn id="fn0003" fn-type="custom" custom-type="reviewed-by"><p>Reviewed by: <ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/2320889/overview">Maikel L&#x00E1;zaro P&#x00E9;rez Gort</ext-link>, Ca' Foscari University of Venice, Italy; <ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/1905746/overview">Qiwei Li</ext-link>, California State University, Fresno, United States</p></fn>
<fn id="fn0001"><p><sup>1</sup><ext-link xlink:href="https://ai.aliyun.com/nls/tts" ext-link-type="uri">https://ai.aliyun.com/nls/tts</ext-link></p></fn>
</fn-group></back>
</article>