<?xml version="1.0" encoding="UTF-8" standalone="no"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xml:lang="EN" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Psychol.</journal-id>
<journal-title>Frontiers in Psychology</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Psychol.</abbrev-journal-title>
<issn pub-type="epub">1664-1078</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fpsyg.2025.1641212</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Psychology</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Optimizing academic engagement and mental health through AI: an experimental study on LLM integration in higher education</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name><surname>Zhang</surname> <given-names>Min</given-names></name>
<xref ref-type="corresp" rid="c001"><sup>&#x0002A;</sup></xref>
<xref ref-type="author-notes" rid="fn001"><sup>&#x02020;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/3070131/overview"/>
<role content-type="https://credit.niso.org/contributor-roles/conceptualization/"/>
<role content-type="https://credit.niso.org/contributor-roles/resources/"/>
<role content-type="https://credit.niso.org/contributor-roles/visualization/"/>
<role content-type="https://credit.niso.org/contributor-roles/validation/"/>
<role content-type="https://credit.niso.org/contributor-roles/formal-analysis/"/>
<role content-type="https://credit.niso.org/contributor-roles/funding-acquisition/"/>
<role content-type="https://credit.niso.org/contributor-roles/project-administration/"/>
<role content-type="https://credit.niso.org/contributor-roles/writing-original-draft/"/>
<role content-type="https://credit.niso.org/contributor-roles/data-curation/"/>
<role content-type="https://credit.niso.org/contributor-roles/supervision/"/>
<role content-type="https://credit.niso.org/contributor-roles/investigation/"/>
<role content-type="https://credit.niso.org/contributor-roles/methodology/"/>
<role content-type="https://credit.niso.org/contributor-roles/writing-review-editing/"/>
<role content-type="https://credit.niso.org/contributor-roles/software/"/>
</contrib>
</contrib-group>
<aff><institution>College of Humanities and Arts, Xi&#x00027;an International University</institution>, <addr-line>Xi&#x02018;an Shaanxi</addr-line>, <country>China</country></aff>
<author-notes>
<fn fn-type="edited-by"><p>Edited by: Xinghua Liu, Shanghai Jiao Tong University, China</p></fn>
<fn fn-type="edited-by"><p>Reviewed by: Joseline Santos, Bulacan State University, Philippines</p>
<p>Vivek Bhardwaj, Manipal University Jaipur, India</p></fn>
<corresp id="c001">&#x0002A;Correspondence: Min Zhang <email>xaiu13176&#x00040;xaiu.edu.cn</email></corresp>
<fn fn-type="other" id="fn001"><p>&#x02020;ORCID: Min Zhang <ext-link ext-link-type="uri" xlink:href="https://orcid.org/0009-0000-7087-7738">orcid.org/0009-0000-7087-7738</ext-link></p></fn></author-notes>
<pub-date pub-type="epub">
<day>01</day>
<month>10</month>
<year>2025</year>
</pub-date>
<pub-date pub-type="collection">
<year>2025</year>
</pub-date>
<volume>16</volume>
<elocation-id>1641212</elocation-id>
<history>
<date date-type="received">
<day>04</day>
<month>06</month>
<year>2025</year>
</date>
<date date-type="accepted">
<day>25</day>
<month>08</month>
<year>2025</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x000A9; 2025 Zhang.</copyright-statement>
<copyright-year>2025</copyright-year>
<copyright-holder>Zhang</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/"><p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p></license>
</permissions>
<abstract>
<sec>
<title>Background:</title>
<p>In alignment with UNESCO&#x00027;s Sustainable Development Goal 4 (SDG4), which advocates for inclusive and equitable quality education, the integration of Artificial Intelligence tools&#x02014;particularly Large Language Models (LLMs)&#x02014;presents promising opportunities for transforming higher education. Despite this potential, empirical research remains scarce regarding the effects of LLM use on students&#x00027; academic performance, mental well-being, and engagement, especially across different modes of implementation.</p></sec>
<sec>
<title>Objective</title>
<p>This experimental study investigated whether a guided, pedagogically grounded use of LLMs enhances students&#x00027; academic writing quality, perceived mental health, and academic engagement more effectively than either unguided use or no exposure to LLMs. The study contributes to UNESCO&#x00027;s &#x0201C;Futures of Education&#x0201D; vision by exploring how structured AI use may foster more inclusive and empowering learning environments.</p></sec>
<sec>
<title>Method</title>
<p>A total of 246 undergraduate students were randomly assigned to one of three conditions: guided LLM use, unguided LLM use, or a control group with no LLM access. Participants completed a critical writing task and standardized instruments measuring academic engagement and mental well-being. Prior academic achievement was controlled for, and writing quality was assessed using Grammarly for Education.</p></sec>
<sec>
<title>Results</title>
<p>Students in the guided LLM condition achieved significantly higher scores in writing quality and academic engagement compared to the control group, with large and moderate effect sizes, respectively. Modest improvements in mental health indicators were also observed. By contrast, unguided use yielded moderate gains in writing quality but did not produce significant effects on engagement or well-being.</p></sec>
<sec>
<title>Conclusion</title>
<p>The findings highlight the critical role of intentional instructional design in the educational integration of AI tools. Structured guidance not only optimizes academic outcomes but also supports students&#x00027; wellbeing and inclusion. This study offers empirical evidence to inform ongoing debates on how digital innovation can contribute to reducing educational disparities and advancing equitable learning in the post-pandemic era.</p></sec></abstract>
<kwd-group>
<kwd>Large Language Models</kwd>
<kwd>academic writing</kwd>
<kwd>mental health</kwd>
<kwd>student engagement</kwd>
<kwd>higher education</kwd>
<kwd>guided instruction</kwd>
<kwd>educational technology</kwd>
<kwd>AI-assisted learning</kwd>
</kwd-group>
<counts>
<fig-count count="0"/>
<table-count count="7"/>
<equation-count count="0"/>
<ref-count count="71"/>
<page-count count="15"/>
<word-count count="12267"/>
</counts>
<custom-meta-wrap>
<custom-meta>
<meta-name>section-at-acceptance</meta-name>
<meta-value>Educational Psychology</meta-value>
</custom-meta>
</custom-meta-wrap>
</article-meta>
</front>
<body>
<sec sec-type="intro" id="s1">
<title>1 Introduction</title>
<p>The accelerated integration of Artificial Intelligence (AI) into higher education is reshaping the academic landscape at an unprecedented pace. In particular, the emergence of Large Language Models (LLMs), such as OpenAI&#x00027;s ChatGPT, has generated both enthusiasm and apprehension among educators and policymakers. While some institutions have embraced these technologies as tools to enhance personalization, accessibility, and innovation in teaching, others have expressed concern about academic integrity, student dependency, and the erosion of critical thinking. The current moment thus presents a pivotal opportunity&#x02014;and challenge&#x02014;for universities to evaluate the pedagogical value of LLMs and their broader impact on student learning (<xref ref-type="bibr" rid="B61">Sharma et al., 2025</xref>).</p>
<p>The urgency of this evaluation is underscored by the widespread and rapid adoption of LLMs in academic contexts. For instance, recent headlines such as &#x0201C;More than half of UK undergraduates say they use AI to help with essays&#x0201D; (<xref ref-type="bibr" rid="B1">Adams, 2024</xref>) reflect a shift in student practices that institutions are still struggling to regulate or harness effectively (<xref ref-type="bibr" rid="B20">Fritz et al., 2024</xref>). Despite this proliferation, empirical evidence remains limited, especially regarding how different modalities of LLM implementation&#x02014;guided versus unguided use&#x02014;affect students&#x00027; academic performance, mental wellbeing, and engagement. Given the scale and speed of adoption, addressing this gap has become an urgent priority for educators and researchers alike.</p>
<p>In this context, international policy frameworks such as UNESCO&#x00027;s Sustainable Development Goal 4 (SDG4), which promotes inclusive, equitable, and quality education, and the &#x0201C;Futures of Education&#x0201D; initiative offer critical guidance. The latter calls for reimagining how knowledge is produced, valued, and shared, with a strong emphasis on human-centered, ethically grounded digital innovation. This vision aligns closely with the need to understand how emerging technologies like LLMs can support not only academic excellence, but also psychological wellbeing and inclusive engagement among students.</p>
<p>Integrating AI into university education is not merely a matter of technological adaptation; it compels a re-examination of core pedagogical processes. Academic writing, for example, remains a central yet often stressful academic demand, both difficult to master and to assess objectively (<xref ref-type="bibr" rid="B3">Ayeni et al., 2024</xref>). Simultaneously, student mental health has emerged as a pressing concern in higher education, particularly within competitive and international environments (<xref ref-type="bibr" rid="B46">Molodynski et al., 2021</xref>). Academic engagement&#x02014;the emotional, cognitive, and behavioral investment in learning&#x02014;is equally critical, yet sensitive to instructional design and motivation (<xref ref-type="bibr" rid="B36">Lin, 2024</xref>).</p>
<p>Although interest in educational applications of LLMs is growing (<xref ref-type="bibr" rid="B48">Ng et al., 2024</xref>), including recent efforts to synthesize their contributions to personalized learning (<xref ref-type="bibr" rid="B61">Sharma et al., 2025</xref>), few studies have experimentally assessed their impact on these three domains within controlled settings (<xref ref-type="bibr" rid="B23">Jungherr, 2023</xref>). Furthermore, how these tools are introduced&#x02014;whether with structured guidance or left to student discretion&#x02014;may significantly influence their effectiveness and students&#x00027; emotional and cognitive responses to academic tasks (<xref ref-type="bibr" rid="B8">Chang, 2024</xref>).</p>
<p>The present study addresses this research gap by experimentally examining the effects of guided versus unguided use of an LLM on undergraduate students&#x00027; academic writing quality, perceived mental health, and academic engagement. Conducted in an international university in China, the study employed a standardized writing task and randomized group assignment (guided use, unguided use, control) to determine whether structured integration enhances learning outcomes and wellbeing. The results aim to inform evidence-based, ethical practices for AI integration in higher education and contribute to global discussions on how digital tools can advance more inclusive, resilient, and human-centered academic environments.</p>
</sec>
<sec id="s2">
<title>2 Theoretical framework and empirical background</title>
<sec>
<title>2.1 LLMs in higher education</title>
<p>LLMs, such as GPT-4, are increasingly present in higher education as tools to assist with language production, research synthesis, and academic writing (<xref ref-type="bibr" rid="B39">Lu et al., 2024</xref>). Their growing use among university students has sparked institutional interest in understanding how these tools influence learning outcomes. However, emerging evidence suggests that the pedagogical value of LLMs depends less on their availability than on the instructional design that accompanies their use (<xref ref-type="bibr" rid="B53">Robleto et al., 2024</xref>).</p>
<p>A useful framework for analyzing the educational integration of technology is the Technological Pedagogical Content Knowledge (TPACK) model developed by (<xref ref-type="bibr" rid="B45">Mishra and Koehler 2006</xref>). TPACK posits that effective technology-enhanced instruction requires the intersection of three types of knowledge: disciplinary content, pedagogical strategies, and technological tools. In this model, the mere introduction of digital resources does not guarantee meaningful learning. Rather, it is the thoughtful alignment of those tools with pedagogical goals and disciplinary content that fosters deep understanding and transferable skills.</p>
<p>This framework is particularly relevant to the use of LLMs. In a guided implementation, students receive explicit instructions on how to use the model to support key aspects of academic writing&#x02014;such as developing argument structure, paraphrasing source material, or revising according to disciplinary conventions (<xref ref-type="bibr" rid="B68">Yan et al., 2024</xref>). This structured use of the tool reflects the TPACK ideal: technology embedded within a coherent pedagogical plan.</p>
<p>By contrast, unguided use of LLMs lacks this intentional alignment. Although students may independently explore the tool&#x00027;s capabilities, they do so without pedagogical framing, which may result in inconsistent outcomes. Unguided users might underuse the tool, rely on it uncritically, or fail to recognize its limitations (<xref ref-type="bibr" rid="B65">Wang, 2022</xref>). Finally, students in the control group, with no access to LLMs, must rely entirely on their prior writing skills and internal resources. While this condition mirrors traditional academic expectations, it may pose additional cognitive and emotional challenges for students with lower confidence or weaker academic preparation (<xref ref-type="bibr" rid="B3">Ayeni et al., 2024</xref>). Building on this theoretical foundation, the present study investigates how different instructional approaches to LLM use&#x02014;guided, unguided, or absent&#x02014;affect academic writing, mental well-being, and engagement. The TPACK framework supports the hypothesis that pedagogically framed LLM use will yield superior outcomes across all domains.</p>
</sec>
</sec>
<sec id="s3">
<title>3 Academic writing quality</title>
<p>Academic writing is a core component of higher education, particularly in the humanities and social sciences. It demands clarity of argument, mastery of disciplinary conventions, and grammatical precision&#x02014;competencies that students often struggle to develop and instructors find difficult to evaluate objectively (<xref ref-type="bibr" rid="B3">Ayeni et al., 2024</xref>). Studies have shown that structured instructional approaches, such as modeling, scaffolding, and feedback, consistently improve students&#x00027; writing skills (<xref ref-type="bibr" rid="B12">De La Paz, 2005</xref>). LLMs offer a new form of writing support, assisting students in generating ideas, organizing content, and refining their language. Early findings suggest that students who use these tools during the planning and revision phases may produce more coherent and technically accurate texts (<xref ref-type="bibr" rid="B29">Lee, 2023</xref>). However, the benefits of LLMs are not automatic. Their effectiveness hinges on how they are introduced and used in educational contexts.</p>
<p>From a TPACK perspective, guided LLM use can enhance academic writing by aligning the tool&#x00027;s features with pedagogical goals. Instructors may, for instance, teach students how to use the model to outline arguments or critically revise text while warning against uncritical copying or overreliance. This structured integration supports metacognitive engagement and allows students to internalize academic writing conventions. In contrast, students in the unguided condition may fail to use the tool optimally. Without pedagogical framing, they might use it only superficially&#x02014;for grammar correction or idea generation&#x02014;without fully engaging with the writing process. Additionally, they may be more prone to accept AI-generated suggestions uncritically, leading to errors in reasoning, style, or source use (<xref ref-type="bibr" rid="B65">Wang, 2022</xref>).</p>
<p>For students in the control condition, the writing task requires managing all stages of composition without external digital support. While this reflects a traditional academic scenario, it may impose greater cognitive demands and limit writing quality, especially for students lacking confidence or fluency in academic writing (<xref ref-type="bibr" rid="B3">Ayeni et al., 2024</xref>). Based on this reasoning, the study hypothesizes that students in the guided LLM condition will produce significantly higher-quality academic writing than those in the unguided and control groups, respectively. These differences are theoretically grounded in the TPACK framework and supported by prior research on instructional scaffolding and technology-mediated writing support.</p>
</sec>
<sec id="s4">
<title>4 Perceived mental health</title>
<p>University students&#x00027; mental health has become a central concern in global higher education, with consistently high levels of anxiety, stress, and emotional exhaustion reported across diverse national contexts (<xref ref-type="bibr" rid="B22">Granieri et al., 2021</xref>). These issues are particularly salient in competitive academic environments, where cognitive demands are high and support structures often limited. Academic writing, in particular, is a cognitively and emotionally taxing task that may exacerbate stress, especially in the absence of timely guidance or feedback.</p>
<p>The Cognitive Load Theory (<xref ref-type="bibr" rid="B63">Sweller, 1988</xref>) provides a relevant framework for understanding how instructional conditions affect students&#x00027; mental wellbeing. According to this theory, cognitive performance is shaped by the interplay of three types of load: intrinsic (task complexity), extraneous (inefficient instructional design), and germane (learning-related processing). Poorly structured tasks tend to increase extraneous load, consuming cognitive resources and contributing to frustration or emotional fatigue (<xref ref-type="bibr" rid="B32">Li et al., 2020</xref>).</p>
<p>LLMs, when properly embedded in instruction, can help reduce extraneous cognitive load by automating lower-level processes such as sentence formulation, grammar correction, or even idea generation. However, this benefit is not automatic. Students need pedagogical guidance to understand how to use the tool effectively and ethically, and how to interpret or revise its suggestions. Without such framing, students may misuse the tool, become overwhelmed by its outputs, or develop dependency without comprehension (<xref ref-type="bibr" rid="B50">Park and Ahn, 2024</xref>).</p>
<p>In the guided condition, students receive structured instructions on how to use the LLM strategically during the writing process&#x02014;e.g., to plan text sections, refine transitions, or paraphrase while maintaining academic integrity. This structure is expected to reduce cognitive overload and enhance students&#x00027; sense of control, which may, in turn, support emotional regulation and perceived well-being. In contrast, the unguided group accesses the tool without clear direction. While they may benefit from its features, they also face the burden of interpreting outputs and deciding when and how to use them. This may increase cognitive load rather than reduce it, particularly for students unfamiliar with AI tools or lacking academic writing experience. Consequently, their perceived mental health may remain unchanged or even be negatively affected.</p>
<p>Finally, students in the control group, without access to any external tool or guidance, must complete the writing task using only their own cognitive and emotional resources. While this mirrors traditional academic practice, it may result in heightened task-related anxiety or emotional exhaustion, especially under time constraints or pressure to perform.</p>
<p>Based on this framework, the present study hypothesizes that students in the guided LLM condition will report significantly better mental well-being than those in the control group, with the unguided group expected to fall somewhere in between. This hypothesis reflects the assumption that instructionally structured technology use, rather than mere access, is the key to supporting psychological outcomes in academic settings.</p>
</sec>
<sec id="s5">
<title>5 Academic engagement</title>
<p>Academic engagement is a multidimensional construct encompassing students&#x00027; behavioral, emotional, and cognitive investment in learning activities (<xref ref-type="bibr" rid="B19">Fredricks et al., 2004</xref>). High levels of engagement have been associated with greater academic achievement, persistence, and satisfaction, particularly in university settings where autonomy and self-regulation are central to success (<xref ref-type="bibr" rid="B65">Wang, 2022</xref>). However, engagement is also sensitive to fluctuations in motivation, task design, and perceived support from instructors or institutional structures (<xref ref-type="bibr" rid="B36">Lin, 2024</xref>).</p>
<p>A useful framework for understanding the mechanisms that foster engagement is the Self-Determination Theory (SDT), proposed by (<xref ref-type="bibr" rid="B14">Deci and Ryan 2000</xref>, <xref ref-type="bibr" rid="B55">Ryan and Deci, 2000</xref>). According to SDT, engagement flourishes when learners experience the fulfillment of three basic psychological needs: competence (feeling effective), autonomy (feeling self-directed), and relatedness (feeling connected and supported). Instructional strategies that enhance these dimensions are more likely to result in sustained engagement and intrinsic motivation.</p>
<p>In this context, the use of LLMs has the potential to support academic engagement&#x02014;but only if implemented thoughtfully. In the guided condition, students receive clear instructions on how to use the tool to improve their writing in ways that foster self-efficacy and control. For example, they may learn to use the model to test different formulations, organize their ideas more efficiently, or revise their text in response to feedback. This structured support not only enhances competence, but also promotes autonomy, as students gain agency in managing complex academic tasks. In contrast, students in the unguided condition are left to navigate the LLM independently. While some may explore the tool productively, others may feel uncertain about how to use it effectively or ethically. This ambiguity can hinder perceived competence and reduce the motivational benefits typically associated with technology-enhanced learning. Without explicit pedagogical framing, LLM use may become a passive or confusing experience, diminishing its capacity to support sustained engagement.</p>
<p>Finally, students in the control group engage in the task without any digital support. Although this may reflect a traditional educational scenario, it offers limited opportunities to enhance autonomy or competence through external scaffolding. For some students, especially those with lower academic confidence, this condition may result in disengagement or surface-level effort. Building on Self-Determination Theory and recent findings on digital learning environments (<xref ref-type="bibr" rid="B65">Wang, 2022</xref>), the present study hypothesizes that students in the guided LLM condition will report the highest levels of academic engagement, followed by those in the unguided condition, with the control group expected to exhibit the lowest levels. This hierarchy reflects the assumption that pedagogically structured AI use can enhance both motivation and investment in academic tasks, provided that it supports students&#x00027; psychological needs.</p>
<p>Taken together, the theoretical and empirical perspectives reviewed above suggest that the impact of AI tools in higher education depends not merely on access to the technology, but critically on how that technology is pedagogically framed and operationalized. While guided use of LLMs has the potential to support students&#x00027; writing development, reduce extraneous cognitive load, and foster meaningful engagement, unguided use may result in uneven or superficial outcomes. Meanwhile, students who receive no digital support may face greater academic pressure and cognitive effort, particularly when completing complex tasks under time constraints.</p>
<p>To examine these assumptions, the present study adopts an experimental design comparing three conditions: guided LLM use, unguided LLM use, and a control group without access to LLMs. The outcomes under investigation&#x02014;academic writing quality, perceived mental wellbeing, and academic engagement&#x02014;were selected because they represent core dimensions of student success and are theoretically linked to instructional design and technological integration. Building on the reviewed literature, it is hypothesized that students in the guided LLM condition will outperform their peers across all three variables, followed by those in the unguided condition, with the control group expected to report the lowest levels of performance and well-being. This hypothesis reflects the view that it is not the technology itself, but rather the pedagogical structuring of its use, that determines its educational value.</p>
<sec>
<title>5.1 Hypotheses</title>
<list list-type="simple">
<list-item><p>H1: Students in the guided LLM use condition will demonstrate significantly higher academic writing quality than those in the unguided LLM use and control groups.</p></list-item>
<list-item><p>H2: Students in the guided LLM use condition will report significantly higher levels of perceived mental health compared to students in the control group.</p></list-item>
<list-item><p>H3: Students in the guided LLM use condition will exhibit significantly greater academic engagement than those in the control group.</p></list-item>
<list-item><p>H4: Students in the unguided LLM use condition will demonstrate intermediate levels of academic writing quality, perceived mental health, and engagement, higher than those in the control group but lower than those in the guided use group.</p></list-item>
</list>
</sec>
</sec>
<sec id="s6">
<title>6 Method</title>
<sec>
<title>6.1 Transparency and openness</title>
<p>In this experimental study, we report how the sample size was determined and all inclusion criteria, manipulations, and outcome measures. All anonymized data are available via the Open Science Framework (<ext-link ext-link-type="uri" xlink:href="https://osf.io/htejm/?view_only=54624dbd9f11467ea26242bae037e713">https://osf.io/htejm/?view_only=54624dbd9f11467ea26242bae037e713</ext-link>). The data were analyzed using SPSS, version 29. No data were collected after the data analysis began. This study was not preregistered.</p>
</sec>
<sec>
<title>6.2 Participants</title>
<p>An a priori power analysis was conducted using G<sup>&#x0002A;</sup>Power (<xref ref-type="bibr" rid="B17">Faul et al., 2009</xref>) to determine the required sample size for a one-way ANCOVA with three groups and one covariate. Setting the alpha level at 0.05, power at 0.80, and anticipating a small to medium effect size (<italic>f</italic> = <sub>0.20</sub>), the estimated minimum sample size was <italic>N</italic> = 246. The final sample consisted of two hundred and eighty eight undergraduate students enrolled in humanities and arts programs at an international university in China. Instructors from four elective courses were initially contacted via internal mailing lists distributed by the College of Humanities and Arts and were invited to authorize data collection during one of their scheduled sessions. Once instructor consent was obtained, students were approached in person during class and invited to participate. Participants were recruited using a convenience sampling strategy, and participation was strictly voluntary. Students were informed that they could decline or withdraw at any point without penalty. No academic credit, compensation, or incentive was offered. Of the approximately three hundred and twenty five students approached across the four courses, two hundred and eighty eight undergraduate (88.6%) agreed to participate and completed all study components. The final sample included one hundred and sixty eight male students (58.3%) and one hundred and twenty female students (41.7%), ranging in age from 18 to 22 years (<italic>M</italic> = 19.88, <italic>SD</italic> = 1.50). All participants completed the writing task and self-report measures under supervised classroom conditions.</p>
</sec>
<sec>
<title>6.3 Ethical approval and informed consent</title>
<p>This study was reviewed and approved by the Institutional Review Board (IRB) of the Xi&#x00027;an International University following the Declaration of Helsinki. Ethical approval was granted on [January 15th, 2024], under the reference number [approval ID, IRB/24/072-HUMARTS]. Before participation, all students received an information sheet outlining the purpose of the study, the nature of the tasks, the voluntary nature of their participation, and their right to withdraw at any point without penalty. Written informed consent was obtained from all participants. The consent form emphasized that participation was anonymous, data would be kept confidential, and results would be used solely for academic research. Participants were also informed that using the LLM was part of an experimental educational intervention and that their course grades would not be affected by their responses or participation.</p>
</sec>
<sec>
<title>6.4 Procedure</title>
<p>Instructors from four elective undergraduate courses in the humanities and arts were first contacted via internal mailing lists distributed by the College of Humanities and Arts. After receiving their consent to conduct the study during scheduled class time, students were invited in person to participate. The study was introduced at the beginning of the session, and all students were informed that participation was entirely voluntary and that they could withdraw at any time without consequences. Students who agreed to participate provided informed consent and completed the study during a supervised class session. Participants were randomly assigned to one of three experimental conditions: guided LLM use, unguided LLM use, or control. Random assignment was conducted at the individual level within each classroom using a pre-generated randomization list.</p>
<p>All participants were asked to complete the same academic writing task: a critical essay on the topic &#x0201C;The impact of globalization on contemporary culture&#x0201D;, to be written in 45 m using a computer. Participants in the two experimental conditions used OpenAI&#x00027;s GPT-4, accessed via a monitored institutional interface. No alternative platforms or personal devices were permitted. All students interacted with the same LLM under identical interface conditions. On average, participants in the experimental conditions spent between 25 and 35 m actively interacting with the LLM during the task. In the control condition, students received the following prompt: &#x0201C;Write a critical essay on the impact of globalization, using the provided readings. Structure your argument and support it with specific examples.&#x0201D; No access to LLMs or external writing tools was provided. In the unguided LLM use condition, students were given access to GPT-4 and instructed: &#x0201C;You may use the language model (LLM) in any way you find useful to complete your essay.&#x0201D; No additional instructions, training, or support were provided.</p>
<p>In the guided LLM use condition, participants received the following prompt: &#x0201C;Write a critical essay on the impact of globalization. Use the language model (LLM) to help you generate ideas, organize your arguments, and improve clarity. You may use it to explore different perspectives, revise paragraphs, or paraphrase content. Ensure that your essay reflects critical thinking, coherence, and academic style.&#x0201D;</p>
<p>Before beginning the writing task, this group received a brief 10-m in-class orientation delivered by the course instructor, based on a script prepared by the research team. The orientation covered three key elements: how to formulate effective prompts, how to evaluate AI-generated outputs critically, and how to use the tool ethically in academic contexts. After completing the writing task, participants responded to standardized self-report measures assessing perceived mental health, academic engagement, and a short demographic questionnaire. All responses were submitted digitally and anonymized prior to analysis. No pilot study was conducted prior to the implementation of the experiment.</p>
</sec>
<sec>
<title>6.5 Instruments</title>
<p>All instructions, writing prompts, the manipulation check, and the self-report measures&#x02014;except for one&#x02014;were administered in English, in accordance with the instructional language of the international university where the study took place. The only exception was the Utrecht Work Engagement Scale &#x02013; Student version (UWES-S), which was administered in its validated Chinese version due to its demonstrated psychometric reliability in Chinese undergraduate populations.</p>
</sec>
<sec>
<title>6.6 Manipulation check</title>
<p>A manipulation check was administered immediately after the writing task to verify participants&#x00027; adherence to their assigned intervention condition. Participants responded to two closed-ended questions: (1) &#x0201C;Did you use the language model (LLM) while completing the essay?&#x0201D; (Yes/No), and (2) &#x0201C;Were you instructed on how to use the LLM?&#x0201D; (Yes/No). These items allowed the researchers to determine whether participants in the experimental groups used the LLM as intended, and whether participants in the control group refrained from doing so.</p>
<p>Only participants whose responses were fully consistent with their assigned condition were retained for the main analyses. Specifically, inclusion criteria required that participants in the guided condition reported using the LLM with instructions, those in the unguided condition reported using the LLM without instructions, and those in the control condition reported not using the LLM. Participants who did not meet these criteria were excluded from the final dataset. As a result, the final sample included two hundred and forty six participants who successfully passed the manipulation check and were eligible for analysis.</p>
</sec>
<sec>
<title>6.7 Text quality</title>
<p>The quality of participants&#x00027; academic writing was assessed using the Grammarly for Education platform (Grammarly, Inc., 2024). After completing the essay, the experimenter uploaded each text under standardized conditions. Grammarly automatically generated a Performance Score, ranging from 0 to 100, which served as the primary indicator of overall text quality. This composite score reflects the extent to which the writing adheres to grammatical norms, clarity, and effective communication, and it can be improved by addressing the platform&#x00027;s suggested revisions. In addition to the performance score, Grammarly provides detailed linguistic metrics, including word count, average word and sentence length, readability score (based on the Flesch scale) (<xref ref-type="bibr" rid="B18">Flesch, 1948</xref>), and vocabulary diversity (e.g., proportion of unique and rare words). These secondary indicators were reviewed to contextualize writing complexity and stylistic variation, though only the Performance Score was used in the statistical analyses. This approach provided a replicable, standardized, and objective method for evaluating the quality of written academic work across all participants, minimizing potential biases associated with human ratings.</p>
</sec>
<sec>
<title>6.8 Mental health</title>
<p>Perceived psychological well-being was assessed using the Ryff Psychological Well-Being Scale (PWBS) (<xref ref-type="bibr" rid="B56">Ryff, 1989</xref>). The scale consists of multiple subdimensions (e.g., autonomy, environmental mastery, personal growth, purpose in life), with responses given on a 5-point Likert scale ranging from 1 (strongly disagree) to 5 (strongly agree). Higher scores indicate greater psychological well-being. The PWBS has been widely validated and used across cross-cultural educational contexts (<xref ref-type="bibr" rid="B57">Ryff and Keyes, 1995</xref>). <xref ref-type="bibr" rid="B33">Li (2014)</xref> tested a shorter version in the Chinese language, and it was used in the present study. In the current sample, internal consistency was acceptable (&#x003B1; = 0.78).</p>
</sec>
<sec>
<title>6.9 Academic engagement</title>
<p>Academic engagement was measured using the Utrecht Work Engagement Scale&#x02014;Student Version (UWES-S) (<xref ref-type="bibr" rid="B60">Schaufeli et al., 2002</xref>). This 17-item scale captures three core dimensions of engagement&#x02014;vigor, dedication, and absorption. Participants rated each item on a 7-point scale from 0 (never) to 6 (always). Total scores were calculated by averaging across all items, with higher scores reflecting greater engagement. The UWES-S has demonstrated strong internal consistency and cross-cultural validity (<xref ref-type="bibr" rid="B59">Schaufeli and Bakker, 2004</xref>). The Chinese version developed by <xref ref-type="bibr" rid="B16">Fang et al. (2008)</xref> was used. In the present study, the scale demonstrated good internal consistency (&#x003B1; = 0.89).</p>
</sec>
<sec>
<title>6.10 Covariate: prior academic performance</title>
<p>All main analyses included participants&#x00027; prior academic performance in a literature-related subject as a covariate. Academic records provided a numerical score on a 100-point scale, reflecting performance in the most recent literature course completed before the intervention. This variable was used to control for potential baseline differences in academic ability related to writing, critical reading, and content familiarity. The scores ranged from 18 to 78, with a mean of 44.37 <italic>(SD</italic> = 13.46), indicating moderate variability across the sample. Controlling for this variable allowed for a more accurate estimation of the intervention effects on the outcome measures.</p>
</sec>
</sec>
<sec sec-type="results" id="s7">
<title>7 Results</title>
<sec>
<title>7.1 Descriptive statistics and correlations</title>
<p>Before conducting the main analyses, descriptive statistics and bivariate correlations were calculated for all continuous variables: prior academic performance, text quality, perceived mental health, and academic engagement. <xref ref-type="table" rid="T1">Table 1</xref> presents the means and standard deviations for each variable. Pearson&#x00027;s correlation coefficients are presented in <xref ref-type="table" rid="T2">Table 2</xref>. All correlations were statistically significant at the <italic>p</italic> &#x0003C; 0.01 level. Academic performance was positively correlated with text quality (<italic>r</italic> = 0.258<italic>, p</italic> &#x0003C; 0.001), mental health <italic>(r</italic> = 0.329, <italic>p</italic> &#x0003C; 0.001), and engagement (r = 0.272, <italic>p</italic> &#x0003C; 0.001). Text quality also showed moderate positive correlations with mental health (<italic>r</italic> =0.406, <italic>p</italic> &#x0003C; 0.001) and engagement (<italic>r</italic> = 0.280, <italic>p</italic> &#x0003C; 0.001). The strongest association was observed between mental health and academic engagement (<italic>r</italic> = 0.568, <italic>p</italic> &#x0003C; 0.001), suggesting a meaningful link between students&#x00027; psychological wellbeing and their engagement with academic tasks.</p>
<table-wrap position="float" id="T1">
<label>Table 1</label>
<caption><p>Descriptive statistics and pearson correlations between study variables <italic>(N</italic> = 288).</p></caption>
<table frame="box" rules="all">
<thead>
<tr>
<th valign="top" align="left"><bold>Variable</bold></th>
<th valign="top" align="center"><bold><italic>M</italic></bold></th>
<th valign="top" align="center"><bold><italic>SD</italic></bold></th>
<th valign="top" align="center"><bold>1</bold></th>
<th valign="top" align="center"><bold>2</bold></th>
<th valign="top" align="center"><bold>3</bold></th>
<th valign="top" align="center"><bold>4</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">1. Academic performance</td>
<td valign="top" align="center">44.37</td>
<td valign="top" align="center">13.46</td>
<td valign="top" align="center">&#x02014;</td>
<td/>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left">2. Text quality</td>
<td valign="top" align="center">123.74</td>
<td valign="top" align="center">22.86</td>
<td valign="top" align="center">.258<sup>&#x0002A;&#x0002A;</sup></td>
<td valign="top" align="center">&#x02014;</td>
<td/>
<td/>
</tr>
<tr>
<td valign="top" align="left">3. Mental health</td>
<td valign="top" align="center">3.82</td>
<td valign="top" align="center">0.82</td>
<td valign="top" align="center">.329<sup>&#x0002A;&#x0002A;</sup></td>
<td valign="top" align="center">.406<sup>&#x0002A;&#x0002A;</sup></td>
<td valign="top" align="center">&#x02014;</td>
<td/>
</tr>
<tr>
<td valign="top" align="left">4. Academic engagement</td>
<td valign="top" align="center">13.64</td>
<td valign="top" align="center">2.86</td>
<td valign="top" align="center">.272<sup>&#x0002A;&#x0002A;</sup></td>
<td valign="top" align="center">.280<sup>&#x0002A;&#x0002A;</sup></td>
<td valign="top" align="center">.568<sup>&#x0002A;&#x0002A;</sup></td>
<td valign="top" align="center">&#x02014;</td>
</tr></tbody>
</table>
<table-wrap-foot>
<p><italic>N</italic> = <italic>288 refers to the total number of participants who completed all measures and were included in the descriptive statistics and Pearson correlations. However, only N</italic> = <italic>246 participants who passed the manipulation check were retained for the main ANCOVA analyses. M</italic> = <italic>mean; SD</italic> = <italic>standard deviation</italic>. <sup>&#x0002A;&#x0002A;</sup> <italic>p</italic>&#x0003C;<italic>0.01</italic>.</p>
</table-wrap-foot>
</table-wrap>
<table-wrap position="float" id="T2">
<label>Table 2</label>
<caption><p>Estimated marginal means for academic writing quality (controlling for academic performance).</p></caption>
<table frame="box" rules="all">
<thead>
<tr>
<th valign="top" align="left"><bold>Intervention</bold></th>
<th valign="top" align="center"><bold>Mean</bold></th>
<th valign="top" align="center"><bold><italic>SE</italic></bold></th>
<th valign="top" align="center"><bold>95% CI lower</bold></th>
<th valign="top" align="center"><bold>95% CI upper</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Control</td>
<td valign="top" align="center">109.61</td>
<td valign="top" align="center">0.70</td>
<td valign="top" align="center">108.23</td>
<td valign="top" align="center">110.99</td>
</tr>
<tr>
<td valign="top" align="left">Unguided LLM use</td>
<td valign="top" align="center">129.61</td>
<td valign="top" align="center">0.77</td>
<td valign="top" align="center">128.10</td>
<td valign="top" align="center">131.13</td>
</tr>
<tr>
<td valign="top" align="left">Guided LLM use</td>
<td valign="top" align="center">151.35</td>
<td valign="top" align="center">0.78</td>
<td valign="top" align="center">149.82</td>
<td valign="top" align="center">152.87</td>
</tr></tbody>
</table>
<table-wrap-foot>
<p><italic>Note</italic>. Means are estimated marginal means adjusted for the covariate (academic performance).</p>
</table-wrap-foot>
</table-wrap>
<p>A series of Univariate Analyses of Covariance (ANCOVA) were conducted to examine the effects of intervention condition on three outcome variables: academic writing quality, perceived mental health, and academic engagement. The independent variable was the type of LLM integration (guided use, unguided use, and control), and prior academic performance was included as a covariate in all models.</p>
</sec>
<sec>
<title>7.2 Academic writing quality</title>
<p>The ANCOVA revealed a significant effect of intervention condition on writing quality, <italic>F</italic> <sub>(2, 253)</sub> = 789.53, <italic>p</italic> &#x0003C; 0.001, with a very large effect size (<italic>R</italic><sup>2</sup> adj = 0.863). This indicates that the intervention condition explained approximately 86% of the variance in writing performance, reflecting a strong and practically meaningful impact of structured LLM integration.</p>
<p>Estimated marginal means showed that students in the guided LLM use condition produced significantly higher quality texts (<italic>M</italic> = 151.35, <italic>SE</italic> = 0.775) than those in the unguided use (<italic>M</italic> = 129.61, <italic>SE</italic> = 0.770) and control groups (<italic>M</italic> = 109.61, <italic>SE</italic> = 0.700), as <xref ref-type="table" rid="T2">Table 2</xref> shows.</p>
<p>Effect sizes computed with pooled standard deviations showed a very large difference between the guided LLM use and control groups (Cohen&#x00027;s d = 5.16), a large difference between the guided and unguided groups <italic>(d</italic> = 2.53), and a large difference between the unguided and control groups (<italic>d</italic> = 3.83). These values highlight the strong impact of guided LLM use on writing performance, and confirm that even unguided use resulted in substantially better outcomes compared to no use. All pairwise comparisons were statistically significant after Bonferroni correction (<italic>p</italic> &#x0003C; 0.001), as <xref ref-type="table" rid="T3">Table 3</xref> shows.</p>
<table-wrap position="float" id="T3">
<label>Table 3</label>
<caption><p>Pairwise comparisons&#x02014;academic writing quality.</p></caption>
<table frame="box" rules="all">
<thead>
<tr>
<th valign="top" align="left"><bold>Comparison</bold></th>
<th valign="top" align="center"><bold>Mean difference (I&#x02013;J)</bold></th>
<th valign="top" align="center"><bold><italic>SE</italic></bold></th>
<th valign="top" align="center"><bold><italic>p</italic></bold></th>
<th valign="top" align="center"><bold>95% <italic>CI</italic> Lower</bold></th>
<th valign="top" align="center"><bold>95% <italic>CI</italic> upper</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Guided LLM use&#x02014;control</td>
<td valign="top" align="center">41.73</td>
<td valign="top" align="center">1.05</td>
<td valign="top" align="center">&#x0003C; 0.001</td>
<td valign="top" align="center">39.20</td>
<td valign="top" align="center">44.27</td>
</tr>
<tr>
<td valign="top" align="left">Guided LLM use&#x02014;unguided use</td>
<td valign="top" align="center">21.73</td>
<td valign="top" align="center">1.09</td>
<td valign="top" align="center">&#x0003C; 0.001</td>
<td valign="top" align="center">19.11</td>
<td valign="top" align="center">24.36</td>
</tr>
<tr>
<td valign="top" align="left">Unguided LLM use&#x02014;control</td>
<td valign="top" align="center">20.00</td>
<td valign="top" align="center">1.05</td>
<td valign="top" align="center">&#x0003C; 0.001</td>
<td valign="top" align="center">17.48</td>
<td valign="top" align="center">22.52</td>
</tr></tbody>
</table>
<table-wrap-foot>
<p><italic>SE</italic> = <italic>standard error; CI</italic> = <italic>confidence interval. Pairwise comparisons were adjusted using the</italic> Bonferroni correction.</p>
</table-wrap-foot>
</table-wrap>
</sec>
<sec>
<title>7.3 Perceived mental health</title>
<p>The analysis also showed a significant effect of intervention condition on perceived mental health, <italic>F</italic> <sub>(2, 253)</sub> = 5.78, <italic>p</italic> = 0.004, with a small to moderate effect size (<italic>R</italic><sup>2</sup> adj = 0.097). This suggests that nearly 10% of the variability in self-reported mental wellbeing was attributable to the different LLM conditions, indicating a modest yet meaningful contribution of guided use to students&#x00027; perceived psychological health. Students in the guided use condition reported significantly higher mental health scores (<italic>M</italic> = 4.15, <italic>SE</italic> = 0.076) than the control group (<italic>M</italic> = 3.81, <italic>SE</italic> = 0.069, <italic>p</italic> = 0.003), as <xref ref-type="table" rid="T4">Table 4</xref> shows.</p>
<table-wrap position="float" id="T4">
<label>Table 4</label>
<caption><p>Estimated marginal means for perceived mental health (controlling for academic performance).</p></caption>
<table frame="box" rules="all">
<thead>
<tr>
<th valign="top" align="left"><bold>Intervention</bold></th>
<th valign="top" align="center"><bold>Mean</bold></th>
<th valign="top" align="center"><bold><italic>SE</italic></bold></th>
<th valign="top" align="center"><bold>95% <italic>CI</italic> lower</bold></th>
<th valign="top" align="center"><bold>95% <italic>CI</italic> upper</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Control</td>
<td valign="top" align="center">3.81</td>
<td valign="top" align="center">0.069</td>
<td valign="top" align="center">3.67</td>
<td valign="top" align="center">3.94</td>
</tr>
<tr>
<td valign="top" align="left">Unguided LLM use</td>
<td valign="top" align="center">3.91</td>
<td valign="top" align="center">0.076</td>
<td valign="top" align="center">3.76</td>
<td valign="top" align="center">4.06</td>
</tr>
<tr>
<td valign="top" align="left">Guided LLM use</td>
<td valign="top" align="center">4.15</td>
<td valign="top" align="center">0.076</td>
<td valign="top" align="center">4.00</td>
<td valign="top" align="center">4.30</td>
</tr></tbody>
</table>
<table-wrap-foot>
<p>Means are estimated marginal means adjusted for the covariate (academic performance).</p>
</table-wrap-foot>
</table-wrap>
<p>As <xref ref-type="table" rid="T5">Table 5</xref> shows, the difference between the guided and unguided groups (<italic>M</italic> = 3.91, <italic>SE</italic> = 0.076) approached statistical significance <italic>(p</italic> = 0.070), whereas no significant difference was observed between the unguided and control conditions (<italic>p</italic> = 0.984). Effect size estimates indicated a moderate difference between the guided LLM use and control conditions (Cohen&#x00027;s <italic>d</italic> = 0.53), a small to moderate effect between guided and unguided use (<italic>d</italic> = 0.31), and a negligible effect between unguided use and control (<italic>d</italic> = 0.14). These findings suggest that only structured use of the LLM led to meaningful psychological benefits.</p>
<table-wrap position="float" id="T5">
<label>Table 5</label>
<caption><p>Pairwise Comparisons &#x02013; Mental Health.</p></caption>
<table frame="box" rules="all">
<thead>
<tr>
<th valign="top" align="left"><bold>Comparison</bold></th>
<th valign="top" align="center"><bold>Mean difference (I&#x02013;J)</bold></th>
<th valign="top" align="center"><bold>SE</bold></th>
<th valign="top" align="center"><bold><italic>p</italic></bold></th>
<th valign="top" align="center"><bold>95% CI Lower</bold></th>
<th valign="top" align="center"><bold>95% CI Upper</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Guided LLM use&#x02014;control</td>
<td valign="top" align="center">0.35</td>
<td valign="top" align="center">0.10</td>
<td valign="top" align="center">.003</td>
<td valign="top" align="center">0.10</td>
<td valign="top" align="center">0.59</td>
</tr>
<tr>
<td valign="top" align="left">Guided LLM use&#x02014;unguided Use</td>
<td valign="top" align="center">0.24</td>
<td valign="top" align="center">0.11</td>
<td valign="top" align="center">0.070</td>
<td valign="top" align="center">&#x02212;0.01</td>
<td valign="top" align="center">0.50</td>
</tr>
<tr>
<td valign="top" align="left">Unguided LLM use&#x02014;control</td>
<td valign="top" align="center">0.10</td>
<td valign="top" align="center">0.10</td>
<td valign="top" align="center">0.984</td>
<td valign="top" align="center">&#x02212;0.35</td>
<td valign="top" align="center">0.15</td>
</tr></tbody>
</table>
<table-wrap-foot>
<p>SE, standard error; CI, confidence interval. Pairwise comparisons were adjusted using the Bonferroni correction.</p>
</table-wrap-foot>
</table-wrap>
</sec>
<sec>
<title>7.4 Academic engagement</title>
<p>Lastly, the ANCOVA for academic engagement indicated a significant effect of intervention condition, <italic>F</italic> <sub>(2, 253)</sub> = 6.70, <italic>p</italic> = 0.001, with a modest effect size (<italic>R</italic><sup>2</sup> adj = 0.101). This means that around 10% of the variance in engagement was explained by the intervention condition, pointing to a small but practically relevant effect of structured LLM integration on students&#x00027; involvement in academic activities. Students in the guided LLM use group reported the highest engagement scores (<italic>M</italic> = 14.67, <italic>SE</italic> = 0.304), significantly higher than the control group (<italic>M</italic> = 13.17, <italic>SE</italic> = 0.275, <italic>p</italic> &#x0003C; 0.001), as <xref ref-type="table" rid="T6">Table 6</xref> shows.</p>
<table-wrap position="float" id="T6">
<label>Table 6</label>
<caption><p>Estimated marginal means for academic engagement (controlling for academic performance).</p></caption>
<table frame="box" rules="all">
<thead>
<tr>
<th valign="top" align="left"><bold>Intervention</bold></th>
<th valign="top" align="center"><bold>Mean</bold></th>
<th valign="top" align="center"><bold><italic>SE</italic></bold></th>
<th valign="top" align="center"><bold>95% <italic>CI</italic> lower</bold></th>
<th valign="top" align="center"><bold>95% CI upper</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Control</td>
<td valign="top" align="center">13.17</td>
<td valign="top" align="center">0.28</td>
<td valign="top" align="center">12.62</td>
<td valign="top" align="center">13.71</td>
</tr>
<tr>
<td valign="top" align="left">Unguided LLM use</td>
<td valign="top" align="center">13.74</td>
<td valign="top" align="center">0.30</td>
<td valign="top" align="center">13.15</td>
<td valign="top" align="center">14.34</td>
</tr>
<tr>
<td valign="top" align="left">Guided LLM use</td>
<td valign="top" align="center">14.67</td>
<td valign="top" align="center">0.30</td>
<td valign="top" align="center">14.07</td>
<td valign="top" align="center">15.27</td>
</tr></tbody>
</table>
<table-wrap-foot>
<p>Means are estimated marginal means adjusted for the covariate (academic performance).</p>
</table-wrap-foot>
</table-wrap>
<p>Although the unguided use group (<italic>M</italic> = 13.74, <italic>SE</italic> = 0.302) scored higher than the control group, this difference was not statistically significant <italic>(p</italic> = 0.485), and the difference between the guided and unguided groups was marginal (<italic>p</italic> = 0.093), as <xref ref-type="table" rid="T7">Table 7</xref> shows. For engagement, the contrast between guided use and control yielded a moderate effect size (Cohen&#x00027;s d = 0.55), while the effect between guided and unguided use was small to moderate (<italic>d</italic> = 0.32), and the difference between unguided use and control was small (<italic>d</italic> = 0.21). These results indicate that guided integration produced a noticeable improvement in students&#x00027; involvement, whereas unguided use led to minimal gains.</p>
<table-wrap position="float" id="T7">
<label>Table 7</label>
<caption><p>Pairwise comparisons&#x02014;academic engagement.</p></caption>
<table frame="box" rules="all">
<thead>
<tr>
<th valign="top" align="left"><bold>Comparison</bold></th>
<th valign="top" align="center"><bold>Mean difference (I&#x02013;J)</bold></th>
<th valign="top" align="center"><bold><italic>SE</italic></bold></th>
<th valign="top" align="center"><bold><italic>p</italic></bold></th>
<th valign="top" align="center"><bold>95% <italic>CI</italic> lower</bold></th>
<th valign="top" align="center"><bold>95% <italic>CI</italic> upper</bold></th>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Guided LLM use&#x02014;control</td>
<td valign="top" align="center">1.51</td>
<td valign="top" align="center">0.41</td>
<td valign="top" align="center">&#x0003C; 0.001</td>
<td valign="top" align="center">0.51</td>
<td valign="top" align="center">2.50</td>
</tr>
<tr>
<td valign="top" align="left">Guided LLM use&#x02014;unguided use</td>
<td valign="top" align="center">0.93</td>
<td valign="top" align="center">0.43</td>
<td valign="top" align="center">0.093</td>
<td valign="top" align="center">&#x02212;0.10</td>
<td valign="top" align="center">1.96</td>
</tr>
<tr>
<td valign="top" align="left">Unguided LLM use&#x02014;control</td>
<td valign="top" align="center">0.58</td>
<td valign="top" align="center">0.41</td>
<td valign="top" align="center">0.485</td>
<td valign="top" align="center">&#x02212;0.41</td>
<td valign="top" align="center">1.57</td>
</tr></tbody>
</table>
<table-wrap-foot>
<p>SE, standard error; CI, confidence interval. Pairwise comparisons were adjusted using the Bonferroni correction.</p>
</table-wrap-foot>
</table-wrap>
<p>These results suggest that the guided integration of LLMs can significantly enhance students&#x00027; academic writing and engagement, and may also support improvements in perceived mental health, compared to both unguided use and no use of LLMs.</p>
<p><italic>Hypothesis 4: Intermediate Outcomes in the Unguided LLM Use Condition</italic></p>
<p>Hypothesis 4 proposed that students in the unguided LLM use condition would demonstrate intermediate levels of academic writing quality, perceived mental health, and academic engagement, higher than those in the control group but lower than those in the guided use group.</p>
<p>The results provided partial support for this hypothesis. In terms of academic writing quality, the unguided group (<italic>M</italic> = 129.61, <italic>SE</italic> = 0.77) scored significantly higher than the control group (<italic>M</italic> = 109.61, <italic>SE</italic> = 0.70, <italic>p</italic> &#x0003C; 0.001), but significantly lower than the guided group (<italic>M</italic> = 151.35, <italic>SE</italic> = 0.78, <italic>p</italic> &#x0003C; 0.001). These findings confirm the predicted ordinal pattern in this domain.</p>
<p>However, for perceived mental health, the unguided group (<italic>M</italic> = 3.91, <italic>SE</italic> = 0.076) did not differ significantly from the control group (<italic>M</italic> = 3.81, <italic>SE</italic> = 0.069, <italic>p</italic> = 0.984), although it trended lower than the guided group (<italic>M</italic> = 4.15, <italic>SE</italic> = 0.076), with this comparison approaching statistical significance (<italic>p</italic> = 0.070).</p>
<p>Similarly, regarding academic engagement, the unguided group (<italic>M</italic> = 13.74, <italic>SE</italic> = 0.30) did not significantly differ from the control group (<italic>M</italic> = 13.17, <italic>SE</italic> = 0.28, <italic>p</italic> = 0.485). The difference between the unguided and guided conditions (<italic>M</italic> = 14.67, <italic>SE</italic> = 0.30) was marginal <italic>(p</italic> = 0.093).</p>
<p>These results indicate that while the unguided LLM condition yielded intermediate outcomes for academic writing quality consistent with Hypothesis 4, the same pattern was not statistically supported in perceived mental health and academic engagement.</p>
</sec></sec>
<sec sec-type="discussion" id="s8">
<title>8 Discussion</title>
<sec>
<title>8.1 H1: Writing quality enhancement through guided LLM use</title>
<p>The results strongly support Hypothesis 1, demonstrating that students who received structured guidance using LLMs achieved significantly higher academic writing quality than those in both the unguided and control groups. The magnitude of the effect was exceptionally large, underscoring the substantial educational potential of guided LLM integration. This finding aligns with the growing body of evidence suggesting that effective integration of LLMs in academic contexts improves the quality of student output and encourages critical engagement with both the writing process and the technology itself (<xref ref-type="bibr" rid="B6">Cash et al., 2025</xref>). The superiority of the guided condition can be interpreted through several converging mechanisms identified in recent research. First, structured frameworks like the Writing Path, which utilize explicit outlines, have been shown to significantly improve text generation quality by aligning outputs with the user&#x00027;s intentions and task-specific goals (<xref ref-type="bibr" rid="B30">Lee et al., 2024</xref>). This alignment is particularly important in academic settings, where coherence, argument structure, and adherence to conventions are critical.</p>
<p>Moreover, the results reflect broader findings in human-AI collaborative writing research. Studies on tasks such as headline generation show that users achieve better outcomes when they can guide and selectively refine LLM outputs. This process enhances quality without compromising user agency or perceived authorship (<xref ref-type="bibr" rid="B15">Ding et al., 2023</xref>). This suggests that guided LLM use in educational settings may strike a productive balance between automation and student ownership. At a cognitive level, guided use of LLMs appears to support key phases in the writing process, particularly translation and revision. <xref ref-type="bibr" rid="B7">Chakrabarty et al. (2024)</xref> found that professional writers benefited most from LLM support during these stages. This insight resonates with our results and further substantiates the utility of guided approaches in educational contexts.</p>
<p>Finally, it is worth noting that the enhanced writing performance observed may not stem solely from the tool&#x00027;s linguistic capabilities, but also from reduced uncertainty and cognitive load due to structured task framing. When students know exactly how to proceed and what is expected of them in using a complex tool like an LLM, their cognitive resources may be more efficiently allocated to higher-order writing concerns, thus improving final output quality.</p>
</sec>
<sec>
<title>8.2 H2: Guided LLM use and perceived mental health</title>
<p>The findings provide empirical support for Hypothesis 2, indicating that students in the guided LLM use condition reported significantly higher levels of perceived mental health compared to those in the control group. Although the effect size was modest, the statistical significance of the difference underscores the potential of guided LLM integration as a psychologically beneficial educational tool. Notably, the comparison between the guided and unguided groups approached significance, suggesting that guidance in LLM use may play a decisive role in how such tools influence users&#x00027; well-being. These results are consistent with growing evidence that conversational AI systems can enhance users&#x00027; subjective mental health experiences when implemented with structured guidance. One contributing factor may be the enhanced user experience associated with anthropomorphically designed systems. For instance, <xref ref-type="bibr" rid="B66">Wu et al. (2024)</xref> showed that agents like Sunnie increased users&#x00027; perceptions of usability and engagement. Such design strategies may foster a more human-like, empathetic interaction, which resonates with students in high-stress academic contexts.</p>
<p>Beyond surface-level interaction quality, systems like VITA have demonstrated that adaptive, behavior-sensitive guidance improves not just perception but also outcomes in mental well-being (<xref ref-type="bibr" rid="B62">Spitale et al., 2025</xref>). These systems personalize responses to individual user profiles and evolving needs, features that align well with the nature of guided LLM use in this study. When students receive structured prompts, reflective exercises, or scaffolded interactions from LLMs, the result is improved engagement and potentially heightened psychological support.</p>
<p>The results also echo the broader literature on LLM-based agents such as Replika, which offer on-demand, judgment-free interactions. <xref ref-type="bibr" rid="B40">Ma et al. (2023)</xref> highlighted how such interactions help individuals engage in self-reflection and develop confidence. In the present context, the structured engagement with LLMs may serve a similar purpose, providing students with an emotionally neutral space to articulate their thoughts and manage academic stress more effectively. Furthermore, <xref ref-type="bibr" rid="B27">Kumar et al. (2024)</xref> noted that AI-guided systems can integrate multiple data sources to detect subtle shifts in mental states and deliver personalized micro-interventions. Although this study did not leverage multimodal inputs, the positive outcome observed in the guided condition suggests that even text-based interventions, when strategically framed, can produce a meaningful uplift in well-being.</p>
<p>At the same time, the non-significant difference between the unguided and control groups raises important questions about the boundary conditions under which LLMs can support mental health. One plausible explanation is that unguided access may generate uncertainty, confusion, or even decision fatigue when students are left to navigate the system without structure. Prior research suggests that the absence of guidance can lead to overwhelming interactions or passive use of the tool, which may fail to produce affective benefits (<xref ref-type="bibr" rid="B70">Zhang et al., 2024</xref>; <xref ref-type="bibr" rid="B71">Zhu, 2024</xref>). It is possible that psychological support through AI requires not only access but also a sense of clarity, safety, and intentionality&#x02014;conditions more likely to be fostered in structured interventions. In sum, the significant improvement in perceived mental health among students in the guided condition supports the hypothesis that structured, intentional interaction with LLMs can enhance psychological experiences in academic settings. These findings reinforce the view that LLMs&#x02014;when deployed thoughtfully&#x02014;can act as supportive companions in educational environments (<xref ref-type="bibr" rid="B69">Youn and Jin, 2021</xref>), particularly when combined with design principles and adaptive features that foster trust, personalization, and emotional safety (<xref ref-type="bibr" rid="B37">Liu et al., 2023</xref>). However, the lack of improvement in the unguided condition highlights the importance of pedagogical framing as a necessary condition for translating technological affordances into emotional gains.</p>
</sec>
<sec>
<title>8.3 H3: Guided LLM use and academic engagement</title>
<p>The results support Hypothesis 3, indicating that students in the guided LLM use condition exhibited significantly greater academic engagement than those in the control group. The ANCOVA revealed a statistically significant effect of condition on engagement scores, with a modest effect size. Students in the guided condition reported the highest levels of engagement, reinforcing the view that structured interaction with LLMs can foster a more involved and focused learning experience. These findings align with a growing body of research highlighting the importance of guidance in shaping students&#x00027; cognitive and emotional engagement in AI-supported learning environments. In particular, structured guidance during LLM use has been shown to reduce off-task behavior, such as random or superficial queries and the indiscriminate use of AI for answer retrieval (<xref ref-type="bibr" rid="B27">Kumar et al., 2024</xref>, <xref ref-type="bibr" rid="B26">2023</xref>). By promoting intentional and reflective engagement, guided LLM interventions encourage students to assume more active roles in their learning processes.</p>
<p>Moreover, research on immersive and AI-integrated educational formats&#x02014;such as Alternate Reality Games (ARGs) augmented with LLM guidance&#x02014;has demonstrated that such designs can enhance behavioral engagement, emotional connection, and control beliefs (<xref ref-type="bibr" rid="B10">Cheng et al., 2022</xref>; <xref ref-type="bibr" rid="B47">Neary and Schueller, 2018</xref>). These findings suggest that when LLMs are embedded in pedagogically grounded frameworks, they can catalyze sustained academic motivation and participation. At the cognitive level, integrating LLMs into virtual teaching assistant roles, such as the Jill Watson system, further supports the idea that guided AI interactions can promote higher-order thinking and intellectual curiosity (<xref ref-type="bibr" rid="B42">Maiti and Goel, 2024</xref>). Students in such systems are not merely passive recipients of information but are actively encouraged to formulate and refine complex inquiries, fostering deeper engagement with content.</p>
<p>In contrast, using unguided LLMs favors quick information retrieval over sustained learning. Although such interactions may yield short-term performance gains, they do not appear to generate the same level of student investment or trust in the learning process (<xref ref-type="bibr" rid="B25">Kumar et al., 2025</xref>). This may help explain why the guided condition in the present study outperformed both the unguided and control groups regarding engagement. Indeed, the absence of clear instructional framing in the unguided condition may have led to uncertainty about how to use the tool productively, diluting its potential benefits for emotional or behavioral engagement. When students are unsure whether they are using a tool &#x0201C;correctly,&#x0201D; this ambiguity can undermine their sense of efficacy and reduce motivation to persist. Thus, although the results confirmed that guided LLM use fosters greater academic engagement, they also suggest that without supportive structure, LLMs may not reliably elicit active academic involvement. By combining technological capabilities with pedagogical intentionality, these systems offer an interactive and supportive environment that encourages students to actively participate, reflect, and persist in their academic work.</p>
</sec>
<sec>
<title>8.4 H4: Intermediate outcomes of unguided LLM use</title>
<p>Hypothesis 4 posited that students in the unguided LLM use condition would exhibit intermediate levels of academic writing quality, perceived mental health, and engagement, greater than those in the control group but lower than those in the guided use condition. The results partially supported this hypothesis: while this expected pattern was observed and statistically confirmed in academic writing quality, it was not replicated in perceived mental health or academic engagement.</p>
<p>The writing results suggest that access to LLMs can meaningfully enhance students&#x00027; output even without structured guidance. The unguided group significantly outperformed the control group, indicating that basic interaction with the tool&#x02014;through prompts, content generation, or surface-level feedback&#x02014;was sufficient to raise writing quality. This aligns with prior findings that when used independently, LLMs can offer valuable assistance in planning, drafting, and refining text (<xref ref-type="bibr" rid="B23">Jungherr, 2023</xref>; <xref ref-type="bibr" rid="B43">Meyer et al., 2024</xref>). Nevertheless, the superior performance of the guided group supports the idea that structured scaffolding, metacognitive prompts, and explicit instruction in tool usage amplify these benefits (<xref ref-type="bibr" rid="B6">Cash et al., 2025</xref>; <xref ref-type="bibr" rid="B58">Salimi and Hajinia, 2025</xref>).</p>
<p>In contrast, the data did not support the predicted intermediate pattern for perceived mental health. Although the unguided group reported slightly higher scores than the control group, this difference was not statistically significant. One possible explanation is that unguided use, while offering on-demand support, lacks the emotional structure and stress regulation strategies typically embedded in guided implementations. Without clear boundaries or reassurance about responsible usage, students may experience friction, uncertainty, or even anxiety about whether they are &#x0201C;using the tool correctly,&#x0201D; which may counteract any potential gains in psychological well-being (<xref ref-type="bibr" rid="B70">Zhang et al., 2024</xref>). In this context, guidance may not only clarify functionality but also serve a regulatory role&#x02014;normalizing AI integration, reducing confusion, and promoting a sense of support and competence (<xref ref-type="bibr" rid="B71">Zhu, 2024</xref>).</p>
<p>A similar pattern emerged with academic engagement. Although the unguided group showed numerically higher engagement than the control group, this difference was not statistically significant. This suggests that, while unguided LLM access may spark curiosity and enable autonomous exploration, it does not consistently produce sustained or deep engagement. One likely reason is that students without instructional scaffolding may remain uncertain about how to engage productively with the tool, leading to hesitant or fragmented interaction. Prior research indicates that without pedagogical framing, students may use LLMs for surface-level information retrieval or task avoidance, limiting the depth of their involvement (<xref ref-type="bibr" rid="B9">Chen and Leitch, 2024</xref>). By contrast, guided use has been associated with stronger emotional and cognitive engagement, as students are trained to leverage LLMs in a reflective and goal-oriented manner (<xref ref-type="bibr" rid="B5">Beurer-Kellner et al., 2024</xref>; <xref ref-type="bibr" rid="B64">Uchendu et al., 2023</xref>).</p>
<p>Altogether, the findings highlight the nuanced role of guidance in realizing the potential of LLMs. While unguided use may yield modest benefits in writing quality, its impact on mental health and engagement appears more limited. The lack of significant differences between the unguided and control groups in two of the three outcome domains suggests that students may underutilize these tools or even encounter friction in their use without structured scaffolding. These results point to the importance of not only providing access to AI tools but also offering appropriate pedagogical frameworks to ensure their effective and psychologically supportive implementation.</p>
</sec>
<sec>
<title>8.5 Limitations</title>
<p>Despite this experimental design&#x00027;s strengths, several limitations should be acknowledged. First, although random assignment was used to allocate participants across conditions, the study relied on self-report manipulation checks to confirm adherence to the assigned use of LLMs. While this approach ensured consistency between reported and intended use, it may introduce response bias or fail to capture subtle variations in how students interpreted and used the tool within each condition (<xref ref-type="bibr" rid="B54">Roshanaei, 2024</xref>). Although technically feasible alternatives such as behavioral logging (e.g., tracking LLM interactions) might offer more objective verification, such data were not collected due to ethical constraints and institutional limitations in access control. Future research could explore ways to integrate such measures in a transparent and privacy-respecting manner.</p>
<p>Second, the study was conducted within a single institutional context and with a relatively homogeneous sample&#x02014;undergraduate students enrolled in humanities and arts programs at an international university in China. This limits the generalizability of the findings to other educational settings, disciplines, and cultural contexts (<xref ref-type="bibr" rid="B58">Salimi and Hajinia, 2025</xref>). In particular, students in STEM fields might interact with LLMs differently not only due to varying levels of digital literacy, but also because of the nature of the writing tasks they face, the disciplinary conventions they follow, and the specific modes of information retrieval their fields require. These differences may influence how beneficial, usable, or trustworthy LLM tools appear in practice.</p>
<p>Third, the primary measure of writing quality relied on the automated Performance Score generated by the Grammarly for Education platform. While this tool offers objectivity and replicability, it prioritizes surface-level features such as grammar, clarity, and lexical variety. As a result, it may not fully capture deeper dimensions of academic writing&#x02014;such as argumentation structure, critical analysis, originality, synthesis of sources, or adherence to disciplinary conventions&#x02014;which are essential in evaluating high-level academic work. Human-rated assessments or rubric-based evaluations could complement automated scoring in future studies to provide a more holistic picture of writing quality (<xref ref-type="bibr" rid="B58">Salimi and Hajinia, 2025</xref>).</p>
<p>Fourth, although the study included a validated measure of prior academic performance as a covariate, this measure was limited to students&#x00027; most recent literature course. While this represents a meaningful control, the category of &#x0201C;literature-related subject&#x0201D; remains broad and may encompass varying levels of complexity and assessment standards. Moreover, other potentially relevant factors&#x02014;such as writing experience in other languages, previous exposure to AI tools, or individual motivation&#x02014;were not controlled and could have influenced the outcomes (<xref ref-type="bibr" rid="B67">Xu et al., 2025</xref>). Fifth, the study used perceived mental health and academic engagement as outcome variables measured through self-report scales. While these instruments are widely validated, self-reported data are subject to social desirability and may not accurately reflect behavioral engagement or psychological functioning (<xref ref-type="bibr" rid="B44">Meyer and Elsweiler, 2025</xref>). Future research could enhance robustness by incorporating behavioral (e.g., time-on-task) or physiological (e.g., stress monitoring) indicators to triangulate self-perceptions with observable evidence (<xref ref-type="bibr" rid="B69">Youn and Jin, 2021</xref>).</p>
<p>To sum up, the findings should be interpreted with an awareness of existing limitations. For example, while guided LLM use markedly improves performance on structured academic tasks, its efficacy may vary across genres or in more creative domains. <xref ref-type="bibr" rid="B21">G&#x000F3;mez-Rodr&#x000ED;guez and Williams (2023)</xref> noted that human writers still maintain an advantage in areas such as humor and originality, which are difficult for LLMs to replicate reliably. Therefore, while our results highlight the transformative potential of guided LLM use, they also reinforce the need for human oversight and creative judgment in academic writing.</p>
</sec>
<sec>
<title>8.6 Practical implications for teaching: structured integration of LLMs in academic instruction</title>
<p>Integrating LLMs into academic writing instruction presents promising opportunities and pressing challenges. The present findings reinforce the importance of structured, guided use of LLMs, particularly in enhancing students&#x00027; writing quality, academic engagement, and, to a certain extent, their perceived mental well-being. These results carry several practical implications for educators, instructional designers, and policymakers in higher education. Importantly, the implications presented here are directly informed by the limitations discussed above, and their placement after the limitations section reflects a deliberate decision to ensure that recommendations are realistic, context-aware, and attuned to the boundaries of the current design.</p>
</sec>
<sec>
<title>8.7 Designing guided LLM integration</title>
<p>The study underscores the pedagogical value of structured engagement with LLMs. Educators should prioritize guided frameworks when introducing LLMs into learning environments (<xref ref-type="bibr" rid="B2">Alsobeh and Woodward, 2023</xref>). This includes providing students with clear instructions on using these tools effectively, offering structured prompts, and integrating LLM interactions into existing learning goals (<xref ref-type="bibr" rid="B11">Chiang and Lee, 2023</xref>). As demonstrated by approaches such as the Writing Path framework (<xref ref-type="bibr" rid="B30">Lee et al., 2024</xref>), guided strategies help align LLM-generated content with academic standards and user intentions, improving writing quality (<xref ref-type="bibr" rid="B21">G&#x000F3;mez-Rodr&#x000ED;guez and Williams, 2023</xref>). Guidance also plays a crucial role in shaping student behavior during AI interaction. Research shows that structured guidance can reduce off-task or random queries and foster a more focused, problem-solving approach to writing (<xref ref-type="bibr" rid="B24">Kulaksiz, 2024</xref>). When embedded in pedagogy, these strategies enhance performance and increase students&#x00027; trust and sense of ownership in the learning process (<xref ref-type="bibr" rid="B4">Bekker, 2024</xref>).</p>
</sec>
<sec>
<title>8.8 The role of educators in mediating LLM use</title>
<p>LLM integration necessitates rediscovering the educator&#x00027;s role&#x02014;from transmitter of knowledge to AI literacy and ethical engagement facilitator (<xref ref-type="bibr" rid="B28">Lazebnik and Rosenfeld, 2024</xref>). Teachers should take an active role in helping students understand the limitations of LLMs, differentiate between responsible use and misuse, and navigate the ethical considerations associated with AI-generated content (<xref ref-type="bibr" rid="B29">Lee, 2023</xref>). Training students to critically evaluate and revise LLM outputs contributes to deeper learning and helps prevent overreliance (<xref ref-type="bibr" rid="B34">Liao et al., 2023</xref>). Instructors can also promote transparency by encouraging students to document their use of LLMs in the writing process, thus reinforcing principles of academic integrity and accountability (<xref ref-type="bibr" rid="B41">Mahmoud and S&#x000F8;rensen, 2024</xref>). This approach fosters a culture of AI-augmented authorship, where students learn to integrate feedback rather than delegate writing tasks to an automated agent. This pedagogical vision also aligns with global policy agendas&#x02014;such as the UNESCO Futures of Education framework and the Sustainable Development Goal 4 (SDG4)&#x02014;by promoting inclusive, equitable, and future-ready higher education that incorporates responsible AI use.</p>
</sec>
<sec>
<title>8.9 Risks of unguided use and the need for policy</title>
<p>While the study found that even unguided use of LLMs can yield some benefits, particularly in writing quality, such benefits are significantly more limited without instructional scaffolding. Unguided use has several risks, including superficial engagement, skill stagnation, and academic integrity concerns. Without explicit instruction, students may bypass the cognitive and metacognitive processes essential to writing, relying instead on the fluency of LLMs to complete tasks (<xref ref-type="bibr" rid="B38">Lopes et al., 2024</xref>). Moreover, the indistinguishability of AI-generated text from human-authored work poses significant challenges for evaluation (<xref ref-type="bibr" rid="B13">De Villiers et al., 2024</xref>; <xref ref-type="bibr" rid="B52">Reinhart et al., 2024</xref>). This complicates the role of assessment and highlights the urgent need for institutional policies that address transparency, disclosure practices, and acceptable uses of generative AI in coursework. As noted in the limitations, behavioral metrics and clearer definitions of disciplinary norms could inform these policies, especially when automated scoring tools like Grammarly are involved.</p>
</sec>
<sec>
<title>8.10 Fostering independent skill development</title>
<p>Finally, educators must balance leveraging the benefits of LLMs and fostering independent writing skills (<xref ref-type="bibr" rid="B51">Patac and Patac Jr, 2025</xref>). While guided use can accelerate learning and reduce barriers, overdependence on AI tools may inhibit students&#x00027; ability to think critically and write autonomously (<xref ref-type="bibr" rid="B49">Ouwehand et al., 2025</xref>). Integrating LLMs should not replace traditional instruction but complement it through strategy-based interventions, peer review, and scaffolded writing tasks (<xref ref-type="bibr" rid="B31">Li et al., 2023</xref>; <xref ref-type="bibr" rid="B35">Lin et al., 2024</xref>). Future research might also explore how these strategies can be adapted to STEM disciplines or interdisciplinary programs, as differences in writing genres and task complexity may shape how students engage with LLMs. In line with SD4&#x02032;s commitment to inclusive and contextually sensitive education, such differentiated approaches are essential to ensuring that AI-enhanced instruction serves diverse learners effectively.</p>
</sec>
</sec>
<sec sec-type="conclusions" id="s9">
<title>9 Conclusion</title>
<p>This study provides empirical evidence for the differential effects of guided and unguided integration of LLMs on students&#x00027; academic writing quality, perceived mental health, and academic engagement in higher education. The findings demonstrate that guided LLM use consistently outperforms both unguided and no use, particularly in improving writing outcomes and fostering student engagement. These effects are most pronounced when LLMs are embedded within structured instructional frameworks that promote critical interaction, strategic thinking, and responsible use.</p>
<p>Notably, unguided LLM use yielded only partial benefits. While it enhanced academic writing quality relative to the control group, it did not significantly improve perceived mental health or engagement. These results suggest that unstructured exposure to generative AI may not be sufficient to produce holistic educational gains. Instead, pedagogical scaffolding and active instructor involvement appear essential to unlock the full potential of LLMs in supporting learning, well-being, and student agency. The study contributes to the growing literature on human&#x02013;AI collaboration in education by underscoring the importance of designing intentional, ethically informed, and learner-centered approaches to AI integration. As educational institutions increasingly adopt LLM-based tools, the distinction between guided and unguided use becomes pedagogically relevant&#x02014;because of its impact on learning quality, engagement, and wellbeing&#x02014;and ethically imperative, as it directly affects student autonomy, academic integrity, and equitable access to meaningful AI-supported education.</p>
<p>Looking ahead, future research should explore how guidance strategies can be tailored to different learning profiles, disciplines, and institutional cultures. Longitudinal and mixed-method designs may further illuminate the evolving relationship between students and AI, providing insights into how LLMs shape academic development. Ultimately, the challenge lies in providing access to powerful technologies and designing meaningful and equitable frameworks for their use&#x02014;frameworks that preserve learning integrity while embracing innovation. The findings of this study call for a thoughtful and pedagogically grounded integration of LLMs in higher education. When embedded in structured learning environments, guided use can enhance academic outcomes, support student wellbeing, and promote ethical use of AI. To realize these benefits, educators must assume an active role in designing, modeling, and monitoring AI engagement, ensuring that LLMs serve as tools for empowerment rather than substitution. Achieving this vision also depends on ongoing professional development for educators, who must be equipped not only with technical competencies but also with pedagogical strategies to guide students in critically and ethically navigating AI-supported academic tasks.</p>
</sec>
<sec id="s10">
<title>10 Declarations</title>
<p>Ethical approval and Informed Consent for participation: This study was reviewed and approved by the Institutional Review Board (IRB) of Xi&#x00027;an International University following the Declaration of Helsinki. Ethical approval was granted on [January 15th, 2024], under the reference number [approval ID, IRB/24/072-HUMARTS]. Before participation, all students received an information sheet outlining the purpose of the study, the nature of the tasks, the voluntary nature of their participation, and their right to withdraw at any point without penalty. Written informed consent was obtained from all participants. The consent form emphasized that participation was anonymous, data would be kept confidential, and results would be used solely for academic research. Participants were also informed that using the LLM was part of an experimental educational intervention and that their course grades would not be affected by their responses or participation.</p>
</sec>
</body>
<back>
<sec sec-type="data-availability" id="s11">
<title>Data availability statement</title>
<p>The datasets presented in this study can be found in online repositories. The names of the repository/repositories and accession number(s) can be found below: <ext-link ext-link-type="uri" xlink:href="https://osf.io/htejm/?view_only=54624dbd9f11467ea26242bae037e713lt;/bgt;">https://osf.io/htejm/?view_only=54624dbd9f11467ea26242bae037e713lt;/bgt;</ext-link>.</p>
</sec>
<sec sec-type="ethics-statement" id="s12">
<title>Ethics statement</title>
<p>The studies involving humans were approved by Institutional Review Board (IRB) of Xi&#x00027;an International University following the Declaration of Helsinki. Ethical approval was granted on [January 15th, 2024], under the reference number [approval ID, IRB/24/072-HUMARTS]. The studies were conducted in accordance with the local legislation and institutional requirements. The participants provided their written informed consent to participate in this study.</p>
</sec>
<sec sec-type="author-contributions" id="s13">
<title>Author contributions</title>
<p>MZ: Conceptualization, Resources, Visualization, Validation, Formal analysis, Funding acquisition, Project administration, Writing &#x02013; original draft, Data curation, Supervision, Investigation, Methodology, Writing &#x02013; review &#x00026; editing, Software.</p>
</sec>
<sec sec-type="funding-information" id="s14">
<title>Funding</title>
<p>The author declares that no financial support was received for the research and/or publication of this article.</p>
</sec>
<sec sec-type="COI-statement" id="conf1">
<title>Conflict of interest</title>
<p>The author declares that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec sec-type="ai-statement" id="s15">
<title>Generative AI statement</title>
<p>The author declares that no Gen AI was used in the creation of this manuscript. Generative AI was used in the preparation of this manuscript for revision of the language and the sentence structure. All the other content (data analyses and theoretical proposals) were written by the authors.</p>
<p>Any alternative text (alt text) provided alongside figures in this article has been generated by Frontiers with the support of artificial intelligence and reasonable efforts have been made to ensure accuracy, including review by the authors wherever possible. If you identify any issues, please contact us.</p>
</sec>
<sec sec-type="disclaimer" id="s16">
<title>Publisher&#x00027;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
<ref-list>
<title>References</title>
<ref id="B1">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Adams</surname> <given-names>R</given-names></name></person-group>. (<year>2024</year>) <source>The New Empire of AI: The Future of Global Inequality</source>. <publisher-loc>London</publisher-loc>. <publisher-name>Polity Press</publisher-name>.</citation>
</ref>
<ref id="B2">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Alsobeh</surname> <given-names>A.</given-names></name> <name><surname>Woodward</surname> <given-names>B.</given-names></name></person-group> (<year>2023</year>). <article-title>&#x0201C;AI as a Partner in Learning: a Novel Student-in-the-Loop Framework for Enhanced Student Engagement and Outcomes in Higher Education&#x0201D;</article-title> <source>Proceedings of the 24th Annual Conference on Information Technology Education, Marietta</source>, GA, <italic>USA</italic>. <pub-id pub-id-type="doi">10.1145/3585059.3611405</pub-id></citation>
</ref>
<ref id="B3">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ayeni</surname> <given-names>O. O.</given-names></name> <name><surname>Al Hamad</surname> <given-names>N. M.</given-names></name> <name><surname>Chisom</surname> <given-names>O. N.</given-names></name> <name><surname>Osawaru</surname> <given-names>B.</given-names></name> <name><surname>Adewusi</surname> <given-names>O. E.</given-names></name></person-group> (<year>2024</year>). <article-title>AI in education: a review of personalized learning and educational technology</article-title>. <source>GSC Adv. Res. Rev</source>. 18, 261-271. <pub-id pub-id-type="doi">10.30574/gscarr.2024.18.2.0062</pub-id></citation>
</ref>
<ref id="B4">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bekker</surname> <given-names>M.</given-names></name></person-group> (<year>2024</year>). <article-title>Large language models and academic writing: five tiers of engagement</article-title>. <source>South Afr. J. Sci</source>. 120, 1-5. <pub-id pub-id-type="doi">10.17159/sajs.2024/17147</pub-id></citation>
</ref>
<ref id="B5">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Beurer-Kellner</surname> <given-names>L.</given-names></name> <name><surname>Fischer</surname> <given-names>M.</given-names></name> <name><surname>Vechev</surname> <given-names>M.</given-names></name></person-group> (<year>2024</year>). <article-title>&#x0201C;Guiding LLMs the right way: fast, non-invasive constrained generation,&#x0201D;</article-title> in <source>Proceedings of the 41st International Conference on Machine Learning (Vienna)</source>, <fpage>3658</fpage>&#x02013;<lpage>3673</lpage>.</citation>
</ref>
<ref id="B6">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cash</surname> <given-names>T. N.</given-names></name> <name><surname>Oppenheimer</surname> <given-names>D. M.</given-names></name> <name><surname>Connell Pensky</surname> <given-names>A. E.</given-names></name></person-group> (<year>2025</year>). <source>You&#x00027;ve Got AI Friend in Me: LLMs as Collaborative Learning Partners.</source> <pub-id pub-id-type="doi">10.31219/osf.io/8q67u_v3</pub-id></citation>
</ref>
<ref id="B7">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chakrabarty</surname> <given-names>T.</given-names></name> <name><surname>Padmakumar</surname> <given-names>V.</given-names></name> <name><surname>Brahman</surname> <given-names>F.</given-names></name> <name><surname>Muresan</surname> <given-names>S.</given-names></name></person-group> (<year>2024</year>). Creativity support in the age of large language models: an empirical study involving professional writers <italic>Proceedings of the 16th Conference on Creativity and Cognition, Chicago: USA</italic>. <pub-id pub-id-type="doi">10.1145/3635636.3656201</pub-id></citation>
</ref>
<ref id="B8">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Chang</surname> <given-names>C.-H.</given-names></name></person-group> (<year>2024</year>). <source>A Study on the Impact of Two AI-Powered Writing Assistants on EFL Learners&#x00027; Writing</source> (<publisher-loc>ProQuest Dissertations and Theses</publisher-loc>), National Taiwan Normal University, Taiwan.</citation>
</ref>
<ref id="B9">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Chen</surname> <given-names>C.</given-names></name> <name><surname>Leitch</surname> <given-names>A.</given-names></name></person-group> (<year>2024</year>). <source>LLMs as Academic Reading Companions: Extending HCI through Synthetic Personae</source>. <publisher-loc>CHI 2024 Workshop</publisher-loc>: <publisher-name>Challenges and Opportunities of LLM-Based Synthetic Personae and Data in HCI, Honolulu, HI</publisher-name>.</citation>
</ref>
<ref id="B10">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cheng</surname> <given-names>X.</given-names></name> <name><surname>Zhang</surname> <given-names>X.</given-names></name> <name><surname>Cohen</surname> <given-names>J.</given-names></name> <name><surname>Mou</surname> <given-names>J.</given-names></name></person-group> (<year>2022</year>). <article-title>Human Vs</article-title>. <source>AI: understanding the impact of anthropomorphism on consumer response to chatbots from the perspective of trust and relationship norms. Inform. Process. Manag</source>. <volume>59</volume>:<fpage>102940</fpage>. <pub-id pub-id-type="doi">10.1016/j.ipm.2022.102940</pub-id></citation>
</ref>
<ref id="B11">
<citation citation-type="book"><person-group person-group-type="author"><name><surname>Chiang</surname> <given-names>C. H.</given-names></name> <name><surname>Lee</surname> <given-names>H. Y.</given-names></name></person-group> (<year>2023</year>). <article-title>&#x0201C;Can large language models be an alternative to human evaluations?,&#x0201D;</article-title> in <source>Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics</source> (<publisher-loc>Vol. 1</publisher-loc>: <publisher-name>Long Papers</publisher-name>), <fpage>15607</fpage>&#x02013;<lpage>15631</lpage>.</citation>
</ref>
<ref id="B12">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>De La Paz</surname> <given-names>S.</given-names></name></person-group> (<year>2005</year>). <article-title>Effects of historical reasoning instruction and writing strategy mastery in culturally and academically diverse middle school classrooms</article-title>. <source>Journal of Educational Psychology, 97(2)</source>, 139&#x02013;156. <pub-id pub-id-type="doi">10.1037/0022-0663.97.2.139</pub-id></citation>
</ref>
<ref id="B13">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>De Villiers</surname> <given-names>C.</given-names></name> <name><surname>Dimes</surname> <given-names>R.</given-names></name> <name><surname>Molinari</surname> <given-names>M.</given-names></name></person-group> (<year>2024</year>). <article-title>How Will AI Text Generation and Processing Impact Sustainability Reporting?</article-title> <source>Critical analysis, a conceptual framework and avenues for future research. Sustain. Account. Manag. Policy J</source>. <volume>15</volume>, <fpage>96</fpage>&#x02013;<lpage>118</lpage>. <pub-id pub-id-type="doi">10.1108/SAMPJ-02-2023-0097</pub-id><pub-id pub-id-type="pmid">38470476</pub-id></citation></ref>
<ref id="B14">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Deci</surname> <given-names>E. L.</given-names></name> <name><surname>Ryan</surname> <given-names>R. M.</given-names></name></person-group> (<year>2000</year>). <article-title>The &#x0201C;what&#x0201D; and &#x0201C;why&#x0201D; of goal pursuits: Human needs and the self-determination of behavior</article-title>. <source>Psychological Inquiry</source>, <volume>11</volume>, <fpage>227</fpage>&#x02013;<lpage>268</lpage>. <pub-id pub-id-type="doi">10.1207/S15327965PLI1104_01</pub-id></citation>
</ref>
<ref id="B15">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ding</surname> <given-names>Z.</given-names></name> <name><surname>Smith-Renner</surname> <given-names>A.</given-names></name> <name><surname>Zhang</surname> <given-names>W.</given-names></name> <name><surname>Tetreault</surname> <given-names>J.</given-names></name> <name><surname>Jaimes</surname> <given-names>A.</given-names></name></person-group> (<year>2023</year>). Harnessing the power of Llms: evaluating human-Ai Text co-creation through the lens of news headline generation.Findings of the Association for Computational Linguistics: EMNLP 2023<italic>, The 2023 Conference on Empirical Methods in Natural Language Processing. Singapur</italic>. <pub-id pub-id-type="doi">10.18653/v1/2023.findings-emnlp.217</pub-id><pub-id pub-id-type="pmid">36568019</pub-id></citation></ref>
<ref id="B16">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fang</surname> <given-names>L. T.</given-names></name> <name><surname>Shi</surname> <given-names>K.</given-names></name> <name><surname>Zhang</surname> <given-names>F.</given-names></name></person-group> (<year>2008</year>). <article-title>Research on reliability and validity of Utrecht work engagement scale-student</article-title>. <source>Chinese J. Clin. Psychol</source>. <volume>16</volume>, <fpage>618</fpage>&#x02013;<lpage>620</lpage>.<pub-id pub-id-type="pmid">39351103</pub-id></citation></ref>
<ref id="B17">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Faul</surname> <given-names>F.</given-names></name> <name><surname>Erdfelder</surname> <given-names>E.</given-names></name> <name><surname>Buchner</surname> <given-names>A.</given-names></name> <name><surname>Lang</surname> <given-names>A.-G.</given-names></name></person-group> (<year>2009</year>). <article-title>Statistical power analyses using G<sup>&#x0002A;</sup> power 3.1: tests for correlation and regression analyses</article-title>. <source>Behav. Res. Methods</source> <volume>41</volume>, <fpage>1149</fpage>&#x02013;<lpage>1160</lpage>. <pub-id pub-id-type="doi">10.3758/BRM.41.4.1149</pub-id><pub-id pub-id-type="pmid">19897823</pub-id></citation></ref>
<ref id="B18">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Flesch</surname> <given-names>R.</given-names></name></person-group> (<year>1948</year>). <article-title>A new readability yardstick</article-title>. <source>J. Appl. Psychol</source>. <volume>32</volume>, <fpage>221</fpage>&#x02013;<lpage>233</lpage>. <pub-id pub-id-type="doi">10.1037/h0057532</pub-id><pub-id pub-id-type="pmid">18867058</pub-id></citation></ref>
<ref id="B19">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fredricks</surname> <given-names>J. A.</given-names></name> <name><surname>Blumenfeld</surname> <given-names>P. C.</given-names></name> <name><surname>Paris</surname> <given-names>A. H.</given-names></name></person-group> (<year>2004</year>). <article-title>School engagement: potential of the concept, state of the evidence</article-title>. <source>Rev. Educ. Res</source>. <volume>74</volume>, <fpage>59</fpage>&#x02013;<lpage>109</lpage>. <pub-id pub-id-type="doi">10.3102/00346543074001059</pub-id><pub-id pub-id-type="pmid">38293548</pub-id></citation></ref>
<ref id="B20">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fritz</surname> <given-names>T.</given-names></name> <name><surname>Gonz&#x000E1;lez Cruz</surname> <given-names>H.</given-names></name> <name><surname>Janke</surname> <given-names>S.</given-names></name> <name><surname>Daumiller</surname> <given-names>M.</given-names></name></person-group> (<year>2024</year>). <article-title>How to best measure academic dishonesty in students: a systematic review of self-report assessment methods and psychometric quality</article-title>. <source>Eur. J. Psychol. Assess</source>. <volume>40</volume>, <fpage>498</fpage>&#x02013;<lpage>514</lpage>. <pub-id pub-id-type="doi">10.1027/1015-5759/a000861</pub-id></citation>
</ref>
<ref id="B21">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>G&#x000F3;mez-Rodr&#x000ED;guez</surname> <given-names>C.</given-names></name> <name><surname>Williams</surname> <given-names>P.</given-names></name></person-group> (<year>2023</year>). <article-title>&#x0201C;A confederacy of models: a comprehensive evaluation of LLMs on creative writing,&#x0201D;</article-title> in <source>Findings of the Association for Computational Linguistics: EMNLP 2023,14504&#x02013;14528, Singapore. Association for Computational Linguistics.</source> <pub-id pub-id-type="doi">10.18653/v1/2023.findings-emnlp.966</pub-id><pub-id pub-id-type="pmid">36568019</pub-id></citation></ref>
<ref id="B22">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Granieri</surname> <given-names>A.</given-names></name> <name><surname>Franzoi</surname> <given-names>I. G.</given-names></name> <name><surname>Chung</surname> <given-names>M. C.</given-names></name></person-group> (<year>2021</year>). <article-title>Editorial: psychological distress among university students</article-title>. <source>Front. Psychol.</source> <volume>12</volume>:<fpage>647940</fpage>. <pub-id pub-id-type="doi">10.3389/fpsyg.2021.647940</pub-id><pub-id pub-id-type="pmid">33828513</pub-id></citation></ref>
<ref id="B23">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jungherr</surname> <given-names>A.</given-names></name></person-group> (<year>2023</year>). <source>Using Chatgpt and Other Large Language Model (LLM) Applications for Academic Paper Assignments</source>. Bamberg: Otto-Friedrich-Universit&#x000E4;t. <pub-id pub-id-type="doi">10.31235/osf.io/d84q6</pub-id><pub-id pub-id-type="pmid">33784477</pub-id></citation></ref>
<ref id="B24">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kulaksiz</surname> <given-names>G. C.</given-names></name></person-group> (<year>2024</year>). <source>Artificial Intelligence-Based Language Modelling: The Effect of Chatgpt Application on Writing Skills in the Context of Teaching English as a Foreign Language.</source> Istambul. Bursa Uludag &#x000DC;niversitesi.</citation>
</ref>
<ref id="B25">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kumar</surname> <given-names>A.</given-names></name> <name><surname>Prol</surname> <given-names>D.</given-names></name> <name><surname>Alipour</surname> <given-names>A.</given-names></name> <name><surname>Ragavan</surname> <given-names>S. S.</given-names></name></person-group> (<year>2025</year>). <article-title>To google or To ChatGPT? A comparison of CS2 students&#x00027; information gathering approaches and outcomes</article-title>. <source>arXiv preprint</source> arXiv:2501.11935.</citation>
</ref>
<ref id="B26">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kumar</surname> <given-names>H.</given-names></name> <name><surname>Musabirov</surname> <given-names>I.</given-names></name> <name><surname>Reza</surname> <given-names>M.</given-names></name> <name><surname>Shi</surname> <given-names>J.</given-names></name> <name><surname>Wang</surname> <given-names>X.</given-names></name> <name><surname>Williams</surname> <given-names>J. J.</given-names></name> <etal/></person-group>. (<year>2023</year>). <article-title>Impact of guidance and interaction strategies for LLM use on Learner Performance and perception</article-title>. <source>arXiv preprint</source> arXiv:2310.13712.</citation>
</ref>
<ref id="B27">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kumar</surname> <given-names>H.</given-names></name> <name><surname>Reza</surname> <given-names>M.</given-names></name> <name><surname>Mitchell</surname> <given-names>J.</given-names></name> <name><surname>Musabirov</surname> <given-names>I.</given-names></name> <name><surname>Zhang</surname> <given-names>L.</given-names></name> <name><surname>Liut</surname> <given-names>M.</given-names></name></person-group> (<year>2024</year>). Understanding help&#x02013;seeking behavior of students using Llms Vs. web search for writing sql queries. PsyAxiv preprint. <pub-id pub-id-type="doi">10.1145/3735091.3737569</pub-id></citation>
</ref>
<ref id="B28">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lazebnik</surname> <given-names>T.</given-names></name> <name><surname>Rosenfeld</surname> <given-names>A.</given-names></name></person-group> (<year>2024</year>). Detecting Llm-assisted writing in scientific communication: are we there yet? <italic>J. Data Inform. Sci</italic>. 9. <pub-id pub-id-type="doi">10.2478/jdis-2024-0020</pub-id></citation>
</ref>
<ref id="B29">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lee</surname> <given-names>S.-M.</given-names></name></person-group> (<year>2023</year>). <article-title>The effectiveness of machine translation in foreign language education: a systematic review and meta-analysis</article-title>. <source>Comput. Assis. Lang. Learn</source>. <volume>36</volume>, <fpage>103</fpage>&#x02013;<lpage>125</lpage>. <pub-id pub-id-type="doi">10.1080/09588221.2021.1901745</pub-id></citation>
</ref>
<ref id="B30">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lee</surname> <given-names>Y.</given-names></name> <name><surname>Ka</surname> <given-names>S.</given-names></name> <name><surname>Son</surname> <given-names>B.</given-names></name> <name><surname>Kang</surname> <given-names>P.</given-names></name> <name><surname>Kang</surname> <given-names>J.</given-names></name></person-group> (<year>2024</year>). <article-title>Navigating the path of writing: outline-guided text generation with large language models</article-title>. <source>PsyAxiv preprint</source>. <pub-id pub-id-type="doi">10.18653/v1/2025.naacl-industry.20</pub-id><pub-id pub-id-type="pmid">36568019</pub-id></citation></ref>
<ref id="B31">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Li</surname> <given-names>C.</given-names></name> <name><surname>Zhang</surname> <given-names>M.</given-names></name> <name><surname>Mei</surname> <given-names>Q.</given-names></name> <name><surname>Wang</surname> <given-names>Y.</given-names></name> <name><surname>Amba Hombaiah</surname> <given-names>S.</given-names></name> <name><surname>Liang</surname> <given-names>Y.</given-names></name> <name><surname>Bendersky</surname> <given-names>M.</given-names></name></person-group> (<year>2023</year>). <source>Teach Llms to Personalize- an Approach Inspired by Writing</source> Education.</citation>
</ref>
<ref id="B32">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Li</surname> <given-names>C.</given-names></name> <name><surname>Zhang</surname> <given-names>Y.</given-names></name> <name><surname>Randhawa</surname> <given-names>A. K.</given-names></name> <name><surname>Madigan</surname> <given-names>D. J.</given-names></name></person-group> (<year>2020</year>). <article-title>Emotional exhaustion and sleep problems in university students: does mental toughness matter?</article-title> <source>Personal. Individ. Diff</source>. <volume>163</volume>:<fpage>110046</fpage>. <pub-id pub-id-type="doi">10.1016/j.paid.2020.110046</pub-id></citation>
</ref>
<ref id="B33">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Li</surname> <given-names>R.-H.</given-names></name></person-group> (<year>2014</year>). <article-title>Reliability and validity of a shorter chinese version for ryff&#x00027;s psychological well-being scale</article-title>. <source>Health Educ. J</source>. <volume>73</volume>, <fpage>446</fpage>&#x02013;<lpage>452</lpage>. <pub-id pub-id-type="doi">10.1177/0017896913485743</pub-id></citation>
</ref>
<ref id="B34">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Liao</surname> <given-names>W.</given-names></name> <name><surname>Liu</surname> <given-names>Z.</given-names></name> <name><surname>Dai</surname> <given-names>H.</given-names></name> <name><surname>Xu</surname> <given-names>S.</given-names></name> <name><surname>Wu</surname> <given-names>Z.</given-names></name> <name><surname>Zhang</surname> <given-names>Y.</given-names></name> <name><surname>Huang</surname> <given-names>X.</given-names></name> <name><surname>Zhu</surname> <given-names>D.</given-names></name> <name><surname>Cai</surname> <given-names>H.</given-names></name> <name><surname>Li</surname> <given-names>Q.</given-names></name></person-group> (<year>2023</year>). <article-title>Differentiating chatgpt-generated and human-written medical texts: quantitative study</article-title>. <source>J. Med. Int. Res. Med. Educ</source>. <volume>9</volume>:<fpage>e48904</fpage>. <pub-id pub-id-type="doi">10.2196/48904</pub-id><pub-id pub-id-type="pmid">38153785</pub-id></citation></ref>
<ref id="B35">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lin</surname> <given-names>H.-C. K.</given-names></name> <name><surname>Lu</surname> <given-names>L.-W.</given-names></name> <name><surname>Lu</surname> <given-names>R.-S.</given-names></name></person-group> (<year>2024</year>). <article-title>Integrating digital technologies and alternate reality games for sustainable education: enhancing cultural heritage awareness and learning engagement</article-title>. <source>Sustainability</source> <volume>16</volume>:<fpage>9451</fpage>. <pub-id pub-id-type="doi">10.3390/su16219451</pub-id></citation>
</ref>
<ref id="B36">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lin</surname> <given-names>X.</given-names></name></person-group> (<year>2024</year>). <article-title>Exploring the role of chatgpt as a facilitator for motivating self-directed learning among adult learners</article-title>. <source>Adult Learn</source>. <volume>35</volume>, <fpage>156</fpage>&#x02013;<lpage>166</lpage>. <pub-id pub-id-type="doi">10.1177/10451595231184928</pub-id></citation>
</ref>
<ref id="B37">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Liu</surname> <given-names>P.</given-names></name> <name><surname>Yuan</surname> <given-names>W.</given-names></name> <name><surname>Fu</surname> <given-names>J.</given-names></name> <name><surname>Jiang</surname> <given-names>Z.</given-names></name> <name><surname>Hayashi</surname> <given-names>H.</given-names></name> <name><surname>Neubig</surname> <given-names>G.</given-names></name></person-group> (<year>2023</year>). <article-title>Pre-Train, prompt, and predict: a systematic survey of prompting methods in natural language processing</article-title>. <source>ACM Comput. Surv</source>. <volume>55</volume>, <fpage>1</fpage>&#x02013;<lpage>35</lpage>. <pub-id pub-id-type="doi">10.1145/3560815</pub-id></citation>
</ref>
<ref id="B38">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lopes</surname> <given-names>R. M.</given-names></name> <name><surname>Silva</surname> <given-names>A. F.</given-names></name> <name><surname>Rodrigues</surname> <given-names>A. C. A.</given-names></name> <name><surname>Melo</surname> <given-names>V.</given-names></name></person-group> (<year>2024</year>). <article-title>Chatbots for well-being: exploring the impact of artificial intelligence on mood enhancement and mental health</article-title>. <source>Eur. Psychiat.</source> <volume>67</volume>, <fpage>S550</fpage>&#x02013;<lpage>S551</lpage>. <pub-id pub-id-type="doi">10.1192/j.eurpsy.2024.1143</pub-id></citation>
</ref>
<ref id="B39">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lu</surname> <given-names>Q.</given-names></name> <name><surname>Yao</surname> <given-names>Y.</given-names></name> <name><surname>Xiao</surname> <given-names>L.</given-names></name> <name><surname>Yuan</surname> <given-names>M.</given-names></name> <name><surname>Wang</surname> <given-names>J.</given-names></name> <name><surname>Zhu</surname> <given-names>X.</given-names></name></person-group> (<year>2024</year>). <article-title>Can chatgpt effectively complement teacher assessment of undergraduate students&#x00027; academic writing?</article-title> <source>Assess. Eval. High. Educ</source>. <volume>49</volume>, <fpage>616</fpage>&#x02013;<lpage>633</lpage>. <pub-id pub-id-type="doi">10.1080/02602938.2024.2301722</pub-id></citation>
</ref>
<ref id="B40">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ma</surname> <given-names>Z.</given-names></name> <name><surname>Mei</surname> <given-names>Y.</given-names></name> <name><surname>Su</surname> <given-names>Z.</given-names></name></person-group> (<year>2023</year>). <article-title>Understanding the benefits and challenges of using large language model-based conversational agents for mental well-being support</article-title>. <source>AMIA Sympos</source>. <volume>2023</volume>, <fpage>1105</fpage>&#x02013;<lpage>1114</lpage>.<pub-id pub-id-type="pmid">38222348</pub-id></citation></ref>
<ref id="B41">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mahmoud</surname> <given-names>C.</given-names></name> <name><surname>S&#x000F8;rensen</surname> <given-names>J.</given-names></name></person-group> (<year>2024</year>). <article-title>Artificial Intelligence in personalized learning with a focus on current developments and future prospects</article-title>. <source>Res. Adv. Educ.</source> <volume>3</volume>, <fpage>25</fpage>-<lpage>31</lpage>. <pub-id pub-id-type="doi">10.56397/RAE.2024.08.04</pub-id></citation>
</ref>
<ref id="B42">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Maiti</surname> <given-names>P.</given-names></name> <name><surname>Goel</surname> <given-names>A. K.</given-names></name></person-group> (<year>2024</year>). <article-title>How do students interact with an LLM-powered virtual teaching assistant in different educational settings?</article-title>. <source>arXiv preprint</source> arXiv:2407.17429.</citation>
</ref>
<ref id="B43">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Meyer</surname> <given-names>J.</given-names></name> <name><surname>Jansen</surname> <given-names>T.</given-names></name> <name><surname>Schiller</surname> <given-names>R.</given-names></name> <name><surname>Liebenow</surname> <given-names>L. W.</given-names></name> <name><surname>Steinbach</surname> <given-names>M.</given-names></name> <name><surname>Horbach</surname> <given-names>A.</given-names></name> <name><surname>Fleckenstein</surname> <given-names>J.</given-names></name></person-group> (<year>2024</year>). <article-title>Using LLMs to bring evidence-based feedback into the classroom: ai-generated feedback increases secondary students&#x00027; text revision, motivation, and positive emotions</article-title>. <source>Comput. Educ. Art. Intell.</source> <volume>6</volume>:<fpage>100199</fpage>. <pub-id pub-id-type="doi">10.1016/j.caeai.2023.100199</pub-id></citation>
</ref>
<ref id="B44">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Meyer</surname> <given-names>S.</given-names></name> <name><surname>Elsweiler</surname> <given-names>D.</given-names></name></person-group> (<year>2025</year>). <article-title>Llm-based conversational agents for behaviour change support: a randomised controlled trial examining efficacy, safety, and the role of user behaviour</article-title>. <source>Int. J. Hum. Comput. Stud.</source> <volume>200</volume>:<fpage>103514</fpage>. <pub-id pub-id-type="doi">10.1016/j.ijhcs.2025.103514</pub-id></citation>
</ref>
<ref id="B45">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mishra</surname> <given-names>P.</given-names></name> <name><surname>Koehler</surname> <given-names>M. J.</given-names></name></person-group> (<year>2006</year>). <article-title>Technological pedagogical content knowledge: a framework for teacher knowledge</article-title>. <source>Teach. Coll. Rec</source>. <volume>108</volume>, <fpage>1017</fpage>&#x02013;<lpage>1054</lpage>. <pub-id pub-id-type="doi">10.1111/j.1467-9620.2006.00684.x</pub-id></citation>
</ref>
<ref id="B46">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Molodynski</surname> <given-names>A.</given-names></name> <name><surname>Lewis</surname> <given-names>T.</given-names></name> <name><surname>Kadhum</surname> <given-names>M.</given-names></name> <name><surname>Farrell</surname> <given-names>S. M.</given-names></name> <name><surname>Lemtiri Chelieh</surname> <given-names>M.</given-names></name> <name><surname>Falc&#x000E3;o De Almeida</surname> <given-names>T.</given-names></name> <name><surname>Masri</surname> <given-names>R.</given-names></name> <name><surname>Kar</surname> <given-names>A.</given-names></name> <name><surname>Volpe</surname> <given-names>U.</given-names></name> <name><surname>Moir</surname> <given-names>F.</given-names></name></person-group> (<year>2021</year>). <article-title>Cultural variations in wellbeing, burnout and substance use amongst medical students in twelve countries</article-title>. <source>Int. Rev. Psychiatr.</source> <volume>33</volume>, <fpage>37</fpage>&#x02013;<lpage>42</lpage>. <pub-id pub-id-type="doi">10.1080/09540261.2020.1738064</pub-id><pub-id pub-id-type="pmid">32186412</pub-id></citation></ref>
<ref id="B47">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Neary</surname> <given-names>M.</given-names></name> <name><surname>Schueller</surname> <given-names>S. M.</given-names></name></person-group> (<year>2018</year>). <article-title>State of the field of mental health apps</article-title>. <source>Cogn. Behav. Pract</source>. <volume>25</volume>, <fpage>531</fpage>&#x02013;<lpage>537</lpage>. <pub-id pub-id-type="doi">10.1016/j.cbpra.2018.01.002</pub-id><pub-id pub-id-type="pmid">33100810</pub-id></citation></ref>
<ref id="B48">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ng</surname> <given-names>D. T. K.</given-names></name> <name><surname>Tan</surname> <given-names>C. W.</given-names></name> <name><surname>Leung</surname> <given-names>J. K. L.</given-names></name></person-group> (<year>2024</year>). <article-title>Empowering student self-regulated learning and science education through chatgpt: a pioneering pilot study</article-title>. <source>Br. J. Educ. Technol.</source> <volume>55</volume>, <fpage>1328</fpage>&#x02013;<lpage>1353</lpage>. <pub-id pub-id-type="doi">10.1111/bjet.13454</pub-id></citation>
</ref>
<ref id="B49">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ouwehand</surname> <given-names>K.</given-names></name> <name><surname>Lespiau</surname> <given-names>F.</given-names></name> <name><surname>Tricot</surname> <given-names>A.</given-names></name> <name><surname>Paas</surname> <given-names>F.</given-names></name></person-group> (<year>2025</year>). <article-title>Cognitive load theory: emerging trends and innovations</article-title>. <source>Educ. Sci</source>. <volume>15</volume>:<fpage>458</fpage>. <pub-id pub-id-type="doi">10.3390/educsci15040458</pub-id></citation>
</ref>
<ref id="B50">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Park</surname> <given-names>H.</given-names></name> <name><surname>Ahn</surname> <given-names>D.</given-names></name></person-group> (<year>2024</year>). <article-title>The promise and peril of chatgpt in higher education: opportunities, challenges, and design implications</article-title>. <source>Proceedings of the 2024 CHI Conference on Human Factors in Computing Systems</source> <volume>271</volume>, <fpage>1</fpage>&#x02013;<lpage>21</lpage>. <pub-id pub-id-type="doi">10.1145/3613904.3642785</pub-id></citation>
</ref>
<ref id="B51">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Patac</surname> <given-names>L. P.</given-names></name> <name><surname>Patac Jr</surname> <given-names>A. V.</given-names></name></person-group> (<year>2025</year>). <article-title>Using chatgpt for academic support: managing cognitive load and enhancing learning efficiency&#x02013;a phenomenological approach</article-title>. <source>Soc. Sci. Human. Open</source> <volume>11</volume>:<fpage>101301</fpage>. <pub-id pub-id-type="doi">10.1016/j.ssaho.2025.101301</pub-id></citation>
</ref>
<ref id="B52">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Reinhart</surname> <given-names>A.</given-names></name> <name><surname>Brown</surname> <given-names>D.</given-names></name> <name><surname>Markey</surname> <given-names>B.</given-names></name> <name><surname>Laudenbach</surname> <given-names>M.</given-names></name> <name><surname>Pantusen</surname> <given-names>K.</given-names></name> <name><surname>Yurko</surname> <given-names>R.</given-names></name> <name><surname>Weinberg</surname> <given-names>G.</given-names></name></person-group> (<year>2024</year>). <source>Do LLMs Write Like Humans? Variation in Grammatical and Rhetorical Styles.</source> <pub-id pub-id-type="doi">10.1073/pnas.2422455122</pub-id></citation>
</ref>
<ref id="B53">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Robleto</surname> <given-names>E.</given-names></name> <name><surname>Habashi</surname> <given-names>A.</given-names></name> <name><surname>Kaplan</surname> <given-names>M.-A. B.</given-names></name> <name><surname>Riley</surname> <given-names>R. L.</given-names></name> <name><surname>Zhang</surname> <given-names>C.</given-names></name> <name><surname>Bianchi</surname> <given-names>L.</given-names></name> <name><surname>Shehadeh</surname> <given-names>L. A.</given-names></name></person-group> (<year>2024</year>). <article-title>Medical Students&#x00027; Perceptions of an Artificial Intelligence (AI) Assisted Diagnosing Program</article-title>. <source>Med. Teach</source>. <volume>46</volume>, <fpage>1180</fpage>&#x02013;<lpage>1186</lpage>. <pub-id pub-id-type="doi">10.1080/0142159X.2024.2305369</pub-id><pub-id pub-id-type="pmid">38306667</pub-id></citation></ref>
<ref id="B54">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Roshanaei</surname> <given-names>M.</given-names></name></person-group> (<year>2024</year>). <article-title>Towards best practices for mitigating artificial intelligence implicit bias in shaping diversity, inclusion and equity in higher education</article-title>. <source>Educ. Inform. Technol</source>. <volume>29</volume>, <fpage>18959</fpage>&#x02013;<lpage>18984</lpage>. <pub-id pub-id-type="doi">10.1007/s10639-024-12605-2</pub-id></citation>
</ref>
<ref id="B55">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ryan</surname> <given-names>R. M.</given-names></name> <name><surname>Deci</surname> <given-names>E. L.</given-names></name></person-group> (<year>2000</year>). <article-title>Self-determination theory and the facilitation of intrinsic motivation, social development, and well-being</article-title>. <source>Am. Psychol</source>. <volume>55</volume>, <fpage>68</fpage>&#x02013;<lpage>78</lpage>. <pub-id pub-id-type="doi">10.1037/0003-066X.55.1.68</pub-id><pub-id pub-id-type="pmid">11392867</pub-id></citation></ref>
<ref id="B56">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ryff</surname> <given-names>C. D.</given-names></name></person-group> (<year>1989</year>). <article-title>Happiness Is everything, or Is It?</article-title> <source>Explorations on the meaning of psychological well-being. J. Personal. Soc. Psychol</source>. <volume>57</volume>, <fpage>1069</fpage>&#x02013;<lpage>1081</lpage>. <pub-id pub-id-type="doi">10.1037/0022-3514.57.6.1069</pub-id></citation>
</ref>
<ref id="B57">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ryff</surname> <given-names>C. D.</given-names></name> <name><surname>Keyes</surname> <given-names>C. L. M.</given-names></name></person-group> (<year>1995</year>). <article-title>The structure of psychological well-being revisited</article-title>. <source>J. Personal. Soc. Psychol.</source> <volume>69</volume>, <fpage>719</fpage>&#x02013;<lpage>727</lpage>. <pub-id pub-id-type="doi">10.1037/0022-3514.69.4.719</pub-id><pub-id pub-id-type="pmid">7473027</pub-id></citation></ref>
<ref id="B58">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Salimi</surname> <given-names>E. A.</given-names></name> <name><surname>Hajinia</surname> <given-names>M.</given-names></name></person-group> (<year>2025</year>). <source>Large Language Models and Academic writing in practice: exploring participants&#x00027; utilization of generative pretrained transformers during an AI-Assisted Course on Writing Research Papers. PsyAxiv preprint.</source> <pub-id pub-id-type="doi">10.21203/rs.3.rs-5534554/v1</pub-id></citation>
</ref>
<ref id="B59">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Schaufeli</surname> <given-names>W. B.</given-names></name> <name><surname>Bakker</surname> <given-names>A. B.</given-names></name></person-group> (<year>2004</year>). <article-title>Job demands, job resources, and their relationship with burnout and engagement: a multi-sample study</article-title>. <source>J. Organ. Behav</source>. <volume>25</volume>, <fpage>293</fpage>&#x02013;<lpage>315</lpage>. <pub-id pub-id-type="doi">10.1002/job.248</pub-id></citation>
</ref>
<ref id="B60">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Schaufeli</surname> <given-names>W. B.</given-names></name> <name><surname>Martinez</surname> <given-names>I. M.</given-names></name> <name><surname>Pinto</surname> <given-names>A. M.</given-names></name> <name><surname>Salanova</surname> <given-names>M.</given-names></name> <name><surname>Bakker</surname> <given-names>A. B.</given-names></name></person-group> (<year>2002</year>). <article-title>Burnout and engagement in university students: a cross-national study</article-title>. <source>J. Cross Cult. Psychol</source>. <volume>33</volume>, <fpage>464</fpage>&#x02013;<lpage>481</lpage>. <pub-id pub-id-type="doi">10.1177/0022022102033005003</pub-id></citation>
</ref>
<ref id="B61">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sharma</surname> <given-names>S.</given-names></name> <name><surname>Mittal</surname> <given-names>P.</given-names></name> <name><surname>Kumar</surname> <given-names>M.</given-names></name> <etal/></person-group>. (<year>2025</year>).The role of Large Language Models in personalized learning: a systematic review of educational impact. <source>Discov. Sustain.</source> <volume>6</volume>:<fpage>243</fpage>. <pub-id pub-id-type="doi">10.1007/s43621-025-01094-z</pub-id></citation>
</ref>
<ref id="B62">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Spitale</surname> <given-names>M.</given-names></name> <name><surname>Axelsson</surname> <given-names>M.</given-names></name> <name><surname>Gunes</surname> <given-names>H.</given-names></name></person-group> (<year>2025</year>). <article-title>Vita: a multi-modal llm-based system for longitudinal, autonomous, and adaptive robotic mental well-being coaching</article-title>. <source>ACM Transact. Hum. Robot Interact.</source> 14. <pub-id pub-id-type="doi">10.1145/3712265</pub-id></citation>
</ref>
<ref id="B63">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sweller</surname> <given-names>J.</given-names></name></person-group> (<year>1988</year>). <article-title>cognitive load during problem solving: effects on learning</article-title>. <source>Cogn. Sci</source>. <volume>12</volume>, <fpage>257</fpage>&#x02013;<lpage>285</lpage>. <pub-id pub-id-type="doi">10.1207/s15516709cog1202_4</pub-id></citation>
</ref>
<ref id="B64">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Uchendu</surname> <given-names>A.</given-names></name> <name><surname>Lee</surname> <given-names>J.</given-names></name> <name><surname>Shen</surname> <given-names>H.</given-names></name> <name><surname>Le</surname> <given-names>T.</given-names></name> <name><surname>Lee</surname> <given-names>D.</given-names></name></person-group> (<year>2023</year>). <article-title>Does human collaboration enhance the accuracy of identifying llm-generated deepfake texts?</article-title> <source>Proceedings of the AAAI Conference on Human Computation and Crowdsourcing</source>, <volume>11</volume>, <fpage>163</fpage>&#x02013;<lpage>174</lpage>. <pub-id pub-id-type="doi">10.1609/hcomp.v11i1.27557</pub-id></citation>
</ref>
<ref id="B65">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wang</surname> <given-names>L.</given-names></name></person-group> (<year>2022</year>). <article-title>Student intrinsic motivation for online creative idea generation: mediating effects of student online learning engagement and moderating effects of teacher emotional support</article-title>. <source>Front. Psychol</source>. <volume>13</volume>:<fpage>954216</fpage>. <pub-id pub-id-type="doi">10.3389/fpsyg.2022.954216</pub-id><pub-id pub-id-type="pmid">35910980</pub-id></citation></ref>
<ref id="B66">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wu</surname> <given-names>S.</given-names></name> <name><surname>Han</surname> <given-names>F.</given-names></name> <name><surname>Yao</surname> <given-names>B.</given-names></name> <name><surname>Xie</surname> <given-names>T.</given-names></name> <name><surname>Zhao</surname> <given-names>X.</given-names></name> <name><surname>Wang</surname> <given-names>D.</given-names></name></person-group> (<year>2024</year>). <article-title>Sunnie: an anthropomorphic LLM-based conversational agent for mental well-being activity recommendation</article-title>. <source>arXiv preprint</source> arXiv:2405.13803.</citation>
</ref>
<ref id="B67">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Xu</surname> <given-names>X.</given-names></name> <name><surname>Qiao</surname> <given-names>L.</given-names></name> <name><surname>Cheng</surname> <given-names>N.</given-names></name> <name><surname>Liu</surname> <given-names>H.</given-names></name> <name><surname>Zhao</surname> <given-names>W.</given-names></name></person-group> (<year>2025</year>). <article-title>Enhancing self-regulated learning and learning experience in generative ai environments: the critical role of metacognitive support</article-title>. <source>Br. J. Educ. Technol. Early View</source> <pub-id pub-id-type="doi">10.1111/bjet.13599</pub-id></citation>
</ref>
<ref id="B68">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Yan</surname> <given-names>L.</given-names></name> <name><surname>Sha</surname> <given-names>L.</given-names></name> <name><surname>Zhao</surname> <given-names>L.</given-names></name> <name><surname>Li</surname> <given-names>Y.</given-names></name> <name><surname>Martinez-Maldonado</surname> <given-names>R.</given-names></name> <name><surname>Chen</surname> <given-names>G.</given-names></name> <name><surname>Li</surname> <given-names>X.</given-names></name> <name><surname>Jin</surname> <given-names>Y.</given-names></name> <name><surname>Ga&#x00161;evi&#x00107;</surname> <given-names>D.</given-names></name></person-group> (<year>2024</year>). <article-title>Practical and ethical challenges of Large Language Models in education: a systematic scoping review</article-title>. <source>Br. J. Educ. Technol</source>. <volume>55</volume>, <fpage>90</fpage>&#x02013;<lpage>112</lpage>. <pub-id pub-id-type="doi">10.1111/bjet.13370</pub-id><pub-id pub-id-type="pmid">38746668</pub-id></citation></ref>
<ref id="B69">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Youn</surname> <given-names>S.</given-names></name> <name><surname>Jin</surname> <given-names>S. V.</given-names></name></person-group> (<year>2021</year>). <article-title>In AI we trust?&#x0201D; the effects of parasocial interaction and technopian vs. luddite ideological views on chatbot-based customer relationship management in the emerging &#x0201C;feeling economy</article-title>. <source>Comput. Hum. Behav.</source> <volume>119</volume>:<fpage>106721</fpage>. <pub-id pub-id-type="doi">10.1016/j.chb.2021.106721</pub-id></citation>
</ref>
<ref id="B70">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zhang</surname> <given-names>O. X.</given-names></name> <name><surname>Zhou</surname> <given-names>S.</given-names></name> <name><surname>Geng</surname> <given-names>J.</given-names></name> <name><surname>Liu</surname> <given-names>Y.</given-names></name> <name><surname>Liu</surname> <given-names>S. X.</given-names></name></person-group> (<year>2024</year>). <article-title>Dr. GPT in campus counseling: understanding higher education students&#x00027; opinions on LLM-assisted mental health services</article-title>. <source>arXiv preprint</source> arXiv:2409.17572.</citation>
</ref>
<ref id="B71">
<citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zhu</surname> <given-names>Y.</given-names></name></person-group> (<year>2024</year>). <article-title>The impact of AI-assisted teaching on students&#x00027; learning and psychology</article-title>. <source>J. Educ. Humanit. Soc. Sci</source>. <volume>38</volume>, <fpage>111</fpage>&#x02013;<lpage>116</lpage>. <pub-id pub-id-type="doi">10.54097/k7a37d11</pub-id></citation>
</ref>
</ref-list>
</back>
</article>