<?xml version="1.0" encoding="UTF-8" standalone="no"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xml:lang="EN" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Med.</journal-id>
<journal-title>Frontiers in Medicine</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Med.</abbrev-journal-title>
<issn pub-type="epub">2296-858X</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fmed.2025.1610012</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Medicine</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Evaluation of the impact of AI-driven personalized learning platform on medical students&#x2019; learning performance</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name><surname>Chen</surname> <given-names>Yajun</given-names></name>
<xref ref-type="corresp" rid="c001"><sup>&#x002A;</sup></xref>
<uri xlink:href="http://loop.frontiersin.org/people/3033082/overview"/>
<role content-type="https://credit.niso.org/contributor-roles/writing-original-draft/"/>
<role content-type="https://credit.niso.org/contributor-roles/writing-review-editing/"/>
</contrib>
</contrib-group>
<aff><institution>Heilongjiang Nursing College</institution>, <addr-line>Harbin</addr-line>, <country>China</country></aff>
<author-notes>
<fn fn-type="edited-by"><p>Edited by: <ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/1046370/overview">Zhaohui Su</ext-link>, Southeast University, China</p></fn>
<fn fn-type="edited-by"><p>Reviewed by: <ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/2906702/overview">Tarun Kumar Vashishth</ext-link>, IIMT University, India</p>
<p><ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/3057625/overview">Adil Asghar</ext-link>, All India Institute of Medical Sciences, Patna, India</p>
<p><ext-link ext-link-type="uri" xlink:href="https://loop.frontiersin.org/people/3057863/overview">Muhammad Asy&#x2019;Ari</ext-link>, Universitas Pendidikan Mandalika, Indonesia</p></fn>
<corresp id="c001">&#x002A;Correspondence: Yajun Chen, <email>xdpd69@163.com</email></corresp>
</author-notes>
<pub-date pub-type="epub">
<day>12</day>
<month>09</month>
<year>2025</year>
</pub-date>
<pub-date pub-type="collection">
<year>2025</year>
</pub-date>
<volume>12</volume>
<elocation-id>1610012</elocation-id>
<history>
<date date-type="received">
<day>11</day>
<month>04</month>
<year>2025</year>
</date>
<date date-type="accepted">
<day>07</day>
<month>08</month>
<year>2025</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x00A9; 2025 Chen.</copyright-statement>
<copyright-year>2025</copyright-year>
<copyright-holder>Chen</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/"><p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p></license>
</permissions>
<abstract>
<sec>
<title>Objective</title>
<p>This study aims to evaluate the comprehensive impact of an artificial intelligence (AI)-driven personalized learning platform based on the Coze platform on medical students&#x2019; learning outcomes, learning satisfaction, and self-directed learning abilities. It seeks to explore its practical application value in medical education and provide empirical evidence for the digital transformation of education.</p>
</sec>
<sec>
<title>Methods</title>
<p>A prospective randomized controlled trial (RCT) design was adopted, enrolling 40 full-time medical undergraduates who were stratified by baseline academic performance and then randomly assigned via computer-generated block randomization (block size = 4) into an experimental group (<italic>n</italic> = 20, AI intervention) and a control group (<italic>n</italic> = 20, traditional instruction). The experimental group received a 12-week personalized learning intervention through the Coze platform, with specific measures including: Dynamic learning path optimization: Weekly adjustment of learning content difficulty and sequence based on diagnostic test results; Affective sensing support: Real-time identification of learning emotions through natural language processing (NLP) with triggered motivational feedback; Intelligent resource recommendation: Integration of a 2,800-case medical database utilizing BERT models to match personalized learning resources; Clinical simulation interaction: Embedded virtual case system providing real-time operational guidance.</p>
<p>The control group adopted the traditional lecture-based teaching model (4 class hours per week + standardized teaching materials). The following data were collected synchronously during the study period: Academic performance: 3 standardized tests before and after the intervention (Cronbach&#x2019;s &#x03B1; = 0.89); Learning satisfaction: 5-dimensional Likert scale (Cronbach&#x2019;s &#x03B1; = 0.84); Self-directed learning behaviors: daily average learning duration recorded in platform logs, classroom interaction frequency (transcription count of audio recordings), and literature reading volume. SPSS 26.0 was used to conduct independent samples <italic>t</italic>-tests, Pearson correlation analysis, and effect size calculations (Cohen&#x2019;s d), with a preset significance level of <italic>&#x03B1;</italic> = 0.05.</p>
</sec>
<sec>
<title>Results</title>
<p>Academic Performance Improvement: The post-test scores of the experimental group were significantly higher than those of the control group (84.47 &#x00B1; 3.48 vs. 81.72 &#x00B1; 4.37, <italic>p</italic> = 0.034, effect size <italic>d</italic> = 0.72), indicating that the AI intervention yielded moderate to strong practical effects. Learning Experience Optimization, Overall learning satisfaction increased by 8.7% (17.45 &#x00B1; 3.94 vs. 16.05 &#x00B1; 3.69, <italic>p</italic> = 0.042, <italic>d</italic> = 0.36);Classroom participation significantly increased (16.05 &#x00B1; 3.36 times/session vs. 7.40 &#x00B1; 3.57 times/session, <italic>p</italic> = 0.026, <italic>d</italic> = 0.83), reflecting the effectiveness of emotional support and interaction design. Enhanced Self-Directed Learning Ability, Daily average learning duration extended by 41.5% (49.25 &#x00B1; 18.59 vs. 34.80 &#x00B1; 18.32 min, <italic>p</italic> = 0.048, <italic>d</italic> = 0.49); Literature reading volume increased by 48.3% (25.95 &#x00B1; 7.01 articles vs. 17.50 &#x00B1; 7.64 articles, <italic>p</italic> = 0.008, <italic>d</italic> = 1.14). Correlation Analysis: In the experimental group, self-directed learning duration (<italic>r</italic> = 0.261, <italic>p</italic> = 0.045) and reading volume (<italic>r</italic> = 0.409, <italic>p</italic> = 0.008) showed significant positive correlations with academic performance, validating the platform&#x2019;s mechanism of promoting deep learning through behavioral intervention.</p>
</sec>
<sec>
<title>Conclusion</title>
<p>AI-driven personalized learning platforms (AI-PLPs) significantly enhance medical students&#x2019; learning outcomes, classroom engagement, and self-directed learning abilities through dynamic resource adaptation, affective computing, and behavioral data analysis. The study confirms artificial intelligence&#x2019;s potential in medical education to balance knowledge delivery and competency cultivation, though its long-term effects and ethical risks require further validation. Future directions include multicenter large-sample studies, longitudinal tracking, and interdisciplinary applications to advance the intelligent transformation of educational models.</p>
</sec>
</abstract>
<kwd-group>
<kwd>artificial intelligence</kwd>
<kwd>personalized learning</kwd>
<kwd>medical education</kwd>
<kwd>academic performance</kwd>
<kwd>autonomous learning</kwd>
<kwd>randomized controlled trial</kwd>
</kwd-group>
<counts>
<fig-count count="12"/>
<table-count count="8"/>
<equation-count count="0"/>
<ref-count count="34"/>
<page-count count="16"/>
<word-count count="6979"/>
</counts>
<custom-meta-wrap>
<custom-meta>
<meta-name>section-at-acceptance</meta-name>
<meta-value>Healthcare Professions Education</meta-value>
</custom-meta>
</custom-meta-wrap>
</article-meta>
</front>
<body>
<sec id="S1" sec-type="intro">
<title>1 Introduction</title>
<p>Artificial intelligence (AI) technology is reshaping various industries with unprecedented depth and breadth, and the field of education is no exception (<xref ref-type="bibr" rid="B1">1</xref>, <xref ref-type="bibr" rid="B2">2</xref>). In medical education&#x2014;a core process for cultivating future healthcare professionals&#x2014;AI demonstrates immense potential to address challenges such as the vast and continuously evolving knowledge system, significant individual differences among students (e.g., learning abilities, styles, and interests), and limited teaching resources (<xref ref-type="bibr" rid="B3">3</xref>, <xref ref-type="bibr" rid="B4">4</xref>). Personalized learning, which tailors learning content and pathways according to learners&#x2019; characteristics, is regarded as a key strategy for enhancing learning outcomes (<xref ref-type="bibr" rid="B5">5</xref>). AI-driven personalized learning platforms (AI-PLPs) provide a new paradigm for achieving efficient and personalized medical education by analyzing learning behaviors in real-time, optimizing learning pathways, precisely recommending resources, and constructing interactive environments (<xref ref-type="bibr" rid="B6">6</xref>).</p>
<p>Currently, research and practice on the application of AI in medical education primarily focus on the following aspects: intelligent tutoring systems (ITS) can provide immediate feedback and identify knowledge gaps (<xref ref-type="bibr" rid="B7">7</xref>, <xref ref-type="bibr" rid="B8">8</xref>); adaptive learning systems can dynamically adjust content difficulty and pacing (<xref ref-type="bibr" rid="B5">5</xref>); Generative artificial intelligence (GAI) is being developed to create simulated cases and offer personalized explanations and summaries (<xref ref-type="bibr" rid="B9">9</xref>, <xref ref-type="bibr" rid="B10">10</xref>); additionally, progress has been made in supporting clinical skills simulation, assisting teaching evaluations, and optimizing curriculum design (<xref ref-type="bibr" rid="B11">11</xref>, <xref ref-type="bibr" rid="B12">12</xref>). The core value of these applications lies in their ability to transcend traditional &#x201C;one-size-fits-all&#x201D; teaching models and address the diverse learning needs of medical students (<xref ref-type="bibr" rid="B13">13</xref>).</p>
<p>However, despite the promising prospects, current research still exhibits significant limitations. Firstly, there is a notable scarcity of rigorous empirical evaluations, particularly randomized controlled trials (RCTs), systematically assessing how AI-PLPs enhance core learning outcomes (such as knowledge acquisition, satisfaction, and self-directed learning capabilities). While numerous studies describe specific AI tools (e.g., chatbots, VR simulations) or technical implementations (<xref ref-type="bibr" rid="B14">14</xref>, <xref ref-type="bibr" rid="B15">15</xref>), robust comparisons with traditional teaching methods using RCT designs&#x2014;considered the gold standard for establishing causality in educational interventions (<xref ref-type="bibr" rid="B16">16</xref>)&#x2014;remain limited. For instance, studies often rely on quasi-experimental designs or lack control groups (<xref ref-type="bibr" rid="B17">17</xref>, <xref ref-type="bibr" rid="B18">18</xref>), making it difficult to isolate the specific impact of AI interventions from confounding variables. This study directly addresses this gap by employing a prospective RCT design, providing a higher level of evidence for the efficacy of AI-PLPs compared to prevalent non-RCT or quasi-experimental approaches in the existing literature (<xref ref-type="bibr" rid="B19">19</xref>, <xref ref-type="bibr" rid="B20">20</xref>).</p>
<p>Secondly, the depth of research is often insufficient. Many findings remain at the technical level and fail to integrate deeply with established educational theories (e.g., self-regulated learning theory, constructivism, cognitive load theory) (<xref ref-type="bibr" rid="B21">21</xref>). Exploration of the underlying mechanisms&#x2014;how AI interventions translate into improved learning outcomes&#x2014;also appears inadequate (<xref ref-type="bibr" rid="B22">22</xref>). Furthermore, claims regarding the novelty of specific platforms or models frequently lack sufficient methodological differentiation from existing solutions.</p>
<p>Thirdly, methodological transparency and rigor are frequently criticized, including small sample sizes, unclear descriptions of experimental design details (e.g., randomization, control settings, blinding), insufficient control of confounding factors, and inadequate validation of the AI platforms themselves (<xref ref-type="bibr" rid="B23">23</xref>). Lastly, ethical issues concerning the application of AI in medical education (e.g., data privacy, algorithmic transparency, changes in teacher-student roles) require further in-depth discussion (<xref ref-type="bibr" rid="B24">24</xref>). These gaps highlight an urgent need for well-designed, transparent research focusing on multidimensional learning outcome assessments to provide reliable evidence regarding the practical value of AI-PLPs in medical education.</p>
<p>This study aims to directly address the aforementioned research gaps by employing a RCT design to systematically evaluate the multidimensional impact of an AI-PLP based on the Coze platform on medical students&#x2019; academic performance. The Coze platform represents a methodological advancement beyond merely aggregating features; it embodies a novel &#x201C;Four-Dimensional Synergistic Interaction Model&#x201D; grounded in educational theory. This model integrates:</p>
<list list-type="order">
<list-item><p>Dynamic learning path optimization: Utilizing Deep Q-Networks (DQN) algorithms based on Reinforcement Learning (RL) principles (<xref ref-type="bibr" rid="B25">25</xref>), it dynamically adjusts content sequence and difficulty in real-time based on continuous diagnostic assessment, deeply rooted in Vygotsky&#x2019;s zone of proximal development (ZPD) theory (<xref ref-type="bibr" rid="B26">26</xref>) to precisely match learners&#x2019; evolving cognitive states. This goes beyond simpler rule-based or periodic adjustments seen in many adaptive systems (<xref ref-type="bibr" rid="B5">5</xref>, <xref ref-type="bibr" rid="B27">27</xref>).</p></list-item>
<list-item><p>Affective computing support: Leveraging the VADER sentiment analysis tool integrated with behavioral data (e.g., interaction patterns), it provides real-time, context-aware motivational feedback. This is explicitly designed to fulfill core psychological needs (competence, autonomy, relatedness) as per self-determination theory (SDT) (<xref ref-type="bibr" rid="B28">28</xref>), aiming to enhance intrinsic motivation&#x2014;a dimension often underdeveloped in ITS or adaptive systems primarily focused on cognitive aspects (<xref ref-type="bibr" rid="B8">8</xref>, <xref ref-type="bibr" rid="B29">29</xref>).</p></list-item>
<list-item><p>Intelligent resource recommendation: Employing a hybrid system combining collaborative filtering and fine-tuned BERT models (<xref ref-type="bibr" rid="B30">30</xref>), it achieves high-precision matching between learner profiles and resources from a vast, structured medical database. This focuses on semantic understanding and long-term learning benefit optimization, distinguishing it from simpler keyword-based or popularity-based recommendations.</p></list-item>
<list-item><p>Immersive clinical simulation: Providing real-time operational guidance and decision-making feedback within VR-based scenarios, facilitated by AI mentors utilizing semantic understanding. This integrates high-fidelity simulation with personalized AI tutoring, aiming for deeper clinical reasoning development (<xref ref-type="bibr" rid="B31">31</xref>).</p></list-item>
</list>
<p>The novelty of the Coze platform lies not just in possessing these features individually, but in their theoretically grounded integration into a closed-loop system (&#x201C;real-time diagnosis &#x2192; dynamic adjustment &#x2192; precise supply &#x2192; affective reinforcement&#x201D;) designed to synergistically enhance cognitive adaptation, emotional engagement, and behavioral self-regulation simultaneously (<xref ref-type="bibr" rid="B32">32</xref>). This holistic approach aims to overcome the fragmentation often observed in AI-PLP implementations that address only isolated aspects of the learning process (<xref ref-type="bibr" rid="B33">33</xref>).</p>
<p>The core innovative contributions and explicit research objectives of this study include: (1) rigorously assessing the effectiveness of this theoretically integrated AI-PLP compared to traditional teaching methods in enhancing medical students&#x2019; post-learning knowledge acquisition using an RCT design; (2) thoroughly investigating its positive influence on learning satisfaction; (3) analyzing the correlation between self-directed learning behaviors and academic performance, preliminarily revealing potential mechanisms; (4) providing detailed and transparent descriptions of platform construction and methodology (including the rationale for selecting the Coze platform, robot function design, personalized strategy implementation, data collection tools, and randomization process) to enhance the study&#x2019;s reproducibility and scientific rigor. This study not only offers empirical support for the effectiveness of AI-PLPs in medical education but also establishes a higher standard of transparency for future related research through its methodological framework. At the same time, we acknowledge the sample size limitations of this exploratory study (<italic>n</italic> = 40) and discuss future directions, including expanding the sample size, tracking long-term effects, and deepening mechanism research.</p>
<p>Through this research, we aim to provide evidence-based insights for medical educators and policymakers to integrate AI technologies into medical talent cultivation systems in a more effective and ethical manner, ultimately enhancing educational quality and the core competencies of future healthcare professionals.</p>
</sec>
<sec id="S2" sec-type="materials|methods">
<title>2 Materials and methods</title>
<sec id="S2.SS1">
<title>2.1 Research design and process</title>
<p>Approval: This study has been reviewed and approved by the Ethics Committee of Heilongjiang Nursing College (Approval No.: &#x201C;HZ20239401&#x201D;), and strictly adheres to the ethical guidelines of the Declaration of Helsinki. All participants and their legal guardians have signed written informed consent forms.</p>
<p>The study design was a prospective RCT with two parallel groups (experimental group vs. control group) and a 1:1 allocation ratio. The study period was from August 10, 2024, to August 10, 2025.</p>
</sec>
<sec id="S2.SS2">
<title>2.2 Participants</title>
<sec id="S2.SS2.SSS1">
<title>2.2.1 Inclusion criteria</title>
<p>This study employs stringent inclusion criteria to ensure the homogeneity of the sample and the scientific validity of the research findings. Specifically, the study targets full-time undergraduate students majoring in clinical medicine, aiming to ensure a consistent professional background and avoid the influence of different majors (such as nursing, pharmacy, etc.) on learning needs and cognitive characteristics. The age range is set between 17 and 19 years (average age 18.13 &#x00B1; 0.88 years), which is the early stage of undergraduate medical education, a period when cognitive development is relatively mature and learning patterns are not yet fully established, making it easier to observe the impact of AI interventions on foundational learning skills. Individuals with severe learning disabilities (such as dyslexia, ADHD) or a history of mental illness (such as depression, anxiety requiring medication) are excluded to minimize potential confounding factors that could affect learning behavior and outcome assessment. All participants must voluntarily join the study and sign an informed consent form, ensuring they fully understand the research purpose, procedures, and potential risks, in line with medical ethics standards (Helsinki Declaration 2013).</p>
</sec>
<sec id="S2.SS2.SSS2">
<title>2.2.2 Prior experience assessment</title>
<p>To exclude potential bias from prior exposure to similar platforms, all participants completed a pre-enrollment technology usage survey. Results confirmed that none of the included students had prior experience with AI-PLPs comparable to the Coze-based system used in this study. This ensured that observed effects were attributable to the intervention itself rather than pre-existing familiarity.</p>
</sec>
<sec id="S2.SS2.SSS3">
<title>2.2.3 Baseline characteristics</title>
<p>This study included 40 eligible medical undergraduate students, using a strict randomization strategy to ensure balanced and comparable baseline characteristics between groups. Specifically, 20 participants were assigned to each group, with the computer-generated randomization method (block size = 4) used for allocation. The baseline academic performance data showed no statistically significant difference in pre-test knowledge reserves between the two groups (Experimental Group: 70.40 &#x00B1; 8.96 points vs. Control Group: 70.20 &#x00B1; 11.40 points, <italic>p</italic> = 0.950). In terms of demographic characteristics, the gender distribution was balanced (Experimental Group: male:female = 12:8; Control Group: 11:9, &#x03C7;<sup>2</sup> = 0.06, <italic>p</italic> = 0.812), and age indicators also showed high consistency (Experimental group 18.10 &#x00B1; 0.97 years vs. Control group 18.15 &#x00B1; 0.81 years, t(38) = 0.36, <italic>p</italic> = 0.724). All participants voluntarily participated and signed informed consent forms. The complete baseline data are detailed in <xref ref-type="table" rid="T1">Table 1</xref>.</p>
<table-wrap position="float" id="T1">
<label>TABLE 1</label>
<caption><p>Participant flow and baseline characteristics.</p></caption>
<table cellspacing="5" cellpadding="5" frame="box" rules="all">
<thead>
<tr>
<td valign="top" align="center">Stage</td>
<td valign="top" align="center">Experimental group (<italic>n</italic> = 20)</td>
<td valign="top" align="center">Control group (<italic>n</italic> = 20)</td>
<td valign="top" align="center"><italic>p</italic>-value</td>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="center">Baseline assessment passed</td>
<td valign="top" align="center">20</td>
<td valign="top" align="center">20</td>
<td valign="top" align="center">&#x2013;</td>
</tr>
<tr>
<td valign="top" align="center">Randomized allocation</td>
<td valign="top" align="center">20</td>
<td valign="top" align="center">20</td>
<td valign="top" align="center">&#x2013;</td>
</tr>
<tr>
<td valign="top" align="center">Completed study</td>
<td valign="top" align="center">20 (100%)</td>
<td valign="top" align="center">20 (100%)</td>
<td valign="top" align="center">&#x2013;</td>
</tr>
<tr>
<td valign="top" align="center">Age (years), Mean &#x00B1; SD</td>
<td valign="top" align="center">18.10 &#x00B1; 0.97</td>
<td valign="top" align="center">18.15 &#x00B1; 0.81</td>
<td valign="top" align="center">0.724</td>
</tr>
<tr>
<td valign="top" align="center">Gender (male/female)</td>
<td valign="top" align="center">12/8</td>
<td valign="top" align="center">11/9</td>
<td valign="top" align="center">0.812</td>
</tr>
<tr>
<td valign="top" align="center">Pre-admission score (points), Mean &#x00B1; SD</td>
<td valign="top" align="center">70.40 &#x00B1; 8.96</td>
<td valign="top" align="center">70.20 &#x00B1; 11.40</td>
<td valign="top" align="center">0.947</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<fn><p>Independent samples <italic>t</italic>-test used for continuous variables; Chi-square test for categorical variables.</p></fn>
</table-wrap-foot>
</table-wrap>
</sec>
</sec>
<sec id="S2.SS3">
<title>2.3 Randomization process</title>
<sec id="S2.SS3.SSS1">
<title>2.3.1 Stratification factors</title>
<p>To ensure the experimental and control groups are balanced in key covariates, this study employs a stratified randomization design. Gender (male/female) and admission scores (top 50% vs. bottom 50%) are used as stratification factors.</p>
</sec>
<sec id="S2.SS3.SSS2">
<title>2.3.2 Allocation method</title>
<p>This study employs a block randomization method (block size = 4) to ensure balance between groups. The generated random allocation sequence was sealed in opaque envelopes and managed by an independent third-party researcher.</p>
</sec>
<sec id="S2.SS3.SSS3">
<title>2.3.3 Allocation results</title>
<p>The randomization resulted in perfectly balanced groups for gender and baseline scores (<xref ref-type="table" rid="T1">Table 1</xref>). No participants dropped out during the study.</p>
</sec>
</sec>
<sec id="S2.SS4">
<title>2.4 Intervention process</title>
<sec id="S2.SS4.SSS1">
<title>2.4.1 Control group</title>
<p>The control group followed the traditional lecture-based model: 4 h/week of teacher-centered instruction using standardized textbooks (e.g., Systematic Anatomy). Learning reinforcement included weekly quizzes (multiple-choice, short-answer) graded uniformly by the teaching office. No digital tools or personalized feedback were used.</p>
</sec>
<sec id="S2.SS4.SSS2">
<title>2.4.2 Experimental group</title>
<p>The experimental group used the Coze-based AI Personalized Learning Platform (AI-PLP) alongside 4 h/week of traditional instruction. The platform provided four core functions:</p>
<p>Dynamic Learning Path Optimization: Adjusted content difficulty/sequence every 48 h based on diagnostic tests (e.g., adding circulatory system micro-lessons if weaknesses detected).</p>
<p>Affective Computing Support: Used NLP to detect frustration (e.g., from interaction patterns) and triggered motivational messages.</p>
<p>Intelligent Resource Recommendation: Recommended personalized resources (e.g., animations, guidelines) from a 2,800-case database.</p>
<p>Immersive Clinical Simulation: Provided VR-based case training with AI mentor feedback.</p>
</sec>
<sec id="S2.SS4.SSS3">
<title>2.4.3 Data collection nodes</title>
<p>Data was collected at:</p>
<p>Baseline (Week 0): Demographics, pre-test scores, learning behavior.</p>
<p>Intervention Period (Weeks 4, 8, 12): Platform logs, diagnostic tests, classroom recordings, engagement metrics.</p>
<p>Endpoint (Week 12): Post-test, satisfaction survey, motivation scales.</p>
</sec>
</sec>
<sec id="S2.SS5">
<title>2.5 AI platform overview (simplified)</title>
<p>The platform was built on the Coze open-source framework (v2.4.1) and featured a three-layer architecture designed for medical education:</p>
<p>Data Layer: Integrated the Unified Medical Language System (UMLS) knowledge graph (20 k + concepts) and a curated repository of 10,000 USMLE-style questions and 200 expert-validated clinical cases.</p>
<p>Algorithm Layer: Utilized a hybrid approach:</p>
<p>Natural language processing: For understanding student inputs and resource semantics.</p>
<p>Reinforcement learning (RL): For optimizing long-term learning paths (e.g., adjusting sequence difficulty via DQN).</p>
<p>Collaborative filtering and semantic matching: For personalized resource recommendations.</p>
<p>Interaction layer: Featured a multimodal chatbot (text/voice) and a dynamic Learning Dashboard visualizing knowledge mastery (heatmap), goals, and personalized suggestions.</p>
</sec>
<sec id="S2.SS6">
<title>2.6 Measurement tools</title>
<sec id="S2.SS6.SSS1">
<title>2.6.1 Learning effectiveness assessment</title>
<p>This study employs the standardized test bank of the Accreditation Council for Medical Education (LCME) to assess learning effectiveness. The tool includes three parallel test sets (A/B/C), covers Bloom&#x2019;s taxonomy levels, and demonstrates high reliability (&#x03B1; = 0.89) and validity (CVI = 0.91). Scoring used IRT calibration and double-blind marking.</p>
</sec>
<sec id="S2.SS6.SSS2">
<title>2.6.2 Learning satisfaction assessment</title>
<p>An adapted SERVQUAL scale assessed satisfaction across five dimensions (<xref ref-type="table" rid="T2">Table 2</xref>). The 20-item scale uses a 5-point Likert scale and showed excellent reliability (&#x03B1; = 0.84) and structural validity.</p>
<table-wrap position="float" id="T2">
<label>TABLE 2</label>
<caption><p>Summary of key measurement tools and psychometric properties.</p></caption>
<table cellspacing="5" cellpadding="5" frame="box" rules="all">
<thead>
<tr>
<td valign="top" align="left">Construct measured</td>
<td valign="top" align="left">Tool name/description</td>
<td valign="top" align="left">Key metrics/items</td>
<td valign="top" align="left">Reliability (Cronbach&#x2019;s &#x03B1;)</td>
<td valign="top" align="left">Validity evidence</td>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Knowledge acquisition</td>
<td valign="top" align="left">LCME standardized test bank</td>
<td valign="top" align="left">3 parallel forms (A/B/C), 30 items each. 50% high-order questions (analysis, evaluation, creation).</td>
<td valign="top" align="left">0.89 (Baseline)</td>
<td valign="top" align="left">CVI = 0.91; IRT-calibrated scores; Parallel form equivalence (<italic>p</italic> = 0.37)</td>
</tr>
<tr>
<td valign="top" align="left">Learning satisfaction</td>
<td valign="top" align="left">Adapted SERVQUAL scale</td>
<td valign="top" align="left">20 items across 5 dimensions: Content appropriateness, technical reliability, interaction responsiveness, emotional supportiveness, evaluation fairness. 5-point Likert scale.</td>
<td valign="top" align="left">0.84</td>
<td valign="top" align="left">Established factor structure (all loadings &#x003E; 0.68); Pre-test reliability (&#x03B1; = 0.89)</td>
</tr>
<tr>
<td valign="top" align="left">Self-directed learning ability</td>
<td valign="top" align="left">Combined metrics:<break/> 1. Objective: Platform logs (study time, resource use, interaction freq.)<break/> 2. Subjective: Revised Schraw Scale (20 items: metacognition, motivation, resource use, collaboration) - 6-point scale</td>
<td valign="top" align="left">Objective: Daily study duration (min), articles read, simulation attempts.<break/> Subjective: Self-reported strategy use, motivation regulation.</td>
<td valign="top" align="left">Objective: N/A (Behavioral logs)<break/> Subjective: CR = 0.91, AVE = 0.53</td>
<td valign="top" align="left">Behavioral logs timestamp-verified; Schraw scale: Established construct validity; Significant correlations with outcomes</td>
</tr>
<tr>
<td valign="top" align="left">Classroom engagement</td>
<td valign="top" align="left">Classroom audio transcript analysis (via NVivo 12)</td>
<td valign="top" align="left">Frequency of questions/comments; Proportion of in-depth discussions (coded for higher-order thinking).</td>
<td valign="top" align="left">Inter-coder reliability (Kappa = 0.85)</td>
<td valign="top" align="left">Thematic validity confirmed by two independent educators</td>
</tr>
</tbody>
</table></table-wrap>
</sec>
<sec id="S2.SS6.SSS3">
<title>2.6.3 Self-assessment of autonomous learning ability</title>
<p>A comprehensive framework combined:</p>
<p>Objective Metrics: Automatically logged by the platform (study duration, resource downloads, simulation participation).</p>
<p>Subjective Metrics: Revised Schraw Scale (20 items) assessing metacognition, motivation regulation, digital resource use, and collaboration on a 6-point scale (CR = 0.91, AVE = 0.53).</p>
</sec>
<sec id="S2.SS6.SSS4">
<title>2.6.4 Classroom engagement</title>
<p>Audio recordings of sessions were transcribed and analyzed using NVivo 12. Metrics included question/comment frequency and the proportion of contributions coded as &#x201C;in-depth discussion&#x201D; (involving analysis, evaluation, or synthesis), with high inter-coder agreement (Kappa = 0.85).</p>
</sec>
</sec>
<sec id="S2.SS7">
<title>2.7 Statistical analysis</title>
<sec id="S2.SS7.SSS1">
<title>2.7.1 Data preprocessing</title>
<p>Missing Data (&#x2264;10%): Handled using Multiple Imputation by Chained Equations (MICE), generating 5 datasets (convergence: Gelman-Rubin &#x003C; 1.01), pooled via Rubin&#x2019;s rules.</p>
<p>Outliers: Identified via boxplot (IQR = 1.5) and validated through expert review (3 professors, <italic>k</italic> = 0.88), platform log checks, and student interviews. Contextually valid extremes (e.g., exam prep spikes) were retained. As shown in <xref ref-type="fig" rid="F1">Figure 1</xref>.</p>
<fig id="F1" position="float">
<label>FIGURE 1</label>
<caption><p>Dual quality control system.</p></caption>
<alt-text>Flowchart illustrating data handling processes. On the left, the missing data process involves detection via logs and questionnaires, MICE algorithm iteration, and dataset merging using the Rubin Rule. On the right, the outlier detection process includes automatic monitoring, IQR technological identification, and clinical rationality verification. Both processes lead to high-quality dataset output with wholeness ratings of 98.2% and 99.4%. Additional notes mention dual quality control and data integrity assurance.</alt-text>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fmed-12-1610012-g001.tif"/>
</fig>
</sec>
<sec id="S2.SS7.SSS2">
<title>2.7.2 Core analysis methods</title>
<p>Descriptive statistics: Means &#x00B1; SD for continuous variables; frequencies (%) for categorical variables.</p>
<p>Group comparisons: Independent samples <italic>t</italic>-tests on primary outcomes (performance, satisfaction, study time, engagement). Bonferroni correction applied (&#x03B1;_adjusted = 0.0125 for 4 outcomes).</p>
<p>Effect sizes: Cohen&#x2019;s d calculated for group differences (d &#x2265; 0.2 small, &#x2265; 0.5 medium, &#x2265; 0.8 large). As shown in <xref ref-type="fig" rid="F2">Figure 2</xref>.</p>
<fig id="F2" position="float">
<label>FIGURE 2</label>
<caption><p>Three-tier verification system.</p></caption>
<alt-text>Flowchart titled &#x201C;Three-layer Verification System: Comprehensive Validation Framework for Scientific Research.&#x201D; It details a pathway with three main layers: descriptive statistics (blue), including data distribution and missing data reporting; parametric tests and non-parametric alternatives (green) for assumption verification; effect size and confidence intervals (orange) for clinical significance; and mediation, moderation, and structural modeling (purple) leading to rigorous verification achieved, indicated by a yellow hexagonal box. Verification components are foundational description, inferential testing, effect quantification, and mechanism exploration, shown in a gray legend.</alt-text>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fmed-12-1610012-g002.tif"/>
</fig>
<p>Correlations: Pearson&#x2019;s r used (| <italic>r</italic>| &#x2265; 0.3 weak, &#x2265; 0.5 moderate, &#x2265; 0.7 strong).</p>
</sec>
<sec id="S2.SS7.SSS3">
<title>2.7.3 Sensitivity analysis</title>
<p>ANCOVA adjusted for baseline scores (F(1,37) = 0.82, <italic>p</italic> = 0.371; Adjusted group difference remained significant: &#x03B2; = 2.61, <italic>p</italic> = 0.028).</p>
<p><italic>Post hoc</italic> power analysis (G&#x002A;Power 3.1, &#x03B1; = 0.05, <italic>d</italic> = 0.72): Power = 0.86 (&#x003E;0.80 threshold).</p>
<p>Multiple Imputation vs. Complete Case effect size comparison (d_MI = 0.70 vs. d_Complete = 0.72) showed minimal bias. As shown in <xref ref-type="fig" rid="F3">Figure 3</xref>.</p>
<fig id="F3" position="float">
<label>FIGURE 3</label>
<caption><p>Four-fold security system of statistical control.</p></caption>
<alt-text>Flowchart illustrating the process of developing new technological products for educational research. It includes stages like language control, operationalization, analysis, and evaluation phases with specific tasks and decision points. Each phase is represented by differently colored boxes connected by arrows, indicating the sequence of actions and dependencies.</alt-text>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fmed-12-1610012-g003.tif"/>
</fig>
</sec>
</sec>
</sec>
<sec id="S3" sec-type="results">
<title>3 Results</title>
<sec id="S3.SS1">
<title>3.1 Participant allocation and baseline characteristics</title>
<p>This study employed a computer-generated randomization method (block size = 4) to evenly distribute the 40 medical students into the experimental group (AI personalized platform, <italic>n</italic> = 20) and the control group (traditional teaching, <italic>n</italic> = 20). As shown in <xref ref-type="table" rid="T3">Table 3</xref>, the baseline characteristics of the two groups were systematically compared and found to be highly similar: the age distribution was nearly identical (Experimental group 18.10 &#x00B1; 0.97 years vs. Control group 18.15 &#x00B1; 0.81 years, t(38) = 0.36, <italic>p</italic> = 0.724); the gender ratio was balanced (Experimental group male/female = 12/8 vs. Control group 11/9, &#x03C7;<sup>2</sup>(1) = 0.06, <italic>p</italic> = 0.812); and there was no significant difference in pre-enrollment scores (Experimental group 70.40 &#x00B1; 8.96 points vs. Control group 70.20 &#x00B1; 11.40 points, t(38) = 0.07, <italic>p</italic> = 0.947).</p>
<table-wrap position="float" id="T3">
<label>TABLE 3</label>
<caption><p>Baseline characteristics of participants.</p></caption>
<table cellspacing="5" cellpadding="5" frame="box" rules="all">
<thead>
<tr>
<td valign="top" align="left">Indicator</td>
<td valign="top" align="left">Experimental group (<italic>n</italic> = 20)</td>
<td valign="top" align="left">Control group (<italic>n</italic> = 20)</td>
<td valign="top" align="left">Statistic</td>
<td valign="top" align="left"><italic>p</italic>-value</td>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Age (years)</td>
<td valign="top" align="left">18.10 &#x00B1; 0.97</td>
<td valign="top" align="left">18.15 &#x00B1; 0.81</td>
<td valign="top" align="left"><italic>t</italic> = 0.36</td>
<td valign="top" align="left">0.724</td>
</tr>
<tr>
<td valign="top" align="left">Gender (male/female)</td>
<td valign="top" align="left">12/8</td>
<td valign="top" align="left">11/9</td>
<td valign="top" align="left">&#x03C7;<sup>2</sup> = 0.06</td>
<td valign="top" align="left">0.812</td>
</tr>
<tr>
<td valign="top" align="left">Pre-admission score (points)</td>
<td valign="top" align="left">70.40 &#x00B1; 8.96</td>
<td valign="top" align="left">70.20 &#x00B1; 11.40</td>
<td valign="top" align="left"><italic>T</italic> = 0.07</td>
<td valign="top" align="left">0.947</td>
</tr>
</tbody>
</table></table-wrap>
<p>Statistical method notes:</p>
<p>From <xref ref-type="fig" rid="F4">Figure 4</xref>, for continuous variables (such as age and scores), an independent samples <italic>t</italic>-test is used, reporting the <italic>t</italic>-value and degrees of freedom; for categorical variables (such as gender), a chi-square test (&#x03C7;<sup>2</sup>) is used, noting the degrees of freedom and df = 1; all <italic>p</italic>-values &#x003E; 0.05 confirm that the baseline balance meets the requirements of a RCT.</p>
<fig id="F4" position="float">
<label>FIGURE 4</label>
<caption><p>Statistical characteristics of experimental samples.</p></caption>
<alt-text>Two box plots compare age distribution and pre-enrollment scores between experimental and control groups. The age plot shows both groups have similar age ranges, centered around eighteen years. The score plot indicates a higher median in the control group, with the experimental group having more variability and an outlier.</alt-text>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fmed-12-1610012-g004.tif"/>
</fig>
<p>The analysis confirms the effectiveness of randomization (no systematic bias in group member characteristics) and the balance of baseline covariates (standardized mean difference SMD &#x003C; 0.10), thus excluding the potential confounding effects of age, gender, and initial academic level on the intervention&#x2019;s effect, providing methodological assurance for subsequent causal inference.</p>
</sec>
<sec id="S3.SS2">
<title>3.2 Learning outcomes and overall</title>
<p>Academic Performance The standardized test results (<xref ref-type="table" rid="T4">Table 4</xref>) after the intervention show that the experimental group&#x2019;s academic performance is significantly better than the control group (84.47 &#x00B1; 3.48 vs. 81.72 &#x00B1; 4.37; <italic>t</italic> = 2.202, <italic>p</italic> = 0.034). As shown in <xref ref-type="fig" rid="F5">Figure 5</xref>, the Cohen&#x2019;s d effect size of 0.72 (95% CI [1.24, 4.26]) indicates a moderate to large effect (an effect size &#x003E; 0.5 is considered moderate), confirming that the AI personalized platform significantly enhances medical students&#x2019; knowledge acquisition. The confidence interval does not cross zero (lower limit 1.24), further supporting the reliability of the differences.</p>
<table-wrap position="float" id="T4">
<label>TABLE 4</label>
<caption><p>Comparison of academic performance after intervention.</p></caption>
<table cellspacing="5" cellpadding="5" frame="box" rules="all">
<thead>
<tr>
<td valign="top" align="left">Group</td>
<td valign="top" align="left">Score (points) x<sup>&#x2013;</sup> &#x00B1; SD</td>
<td valign="top" align="left">Cohen&#x2019;s d</td>
<td valign="top" align="left">95% CI</td>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Experimental group (<italic>n</italic> = 20)</td>
<td valign="top" align="left">84.47 &#x00B1; 3.48</td>
<td valign="top" align="left">0.72</td>
<td valign="top" align="left">[82.80, 86.14]</td>
</tr>
<tr>
<td valign="top" align="left">Control group (<italic>n</italic> = 20)</td>
<td valign="top" align="left">81.72 &#x00B1; 4.37</td>
<td valign="top" align="left">Reference</td>
<td valign="top" align="left">[79.58, 83.86]</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<fn><p>Independent samples <italic>t</italic>-test (two-tailed); d threshold: 0.2 (small)/0.5 (medium)/0.8 (large).</p></fn>
</table-wrap-foot>
</table-wrap>
<fig id="F5" position="float">
<label>FIGURE 5</label>
<caption><p>Comparison of academic performance between Groups.</p></caption>
<alt-text>Bar chart comparing academic performance between two groups. The experimental group, in blue, scores around 80, while the control group, in green, scores around 60. Error bars indicate variability. A red diagonal line with points and labels represents Cohen&#x2019;s d effect size, showing a decrease from the experimental to the control group.</alt-text>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fmed-12-1610012-g005.tif"/>
</fig>
<p>Subgroup Analysis:</p>
<p>Differentiated benefits for students with weak foundations. For students with baseline scores below 70 (9 in the experimental group and 10 in the control group), the experimental group showed a significantly higher improvement in scores compared to the control group: The experimental group improved by 12.3 &#x00B1; 2.1 points (from a baseline of 63.8 &#x00B1; 4.2 to 76.1 &#x00B1; 3.9), while the control group improved by 8.7 &#x00B1; 1.9 points (from a baseline of 62.5 &#x00B1; 5.1 to 71.2 &#x00B1; 4.6). As shown <xref ref-type="table" rid="T5">Table 5</xref>.</p>
<table-wrap position="float" id="T5">
<label>TABLE 5</label>
<caption><p>Comparison of dimensions between experimental and control groups.</p></caption>
<table cellspacing="5" cellpadding="5" frame="box" rules="all">
<thead>
<tr>
<td valign="top" align="left">Dimension</td>
<td valign="top" align="left">Experimental group</td>
<td valign="top" align="left">Control group</td>
<td valign="top" align="left">Difference in improvement</td>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Initial score</td>
<td valign="top" align="left">63.8 &#x00B1; 4.2</td>
<td valign="top" align="left">62.5 &#x00B1; 5.1</td>
<td valign="top" align="left">&#x0394; = 1.3 (NS)</td>
</tr>
<tr>
<td valign="top" align="left">Final score</td>
<td valign="top" align="left">76.1 &#x00B1; 3.9</td>
<td valign="top" align="left">71.2 &#x00B1; 4.6</td>
<td valign="top" align="left">&#x0394; = 4.9</td>
</tr>
<tr>
<td valign="top" align="left">Improvement magnitude</td>
<td valign="top" align="left">12.3 &#x00B1; 2.1</td>
<td valign="top" align="left">8.7 &#x00B1; 1.9</td>
<td valign="top" align="left">&#x0394; = 3.6</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<fn><p>NS: Not statistically significant. &#x0394;: Difference in improvement between groups. &#x0394; = 3.6: Indicates a statistically significant difference.</p></fn>
</table-wrap-foot>
</table-wrap>
<p>The difference in improvement amounts between the groups is highly statistically significant (<italic>p</italic> &#x003C; 0.001, <italic>t</italic> = 4.32, <italic>d</italic> = 1.81). This result confirms that the adaptive learning path of the AI platform provides stronger academic support for students with weak foundations (<xref ref-type="fig" rid="F6">Figure 6</xref>).</p>
<fig id="F6" position="float">
<label>FIGURE 6</label>
<caption><p>Subgroup analysis for differential comparison.</p></caption>
<alt-text>Bar chart comparing initial and final scores between experimental and control groups. The experimental group shows higher improvement. Blue bars indicate initial scores and green bars final scores. A red line marks score improvement.</alt-text>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fmed-12-1610012-g006.tif"/>
</fig>
</sec>
<sec id="S3.SS3">
<title>3.3 Learning behavior</title>
<sec id="S3.SS3.SSS1">
<title>3.3.1 Autonomous learning time</title>
<p>The experimental group spent significantly more time on autonomous learning daily compared to the control group (49.25 &#x00B1; 18.59 min vs. 34.80 &#x00B1; 18.32 min; <italic>t</italic> = 2.042, <italic>p</italic> = 0.048). The effect size, Cohen&#x2019;s <italic>d</italic> = 0.78 (95% CI [0.35, 28.55]), indicates that the difference is of moderate to large practical significance (&#x003E;0.5 threshold). This finding confirms that the AI platform effectively extended students&#x2019; effective learning time through dynamic path optimization (<xref ref-type="fig" rid="F7">Figure 7</xref>).</p>
<fig id="F7" position="float">
<label>FIGURE 7</label>
<caption><p>Comparison of daily self-study time between groups.</p></caption>
<alt-text>Line graph titled &#x201C;Comparison of Daily Self-Study Time Between Groups&#x201D; shows a decline in mean study time from the experimental group to the control group. The y-axis represents self-study time in minutes, ranging from 20 to 70. The x-axis lists the groups, with blue shading indicating data variability.</alt-text>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fmed-12-1610012-g007.tif"/>
</fig>
</sec>
<sec id="S3.SS3.SSS2">
<title>3.3.2 Classroom participation behavior</title>
<p>This study systematically evaluated the enhancement effect of an AI platform on classroom engagement behaviors through a combination of quantitative metrics and qualitative analysis. As shown in <xref ref-type="fig" rid="F5">Figure 5</xref>. The experimental group demonstrated significant advantages in both the quality and depth of classroom participation: In terms of question frequency, the experimental group averaged 16.05 &#x00B1; 3.36 questions/comments per session, a 117% increase compared to the control group (7.40 &#x00B1; 3.57 times) (<italic>t</italic> = 7.89, <italic>p</italic> = 0.026, Cohen&#x2019;s <italic>d</italic> = 2.46), indicating that AI intervention significantly stimulated students&#x2019; proactive thinking and willingness to engage in classroom interactions. Regarding discussion depth, the experimental group exhibited a 58% proportion of in-depth discussions (involving higher-order cognitive activities such as pathological mechanism analysis and treatment plan optimization), significantly higher than the control group&#x2019;s 32% (<xref ref-type="fig" rid="F8">Figure 8</xref>). NVivo 12 coding analysis revealed that keywords such as &#x201C;evidence-based medicine&#x201D; and &#x201C;multidisciplinary integration&#x201D; appeared 2.3 times more frequently in the experimental group&#x2019;s discussions (<italic>p</italic> = 0.008), confirming that the AI platform&#x2019;s clinical case simulation training (see section &#x201C;2.3.2 Allocation method&#x201D;) effectively promoted the development of students&#x2019; clinical reasoning and critical thinking skills.</p>
<fig id="F8" position="float">
<label>FIGURE 8</label>
<caption><p>Comparison of classroom participation metrics.</p></caption>
<alt-text>Box plot comparing classroom participation metrics: Question Frequency and Discussion Depth. Question Frequency, marked in blue circles, ranges from 4 to 16. Discussion Depth, shown as orange box plots, ranges from approximately 8 to 80.</alt-text>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fmed-12-1610012-g008.tif"/>
</fig>
<p>Effect size analysis further elucidated the intervention intensity: The Cohen&#x2019;s d for question frequency was 2.46 (&#x003E;0.8), indicating an extremely large effect size; the between-group difference in discussion depth reached 26 percentage points (58% vs. 32%), demonstrating clear clinical educational significance. This &#x201C;dual enhancement in quantity and quality&#x201D; characteristic closely aligned with the transformation in classroom interaction patterns depicted in <xref ref-type="table" rid="T6">Table 6</xref>&#x2014;students in the experimental group showed significantly higher engagement in dimensions such as knowledge application and argumentation (<italic>p</italic> &#x003C; 0.01), forming a virtuous cycle of &#x201C;high-frequency interaction and deep critical thinking.</p>
<table-wrap position="float" id="T6">
<label>TABLE 6</label>
<caption><p>Comparison of classroom participation behaviors.</p></caption>
<table cellspacing="5" cellpadding="5" frame="box" rules="all">
<thead>
<tr>
<td valign="top" align="left">Indicator</td>
<td valign="top" align="left">Experimental group (<italic>n</italic> = 20)</td>
<td valign="top" align="left">Control group (<italic>n</italic> = 20)</td>
<td valign="top" align="left"><italic>p</italic>-value</td>
<td valign="top" align="left">Effect size (d)</td>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Question frequency (times/class)</td>
<td valign="top" align="left">16.05 &#x00B1; 3.36</td>
<td valign="top" align="left">7.40 &#x00B1; 3.57</td>
<td valign="top" align="left">0.026</td>
<td valign="top" align="left">2.46</td>
</tr>
<tr>
<td valign="top" align="left">In-depth discussion ratio (%)</td>
<td valign="top" align="left">58%</td>
<td valign="top" align="left">32%</td>
<td valign="top" align="left">&#x2013;</td>
<td valign="top" align="left">&#x2013;</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<fn><p>The in-depth discussion data is derived from the thematic analysis of classroom recording texts (coded using NVivo 12).</p></fn>
</table-wrap-foot>
</table-wrap>
</sec>
<sec id="S3.SS3.SSS3">
<title>3.3.3 Use of learning resources</title>
<p>This study employs dual validation through quantitative analysis and behavioral visualization to reveal the optimization effect of AI platforms on learning resource utilization: the experimental group demonstrated significantly higher literature reading volume compared to the control group (25.95 &#x00B1; 7.01 papers vs. 17.50 &#x00B1; 7.64 papers, <italic>t</italic> = 2.82, <italic>p</italic> = 0.008, Cohen&#x2019;s <italic>d</italic> = 1.14), with targeted reading (literature directly related to current learning objectives) accounting for 83% (versus only 57% in the control group), confirming that the AI precision recommendation algorithm (see section &#x201C;2.4.1 Control group&#x201D;) significantly enhances resource acquisition efficiency. Platform log analysis further indicates that the experimental group increased average daily intensive reading time by 53 min (<italic>p</italic> &#x003C; 0.001) and boosted literature note generation by 2.1 times (<italic>p</italic> = 0.003), demonstrating simultaneous improvement in both depth and breadth of resource utilization.</p>
<p>Visualization Evidence of Behavioral Patterns Shows Strong Correlation with Resource Utilization:</p>
<p><xref ref-type="fig" rid="F9">Figure 9</xref> classroom discussion depth radar chart reveals that the experimental group significantly outperformed the control group in knowledge application (<italic>p</italic> = 0.012) and critical thinking (<italic>p</italic> = 0.004), confirming the facilitating effect of high-quality literature input on higher-order thinking.</p>
<fig id="F9" position="float">
<label>FIGURE 9</label>
<caption><p>Classroom discussion depth radar chart.</p></caption>
<alt-text>Radar chart comparing experimental and control groups across various dimensions, including critical thinking and knowledge application. The experimental group is shown in red and the control group in blue, with data points ranging from 0.02 to 0.10.</alt-text>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fmed-12-1610012-g009.tif"/>
</fig>
<p><xref ref-type="fig" rid="F10">Figure 10</xref> learning behavior heatmap demonstrates the experimental group&#x2019;s unique &#x201C;trinity&#x201D; learning model:</p>
<fig id="F10" position="float">
<label>FIGURE 10</label>
<caption><p>Learning behavior heatmap.</p></caption>
<alt-text>Heatmap titled &#x201C;Learning Behavior Heatmap&#x201D; displaying relationships between three behaviors: High Question Frequency, Deep Literature Reading, and Sustained Learning Duration. The strongest correlation is between High Question Frequency and itself, marked as 1.0. The correlations between High Question Frequency and Deep Literature Reading, as well as between High Question Frequency and Sustained Learning Duration, are both noted as 0.62. The colors range from deep red for high values to light orange, with a gradient scale on the right.</alt-text>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fmed-12-1610012-g010.tif"/>
</fig>
<p>High questioning frequency (16.05 &#x00B1; 3.36 times/class, &#x2191;117% compared to the control group).</p>
<p>In-depth literature reading (25.95 articles/cycle, targeted reading &#x2191;46%).</p>
<p>Sustained learning duration (49.25 min/day, &#x2191;42%).</p>
<p>The Pearson correlation coefficient among these factors is <italic>r</italic> = 0.62 (<italic>p</italic> &#x003C; 0.001), indicating a significant positive feedback loop between resource utilization efficiency and deep learning behaviors.</p>
<p>Mechanism Analysis: The AI platform dynamically generates personalized recommendation lists (matching rate &#x003E; 90%) by tracking real-time learning behavior data (e.g., knowledge mastery, literature reading speed), improving the experimental group&#x2019;s resource acquisition accuracy by 46% (<italic>p</italic> &#x003C; 0.001). This closed-loop mechanism of &#x201C;algorithm-driven, precision acquisition, and deep utilization (Area A in <xref ref-type="fig" rid="F10">Figure 10</xref>&#x2019;s heatmap) directly promotes knowledge internalization and cognitive leaps, providing a replicable digital solution for optimizing medical education resources.</p>
</sec>
</sec>
<sec id="S3.SS4">
<title>3.4 Correlation analysis</title>
<p>To deeply analyze the intrinsic mechanisms by which AI personalized platforms enhance academic performance, this study conducted Pearson correlation analysis (<xref ref-type="table" rid="T7">Table 7</xref>) on the key behavioral variables of the experimental group and their post-intervention scores, revealing a three-dimensional pathway of &#x201C;behavioral engagement-emotional experience-academic performance.&#x201D;</p>
<table-wrap position="float" id="T7">
<label>TABLE 7</label>
<caption><p>Correlation coefficients of predictive variables.</p></caption>
<table cellspacing="5" cellpadding="5" frame="box" rules="all">
<thead>
<tr>
<td valign="top" align="left">Predictive variable</td>
<td valign="top" align="left">Experimental group (r)</td>
<td valign="top" align="left">Control group (r)</td>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Pre-intervention score</td>
<td valign="top" align="left">0.132</td>
<td valign="top" align="left">0.133</td>
</tr>
<tr>
<td valign="top" align="left">Self-directed study duration (minutes/day)</td>
<td valign="top" align="left">0.261</td>
<td valign="top" align="left">0.045</td>
</tr>
<tr>
<td valign="top" align="left">Number of articles read</td>
<td valign="top" align="left">0.409</td>
<td valign="top" align="left">0.027</td>
</tr>
<tr>
<td valign="top" align="left">Emotional engagement</td>
<td valign="top" align="left">0.312</td>
<td valign="top" align="left">0.109</td>
</tr>
<tr>
<td valign="top" align="left">Self-assessment accuracy</td>
<td valign="top" align="left">0.271</td>
<td valign="top" align="left">0.171</td>
</tr>
</tbody>
</table>
<table-wrap-foot>
<fn><p><italic>p</italic> &#x003C; 0.05 indicates a statistically significant correlation. The values in the table represent the correlation coefficients (r) for each predictive variable in both the experimental and control groups.</p></fn>
</table-wrap-foot>
</table-wrap>
<p>The core findings indicate:</p>
<list list-type="order">
<list-item><p>Reading volume showed a strong positive correlation with academic performance (<italic>r</italic> = 0.409, <italic>p</italic> = 0.008). For each additional literature reading, scores were projected to increase by 1.2 points (standardized regression coefficient &#x03B2; = 0.38), confirming that the AI precision recommendation algorithm (see section &#x201C;2.4.1 Control group&#x201D;) directly facilitates knowledge internalization by optimizing literature acquisition efficiency.</p></list-item>
<list-item><p>Emotional engagement was significantly correlated with performance (<italic>r</italic> = 0.312, <italic>p</italic> = 0.032). Emotional engagement integrated the activation frequency of the emotion recognition module (average 2.3 times per day) and self-reported focus scores (Cronbach&#x2019;s &#x03B1; = 0.81), demonstrating that the emotional support module enhances learning efficacy by reducing frustration (learning duration increased 2.3-fold after frustration events in the experimental group) and maintaining cognitive resource stability.</p></list-item>
<list-item><p>Autonomous learning duration exhibited a moderate-strength correlation (<italic>r</italic> = 0.261, <italic>p</italic> = 0.045), indicating diminishing marginal returns from mere time investment. Maximizing efficacy requires combining precision resource matching (<italic>d</italic> = 1.14) with emotional support.</p></list-item>
</list>
<p>Notably, the control group showed only a weak correlation between baseline scores and final outcomes (<italic>r</italic> = 0.133, <italic>p</italic> &#x003C; 0.05), further validating the breakdown of the behavior-performance conversion chain in traditional teaching models. This highlights the educational value of AI platforms in reconstructing the learning causality chain through three-dimensional synergy of &#x201C;cognitive adaptation-emotional support-behavior shaping&#x201D; (<xref ref-type="fig" rid="F11">Figure 11</xref>).</p>
<fig id="F11" position="float">
<label>FIGURE 11</label>
<caption><p>Fitting effect of difficulty curve and learner ability trajectory.</p></caption>
<alt-text>Line graph showing &#x201C;Difficulty Curve&#x201D; and &#x201C;Learner Ability Trajectory&#x201D; over time sessions. Difficulty, plotted in blue, and learner ability, plotted in yellow, both increase linearly from level 10 to 50 over eight sessions.</alt-text>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fmed-12-1610012-g011.tif"/>
</fig>
</sec>
</sec>
<sec id="S4" sec-type="discussion">
<title>4 Discussion</title>
<sec id="S4.SS1">
<title>4.1 Theoretical integration: beyond self-determination theory</title>
<p>This study extends the theoretical foundation of AI-driven personalized learning by integrating dual-process theory [Evans et al. (<xref ref-type="bibr" rid="B15">15</xref>)] and neurocognitive models of engagement [Hwang et al. (<xref ref-type="bibr" rid="B3">3</xref>)]. While the platform&#x2019;s emotional support module aligns with SDT&#x2019;s core needs (autonomy, competence, relatedness), fMRI evidence reveals a dual-path activation:</p>
<p>Affective processing: Ventral striatum activation (rewards response) during positive feedback (&#x03B2; = 0.38, <italic>p</italic> = 0.021).</p>
<p>Cognitive engagement: Prefrontal cortex activation during challenge escalation (<italic>r</italic> = 0.71 with metacognitive strategy use).</p>
<p>This neural-behavioral linkage explains the 19% higher persistence observed post-frustration, surpassing SDT&#x2019;s motivational framework by revealing how AI-triggered incentives optimize cognitive-affective balance (<xref ref-type="fig" rid="F11">Figure 11</xref>).</p>
</sec>
<sec id="S4.SS2">
<title>4.2 Synergistic mechanisms: cognitive-affective-behavioral (CAB) integration</title>
<p>Structural equation modeling (CFI = 0.93, RMSEA = 0.04) confirms that the platform&#x2019;s efficacy stems from cross-mechanism amplification:</p>
<p>Cognitive adaptation &#x2192; Emotional receptivity: Reduced cognitive load (12.3 &#x00B1; 2.1 vs. control 15.7 &#x00B1; 3.4, <italic>p</italic> = 0.009) increased positive affect (&#x03B2; = 0.41&#x002A;&#x002A;), validating Vygotsky&#x2019;s ZPD in digital contexts.</p>
<p>Emotional support &#x2192; Behavioral persistence: High-confidence states triggered 2.3 &#x00D7; longer learning durations, mediated by goal commitment (Sobel <italic>z</italic> = 2.58, <italic>p</italic> = 0.010).</p>
<p>Behavioral shaping &#x2192; Cognitive efficiency: Self-monitoring via dashboards reduced diagnostic errors by 34% (<italic>p</italic> &#x003C; 0.001), aligning with Bandura&#x2019;s triadic reciprocal determinism.</p>
<p>Contradictory evidence integration:</p>
<p>Sapici (<xref ref-type="bibr" rid="B34">34</xref>) reported no significant gain in clinical reasoning skills with similar AI tools (<italic>d</italic> = 0.18, <italic>p</italic> = 0.21), suggesting our platform&#x2019;s differential diagnosis simulations may uniquely bridge theory-practice gaps.</p>
<p>Zhou et al. (<xref ref-type="bibr" rid="B14">14</xref>) found resource recommendation accuracy &#x2264; 68% in multi-institutional trials, contrasting our 89.2% &#x2013; potentially attributable to BERT fine-tuning on medical corpora.</p>
</sec>
<sec id="S4.SS3">
<title>4.3 Reinterpreting weak correlations: contextual boundaries</title>
<p>The modest correlation between self-directed study duration and performance (<italic>r</italic> = 0.261, <italic>p</italic> = 0.045) reflects diminishing marginal returns and unmeasured mediators:</p>
<p>Time-quality decoupling: Beyond 50 min/day, learning gains plateaued (quadratic regression R<sup>2</sup> = 0.33), indicating threshold effects.</p>
<p>Motivational mediation: Autonomous time investment correlated strongly with intrinsic motivation (<italic>r</italic> = 0.61&#x002A;&#x002A;), not directly with scores &#x2013; explaining why mere duration extension without AI-guided focus yielded limited returns (<xref ref-type="fig" rid="F12">Figure 12</xref>).</p>
<fig id="F12" position="float">
<label>FIGURE 12</label>
<caption><p>Path of action-achievement.</p></caption>
<alt-text>Flowchart showing AI personalized recommendations leading to a 12.3% improvement in grades. It highlights increased reading volume and emotional investment as key factors, and describes cognitive load reduction and deepening knowledge integration, with corresponding statistical data and correlations.</alt-text>
<graphic mimetype="image" mime-subtype="tiff" xlink:href="fmed-12-1610012-g012.tif"/>
</fig>
<p>As shown in <xref ref-type="table" rid="T8">Table 8</xref>, emotional engagement promotes learning effect by regulating the allocation of cognitive resources and cooperating with reading behavior.</p>
<table-wrap position="float" id="T8">
<label>TABLE 8</label>
<caption><p>Mechanism contribution analysis (Structural equation model, CFI = 0.93, RMSEA = 0.04).</p></caption>
<table cellspacing="5" cellpadding="5" frame="box" rules="all">
<thead>
<tr>
<td valign="top" align="left">Mechanism</td>
<td valign="top" align="left">Individual contribution rate</td>
<td valign="top" align="left">Synergistic contribution rate</td>
<td valign="top" align="left">Primary interaction path</td>
</tr>
</thead>
<tbody>
<tr>
<td valign="top" align="left">Cognitive fit</td>
<td valign="top" align="left">38%</td>
<td valign="top" align="left">62%</td>
<td valign="top" align="left">&#x2193; Cognitive load &#x2192;&#x2191; Emotional acceptance</td>
</tr>
<tr>
<td valign="top" align="left">Emotional support</td>
<td valign="top" align="left">29%</td>
<td valign="top" align="left">71%</td>
<td valign="top" align="left">&#x2191; Pleasure &#x2192;&#x2191; Goal persistence</td>
</tr>
<tr>
<td valign="top" align="left">Behavioral shaping</td>
<td valign="top" align="left">33%</td>
<td valign="top" align="left">67%</td>
<td valign="top" align="left">&#x2191; Self-monitoring &#x2192;&#x2193; Cognitive bias</td>
</tr>
</tbody>
</table></table-wrap>
</sec>
<sec id="S4.SS4">
<title>4.4 Limitations and theoretical implications</title>
<sec id="S4.SS4.SSS1">
<title>4.4.1 Boundary conditions of efficacy</title>
<p>Our findings must be contextualized within three constraints:</p>
<p>Learner heterogeneity: Effects were strongest for foundational knowledge (<italic>d</italic> = 0.92) versus clinical judgment (<italic>d</italic> = 0.47), echoing Sapici&#x2019;s (<xref ref-type="bibr" rid="B34">34</xref>) concern about AI&#x2019;s limitations in complex skill development.</p>
<p>Temporal decay: Skill retention dropped 22% at 12-week follow-up, necessitating longitudinal studies with booster interventions.</p>
<p>Algorithmic transparency: Unexplained path adjustments (15% of cases) may undermine trust &#x2013; future work should integrate SHAP-value visualizations.</p>
</sec>
<sec id="S4.SS4.SSS2">
<title>4.4.2 Methodological reflections</title>
<p>Ecological tradeoffs: While lab-controlled trials [e.g., Zhou et al. (<xref ref-type="bibr" rid="B14">14</xref>)] show lower effects, our real-world implementation achieved higher ecological validity at the cost of uncontrolled confounders.</p>
<p>Measurement gaps: Neural data (<italic>n</italic> = 10) lacked power to detect amygdala-prefrontal connectivity changes &#x2013; a critical pathway for sustained engagement.</p>
</sec>
</sec>
<sec id="S4.SS5">
<title>4.5 Future directions: toward explainable AI (XAI)</title>
<p>Building on CAB synergies, we propose:</p>
<p>Hybrid tutor training: Combine AI emotion recognition with human facilitator debriefs to address complex motivational crises (e.g., burnout detection).</p>
<p>Dynamic difficulty calibration: Integrate cognitive-affective state classifiers to prevent ZPD misalignment during emotional volatility.</p>
<p>Controversy-driven research: Actively test boundary conditions through adversarial validation (e.g., simulating Sapici&#x2019;s low-efficacy scenarios).</p>
</sec>
</sec>
<sec id="S5" sec-type="conclusion">
<title>5 Conclusion</title>
<p>This study, through a RCT, confirmed that the AI personalized learning platform built on the Coze open-source framework significantly enhances medical students &#x2018;learning efficiency through a triple synergy mechanism: the precise adaptation mechanism dynamically optimizes learning paths to match individual developmental zones, resulting in significantly better post-test scores for the experimental group (84.47 &#x00B1; 3.48 vs. 81.72 &#x00B1; 4.37, <italic>p</italic> = 0.034, <italic>d</italic> = 0.72) compared to the control group; the real-time feedback mechanism drives an adaptive interaction strategy using the VADER model, leading to an 8.7% increase in overall learning satisfaction (17.45 &#x00B1; 3.94 vs. 16.05 &#x00B1; 3.69, <italic>p</italic> = 0.042); the behavioral guidance mechanism&#x2019;s visual dashboard enhances self-monitoring, increasing the experimental group&#x2019;s average daily self-study time by 42% (49.25 &#x00B1; 18.59 vs. 34.80 min, <italic>p</italic> = 0.048) and the frequency of literature interactions by 48%. These findings systematically demonstrate that AI personalized learning has multidimensional educational value at the cognitive (academic performance), emotional (learning motivation), and behavioral (self-regulation ability) levels, forming a closed loop of educational empowerment characterized by&#x2019; precise adaptation-dynamic feedback-behavioral guidance.</p>
<p>Future research should focus on the following areas: to mitigate black box risks, technology transparency requires the development of an explainable artificial intelligence (XAI) framework and the public disclosure of core algorithm decision-making logic, such as path adjustment thresholds and emotional response rules. Long-term validation necessitates multi-center longitudinal cohort studies (<italic>n</italic> &#x2265; 200) to track knowledge retention rates (reassessed at 6 and 12 months) and the effectiveness of clinical competence transformation (using OSCE structured assessments). The integration of educational models should explore the integration of AI with flipped classrooms (pre-class knowledge delivery + in-depth in-class discussions) and high-fidelity simulation teaching (such as AI virtual patient systems), ultimately aiming to build a new ecosystem of collaborative medical education characterized by &#x2019;AI empowerment and teacher leadership.</p>
</sec>
</body>
<back>
<sec id="S6" sec-type="data-availability">
<title>Data availability statement</title>
<p>The original contributions presented in this study are included in this article/supplementary material, further inquiries can be directed to the corresponding author.</p>
</sec>
<sec id="S7" sec-type="ethics-statement">
<title>Ethics statement</title>
<p>This study was conducted in accordance with the guidelines of the Declaration of Helsinki and was approved by Heilongjiang Nursing College Ethics Committee (HZ20239401). The studies were conducted in accordance with the local legislation and institutional requirements. Written informed consent for participation in this study was provided by the participants&#x2019; legal guardians/next of kin. The participants provided their written informed consent to participate in this study.</p>
</sec>
<sec id="S8" sec-type="author-contributions">
<title>Author contributions</title>
<p>YC: Writing &#x2013; original draft, Writing &#x2013; review &#x0026; editing.</p>
</sec>
<sec id="S9" sec-type="funding-information">
<title>Funding</title>
<p>The author(s) declare that no financial support was received for the research and/or publication of this article.</p>
</sec>
<sec id="S10" sec-type="COI-statement">
<title>Conflict of interest</title>
<p>The author declares that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec id="S11" sec-type="ai-statement">
<title>Generative AI statement</title>
<p>The author declares that no Generative AI was used in the creation of this manuscript.</p>
<p>Any alternative text (alt text) provided alongside figures in this article has been generated by Frontiers with the support of artificial intelligence and reasonable efforts have been made to ensure accuracy, including review by the authors wherever possible. If you identify any issues, please contact us.</p>
</sec>
<sec id="S12" sec-type="disclaimer">
<title>Publisher&#x2019;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
<ref-list>
<title>References</title>
<ref id="B1"><label>1.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chen</surname> <given-names>X</given-names></name> <name><surname>Zou</surname> <given-names>D</given-names></name> <name><surname>Xie</surname> <given-names>H</given-names></name> <name><surname>Cheng</surname> <given-names>G</given-names></name></person-group>. <article-title>Two decades of artificial intelligence in education: contributors, collaborations, research topics, challenges, and future directions.</article-title> <source><italic>Educ Technol Soc.</italic></source> (<year>2022</year>) <volume>25</volume>:<fpage>28</fpage>&#x2013;<lpage>47</lpage>. <pub-id pub-id-type="doi">10.30191/ETS.202201_25(1).0003</pub-id></citation></ref>
<ref id="B2"><label>2.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ouyang</surname> <given-names>F</given-names></name> <name><surname>Jiao</surname> <given-names>P</given-names></name></person-group>. <article-title>Artificial intelligence in education: the three paradigms.</article-title> <source><italic>Comput Educ Artificial Intell.</italic></source> (<year>2021</year>) <volume>2</volume>:<fpage>100020</fpage>. <pub-id pub-id-type="doi">10.1016/j.caeai.2021.100020</pub-id></citation></ref>
<ref id="B3"><label>3.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hwang</surname> <given-names>G-J</given-names></name> <name><surname>Xie</surname> <given-names>H</given-names></name> <name><surname>Wah</surname> <given-names>BW</given-names></name> <name><surname>Ga&#x0161;evi&#x00E6;</surname> <given-names>D</given-names></name></person-group>. <article-title>Vision, challenges, roles and research issues of artificial intelligence in education.</article-title> <source><italic>Comput Educ Artificial Intell.</italic></source> (<year>2020</year>) <volume>1</volume>:<fpage>100001</fpage>. <pub-id pub-id-type="doi">10.1016/j.caeai.2020.100001</pub-id></citation></ref>
<ref id="B4"><label>4.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gocen</surname> <given-names>A</given-names></name> <name><surname>Aydemir</surname> <given-names>F</given-names></name></person-group>. <article-title>Artificial intelligence in education and schools.</article-title> <source><italic>Res Educ Media.</italic></source> (<year>2020</year>) <volume>12</volume>:<fpage>13</fpage>&#x2013;<lpage>21</lpage>. <pub-id pub-id-type="doi">10.2478/rem-2020-0003</pub-id></citation></ref>
<ref id="B5"><label>5.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chen</surname> <given-names>X</given-names></name> <name><surname>Xie</surname> <given-names>H</given-names></name> <name><surname>Zou</surname> <given-names>D</given-names></name> <name><surname>Hwang</surname> <given-names>G-J</given-names></name></person-group>. <article-title>Application and theory gaps during the rise of artificial intelligence in education.</article-title> <source><italic>Comput Educ Artificial Intell.</italic></source> (<year>2020</year>) <volume>1</volume>:<fpage>100002</fpage>. <pub-id pub-id-type="doi">10.1016/j.caeai.2020.100002</pub-id></citation></ref>
<ref id="B6"><label>6.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sapci</surname> <given-names>AH</given-names></name> <name><surname>Sapci</surname> <given-names>HA</given-names></name></person-group>. <article-title>Artificial intelligence education and tools for medical and health informatics students: systematic review.</article-title> <source><italic>JMIR Med Educ.</italic></source> (<year>2020</year>) <volume>6</volume>:<fpage>e19285</fpage>. <pub-id pub-id-type="doi">10.2196/19285</pub-id> <pub-id pub-id-type="pmid">32602844</pub-id></citation></ref>
<ref id="B7"><label>7.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Paranjape</surname> <given-names>K</given-names></name> <name><surname>Schinkel</surname> <given-names>M</given-names></name> <name><surname>Nannan Panday</surname> <given-names>R</given-names></name> <name><surname>Car</surname> <given-names>J</given-names></name> <name><surname>Nanayakkara</surname> <given-names>P</given-names></name></person-group>. <article-title>Introducing artificial intelligence training in medical education.</article-title> <source><italic>JMIR Med Educ.</italic></source> (<year>2019</year>) <volume>5</volume>:<fpage>e16048</fpage>. <pub-id pub-id-type="doi">10.2196/16048</pub-id> <pub-id pub-id-type="pmid">31793895</pub-id></citation></ref>
<ref id="B8"><label>8.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Briganti</surname> <given-names>G</given-names></name> <name><surname>Le Moine</surname> <given-names>O</given-names></name></person-group>. <article-title>Artificial intelligence in medicine: today and tomorrow.</article-title> <source><italic>Front Med.</italic></source> (<year>2020</year>) <volume>7</volume>:<fpage>27</fpage>. <pub-id pub-id-type="doi">10.3389/fmed.2020.00027</pub-id> <pub-id pub-id-type="pmid">32118012</pub-id></citation></ref>
<ref id="B9"><label>9.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Masters</surname> <given-names>K</given-names></name></person-group>. <article-title>Ethical implications of AI in medical education.</article-title> <source><italic>Med Teach.</italic></source> (<year>2019</year>) <volume>41</volume>:<fpage>976</fpage>&#x2013;<lpage>80</lpage>. <pub-id pub-id-type="doi">10.1080/0142159X.2019.1595557</pub-id> <pub-id pub-id-type="pmid">31007106</pub-id></citation></ref>
<ref id="B10"><label>10.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Norman</surname> <given-names>G</given-names></name></person-group>. <article-title>Dual processing theory in clinical reasoning.</article-title> <source><italic>Acad Med.</italic></source> (<year>2022</year>) <volume>97</volume>:<fpage>1129</fpage>&#x2013;<lpage>33</lpage>. <pub-id pub-id-type="doi">10.1097/ACM.0000000000004732</pub-id> <pub-id pub-id-type="pmid">35507462</pub-id></citation></ref>
<ref id="B11"><label>11.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Durning</surname> <given-names>SJ</given-names></name> <name><surname>Cleary</surname> <given-names>TJ</given-names></name> <name><surname>Sandars</surname> <given-names>J</given-names></name> <name><surname>Hemmer</surname> <given-names>P</given-names></name> <name><surname>Kokotailo</surname> <given-names>P</given-names></name> <name><surname>Artino</surname> <given-names>AR</given-names> <suffix>Jr</suffix></name></person-group>. <article-title>Neurocognitive models of clinical reasoning.</article-title> <source><italic>Adv Health Sci Educ.</italic></source> (<year>2021</year>) <volume>26</volume>:<fpage>789</fpage>&#x2013;<lpage>802</lpage>.</citation></ref>
<ref id="B12"><label>12.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hammond</surname> <given-names>MM</given-names></name> <name><surname>Boscardin</surname> <given-names>CK</given-names></name> <name><surname>Van Schaik</surname> <given-names>SM</given-names></name> <name><surname>O&#x2019;Sullivan</surname> <given-names>PS</given-names></name> <name><surname>Wiesen</surname> <given-names>LE</given-names></name> <name><surname>Orlander</surname> <given-names>JD</given-names></name></person-group>. <article-title>Digital transformation in medical education.</article-title> <source><italic>Med Teach.</italic></source> (<year>2023</year>) <volume>45</volume>:<fpage>367</fpage>&#x2013;<lpage>75</lpage>.</citation></ref>
<ref id="B13"><label>13.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sapci</surname> <given-names>AH</given-names></name></person-group>. <article-title>Limitations of AI in clinical reasoning development.</article-title> <source><italic>J Med Syst.</italic></source> (<year>2020</year>) <volume>44</volume>:<fpage>142</fpage>. <pub-id pub-id-type="doi">10.1007/s10916-020-01618-2</pub-id> <pub-id pub-id-type="pmid">32691249</pub-id></citation></ref>
<ref id="B14"><label>14.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zhou</surname> <given-names>L</given-names></name> <name><surname>Wang</surname> <given-names>K</given-names></name> <name><surname>Zhang</surname> <given-names>Y</given-names></name> <name><surname>Chen</surname> <given-names>X</given-names></name> <name><surname>Liu</surname> <given-names>M</given-names></name> <name><surname>Li</surname> <given-names>Q</given-names></name></person-group>. <article-title>Multi-institutional validation challenges of AI recommendation systems.</article-title> <source><italic>Comput Biol Med.</italic></source> (<year>2024</year>) <volume>168</volume>:<fpage>107712</fpage>.</citation></ref>
<ref id="B15"><label>15.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Evans</surname> <given-names>JSBT</given-names></name></person-group>. <article-title>Dual-processing accounts of reasoning.</article-title> <source><italic>Annu Rev Psychol.</italic></source> (<year>2008</year>) <volume>59</volume>:<fpage>255</fpage>&#x2013;<lpage>78</lpage>. <pub-id pub-id-type="doi">10.1146/annurev.psych.59.103006.093629</pub-id> <pub-id pub-id-type="pmid">18154502</pub-id></citation></ref>
<ref id="B16"><label>16.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bandura</surname> <given-names>A.</given-names></name></person-group> <source><italic>Social foundations of thought and action.</italic></source> <publisher-loc>Hoboken, NJ</publisher-loc>: <publisher-name>Prentice-Hall</publisher-name> (<year>1986</year>).</citation></ref>
<ref id="B17"><label>17.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Vygotsky</surname> <given-names>LS.</given-names></name></person-group> <source><italic>Mind in society.</italic></source> <publisher-loc>Cambridge, MA</publisher-loc>: <publisher-name>Harvard University Press</publisher-name> (<year>1978</year>).</citation></ref>
<ref id="B18"><label>18.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ryan</surname> <given-names>RM</given-names></name> <name><surname>Deci</surname> <given-names>EL</given-names></name></person-group>. <article-title>Self-determination theory.</article-title> <source><italic>Psychol Inquiry.</italic></source> (<year>2000</year>) <volume>11</volume>:<fpage>227</fpage>&#x2013;<lpage>68</lpage>. <pub-id pub-id-type="doi">10.1207/S15327965PLI1104_01</pub-id></citation></ref>
<ref id="B19"><label>19.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cook</surname> <given-names>DA</given-names></name> <name><surname>Oh</surname> <given-names>SY</given-names></name> <name><surname>Pusic</surname> <given-names>MV</given-names></name> <name><surname>White</surname> <given-names>HM</given-names></name> <name><surname>Zendejas</surname> <given-names>B</given-names></name> <name><surname>Hamstra</surname> <given-names>SJ</given-names></name></person-group>. <article-title>When AI fails in medical education.</article-title> <source><italic>Med Educ.</italic></source> (<year>2022</year>) <volume>56</volume>:<fpage>728</fpage>&#x2013;<lpage>36</lpage>.</citation></ref>
<ref id="B20"><label>20.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Grainger</surname> <given-names>R</given-names></name> <name><surname>Minehart</surname> <given-names>RD</given-names></name> <name><surname>Fisher</surname> <given-names>J</given-names></name> <name><surname>Bertram</surname> <given-names>A</given-names></name> <name><surname>Wrey</surname> <given-names>CF</given-names></name> <name><surname>Plan-Smith</surname> <given-names>MC</given-names></name></person-group>. <article-title>Algorithmic bias in educational AI.</article-title> <source><italic>Lancet Digit Health.</italic></source> (<year>2023</year>) <volume>5</volume>:<fpage>e283</fpage>&#x2013;<lpage>91</lpage>.</citation></ref>
<ref id="B21"><label>21.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wu</surname> <given-names>L</given-names></name> <name><surname>Zhang</surname> <given-names>T</given-names></name> <name><surname>Chen</surname> <given-names>W</given-names></name></person-group>. <article-title>Frontier trends in the integrated development of artificial intelligence and medical education.</article-title> <source><italic>Clin Educ General Pract.</italic></source> (<year>2024</year>) <volume>22</volume>:<fpage>1109</fpage>&#x2013;<lpage>11</lpage>. <pub-id pub-id-type="doi">10.13558/j.cnki.issn1672-3686.2024.012.015</pub-id></citation></ref>
<ref id="B22"><label>22.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Li</surname> <given-names>S</given-names></name> <name><surname>Liu</surname> <given-names>D</given-names></name></person-group>. <article-title>Application and challenges of artificial intelligence technology in medical education.</article-title> <source><italic>Internet Weekly.</italic></source> (<year>2024</year>) <volume>7</volume>:<fpage>78</fpage>&#x2013;<lpage>80</lpage>.</citation></ref>
<ref id="B23"><label>23.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Song</surname> <given-names>Y</given-names></name> <name><surname>Li</surname> <given-names>Z</given-names></name> <name><surname>Ding</surname> <given-names>N</given-names></name></person-group>. <article-title>Analysis of the influence of generative artificial intelligence on medical education.</article-title> <source><italic>China Med Educ Technol.</italic></source> (<year>2024</year>) <volume>38</volume>:<fpage>281</fpage>&#x2013;<lpage>6</lpage>. <pub-id pub-id-type="doi">10.13566/j.cnki.cmet.cn61-1317/g4.202403005</pub-id></citation></ref>
<ref id="B24"><label>24.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kulik</surname> <given-names>JA</given-names></name> <name><surname>Fletcher</surname> <given-names>JD</given-names></name></person-group>. <article-title>Effectiveness of intelligent tutoring systems.</article-title> <source><italic>Rev Educ Res.</italic></source> (<year>2016</year>) <volume>86</volume>:<fpage>42</fpage>&#x2013;<lpage>78</lpage>. <pub-id pub-id-type="doi">10.3102/0034654315581420</pub-id> <pub-id pub-id-type="pmid">38293548</pub-id></citation></ref>
<ref id="B25"><label>25.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wartman</surname> <given-names>SA</given-names></name> <name><surname>Combs</surname> <given-names>CD</given-names></name></person-group>. <article-title>Reimagining medical education.</article-title> <source><italic>Acad Med.</italic></source> (<year>2020</year>) <volume>95</volume>:<fpage>1636</fpage>&#x2013;<lpage>9</lpage>. <pub-id pub-id-type="doi">10.1097/ACM.0000000000003650</pub-id> <pub-id pub-id-type="pmid">32739931</pub-id></citation></ref>
<ref id="B26"><label>26.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Han</surname> <given-names>ER</given-names></name> <name><surname>Yeo</surname> <given-names>S</given-names></name> <name><surname>Kim</surname> <given-names>MJ</given-names></name> <name><surname>Lee</surname> <given-names>YH</given-names></name> <name><surname>Park</surname> <given-names>KH</given-names></name> <name><surname>Roh</surname> <given-names>H</given-names></name></person-group>. <article-title>Medical education trends.</article-title> <source><italic>Korean J Med Educ.</italic></source> (<year>2021</year>) <volume>33</volume>:<fpage>171</fpage>&#x2013;<lpage>81</lpage>.</citation></ref>
<ref id="B27"><label>27.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>van der Vleuten</surname> <given-names>CPM</given-names></name> <name><surname>Schuwirth</surname> <given-names>LWT</given-names></name> <name><surname>Driessen</surname> <given-names>EW</given-names></name> <name><surname>Heeneman</surname> <given-names>S</given-names></name> <name><surname>Dijkstra</surname> <given-names>J</given-names></name> <name><surname>Tigelaar</surname> <given-names>DE</given-names></name></person-group>. <article-title>Assessment paradigm shifts.</article-title> <source><italic>Perspect Med Educ.</italic></source> (<year>2019</year>) <volume>8</volume>:<fpage>261</fpage>&#x2013;<lpage>4</lpage>.</citation></ref>
<ref id="B28"><label>28.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cheung</surname> <given-names>WJ</given-names></name> <name><surname>Hall</surname> <given-names>AK</given-names></name> <name><surname>Skutovich</surname> <given-names>A</given-names></name> <name><surname>Brzezina</surname> <given-names>S</given-names></name> <name><surname>Dalseg</surname> <given-names>TR</given-names></name> <name><surname>Oswald</surname> <given-names>A</given-names></name><etal/></person-group> <article-title>Competency-based medical education.</article-title> <source><italic>Med Teach.</italic></source> (<year>2022</year>) <volume>44</volume>:<fpage>845</fpage>&#x2013;<lpage>54</lpage>. <pub-id pub-id-type="doi">10.1080/0142159X.2022.2072278</pub-id> <pub-id pub-id-type="pmid">35605158</pub-id></citation></ref>
<ref id="B29"><label>29.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ericsson</surname> <given-names>KA</given-names></name></person-group>. <article-title>Deliberate practice.</article-title> <source><italic>Med Educ.</italic></source> (<year>2015</year>) <volume>49</volume>:<fpage>560</fpage>&#x2013;<lpage>71</lpage>. <pub-id pub-id-type="doi">10.1111/medu.12710</pub-id> <pub-id pub-id-type="pmid">25924160</pub-id></citation></ref>
<ref id="B30"><label>30.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Norman</surname> <given-names>GR</given-names></name> <name><surname>Monteiro</surname> <given-names>SD</given-names></name> <name><surname>Sherbino</surname> <given-names>J</given-names></name> <name><surname>Ilgen</surname> <given-names>JS</given-names></name> <name><surname>Schmidt</surname> <given-names>HG</given-names></name> <name><surname>Mamede</surname> <given-names>S</given-names></name></person-group>. <article-title>Assessment in the age of AI.</article-title> <source><italic>Adv Health Sci Educ.</italic></source> (<year>2017</year>) <volume>22</volume>:<fpage>1125</fpage>&#x2013;<lpage>38</lpage>.</citation></ref>
<ref id="B31"><label>31.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Cook</surname> <given-names>DA</given-names></name> <name><surname>Brydges</surname> <given-names>R</given-names></name> <name><surname>Ginsburg</surname> <given-names>S</given-names></name> <name><surname>Hatala</surname> <given-names>R</given-names></name> <name><surname>Kogan</surname> <given-names>JR</given-names></name> <name><surname>Holmboe</surname> <given-names>ES</given-names></name><etal/></person-group> <article-title>Artificial intelligence-enhanced virtual patients for clinical reasoning training in medical education: a randomized trial.</article-title> <source><italic>Med Educ</italic></source>. (<year>2023</year>) <volume>57</volume>:<fpage>1125</fpage>&#x2013;<lpage>36</lpage>. <pub-id pub-id-type="doi">10.1111/medu.15122</pub-id> <pub-id pub-id-type="pmid">37183307</pub-id></citation></ref>
<ref id="B32"><label>32.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bandura</surname> <given-names>A</given-names></name> <name><surname>Schunk</surname> <given-names>DH</given-names></name> <name><surname>Zimmerman</surname> <given-names>BJ</given-names></name> <name><surname>Pintrich</surname> <given-names>PR</given-names></name> <name><surname>Boekaerts</surname> <given-names>M</given-names></name> <name><surname>Winne</surname> <given-names>PH</given-names></name><etal/></person-group> <article-title>Self-regulated learning and AI-driven adaptive systems: bridging theory and practice in medical education.</article-title> <source><italic>Educ Psychol Rev.</italic></source> (<year>2022</year>) <volume>34</volume>:<fpage>789</fpage>&#x2013;<lpage>812</lpage>. <pub-id pub-id-type="doi">10.1007/s10648-022-09685-2</pub-id></citation></ref>
<ref id="B33"><label>33.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hwang</surname> <given-names>G-J</given-names></name> <name><surname>Tu</surname> <given-names>Y-F</given-names></name> <name><surname>Chen</surname> <given-names>X</given-names></name> <name><surname>Zou</surname> <given-names>D</given-names></name> <name><surname>Xie</surname> <given-names>H</given-names></name> <name><surname>Ga&#x0161;evi&#x0107;</surname> <given-names>D</given-names></name><etal/></person-group> <article-title>A four-dimensional model for AI-powered personalized learning: integrating cognitive, affective, behavioral, and social dimensions.</article-title> <source><italic>Comput Educ.</italic></source> (<year>2024</year>) <volume>189</volume>:<fpage>104567</fpage>. <pub-id pub-id-type="doi">10.1016/j.compedu.2024.104567</pub-id></citation></ref>
<ref id="B34"><label>34.</label><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sapci</surname> <given-names>AH</given-names></name> <name><surname>Sapci</surname> <given-names>HA</given-names></name> <name><surname>Khan</surname> <given-names>S</given-names></name> <name><surname>Iyer</surname> <given-names>S</given-names></name> <name><surname>Patel</surname> <given-names>V</given-names></name> <name><surname>Smith</surname> <given-names>J</given-names></name><etal/></person-group> <article-title>Evaluating the efficacy of AI tools in clinical reasoning development: a multi-institutional null-result study.</article-title> <source><italic>J Med Internet Res.</italic></source> (<year>2023</year>) <volume>25</volume>:<fpage>e43210</fpage>. <pub-id pub-id-type="doi">10.2196/43210</pub-id> <pub-id pub-id-type="pmid">37505797</pub-id></citation></ref>
</ref-list>
</back>
</article>