<?xml version="1.0" encoding="utf-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" article-type="research-article" dtd-version="2.3" xml:lang="EN">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Educ.</journal-id>
<journal-title>Frontiers in Education</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Educ.</abbrev-journal-title>
<issn pub-type="epub">2504-284X</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/feduc.2025.1624516</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Education</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Experiment with ChatGPT: methodology of first simulation</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name><surname>Shvets</surname> <given-names>Oleg</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="corresp" rid="c001"><sup>&#x002A;</sup></xref>
<uri xlink:href="https://loop.frontiersin.org/people/3041639/overview"/>
<role content-type="https://credit.niso.org/contributor-roles/writing-review-editing/"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Murtazin</surname> <given-names>Kristina</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<uri xlink:href="https://loop.frontiersin.org/people/3047646/overview"/>
<role content-type="https://credit.niso.org/contributor-roles/writing-review-editing/"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Piho</surname> <given-names>Gunnar</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<uri xlink:href="https://loop.frontiersin.org/people/2531011/overview"/>
<role content-type="https://credit.niso.org/contributor-roles/funding-acquisition/"/>
<role content-type="https://credit.niso.org/contributor-roles/writing-review-editing/"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Meeter</surname> <given-names>Martijn</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="aff" rid="aff2"><sup>2</sup></xref>
<uri xlink:href="https://loop.frontiersin.org/people/61864/overview"/>
<role content-type="https://credit.niso.org/contributor-roles/writing-review-editing/"/>
</contrib>
</contrib-group>
<aff id="aff1"><sup>1</sup><institution>Department of Software Science, Tallinn University of Technology (TalTech)</institution>, <addr-line>Tallinn</addr-line>, <country>Estonia</country></aff>
<aff id="aff2"><sup>2</sup><institution>LEARN! Research Institute, Vrije Universiteit Amsterdam</institution>, <addr-line>Amsterdam</addr-line>, <country>Netherlands</country></aff>
<author-notes>
<fn fn-type="edited-by" id="fn0002">
<p>Edited by: Eug&#x00E8;ne Loos, Utrecht University, Netherlands</p>
</fn>
<fn fn-type="edited-by" id="fn0003">
<p>Reviewed by: Noble Lo, Lancaster University, United Kingdom</p>
<p>Deepika Dhamija, Manipal University Jaipur, India</p>
</fn>
<corresp id="c001">&#x002A;Correspondence: Oleg Shvets, <email>oleg.shvets@taltech.ee</email></corresp>
</author-notes>
<pub-date pub-type="epub">
<day>08</day>
<month>08</month>
<year>2025</year>
</pub-date>
<pub-date pub-type="collection">
<year>2025</year>
</pub-date>
<volume>10</volume>
<elocation-id>1624516</elocation-id>
<history>
<date date-type="received">
<day>07</day>
<month>05</month>
<year>2025</year>
</date>
<date date-type="accepted">
<day>28</day>
<month>07</month>
<year>2025</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x00A9; 2025 Shvets, Murtazin, Piho and Meeter.</copyright-statement>
<copyright-year>2025</copyright-year>
<copyright-holder>Shvets, Murtazin, Piho and Meeter</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/">
<p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p>
</license>
</permissions>
<abstract>
<p>Providing timely and effective feedback is a crucial element of the educational process, directly impacting student engagement, comprehension, and academic achievement. However, even within small groups, delivering personalized feedback presents a significant challenge for educators, especially when opportunities for individual interaction are limited. As a result, there is growing interest in the use of AI-based feedback systems as a potential solution to this problem. This study examines the impact of AI-generated feedback, specifically from ChatGPT 3.5, compared to traditional feedback provided by a supervisor. The aim of the research is to assess students&#x2019; perceptions of both types of feedback, their satisfaction levels, and the effectiveness of each in supporting academic progress. As part of our broader research agenda, we also aim to evaluate the relevance of the domain model currently under development for supporting automated feedback. This model is intended, among other functions, to facilitate the integration of AI-driven mechanisms with student-centered feedback in order to enhance the quality of learning. At this stage, the domain model is employed at a conceptual level to define key actors in the educational process and the relationships between them, to describe the feedback process within a course, and to structure assignment content and assessment criteria. The experiment presented in this study serves as a preparatory step toward the implementation and integration of the model into the educational process, highlighting its function as a conceptual framework for feedback design. Our results indicate that both types of feedback were generally perceived positively, but differences were observed in how their quality was evaluated. In one group, supervisor-provided feedback received higher ratings for clarity, depth, and relevance. At the same time, students in the other group showed a slight preference for feedback from ChatGPT 3.5, particularly in terms of improving their understanding of the assignment topics. The speed and consistency of AI-generated feedback were highlighted as key advantages, indicating its potential value in educational environments where personalized feedback from instructors is limited.</p>
</abstract>
<kwd-group>
<kwd>AI</kwd>
<kwd>ChatGPT 3.5</kwd>
<kwd>domain model</kwd>
<kwd>experiment</kwd>
<kwd>personalized feedback</kwd>
</kwd-group>
<counts>
<fig-count count="7"/>
<table-count count="1"/>
<equation-count count="0"/>
<ref-count count="18"/>
<page-count count="8"/>
<word-count count="5140"/>
</counts>
<custom-meta-wrap>
<custom-meta>
<meta-name>section-at-acceptance</meta-name>
<meta-value>Higher Education</meta-value>
</custom-meta>
</custom-meta-wrap>
</article-meta>
</front>
<body>
<sec sec-type="intro" id="sec1">
<label>1</label>
<title>Introduction</title>
<p>Digitalisation has transformed the way students learn and teachers teach. It has enabled students to access educational resources from anywhere in the world, at any time, and at their own pace (<xref ref-type="bibr" rid="ref13">Steriu and St&#x0103;nescu, 2023</xref>). Automating the learning process may revolutionize learning and teaching by reducing teachers&#x2019; workload and providing students with personalized learning experiences and feedback (<xref ref-type="bibr" rid="ref6">Dhananjaya et al., 2024</xref>). Bill Gates says AI will replace doctors and teachers within 10&#x202F;years and claims that humans will not be needed &#x201C;for most things&#x201D; (<xref ref-type="bibr" rid="ref18">Zilber, 2025</xref>).</p>
<p>The presented work is part of our research program focused on developing a general and universal domain model for learning and teaching in higher education institutions, based on an abstract process domain model. This model should not only encompass process objectives and outcomes but also integrate a reliable feedback mechanism. The proposed domain model serves as a conceptual framework that defines and organizes the key concepts and their interrelationships related to the teaching and learning of a specific subject or topic in higher education. It provides a foundation for describing the learning content, including the set of skills, knowledge and strategies related to the tutored topic, expert knowledge, and potential misconceptions students may have. Additionally, it facilitates understanding and formalization of educational processes, including personalized monitoring and support for students as well as personalized feedback (<xref ref-type="bibr" rid="ref12">Shvets et al., 2024</xref>).</p>
<p>One of the stages of this research program is the evaluation of the domain model for automated personalized feedback to students in the learning process. To evaluate our model, we utilize several methods:</p>
<list list-type="bullet">
<list-item>
<p>Evaluating the correctness of the model using real-life use case scenarios (<xref ref-type="bibr" rid="ref7">Laskar et al., 2023</xref>; <xref ref-type="bibr" rid="ref8">Nasiri et al., 2021</xref>; <xref ref-type="bibr" rid="ref14">Suhail et al., 2022</xref>)</p>
</list-item>
<list-item>
<p>Evaluating the domain model with the involvement of stakeholders to verify its alignment with the needs of users and its practical application in the educational process (<xref ref-type="bibr" rid="ref3">Arora et al., 2024</xref>; <xref ref-type="bibr" rid="ref4">Chowdhury et al., 2022</xref>; <xref ref-type="bibr" rid="ref11">Owan et al., 2023</xref>)</p>
</list-item>
<list-item>
<p>Conducting an experiment that simulates automated feedback according to the proposed domain model under real conditions within a course to evaluate the potential pedagogical effect of the automated feedback (<xref ref-type="bibr" rid="ref16">Wang et al., 2023</xref>; <xref ref-type="bibr" rid="ref17">Yu et al., 2021</xref>)</p>
</list-item>
</list>
<p>In the presented study, we focus on the last of the three points. We conduct an experiment in field conditions within a single academic course to test how our domain model can assist in automating personalized feedback for students. The aim of this experiment is to examine the impact of instant automated personalized feedback on students&#x2019; performance. In one case, the feedback will be provided by the ChatGPT 3.5 chatbot. In the other case, the feedback will be given by a supervisor in the usual manner.</p>
<p>The structure of the paper is as follows: The research methodology is presented in section 2. In section 3, we present the results and conduct an analysis of students&#x2019; performance and their satisfaction with the feedback. Finally, section 4 outlines the main findings, the study&#x2019;s limitations, and general recommendations for future research.</p>
</sec>
<sec sec-type="materials|methods" id="sec2">
<label>2</label>
<title>Materials and methods</title>
<sec id="sec3">
<label>2.1</label>
<title>The approach</title>
<p>The central element of any research is the selection of strategies, methods, or techniques that are applied to conduct the study (<xref ref-type="bibr" rid="ref15">van Thiel, 2014</xref>). As part of our research, we decided to conduct an experiment in which we analyse students&#x2019; satisfaction with the feedback received from both the teacher and ChatGPT 3.5 within the context of the learning course.</p>
</sec>
<sec id="sec4">
<label>2.2</label>
<title>The procedure</title>
<p>The experiment involves 36 students from Taltech Virumaa College, with an average age of 30&#x202F;years. The syllabus for the RAM0800 Digital Logic and Digital Systems study course (6 credits) in the Telematics bachelor&#x2019;s degree program was used for the experiment. The theoretical and practical materials of the course are studied during weeks 1 and 14 of the 17-week course. The experiment was conducted during the last 4 weeks of the course (weeks 14&#x2013;17).</p>
<p>Students&#x2019; tasks consist of four sets of assessment questions. The revised Bloom&#x2019;s taxonomy by <xref ref-type="bibr" rid="ref2">Anderson et al. (2001)</xref> was used for the formulation and design of sets of assessment questions.</p>
<p>Each set of questions for each learning outcome is designed to include four questions: an A, a B, a C, and a D question. The A question covers the &#x201C;Remember&#x201D; level of Bloom&#x2019;s Taxonomy, the B question the &#x201C;Understand&#x201D; level, and the C question the &#x201C;Apply&#x201D; level. The D question covers the three higher levels (analysis, synthesis, and evaluation) of Bloom&#x2019;s taxonomy. The A questions are threshold questions necessary for passing the study course. Students achieving correct answers to questions A, B, and C receive an excellent grade. D questions are designed to indicate to students that there is still much to be learned. The number of attempts for solving each question is unlimited. The time for each attempt is limited depending on the complexity of the question.</p>
<p>As a mechanism for automated personalized feedback on these questions, we utilize the capabilities of ChatGPT 3.5. This innovative AI platform offers significant opportunities for enhancing personalized learning experiences and shaping the future of education (<xref ref-type="bibr" rid="ref1">Albdrani and Al-Shargabi, 2023</xref>). ChatGPT 3.5 analyses students&#x2019; responses and determines whether the correct answer has been provided, assigns grades, and offers feedback recommendations to the student.</p>
<p>While this feedback will be automated in the domain model we are developing, in the experiment, it was provided by the teacher, simulating the future automated feedback algorithm. To maintain a manageable workload for the teacher and simultaneously create high-quality research design, feedback using ChatGPT 3.5 was provided to only half of the students, while the other half received feedback in the usual manner. These groups alternated weekly to ensure fairness. Overall, the experiment consisted of four consecutive stages.</p>
<p>Stage 1: The students were divided into two groups (Group A and Group B, with 18 students each) using simple random sampling facilitated by a random number generator in Microsoft Excel. Due to the small sample size, stratified randomization was not employed. To ensure the equivalence of the groups at baseline, we conducted a comparative analysis of participants&#x2019; key characteristics (age, knowledge level, academic performance) and confirmed that no statistically significant differences were present. This allowed us to consider the groups comparable prior to the start of the experiment.</p>
<p>Stage 2: All students were informed in advance that they would be participating in this experiment. Each student had the option to decline participation, in which case they could complete the course in the usual manner.</p>
<p>During the 14th and 15th weeks of the semester, the teacher provided the first and second sets of control questions (four questions in each set). The task for all students was to complete these control assignments. Upon completion, the task had to be uploaded as a text file to the Moodle platform. The response could be in the form of a text-based essay or specific program code in the Falstad simulator.<xref ref-type="fn" rid="fn0001"><sup>1</sup></xref> Once a student uploaded their work to Moodle, the teacher received a notification.</p>
<p>Group A served as the focus group, receiving additional feedback. Additional feedback meant that students received a response from the teacher immediately after submitting their assignment. To generate this feedback, the teacher used ChatGPT 3.5. Group B acted as the control group and received feedback directly from the teacher within the standard timeframe (no later than 48&#x202F;h after submission) during the same period.</p>
<p>The final deadline for completing the assignments was set at the end of the 15th week.</p>
<p>Stage 3: During the 16th and 17th weeks of the semester, the teacher provided the third and fourth sets of control questions (four questions in each set). At the beginning of week 16, Group B became the focus group, and Group A became the control group. The strict deadline for completing the tasks was set for the end of week 17.</p>
<p>Stage 4: At the end of the 15th and 17th weeks, a survey was conducted among the students. This survey gathered the students&#x2019; opinions on the automated personalized feedback provided during the experiment.</p>
</sec>
<sec id="sec5">
<label>2.3</label>
<title>Data collection procedures</title>
<sec id="sec6">
<label>2.3.1</label>
<title>Qualitative data collection</title>
<p>During the qualitative phase of this study, a carefully designed online survey was conducted using the Moodle survey tool. The aim of this survey was to assess the quality of feedback received by students from both the supervisor and ChatGPT 3.5. The survey questions underwent thorough review and refinement to ensure their clarity and comprehensiveness. Data collection took place over 2 weeks in the spring semester of 2024, resulting in 20 valid responses after Stage 2 and 23 valid responses after Stage 3. The difference in the number of valid responses across the two stages is explained by the absence or incompleteness of some questionnaires. There were no formal refusals to participate; the survey was voluntary and anonymous. However, some students did not complete the survey at one or both stages, while others submitted incomplete responses, which were excluded from the final analysis. Specifically, questionnaires with more than 20% of mandatory questions left unanswered, as well as those displaying inconsistent response patterns (e.g., identical ratings across all items without justification in open-ended fields), were excluded. Incomplete or incorrectly filled questionnaires were filtered out prior to analysis to ensure the reliability of the data.</p>
<p>Our survey is inspired by the work of <xref ref-type="bibr" rid="ref5">Dawson et al. (2019)</xref> and <xref ref-type="bibr" rid="ref10">Olsen and Hunnes (2024)</xref>. However, in our study, we used only the abbreviated version of the survey, focusing on written feedback. Within the course, students primarily received written feedback through electronic annotations, i.e., typed comments in Moodle. Both Group A and Group B received feedback from both the supervisor and ChatGPT 3.5.</p>
<p>For data collection, the survey was conducted within the Moodle learning environment, and the data was exported and analyzed in MS Excel. The survey was conducted at the end of weeks 15 and 17 of the academic term. The survey was personalized and conducted for all participants in the experiment (<xref ref-type="supplementary-material" rid="SM1">Supplementary Appendix A</xref>).</p>
</sec>
<sec id="sec7">
<label>2.3.2</label>
<title>Quantitative data collection</title>
<p>The results of the sets of assessment questions completed by students were used as quantitative data.</p>
<p>We created a form for conducting the experiment and interacting with ChatGPT 3.5. The key feature of this form is that the AI must first generate its own solution and then compare it with the student&#x2019;s solution. Our task for ChatGPT 3.5 is to assess the student&#x2019;s solution. The form for conducting the experiment data is presented in <xref ref-type="table" rid="tab1">Table 1</xref>.</p>
<table-wrap position="float" id="tab1">
<label>Table 1</label>
<caption>
<p>Form for conducting the experiment data.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top">Steps</th>
<th align="left" valign="top">The content of the used form</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="top">1</td>
<td align="left" valign="top">Question: question here</td>
</tr>
<tr>
<td align="left" valign="top">2</td>
<td align="left" valign="top">Student&#x2019;s solution: student&#x2019;s solution here</td>
</tr>
<tr>
<td align="left" valign="top">3</td>
<td align="left" valign="top">Actual solution: steps to work out the solution and your solution here (from teacher and from ChatGPT 3.5)</td>
</tr>
<tr>
<td align="left" valign="top">4</td>
<td align="left" valign="top">If the student&#x2019;s solution is the same as the actual solution just calculated: &#x201C;yes&#x201D; or &#x201C;no&#x201D;</td>
</tr>
<tr>
<td align="left" valign="top">5</td>
<td align="left" valign="top">Student grade: correct or almost correct or incorrect (if correct or almost correct then put points and give feedback, if incorrect the reject and give feedback)</td>
</tr>
<tr>
<td align="left" valign="top">6</td>
<td align="left" valign="top">Feedback: If extra feedback then generated by ChatGPT 3.5, else feedback from teacher.</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>After the experiment was completed, the results were exported to .xls files.</p>
</sec>
</sec>
<sec id="sec8">
<label>2.4</label>
<title>Data analysis procedures</title>
<p>After the qualitative data analysis, we intended to assess the quality of feedback. The analysis of students&#x2019; responses to the questions evaluates the clarity, completeness, comprehensibility, and relevance of the feedback provided by the supervisor and ChatGPT 3.5 in the student survey conducted at the end of the experiment. First, we assessed the percentage of respondents for each question in the proposed survey. Second, we evaluated the average weight values of each question.</p>
<p>As a quantitative analysis, we conducted a comparison of the mean scores assigned by the supervisor and ChatGPT 3.5 for each assignment. We also compared the number of attempts made by students when feedback was provided by the supervisor and ChatGPT 3.5 for each assignment.</p>
</sec>
</sec>
<sec sec-type="results" id="sec9">
<label>3</label>
<title>Results</title>
<sec id="sec10">
<label>3.1</label>
<title>Student evaluation of feedback</title>
<p>Students were asked to evaluate the feedback by indicating their level of agreement with six statements. Each response in the conducted survey has its own weight: 1&#x2014;strongly disagree, 2&#x2014;disagree, 3&#x2014;neither disagree nor agree, 4&#x2014;agree, 5&#x2014;strongly agree, 6&#x2014;not able to judge. In the first survey (after the 15th week), 56% of the total number of students participated. In the second survey (after the 17th week), 64% of the students who took part in the experiment participated.</p>
<p>The results obtained after surveying Group A are presented in <xref ref-type="fig" rid="fig1">Figures 1</xref>, <xref ref-type="fig" rid="fig2">2</xref>. <xref ref-type="fig" rid="fig3">Figures 3</xref>, <xref ref-type="fig" rid="fig4">4</xref> illustrate how students assessed the feedback in Group B after weeks 15 and 17, respectively. Group B was more satisfied on all aspects than Group A. This was the case regardless of whether they were rating the feedback received from the supervisor or from ChatGPT 3.5.</p>
<fig position="float" id="fig1">
<label>Figure 1</label>
<caption>
<p>Degree of agreement regarding feedback from the ChatGPT 3.5 in the section of the course where written feedback was utilized. First stage (14&#x2013;15&#x202F;weeks). Group A. Percentage.</p>
</caption>
<graphic xlink:href="feduc-10-1624516-g001.tif" mimetype="image" mime-subtype="tiff">
<alt-text content-type="machine-generated">Bar chart showing responses to feedback comments on assignments. Most respondents agree or strongly agree that feedback helped them understand topics, improve solutions, and was detailed and consistent. Few strongly disagree or are unable to judge.</alt-text>
</graphic>
</fig>
<fig position="float" id="fig2">
<label>Figure 2</label>
<caption>
<p>Degree of agreement regarding feedback from supervisor in the section of the course where written feedback was utilized. Second stage (16&#x2013;17&#x202F;weeks). Group A. Percentage.</p>
</caption>
<graphic xlink:href="feduc-10-1624516-g002.tif" mimetype="image" mime-subtype="tiff">
<alt-text content-type="machine-generated">Bar chart showing responses to statements about feedback comments on assignments. Categories include: "Strongly disagree," "Disagree," "Neither disagree nor agree," "Agree," "Strongly agree," and "Not able to judge." Most responses are positive, with high percentages in "Agree" and "Strongly agree" for understanding and using feedback.</alt-text>
</graphic>
</fig>
<fig position="float" id="fig3">
<label>Figure 3</label>
<caption>
<p>Degree of agreement regarding feedback from supervisor in the section of the course where written feedback was utilized. First stage (14&#x2013;15&#x202F;weeks). Group B. Percentage.</p>
</caption>
<graphic xlink:href="feduc-10-1624516-g003.tif" mimetype="image" mime-subtype="tiff">
<alt-text content-type="machine-generated">Bar chart depicting survey responses about feedback comments on assignments. The bars show varying degrees of agreement with statements about understanding and using feedback. Colors represent responses from strongly disagree to strongly agree. Most responses indicate agreement or strong agreement across statements.</alt-text>
</graphic>
</fig>
<fig position="float" id="fig4">
<label>Figure 4</label>
<caption>
<p>Degree of agreement regarding feedback from the ChatGPT 3.5 in the section of the course where written feedback was utilized. Second stage (16&#x2013;17&#x202F;weeks). Group B. Percentage.</p>
</caption>
<graphic xlink:href="feduc-10-1624516-g004.tif" mimetype="image" mime-subtype="tiff">
<alt-text content-type="machine-generated">Bar chart showing feedback on assignment comments. Categories include understanding, improvement, detail, and consistency. Most responses are strongly agree, particularly for understanding feedback, with agree as the second most common response. Key colors: strongly disagree (black), disagree (orange), neither (gray), agree (yellow), strongly agree (blue), not able to judge (white).</alt-text>
</graphic>
</fig>
<p>We compared the average satisfaction with feedback received from the supervisor and from ChatGPT 3.5. The result of this comparison is presented in <xref ref-type="fig" rid="fig5">Figure 5</xref>. Again, the differences between Group A and Group B are striking, with Group A (receiving ChatGPT 3.5 feedback in week 15 and supervisor feedback in week 17) being less satisfied than Group B. This was confirmed by a linear mixed models analysis with condition, week and group as fixed effects and student ID as random effect. While there was a trend toward a difference between groups A and B, <italic>F</italic>(1, 24.19)&#x202F;=&#x202F;2.95, <italic>p</italic>&#x202F;=&#x202F;0.095, no effect of week, <italic>F</italic>(1, 20.87)&#x202F;=&#x202F;0.086, <italic>p</italic>&#x202F;=&#x202F;0.773 or condition, <italic>F</italic>(1, 20.87)&#x202F;=&#x202F;0.42, <italic>p</italic>&#x202F;=&#x202F;0.523, could be discerned. Entering group as random factor instead of as fixed effect did not change the results for either week or condition. In other words, students seemed equally satisfied with both types of feedback.</p>
<fig position="float" id="fig5">
<label>Figure 5</label>
<caption>
<p>Average satisfaction with feedback received from the supervisor and from ChatGPT 3.5.</p>
</caption>
<graphic xlink:href="feduc-10-1624516-g005.tif" mimetype="image" mime-subtype="tiff">
<alt-text content-type="machine-generated">Bar chart comparing ratings after weeks 15 and 17 from ChatGPT and a supervisor. Ratings from ChatGPT increased from around 4.2 to 4.8, while ratings from the supervisor decreased slightly from 4.6 to 4.5.</alt-text>
</graphic>
</fig>
</sec>
<sec id="sec11">
<label>3.2</label>
<title>Comparison of mean scores assigned by the supervisor and ChatGPT 3.5</title>
<p>We conducted a comparative analysis of the average scores obtained by students for the completion of control assignments as part of the experiment. Initially, all assignments were graded using a 100-point scale, which allowed for more precise differentiation of student performance. Subsequently, to ensure consistency and facilitate interpretation of the results, the scores were converted into the traditional five-point grading scale commonly used in educational practice.</p>
<p>We compared the average marks awarded to students depending on whether the assessment and feedback were provided by the human supervisor or ChatGPT 3.5. This approach enabled us to analyse the impact of different types of feedback on students&#x2019; academic outcomes.</p>
<p>The comparison of average scores given by the supervisor and ChatGPT 3.5 is presented in <xref ref-type="fig" rid="fig6">Figure 6</xref>. Again, linear mixed-model analyses of differences between grades found a trend for group, <italic>F</italic>(1, 28.20)&#x202F;=&#x202F;3.02; <italic>p</italic>&#x202F;=&#x202F;0.093, this time with an advantage for group A. However, no effect was found for either week, <italic>F</italic>(1, 27.63)&#x202F;=&#x202F;0.07, <italic>p</italic>&#x202F;=&#x202F;0.79, or for condition, <italic>F</italic>(1, 27.63)&#x202F;=&#x202F;0.02, <italic>p</italic>&#x202F;=&#x202F;0.83. Again, the same results for week and condition held when group was entered as a random instead of fixed effect. There was thus no evidence of any change in grades as a function of the kind of feedback.</p>
<fig position="float" id="fig6">
<label>Figure 6</label>
<caption>
<p>Average grade on assignments from the supervisor and ChatGPT 3.5.</p>
</caption>
<graphic xlink:href="feduc-10-1624516-g006.tif" mimetype="image" mime-subtype="tiff">
<alt-text content-type="machine-generated">Bar chart comparing values labeled "From ChatGPT" and "From supervisor" after week fifteen and week seventeen. After week fifteen, "From ChatGPT" is higher at about ninety, while "From supervisor" is around seventy. After week seventeen, "From supervisor" increases to ninety-five, surpassing "From ChatGPT" at seventy-five.</alt-text>
</graphic>
</fig>
</sec>
<sec id="sec12">
<label>3.3</label>
<title>Number of attempts made by students</title>
<p>We compared the number of student response attempts made to achieve the highest result. Given that each attempt rejected by the teacher was necessarily accompanied by a comment with personalized feedback for the student, this means that the number of attempts is directly related to the number of feedback comments.</p>
<p>The comparison of the average number of attempts made by students to complete tasks, when working with a supervisor and with ChatGPT 3.5, is presented in <xref ref-type="fig" rid="fig7">Figure 7</xref>. No difference between the two conditions was found, <italic>F</italic>(1, 27.55)&#x202F;=&#x202F;1.23, <italic>p</italic>&#x202F;=&#x202F;0.277, nor of week, <italic>F</italic>(1, 27.55)&#x202F;=&#x202F;1.23, <italic>p</italic>&#x202F;=&#x202F;0.277, nor of group, <italic>F</italic>(1, 28.10)&#x202F;=&#x202F;0.001, <italic>p</italic>&#x202F;=&#x202F;0.97.</p>
<fig position="float" id="fig7">
<label>Figure 7</label>
<caption>
<p>Average number of attempts from the supervisor and ChatGPT 3.5.</p>
</caption>
<graphic xlink:href="feduc-10-1624516-g007.tif" mimetype="image" mime-subtype="tiff">
<alt-text content-type="machine-generated">Bar chart comparing feedback received after week 15 and week 17. After week 15: ChatGPT feedback is 8, supervisor feedback is 6. After week 17: ChatGPT feedback is 7, supervisor feedback is 7.</alt-text>
</graphic>
</fig>
</sec>
</sec>
<sec sec-type="discussion" id="sec13">
<label>4</label>
<title>Discussion</title>
<p>The results of our study indicate that both supervisor-provided feedback and feedback from ChatGPT 3.5 were generally well received by students; no differences were observed in student satisfaction and their perception of feedback quality. Overall, students from both groups (A and B) demonstrated a positive attitude toward both types of feedback, acknowledging their impact on improving their work, enhancing their understanding of assignment topics, and motivating them in their studies.</p>
<p>Regarding feedback quality, our findings show that students were generally equally satisfied with the clarity and level of detail in the comments provided by their supervisors.</p>
<p>A comparison of the average grades received by students based on the source of feedback (<xref ref-type="fig" rid="fig6">Figure 6</xref>) showed no differences in academic performance. This suggests that in line with the similarity in the perceived quality of feedback, both types of comments led to similar academic outcomes. Furthermore, a comparison of the number of attempts made by students to achieve their best results (<xref ref-type="fig" rid="fig7">Figure 7</xref>) revealed a similar level of engagement with both types of feedback. This confirms that the effectiveness of feedback in encouraging students to revise and improve their work was comparable for both supervisor and ChatGPT 3.5 feedback.</p>
<p>Several limitations of this study warrant consideration. Firstly, the relatively small sample size (<italic>N</italic>&#x202F;=&#x202F;36) restricts the generalizability of the findings. Although baseline equivalence between groups was confirmed, future research involving larger and more diverse samples would strengthen the reliability of the results and enhance their relevance across diverse educational contexts and disciplines.</p>
<p>Secondly, the AI-generated feedback was delivered exclusively via ChatGPT 3.5 and was not fully integrated with the domain model currently under development, nor supported by a specialized automated feedback system. Consequently, the feedback may have lacked the depth, adaptability, and contextual sensitivity that our model is intended to provide in the future.</p>
<p>Thirdly, the potential for experimenter bias warrants consideration. Although random allocation to groups was employed to minimize systematic bias and ensure representativeness, the supervisor&#x2019;s dual role&#x2014;as both the provider of feedback in the control group and the assessor&#x2014;may have inadvertently influenced students&#x2019; perceptions or assessment outcomes. The involvement of independent evaluators, as well as greater automation in the delivery of feedback, would likely mitigate this risk.</p>
<p>Finally, this study focused on short-term effects within a single course. Further research involving multiple courses and curricula is required to develop a more comprehensive understanding of the role that AI-generated feedback plays in supporting student learning and academic development.</p>
<p>The results of our study can be further contextualized within recent empirical research on the application of generative AI in higher education. For example, a randomized controlled trial conducted by <xref ref-type="bibr" rid="ref9">Noble and Wong Alan (2025)</xref> at the University of Hong Kong investigated the effects of AI-generated feedback on students&#x2019; written assignments, motivation, and emotional responses. Their study reported a significantly greater improvement in essay quality among students who received AI-generated feedback compared to those in the control group, underscoring the potential of such systems to enhance academic performance.</p>
<p>These findings align with our own and support the view that AI-generated feedback (e.g., via ChatGPT 3.5) can be as effective as feedback provided by human supervisors. This reinforces the broader hypothesis that generative AI tools can successfully complement or partially replace traditional feedback mechanisms, particularly under constraints of limited time and teaching resources. Our results further validate the pedagogical value of AI-driven feedback and its role in enhancing the human element of the learning process.</p>
<p>The experiment we conducted validates the functionality of the feedback component within our domain model for learning and teaching. This model enables an automated system to determine whether a student is prepared to begin a task and whether the outcome meets specified criteria. During task execution, the student&#x2019;s actions are aggregated into an actual result, which is then used to generate feedback. Although this architecture supports fully automated feedback, in the present study the process was manually simulated with the involvement of a teacher.</p>
<p>Ultimately, our findings highlight the potential of AI-based feedback systems such as ChatGPT 3.5 as a valuable complement to traditional supervisor feedback, particularly in large-scale educational settings, where providing personalized comments to each student within a short timeframe can be challenging. However, the importance of supervisor feedback should not be underestimated, especially in situations requiring deep understanding and context-specific guidance.</p>
<p>Our findings inspire future research. The domain model for learning and teaching that we are developing should be universal, aimed at optimizing the combination of human and automated feedback to enhance student engagement, learning outcomes, and overall satisfaction. Our experiment demonstrated that automated feedback is not inferior to manual feedback, which requires significantly more effort. This further underscores the importance of exploring processes for automating and personalizing feedback.</p>
</sec>
</body>
<back>
<sec sec-type="data-availability" id="sec14">
<title>Data availability statement</title>
<p>The original contributions presented in the study are included in the article/<xref ref-type="supplementary-material" rid="SM1">Supplementary material</xref>, further inquiries can be directed to the corresponding author.</p>
</sec>
<sec sec-type="ethics-statement" id="sec15">
<title>Ethics statement</title>
<p>Ethical approval was not required for the studies involving humans because the study was conducted anonymously for the participants. All participants were informed that they were part of an experiment. Each participant had the option to withdraw from the experiment. The presented results do not include any personal data of the participants. The studies were conducted in accordance with the local legislation and institutional requirements. Written informed consent for participation was not required from the participants or the participants&#x2019; legal guardians/next of kin in accordance with the national legislation and institutional requirements.</p>
</sec>
<sec sec-type="author-contributions" id="sec16">
<title>Author contributions</title>
<p>OS: Writing &#x2013; review &#x0026; editing. KM: Writing &#x2013; review &#x0026; editing. GP: Funding acquisition, Writing &#x2013; review &#x0026; editing. MM: Writing &#x2013; review &#x0026; editing.</p>
</sec>
<sec sec-type="funding-information" id="sec17">
<title>Funding</title>
<p>The author(s) declare that no financial support was received for the research and/or publication of this article.</p>
</sec>
<sec sec-type="COI-statement" id="sec18">
<title>Conflict of interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec sec-type="ai-statement" id="sec19">
<title>Generative AI statement</title>
<p>The authors declare that Gen AI was used in the creation of this manuscript for conducting the experiment. ChatGPT 3.5 was part of the experiment conducted.</p>
</sec>
<sec sec-type="disclaimer" id="sec20">
<title>Publisher&#x2019;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
<sec sec-type="supplementary-material" id="sec21">
<title>Supplementary material</title>
<p>The Supplementary material for this article can be found online at: <ext-link xlink:href="https://www.frontiersin.org/articles/10.3389/feduc.2025.1624516/full#supplementary-material" ext-link-type="uri">https://www.frontiersin.org/articles/10.3389/feduc.2025.1624516/full#supplementary-material</ext-link></p>
<supplementary-material xlink:href="Data_Sheet_1.docx" id="SM1" mimetype="application/vnd.openxmlformats-officedocument.wordprocessingml.document" xmlns:xlink="http://www.w3.org/1999/xlink"/>
</sec>
<fn-group>
<fn id="fn0001"><p><sup>1</sup><ext-link xlink:href="https://www.falstad.com/circuit/" ext-link-type="uri">https://www.falstad.com/circuit/</ext-link></p></fn>
</fn-group>
<ref-list>
<title>References</title>
<ref id="ref1"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Albdrani</surname> <given-names>R. N.</given-names></name> <name><surname>Al-Shargabi</surname> <given-names>A. A.</given-names></name></person-group> (<year>2023</year>). <article-title>Investigating the effectiveness of ChatGPT for providing personalized learning experience: a case study</article-title>. <source>Int. J. Adv. Comput. Sci. Appl.</source> <volume>14</volume>, <fpage>1208</fpage>&#x2013;<lpage>1213</lpage>. doi: <pub-id pub-id-type="doi">10.14569/IJACSA.2023.01411122</pub-id></citation></ref>
<ref id="ref2"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Anderson</surname> <given-names>L. W.</given-names></name> <name><surname>Krathwohl</surname> <given-names>D. R.</given-names></name> <name><surname>Bloom</surname> <given-names>B. S.</given-names></name></person-group> (<year>2001</year>). <source>A taxonomy for learning, teaching, and assessing: a revision of Bloom&#x2019;s taxonomy of educational objectives</source>. New York: <publisher-name>Addison Wesley Longman, Inc</publisher-name>.</citation></ref>
<ref id="ref3"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Arora</surname> <given-names>C.</given-names></name> <name><surname>Grundy</surname> <given-names>J.</given-names></name> <name><surname>Abdelrazek</surname> <given-names>M.</given-names></name></person-group> (<year>2024</year>). &#x201C;<article-title>Advancing requirements engineering through generative AI: assessing the role of LLMs</article-title>&#x201D; in <source>Generative AI for effective software development</source> (<publisher-name>Springer</publisher-name>), <fpage>129</fpage>&#x2013;<lpage>148</lpage>. doi: <pub-id pub-id-type="doi">10.1007/978-3-031-55642-5_6</pub-id></citation></ref>
<ref id="ref4"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chowdhury</surname> <given-names>N.</given-names></name> <name><surname>Katsikas</surname> <given-names>S.</given-names></name> <name><surname>Gkioulos</surname> <given-names>V.</given-names></name></person-group> (<year>2022</year>). <article-title>Modeling effective cybersecurity training frameworks: a Delphi method-based study</article-title>. <source>Comput. Secur.</source> <volume>113</volume>:<fpage>102551</fpage>. doi: <pub-id pub-id-type="doi">10.1016/j.cose.2021.102551</pub-id></citation></ref>
<ref id="ref5"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dawson</surname> <given-names>P.</given-names></name> <name><surname>Henderson</surname> <given-names>M.</given-names></name> <name><surname>Mahoney</surname> <given-names>P.</given-names></name> <name><surname>Phillips</surname> <given-names>M.</given-names></name> <name><surname>Ryan</surname> <given-names>T.</given-names></name> <name><surname>Boud</surname> <given-names>D.</given-names></name> <etal/></person-group>. (<year>2019</year>). <article-title>What makes for effective feedback: staff and student perspectives</article-title>. <source>Assess. Eval. High. Educ.</source> <volume>44</volume>, <fpage>25</fpage>&#x2013;<lpage>36</lpage>. doi: <pub-id pub-id-type="doi">10.1080/02602938.2018.1467877</pub-id></citation></ref>
<ref id="ref6"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dhananjaya</surname> <given-names>G. M.</given-names></name> <name><surname>Goudar</surname> <given-names>R. H.</given-names></name> <name><surname>Kulkarni</surname> <given-names>A. A.</given-names></name> <name><surname>Rathod</surname> <given-names>V. N.</given-names></name> <name><surname>Hukkeri</surname> <given-names>G. S.</given-names></name></person-group> (<year>2024</year>). <article-title>A digital recommendation system for personalized learning to enhance online education: a review</article-title>. <source>IEEE Access</source> <volume>12</volume>, <fpage>34019</fpage>&#x2013;<lpage>34041</lpage>. doi: <pub-id pub-id-type="doi">10.1109/ACCESS.2024.3369901</pub-id></citation></ref>
<ref id="ref7"><citation citation-type="other"><person-group person-group-type="author"><name><surname>Laskar</surname> <given-names>M. T. R.</given-names></name> <name><surname>Bari</surname> <given-names>M. S.</given-names></name> <name><surname>Rahman</surname> <given-names>M.</given-names></name> <name><surname>Bhuiyan</surname> <given-names>M. A. H.</given-names></name> <name><surname>Joty</surname> <given-names>S.</given-names></name> <name><surname>Huang</surname> <given-names>J. X.</given-names></name></person-group> (<year>2023</year>). &#x201C;<article-title>A systematic study and comprehensive evaluation of ChatGPT on benchmark datasets</article-title>&#x201D; in <source>Proceedings of the Annual Meeting of the Association for Computational Linguistics</source>, <fpage>431</fpage>&#x2013;<lpage>469</lpage>. doi: <pub-id pub-id-type="doi">10.48550/arXiv.2305.18486</pub-id></citation></ref>
<ref id="ref8"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Nasiri</surname> <given-names>S.</given-names></name> <name><surname>Rhazali</surname> <given-names>Y.</given-names></name> <name><surname>Lahmer</surname> <given-names>M.</given-names></name> <name><surname>Adadi</surname> <given-names>A.</given-names></name></person-group> (<year>2021</year>). <article-title>From user stories to UML diagrams driven by ontological and production model</article-title>. <source>Int. J. Adv. Comput. Sci. Appl.</source> <volume>12</volume>, <fpage>333</fpage>&#x2013;<lpage>340</lpage>. doi: <pub-id pub-id-type="doi">10.14569/IJACSA.2021.0120637</pub-id></citation></ref>
<ref id="ref9"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Noble</surname> <given-names>L.</given-names></name> <name><surname>Wong Alan</surname> <given-names>C. S.</given-names></name></person-group> (<year>2025</year>). <article-title>The impact of generative AI on essay revisions and student engagement</article-title>. <source>Comput. Educ. Open</source>:<fpage>100249</fpage>. doi: <pub-id pub-id-type="doi">10.1016/j.caeo.2025.100249</pub-id></citation></ref>
<ref id="ref10"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Olsen</surname> <given-names>T.</given-names></name> <name><surname>Hunnes</surname> <given-names>J.</given-names></name></person-group> (<year>2024</year>). <article-title>Improving students&#x2019; learning&#x2014;the role of formative feedback: experiences from a crash course for business students in academic writing</article-title>. <source>Assess. Eval. High. Educ.</source> <volume>49</volume>, <fpage>129</fpage>&#x2013;<lpage>141</lpage>. doi: <pub-id pub-id-type="doi">10.1080/02602938.2023.2187744</pub-id></citation></ref>
<ref id="ref11"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Owan</surname> <given-names>V. J.</given-names></name> <name><surname>Abang</surname> <given-names>K. B.</given-names></name> <name><surname>Idika</surname> <given-names>D. O.</given-names></name> <name><surname>Etta</surname> <given-names>E. O.</given-names></name> <name><surname>Bassey</surname> <given-names>B. A.</given-names></name></person-group> (<year>2023</year>). <article-title>Exploring the potential of artificial intelligence tools in educational measurement and assessment</article-title>. <source>Eurasia J. Math. Sci. Technol. Educ</source> <volume>19</volume>:<fpage>em2307</fpage>. doi: <pub-id pub-id-type="doi">10.29333/ejmste/13428</pub-id></citation></ref>
<ref id="ref12"><citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Shvets</surname> <given-names>O.</given-names></name> <name><surname>Murtazin</surname> <given-names>K.</given-names></name> <name><surname>Meeter</surname> <given-names>M.</given-names></name> <name><surname>Piho</surname> <given-names>G.</given-names></name></person-group>, (<year>2024</year>). <article-title>Towards a domain model for learning and teaching</article-title>. In: <conf-name>International Conference on Model-Driven Engineering and Software Development</conf-name>. pp. <fpage>288</fpage>&#x2013;<lpage>296</lpage>. doi: <pub-id pub-id-type="doi">10.5220/0012471400003645</pub-id></citation></ref>
<ref id="ref13"><citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Steriu</surname> <given-names>I.</given-names></name> <name><surname>St&#x0103;nescu</surname> <given-names>A.</given-names></name></person-group> (<year>2023</year>). <article-title>Digitalization in education: navigating the future of learning</article-title>. In: <conf-name>Proceedings of the International Conference on Virtual Learning</conf-name>, pp. <fpage>169</fpage>&#x2013;<lpage>182</lpage>. doi: <pub-id pub-id-type="doi">10.58503/icvl-v18y202314</pub-id></citation></ref>
<ref id="ref14"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Suhail</surname> <given-names>S.</given-names></name> <name><surname>Malik</surname> <given-names>S. U. R.</given-names></name> <name><surname>Jurdak</surname> <given-names>R.</given-names></name> <name><surname>Hussain</surname> <given-names>R.</given-names></name> <name><surname>Matulevi&#x010D;ius</surname> <given-names>R.</given-names></name> <name><surname>Svetinovic</surname> <given-names>D.</given-names></name></person-group> (<year>2022</year>). <article-title>Towards situational aware cyber-physical systems: a security-enhancing use case of blockchain-based digital twins</article-title>. <source>Comput. Ind.</source> <volume>141</volume>:<fpage>103699</fpage>. doi: <pub-id pub-id-type="doi">10.1016/j.compind.2022.103699</pub-id></citation></ref>
<ref id="ref15"><citation citation-type="book"><person-group person-group-type="author"><name><surname>van Thiel</surname> <given-names>S.</given-names></name></person-group> (<year>2014</year>). <source>Research methods in public administration and public management: an introduction.</source> London: <publisher-name>Routledge</publisher-name>. doi: <pub-id pub-id-type="doi">10.4324/9780203078525</pub-id></citation></ref>
<ref id="ref16"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wang</surname> <given-names>H.</given-names></name> <name><surname>Tlili</surname> <given-names>A.</given-names></name> <name><surname>Huang</surname> <given-names>R.</given-names></name> <name><surname>Cai</surname> <given-names>Z.</given-names></name> <name><surname>Li</surname> <given-names>M.</given-names></name> <name><surname>Cheng</surname> <given-names>Z.</given-names></name> <etal/></person-group>. (<year>2023</year>). <article-title>Examining the applications of intelligent tutoring systems in real educational contexts: a systematic literature review from the social experiment perspective</article-title>. <source>Educ. Inf. Technol.</source> <volume>28</volume>, <fpage>9113</fpage>&#x2013;<lpage>9148</lpage>. doi: <pub-id pub-id-type="doi">10.1007/s10639-022-11555-x</pub-id>, PMID: <pub-id pub-id-type="pmid">36643383</pub-id></citation></ref>
<ref id="ref17"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Yu</surname> <given-names>S. J.</given-names></name> <name><surname>Hsueh</surname> <given-names>Y. L.</given-names></name> <name><surname>Sun</surname> <given-names>J. C. Y.</given-names></name> <name><surname>Liu</surname> <given-names>H. Z.</given-names></name></person-group> (<year>2021</year>). <article-title>Developing an intelligent virtual reality interactive system based on the ADDIE model for learning pour-over coffee brewing</article-title>. <source>Comput. Educ. Artif. Intell.</source> <volume>2</volume>:<fpage>100030</fpage>. doi: <pub-id pub-id-type="doi">10.1016/j.caeai.2021.100030</pub-id></citation></ref>
<ref id="ref18"><citation citation-type="other"><person-group person-group-type="author"><name><surname>Zilber</surname> <given-names>A.</given-names></name></person-group>, (<year>2025</year>). <article-title>No title [WWW document]</article-title>. <source>New York Times</source> Available online at: <ext-link xlink:href="https://nypost.com/2025/03/27/business/bill-gates-said-ai-will-replace-doctors-teachers-within-10-years/" ext-link-type="uri">https://nypost.com/2025/03/27/business/bill-gates-said-ai-will-replace-doctors-teachers-within-10-years/</ext-link></citation></ref>
</ref-list>
</back>
</article>