Skip Navigation
Skip to contents

JEEHP : Journal of Educational Evaluation for Health Professions

OPEN ACCESS
SEARCH
Search

Search

Page Path
HOME > Search
31 "Artificial intelligence"
Filter
Filter
Article category
Keywords
Publication year
Authors
Funded articles
Research articles
Refusal-aware evaluation of frontier AI models available in June 2026 using the Japanese National License Examination for Pharmacists: a comparative study
Hiroyasu Sato, Katsuhiko Ogasawara, Hidehiko Sakurai
J Educ Eval Health Prof. 2026;23:29.   Published online September 3, 2026
DOI: https://doi.org/10.3352/jeehp.2026.23.29    [Epub ahead of print]
  • 223 View
  • 23 Download
AbstractAbstract PDF
Purpose
Conventional single-run accuracy may be insufficient for evaluating frontier generative artificial intelligence (AI) models when safety-related refusals occur. This 4-model benchmark examined the need for repeated, refusal-aware evaluation using the Japanese National License Examination for Pharmacists (JNLEP).
Methods
ChatGPT GPT-5.5, Gemini 3.5 Flash, Claude Opus 4.8, and Claude Fable 5 were evaluated using all 345 questions from the 107th JNLEP. The original Japanese questions, including image-containing items, were submitted through application programming interfaces (APIs) in 3 independent runs. Refusals were treated as incorrect when overall accuracy was calculated. For Fable 5, accuracy excluding refusals, refusal rate, refusal consistency across runs, the subject-wise distribution of refusals, and system-assigned refusal categories were also evaluated.
Results
The mean overall accuracies were 98.7% for GPT-5.5, 98.3% for Gemini 3.5 Flash, 96.1% for Claude Opus 4.8, and 70.3% for Claude Fable 5. Fable 5 had a mean refusal rate of 29.0%, whereas its mean accuracy excluding refusals was 99.0%. Among the 345 items, 93 were refused in all 3 runs, 14 were refused inconsistently across runs, and 238 were never refused. All 300 refusal responses were assigned to the bio category. Refusals were most frequent in Biology (90.0%) and Pharmacology (69.2%) but uncommon in Practice (2.5%).
Conclusion
Near-saturation benchmark performance coexisted with frequent and partly run-dependent refusals in a safeguard-equipped model. Overall accuracy, accuracy excluding refusals, refusal rate, and refusal consistency describe complementary aspects of performance; therefore, repeated, refusal-aware evaluation is needed to interpret frontier AI models in pharmacy education.
Practicing pediatric interviews with parents through conversational AI-based virtual patients: an observational study in Spanish undergraduate medical education
César Fernández, Francisco Sánchez-Ferrer, María Asunción Vicente
J Educ Eval Health Prof. 2026;23:27.   Published online August 31, 2026
DOI: https://doi.org/10.3352/jeehp.2026.23.27    [Epub ahead of print]
  • 242 View
  • 27 Download
AbstractAbstract PDFSupplementary Material
Purpose
To evaluate students’ acceptance of training via artificial intelligence (AI)-based virtual patients (VPs) in pediatrics; to analyze the main factors that affect student satisfaction; and to compare student satisfaction with that reported in the literature.
Methods
Two observational studies were carried out. Study S1 analyzed students’ interactions with the platform and study S2 analyzed their answers to a final questionnaire. All students enrolled in the “Pediatrics II and Pediatric Surgery” course were invited to participate during a continuous session on May 7, 2025. Study S1 analyzed whether the case/session order, the number of interactions, the time spent, and student gender were associated with student satisfaction (by fitting a linear mixed-effects model); and also performed a qualitative analysis of students’ open-ended comments. Study S2 analyzed the influence of demographic data on students’ feedback (by fitting proportional odds ordinal logistic regression models and linear models). All experimental data, code and results are available.
Results
Concerning study S1, 70 students participated. Platform rating was high (mean=9.04, standard deviation=1.09) and was positively associated with the number of interactions with the VPs (P=0.038, β=0.022 [0.001–0.043]). Concerning study S2, 60 students participated. No statistically significant influence of demographics was found either for answers to individual questionnaire items or for grouped answers.
Conclusion
Student acceptance of training with AI-based VPs was supported by the satisfaction outcomes measured and by comparison with previous studies. However, further research is needed to confirm these findings.
Development and psychometric validation of the Thai AI Literacy Scale for Nursing Students in Thailand: a methodological study
Wannaporn Jongchidklang, Supalak Phonphithak, Porntep Amornritvanich
J Educ Eval Health Prof. 2026;23:22.   Published online August 4, 2026
DOI: https://doi.org/10.3352/jeehp.2026.23.22    [Epub ahead of print]
  • 707 View
  • 109 Download
AbstractAbstract PDF
Purpose
This study aimed to develop and validate the Thai AI Literacy Scale for Nursing Students (TAILS-NS).
Methods
A cross-sectional study was conducted across multiple nursing institutions in Thailand from March to May 2026 to address the absence of a validated artificial intelligence (AI) literacy instrument for nursing students in Thai or Southeast Asian contexts. A total of 410 nursing students participated, yielding a response rate of 94.5%. The TAILS-NS was developed through item generation, expert content-validity assessment using the item-objective congruence index, and a 2-phase pilot study, resulting in a 40-item instrument comprising a 15-item knowledge test and 25 Likert-scale items across 6 domains. Exploratory factor analysis using maximum likelihood estimation and confirmatory factor analysis using the weighted least squares mean and variance-adjusted estimator, as well as internal consistency, convergent validity, and discriminant validity, were assessed. An independent validation sample (n=157) was recruited for cross-validation confirmatory factor analysis. Raw data are available as a supplement.
Results
Exploratory factor analysis supported a 6-factor structure: AI awareness, skills, ethics and professionalism, positive attitude, AI anxiety, and readiness. Confirmatory factor analysis showed acceptable model fit, although the root mean square error of approximation (RMSEA) indicated marginal fit (comparative fit index [CFI]=0.919, Tucker-Lewis index [TLI]=0.906, RMSEA=0.090). Reliability of the knowledge subscale was acceptable (Kuder-Richardson Formula 20=0.758). Internal consistency was excellent (α=0.810–0.934; total α=0.974; ω=0.989). Average variance extracted exceeded 0.50 for all factors, supporting convergent validity. Most heterotrait-monotrait ratios were below 0.90, supporting discriminant validity. Cross-validation confirmed factorial replicability (CFI=0.917, TLI=0.904).
Conclusion
The TAILS-NS provides evidence of validity and reliability for measuring AI literacy among Thai nursing students and may serve as a standardized tool for curriculum evaluation. Future studies should expand the AI anxiety subscale and explore cross-cultural applicability in Southeast Asian nursing contexts.
Review
Analysis of digital twin applications in nursing practice and education: a scoping review  
A Reum Lim, Hyun Kyoung Kim
J Educ Eval Health Prof. 2026;23:13.   Published online June 9, 2026
DOI: https://doi.org/10.3352/jeehp.2026.23.13
  • 1,309 View
  • 112 Download
  • 1 Web of Science
AbstractAbstract PDFSupplementary Material
This scoping review examined research applying digital twins in nursing practice and education and summarized their application domains, methods, outcomes, and implications. A human digital twin is a virtual health replica modeled from real-world data. This study followed the 5-stage scoping review process proposed by Arksey and O’Malley. Two researchers independently conducted the literature search without restrictions on publication year. From April 1 to 15, 2026, the Cochrane Library, PubMed, Embase, CINAHL, ERIC, and RISS databases were searched, and 15 studies were ultimately included. Digital twin applications were identified in 3 major domains: clinical practice and patient-centered care, education and training, and decision-making and workflow management. Application methods and outcomes varied according to technological implementation and included (1) modeling and data-driven prediction, (2) development of immersive learning and practice-training environments, and (3) system integration and decision-support frameworks. In clinical settings, multimodal patient data can be analyzed using artificial intelligence and machine learning to generate a virtual persona resembling the patient, thereby facilitating real-time personalized nursing care and self-management. In educational settings, digital twins can provide realistic and safe learning environments that enhance training effectiveness. Digital twins show substantial potential to advance predictive and personalized nursing in both clinical practice and education. Their data-driven capabilities are expected to contribute to innovative applications in future nursing practice and educational environments.
Brief report
Large language model-generated versus teacher-written objective structured clinical examination stations for medical students: a blinded comparative pilot study  
Piotr Szychowiak, Jonathan Wong-So, Hélène Messet, Mélanie Faure, Isaure Breteau, Simon Jamard, François Barbier, Maxime Desgrouas
J Educ Eval Health Prof. 2026;23:9.   Published online May 26, 2026
DOI: https://doi.org/10.3352/jeehp.2026.23.9
  • 1,216 View
  • 93 Download
AbstractAbstract PDFSupplementary Material
Developing objective structured clinical examination (OSCE) stations is time-consuming for medical teachers. We aimed to evaluate the ability of a large language model (LLM) to generate ready-to-use OSCE stations. Five OSCE stations generated by the LLM GPT-4o were evaluated by 7 expert assessors using a 5-point Likert scale and compared with 5 teacher-written stations targeting similar learning objectives. A station was considered to be of good quality if most assessors responded “agree” or “strongly agree” to the statement “The station is good enough to be used by students.” All teacher-written stations were rated as being of good quality, compared with only one GPT-4o-generated station. The LLM produced adequate clinical scenarios when reference knowledge was provided and tasks were clearly ordered, but it failed to generate reliable assessment grids. Careful review by teachers remained essential. GPT-4o failed to consistently produce fully ready-to-use OSCE stations.
Reviews
The impact of artificial intelligence-driven simulation on the development of non-technical skills in medical education: a systematic review  
Sana Loubbairi, Yasmine El Moussaoui, Laila Lahlou, Imad Chakri, Hicham Nassik
J Educ Eval Health Prof. 2025;22:37.   Published online November 24, 2025
DOI: https://doi.org/10.3352/jeehp.2025.22.37
  • 5,641 View
  • 526 Download
  • 3 Web of Science
  • 6 Crossref
AbstractAbstract PDFSupplementary Material
Purpose
Artificial intelligence (AI)-driven simulation is an emerging approach in healthcare education that enhances learning effectiveness. This review examined its impact on the development of non-technical skills among medical learners.
Methods
Following the PRISMA (Preferred Reporting Items for Systematic Reviews and Meta-Analyses) guidelines, a systematic review was conducted using the following databases: Web of Science, ScienceDirect, Scopus, and PubMed. The quality of the included studies was assessed using the Mixed Methods Appraisal Tool. The protocol was previously registered in PROSPERO (CRD420251038024).
Results
Of the 1,442 studies identified in the initial search, 20 met the inclusion criteria, involving 2,535 participants. The simulators varied considerably, ranging from platforms built on symbolic AI methods to social robots powered by computational AI. Among the 15 AI-driven simulators, 10 used ChatGPT or its variants as virtual patients. Several studies evaluated multiple non-technical skills simultaneously. Communication and clinical reasoning were the most frequently assessed skills, appearing in 12 and 6 studies, respectively, which generally reported positive outcomes. Improvements were also noted in decision-making, empathy, self-confidence, critical thinking, and problem-solving. In contrast, emotional regulation, assessed in a single study, showed no significant difference. Notably, none of the studies examined reflection, reflective practice, teamwork, or leadership.
Conclusion
AI-driven simulation shows substantial potential for enhancing non-technical skills in medical education, particularly communication and clinical reasoning. However, its effects on several other non-technical skills remain unclear. Given heterogeneity in study designs and outcome measures, these findings should be interpreted cautiously. These considerations highlight the need for further research to support integrating this innovative approach into medical curricula.

Citations

Citations to this article as recorded by  
  • The Effect of Infection Control Session on Nursing Students' Knowledge and Compliance with Standard Precautions at Hassan College of Nursing, Swat
    Ahmad Ullah, Hammad Ullah Khan, Zia Ullah Khan, Numan Khan, Shah Hussain
    medtigo Journal of Medicine.2026;[Epub]     CrossRef
  • Empowering Surgical Training through Artificial Intelligence: A Cross-Sectional Study on Residents’ Acceptance and Perceived Usefulness of AI-Based Simulation
    Sami Ur Rahman, Kulsoom Nadir, Muhammad Ilyas, Mehar Nigar, Anwar Khan, Abdur Rahman, Shah Hussain
    medtigo Journal of Medicine.2026;[Epub]     CrossRef
  • Advances in evaluating and delivering nontechnical skills training: The use of simulation, robotics, artificial intelligence and virtual reality
    Ravanth Baskaran, Aditya Singh, Bhaskar Kumar Somani
    Current Opinion in Urology.2026; 36(5): 485.     CrossRef
  • La formación en medicina interna en la era de la inteligencia artificial
    J. García-Alegría, D. Ruiz-Hidalgo, M. Rodríguez-Carballeira
    Revista Clínica Española.2026; : 502609.     CrossRef
  • Effectiveness of a “5E+AI” teaching model on science communication in medical students: a contribution to global health literacy
    Huiru Dai, Minling Liu, Tingwei Li, Jiancheng Wang, Shuo Fang
    Frontiers in Public Health.2026;[Epub]     CrossRef
  • Training in internal medicine in the age of artificial intelligence
    J. García-Alegría, D. Ruiz-Hidalgo, M. Rodríguez-Carballeira
    Revista Clínica Española (English Edition).2026; : 502609.     CrossRef
Performance of large language models in medical licensing examinations: a systematic review and meta-analysis  
Haniyeh Nouri, Abdollah Mahdavi, Ali Abedi, Alireza Mohammadnia, Mahnaz Hamedan, Masoud Amanzadeh
J Educ Eval Health Prof. 2025;22:36.   Published online November 18, 2025
DOI: https://doi.org/10.3352/jeehp.2025.22.36
  • 4,919 View
  • 258 Download
  • 5 Web of Science
  • 10 Crossref
AbstractAbstract PDFSupplementary Material
Purpose
This study systematically evaluates and compares the performance of large language models (LLMs) in answering medical licensing examination questions. By conducting subgroup analyses based on language, question format, and model type, this meta-analysis aims to provide a comprehensive overview of LLM capabilities in medical education and clinical decision-making.
Methods
This systematic review, registered in PROSPERO and following PRISMA (Preferred Reporting Items for Systematic Reviews and Meta-Analyses) guidelines, searched MEDLINE (PubMed), Scopus, and Web of Science for relevant articles published up to February 1, 2025. The search strategy included Medical Subject Headings (MeSH) terms and keywords related to (“ChatGPT” OR “GPT” OR “LLM variants”) AND (“medical licensing exam*” OR “medical exam*” OR “medical education” OR “radiology exam*”). Eligible studies evaluated LLM accuracy on medical licensing examination questions. Pooled accuracy was estimated using a random-effects model, with subgroup analyses by LLM type, language, and question format. Publication bias was assessed using Egger’s regression test.
Results
This systematic review identified 2,404 studies. After removing duplicates and excluding irrelevant articles through title and abstract screening, 36 studies were included after full-text review. The pooled accuracy was 72% (95% confidence interval, 70.0% to 75.0%) with high heterogeneity (I2=99%, P<0.001). Among LLMs, GPT-4 achieved the highest accuracy (81%), followed by Bing (79%), Claude (74%), Gemini/Bard (70%), and GPT-3.5 (60%) (P=0.001). Performance differences across languages (range, 62% in Polish to 77% in German) were not statistically significant (P=0.170).
Conclusion
LLMs, particularly GPT-4, can match or exceed medical students’ examination performance and may serve as supportive educational tools. However, due to variability and the risk of errors, they should be used cautiously as complements rather than replacements for traditional learning methods.

Citations

Citations to this article as recorded by  
  • Effective prompt design for large language models in clinical practice
    Steven Callens
    Acta Clinica Belgica.2026; 81(2): 118.     CrossRef
  • Examination of Gemini's ability to answer anatomical questions: An overview
    D. Chytas, G. Noussios, M.-K. Kaseta, D. Chrysikos, A.V. Vasiliadis, C. Lyrtzis, T. Troupis
    Morphologie.2026; 110(369): 101118.     CrossRef
  • Evaluación comparativa de modelos de inteligencia artificial de última generación frente a psiquiatras humanos en el examen nacional de subespecialidad en Perú: un estudio transversal
    Javier A. Flores-Cohaila, Jeff Huarcaya-Victoria, Cesar Copaja-Corzo
    Educación Médica.2026; 27(3): 101179.     CrossRef
  • ChatGPT vs Claude: Scoping Review with ☸️SAIMSARA

    SAIMSARA Journal.2026;[Epub]     CrossRef
  • PeruMedQA: A Stress Evaluation Using Ten Large Language Models to Answer Medical Exams
    Rodrigo M. Carrillo-Larco
    Medical Science Educator.2026; 36(3): 1091.     CrossRef
  • Evaluating the accuracy and communication quality of large language models in Ewing sarcoma: a comparative analysis of ChatGPT, Claude, Gemini, DeepSeek, and Grok
    Cihan Ünyılmaz
    Frontiers in Pediatrics.2026;[Epub]     CrossRef
  • The Promises and Perils of Clinical Decision Support Artificial Intelligence
    Jorge Cervantes, Bhavya Vashi
    The Clinical Teacher.2026;[Epub]     CrossRef
  • NASA TLX workload profiles of multimodal artificial intelligence models and dental students during objective structured practical dental examinations
    Sanaa N. Al-Haj Ali, Ra’fat I. Farah
    Discover Education.2026;[Epub]     CrossRef
  • Artificial Intelligence and Sleep
    Logan Douglas Schneider, John Hernandez, Conor Heneghan
    Neurologic Clinics.2026;[Epub]     CrossRef
  • Generative artificial intelligence in medical education: from knowledge assessment to clinical reasoning and professional competence
    Renxian Xie, Beien Zhang, Lifeng Xiao
    Frontiers in Medicine.2026;[Epub]     CrossRef
Prompt engineering for single-best-answer multiple-choice questions in licensing examinations: a narrative review with a case study involving the Korean Medical Licensing Examination  
Bokyoung Kim, Junseok Kang, Min-Young Kim, Jihyun Ahn
J Educ Eval Health Prof. 2025;22:34.   Published online October 27, 2025
DOI: https://doi.org/10.3352/jeehp.2025.22.34
  • 2,966 View
  • 245 Download
  • 1 Crossref
AbstractAbstract PDFSupplementary Material
The emergence of large language models (LLMs) has generated growing interest in their potential applications for medical assessment and item development. This practice-oriented narrative review examines the potential of LLMs, particularly ChatGPT, for generating and validating single-best-answer multiple-choice questions in health professions licensing examinations, using a Korean Medical Licensing Examination (KMLE)-focused case perspective. We frame LLMs as human-in-the-loop tools rather than replacements for high-stakes testing. Recent applications of LLMs in assessment were reviewed, including prompting strategies such as few-shot, multi-stage, and chain-of-thought methods, as well as retrieval-augmented generation (RAG) to align outputs with exam blueprints. Approaches to enforcing formatting rules, checklist-based self-validation, and iterative refinement were analyzed for their role in supporting item development. Findings indicate that LLMs can perform near passing thresholds on high-stakes exams and assist with grading and feedback tasks. Prompt engineering enhances structural fidelity and clinical plausibility, while human oversight remains critical for accuracy, cultural appropriateness, and psychometric defensibility. The emerging multimodal generation of images, audio, and video suggests the feasibility of new item formats, provided robust validation safeguards are implemented. The most effective approach is a human-in-the-loop workflow that leverages artificial intelligence efficiency while embedding expert judgment, psychometric evaluation, and ethical governance. This practice-oriented roadmap—integrating strategic prompt selection, RAG-based blueprint alignment, rigorous validation gates, and KMLE-specific formatting—offers an implementable and methodologically defensible approach for licensing examinations.

Citations

Citations to this article as recorded by  
  • Generative artificial intelligence and large language models in competency-based medical education: applications, challenges, and future directions
    In Hwa Jeong, Heeyoung Kim, Hyunyong Hwang
    Kosin Medical Journal.2026; 41(2): 114.     CrossRef
Research articles
Performance of GPT-4o and o1-Pro on United Kingdom Medical Licensing Assessment-style items: a comparative study  
Behrad Vakili, Aadam Ahmad, Mahsa Zolfaghari
J Educ Eval Health Prof. 2025;22:30.   Published online October 10, 2025
DOI: https://doi.org/10.3352/jeehp.2025.22.30
  • 3,207 View
  • 254 Download
AbstractAbstract PDFSupplementary Material
Purpose
Large language models (LLMs) such as ChatGPT, and their potential to support autonomous learning for licensing exams like the UK Medical Licensing Assessment (UKMLA), are of growing interest. However, empirical evaluations of artificial intelligence (AI) performance against the UKMLA standard remain limited.
Methods
We evaluated the performance of 2 recent ChatGPT versions, GPT-4o and o1-Pro, on a curated set of 374 UKMLA-style single-best-answer items spanning diverse medical specialties. Statistical comparisons using McNemar’s test assessed the significance of differences between the 2 models. Specialties were analyzed to identify domain-specific variation. In addition, 20 image-based items were evaluated.
Results
GPT-4o achieved an accuracy of 88.8%, while o1-Pro achieved 93.0%. McNemar’s test revealed a statistically significant difference in favor of o1-Pro. Across specialties, both models demonstrated excellent performance in surgery, psychiatry, and infectious diseases. Notable differences arose in dermatology, respiratory medicine, and imaging, where o1-Pro consistently outperformed GPT-4o. Nevertheless, isolated weaknesses in general practice were observed. The analysis of image-based items showed 75% accuracy for GPT-4o and 90% for o1-Pro (P=0.25).
Conclusion
ChatGPT shows strong potential as an adjunct learning tool for UKMLA preparation, with both models achieving scores above the calculated pass mark. This underscores the promise of advanced AI models in medical education. However, specialty-specific inconsistencies suggest AI tools should complement, rather than replace, traditional study methods.
Performance of ChatGPT-4 on the French Board of Plastic Reconstructive and Aesthetic Surgery written exam: a descriptive study
Emma Dejean-Bouyer, Anoujat Kanlagna, François Thuau, Pierre Perrot, Ugo Lancien
J Educ Eval Health Prof. 2025;22:27.   Published online September 30, 2025
DOI: https://doi.org/10.3352/jeehp.2025.22.27
  • 2,048 View
  • 176 Download
AbstractAbstract PDFSupplementary Material
Purpose
This study aims to evaluate the performance of Chat Generative Pre-Trained Transformer 4 (ChatGPT-4) on the French Board of Plastic, Reconstructive, and Aesthetic Surgery written examination and to assess its role as a supplementary resource in helping residents prepare for the qualification examination in plastic surgery.
Methods
This descriptive study evaluated ChatGPT-4’s performance on 213 items from the October 2024 French Board of Plastic, Reconstructive, and Aesthetic Surgery written examination. Responses were assessed for accuracy, logical reasoning, internal and external information use, and were categorized for fallacies by independent reviewers. Statistical analyses included chi-square tests and Fisher’s exact test for significance.
Results
ChatGPT-4 answered all questions across the 10 modules, achieving an overall accuracy rate of 77.5%. The model applied logical reasoning in 98.1% of the questions, utilized internal information in 94.4%, and incorporated external information in 91.1%.
Conclusion
ChatGPT-4 performs satisfactorily on the French Board of Plastic, Reconstructive, and Aesthetic Surgery written examination. Its accuracy met the minimum passing standards for the exam. While responses generally align with expected knowledge, careful verification remains necessary, particularly for questions involving image interpretation. As artificial intelligence continues to evolve, ChatGPT-4 is expected to become an increasingly reliable tool for medical education. At present, it remains a valuable resource for assisting plastic surgery residents in their training.
Comparing generative artificial intelligence platforms and nursing student performance on a women’s health nursing examination in Korea: a Rasch model approach  
Eun Jeong Ko, Tae Kyung Lee, Geum Hee Jeong
J Educ Eval Health Prof. 2025;22:23.   Published online September 5, 2025
DOI: https://doi.org/10.3352/jeehp.2025.22.23
  • 2,735 View
  • 243 Download
  • 2 Web of Science
  • 1 Crossref
AbstractAbstract PDFSupplementary Material
Purpose
This psychometric study aimed to compare the ability parameter estimates of generative artificial intelligence (AI) platforms with those of nursing students on a 50-item women’s health nursing examination at Hallym University, Korea, using the Rasch model. It also sought to estimate item difficulty parameters and evaluate AI performance across varying difficulty levels.
Methods
The exam, consisting of 39 multiple-choice items and 11 true/false items, was administered to 111 fourth-year nursing students in June 2023. In December 2024, 6 generative AI platforms (GPT-4o, ChatGPT free version, Claude.ai, Clova X, Mistral.ai, Google Gemini) completed the same items. The responses were analyzed using the Rasch model to estimate the ability and difficulty parameters. Unidimensionality was verified by the Dimensionality Evaluation to Enumerate Contributing Traits (DETECT), and analyses were conducted using the R packages irtQ and TAM.
Results
The items satisfied unidimensionality (DETECT=–0.16). Item difficulty parameter estimates ranged from –3.87 to 1.96 logits (mean=–0.61), with a mean difficulty index of 0.79. Examinees’ ability parameter estimates ranged from –0.71 to 3.15 logits (mean=1.17). GPT-4o, ChatGPT free version, and Claude.ai outperformed the median student ability (1.09 logits), scoring 2.68, 2.34, and 2.34, respectively, while Clova X, Mistral.ai, and Google Gemini exhibited lower scores (0.20, –0.12, 0.80). The test information curve peaked below θ=0, indicating suitability for examinees with low to average ability.
Conclusion
Advanced generative AI platforms approximated the performance of high-performing students, but outcomes varied. The Rasch model effectively evaluated AI competency, supporting its potential utility for future AI performance assessments in nursing education.

Citations

Citations to this article as recorded by  
  • Eroding scholarly integrity: Confronting the misuse of generative AI in nursing education
    Kechi Iheduru-Anderson
    Nurse Education Today.2026; 164: 107145.     CrossRef
Comparison between GPT-4 and human raters in grading pharmacy students’ exam responses in Malaysia: a cross-sectional study
Wuan Shuen Yap, Pui San Saw, Li Ling Yeap, Shaun Wen Huey Lee, Wei Jin Wong, Ronald Fook Seng Lee
J Educ Eval Health Prof. 2025;22:20.   Published online July 28, 2025
DOI: https://doi.org/10.3352/jeehp.2025.22.20
  • 5,159 View
  • 321 Download
  • 1 Web of Science
AbstractAbstract PDFSupplementary Material
Purpose
Manual grading is time-consuming and prone to inconsistencies, prompting the exploration of generative artificial intelligence tools such as GPT-4 to enhance efficiency and reliability. This study investigated GPT-4’s potential in grading pharmacy students’ exam responses, focusing on the impact of optimized prompts. Specifically, it evaluated the alignment between GPT-4 and human raters, assessed GPT-4’s consistency over time, and determined its error rates in grading pharmacy students’ exam responses.
Methods
We conducted a comparative study using past exam responses graded by university-trained raters and by GPT-4. Responses were randomized before evaluation by GPT-4, accessed via a Plus account between April and September 2024. Prompt optimization was performed on 16 responses, followed by evaluation of 3 prompt delivery methods. We then applied the optimized approach across 4 item types. Intraclass correlation coefficients and error analyses were used to assess consistency and agreement between GPT-4 and human ratings.
Results
GPT-4’s ratings aligned reasonably well with human raters, demonstrating moderate to excellent reliability (intraclass correlation coefficient=0.617–0.933), depending on item type and the optimized prompt. When stratified by grade bands, GPT-4 was less consistent in marking high-scoring responses (Z=–5.71–4.62, P<0.001). Overall, despite achieving substantial alignment with human raters in many cases, discrepancies across item types and a tendency to commit basic errors necessitate continued educator involvement to ensure grading accuracy.
Conclusion
With optimized prompts, GPT-4 shows promise as a supportive tool for grading pharmacy students’ exam responses, particularly for objective tasks. However, its limitations—including errors and variability in grading high-scoring responses—require ongoing human oversight. Future research should explore advanced generative artificial intelligence models and broader assessment formats to further enhance grading reliability.
Performance of large language models on Thailand’s national medical licensing examination: a cross-sectional study  
Prut Saowaprut, Romen Samuel Wabina, Junwei Yang, Lertboon Siriwat
J Educ Eval Health Prof. 2025;22:16.   Published online May 12, 2025
DOI: https://doi.org/10.3352/jeehp.2025.22.16
  • 7,625 View
  • 376 Download
  • 6 Web of Science
  • 5 Crossref
AbstractAbstract PDFSupplementary Material
Purpose
This study aimed to evaluate the feasibility of general-purpose large language models (LLMs) in addressing inequities in medical licensure exam preparation for Thailand’s National Medical Licensing Examination (ThaiNLE), which currently lacks standardized public study materials.
Methods
We assessed 4 multi-modal LLMs (GPT-4, Claude 3 Opus, Gemini 1.0/1.5 Pro) using a 304-question ThaiNLE Step 1 mock examination (10.2% image-based), applying deterministic API configurations and 5 inference repetitions per model. Performance was measured via micro- and macro-accuracy metrics compared against historical passing thresholds.
Results
All models exceeded passing scores, with GPT-4 achieving the highest accuracy (88.9%; 95% confidence interval, 88.7–89.1), surpassing Thailand’s national average by more than 2 standard deviations. Claude 3.5 Sonnet (80.1%) and Gemini 1.5 Pro (72.8%) followed hierarchically. Models demonstrated robustness across 17 of 20 medical domains, but variability was noted in genetics (74.0%) and cardiovascular topics (58.3%). While models demonstrated proficiency with images (Gemini 1.0 Pro: +9.9% vs. text), text-only accuracy remained superior (GPT-4o: 90.0% vs. 82.6%).
Conclusion
General-purpose LLMs show promise as equitable preparatory tools for ThaiNLE Step 1. However, domain-specific knowledge gaps and inconsistent multi-modal integration warrant refinement before clinical deployment.

Citations

Citations to this article as recorded by  
  • Is artificial intelligence getting better at anatomy? A two‐year review of ChatGPT's free public versions
    Bahattin Paslı, Ceren Günenç Beşer
    Anatomical Sciences Education.2026; 19(8): 1279.     CrossRef
  • The performance of ChatGPT and other large language models on multiple‐choice questions in biomedical disciplines: A meta‐analysis
    Colleen M. Cheverko, Volodymyr Mavrych, Olena Bolgova, Fathima Raahima Riyas Mohamed, Jennifer Westrick, Lorena Juarez, Emily Rush, Kathryn A. Solka, Alison F. Doubleday, Jessica N. Byram, Robert Becker, Victoria Gomez, Brenda K. Anak Ganeng, Leslie A. Ho
    Anatomical Sciences Education.2026; 19(9): 1556.     CrossRef
  • Performance of GPT-4o and o1-Pro on United Kingdom Medical Licensing Assessment-style items: a comparative study
    Behrad Vakili, Aadam Ahmad, Mahsa Zolfaghari
    Journal of Educational Evaluation for Health Professions.2025; 22: 30.     CrossRef
  • Large Language Models for the National Radiological Technologist Licensure Examination in Japan: Cross-Sectional Comparative Benchmarking and Evaluation of Model-Generated Items Study
    Toshimune Ito, Toru Ishibashi, Tatsuya Hayashi, Shinya Kojima, Kazumi Sogabe
    JMIR Medical Education.2025; 11: e81807.     CrossRef
  • Technologies, opportunities, challenges, and future directions for integrating generative artificial intelligence into medical education: a narrative review
    Junseok Kang, Jihyun Ahn
    Ewha Medical Journal.2025; 48(4): e53.     CrossRef
Educational/Faculty development material
The role of large language models in the peer-review process: opportunities and challenges for medical journal reviewers and editors  
Jisoo Lee, Jieun Lee, Jeong-Ju Yoo
J Educ Eval Health Prof. 2025;22:4.   Published online January 16, 2025
DOI: https://doi.org/10.3352/jeehp.2025.22.4
  • 14,897 View
  • 563 Download
  • 23 Web of Science
  • 27 Crossref
AbstractAbstract PDFSupplementary Material
The peer review process ensures the integrity of scientific research. This is particularly important in the medical field, where research findings directly impact patient care. However, the rapid growth of publications has strained reviewers, causing delays and potential declines in quality. Generative artificial intelligence, especially large language models (LLMs) such as ChatGPT, may assist researchers with efficient, high-quality reviews. This review explores the integration of LLMs into peer review, highlighting their strengths in linguistic tasks and challenges in assessing scientific validity, particularly in clinical medicine. Key points for integration include initial screening, reviewer matching, feedback support, and language review. However, implementing LLMs for these purposes will necessitate addressing biases, privacy concerns, and data confidentiality. We recommend using LLMs as complementary tools under clear guidelines to support, not replace, human expertise in maintaining rigorous peer review standards.

Citations

Citations to this article as recorded by  
  • ChatGPT-4.0 as a Tool for Automated Review of Ethics and Transparency in Biomedical Literature
    Bohdana Doskaliuk, Birzhan Seiil, Ainur Qumar
    Journal of Korean Medical Science.2026;[Epub]     CrossRef
  • A Cross‐Disciplinary Analysis of AI Policies in Academic Peer Review
    Zhongshi Wang, Mengyue Gong
    Learned Publishing.2026;[Epub]     CrossRef
  • The role of artificial intelligence in shaping dentistry through advancement in data acquisition, clinical practice, education, and research
    Franklin R. Tay, Reid Loveless, Theodore D. Ravenel
    Dental Research.2026; 1(1): 100005.     CrossRef
  • A proof-of-concept study on the use of large language models for assessing research methodology in neuroimaging
    Brock Pluimer, Apeksha Sridhar, Ishtiaq Mawla, Helen Mengxuan Wu, Roshni Lulla, Sarah Hennessy, Patrick Sadil, Rishab Iyer, Eric Ichesco, Anson Kairys, Max Egan, Jonas Kaplan, Richard E. Harris
    Neuroscience Informatics.2026; 6(1): 100262.     CrossRef
  • BIOTECNOLOGIA APLICADA À BIOECONOMIA AMAZÔNICA: POTENCIAL E DESAFIOS CIENTÍFICOS
    Andre de Oliveira Melo, Ágata Chris Gonzales Diaz
    ARACÊ .2026; 8(2): e12010.     CrossRef
  • Human writing and machine patterns: analyzing a decade of convergence
    Eunsuk Chang
    Scientometrics.2026; 131(3): 1635.     CrossRef
  • How editors perceive the use of generative artificial intelligence in writing academic papers: a narrative review
    Sun Huh
    Journal of the Korean Medical Association.2026; 69(2): 111.     CrossRef
  • Automated Assessment of Method Reporting in Obstetrics and Gynecology: A Pilot Study Using ChatGPT 5.0
    Gozde Miray Yilmaz, Serdar Aykut, Tunahan Ates
    The Journal of Obstetrics and Gynecology of India.2026;[Epub]     CrossRef
  • Artificial intelligence in manuscript peer review: Opportunities, risks, and the continuing role of human judgement
    Kaushik Bhattacharya, Surajit Bhattacharya, Dhananjaya Sharma, Michael Cotton
    Tropical Doctor.2026; 56(3): 449.     CrossRef
  • Artificial intelligence in scholarly peer review: a scoping review of applications, risks, and governance challenges
    Ali Nabavi, Farima Safari, Abdel Hadi Shmoury, Salam Tabet, Camilo Perdomo-Luna, Leo Anthony Celi
    International Journal of Medical Informatics.2026; 214: 106418.     CrossRef
  • Artificial Intelligence and Peer Review: Preserving Integrity in the Pursuit of Efficiency
    José de Bessa Jr., Cristiano Mendes Gomes
    International braz j urol.2026;[Epub]     CrossRef
  • Evaluating large language models for abstract evaluation tasks: an empirical study
    Yinuo Liu, Emre Sezgin, Eric A. Youngstrom
    Frontiers in Research Metrics and Analytics.2026;[Epub]     CrossRef
  • Health Governance Review Volume 31, Issue 2: The role of peer review
    Irina Ibragimova
    International Journal of Health Governance.2026; 31(2): 149.     CrossRef
  • Artificial Intelligence in Academic Publishing: Important Dynamic Considerations For Authors, Reviewers, Editors, and Publishers
    John G. Augoustides
    Journal of Cardiothoracic and Vascular Anesthesia.2026; 40(9): 2643.     CrossRef
  • AI in peer review: the elephant in the editorial room
    Giusy Rita Maria La Rosa, Mona Nasser
    Evidence-Based Dentistry.2026; 27(2): 23.     CrossRef
  • Exploiting large language models in peer review: indirect prompt injection attacks and integrity probes
    Federico Torrielli, Stefano Locci, Amon Rapp, Luigi Di Caro
    Scientometrics.2026; 131(7): 4889.     CrossRef
  • The Role of Artificial Intelligence in the Lifecycle of Scientific Manuscripts: Authoring, Reviewing, and Editorial Selection
    Jose L. Domingo
    Qeios.2026;[Epub]     CrossRef
  • Can large language models provide high-quality desk review decisions in an orthopaedic surgery journal? A concordance study comparing three AI models to human editorial decisions
    Elise Lupon, Quentin Bérard, Henri Migaud, Philippe Clavert, Grégoire Micicoi
    Orthopaedics & Traumatology: Surgery & Research.2026; : 104801.     CrossRef
  • AI will reorganize science. Will research remain a human enterprise?
    Michael E. Hochberg, Peter H. Thrall
    Proceedings of the National Academy of Sciences.2026;[Epub]     CrossRef
  • Recommendations for stakeholders in journal publishing regarding the use and development of artificial intelligence platforms
    Sang-Jun Kim
    Science Editing.2026; 13(2): 188.     CrossRef
  • Presence and content of policies on the use of generative artificial intelligence in nursing journals indexed in PubMed: a descriptive study
    Geum Hee Jeong, Eun Jeong Ko
    Science Editing.2026; 13(2): 103.     CrossRef
  • Comments: Can Statistics and AI Technologies Help Our Troubled Review System?
    Xiao-Li Meng
    Journal of the American Statistical Association.2026; 121(554): 855.     CrossRef
  • Les grands modèles de langage peuvent-ils fournir des décisions de revue éditoriale de haute qualité dans un journal de chirurgie orthopédique ? Une étude de concordance comparant trois modèles d’IA aux décisions éditoriales humaines
    Élise Lupon, Quentin Bérard, Henri Migaud, Philippe Clavert, Grégoire Micicoi
    Revue de Chirurgie Orthopédique et Traumatologique.2026;[Epub]     CrossRef
  • Large Language Models as Peer Reviewers: Prompt Sensitivity and Model-Dependent Reproducibility
    Sukru Mehmet Erturk, Mustafa Durmaz
    Academic Radiology.2026; 33(9): 3642.     CrossRef
  • A reviewer identification using machine learning methods
    Denis Yu. Bolshakov
    Science Editor and Publisher.2025; 10(1): 32.     CrossRef
  • Beyond the Review: The Editorial Duty to Uphold Professional Conduct
    Stephen A. Bustin
    Publications.2025; 13(4): 48.     CrossRef
  • Role of Medical Editors in the Age of Generative Artificial Intelligence
    Sun Huh
    Healthcare Informatics Research.2025; 31(4): 317.     CrossRef
Research article
Effectiveness of ChatGPT-4o in developing continuing professional development plans for graduate radiographers: a descriptive study  
Minh Chau, Elio Stefan Arruzza, Kelly Spuur
J Educ Eval Health Prof. 2024;21:34.   Published online November 18, 2024
DOI: https://doi.org/10.3352/jeehp.2024.21.34
  • 5,899 View
  • 267 Download
  • 6 Web of Science
  • 8 Crossref
AbstractAbstract PDFSupplementary Material
Purpose
This study evaluates the use of ChatGPT-4o in creating tailored continuing professional development (CPD) plans for radiography students, addressing the challenge of aligning CPD with Medical Radiation Practice Board of Australia (MRPBA) requirements. We hypothesized that ChatGPT-4o could support students in CPD planning while meeting regulatory standards.
Methods
A descriptive, experimental design was used to generate 3 unique CPD plans using ChatGPT-4o, each tailored to hypothetical graduate radiographers in varied clinical settings. Each plan followed MRPBA guidelines, focusing on computed tomography specialization by the second year. Three MRPBA-registered academics assessed the plans using criteria of appropriateness, timeliness, relevance, reflection, and completeness from October 2024 to November 2024. Ratings underwent analysis using the Friedman test and intraclass correlation coefficient (ICC) to measure consistency among evaluators.
Results
ChatGPT-4o generated CPD plans generally adhered to regulatory standards across scenarios. The Friedman test indicated no significant differences among raters (P=0.420, 0.761, and 0.807 for each scenario), suggesting consistent scores within scenarios. However, ICC values were low (–0.96, 0.41, and 0.058 for scenarios 1, 2, and 3), revealing variability among raters, particularly in timeliness and completeness criteria, suggesting limitations in the ChatGPT-4o’s ability to address individualized and context-specific needs.
Conclusion
ChatGPT-4o demonstrates the potential to ease the cognitive demands of CPD planning, offering structured support in CPD development. However, human oversight remains essential to ensure plans are contextually relevant and deeply reflective. Future research should focus on enhancing artificial intelligence’s personalization for CPD evaluation, highlighting ChatGPT-4o’s potential and limitations as a tool in professional education.

Citations

Citations to this article as recorded by  
  • Shaping the Future of Radiography Education: Lessons From ChatGPT and Generative AI
    Minh T. Chau, Haydn Kerr, Clare L. Singh, Bismark Ofori‐Manteaw, Elio Arruzza, Kelly Bentley‐Spuur
    Journal of Medical Radiation Sciences.2026; 73(3): 327.     CrossRef
  • Applied insights for using Generative Artificial Intelligence in Faculty Development in Health Professions Education
    Melchor Sánchez-Mendiola, Megan Anakin, Ardi Findyartini, Rachel Levine, Ana Da Silva, Farhan Saeed Vakani
    MedEdPublish.2026; 15: 279.     CrossRef
  • Halted medical education and medical residents’ training in Korea, journal metrics, and appreciation to reviewers and volunteers
    Sun Huh
    Journal of Educational Evaluation for Health Professions.2025; 22: 1.     CrossRef
  • The ‘Negotiator’: Assessing artificial intelligence (AI) interview preparation for graduate radiographers
    M. Chau, E. Arruzza, C.L. Singh
    Journal of Medical Imaging and Radiation Sciences.2025; 56(5): 101982.     CrossRef
  • ‘Bill’: An artificial intelligence (AI) clinical scenario coach for medical radiation science education
    M. Chau, G. Higgins, E. Arruzza, C.L. Singh
    Radiography.2025; 31(5): 103002.     CrossRef
  • Exploring ChatGPT-4o-generated reflections: Alignment with professional standards in diagnostic radiography: A pilot experiment
    C Nabasenja, M Chau, E Green
    Journal of Medical Imaging and Radiation Sciences.2025; 56(6): 102082.     CrossRef
  • Applied insights for using Generative Artificial Intelligence in Faculty Development in Health Professions Education
    Melchor Sánchez-Mendiola, Megan Anakin, Ardi Findyartini, Rachel Levine, Ana Da Silva, Farhan Saeed Vakani
    MedEdPublish.2025; 15: 279.     CrossRef
  • A research roadmap for AI opportunities in student assessment for medical education
    Morteza Rezaei-Zadeh, Magdalena Cerbin-Koczorowska
    BMC Medical Education.2025;[Epub]     CrossRef
Educational/Faculty development material
The performance of ChatGPT-4.0o in medical imaging evaluation: a cross-sectional study  
Elio Stefan Arruzza, Carla Marie Evangelista, Minh Chau
J Educ Eval Health Prof. 2024;21:29.   Published online October 31, 2024
DOI: https://doi.org/10.3352/jeehp.2024.21.29
  • 7,747 View
  • 326 Download
  • 14 Web of Science
  • 18 Crossref
AbstractAbstract PDFSupplementary Material
This study investigated the performance of ChatGPT-4.0o in evaluating the quality of positioning in radiographic images. Thirty radiographs depicting a variety of knee, elbow, ankle, hand, pelvis, and shoulder projections were produced using anthropomorphic phantoms and uploaded to ChatGPT-4.0o. The model was prompted to provide a solution to identify any positioning errors with justification and offer improvements. A panel of radiographers assessed the solutions for radiographic quality based on established positioning criteria, with a grading scale of 1–5. In only 20% of projections, ChatGPT-4.0o correctly recognized all errors with justifications and offered correct suggestions for improvement. The most commonly occurring score was 3 (9 cases, 30%), wherein the model recognized at least 1 specific error and provided a correct improvement. The mean score was 2.9. Overall, low accuracy was demonstrated, with most projections receiving only partially correct solutions. The findings reinforce the importance of robust radiography education and clinical experience.

Citations

Citations to this article as recorded by  
  • Artificial Intelligence Versus Human Dental Expertise in Diagnosing Periapical Pathosis on Periapical Radiographs: A Multicenter Study
    Fatma E. A. Hassanein, Radwa R. Hussein, Mohamed Riad Elgarhy, Shaymaa Mohamed Maher, Ahmed Hassen, Sherif Heidar, Marwa Ezz El Arab, Amr Edress, Asmaa Abou-Bakr, Mohamed Mekhemar
    Bioengineering.2026; 13(2): 232.     CrossRef
  • Shaping the Future of Radiography Education: Lessons From ChatGPT and Generative AI
    Minh T. Chau, Haydn Kerr, Clare L. Singh, Bismark Ofori‐Manteaw, Elio Arruzza, Kelly Bentley‐Spuur
    Journal of Medical Radiation Sciences.2026; 73(3): 327.     CrossRef
  • The Convergence of ChatGPT‐4 and Nanotechnology for Transforming the Future of Radiological Imaging: A Comprehensive Narrative Review
    Biruk Demisse Ayalew, Maria Qadri, Muhammad Areeb Ul Haq, Lintha Zafar Khattak, Aayat Kashif, Ali Dheyaa Marsool, Nuradin Abdi Ali, Samra Solomon Wondemu, Temesgen Mamo Sharew, Getnet Bimer Kelemu, Michael Teklehaimanot Abera, Alaa Ragab Hani
    iRADIOLOGY.2026; 4(2): 137.     CrossRef
  • Responsible AI in healthcare: Mitigating hallucinations and enhancing multimodal fusion - based reasoning in medical imaging
    Saeed Iqbal, Xiaopin Zhong, Muhammad Attique Khan, Zongze Wu, Nouf Abdullah Almujally, Weixiang Liu, Erik Cambria, Amir Hussain
    Information Fusion.2026; 136: 104483.     CrossRef
  • The Accuracy of ChatGPT in Classifying Lumbar Spondylolisthesis and Compression Fractures
    Justin Chung, Rowen Lin, Evan Dunn, Grace Kim, Caleb Choi, Kevin Mo, William Fang, Daniel Lee
    Journal of the American Osteopathic Academy of Orthopedics.2026;[Epub]     CrossRef
  • Diagnostic accuracy and repeatability of ChatGPT using textual and radiographic data in reversible pulpitis: a retrospective diagnostic study
    María Llorente de Pedro, Yolanda Freire, Natalia Moneo, Cristina Andreu-Vázquez, Roberto Estévez, Víctor Díaz-Flores García, Ana Suárez
    Frontiers in Bioinformatics.2026;[Epub]     CrossRef
  • A Cautious Integration With AI in the Clinic: A Standardized-Patient Pilot Study of ChatGPT’s Reliability in Hamilton Depression Rating Scale Scoring
    Chun-Hung Chang, Szu-Wei Cheng, Wei-Jen Chen, Chung-Wen Chang, Ting-Hui Liu, Jia-Hau Lee, Sheng-Che Lin, Kuan-Pin Su
    Alpha Psychiatry.2026;[Epub]     CrossRef
  • Evaluating Large Language Models for Burning Mouth Syndrome Diagnosis
    Takayuki Suga, Osamu Uehara, Yoshihiro Abiko, Akira Toyofuku
    Journal of Pain Research.2025; Volume 18: 1387.     CrossRef
  • Evaluating the performance of GPT-3.5, GPT-4, and GPT-4o in the Chinese National Medical Licensing Examination
    Dingyuan Luo, Mengke Liu, Runyuan Yu, Yulian Liu, Wenjun Jiang, Qi Fan, Naifeng Kuang, Qiang Gao, Tao Yin, Zuncheng Zheng
    Scientific Reports.2025;[Epub]     CrossRef
  • The ‘Negotiator’: Assessing artificial intelligence (AI) interview preparation for graduate radiographers
    M. Chau, E. Arruzza, C.L. Singh
    Journal of Medical Imaging and Radiation Sciences.2025; 56(5): 101982.     CrossRef
  • Transforming behavioral intention and academic performance: ChatGPT-4.0 insights through SEM, ANN, and cIPMA analysis
    Fazeelat Aziz, Cai Li, Asad Ullah Khan
    Information Development.2025; 41(3): 933.     CrossRef
  • Can ChatGPT Aid in Musculoskeletal Intervention?
    Mohamed Ashiq Shazahan, Saavi Reddy Pellakuru, Sonal Saran, Shashank Chapala, Sindhura Mettu, Rajesh Botchu
    Journal of Clinical Interventional Radiology ISVIR.2025; 09(03): 148.     CrossRef
  • ‘Bill’: An artificial intelligence (AI) clinical scenario coach for medical radiation science education
    M. Chau, G. Higgins, E. Arruzza, C.L. Singh
    Radiography.2025; 31(5): 103002.     CrossRef
  • Correlates of Trust of Generative Artificial Intelligence Tools Among Patients and Caregivers: A Review of Empirical Research
    Oliver T. Nguyen, Arsalan Ahmad, Douglas A. Wiegmann
    Proceedings of the Human Factors and Ergonomics Society Annual Meeting.2025; 69(1): 1656.     CrossRef
  • The performance of ChatGPT on medical image-based assessments and implications for medical education
    Xiang Yang, Wei Chen
    BMC Medical Education.2025;[Epub]     CrossRef
  • Technologies, opportunities, challenges, and future directions for integrating generative artificial intelligence into medical education: a narrative review
    Junseok Kang, Jihyun Ahn
    Ewha Medical Journal.2025; 48(4): e53.     CrossRef
  • Conversational LLM Chatbot ChatGPT-4 for Colonoscopy Boston Bowel Preparation Scoring: An Artificial Intelligence-to-Head Concordance Analysis
    Raffaele Pellegrino, Alessandro Federico, Antonietta Gerarda Gravina
    Diagnostics.2024; 14(22): 2537.     CrossRef
  • Effectiveness of ChatGPT-4o in developing continuing professional development plans for graduate radiographers: a descriptive study
    Minh Chau, Elio Stefan Arruzza, Kelly Spuur
    Journal of Educational Evaluation for Health Professions.2024; 21: 34.     CrossRef
Research articles
GPT-4o’s competency in answering the simulated written European Board of Interventional Radiology exam compared to a medical student and experts in Germany and its ability to generate exam items on interventional radiology: a descriptive study
Sebastian Ebel, Constantin Ehrengut, Timm Denecke, Holger Gößmann, Anne Bettina Beeskow
J Educ Eval Health Prof. 2024;21:21.   Published online August 20, 2024
DOI: https://doi.org/10.3352/jeehp.2024.21.21
  • 6,793 View
  • 361 Download
  • 17 Web of Science
  • 16 Crossref
AbstractAbstract PDFSupplementary Material
Purpose
This study aimed to determine whether ChatGPT-4o, a generative artificial intelligence (AI) platform, was able to pass a simulated written European Board of Interventional Radiology (EBIR) exam and whether GPT-4o can be used to train medical students and interventional radiologists of different levels of expertise by generating exam items on interventional radiology.
Methods
GPT-4o was asked to answer 370 simulated exam items of the Cardiovascular and Interventional Radiology Society of Europe (CIRSE) for EBIR preparation (CIRSE Prep). Subsequently, GPT-4o was requested to generate exam items on interventional radiology topics at levels of difficulty suitable for medical students and the EBIR exam. Those generated items were answered by 4 participants, including a medical student, a resident, a consultant, and an EBIR holder. The correctly answered items were counted. One investigator checked the answers and items generated by GPT-4o for correctness and relevance. This work was done from April to July 2024.
Results
GPT-4o correctly answered 248 of the 370 CIRSE Prep items (67.0%). For 50 CIRSE Prep items, the medical student answered 46.0%, the resident 42.0%, the consultant 50.0%, and the EBIR holder 74.0% correctly. All participants answered 82.0% to 92.0% of the 50 GPT-4o generated items at the student level correctly. For the 50 GPT-4o items at the EBIR level, the medical student answered 32.0%, the resident 44.0%, the consultant 48.0%, and the EBIR holder 66.0% correctly. All participants could pass the GPT-4o-generated items for the student level; while the EBIR holder could pass the GPT-4o-generated items for the EBIR level. Two items (0.3%) out of 150 generated by the GPT-4o were assessed as implausible.
Conclusion
GPT-4o could pass the simulated written EBIR exam and create exam items of varying difficulty to train medical students and interventional radiologists.

Citations

Citations to this article as recorded by  
  • When AI meets medical assessment: A comparative study of GPT-4O and Claude 3 Opus in China's standardized resident physician examinations
    Wenjie Zhong, Yidan Hu, Ruiqiang Su, Lingcong Xu, Yating Chen, Niezhenghao He, Caiyuan Liu, Ke Xu, Mao Zhao, Wenao Liao, Wei Zhang, Jiang Hu, Fei Wang, Haowen Cui
    Computers in Human Behavior Reports.2026; 21: 100974.     CrossRef
  • The current status and future prospects of artificial intelligence education in residency training
    Hongsen Zhang, Kun Qian, Jing Wang, Chuansheng Zheng
    Frontiers in Education.2026;[Epub]     CrossRef
  • Comparative analysis of multimodal large language models GPT-4o and o1 versus clinicians in clinical case challenge questions: Retrospective cross-sectional study
    Jaewon Jung, Hyunjae Kim, SungA Bae, Jin Young Park
    Medicine.2026; 105(4): e47071.     CrossRef
  • Validity of AI-generated multiple-choice questions in medical education: a systematic review
    Yavuz Selim Kıyak, Abdullah Bedir Kaya, Emre Emekli
    Postgraduate Medical Journal.2026;[Epub]     CrossRef
  • Evaluation of large language models in a national orthopaedic proficiency examination: Implications for health informatics and medical education
    Bünyamin Arı
    Health Informatics Journal.2026;[Epub]     CrossRef
  • Evaluating the performance of ChatGPT in patient consultation and image-based preliminary diagnosis in thyroid eye disease
    Yue Wang, Shuo Yang, Chengcheng Zeng, Yingwei Xie, Ya Shen, Jian Li, Xiao Huang, Ruili Wei, Yuqing Chen
    Frontiers in Medicine.2025;[Epub]     CrossRef
  • Solving Complex Pediatric Surgical Case Studies: A Comparative Analysis of Copilot, ChatGPT-4, and Experienced Pediatric Surgeons' Performance
    Richard Gnatzy, Martin Lacher, Michael Berger, Michael Boettcher, Oliver J. Deffaa, Joachim Kübler, Omid Madadi-Sanjani, Illya Martynov, Steffi Mayer, Mikko P. Pakarinen, Richard Wagner, Tomas Wester, Augusto Zani, Ophelia Aubert
    European Journal of Pediatric Surgery.2025; 35(05): 382.     CrossRef
  • Preliminary assessment of large language models’ performance in answering questions on developmental dysplasia of the hip
    Shiwei Li, Jun Jiang, Xiaodong Yang
    Journal of Children's Orthopaedics.2025; 19(3): 207.     CrossRef
  • AI and Interventional Radiology: A Narrative Review of Reviews on Opportunities, Challenges, and Future Directions
    Andrea Lastrucci, Nicola Iosca, Yannick Wandael, Angelo Barra, Graziano Lepri, Nevio Forini, Renzo Ricci, Vittorio Miele, Daniele Giansanti
    Diagnostics.2025; 15(7): 893.     CrossRef
  • Evaluating the performance of GPT-3.5, GPT-4, and GPT-4o in the Chinese National Medical Licensing Examination
    Dingyuan Luo, Mengke Liu, Runyuan Yu, Yulian Liu, Wenjun Jiang, Qi Fan, Naifeng Kuang, Qiang Gao, Tao Yin, Zuncheng Zheng
    Scientific Reports.2025;[Epub]     CrossRef
  • Evaluating Large Language Models for Preoperative Patient Education in Superior Capsular Reconstruction: Comparative Study of Claude, GPT, and Gemini
    Yukang Liu, Hua Li, Jianfeng Ouyang, Zhaowen Xue, Min Wang, Hebei He, Bin Song, Xiaofei Zheng, Wenyi Gan
    JMIR Perioperative Medicine.2025; 8: e70047.     CrossRef
  • Evaluating ChatGPT's performance across radiology subspecialties: A meta-analysis of board-style examination accuracy and variability
    Dan Nguyen, Grace Hyun J. Kim, Arash Bedayat
    Clinical Imaging.2025; 125: 110551.     CrossRef
  • Performance of ChatGPT-4 on the French Board of Plastic Reconstructive and Aesthetic Surgery written exam: a descriptive study
    Emma Dejean-Bouyer, Anoujat Kanlagna, François Thuau, Pierre Perrot, Ugo Lancien
    Journal of Educational Evaluation for Health Professions.2025; 22: 27.     CrossRef
  • Technologies, opportunities, challenges, and future directions for integrating generative artificial intelligence into medical education: a narrative review
    Junseok Kang, Jihyun Ahn
    Ewha Medical Journal.2025; 48(4): e53.     CrossRef
  • From GPT-3.5 to GPT-4.o: A Leap in AI’s Medical Exam Performance
    Markus Kipp
    Information.2024; 15(9): 543.     CrossRef
  • Performance of ChatGPT and Bard on the medical licensing examinations varies across different cultures: a comparison study
    Yikai Chen, Xiujie Huang, Fangjie Yang, Haiming Lin, Haoyu Lin, Zhuoqun Zheng, Qifeng Liang, Jinhai Zhang, Xinxin Li
    BMC Medical Education.2024;[Epub]     CrossRef
Performance of GPT-3.5 and GPT-4 on standardized urology knowledge assessment items in the United States: a descriptive study  
Max Samuel Yudovich, Elizaveta Makarova, Christian Michael Hague, Jay Dilip Raman
J Educ Eval Health Prof. 2024;21:17.   Published online July 8, 2024
DOI: https://doi.org/10.3352/jeehp.2024.21.17
  • 9,744 View
  • 373 Download
  • 22 Web of Science
  • 23 Crossref
AbstractAbstract PDFSupplementary Material
Purpose
This study aimed to evaluate the performance of Chat Generative Pre-Trained Transformer (ChatGPT) with respect to standardized urology multiple-choice items in the United States.
Methods
In total, 700 multiple-choice urology board exam-style items were submitted to GPT-3.5 and GPT-4, and responses were recorded. Items were categorized based on topic and question complexity (recall, interpretation, and problem-solving). The accuracy of GPT-3.5 and GPT-4 was compared across item types in February 2024.
Results
GPT-4 answered 44.4% of items correctly compared to 30.9% for GPT-3.5 (P<0.00001). GPT-4 (vs. GPT-3.5) had higher accuracy with urologic oncology (43.8% vs. 33.9%, P=0.03), sexual medicine (44.3% vs. 27.8%, P=0.046), and pediatric urology (47.1% vs. 27.1%, P=0.012) items. Endourology (38.0% vs. 25.7%, P=0.15), reconstruction and trauma (29.0% vs. 21.0%, P=0.41), and neurourology (49.0% vs. 33.3%, P=0.11) items did not show significant differences in performance across versions. GPT-4 also outperformed GPT-3.5 with respect to recall (45.9% vs. 27.4%, P<0.00001), interpretation (45.6% vs. 31.5%, P=0.0005), and problem-solving (41.8% vs. 34.5%, P=0.56) type items. This difference was not significant for the higher-complexity items.
Conclusions
ChatGPT performs relatively poorly on standardized multiple-choice urology board exam-style items, with GPT-4 outperforming GPT-3.5. The accuracy was below the proposed minimum passing standards for the American Board of Urology’s Continuing Urologic Certification knowledge reinforcement activity (60%). As artificial intelligence progresses in complexity, ChatGPT may become more capable and accurate with respect to board examination items. For now, its responses should be scrutinized.

Citations

Citations to this article as recorded by  
  • Response to letter to the editor Re: Advancements in large language model accuracy for answering physical medicine and rehabilitation board review questions
    Jason Bitterman, Alexander D'Angelo, Alexandra Holachek, James E. Eubanks
    PM&R.2026; 18(1): 109.     CrossRef
  • Assessing the utility of a natural language processing model in answering common urological questions
    Wyatt MacNevin, John‐David Brown, Nicholas Dawe, Jesse T. R. Spooner, Nicholas R. Paterson, Daniel T. Keefe, David G. Bell
    UroPrecision.2026; 4(2): 98.     CrossRef
  • Artificial Intelligence as a Drug Information Resource: Limitations and Strategies to Optimize in Pharmacy Practice
    Christopher Soujah, Carole Bejjani, Nour Adra, Laura Blackburn
    Hospital Pharmacy.2026; 61(2): 117.     CrossRef
  • Assessing multiple chatGPT versions on novel content in the Taiwan urology board examination: accuracy, speed, and domain-specific performance
    Ho Li, Hui-Kung Ting, Yu-Cing Jhuo, Chin-Li Chen, Chien-Chang Kao, Ming-Hsin Yang, Chih-Wei Tsao, En Meng, Sheng-Tang Wu, Pei-Jhang Chiang
    World Journal of Urology.2026;[Epub]     CrossRef
  • Large Language Models Provide Accurate but Potentially Unsafe Answers to Multimodal Critical Care Medicine Board Review Questions
    Ish Sethi, Sharaf Khan, Patrick G. Lyons, Catherine A. Gao, Yuan Luo, Danielle Miltz, Santiago Tovar, Caitlin ten Lohuis, Derek M. Polly, Sagar B. Dave, Michael Sterling, Juan C. Rojas, Xuan Han, Ankit Sakhuja, Leo A. Celi, Greg S. Martin, Craig M. Cooper
    Critical Care Medicine.2026;[Epub]     CrossRef
  • Evaluating the Performance of ChatGPT4.0 Versus ChatGPT3.5 on the Hand Surgery Self-Assessment Exam: A Comparative Analysis of Performance on Image-Based Questions
    Kiera L Vrindten, Megan Hsu, Yuri Han, Brian Rust, Heili Truumees, Brian M Katt
    Cureus.2025;[Epub]     CrossRef
  • Assessing the performance of large language models (GPT-3.5 and GPT-4) and accurate clinical information for pediatric nephrology
    Nadide Melike Sav
    Pediatric Nephrology.2025; 40(9): 2879.     CrossRef
  • Retrieval-augmented generation enhances large language model performance on the Japanese orthopedic board examination
    Juntaro Maruyama, Satoshi Maki, Takeo Furuya, Yuki Nagashima, Kyota Kitagawa, Yasunori Toki, Shuhei Iwata, Megumi Yazaki, Takaki Kitamura, Sho Gushiken, Yuji Noguchi, Masataka Miura, Masahiro Inoue, Yasuhiro Shiga, Kazuhide Inage, Sumihisa Orita, Seiji Oh
    Journal of Orthopaedic Science.2025; 30(6): 1193.     CrossRef
  • Advancements in large language model accuracy for answering physical medicine and rehabilitation board review questions
    Jason Bitterman, Alexander D'Angelo, Alexandra Holachek, James E. Eubanks
    PM&R.2025; 17(9): 1091.     CrossRef
  • Accuracy of Large Language Models When Answering Clinical Research Questions: Systematic Review and Network Meta-Analysis
    Ling Wang, Jinglin Li, Boyang Zhuang, Shasha Huang, Meilin Fang, Cunze Wang, Wen Li, Mohan Zhang, Shurong Gong
    Journal of Medical Internet Research.2025; 27: e64486.     CrossRef
  • OpenAI o1 Large Language Model Outperforms GPT-4o, Gemini 1.5 Flash, and Human Test Takers on Ophthalmology Board–Style Questions
    Ryan Shean, Tathya Shah, Sina Sobhani, Alan Tang, Ali Setayesh, Kyle Bolo, Van Nguyen, Benjamin Xu
    Ophthalmology Science.2025; 5(6): 100844.     CrossRef
  • A comparative analysis of DeepSeek R1, DeepSeek-R1-Lite, OpenAi o1 Pro, and Grok 3 performance on ophthalmology board-style questions
    Ryan Shean, Tathya Shah, Aditya Pandiarajan, Alan Tang, Kyle Bolo, Van Nguyen, Benjamin Xu
    Scientific Reports.2025;[Epub]     CrossRef
  • The performance of ChatGPT on medical image-based assessments and implications for medical education
    Xiang Yang, Wei Chen
    BMC Medical Education.2025;[Epub]     CrossRef
  • Applications, Challenges, and Prospects of Generative Artificial Intelligence Empowering Medical Education: Scoping Review
    Yuhang Lin, Zhiheng Luo, Zicheng Ye, Nuoxi Zhong, Lijian Zhao, Long Zhang, Xiaolan Li, Zetao Chen, Yijia Chen
    JMIR Medical Education.2025; 11: e71125.     CrossRef
  • ChatGPT’s role in the rapidly evolving hematologic cancer landscape
    Tiffany Nong, Sean Britton, Viralkumar Bhanderi, Justin Taylor
    Future Science OA.2025;[Epub]     CrossRef
  • Performance of ChatGPT-4 on the French Board of Plastic Reconstructive and Aesthetic Surgery written exam: a descriptive study
    Emma Dejean-Bouyer, Anoujat Kanlagna, François Thuau, Pierre Perrot, Ugo Lancien
    Journal of Educational Evaluation for Health Professions.2025; 22: 27.     CrossRef
  • Technologies, opportunities, challenges, and future directions for integrating generative artificial intelligence into medical education: a narrative review
    Junseok Kang, Jihyun Ahn
    Ewha Medical Journal.2025; 48(4): e53.     CrossRef
  • Performance of the ChatGPT-5 Language Model in Solving a Specialty Examination in Balneology and Physical Medicine
    Michalina Loson-Kawalec, Anna Kowalczyk, Dawid Szymanski, Patrycja Dadynska, Aleksander Tabor, Dawid Bartosik , Marta Zerek, Gracjan Sitarek, Bartosz Starzynski, Alina Keska, Bartlomiej Cwikla, Piotr Sawina, Tomasz Dolata, Adrianna Pielech, Maciej Majchrz
    Cureus.2025;[Epub]     CrossRef
  • Diagnostic accuracy and bias in open access and subscription-based large language models for multiple sclerosis and neuromyelitis optica spectrum disorder
    Tom G. Punnen, Kevin S. Shan, Mahi A. Patel, Morgan C. McCreary, Diem H. Tran, Jose R. Santoyo, Katy W. Burgess, Tatum M. Moog, Alexander D. Smith, Darin T. Okuda
    Intelligence-Based Medicine.2025; 12: 100314.     CrossRef
  • Privacy-by-Design Framework for Large Language Model Chatbots in Urology
    Eun Joung Kim, JungYoon Kim
    International Neurourology Journal.2025; 29(Suppl 2): S65.     CrossRef
  • Potential and pitfalls: accuracy versus adequacy of ChatGPT’s performance on surgery shelf examination
    Baylee Brochu, Michael D. Cobler-Lichter, Talia R. Arcieri, Nikita M. Shah, Jessica M. Delamater, Ana M. Reyes, Matthew S. Sussman, Edward B. Lineen, Laurence R. Sands, Vanessa W. Hui, Steven E. Rodgers, Chad M. Thorson
    Global Surgical Education - Journal of the Association for Surgical Education.2025;[Epub]     CrossRef
  • From GPT-3.5 to GPT-4.o: A Leap in AI’s Medical Exam Performance
    Markus Kipp
    Information.2024; 15(9): 543.     CrossRef
  • Artificial Intelligence can Facilitate Application of Risk Stratification Algorithms to Bladder Cancer Patient Case Scenarios
    Max S Yudovich, Ahmad N Alzubaidi, Jay D Raman
    Clinical Medicine Insights: Oncology.2024;[Epub]     CrossRef
Review
Opportunities, challenges, and future directions of large language models, including ChatGPT in medical education: a systematic scoping review  
Xiaojun Xu, Yixiao Chen, Jing Miao
J Educ Eval Health Prof. 2024;21:6.   Published online March 15, 2024
DOI: https://doi.org/10.3352/jeehp.2024.21.6
  • 26,847 View
  • 1,001 Download
  • 110 Web of Science
  • 134 Crossref
AbstractAbstract PDFSupplementary Material
Background
ChatGPT is a large language model (LLM) based on artificial intelligence (AI) capable of responding in multiple languages and generating nuanced and highly complex responses. While ChatGPT holds promising applications in medical education, its limitations and potential risks cannot be ignored.
Methods
A scoping review was conducted for English articles discussing ChatGPT in the context of medical education published after 2022. A literature search was performed using PubMed/MEDLINE, Embase, and Web of Science databases, and information was extracted from the relevant studies that were ultimately included.
Results
ChatGPT exhibits various potential applications in medical education, such as providing personalized learning plans and materials, creating clinical practice simulation scenarios, and assisting in writing articles. However, challenges associated with academic integrity, data accuracy, and potential harm to learning were also highlighted in the literature. The paper emphasizes certain recommendations for using ChatGPT, including the establishment of guidelines. Based on the review, 3 key research areas were proposed: cultivating the ability of medical students to use ChatGPT correctly, integrating ChatGPT into teaching activities and processes, and proposing standards for the use of AI by medical students.
Conclusion
ChatGPT has the potential to transform medical education, but careful consideration is required for its full integration. To harness the full potential of ChatGPT in medical education, attention should not only be given to the capabilities of AI but also to its impact on students and teachers.

Citations

Citations to this article as recorded by  
  • Reliability of ChatGPT-4o in analysing medical data: a test case study on patients at risk for limb amputation
    Liat Toderis, Iris Reychav, Roger McHaney, Bernice Oberman, Chen Speter, Ronen Loebstein
    Health Systems.2026; 15(2): 125.     CrossRef
  • Large language models in education: a systematic review of empirical applications, benefits, and challenges
    Yuhong Shi, Kun Yu, Yifei Dong, Fang Chen
    Computers and Education: Artificial Intelligence.2026; 10: 100529.     CrossRef
  • Evaluating the Insights of ChatGPT, Gemini and Expert Surgeons in Revision Rhinoplasty Consultation
    Shahin Bastaninejad, Samira Alipour, Luiz Carlos Ishida, Mahdieh Mohebbi, Benyamin Mousavi-asl, Farrokh Heidari, Habib Azimi
    Aesthetic Plastic Surgery.2026; 50(11): 3831.     CrossRef
  • Large Language Model Applications in the Algebra Domain: A Systematic Review
    Yajie Song, Yimei Zhang, Doina Precup, Reihaneh Rabbany, Maria Cutumisu
    Technology, Knowledge and Learning.2026; 31(2): 685.     CrossRef
  • Conversational AI agents in education: an umbrella review of current utilization, challenges, and future directions for ethical and responsible use
    Amrita Ganguly, Nafisa Mehjabin, Aqdas Malik, Aditya Johri
    AI and Ethics.2026;[Epub]     CrossRef
  • Generative artificial intelligence in mental health: A preliminary study on automating materials development for cognitive bias modification
    Che-Wei Hsu, Mia Cochrane, Sasini Bambarawana
    International Journal of Mental Health.2026; : 1.     CrossRef
  • Medical Student Experiences With ChatGPT: National Cross-Sectional Study
    Alan Yuesheng Xu, Skye Speakman, Vincent Salvatore Piranio, Robert Medina, Michelle Liu, Chris Lamprecht, Nicolas Abchee, Meghan Brennan
    JMIR Formative Research.2026; 10: e76838.     CrossRef
  • Large Language Models Evaluation of Medical Licensing Examination Using GPT-4.0, ERNIE Bot 4.0, and GPT-4o
    Luoyu Lian, Xin Luo, Kavimbi Chipusu, Muhammad Awais Ashraf, Kelvin K. L. Wong, Wenjun Zhang
    Bioengineering.2026; 13(1): 113.     CrossRef
  • From “teaching by word and deed” to “intelligent mentorship”: ethical reconsiderations of AI-enabled medical education — lessons from China
    Zhitao Hou, Jing Chen, Hongwei Guo
    Frontiers in Medicine.2026;[Epub]     CrossRef
  • Artificial intelligence in medical ethics education: a descriptive study of eight models in multiple choice question generation
    John Obeid, Christopher Bobier, Alex Gillham, Adam Omelianchuk, Daniel Hurst
    BMC Medical Education.2026;[Epub]     CrossRef
  • Performance of ChatGPT-4o, Gemini 2.0 Pro, and DeepSeek-V3 in Patient-Facing Information on Chest Wall Deformities: A Comparative Evaluation of Accuracy, RELIABILITY, and Reproducibility
    Deniz Oke, Ozge Gulsum Illeez, Esra Giray, Betül Çiftçi
    Diagnostics.2026; 16(4): 589.     CrossRef
  • Current concerns and future directions of large language model ChatGPT in medicine: a machine-learning-driven global-scale bibliometric analysis
    Song-Bin Guo, Deng-Yao Liu, Xiao-Jie Fang, Yuan Meng, Zhen-Zhong Zhou, Jing Li, Mei Li, Li-Ling Luo, Hai-Long Li, Xiu-Yu Cai, Wei-Juan Huang, Xiao-Peng Tian
    International Journal of Surgery.2026; 112(2): 2805.     CrossRef
  • Postgraduate General Practice Training Under Early Clinical Responsibility: A Narrative Review on System-Based Supervision and the Supportive Role of Artificial Intelligence
    Christian J. Wiedermann, Giuliano Piccoliori, Pietro Murali, Cristina Pizzini, Doris Hager von Strobele Prainsack
    Healthcare.2026; 14(4): 503.     CrossRef
  • Comparing ChatGPT, Gemini, and Emerging LLMs in Low-Resource Educational Settings: Reasoning Quality, Consistency, and Explainability
    Aytug Onan, Arbi Haza Nasution, Tugba Celikten, Pinar Cetin
    IEEE Access.2026; 14: 32807.     CrossRef
  • From Hallucination to Precision: A Longitudinal Analysis of Reference Accuracy and Plagiarism in AI-Generated Medical Literature (2024–2026)
    Mevlüt Okan Aydin, Alper Vatansever, Sezer Erer Kafa
    Uludağ Üniversitesi Tıp Fakültesi Dergisi.2026; 52: 1870116.     CrossRef
  • Comparison and Review of Different Versions of OpenAI ChatGPT, Anthropic Claude and Google Gemini Large Language Models' Performance on Endodontics Questions in the Turkish Dentistry Specialization Exam
    Anil Ozgun Karatekin
    European Journal of Dental Education.2026;[Epub]     CrossRef
  • Generative AI's Impact on the Mental Health of Medical Students: Scenario Analysis
    Nora Arvai, Bertalan Meskó, Gellért Katonai
    JMIR Medical Education.2026; 12: e85373.     CrossRef
  • From policy to practice: editorial leadership in the age of AI-assisted science
    Whitley Stone, Matthew Garver, Samantha Johnson, Jennifer Caputo, Adam Ibrahim, Dustin Davis, James Navalta
    Sport, Education and Society.2026; : 1.     CrossRef
  • Implementation of artificial intelligence in the 2025 medical parasitology course at Hallym University
    Eun Hee Ha
    Journal of Educational Evaluation for Health Professions.2026; 23: 4.     CrossRef
  • Evaluating the clinical decision-making performance of large language models in clinically oriented thoracic anatomy scenarios: a comparative evaluation study
    Zeynep Nisa Karakoyun, Mustafa Deniz Yörük, Mehmed Emre Özdemir, Mehmet İlkay Koşar
    BMC Medical Education.2026;[Epub]     CrossRef
  • The Evolving Role of AI in Simulation-Based Medical Education: A Narrative Review
    Selina Hasan, Ayesha Ahmed, Faisal Ismail
    Advances in Medical Education and Practice.2026; Volume 17: 1.     CrossRef
  • Evaluating the performance of large language models versus prosthodontic residents on the 2024 and 2025 National Prosthodontic Resident Examination
    Soni Prasad, Qiao Fang, Merve Koseoglu, Marwa Shembesh, Maryam Gheisarifar, Cortino Sukotjo
    The Journal of Prosthetic Dentistry.2026; 136(2): e335.     CrossRef
  • Deciphering AI-chatbots utilization in clinical practice, research and education among Chinese medical students in generative AI era: a cross-sectional study
    Wenyi Gan, Yukang Liu, Anqi Lin, Bufu Tang, Aimin Jiang, Chang Qi, Lingxuan Zhu, Weiming Mou, Dongqiang Zeng, Mingjia Xiao, Guangdi Chu, Zaoqu Liu, Jiarui Xie, Quan Chen, Yiyi Zhang, Peng Luo
    BMC Medical Education.2026;[Epub]     CrossRef
  • AI at the bedside: Randomised controlled trial of ChatGPT’s impact on student performance in real-patient clinical exams
    Haroon Saloojee, Michael C. Gramanie, Rhodasi Mwale, Bruce A. Bassett, Shabir A. Madhi, Ismail S. Kalla
    Medical Teacher.2026; : 1.     CrossRef
  • Benchmarking publicly accessible large language models for high-myopia multiple-choice question generation in digital ophthalmic education and public health training
    Ligang Jiang, Xin Jiang, Wencan Wu, Fangzheng Jiang
    Frontiers in Public Health.2026;[Epub]     CrossRef
  • Generative Artificial Intelligence and the Problem of Authorship and Personal Voice in Academic Writing
    Sousen Elbouri, Wahbi Albasyouni
    Review of Artificial Intelligence in Education.2026; 7: e083.     CrossRef
  • Mitigating hallucinations in healthcare LLMs with granular fact-checking and domain-specific adaptation
    Musarrat Zeba, Abdullah Al Mamun, Kishoar Jahan Tithee, Debopom Sutradhar, Mohaimenul Azam Khan Raiaan, Saddam Mukta, Reem E. Mohamed, Md Rafiqul Islam, Yakub Sebastian, Mukhtar Hussain, Sami Azam
    Expert Systems with Applications.2026; 329: 132966.     CrossRef
  • The use of generative AI to improve the speaking skills of Chinese EFL learners in the vocational education context
    Xingju Chen, Bin Zou, Chenghao Wang
    Journal of China Computer-Assisted Language Learning.2026;[Epub]     CrossRef
  • A bibliometric analysis of artificial intelligence in anatomy education: Current situation, hot spots, and global trends
    Yuanyuan Fan, Xun Zheng, Guang Yang, Pengyu Li, Yanhao Ran, Tianfeng Xu, Tao Wei
    Medicine.2026; 105(22): e49128.     CrossRef
  • Clinical Evaluation of ChatGPT-5.3 Responses to Patient-Oriented Questions on Scoliosis: A Multidimensional Expert Analysis
    Muhsin Doran
    Healthcare.2026; 14(11): 1563.     CrossRef
  • Sentiment Differences in Resolving Moral Dilemmas: An Analysis of Student Attitudes and Large Language Models
    Agnieszka Zok, Jadwiga Wiertlweska-Bielarz, Marcin Moskalewicz
    Journal of Academic Ethics.2026;[Epub]     CrossRef
  • Generative Artificial Intelligence-driven orthodontic education practices
    Menghan Zhang, Yuzhi Yang, Yan lv, Yanfang Yu, Sihui Hu, Ziyuan Yang, Zhiwei Wang, Mengjie Wu
    BMC Medical Education.2026;[Epub]     CrossRef
  • Large language models as data-driven engines for benchmarking preventive and clinical knowledge in Chinese dental examinations
    Yong Zeng, Xinyi Hu, Wei Liu, Ke Deng, Meiqin Zhou, Yao Wang, Ling Ma, Qi Liu, He Meng
    Frontiers in Oral Health.2026;[Epub]     CrossRef
  • Effectiveness of a generative AI-powered digital tutor integrated with a knowledge graph in anatomy education for nursing students: a randomized controlled trial
    Can Zhao, Jianzhong Zhu, Jianhui Liu, Wentao Zhao, Yin Pang
    BMC Medical Education.2026;[Epub]     CrossRef
  • Applications, Challenges, and Future Directions of Large Language Models in Health Care Communication: Scoping Review
    Jing Chang, Ruotong Peng, Xi Chen, Yishu Zhu, Ruting Miao, Zeng Cao, Hui Feng
    Journal of Medical Internet Research.2026; 28: e84726.     CrossRef
  • An Examination of Mathematics Teachers’ Views on Differentiated and Individualized Education Programs for Gifted Students Developed with Artificial Intelligence
    Burak Karabey, Arzunur Çakmak, Yağmur Günay
    Necatibey Eğitim Fakültesi Elektronik Fen ve Matematik Eğitimi Dergisi.2026; 20(1): 357.     CrossRef
  • The Use of Artificial Intelligence Technologies in Higher Medical Education: Benefits, Possible Risks and Ways to Improve
    O. Voloshyna
    Lviv clinical bulletin.2026; (2 (54)): 44.     CrossRef
  • The Role of Artificial Intelligence in Nephrology Education
    Jing Miao, Charat Thongprayoon, Wisit Cheungpasitporn
    Advances in Kidney Disease and Health.2026;[Epub]     CrossRef
  • The digital therapist? LLMs and the future of clinical dialogue
    Edward Ruoyang Shi, Ran Pei, Lluís Barceló-Coblijn
    Language and Health.2026; 4(2): 100086.     CrossRef
  • Application effect and teaching evaluation of case-based learning combined with ChatGPT in ophthalmology clinical teaching
    Yuan Shen, Yuxiao Chen, Mengyao Li, Xiang Gu, Ai Zhuang
    Frontiers in Medicine.2026;[Epub]     CrossRef
  • The impact of generative artificial intelligence on clinical skills and knowledge acquisition in medical undergraduates: a systematic review and meta-analysis
    Guanli Xie, Jianglong Liao, Xiaoxia Tang, Han Fu, Duo Liu, Li Deng, Tao Wang, Xin Hong
    BMC Medical Education.2026;[Epub]     CrossRef
  • The Perceived Effectiveness of AI‐Powered Tools in Undergraduate Anatomy Education: A Cross‐Sectional Multifaceted Evaluation
    Ayisha Almamari, Mohamed Al‐Farsi, Abdul‐Rahman Almamari, Hussein Abdellatif
    Clinical Anatomy.2026;[Epub]     CrossRef
  • An exploratory study of the use of artificial intelligence-based virtual patients to enhance dentist-patient communication training
    Yixuan Xie, Zhanpeng Ou, Yuanding Huang, Hong Huang, Xiongwen Ran, Hongwei Dai, Bo Huang, Linjing Shu
    BMC Medical Education.2026;[Epub]     CrossRef
  • Faculty development and institutional readiness for incorporating AI in health professions education- A narrative review
    Prathibha Prasad, Mahinour Amin, Elizabeth Fitriana Sari, Lovely M. Annamma, Jayaraj Narayanan, Dinesh Yasothkumar, Mehzabin Ahmed
    Frontiers in Dental Medicine.2026;[Epub]     CrossRef
  • The temporal changes in GPT-4 performance on UKMLA practice questions: educational and clinical implications
    Ravanth Baskaran, Sai Sirikonda, Aditya Singh, Sripradha Srinivasan, Becky Leveridge, Octavi Casals-Farre, Harmeena Kaur, Susruta Manivannan, Kabilan Elangovan, Chrystie Wan Ning Quek, Kanae Fukutsu, Judith Cave, Bhaskar Kumar Somani, Daniel Shu Wei Ting,
    Frontiers in Medicine.2026;[Epub]     CrossRef
  • Artificial intelligence teaching assistants: a scalable solution for supporting struggling medical students
    Alina Sami, Mark Adkins, Anne McLeod, Fok-Han Leung, Chris Gilchrist
    Academic Medicine.2026;[Epub]     CrossRef
  • ARE AI MODELS READY FOR CLINICAL DECISION-MAKING IN DENTAL BLEACHING? AN EVIDENCE-BASED EVALUATION OF ACCURACY, COMPLETENESS, AND READABILITY OF CHATGPT-5, GEMINI 2.5 PRO, DEEPSEEK V3.2, AND CLAUDE SONNET 4.5
    Suzan Cangül, Özkan Adıgüzel, Tuba Tunç, Makbule Taşyürek, Hatice Ortaç
    Acta Medica Nicomedia.2026; 9: 1.     CrossRef
  • Artificial Intelligence and our journal, "Einstein São Paulo"
    Jacyr Pasternak
    einstein (São Paulo).2026;[Epub]     CrossRef
  • Exploring communication self-efficacy and artificial intelligence generated assessment tools in primary care education
    Constanze Dietzsch, Johanna Klutmann, Aline Köhler, Sara Volz-Willems, Johannes Jäger, Fabian Dupont
    Discover Education.2026;[Epub]     CrossRef
  • Hematology-Oncology Fellows’ Use of Artificial Intelligence: A Multicenter Educational Practice and Needs Assessment Survey
    Ariela L. Marshall, Richard C. Godby, Marc Braunstein, Soo Young Kim, Scott Moerdler, Layla Van Doren, Juan Jose Chango Azanza, Ronak Mistry
    Journal of Cancer Education.2026;[Epub]     CrossRef
  • Generative AI in scenario-based healthcare education: A systematic review of applications, validation practices, and pedagogical integration
    Mariana Neto, Rui Pinto, João Reis, Liliana Antão
    Computers and Education: Artificial Intelligence.2026; 11: 100654.     CrossRef
  • Large language models in emergency medicine education: opportunities, challenges, and implementation pathways
    Tingting Fan, Tianle Gao, Wan Tang, Qian Xu, Xingyou Wang, Qiaoli Su, Qingguo Lyu
    Frontiers in Public Health.2026;[Epub]     CrossRef
  • The Future of Anesthesiology Education
    Fei Chen, Susan M. Martinelli, Robert Isaak
    Anesthesiology Clinics.2026;[Epub]     CrossRef
  • Artificial Intelligence and Psychiatric Training: Opportunities, Challenges, and the Future of Mental Health Education
    Nazar Muhammad, Genna Sharp, Ankit Gautam, Sagarika Ray
    Cureus.2026;[Epub]     CrossRef
  • Embracing Large Language Models for Medical Applications, Part II: Building a Framework for Clinical Stewardship
    Shiv Patil, Mert Karabacak, Matthew Southerby, Konstantinos Margetis
    Cureus.2026;[Epub]     CrossRef
  • Integrating AI-Assisted World Café Discussion Into Clinical Reasoning Training: an Exploratory Study in Medical Education
    Xia Li, Yan Dan Lu, Hai Yun Luo, Yun Xia Xiong, Fei Li, Qian Wang, Yue Cui, Heng Tan, Jian Ping Xie, Ying Guo
    Medical Science Educator.2026;[Epub]     CrossRef
  • Inteligencia Artificial como recurso didáctico y rendimiento académico en estudiantes de Educación General Básica y Bachillerato.
    Mayra Elizabeth Ramón Ordoñez , Norma Piedad Castillo Castillo , Grecia Greolandia Lucas Muñoz , Johanna Rosa Barrera Salazar , Bryan Ricardo Morales Naranjo
    Saber Estratégico Internacional.2026; 4(3): 768.     CrossRef
  • AI-assisted patient education: Challenges and solutions in pediatric kidney transplantation
    MZ Ihsan, Dony Apriatama, Pithriani, Riza Amalia
    Patient Education and Counseling.2025; 131: 108575.     CrossRef
  • Exploring predictors of AI chatbot usage intensity among students: Within- and between-person relationships based on the technology acceptance model
    Anne-Kathrin Kleine, Insa Schaffernak, Eva Lermer
    Computers in Human Behavior: Artificial Humans.2025; 3: 100113.     CrossRef
  • Artificial intelligence in medical problem-based learning: opportunities and challenges
    Yaoxing Chen, Hong Qi, Yu Qiu, Juan Li, Liang Zhu, Xiaoling Gao, Hao Wang, Gan Jiang
    Global Medical Education.2025; 2(1): 5.     CrossRef
  • AI-powered standardised patients: evaluating ChatGPT-4o’s impact on clinical case management in intern physicians
    Selcen Öncü, Fulya Torun, Hilal Hatice Ülkü
    BMC Medical Education.2025;[Epub]     CrossRef
  • UsmleGPT: An AI application for developing MCQs via multi-agent system
    Zhehan Jiang, Shicong Feng
    Software Impacts.2025; 23: 100742.     CrossRef
  • ChatGPT’s Performance on Portuguese Medical Examination Questions: Comparative Analysis of ChatGPT-3.5 Turbo and ChatGPT-4o Mini
    Filipe Prazeres
    JMIR Medical Education.2025; 11: e65108.     CrossRef
  • Transforming medical education: leveraging large language models to enhance PBL—a proof-of-concept study
    Shoukat Ali Arain, Shahid Akhtar Akhund, Muhammad Abrar Barakzai, Sultan Ayoub Meo
    Advances in Physiology Education.2025; 49(2): 398.     CrossRef
  • Integrating artificial intelligence into pre-clinical medical education: challenges, opportunities, and recommendations
    Birgit Pohn, Lars Mehnen, Sebastian Fitzek, Kyung-Eun (Anna) Choi, Ralf J. Braun, Sepideh Hatamikia
    Frontiers in Education.2025;[Epub]     CrossRef
  • Evaluating the Accuracy and Reliability of Large Language Models (ChatGPT, Claude, DeepSeek, Gemini, Grok, and Le Chat) in Answering Item-Analyzed Multiple-Choice Questions on Blood Physiology
    Mayank Agarwal, Priyanka Sharma, Pinaki Wani
    Cureus.2025;[Epub]     CrossRef
  • Artificial intelligence-assisted academic writing: recommendations for ethical use
    Adam Cheng, Aaron Calhoun, Gabriel Reedy
    Advances in Simulation.2025;[Epub]     CrossRef
  • University Educators Perspectives on ChatGPT: A Technology Acceptance Model-Based Study
    Muna Barakat, Nesreen A. Salim, Malik Sallam
    Open Praxis.2025; 17(1): 129.     CrossRef
  • Knowledge and use, perceptions of benefits and limitations of artificial intelligence chatbots among Italian physiotherapy students: a cross-sectional national study
    Fabio Tortella, Alvisa Palese, Andrea Turolla, Greta Castellini, Paolo Pillastrini, Maria Gabriella Landuzzi, Chad Cook, Giovanni Galeoto, Giuseppe Giovannico, Lia Rodeghiero, Silvia Gianola, Giacomo Rossettini
    BMC Medical Education.2025;[Epub]     CrossRef
  • Digital and Intelligence Education in Medicine: A Bibliometric and Visualization Analysis Using CiteSpace and VOSviewer
    Bing Xiang Yang, FuLing Zhou, Nan Bai, Sichen Zhou, Chunyan Luo, Qing Wang, Arkers Kwan Ching Wong, Frances Lin
    Frontiers of Digital Education.2025;[Epub]     CrossRef
  • Prompts, privacy, and personalized learning: integrating AI into nursing education—a qualitative study
    Mingyan Shen, Yanping Shen, Fangchi Liu, Jiawen Jin
    BMC Nursing.2025;[Epub]     CrossRef
  • The role of ChatGPT-4o in differential diagnosis and management of vertigo-related disorders
    Xu Liu, Suming Shi, Xin Zhang, Qianwen Gao, Wuqing Wang
    Scientific Reports.2025;[Epub]     CrossRef
  • Situating governance and regulatory concerns for generative artificial intelligence and large language models in medical education
    Michael Tran, Chinthaka Balasooriya, Jitendra Jonnagaddala, Gilberto Ka-Kit Leung, Neeraj Mahboobani, Subha Ramani, Joel Rhee, Lambert Schuwirth, Neysan Sedaghat Najafzadeh-Tabrizi, Carolyn Semmler, Zoie SY Wong
    npj Digital Medicine.2025;[Epub]     CrossRef
  • Real-world implementation of an AI learning tool-MetaGP-Edu in medical education: A multi-center cohort study
    Yili Sun, Fei Liu
    Computers & Education.2025; 237: 105388.     CrossRef
  • A Comparative Analysis of the Accuracy and Readability of Popular Artificial Intelligence-Chat Bots for Inguinal Hernia Management
    Thisun Udagedara, Ashley Tran, Sumaya Bokhari, Sharon Shiraga, Stuart Abel, Caitlin Houghton, Katie Galvin, Kamran Samakar, Luke R. Putnam
    The American Surgeon™.2025; 91(10): 1729.     CrossRef
  • Exploring the Boundaries of AI: ChatGPT’s Accuracy in Anatomical Image Generation & Bone Identification of the Foot
    Simran Shamith, Neshal K. Kothari, Serena K. Kothari, Carolyn Giordano
    Journal of Orthopaedic Experience & Innovation.2025;[Epub]     CrossRef
  • Intelligent system for foreign students preparation for postgraduate exams in surgery
    A. A. Litvin, V. V. Bereshchenko, S. A. Anashkina, A. M. Karamyshau, V. S. Ivanov
    Health and Ecology Issues.2025; 22(2): 140.     CrossRef
  • Generative AI in higher education: A cross-sector analysis of ChatGPT's impact on STEM, social sciences, and healthcare
    Rouba Jamal Eddine, Ergun Gide, Abdallah Al-Sabbagh
    STEM Education.2025; 5(5): 757.     CrossRef
  • Feasibility study of using GPT for history-taking training in medical education: a randomized clinical trial
    Zhen Wang, Ting-Ting Fan, Meng-Li Li, Nin-Jun Zhu, Xiao-Chen Wang
    BMC Medical Education.2025;[Epub]     CrossRef
  • Patterns, advances, and gaps in using ChatGPT and similar technologies in nursing education: A PAGER scoping review
    Isaac Amankwaa, Emmanuel Ekpor, Daniel Cudjoe, Emmanuel Kobiah, Abdul-Karim Jebuni Fuseini, Maximous Diebieri, Sabastin Gyamfi, Sharon Brownie
    Nurse Education Today.2025; : 106822.     CrossRef
  • Comparison of the Performance of Five Generative Artificial Intelligence Models on a Medical Molecular Biology Examination
    Xiaoying Jiang
    Cureus.2025;[Epub]     CrossRef
  • ChatGPT and AI Chatbots in Education: An Umbrella Review of Systematic Reviews, Scoping Reviews, and Meta-Analyses
    Emmanouil D. Milakis, Constantina Corazon Argyrakou, Alexandros Melidis, John Vrettaros
    International Journal of Education and Information Technologies.2025; 19: 100.     CrossRef
  • Evaluating the efficacy of using large language models in preoperative prediction of microvascular invasion in HCC: a multicenter study
    Zongren Ding, Jianxing Zeng, Guoxu Fang, Pengfei Guo, Weiping Zhou, Yongyi Zeng
    Scientific Reports.2025;[Epub]     CrossRef
  • A Comparative Study of ChatGPT-4o and DeepSeek Responses to Mandibular Angle Osteotomy Questions
    Chenshan Jiang, Wenjie Cheng, Xinyi Jiang, Jianlin Zhang, Xiaojun Tang
    Journal of Craniofacial Surgery.2025; 36(7): e1113.     CrossRef
  • AI‐scending the scope: Perspectives on the integration and utilization of artificial intelligence and machine learning in genetic counseling graduate programs
    Ofir Feuer, Kyla Holmes, Sarah Kane, Kathryn M. Curry, Daria Ma, Chloe A. Chatwin, Marie Chuldzhyan, Emily Quinn, Nicholas Gorman
    Journal of Genetic Counseling.2025;[Epub]     CrossRef
  • How generative artificial intelligence transforms teaching and influences student wellbeing in future education
    Marcin Jukiewicz
    Frontiers in Education.2025;[Epub]     CrossRef
  • Application and Development of Large Language Models in Smart Inhalers
    Wenxu Guo, Zhihong Cheng, Jian Wang
    Pharmaceutical Fronts.2025; 07(03): e158.     CrossRef
  • Comparing Large Language Models as Health Literacy Tools: Evaluating and Simplifying Texts on gender-Affirming Surgery
    VICTORIA N. Yi, Angel P. Scialdone, Ann Marie Flusche, Kendall Reitz, Holly C. Lewis, William M. Tian, Elda Fisher, Kristen Rezak, Ash Patel
    Journal of Health Communication.2025; 30(10-12): 296.     CrossRef
  • Large Language Models in Lung Cancer: A Systematic Review (Preprint)
    Ruikang Zhong, Siyi Chen, Zexing Li, Tangke Gao, Yisha Su, Wenzheng Zhang, Dianna Liu, Lei Gao, Kaiwen Hu
    Journal of Medical Internet Research.2025;[Epub]     CrossRef
  • Comparative evaluation of large language models performance in medical education using urinary system histology assessment
    Anikó Szabó, Ghasem Dolatkhah Laein
    Scientific Reports.2025;[Epub]     CrossRef
  • Exploring the nexus of academic integrity and artificial intelligence in higher education: a bibliometric analysis
    Daniela Avello, Samuel Aranguren Zurita
    International Journal for Educational Integrity.2025;[Epub]     CrossRef
  • AI-powered platform revolutionizing blood cell morphology education for medical students
    Xuekai Liu, Lei Shang, Chen Liu, Yongpei Yu, Donghua Shao, Meilin He, Guowei Liang
    BMC Medical Education.2025;[Epub]     CrossRef
  • The strengths, weaknesses, opportunities, and threats of generative artificial intelligence: a qualitative study of undergraduate nursing students
    You Yuan, Jing Fu, Lanlan Leng, Zhuosi Wen, Xiaoman Wei, Die Han, Xinyang Hu, Yu Liang, Qian Luo, Xia Zhang, Rujun Hu
    Frontiers in Public Health.2025;[Epub]     CrossRef
  • ChatGPT’s progress over time: A longitudinal enhancing biostatistical problem-solving in medical education
    Aleksandra Ignjatović, Marija Anđelković Apostolović, Lazar Stevanović, Pavle Radovanović, Marija Topalović, Tamara Filipović, Suzana Otašević
    Health Informatics Journal.2025;[Epub]     CrossRef
  • Applications of artificial intelligence in healthcare simulation: a model of thinking
    Adam Cheng, Carolyn McGregor
    Advances in Simulation.2025;[Epub]     CrossRef
  • The current status, knowledge, attitudes, and challenges of generative artificial intelligence use among undergraduate nursing students: a single-center cross-sectional survey of western China
    Yuanyuan Zhao, You Yuan, Zhuosi Wen, Lanlan Leng, Lei Shi, Xinyang Hu, Xiaoman Wei, Meng Zuo, Jianghong Mou, Qian Luo, Mei Chen, Rujun Hu, Huiming Gao
    Frontiers in Public Health.2025;[Epub]     CrossRef
  • Evaluating large language models for mild cognitive impairment among older adults: A bilingual comparison of ChatGPT, Gemini, and Kimi
    Yexuan Xiao, Qianhui Pan, Haoyuan Liu, Yilin He, Yuhe Zhang, Nan Jiang
    Health Informatics Journal.2025;[Epub]     CrossRef
  • ChatGPT Utility in Medical Education: A Systematic Review and Meta-Analysis
    Fariba Hosseinzadegan, Seyedeh Masumeh Hashemi, Seyed Vahid Sharifi, Mehrnoosh Khoshnoodifar
    Health Science Monitor.2025; 4(3): 173.     CrossRef
  • A systematic review of the impact of generative AI on postgraduate research: opportunities, challenges, and ethical implications
    Vicent Mabirizi, Calorine Katushabe, Gloria Muhoza, Jack Rugasira
    Discover Artificial Intelligence.2025;[Epub]     CrossRef
  • Evaluating Science Readiness of Pre-Service Elementary Teachers Through Diagnostic Assessment and Parental Feedback: Implications for Teacher Education
    Zydrick L Avelino
    Integrated Science Education Journal.2025; 6(3): 185.     CrossRef
  • How can artificial intelligence transform the training of medical students and physicians?
    Yilin Ning, Jasmine Chiat Ling Ong, Haoran Cheng, Haibo Wang, Daniel Shu Wei Ting, Yih Chung Tham, Tien Yin Wong, Nan Liu
    The Lancet Digital Health.2025; 7(10): 100900.     CrossRef
  • Poor Performance of Large Language Models Based on the Diabetes and Endocrinology Specialty Certificate Examination of the United Kingdom
    Ka Siu Fan, Jeffrey Gan, Isabelle X Zou, Maja Kaladjiska, Monique B Inguanez, Gillian L Garden
    Cureus.2025;[Epub]     CrossRef
  • Perceptions and Use of Generative Artificial Intelligence in Medical Students: A Multicenter Survey
    Cecilia Tran, Brett N. Hryciw, Sean William Moore, Alan Chaput, Andrew John Ervine Seely
    Journal of Medical Education and Curricular Development.2025;[Epub]     CrossRef
  • Effect of artificial intelligence-assisted personalized feedback on radiographic diagnostic performance of dental students: a controlled study
    Busra Nur Gokkurt Yilmaz, Furkan Ozbey, Birkan Eyup Yilmaz
    BMC Medical Education.2025;[Epub]     CrossRef
  • Academic misconduct and artificial intelligence use by medical students, interns and PhD students in Ukraine: a cross-sectional study
    Lesya Lymar, Iurii Kuchyn, Kateryna Bielka, Livia Puljak
    BMC Medical Education.2025;[Epub]     CrossRef
  • Medical multimodal large language models: A systematic review
    Yuan Hu, Chenhan Xu, Bo Lin, Weibin Yang, Yuan Yan Tang
    Intelligent Oncology.2025; 1(4): 308.     CrossRef
  • Research on the training strategy of college students' design thinking and innovation ability based on multimodal large model
    Qing Liu, Wei Xue, Lingbo Meng, Yilin Zhu, Jixin Li
    Frontiers in Education.2025;[Epub]     CrossRef
  • Clinical Assessment of Large Language Models: A Comprehensive Multi-domain Performance Study for Healthcare Applications
    Harsh Hirani, Bharat Saboo, Alok Modi, Palak Modi, Shambo Samrat Samajdar, Banshi Saboo, Manoj Chawla, Anuj Maheshwari, Amit Gupta, Rakesh Parikh, S. S. Dariya, Ritu Johari, Rutul Gokhlani, Param Kadam
    International Journal of Diabetes and Technology.2025; 4(4): 159.     CrossRef
  • Evaluating the impact of AI-tutoring versus expert human instruction on surgical skills in medical students
    Yili Sun, Fei Liu
    Education and Information Technologies.2025; 30(18): 26413.     CrossRef
  • AI adoption in higher education: Exploring attitudes and perceived benefits between users and non-users
    Nevenka Popović Šević, Aleksandar Šević, Milica Slijepčević, Jelena Krstić
    Online Journal of Communication and Media Technologies.2025; 15(4): e202528.     CrossRef
  • Integrating generative Artificial Intelligence into student learning: A systematic review from a TPACK perspective
    Xiaofan Liu, Baichang Zhong
    Educational Research Review.2025; 49: 100741.     CrossRef
  • Large language models as educational collaborators: developing non-conventional teaching aids in pharmacology & therapeutics
    Kannan Sridharan, Gowri Sivaramakrishnan
    BMC Medical Education.2025;[Epub]     CrossRef
  • Technologies, opportunities, challenges, and future directions for integrating generative artificial intelligence into medical education: a narrative review
    Junseok Kang, Jihyun Ahn
    Ewha Medical Journal.2025; 48(4): e53.     CrossRef
  • A bibliometric analysis of artificial intelligence in medical education (2015–2025)
    Wendan Cheng, Zhongyao Hu, Haoran Yu
    Medicine.2025; 104(46): e45684.     CrossRef
  • Embracing the Future of Medical Education With Large Language Model–Based Virtual Patients: Scoping Review
    Jianwen Zeng, Wenhao Qi, Shiying Shen, Xin Liu, Sixie Li, Bing Wang, Chaoqun Dong, Xiaohong Zhu, Yankai Shi, Xiajing Lou, Bingsheng Wang, Jiani Yao, Guowei Jiang, Qiong Zhang, Shihua Cao
    Journal of Medical Internet Research.2025; 27: e79091.     CrossRef
  • Artificial intelligence in undergraduate medical education: an updated scoping review
    Jennifer Simoni, Judith Urtubia-Fernandez, Elisa Mengual, Diglio A. Simoni, Montserrat Royo, Diego Egaña-Yin, Oliver L. A. Hertog, Lourdes López-Ortiz, Adrián Muñoz-Tomás, Paula Santiago-Martínez, Adrián Vahamaki, José Luis Pereira
    BMC Medical Education.2025;[Epub]     CrossRef
  • Considerations for Patient Privacy of Large Language Models in Health Care: Scoping Review
    Xiaoying Zhong, Siyi Li, Zhao Chen, Long Ge, Dongdong Yu, Shijia Wang, Liangzhen You, Hongcai Shang
    Journal of Medical Internet Research.2025; 27: e76571.     CrossRef
  • The Growing Importance of Soft Skills in Medical Education in the AI Era: Balancing Humanistic Care and Artificial Intelligence
    Effie Simou
    International Medical Education.2025; 4(4): 50.     CrossRef
  • Integrating Generative AI in Health Education: A Scoping Review and Implementation Framework
    Kellie Toohey, Zach Quince, Felicity Walker, Linda Furness, Michelle Bissett, Carlie Daley, Kachina Allen, Natalie Munro, Andy Smidt, Jodie Cochrane Wilkie, Louise Horstmanshof, Kathryn Baltrotsky, Fiona Naumann
    Medical Science Educator.2025; 35(6): 2751.     CrossRef
  • Beyond technical efficacy: challenges and critical concerns of large language model’s impact on medical education in China: a systematic review
    Puwen Shen, Yongxiang Yuan, Xinyao He, Fang Wang
    Global Medical Education.2025; 2(1): 113.     CrossRef
  • The Application of Flipped Classroom Integrated with ChatGPT in Improving Graduate Education on Choroidal Melanoma
    Shengyu Tan, Qijian Deng, Qiaoyan Wei, Xuan Zhu, Shengguo Li
    Journal of Cancer Education.2025;[Epub]     CrossRef
  • Case study: creating an ‘AI for Academic Writing Skills’ induction session for postgraduate life science courses
    Jennifer Carter, Anne Ferrey, Hubert Lam, Kelly Webb-Davies, Damion Young, Barbara Zonta, Delia O' Rourke
    Emerging Topics in Life Sciences.2025;[Epub]     CrossRef
  • Chatbots in neurology and neuroscience: Interactions with students, patients and neurologists
    Stefano Sandrone
    Brain Disorders.2024; 15: 100145.     CrossRef
  • ChatGPT in education: unveiling frontiers and future directions through systematic literature review and bibliometric analysis
    Buddhini Amarathunga
    Asian Education and Development Studies.2024; 13(5): 412.     CrossRef
  • Evaluating the performance of ChatGPT-3.5 and ChatGPT-4 on the Taiwan plastic surgery board examination
    Ching-Hua Hsieh, Hsiao-Yun Hsieh, Hui-Ping Lin
    Heliyon.2024; 10(14): e34851.     CrossRef
  • Preparing for Artificial General Intelligence (AGI) in Health Professions Education: AMEE Guide No. 172
    Ken Masters, Anne Herrmann-Werner, Teresa Festl-Wietek, David Taylor
    Medical Teacher.2024; 46(10): 1258.     CrossRef
  • A Comparative Analysis of ChatGPT and Medical Faculty Graduates in Medical Specialization Exams: Uncovering the Potential of Artificial Intelligence in Medical Education
    Gülcan Gencer, Kerem Gencer
    Cureus.2024;[Epub]     CrossRef
  • Research ethics and issues regarding the use of ChatGPT-like artificial intelligence platforms by authors and reviewers: a narrative review
    Sang-Jun Kim
    Science Editing.2024; 11(2): 96.     CrossRef
  • Innovation Off the Bat: Bridging the ChatGPT Gap in Digital Competence among English as a Foreign Language Teachers
    Gulsara Urazbayeva, Raisa Kussainova, Aikumis Aibergen, Assel Kaliyeva, Gulnur Kantayeva
    Education Sciences.2024; 14(9): 946.     CrossRef
  • Exploring the perceptions of Chinese pre-service teachers on the integration of generative AI in English language teaching: Benefits, challenges, and educational implications
    Ji Young Chung, Seung-Hoon Jeong
    Online Journal of Communication and Media Technologies.2024; 14(4): e202457.     CrossRef
  • Unveiling the bright side and dark side of AI-based ChatGPT : a bibliographic and thematic approach
    Chandan Kumar Tiwari, Mohd. Abass Bhat, Abel Dula Wedajo, Shagufta Tariq Khan
    Journal of Decision Systems.2024; : 1.     CrossRef
  • Artificial Intelligence in Medical Education and Mentoring in Rehabilitation Medicine
    Julie K. Silver, Mustafa Reha Dodurgali, Nara Gavini
    American Journal of Physical Medicine & Rehabilitation.2024; 103(11): 1039.     CrossRef
  • The Potential of Artificial Intelligence Tools for Reducing Uncertainty in Medicine and Directions for Medical Education
    Sauliha Rabia Alli, Soaad Qahhār Hossain, Sunit Das, Ross Upshur
    JMIR Medical Education.2024; 10: e51446.     CrossRef
  • A Systematic Literature Review of Empirical Research on Applying Generative Artificial Intelligence in Education
    Xin Zhang, Peng Zhang, Yuan Shen, Min Liu, Qiong Wang, Dragan Gašević, Yizhou Fan
    Frontiers of Digital Education.2024; 1(3): 223.     CrossRef
Research articles
ChatGPT (GPT-4) passed the Japanese National License Examination for Pharmacists in 2022, answering all items including those with diagrams: a descriptive study  
Hiroyasu Sato, Katsuhiko Ogasawara
J Educ Eval Health Prof. 2024;21:4.   Published online February 28, 2024
DOI: https://doi.org/10.3352/jeehp.2024.21.4
  • 11,930 View
  • 361 Download
  • 13 Web of Science
  • 15 Crossref
AbstractAbstract PDFSupplementary Material
Purpose
The objective of this study was to assess the performance of ChatGPT (GPT-4) on all items, including those with diagrams, in the Japanese National License Examination for Pharmacists (JNLEP) and compare it with the previous GPT-3.5 model’s performance.
Methods
The 107th JNLEP, conducted in 2022, with 344 items input into the GPT-4 model, was targeted for this study. Separately, 284 items, excluding those with diagrams, were entered into the GPT-3.5 model. The answers were categorized and analyzed to determine accuracy rates based on categories, subjects, and presence or absence of diagrams. The accuracy rates were compared to the main passing criteria (overall accuracy rate ≥62.9%).
Results
The overall accuracy rate for all items in the 107th JNLEP in GPT-4 was 72.5%, successfully meeting all the passing criteria. For the set of items without diagrams, the accuracy rate was 80.0%, which was significantly higher than that of the GPT-3.5 model (43.5%). The GPT-4 model demonstrated an accuracy rate of 36.1% for items that included diagrams.
Conclusion
Advancements that allow GPT-4 to process images have made it possible for LLMs to answer all items in medical-related license examinations. This study’s findings confirm that ChatGPT (GPT-4) possesses sufficient knowledge to meet the passing criteria.

Citations

Citations to this article as recorded by  
  • Applications and potential of ChatGPT in dentistry: Scoping review of research perspectives
    Masakazu Hamada, Sumire Kikuchi, Tatsuya Akitomo, Satoru Kusaka, Yuko Iwamoto, Ryota Nomura
    Journal of Dental Sciences.2026; 21(1): 1.     CrossRef
  • Performance of ChatGPT‐3.5 and ChatGPT‐4o in the Japanese National Dental Examination
    Osamu Uehara, Tetsuro Morikawa, Fumiya Harada, Nodoka Sugiyama, Yuko Matsuki, Daichi Hiraki, Hinako Sakurai, Takashi Kado, Koki Yoshida, Yukie Murata, Hirofumi Matsuoka, Toshiyuki Nagasawa, Yasushi Furuichi, Yoshihiro Abiko, Hiroko Miura
    Journal of Dental Education.2025; 89(4): 459.     CrossRef
  • Qwen-2.5 Outperforms Other Large Language Models in the Chinese National Nursing Licensing Examination: Retrospective Cross-Sectional Comparative Study
    Shiben Zhu, Wanqin Hu, Zhi Yang, Jiani Yan, Fang Zhang
    JMIR Medical Informatics.2025; 13: e63731.     CrossRef
  • ChatGPT (GPT-4V) Performance on the Healthcare Information Technologist Examination in Japan
    Kai Ishida, Eisuke Hanada
    Cureus.2025;[Epub]     CrossRef
  • Medication counseling for OTC drugs using customized ChatGPT-4: Comparison with ChatGPT-3.5 and ChatGPT-4o
    Keisuke Kiyomiya, Tohru Aomori, Hisakazu Ohtani
    DIGITAL HEALTH.2025;[Epub]     CrossRef
  • Current Use of Generative Artificial Intelligence in Pharmacy Practice: A Literature Mini-review
    Keisuke Kiyomiya, Tohru Aomori, Hitoshi Kawazoe, Hisakazu Ohtani
    Iryo Yakugaku (Japanese Journal of Pharmaceutical Health Care and Sciences).2025; 51(4): 177.     CrossRef
  • Performance evaluation of large language models for the national nursing examination in Japan
    Tomoki Kuribara, Kengo Hirayama, Kenji Hirata
    DIGITAL HEALTH.2025;[Epub]     CrossRef
  • Harnessing ChatGPT for digital tools in pharmacy practice
    Reginald Amin Yakob, Adeola Bamgboje-Ayodele, Jack C. Collins, Parisa Aslani
    Research in Social and Administrative Pharmacy.2025; 21(11): 943.     CrossRef
  • Performance Evaluation of 18 Generative AI Models (ChatGPT, Gemini, Claude, and Perplexity) in 2024 Japanese Pharmacist Licensing Examination: Comparative Study
    Hiroyasu Sato, Katsuhiko Ogasawara, Hidehiko Sakurai
    JMIR Medical Education.2025; 11: e76925.     CrossRef
  • Evaluation of the Accuracy and Reliability of Responses Generated by Artificial Intelligence Related to Clinical Pharmacology
    Michal Ordak, Julia Adamczyk, Agata Oskroba, Michal Majewski, Tadeusz Nasierowski
    Journal of Clinical Medicine.2025; 14(21): 7563.     CrossRef
  • Performance of ChatGPT-4 on the French Board of Plastic Reconstructive and Aesthetic Surgery written exam: a descriptive study
    Emma Dejean-Bouyer, Anoujat Kanlagna, François Thuau, Pierre Perrot, Ugo Lancien
    Journal of Educational Evaluation for Health Professions.2025; 22: 27.     CrossRef
  • Potential of ChatGPT to Pass the Japanese Medical and Healthcare Professional National Licenses: A Literature Review
    Kai Ishida, Eisuke Hanada
    Cureus.2024;[Epub]     CrossRef
  • Performance of Generative Pre-trained Transformer (GPT)-4 and Gemini Advanced on the First-Class Radiation Protection Supervisor Examination in Japan
    Hiroki Goto, Yoshioki Shiraishi, Seiji Okada
    Cureus.2024;[Epub]     CrossRef
  • An exploratory assessment of GPT-4o and GPT-4 performance on the Japanese National Dental Examination
    Masaki Morishita, Hikaru Fukuda, Shino Yamaguchi, Kosuke Muraoka, Taiji Nakamura, Masanari Hayashi, Izumi Yoshioka, Kentaro Ono, Shuji Awano
    The Saudi Dental Journal.2024; 36(12): 1577.     CrossRef
  • Evaluating the Accuracy of ChatGPT in the Japanese Board-Certified Physiatrist Examination
    Yuki Kato, Kenta Ushida, Ryo Momosaki
    Cureus.2024;[Epub]     CrossRef
Information amount, accuracy, and relevance of generative artificial intelligence platforms’ answers regarding learning objectives of medical arthropodology evaluated in English and Korean queries in December 2023: a descriptive study
Hyunju Lee, Soobin Park
J Educ Eval Health Prof. 2023;20:39.   Published online December 28, 2023
DOI: https://doi.org/10.3352/jeehp.2023.20.39
  • 6,531 View
  • 287 Download
  • 7 Web of Science
  • 6 Crossref
AbstractAbstract PDFSupplementary Material
Purpose
This study assessed the performance of 6 generative artificial intelligence (AI) platforms on the learning objectives of medical arthropodology in a parasitology class in Korea. We examined the AI platforms’ performance by querying in Korean and English to determine their information amount, accuracy, and relevance in prompts in both languages.
Methods
From December 15 to 17, 2023, 6 generative AI platforms—Bard, Bing, Claude, Clova X, GPT-4, and Wrtn—were tested on 7 medical arthropodology learning objectives in English and Korean. Clova X and Wrtn are platforms from Korean companies. Responses were evaluated using specific criteria for the English and Korean queries.
Results
Bard had abundant information but was fourth in accuracy and relevance. GPT-4, with high information content, ranked first in accuracy and relevance. Clova X was 4th in amount but 2nd in accuracy and relevance. Bing provided less information, with moderate accuracy and relevance. Wrtn’s answers were short, with average accuracy and relevance. Claude AI had reasonable information, but lower accuracy and relevance. The responses in English were superior in all aspects. Clova X was notably optimized for Korean, leading in relevance.
Conclusion
In a study of 6 generative AI platforms applied to medical arthropodology, GPT-4 excelled overall, while Clova X, a Korea-based AI product, achieved 100% relevance in Korean queries, the highest among its peers. Utilizing these AI platforms in classrooms improved the authors’ self-efficacy and interest in the subject, offering a positive experience of interacting with generative AI platforms to question and receive information.

Citations

Citations to this article as recorded by  
  • Implementation of artificial intelligence in the 2025 medical parasitology course at Hallym University
    Eun Hee Ha
    Journal of Educational Evaluation for Health Professions.2026; 23: 4.     CrossRef
  • How appropriately can generative artificial intelligence platforms, including GPT-4, Gemini, Bing, and Wrtn, answer questions about colon cancer in the Korean language?
    Sun Huh
    Annals of Coloproctology.2025; 41(3): 190.     CrossRef
  • How Can Clinicians Leverage Vibe Coding for Machine Learning and Deep Learning Research?
    Yoonhwan Lee, Sun Huh
    Endocrinology and Metabolism.2025; 40(5): 659.     CrossRef
  • Opportunities, challenges, and future directions of large language models, including ChatGPT in medical education: a systematic scoping review
    Xiaojun Xu, Yixiao Chen, Jing Miao
    Journal of Educational Evaluation for Health Professions.2024; 21: 6.     CrossRef
  • The emergence of generative artificial intelligence platforms in 2023, journal metrics, appreciation to reviewers and volunteers, and obituary
    Sun Huh
    Journal of Educational Evaluation for Health Professions.2024; 21: 9.     CrossRef
  • Comparison of the Performance of ChatGPT, Claude and Bard in Support of Myopia Prevention and Control
    Yan Wang, Lihua Liang, Ran Li, Yihua Wang, Changfu Hao
    Journal of Multidisciplinary Healthcare.2024; Volume 17: 3917.     CrossRef
Review
Application of artificial intelligence chatbots, including ChatGPT, in education, scholarly work, programming, and content generation and its prospects: a narrative review
Tae Won Kim
J Educ Eval Health Prof. 2023;20:38.   Published online December 27, 2023
DOI: https://doi.org/10.3352/jeehp.2023.20.38
  • 32,762 View
  • 1,578 Download
  • 69 Web of Science
  • 80 Crossref
AbstractAbstract PDFSupplementary Material
This study aims to explore ChatGPT’s (GPT-3.5 version) functionalities, including reinforcement learning, diverse applications, and limitations. ChatGPT is an artificial intelligence (AI) chatbot powered by OpenAI’s Generative Pre-trained Transformer (GPT) model. The chatbot’s applications span education, programming, content generation, and more, demonstrating its versatility. ChatGPT can improve education by creating assignments and offering personalized feedback, as shown by its notable performance in medical exams and the United States Medical Licensing Exam. However, concerns include plagiarism, reliability, and educational disparities. It aids in various research tasks, from design to writing, and has shown proficiency in summarizing and suggesting titles. Its use in scientific writing and language translation is promising, but professional oversight is needed for accuracy and originality. It assists in programming tasks like writing code, debugging, and guiding installation and updates. It offers diverse applications, from cheering up individuals to generating creative content like essays, news articles, and business plans. Unlike search engines, ChatGPT provides interactive, generative responses and understands context, making it more akin to human conversation, in contrast to conventional search engines’ keyword-based, non-interactive nature. ChatGPT has limitations, such as potential bias, dependence on outdated data, and revenue generation challenges. Nonetheless, ChatGPT is considered to be a transformative AI tool poised to redefine the future of generative technology. In conclusion, advancements in AI, such as ChatGPT, are altering how knowledge is acquired and applied, marking a shift from search engines to creativity engines. This transformation highlights the increasing importance of AI literacy and the ability to effectively utilize AI in various domains of life.

Citations

Citations to this article as recorded by  
  • Investigating AI Chatbot Dependence: Associations with Internet and Smartphone Dependence, Mental Health Outcomes, and the Moderating Role of Usage Purposes
    Xing Zhang, Hansen Li, Mingyue Yin, Mingyang Zhang, Zhaoqian Li, Zongwei Chen
    International Journal of Human–Computer Interaction.2026; 42(7): 5592.     CrossRef
  • Exploring Early Perceptions and Experiences of ChatGPT in Pediatric Critical Care: A Qualitative Study Among Health Care Professionals
    Mohamad-Hani Temsah, Noura Abouammoh, Mohammed Alsatrawi, Muneera Al-Jelaify, Ibraheem Altamimi, PICU ChatGPT Research Consortium, Khalid Alhasan, Khalid H. Malki, Jaffar A. Al-Tawfiq, Ayman Al-Eyadhy
    Clinical Pediatrics.2026; 65(1): 11.     CrossRef
  • Social-Ecological Factors Associated with AI Chatbot Use and Dependence
    Hansen Li, Jie Tian, Mingyue Yin, Xing Zhang
    International Journal of Human–Computer Interaction.2026; 42(11): 7776.     CrossRef
  • Use of artificial intelligence and big data in transfusion medicine: An exploratory assessment of status in the Eastern Mediterranean and North Africa region
    Arwa Z. Al‐Riyami, Suha Herjes
    Vox Sanguinis.2026; 121(4): 511.     CrossRef
  • The legacy and future of recurrent neural networks in personalized medicine: A reflection on the 2024 Nobel Physics Prize
    Emily Wittrup, Alan Kay, Jett Rosen, Kuan-Fu Chen, Kayvan Najarian
    Biomedical Journal.2026; 49(1): 100933.     CrossRef
  • Designing intelligent chatbots with ChatGPT: a framework for development and implementation
    Sajjad Hyder, Javeed Kittur
    Frontiers in Artificial Intelligence.2026;[Epub]     CrossRef
  • Navigating academic integrity in biomedical research: the impact of large language models on current practices and future directions
    Anqi Lin, Zuwei Chen, Aimin Jiang, Bufu Tang, Chang Qi, Lingxuan Zhu, Weiming Mou, Wenyi Gan, Dongqiang Zeng, Mingjia Xiao, Guangdi Chu, Shengkun Peng, Hank Z.H. Wong, Lin Zhang, Hengguo Zhang, Xinpei Deng, Quan Cheng, Jian Zhang, Peng Luo
    International Journal of Surgery.2026; 112(2): 4418.     CrossRef
  • When Generative AI Meets Socratic Method: Investigating Programming Learning Dynamics Through Behaviours, Interaction Qualities and Perceptions
    Dan Sun, Yi Zheng, Jie Xu, Zhanshan Yang
    Journal of Computer Assisted Learning.2026;[Epub]     CrossRef
  • ADAPTIVE LEARNING PLATFORMS FOR PERFORMING ARTS
    Jyoti M. Shinde, Swati Chaudhary, Amol Bhilare, Rajendra Subhas Jarad, Anitha K, Shubhangi Sunil Satav
    ShodhKosh: Journal of Visual and Performing Arts.2026; 7(1s): 264.     CrossRef
  • Temporal Trends in Orthopaedic Scientific Publishing Before and After the Release of ChatGPT: A Bibliometric Analysis
    Gabriel Ferraz Ferreira, Thomas Lorchan Lewis, Miguel Viana Pereira Filho, Fabricio Machado, Gustavo Gonçalves Arliani, Noel Oizerovici Foni, Roberto Zambeli, Caio Nery, Rodrigo Pagnano
    Qeios.2026;[Epub]     CrossRef
  • Integration of AI Content Generation-Enabled Virtual Museums into University History Education
    Shirong Tan, Yuchun Liu, Lei Wang
    Applied System Innovation.2026; 9(3): 64.     CrossRef
  • Integrating generative artificial intelligence in African higher education: university students’ awareness, attitudes, and use of ChatGPT in Zambia
    Steward Mudenda, Moses Mukosha, Ruth Lindizyani Mfune, Bernard Kathewera, Imukusi Mutanekelwa, Boris Mwanza, Webrod Mufwambi, Muchindu Hampango, Kingsley Kamvuma, Martha Mwaba, Tumelo Muyenga, Chikwanda Chileshe, Mildred Zulu, Rabecca Tembo, Florence Mwab
    Frontiers in Education.2026;[Epub]     CrossRef
  • Evaluation of a Hybrid Case-Based Learning and Small-Group Teaching Model Using Artificial Intelligence–Assisted Case Generation: An Innovative Dental Education Study
    Sushmitha Malpe Gopal, Shashidhar Venkatesh Murthy, ShanthaKumari B R, Vathsala Patil, Tanvi Shetty, Komattil Ramnarayan
    Advances in Medical Education and Practice.2026; Volume 17: 1.     CrossRef
  • Matriz de Competências em Informática em Saúde e Saúde Digital: um framework interprofissional para sistemas de saúde inteligentes
    Grace Marcon Dal Sasso, Antonio Carlos Onofre de Lira, Heimar de Fatima Marin, Juliano de Souza Gaspar
    Journal of Health Informatics.2026; 18: 1635.     CrossRef
  • Are You Ready for Human-like AI Service Agents: Consumers’ Willingness to Use Substitute Versus Assist AI on OTA Platforms
    Wenqiu Guo, Yenchen Liu, Banggang Wu, Xiaoyu Deng
    Journal of Theoretical and Applied Electronic Commerce Research.2026; 21(6): 160.     CrossRef
  • Let’s Talk Systems - What Do We Really Mean by a Systems Approach in AI-Enhanced Medical Education?
    Krishna Mohan Surapaneni
    Medical Science Educator.2026;[Epub]     CrossRef
  • Teaching AI-Native Learners: Reimagining Medical Education for the Post-ChatGPT Era
    Noor-i-Kiran Naeem
    Horizon.2026; 1(3): 1.     CrossRef
  • Modern approaches to studying human interaction with generative artificial intelligence
    N.I. Skrylnikova, M.A. Kholodnaya
    Journal of Modern Foreign Psychology.2026; 15(2): 17.     CrossRef
  • Comparing Speed and Accuracy of Artificial Intelligence Large Language Models on the Orthopedic In-Training Examination
    Fahad Nadeem, Saad Ibrahim, Sean Taylor, Saurabh Rawall, Zuhair Mohammad, Humza Pirzadah, José Ayala-Ortiz, Sakthivel Rajaram
    Southern Medical Journal.2026; 119(7): 362.     CrossRef
  • Exploring the role of generative artificial intelligence in enhancing clinical skills training: A bibliometric analysis
    Jia Zhang, Yuanzhou Liu, Bo Wang
    Medicine.2026; 105(29): e49849.     CrossRef
  • Performance of DeepSeek V3 and ChatGPT-4o in answering esophageal cancer-related questions
    Qian Yang, Jing Yi, Anran Gong, Yanyan Li, Qingyan Feng, Aiyan Fu, Jiangtao Li, Yunkai Zhan
    Medicine.2026; 105(30): e49896.     CrossRef
  • Disclosure is not documentation: an open science framework for documenting generative AI use in scholarly research and publication workflows
    Jan Christopher Cwik
    Research Integrity and Peer Review.2026;[Epub]     CrossRef
  • Beyond Pattern Recognition: Call for Functionally Aware AI for Anatomical Illustration
    Agata Maria Kawalec-Rutkowska, Marian Simka
    JMIR Medical Education.2026; 12: e92700.     CrossRef
  • Evaluating AI-generated patient education materials for endometrial cancer surgery: a comparative analysis of response quality, reliability, and readability between ChatGPT and DeepSeek models
    Ling Tian, Long-yu Tang, Ming-tao Yang, Yan Wang, Hong-ni He, Tian-wen He, Ying Tang, Chuan Lin, Hui-quan Hu, Jun Li
    Frontiers in Public Health.2026;[Epub]     CrossRef
  • The Development and Validation of an Artificial Intelligence Chatbot Dependence Scale
    Xing Zhang, Mingyue Yin, Mingyang Zhang, Zhaoqian Li, Hansen Li
    Cyberpsychology, Behavior, and Social Networking.2025; 28(2): 126.     CrossRef
  • Readability, quality and accuracy of generative artificial intelligence chatbots for commonly asked questions about labor epidurals: a comparison of ChatGPT and Bard
    D. Lee, M. Brown, J. Hammond, M. Zakowski
    International Journal of Obstetric Anesthesia.2025; 61: 104317.     CrossRef
  • ChatGPT-4 Performance on German Continuing Medical Education—Friend or Foe (Trick or Treat)? Protocol for a Randomized Controlled Trial
    Christian Burisch, Abhav Bellary, Frank Breuckmann, Jan Ehlers, Serge C Thal, Timur Sellmann, Daniel Gödde
    JMIR Research Protocols.2025; 14: e63887.     CrossRef
  • The effect of incorporating large language models into the teaching on critical thinking disposition: An “AI + Constructivism Learning Theory” attempt
    Peng Wang, Kexin Yin, Mingzhu Zhang, Yuanxin Zheng, Tong Zhang, Yanjun Kang, Xun Feng
    Education and Information Technologies.2025; 30(9): 11625.     CrossRef
  • The Impact of Adaptive Learning Technologies, Personalized Feedback, and Interactive AI Tools on Student Engagement: The Moderating Role of Digital Literacy
    Husam Yaseen, Abdelaziz Saleh Mohammad, Najwa Ashal, Hesham Abusaimeh, Ahmad Ali, Abdel-Aziz Ahmad Sharabati
    Sustainability.2025; 17(3): 1133.     CrossRef
  • Artificial Intelligence in Nursing: New Opportunities and Challenges
    Estel·la Ramírez‐Baraldes, Daniel García‐Gutiérrez, Cristina García‐Salido
    European Journal of Education.2025;[Epub]     CrossRef
  • Can ChatGPT be used as a scientific source of information on tooth extraction?
    Shiori Yamamoto, Masakazu Hamada, Kyoko Nishiyama, Ayako Motoki, Yusei Fujita, Narikazu Uzawa
    Journal of Oral and Maxillofacial Surgery, Medicine, and Pathology.2025; 37(5): 877.     CrossRef
  • Exploring artificial intelligence (AI) Chatbot usage behaviors and their association with mental health outcomes in Chinese university students
    Xing Zhang, Zhaoqian Li, Mingyang Zhang, Mingyue Yin, Zhangyu Yang, Dong Gao, Hansen Li
    Journal of Affective Disorders.2025; 380: 394.     CrossRef
  • The analysis of optimization in music aesthetic education under artificial intelligence
    Yixuan Peng
    Scientific Reports.2025;[Epub]     CrossRef
  • The Role of Artificial Intelligence in Computer Science Education: A Systematic Review with a Focus on Database Instruction
    Alkmini Gaitantzi, Ioannis Kazanidis
    Applied Sciences.2025; 15(7): 3960.     CrossRef
  • A Bibliometric Exposition and Review on Leveraging LLMs for Programming Education
    Joanah Pwanedo Amos, Oluwatosin Ahmed Amodu, Raja Azlina Raja Mahmood, Akanbi Bolakale Abdulqudus, Anies Faziehan Zakaria, Abimbola Rhoda Iyanda, Umar Ali Bukar, Zurina Mohd Hanapi
    IEEE Access.2025; 13: 58364.     CrossRef
  • Can ChatGPT be trusted as a resource for a scholarly article on treatment planning implant-supported prostheses?
    Steven J. Sadowsky
    The Journal of Prosthetic Dentistry.2025; 134(2): 438.     CrossRef
  • Use of machine translation in foreign language education
    Blanka Klimova
    Cogent Arts & Humanities.2025;[Epub]     CrossRef
  • Comparison of triage performance among DRP tool, ChatGPT, and outpatient rehabilitation doctors
    Yucong Zou, Ruixue Ye, Yan Gao, Jing Zhou, Yawei Li, Wenshi Chen, Fubing Zha, Yulong Wang
    Scientific Reports.2025;[Epub]     CrossRef
  • A Training Needs Analysis for AI and Generative AI in Medical Education: Perspectives of Faculty and Students
    Lise McCoy, Natarajan Ganesan, Viswanathan Rajagopalan, Douglas McKell, Diego F. Niño, Mary Claire Swaim
    Journal of Medical Education and Curricular Development.2025;[Epub]     CrossRef
  • Application of ChatGPT 4.0 in radiological dose management: Perceptions of radiographers with varying expertise
    L. Federico, D.D. Fusaro, G.C. Coppola, M. Gregori, S. Durante
    Radiography.2025; 31(4): 102972.     CrossRef
  • Status and perceptions of ChatGPT utilization among medical students: a survey-based study
    Na Hu, Xiao Qin Jiang, Yi Da Wang, Yan Ming Kang, Zhen Xia, Hao Hui Chen, Sai Nan Duan, Dong Xu Chen
    BMC Medical Education.2025;[Epub]     CrossRef
  • Artificial Intelligence and the Scientific Process: A Review of ChatGPT’s Role to Foster Experimental Thinking in Physics Education
    Konstantinos T. Kotsis
    European Journal of Contemporary Education and E-Learning.2025; 3(3): 183.     CrossRef
  • Performance of GPT-4 in oral and maxillofacial surgery board exams: challenges in specialized questions
    Felix Benjamin Warwas, Nils Heim
    Oral and Maxillofacial Surgery.2025;[Epub]     CrossRef
  • Evaluating Authorship Guidelines of Top Medical Schools and Plastic Surgery Journals
    Nicholas A. Mirsky, Sara E. Munkwitz, Wrood M. Kassira, Pawan Pathagamage, Paulo G. Coelho, Seth R. Thaller
    Annals of Plastic Surgery.2025; 95(5): e53.     CrossRef
  • How appropriately can generative artificial intelligence platforms, including GPT-4, Gemini, Bing, and Wrtn, answer questions about colon cancer in the Korean language?
    Sun Huh
    Annals of Coloproctology.2025; 41(3): 190.     CrossRef
  • The application of problem-based learning (PBL) guided by ChatGPT in clinical education in the Department of Nephrology
    Xiaoya Tong, Ying Hu, Yanjun Long, Qian Zhang, Yao Yang, Jing Yuan, Yan Zha
    BMC Medical Education.2025;[Epub]     CrossRef
  • Mental load, ChatGPT, and work-study balance: A TAM-based study of AI adoption by employees in non-traditional education
    Humayyun Bashir, Rongting Zhou
    Telematics and Informatics Reports.2025; 19: 100216.     CrossRef
  • ChatGPT and AI Chatbots in Education: An Umbrella Review of Systematic Reviews, Scoping Reviews, and Meta-Analyses
    Emmanouil D. Milakis, Constantina Corazon Argyrakou, Alexandros Melidis, John Vrettaros
    International Journal of Education and Information Technologies.2025; 19: 100.     CrossRef
  • Battle of the artificial intelligence: a comprehensive comparative analysis of DeepSeek and ChatGPT for urinary incontinence-related questions
    Huawei Cao, Changzhen Hao, Tao Zhang, Xiang Zheng, Zihao Gao, Jiyue Wu, Lijian Gan, Yu Liu, Xiangjun Zeng, Wei Wang
    Frontiers in Public Health.2025;[Epub]     CrossRef
  • A review of ChatGPT in medical education: exploring advantages and limitations
    Yuan Cheng, Lingling Zhu
    International Journal of Surgery.2025; 111(7): 4586.     CrossRef
  • Artificial Intelligence and Large Language Models in the Fight Against Superficial Fungal Infections: Friend or Foe?
    Aditya Gupta, Vasiliki Economopoulos
    Clinical, Cosmetic and Investigational Dermatology.2025; Volume 18: 1959.     CrossRef
  • Exploring the Effects of the CER Model‐Based GenAI Learning System to Cultivate Elementary School Students' Computational Thinking Core Skills in Science Courses
    Jia‐Hua Zhao, Shu‐Tao Shangguan, Ying Wang
    Journal of Computer Assisted Learning.2025;[Epub]     CrossRef
  • Artificial intelligence in preclinical epilepsy research: Current state, potential, and challenges
    Jesús Servando Medel‐Matus, Cesar Santana‐Gomez, Ruby G. Escalante, Dominique Duncan, Pedro F. Viana, Giulia Sofia Cereda, Naoto Kuroda, Aristea S. Galanopoulou
    Epilepsia Open.2025;[Epub]     CrossRef
  • Exploring the Perceived Usefulness and Effect of AI Writing Tools in Enhancing the Quality of Written Outputs of Senior High School Students
    Roy A. Discutido
    International Journal of Multidisciplinary: Applied Business and Education Research.2025; 6(9): 4267.     CrossRef
  • Üniversite Öğrencilerinin Yapay Zekâ Okuryazarlığı Üzerine Bir Araştırma
    Ela OĞAN
    Üçüncü Sektör Sosyal Ekonomi Dergisi.2025; 60(3): 2670.     CrossRef
  • Generative artificial intelligence in dentistry: A narrative review of current approaches and future challenges
    Fabián Villena, Claudia Véliz, Rosario García-Huidobro, Sebastian Aguayo
    Dentistry Review.2025; 5(4): 100160.     CrossRef
  • A product appearance design method based on artificial intelligence generated content
    Yanpu Yang, Yueming Zhuo, Zhihong Wu, Jialing Liu, Wenhao Meng
    Advanced Design Research.2025; 3(1): 64.     CrossRef
  • ChatGPT in Medical Education: Bibliometric and Visual Analysis
    Yuning Zhang, Xiaolu Xie, Qi Xu
    JMIR Medical Education.2025; 11: e72356.     CrossRef
  • Collaboration, Digital Tools, and AI in Academic Writing: Student Experiences, Challenges, and Perspectives
    Mehmet Kanık
    Dil Eğitimi ve Araştırmaları Dergisi.2025; 11(2): 1038.     CrossRef
  • Detection of Medical Misinformation in Hemangioma Patient Education: Comparative Study of ChatGPT-4o and DeepSeek-R1 Large Language Models
    Guoyong Wang, Ye Zhang, Weixin Wang, Yingjie Zhu, Wei Lu, Chaonan Wang, Hui Bi, Xiaonan Yang
    JMIR AI.2025; 4: e76372.     CrossRef
  • AI algorithm bias awareness, ethical concerns and use of AI for information retrieval among LIS students in Nigeria
    Omorodion Okuonghae, Ifeyinwa Nkechi Okonkwo, Nosakhare Okuonghae, Magnus Osahon Igbinovia
    Alexandria: The Journal of National and International Library and Information Issues.2025;[Epub]     CrossRef
  • Perceptions of AI-based tools among polish medical university students
    Piotr Ratajczak, Oliwia Słowik, Julia Cynar, Dorota Kopciuch, Anna Paczkowska, Tomasz Zaprutko, Krzysztof Kus
    BMC Medical Education.2025;[Epub]     CrossRef
  • Artificial intelligence: a new era in editing academic and popular science writing
    Mykola Zhelezniak, Yuliya Mishura
    Entsyklopedychnyi visnyk Ukrainy [The Encyclopedia Herald of Ukraine].2025; 17: 9.     CrossRef
  • Factors influencing users’ attitudes towards intelligent chatbots in Chinese academic libraries: the role of algorithm literacy
    Heng Lu, Xin Li
    Humanities and Social Sciences Communications.2025;[Epub]     CrossRef
  • Uso de la inteligencia artificial en estrategias de repetición espaciada para la educación médica y el aprendizaje significativo: revisión sistemática.
    Jennifer Carolina Duque Espinel, Andrés Leonardo García Casquete, Yoiler Batista Garcet
    Revista Española de Educación Médica.2025;[Epub]     CrossRef
  • Python’s Contribution to Artificial Intelligence in Education: A state-of-the-art review
    Alexandros Papadimitriou
    Journal of Future Artificial Intelligence and Technologies.2025; 2(3): 445.     CrossRef
  • Awareness and practice of using public generative AI solutions (such as ChatGPT) and social media among psychiatrists compared to other professionals: A pilot study
    Ema Nicea Gruber, Lucija Gruber Zlatec, Sanja Martić Biočina
    PSYCHIATRIA DANUBINA.2025; 37(2): 141.     CrossRef
  • Impact of Artificial Intelligence on College and University Students: A Global Transformation of the Education System
    Steward Mudenda
    Creative Education.2025; 16(11): 1801.     CrossRef
  • Opportunities, challenges, and future directions of large language models, including ChatGPT in medical education: a systematic scoping review
    Xiaojun Xu, Yixiao Chen, Jing Miao
    Journal of Educational Evaluation for Health Professions.2024; 21: 6.     CrossRef
  • Artificial Intelligence: Fundamentals and Breakthrough Applications in Epilepsy
    Wesley Kerr, Sandra Acosta, Patrick Kwan, Gregory Worrell, Mohamad A. Mikati
    Epilepsy Currents.2024;[Epub]     CrossRef
  • A Developed Graphical User Interface-Based on Different Generative Pre-trained Transformers Models
    Ekrem Küçük, İpek Balıkçı Çiçek, Zeynep Küçükakçalı, Cihan Yetiş, Cemil Çolak
    ODÜ Tıp Dergisi.2024; 11(1): 18.     CrossRef
  • Art or Artifact: Evaluating the Accuracy, Appeal, and Educational Value of AI-Generated Imagery in DALL·E 3 for Illustrating Congenital Heart Diseases
    Mohamad-Hani Temsah, Abdullah N. Alhuzaimi, Mohammed Almansour, Fadi Aljamaan, Khalid Alhasan, Munirah A. Batarfi, Ibraheem Altamimi, Amani Alharbi, Adel Abdulaziz Alsuhaibani, Leena Alwakeel, Abdulrahman Abdulkhaliq Alzahrani, Khaled B. Alsulaim, Amr Jam
    Journal of Medical Systems.2024;[Epub]     CrossRef
  • Authentic assessment in medical education: exploring AI integration and student-as-partners collaboration
    Syeda Sadia Fatima, Nabeel Ashfaque Sheikh, Athar Osama
    Postgraduate Medical Journal.2024; 100(1190): 959.     CrossRef
  • Comparative performance analysis of large language models: ChatGPT-3.5, ChatGPT-4 and Google Gemini in glucocorticoid-induced osteoporosis
    Linjian Tong, Chaoyang Zhang, Rui Liu, Jia Yang, Zhiming Sun
    Journal of Orthopaedic Surgery and Research.2024;[Epub]     CrossRef
  • Can AI-Generated Clinical Vignettes in Japanese Be Used Medically and Linguistically?
    Yasutaka Yanagita, Daiki Yokokawa, Shun Uchida, Yu Li, Takanori Uehara, Masatomi Ikusaka
    Journal of General Internal Medicine.2024; 39(16): 3282.     CrossRef
  • ChatGPT vs. sleep disorder specialist responses to common sleep queries: Ratings by experts and laypeople
    Jiyoung Kim, Seo-Young Lee, Jee Hyun Kim, Dong-Hyeon Shin, Eun Hye Oh, Jin A Kim, Jae Wook Cho
    Sleep Health.2024; 10(6): 665.     CrossRef
  • Technology integration into Chinese as a foreign language learning in higher education: An integrated bibliometric analysis and systematic review (2000–2024)
    Binze Xu
    Language Teaching Research.2024;[Epub]     CrossRef
  • The Transformative Power of Generative Artificial Intelligence for Achieving the Sustainable Development Goal of Quality Education
    Prema Nedungadi, Kai-Yu Tang, Raghu Raman
    Sustainability.2024; 16(22): 9779.     CrossRef
  • Is AI the new course creator
    Sheri Conklin, Tom Dorgan, Daisyane Barreto
    Discover Education.2024;[Epub]     CrossRef
  • Emergency Medicine Assistants in the Field of Toxicology, Comparison of ChatGPT-3.5 and GEMINI Artificial Intelligence Systems
    Hatice Aslı Bedel, Cihan Bedel, Fatih Selvi, Ökkeş Zortuk, Yusuf Karanci
    Acta medica Lituanica.2024; 31(2): 294.     CrossRef
Brief report
ChatGPT (GPT-3.5) as an assistant tool in microbial pathogenesis studies in Sweden: a cross-sectional comparative study  
Catharina Hultgren, Annica Lindkvist, Volkan Özenci, Sophie Curbo
J Educ Eval Health Prof. 2023;20:32.   Published online November 22, 2023
DOI: https://doi.org/10.3352/jeehp.2023.20.32
  • 5,285 View
  • 184 Download
  • 7 Web of Science
  • 7 Crossref
AbstractAbstract PDFSupplementary Material
ChatGPT (GPT-3.5) has entered higher education and there is a need to determine how to use it effectively. This descriptive study compared the ability of GPT-3.5 and teachers to answer questions from dental students and construct detailed intended learning outcomes. When analyzed according to a Likert scale, we found that GPT-3.5 answered the questions from dental students in a similar or even more elaborate way compared to the answers that had previously been provided by a teacher. GPT-3.5 was also asked to construct detailed intended learning outcomes for a course in microbial pathogenesis, and when these were analyzed according to a Likert scale they were, to a large degree, found irrelevant. Since students are using GPT-3.5, it is important that instructors learn how to make the best use of it both to be able to advise students and to benefit from its potential.

Citations

Citations to this article as recorded by  
  • Global Trends in the Use of Artificial Intelligence in Dental Education: A Bibliometric Analysis
    Margarita Iniesta, Juan José Pérez‐Higueras
    European Journal of Dental Education.2026; 30(2): 427.     CrossRef
  • Educational Applications of ChatGPT in University‐Based Dental Education. A Systematic Review
    Juan Ignacio Aura‐Tormos, Maria Llacer‐Martinez, Ines Torres‐Osca
    European Journal of Dental Education.2026; 30(2): 644.     CrossRef
  • Choosing not to use generative AI in higher education: a mixed methods study of student reasoning, assessment design, and AI literacy
    Annica Lindkvist, Sophie Curbo, Catharina Hultgren
    Frontiers in Education.2026;[Epub]     CrossRef
  • Unlocking learning: exploring take-home examinations and viva voce examinations in microbiology education for biomedical laboratory science students
    Sophie Curbo, Annica Lindkvist, Catharina Hultgren, Jorge Cervantes
    Journal of Microbiology & Biology Education.2025;[Epub]     CrossRef
  • How Can Clinicians Leverage Vibe Coding for Machine Learning and Deep Learning Research?
    Yoonhwan Lee, Sun Huh
    Endocrinology and Metabolism.2025; 40(5): 659.     CrossRef
  • Opportunities, challenges, and future directions of large language models, including ChatGPT in medical education: a systematic scoping review
    Xiaojun Xu, Yixiao Chen, Jing Miao
    Journal of Educational Evaluation for Health Professions.2024; 21: 6.     CrossRef
  • Information amount, accuracy, and relevance of generative artificial intelligence platforms’ answers regarding learning objectives of medical arthropodology evaluated in English and Korean queries in December 2023: a descriptive study
    Hyunju Lee, Soobin Park
    Journal of Educational Evaluation for Health Professions.2023; 20: 39.     CrossRef
Research articles
Performance of ChatGPT, Bard, Claude, and Bing on the Peruvian National Licensing Medical Examination: a cross-sectional study  
Betzy Clariza Torres-Zegarra, Wagner Rios-Garcia, Alvaro Micael Ñaña-Cordova, Karen Fatima Arteaga-Cisneros, Xiomara Cristina Benavente Chalco, Marina Atena Bustamante Ordoñez, Carlos Jesus Gutierrez Rios, Carlos Alberto Ramos Godoy, Kristell Luisa Teresa Panta Quezada, Jesus Daniel Gutierrez-Arratia, Javier Alejandro Flores-Cohaila
J Educ Eval Health Prof. 2023;20:30.   Published online November 20, 2023
DOI: https://doi.org/10.3352/jeehp.2023.20.30
  • 10,526 View
  • 306 Download
  • 45 Web of Science
  • 47 Crossref
AbstractAbstract PDFSupplementary Material
Purpose
We aimed to describe the performance and evaluate the educational value of justifications provided by artificial intelligence chatbots, including GPT-3.5, GPT-4, Bard, Claude, and Bing, on the Peruvian National Medical Licensing Examination (P-NLME).
Methods
This was a cross-sectional analytical study. On July 25, 2023, each multiple-choice question (MCQ) from the P-NLME was entered into each chatbot (GPT-3, GPT-4, Bing, Bard, and Claude) 3 times. Then, 4 medical educators categorized the MCQs in terms of medical area, item type, and whether the MCQ required Peru-specific knowledge. They assessed the educational value of the justifications from the 2 top performers (GPT-4 and Bing).
Results
GPT-4 scored 86.7% and Bing scored 82.2%, followed by Bard and Claude, and the historical performance of Peruvian examinees was 55%. Among the factors associated with correct answers, only MCQs that required Peru-specific knowledge had lower odds (odds ratio, 0.23; 95% confidence interval, 0.09–0.61), whereas the remaining factors showed no associations. In assessing the educational value of justifications provided by GPT-4 and Bing, neither showed any significant differences in certainty, usefulness, or potential use in the classroom.
Conclusion
Among chatbots, GPT-4 and Bing were the top performers, with Bing performing better at Peru-specific MCQs. Moreover, the educational value of justifications provided by the GPT-4 and Bing could be deemed appropriate. However, it is essential to start addressing the educational value of these chatbots, rather than merely their performance on examinations.

Citations

Citations to this article as recorded by  
  • ChatGPT vs Claude: Scoping Review with ☸️SAIMSARA

    SAIMSARA Journal.2026;[Epub]     CrossRef
  • Consistency over accuracy: run-to-run stability of contemporary large language models on Turkish curriculum-aligned theoretical anatomy multiple-choice questions
    Ömer Alperen Gürses, İsmail Ceylan
    BMC Medical Education.2026;[Epub]     CrossRef
  • Evaluation of large language models in medical examinations: A scoping review protocol
    Weiqi Wang, Baifeng Wang, Yan Zhu, Zhe Wang, Suyuan Peng, Haiyang Chen
    PLOS One.2026; 21(4): e0347539.     CrossRef
  • Comparative performance of artificial intelligence models in intensive care nursing questions: an evaluation of ChatGPT, DeepSeek, and Google Gemini
    Seçil Gülhan Güner, Ziya Tan, Serhat Gülpınar
    BMC Nursing.2026;[Epub]     CrossRef
  • Evaluation of large language models in cardiovascular surgery: a comparative study of board-level clinical question answering and generation
    Mehmet Inanc Yesilkaya, Uzeyir Yilmaz
    Journal of Cardiothoracic Surgery.2026;[Epub]     CrossRef
  • PICOT questions and search strategies formulation: A novel approach using artificial intelligence automation
    Lucija Gosak, Gregor Štiglic, Lisiane Pruinelli, Dominika Vrbnjak
    Journal of Nursing Scholarship.2025; 57(1): 5.     CrossRef
  • Using large language models (ChatGPT, Copilot, PaLM, Bard, and Gemini) in Gross Anatomy course: Comparative analysis
    Volodymyr Mavrych, Paul Ganguly, Olena Bolgova
    Clinical Anatomy.2025; 38(2): 200.     CrossRef
  • Capable exam-taker and question-generator: the dual role of generative AI in medical education assessment
    Yihong Qiu, Chang Liu
    Global Medical Education.2025; 2(1): 135.     CrossRef
  • Comparison of artificial intelligence systems in answering prosthodontics questions from the dental specialty exam in Turkey
    Busra Tosun, Zeynep Sen Yilmaz
    Journal of Dental Sciences.2025; 20(3): 1454.     CrossRef
  • Benchmarking LLM chatbots’ oncological knowledge with the Turkish Society of Medical Oncology’s annual board examination questions
    Efe Cem Erdat, Engin Eren Kavak
    BMC Cancer.2025;[Epub]     CrossRef
  • Evaluating the Performance of Large Language Models in Anatomy Education Advancing Anatomy Learning with ChatGPT-4o
    Fatma Ok, Burak Karip, Fulya Temizsoy Korkmaz
    European Journal of Therapeutics.2025; 31(1): 35.     CrossRef
  • Large Language Models in Biochemistry Education: Comparative Evaluation of Performance
    Olena Bolgova, Inna Shypilova, Volodymyr Mavrych
    JMIR Medical Education.2025; 11: e67244.     CrossRef
  • Attributional patterns toward students with and without learning disabilities: Artificial intelligence models vs. trainee teachers
    Inbar Levkovich, Eyal Rabin, Rania Hussein Farraj, Zohar Elyoseph
    Research in Developmental Disabilities.2025; 160: 104970.     CrossRef
  • The double-edged sword of generative AI: surpassing an expert or a deceptive “false friend”?
    Franziska C.S. Altorfer, Michael J. Kelly, Fedan Avrumova, Varun Rohatgi, Jiaqi Zhu, Christopher M. Bono, Darren R. Lebl
    The Spine Journal.2025; 25(8): 1635.     CrossRef
  • Claude, ChatGPT, Copilot, and Gemini performance versus students in different topics of neuroscience
    Volodymyr Mavrych, Ahmed Yaqinuddin, Olena Bolgova
    Advances in Physiology Education.2025; 49(2): 430.     CrossRef
  • Large Language Models Take on the AAMC Situational Judgment Test: Evaluating Dilemma-Based Scenarios
    Angelo Cadiente, Jamie Chen, Lora J. Kasselman, Bryan Pilkington
    International Journal of Artificial Intelligence in Education.2025; 35(4): 2202.     CrossRef
  • Validation of a generative artificial intelligence tool for the critical appraisal of articles on the epidemiology of mental health: Its application in the Middle East and North Africa
    Cheima Moussa, Sarah Altayyar, Marion Vergonjeanne, Thibaut Gelle, Pierre-Marie Preux
    Journal of Epidemiology and Population Health.2025; 73(2): 202990.     CrossRef
  • Temporal Association Between ChatGPT-Generated Diarrhea Synonyms in Internet Search Queries and Emergency Department Visits for Diarrhea-Related Symptoms in South Korea: Exploratory Study
    Jinsoo Kim, Ansun Jeong, Juseong Jin, Sangjun Lee, Do Kyoon Yoon, Soyeoun Kim
    Journal of Medical Internet Research.2025; 27: e65101.     CrossRef
  • Artificial intelligence (AI) performance on pharmacy skills laboratory course assignments
    Vivian Do, Krista L. Donohoe, Apryl N. Peddi, Eleanor Carr, Christina Kim, Virginia Mele, Dhruv Patel, Alexis N. Crawford
    Currents in Pharmacy Teaching and Learning.2025; 17(7): 102367.     CrossRef
  • Accuracy of Large Language Models When Answering Clinical Research Questions: Systematic Review and Network Meta-Analysis
    Ling Wang, Jinglin Li, Boyang Zhuang, Shasha Huang, Meilin Fang, Cunze Wang, Wen Li, Mohan Zhang, Shurong Gong
    Journal of Medical Internet Research.2025; 27: e64486.     CrossRef
  • Performance of artificial intelligence chatbots in National dental licensing examination
    Chad Chan-Chia Lin, Jui-Sheng Sun, Chin-Hao Chang, Yu-Han Chang, Jenny Zwei-Chieng Chang
    Journal of Dental Sciences.2025; 20(4): 2307.     CrossRef
  • Does AI have utility in medical student surgical education? A comparative analysis of chatbots in answering standardized surgical multiple-choice questions
    Natalia DaFonte, Angelo Cadiente, Catherine Implicito, Natasha Becker, Burton Surick
    Global Surgical Education - Journal of the Association for Surgical Education.2025;[Epub]     CrossRef
  • A review of ChatGPT in medical education: exploring advantages and limitations
    Yuan Cheng, Lingling Zhu
    International Journal of Surgery.2025; 111(7): 4586.     CrossRef
  • Applications, Challenges, and Prospects of Generative Artificial Intelligence Empowering Medical Education: Scoping Review
    Yuhang Lin, Zhiheng Luo, Zicheng Ye, Nuoxi Zhong, Lijian Zhao, Long Zhang, Xiaolan Li, Zetao Chen, Yijia Chen
    JMIR Medical Education.2025; 11: e71125.     CrossRef
  • Collaborative intelligence in AI: Evaluating the performance of a council of AIs on the USMLE
    Yahya Shaikh, Zainab Asiyah Jeelani-Shaikh, Muzamillah Mushtaq Jeelani, Aamir Javaid, Tauhid Mahmud, Shiv Gaglani, Michael Christopher Gibbons, Minahil Cheema, Amanda Cross, Denisa Livingston, Morgan Cheatham, Elahe Nezami, Ronald Dixon, Ashwini Niranjan-
    PLOS Digital Health.2025; 4(10): e0000787.     CrossRef
  • ChatGPT in Medical Education: Bibliometric and Visual Analysis
    Yuning Zhang, Xiaolu Xie, Qi Xu
    JMIR Medical Education.2025; 11: e72356.     CrossRef
  • Performance of ChatGPT and Large Language Models on Medical Licensing Exams Worldwide: A Systematic Review and Network Meta-Analysis With Meta-Regression
    Alousious Kasagga, Aayam Sapkota, Gichin Changaramkumarath, Jane M Abucha, Mekdes M Wollel, Nethra Somannagari, Malik Y Husami, Kirubel T Hailu, Ecem Kasagga
    Cureus.2025;[Epub]     CrossRef
  • Performance of GPT-4o and o1-Pro on United Kingdom Medical Licensing Assessment-style items: a comparative study
    Behrad Vakili, Aadam Ahmad, Mahsa Zolfaghari
    Journal of Educational Evaluation for Health Professions.2025; 22: 30.     CrossRef
  • Performance of large language models in fluoride-related dental knowledge: a comparative evaluation study of ChatGPT-4, Claude 3.5 Sonnet, Copilot, and Grok 3
    Raju Biswas, Atanu Mukhopadhyay, Santanu Mukhopadhyay
    Journal of Yeungnam Medical Science.2025; 42: 53.     CrossRef
  • Comparative performance of large language models in answering periodontology questions from the Turkish Dental Specialty Examination: a cross-sectional study on accuracy and coverage
    Muzeyyen Kandemir, Ebru Ece Sarıbaş
    BMC Oral Health.2025;[Epub]     CrossRef
  • Comparison of chatbots’ accuracy in endodontics questions in dentistry specialization exam in Türkiye: ChatGPT-4o, Gemini Advanced, Copilot, and Claude
    Ceren Turan Gökduman, Esra Arılı Öztürk, Şule Aktaş, Burhan Can Çanakçi̇
    BMC Oral Health.2025;[Epub]     CrossRef
  • Performance comparison of large language models on pediatric dentistry questions in the Turkish dentistry specialization examination
    Hatice Kübra Başkan, Beyhan Başkan
    BMC Medical Education.2025;[Epub]     CrossRef
  • Assessment of artificial intelligence chatbots in responding to dental occlusion questions: a comparative study
    Hamod Alqahtani
    BMC Oral Health.2025;[Epub]     CrossRef
  • Performance of Artificial Intelligence Chatbots on Standardized Medical Examination Questions in Obstetrics & Gynecology
    Angelo Cadiente, Natalia DaFonte, Jonathan D. Baum
    Open Journal of Obstetrics and Gynecology.2025; 15(01): 1.     CrossRef
  • Performance of GPT-4V in Answering the Japanese Otolaryngology Board Certification Examination Questions: Evaluation Study
    Masao Noda, Takayoshi Ueno, Ryota Koshu, Yuji Takaso, Mari Dias Shimada, Chizu Saito, Hisashi Sugimoto, Hiroaki Fushiki, Makoto Ito, Akihiro Nomura, Tomokazu Yoshizaki
    JMIR Medical Education.2024; 10: e57054.     CrossRef
  • Response to Letter to the Editor re: “Artificial Intelligence Versus Expert Plastic Surgeon: Comparative Study Shows ChatGPT ‘Wins' Rhinoplasty Consultations: Should We Be Worried? [1]” by Durairaj et al
    Kay Durairaj, Omer Baker
    Facial Plastic Surgery & Aesthetic Medicine.2024; 26(3): 276.     CrossRef
  • Opportunities, challenges, and future directions of large language models, including ChatGPT in medical education: a systematic scoping review
    Xiaojun Xu, Yixiao Chen, Jing Miao
    Journal of Educational Evaluation for Health Professions.2024; 21: 6.     CrossRef
  • Performance of ChatGPT Across Different Versions in Medical Licensing Examinations Worldwide: Systematic Review and Meta-Analysis
    Mingxin Liu, Tsuyoshi Okuhara, XinYi Chang, Ritsuko Shirabe, Yuriko Nishiie, Hiroko Okada, Takahiro Kiuchi
    Journal of Medical Internet Research.2024; 26: e60807.     CrossRef
  • Comparative accuracy of ChatGPT-4, Microsoft Copilot and Google Gemini in the Italian entrance test for healthcare sciences degrees: a cross-sectional study
    Giacomo Rossettini, Lia Rodeghiero, Federica Corradi, Chad Cook, Paolo Pillastrini, Andrea Turolla, Greta Castellini, Stefania Chiappinotto, Silvia Gianola, Alvisa Palese
    BMC Medical Education.2024;[Epub]     CrossRef
  • Evaluating the competency of ChatGPT in MRCP Part 1 and a systematic literature review of its capabilities in postgraduate medical assessments
    Oliver Vij, Henry Calver, Nikki Myall, Mrinalini Dey, Koushan Kouranloo, Thiago P. Fernandes
    PLOS ONE.2024; 19(7): e0307372.     CrossRef
  • Large Language Models in Pediatric Education: Current Uses and Future Potential
    Srinivasan Suresh, Sanghamitra M. Misra
    Pediatrics.2024;[Epub]     CrossRef
  • Comparison of the Performance of ChatGPT, Claude and Bard in Support of Myopia Prevention and Control
    Yan Wang, Lihua Liang, Ran Li, Yihua Wang, Changfu Hao
    Journal of Multidisciplinary Healthcare.2024; Volume 17: 3917.     CrossRef
  • Evaluating Large Language Models in Dental Anesthesiology: A Comparative Analysis of ChatGPT-4, Claude 3 Opus, and Gemini 1.0 on the Japanese Dental Society of Anesthesiology Board Certification Exam
    Misaki Fujimoto, Hidetaka Kuroda, Tomomi Katayama, Atsuki Yamaguchi, Norika Katagiri, Keita Kagawa, Shota Tsukimoto, Akito Nakano, Uno Imaizumi, Aiji Sato-Boku, Naotaka Kishimoto, Tomoki Itamiya, Kanta Kido, Takuro Sanuki
    Cureus.2024;[Epub]     CrossRef
  • Dermatological Knowledge and Image Analysis Performance of Large Language Models Based on Specialty Certificate Examination in Dermatology
    Ka Siu Fan, Ka Hay Fan
    Dermato.2024; 4(4): 124.     CrossRef
  • ChatGPT and Other Large Language Models in Medical Education — Scoping Literature Review
    Alexandra Aster, Matthias Carl Laupichler, Tamina Rockwell-Kollmann, Gilda Masala, Ebru Bala, Tobias Raupach
    Medical Science Educator.2024; 35(1): 555.     CrossRef
  • Performance of ChatGPT and Bard on the medical licensing examinations varies across different cultures: a comparison study
    Yikai Chen, Xiujie Huang, Fangjie Yang, Haiming Lin, Haoyu Lin, Zhuoqun Zheng, Qifeng Liang, Jinhai Zhang, Xinxin Li
    BMC Medical Education.2024;[Epub]     CrossRef
  • Information amount, accuracy, and relevance of generative artificial intelligence platforms’ answers regarding learning objectives of medical arthropodology evaluated in English and Korean queries in December 2023: a descriptive study
    Hyunju Lee, Soobin Park
    Journal of Educational Evaluation for Health Professions.2023; 20: 39.     CrossRef
Medical students’ patterns of using ChatGPT as a feedback tool and perceptions of ChatGPT in a Leadership and Communication course in Korea: a cross-sectional study  
Janghee Park
J Educ Eval Health Prof. 2023;20:29.   Published online November 10, 2023
DOI: https://doi.org/10.3352/jeehp.2023.20.29
  • 9,985 View
  • 332 Download
  • 20 Web of Science
  • 24 Crossref
AbstractAbstract PDFSupplementary Material
Purpose
This study aimed to analyze patterns of using ChatGPT before and after group activities and to explore medical students’ perceptions of ChatGPT as a feedback tool in the classroom.
Methods
The study included 99 2nd-year pre-medical students who participated in a “Leadership and Communication” course from March to June 2023. Students engaged in both individual and group activities related to negotiation strategies. ChatGPT was used to provide feedback on their solutions. A survey was administered to assess students’ perceptions of ChatGPT’s feedback, its use in the classroom, and the strengths and challenges of ChatGPT from May 17 to 19, 2023.
Results
The students responded by indicating that ChatGPT’s feedback was helpful, and revised and resubmitted their group answers in various ways after receiving feedback. The majority of respondents expressed agreement with the use of ChatGPT during class. The most common response concerning the appropriate context of using ChatGPT’s feedback was “after the first round of discussion, for revisions.” There was a significant difference in satisfaction with ChatGPT’s feedback, including correctness, usefulness, and ethics, depending on whether or not ChatGPT was used during class, but there was no significant difference according to gender or whether students had previous experience with ChatGPT. The strongest advantages were “providing answers to questions” and “summarizing information,” and the worst disadvantage was “producing information without supporting evidence.”
Conclusion
The students were aware of the advantages and disadvantages of ChatGPT, and they had a positive attitude toward using ChatGPT in the classroom.

Citations

Citations to this article as recorded by  
  • An Alternative Approach in Anatomy Education: Design of a Learning Environment Based on Artificial Intelligence‐Supported Virtual Manipulatives and Investigation of Its Effectiveness
    Gunes Bolatli, Salih Birisci, Zafer Bolatli
    Clinical Anatomy.2026; 39(1): 30.     CrossRef
  • Applications and Outcomes of Large‑Language‑Model‑Generated Feedback in Undergraduate Medical Education: A Scoping Review
    Yavuz Selim Kıyak, Tuğba İş-Kara, Emre Emekli
    Medical Science Educator.2026; 36(1): 81.     CrossRef
  • Attitudes and perceptions of the application of large language models among health professionals: A mixed-methods systematic review
    Wen Luo, Tao Feng, Ting Zhang, Xinyu Chen, Xianying Lu, Yuhang Li, Chaoming Hou, Jing Gao
    Public Health.2026; 254: 106252.     CrossRef
  • Generative AI's Impact on the Mental Health of Medical Students: Scenario Analysis
    Nora Arvai, Bertalan Meskó, Gellért Katonai
    JMIR Medical Education.2026; 12: e85373.     CrossRef
  • Mapping the landscape of AI-driven feedback in education: a scoping review
    Anastasiya A. Lipnevich, Carmen A. Taranto, Christopher DeLuca, Ephraim Nukpetsi, Nathan Rickey, Therese Hopfenbeck, Emma Carter, Joshua McGrane
    Frontiers in Education.2026;[Epub]     CrossRef
  • The Impact of ChatGPT on Higher Education: A Systematic Review of Global Opportunities, Perceptions, and Challenges
    Olukayode Emmanuel Apata, Oi‐Man Kwok, Segun Timothy Ajose
    Journal of Computer Assisted Learning.2026;[Epub]     CrossRef
  • Student Perceptions of Virtual Counseling Practice Using ChatGPT Voice Mode in Pharmacy Education
    Nuntapong Boonrit, Ashley M. Hopkins, Warit Ruanglertboon
    JACCP: JOURNAL OF THE AMERICAN COLLEGE OF CLINICAL PHARMACY.2026;[Epub]     CrossRef
  • Education Research: Integrating AI-Enabled Interactive Case-Based Learning in a Preclinical Neurosciences Course
    Tamara B. Kaplan, Kaiying Wang, Oliver Bichsel, Stephen Bacchi, Galina Gheihman
    Neurology Education.2026;[Epub]     CrossRef
  • A comparative evaluation of ChatGPT-assisted and traditional methods of teaching diagnostic assessments to medical students
    Hua Xu, Yu Sun, Zhi-Wen Mo, Yu-Nan Man, San-Mao Liu, Yu-Ning Wei, Mao-Lin He
    Medicine.2026; 105(36): e50512.     CrossRef
  • Higher education students’ perceptions of ChatGPT: A global study of early reactions
    Dejan Ravšelj, Damijana Keržič, Nina Tomaževič, Lan Umek, Nejc Brezovar, Noorminshah A. Iahad, Ali Abdulla Abdulla, Anait Akopyan, Magdalena Waleska Aldana Segura, Jehan AlHumaid, Mohamed Farouk Allam, Maria Alló, Raphael Papa Kweku Andoh, Octavian Andron
    PLOS ONE.2025; 20(2): e0315011.     CrossRef
  • Generative AI in Otolaryngology Residency Personal Statement Writing: A Mixed‐Methods Analysis
    Jacob G. J. Wihlidal, Nikolaus E. Wolter, Evan J. Propst, Vincent Lin, Michael Au, Shaunak Amin, Jennifer M. Siu
    The Laryngoscope.2025; 135(10): 3570.     CrossRef
  • Feasibility of a Randomized Controlled Trial of Large AI-Based Linguistic Models for Clinical Reasoning Training of Physical Therapy Students: Pilot Randomized Parallel-Group Study
    Raúl Ferrer-Peña, Silvia Di-Bonaventura, Alberto Pérez-González, Alfredo Lerín-Calvo
    JMIR Formative Research.2025; 9: e66126.     CrossRef
  • Applications of Artificial Intelligence for Nonpsychomotor Skills Training in Health Professions Education: A Scoping Review
    Kenya A Costa-Dookhan, Zachary Adirim, Marta Maslej, Kayle Donner, Terri Rodak, Sophie Soklaridis, Sanjeev Sockalingam, Anupam Thakur
    Academic Medicine.2025; 100(5): 635.     CrossRef
  • MD Student Perceptions of ChatGPT for Reflective Writing Feedback in Undergraduate Medical Education
    Nabil Haider, Leo Morjaria, Urmi Sheth, Nujud Al-Jabouri, Matthew Sibbald
    International Medical Education.2025; 4(3): 27.     CrossRef
  • Exploring medical students’ attitudes and perceptions toward artificial intelligence in medicine in Shandong Province, China
    Mingchan Liu, Yi Cheng, Shu Li, Shanshan Wang, Feng Du, Xiaonan Wei, Zhiying Ai, Siyuan Yan
    BMC Medical Education.2025;[Epub]     CrossRef
  • How Can Clinicians Leverage Vibe Coding for Machine Learning and Deep Learning Research?
    Yoonhwan Lee, Sun Huh
    Endocrinology and Metabolism.2025; 40(5): 659.     CrossRef
  • Opportunities, challenges, and future directions of large language models, including ChatGPT in medical education: a systematic scoping review
    Xiaojun Xu, Yixiao Chen, Jing Miao
    Journal of Educational Evaluation for Health Professions.2024; 21: 6.     CrossRef
  • Embracing ChatGPT for Medical Education: Exploring Its Impact on Doctors and Medical Students
    Yijun Wu, Yue Zheng, Baijie Feng, Yuqi Yang, Kai Kang, Ailin Zhao
    JMIR Medical Education.2024; 10: e52483.     CrossRef
  • Integration of ChatGPT Into a Course for Medical Students: Explorative Study on Teaching Scenarios, Students’ Perception, and Applications
    Anita V Thomae, Claudia M Witt, Jürgen Barth
    JMIR Medical Education.2024; 10: e50545.     CrossRef
  • A cross sectional investigation of ChatGPT-like large language models application among medical students in China
    Guixia Pan, Jing Ni
    BMC Medical Education.2024;[Epub]     CrossRef
  • A Pilot Study of Medical Student Opinions on Large Language Models
    Alan Y Xu, Vincent S Piranio, Skye Speakman, Chelsea D Rosen, Sally Lu, Chris Lamprecht, Robert E Medina, Maisha Corrielus, Ian T Griffin, Corinne E Chatham, Nicolas J Abchee, Daniel Stribling, Phuong B Huynh, Heather Harrell, Benjamin Shickel, Meghan Bre
    Cureus.2024;[Epub]     CrossRef
  • The intent of ChatGPT usage and its robustness in medical proficiency exams: a systematic review
    Tatiana Chaiban, Zeinab Nahle, Ghaith Assi, Michelle Cherfane
    Discover Education.2024;[Epub]     CrossRef
  • ChatGPT and Clinical Training: Perception, Concerns, and Practice of Pharm-D Students
    Mohammed Zawiah, Fahmi Al-Ashwal, Lobna Gharaibeh, Rana Abu Farha, Karem Alzoubi, Khawla Abu Hammour, Qutaiba A Qasim, Fahd Abrah
    Journal of Multidisciplinary Healthcare.2023; Volume 16: 4099.     CrossRef
  • Information amount, accuracy, and relevance of generative artificial intelligence platforms’ answers regarding learning objectives of medical arthropodology evaluated in English and Korean queries in December 2023: a descriptive study
    Hyunju Lee, Soobin Park
    Journal of Educational Evaluation for Health Professions.2023; 20: 39.     CrossRef
Efficacy and limitations of ChatGPT as a biostatistical problem-solving tool in medical education in Serbia: a descriptive study  
Aleksandra Ignjatović, Lazar Stevanović
J Educ Eval Health Prof. 2023;20:28.   Published online October 16, 2023
DOI: https://doi.org/10.3352/jeehp.2023.20.28
  • 11,123 View
  • 301 Download
  • 32 Web of Science
  • 33 Crossref
AbstractAbstract PDFSupplementary Material
Purpose
This study aimed to assess the performance of ChatGPT (GPT-3.5 and GPT-4) as a study tool in solving biostatistical problems and to identify any potential drawbacks that might arise from using ChatGPT in medical education, particularly in solving practical biostatistical problems.
Methods
ChatGPT was tested to evaluate its ability to solve biostatistical problems from the Handbook of Medical Statistics by Peacock and Peacock in this descriptive study. Tables from the problems were transformed into textual questions. Ten biostatistical problems were randomly chosen and used as text-based input for conversation with ChatGPT (versions 3.5 and 4).
Results
GPT-3.5 solved 5 practical problems in the first attempt, related to categorical data, cross-sectional study, measuring reliability, probability properties, and the t-test. GPT-3.5 failed to provide correct answers regarding analysis of variance, the chi-square test, and sample size within 3 attempts. GPT-4 also solved a task related to the confidence interval in the first attempt and solved all questions within 3 attempts, with precise guidance and monitoring.
Conclusion
The assessment of both versions of ChatGPT performance in 10 biostatistical problems revealed that GPT-3.5 and 4’s performance was below average, with correct response rates of 5 and 6 out of 10 on the first attempt. GPT-4 succeeded in providing all correct answers within 3 attempts. These findings indicate that students must be aware that this tool, even when providing and calculating different statistical analyses, can be wrong, and they should be aware of ChatGPT’s limitations and be careful when incorporating this model into medical education.

Citations

Citations to this article as recorded by  
  • The Pediatric Surgeon's AI Toolbox: How Large Language Models Like ChatGPT Are Simplifying Practice and Expanding Global Access
    Carlos Andres Colunga Tinajero
    European Journal of Pediatric Surgery.2026; 36(03): 190.     CrossRef
  • Reliability of ChatGPT-4o in analysing medical data: a test case study on patients at risk for limb amputation
    Liat Toderis, Iris Reychav, Roger McHaney, Bernice Oberman, Chen Speter, Ronen Loebstein
    Health Systems.2026; 15(2): 125.     CrossRef
  • Editorial: Generative AI, human authorship and the transformation of scholarly communication
    Luis Hernan Contreras Pinochet, Kavita Miadaira Hamza, Yogesh Kumar Dwivedi
    Revista de Gestão.2026; 33: 31.     CrossRef
  • Will Artificial Intelligence Replace Biostatisticians? Evolving Tools and Enduring Responsibilities
    Özge Pasin
    Hamidiye Medical Journal.2026;[Epub]     CrossRef
  • Engineering Students' Critical Engagement With ChatGPT: Effects of Output Accuracy on Answers and Confidence in a Real-World Probability Context
    Marija Kaplar, Zorana Luzanin, Milos Vucic, Lidija Ivanovic, Sebastijan Kaplar
    IEEE Transactions on Education.2026; 69(4): 246.     CrossRef
  • Can Generative AI and ChatGPT Outperform Humans on Cognitive-Demanding Problem-Solving Tasks in Science?
    Xiaoming Zhai, Matthew Nyaaba, Wenchao Ma
    Science & Education.2025; 34(2): 649.     CrossRef
  • From statistics to deep learning: Using large language models in psychiatric research
    Yining Hua, Andrew Beam, Lori B. Chibnik, John Torous
    International Journal of Methods in Psychiatric Research.2025;[Epub]     CrossRef
  • Assessing the Current Limitations of Large Language Models in Advancing Health Care Education
    JaeYong Kim, Bathri Narayan Vajravelu
    JMIR Formative Research.2025; 9: e51319.     CrossRef
  • ChatGPT for Univariate Statistics: Validation of AI-Assisted Data Analysis in Healthcare Research
    Michael R Ruta, Tony Gaidici, Chase Irwin, Jonathan Lifshitz
    Journal of Medical Internet Research.2025; 27: e63550.     CrossRef
  • ChatGPT-Assisted Deep Learning Models for Influenza-Like Illness Prediction in Mainland China: Time Series Analysis
    Weihong Huang, Wudi Wei, Xiaotao He, Baili Zhan, Xiaoting Xie, Meng Zhang, Shiyi Lai, Zongxiang Yuan, Jingzhen Lai, Rongfeng Chen, Junjun Jiang, Li Ye, Hao Liang
    Journal of Medical Internet Research.2025; 27: e74423.     CrossRef
  • Confirming SPSS Results With ChatGPT-4 and o3-mini Models
    Frederick Strale, Isaac Riddle, Bowen Geng, Blake Oxford, Malia Kah, Robert Sherwin
    Cureus.2025;[Epub]     CrossRef
  • A whole new world, a new fantastic point of view: Charting unexplored territories in consumer research with generative artificial intelligence
    Kiwoong Yoo, Michael Haenlein, Kelly Hewett
    Journal of the Academy of Marketing Science.2025; 53(3): 723.     CrossRef
  • One year in the classroom with ChatGPT: empirical insights and transformative impacts
    Feng Guo, Tian Li, Christopher J. L. Cunningham
    Frontiers in Education.2025;[Epub]     CrossRef
  • A Comparative Study of the Advantages and Disadvantages of DeepSeek and SPSS in Statistical Analysis
    沙沙 庞
    Statistics and Application.2025; 14(06): 172.     CrossRef
  • AI‐Assisted Statistical Review: Could It Have Averted Retractions? A Case‐Based Perspective From Immunology
    Michal Ordak
    Allergy.2025; 80(12): 3441.     CrossRef
  • The impact of generative AI on critical thinking skills: a systematic review, conceptual framework and future research directions
    Mohamed Y. I. Helal, Ibrahim A. Elgendy, Mousa Ahmed Albashrawi, Yogesh K. Dwivedi, Mohammad S. Al-Ahmadi, Il Jeon
    Information Discovery and Delivery.2025;[Epub]     CrossRef
  • ChatGPT's performance in sample size estimation: a preliminary study on the capabilities of artificial intelligence
    Paul Sebo, Ting Wang
    Family Practice.2025;[Epub]     CrossRef
  • ChatGPT in Medical Education: Bibliometric and Visual Analysis
    Yuning Zhang, Xiaolu Xie, Qi Xu
    JMIR Medical Education.2025; 11: e72356.     CrossRef
  • ChatGPT’s progress over time: A longitudinal enhancing biostatistical problem-solving in medical education
    Aleksandra Ignjatović, Marija Anđelković Apostolović, Lazar Stevanović, Pavle Radovanović, Marija Topalović, Tamara Filipović, Suzana Otašević
    Health Informatics Journal.2025;[Epub]     CrossRef
  • AI-assisted statistical review of 100 oncology research articles: compliance with SAMPL guidelines
    Michal Ordak
    Current Research in Translational Medicine.2025; 73(4): 103544.     CrossRef
  • Applications, Challenges, and Prospects of Generative Artificial Intelligence Empowering Medical Education: Scoping Review
    Yuhang Lin, Zhiheng Luo, Zicheng Ye, Nuoxi Zhong, Lijian Zhao, Long Zhang, Xiaolan Li, Zetao Chen, Yijia Chen
    JMIR Medical Education.2025; 11: e71125.     CrossRef
  • Generative Artificial Intelligence for Data Analysis: A Randomised Controlled Trial in a Public Health Research Institute
    Tafadzwa Dhokotera, Nandi Joubert, Aline Veillat, Christoph Pimmer, Karin Gross, Marco Waser, Jan Hattendorf, Julia Bohlius
    International Journal of Public Health.2025;[Epub]     CrossRef
  • ChatGPT as a Tool for Biostatisticians: A Tutorial on Applications, Opportunities, and Limitations
    Dennis Dobler, Harald Binder, Anne‐Laure Boulesteix, Jan‐Bernd Igelmann, David Köhler, Ulrich Mansmann, Markus Pauly, André Scherag, Matthias Schmid, Amani Al Tawil, Susanne Weber
    Statistics in Medicine.2025;[Epub]     CrossRef
  • Statistical analysis using ChatGPT in medical research
    Soo-Nyung Kim
    Obstetrics & Gynecology Science.2025; 68(6): 467.     CrossRef
  • Opportunities, challenges, and future directions of large language models, including ChatGPT in medical education: a systematic scoping review
    Xiaojun Xu, Yixiao Chen, Jing Miao
    Journal of Educational Evaluation for Health Professions.2024; 21: 6.     CrossRef
  • Comparing the Performance of ChatGPT-4 and Medical Students on MCQs at Varied Levels of Bloom’s Taxonomy
    Ambadasu Bharatha, Nkemcho Ojeh, Ahbab Mohammad Fazle Rabbi, Michael Campbell, Kandamaran Krishnamurthy, Rhaheem Layne-Yarde, Alok Kumar, Dale Springer, Kenneth Connell, Md Anwarul Majumder
    Advances in Medical Education and Practice.2024; Volume 15: 393.     CrossRef
  • Revolutionizing Cardiology With Words: Unveiling the Impact of Large Language Models in Medical Science Writing
    Abhijit Bhattaru, Naveena Yanamala, Partho P. Sengupta
    Canadian Journal of Cardiology.2024; 40(10): 1950.     CrossRef
  • ChatGPT in medicine: prospects and challenges: a review article
    Songtao Tan, Xin Xin, Di Wu
    International Journal of Surgery.2024; 110(6): 3701.     CrossRef
  • In-depth analysis of ChatGPT’s performance based on specific signaling words and phrases in the question stem of 2377 USMLE step 1 style questions
    Leonard Knoedler, Samuel Knoedler, Cosima C. Hoch, Lukas Prantl, Konstantin Frank, Laura Soiderer, Sebastian Cotofana, Amir H. Dorafshar, Thilo Schenck, Felix Vollbach, Giuseppe Sofo, Michael Alfertshofer
    Scientific Reports.2024;[Epub]     CrossRef
  • Evaluating the quality of responses generated by ChatGPT
    Danimir Mandić, Gordana Miščević, Ljiljana Bujišić
    Metodicka praksa.2024; 27(1): 5.     CrossRef
  • A Comparative Evaluation of Statistical Product and Service Solutions (SPSS) and ChatGPT-4 in Statistical Analyses
    Al Imran Shahrul, Alizae Marny F Syed Mohamed
    Cureus.2024;[Epub]     CrossRef
  • ChatGPT and Other Large Language Models in Medical Education — Scoping Literature Review
    Alexandra Aster, Matthias Carl Laupichler, Tamina Rockwell-Kollmann, Gilda Masala, Ebru Bala, Tobias Raupach
    Medical Science Educator.2024; 35(1): 555.     CrossRef
  • Exploring the potential of large language models for integration into an academic statistical consulting service–the EXPOLS study protocol
    Urs Alexander Fichtner, Jochen Knaus, Erika Graf, Georg Koch, Jörg Sahlmann, Dominikus Stelzer, Martin Wolkewitz, Harald Binder, Susanne Weber, Bekalu Tadesse Moges
    PLOS ONE.2024; 19(12): e0308375.     CrossRef
Brief report
Comparing ChatGPT’s ability to rate the degree of stereotypes and the consistency of stereotype attribution with those of medical students in New Zealand in developing a similarity rating test: a methodological study  
Chao-Cheng Lin, Zaine Akuhata-Huntington, Che-Wei Hsu
J Educ Eval Health Prof. 2023;20:17.   Published online June 12, 2023
DOI: https://doi.org/10.3352/jeehp.2023.20.17
  • 7,349 View
  • 194 Download
  • 9 Web of Science
  • 10 Crossref
AbstractAbstract PDFSupplementary Material
Learning about one’s implicit bias is crucial for improving one’s cultural competency and thereby reducing health inequity. To evaluate bias among medical students following a previously developed cultural training program targeting New Zealand Māori, we developed a text-based, self-evaluation tool called the Similarity Rating Test (SRT). The development process of the SRT was resource-intensive, limiting its generalizability and applicability. Here, we explored the potential of ChatGPT, an automated chatbot, to assist in the development process of the SRT by comparing ChatGPT’s and students’ evaluations of the SRT. Despite results showing non-significant equivalence and difference between ChatGPT’s and students’ ratings, ChatGPT’s ratings were more consistent than students’ ratings. The consistency rate was higher for non-stereotypical than for stereotypical statements, regardless of rater type. Further studies are warranted to validate ChatGPT’s potential for assisting in SRT development for implementation in medical education and evaluation of ethnic stereotypes and related topics.

Citations

Citations to this article as recorded by  
  • Development and validation of a GPT-based rater for assessing communication skills using the Gap-Kalamazoo Communication Skills Assessment Form
    Yu-Jeng Ju, Yi-Ching Wang, Shih-Chieh Lee, Cheng‐Heng Liu, Meng-Lin Lee, Chieh-Yi Hou, Chih-Wei Yang, Ching-Lin Hsieh
    Medical Teacher.2026; 48(1): 93.     CrossRef
  • Generative artificial intelligence in mental health: A preliminary study on automating materials development for cognitive bias modification
    Che-Wei Hsu, Mia Cochrane, Sasini Bambarawana
    International Journal of Mental Health.2026; : 1.     CrossRef
  • Reducing negative interpretation bias in depression using cognitive bias modification with AI-generated training materials: a proof-of-principle study
    Che-Wei Hsu, Azariah Drummond, Adrienne Buckingham, Katrina Le Cong, Kerryn Carson
    International Journal of Mental Health.2026; : 1.     CrossRef
  • Applications of Artificial Intelligence in Medical Education: A Systematic Review
    Eric Hallquist, Ishank Gupta, Michael Montalbano, Marios Loukas
    Cureus.2025;[Epub]     CrossRef
  • One year in the classroom with ChatGPT: empirical insights and transformative impacts
    Feng Guo, Tian Li, Christopher J. L. Cunningham
    Frontiers in Education.2025;[Epub]     CrossRef
  • AI-driven network biology identifies SRC as a therapeutic target in metastatic pancreatic adenocarcinoma
    Ayla Zhang, Jake Y. Chen
    Intelligent Oncology.2025; 1(3): 233.     CrossRef
  • The Performance of ChatGPT on Short-answer Questions in a Psychiatry Examination: A Pilot Study
    Chao-Cheng Lin, Kobus du Plooy, Andrew Gray, Deirdre Brown, Linda Hobbs, Tess Patterson, Valerie Tan, Daniel Fridberg, Che-Wei Hsu
    Taiwanese Journal of Psychiatry.2024; 38(2): 94.     CrossRef
  • ChatGPT and Other Large Language Models in Medical Education — Scoping Literature Review
    Alexandra Aster, Matthias Carl Laupichler, Tamina Rockwell-Kollmann, Gilda Masala, Ebru Bala, Tobias Raupach
    Medical Science Educator.2024; 35(1): 555.     CrossRef
  • Psychiatric Care, Training and Research in Aotearoa New Zealand
    Chao-Cheng (Chris) Lin, Charlotte Mentzel, Maria Luz C. Querubin
    Taiwanese Journal of Psychiatry.2024; 38(4): 161.     CrossRef
  • Efficacy and limitations of ChatGPT as a biostatistical problem-solving tool in medical education in Serbia: a descriptive study
    Aleksandra Ignjatović, Lazar Stevanović
    Journal of Educational Evaluation for Health Professions.2023; 20: 28.     CrossRef
Review
Can an artificial intelligence chatbot be the author of a scholarly article?  
Ju Yoen Lee
J Educ Eval Health Prof. 2023;20:6.   Published online February 27, 2023
DOI: https://doi.org/10.3352/jeehp.2023.20.6
  • 66,737 View
  • 1,035 Download
  • 78 Web of Science
  • 84 Crossref
AbstractAbstract PDFSupplementary Material
At the end of 2022, the appearance of ChatGPT, an artificial intelligence (AI) chatbot with amazing writing ability, caused a great sensation in academia. The chatbot turned out to be very capable, but also capable of deception, and the news broke that several researchers had listed the chatbot (including its earlier version) as co-authors of their academic papers. In response, Nature and Science expressed their position that this chatbot cannot be listed as an author in the papers they publish. Since an AI chatbot is not a human being, in the current legal system, the text automatically generated by an AI chatbot cannot be a copyrighted work; thus, an AI chatbot cannot be an author of a copyrighted work. Current AI chatbots such as ChatGPT are much more advanced than search engines in that they produce original text, but they still remain at the level of a search engine in that they cannot take responsibility for their writing. For this reason, they also cannot be authors from the perspective of research ethics.

Citations

Citations to this article as recorded by  
  • Integrating Artificial Intelligence in Medical Writing: Balancing Technological Innovation and Human Expertise, with Practical Applications in Lower Extremity Wounds Care
    Pak Thaichana, Myo Zin Oo, Gabriel Leiden Thorup, Chayatorn Chansakaow, Supapong Arworn, Kittipan Rerkasem
    The International Journal of Lower Extremity Wounds.2026; 25(2): 407.     CrossRef
  • Generative artificial intelligence in ophthalmology research writing: A comprehensive review of applications, detection tools, and ethical considerations
    Pin-Jung Cheng, Fang-Yu Hu, Le-Yu Chen, Jen-Yu Liu, Jo-Hsuan Wu, Wei-Li Chen
    Taiwan Journal of Ophthalmology.2026; 16(1): 68.     CrossRef
  • Peer-reviewed by human experts: AI failed in key steps to generate a scoping review on the neural mechanisms of cross-education
    M. Morrone, T. Hortobágyi, D. Kidgell, J. P. Farthing, F. Deriu, A. Manca
    European Journal of Applied Physiology.2026; 126(4): 1907.     CrossRef
  • Authors, Academics, and AI: Questions About Research and Publishing in the World of Artificial Intelligence
    Christopher J. Peterson, Caleb Anderson, Gilbert Berdine, Kenneth Nugent
    Journal of Electronic Resources in Medical Libraries.2026; 23(1): 12.     CrossRef
  • Navigating academic integrity in biomedical research: the impact of large language models on current practices and future directions
    Anqi Lin, Zuwei Chen, Aimin Jiang, Bufu Tang, Chang Qi, Lingxuan Zhu, Weiming Mou, Wenyi Gan, Dongqiang Zeng, Mingjia Xiao, Guangdi Chu, Shengkun Peng, Hank Z.H. Wong, Lin Zhang, Hengguo Zhang, Xinpei Deng, Quan Cheng, Jian Zhang, Peng Luo
    International Journal of Surgery.2026; 112(2): 4418.     CrossRef
  • Artificial Intelligence (AI) guidance for authors, peer reviewers, and editors: A content analysis of journal policies
    Deborah H. Charbonneau, Mei Zhang
    Accountability in Research.2026;[Epub]     CrossRef
  • An overview of artificial intelligence approaches for automating evidence synthesis
    Sashika Harasgama, Helen Pearce, Liam Loftus, Helena Painter, John Ford
    Public Health.2026; 254: 106220.     CrossRef
  • From Hallucination to Precision: A Longitudinal Analysis of Reference Accuracy and Plagiarism in AI-Generated Medical Literature (2024–2026)
    Mevlüt Okan Aydin, Alper Vatansever, Sezer Erer Kafa
    Uludağ Üniversitesi Tıp Fakültesi Dergisi.2026; 52: 1870116.     CrossRef
  • Artificial intelligence policies in disaster research: a review of journals and their publishers
    Richard Armitage
    International Journal of Disaster Risk Reduction.2026; 137: 106078.     CrossRef
  • An Enhanced AI Chatbot to Facilitate Learning Performance in a Virtual 3D Chemistry Laboratory
    Chih-Ming Chen, Bo-Jin Chen, Chieh-Ling Huang
    Journal of Science Education and Technology.2026;[Epub]     CrossRef
  • AI Collaboration and the Decentering of Human Creativity
    Azmine Toushik Wasi, Sadia Tasnim Meem
    ACM AI Letters.2026; 1(2): 1.     CrossRef
  • Is AI-Produced Humanities Scholarship a Case of Research Misconduct?
    Thomas Metcalf
    Journal of Academic Ethics.2026;[Epub]     CrossRef
  • Generative Artificial Intelligence in Scientific Publishing: Ethical Governance, Challenges, and Responsibilities
    Dawid Gruszczyński, Kacper Nijakowski, Jowita Halupczok-Żyła, Marek Ruchała, Jarosław Walkowiak, Nadia Sawicka-Gutaj
    Journal of Medical Science.2026; 95(2): e1571.     CrossRef
  • Lens of Generative Odyssey: Reframing Artificial Intelligence in Ophthalmic Manuscripts—The Good, the Bad, and the Ugly
    Prasanna Venkatesh Ramesh
    TNOA Journal of Ophthalmic Science and Research.2026; 64(3): 337.     CrossRef
  • Beyond text generation: a comprehensive evaluation of ChatGPT-5.5 for scientific narrative review writing on sleeve gastrectomy
    Akin Calisir, Mustafa Sahin, Fatih Turkoglu, Sinan Sener, Abdullah Gurhan Duyan, Husnu Alptekin
    Updates in Surgery.2026;[Epub]     CrossRef
  • Who wrote this? An EAP think-aloud study on AI detection
    Zainab Teraif
    Language Learning & Technology.2026; 30(1): 1.     CrossRef
  • Identification of ChatGPT‐Generated Abstracts Within Shoulder and Elbow Surgery Poses a Challenge for Reviewers
    Ryan D. Stadler, Suleiman Y. Sudah, Michael A. Moverman, Patrick J. Denard, Xavier A. Duralde, Grant E. Garrigues, Christopher S. Klifto, Jonathan C. Levy, Surena Namdari, Joaquin Sanchez‐Sotelo, Mariano E. Menendez
    Arthroscopy.2025; 41(4): 916.     CrossRef
  • ChatGPT or Gemini: Who Makes the Better Scientific Writing Assistant?
    Hatoon S. AlSagri, Faiza Farhat, Shahab Saquib Sohail, Abdul Khader Jilani Saudagar
    Journal of Academic Ethics.2025; 23(3): 1121.     CrossRef
  • Let stochastic parrots squawk: why academic journals should allow large language models to coauthor articles
    Nicholas J. Abernethy
    AI and Ethics.2025; 5(5): 4535.     CrossRef
  • Can ChatGPT be an author? Generative AI creative writing assistance and perceptions of authorship, creatorship, responsibility, and disclosure
    Paul Formosa, Sarah Bankins, Rita Matulionyte, Omid Ghasemi
    AI & SOCIETY.2025; 40(5): 3405.     CrossRef
  • Attitudes and perceptions of medical researchers towards the use of artificial intelligence chatbots in the scientific process: an international cross-sectional survey
    Jeremy Y Ng, Sharleen G Maduranayagam, Nirekah Suthakar, Amy Li, Cynthia Lokker, Alfonso Iorio, R Brian Haynes, David Moher
    The Lancet Digital Health.2025; 7(1): e94.     CrossRef
  • Introducing Our Custom GPT: An Example of the Potential Impact of Personalized GPT Builders on Scientific Writing
    Aymen Kabir, Suraj Shah, Alexander Haddad, Daniel M.S. Raper
    World Neurosurgery.2025; 193: 461.     CrossRef
  • Ethical issues and violations in using chatbots in academic writing and publishing: the answers from ChatGPT
    Eren Erkılıç, Ibrahim Cifci
    Journal of Multidisciplinary Academic Tourism.2025; 10(1): 111.     CrossRef
  • Chat GPT vs an experienced ophthalmologist: evaluating chatbot writing performance in ophthalmology
    Gabriel Katz, Ofira Zloto, Avner Hostovsky, Ruth Huna-Baron, Iris Ben-Bassat Mizrachi, Zvia Burgansky, Alon Skaat, Vicktoria Vishnevskia-Dai, Ido Didi Fabian, Oded Sagiv, Ayelet Priel, Benjamin S. Glicksberg, Eyal Klang
    Eye.2025; 39(10): 1948.     CrossRef
  • Comparison of hand surgery certification exams in Europe and the United States using ChatGPT 4.0
    Salman Hasan, Kyros Ipaktchi, Nicolas Meyer, Philippe Liverneaux
    Journal of Hand and Microsurgery.2025; 17(4): 100258.     CrossRef
  • Stop citing ChatGPT and other LLMs as academic references
    Louie Giray
    Public Services Quarterly.2025; 21(3): 197.     CrossRef
  • Editorial policies for use and acknowledgment of artificial intelligence in dental journals
    Ana Beatriz L. Queiroz, Letícia Regina Morello Sartori, Giana da Silveira Lima, Rafael R. Moraes
    Journal of Dentistry.2025; 161: 105923.     CrossRef
  • Identification and Categorization of the Top 100 Articles and the Future of Large Language Models: Thematic Analysis Using Bibliometric Analysis
    Ethan Bernstein, Anya Ramsamooj, Kelsey L Millar, Zachary C Lum
    JMIR AI.2025; 4: e68603.     CrossRef
  • Large language models in medicine: current ethical challenges
    SA Kostrov, MP Potapov
    Медицинская этика.2025;[Epub]     CrossRef
  • Generative AI Governance Model in Educational Research
    Isabel Pinho, António Pedro Costa, Cláudia Pinho
    Frontiers in Education.2025;[Epub]     CrossRef
  • Comparative performance of neurosurgery-specific, peer-reviewed versus general AI chatbots in bilingual board examinations: evaluating accuracy, consistency, and error minimization strategies
    Mahmut Çamlar, Umut Tan Sevgi, Gökberk Erol, Furkan Karakaş, Yücel Doğruel, Abuzer Güngör
    Acta Neurochirurgica.2025;[Epub]     CrossRef
  • ¿Cómo está transformando la inteligencia artificial la comunicación científica? Desafíos, oportunidades y el papel de los actores involucrados: una revisión de alcance
    Jairo Buitrago-Ciro, Estela Morales Campos, César Leonardo Villamizar Romero
    Investigación Bibliotecológica: archivonomía, bibliotecología e información.2025; 39(104): 111.     CrossRef
  • What fifty-one years of linguistics and artificial intelligence research tell us about their correlation: A scientometric analysis
    Mohammed Q. Shormani
    Artificial Intelligence Review.2025;[Epub]     CrossRef
  • A holistic exploration of student attitudes toward AI use in higher education: an international comparison
    David Vaněček, Yilmaz Ilker Yorulmaz, Dana Dobrovská
    Cogent Education.2025;[Epub]     CrossRef
  • GENERATIVE ARTIFICIAL INTELLIGENCE: LEGAL CHALLENGES REGARDING COPYRIGHT IN THE USE OF CHATBOTS
    José dos Santos Machado, Francisco Sandro Rodrigues Holanda, Valdir Ribeiro Pimenta Neto, Adauto Cavalcante Menezes
    REVISTA FOCO.2025; 18(11): e10608.     CrossRef
  • The ethics of using artificial intelligence in writing medical research papers
    Shinae Yu, Hyunyong Hwang
    Kosin Medical Journal.2025; 40(4): 270.     CrossRef
  • Risks of abuse of large language models, like ChatGPT, in scientific publishing: Authorship, predatory publishing, and paper mills
    Graham Kendall, Jaime A. Teixeira da Silva
    Learned Publishing.2024; 37(1): 55.     CrossRef
  • Can ChatGPT be an author? A study of artificial intelligence authorship policies in top academic journals
    Brady D. Lund, K.T. Naheem
    Learned Publishing.2024; 37(1): 13.     CrossRef
  • The Role of AI in Writing an Article and Whether it Can Be a Co-author: What if it Gets Support From 2 Different AIs Like ChatGPT and Google Bard for the Same Theme?
    İlhan Bahşi, Ayşe Balat
    Journal of Craniofacial Surgery.2024; 35(1): 274.     CrossRef
  • Artificial Intelligence–Generated Scientific Literature: A Critical Appraisal
    Justyna Zybaczynska, Matthew Norris, Sunjay Modi, Jennifer Brennan, Pooja Jhaveri, Timothy J. Craig, Taha Al-Shaikhly
    The Journal of Allergy and Clinical Immunology: In Practice.2024; 12(1): 106.     CrossRef
  • Does Google’s Bard Chatbot perform better than ChatGPT on the European hand surgery exam?
    Goetsch Thibaut, Armaghan Dabbagh, Philippe Liverneaux
    International Orthopaedics.2024; 48(1): 151.     CrossRef
  • ChatGPT in medical writing: A game-changer or a gimmick?
    Shital Sarah Ahaley, Ankita Pandey, Simran Kaur Juneja, Tanvi Suhane Gupta, Sujatha Vijayakumar
    Perspectives in Clinical Research.2024; 15(4): 165.     CrossRef
  • A Brief Review of the Efficacy in Artificial Intelligence and Chatbot-Generated Personalized Fitness Regimens
    Daniel K. Bays, Cole Verble, Kalyn M. Powers Verble
    Strength & Conditioning Journal.2024; 46(4): 485.     CrossRef
  • Academic publisher guidelines on AI usage: A ChatGPT supported thematic analysis
    Mike Perkins, Jasper Roe
    F1000Research.2024; 12: 1398.     CrossRef
  • The Use of Artificial Intelligence in Writing Scientific Review Articles
    Melissa A. Kacena, Lilian I. Plotkin, Jill C. Fehrenbacher
    Current Osteoporosis Reports.2024; 22(1): 115.     CrossRef
  • Using AI to Write a Review Article Examining the Role of the Nervous System on Skeletal Homeostasis and Fracture Healing
    Murad K. Nazzal, Ashlyn J. Morris, Reginald S. Parker, Fletcher A. White, Roman M. Natoli, Jill C. Fehrenbacher, Melissa A. Kacena
    Current Osteoporosis Reports.2024; 22(1): 217.     CrossRef
  • GenAI et al.: Cocreation, Authorship, Ownership, Academic Ethics and Integrity in a Time of Generative AI
    Aras Bozkurt
    Open Praxis.2024; 16(1): 1.     CrossRef
  • An integrative decision-making framework to guide policies on regulating ChatGPT usage
    Umar Ali Bukar, Md Shohel Sayeed, Siti Fatimah Abdul Razak, Sumendra Yogarayan, Oluwatosin Ahmed Amodu
    PeerJ Computer Science.2024; 10: e1845.     CrossRef
  • Artificial Intelligence and Its Role in Medical Research
    Anurag Gola, Ambarish Das, Amar B. Gumataj, S. Amirdhavarshini, J. Venkatachalam
    Current Medical Issues.2024; 22(2): 97.     CrossRef
  • From advancements to ethics: Assessing ChatGPT’s role in writing research paper
    Vasu Gupta, Fnu Anamika, Kinna Parikh, Meet A Patel, Rahul Jain, Rohit Jain
    Turkish Journal of Internal Medicine.2024; 6(2): 74.     CrossRef
  • Yapay Zekânın Edebiyatta Kullanım Serüveni
    Nesime Ceyhan Akça, Serap Aslan Cobutoğlu, Özlem Yeşim Özbek, Mehmet Furkan Akça
    RumeliDE Dil ve Edebiyat Araştırmaları Dergisi.2024; (39): 283.     CrossRef
  • ChatGPT's Gastrointestinal Tumor Board Tango: A limping dance partner?
    Ughur Aghamaliyev, Javad Karimbayli, Clemens Giessen-Jung, Matthias Ilmer, Kristian Unger, Dorian Andrade, Felix O. Hofmann, Maximilian Weniger, Martin K. Angele, C. Benedikt Westphalen, Jens Werner, Bernhard W. Renz
    European Journal of Cancer.2024; 205: 114100.     CrossRef
  • Gout and Gout-Related Comorbidities: Insight and Limitations from Population-Based Registers in Sweden
    Panagiota Drivelegka, Lennart TH Jacobsson, Mats Dehlin
    Gout, Urate, and Crystal Deposition Disease.2024; 2(2): 144.     CrossRef
  • Artificial intelligence in academic cardiothoracic surgery
    Adham AHMED, Irbaz HAMEED
    The Journal of Cardiovascular Surgery.2024;[Epub]     CrossRef
  • The emergence of generative artificial intelligence platforms in 2023, journal metrics, appreciation to reviewers and volunteers, and obituary
    Sun Huh
    Journal of Educational Evaluation for Health Professions.2024; 21: 9.     CrossRef
  • A survey of safety and trustworthiness of large language models through the lens of verification and validation
    Xiaowei Huang, Wenjie Ruan, Wei Huang, Gaojie Jin, Yi Dong, Changshun Wu, Saddek Bensalem, Ronghui Mu, Yi Qi, Xingyu Zhao, Kaiwen Cai, Yanghao Zhang, Sihao Wu, Peipei Xu, Dengyu Wu, Andre Freitas, Mustafa A. Mustafa
    Artificial Intelligence Review.2024;[Epub]     CrossRef
  • Decision-Making Framework for the Utilization of Generative Artificial Intelligence in Education: A Case Study of ChatGPT
    Umar Ali Bukar, Md. Shohel Sayeed, Siti Fatimah Abdul Razak, Sumendra Yogarayan, Radhwan Sneesl
    IEEE Access.2024; 12: 95368.     CrossRef
  • The Syntax of Smart Writing: Artificial Intelligence Unveiled
    Balaji Arumugam, Arun Murugan, Kirubakaran S., Saranya Rajamanickam
    International Journal of Preventative & Evidence Based Medicine.2024; : 1.     CrossRef
  • Generative artificial intelligence usage by researchers at work: Effects of gender, career stage, type of workplace, and perceived barriers
    Pablo Dorta-González, Alexis Jorge López-Puig, María Isabel Dorta-González, Sara M. González-Betancor
    Telematics and Informatics.2024; 94: 102187.     CrossRef
  • Leveraging Artificial Intelligence In Project-Based Service Learning To Advance Sustainable Development: A Pedagogical Approach For Marketing Education
    C. M. Dubay, Melanie B. Richards
    Marketing Education Review.2024; 34(4): 307.     CrossRef
  • Strategies for integrating ChatGPT and generative AI into clinical studies
    Jeong-Moo Lee
    Blood Research.2024;[Epub]     CrossRef
  • Universal skepticism of ChatGPT: a review of early literature on chat generative pre-trained transformer
    Casey Watters, Michal K. Lemanski
    Frontiers in Big Data.2023;[Epub]     CrossRef
  • The importance of human supervision in the use of ChatGPT as a support tool in scientific writing
    William Castillo-González
    Metaverse Basic and Applied Research.2023;[Epub]     CrossRef
  • ChatGPT for Future Medical and Dental Research
    Bader Fatani
    Cureus.2023;[Epub]     CrossRef
  • Chatbots in Medical Research
    Punit Sharma
    Clinical Nuclear Medicine.2023; 48(9): 838.     CrossRef
  • Potential applications of ChatGPT in dermatology
    Nicolas Kluger
    Journal of the European Academy of Dermatology and Venereology.2023;[Epub]     CrossRef
  • The emergent role of artificial intelligence, natural learning processing, and large language models in higher education and research
    Tariq Alqahtani, Hisham A. Badreldin, Mohammed Alrashed, Abdulrahman I. Alshaya, Sahar S. Alghamdi, Khalid bin Saleh, Shuroug A. Alowais, Omar A. Alshaya, Ishrat Rahman, Majed S. Al Yami, Abdulkareem M. Albekairy
    Research in Social and Administrative Pharmacy.2023; 19(8): 1236.     CrossRef
  • ChatGPT Performance on the American Urological Association Self-assessment Study Program and the Potential Influence of Artificial Intelligence in Urologic Training
    Nicholas A. Deebel, Ryan Terlecki
    Urology.2023; 177: 29.     CrossRef
  • Intelligence or artificial intelligence? More hard problems for authors of Biological Psychology, the neurosciences, and everyone else
    Thomas Ritz
    Biological Psychology.2023; 181: 108590.     CrossRef
  • The ethics of disclosing the use of artificial intelligence tools in writing scholarly manuscripts
    Mohammad Hosseini, David B Resnik, Kristi Holmes
    Research Ethics.2023; 19(4): 449.     CrossRef
  • How trustworthy is ChatGPT? The case of bibliometric analyses
    Faiza Farhat, Shahab Saquib Sohail, Dag Øivind Madsen
    Cogent Engineering.2023;[Epub]     CrossRef
  • Disclosing use of Artificial Intelligence: Promoting transparency in publishing
    Parvaiz A. Koul
    Lung India.2023; 40(5): 401.     CrossRef
  • ChatGPT in medical research: challenging time ahead
    Daideepya C Bhargava, Devendra Jadav, Vikas P Meshram, Tanuj Kanchan
    Medico-Legal Journal.2023; 91(4): 223.     CrossRef
  • Academic publisher guidelines on AI usage: A ChatGPT supported thematic analysis
    Mike Perkins, Jasper Roe
    F1000Research.2023; 12: 1398.     CrossRef
  • Ethical consideration of the use of generative artificial intelligence, including ChatGPT in writing a nursing article
    Sun Huh
    Child Health Nursing Research.2023; 29(4): 249.     CrossRef
  • Artificial Intelligence-Supported Systems in Anesthesiology and Its Standpoint to Date—A Review
    Fiona M. P. Pham
    Open Journal of Anesthesiology.2023; 13(07): 140.     CrossRef
  • ChatGPT as an innovative tool for increasing sales in online stores
    Michał Orzoł, Katarzyna Szopik-Depczyńska
    Procedia Computer Science.2023; 225: 3450.     CrossRef
  • Intelligent Plagiarism as a Misconduct in Academic Integrity
    Jesús Miguel Muñoz-Cantero, Eva Maria Espiñeira-Bellón
    Acta Médica Portuguesa.2023; 37(1): 1.     CrossRef
  • Follow-up of Artificial Intelligence Development and its Controlled Contribution to the Article: Step to the Authorship?
    Ekrem Solmaz
    European Journal of Therapeutics.2023;[Epub]     CrossRef
  • May Artificial Intelligence Be a Co-Author on an Academic Paper?
    Ayşe Balat, İlhan Bahşi
    European Journal of Therapeutics.2023; 29(3): e12.     CrossRef
  • Opportunities and challenges for ChatGPT and large language models in biomedicine and health
    Shubo Tian, Qiao Jin, Lana Yeganova, Po-Ting Lai, Qingqing Zhu, Xiuying Chen, Yifan Yang, Qingyu Chen, Won Kim, Donald C Comeau, Rezarta Islamaj, Aadit Kapoor, Xin Gao, Zhiyong Lu
    Briefings in Bioinformatics.2023;[Epub]     CrossRef
  • ChatGPT: "To be or not to be" ... in academic research. The human mind's analytical rigor and capacity to discriminate between AI bots' truths and hallucinations
    Aurelian Anghelescu, Ilinca Ciobanu, Constantin Munteanu, Lucia Ana Maria Anghelescu, Gelu Onose
    Balneo and PRM Research Journal.2023; 14(Vol.14, no): 614.     CrossRef
  • Editorial policies of Journal of Educational Evaluation for Health Professions on the use of generative artificial intelligence in article writing and peer review
    Sun Huh
    Journal of Educational Evaluation for Health Professions.2023; 20: 40.     CrossRef
  • Should We Wait for Major Frauds to Unveil to Plan an AI Use License?
    Istemihan Coban
    European Journal of Therapeutics.2023; 30(2): 198.     CrossRef
Brief report
Are ChatGPT’s knowledge and interpretation ability comparable to those of medical students in Korea for taking a parasitology examination?: a descriptive study  
Sun Huh
J Educ Eval Health Prof. 2023;20:1.   Published online January 11, 2023
DOI: https://doi.org/10.3352/jeehp.2023.20.1
  • 25,011 View
  • 1,262 Download
  • 272 Web of Science
  • 133 Crossref
AbstractAbstract PDFSupplementary Material
This study aimed to compare the knowledge and interpretation ability of ChatGPT, a language model of artificial general intelligence, with those of medical students in Korea by administering a parasitology examination to both ChatGPT and medical students. The examination consisted of 79 items and was administered to ChatGPT on January 1, 2023. The examination results were analyzed in terms of ChatGPT’s overall performance score, its correct answer rate by the items’ knowledge level, and the acceptability of its explanations of the items. ChatGPT’s performance was lower than that of the medical students, and ChatGPT’s correct answer rate was not related to the items’ knowledge level. However, there was a relationship between acceptable explanations and correct answers. In conclusion, ChatGPT’s knowledge and interpretation ability for this parasitology examination were not yet comparable to those of medical students in Korea.

Citations

Citations to this article as recorded by  
  • Is ChatGPT Enhancing Youth’s Learning, Engagement and Satisfaction?
    Christina Sanchita Shah, Smriti Mathur, Sushant Kr. Vishnoi
    Journal of Computer Information Systems.2026; 66(3): 353.     CrossRef
  • Unveiling the impact of ChatGPT: investigating self-efficacy, anxiety and motivation on student performance in blended learning environments
    Ridwan Daud Mahande, M. Miftach Fakhri, Irwansyah Suwahyu, Dwi Rezky Anandari Sulaiman
    Journal of Applied Research in Higher Education.2026; 18(1): 282.     CrossRef
  • Evaluation of the performance of ChatGPT‐4 and ChatGPT‐4o as a learning tool in endodontics
    Esra Arılı Öztürk, Ceren Turan Gökduman, Burhan Can Çanakçi
    International Endodontic Journal.2026; 59(6): 1057.     CrossRef
  • Empirical investigation of critical factors influencing ChatGPT integration in contemporary technology education
    Ming Yuan Hsieh
    Research in Science & Technological Education.2026; 44(2): 641.     CrossRef
  • Comparison of Human and Artificial Intelligence (AI) in Writing and Rating Restorative Dentistry Essays
    Afnan O. Al‐Zain, Abdulrahman A. Alghamdi, Bashair Alansari, Alanoud Alamoudi, Heba El‐Deeb, Eman H. Ismail
    European Journal of Dental Education.2026; 30(3): 1026.     CrossRef
  • ChatGPT usage among nursing students in Taiwan: a qualitative study
    Hui-Man Huang
    BMC Nursing.2026;[Epub]     CrossRef
  • Comparative performance of ChatGPT, Gemini, and Deepseek on endodontic exam questions in Turkish and English
    Eda Gürsu Şahin
    BMC Oral Health.2026;[Epub]     CrossRef
  • Generative artificial intelligence in mental health: A preliminary study on automating materials development for cognitive bias modification
    Che-Wei Hsu, Mia Cochrane, Sasini Bambarawana
    International Journal of Mental Health.2026; : 1.     CrossRef
  • Multi-metric comparative evaluation of DeepSeek and ChatGPT in USMLE versus CNMLE for medical education
    Qing Wang, Junlian Li, Xiaoying Li, Panpan Deng
    Scientific Reports.2026;[Epub]     CrossRef
  • Assessing the accuracy and educational value of ChatGPT-generated content for core topics in cardiology: a descriptive analysis at Selçuk University Cardiology Clinic
    Hüseyin Tezcan, Abdullah Tunçez, Kadri Murat Gürses, Yasin Özen, Muhammed Ulvi Yalçın
    BMC Medical Education.2026;[Epub]     CrossRef
  • Evaluation of the approaches of different specialties and artificial intelligence systems on traumatic dental injuries
    Gonca Zelal Şahin, Ayhan Eymirli
    BMC Oral Health.2026;[Epub]     CrossRef
  • Artificial Intelligence for Assessment and Feedback in Medical Education: Bibliometric Mapping Study and Thematic Evidence Map
    Zihang Zhao, Zihan Liu, Liang Guo, Teng Pan, Yousheng Zhang, Chenxiang Miao, Yiting Ge, Yipeng Wang, Xin Hu, Xin Wang, Ruipeng Zhang, Zhiyong Hou
    JMIR Medical Education.2026; 12: e98949.     CrossRef
  • Bibliometric mapping and evolutionary logic of large language models reshaping medical education 2022–2026
    Xi Ling, Chengliang Wang
    Discover Computing.2026;[Epub]     CrossRef
  • Artificial Intelligence In Dental Exams: A Performance Comparison Of ChatGPT, DeepSeek, And Students
    Şükran Ayran, Selma Ece Karabıyıkoğlu, Neşe Oral, Hilal Peker Öztürk, Barış Filiz Erol, Hakan Avsever
    Acta Odontologica Turcica.2026; 43(3): 160.     CrossRef
  • Integration of AI into Education: Teachers' Perspectives on Competencies, Challenges, and Needs
    Feray Uğur Erdoğmuş
    Participatory Educational Research.2026; 13(5): 115.     CrossRef
  • A Multidisciplinary Evaluation of ChatGPT and DeepSeek's Effectiveness in Answering Questions on Antibiotic Prophylaxis
    Şükran Acıpınar, Mehmet Şahinbaş, İrem Sude Aydın, Merve Yalçın
    Sağlık Bilimleri Dergisi.2026; 35(2): 427.     CrossRef
  • ChatGPT and the AI revolution: a comprehensive investigation of its multidimensional impact and potential
    Mohd Afjal
    Library Hi Tech.2025; 43(1): 353.     CrossRef
  • A comparative vignette study: Evaluating the potential role of a generative AI model in enhancing clinical decision‐making in nursing
    Mor Saban, Ilana Dubovi
    Journal of Advanced Nursing.2025; 81(11): 7489.     CrossRef
  • Augmenting intensive care unit nursing practice with generative AI: A formative study of diagnostic synergies using simulation‐based clinical cases
    Chedva Levin, Moriya Suliman, Etti Naimi, Mor Saban
    Journal of Clinical Nursing.2025; 34(7): 2898.     CrossRef
  • Exploring the Current Applications and Effectiveness of ChatGPT in Nursing: An Integrative Review
    Yuan Luo, Yiqun Miao, Yuhan Zhao, Jiawei Li, Ying Wu
    Journal of Advanced Nursing.2025; 81(7): 3473.     CrossRef
  • ChatGPT-Produced Content as a Resource in the Language Education Classroom: A Guiding Hand
    Rod E. Case, Leping Liu
    Computers in the Schools.2025; 42(2): 187.     CrossRef
  • Utility of ChatGPT as a preparation tool for the Orthopaedic In‐Training Examination
    Dhruv Mendiratta, Isabel Herzog, Rohan Singh, Ashok Para, Tej Joshi, Michael Vosbikian, Neil Kaushal
    Journal of Experimental Orthopaedics.2025;[Epub]     CrossRef
  • Exploring knowledge, attitudes, and practices of academics in the field of educational sciences towards using ChatGPT
    Burcu Karafil, Ahmet Uyar
    Education and Information Technologies.2025; 30(9): 11649.     CrossRef
  • Factors influencing Chinese pre-service teachers’ adoption of generative AI in teaching: an empirical study based on UTAUT2 and PLS-SEM
    Linlin Hu, Hao Wang, Yunfei Xin
    Education and Information Technologies.2025; 30(9): 12609.     CrossRef
  • Integrating AI Technology Into Language Teacher Education: Challenges, Potentials, and Assumptions
    Rod Case, Leping Liu, Joseph Mintz
    Computers in the Schools.2025; 42(2): 93.     CrossRef
  • Performance of ChatGPT-3.5 and ChatGPT-4 in the Taiwan National Pharmacist Licensing Examination: Comparative Evaluation Study
    Ying-Mei Wang, Hung-Wei Shen, Tzeng-Ji Chen, Shu-Chiung Chiang, Ting-Guan Lin
    JMIR Medical Education.2025; 11: e56850.     CrossRef
  • Eight Months into Reality: A Scoping Review of the Application of ChatGPT in Higher Education Teaching and Learning
    Qian Liu, Anjin Hu, Tehmina Gladman, Steve Gallagher
    Innovative Higher Education.2025; 50(5): 1677.     CrossRef
  • Performance of artificial intelligence on Turkish dental specialization exam: can ChatGPT-4.0 and gemini advanced achieve comparable results to humans?
    Soner Sismanoglu, Belen Sirinoglu Capan
    BMC Medical Education.2025;[Epub]     CrossRef
  • Applications of Artificial Intelligence in Medical Education: A Systematic Review
    Eric Hallquist, Ishank Gupta, Michael Montalbano, Marios Loukas
    Cureus.2025;[Epub]     CrossRef
  • Performance of ChatGPT-4 on Taiwanese Traditional Chinese Medicine Licensing Examinations: Cross-Sectional Study
    Liang-Wei Tseng, Yi-Chin Lu, Liang-Chi Tseng, Yu-Chun Chen, Hsing-Yu Chen
    JMIR Medical Education.2025; 11: e58897.     CrossRef
  • Comparing diagnostic skills in endodontic cases: dental students versus ChatGPT-4o
    Parla Meva Durmazpinar, Ece Ekmekci
    BMC Oral Health.2025;[Epub]     CrossRef
  • Evaluating the agreement between ChatGPT-4 and validated questionnaires in screening for anxiety and depression in college students: a cross-sectional study
    Jiali Liu, Juan Gu, Mengjie Tong, Yake Yue, Yufei Qiu, Lijuan Zeng, Yiqing Yu, Fen Yang, Shuyan Zhao
    BMC Psychiatry.2025;[Epub]     CrossRef
  • Factors Associated with Lower Performance of Artificial Intelligence on Answering Undergraduate Medical Education Multiple-Choice Questions
    Renato Ferretti, Joyce Santana Rizzi, Lorraine Silva Requena, Angelica Maria Bicudo, Pedro Tadao Hamamoto Filho
    Medical Science Educator.2025; 35(4): 2145.     CrossRef
  • ChatGPT는 건강운동관리사 이론시험에 합격할 수 있을까?
    상훈 김
    The Korean Journal of Physical Education.2025; 64(2): 15.     CrossRef
  • Assessing Information Provided by ChatGPT: Heart Failure Versus Patent Ductus Arteriosus
    Meghana Bhupathi, Jaza Mehweish Kareem, Anjali Mediboina, Keerthana Janapareddy
    Cureus.2025;[Epub]     CrossRef
  • How appropriately can generative artificial intelligence platforms, including GPT-4, Gemini, Bing, and Wrtn, answer questions about colon cancer in the Korean language?
    Sun Huh
    Annals of Coloproctology.2025; 41(3): 190.     CrossRef
  • Identification and Categorization of the Top 100 Articles and the Future of Large Language Models: Thematic Analysis Using Bibliometric Analysis
    Ethan Bernstein, Anya Ramsamooj, Kelsey L Millar, Zachary C Lum
    JMIR AI.2025; 4: e68603.     CrossRef
  • Determinants of researchers’ intentions to utilize ChatGPT: insights from Türkiye and the United Kingdom
    Nazli Yuceol, Ayse Merve Urfa Yilmaz, Meryem Akin
    Journal of Cultural Cognitive Science.2025; 9(2): 331.     CrossRef
  • The application of artificial intelligence-generated content in ophthalmology education
    Yinzongxiao Wang, Yun Zhao, Jia Li
    Frontiers in Medicine.2025;[Epub]     CrossRef
  • Evaluating ChatGPT's Diagnostic Accuracy in Oral Mucosal Lesions: A Comparative Study with a Maxillofacial Surgeon
    Hacer Eberliköse, Arif Yiğit Güler, Raha Akbarihamed, Caner Öztürk, Hakan Alpay Karasu
    European Annals of Dental Sciences.2025; 52(2): 92.     CrossRef
  • ChatGPT in Medical Education: Bibliometric and Visual Analysis
    Yuning Zhang, Xiaolu Xie, Qi Xu
    JMIR Medical Education.2025; 11: e72356.     CrossRef
  • Applications, Challenges, and Prospects of Generative Artificial Intelligence Empowering Medical Education: Scoping Review
    Yuhang Lin, Zhiheng Luo, Zicheng Ye, Nuoxi Zhong, Lijian Zhao, Long Zhang, Xiaolan Li, Zetao Chen, Yijia Chen
    JMIR Medical Education.2025; 11: e71125.     CrossRef
  • How AI Is Transforming Medical Education: Bibliometric Analysis
    Youyang Wang, Chuheng Chang, Wen Shi, Huiting Liu, Xiaoming Huang, Yang Jiao
    JMIR Medical Education.2025; 11: e75911.     CrossRef
  • Chatbot Underperformance in Biology and Image-Based Questions in Medical Education
    Joyce Santana Rizzi, Lorraine Silva Requena, Angelica Maria Bicudo, Pedro Tadao Hamamoto Filho, Renato Ferretti
    Journal of CME.2025;[Epub]     CrossRef
  • Evaluation of the accuracy of large language models in answering bone cancer-related questions
    Qilin Pan, Leiwen Huang, Ning Liu, Feixiang Lin, Shuxi Ye
    Frontiers in Public Health.2025;[Epub]     CrossRef
  • Evaluation of ChatGPT-4o and DeepSeek as tools for orthodontic health literacy in public dental education
    Zhaoxiang Wen, Jiaxin Huang, Keer Yu, Yaqi Li, Zhenhui Wang, Xiaozhu Liao, Biao Li, Zhendong Tao, Hong He
    Scientific Reports.2025;[Epub]     CrossRef
  • Comparison of chatbots’ accuracy in endodontics questions in dentistry specialization exam in Türkiye: ChatGPT-4o, Gemini Advanced, Copilot, and Claude
    Ceren Turan Gökduman, Esra Arılı Öztürk, Şule Aktaş, Burhan Can Çanakçi̇
    BMC Oral Health.2025;[Epub]     CrossRef
  • Multiple Large Language Models’ Performance on the Chinese Medical Licensing Examination: Quantitative Comparative Study
    Yanyu Diao, Mengyuan Wu, Jingwen Xu, Yifeng Pan
    JMIR Human Factors.2025; 12: e77978.     CrossRef
  • Performance of ChatGPT on the India Undergraduate Community Medicine Examination: Cross-Sectional Study
    Aravind P Gandhi, Felista Karen Joesph, Vineeth Rajagopal, P Aparnavi, Sushma Katkuri, Sonal Dayama, Prakasini Satapathy, Mahalaqua Nazli Khatib, Shilpa Gaidhane, Quazi Syed Zahiruddin, Ashish Behera
    JMIR Formative Research.2024; 8: e49964.     CrossRef
  • Large Language Models and Artificial Intelligence: A Primer for Plastic Surgeons on the Demonstrated and Potential Applications, Promises, and Limitations of ChatGPT
    Jad Abi-Rafeh, Hong Hao Xu, Roy Kazan, Ruth Tevlin, Heather Furnas
    Aesthetic Surgery Journal.2024; 44(3): 329.     CrossRef
  • Redesigning Tertiary Educational Evaluation with AI: A Task-Based Analysis of LIS Students’ Assessment on Written Tests and Utilizing ChatGPT at NSTU
    Shamima Yesmin
    Science & Technology Libraries.2024; 43(4): 355.     CrossRef
  • Unveiling the ChatGPT phenomenon: Evaluating the consistency and accuracy of endodontic question answers
    Ana Suárez, Víctor Díaz‐Flores García, Juan Algar, Margarita Gómez Sánchez, María Llorente de Pedro, Yolanda Freire
    International Endodontic Journal.2024; 57(1): 108.     CrossRef
  • Bob or Bot: Exploring ChatGPT's Answers to University Computer Science Assessment
    Mike Richards, Kevin Waugh, Mark Slaymaker, Marian Petre, John Woodthorpe, Daniel Gooch
    ACM Transactions on Computing Education.2024; 24(1): 1.     CrossRef
  • A systematic review of ChatGPT use in K‐12 education
    Peng Zhang, Gemma Tur
    European Journal of Education.2024;[Epub]     CrossRef
  • Evaluating ChatGPT as a self‐learning tool in medical biochemistry: A performance assessment in undergraduate medical university examination
    Krishna Mohan Surapaneni, Anusha Rajajagadeesan, Lakshmi Goudhaman, Shalini Lakshmanan, Saranya Sundaramoorthi, Dineshkumar Ravi, Kalaiselvi Rajendiran, Porchelvan Swaminathan
    Biochemistry and Molecular Biology Education.2024; 52(2): 237.     CrossRef
  • Examining the use of ChatGPT in public universities in Hong Kong: a case study of restricted access areas
    Michelle W. T. Cheng, Iris H. Y. YIM
    Discover Education.2024;[Epub]     CrossRef
  • Performance of ChatGPT on Ophthalmology-Related Questions Across Various Examination Levels: Observational Study
    Firas Haddad, Joanna S Saade
    JMIR Medical Education.2024; 10: e50842.     CrossRef
  • Assessment of Artificial Intelligence Platforms With Regard to Medical Microbiology Knowledge: An Analysis of ChatGPT and Gemini
    Jai Ranjan, Absar Ahmad, Monalisa Subudhi, Ajay Kumar
    Cureus.2024;[Epub]     CrossRef
  • Comparison of the Performance of GPT-3.5 and GPT-4 With That of Medical Students on the Written German Medical Licensing Examination: Observational Study
    Annika Meyer, Janik Riese, Thomas Streichert
    JMIR Medical Education.2024; 10: e50965.     CrossRef
  • From hype to insight: Exploring ChatGPT's early footprint in education via altmetrics and bibliometrics
    Lung‐Hsiang Wong, Hyejin Park, Chee‐Kit Looi
    Journal of Computer Assisted Learning.2024; 40(4): 1428.     CrossRef
  • A scoping review of artificial intelligence in medical education: BEME Guide No. 84
    Morris Gordon, Michelle Daniel, Aderonke Ajiboye, Hussein Uraiby, Nicole Y. Xu, Rangana Bartlett, Janice Hanson, Mary Haas, Maxwell Spadafore, Ciaran Grafton-Clarke, Rayhan Yousef Gasiea, Colin Michie, Janet Corral, Brian Kwan, Diana Dolmans, Satid Thamma
    Medical Teacher.2024; 46(4): 446.     CrossRef
  • Üniversite Öğrencilerinin ChatGPT 3,5 Deneyimleri: Yapay Zekâyla Yazılmış Masal Varyantları
    Bilge GÖK, Fahri TEMİZYÜREK, Özlem BAŞ
    Korkut Ata Türkiyat Araştırmaları Dergisi.2024; (14): 1040.     CrossRef
  • Tracking ChatGPT Research: Insights From the Literature and the Web
    Omar Mubin, Fady Alnajjar, Zouheir Trabelsi, Luqman Ali, Medha Mohan Ambali Parambil, Zhao Zou
    IEEE Access.2024; 12: 30518.     CrossRef
  • Potential applications of ChatGPT in obstetrics and gynecology in Korea: a review article
    YooKyung Lee, So Yun Kim
    Obstetrics & Gynecology Science.2024; 67(2): 153.     CrossRef
  • Application of generative language models to orthopaedic practice
    Jessica Caterson, Olivia Ambler, Nicholas Cereceda-Monteoliva, Matthew Horner, Andrew Jones, Arwel Tomos Poacher
    BMJ Open.2024; 14(3): e076484.     CrossRef
  • Opportunities, challenges, and future directions of large language models, including ChatGPT in medical education: a systematic scoping review
    Xiaojun Xu, Yixiao Chen, Jing Miao
    Journal of Educational Evaluation for Health Professions.2024; 21: 6.     CrossRef
  • The advent of ChatGPT: Job Made Easy or Job Loss to Data Analysts
    Abiola Timothy Owolabi, Oluwaseyi Oluwadamilare Okunlola, Emmanuel Taiwo Adewuyi, Janet Iyabo Idowu, Olasunkanmi James Oladapo
    WSEAS TRANSACTIONS ON COMPUTERS.2024; 23: 24.     CrossRef
  • ChatGPT in dentomaxillofacial radiology education
    Hilal Peker Öztürk, Hakan Avsever, Buğra Şenel, Şükran Ayran, Mustafa Çağrı Peker, Hatice Seda Özgedik, Nurten Baysal
    Journal of Health Sciences and Medicine.2024; 7(2): 224.     CrossRef
  • Performance of ChatGPT on the Korean National Examination for Dental Hygienists
    Soo-Myoung Bae, Hye-Rim Jeon, Gyoung-Nam Kim, Seon-Hui Kwak, Hyo-Jin Lee
    Journal of Dental Hygiene Science.2024; 24(1): 62.     CrossRef
  • Medical knowledge of ChatGPT in public health, infectious diseases, COVID-19 pandemic, and vaccines: multiple choice questions examination based performance
    Sultan Ayoub Meo, Metib Alotaibi, Muhammad Zain Sultan Meo, Muhammad Omair Sultan Meo, Mashhood Hamid
    Frontiers in Public Health.2024;[Epub]     CrossRef
  • Unlock the potential for Saudi Arabian higher education: a systematic review of the benefits of ChatGPT
    Eman Faisal
    Frontiers in Education.2024;[Epub]     CrossRef
  • Does the Information Quality of ChatGPT Meet the Requirements of Orthopedics and Trauma Surgery?
    Adnan Kasapovic, Thaer Ali, Mari Babasiz, Jessica Bojko, Martin Gathen, Robert Kaczmarczyk, Jonas Roos
    Cureus.2024;[Epub]     CrossRef
  • Exploring the Profile of University Assessments Flagged as Containing AI-Generated Material
    Daniel Gooch, Kevin Waugh, Mike Richards, Mark Slaymaker, John Woodthorpe
    ACM Inroads.2024; 15(2): 39.     CrossRef
  • Comparing the Performance of ChatGPT-4 and Medical Students on MCQs at Varied Levels of Bloom’s Taxonomy
    Ambadasu Bharatha, Nkemcho Ojeh, Ahbab Mohammad Fazle Rabbi, Michael Campbell, Kandamaran Krishnamurthy, Rhaheem Layne-Yarde, Alok Kumar, Dale Springer, Kenneth Connell, Md Anwarul Majumder
    Advances in Medical Education and Practice.2024; Volume 15: 393.     CrossRef
  • The emergence of generative artificial intelligence platforms in 2023, journal metrics, appreciation to reviewers and volunteers, and obituary
    Sun Huh
    Journal of Educational Evaluation for Health Professions.2024; 21: 9.     CrossRef
  • ChatGPT, a Friend or a Foe in Medical Education: A Review of Strengths, Challenges, and Opportunities
    Mahdi Zarei, Maryam Zarei, Sina Hamzehzadeh, Sepehr Shakeri Bavil Oliyaei, Mohammad-Salar Hosseini
    Shiraz E-Medical Journal.2024;[Epub]     CrossRef
  • Artificial intelligence chatbots for the nutrition management of diabetes and the metabolic syndrome
    Farah Naja, Mandy Taktouk, Dana Matbouli, Sharfa Khaleel, Ayah Maher, Berna Uzun, Maryam Alameddine, Lara Nasreddine
    European Journal of Clinical Nutrition.2024; 78(10): 887.     CrossRef
  • Large language models in healthcare: from a systematic review on medical examinations to a comparative analysis on fundamentals of robotic surgery online test
    Andrea Moglia, Konstantinos Georgiou, Pietro Cerveri, Luca Mainardi, Richard M. Satava, Alfred Cuschieri
    Artificial Intelligence Review.2024;[Epub]     CrossRef
  • Comparison of ChatGPT, Gemini, and Le Chat with physician interpretations of medical laboratory questions from an online health forum
    Annika Meyer, Ari Soleman, Janik Riese, Thomas Streichert
    Clinical Chemistry and Laboratory Medicine (CCLM).2024;[Epub]     CrossRef
  • Performance of ChatGPT-3.5 and GPT-4 in national licensing examinations for medicine, pharmacy, dentistry, and nursing: a systematic review and meta-analysis
    Hye Kyung Jin, Ha Eun Lee, EunYoung Kim
    BMC Medical Education.2024;[Epub]     CrossRef
  • Role of ChatGPT in Dentistry: A Review
    Pratik Surana, Priyanka P. Ostwal, Shruti Vishal Dev, Jayesh Tiwari, Kadire Shiva Charan Yadav, Gajji Renuka
    Research Journal of Pharmacy and Technology.2024; : 3489.     CrossRef
  • A Scoping Review on the Educational Applications of Generative AI in Primary and Secondary Education
    Solmoe Ahn, Jeongyoon Lee, Jungmin Park, Soyoung Jung, Jihoon Song
    The Journal of Korean Association of Computer Education.2024; 27(6): 11.     CrossRef
  • Performance of GPT-3.5 and GPT-4 on the Korean Pharmacist Licensing Examination: Comparison Study
    Hye Kyung Jin, EunYoung Kim
    JMIR Medical Education.2024; 10: e57451.     CrossRef
  • Evaluating the Feasibility of ChatGPT in Dental Morphology Education: A Pilot Study on AI-Assisted Learning in Dental Morphology
    Eun-Young Jeon, Hyun-Na Ahn, Jeong-Hyun Lee
    Journal of Dental Hygiene Science.2024; 24(4): 309.     CrossRef
  • Detecting AI- generated versus human- written medical student essays: a semi-randomized controlled study (Preprint)
    Berin Doru, Christoph Maier, Johanna Sophie Busse, Thomas Lücke, Judith Schönhoff, Elena Enax- Krumova, Steffen Hessler, Maria Berger, Marianne Tokic
    JMIR Medical Education.2024;[Epub]     CrossRef
  • Is ChatGPT reliable in education?
    Amal Abdullah Alibrahim
    South African Journal of Education.2024; 44(4): 1.     CrossRef
  • Advancing Scholarly Publishing Through Artificial Intelligence: A Paradigm Shift
    Muskan Dubey, Arun Kumar Dubey , Ravindra P Veeranna
    Trends in Scholarly Publishing.2024; 3(1): 1.     CrossRef
  • Comparative analysis of diagnostic accuracy in endodontic assessments: dental students vs. artificial intelligence
    Abubaker Qutieshat, Alreem Al Rusheidi, Samiya Al Ghammari, Abdulghani Alarabi, Abdurahman Salem, Maja Zelihic
    Diagnosis.2024; 11(3): 259.     CrossRef
  • Applicability of ChatGPT in Assisting to Solve Higher Order Problems in Pathology
    Ranwir K Sinha, Asitava Deb Roy, Nikhil Kumar, Himel Mondal
    Cureus.2023;[Epub]     CrossRef
  • Issues in the 3rd year of the COVID-19 pandemic, including computer-based testing, study design, ChatGPT, journal metrics, and appreciation to reviewers
    Sun Huh
    Journal of Educational Evaluation for Health Professions.2023; 20: 5.     CrossRef
  • Emergence of the metaverse and ChatGPT in journal publishing after the COVID-19 pandemic
    Sun Huh
    Science Editing.2023; 10(1): 1.     CrossRef
  • Assessing the Capability of ChatGPT in Answering First- and Second-Order Knowledge Questions on Microbiology as per Competency-Based Medical Education Curriculum
    Dipmala Das, Nikhil Kumar, Langamba Angom Longjam, Ranwir Sinha, Asitava Deb Roy, Himel Mondal, Pratima Gupta
    Cureus.2023;[Epub]     CrossRef
  • Evaluating ChatGPT's Ability to Solve Higher-Order Questions on the Competency-Based Medical Education Curriculum in Medical Biochemistry
    Arindam Ghosh, Aritri Bir
    Cureus.2023;[Epub]     CrossRef
  • Overview of Early ChatGPT’s Presence in Medical Literature: Insights From a Hybrid Literature Review by ChatGPT and Human Experts
    Omar Temsah, Samina A Khan, Yazan Chaiah, Abdulrahman Senjab, Khalid Alhasan, Amr Jamal, Fadi Aljamaan, Khalid H Malki, Rabih Halwani, Jaffar A Al-Tawfiq, Mohamad-Hani Temsah, Ayman Al-Eyadhy
    Cureus.2023;[Epub]     CrossRef
  • ChatGPT for Future Medical and Dental Research
    Bader Fatani
    Cureus.2023;[Epub]     CrossRef
  • ChatGPT in Dentistry: A Comprehensive Review
    Hind M Alhaidry, Bader Fatani, Jenan O Alrayes, Aljowhara M Almana, Nawaf K Alfhaed
    Cureus.2023;[Epub]     CrossRef
  • Can we trust AI chatbots’ answers about disease diagnosis and patient care?
    Sun Huh
    Journal of the Korean Medical Association.2023; 66(4): 218.     CrossRef
  • Large Language Models in Medical Education: Opportunities, Challenges, and Future Directions
    Alaa Abd-alrazaq, Rawan AlSaad, Dari Alhuwail, Arfan Ahmed, Padraig Mark Healy, Syed Latifi, Sarah Aziz, Rafat Damseh, Sadam Alabed Alrazak, Javaid Sheikh
    JMIR Medical Education.2023; 9: e48291.     CrossRef
  • Early applications of ChatGPT in medical practice, education and research
    Sam Sedaghat
    Clinical Medicine.2023; 23(3): 278.     CrossRef
  • A Review of Research on Teaching and Learning Transformation under the Influence of ChatGPT Technology
    璇 师
    Advances in Education.2023; 13(05): 2617.     CrossRef
  • Performance of GPT-3.5 and GPT-4 on the Japanese Medical Licensing Examination: Comparison Study
    Soshi Takagi, Takashi Watari, Ayano Erabi, Kota Sakaguchi
    JMIR Medical Education.2023; 9: e48002.     CrossRef
  • ChatGPT’s quiz skills in different otolaryngology subspecialties: an analysis of 2576 single-choice and multiple-choice board certification preparation questions
    Cosima C. Hoch, Barbara Wollenberg, Jan-Christoffer Lüers, Samuel Knoedler, Leonard Knoedler, Konstantin Frank, Sebastian Cotofana, Michael Alfertshofer
    European Archives of Oto-Rhino-Laryngology.2023; 280(9): 4271.     CrossRef
  • Analysing the Applicability of ChatGPT, Bard, and Bing to Generate Reasoning-Based Multiple-Choice Questions in Medical Physiology
    Mayank Agarwal, Priyanka Sharma, Ayan Goswami
    Cureus.2023;[Epub]     CrossRef
  • The Intersection of ChatGPT, Clinical Medicine, and Medical Education
    Rebecca Shin-Yee Wong, Long Chiau Ming, Raja Affendi Raja Ali
    JMIR Medical Education.2023; 9: e47274.     CrossRef
  • The Role of Artificial Intelligence in Higher Education: ChatGPT Assessment for Anatomy Course
    Tarık TALAN, Yusuf KALINKARA
    Uluslararası Yönetim Bilişim Sistemleri ve Bilgisayar Bilimleri Dergisi.2023; 7(1): 33.     CrossRef
  • Comparing ChatGPT’s ability to rate the degree of stereotypes and the consistency of stereotype attribution with those of medical students in New Zealand in developing a similarity rating test: a methodological study
    Chao-Cheng Lin, Zaine Akuhata-Huntington, Che-Wei Hsu
    Journal of Educational Evaluation for Health Professions.2023; 20: 17.     CrossRef
  • Examining Real-World Medication Consultations and Drug-Herb Interactions: ChatGPT Performance Evaluation
    Hsing-Yu Hsu, Kai-Cheng Hsu, Shih-Yen Hou, Ching-Lung Wu, Yow-Wen Hsieh, Yih-Dih Cheng
    JMIR Medical Education.2023; 9: e48433.     CrossRef
  • Assessing the Efficacy of ChatGPT in Solving Questions Based on the Core Concepts in Physiology
    Arijita Banerjee, Aquil Ahmad, Payal Bhalla, Kavita Goyal
    Cureus.2023;[Epub]     CrossRef
  • ChatGPT Performs on the Chinese National Medical Licensing Examination
    Xinyi Wang, Zhenye Gong, Guoxin Wang, Jingdan Jia, Ying Xu, Jialu Zhao, Qingye Fan, Shaun Wu, Weiguo Hu, Xiaoyang Li
    Journal of Medical Systems.2023;[Epub]     CrossRef
  • Artificial intelligence and its impact on job opportunities among university students in North Lima, 2023
    Doris Ruiz-Talavera, Jaime Enrique De la Cruz-Aguero, Nereo García-Palomino, Renzo Calderón-Espinoza, William Joel Marín-Rodriguez
    ICST Transactions on Scalable Information Systems.2023;[Epub]     CrossRef
  • Revolutionizing Dental Care: A Comprehensive Review of Artificial Intelligence Applications Among Various Dental Specialties
    Najd Alzaid, Omar Ghulam, Modhi Albani, Rafa Alharbi, Mayan Othman, Hasan Taher, Saleem Albaradie, Suhael Ahmed
    Cureus.2023;[Epub]     CrossRef
  • Opportunities, Challenges, and Future Directions of Generative Artificial Intelligence in Medical Education: Scoping Review
    Carl Preiksaitis, Christian Rose
    JMIR Medical Education.2023; 9: e48785.     CrossRef
  • Exploring the impact of language models, such as ChatGPT, on student learning and assessment
    Araz Zirar
    Review of Education.2023;[Epub]     CrossRef
  • Evaluating the reliability of ChatGPT as a tool for imaging test referral: a comparative study with a clinical decision support system
    Shani Rosen, Mor Saban
    European Radiology.2023; 34(5): 2826.     CrossRef
  • The Significance of Artificial Intelligence Platforms in Anatomy Education: An Experience With ChatGPT and Google Bard
    Hasan B Ilgaz, Zehra Çelik
    Cureus.2023;[Epub]     CrossRef
  • Is ChatGPT’s Knowledge and Interpretative Ability Comparable to First Professional MBBS (Bachelor of Medicine, Bachelor of Surgery) Students of India in Taking a Medical Biochemistry Examination?
    Abhra Ghosh, Nandita Maini Jindal, Vikram K Gupta, Ekta Bansal, Navjot Kaur Bajwa, Abhishek Sett
    Cureus.2023;[Epub]     CrossRef
  • Ethical consideration of the use of generative artificial intelligence, including ChatGPT in writing a nursing article
    Sun Huh
    Child Health Nursing Research.2023; 29(4): 249.     CrossRef
  • Potential Use of ChatGPT for Patient Information in Periodontology: A Descriptive Pilot Study
    Osman Babayiğit, Zeynep Tastan Eroglu, Dilek Ozkan Sen, Fatma Ucan Yarkac
    Cureus.2023;[Epub]     CrossRef
  • Efficacy and limitations of ChatGPT as a biostatistical problem-solving tool in medical education in Serbia: a descriptive study
    Aleksandra Ignjatović, Lazar Stevanović
    Journal of Educational Evaluation for Health Professions.2023; 20: 28.     CrossRef
  • Assessing the Performance of ChatGPT in Medical Biochemistry Using Clinical Case Vignettes: Observational Study
    Krishna Mohan Surapaneni
    JMIR Medical Education.2023; 9: e47191.     CrossRef
  • Performance of ChatGPT, Bard, Claude, and Bing on the Peruvian National Licensing Medical Examination: a cross-sectional study
    Betzy Clariza Torres-Zegarra, Wagner Rios-Garcia, Alvaro Micael Ñaña-Cordova, Karen Fatima Arteaga-Cisneros, Xiomara Cristina Benavente Chalco, Marina Atena Bustamante Ordoñez, Carlos Jesus Gutierrez Rios, Carlos Alberto Ramos Godoy, Kristell Luisa Teresa
    Journal of Educational Evaluation for Health Professions.2023; 20: 30.     CrossRef
  • ChatGPT’s performance in German OB/GYN exams – paving the way for AI-enhanced medical education and clinical practice
    Maximilian Riedel, Katharina Kaefinger, Antonia Stuehrenberg, Viktoria Ritter, Niklas Amann, Anna Graf, Florian Recker, Evelyn Klein, Marion Kiechle, Fabian Riedel, Bastian Meyer
    Frontiers in Medicine.2023;[Epub]     CrossRef
  • Medical students’ patterns of using ChatGPT as a feedback tool and perceptions of ChatGPT in a Leadership and Communication course in Korea: a cross-sectional study
    Janghee Park
    Journal of Educational Evaluation for Health Professions.2023; 20: 29.     CrossRef
  • FROM TEXT TO DIAGNOSE: CHATGPT’S EFFICACY IN MEDICAL DECISION-MAKING
    Yaroslav Mykhalko, Pavlo Kish, Yelyzaveta Rubtsova, Oleksandr Kutsyn, Valentyna Koval
    Wiadomości Lekarskie.2023; 76(11): 2345.     CrossRef
  • Using ChatGPT for Clinical Practice and Medical Education: Cross-Sectional Survey of Medical Students’ and Physicians’ Perceptions
    Pasin Tangadulrat, Supinya Sono, Boonsin Tangtrakulwanich
    JMIR Medical Education.2023; 9: e50658.     CrossRef
  • Below average ChatGPT performance in medical microbiology exam compared to university students
    Malik Sallam, Khaled Al-Salahat
    Frontiers in Education.2023;[Epub]     CrossRef
  • ChatGPT: "To be or not to be" ... in academic research. The human mind's analytical rigor and capacity to discriminate between AI bots' truths and hallucinations
    Aurelian Anghelescu, Ilinca Ciobanu, Constantin Munteanu, Lucia Ana Maria Anghelescu, Gelu Onose
    Balneo and PRM Research Journal.2023; 14(Vol.14, no): 614.     CrossRef
  • ChatGPT Review: A Sophisticated Chatbot Models in Medical & Health-related Teaching and Learning
    Nur Izah Ab Razak, Muhammad Fawwaz Muhammad Yusoff, Rahmita Wirza O.K. Rahmat
    Malaysian Journal of Medicine and Health Sciences.2023; 19(s12): 98.     CrossRef
  • Application of artificial intelligence chatbots, including ChatGPT, in education, scholarly work, programming, and content generation and its prospects: a narrative review
    Tae Won Kim
    Journal of Educational Evaluation for Health Professions.2023; 20: 38.     CrossRef
  • Trends in research on ChatGPT and adoption-related issues discussed in articles: a narrative review
    Sang-Jun Kim
    Science Editing.2023; 11(1): 3.     CrossRef
  • Information amount, accuracy, and relevance of generative artificial intelligence platforms’ answers regarding learning objectives of medical arthropodology evaluated in English and Korean queries in December 2023: a descriptive study
    Hyunju Lee, Soobin Park
    Journal of Educational Evaluation for Health Professions.2023; 20: 39.     CrossRef
  • What will ChatGPT revolutionize in the financial industry?
    Hassnian Ali, Ahmet Faruk Aysan
    Modern Finance.2023; 1(1): 116.     CrossRef
  • Performance of a Large Language Model in Medical Pharmacology Education: An Assessment Using Multiple-Choice Questions
    Benjamin S. Wright, Laura J. Kim, Nathan R. Coleman
    Annals of Pharmacy Education, Safety, and Public Health Advocacy.2023; 3(1): 232.     CrossRef
Review
What should medical students know about artificial intelligence in medicine?  
Seong Ho Park, Kyung-Hyun Do, Sungwon Kim, Joo Hyun Park, Young-Suk Lim
J Educ Eval Health Prof. 2019;16:18.   Published online July 3, 2019
DOI: https://doi.org/10.3352/jeehp.2019.16.18
  • 31,565 View
  • 783 Download
  • 116 Web of Science
  • 128 Crossref
AbstractAbstract PDFSupplementary Material
Artificial intelligence (AI) is expected to affect various fields of medicine substantially and has the potential to improve many aspects of healthcare. However, AI has been creating much hype, too. In applying AI technology to patients, medical professionals should be able to resolve any anxiety, confusion, and questions that patients and the public may have. Also, they are responsible for ensuring that AI becomes a technology beneficial for patient care. These make the acquisition of sound knowledge and experience about AI a task of high importance for medical students. Preparing for AI does not merely mean learning information technology such as computer programming. One should acquire sufficient knowledge of basic and clinical medicines, data science, biostatistics, and evidence-based medicine. As a medical student, one should not passively accept stories related to AI in medicine in the media and on the Internet. Medical students should try to develop abilities to distinguish correct information from hype and spin and even capabilities to create thoroughly validated, trustworthy information for patients and the public.

Citations

Citations to this article as recorded by  
  • The Impact of Artificial Intelligence Technologies on Nutritional Care in Patients With Chronic Kidney Disease: A Systematic Review
    Sara Morales Palomares, Gaetano Ferrara, Marco Sguanci, Domenica Gazineo, Lea Godino, Addolorata Palmisano, Alberto Paderno, Giada Vrenna, Eleonora Faraglia, Fabio Petrelli, Giovanni Cangelosi, Francesco Gravante, Stefano Mancin
    Journal of Renal Nutrition.2026; 36(1): 13.     CrossRef
  • A community-based AI and data science practicum: enhancing health information science education in Tanzania’s healthcare
    Rajabu Simba, Haruna Hussein, Augustino Mwogosi
    Information and Learning Sciences.2026; 127(1-2): 92.     CrossRef
  • Digital health literacy in medical education: a scoping review of current challenges and development strategies
    Guanli Xie, Jianglong Liao, Xiaoxia Tang, Yanfang Yang, Fu Han, Duo Liu, Deng Li, Yaju Jin, Tao Wang
    BMC Medical Education.2026;[Epub]     CrossRef
  • THE USE OF ARTIFICIAL INTELLIGENCE IN TRAUMATOLOGY: A SYSTEMATIC REVIEW AND RECOMMENDATIONS FOR CLINICAL PRACTICE
    V. V. Savgachev, L. B. Shubin
    Bulletin of Pirogov National Medical & Surgical Center.2026; 21(1): 127.     CrossRef
  • Awareness, attitudes, and educational use of artificial intelligence among medical students: a large cross-sectional survey
    Ali Veysel Kara, Hatice Harmancı, Yusuf Yılmaz
    BMC Medical Education.2026;[Epub]     CrossRef
  • Perceived knowledge, attitude, and practice of artificial intelligence among medical students in Guangxi: a cross-sectional study
    Lulin Chen, Wei Liu, Yanting Zhou
    Frontiers in Public Health.2026;[Epub]     CrossRef
  • Future-Ready Doctors: A Cross-Sectional Study of Undergraduate Knowledge of Artificial Intelligence in Clinical Biochemistry
    Susanna Theophilus Yesupatham, Ankita Kumari, Ravishankar Suryanarayana
    Cureus.2026;[Epub]     CrossRef
  • Artificial intelligence in dentistry: knowledge, attitudes, and educational readiness among dental students and dentists
    Ali Altındağ, Sultan Uzun, Ömer Altındağ, Kaan Orhan
    BMC Medical Education.2026;[Epub]     CrossRef
  • Implementation of a Novel Case-Based Session for Medical Students Focused on Artificial Intelligence Ethics
    Danielle M. Fernandes, Talya Lisker, Shitij Arora, Aaron Hui, Adira Hulkower, Sunit Jariwala, Janice Thomas John
    MedEdPORTAL.2026;[Epub]     CrossRef
  • Medical students’ knowledge and attitude towards using artificial intelligence in medical education and practice: a pre-post study
    Heba Tarek Emara, Mohamed Azmy Khafagy, Nermeen Ahmed Niazy, Sherehan Adel Abdel-Salam
    BMC Medical Education.2026;[Epub]     CrossRef
  • Potential of artificial intelligence methods in diseases of the venous system of the lower extremities
    S. Е. Katorkin
    Ambulatornaya khirurgiya = Ambulatory Surgery (Russia).2026; 23(1): 15.     CrossRef
  • Artificial Intelligence for Assessment and Feedback in Medical Education: Bibliometric Mapping Study and Thematic Evidence Map
    Zihang Zhao, Zihan Liu, Liang Guo, Teng Pan, Yousheng Zhang, Chenxiang Miao, Yiting Ge, Yipeng Wang, Xin Hu, Xin Wang, Ruipeng Zhang, Zhiyong Hou
    JMIR Medical Education.2026; 12: e98949.     CrossRef
  • ÖĞRETMEN ADAYLARININ YAPAY ZEKÂ HAZIR BULUNUŞLUK DÜZEYLERİNİN ÇEŞİTLİ DEĞİŞKENLER AÇISINDAN İNCELENMESİ
    Tülin Hündür, Erol Taş
    Trakya Eğitim Dergisi.2026; 16(3): 1472.     CrossRef
  • Healthcare sciences lecturers' views on the use of artificial intelligence for patient diagnosis in Gauteng province, South Africa: qualitative study
    Raikane James Seretlo, Aminat Oluwatoyin Adelowotan
    Frontiers in Education.2026;[Epub]     CrossRef
  • Artificial intelligence in medical education: Curriculum gaps and medical student perceptions at Qassim university
    Ahmad S. Alamro
    Saudi Journal for Health Sciences.2026; 15(2): 140.     CrossRef
  • Artificial Intelligence In Dental Exams: A Performance Comparison Of ChatGPT, DeepSeek, And Students
    Şükran Ayran, Selma Ece Karabıyıkoğlu, Neşe Oral, Hilal Peker Öztürk, Barış Filiz Erol, Hakan Avsever
    Acta Odontologica Turcica.2026; 43(3): 160.     CrossRef
  • The Promise of Artificial Intelligence and Machine Learning in Geriatric Anesthesiology Education: An Idea Whose Time Has Come
    Larry F. Chu, Viji Kurup
    Current Anesthesiology Reports.2025;[Epub]     CrossRef
  • Exploring the potential of acupuncture practice education using artificial intelligence
    Kyeong Han Kim, Hyein Jeong, Gyeong Seo Lee, Seung-Hee Lee
    Integrative Medicine Research.2025; 14(1): 101123.     CrossRef
  • The data-intensive research paradigm: challenges and responses in clinical professional graduate education
    Chunhong Yang, Yijing Chen, Changshun Qian, Fangmin Shi, You Guo
    Frontiers in Medicine.2025;[Epub]     CrossRef
  • Factors affecting medical artificial intelligence (AI) readiness among medical students: taking stock and looking forward
    Arash Ziapour, Fatemeh Darabi, Parisa Janjani, Mohammad Amin Amani, Murat Yıldırım, Sayeh Motevaseli
    BMC Medical Education.2025;[Epub]     CrossRef
  • Artificial Intelligence Literacy Levels of Perioperative Nurses: The Case of Türkiye
    Hilal Kahraman, Seda Akutay, Hatice Yüceler Kaçmaz, Sultan Taşci
    Nursing & Health Sciences.2025;[Epub]     CrossRef
  • Advantages and limitations of large language models for antibiotic prescribing and antimicrobial stewardship
    Daniele Roberto Giacobbe, Cristina Marelli, Bianca La Manna, Donatella Padua, Alberto Malva, Sabrina Guastavino, Alessio Signori, Sara Mora, Nicola Rosso, Cristina Campi, Michele Piana, Ylenia Murgia, Mauro Giacomini, Matteo Bassetti
    npj Antimicrobials and Resistance.2025;[Epub]     CrossRef
  • Essential competencies of nurses working with AI-driven lifestyle monitoring in long-term care: A modified Delphi study
    S.W.M. Groeneveld, H. van Os-Medendorp, J.E.W.C. van Gemert-Pijnen, R.M. Verdaasdonk, T. van Houwelingen, T. Dekkers, M.E.M. den Ouden
    Nurse Education Today.2025; 149: 106659.     CrossRef
  • The Relationship Between Anxiety and Readiness Levels Regarding Artificial Intelligence in Midwives
    Ayşe Nur Yilmaz, Sümeyye Altiparmak, Remziye Sökmen
    CIN: Computers, Informatics, Nursing.2025;[Epub]     CrossRef
  • Assessing artificial intelligence knowledge among Al-Zahraa university students: A cross-sectional study
    Hassan Hadi Al kazzaz, Ahmad Hassan Kazzaz, Sarah Kazzaz, Ali Al Mousawi
    F1000Research.2025; 14: 405.     CrossRef
  • Exploring Filipino Medical Students’ Attitudes and Perceptions of Artificial Intelligence in Medical Education: A Mixed-Methods Study
    Robbi Miguel G. Falcon, Renne Margaret U. Alcazar, Hannah G. Babaran, Beatrice Dominique B. Caragay, Cheenie Ann A. Corpuz, Maegan Victoria S. Kho, Aleisha Claire N. Perez, Iris Thiele C. Isip-Tan
    MedEdPublish.2025; 14: 282.     CrossRef
  • Novel Blended Learning on Artificial Intelligence for Medical Students: Qualitative Interview Study
    Zoe S Oftring, Kim Deutsch, Daniel Tolks, Florian Jungmann, Sebastian Kuhn
    JMIR Medical Education.2025; 11: e65220.     CrossRef
  • Future-ready medicine: Assessing the need for A.I. education in Indian undergraduate medical curriculum: A mixed method survey of student perspectives
    Smita R. Sorte, Alka T. Rawekar, Sachin B. Rathod, Nisha Surana Gandhi
    Journal of Education and Health Promotion.2025;[Epub]     CrossRef
  • Inteligencia Artificial en Educación Médica: ¿hacia dónde?
    Federico Antillón
    Revista de la Facultad de Medicina.2025; 3(1): 4.     CrossRef
  • Semantic and Visual Pathways to Artificial intelligence Literacy. Challenges and Lessons Learned in the Medical Domain
    Pablo Pérez-Sánchez, Andrea Vázquez-Ingelmo, Marco Terzo Zani, Víctor Vicente-Palacios, Antonio Sánchez-Puente, Francisco José García-Peñalvo, Pedro Luis Sánchez
    International Journal on Semantic Web and Information Systems.2025; 21(1): 1.     CrossRef
  • Exploring the Impact of Female Student’s Digital Intelligence on Sustainable Learning and Digital Mental Well-Being: A Case Study of Saudi Arabia
    Norah Muflih Alruwaili, Zaiba Ali, Mohd Shuaib Siddiqui, Asad Hassan Butt, Hassan Ahmad, Rahila Ali, Shaden Hamad Alsalem
    Sustainability.2025; 17(14): 6632.     CrossRef
  • Tıp ve diş hekimliği fakültesi öğrencilerinin yapay zekaya yönelik genel tutumu ve yapay zeka okuryazarlık seviyelerinin belirlenmesi
    Yunus Emre Kaban, Danış Aygün, Ayşen Til
    Tıp Eğitimi Dünyası.2025; 24(73): 19.     CrossRef
  • Knowledge Attitudes and Ethical Concerns About Artificial Intelligence Among Medical Students at Taibah University: A Cross-Sectional Study
    Samah Alfahl
    Advances in Medical Education and Practice.2025; Volume 16: 1609.     CrossRef
  • Exploring AI literacy, attitudes toward AI, and intentions to use AI in clinical contexts among healthcare students in Korea: a cross-sectional study
    Jihyun Si
    BMC Medical Education.2025;[Epub]     CrossRef
  • Medical Students’ Attitudes towards Artificial Intelligence and Educational Needs: A Cross-sectional Study
    Ehsan Moallem, Vahid Ghavami, Javad Moghri, Abolfazl Marvi, Mahboobe Najafi, Seyed saeed Tabatabaee
    Journal of Health Administration.2025; 28(1): 40.     CrossRef
  • Tıp fakültesi öğrencilerinin sağlıkta yapay zekanın uygulanabilirliği ve etiği hakkındaki görüşlerinin araştırılması
    Maide Barış, Kerim Kağıt, Zeynep Betül Yazıcı
    Anadolu Kliniği Tıp Bilimleri Dergisi.2025; 30(3): 404.     CrossRef
  • Perception of Medical Undergraduates on Artificial Intelligence in Medical Education: Qualitative Exploration
    Thilanka Seneviratne, Kaumudee Kodikara, Isuru Abeykoon, Wathsala Palpola
    JMIR Medical Education.2025; 11: e73798.     CrossRef
  • Clinical utility of artificial intelligence models in radiology: a systemic scoping review of diagnostic and endovascular applications
    Som P. Singh, Aarya Ramprasad, Mina S. Makary
    CVIR Endovascular.2025;[Epub]     CrossRef
  • Generative Artificial Intelligence in Urology: Navigating the Frontier of Ethical, Legal, and Clinical Challenges
    Waqas Khalil, Mazhar Sheikh, Jawad U Islam
    Cureus.2025;[Epub]     CrossRef
  • Content and structural needs assessment for an artificial intelligence education mobile app in healthcare: a mixed methods study
    Seyyedeh Fatemeh Mousavi Baigi, Reyhane Norouzi Aval, Masoumeh Sarbaz, Seyyed Mohammad Tabatabaei, Khalil Kimiafar
    BMC Medical Education.2025;[Epub]     CrossRef
  • Performance and risks of ChatGPT used in drug information: an exploratory real-world analysis
    Benedict Morath, Ute Chiriac, Elena Jaszkowski, Carolin Deiß, Hannah Nürnberg, Katrin Hörth, Torsten Hoppe-Tichy, Kim Green
    European Journal of Hospital Pharmacy.2024; 31(6): 491.     CrossRef
  • Radiology as a Specialty in the Era of Artificial Intelligence: A Systematic Review and Meta-analysis on Medical Students, Radiology Trainees, and Radiologists
    Amir Hassankhani, Melika Amoukhteh, Parya Valizadeh, Payam Jannatdoust, Paniz Sabeghi, Ali Gholamrezanezhad
    Academic Radiology.2024; 31(1): 306.     CrossRef
  • Views of veterinary faculty students on the concept of Artificial Intelligence and its use in Veterinary Medicine practices: An example of Ankara University Faculty of Veterinary Medicine
    Nigar Yerlikaya, Özgül Küçükaslan
    Ankara Üniversitesi Veteriner Fakültesi Dergisi.2024; 71(3): 249.     CrossRef
  • Strategies for Implementing Machine Learning Algorithms in the Clinical Practice of Radiology
    Allison Chae, Michael S. Yao, Hersh Sagreiya, Ari D. Goldberg, Neil Chatterjee, Matthew T. MacLean, Jeffrey Duda, Ameena Elahi, Arijitt Borthakur, Marylyn D. Ritchie, Daniel Rader, Charles E. Kahn, Walter R. Witschey, James C. Gee
    Radiology.2024;[Epub]     CrossRef
  • Towards integration of artificial intelligence into medical devices as a real-time recommender system for personalised healthcare: State-of-the-art and future prospects
    Talha Iqbal, Mehedi Masud, Bilal Amin, Conor Feely, Mary Faherty, Tim Jones, Michelle Tierney, Atif Shahzad, Patricia Vazquez
    Health Sciences Review.2024; 10: 100150.     CrossRef
  • The Knowledge of Students at Bursa Faculty of Medicine towards Artificial Intelligence: A Survey Study
    Deniz GÜVEN, Elif Güler KAZANCI, Ayşe ÖREN, Livanur SEVER, Pelin ÜNLÜ
    Journal of Bursa Faculty of Medicine.2024; 2(1): 20.     CrossRef
  • Preparing healthcare leaders of the digital age with an integrative artificial intelligence curriculum: a pilot study
    Soo Hwan Park, Roshini Pinto-Powell, Thomas Thesen, Alexander Lindqwister, Joshua Levy, Rachael Chacko, Devina Gonzalez, Connor Bridges, Adam Schwendt, Travis Byrum, Justin Fong, Shahin Shahsavari, Saeed Hassanpour
    Medical Education Online.2024;[Epub]     CrossRef
  • A scoping review of artificial intelligence in medical education: BEME Guide No. 84
    Morris Gordon, Michelle Daniel, Aderonke Ajiboye, Hussein Uraiby, Nicole Y. Xu, Rangana Bartlett, Janice Hanson, Mary Haas, Maxwell Spadafore, Ciaran Grafton-Clarke, Rayhan Yousef Gasiea, Colin Michie, Janet Corral, Brian Kwan, Diana Dolmans, Satid Thamma
    Medical Teacher.2024; 46(4): 446.     CrossRef
  • Artificial Intelligence Readiness Status of Medical Faculty Students
    Büşra EMİR, Tulin YURDEM, Tulin OZEL, Toygar SAYAR, Teoman Atalay UZUN, Umit AKAR, Unal Arda COLAK
    Konuralp Tıp Dergisi.2024; 16(1): 88.     CrossRef
  • Potential applications of ChatGPT in obstetrics and gynecology in Korea: a review article
    YooKyung Lee, So Yun Kim
    Obstetrics & Gynecology Science.2024; 67(2): 153.     CrossRef
  • ChatGPT in dentomaxillofacial radiology education
    Hilal Peker Öztürk, Hakan Avsever, Buğra Şenel, Şükran Ayran, Mustafa Çağrı Peker, Hatice Seda Özgedik, Nurten Baysal
    Journal of Health Sciences and Medicine.2024; 7(2): 224.     CrossRef
  • Twelve tips for addressing ethical concerns in the implementation of artificial intelligence in medical education
    Russell Franco D’Souza, Mary Mathew, Vedprakash Mishra, Krishna Mohan Surapaneni
    Medical Education Online.2024;[Epub]     CrossRef
  • Examining labelling guidelines for AI‐based software as a medical device: A review and analysis of dermatology mobile applications in Australia
    Ayooluwatomiwa Oloruntoba, Åsa Ingvar, Maithili Sashindranath, Ojochonu Anthony, Lisa Abbott, Pascale Guitera, Tony Caccetta, Monika Janda, H. Peter Soyer, Victoria Mar
    Australasian Journal of Dermatology.2024; 65(5): 409.     CrossRef
  • Comparing the Performance of ChatGPT-4 and Medical Students on MCQs at Varied Levels of Bloom’s Taxonomy
    Ambadasu Bharatha, Nkemcho Ojeh, Ahbab Mohammad Fazle Rabbi, Michael Campbell, Kandamaran Krishnamurthy, Rhaheem Layne-Yarde, Alok Kumar, Dale Springer, Kenneth Connell, Md Anwarul Majumder
    Advances in Medical Education and Practice.2024; Volume 15: 393.     CrossRef
  • Artificial intelligence and learning environment: Human considerations
    Esmaeil Jafari
    Journal of Computer Assisted Learning.2024; 40(5): 2135.     CrossRef
  • ChatGPT, a Friend or a Foe in Medical Education: A Review of Strengths, Challenges, and Opportunities
    Mahdi Zarei, Maryam Zarei, Sina Hamzehzadeh, Sepehr Shakeri Bavil Oliyaei, Mohammad-Salar Hosseini
    Shiraz E-Medical Journal.2024;[Epub]     CrossRef
  • Design and validation of an artificial intelligence-powered instrument for the assessment of migraine risk in university students in Lebanon
    Zahraa Tahhan, Georges Hatem, Ahmed M. Abouelmaty, Zad Rafei, Sanaa Awada
    Computers in Human Behavior Reports.2024; 15: 100453.     CrossRef
  • Curriculum Frameworks and Educational Programs in AI for Medical Students, Residents, and Practicing Physicians: Scoping Review
    Raymond Tolentino, Ashkan Baradaran, Genevieve Gore, Pierre Pluye, Samira Abbasgholizadeh-Rahimi
    JMIR Medical Education.2024; 10: e54793.     CrossRef
  • Artificial intelligence in medical education - perception among medical students
    Preetha Jackson, Gayathri Ponath Sukumaran, Chikku Babu, M. Christa Tony, Deen Stephano Jack, V. R. Reshma, Dency Davis, Nisha Kurian, Anjum John
    BMC Medical Education.2024;[Epub]     CrossRef
  • The “Magical Theory” of AI in Medicine: Thematic Narrative Analysis
    Giorgia Lorenzini, Laura Arbelaez Ossa, Stephen Milford, Bernice Simone Elger, David Martin Shaw, Eva De Clercq
    JMIR AI.2024; 3: e49795.     CrossRef
  • Encompassing trust in medical AI from the perspective of medical students: a quantitative comparative study
    Anamaria Malešević, Mária Kolesárová, Anto Čartolovni
    BMC Medical Ethics.2024;[Epub]     CrossRef
  • Patient Autonomy in Medical Education: Navigating Ethical Challenges in the Age of Artificial Intelligence
    Hui Lu, Ahmad Alhaskawi, Yanzhao Dong, Xiaodi Zou, Haiying Zhou, Sohaib Hasan Abdullah Ezzi, Vishnu Goutham Kota, Mohamed Hasan Abdulla Hasan Abdulla, Sahar Ahmed Abdalbary
    INQUIRY: The Journal of Health Care Organization, Provision, and Financing.2024;[Epub]     CrossRef
  • Medical Education and Artificial Intelligence: Web of Science–Based Bibliometric Analysis (2013-2022)
    Shuang Wang, Liuying Yang, Min Li, Xinghe Zhang, Xiantao Tai
    JMIR Medical Education.2024; 10: e51411.     CrossRef
  • Correlates of Medical and Allied Health Students’ Engagement with Generative AI in Nigeria
    Zubairu Iliyasu, Hameedat O. Abdullahi, Bilkisu Z. Iliyasu, Humayra A. Bashir, Taiwo G. Amole, Hadiza M. Abdullahi, Amina U. Abdullahi, Aminatu A. Kwaku, Tahir Dahir, Fatimah I. Tsiga-Ahmed, Abubakar M. Jibo, Hamisu M. Salihu, Muktar H. Aliyu
    Medical Science Educator.2024; 35(1): 269.     CrossRef
  • Going beyond competencies: Building blocks for a patient- and population-centered medical curriculum
    Mohi Eldin Magzoub, Mohammed Hassan Taha, Susan Waller, Awad Mansour Al Eissa, Hossam Hamdy, John Norcini, Saeeda Al Marzooqi, Sami Shaban, Mohammed Elhassan Abdalla, Henk Schmidt
    Medical Teacher.2024; 46(12): 1568.     CrossRef
  • Attitudes and perceptions of Thai medical students regarding artificial intelligence in radiology and medicine
    Salita Angkurawaranon, Nakarin Inmutto, Kittipitch Bannangkoon, Surapat Wonghan, Thanawat Kham-ai, Porched Khumma, Kanvijit Daengpisut, Phattanun Thabarsa, Chaisiri Angkurawaranon
    BMC Medical Education.2024;[Epub]     CrossRef
  • Exploring Filipino Medical Students’ Attitudes and Perceptions of Artificial Intelligence in Medical Education: A Mixed-Methods Study
    Robbi Miguel G. Falcon, Renne Margaret U. Alcazar, Hannah G. Babaran, Beatrice Dominique B. Caragay, Cheenie Ann A. Corpuz, Maegan Victoria S. Kho, Aleisha Claire N. Perez, Iris Thiele C. Isip-Tan
    MedEdPublish.2024; 14: 282.     CrossRef
  • Exploring the Ethical Implications of ChatGPT in Medical Education: Privacy, Accuracy, and Professional Integrity in a Cross-Sectional Survey
    Hafiz Muhammad Amad Abdullah, Noor-i-Kiran Naeem, Ghufran Ali Malkana, Masib Javed, Malik Muhammad Shahzaib
    Cureus.2024;[Epub]     CrossRef
  • Medical Students’ Perspectives on Trust in Medical AI: A Quantitative Comparative Study
    Jana Kajanova, Anamaria Badrov
    Asian Journal of Ethics in Health and Medicine.2024; 4(1): 44.     CrossRef
  • A novel adaptive cubic quasi‐Newton optimizer for deep learning based medical image analysis tasks, validated on detection of COVID‐19 and segmentation for COVID‐19 lung infection, liver tumor, and optic disc/cup
    Yan Liu, Maojun Zhang, Zhiwei Zhong, Xiangrong Zeng
    Medical Physics.2023; 50(3): 1528.     CrossRef
  • Clinical informatics training in medical school education curricula: a scoping review
    Humairah Zainal, Joshua Kuan Tan, Xin Xiaohui, Julian Thumboo, Fong Kok Yong
    Journal of the American Medical Informatics Association.2023; 30(3): 604.     CrossRef
  • Are ChatGPT’s knowledge and interpretation ability comparable to those of medical students in Korea for taking a parasitology examination?: a descriptive study
    Sun Huh
    Journal of Educational Evaluation for Health Professions.2023; 20: 1.     CrossRef
  • Exploring the views of Singapore junior doctors on medical curricula for the digital age: A case study
    Humairah Zainal, Xin Xiaohui, Julian Thumboo, Fong Kok Yong, Conor Gilligan
    PLOS ONE.2023; 18(3): e0281108.     CrossRef
  • Artificial Intelligence Teaching as Part of Medical Education: Qualitative Analysis of Expert Interviews
    Lukas Weidener, Michael Fischer
    JMIR Medical Education.2023; 9: e46428.     CrossRef
  • Investigating Students’ Perceptions towards Artificial Intelligence in Medical Education
    Ali Jasem Buabbas, Brouj Miskin, Amar Ali Alnaqi, Adel K. Ayed, Abrar Abdulmohsen Shehab, Shabbir Syed-Abdul, Mohy Uddin
    Healthcare.2023; 11(9): 1298.     CrossRef
  • A closer look at the current knowledge and prospects of artificial intelligence integration in dentistry practice: A cross-sectional study
    Zuhal Y. Hamd, Wiam Elshami, Sausan Al Kawas, Hanan Aljuaid, Mohamed M. Abuzaid
    Heliyon.2023; 9(6): e17089.     CrossRef
  • ChatGPT and the Future of Digital Health: A Study on Healthcare Workers’ Perceptions and Expectations
    Mohamad-Hani Temsah, Fadi Aljamaan, Khalid H. Malki, Khalid Alhasan, Ibraheem Altamimi, Razan Aljarbou, Faisal Bazuhair, Abdulmajeed Alsubaihin, Naif Abdulmajeed, Fatimah S. Alshahrani, Reem Temsah, Turki Alshahrani, Lama Al-Eyadhy, Serin Mohammed Alkhate
    Healthcare.2023; 11(13): 1812.     CrossRef
  • The Impact of Artificial Intelligence on the Preference of Radiology as a Future Specialty Among Medical Students at Jazan University, Saudi Arabia: A Cross-Sectional Study
    Khalid M Hakami, Mohammed Alameer, Essa Jaawna, Abdulrahman Sudi, Bahiyyah Bahkali, Amnah Mohammed, Abdulaziz Hakami, Mohamed Salih Mahfouz, Abdulaziz H Alhazmi, Turki M Dhayihi
    Cureus.2023;[Epub]     CrossRef
  • Application of artificial intelligence in medical education: focus on the application of ChatGPT for clinical medical education
    Hyeonmi Hong, Youngjoon Kang, Youngjon Kim, Bomsol Kim
    Journal of Medicine and Life Science.2023; 20(2): 53.     CrossRef
  • Medical Students’ Perspectives on Artificial Intelligence in Radiology: The Current Understanding and Impact on Radiology as a Future Specialty Choice
    Ali Alamer
    Current Medical Imaging Formerly Current Medical Imaging Reviews.2023;[Epub]     CrossRef
  • Psychometric properties of the persian version of the Medical Artificial Intelligence Readiness Scale for Medical Students (MAIRS-MS)
    AmirAli Moodi Ghalibaf, Maryam Moghadasin, Ali Emadzadeh, Haniye Mastour
    BMC Medical Education.2023;[Epub]     CrossRef
  • A Pilot Remote Curriculum to Enhance Resident and Medical Student Understanding of Machine Learning in Healthcare
    Seth M. Meade, Sebastian Salas-Vega, Matthew R. Nagy, Swetha J. Sundar, Michael P. Steinmetz, Edward C. Benzel, Ghaith Habboub
    World Neurosurgery.2023; 180: e142.     CrossRef
  • Medical Students’ Knowledge and Attitudes about Artificial Intelligence: A Cross-Sectional Survey
    Amber EKER, Ahmet Asım ÇALIŞKAN, Aysel ZORALİ, Bensu KAYNAK, Mehmet Erhan DERİN
    Tıp Eğitimi Dünyası.2023; 22(68): 41.     CrossRef
  • El camino a futuro de la pediatría: Nuevas oportunidades con la inteligencia artificial en la atención infantil
    Wagner Rios-Garcia, Mayli M. Condori-Orosco, Cyntia J. Huasasquiche
    Investigación e Innovación Clínica y Quirúrgica Pediátrica.2023; 1(2): 71.     CrossRef
  • Generative Artificial Intelligence: Enhancing Patient Education in Cardiovascular Imaging
    Ahmed Marey, Abdelrahman M Saad, Benjamin D Killeen, Catalina Gomez, Mariia Tregubova, Mathias Unberath, Muhammad Umair
    BJR|Open.2023;[Epub]     CrossRef
  • Percepciones de estudiantes de Medicina sobre el impacto de la inteligencia artificial en radiología
    G. Caparrós Galán, F. Sendra Portero
    Radiología.2022; 64(6): 516.     CrossRef
  • Finding the needle by modeling the haystack: Pulmonary embolism in an emergency patient with cardiorespiratory manifestations
    Davide Luciani, Alessandro Magrini, Carlo Berzuini, Antonello Gavazzi, Paolo Canova, Tiziano Barbui, Guido Bertolini
    Expert Systems with Applications.2022; 189: 116066.     CrossRef
  • SHIFTing artificial intelligence to be responsible in healthcare: A systematic review
    Haytham Siala, Yichuan Wang
    Social Science & Medicine.2022; 296: 114782.     CrossRef
  • AUGMENTING CBME CURRICULUM WITH ARTIFICIAL INTELLIGENCE COURSES – A FUTURISTIC APPROACH.
    Yogesh Bahurupi, Ashwini A Mahadule, Prashant M Patil, Vartika Saxena
    INDIAN JOURNAL OF APPLIED RESEARCH.2022; : 46.     CrossRef
  • Artificial Intelligence in Pediatric Pathology: The Extinction of a Medical Profession or the Key to a Bright Future?
    Ananda van der Kamp, Tomas J. Waterlander, Thomas de Bel, Jeroen van der Laak, Marry M. van den Heuvel-Eibrink, Annelies M. C. Mavinkurve-Groothuis, Ronald R. de Krijger
    Pediatric and Developmental Pathology.2022; 25(4): 380.     CrossRef
  • Artificial Intelligence Education for the Health Workforce: Expert Survey of Approaches and Needs
    Kathleen Gray, John Slavotinek, Gerardo Luis Dimaguila, Dawn Choo
    JMIR Medical Education.2022; 8(2): e35223.     CrossRef
  • Advancements in Oncology with Artificial Intelligence—A Review Article
    Nikitha Vobugari, Vikranth Raja, Udhav Sethi, Kejal Gandhi, Kishore Raja, Salim R. Surani
    Cancers.2022; 14(5): 1349.     CrossRef
  • Needs, Challenges, and Applications of Artificial Intelligence in Medical Education Curriculum
    Joel Grunhut, Oge Marques, Adam T M Wyatt
    JMIR Medical Education.2022; 8(2): e35587.     CrossRef
  • Promoting Research, Awareness, and Discussion on AI in Medicine Using #MedTwitterAI: A Longitudinal Twitter Hashtag Analysis
    Faisal A. Nawaz, Austin A. Barr, Monali Y. Desai, Christos Tsagkaris, Romil Singh, Elisabeth Klager, Fabian Eibensteiner, Emil D. Parvanov, Mojca Hribersek, Maria Kletecka-Pulker, Harald Willschke, Atanas G. Atanasov
    Frontiers in Public Health.2022;[Epub]     CrossRef
  • Communication training for pharmacy students with standard patients using artificial intelligence
    Naoto Nakagawa, Keita Odanaka, Hiroshi Ohara, Shigeki Kisara
    Currents in Pharmacy Teaching and Learning.2022; 14(7): 854.     CrossRef
  • Artificial intelligence in healthcare: Should it be included in the medical curriculum? A students’ perspective
    MANISHI BANSAL, ANKUSH JINDAL
    The National Medical Journal of India.2022; 35: 56.     CrossRef
  • Undergraduate Medical Students’ and Interns’ Knowledge and Perception of Artificial Intelligence in Medicine
    Nisha Jha, Pathiyil Ravi Shankar, Mohammed Azmi Al-Betar, Rupesh Mukhia, Kabita Hada, Subish Palaian
    Advances in Medical Education and Practice.2022; Volume 13: 927.     CrossRef
  • Perceptions of US Medical Students on Artificial Intelligence in Medicine: Mixed Methods Survey Study
    David Shalom Liu, Jake Sawyer, Alexander Luna, Jihad Aoun, Janet Wang, Lord Boachie, Safwan Halabi, Bina Joe
    JMIR Medical Education.2022; 8(4): e38325.     CrossRef
  • Artificial intelligence in medical education: a cross-sectional needs assessment
    M. Murat Civaner, Yeşim Uncu, Filiz Bulut, Esra Giounous Chalil, Abdülhamit Tatli
    BMC Medical Education.2022;[Epub]     CrossRef
  • Medical students’ perceptions of the impact of artificial intelligence in radiology
    G. Caparrós Galán, F. Sendra Portero
    Radiología (English Edition).2022; 64(6): 516.     CrossRef
  • Medical Education 4.0: A Neurology Perspective
    Zaitoon Zafar, Muhammad Umair, Filzah Faheem, Danish Bhatti , Junaid S Kalia
    Cureus.2022;[Epub]     CrossRef
  • AI in the hands of imperfect users
    Kristin M. Kostick-Quenet, Sara Gerke
    npj Digital Medicine.2022;[Epub]     CrossRef
  • Trust and medical AI: the challenges we face and the expertise needed to overcome them
    Thomas P Quinn, Manisha Senadeera, Stephan Jacobs, Simon Coghlan, Vuong Le
    Journal of the American Medical Informatics Association.2021; 28(4): 890.     CrossRef
  • Attitude of Brazilian dentists and dental students regarding the future role of artificial intelligence in oral radiology: a multicenter survey
    Ruben Pauwels, Yumi Chokyu Del Rey
    Dentomaxillofacial Radiology.2021; 50(5): 20200461.     CrossRef
  • Key Principles of Clinical Validation, Device Approval, and Insurance Coverage Decisions of Artificial Intelligence
    Seong Ho Park, Jaesoon Choi, Jeong-Sik Byeon
    Korean Journal of Radiology.2021; 22(3): 442.     CrossRef
  • Basic of machine learning and deep learning in imaging for medical physicists
    Luigi Manco, Nicola Maffei, Silvia Strolin, Sara Vichi, Luca Bottazzi, Lidia Strigari
    Physica Medica.2021; 83: 194.     CrossRef
  • Inteligencia artificial y simulación en urología
    J. Gómez Rivas, C. Toribio Vázquez, C. Ballesteros Ruiz, M. Taratkin, J.L. Marenco, G.E. Cacciamani, E. Checcucci, Z. Okhunov, D. Enikeev, F. Esperto, R. Grossmann, B. Somani, D. Veneziano
    Actas Urológicas Españolas.2021; 45(8): 524.     CrossRef
  • Regulating AI in Health Care: The Challenges of Informed User Engagement
    Olya Kudina
    Hastings Center Report.2021; 51(5): 6.     CrossRef
  • Are We Ready to Integrate Artificial Intelligence Literacy into Medical School Curriculum: Students and Faculty Survey
    Elena A Wood, Brittany L Ange, D Douglas Miller
    Journal of Medical Education and Curricular Development.2021;[Epub]     CrossRef
  • A Conference-Friendly, Hands-on Introduction to Deep Learning for Radiology Trainees
    Walter F. Wiggins, M. Travis Caton, Kirti Magudia, Michael H. Rosenthal, Katherine P. Andriole
    Journal of Digital Imaging.2021; 34(4): 1026.     CrossRef
  • Artificial intelligence and simulation in urology
    J. Gómez Rivas, C. Toribio Vázquez, C. Ballesteros Ruiz, M. Taratkin, J.L. Marenco, G.E. Cacciamani, E. Checcucci, Z. Okhunov, D. Enikeev, F. Esperto, R. Grossmann, B. Somani, D. Veneziano
    Actas Urológicas Españolas (English Edition).2021; 45(8): 524.     CrossRef
  • Accelerating the Appropriate Adoption of Artificial Intelligence in Health Care: Protocol for a Multistepped Approach
    David Wiljer, Mohammad Salhia, Elham Dolatabadi, Azra Dhalla, Caitlin Gillan, Dalia Al-Mouaswas, Ethan Jackson, Jacqueline Waldorf, Jane Mattson, Megan Clare, Nadim Lalani, Rebecca Charow, Sarmini Balakumar, Sarah Younus, Tharshini Jeyakumar, Wanda Petean
    JMIR Research Protocols.2021; 10(10): e30940.     CrossRef
  • Artificial Intelligence in Undergraduate Medical Education: A Scoping Review
    Juehea Lee, Annie Siyu Wu, David Li, Kulamakan (Mahan) Kulasegaram
    Academic Medicine.2021; 96(11S): S62.     CrossRef
  • Artificial Intelligence Evidence-Based Current Status and Potential for Lower Limb Vascular Management
    Xenia Butova, Sergey Shayakhmetov, Maxim Fedin, Igor Zolotukhin, Sergio Gianesini
    Journal of Personalized Medicine.2021; 11(12): 1280.     CrossRef
  • Artificial Intelligence Education Programs for Health Care Professionals: Scoping Review
    Rebecca Charow, Tharshini Jeyakumar, Sarah Younus, Elham Dolatabadi, Mohammad Salhia, Dalia Al-Mouaswas, Melanie Anderson, Sarmini Balakumar, Megan Clare, Azra Dhalla, Caitlin Gillan, Shabnam Haghzare, Ethan Jackson, Nadim Lalani, Jane Mattson, Wanda Pete
    JMIR Medical Education.2021; 7(4): e31043.     CrossRef
  • The Journal Citation Indicator has arrived for Emerging Sources Citation Index journals, including the Journal of Educational Evaluation for Health Professions, in June 2021
    Sun Huh
    Journal of Educational Evaluation for Health Professions.2021; 18: 20.     CrossRef
  • Ethical Challenges of Artificial Intelligence in Health Care: A Narrative Review
    Aaron T. Hui, Shawn S. Ahn, Carolyn T. Lye, Jun Deng
    Ethics in Biology, Engineering and Medicine: An International Journal.2021; 12(1): 55.     CrossRef
  • Bayesian networks: Making the most of a history
    Rami Abbass, Usmaan Bhatti, Shad Asinger
    The Clinical Teacher.2021; 18(2): 140.     CrossRef
  • Fundamentals in Artificial Intelligence for Vascular Surgeons
    Juliette Raffort, Cédric Adam, Marion Carrier, Fabien Lareyre
    Annals of Vascular Surgery.2020; 65: 254.     CrossRef
  • Extending capabilities of artificial intelligence for decision-making and healthcare education
    Mohd Javaid, Abid Haleem, IbrahimHaleem Khan, Raju Vaishya, Abhishek Vaish
    Apollo Medicine.2020; 17(1): 53.     CrossRef
  • Artificial intelligence with multi-functional machine learning platform development for better healthcare and precision medicine
    Zeeshan Ahmed, Khalid Mohamed, Saman Zeeshan, XinQi Dong
    Database.2020;[Epub]     CrossRef
  • Artificial Intelligence Education and Tools for Medical and Health Informatics Students: Systematic Review
    A Hasan Sapci, H Aylin Sapci
    JMIR Medical Education.2020; 6(1): e19285.     CrossRef
  • Evaluation of epidemiological lectures using peer instruction: focusing on the importance of ConcepTests
    Toshiharu Mitsuhashi
    PeerJ.2020; 8: e9640.     CrossRef
  • Artificial Intelligence in Small Bowel Endoscopy: Current Perspectives and Future Directions
    Dinesh Meher, Mrinal Gogoi, Pankaj Bharali, Prajna Anirvan, Shivaram Prasad Singh
    Journal of Digestive Endoscopy.2020; 11(04): 245.     CrossRef
  • Key principles of clinical validation, device approval, and insurance coverage decisions of artificial intelligence
    Seong Ho Park, Jaesoon Choi, Jeong-Sik Byeon
    Journal of the Korean Medical Association.2020; 63(11): 696.     CrossRef
  • Artificial intelligence-based education assists medical students’ interpretation of hip fracture
    Chi-Tung Cheng, Chih-Chi Chen, Chih-Yuan Fu, Chung-Hsien Chaou, Yu-Tung Wu, Chih-Po Hsu, Chih-Chen Chang, I-Fang Chung, Chi-Hsun Hsieh, Ming-Ju Hsieh, Chien-Hung Liao
    Insights into Imaging.2020;[Epub]     CrossRef
  • Current Status and Future Direction of Artificial Intelligence in Healthcare and Medical Education
    Jin Sup Jung
    Korean Medical Education Review.2020; 22(2): 99.     CrossRef
  • Introducing Artificial Intelligence Training in Medical Education
    Ketan Paranjape, Michiel Schinkel, Rishi Nannan Panday, Josip Car, Prabath Nanayakkara
    JMIR Medical Education.2019; 5(2): e16048.     CrossRef
Research article
An expert-led and artificial intelligence system-assisted tutoring course to improve the confidence of Chinese medical interns in suturing and ligature skills: a prospective pilot study  
Ying-Ying Yang, Boaz Shulruf
J Educ Eval Health Prof. 2019;16:7.   Published online April 10, 2019
DOI: https://doi.org/10.3352/jeehp.2019.16.7
  • 23,880 View
  • 387 Download
  • 44 Web of Science
  • 48 Crossref
AbstractAbstract PDFSupplementary Material
Purpose
Lack of confidence in suturing/ligature skills due to insufficient practice and assessments is common among novice Chinese medical interns. This study aimed to improve the skill acquisition of medical interns through a new intervention program.
Methods
In addition to regular clinical training, expert-led or expert-led plus artificial intelligence (AI) system tutoring courses were implemented during the first 2 weeks of the surgical block. Interns could voluntarily join the regular (no additional tutoring), expert-led tutoring, or expert-led+AI tutoring groups freely. In the regular group, interns (n=25) did not receive additional tutoring. The expert-led group received 3-hour expert-led tutoring and in-training formative assessments after 2 practice sessions. After a similar expert-led course, the expert-led+AI group (n=23) practiced and assessed their skills on an AI system. Through a comparison with the internal standard, the system automatically recorded and evaluated every intern’s suturing/ligature skills. In the expert-led+AI group, performance and confidence were compared between interns who participated in 1, 2, or 3 AI practice sessions.
Results
The end-of-surgical block objective structured clinical examination (OSCE) performance and self-assessed confidence in suturing/ligature skills were highest in the expert-led+AI group. In comparison with the expert-led group, the expert-led+AI group showed similar performance in the in-training assessment and greater improvement in the end-of-surgical block OSCE. In the expert-led+AI group, the best performance and highest post-OSCE confidence were noted in those who engaged in 3 AI practice sessions.
Conclusion
This pilot study demonstrated the potential value of incorporating an additional expert-led+AI system–assisted tutoring course into the regular surgical curriculum.

Citations

Citations to this article as recorded by  
  • Validation and Reliability of the Turkish Adaptation of the Artificial Intelligence Literacy Scale (AILS) for Healthcare Professionals
    Birgül Yabana Kiremit
    International Journal of Human–Computer Interaction.2026; 42(10): 7087.     CrossRef
  • Exploring the relationship between health professionals’ artificial intelligence literacy and their attitudes toward artificial intelligence
    Hüseyin Çapuk, Muhammet Faruk Yiğit, Mehmet Uçar
    Informatics for Health and Social Care.2026; 51(2): 139.     CrossRef
  • Effects of virtual peer based on generative artificial Intelligence on pre-service teachers’ informational instructional design ability
    Rongping Que, Xue Zhang, Haipeng Wan
    Asia Pacific Journal of Education.2026; : 1.     CrossRef
  • Educational Impact of Automated Feedback Systems in Surgical Training: A Systematic Review With Quantitative Synthesis
    Gauri Harshawardhan Godbole, Daniel Hawkins, Kingsley Ewool, Mauro Henrique Batista Camacho, Rezaul Karim, Bijendra Patel
    Journal of Surgical Education.2026; 83(4): 103879.     CrossRef
  • Attitude, perception, and knowledge toward artificial intelligence among dental hygiene students and alumni: a cross-sectional survey study
    Youssef ElKhyatt, Mohamad Hassan Fadi Hijab, Dena Al-Thani
    Frontiers in Education.2026;[Epub]     CrossRef
  • Performance evaluation of generative pre-trained transformer on the National Veterinary Licensing Examination in Japan
    Takahiro Kako, Daiki Kato, Takaaki Iguchi, Shiyu Qin, Miki Ando, Shoma Koseki, Hayato Shibahara, Haruka Motoi, Rin Isaka, Namiko Ikeda, Hiroto Toyoda, Takayuki Nakagawa
    Scientific Reports.2026;[Epub]     CrossRef
  • Associations between stressors and leave-taking behavior among nursing interns: a cross-sectional quantitative survey
    Huawen Song, Yanping Liu
    Frontiers in Medicine.2026;[Epub]     CrossRef
  • Adaptability and Innovation of Artificial Intelligence in Educational Contexts: A Conceptual Framework
    Prof (Dr) Sajna Jaleel, Mariya George
    International Journal of Latest Technology in Engineering Management & Applied Science.2026; 15(5): 1965.     CrossRef
  • Artificial intelligence augmented tutoring vs expert instruction on learning simulated general surgical skills: a systematic review and meta-analysis
    Fouad Hanna, Mahmoud Mohamed Gad, Mohamed Bassiouny Fouad Helmy, Mohamed Mostafa Eisa, Mohamed Wael Z. Omran, Muhammad Youssef, Ahmed Youssef Hassan, Abdelrhman Waleed Kotb, Mariam A. Abusalah, Abdelrahman El-Helbawy, Mina Gamil Zekri Basta, Abd-Elfattah
    BMC Medical Education.2026;[Epub]     CrossRef
  • Ortopedi Alanındaki Yapay Zeka Uygulamaları
    Mehmet Kurt
    Arşiv Kaynak Tarama Dergisi.2026; 35(2): 97.     CrossRef
  • A systematic review of the impact of artificial intelligence on educational outcomes in health professions education
    Eva Feigerlova, Hind Hani, Ellie Hothersall-Davies
    BMC Medical Education.2025;[Epub]     CrossRef
  • Exploring the Role of Artificial Intelligence (AI)-Driven Training in Laparoscopic Suturing: A Systematic Review of Skills Mastery, Retention, and Clinical Performance in Surgical Education
    Chidozie N. Ogbonnaya, Shizhou Li, Changshi Tang, Baobing Zhang, Paul Sullivan, Mustafa Suphi Erden, Benjie Tang
    Healthcare.2025; 13(5): 571.     CrossRef
  • Applications of Artificial Intelligence in Medical Education: A Systematic Review
    Eric Hallquist, Ishank Gupta, Michael Montalbano, Marios Loukas
    Cureus.2025;[Epub]     CrossRef
  • Artificial Intelligence in Medical Education: a Scoping Review of the Evidence for Efficacy and Future Directions
    Kody Shaw, Marcus A. Henning, Craig S. Webster
    Medical Science Educator.2025; 35(3): 1803.     CrossRef
  • A systematic review and sequential explanatory synthesis: Artificial intelligence in healthcare education, a case of nursing
    S. Aslı Bozkurt, Sinan Aydoğan, Fatma Dursun Ergezen, Aykut Türkoğlu
    International Nursing Review.2025;[Epub]     CrossRef
  • Cross-cultural perspectives on AI adoption in teacher education: a comparative study of pre-service teachers in Turkey and the United Arab Emirates
    Ahmet Sami Konca, Ahmet Simsar, Reem Alhajji, Afra Al Mansoori
    Interactive Learning Environments.2025; 33(10): 5820.     CrossRef
  • A Comparative Bicentric Study on Ultrasound Education for Students: App- and AI-Supported Learning Versus Traditional Hands-on Instruction (AI-Teach Study)
    Elena Höhne, Eva Bauer, Claus Bauer, Valentin Schäfer, Jennifer Gotta, Philipp Reschke, Thomas Vogl, Ibrahim Yel, Johannes Weimer, Agnes Wittek, Florian Recker
    Academic Radiology.2025; 32(8): 4930.     CrossRef
  • Real-world implementation of an AI learning tool-MetaGP-Edu in medical education: A multi-center cohort study
    Yili Sun, Fei Liu
    Computers & Education.2025; 237: 105388.     CrossRef
  • Design and assessment of AI-based learning tools in higher education: a systematic review
    Jihao Luo, Chenxu Zheng, Jiamin Yin, Hock Hai Teo
    International Journal of Educational Technology in Higher Education.2025;[Epub]     CrossRef
  • Effectiveness of AI-assisted medical education for Chinese undergraduate medical students: a meta-analysis
    Jingyue Peng, Hongying Zhang, Xiaohua Tu, Xuemei Zhang, Qiuhan Wu, Yijun Wang, Deng Xiao
    BMC Medical Education.2025;[Epub]     CrossRef
  • The strengths, weaknesses, opportunities, and threats of generative artificial intelligence: a qualitative study of undergraduate nursing students
    You Yuan, Jing Fu, Lanlan Leng, Zhuosi Wen, Xiaoman Wei, Die Han, Xinyang Hu, Yu Liang, Qian Luo, Xia Zhang, Rujun Hu
    Frontiers in Public Health.2025;[Epub]     CrossRef
  • How University students in Bangladesh engage with ChatGPT: A qualitative study
    Mir Hasib, Md. Shariful Islam, Mu-Hsuan Huang
    PLOS One.2025; 20(9): e0333089.     CrossRef
  • Do generative artificial intelligence (GenAI) and science education mix? A systematic review of the literature
    Kason Ka Ching Cheung, Amina Zerouali, Jenna Koenen, Sibel Erduran
    Studies in Science Education.2025; : 1.     CrossRef
  • A holistic exploration of student attitudes toward AI use in higher education: an international comparison
    David Vaněček, Yilmaz Ilker Yorulmaz, Dana Dobrovská
    Cogent Education.2025;[Epub]     CrossRef
  • Bridging artificial intelligence and ethics in medical education: a comprehensive perspective from students and faculty
    Manjiri Vilas Hawal, Miriam Archana Simon, Noor Hasan Abdulla Husain, Salima Khamis Rashid Al-Harrasi, Lora Nasser Said Al-Hinai
    Journal of Medical Education Development.2025; 18(3): 14.     CrossRef
  • De la tiza al silicio: guía práctica para integrar la IA en docencia médica
    Pedro Errázuriz G.
    Revista Chilena de Reumatología.2025; 41(3): 72.     CrossRef
  • The impact of Generative AI (GenAI) on practices, policies and research direction in education: a case of ChatGPT and Midjourney
    Thomas K. F. Chiu
    Interactive Learning Environments.2024; 32(10): 6187.     CrossRef
  • Application value of an artificial intelligence-based diagnosis and recognition system in gastroscopy training for graduate students in gastroenterology: a preliminary study
    Peng An, Zhongqiu Wang
    Wiener Medizinische Wochenschrift.2024; 174(9-10): 173.     CrossRef
  • Automated measurement extraction for assessing simple suture quality in medical education
    Thanapon Noraset, Prawej Mahawithitwong, Wethit Dumronggittigule, Pongthep Pisarnturakit, Cherdsak Iramaneerat, Chanean Ruansetakit, Irin Chaikangwan, Nattanit Poungjantaradej, Nutcha Yodrabum
    Expert Systems with Applications.2024; 241: 122722.     CrossRef
  • Dental student application of artificial intelligence technology in detecting proximal caries lesions
    Enes Ayan, Yusuf Bayraktar, Çiğdem Çelik, Baturalp Ayhan
    Journal of Dental Education.2024; 88(4): 490.     CrossRef
  • Development of an Artificial Intelligence Teaching Assistant System for Undergraduate Nursing Students
    Yanika Kowitlawakul, Jocelyn Jie Min Tan, Siriwan Suebnukarn, Hoang D. Nguyen, Danny Chiang Choon Poo, Joseph Chai, Devi M. Kamala, Wenru Wang
    CIN: Computers, Informatics, Nursing.2024; 42(5): 334.     CrossRef
  • The Role of Artificial Intelligence in Medical Education: A Systematic Review
    Atinc Tozsin, Harun Ucmak, Selim Soyturk, Abdullatif Aydin, Ali Serdar Gozen, Maha Al Fahim, Selcuk Güven, Kamran Ahmed
    Surgical Innovation.2024; 31(4): 415.     CrossRef
  • Artificial Intelligence in Medical Education and Mentoring in Rehabilitation Medicine
    Julie K. Silver, Mustafa Reha Dodurgali, Nara Gavini
    American Journal of Physical Medicine & Rehabilitation.2024; 103(11): 1039.     CrossRef
  • The use of generative artificial intelligence in surgical education: a narrative review
    Lavina Rao, Eric Yang, Savannah Dissanayake, Roberto Cuomo, Ishith Seth, Warren M. Rozen
    Plastic and Aesthetic Research.2024;[Epub]     CrossRef
  • Topics and Trends of Health Informatics Education Research: Scientometric Analysis
    Qing Han
    JMIR Medical Education.2024; 10: e58165.     CrossRef
  • Uses of Artificial Intelligence in Medicine
    J. Vijay Rao
    Telangana Journal of IMA.2024; 4(2): 79.     CrossRef
  • AI-Driven Learning Management Systems: Modern Developments, Challenges and Future Trends during the Age of ChatGPT
    Sameer Qazi, Muhammad Bilal Kadri, Muhammad Naveed, Bilal A. Khawaja, Sohaib Zia Khan, Muhammad Mansoor Alam, Mazliham Mohd Su’ud
    Computers, Materials & Continua.2024; 80(2): 3289.     CrossRef
  • Systematic literature review on opportunities, challenges, and future research recommendations of artificial intelligence in education
    Thomas K.F. Chiu, Qi Xia, Xinyan Zhou, Ching Sing Chai, Miaoting Cheng
    Computers and Education: Artificial Intelligence.2023; 4: 100118.     CrossRef
  • Technological advancements in surgical laparoscopy considering artificial intelligence: a survey among surgeons in Germany
    Sebastian Lünse, Eric L. Wisotzky, Sophie Beckmann, Christoph Paasch, Richard Hunger, René Mantke
    Langenbeck's Archives of Surgery.2023;[Epub]     CrossRef
  • Artificial intelligence (AI) integration in medical education: A pan-India cross-sectional observation of acceptance and understanding among students
    Vipul Sharma, Uddhave Saini, Varun Pareek, Lokendra Sharma, Susheel Kumar
    Scripta Medica.2023; 54(4): 343.     CrossRef
  • Artificial Intelligence Methods and Artificial Intelligence-Enabled Metrics for Surgical Education: A Multidisciplinary Consensus
    S Swaroop Vedula, Ahmed Ghazi, Justin W Collins, Carla Pugh, Dimitrios Stefanidis, Ozanan Meireles, Andrew J Hung, Steven Schwaitzberg, Jeffrey S Levy, Ajit K Sachdeva
    Journal of the American College of Surgeons.2022; 234(6): 1181.     CrossRef
  • The use and future perspective of Artificial Intelligence—A survey among German surgeons
    Mathieu Pecqueux, Carina Riediger, Marius Distler, Florian Oehme, Ulrich Bork, Fiona R. Kolbinger, Oliver Schöffski, Peter van Wijngaarden, Jürgen Weitz, Johannes Schweipert, Christoph Kahlert
    Frontiers in Public Health.2022;[Epub]     CrossRef
  • TIPTA YAPAY ZEKA UYGULAMALARI
    Hatice KELEŞ
    Kırıkkale Üniversitesi Tıp Fakültesi Dergisi.2022; 24(3): 604.     CrossRef
  • Application of Artificial Intelligence in Medicine: An Overview
    Peng-ran Liu, Lin Lu, Jia-yao Zhang, Tong-tong Huo, Song-xiang Liu, Zhe-wei Ye
    Current Medical Science.2021; 41(6): 1105.     CrossRef
  • Applications and Effects of EdTech in Medical Education
    Hyeonmi Hong, Youngjon Kim
    Korean Medical Education Review.2021; 23(3): 160.     CrossRef
  • Artificial Intelligence Education and Tools for Medical and Health Informatics Students: Systematic Review
    A Hasan Sapci, H Aylin Sapci
    JMIR Medical Education.2020; 6(1): e19285.     CrossRef
  • Scientific Development of Educational Artificial Intelligence in Web of Science
    Antonio-José Moreno-Guerrero, Jesús López-Belmonte, José-Antonio Marín-Marín, Rebeca Soler-Costa
    Future Internet.2020; 12(8): 124.     CrossRef
  • An Educational Network for Surgical Education Supported by Gamification Elements: Protocol for a Randomized Controlled Trial
    Natasha Guérard-Poirier, Michèle Beniey, Léamarie Meloche-Dumas, Florence Lebel-Guay, Bojana Misheva, Myriam Abbas, Malek Dhane, Myriam Elraheb, Adam Dubrowski, Erica Patocskai
    JMIR Research Protocols.2020; 9(12): e21273.     CrossRef

JEEHP : Journal of Educational Evaluation for Health Professions
TOP