ChatGPT Conquers the Saudi Medical Licensing Exam: Exploring the Accuracy of Artificial Intelligence in Medical Knowledge Assessment and Implications for Modern Medical Education.

Fahad K Aljindan,Omar A Aldamigh,Abeer Mohammed M Alanazi,Faisal Falah Almutairi,Abdullah A Al Qurashi,Subhi M K Zino Alarki,Ibrahim R Halawani,Ibrahim Abdullah S Albalawi,Hussam Abdulkhaliq M Aljuhani

doi:10.7759/cureus.45043

Abstract

Background The application of artificial intelligence (AI) in education is undergoing rapid advancements, with models such as ChatGPT-4 showing potential in medical education. This study aims to evaluate the proficiency of ChatGPT-4 in answering Saudi Medical Licensing Exam (SMLE) questions. Methodology A dataset of 220 questions across four medical disciplines was used. The model was trained using a specific code to answer the questions accurately, and its performance was assessed using key performance indicators, difficulty level, and exam sections. Results ChatGPT-4 demonstrated an overall accuracy of 88.6%. It showed high proficiency with Easy and Average questions, but accuracy decreased for Hard questions. Performance was consistent across all disciplines, indicating a broad knowledge base. However, an error analysis revealed areas for further refinement, particularly with category (Option) A questions across all sections. Conclusions This study underscores the potential of ChatGPT-4 as an AI-assisted tool in medical education, demonstrating high proficiency in answering SMLE questions. Future research is recommended to expand the scope of training and evaluation as well as to enhance the model's performance on complex clinical questions.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

ChatGPT Conquers the Saudi Medical Licensing Exam: Exploring the Accuracy of Artificial Intelligence in Medical Knowledge Assessment and Implications for Modern Medical Education.

Abstract

Talk to us

Similar Papers

More From: Cureus

Lead the way for us

Similar Papers

Success of ChatGPT, an AI language model, in taking the French language version of the European Board of Ophthalmology examination: A novel approach to medical knowledge assessment
C Panthier ... D Gatinel
Journal Français d'Ophtalmologie | VOL. 46
C Panthier, et. al.C Panthier ... D Gatinel
01 Aug 2023
Journal Français d'Ophtalmologie | VOL. 46

Knowledge and attitudes of medical students in Lebanon toward artificial intelligence: A national survey study.
George Doumat ... Nadim-Nicolas Ghanem
Frontiers in Artificial Intelligence | VOL. 5
George Doumat, et. al.George Doumat ... Nadim-Nicolas Ghanem
02 Nov 2022
Frontiers in Artificial Intelligence | VOL. 5

The role of artificial intelligence in higher medical education and the ethical challenges of its implementation
Mark Perkins ... Agnieszka Pregowska
Artificial Intelligence in Health | VOL. 0
Mark Perkins, et. al.Mark Perkins ... Agnieszka Pregowska
21 Oct 2024
Artificial Intelligence in Health | VOL. 0

Reshaping medical education: Performance of ChatGPT on a PES medical examination.
Simona Wójcik ... Marcin Poboży
Cardiology journal | VOL. 31
Simona Wójcik, et. al.Simona Wójcik ... Marcin Poboży
28 Jun 2024
Cardiology journal | VOL. 31

Journal: Cureus	Publication Date: Sep 11, 2023
Citations: 12

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

ChatGPT Conquers the Saudi Medical Licensing Exam: Exploring the Accuracy of Artificial Intelligence in Medical Knowledge Assessment and Implications for Modern Medical Education.

Abstract

Talk to us

Similar Papers

More From: Cureus