Saved in:
Bibliographic Details
Main Authors: Aygul, Yesim, Olucoglu, Muge, Alpkocak, Adil
Format: Preprint
Published: 2024
Subjects:
Online Access:https://arxiv.org/abs/2408.12305
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866917759649054720
author Aygul, Yesim
Olucoglu, Muge
Alpkocak, Adil
author_facet Aygul, Yesim
Olucoglu, Muge
Alpkocak, Adil
contents The potential of artificial intelligence in medical education and assessment has been made evident by recent developments in natural language processing and artificial intelligence. Medical questions can now be successfully answered by artificial intelligence algorithms. It can help medical practitioners. This study evaluates the performance of three different artificial intelligence models in answering Turkish medical questions in the 2021 1st Term Medical Specialization Examination (MSE). MSE consists of a total of 240 questions across clinical (CMST) and basic (BMST) medical sciences. According to the results in CMST, it was concluded that Gemini correctly answered 82 questions, ChatGPT-4 answered 105 questions and ChatGPT-4o answered 117 questions. In BMST, Gemini and ChatGPT-4 answered 93 questions and ChatGPT-4o answered 107 questions correctly according to the answer key. ChatGPT-4o outperformed the candidate with the highest scores of 113 and 106 according to CMST and BMST respectively. This study highlights the importance of the potential of artificial intelligence in medical education and assessment. It demonstrates that advanced models can achieve high accuracy and contextual understanding, demonstrating their potential role in medical education and evaluation.
format Preprint
id arxiv_https___arxiv_org_abs_2408_12305
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Tipta uzmanlik sinavinda (tus) buyuk dil modelleri insanlardan daha mi basarili?
Aygul, Yesim
Olucoglu, Muge
Alpkocak, Adil
Artificial Intelligence
The potential of artificial intelligence in medical education and assessment has been made evident by recent developments in natural language processing and artificial intelligence. Medical questions can now be successfully answered by artificial intelligence algorithms. It can help medical practitioners. This study evaluates the performance of three different artificial intelligence models in answering Turkish medical questions in the 2021 1st Term Medical Specialization Examination (MSE). MSE consists of a total of 240 questions across clinical (CMST) and basic (BMST) medical sciences. According to the results in CMST, it was concluded that Gemini correctly answered 82 questions, ChatGPT-4 answered 105 questions and ChatGPT-4o answered 117 questions. In BMST, Gemini and ChatGPT-4 answered 93 questions and ChatGPT-4o answered 107 questions correctly according to the answer key. ChatGPT-4o outperformed the candidate with the highest scores of 113 and 106 according to CMST and BMST respectively. This study highlights the importance of the potential of artificial intelligence in medical education and assessment. It demonstrates that advanced models can achieve high accuracy and contextual understanding, demonstrating their potential role in medical education and evaluation.
title Tipta uzmanlik sinavinda (tus) buyuk dil modelleri insanlardan daha mi basarili?
topic Artificial Intelligence
url https://arxiv.org/abs/2408.12305