Cargando…

How does ChatGPT-4 preform on non-English national medical licensing examination? An evaluation in Chinese language

ChatGPT, an artificial intelligence (AI) system powered by large-scale language models, has garnered significant interest in healthcare. Its performance dependent on the quality and quantity of training data available for a specific language, with the majority of it being in English. Therefore, its...

Descripción completa

Detalles Bibliográficos
Autores principales:	Fang, Changchang, Wu, Yuting, Fu, Wanying, Ling, Jitao, Wang, Yue, Liu, Xiaolin, Jiang, Yuan, Wu, Yifan, Chen, Yixuan, Zhou, Jing, Zhu, Zhichen, Yan, Zhiwei, Yu, Peng, Liu, Xiao
Formato:	Online Artículo Texto
Lenguaje:	English
Publicado:	Public Library of Science 2023
Materias:	Research Article
Acceso en línea:	https://www.ncbi.nlm.nih.gov/pmc/articles/PMC10691691/ https://www.ncbi.nlm.nih.gov/pubmed/38039286 http://dx.doi.org/10.1371/journal.pdig.0000397

_version_	1785152788365312000
author	Fang, Changchang Wu, Yuting Fu, Wanying Ling, Jitao Wang, Yue Liu, Xiaolin Jiang, Yuan Wu, Yifan Chen, Yixuan Zhou, Jing Zhu, Zhichen Yan, Zhiwei Yu, Peng Liu, Xiao
author_facet	Fang, Changchang Wu, Yuting Fu, Wanying Ling, Jitao Wang, Yue Liu, Xiaolin Jiang, Yuan Wu, Yifan Chen, Yixuan Zhou, Jing Zhu, Zhichen Yan, Zhiwei Yu, Peng Liu, Xiao
author_sort	Fang, Changchang
collection	PubMed
description	ChatGPT, an artificial intelligence (AI) system powered by large-scale language models, has garnered significant interest in healthcare. Its performance dependent on the quality and quantity of training data available for a specific language, with the majority of it being in English. Therefore, its effectiveness in processing the Chinese language, which has fewer data available, warrants further investigation. This study aims to assess the of ChatGPT’s ability in medical education and clinical decision-making within the Chinese context. We utilized a dataset from the Chinese National Medical Licensing Examination (NMLE) to assess ChatGPT-4’s proficiency in medical knowledge in Chinese. Performance indicators, including score, accuracy, and concordance (confirmation of answers through explanation), were employed to evaluate ChatGPT’s effectiveness in both original and encoded medical questions. Additionally, we translated the original Chinese questions into English to explore potential avenues for improvement. ChatGPT scored 442/600 for original questions in Chinese, surpassing the passing threshold of 360/600. However, ChatGPT demonstrated reduced accuracy in addressing open-ended questions, with an overall accuracy rate of 47.7%. Despite this, ChatGPT displayed commendable consistency, achieving a 75% concordance rate across all case analysis questions. Moreover, translating Chinese case analysis questions into English yielded only marginal improvements in ChatGPT’s performance (p = 0.728). ChatGPT exhibits remarkable precision and reliability when handling the NMLE in Chinese. Translation of NMLE questions from Chinese to English does not yield an improvement in ChatGPT’s performance.
format	Online Article Text
id	pubmed-10691691
institution	National Center for Biotechnology Information
language	English
publishDate	2023
publisher	Public Library of Science
record_format	MEDLINE/PubMed
spelling	pubmed-106916912023-12-02 How does ChatGPT-4 preform on non-English national medical licensing examination? An evaluation in Chinese language Fang, Changchang Wu, Yuting Fu, Wanying Ling, Jitao Wang, Yue Liu, Xiaolin Jiang, Yuan Wu, Yifan Chen, Yixuan Zhou, Jing Zhu, Zhichen Yan, Zhiwei Yu, Peng Liu, Xiao PLOS Digit Health Research Article ChatGPT, an artificial intelligence (AI) system powered by large-scale language models, has garnered significant interest in healthcare. Its performance dependent on the quality and quantity of training data available for a specific language, with the majority of it being in English. Therefore, its effectiveness in processing the Chinese language, which has fewer data available, warrants further investigation. This study aims to assess the of ChatGPT’s ability in medical education and clinical decision-making within the Chinese context. We utilized a dataset from the Chinese National Medical Licensing Examination (NMLE) to assess ChatGPT-4’s proficiency in medical knowledge in Chinese. Performance indicators, including score, accuracy, and concordance (confirmation of answers through explanation), were employed to evaluate ChatGPT’s effectiveness in both original and encoded medical questions. Additionally, we translated the original Chinese questions into English to explore potential avenues for improvement. ChatGPT scored 442/600 for original questions in Chinese, surpassing the passing threshold of 360/600. However, ChatGPT demonstrated reduced accuracy in addressing open-ended questions, with an overall accuracy rate of 47.7%. Despite this, ChatGPT displayed commendable consistency, achieving a 75% concordance rate across all case analysis questions. Moreover, translating Chinese case analysis questions into English yielded only marginal improvements in ChatGPT’s performance (p = 0.728). ChatGPT exhibits remarkable precision and reliability when handling the NMLE in Chinese. Translation of NMLE questions from Chinese to English does not yield an improvement in ChatGPT’s performance. Public Library of Science 2023-12-01 /pmc/articles/PMC10691691/ /pubmed/38039286 http://dx.doi.org/10.1371/journal.pdig.0000397 Text en © 2023 Fang et al https://creativecommons.org/licenses/by/4.0/This is an open access article distributed under the terms of the Creative Commons Attribution License (https://creativecommons.org/licenses/by/4.0/) , which permits unrestricted use, distribution, and reproduction in any medium, provided the original author and source are credited.
spellingShingle	Research Article Fang, Changchang Wu, Yuting Fu, Wanying Ling, Jitao Wang, Yue Liu, Xiaolin Jiang, Yuan Wu, Yifan Chen, Yixuan Zhou, Jing Zhu, Zhichen Yan, Zhiwei Yu, Peng Liu, Xiao How does ChatGPT-4 preform on non-English national medical licensing examination? An evaluation in Chinese language
title	How does ChatGPT-4 preform on non-English national medical licensing examination? An evaluation in Chinese language
title_full	How does ChatGPT-4 preform on non-English national medical licensing examination? An evaluation in Chinese language
title_fullStr	How does ChatGPT-4 preform on non-English national medical licensing examination? An evaluation in Chinese language
title_full_unstemmed	How does ChatGPT-4 preform on non-English national medical licensing examination? An evaluation in Chinese language
title_short	How does ChatGPT-4 preform on non-English national medical licensing examination? An evaluation in Chinese language
title_sort	how does chatgpt-4 preform on non-english national medical licensing examination? an evaluation in chinese language
topic	Research Article
url	https://www.ncbi.nlm.nih.gov/pmc/articles/PMC10691691/ https://www.ncbi.nlm.nih.gov/pubmed/38039286 http://dx.doi.org/10.1371/journal.pdig.0000397
work_keys_str_mv	AT fangchangchang howdoeschatgpt4preformonnonenglishnationalmedicallicensingexaminationanevaluationinchineselanguage AT wuyuting howdoeschatgpt4preformonnonenglishnationalmedicallicensingexaminationanevaluationinchineselanguage AT fuwanying howdoeschatgpt4preformonnonenglishnationalmedicallicensingexaminationanevaluationinchineselanguage AT lingjitao howdoeschatgpt4preformonnonenglishnationalmedicallicensingexaminationanevaluationinchineselanguage AT wangyue howdoeschatgpt4preformonnonenglishnationalmedicallicensingexaminationanevaluationinchineselanguage AT liuxiaolin howdoeschatgpt4preformonnonenglishnationalmedicallicensingexaminationanevaluationinchineselanguage AT jiangyuan howdoeschatgpt4preformonnonenglishnationalmedicallicensingexaminationanevaluationinchineselanguage AT wuyifan howdoeschatgpt4preformonnonenglishnationalmedicallicensingexaminationanevaluationinchineselanguage AT chenyixuan howdoeschatgpt4preformonnonenglishnationalmedicallicensingexaminationanevaluationinchineselanguage AT zhoujing howdoeschatgpt4preformonnonenglishnationalmedicallicensingexaminationanevaluationinchineselanguage AT zhuzhichen howdoeschatgpt4preformonnonenglishnationalmedicallicensingexaminationanevaluationinchineselanguage AT yanzhiwei howdoeschatgpt4preformonnonenglishnationalmedicallicensingexaminationanevaluationinchineselanguage AT yupeng howdoeschatgpt4preformonnonenglishnationalmedicallicensingexaminationanevaluationinchineselanguage AT liuxiao howdoeschatgpt4preformonnonenglishnationalmedicallicensingexaminationanevaluationinchineselanguage

How does ChatGPT-4 preform on non-English national medical licensing examination? An evaluation in Chinese language

Ejemplares similares