Generative pretrained transformer 4:an innovative approach to facilitate value-based healthcare  

在线阅读下载全文

作  者:Han Lyu Zhixiang Wang Jia Li Jing Sun Xinghao Wang Pengling Ren Linkun Cai Zhenchang Wang Max Wintermark 

机构地区:[1]Department of Radiology,Capital Medical University Affiliated Beijing Friendship Hospital,Beijing 100050,China [2]Department of Medical Imaging,Capital Medical University Affiliated Beijing Friendship Hospital,Beijing 100050,China [3]School of Biological Science and Medical Engineering,Beihang University,Beijing 100191,China [4]Department of Neuroradiology,University of Texas MD Anderson Center,Houston,TX 77030,USA

出  处:《Intelligent Medicine》2024年第1期10-15,共6页智慧医学(英文)

基  金:National Natural Science Foundation of China(Grant Nos.62171297 and 61931013).

摘  要:Objective Appropriate medical imaging is important for value-based care.We aim to evaluate the performance of generative pretrained transformer 4(GPT-4),an innovative natural language processing model,providing appropriate medical imaging automatically in different clinical scenarios.Methods Institutional Review Boards(IRB)approval was not required due to the use of nonidentifiable data.Instead,we used 112 questions from the American College of Radiology(ACR)Radiology-TEACHES Program as prompts,which is an open-sourced question and answer program to guide appropriate medical imaging.We included 69 free-text case vignettes and 43 simplified cases.For the performance evaluation of GPT-4 and GPT-3.5,we considered the recommendations of ACR guidelines as the gold standard,and then three radiologists analyzed the consistency of the responses from the GPT models with those of the ACR.We set a five-score criterion for the evaluation of the consistency.A paired t-test was applied to assess the statistical significance of the findings.Results For the performance of the GPT models in free-text case vignettes,the accuracy of GPT-4 was 92.9%,whereas the accuracy of GPT-3.5 was just 78.3%.GPT-4 can provide more appropriate suggestions to reduce the overutilization of medical imaging than GPT-3.5(t=3.429,P=0.001).For the performance of the GPT models in simplified scenarios,the accuracy of GPT-4 and GPT-3.5 was 66.5%and 60.0%,respectively.The differences were not statistically significant(t=1.858,P=0.070).GPT-4 was characterized by longer reaction times(27.1 s in average)and extensive responses(137.1 words on average)than GPT-3.5.Conclusion As an advanced tool for improving value-based healthcare in clinics,GPT-4 may guide appropriate medical imaging accurately and efficiently。

关 键 词:Generative pretrained transformer 4 model Natural language processing Medical imaging APPROPRIATENESS 

分 类 号:R197.1[医药卫生—卫生事业管理]

 

参考文献:

正在载入数据...

 

二级参考文献:

正在载入数据...

 

耦合文献:

正在载入数据...

 

引证文献:

正在载入数据...

 

二级引证文献:

正在载入数据...

 

同被引文献:

正在载入数据...

 

相关期刊文献:

正在载入数据...

相关的主题
相关的作者对象
相关的机构对象