Can ChatGPT, an Artificial Intelligence Language Model, Provide Accurate and High-quality Patient Information on Prostate Cancer?

被引:64
作者
Coskun, Burhan [2 ]
Ocakoglu, Gokhan [1 ]
Yetemen, Melih [1 ]
Kaygisiz, Onur [1 ]
机构
[1] Bursa Uludag Univ, Dept Urol, Bursa, Turkiye
[2] Bursa Uludag Univ, Sch Med, Dept Urol, Gorukle Kampusu, TR-16059 Bursa, Turkiye
关键词
HEALTH INFORMATION; INTERNET;
D O I
10.1016/j.urology.2023.05.040
中图分类号
R5 [内科学]; R69 [泌尿科学(泌尿生殖系疾病)];
学科分类号
1002 ; 100201 ;
摘要
OBJECTIVE To evaluate the performance of ChatGPT, an artificial intelligence (AI) language model, in providing patient information on prostate cancer, and to compare the accuracy, similarity, and quality of the information to a reference source. METHODS Patient information material on prostate cancer was used as a reference source from the website of the European Association of Urology Patient Information. This was used to generate 59 queries. The accuracy of the model's content was determined with F1, precision, and recall scores. The similarity was assessed with cosine similarity, and the quality was evaluated using a 5RESULTS ChatGPT was able to respond to all prostate cancer-related queries. The average F1 score was 0.426 (range: 0-1), precision score was 0.349 (range: 0-1), recall score was 0.549 (range: 0-1), and cosine similarity was 0.609 (range: 0-1). The average GQS was 3.62 +/- 0.49 (range: 1-5), with no answers achieving the maximum GQS of 5. While ChatGPT produced a larger amount of information compared to the reference, the accuracy and quality of the content were not optimal, with all scores indicating need for improvement in the model's performance. CONCLUSION Caution should be exercised when using ChatGPT as a patient information source for prostate cancer due to limitations in its performance, which may lead to inaccuracies and potential misunderstandings. Further studies, using different topics and language models, are needed to fully understand the capabilities and limitations of AI-generated patient information. UROLOGY 180: 35-58, 2023. (c) 2023 Elsevier Inc. All rights reserved.
引用
收藏
页码:35 / 58
页数:24
相关论文
共 23 条
[1]   Development and Evaluation of Patient Information Leaflets (PIL) Usefulness [J].
Adepu, R. ;
Swamy, M. K. .
INDIAN JOURNAL OF PHARMACEUTICAL SCIENCES, 2012, 74 (02) :174-U1178
[2]  
Brown TB, 2020, Arxiv, DOI [arXiv:2005.14165, DOI 10.48550/ARXIV.2005.14165]
[3]  
Borji A, 2023, Arxiv, DOI [arXiv:2302.03494, DOI 10.48550/ARXIV.2302.03494, 10.48550/arxiv.2302.03494]
[4]   Early Detection of Prostate Cancer: AUA Guideline [J].
Carter, H. Ballentine ;
Albertsen, Peter C. ;
Barry, Michael J. ;
Etzioni, Ruth ;
Freedland, Stephen J. ;
Greene, Kirsten Lynn ;
Holmberg, Lars ;
Kantoff, Philip ;
Konety, Badrinath R. ;
Murad, Mohammad Hassan ;
Penson, David F. ;
Zietman, Anthony L. .
JOURNAL OF UROLOGY, 2013, 190 (02) :419-426
[5]   Prostate Cancer Screening [J].
Catalona, William J. .
MEDICAL CLINICS OF NORTH AMERICA, 2018, 102 (02) :199-+
[6]   DISCERN: an instrument for judging the quality of written consumer health information on treatment choices [J].
Charnock, D ;
Shepperd, S ;
Needham, G ;
Gann, R .
JOURNAL OF EPIDEMIOLOGY AND COMMUNITY HEALTH, 1999, 53 (02) :105-111
[7]   Plausible conditions and mechanisms for increasing physical activity behavior in men with prostate cancer using patient education interventions: sequential explanatory mixed studies synthesis [J].
Ezenwankwo, Elochukwu Fortune ;
Motsoeneng, Portia ;
Atterbury, Elizabeth Maria ;
Albertus, Yumna ;
Lambert, Estelle Victoria ;
Shamley, Delva .
SUPPORTIVE CARE IN CANCER, 2022, 30 (06) :4617-4633
[8]   YouTube as a Source of Information About Premature Ejaculation Treatment [J].
Gul, Murat ;
Diri, Mehmet Akif .
JOURNAL OF SEXUAL MEDICINE, 2019, 16 (11) :1734-1740
[9]   Survey of Hallucination in Natural Language Generation [J].
Ji, Ziwei ;
Lee, Nayeon ;
Frieske, Rita ;
Yu, Tiezheng ;
Su, Dan ;
Xu, Yan ;
Ishii, Etsuko ;
Bang, Ye Jin ;
Madotto, Andrea ;
Fung, Pascale .
ACM COMPUTING SURVEYS, 2023, 55 (12)
[10]  
Kadhim AI, 2014, IEEE ST CONF RES DEV