• Home
  • Search
  • Comparing Diagnostic Accuracy of Clinical Professionals and Large Language Models: Systematic Review and Meta-Analysis
  • Cite Icon27
  • https://doi.org/10.2196/64963Copy DOI Icon

Comparing Diagnostic Accuracy of Clinical Professionals and Large Language Models: Systematic Review and Meta-Analysis

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

BackgroundWith the rapid development of artificial intelligence (AI) technology, especially generative AI, large language models (LLMs) have shown great potential in the medical field. Through massive medical data training, it can understand complex medical texts and can quickly analyze medical records and provide health counseling and diagnostic advice directly, especially in rare diseases. However, no study has yet compared and extensively discussed the diagnostic performance of LLMs with that of physicians.ObjectiveThis study systematically reviewed the accuracy of LLMs in clinical diagnosis and provided reference for further clinical application.MethodsWe conducted searches in CNKI (China National Knowledge Infrastructure), VIP Database, SinoMed, PubMed, Web of Science, Embase, and CINAHL (Cumulative Index to Nursing and Allied Health Literature) from January 1, 2017, to the present. A total of 2 reviewers independently screened the literature and extracted relevant information. The risk of bias was assessed using the Prediction Model Risk of Bias Assessment Tool (PROBAST), which evaluates both the risk of bias and the applicability of included studies.ResultsA total of 30 studies involving 19 LLMs and a total of 4762 cases were included. The quality assessment indicated a high risk of bias in the majority of studies, primary cause is known case diagnosis. For the optimal model, the accuracy of the primary diagnosis ranged from 25% to 97.8%, while the triage accuracy ranged from 66.5% to 98%.ConclusionsLLMs have demonstrated considerable diagnostic capabilities and significant potential for application across various clinical cases. Although their accuracy still falls short of that of clinical professionals, if used cautiously, they have the potential to become one of the best intelligent assistants in the field of human health care.

Similar Papers
  • Research Article

Risk prediction models for permanent pacemaker implantation following transcatheter aortic valve replacement: a systematic review and meta-analysis

  • Sep 25, 2025
  • Frontiers in Cardiovascular Medicine
  • Yijun Mao +8
  • Research Article
  • Citations6

Risk prediction models for dental caries in children and adolescents: a systematic review and meta-analysis

  • Mar 01, 2025
  • BMJ Open
  • Xijia Wang +8
  • PDF
  • Research Article
  • Citations17

Prediction model for cognitive frailty in older adults: A systematic review and critical appraisal

  • Apr 12, 2023
  • Frontiers in Aging Neuroscience
  • Jundan Huang +6
  • Research Article
  • Citations3

Risk prediction model for chemotherapy-induced nausea and vomiting in cancer patients: a systematic review.

  • Aug 01, 2025
  • International journal of nursing studies
  • Yongjian Wang +7
  • Research Article
  • Citations8

Risk prediction models for postoperative delirium in elderly patients with fragility hip fracture: A systematic review and critical appraisal

  • Dec 10, 2023
  • International journal of orthopaedic and trauma nursing
  • Bingqian Zhou +2
  • Research Article
  • Citations2

Risk prediction models for depression in patients with coronary heart disease: a systematic review and meta-analysis

  • Jan 15, 2025
  • Frontiers in Cardiovascular Medicine
  • Jie Zhang +5
  • PDF
  • Supplementary Content
  • Citations13

A pooled analysis of the risk prediction models for mortality in acute exacerbation of chronic obstructive pulmonary disease

  • Mar 21, 2023
  • The Clinical Respiratory Journal
  • Zile Ji +4
  • Research Article
  • Citations62

Risk of bias of prognostic models developed using machine learning: a systematic review in oncology

  • Jul 07, 2022
  • Diagnostic and Prognostic Research
  • Paula Dhiman +11
  • Supplementary Content

Bias and Reporting Quality of Clinical Prognostic Models for Idiopathic Pulmonary Fibrosis: A Cross-Sectional Study

  • Jun 08, 2022
  • Risk Management and Healthcare Policy
  • Jiaqi Di +4
  • Research Article

Predictive models for ICU patient readmission based on machine learning: A systematic review.

  • Apr 03, 2026
  • Journal of the Intensive Care Society
  • Zhixiang Zheng +3
  • Supplementary Content

Are We Accurately Predicting Mortality in Renal Cancer? A Systematic Review of Prognostic Models

  • Aug 19, 2025
  • Journal of Clinical Medicine
  • Laura Martinez-Cayuelas +5
  • PDF
  • Research Article
  • Citations7

Optimal surveillance strategies for patients with stage 1 cutaneous melanoma post primary tumour excision: three systematic reviews and an economic model.

  • Nov 01, 2021
  • Health Technology Assessment
  • Luke Vale +23
  • PDF
  • Research Article
  • Citations34

Cardiovascular disease risk prediction models in the Chinese population- a systematic review and meta-analysis

  • Aug 24, 2022
  • BMC Public Health
  • Guo Zhiting +5
  • Research Article

Risk prediction models for extubation failure in critically ill patients on mechanical ventilation: a systematic review

  • Nov 20, 2025
  • Frontiers in Medicine
  • Xiang Zeng +5
  • Research Article

Risk Prediction Models for Sarcopenia in Patients Undergoing Maintenance Haemodialysis: A Systematic Review and Meta-Analysis.

  • Apr 04, 2025
  • Journal of clinical nursing
  • Qing Yang +6
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.