Aller au contenu principal
Accès ouvert déclaré 2025 article

Real-world feasibility of generative large language models for clinical decision support in benign prostatic hyperplasia

1Citations signalées, ce qui n’est pas une note de qualité
4Institutions déclarées
1Pays d’affiliation déclarés

Rattachement africain : cn. Niveau de preuve : code pays fourni par la source.

Le résumé fourni par la source

BACKGROUND: Benign prostatic hyperplasia (BPH) is a common condition among middle-aged and elderly men, often accompanied by lower urinary tract symptoms that significantly impact patients' quality of life. It has emerged as a major public health challenge worldwide. In recent years, artificial intelligence (AI), particularly large language models (LLMs), has shown great potential in supporting clinical decision-making. This study aimed to systematically evaluate the level of clinical knowledge mastery demonstrated by mainstream generative LLMs in the context of BPH and to further explore their decision-support capabilities in real-world clinical scenarios. METHODS: We assessed the clinical knowledge and decision-making capabilities of ChatGPT o1 and DeepSeek R1 in the field of BPH. A set of 30 clinically relevant questions was developed and submitted to both AI models. For comparison, clinical physicians from Chinese medical institutions answered the same questions under closed-book conditions. Additionally, six real-world BPH cases were used to simulate clinical diagnostic and treatment scenarios to further evaluate the performance of the two AI models in clinical decision-making. RESULTS: In the clinical knowledge assessment, both ChatGPT o1 and DeepSeek R1 outperformed the physician group, with no significant difference between the two models. Subgroup analyses revealed performance differences based on clinical knowledge categories and physician experience levels. DeepSeek R1 and ChatGPT o1 generally outperformed resident physicians and performed comparably to attending physicians. In terms of clinical decision support, DeepSeek R1 outperformed ChatGPT o1 in both accuracy of medical knowledge and logical coherence. CONCLUSION: Studies have shown that in the domain of clinical knowledge related to BPH, ChatGPT o1 and DeepSeek R1 perform at levels approaching those of attending physicians. Both models demonstrate substantial potential in supporting clinical decision-making, with DeepSeek R1 performing particularly well. Despite existing limitations, the application of AI in healthcare holds considerable promise and is highly anticipated.

Ce résumé expose les affirmations des auteurs. BNTIC ne l’interprète pas comme une validation indépendante des résultats.

Le contrôle bibliographique ouvert

DOI retrouvé dans Crossref DOI retrouvé ; titre concordant.

Titre Crossref
Real-world feasibility of generative large language models for clinical decision support in benign prostatic hyperplasia
Date Crossref
12/11/2025
Éditeur
Ovid Technologies (Wolters Kluwer Health)
Type
journal-article

Ce recoupement confirme des métadonnées liées au DOI. Il ne confirme ni la méthode ni les conclusions de l’étude, et il ne compte pas comme une seconde source scientifique indépendante.

Où se fait cette recherche

  • Affiliated Hospital of Southwest Medical University Department of Urology pays non établi dans la notice
    Établissement de santé
  • Southern Medical University Shenzhen Hospital pays non établi dans la notice
    Établissement de santé
  • Sichuan Mianyang 404 Hospital pays non établi dans la notice
    Établissement de santé
  • Affiliated Hospital of North Sichuan Medical College pays non établi dans la notice
    Établissement de santé
  • Shenzhen University Department of Urology pays non établi dans la notice
    Université ou école supérieure
  • Santai Hospital Affiliated to North Sichuan Medical College Department of Urology pays non établi dans la notice
    Université ou école supérieure

Department of Urology — Affiliated Hospital of Southwest Medical University, Southern Medical University Shenzhen Hospital et Sichuan Mianyang 404 Hospital, avec 3 autres affiliations.

Une affiliation ne permet pas de déduire la nationalité d’un auteur.

Les sujets associés

Machine Learning in HealthcareUrinary Bladder and Prostate ResearchArtificial Intelligence in Healthcare and Education

BNTIC News n’est pas le producteur de ces données. Les publications sont interrogées à la demande dans Crossref, OpenAIRE, DOAJ, Europe PMC, HAL, DataCite, AfricArXiv, ROR et la Banque mondiale, sans clé d’accès. OpenAlex reste optionnel. Aucun service payant n’est nécessaire et aucune donnée externe n’est enregistrée en base. Consulter les sources et leurs limites.