Aller au contenu principal
Accès ouvert déclaré 2025 article

Comparison of GPT-5 Responses With the Official Results of the Polish Specialized Psychiatric Examination in Child and Adolescent Psychiatry

3Citations signalées, ce qui n’est pas une note de qualité
6Institutions déclarées
4Pays d’affiliation déclarés

Rattachement africain : pl, be, ir, it. Niveau de preuve : code pays fourni par la source.

Le résumé fourni par la source

Introduction Artificial intelligence (AI), particularly language models such as ChatGPT (OpenAI, San Francisco, CA, USA), is becoming increasingly important in medical education and knowledge assessment. Prior studies have demonstrated the growing effectiveness of AI in preparing students for medical examinations, including the Medical Final Examination (Lekarski Egzamin Końcowy (LEK)) of Poland and the National Specialty Examination across various disciplines. This raises important questions regarding its potential role as a tool to support specialist training. Objective The aim of this study is to evaluate the effectiveness of the advanced GPT-5 model in addressing problems in child and adolescent psychiatry. The focus is on the accuracy of answers, their correctness, and the model's self-declared confidence levels to assess its potential efficacy in education. Methodology The study analyzed the official spring 2025 National Specialty Examination (Państwowy Egzamin Specjalizacyjny (PES)) of Poland in child and adolescent psychiatry. The exam consisted of 120 multiple-choice questions with a single correct answer. GPT-5 was familiarized with the examination rules and then presented with the questions in the Polish language. Answers were evaluated using the official Centre for Medical Examination (CEM) key. In addition, the model provided a confidence rating for each answer on a five-point scale. Questions were categorized as either clinical or theoretical. Statistical analysis was conducted using the chi-square test and the Mann-Whitney U test. Results GPT-5 answered 97 questions correctly (80.8%), surpassing the required passing threshold. No significant difference was observed between the accuracy of responses to clinical versus theoretical questions (p = 0.399). However, correct answers were significantly more likely when the model reported higher confidence levels (p = 0.012). Conclusions GPT-5 demonstrated strong performance in the National Specialty Examination of Poland in child and adolescent psychiatry, supporting its potential as a supplementary tool in specialist education. Confidence ratings may provide an additional metric for evaluating the reliability of answers. Nevertheless, broader integration of AI in medical education requires experts overseeing the process and further research across diverse medical disciplines.

Ce résumé expose les affirmations des auteurs. BNTIC ne l’interprète pas comme une validation indépendante des résultats.

Le contrôle bibliographique ouvert

DOI retrouvé dans Crossref DOI retrouvé ; titre concordant.

Titre Crossref
Comparison of GPT-5 Responses With the Official Results of the Polish Specialized Psychiatric Examination in Child and Adolescent Psychiatry
Date Crossref
22/09/2025
Éditeur
Springer Science and Business Media LLC
Type
journal-article

Ce recoupement confirme des métadonnées liées au DOI. Il ne confirme ni la méthode ni les conclusions de l’étude, et il ne compte pas comme une seconde source scientifique indépendante.

Les institutions déclarées

Une affiliation ne permet pas de déduire la nationalité d’un auteur.

Les sujets associés

Artificial Intelligence in Healthcare and EducationHealthcare cost, quality, practicesClinical Reasoning and Diagnostic Skills

BNTIC News n’est pas le producteur de ces données. Les publications sont interrogées à la demande dans Crossref, OpenAIRE, DOAJ, Europe PMC, HAL, DataCite, AfricArXiv, ROR et la Banque mondiale, sans clé d’accès. OpenAlex reste optionnel. Aucun service payant n’est nécessaire et aucune donnée externe n’est enregistrée en base. Consulter les sources et leurs limites.